AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Model Releases

LoMime: Query-Efficient Membership Inference using Model Extraction in Label-Only Settings

DGX agent

arXiv:2602.18934v2 Announce Type: replace Abstract: Membership inference attacks (MIAs) threaten the privacy of machine learning models by revealing whether a specific data point was used during train

model-releasesarxiv-cs-lg
24 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Model selection with proper scoring rules on data sets of time series

DGX agent

arXiv:2606.24715v1 Announce Type: cross Abstract: We consider the problem of model selection between probabilistic models on data sets of time series. Chosen a proper scoring rule, we denote by the te

researcharxiv-cs-lg
24 Jun 2026
Model Releases

Assistron: Bayesian Shared Autonomy with Off-the-shelf Vision-Language-Action Models

DGX agent

arXiv:2606.23147v1 Announce Type: new Abstract: We propose Assistron, a shared autonomy model that leverages Vision-Language-Action (VLA) models to assist the user in daily activities. Our approach is

model-releasesarxiv-cs-ro
23 Jun 2026
Model Releases

Bayesian Adaptation Gym: A Benchmark for the Bayesian Low-Rank Adaptation of Multi-Modal Language Models

DGX agent

arXiv:2606.22188v1 Announce Type: new Abstract: Large multi-modal language models are increasingly deployed in high-stakes domains, making well-calibrated uncertainty essential. Traditional Bayesian m

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

CAT-Translate: Building Compact Open-Source Models for Japanese-English Translation

DGX agent

arXiv:2606.21413v1 Announce Type: cross Abstract: Nowadays, large multilingual translation models demonstrate impressive translation capabilities in the machine translation benchmarks. This raises a p

applicationsarxiv-cs-lg
23 Jun 2026
Research

Frequency-Domain Neural ODEs for Modeling Non-Linear Dynamical Systems

DGX agent

arXiv:2606.22075v1 Announce Type: new Abstract: Standard continuous-depth models, such as Neural Ordinary Differential Equations (NODEs), offer significant advantages in modeling physical systems by l

researcharxiv-cs-lg
23 Jun 2026
Model Releases

IndicGuard: A Multilingual Safety Guard Model and Dataset for Indic Languages

DGX agent

arXiv:2606.22841v1 Announce Type: cross Abstract: As Large Language Models (LLMs) achieve widespread integration across diverse linguistic landscapes, ensuring their safety and alignment with regional

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

MOOZY: A Patient-First Foundation Model for Computational Pathology

DGX agent

arXiv:2603.27048v3 Announce Type: replace Abstract: Computational pathology needs whole-slide image (WSI) foundation models that transfer across diverse clinical tasks, yet current approaches remain l

model-releasesarxiv-cs-cv
23 Jun 2026
Research

One-Step Flow Matching for Generative Modeling of Path-Dependent Physical Fields

DGX agent

arXiv:2606.22752v1 Announce Type: new Abstract: Physical simulations for intricate geometries with path-dependent constitutive models face difficulties due to the enormous computational cost they requ

researcharxiv-cs-lg
23 Jun 2026
Tutorials

Provable Benefits of RLVR over SFT for Reasoning Models: Learning to Backtrack Efficiently

DGX agent

arXiv:2606.22938v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have demonstrated that reinforcement fine-tuning of pretrained base models can lead to significant gains

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Training-free Task Classification for Multi-Task Model Merging

DGX agent

arXiv:2606.22589v1 Announce Type: new Abstract: Ever since the advent of foundation models and the pre-training-finetuning paradigm, there have been numerous efforts to merge multiple task-specific ex

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

A theory of learning data statistics in diffusion models, from easy to hard

DGX agent

arXiv:2603.12901v2 Announce Type: replace-cross Abstract: While diffusion models have emerged as a powerful class of generative models, their learning dynamics remain poorly understood. We address thi

safetyarxiv-cs-lg
11 Jun 2026
Model Releases

Calibration Drift Under Reasoning: How Chain-of-Thought Budgets Induce Overconfidence in Large Language Models

DGX agent

arXiv:2606.11211v1 Announce Type: cross Abstract: The ability of large language models (LLMs) to express calibrated uncertainty is important for safe deployment. Chain-of-thought (CoT) reasoning is wi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models

DGX agent

arXiv:2606.11324v1 Announce Type: cross Abstract: We introduce Embodied-R1.5, a unified Embodied Foundation Model (EFM) that integrates comprehensive embodied reasoning capabilities, spanning embodied

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

Making Foresight Actionable: Repurposing Representation Alignment in World Action Models

DGX agent

arXiv:2606.12217v1 Announce Type: cross Abstract: World Action Models (WAMs) offer a promising route for robot manipulation by using video generation models to model future scene evolution before prod

local-aiarxiv-cs-ai
11 Jun 2026
Model Releases

Physically Constrained Ensemble Gaussian Process Modelling for Expensive Quantum Systems with Heteroskedastic Noise

DGX agent

arXiv:2606.11240v1 Announce Type: cross Abstract: Accurate modeling of quantum many-body systems often requires computationally expensive simulations such as Density Matrix Renormalization Group (DMRG

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

To Intervene or Not: Guiding Inference-time Alignment with Probabilistic Model Blending

DGX agent

arXiv:2606.11201v1 Announce Type: cross Abstract: The wide deployment of LLMs has made model alignment necessary to make newly trained models safely and effectively respond to user instructions. Among

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

CableRobotGraphSim: A Graph Neural Network for Modeling Partially Observable Cable-Driven Robot Dynamics

DGX agent

arXiv:2602.21331v2 Announce Type: replace Abstract: General-purpose simulators have accelerated the development of robots. Traditional simulators based on first-principles, however, typically require

model-releasesarxiv-cs-ro
10 Jun 2026
Research

Data assimilation for subsurface flow using latent diffusion model parameterization: performance of ensemble-Kalman and Monte Carlo techniques

DGX agent

arXiv:2606.11140v1 Announce Type: cross Abstract: Data assimilation (DA) in subsurface flow entails calibrating model parameters to match observed data, typically at wells, while preserving geological

researcharxiv-cs-ai
10 Jun 2026
Model Releases

Do Vision-Language Models See or Guess? Measuring and Reducing Textual-Prior Reliance with a Phrasing-Controlled Benchmark

DGX agent

arXiv:2606.10400v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed where answers must follow from what is in the image, yet they often answer from textual priors,

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

FailureScope: Cross-Regime Behavioral Diagnosis of Language Model Weaknesses

DGX agent

arXiv:2606.09878v1 Announce Type: new Abstract: Standard benchmarks report aggregate accuracy, but practitioners need to know which specific capabilities a model lacks. We introduce FailureScope, a be

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

Rotate2Think: Geometric Priming via Orthogonal Rotation to Improve Language Model Reasoning

DGX agent

arXiv:2606.09873v1 Announce Type: cross Abstract: Reasoning models achieve strong performance on challenging tasks by generating explicit intermediate reasoning traces before producing a final answer.

model-releasesarxiv-cs-ai
10 Jun 2026
Research

What Really Matters for Table LLMs? A Meta-Evaluation of Model and Data Effects

DGX agent

arXiv:2501.14717v2 Announce Type: replace Abstract: Table modeling has progressed for decades. In this work, we revisit this trajectory and highlight emerging challenges in the LLM era, particularly t

researcharxiv-cs-cl
10 Jun 2026
Model Releases

A Dataset for Dynamic Human Preferences for Vision Language Models

DGX agent

arXiv:2606.07653v1 Announce Type: cross Abstract: Given the increased adoption of Vision Language Models (VLMs) in human-interactive settings, it is important that we evaluate how well these models ca

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

A retrieval conditioned rebinding circuit for dynamic entity tracking in large language models

DGX agent

arXiv:2606.08644v1 Announce Type: cross Abstract: To interpret context correctly and retrieve relevant information, large language models must bind entities to their attributes and update these bindin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Evaluating Hallucinations in Domain-Adapted Large Language Models

DGX agent

arXiv:2606.07521v1 Announce Type: cross Abstract: This study investigates the phenomenon of hallucinations in domain-adapted Large Language Models (LLMs), focusing on the fine-tuning of the Llama-2 mo

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Finite Certificates for In-Context Determinacy and a Threshold Theory of Emergence in Language Models

DGX agent

arXiv:2606.07623v1 Announce Type: new Abstract: This paper develops a model-theoretic framework for verifying context-conditioned language-model behavior by replacing benchmark labels with finite sema

model-releasesarxiv-cs-lg
9 Jun 2026
Model Releases

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

DGX agent

arXiv:2510.12171v2 Announce Type: replace Abstract: Large Language Models have shown strong scientific reasoning ability, but their performance on materials science problems remains less studied. To f

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Language Models

DGX agent

arXiv:2606.07532v1 Announce Type: cross Abstract: RLHF-trained models are systematically biased toward agreement over accuracy, a structural property of the training process. We present Principled Age

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Stress-testing medical large language models reveals latent safety pathology beyond benchmark accuracy

DGX agent

arXiv:2606.07929v1 Announce Type: new Abstract: Large language models (LLMs) are entering clinical practice based on benchmark accuracy that may fail to detect safety-relevant failure modes. Here we p

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

DGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Audio-Visual World Models: Grounding Multisensory Imagination for Embodied Agents

DGX agent

arXiv:2512.00883v3 Announce Type: replace-cross Abstract: World models simulate environmental dynamics to enable agents to plan and reason about future states. While existing approaches have primarily

model-releasesarxiv-cs-cv
8 Jun 2026
Model Releases

Endogenous Resistance to Activation Steering in Language Models

DGX agent

arXiv:2602.06941v2 Announce Type: replace-cross Abstract: Large language models can recover mid-generation from task-misaligned activation steering, producing explicit verbal restarts (e.g., ``wait, t

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Evidence Graph Consistency in Retrieval-Augmented Generation: A Model-Dependent Analysis of Hallucination Detection

DGX agent

arXiv:2606.06748v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) reduces but does not eliminate hallucination in large language models. Existing detection methods rely on flat si

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Explain Like I'm 5 or Whatever I Choose: Evaluating the Interactive Potential of Language Model Responses

DGX agent

arXiv:2606.06788v1 Announce Type: new Abstract: Evaluations of large language models (LLMs) in scientific information seeking tasks have become increasingly use-centric, such as conducting live or mul

model-releasesarxiv-cs-cl
8 Jun 2026
Research

Modular Monolingual Adaptation using Pretrained Language Models

DGX agent

arXiv:2606.06738v1 Announce Type: new Abstract: Building monolingual language models (LMs) for low-resource languages typically relies on adapting pretrained language models (PLMs) by finetuning the w

researcharxiv-cs-cl
8 Jun 2026
Model Releases

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

DGX agent

arXiv:2606.07237v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summariz

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

No Need to Train Your RDB Foundation Model

DGX agent

arXiv:2602.13697v2 Announce Type: replace Abstract: Relational databases (RDBs) contain vast amounts of heterogeneous tabular information that can be exploited for predictive modeling purposes. But si

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Humans' ALMANAC: A Human Collaboration Dataset of Action-Level Mental Model Annotations for Agent Collaboration

DGX agent

arXiv:2606.06388v1 Announce Type: cross Abstract: Recent advances in LLM agents have enabled complex cognitive capabilities, such as multi-step reasoning, planning, and tool use, that increasingly pos

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Knowledge Distillation for Visual Autoregressive Models

DGX agent

arXiv:2606.06078v1 Announce Type: new Abstract: Autoregressive (AR) image generation models are highly expressive but computationally intensive, motivating effective model compression. Knowledge disti

researcharxiv-cs-cv
5 Jun 2026
Model Releases

ChessMimic: Per-Rating Transformer Models for Human Move, Clock, and Outcome Prediction in Online Blitz Chess

DGX agent

arXiv:2606.04473v1 Announce Type: cross Abstract: We present ChessMimic, a system of three small encoder-only transformers - for move, thinking-time, and outcome prediction - conditioned on the positi

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models

DGX agent

arXiv:2510.20042v3 Announce Type: replace Abstract: Generative image models produce striking visuals yet often misrepresent culture. Prior work has examined cultural bias mainly in text-to-image (T2I)

model-releasesarxiv-cs-cv
4 Jun 2026
Model Releases

Flow Matching Calibration for Simulation-Based Inference under Model Misspecification

DGX agent

arXiv:2509.23385v5 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) is transforming experimental sciences by enabling parameter estimation in complex non-linear models from simu

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

From Symbolic to Geometric: Enabling Spatial Reasoning in Large Language Models

DGX agent

arXiv:2606.04381v1 Announce Type: cross Abstract: Recent large language models (LLMs) often appear to exhibit spatial reasoning ability; however, this capability is largely symbolic, arising from patt

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

A 3D Isovist World Model -- Revealing a City's Unseen Geometry and Its Emergent Cross-City Signature

DGX agent

arXiv:2606.03609v1 Announce Type: cross Abstract: Embodied agents that navigate cities rely on world models that predict how their surroundings will change as they move. But for navigation, what matte

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

CoEval: Ranking Language Models for Custom Tasks Without Labeled Data or Trustworthy Benchmarks

DGX agent

arXiv:2606.03650v1 Announce Type: cross Abstract: Choosing or ranking language models for a specific application is hardest when no task-specific labeled data exists, and standard public benchmarks ca

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

Knowledge Editing in Masked Diffusion Language Models

DGX agent

arXiv:2606.03924v1 Announce Type: new Abstract: Knowledge editing aims to update or correct factual knowledge in a language model. A widely used approach, locate-then-edit, does this in two steps: it

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

Staying Alive: Uncensored Survival Analysis with Tabular Foundation Models

DGX agent

arXiv:2606.03689v1 Announce Type: cross Abstract: Survival Analysis (SA) is a statistical framework that models the time span until some event of interest occurs. Widely used in several domains, inclu

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…3435363738…1021
Next →