AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,404 results
Model Releases

Benchmarking Attribute Discrimination in Infant-Scale Vision-Language Models

DGX agent

arXiv:2512.18951v3 Announce Type: replace Abstract: Infants learn not only object categories but also fine-grained visual attributes such as color, size, and texture from limited experience. Prior inf

model-releasesarxiv-cs-lg
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Characterizing Universal Object Representations Across Vision Models

DGX agent

arXiv:2605.13675v1 Announce Type: new Abstract: Deep neural networks trained with different architectures, objectives, and datasets have been reported to converge on similar visual representations. Ho

safetyarxiv-cs-cv
14 May 2026
Model Releases

DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions

DGX agent

arXiv:2509.19538v2 Announce Type: replace-cross Abstract: Diffusion-based world models have demonstrated strong capabilities in synthesizing realistic long-horizon trajectories for offline reinforceme

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Domain Adaptation of Large Language Models for Polymer-Composite Additive Manufacturing Using Retrieval-Augmented Generation and Fine-Tuning

DGX agent

arXiv:2605.12516v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) often struggle to generate reliable responses in specialized engineering domains due to limited domain gr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Entropy Aware Reward Guidance for Diffusion Language Model Alignment

DGX agent

arXiv:2602.05000v2 Announce Type: replace-cross Abstract: Reward guidance, also known as posterior sampling, is a popular method for test-time adaptation and post-training in continuous diffusion mode

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

TiCo: Time-Controllable Spoken Dialogue Model

DGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

UNIV: Unified Foundation Model for Infrared and Visible Modalities

DGX agent

arXiv:2509.15642v3 Announce Type: replace Abstract: Joint RGB-infrared perception is essential for achieving robustness under diverse weather and illumination conditions. Although foundation models ex

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

DGX agent

arXiv:2605.12178v1 Announce Type: cross Abstract: World models enable agents to anticipate the effects of their actions by internalizing environment dynamics. In enterprise systems, however, these dyn

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

DGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model

DGX agent

arXiv:2605.11255v1 Announce Type: new Abstract: We present Hebatron, a Hebrew-specialized open-weight large language model built on the NVIDIA Nemotron-3 sparse Mixture-of-Experts architecture. Traini

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

DGX agent

arXiv:2605.11061v1 Announce Type: new Abstract: The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In

model-releasesarxiv-cs-cv
13 May 2026
Safety

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer?

DGX agent

arXiv:2605.11301v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have heterogeneous strengths across OCR, chart understanding, spatial reasoning, visual question answering, c

safetyarxiv-cs-cl
13 May 2026
Applications

Mechanistic Interpretability of ASR models using Sparse Autoencoders

DGX agent

arXiv:2605.12225v1 Announce Type: new Abstract: Understanding the internal machinations of deep Transformer-based NLP models is more crucial than ever as these models see widespread use in various dom

applicationsarxiv-cs-cl
13 May 2026
Model Releases

MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models

DGX agent

arXiv:2501.02955v2 Announce Type: replace Abstract: In recent years, vision language models (VLMs) have made significant advancements in video understanding. However, a crucial capability - fine-grain

model-releasesarxiv-cs-cv
13 May 2026
Agents

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

DGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

agentsarxiv-cs-cl
13 May 2026
Research

Pretraining Strategies and Scaling for ECG Foundation Models: A Systematic Study

DGX agent

arXiv:2605.12241v1 Announce Type: cross Abstract: Specialized foundation models are beginning to emerge in various medical subdomains, but pretraining methodologies and parametric scaling with the siz

researcharxiv-cs-lg
13 May 2026
Model Releases

Probabilistic Calibration Is a Trainable Capability in Language Models

DGX agent

arXiv:2605.11845v1 Announce Type: new Abstract: Language models are increasingly used in settings where outputs must satisfy user-specified randomness constraints, yet their generation probabilities a

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

DGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

model-releasesarxiv-cs-cl
13 May 2026
Safety

World Action Models: The Next Frontier in Embodied AI

DGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

safetyarxiv-cs-cl
13 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CLEF: EEG Foundation Model for Learning Clinical Semantics

DGX agent

arXiv:2605.10817v1 Announce Type: new Abstract: Clinical EEG interpretation requires reasoning over full EEG sessions and integrating signal patterns with clinical context. Existing EEG foundation mod

model-releasesarxiv-cs-ai
12 May 2026
Safety

Continuity Laws for Sequential Models

DGX agent

arXiv:2605.08539v1 Announce Type: cross Abstract: Inductive biases influence the behavior and performance of sequential models. In this work, we study an underexplored inductive bias in sequential mod

safetyarxiv-cs-ai
12 May 2026
Research

DeltaRubric: Generative Multimodal Reward Modeling via Joint Planning and Verification

DGX agent

arXiv:2605.09269v1 Announce Type: new Abstract: Aligning Multimodal Large Language Models (MLLMs) requires reliable reward models, yet existing single-step evaluators can suffer from lazy judging, exp

researcharxiv-cs-cl
12 May 2026
Tutorials

Efficient Estimation of Kernel Surrogate Models for Task Attribution

DGX agent

arXiv:2602.03783v2 Announce Type: replace-cross Abstract: Modern AI agents such as large language models are trained on diverse tasks -- translation, code generation, mathematical reasoning, and text

tutorialsarxiv-cs-ai
12 May 2026
Research

Evaluating Pragmatic Reasoning in Large Language Models: Evidence from Scalar Diversity

DGX agent

arXiv:2605.09042v1 Announce Type: new Abstract: Evaluating pragmatic reasoning in large language models (LLMs) remains challenging because model behavior can vary depending on evaluation methods. Prev

researcharxiv-cs-cl
12 May 2026
Model Releases

FormalRewardBench: A Benchmark for Formal Theorem Proving Reward Models

DGX agent

arXiv:2605.10141v1 Announce Type: new Abstract: Recent neural theorem provers use reinforcement learning with verifiable rewards (RLVR), where proof assistants provide binary correctness signals. Whil

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

H-POPE: Hierarchical Polling-based Probing Evaluation of Hallucinations in Large Vision-Language Models

DGX agent

arXiv:2411.04077v2 Announce Type: replace Abstract: By leveraging both texts and images, large vision language models (LVLMs) have shown significant progress in various multi-modal tasks. Nevertheless

model-releasesarxiv-cs-cv
12 May 2026
Safety

LaWM: Least Action World Models for Long-Horizon Physical Consistency from Visual Observations

DGX agent

arXiv:2605.08279v1 Announce Type: cross Abstract: Learning predictive world models from visual observations is a core problem in embodied AI, with applications to model-based reinforcement learning an

safetyarxiv-cs-ai
12 May 2026
Research

Modeling Atomic Conformational Ensembles of Proteins via Test-Time Supervision of Boltz-2 on Cryo-EM Density Maps

DGX agent

arXiv:2605.09832v1 Announce Type: new Abstract: Knowledge of a protein's atomic conformational ensemble is critical to determining its function, yet state-of-the-art ensemble prediction models are lim

researcharxiv-cs-lg
12 May 2026
Model Releases

PhyGround: Benchmarking Physical Reasoning in Generative World Models

DGX agent

arXiv:2605.10806v1 Announce Type: cross Abstract: Generative world models are increasingly used for video generation, where learned simulators are expected to capture the physical rules that govern re

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Position: AI Security Policy Should Target Systems, Not Models

DGX agent

arXiv:2605.09504v1 Announce Type: cross Abstract: We present swarm-attack, an open-source adversarial testing framework in which multiple lightweight LLM agents coordinate through shared memory, paral

model-releasesarxiv-cs-ai
12 May 2026
Research

Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models

DGX agent

arXiv:2605.09630v1 Announce Type: new Abstract: Tokenizer-free language models eliminate the tokenizer step of the language modeling pipeline by operating directly on bytes; patch-based variants furth

researcharxiv-cs-cl
12 May 2026
Model Releases

SMIXAE: Towards Unsupervised Manifold Discovery in Language Models

DGX agent

arXiv:2605.09224v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have been used widely to decompose and interpret neural network activations, especially those of transformer language models.

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild?

DGX agent

arXiv:2602.03916v3 Announce Type: replace-cross Abstract: Spatial reasoning is a fundamental aspect of human cognition, yet it remains a major challenge for contemporary vision-language models (VLMs).

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

TabPFN-3 just released: a pre-trained tabular foundation model for up to 1M rows [R][N]

DGX agent

TabPFN-3 is a pre-trained tabular foundation model that supports datasets up to 1,000,000 rows × 200 features , representing a significant scaling improvement for the TabPFN family. The model delivers

model-releasesr-machinelearning
12 May 2026
Model Releases

TIDES: Implicit Time-Awareness in Selective State Space Models

DGX agent

arXiv:2605.09742v1 Announce Type: cross Abstract: Selective state space models (SSMs), such as Mamba, achieve strong per-token expressivity by making the time discretization step Tilde{Delta} a learne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Towards Generative Predictive Display for Vision-Based Teleoperation: A Zero-Shot Benchmark of Off-the-Shelf Video Models

DGX agent

arXiv:2605.09670v1 Announce Type: cross Abstract: Teleoperation systems are fundamentally limited by communication latency, which degrades situational awareness and control performance. Predictive dis

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Transformation-Augmented GRPO for Enhancing Exploration in Reasoning of Large Language Models

DGX agent

arXiv:2601.22478v3 Announce Type: replace Abstract: Group Relative Policy Optimization (GRPO) has become the dominant method for reinforcement learning with verifiable rewards in large language models

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

VC-Soup: Value-Consistency Guided Multi-Value Alignment for Large Language Models

DGX agent

arXiv:2603.18113v2 Announce Type: replace-cross Abstract: As large language models (LLMs) increasingly shape content generation, interaction, and decision-making across the Web, aligning them with hum

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models

DGX agent

arXiv:2605.10485v1 Announce Type: new Abstract: Precise spatial reasoning is fundamental to robotic manipulation, yet the visual backbones of current vision-language-action (VLA) models are predominan

model-releasesarxiv-cs-ro
12 May 2026
Research

Weighted Rules under the Stable Model Semantics

DGX agent

arXiv:2605.09519v1 Announce Type: new Abstract: We introduce the concept of weighted rules under the stable model semantics following the log-linear models of Markov Logic. This provides versatile met

researcharxiv-cs-ai
12 May 2026
Research

XPERT: Expert Knowledge Transfer for Effective Training of Language Models

DGX agent

arXiv:2605.08842v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) language models organize knowledge into explicitly routed expert modules, making expert-level representations traceable and ana

researcharxiv-cs-cl
12 May 2026
Model Releases

Continually Evolving Skill Knowledge in Vision Language Action Model

DGX agent

arXiv:2511.18085v4 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models show promising knowledge accumulation ability from pretraining, yet continual learning in VLA remains chal

model-releasesarxiv-cs-ai
11 May 2026
Research

Coupling Models for One-Step Discrete Generation

DGX agent

arXiv:2605.07193v1 Announce Type: new Abstract: Generative modeling over discrete structures underpins applications across deep learning, from biological sequence design and code generation to large l

researcharxiv-cs-lg
11 May 2026
Model Releases

From Pixels to Prompts: Vision-Language Models

DGX agent

arXiv:2605.07544v1 Announce Type: new Abstract: When you read a paper about a new Vision-Language Model today, it can be easy to forget how strange this idea would have sounded not so long ago. Teachi

model-releasesarxiv-cs-ai
11 May 2026
Research

Generative Modeling with Flux Matching

DGX agent

arXiv:2605.07319v1 Announce Type: cross Abstract: We introduce Flux Matching, a new paradigm for generative modeling that generalizes existing score-based models to a broader family of vector fields t

researcharxiv-cs-ai
11 May 2026
Agents

Give our early preview of Computer Use (with ANY model) a try today! Built into the latest Hermes Agent and powered by @trycua - opens the d…

DGX agent

Give our early preview of Computer Use (with ANY model) a try today! Built into the latest Hermes Agent and powered by @trycua - opens the door to any model, not just the frontier models in special mo

agentsnous-research--x
11 May 2026
Research

Inference of Qualitative Models from Steady-State Data via Weighted MaxSMT

DGX agent

arXiv:2605.07433v1 Announce Type: cross Abstract: Qualitative models provide crucial instruments for modelling complex biological systems. While advances in automated reasoning and symbolic encodings

researcharxiv-cs-lg
11 May 2026
← Previous
1…6768697071…1259
Next →