AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,975 results
Local Ai

Locally Coherent Parallel Decoding in Diffusion Language Models

DGX agent

arXiv:2603.20216v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have emerged as a promising alternative to autoregressive (AR) models, offering sub-linear generation latency

local-aiarxiv-cs-ai
19 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

DGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Membership Inference Attacks on Discrete Diffusion Language Models

DGX agent

arXiv:2605.16445v1 Announce Type: cross Abstract: Masked Diffusion Language Models MDLMs replace autoregressive generation with iterative demasking and their privacy properties are largely unstudied.

model-releasesarxiv-cs-ai
19 May 2026
Research

Preference Instability in Reward Models: Detection and Mitigation via Sparse Autoencoders

DGX agent

arXiv:2605.16339v1 Announce Type: new Abstract: Preference learning in large language models relies on reward models as proxies for human judgment. However, these models frequently exhibit preference

researcharxiv-cs-lg
19 May 2026
Model Releases

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

DGX agent

arXiv:2605.18380v1 Announce Type: new Abstract: We introduce an extensive qualitative spatial and temporal reasoning (QSTR) benchmark for evaluating large language models (LLMs). We pose questions con

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TabH2O: A Unified Foundation Model for Tabular Prediction

DGX agent

arXiv:2605.18383v1 Announce Type: new Abstract: We present TabH2O, a foundation model for tabular data that performs classification and regression in a single forward pass via in-context learning. Tab

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

The Illusion of Specialization: Unveiling the Domain-Invariant 'Standing Committee' in Mixture-of-Experts Models

DGX agent

arXiv:2601.03425v2 Announce Type: replace-cross Abstract: Mixture of Experts models are widely assumed to achieve domain specialization through sparse routing. In this work, we question this assumptio

model-releasesarxiv-cs-ai
19 May 2026
Safety

UniAlign: A Model-Agnostic Framework for Robust Network Traffic Classification under Distribution Shifts

DGX agent

arXiv:2605.17575v1 Announce Type: cross Abstract: Network traffic classification (NTC) models often suffer severe performance degradation when deployed in real-world environments due to distribution s

safetyarxiv-cs-ai
19 May 2026
Model Releases

What Does the AI Doctor Value? Auditing Pluralism in the Clinical Ethics of Language Models

DGX agent

arXiv:2605.18738v1 Announce Type: new Abstract: Medicine is inherently pluralistic. Principles such as autonomy, beneficence, nonmaleficence, and justice routinely conflict, and such ethical dilemmas

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Who Generated This 3D Asset? Learning Source Attribution for Generative 3D Models

DGX agent

arXiv:2605.18132v1 Announce Type: cross Abstract: Generative 3D models are deployed in gaming, robotics, and immersive creation, making source attribution critical: given a 3D asset, can we identify w

model-releasesarxiv-cs-ai
19 May 2026
Research

DiLA: Disentangled Latent Action World Models

DGX agent

arXiv:2605.15725v1 Announce Type: cross Abstract: Latent Action Models (LAMs) enable the learning of world models from unlabeled video by inferring abstract actions between consecutive frames. However

researcharxiv-cs-ai
18 May 2026
Research

Do Chinese models speak Chinese languages?

DGX agent

arXiv:2504.00289v3 Announce Type: replace-cross Abstract: The release of top-performing open-weight LLMs has cemented China's role as a leading force in AI development. Do these models support languag

researcharxiv-cs-ai
18 May 2026
Safety

Imperfect World Models are Exploitable

DGX agent

arXiv:2605.15960v1 Announce Type: new Abstract: We propose a novel definition of model exploitation in reinforcement learning. Informally, a world model is exploitable if it implies that one policy sh

safetyarxiv-cs-ai
18 May 2026
Model Releases

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

DGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

DGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkyLink: A Large Vision-Language Model Driven Re-ranking Framework for Cross-View UAV geolocalization

DGX agent

arXiv:2603.08063v3 Announce Type: replace Abstract: Cross-view UAV geolocalization is fundamentally a challenging large-scale image retrieval task, aiming to determine the geographic coordinates of Un

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

DGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Tadpole: Autoencoders as Foundation Models for 3D PDEs with Online Learning

DGX agent

arXiv:2605.15284v1 Announce Type: new Abstract: We introduce Tadpole, a novel foundation model for three-dimensional partial differential equations (PDEs) that addresses key challenges in transferabil

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Breaking Dual Bottlenecks: Evolving Unified Multimodal Models into Self-Adaptive Interleaved Visual Reasoners

DGX agent

arXiv:2605.14709v1 Announce Type: new Abstract: Recent unified models integrate multimodal understanding and generation within a single framework. However, an 'understanding-generation gap' persists,

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

DGX agent

arXiv:2509.23023v3 Announce Type: replace Abstract: Large language models are increasingly deployed in multi-agent settings whose outcomes hinge on social intelligence, motivating evaluations of their

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

GFMate: Empowering Graph Foundation Models with Test-time Prompt Tuning

DGX agent

arXiv:2605.14809v1 Announce Type: new Abstract: Graph prompt tuning has shown great potential in graph learning by introducing trainable prompts to enhance the model performance in conventional single

model-releasesarxiv-cs-lg
15 May 2026
Research

HERO: Hierarchical Extrapolation and Refresh for Efficient World Models

DGX agent

arXiv:2508.17588v2 Announce Type: replace Abstract: Generation-driven world models create immersive virtual environments but suffer slow inference due to the iterative nature of diffusion models. Whil

researcharxiv-cs-cv
15 May 2026
Safety

Proxy Compression for Language Modeling

DGX agent

arXiv:2602.04289v2 Announce Type: replace Abstract: Modern language models are trained almost exclusively on token sequences produced by a fixed tokenizer, an external lossless compressor often over U

safetyarxiv-cs-cl
15 May 2026
Model Releases

REALM: Retrospective Encoder Alignment for LFP Modeling

DGX agent

arXiv:2605.14867v1 Announce Type: cross Abstract: Spike activity has been the dominant neural signal for behavior decoding due to its high spatial and temporal resolution. However, as brain-computer i

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

SANA-WM: Efficient Minute-Scale World Modeling with Hybrid Linear Diffusion Transformer

DGX agent

arXiv:2605.15178v1 Announce Type: new Abstract: We introduce SANA-WM, an efficient 2.6B-parameter open-source world model natively trained for one-minute generation, synthesizing high-fidelity, 720p,

model-releasesarxiv-cs-cv
15 May 2026
Hardware

Synthetic Sociality: How Generative Models Privatize the Social Fabric

DGX agent

arXiv:2605.14090v1 Announce Type: cross Abstract: We put forth a critical theoretical framework for analyzing generative models both descriptively and normatively. Our thesis is that generative models

hardwarearxiv-cs-lg
15 May 2026
Model Releases

Unsupervised learning of acquisition variability in structural connectomes via hybrid latent space modeling

DGX agent

arXiv:2605.13933v1 Announce Type: cross Abstract: Acquisition differences across sites, scanners, and protocols in dMRI introduce variability that complicates structural connectome analysis. This moti

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Benchmarking Attribute Discrimination in Infant-Scale Vision-Language Models

DGX agent

arXiv:2512.18951v3 Announce Type: replace Abstract: Infants learn not only object categories but also fine-grained visual attributes such as color, size, and texture from limited experience. Prior inf

model-releasesarxiv-cs-lg
14 May 2026
Safety

Characterizing Universal Object Representations Across Vision Models

DGX agent

arXiv:2605.13675v1 Announce Type: new Abstract: Deep neural networks trained with different architectures, objectives, and datasets have been reported to converge on similar visual representations. Ho

safetyarxiv-cs-cv
14 May 2026
Model Releases

DAWM: Diffusion Action World Models for Offline Reinforcement Learning via Action-Inferred Transitions

DGX agent

arXiv:2509.19538v2 Announce Type: replace-cross Abstract: Diffusion-based world models have demonstrated strong capabilities in synthesizing realistic long-horizon trajectories for offline reinforceme

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Domain Adaptation of Large Language Models for Polymer-Composite Additive Manufacturing Using Retrieval-Augmented Generation and Fine-Tuning

DGX agent

arXiv:2605.12516v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) often struggle to generate reliable responses in specialized engineering domains due to limited domain gr

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Entropy Aware Reward Guidance for Diffusion Language Model Alignment

DGX agent

arXiv:2602.05000v2 Announce Type: replace-cross Abstract: Reward guidance, also known as posterior sampling, is a popular method for test-time adaptation and post-training in continuous diffusion mode

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

TiCo: Time-Controllable Spoken Dialogue Model

DGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

UNIV: Unified Foundation Model for Infrared and Visible Modalities

DGX agent

arXiv:2509.15642v3 Announce Type: replace Abstract: Joint RGB-infrared perception is essential for achieving robustness under diverse weather and illumination conditions. Although foundation models ex

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

DGX agent

arXiv:2605.12178v1 Announce Type: cross Abstract: World models enable agents to anticipate the effects of their actions by internalizing environment dynamics. In enterprise systems, however, these dyn

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

EsoLang-Bench: Evaluating Genuine Reasoning in Large Language Models via Esoteric Programming Languages

DGX agent

arXiv:2603.09678v2 Announce Type: replace-cross Abstract: Large language models achieve near-ceiling performance on code generation benchmarks, yet most of the programming languages used by popular be

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

HEBATRON: A Hebrew-Specialized Open-Weight Mixture-of-Experts Language Model

DGX agent

arXiv:2605.11255v1 Announce Type: new Abstract: We present Hebatron, a Hebrew-specialized open-weight large language model built on the NVIDIA Nemotron-3 sparse Mixture-of-Experts architecture. Traini

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

DGX agent

arXiv:2605.11061v1 Announce Type: new Abstract: The evolution of visual generative models has long been constrained by fragmented architectures relying on disjoint text encoders and external VAEs. In

model-releasesarxiv-cs-cv
13 May 2026
Safety

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer?

DGX agent

arXiv:2605.11301v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have heterogeneous strengths across OCR, chart understanding, spatial reasoning, visual question answering, c

safetyarxiv-cs-cl
13 May 2026
Applications

Mechanistic Interpretability of ASR models using Sparse Autoencoders

DGX agent

arXiv:2605.12225v1 Announce Type: new Abstract: Understanding the internal machinations of deep Transformer-based NLP models is more crucial than ever as these models see widespread use in various dom

applicationsarxiv-cs-cl
13 May 2026
Model Releases

MotionBench: Benchmarking and Improving Fine-grained Video Motion Understanding for Vision Language Models

DGX agent

arXiv:2501.02955v2 Announce Type: replace Abstract: In recent years, vision language models (VLMs) have made significant advancements in video understanding. However, a crucial capability - fine-grain

model-releasesarxiv-cs-cv
13 May 2026
Agents

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

DGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

agentsarxiv-cs-cl
13 May 2026
Research

Pretraining Strategies and Scaling for ECG Foundation Models: A Systematic Study

DGX agent

arXiv:2605.12241v1 Announce Type: cross Abstract: Specialized foundation models are beginning to emerge in various medical subdomains, but pretraining methodologies and parametric scaling with the siz

researcharxiv-cs-lg
13 May 2026
Model Releases

Probabilistic Calibration Is a Trainable Capability in Language Models

DGX agent

arXiv:2605.11845v1 Announce Type: new Abstract: Language models are increasingly used in settings where outputs must satisfy user-specified randomness constraints, yet their generation probabilities a

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

StoicLLM: Preference Optimization for Philosophical Alignment in Small Language Models

DGX agent

arXiv:2605.11483v1 Announce Type: new Abstract: While large language models excel at factual adaptation, their ability to internalize nuanced philosophical frameworks under severe data constraints rem

model-releasesarxiv-cs-cl
13 May 2026
Safety

World Action Models: The Next Frontier in Embodied AI

DGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

safetyarxiv-cs-cl
13 May 2026
Model Releases

A Geometric Perspective on Next-Token Prediction in Large Language Models: Three Emerging Phases

DGX agent

arXiv:2605.09011v1 Announce Type: cross Abstract: We investigate the geometry of predictive information across the layers of large language models (LLMs). We repurpose representation lenses-learned af

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

CLEF: EEG Foundation Model for Learning Clinical Semantics

DGX agent

arXiv:2605.10817v1 Announce Type: new Abstract: Clinical EEG interpretation requires reasoning over full EEG sessions and integrating signal patterns with clinical context. Existing EEG foundation mod

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…5051525354…1021
Next →