AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Learning Long Range Spatio-Temporal Representations over Continuous Time Dynamic Graphs with State Space Models

DGX agent

arXiv:2606.04672v1 Announce Type: cross Abstract: Continuous-time dynamic graphs (CTDGs) provide a richer framework to capture fine-grained temporal patterns in evolving relational data. Long-range in

model-releasesarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

DGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Research

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification

DGX agent

arXiv:2606.04579v1 Announce Type: new Abstract: While Process Reward Models (PRMs) have achieved remarkable success in mathematical reasoning, their application in complex scientific domains-such as b

researcharxiv-cs-ai
4 Jun 2026
Safety

Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

DGX agent

arXiv:2602.19101v2 Announce Type: replace-cross Abstract: Value alignment of Large Language Models (LLMs) requires us to empirically measure these models' actual, acquired representation of value. Amo

safetyarxiv-cs-ai
4 Jun 2026
Applications

AugMask: Training Diffusion Models on Incomplete Tabular Data via Stochastic Augmentation and Masking

DGX agent

arXiv:2606.03347v1 Announce Type: cross Abstract: Score-based diffusion models have emerged as prominent deep generative models; however, their application to tabular data remains challenging because

applicationsarxiv-cs-ai
3 Jun 2026
Research

Bridging Auxiliary Constraints to Resolve Instruction Following in Large Reasoning Models

DGX agent

arXiv:2606.03624v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated impressive capabilities in many tasks, yet they struggle with reliably following multiple instructions,

researcharxiv-cs-ai
3 Jun 2026
Model Releases

ClinicalMC: A Benchmark for Multi-Course Clinical Decision-Making with Large Language Models

DGX agent

arXiv:2606.03157v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted in healthcare, yet they still encounter significant challenges in complex clinical decision-making

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction

DGX agent

arXiv:2606.03874v1 Announce Type: new Abstract: We present DyaPlex, a streaming, full-duplex speech-and-motion model designed for dyadic interaction. To capture the continuous and reciprocal nature of

agentsarxiv-cs-cv
3 Jun 2026
Local Ai

Expert-Aware Causal Tracing of Factual Recall in Sparse MoE Language Models

DGX agent

arXiv:2606.03780v1 Announce Type: new Abstract: Causal tracing of factual recall has been studied predominantly in dense transformer language models, where interventions localize information flow to l

local-aiarxiv-cs-cl
3 Jun 2026
Model Releases

Instant Personalized Large Language Model Adaptation via Hypernetwork

DGX agent

arXiv:2510.16282v2 Announce Type: replace Abstract: Personalized large language models (LLMs) tailor content to individual preferences using user profiles or histories. However, existing parameter-eff

model-releasesarxiv-cs-cl
3 Jun 2026
Safety

Large Language Models Are Overconfident in Their Own Responses

DGX agent

arXiv:2606.03437v1 Announce Type: new Abstract: Prior work has shown that instruction-tuned large language models (LLMs) are less well calibrated than their base pre-trained counterparts. However, lit

safetyarxiv-cs-cl
3 Jun 2026
Safety

Measuring Weak-to-Strong Legibility of Reasoning Models

DGX agent

arXiv:2603.20508v2 Announce Type: replace-cross Abstract: Reasoning language models (RLMs) and the intermediate chains of thought they emit play an increasingly central role in multi-agent setups such

safetyarxiv-cs-ai
3 Jun 2026
Research

Non-Identical Diffusion Models in MIMO-OFDM Channel Generation

DGX agent

arXiv:2509.01641v3 Announce Type: replace-cross Abstract: We propose a novel diffusion model, termed the non-identical diffusion model, and investigate its application to wireless orthogonal frequency

researcharxiv-cs-ai
3 Jun 2026
Safety

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

DGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Models

DGX agent

arXiv:2601.12247v3 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) present a promising non-sequential paradigm for text generation, distinct from standard autoregressive (AR) a

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Position: Prioritize Identifying Structure, Not Complex Models, for Scientific Discovery

DGX agent

arXiv:2606.02632v1 Announce Type: cross Abstract: Modern Machine Learning (ML) and Artificial Intelligence (AI) models, especially large language models (LLMs), are increasingly used to generate scien

researcharxiv-cs-ai
3 Jun 2026
Model Releases

PyraMathBench: Evaluating and Improving Mathematical Capability in Large Language Models

DGX agent

arXiv:2606.03858v1 Announce Type: new Abstract: Despite the pivotal role of numerical reasoning as the cornerstone of mathematical capabilities in large language models (LLMs) across applications, few

model-releasesarxiv-cs-ai
3 Jun 2026
Tutorials

Text-to-Image Models Need Less from Text Encoders Than You Think

DGX agent

arXiv:2606.03715v1 Announce Type: new Abstract: Text-to-image models rely on text prompts as their primary interface to human intent. Prompts are encoded by a text encoder into embeddings that conditi

tutorialsarxiv-cs-cv
3 Jun 2026
Research

Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models

DGX agent

arXiv:2606.02835v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by generating explicit intermediate reasoning traces through increased test-time compute, yet the assu

researcharxiv-cs-ai
3 Jun 2026
Research

TimeOmni-VL: Unified Models for Time Series Understanding and Generation

DGX agent

arXiv:2602.17149v2 Announce Type: replace-cross Abstract: Recent time series modeling faces a sharp divide between numerical generation and semantic understanding, with research showing that generatio

researcharxiv-cs-ai
3 Jun 2026
Tutorials

Visual Graph Scaffolds for Structural Reasoning in Large Language Models

DGX agent

arXiv:2606.02673v1 Announce Type: new Abstract: Graphs have been used to enhance large language models (LLMs) for structured reasoning, mostly as external knowledge sources are provided to models at t

tutorialsarxiv-cs-ai
3 Jun 2026
Safety

When Graph Tokens Sink: A Mechanistic Analysis of Graph Language Models

DGX agent

arXiv:2606.03712v1 Announce Type: new Abstract: Graph Language Models (GLMs) have become a promising direction for adapting Large Language Models (LLMs) to graph learning tasks. By transforming graph

safetyarxiv-cs-lg
3 Jun 2026
Model Releases

When Model Merging Breaks Routing: Training-Free Calibration for MoE

DGX agent

arXiv:2606.03391v1 Announce Type: cross Abstract: Model merging has emerged as a cost-effective approach for consolidating the capabilities of multiple LLMs without retraining. However, existing mergi

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

3D Segment Anything Model with Visual Mamba for Diagnosing Placenta Accreta Spectrum

DGX agent

arXiv:2606.00489v1 Announce Type: new Abstract: Placenta Accreta Spectrum (PAS) is a rare but highly dangerous obstetric disease. Early and accurate PAS diagnosis is critical for maternal health. Trad

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

AgentPLM: Agentic Protein Language Models with Reasoning-Augmented Decoding for Protein Sequence Design

DGX agent

arXiv:2606.02386v1 Announce Type: new Abstract: Protein language models (PLMs) are passive oracles: they generate sequences in a single forward pass with no mechanism to consult external biophysical f

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

Are Large Reasoning Models Interruptible?

DGX agent

arXiv:2510.11713v4 Announce Type: replace Abstract: Real-world applications of Large Reasoning Models (LRMs) often require reasoning about changing prompts or environments. In this work, we challenge

applicationsarxiv-cs-cl
2 Jun 2026
Safety

au_0-WM: A Unified Video-Action World Model for Robotic Manipulation

DGX agent

arXiv:2606.01027v1 Announce Type: new Abstract: Robotic manipulation requires models that generate executable actions while anticipating and evaluating their future consequences before physical execut

safetyarxiv-cs-ro
2 Jun 2026
Model Releases

AutoEval Done Right: Using Synthetic Data for Model Evaluation

DGX agent

arXiv:2403.07008v3 Announce Type: replace-cross Abstract: The evaluation of machine learning models using human-labeled validation data can be expensive and time-consuming. AI-labeled synthetic data c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Benchmarking Large Language Models for Cryptanalysis and Side-Channel Vulnerabilities

DGX agent

arXiv:2505.24621v3 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have transformed natural language understanding and generation, leading to extensive benchmarkin

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

Beyond Categories of Caste: Examining Caste Bias and Morality in Text-to-Image AI Models

DGX agent

arXiv:2606.00039v1 Announce Type: cross Abstract: Text-to-Image (T2I) models have shown promising utility across various domains. However, such models are also amplifying harmful societal biases in th

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Beyond Text and Tables: Vision-Language Model Integration in ComProScanner for Extracting Materials Data from Scientific Figures with High Accuracy

DGX agent

arXiv:2606.00065v1 Announce Type: cross Abstract: Automated extraction of materials composition-property data from scientific literature has advanced considerably with the development of large languag

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Consistency Deep Equilibrium Models

DGX agent

arXiv:2602.03024v2 Announce Type: replace-cross Abstract: Deep Equilibrium Models (DEQs) have emerged as a powerful paradigm in deep learning, offering the ability to model infinite-depth networks wit

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Controllable Value Alignment in Large Language Models through Neuron-Level Editing

DGX agent

arXiv:2602.07356v2 Announce Type: replace Abstract: Aligning large language models (LLMs) with human values has become increasingly important as their influence on human behavior and decision-making e

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

DASH: Dual-Branch Score Distillation for Guidance-Calibrated Compact Diffusion Models

DGX agent

arXiv:2606.00798v1 Announce Type: cross Abstract: Parameter compression of class-conditional diffusion models reveals an underexplored limitation in output-level distillation: the unconditional score

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Diffusion Image Generation with Explicit Modeling of Data Manifold Geometry

DGX agent

arXiv:2606.00094v1 Announce Type: cross Abstract: Image generative models aim to sample data points from the underlying data manifold, a task that requires learning and decoding a dense, low-dimension

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Distillation of Large Language Models via Concrete Score Matching

DGX agent

arXiv:2509.25837v3 Announce Type: replace-cross Abstract: Large language models (LLMs) deliver remarkable performance but are costly to deploy, motivating knowledge distillation (KD) for efficient inf

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

DLLM-JEPA: Joint Embedding Predictive Architectures for Masked Diffusion Language Models

DGX agent

arXiv:2606.00091v1 Announce Type: cross Abstract: Joint Embedding Predictive Architectures (JEPAs) have reshaped self-supervised representation learning in vision. The recent LLM-JEPA ported JEPA to a

model-releasesarxiv-cs-ai
2 Jun 2026
Research

ELF: A Family of Encoder-Free ECG-Language Models

DGX agent

arXiv:2601.18798v2 Announce Type: replace-cross Abstract: ECG-Language Models (ELMs) extend recent advances in Multimodal Large Language Models (MLLMs) to automated ECG interpretation. However, most e

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Empathy Applicability Modeling for General Health Queries

DGX agent

arXiv:2601.09696v2 Announce Type: replace Abstract: LLMs are increasingly being integrated into clinical workflows, yet they often lack clinical empathy, an essential aspect of effective doctor-patien

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Escaping the Mode Lottery: Multi-Response Training Improves Language Model Generalization

DGX agent

arXiv:2606.00544v1 Announce Type: cross Abstract: Modern language-model fine-tuning typically pairs each prompt with a single response, even though many prompts admit multiple valid completions. This

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

From 'Weak' Signals to Strong Models: Preference Delta Aggregation with LoRA Merging

DGX agent

arXiv:2606.00357v1 Announce Type: new Abstract: Training strong large language models (LLMs) requires high-quality supervision, which is often scarce. Recent work shows that paired preference data fro

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

GLENS: Global Search via Learning from Solver Iterates with Diffusion Models

DGX agent

arXiv:2606.00366v1 Announce Type: new Abstract: We consider the problem of generating a large collection of initial guesses for local minima of multimodal non-convex continuous optimization problems.

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

GNMR: Runtime Stability Control for Low-Precision Large Language Model Training

DGX agent

arXiv:2606.00539v1 Announce Type: new Abstract: Training stability is a key bottleneck in low-precision language model training: efficient low-cost paths can still produce short-lived numerical risks

model-releasesarxiv-cs-lg
2 Jun 2026
Research

IDLM: Inverse-distilled Diffusion Language Models

DGX agent

arXiv:2602.19066v2 Announce Type: replace-cross Abstract: Diffusion Language Models (DLMs) have recently achieved strong results in text generation. However, their multi-step sampling leads to slow in

researcharxiv-cs-ai
2 Jun 2026
Model Releases

InPhyRe Discovers: Large Multimodal Models Struggle in Inductive Physical Reasoning

DGX agent

arXiv:2509.12263v3 Announce Type: replace Abstract: Large multimodal models (LMMs) encode physical laws observed during training, such as momentum conservation, as parametric knowledge. It allows LMMs

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Interpretability in Deep Time Series Models Demands Semantic Alignment

DGX agent

arXiv:2602.02239v2 Announce Type: replace Abstract: Deep time series models continue to improve predictive performance, yet their deployment remains limited by their black-box nature. In response, exi

safetyarxiv-cs-lg
2 Jun 2026
Model Releases

Measuring and Mitigating Bias in Code Generated by Large Language Models

DGX agent

arXiv:2606.00049v1 Announce Type: cross Abstract: Large language models (LLMs) are widely recognised for their applications in natural language generation and are increasingly used for code generation

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…106107108109110…1030
Next →