AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Tutorials

Emergent Semantic Role Understanding in Language Models

DGX agent

arXiv:2605.09187v1 Announce Type: new Abstract: Understanding how linguistic structure emerges in language models is central to interpreting what these systems learn from data and how much supervision

tutorialsarxiv-cs-ai
12 May 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Evidence-based Decision Modeling for Synthetic Face Detection with Uncertainty-driven Active Learning

DGX agent

arXiv:2605.09935v1 Announce Type: new Abstract: With the rapid development of deep generative models, forged facial images are massively exploited for illegal activities. Although existing synthetic f

researcharxiv-cs-cv
12 May 2026
Model Releases

Explanation Fairness in Large Language Models: An Empirical Analysis of Disparities in How LLMs Justify Decisions Across Demographic Groups

DGX agent

arXiv:2605.08671v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed not only to make decisions but to explain them. While AI decision fairness has been studied ext

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Federated Language Models Under Bandwidth Budgets: Distillation Rates and Conformal Coverage

DGX agent

arXiv:2605.09986v1 Announce Type: cross Abstract: Training a language model on data scattered across bandwidth-limited nodes that cannot be centralized is a setting that arises in clinical networks, e

model-releasesarxiv-cs-cl
12 May 2026
Tutorials

IntroLM: Introspective Language Models via Prefilling-Time Self-Evaluation

DGX agent

arXiv:2601.03511v2 Announce Type: replace-cross Abstract: A major challenge for the operation of large language models (LLMs) is how to predict whether a specific LLM will produce sufficiently high-qu

tutorialsarxiv-cs-ai
12 May 2026
Research

Kernel-Gradient Drifting Models

DGX agent

arXiv:2605.10727v1 Announce Type: new Abstract: We propose kernel-gradient drifting, a one-step generative modeling framework that replaces the fixed Euclidean displacement direction in drifting model

researcharxiv-cs-lg
12 May 2026
Model Releases

Latent Geometry Beyond Search: Amortizing Planning in World Models

DGX agent

arXiv:2605.08732v1 Announce Type: cross Abstract: Modern vision-based world models can represent observations as compact yet expressive latent manifolds, but fast goal-oriented planning in these space

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

Lost in Volume: The CT-SpatialVQA Benchmark for Evaluating Semantic-Spatial Understanding of 3D Medical Vision-Language Models

DGX agent

arXiv:2605.08787v1 Announce Type: new Abstract: Recent advances in 3D medical vision-language models have enabled joint reasoning over volumetric images and text, showing strong performance in medical

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

MicroWorld: Empowering Multimodal Large Language Models to Bridge the Microscopic Domain Gap with Multimodal Attribute Graph

DGX agent

arXiv:2605.10120v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) show remarkable potential for scientific reasoning, yet their performance in specialized domains such as micr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models

DGX agent

arXiv:2510.09592v2 Announce Type: replace Abstract: Real-time Spoken Language Models (SLMs) struggle to leverage Chain-of-Thought (CoT) reasoning due to the prohibitive latency of generating the entir

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

MolRGen: A Training and Evaluation Setting for De Novo Molecular Generation with Reasonning Models

DGX agent

arXiv:2603.18256v2 Announce Type: replace-cross Abstract: Recent reasoning-based large language models have shown strong performance on tasks with verifiable outcomes, but their use in de novo molecul

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Personal Visual Context Learning in Large Multimodal Models

DGX agent

arXiv:2605.10936v1 Announce Type: new Abstract: As wearable devices like smart glasses integrate Large Multimodal Models (LMMs) into the continuous first-person visual streams of individual users, the

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Product-of-Gaussian-Mixture Diffusion Models for Joint Nonlinear MRI Reconstruction

DGX agent

arXiv:2605.10629v1 Announce Type: new Abstract: Recently, diffusion models have attracted considerable attention for magnetic resonance image reconstruction due to their high sample quality. However,

model-releasesarxiv-cs-cv
12 May 2026
Safety

PromptGuard: Soft Prompt-Guided Unsafe Content Moderation for Text-to-Image Models

DGX agent

arXiv:2501.03544v5 Announce Type: replace-cross Abstract: Recent text-to-image (T2I) models have exhibited remarkable performance in generating high-quality images from text descriptions. However, the

safetyarxiv-cs-ai
12 May 2026
Model Releases

Reasoning emerges from constrained inference manifolds in large language models

DGX agent

arXiv:2605.08142v1 Announce Type: cross Abstract: Reasoning in large language models is predominantly evaluated through labeled benchmarks, conflating task performance with the quality of internal inf

model-releasesarxiv-cs-cl
12 May 2026
Research

Reinforce Adjoint Matching: Scaling RL Post-Training of Diffusion and Flow-Matching Models

DGX agent

arXiv:2605.10759v1 Announce Type: cross Abstract: Diffusion and flow-matching models scale because pretraining is supervised regression: a clean sample is noised analytically, and a model regresses ag

researcharxiv-cs-cv
12 May 2026
Safety

Relational reasoning and inductive bias in transformers and large language models

DGX agent

arXiv:2506.04289v3 Announce Type: replace Abstract: Transformer-based models have demonstrated remarkable reasoning abilities, but the mechanisms underlying relational reasoning remain poorly understo

safetyarxiv-cs-lg
12 May 2026
Model Releases

Relative Kinetic Utility for Reasoning-Aware Structural Pruning in Large Language Models

DGX agent

arXiv:2605.09008v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) prompting symbolized a huge improvement of reasoning capabilities of Large Language Models (LLMs). However, scaling up test-tim

model-releasesarxiv-cs-cl
12 May 2026
Applications

Sundial: A Family of Highly Capable Time Series Foundation Models

DGX agent

arXiv:2502.00816v4 Announce Type: replace Abstract: We introduce Sundial, a family of native, flexible, and scalable time series foundation models. To predict the next-patch's distribution, we propose

applicationsarxiv-cs-lg
12 May 2026
Model Releases

TFM-Retouche: A Lightweight Input-Space Adapter for Tabular Foundation Models

DGX agent

arXiv:2605.06047v2 Announce Type: replace-cross Abstract: Tabular foundation models (TFMs), such as TabPFN-2.6, TabICLv2, ConTextTab, Mitra, LimiX, and TabDPT, achieve strong zero-shot performance thr

model-releasesarxiv-cs-ai
12 May 2026
Safety

The Safety-Aware Denoiser for Text Diffusion Models

DGX agent

arXiv:2605.08116v1 Announce Type: cross Abstract: Recent work on text diffusion models offers a promising alternative to autoregressive generation, but controlling their safety remains underexplored.

safetyarxiv-cs-ai
12 May 2026
Tutorials

The two clocks and the innovation window: When and how generative models learn rules

DGX agent

arXiv:2605.10019v1 Announce Type: cross Abstract: Generative models trained on finite data face a fundamental tension: their score-matching or next-token objective converges to the empirical training

tutorialsarxiv-cs-ai
12 May 2026
Model Releases

VLADriver-RAG: Retrieval-Augmented Vision-Language-Action Models for Autonomous Driving

DGX agent

arXiv:2605.08133v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for end-to-end autonomous driving, yet their reliance on implicit parametric

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

When to Trust Imagination: Adaptive Action Execution for World Action Models

DGX agent

arXiv:2605.06222v2 Announce Type: replace-cross Abstract: World Action Models (WAMs) have recently emerged as a promising paradigm for robotic manipulation by jointly predicting future visual observat

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

A Reproducible Optimisation Protocol for Calibrating Prompt-Based Large Language Model Workflows in Evidence Synthesis

DGX agent

arXiv:2605.06937v1 Announce Type: new Abstract: This methods article presents a reproducible calibration workflow for prompt-based large language models (LLMs) in structured evidence-synthesis tasks.

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Beyond Retrieval: A Multitask Benchmark and Model for Code Search

DGX agent

arXiv:2605.04615v2 Announce Type: replace-cross Abstract: Code search has usually been evaluated as first-stage retrieval, even though production systems rely on broader pipelines with reranking and d

model-releasesarxiv-cs-ai
11 May 2026
Applications

Causal-Aware Foundation-Model for Bilevel Optimization in Discrete Choice Settings

DGX agent

arXiv:2605.06941v1 Announce Type: new Abstract: We introduce a causal aware foundation-model framework for real time optimal decision making in discrete choice environments. We propose a constrained t

applicationsarxiv-cs-lg
11 May 2026
Model Releases

MIND: Monge Inception Distance for Generative Models Evaluation

DGX agent

arXiv:2605.06797v1 Announce Type: new Abstract: We propose the Monge Inception Distance (MIND), a metric for evaluating generative models that addresses key limitations of the widely adopted Frechet I

model-releasesarxiv-cs-lg
11 May 2026
Local Ai

On the Tradeoffs of On-Device Generative Models in Federated Predictive Maintenance Systems

DGX agent

arXiv:2605.07860v1 Announce Type: cross Abstract: Federated Learning (FL) has emerged as a promising paradigm for preserving client data ownership and control over distributed Internet of Things (IoT)

local-aiarxiv-cs-ai
11 May 2026
Local Ai

Predictive but Not Plannable: RC-aux for Latent World Models

DGX agent

arXiv:2605.07278v1 Announce Type: cross Abstract: A latent world model may achieve accurate short-horizon prediction while still inducing a latent space that is poorly aligned with planning. A key iss

local-aiarxiv-cs-ai
11 May 2026
Research

Self-Consolidating Language Models: Continual Knowledge Incorporation from Context

DGX agent

arXiv:2605.07076v1 Announce Type: new Abstract: Large language models (LLMs) increasingly receive information as streams of passages, conversations, and long-context workflows. While longer context wi

researcharxiv-cs-cl
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Local Ai

ST-Gen4D: Embedding 4D Spatiotemporal Cognition into World Model for 4D Generation

DGX agent

arXiv:2605.07390v1 Announce Type: new Abstract: Generative models have achieved success in producing apparently coherent 2D videos, but remain challenging in the physical world due to lack of 4D spati

local-aiarxiv-cs-cv
11 May 2026
Model Releases

The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass

DGX agent

arXiv:2605.07282v1 Announce Type: new Abstract: Final outputs hide when a checkpoint commits to its next-token prediction. We introduce the convergence gap, a model-diffing diagnostic that decodes eac

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Tracing Uncertainty in Language Model 'Reasoning'

DGX agent

arXiv:2605.07776v1 Announce Type: cross Abstract: Language model (LM) 'reasoning', commonly described as Chain-of-Thought or test-time scaling, often improves benchmark performance, but the dynamics u

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TSRBench: A Comprehensive Multi-task Multi-modal Time Series Reasoning Benchmark for Generalist Models

DGX agent

arXiv:2601.18744v2 Announce Type: replace Abstract: Time series are ubiquitous in real-world scenarios and crucial for applications ranging from energy management to traffic control. Consequently, the

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Where's the Plan? Locating Latent Planning in Language Models with Lightweight Mechanistic Interventions

DGX agent

arXiv:2605.07984v1 Announce Type: cross Abstract: We study planning site formation in language models -- where internal representations of structurally-constrained future tokens form during the forwar

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Driver-WM: A Driver-Centric Traffic-Conditioned Latent World Model for In-Cabin Dynamics Rollout

DGX agent

arXiv:2605.05092v1 Announce Type: cross Abstract: Safe L2/L3 driving automation requires anticipating human-in-the-loop reactions during shared-control transitions. While most driving world models for

model-releasesarxiv-cs-cv
7 May 2026
Safety

Efficient Model-Based Reinforcement Learning for Robot Control via Online Optimization

DGX agent

arXiv:2510.18518v2 Announce Type: replace Abstract: We present an online model-based reinforcement learning algorithm suitable for controlling complex robotic systems directly in the real world. Unlik

safetyarxiv-cs-ro
7 May 2026
Research

External Validation of Deep Learning Models for BI-RADS Breast Density Prediction from Ultrasound Images

DGX agent

arXiv:2605.05082v1 Announce Type: cross Abstract: We externally validated three deep learning models (DenseNet121, ViT-B/32, and ResNet50) for predicting mammographic breast density from breast ultras

researcharxiv-cs-cv
7 May 2026
Research

Full-chip CMP modelling based on Fully Convolutional Network leveraging White Light Interferometry

DGX agent

arXiv:2605.05062v1 Announce Type: new Abstract: As time-to-market is crucial in the Integrated Circuit (IC) industry, speeding up layout manufacturability verifi-cation is essential. Chemical-Mechanic

researcharxiv-cs-lg
7 May 2026
Model Releases

LoViF 2026 The First Challenge on Holistic Quality Assessment for 4D World Model (PhyScore)

DGX agent

arXiv:2605.05187v1 Announce Type: new Abstract: This paper reports on the LoViF 2026 PhyScore challenge, a competition on holistic quality assessment of world-model-generated videos across both 2D and

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Manifold of Failure: Behavioral Attraction Basins in Language Models

DGX agent

arXiv:2602.22291v3 Announce Type: replace Abstract: While prior work has focused on projecting adversarial examples back onto the manifold of natural data to restore safety, we argue that a comprehens

model-releasesarxiv-cs-lg
7 May 2026
Research

Multi-site modelling and reconstruction of past extreme skew surges along the French Atlantic coast

DGX agent

arXiv:2505.00835v2 Announce Type: replace-cross Abstract: Appropriate modelling of extreme skew surges is crucial, particularly for coastal risk management. Our study focuses on modelling extreme skew

researcharxiv-cs-lg
7 May 2026
Model Releases

CC-OCR V2: Benchmarking Large Multimodal Models for Literacy in Real-world Document Processing

DGX agent

arXiv:2605.03903v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently shown strong performance on Optical Character Recognition (OCR) tasks, demonstrating their promising capabi

model-releasesarxiv-cs-cl
6 May 2026
Research

Code World Model Preparedness Report

DGX agent

arXiv:2605.00932v1 Announce Type: cross Abstract: This report documents the preparedness assessment of Code World Model (CWM), a model for code generation and reasoning about code from Meta. We conduc

researcharxiv-cs-ai
6 May 2026
Agents

Complexity Horizons of Compressed Models in Analog Circuit Analysis

DGX agent

arXiv:2605.02285v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) for specialized engineering domains, such as circuit analysis, often faces a trade-off between reasoning

agentsarxiv-cs-ai
6 May 2026
Model Releases

Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models

DGX agent

arXiv:2605.03547v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs), trained on web-scale data, risk memorizing and regenerating copyrighted visual content such as characters and logo

model-releasesarxiv-cs-cv
6 May 2026
← Previous
1…8990919293…1030
Next →