AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Model Releases

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

DGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

model-releasesarxiv-cs-cl
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DEL: Digit Entropy Loss for Numerical Learning of Large Language Models

DGX agent

arXiv:2605.20369v1 Announce Type: new Abstract: Number prediction stands as a fundamental capability of large language models (LLMs) in mathematical problem-solving and code generation. The widely ado

model-releasesarxiv-cs-cl
21 May 2026
Research

Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models

DGX agent

arXiv:2411.01141v2 Announce Type: replace Abstract: There are two shortages in the current Large Language Models (LLMs) era. The first is short of multilingual models, where most LLMs are English-cent

researcharxiv-cs-cl
21 May 2026
Research

Efficient training for compact compression models via sequential distillation

DGX agent

arXiv:2601.05639v2 Announce Type: replace Abstract: Deep learning models for image compression often face practical limitations in hardware-constrained applications. Although these models achieve high

researcharxiv-cs-cv
21 May 2026
Research

How Open Must Language Models be to Enable Reliable Scientific Inference?

DGX agent

arXiv:2603.26539v2 Announce Type: replace Abstract: How does the extent to which a model is open or closed impact the scientific inferences that can be drawn from research that involves it? In this pa

researcharxiv-cs-cl
21 May 2026
Model Releases

LamPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.21235v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving reasoning language models on tasks such as mathemat

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Leveraging Vision-Language Models to Detect Attention in Educational Videos

DGX agent

arXiv:2605.20211v1 Announce Type: new Abstract: Educational videos are a cornerstone of remote and blended learning. However, learners' fluctuating attention remains a significant barrier to effective

model-releasesarxiv-cs-cv
21 May 2026
Applications

WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems

DGX agent

arXiv:2603.14392v2 Announce Type: replace Abstract: Trajectory world models play a crucial role in robotic dynamics learning, planning, and control. While recent works have explored trajectory world m

applicationsarxiv-cs-lg
21 May 2026
Research

Bayesian Joint Model of Multi-Sensor and Failure Event Data for Multi-Mode Failure Prediction

DGX agent

arXiv:2506.17036v2 Announce Type: replace-cross Abstract: Modern industrial systems are often subject to multiple failure modes, and their conditions are monitored by multiple sensors, generating mult

researcharxiv-cs-lg
20 May 2026
Safety

Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay

DGX agent

arXiv:2605.19352v1 Announce Type: cross Abstract: Understanding how humans and artificial intelligence systems predict and plan by interacting with their environment is a fundamental challenge at the

safetyarxiv-cs-ai
20 May 2026
Safety

Chessformer: A Unified Architecture for Chess Modeling

DGX agent

arXiv:2605.19091v1 Announce Type: new Abstract: Chess has long served as a canonical testbed for artificial intelligence, but modeling approaches for its central tasks have diverged. Maximizing playin

safetyarxiv-cs-lg
20 May 2026
Research

CPC-VAR:Continual Personalized and Compositional Generation in Visual Autoregressive Models

DGX agent

arXiv:2605.19750v1 Announce Type: new Abstract: Visual autoregressive (VAR) models have recently emerged as an efficient paradigm for text-to-image generation. Despite their strong generative capabili

researcharxiv-cs-cv
20 May 2026
Safety

D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models

DGX agent

arXiv:2605.19690v1 Announce Type: new Abstract: Navigation Foundation Models (NFMs) trained on large cross-embodied datasets have demonstrated powerful generalizability in various scenarios. Adopting

safetyarxiv-cs-ro
20 May 2026
Model Releases

Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings

DGX agent

arXiv:2507.03122v2 Announce Type: replace-cross Abstract: This study investigates the feasibility and performance of federated learning (FL) for multi-label ICD code classification using clinical note

model-releasesarxiv-cs-cl
20 May 2026
Agents

High-quality generation of dynamic game content via small language models: A proof of concept

DGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

agentsarxiv-cs-ai
20 May 2026
Model Releases

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

DGX agent

arXiv:2605.19301v1 Announce Type: new Abstract: Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastroph

model-releasesarxiv-cs-cv
20 May 2026
Research

Neural Operators for Design-Space Surrogate Modeling of Tendon-Actuated Continuum Robots

DGX agent

arXiv:2605.19104v1 Announce Type: cross Abstract: Continuum robots enable dexterous manipulation in constrained environments, but require accurate and efficient models for real-time manipulation and c

researcharxiv-cs-ai
20 May 2026
Research

Probabilistic Tiny Recursive Model

DGX agent

arXiv:2605.19943v1 Announce Type: new Abstract: Tiny Recursive Models (TRM) solve complex reasoning tasks with a fraction of the parameters of modern large language models (LLMs) by iteratively refini

researcharxiv-cs-ai
20 May 2026
Model Releases

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

DGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

model-releasesarxiv-cs-ai
20 May 2026
Local Ai

Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models

DGX agent

arXiv:2605.20158v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) show promise in medical applications, but their inability to faithfully ground responses in visual evidence raise

local-aiarxiv-cs-ai
20 May 2026
Model Releases

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

DGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

model-releasesarxiv-cs-ai
20 May 2026
Research

Shaping the Prior: How Synthetic Task Distributions Determine Tabular Foundation Model Quality

DGX agent

arXiv:2605.18971v1 Announce Type: cross Abstract: What determines the quality of a tabular foundation model? Unlike language or vision, tabular foundation models acquire their inductive biases almost

researcharxiv-cs-ai
20 May 2026
Model Releases

SLoW: Select Low-frequency Words! Automatic Dictionary Selection for Translation on Large Language Models

DGX agent

arXiv:2507.18902v2 Announce Type: replace Abstract: There are more than 7,000 languages around the world, and current Large Language Models (LLMs) only support hundreds of languages. Dictionary-based

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TADA! Tuning Audio Diffusion Models through Activation Steering

DGX agent

arXiv:2602.11910v2 Announce Type: replace-cross Abstract: Audio diffusion models can synthesize high-fidelity music from text, yet achieving fine-grained control over specific musical attributes remai

model-releasesarxiv-cs-lg
20 May 2026
Tutorials

Tweedie's Formulae and Diffusion Generative Models Beyond Gaussian

DGX agent

arXiv:2605.19391v1 Announce Type: cross Abstract: Diffusion models have achieved remarkable success in generating samples from unknown data distributions. Most popular stochastic differential equation

tutorialsarxiv-cs-lg
20 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Safety

Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States

DGX agent

arXiv:2605.17144v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models leverage powerful perceptual priors from web-scale Vision-Language Model (VLM) pre-training, yet they remain surpr

safetyarxiv-cs-ai
19 May 2026
Model Releases

CrossView Suite: Harnessing Cross-view Spatial Intelligence of MLLMs with Dataset, Model and Benchmark

DGX agent

arXiv:2605.18621v1 Announce Type: cross Abstract: Spatial intelligence requires multimodal large language models (MLLMs) to move beyond single-view perception and reason consistently about objects, vi

model-releasesarxiv-cs-ai
19 May 2026
Research

Data-Driven Dynamic Modeling of a Tendon-Actuated Continuum Robot

DGX agent

arXiv:2605.18720v1 Announce Type: new Abstract: Developing dynamic models for tendon-driven continuum robots is challenging due to their nonlinear, high-dimensional, and friction-dominated dynamics. T

researcharxiv-cs-ro
19 May 2026
Safety

Diffusion Models, Denoiser Architecture and Creativity

DGX agent

arXiv:2605.16415v1 Announce Type: new Abstract: The creativity of diffusion models refers to their ability to generate highly realistic images that are different from their training data. Creativity i

safetyarxiv-cs-cv
19 May 2026
Applications

DiffWind: Physics-Informed Differentiable Modeling of Wind-Driven Object Dynamics

DGX agent

arXiv:2603.09668v2 Announce Type: replace Abstract: Modeling wind-driven object dynamics from video observations is highly challenging due to the invisibility and spatio-temporal variability of wind,

applicationsarxiv-cs-cv
19 May 2026
Safety

Dual-Space Knowledge Distillation with Key-Query Matching for Large Language Models with Vocabulary Mismatch

DGX agent

arXiv:2603.22056v2 Announce Type: replace Abstract: Large language models (LLMs) achieve state-of-the-art (SOTA) performance across language tasks, but are costly to deploy due to their size and resou

safetyarxiv-cs-cl
19 May 2026
Model Releases

Dynamic Elliptical Graph Factor Models via Riemannian Optimization with Geodesic Temporal Regularization

DGX agent

arXiv:2605.18316v1 Announce Type: new Abstract: Inferring time-varying graph structures from high-dimensional nodal observations is a fundamental problem arising in neuroscience, finance, climatology,

model-releasesarxiv-cs-lg
19 May 2026
Safety

ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation

DGX agent

arXiv:2605.17580v1 Announce Type: new Abstract: Electrocardiogram (ECG)-based models have achieved strong performance in diagnostic tasks, yet they remain limited in modeling how cardiac dynamics evol

safetyarxiv-cs-ai
19 May 2026
Model Releases

Embodied Task Planning via Graph-Informed Action Generation with Large Language Models

DGX agent

arXiv:2601.21841v3 Announce Type: replace Abstract: While Large Language Models (LLMs) have demonstrated strong zero-shot reasoning capabilities, their deployment as embodied agents still faces fundam

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Ensembling Tabular Foundation Models - A Diversity Ceiling And A Calibration Trap

DGX agent

arXiv:2605.18696v1 Announce Type: cross Abstract: Tabular foundation models (TFMs) now match or beat tuned gradient-boosted trees on a growing fraction of tabular tasks, but no single TFM wins on ever

model-releasesarxiv-cs-ai
19 May 2026
Safety

FLAG: Foundation model representation with Latent diffusion Alignment via Graph for spatial gene expression prediction

DGX agent

arXiv:2605.18055v1 Announce Type: cross Abstract: Predicting spatial gene expression from routine H&E enables large-scale molecular profiling, yet current models treat this as isolated pointwise tasks

safetyarxiv-cs-ai
19 May 2026
Model Releases

Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models

DGX agent

arXiv:2508.06038v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) incur substantial computational overhead and inference latency due to the large number of vision tokens introduc

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

FrequencyBooster: Full-Frequency Modeling for High-Fidelity Pixel Diffusion

DGX agent

arXiv:2605.17759v1 Announce Type: new Abstract: To circumvent the inherent fidelity bottlenecks and optimization misalignment of VAE-based latent diffusion, pixel-space diffusion models have emerged a

local-aiarxiv-cs-cv
19 May 2026
Model Releases

From Pixels to Places: A Systematic Benchmark for Evaluating Image Geolocalization Ability in Large Language Models

DGX agent

arXiv:2508.01608v2 Announce Type: replace Abstract: Image geolocalization, the task of identifying the geographic location depicted in an image, is important for applications in crisis response, digit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

From Static Risk to Dynamic Trajectories: Toward World-Model-Inspired Clinical Prediction

DGX agent

arXiv:2605.16927v1 Announce Type: new Abstract: Clinical decision-making is a feedback system where risk estimates influence treatment, which in turn changes disease trajectories, and both shape clini

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics

DGX agent

arXiv:2603.02531v2 Announce Type: replace-cross Abstract: Classifier-Free Guidance (CFG) improves sample quality in diffusion models, but its dual-pass inference and reliance on null-condition trainin

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

How Many Visual Tokens Do Multimodal Language Models Need? Scaling Visual Token Pruning with F^3A

DGX agent

arXiv:2605.16359v1 Announce Type: cross Abstract: Vision-language models improve perception by feeding increasingly long visual token sequences into language backbones, but the resulting inference cos

local-aiarxiv-cs-ai
19 May 2026
Safety

Lance: Unified Multimodal Modeling by Multi-Task Synergy

DGX agent

arXiv:2605.18678v1 Announce Type: cross Abstract: We present Lance, a lightweight native unified model supporting multimodal understanding, generation, and editing for both images and videos. Rather t

safetyarxiv-cs-ai
19 May 2026
Research

Large Language Models and Impossible Language Acquisition: 'False Promise' or an Overturn of our Current Perspective towards AI

DGX agent

arXiv:2602.08437v5 Announce Type: replace Abstract: In Chomsky's provocative critique 'The False Promise of CHATGPT,' Large Language Models (LLMs) are characterized as mere pattern predictors that do

researcharxiv-cs-cl
19 May 2026
Agents

Learning from Self-Debate: Preparing Reasoning Models for Multi-Agent Debate

DGX agent

arXiv:2601.22297v2 Announce Type: replace Abstract: The reasoning abilities of large language models (LLMs) have been substantially improved by reinforcement learning with verifiable rewards (RLVR). A

agentsarxiv-cs-cl
19 May 2026
Model Releases

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

DGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

DGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…110111112113114…1030
Next →