AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Research

The Pre-Training Study of Expanded-SPLADE Models on Web Document Titles

DGX agent

arXiv:2605.01407v1 Announce Type: cross Abstract: Masked Language Modeling (MLM) pre-training is one of the primary ways to initialize Neural Information Retrieval (IR) models prior to retrieval fine-

researcharxiv-cs-cl
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

When Correct Isn't Usable: Improving Structured Output Reliability in Small Language Models

DGX agent

arXiv:2605.02363v1 Announce Type: new Abstract: Deployed language models must produce outputs that are both correct and format-compliant. We study this structured-output reliability gap using two math

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks

DGX agent

arXiv:2507.01955v3 Announce Type: replace Abstract: Multimodal foundation models (MFMs), such as GPT-4o, have recently made remarkable progress. However, their detailed visual understanding beyond que

model-releasesarxiv-cs-cv
4 May 2026
Model Releases

Jailbroken Frontier Models Retain Their Capabilities

DGX agent

arXiv:2605.00267v1 Announce Type: new Abstract: As language model safeguards become more robust, attackers are pushed toward developing increasingly complex jailbreaks. Prior work has found that this

model-releasesarxiv-cs-lg
4 May 2026
Safety

MotuBrain: An Advanced World Action Model for Robot Control

DGX agent

arXiv:2604.27792v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models achieve strong semantic generalization but often lack fine-grained modeling of world dynamics. Recent work explores

safetyarxiv-cs-ro
1 May 2026
Research

Learning Illumination Control in Diffusion Models

DGX agent

arXiv:2604.24877v1 Announce Type: new Abstract: Controlling illumination in images is essential for photography and visual content creation. While closed-source models have demonstrated impressive ill

researcharxiv-cs-cv
29 Apr 2026
Research

Revisiting the Past: Data Unlearning with Model State History

DGX agent

arXiv:2506.20941v3 Announce Type: replace Abstract: Large language models are trained on massive corpora of web data, which may include private data, copyrighted material, factually inaccurate data, o

researcharxiv-cs-lg
29 Apr 2026
Model Releases

EgoDyn-Bench: Evaluating Ego-Motion Understanding in Vision-Centric Foundation Models for Autonomous Driving

DGX agent

arXiv:2604.22851v1 Announce Type: cross Abstract: While Vision-Language Models (VLMs) have advanced highlevel reasoning in autonomous driving, their ability to ground this reasoning in the underlying

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

Evaluating whether AI models would sabotage AI safety research

DGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

LinguDistill: Recovering Linguistic Ability in Vision-Language Models via Selective Cross-Modal Distillation

DGX agent

arXiv:2604.00829v3 Announce Type: replace-cross Abstract: Adapting pretrained language models (LMs) into vision-language models (VLMs) can degrade their native linguistic capability due to representat

safetyarxiv-cs-cl
28 Apr 2026
Research

On the Memorization of Consistency Distillation for Diffusion Models

DGX agent

arXiv:2604.23552v1 Announce Type: cross Abstract: Diffusion models are central to modern generative modeling, and understanding how they balance memorization and generalization is critical for reliabl

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Patching LLM Like Software: A Lightweight Method for Improving Safety Policy in Large Language Models

DGX agent

arXiv:2511.08484v2 Announce Type: replace Abstract: We propose patching for large language models (LLMs) like software versions, a lightweight and modular approach for addressing safety vulnerabilitie

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Are Video Models Emerging as Zero-Shot Learners and Reasoners in Medical Imaging?

DGX agent

arXiv:2510.10254v2 Announce Type: replace Abstract: Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-sh

researcharxiv-cs-cv
27 Apr 2026
Model Releases

Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models

DGX agent

arXiv:2507.04023v3 Announce Type: replace Abstract: Large language models (LLMs) achieve impressive performance on complex mathematical benchmarks yet sometimes fail on basic math reasoning while gene

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

OMIBench: Benchmarking Olympiad-Level Multi-Image Reasoning in Large Vision-Language Model

DGX agent

arXiv:2604.20806v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) have made substantial advances in reasoning tasks at the Olympiad level. Nevertheless, current Olympiad-level mul

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Surrogate modeling for interpreting black-box LLMs in medical predictions

DGX agent

arXiv:2604.20331v1 Announce Type: cross Abstract: Large language models (LLMs), trained on vast datasets, encode extensive real-world knowledge within their parameters, yet their black-box nature obsc

applicationsarxiv-cs-ai
23 Apr 2026
Applications

Beyond Coefficients: Forecast-Necessity Testing for Interpretable Causal Discovery in Nonlinear Time-Series Models

DGX agent

arXiv:2604.18751v1 Announce Type: cross Abstract: Nonlinear machine-learning models are increasingly used to discover causal relationships in time-series data, yet the interpretation of their outputs

applicationsarxiv-cs-ai
22 Apr 2026
Applications

Handling and Interpreting Missing Modalities in Patient Clinical Trajectories via Autoregressive Sequence Modeling

DGX agent

arXiv:2604.18753v1 Announce Type: cross Abstract: An active challenge in developing multimodal machine learning (ML) models for healthcare is handling missing modalities during training and deployment

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

IndiaFinBench: An Evaluation Benchmark for Large Language Model Performance on Indian Financial Regulatory Text

DGX agent

arXiv:2604.19298v1 Announce Type: cross Abstract: We introduce IndiaFinBench, to our knowledge the first publicly available evaluation benchmark for assessing large language model (LLM) performance on

model-releasesarxiv-cs-ai
22 Apr 2026
Research

OmniVoice: Towards Omnilingual Zero-Shot Text-to-Speech with Diffusion Language Models

DGX agent

arXiv:2604.00688v3 Announce Type: replace Abstract: We present OmniVoice, a massively multilingual zero-shot text-to-speech (TTS) model that scales to over 600 languages. At its core is a novel diffus

researcharxiv-cs-cl
22 Apr 2026
Model Releases

Finding Culture-Sensitive Neurons in Vision-Language Models

DGX agent

arXiv:2510.24942v2 Announce Type: replace-cross Abstract: Despite their impressive performance, vision-language models (VLMs) still struggle on culturally situated inputs. To understand how VLMs proce

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

From Heads to Neurons: Causal Attribution and Steering in Multi-Task Vision-Language Models

DGX agent

arXiv:2604.17941v1 Announce Type: cross Abstract: Recent work has increasingly explored neuron-level interpretation in vision-language models (VLMs) to identify neurons critical to final predictions.

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Heterogeneity in Formal Linguistic Competence of Language Models: Is Data the Real Bottleneck?

DGX agent

arXiv:2604.17930v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit a puzzling disparity in their formal linguistic competence: while they learn some linguistic phenomena with near-pe

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

DGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Reciprocal Co-Training (RCT): Coupling Gradient-Based and Non-Differentiable Models via Reinforcement Learning

DGX agent

arXiv:2604.16378v1 Announce Type: new Abstract: Large language models (LLMs) and classical machine learning methods offer complementary strengths for predictive modeling, yet their fundamentally diffe

tutorialsarxiv-cs-cl
21 Apr 2026
Model Releases

ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinforcement Learning

DGX agent

arXiv:2603.05863v2 Announce Type: replace Abstract: While Large Language Models (LLMs) have revolutionized code generation, standard ``System 1'' approaches that generate solutions in a single forward

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection

DGX agent

arXiv:2508.11281v3 Announce Type: replace Abstract: Detecting toxic content using language models is crucial yet challenging. While substantial progress has been made in English, toxicity detection in

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Understanding Counting Mechanisms in Large Language and Vision-Language Models

DGX agent

arXiv:2511.17699v2 Announce Type: replace Abstract: Counting is one of the fundamental abilities of large language models (LLMs) and large vision-language models (LVLMs). This paper examines how these

researcharxiv-cs-cv
21 Apr 2026
Model Releases

BAGEL: Benchmarking Animal Knowledge Expertise in Language Models

DGX agent

arXiv:2604.16241v1 Announce Type: cross Abstract: Large language models have shown strong performance on broad-domain knowledge and reasoning benchmarks, but it remains unclear how well language model

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Cost-Aware Model Orchestration for LLM-based Systems

DGX agent

arXiv:2512.01099v2 Announce Type: replace Abstract: As modern artificial intelligence (AI) systems become more advanced and capable, they can leverage a wide range of tools and models to perform compl

researcharxiv-cs-ai
20 Apr 2026
Model Releases

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

DGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Using Large Language Models and Knowledge Graphs to Improve the Interpretability of Machine Learning Models in Manufacturing

DGX agent

arXiv:2604.16280v1 Announce Type: new Abstract: Explaining Machine Learning (ML) results in a transparent and user-friendly manner remains a challenging task of Explainable Artificial Intelligence (XA

applicationsarxiv-cs-ai
20 Apr 2026
Research

Cornfigurator: Automated Planning for Any-to-Any Multimodal Model Serving

DGX agent

arXiv:2512.14098v3 Announce Type: replace Abstract: Any-to-Any models are an emerging class of multimodal models that accept combinations of text and multimodal data as input and generate them as outp

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Internal Knowledge Without External Expression: Probing the Generalization Boundary of a Classical Chinese Language Model

DGX agent

arXiv:2604.14180v1 Announce Type: new Abstract: We train a 318M-parameter Transformer language model from scratch on a curated corpus of 1.56 billion tokens of pure Classical Chinese, with zero Englis

model-releasesarxiv-cs-cl
17 Apr 2026
Local Ai

Logo-LLM: Local and Global Modeling with Large Language Models for Time Series Forecasting

DGX agent

arXiv:2505.11017v2 Announce Type: replace Abstract: Time series forecasting is critical across multiple domains, where time series data exhibit both local patterns and global dependencies. While Trans

local-aiarxiv-cs-lg
17 Apr 2026
Research

Dataset-Level Metrics Attenuate Non-Determinism: A Fine-Grained Non-Determinism Evaluation in Diffusion Language Models

DGX agent

arXiv:2604.13413v1 Announce Type: new Abstract: Diffusion language models (DLMs) have emerged as a promising paradigm for large language models (LLMs), yet the non-deterministic behavior of DLMs remai

researcharxiv-cs-lg
16 Apr 2026
Model Releases

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

DGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

model-releasesarxiv-cs-cv
16 Apr 2026
Research

Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model

DGX agent

arXiv:2510.18165v2 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) are emerging as a powerful and promising alternative to the dominant autoregressive paradigm, offering inhere

researcharxiv-cs-cl
16 Apr 2026
Model Releases

Transcriptomic Models for Immunotherapy Response Prediction Show Limited Cross-cohort Generalisability

DGX agent

arXiv:2604.05478v2 Announce Type: replace-cross Abstract: Immune checkpoint inhibitors (ICIs) have transformed cancer therapy; yet substantial proportion of patients exhibit intrinsic or acquired resi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Latent Chain-of-Thought World Modeling for End-to-End Driving

DGX agent

arXiv:2512.10226v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as a way to improve driving performance and safet

model-releasesarxiv-cs-cv
15 Apr 2026
Tutorials

NoisePrints: Distortion-Free Watermarks for Authorship in Private Diffusion Models

DGX agent

arXiv:2510.13793v2 Announce Type: replace Abstract: With the rapid adoption of diffusion models for visual content generation, proving authorship and protecting copyright have become critical. This ch

tutorialsarxiv-cs-cv
15 Apr 2026
Model Releases

Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework

DGX agent

arXiv:2509.18127v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) enable interpretability research by decomposing entangled model activations into monosemantic features. However, un

model-releasesarxiv-cs-ai
15 Apr 2026
Research

Variational Autoencoding Discrete Diffusion with Enhanced Dimensional Correlations Modeling

DGX agent

arXiv:2505.17384v2 Announce Type: replace-cross Abstract: Discrete diffusion models have recently shown great promise for modeling complex discrete data, with masked diffusion models (MDMs) offering a

researcharxiv-cs-cv
15 Apr 2026
Model Releases

CArtBench: Evaluating Vision-Language Models on Chinese Art Understanding, Interpretation, and Authenticity

DGX agent

arXiv:2604.11632v1 Announce Type: new Abstract: We introduce CARTBENCH, a museum-grounded benchmark for evaluating vision-language models (VLMs) on Chinese artworks beyond short-form recognition and Q

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Design Principles for Sequence Models via Coefficient Dynamics

DGX agent

arXiv:2510.09389v2 Announce Type: replace-cross Abstract: Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamental

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Do LLMs Build Spatial World Models? Evidence from Grid-World Maze Tasks

DGX agent

arXiv:2604.10690v1 Announce Type: new Abstract: Foundation models have shown remarkable performance across diverse tasks, yet their ability to construct internal spatial world models for reasoning and

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Edu-MMBias: A Three-Tier Multimodal Benchmark for Auditing Social Bias in Vision-Language Models under Educational Contexts

DGX agent

arXiv:2604.10200v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become integral to educational decision-making, ensuring their fairness is paramount. However, current text-centric eva

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ExecTune: Effective Steering of Black-Box LLMs with Guide Models

DGX agent

arXiv:2604.09741v1 Announce Type: cross Abstract: For large language models deployed through black-box APIs, recurring inference costs often exceed one-time training costs. This motivates composed age

model-releasesarxiv-cs-ai
14 Apr 2026
← Previous
1…1415161718…1012
Next →