AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Model Releases

Sharpness-Aware Hybrid Model Learning for Architecture-Agnostic Parameter Estimation

DGX agent

arXiv:2602.06837v2 Announce Type: replace Abstract: Hybrid modeling, the combination of machine learning models and scientific mathematical models, enables flexible and robust data-driven prediction w

model-releasesarxiv-cs-lg
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

DGX agent

arXiv:2604.22548v1 Announce Type: cross Abstract: Problem definition: Data-driven models in machine learning have enabled efficient management of production systems. However, a majority of machine lea

model-releasesarxiv-cs-lg
27 Apr 2026
Safety

Toward Reusability of AI Models Using Dynamic Updates of AI Documentation

DGX agent

arXiv:2604.17626v1 Announce Type: cross Abstract: This work addresses the challenge of disseminating reusable artificial intelligence (AI) models accompanied by AI documentation (a.k.a., AI model card

safetyarxiv-cs-cl
21 Apr 2026
Research

C2: Scalable Rubric-Augmented Reward Modeling from Binary Preferences

DGX agent

arXiv:2604.13618v1 Announce Type: new Abstract: Rubric-augmented verification guides reward models with explicit evaluation criteria, yielding more reliable judgments than single-model verification. H

researcharxiv-cs-cl
16 Apr 2026
Research

A2-DIDM: Privacy-preserving Accumulator-enabled Auditing for Distributed Identity of DNN Model

DGX agent

arXiv:2405.04108v2 Announce Type: replace-cross Abstract: Recent booming development of Generative Artificial Intelligence (GenAI) has facilitated model commercialization to reinforce the model perfor

researcharxiv-cs-ai
15 Apr 2026
Research

Navigating the Accuracy-Size Trade-Off with Flexible Model Merging

DGX agent

arXiv:2505.23209v3 Announce Type: replace Abstract: Model merging has emerged as an efficient method to combine multiple single-task fine-tuned models. The merged model can enjoy multi-task capabiliti

researcharxiv-cs-cv
15 Apr 2026
Safety

Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges

DGX agent

arXiv:2607.28636v1 Announce Type: new Abstract: LLMs increasingly serve as automated judges, but their judgments remain vulnerable to cognitive biases. Existing mitigations mostly rely on prompt-drive

safetyarxiv-cs-cl
3 Aug 2026
Model Releases

Model Gateway: Management Platform for Model-Driven Drug Discovery

DGX agent

arXiv:2512.05462v2 Announce Type: replace-cross Abstract: Pharmaceutical drug discovery demands machine learning (ML) infrastructure that goes beyond general-purpose Machine Learning Operations (MLOps

model-releasesarxiv-cs-lg
23 Jul 2026
Research

Knowledgeless Language Models: Suppressing Parametric Recall for Evidence-Grounded Language Modeling

DGX agent

arXiv:2607.12831v1 Announce Type: new Abstract: Language models encode substantial factual knowledge in their parameters, which can lead to unreliable behavior when this knowledge is outdated, incompl

researcharxiv-cs-cl
15 Jul 2026
Model Releases

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models

DGX agent

arXiv:2607.08665v1 Announce Type: new Abstract: Routing among large language models (LLMs) trades response quality against serving cost, motivated by the reported gap between deployed routers and a pe

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

A Systematic Survey on Large Language Models for Evolutionary Optimization: From Modeling to Solving

DGX agent

arXiv:2509.08269v5 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly integrated with evolutionary computation to support optimization tasks. This survey primarily fo

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology

DGX agent

arXiv:2607.01033v1 Announce Type: new Abstract: Model organisms (MOs) - language models trained to exhibit undesired or unnatural behaviours - are frequently used as testbeds for evaluating white-box

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

AutoTrainess: Teaching Language Models to Improve Language Models Autonomously

DGX agent

arXiv:2606.31551v1 Announce Type: new Abstract: Training language models (LMs) remains a highly human-intensive process, even as frontier language model agents become increasingly capable at software

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Training for the Model You Return: Improving Optimization for Iterate-Averaged Language Models

DGX agent

arXiv:2606.25086v1 Announce Type: new Abstract: Many modern Language Model (LM) pipelines return an averaged model, such as an exponential moving average of the training iterates, rather than the fina

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Harmonic: Hierarchical State Space Models for Efficient Long-Context Language Modeling

DGX agent

arXiv:2606.24650v1 Announce Type: new Abstract: We present Harmonic, a hierarchical state space model (SSM) for language modeling. The architecture stacks three recurrent levels at progressively slowe

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

World Model Self-Distillation: Training World Models to Solve General Tasks

DGX agent

arXiv:2606.12072v1 Announce Type: new Abstract: Pretrained video generators are promising visual world models that exhibit emergent task-solving abilities; however, their reliance on detailed textual

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency

DGX agent

arXiv:2606.10620v1 Announce Type: cross Abstract: Image generation models now produce high-quality static images, yet their ability to represent how a visual world changes over time remains poorly und

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Solving Inverse Problems with Flow-based Models via Model Predictive Control

DGX agent

arXiv:2601.23231v2 Announce Type: replace-cross Abstract: Flow-based generative models provide strong unconditional priors for inverse problems, but guiding their dynamics for conditional generation r

model-releasesarxiv-cs-lg
9 Jun 2026
Research

Revisiting Model Stitching In the Foundation Model Era

DGX agent

arXiv:2603.12433v3 Announce Type: replace-cross Abstract: Model stitching, connecting early layers of one model (source) to later layers of another (target) via a light stitch layer, has served as a p

researcharxiv-cs-ai
4 Jun 2026
Model Releases

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models

DGX agent

arXiv:2605.29585v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to answer questions about physical scenes, yet most evaluations reduce performance to a final answer

model-releasesarxiv-cs-cl
29 May 2026
Research

Multi-Mixer Models: Flexible Sequence Modeling with Shared Representations

DGX agent

arXiv:2605.28769v1 Announce Type: new Abstract: Softmax attention is the cornerstone of modern large language models, but its memory scales linearly and compute quadratically with sequence length. Lin

researcharxiv-cs-lg
28 May 2026
Safety

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

DGX agent

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

safetyarxiv-cs-lg
26 May 2026
Safety

Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuning

DGX agent

arXiv:2605.22356v1 Announce Type: new Abstract: Large language models are increasingly used as computational tools for modeling human-like behavior. We introduce a behavioral induction framework that

safetyarxiv-cs-cl
22 May 2026
Research

Flow Map Language Models: One-step Language Modeling via Continuous Denoising

DGX agent

arXiv:2602.16813v3 Announce Type: replace Abstract: Language models based on discrete diffusion have attracted widespread interest for their potential to provide faster generation than autoregressive

researcharxiv-cs-cl
21 May 2026
Model Releases

Could Large Language Models work as Post-hoc Explainability Tools in Credit Risk Models?

DGX agent

arXiv:2602.18895v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in translating model-based explanations into human-readable narratives. This study evaluates w

model-releasesarxiv-cs-lg
19 May 2026
Safety

GeoWorld-VLM: Geometry from World Models for Vision-Language Models

DGX agent

arXiv:2605.16713v1 Announce Type: cross Abstract: Modern Vision-Language Models (VLMs) achieve strong semantic recognition, yet remain brittle on elementary spatial relations such as left of, on, behi

safetyarxiv-cs-ai
19 May 2026
Applications

ExplainerPFN: Towards tabular foundation models for model-free zero-shot feature importance estimations

DGX agent

arXiv:2601.23068v2 Announce Type: replace-cross Abstract: Computing the importance of features in supervised classification tasks is critical for model interpretability. Shapley values are a widely us

applicationsarxiv-cs-ai
18 May 2026
Local Ai

JEDI: Joint Embedding Diffusion World Model for Online Model-Based Reinforcement Learning

DGX agent

arXiv:2605.13013v1 Announce Type: new Abstract: Diffusion world models have recently become competitive for online model-based reinforcement learning, but current approaches expose a tension: pixel di

local-aiarxiv-cs-lg
14 May 2026
Local Ai

COSMOS: Model-Agnostic Personalized Federated Learning with Clustered Server Models and Pseudo-Label-Only Communication

DGX agent

arXiv:2605.11165v1 Announce Type: new Abstract: Federated learning (FL) in heterogeneous environments remains challenging because client models often differ in both architecture and data distribution.

local-aiarxiv-cs-lg
13 May 2026
Research

The Bicameral Model: Bidirectional Hidden-State Coupling Between Parallel Language Models

DGX agent

arXiv:2605.11167v1 Announce Type: new Abstract: Existing multi-model and tool-augmented systems communicate by generating text, serializing every exchange through the output vocabulary. Can two pretra

researcharxiv-cs-cl
13 May 2026
Model Releases

Setting-Matched and Semantics-Scaled Benchmarking of One-Step Generative Models Against Multistep Diffusion and Flow Models

DGX agent

arXiv:2603.14186v4 Announce Type: replace Abstract: State-of-the-art text-to-image models produce high-quality images, but inference remains expensive as generation requires several sequential ODE or

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Fin-PRM: A Domain-Specialized Process Reward Model for Financial Reasoning in Large Language Models

DGX agent

arXiv:2508.15202v2 Announce Type: replace Abstract: Process Reward Models (PRMs) supervise intermediate reasoning steps in large language models (LLMs), but existing PRMs are mainly trained on general

model-releasesarxiv-cs-cl
5 May 2026
Research

Poodle: Seamlessly Scaling Down Large Language Models with Just-in-Time Model Replacement

DGX agent

arXiv:2512.05525v2 Announce Type: replace-cross Abstract: Businesses increasingly rely on large language models (LLMs) to automate simple repetitive tasks instead of developing custom machine learning

researcharxiv-cs-lg
5 May 2026
Research

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling

DGX agent

arXiv:2604.27039v1 Announce Type: new Abstract: Token serves as the fundamental unit of computation in modern autoregressive models, and generation length directly influences both inference cost and r

researcharxiv-cs-cl
1 May 2026
Research

Iterative Model-Learning Scheme via Gaussian Processes for Nonlinear Model Predictive Control of (Semi-)Batch Processes

DGX agent

arXiv:2604.22672v1 Announce Type: new Abstract: Batch processes are inherently transient and typically nonlinear, motivating nonlinear model predictive control (NMPC). However, adopting NMPC is hinder

researcharxiv-cs-lg
27 Apr 2026
Research

Do We Need Bigger Models for Science? Task-Aware Retrieval with Small Language Models

DGX agent

arXiv:2604.01965v2 Announce Type: replace-cross Abstract: Scientific knowledge discovery increasingly relies on large language models, yet many existing scholarly assistants depend on proprietary syst

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Cross-Model Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Across Three Large Language Models

DGX agent

arXiv:2604.19598v1 Announce Type: cross Abstract: This study compared repeated generation consistency of exercise prescription outputs across three large language models (LLMs), specifically GPT-4.1,

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Multi-Model Synthetic Training for Mission-Critical Small Language Models

DGX agent

arXiv:2509.13047v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across many domains, yet their application to specialized fields remain

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging

DGX agent

arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that generate step-by-step reasoning before making predi

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Test, then Route: How Language Models Execute In-Context Conditional Rules Across Models and Languages

DGX agent

arXiv:2608.04183v1 Announce Type: new Abstract: When a language model follows an in-context conditional rule such as 'if P(x) then A else B,' does it assemble a runtime circuit with one module that te

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

Simulation Code Generation for Fluid Systems using Large Language Models: Benchmarking Models and Prompting Strategies

DGX agent

arXiv:2607.29389v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated a strong ability to generate syntactically correct code from natural-language specifications. In this stu

model-releasesarxiv-cs-lg
3 Aug 2026
Model Releases

A Characterization of the Orthocomplement of the Tangent Space of Semiparametric Markov Models

DGX agent

arXiv:2607.23439v1 Announce Type: cross Abstract: Graphical models are ubiquitous in social and empirical science as they are intuitive and easy to use. These models belong to the broader class of Mar

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Medical-Checklist: Assessing the Comprehension of Medical Images by Multimodal Models

DGX agent

arXiv:2607.21998v1 Announce Type: new Abstract: This paper introduces a new benchmark test, Medical-Checklist, for assessing medical multimodal models. The recent advancements in multimodal models hav

model-releasesarxiv-cs-cv
27 Jul 2026
Tutorials

STeMP: Spatio-Temporal Modelling Protocol

DGX agent

arXiv:2607.20592v1 Announce Type: new Abstract: Spatio-temporal machine-learning modelling is an important tool in environmental research. However, machine-learning models are highly sensitive to both

tutorialsarxiv-cs-lg
24 Jul 2026
Model Releases

The World Model Remembers, the Actor Forgets: Dream Rehearsal for Continual Model-Based RL

DGX agent

arXiv:2607.19749v1 Announce Type: cross Abstract: Model-based reinforcement-learning agents of the DreamerV3 family forget catastrophically when trained on task sequences, even when an unbounded repla

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders

DGX agent

arXiv:2607.14088v1 Announce Type: new Abstract: Video generative models commonly rely on latent spaces learned by 3D Variational Autoencoders (3D-VAEs). However, conventional 3D-VAEs are mainly optimi

model-releasesarxiv-cs-cv
16 Jul 2026
Research

Reduced-Order Models: The Mother of World Models

DGX agent

arXiv:2607.03198v1 Announce Type: new Abstract: World models -- compressed latent representations of an environment that support action-conditioned prediction and planning -- are typically presented a

researcharxiv-cs-lg
7 Jul 2026
Model Releases

Embodied.cpp: A Portable Inference Runtime of Embodied AI Models on Heterogeneous Robots

DGX agent

arXiv:2607.02501v1 Announce Type: new Abstract: Embodied AI models now span vision-language-action (VLA) models and world-action models (WAMs), but practical deployment remains fragmented across model

model-releasesarxiv-cs-ro
3 Jul 2026
← Previous
12345…1012
Next →