AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Model Releases

Goldfish: Monolingual Language Models for 350 Languages

DGX agent

arXiv:2408.10441v3 Announce Type: replace Abstract: For many low-resource languages, the only available language models are large multilingual models trained on many languages simultaneously. Despite

model-releasesarxiv-cs-cl
1 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Variance-Reduced Model Predictive Path Integral via Quadratic Model Approximation

DGX agent

arXiv:2602.03639v2 Announce Type: replace Abstract: Sampling-based controllers, such as Model Predictive Path Integral (MPPI) methods, offer substantial flexibility but often suffer from high variance

researcharxiv-cs-ro
1 Jun 2026
Model Releases

Weight Decay Improves Language Model Plasticity

DGX agent

arXiv:2602.11137v2 Announce Type: replace-cross Abstract: Large language models are typically trained in two broad phases: pretraining to produce a base model, followed by further training to improve

model-releasesarxiv-cs-ai
1 Jun 2026
Tutorials

Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

DGX agent

arXiv:2605.29548v1 Announce Type: new Abstract: Larger models learn tasks smaller models do not. What drives this phenomenon? We develop a simple phenomenological argument that power-law scaling alrea

tutorialsarxiv-cs-lg
29 May 2026
Research

Generating 3D models from sketches of human faces using a combined approach of Convolutional Neural Networks, Procedural Modeling, and Contour Mapping

DGX agent

arXiv:2605.25418v1 Announce Type: cross Abstract: Generating 3D models from face sketches is an active topic of research in Computer Graphics due to its potential to tremendously facilitate the modeli

researcharxiv-cs-lg
26 May 2026
Model Releases

Reverse-Engineering Model Editing on Language Models

DGX agent

arXiv:2602.10134v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are pretrained on corpora containing trillions of tokens and, therefore, inevitably memorize sensitive informatio

model-releasesarxiv-cs-ai
19 May 2026
Agents

Learning POMDP World Models from Observations with Language-Model Priors

DGX agent

arXiv:2605.13740v1 Announce Type: new Abstract: Whether navigating a building, operating a robot, or playing a game, an agent that acts effectively in an environment must first learn an internal model

agentsarxiv-cs-lg
14 May 2026
Research

Self-Distilled Trajectory-Aware Boltzmann Modeling: Bridging the Training-Inference Discrepancy in Diffusion Language Models

DGX agent

arXiv:2605.11854v1 Announce Type: new Abstract: Diffusion Language Models (DLMs) have recently emerged as a promising alternative to autoregressive language models, offering stronger global awareness

researcharxiv-cs-cl
13 May 2026
Model Releases

Statistical Model Checking of the Keynes+Schumpeter Model: A Transient Sensitivity Analysis of a Macroeconomic ABM

DGX agent

arXiv:2605.10447v1 Announce Type: cross Abstract: Agent-based models (ABMs) are increasingly used in macroeconomics, but their analysis still often relies on ad hoc Monte Carlo campaigns with heteroge

model-releasesarxiv-cs-ai
12 May 2026
Research

Model Organisms Are Leaky: Perplexity Differencing Often Reveals Finetuning Objectives

DGX agent

arXiv:2605.00994v1 Announce Type: new Abstract: Finetuning can significantly modify the behavior of large language models, including introducing harmful or unsafe behaviors. To study these risks, rese

researcharxiv-cs-cl
5 May 2026
Model Releases

Spectral Model eXplainer: a chemically-grounded explainability framework for spectral-based machine learning models

DGX agent

arXiv:2605.02684v1 Announce Type: new Abstract: Spectral-based machine learning models have been increasingly deployed in chemometrics and spectroscopy, where predictive accuracy is as important as ex

model-releasesarxiv-cs-lg
5 May 2026
Research

Foundation models for discovering robust biomarkers of neurological disorders from dynamic functional connectivity

DGX agent

arXiv:2604.22018v1 Announce Type: cross Abstract: Several brain foundation models (FM) have recently been proposed to predict brain disorders by modelling dynamic functional connectivity (FC). While t

researcharxiv-cs-ai
27 Apr 2026
Safety

SafeMERGE: Preserving Safety Alignment in Fine-Tuned Large Language Models via Selective Layer-Wise Model Merging

DGX agent

arXiv:2503.17239v3 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) is a common practice to adapt generalist models to specialized domains. However, recent studies show

safetyarxiv-cs-ai
24 Apr 2026
Safety

VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models

DGX agent

arXiv:2510.18457v3 Announce Type: replace Abstract: The performance of Latent Diffusion Models (LDMs) is critically dependent on the quality of their visual tokenizers. While recent works have explore

safetyarxiv-cs-cv
24 Apr 2026
Safety

Cloning Deterministic Worlds: The Critical Role of Latent Geometry in Long-Horizon World Models

DGX agent

arXiv:2510.26782v3 Announce Type: replace-cross Abstract: A world model is an internal model that simulates how the world evolves. Given past observations and actions, it predicts the future physical

safetyarxiv-cs-ai
22 Apr 2026
Agents

Don't Adapt Small Language Models for Tools; Adapt Tool Schemas to the Models

DGX agent

arXiv:2510.07248v3 Announce Type: replace Abstract: Small language models (SLMs) enable scalable tool-augmented multi-agent systems where multiple SLMs handle subtasks orchestrated by a powerful coord

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training

DGX agent

arXiv:2604.07484v1 Announce Type: cross Abstract: Generative reward models (GRMs) have emerged as a promising approach for aligning Large Language Models (LLMs) with human preferences by offering grea

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Distributed Interpretability and Control for Large Language Models

DGX agent

arXiv:2604.06483v1 Announce Type: cross Abstract: Large language models that require multiple GPU cards to host are usually the most capable models. It is necessary to understand and steer these model

model-releasesarxiv-cs-ai
10 Apr 2026
Research

Towards Hierarchical Multi-Step Reward Models for Enhanced Reasoning in Large Language Models

DGX agent

arXiv:2503.13551v5 Announce Type: replace Abstract: Recent studies show that Large Language Models (LLMs) achieve strong reasoning capabilities through supervised fine-tuning or reinforcement learning

researcharxiv-cs-cl
10 Apr 2026
Agents

4D-WAM: 4D Consistent World Modeling for Autonomous Driving

DGX agent

arXiv:2608.10107v1 Announce Type: new Abstract: Emerging World-Action Models (WAMs) have demonstrated promising performance in autonomous driving by jointly modeling future driving scene evolution and

agentsarxiv-cs-cv
12 Aug 2026
Model Releases

TACTICL: Task-Aware Compression of Tabular ICL Models

DGX agent

arXiv:2608.10837v1 Announce Type: cross Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Revisiting Predictive Process Monitoring in the Age of Foundation Models: A Comparative Study of Sequence, Tabular, and LLM Approaches

DGX agent

arXiv:2607.27797v1 Announce Type: new Abstract: Predictive process monitoring (PPM) leverages event logs to forecast the future of running process instances, for instance, predicting the next activity

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Tag Questions and the Generational Reversal of Sycophancy Across 45 Language Models

DGX agent

arXiv:2607.23976v1 Announce Type: cross Abstract: Appending a two-word confirmation tag to a decision question -- 'Is X the better choice?' versus 'X is the better choice, right?' -- changes whether a

model-releasesarxiv-cs-ai
28 Jul 2026
Tutorials

Robot-Factored World Models via Robot Rendering

DGX agent

arXiv:2607.22535v1 Announce Type: cross Abstract: Action-conditioned video world models predict future observations from an initial observation and an action signal. In robotics, actions influence fut

tutorialsarxiv-cs-cv
27 Jul 2026
Research

Small Vision-Language Models Know When They Are Wrong But Cannot Say So: A Two-Model Study of Stated versus Internal Confidence Under Realistic Image Degradation

DGX agent

arXiv:2607.22034v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on consumer hardware where input images are degraded by compression, camera shake, and poor li

researcharxiv-cs-cl
27 Jul 2026
Model Releases

Collective Intelligence with Foundation Models

DGX agent

arXiv:2607.07729v1 Announce Type: cross Abstract: As foundation models grow in scale and diversity, coordinating multiple models into cooperative reasoning systems offers a path toward safer, more rel

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation

DGX agent

arXiv:2607.02642v1 Announce Type: new Abstract: Evaluating embodied robot foundation models remains a critical bottleneck; unlike large language models efficiently assessed via digital benchmarks, rob

model-releasesarxiv-cs-ro
7 Jul 2026
Research

Optimal Mixture-of-Experts Model Averaging for Conditional Generative Models

DGX agent

arXiv:2607.04360v1 Announce Type: cross Abstract: Conditional generative models have emerged as powerful tools for sampling from target conditional distributions, driving substantial advances across a

researcharxiv-cs-lg
7 Jul 2026
Local Ai

The Multiscale Single-Index Model: A Stylized Model for Hierarchical Feature Learning

DGX agent

arXiv:2607.03347v1 Announce Type: new Abstract: We consider the Multiscale Single-Index Model (MSIM), first introduced in ite{oymak2021learning}, as a stylized model for hierarchical learning with sca

local-aiarxiv-cs-lg
7 Jul 2026
Tutorials

When Models Know When They Do Not Know: Calibration, Cascading, and Cleaning

DGX agent

arXiv:2601.07965v2 Announce Type: replace Abstract: When a model knows when it does not know, many possibilities emerge. The first question is how to enable a model to recognize that it does not know.

tutorialsarxiv-cs-ai
30 Jun 2026
Research

Embedding Foundation Model Predictions in Discrete-Choice Models with Structural Guarantees

DGX agent

arXiv:2606.26432v1 Announce Type: new Abstract: Tabular foundation models achieve strong accuracy on choice prediction tasks, but their predictions often violate the economic logic those tasks require

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Atomistic Language Models Understand and Generate Materials

DGX agent

arXiv:2606.21395v1 Announce Type: new Abstract: Atomistic structure and natural language have long been modeled separately, with language models either calling atomistic models as tools or being fine-

model-releasesarxiv-cs-lg
23 Jun 2026
Research

Turning Tabular Foundation Models into Graph Foundation Models

DGX agent

arXiv:2508.20906v3 Announce Type: replace Abstract: While foundation models have revolutionized fields such as natural language processing and computer vision, their potential in graph machine learnin

researcharxiv-cs-lg
23 Jun 2026
Safety

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

DGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

i1: A Simple and Fully Open Recipe for Strong Text-to-Image Models

DGX agent

arXiv:2606.11289v1 Announce Type: new Abstract: Diffusion models have consistently driven progress in text-to-image generation. However, it is challenging to attribute recent progress to specific mode

model-releasesarxiv-cs-cv
11 Jun 2026
Model Releases

Scaling Laws of Global Weather Models

DGX agent

arXiv:2602.22962v2 Announce Type: replace Abstract: Data-driven models are revolutionizing weather forecasting. To optimize training efficiency and model performance, this paper analyzes empirical sca

model-releasesarxiv-cs-lg
11 Jun 2026
Safety

Toward Preference-aligned Large Language Models via Residual-based Model Steering

DGX agent

arXiv:2509.23982v2 Announce Type: replace-cross Abstract: Preference alignment is a critical step in making Large Language Models (LLMs) useful and aligned with (human) preferences. Existing approache

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

DGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

model-releasesarxiv-cs-cl
11 Jun 2026
Research

The Emergence of Reproducibility and Generalizability in Diffusion Models

DGX agent

arXiv:2310.05264v5 Announce Type: replace-cross Abstract: In this work, we investigate an intriguing and prevalent phenomenon of diffusion models which we term as 'consistent model reproducibility': g

researcharxiv-cs-cv
10 Jun 2026
Local Ai

Model Poisoning Against Federated Model Adaptation with Chain of Bit-Flips

DGX agent

arXiv:2606.09548v1 Announce Type: cross Abstract: Federated Learning (FL) allows a set of clients to collectively train a global model without sharing local training data. Giving the responsibility of

local-aiarxiv-cs-ai
9 Jun 2026
Applications

Attend to Anything: Foundation Model for Unified Human Attention Modeling

DGX agent

arXiv:2606.03540v1 Announce Type: new Abstract: Existing human attention (saliency) modeling methods persist as highly fragmented across modalities, scenes, and task formulations. Consequently, even w

applicationsarxiv-cs-cv
3 Jun 2026
Research

Representational Capacity: Geometric Limits on Feature Representation in Transformer Language Models

DGX agent

arXiv:2606.02765v1 Announce Type: cross Abstract: Model dimension (d_{model}) is a fundamental hyperparameter in transformer language models, yet its role in setting the geometric limits of feature re

researcharxiv-cs-ai
3 Jun 2026
Safety

World Models Meet Language Models: On the Complementarity of Concrete and Abstract Reasoning

DGX agent

arXiv:2606.03603v1 Announce Type: cross Abstract: World models and multimodal large language models (MLLMs) provide complementary capabilities for predicting future outcomes from static visual observa

safetyarxiv-cs-cl
3 Jun 2026
Model Releases

AMix-2: Establishing Protein as a Native Modality in Large Language Models

DGX agent

arXiv:2605.30963v1 Announce Type: cross Abstract: We present AMix-2, a protein-text foundation model that establishes protein as a native modality in large language models (LLMs), unifying protein und

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

Dual-Stream Diffusion for World-Model Augmented Vision-Language-Action Model

DGX agent

arXiv:2510.27607v3 Announce Type: replace Abstract: Augmenting vision-language-action models (VLAs) with world models is promising for robotic policy learning but faces challenges in jointly predictin

safetyarxiv-cs-cv
29 May 2026
Model Releases

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

DGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Falcon-X: A Time Series Foundation Model for Heterogeneous Multivariate Modeling

DGX agent

arXiv:2605.27286v1 Announce Type: cross Abstract: Time series foundation models (TSFMs) are transforming the forecasting paradigm through large-scale cross-domain pretraining. However, most existing T

model-releasesarxiv-cs-ai
27 May 2026
Tutorials

GT-SVJ: Generative-Transformer-Based Self-Supervised Video Judge For Efficient Video Reward Modeling

DGX agent

arXiv:2602.05202v2 Announce Type: replace Abstract: Aligning video generative models with human preferences remains challenging: current approaches rely on Vision-Language Models (VLMs) for reward mod

tutorialsarxiv-cs-cv
25 May 2026
← Previous
1…34567…1012
Next →