AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
48,543 results
Safety

Rethinking On-Policy Self-Distillation for Thinking Models

DGX agent

arXiv:2607.05184v1 Announce Type: new Abstract: Self-distillation is a promising recipe for self-improvement in language models. In this setting, a model can serve as its own teacher when given privil

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Spectral Signatures of Large Language Models

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.03377v1 Announce Type: cross Abstract: The rapidly growing repository of publicly available large language models (LLMs) presents significant challenges for systematic management and quanti

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Labeling Training Data for Entity Matching Using Large Language Models

DGX agent

arXiv:2606.28823v1 Announce Type: new Abstract: Recent large language models (LLMs) achieve strong performance on entity matching without requiring task-specific training data. However, applying these

model-releasesarxiv-cs-cl
30 Jun 2026
Research

MuonSSM: Orthogonalizing State Space Models for Sequence Modeling

DGX agent

arXiv:2606.30461v1 Announce Type: new Abstract: State space models (SSMs) have emerged as efficient linear-time alternatives to attention for long-sequence modeling. However, existing SSMs often suffe

researcharxiv-cs-lg
30 Jun 2026
Research

Proofs of Ownership for Machine Learning Models

DGX agent

arXiv:2606.30423v1 Announce Type: new Abstract: With the increasing adoption of Machine Learning, protecting model ownership has become an essential challenge. We initiate a formal study of Proof of O

researcharxiv-cs-lg
30 Jun 2026
Tutorials

Temporal Feature Extractors in EEG Foundation Models: A Controlled Comparison Including a Pretrained Time-Series Model

DGX agent

arXiv:2606.30104v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models aim to learn generalizable representations from large-scale brain recordings. However, the role of tempor

tutorialsarxiv-cs-ai
30 Jun 2026
Local Ai

AdaReP:Adaptive Re-Planning under Model Mismatch for Neural World-Model Predictive Control

DGX agent

arXiv:2606.23079v1 Announce Type: new Abstract: Neural world models coupled with model predictive control (MPC) replan at every environment step to bound accumulated prediction error, but this incurs

local-aiarxiv-cs-ro
23 Jun 2026
Model Releases

Model Merging in the Essential Subspace

DGX agent

arXiv:2602.20208v2 Announce Type: replace Abstract: Model merging aims to integrate multiple task-specific fine-tuned models derived from a shared pre-trained checkpoint into a single multi-task model

model-releasesarxiv-cs-lg
23 Jun 2026
Tutorials

Scalable Bayesian Additive Models for Stellar Flare Detection via Amortized Gaussian Process Inference and Hidden Markov Models

DGX agent

arXiv:2606.22601v1 Announce Type: cross Abstract: Gaussian Processes (GPs) are a powerful tool for Bayesian time-series modeling, yet their cubic computational cost remains a severe barrier for applic

tutorialsarxiv-cs-lg
23 Jun 2026
Model Releases

Building Social World Models with Large Language Models

DGX agent

arXiv:2606.11482v1 Announce Type: cross Abstract: Understanding and predicting how social beliefs evolve in response to events -- from policy changes to scientific breakthroughs -- remains a fundament

model-releasesarxiv-cs-cl
11 Jun 2026
Research

T2S: A Rehearsal-Based Approach for Extraction-Resistant Model Watermarking

DGX agent

arXiv:2606.11698v1 Announce Type: cross Abstract: Model watermarking safeguards AI model intellectual property by embedding distinctive knowledge that induces unique behavioral signatures. The primary

researcharxiv-cs-ai
11 Jun 2026
Model Releases

ABLE: Representing and Mapping LLMs via Attribution-Based Large-model Embedding

DGX agent

arXiv:2606.07524v1 Announce Type: cross Abstract: The explosive growth of large language models (LLMs) has created a heterogeneous and poorly documented ecosystem, making systematic model comparison i

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

FIT-Print: Towards False-claim-resistant Model Ownership Verification via Targeted Fingerprint

DGX agent

arXiv:2501.15509v5 Announce Type: replace-cross Abstract: Model fingerprinting has emerged as a crucial mechanism for safeguarding the intellectual property of open-source models, offering a non-intru

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

GlucoFM-Bench: Benchmarking Time-Series Foundation Models for Blood Glucose Forecasting

DGX agent

arXiv:2606.06881v1 Announce Type: new Abstract: Blood glucose forecasting models are foundational for modern diabetes management systems, as reliable short-term predictions can enable proactive interv

model-releasesarxiv-cs-lg
8 Jun 2026
Safety

Dynamic Policy Learning for Legged Robot with Simplified Model Pretraining and Model-Homotopy-Inspired Transfer

DGX agent

arXiv:2512.24698v2 Announce Type: replace Abstract: Generating dynamic motions for legged robots remains a challenging problem. While reinforcement learning has achieved notable success in various leg

safetyarxiv-cs-ro
4 Jun 2026
Model Releases

Physically Viable World Models: A Case for Query-Conditioned Embodied AI

DGX agent

arXiv:2605.30542v1 Announce Type: new Abstract: World models for embodied AI must be physically viable: constructed to answer intervention queries by representing the physical structure governing acti

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Cert-LAS: Toward Certified Model Ownership Verification for Text-to-Image Diffusion Models via Layer-Adaptive Smoothing

DGX agent

arXiv:2605.29809v1 Announce Type: cross Abstract: Large-scale text-to-image (T2I) diffusion models have enabled unprecedented creative applications, but their unauthorized use has raised serious intel

researcharxiv-cs-cv
29 May 2026
Safety

When and How Human Curation Backfires: Preference Alignment under Multi-Model Self-Consuming Loop

DGX agent

arXiv:2605.29267v1 Announce Type: new Abstract: Foundation models are increasingly trained on synthetic data generated by prior model iterations rather than exclusively on real data. This self-consumi

safetyarxiv-cs-ai
29 May 2026
Tutorials

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

DGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

tutorialsarxiv-cs-ai
29 May 2026
Research

Bridging Maximum Likelihood and Optimal Transport for Efficient Inference and Model Selection in Stochastic Block Models

DGX agent

arXiv:2605.28488v1 Announce Type: cross Abstract: We study inference in stochastic block models (SBMs) through the lens of optimal transport (OT). We first establish that maximum likelihood variationa

researcharxiv-cs-lg
28 May 2026
Research

Generating Robust Portfolios of Optimization Models using Large Language Models

DGX agent

arXiv:2605.27013v1 Announce Type: new Abstract: Mathematical optimization is a powerful tool for structured decision-making across domains such as resource allocation and planning. Formulating optimiz

researcharxiv-cs-ai
27 May 2026
Research

Message-Passing State-Space Models: Improving Graph Learning with Modern Sequence Modeling

DGX agent

arXiv:2505.18728v2 Announce Type: replace-cross Abstract: The recent success of State-Space Models (SSMs) in sequence modeling has motivated their adaptation to graph learning, giving rise to Graph St

researcharxiv-cs-ai
27 May 2026
Local Ai

The Model Parking Tax: Quantifying the Hidden Energy Cost of Always-On GPU Model Deployment

DGX agent

arXiv:2605.23918v1 Announce Type: cross Abstract: The AI inference industry keeps models loaded in GPU memory around the clock to avoid cold-start latency, implicitly treating idle power as a fixed co

local-aiarxiv-cs-lg
26 May 2026
Model Releases

Cost-Effective Model Evaluation with Meta-Learning

DGX agent

arXiv:2605.23595v1 Announce Type: cross Abstract: The rapid growth of machine learning has produced an ever-expanding ecosystem of models, making it increasingly challenging to verify the reliability

model-releasesarxiv-cs-ai
25 May 2026
Model Releases

Recursive Block-Diagonal Coupling for Resource-Efficient Training of Vision Models

DGX agent

arXiv:2605.23656v1 Announce Type: new Abstract: Training high-capacity vision models from scratch requires substantial computational resources. To improve training efficiency of a wide target model, e

model-releasesarxiv-cs-cv
25 May 2026
Research

Enhancing Causal Reasoning in Large Language Models: A Causal Attribution Model for Precision Fine-Tuning

DGX agent

arXiv:2401.00139v3 Announce Type: replace-cross Abstract: This paper introduces a causal attribution model to enhance the interpretability of large language models (LLMs) and improve their causal reas

researcharxiv-cs-cl
22 May 2026
Model Releases

LAION-C: An Out-of-Distribution Benchmark for Web-Scale Vision Models

DGX agent

arXiv:2506.16950v2 Announce Type: replace Abstract: Out-of-distribution (OOD) robustness is a desired property of computer vision models. Improving model robustness requires high-quality signals from

model-releasesarxiv-cs-cv
21 May 2026
Tutorials

Language Model Memory and Memory Models for Language

DGX agent

arXiv:2602.13466v2 Announce Type: replace-cross Abstract: The ability of machine learning models to store input information in hidden layer vector embeddings, analogous to the concept of `memory', is

tutorialsarxiv-cs-ai
20 May 2026
Research

When Does Model Collapse Occur in Structured Interactive Learning?

DGX agent

arXiv:2605.20151v1 Announce Type: new Abstract: The proliferation of generative artificial intelligence has given rise to an interactive learning environment, where model parameters are continuously u

researcharxiv-cs-lg
20 May 2026
Model Releases

TinySAM 2: Extreme Memory Compression for Efficient Track Anything Model

DGX agent

arXiv:2605.18013v1 Announce Type: cross Abstract: Segment Anything Model 2 (SAM 2) serves as a core foundation model in the field of video segmentation. Building upon the original SAM model, it introd

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

VideoVerse: Does Your T2V Generator Have World Model Capability to Synthesize Videos?

DGX agent

arXiv:2510.08398v4 Announce Type: replace Abstract: The recent rapid advancement of Text-to-Video (T2V) generation technologies are engaging the trained models with more world model ability, making th

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Persona-Model Collapse in Emergent Misalignment

DGX agent

arXiv:2605.12850v1 Announce Type: cross Abstract: Fine-tuning large language models on narrow data with harmful content produces broadly misaligned behavior on unrelated prompts, a phenomenon known as

model-releasesarxiv-cs-ai
14 May 2026
Research

Dirichlet process mixtures of block g priors for model selection and prediction in linear models

DGX agent

arXiv:2411.00471v3 Announce Type: replace-cross Abstract: This paper introduces Dirichlet process mixtures of block g priors for model selection and prediction in linear models. These priors are exten

researcharxiv-cs-lg
13 May 2026
Research

Why do Large Language Models Fail in Low-resource Translation? Unraveling the Token Dynamics of Large Language Models for Machine Translation

DGX agent

arXiv:2605.07533v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently demonstrated strong performance in machine translation (MT). However, most prior work focuses on improving or

researcharxiv-cs-cl
11 May 2026
Applications

A Consistency-Centric Approach to Set-Based Optimization with Multiple Models of Unranked Fidelity

DGX agent

arXiv:2605.04051v1 Announce Type: cross Abstract: In complex real-world settings, optimization is challenged by the presence of diverse models of differing fidelity. In many optimization problems, a s

applicationsarxiv-cs-lg
7 May 2026
Model Releases

Deep Reprogramming Distillation for Medical Foundation Models

DGX agent

arXiv:2605.04447v1 Announce Type: new Abstract: Medical foundation models pre-trained on large-scale datasets have shown powerful versatile performance. However, when adapting medical foundation model

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

Deployment-Relevant Alignment Cannot Be Inferred from Model-Level Evaluation Alone

DGX agent

arXiv:2605.04454v1 Announce Type: cross Abstract: Alignment evaluation in machine learning has largely become evaluation of models. Influential benchmarks score model outputs under fixed inputs, such

model-releasesarxiv-cs-lg
7 May 2026
Model Releases

AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages

DGX agent

arXiv:2601.06395v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly multilingual, yet open models continue to underperform relative to proprietary systems, with the gap m

model-releasesarxiv-cs-cl
6 May 2026
Safety

LLM-XTM: Enhancing Cross-Lingual Topic Models with Large Language Models

DGX agent

arXiv:2605.03299v1 Announce Type: new Abstract: Cross-lingual topic modeling aims to discover shared semantic structures across languages, yet existing models depend on sparse bilingual resources and

safetyarxiv-cs-cl
6 May 2026
Agents

Video Generation Models as World Models: Efficient Paradigms, Architectures and Algorithms

DGX agent

arXiv:2603.28489v2 Announce Type: cross Abstract: The rapid evolution of video generation has enabled models to simulate complex physical dynamics and long-horizon causalities, positioning them as pot

agentsarxiv-cs-cv
6 May 2026
Hardware

Maistros: A Greek Large Language Model Adapted Through Knowledge Distillation From Large Reasoning Models

DGX agent

arXiv:2605.01870v1 Announce Type: new Abstract: Large Language Models (LLMs) have substantially advanced the field of Natural Language Processing (NLP), achieving state-of-the-art performance across a

hardwarearxiv-cs-cl
5 May 2026
Model Releases

Training-Free Adaptation of New-Generation LLMs using Legacy Clinical Models

DGX agent

arXiv:2601.03423v3 Announce Type: replace-cross Abstract: Adapting language models to the clinical domain through continued pretraining and instruction tuning requires costly retraining for each new m

model-releasesarxiv-cs-ai
30 Apr 2026
Research

Conditional Score-Based Modeling of Effective Langevin Dynamics

DGX agent

arXiv:2604.23952v1 Announce Type: cross Abstract: Stochastic reduced-order models are widely used to represent the effective dynamics of complex systems, but estimating their drift and diffusion coeff

researcharxiv-cs-lg
28 Apr 2026
Model Releases

MTRouter: Cost-Aware Multi-Turn LLM Routing with History-Model Joint Embeddings

DGX agent

arXiv:2604.23530v1 Announce Type: cross Abstract: Multi-turn, long-horizon tasks are increasingly common for large language models (LLMs), but solving them typically requires many sequential model inv

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modeling

DGX agent

arXiv:2604.23586v1 Announce Type: cross Abstract: Joint audio-video generation models have shown that unified generation yields stronger cross-modal coherence than cascaded approaches. However, existi

researcharxiv-cs-cl
28 Apr 2026
Model Releases

Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models

DGX agent

arXiv:2510.07632v2 Announce Type: replace Abstract: Frontier AI models have achieved remarkable progress, yet recent studies suggest they struggle with compositional reasoning, often performing at or

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Preserving Knowledge in Large Language Model with Model-Agnostic Self-Decompression

DGX agent

arXiv:2406.11354v3 Announce Type: replace-cross Abstract: Humans can retain old knowledge while learning new information, but Large Language Models (LLMs) often suffer from catastrophic forgetting whe

researcharxiv-cs-ai
24 Apr 2026
Model Releases

RewardBench 2: Advancing Reward Model Evaluation

DGX agent

arXiv:2506.01937v2 Announce Type: replace Abstract: Reward models are used throughout the post-training of language models to capture nuanced signals from preference data and provide a training target

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…56789…1012
Next →