AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,490 results
Model Releases

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving

DGX agent

arXiv:2605.10564v1 Announce Type: new Abstract: End-to-end autonomous driving systems are increasingly integrating Vision-Language Model (VLM) architectures, incorporating text reasoning or visual rea

model-releasesarxiv-cs-cv
12 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

DUALFloodGNN: Physics-informed Graph Neural Network for Operational Flood Modeling

DGX agent

arXiv:2512.23964v2 Announce Type: replace-cross Abstract: Flood models inform strategic disaster management by simulating the spatiotemporal hydrodynamics of flooding. While physics-based numerical fl

local-aiarxiv-cs-ai
12 May 2026
Model Releases

EnergyLens: Interpretable Closed-Form Energy Models for Multimodal LLM Inference Serving

DGX agent

arXiv:2605.10556v1 Announce Type: new Abstract: As large language models span dense, mixture-of-experts, and state-space architectures and are deployed on heterogeneous accelerators under increasingly

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

DGX agent

arXiv:2605.10410v1 Announce Type: new Abstract: Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We

model-releasesarxiv-cs-lg
12 May 2026
Local Ai

Event Fields: Learning Latent Event Structure for Waveform Foundation Models

DGX agent

arXiv:2605.08685v1 Announce Type: cross Abstract: We propose a new class of waveform foundation models that departs from conventional sequence based representations by modeling physiological time seri

local-aiarxiv-cs-ai
12 May 2026
Model Releases

Holmes: A Benchmark to Assess the Linguistic Competence of Language Models

DGX agent

arXiv:2404.18923v5 Announce Type: replace Abstract: We introduce Holmes, a new benchmark designed to assess language models (LMs) linguistic competence - their unconscious understanding of linguistic

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

HoReN: Normalized Hopfield Retrieval for Large-Scale Sequential Model Editing

DGX agent

arXiv:2605.08143v1 Announce Type: cross Abstract: Large language models encode vast factual knowledge that inevitably becomes outdated or incorrect after deployment, yet retraining is costly prohibiti

model-releasesarxiv-cs-ai
12 May 2026
Research

How Mobile World Model Guides GUI Agents?

DGX agent

arXiv:2605.10347v1 Announce Type: new Abstract: Recent advances in vision-language models have enabled mobile GUI agents to perceive visual interfaces and execute user instructions, but reliable predi

researcharxiv-cs-ai
12 May 2026
Research

HyperTransport: Amortized Conditioning of T2I Generative Models

DGX agent

arXiv:2605.08254v1 Announce Type: cross Abstract: As foundation models grow in capability, the ability to efficiently and reliably control their behavior becomes critical. Fine-tuning these models can

researcharxiv-cs-ai
12 May 2026
Model Releases

Improving Generalization by Permutation Routing Across Model Copies

DGX agent

arXiv:2605.09256v1 Announce Type: cross Abstract: We introduce a use of the (M)-cover (or (M)-layer) transform for machine learning. The method replicates a model (M) times, but instead of coupling th

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models

DGX agent

arXiv:2605.09806v1 Announce Type: cross Abstract: Large reasoning models, such as OpenAI o1 and DeepSeek-R1, tend to become increasingly verbose as their reasoning capabilities improve. These inflated

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LLM Jaggedness Unlocks Scientific Creativity

DGX agent

arXiv:2605.10574v1 Announce Type: new Abstract: As artificial intelligence advances, models are not improving uniformly. Instead, progress unfolds in a jagged fashion, with capabilities growing uneven

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

M2A: Synergizing Mathematical and Agentic Reasoning in Large Language Models

DGX agent

arXiv:2605.09879v1 Announce Type: new Abstract: While reasoning has become a central capability of large language models (LLMs), the reasoning patterns required for different scenarios are often misal

model-releasesarxiv-cs-ai
12 May 2026
Applications

Marrying Generative Model of Healthcare Events with Digital Twin of Social Determinants of Health for Disease Reasoning

DGX agent

arXiv:2605.09771v1 Announce Type: new Abstract: Despite the central role of sensor-derived measurements such as imaging traits and plasma biomarkers in biomedical research and clinical practice, exist

applicationsarxiv-cs-ai
12 May 2026
Research

PoDAR: Power-Disentangled Audio Representation for Generative Modeling

DGX agent

arXiv:2605.10084v1 Announce Type: cross Abstract: The performance of audio latent diffusion models is primarily governed by generator expressivity and the modelability of the underlying latent space.

researcharxiv-cs-ai
12 May 2026
Safety

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

DGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

safetyarxiv-cs-ai
12 May 2026
Model Releases

PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models

DGX agent

arXiv:2605.08800v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) may memorize sensitive cross-modal information during pretraining. However, existing MLLM unlearning benchmar

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model

DGX agent

arXiv:2504.12334v2 Announce Type: replace Abstract: Large language models (LLMs) face significant challenges in specialized biomedical tasks due to the inherent complexity of medical reasoning and the

model-releasesarxiv-cs-cl
12 May 2026
Research

Quantitative Clustering in Mean-Field Transformer Models

DGX agent

arXiv:2504.14697v3 Announce Type: replace Abstract: The evolution of tokens through deep transformer models can be modeled as an interacting particle system that has been shown to exhibit an asymptoti

researcharxiv-cs-lg
12 May 2026
Safety

Revitalizing the Beginning: Avoiding Storage Dependency for Model Merging in Continual Learning

DGX agent

arXiv:2605.08311v1 Announce Type: cross Abstract: Model merging provides a compelling paradigm for integrating specialized expertise into a unified multi-task model, a goal that aligns naturally with

safetyarxiv-cs-cv
12 May 2026
Safety

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

DGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

safetyarxiv-cs-cl
12 May 2026
Model Releases

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

DGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

DGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

DGX agent

arXiv:2601.02954v3 Announce Type: replace-cross Abstract: Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models

DGX agent

arXiv:2605.09904v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have achieved remarkable progress in general video understanding, yet their ability to maintain temporal object

model-releasesarxiv-cs-cv
12 May 2026
Model Releases

Tracing Moral Foundations in Large Language Models

DGX agent

arXiv:2601.05437v2 Announce Type: replace-cross Abstract: Large language models often produce human-like moral judgments, but it is unclear whether this reflects an internal conceptual structure or su

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

When Large Vision-Language Models Meet Person Re-Identification

DGX agent

arXiv:2411.18111v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) that incorporate visual models and large language models have achieved impressive results across cross-modal un

tutorialsarxiv-cs-cv
12 May 2026
Model Releases

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

DGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

An Interpretable and Scalable Framework for Evaluating Large Language Models

DGX agent

arXiv:2605.07046v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning

DGX agent

arXiv:2602.19612v3 Announce Type: replace Abstract: Machine Unlearning (MU) enables Large Language Models (LLMs) to remove unsafe or outdated information. However, existing work assumes that all facts

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Benchmarking Foundation Models for Renal Lesion Stratification in CT

DGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, y…

DGX agent

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, your evals, your edge cases is a strong moat. OpenAI is windi

model-releasesfireworks-ai--x
11 May 2026
Safety

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

DGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

safetyarxiv-cs-cv
11 May 2026
Model Releases

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

DGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition

DGX agent

arXiv:2605.07766v1 Announce Type: new Abstract: Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are mi

model-releasesarxiv-cs-cv
11 May 2026
Tutorials

How to Train Your Latent Diffusion Language Model Jointly With the Latent Space

DGX agent

arXiv:2605.07933v1 Announce Type: new Abstract: Latent diffusion models offer an attractive alternative to discrete diffusion for non-autoregressive text generation by operating on continuous text rep

tutorialsarxiv-cs-cl
11 May 2026
Safety

Learning Visual Feature-Based World Models via Residual Latent Action

DGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

safetyarxiv-cs-ai
11 May 2026
Model Releases

NSMQ Riddles: A Benchmark of Scientific and Mathematical Riddles for Quizzing Large Language Models

DGX agent

arXiv:2605.07051v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown good performance on various science educational benchmarks, demonstrating their potential for use in science and

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Optimizing Language Models for Crosslingual Knowledge Consistency

DGX agent

arXiv:2603.04678v2 Announce Type: replace-cross Abstract: Large language models are known to often exhibit inconsistent knowledge. This is particularly problematic in multilingual scenarios, where mod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

PerfCoder: Large Language Models for Interpretable Code Performance Optimization

DGX agent

arXiv:2512.14018v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress in automatic code generation, yet their ability to produce high-performance cod

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion

DGX agent

arXiv:2503.06223v5 Announce Type: replace Abstract: Large Vision-Language Models (VLMs) are increasingly deployed in open-ended environments, where ensuring reliable safety under multimodal inputs is

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

DGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

model-releasesarxiv-cs-cl
11 May 2026
Agents

Switchcraft: AI Model Router for Agentic Tool Calling

DGX agent

arXiv:2605.07112v1 Announce Type: new Abstract: Agentic AI systems that invoke external tools are powerful but costly, leading developers to default to large models and overspend inference budgets. Mo

agentsarxiv-cs-ai
11 May 2026
Model Releases

Toeplitz MLP Mixers are Low Complexity, Information-Rich Sequence Models

DGX agent

arXiv:2605.06683v1 Announce Type: cross Abstract: Transformer-based large language models are in some respects limited by the quadratic time and space computational complexity of attention. We introdu

model-releasesarxiv-cs-ai
11 May 2026
Research

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

DGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

researcharxiv-cs-ai
11 May 2026
Agents

open models got good enough right as frontier inference pricing started creeping up

DGX agent

Open-source language models have reached sufficient quality and capability levels just as commercial frontier model providers have begun increasing their inference API pricing. This timing creates a p

agentsharrison-chase--x
8 May 2026
Research

A foundation model of vision, audition, and language for in-silico neuroscience

DGX agent

arXiv:2605.04326v1 Announce Type: cross Abstract: Cognitive neuroscience is fragmented into specialized models, each tailored to specific experimental paradigms, hence preventing a unified model of co

researcharxiv-cs-lg
7 May 2026
Model Releases

A Scalable Multi-Task Model for Virtual Sensors

DGX agent

arXiv:2601.20634v2 Announce Type: replace Abstract: Virtual sensors replace expensive physical sensors in critical applications through machine learning by predicting target signals from available mea

model-releasesarxiv-cs-lg
7 May 2026
← Previous
1…8788899091…1261
Next →