AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
13 May 2026

DriftXpress: Faster Drifting Models via Projected RKHS Fields

ResearchDGX agent

arXiv:2605.12183v1 Announce Type: new Abstract: Drifting Models have emerged as a new paradigm for one-step generative modeling, achieving strong image quality without iterative inference. The premise

GeneZip: Region-Aware Compression for Long Context DNA Modeling

Model ReleasesDGX agent

arXiv:2602.17739v3 Announce Type: replace-cross Abstract: Long-context DNA models are limited by token-mixing cost and by how compression allocates representational budget across the genome. Existing

Language Modeling with Hyperspherical Flows

ResearchDGX agent

arXiv:2605.11125v1 Announce Type: new Abstract: Discrete Diffusion Language Models progressed rapidly as an alternative to autoregressive (AR) models, motivated by their parallel generation abilities.

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

OmniThoughtVis: A Scalable Distillation Pipeline for Deployable Multimodal Reasoning Models

ApplicationsDGX agent

arXiv:2605.11629v1 Announce Type: new Abstract: Recent multimodal large language models (MLLMs) have shown strong chain-of-thought (CoT) reasoning ability on vision-language tasks, but their direct de

Overcoming Dynamics-Blindness: Training-Free Pace-and-Path Correction for VLA Models

Model ReleasesDGX agent

arXiv:2605.11459v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models achieve remarkable flexibility and generalization beyond classical control paradigms. However, most prevailing VLA

Parameter-Efficient Adaptation of Pre-Trained Vision Foundation Models for Active and Passive Seismic Data Denoising

Model ReleasesDGX agent

arXiv:2605.10953v1 Announce Type: cross Abstract: The demand for high-resolution subsurface imaging and continuous Earth monitoring has driven rapid growth in active and passive seismic data from dens

ReasonEdit: Editing Vision-Language Models using Human Reasoning

ResearchDGX agent

arXiv:2602.02408v4 Announce Type: replace Abstract: Model editing aims to correct errors in large, pretrained models without altering unrelated behaviors. While some recent works have edited vision-la

Rethinking external validation for the target population: Capturing patient-level similarity with a generative model

SafetyDGX agent

arXiv:2605.11284v1 Announce Type: cross Abstract: Background: External validation is essential for assessing the transportability of predictive models. However, its interpretation is often confounded

STRUM: A Spectral Transcription and Rhythm Understanding Model for End-to-End Generation of Playable Rhythm-Game Charts

Model ReleasesDGX agent

arXiv:2605.12135v1 Announce Type: cross Abstract: We present STRUM (Spectral Transcription and Rhythm Understanding Model), an audio-to-chart pipeline that converts raw recordings into playable Clone

TCP-SSM: Efficient Vision State Space Models with Token-Conditioned Poles

ResearchDGX agent

arXiv:2605.11563v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as a compelling alternative to attention models for long-range vision tasks, offering input-dependent recurrence

12 May 2026

Alice v1: Distillation-Enhanced Video Generation Surpassing Closed-Source Models

Model ReleasesDGX agent

arXiv:2605.08115v1 Announce Type: cross Abstract: Wepresent Alice v1, a 14-billion parameter open-source video generation model that achieves state-of-the-art quality through consistency distillation

Architecture, Not Scale: Circuit Localization in Large Language Models

Model ReleasesDGX agent

arXiv:2605.08853v1 Announce Type: new Abstract: Mechanistic interpretability assumes that circuit analysis becomes harder as models scale. We challenge this assumption by showing that the attention ar

Beyond Language: Format-Agnostic Reasoning Subspaces in Large Language Models

Model ReleasesDGX agent

arXiv:2605.09496v1 Announce Type: new Abstract: Large language models represent the same reasoning in vastly different surface forms -- English prose, Python code, mathematical notation -- yet whether

Beyond Local Edits: Embedding-Virtualized Knowledge for Broader Evaluation and Preservation of Model Editing

Model ReleasesDGX agent

arXiv:2602.01977v2 Announce Type: replace Abstract: Knowledge editing methods for large language models are commonly evaluated using predefined benchmarks that assess edited facts together with a limi

Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models

Model ReleasesDGX agent

arXiv:2605.08404v1 Announce Type: cross Abstract: This work investigates the use of large language models (LLMs) for tasks in smart cities. The core idea is to leverage remote sensing imagery to chara

DeepSight: Long-Horizon World Modeling via Latent States Prediction for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2605.10564v1 Announce Type: new Abstract: End-to-end autonomous driving systems are increasingly integrating Vision-Language Model (VLM) architectures, incorporating text reasoning or visual rea

DUALFloodGNN: Physics-informed Graph Neural Network for Operational Flood Modeling

Local AiDGX agent

arXiv:2512.23964v2 Announce Type: replace-cross Abstract: Flood models inform strategic disaster management by simulating the spatiotemporal hydrodynamics of flooding. While physics-based numerical fl

EnergyLens: Interpretable Closed-Form Energy Models for Multimodal LLM Inference Serving

Model ReleasesDGX agent

arXiv:2605.10556v1 Announce Type: new Abstract: As large language models span dense, mixture-of-experts, and state-space architectures and are deployed on heterogeneous accelerators under increasingly

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

Model ReleasesDGX agent

arXiv:2605.10410v1 Announce Type: new Abstract: Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We

Event Fields: Learning Latent Event Structure for Waveform Foundation Models

Local AiDGX agent

arXiv:2605.08685v1 Announce Type: cross Abstract: We propose a new class of waveform foundation models that departs from conventional sequence based representations by modeling physiological time seri

Holmes: A Benchmark to Assess the Linguistic Competence of Language Models

Model ReleasesDGX agent

arXiv:2404.18923v5 Announce Type: replace Abstract: We introduce Holmes, a new benchmark designed to assess language models (LMs) linguistic competence - their unconscious understanding of linguistic

HoReN: Normalized Hopfield Retrieval for Large-Scale Sequential Model Editing

Model ReleasesDGX agent

arXiv:2605.08143v1 Announce Type: cross Abstract: Large language models encode vast factual knowledge that inevitably becomes outdated or incorrect after deployment, yet retraining is costly prohibiti

How Mobile World Model Guides GUI Agents?

ResearchDGX agent

arXiv:2605.10347v1 Announce Type: new Abstract: Recent advances in vision-language models have enabled mobile GUI agents to perceive visual interfaces and execute user instructions, but reliable predi

HyperTransport: Amortized Conditioning of T2I Generative Models

ResearchDGX agent

arXiv:2605.08254v1 Announce Type: cross Abstract: As foundation models grow in capability, the ability to efficiently and reliably control their behavior becomes critical. Fine-tuning these models can

Improving Generalization by Permutation Routing Across Model Copies

Model ReleasesDGX agent

arXiv:2605.09256v1 Announce Type: cross Abstract: We introduce a use of the (M)-cover (or (M)-layer) transform for machine learning. The method replicates a model (M) times, but instead of coupling th

LEAD: Length-Efficient Adaptive and Dynamic Reasoning for Large Language Models

Model ReleasesDGX agent

arXiv:2605.09806v1 Announce Type: cross Abstract: Large reasoning models, such as OpenAI o1 and DeepSeek-R1, tend to become increasingly verbose as their reasoning capabilities improve. These inflated

LLM Jaggedness Unlocks Scientific Creativity

Model ReleasesDGX agent

arXiv:2605.10574v1 Announce Type: new Abstract: As artificial intelligence advances, models are not improving uniformly. Instead, progress unfolds in a jagged fashion, with capabilities growing uneven

M2A: Synergizing Mathematical and Agentic Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2605.09879v1 Announce Type: new Abstract: While reasoning has become a central capability of large language models (LLMs), the reasoning patterns required for different scenarios are often misal

Marrying Generative Model of Healthcare Events with Digital Twin of Social Determinants of Health for Disease Reasoning

ApplicationsDGX agent

arXiv:2605.09771v1 Announce Type: new Abstract: Despite the central role of sensor-derived measurements such as imaging traits and plasma biomarkers in biomedical research and clinical practice, exist

PoDAR: Power-Disentangled Audio Representation for Generative Modeling

ResearchDGX agent

arXiv:2605.10084v1 Announce Type: cross Abstract: The performance of audio latent diffusion models is primarily governed by generator expressivity and the modelability of the underlying latent space.

Political Plasticity: An Analysis of Ideological Adaptability in Large Language Models

SafetyDGX agent

arXiv:2605.08415v1 Announce Type: new Abstract: Since the advent of Large Language Models (LLMs), a significant area of research has focused on their intrinsic biases, particularly in political discou

PPU-Bench:Real World Benchmark for Personalized Partial Unlearning in Vision Language Models

Model ReleasesDGX agent

arXiv:2605.08800v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) may memorize sensitive cross-modal information during pretraining. However, existing MLLM unlearning benchmar

QM-ToT: A Medical Tree of Thoughts Reasoning Framework for Quantized Model

Model ReleasesDGX agent

arXiv:2504.12334v2 Announce Type: replace Abstract: Large language models (LLMs) face significant challenges in specialized biomedical tasks due to the inherent complexity of medical reasoning and the

Quantitative Clustering in Mean-Field Transformer Models

ResearchDGX agent

arXiv:2504.14697v3 Announce Type: replace Abstract: The evolution of tokens through deep transformer models can be modeled as an interacting particle system that has been shown to exhibit an asymptoti

Revitalizing the Beginning: Avoiding Storage Dependency for Model Merging in Continual Learning

SafetyDGX agent

arXiv:2605.08311v1 Announce Type: cross Abstract: Model merging provides a compelling paradigm for integrating specialized expertise into a unified multi-task model, a goal that aligns naturally with

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

SafetyDGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

Model ReleasesDGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

The Echo Amplifies the Knowledge: Somatic Marker Analogues in Language Models via Emotion Vector Re-Injection

Model ReleasesDGX agent

arXiv:2605.08611v1 Announce Type: new Abstract: Current language model memory systems store what happened but not how it felt. This distinction -- between semantic memory (knowing about a past event)

The World is Not Mono: Enabling Spatial Understanding in Large Audio-Language Models

Model ReleasesDGX agent

arXiv:2601.02954v3 Announce Type: replace-cross Abstract: Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding

TOC-Bench: A Temporal Object Consistency Benchmark for Video Large Language Models

Model ReleasesDGX agent

arXiv:2605.09904v1 Announce Type: new Abstract: Video large language models (Video-LLMs) have achieved remarkable progress in general video understanding, yet their ability to maintain temporal object

Tracing Moral Foundations in Large Language Models

Model ReleasesDGX agent

arXiv:2601.05437v2 Announce Type: replace-cross Abstract: Large language models often produce human-like moral judgments, but it is unclear whether this reflects an internal conceptual structure or su

When Large Vision-Language Models Meet Person Re-Identification

TutorialsDGX agent

arXiv:2411.18111v2 Announce Type: replace Abstract: Large Vision-Language Models (LVLMs) that incorporate visual models and large language models have achieved impressive results across cross-modal un

When Prompts Become Payloads: A Framework for Mitigating SQL Injection Attacks in Large Language Model-Driven Applications

Model ReleasesDGX agent

arXiv:2605.10176v1 Announce Type: cross Abstract: Natural language interfaces to structured databases are becoming increasingly common, largely due to advances in large language models (LLMs) that ena

11 May 2026

An Interpretable and Scalable Framework for Evaluating Large Language Models

Model ReleasesDGX agent

arXiv:2605.07046v1 Announce Type: cross Abstract: Evaluation of large language models (LLMs) is increasingly critical, yet standard benchmarking methods rely on average accuracy, overlooking both the

Anatomy of Unlearning: The Dual Impact of Fact Salience and Model Fine-Tuning

Model ReleasesDGX agent

arXiv:2602.19612v3 Announce Type: replace Abstract: Machine Unlearning (MU) enables Large Language Models (LLMs) to remove unsafe or outdated information. However, existing work assumes that all facts

Benchmarking Foundation Models for Renal Lesion Stratification in CT

Model ReleasesDGX agent

arXiv:2605.07749v1 Announce Type: new Abstract: The rapid proliferation of open-source medical foundation models (FMs) raises a practical question: how well do their pre-trained representations transf

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, y…

Model ReleasesDGX agent

Fine-tuning on your proprietary data is the highest leverage thing you can do. Prompts get copied overnight. A model trained on your data, your evals, your edge cases is a strong moat. OpenAI is windi

GPO-V: Jailbreak Diffusion Vision Language Model by Global Probability Optimization

SafetyDGX agent

arXiv:2605.07399v1 Announce Type: new Abstract: Diffusion Vision-Language Models (dVLMs), built upon the non-causal foundations of Diffusion Large Language Models (dLLMs), have demonstrated remarkable

Graph Representation Learning Augmented Model Manipulation on Federated Fine-Tuning of LLMs

Model ReleasesDGX agent

arXiv:2605.07961v1 Announce Type: new Abstract: Federated fine-tuning (FFT) has emerged as a privacy-preserving paradigm for collaboratively adapting large language models (LLMs). Built upon federated

Head Similarity: Modeling Structured Whole-Head Appearance Beyond Face Recognition

Model ReleasesDGX agent

arXiv:2605.07766v1 Announce Type: new Abstract: Many vision applications require identity consistency beyond strict biometric recognition, especially under non-frontal views or when facial cues are mi

How to Train Your Latent Diffusion Language Model Jointly With the Latent Space

TutorialsDGX agent

arXiv:2605.07933v1 Announce Type: new Abstract: Latent diffusion models offer an attractive alternative to discrete diffusion for non-autoregressive text generation by operating on continuous text rep

Learning Visual Feature-Based World Models via Residual Latent Action

SafetyDGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

NSMQ Riddles: A Benchmark of Scientific and Mathematical Riddles for Quizzing Large Language Models

Model ReleasesDGX agent

arXiv:2605.07051v1 Announce Type: new Abstract: Large Language Models (LLMs) have shown good performance on various science educational benchmarks, demonstrating their potential for use in science and

Optimizing Language Models for Crosslingual Knowledge Consistency

Model ReleasesDGX agent

arXiv:2603.04678v2 Announce Type: replace-cross Abstract: Large language models are known to often exhibit inconsistent knowledge. This is particularly problematic in multilingual scenarios, where mod

PerfCoder: Large Language Models for Interpretable Code Performance Optimization

Model ReleasesDGX agent

arXiv:2512.14018v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable progress in automatic code generation, yet their ability to produce high-performance cod

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion

Model ReleasesDGX agent

arXiv:2503.06223v5 Announce Type: replace Abstract: Large Vision-Language Models (VLMs) are increasingly deployed in open-ended environments, where ensuring reliable safety under multimodal inputs is

S2S-Arena: Evaluating Paralinguistic Instruction Following in Speech-to-Speech Models

Model ReleasesDGX agent

arXiv:2503.05085v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have fundamentally reshaped speech-to-speech (S2S) systems, enabling increasingly natural spoken int

Switchcraft: AI Model Router for Agentic Tool Calling

AgentsDGX agent

arXiv:2605.07112v1 Announce Type: new Abstract: Agentic AI systems that invoke external tools are powerful but costly, leading developers to default to large models and overspend inference budgets. Mo

Toeplitz MLP Mixers are Low Complexity, Information-Rich Sequence Models

Model ReleasesDGX agent

arXiv:2605.06683v1 Announce Type: cross Abstract: Transformer-based large language models are in some respects limited by the quadratic time and space computational complexity of attention. We introdu

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

ResearchDGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

← Previous
1…6970717273…1007
Next →