AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
49,435 results
Safety

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models

DGX agent

arXiv:2605.15055v1 Announce Type: cross Abstract: Reinforcement learning has emerged as a powerful tool for improving diffusion-based text-to-image models, but existing methods are largely limited to

safetyarxiv-cs-cv
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

DGX agent

arXiv:2605.14517v1 Announce Type: cross Abstract: Holistic evaluation scores capture overall output quality but do not distinguish whether a model reproduced the structural form of a user's request fr

safetyarxiv-cs-ai
15 May 2026
Safety

Distribution Corrected Offline Data Distillation for Large Language Models

DGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

safetyarxiv-cs-cl
15 May 2026
Research

Factorization-Error-Free Discrete Diffusion Language Model via Speculative Decoding

DGX agent

arXiv:2605.14305v1 Announce Type: new Abstract: Discrete diffusion language models improve generation efficiency through parallel token prediction, but standard X_0 prediction methods introduce factor

researcharxiv-cs-cl
15 May 2026
Research

Geometry-Aware Decoding with Wasserstein-Regularized Truncation and Mass Penalties for Large Language Models

DGX agent

arXiv:2602.10346v2 Announce Type: replace Abstract: Large language models (LLMs) must balance diversity and creativity against logical coherence in open-ended generation. Existing truncation-based sam

researcharxiv-cs-cl
15 May 2026
Model Releases

IntentVLA: Short-Horizon Intent Modeling for Aliased Robot Manipulation

DGX agent

arXiv:2605.14712v1 Announce Type: cross Abstract: Robot imitation data are often multimodal: similar visual-language observations may be followed by different action chunks because human demonstrators

model-releasesarxiv-cs-ai
15 May 2026
Tutorials

Large Language Models for Web Accessibility: A Systematic Literature Review

DGX agent

arXiv:2605.13873v1 Announce Type: cross Abstract: Web accessibility aims to ensure that web content and services are usable by people with diverse abilities. In recent years, Large Language Models (LL

tutorialsarxiv-cs-ai
15 May 2026
Model Releases

Learning Cross-Coupled and Regime Dependent Dynamics for Aerial Manipulation

DGX agent

arXiv:2605.14805v1 Announce Type: new Abstract: Accurate dynamics models are critical for aerial manipulators operating under complex tasks such as payload transport. However, modeling these systems r

model-releasesarxiv-cs-ro
15 May 2026
Model Releases

LLMs Should Express Uncertainty Explicitly

DGX agent

arXiv:2604.05306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often produce confident yet incorrect answers, which can lead to risky failures in real-world applications. We st

model-releasesarxiv-cs-ai
15 May 2026
Research

MPU: Towards Secure and Privacy-Preserving Knowledge Unlearning for Large Language Models

DGX agent

arXiv:2602.23798v2 Announce Type: replace-cross Abstract: Machine unlearning for large language models often faces a privacy dilemma in which strict constraints prohibit sharing either the server's pa

researcharxiv-cs-ai
15 May 2026
Research

Non-linear Interventions on Large Language Models

DGX agent

arXiv:2605.14749v1 Announce Type: cross Abstract: Intervention is one of the most representative and widely used methods for understanding the internal representations of large language models (LLMs).

researcharxiv-cs-ai
15 May 2026
Safety

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music

DGX agent

arXiv:2605.14765v1 Announce Type: cross Abstract: Persian music, with its unique tonalities, modal systems (Dastgah), and rhythmic structures, presents significant challenges for music generation mode

safetyarxiv-cs-cl
15 May 2026
Applications

ReasonCache: Accelerating Large Reasoning Model Serving through KV Cache Sharing

DGX agent

arXiv:2507.21433v3 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) are becoming integral to many AI inference systems, enhancing their capabilities with advanced reasoning. Howeve

applicationsarxiv-cs-ai
15 May 2026
Safety

Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases

DGX agent

arXiv:2601.03630v2 Announce Type: replace Abstract: This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. O

safetyarxiv-cs-cl
15 May 2026
Research

Scalable Subset Selection in Linear Mixed Models

DGX agent

arXiv:2506.20425v3 Announce Type: replace-cross Abstract: Linear mixed models (LMMs), which incorporate fixed and random effects, are key tools for analyzing heterogeneous data, such as in personalize

researcharxiv-cs-lg
15 May 2026
Tutorials

SeesawNet: Towards Non-stationary Time Series Forecasting with Balanced Modeling of Common and Specific Dependencies

DGX agent

arXiv:2605.14551v1 Announce Type: new Abstract: Instance normalization (IN) is widely used in non-stationary multivariate time series forecasting to reduce distribution shifts and highlight common pat

tutorialsarxiv-cs-lg
15 May 2026
Research

Separating Intrinsic Ambiguity from Estimation Uncertainty in Deep Generative Models for Linear Inverse Problems

DGX agent

arXiv:2605.15050v1 Announce Type: new Abstract: Recently, deep generative models have been used for posterior inference in inverse problems, including high-stakes applications in medical imaging and s

researcharxiv-cs-lg
15 May 2026
Model Releases

Thinking Ahead: Prospection-Guided Retrieval of Memory with Language Models

DGX agent

arXiv:2605.14177v1 Announce Type: cross Abstract: Long-horizon personalization requires dialogue assistants to retrieve user-specific facts from extended interaction histories. In practice, many relev

model-releasesarxiv-cs-ai
15 May 2026
Safety

Training ML Models with Predictable Failures

DGX agent

arXiv:2605.15134v1 Announce Type: new Abstract: Estimating how often an ML model will fail at deployment scale is central to pre-deployment safety assessment, but a feasible evaluation set is rarely l

safetyarxiv-cs-lg
15 May 2026
Research

Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacement

DGX agent

arXiv:2605.14368v1 Announce Type: cross Abstract: Continuous diffusion language models lag behind autoregressive transformers, partly because diffusion is applied in spaces poorly suited to language d

researcharxiv-cs-ai
15 May 2026
Safety

AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation

DGX agent

arXiv:2605.13724v1 Announce Type: cross Abstract: Few-step video generation has been significantly advanced by consistency distillation. However, the performance of consistency-distilled models often

safetyarxiv-cs-ai
14 May 2026
Research

Bridging the Missing-Modality Gap: Improving Text-Only Calibration of Vision Language Models

DGX agent

arXiv:2605.12517v1 Announce Type: cross Abstract: Vision-language models (VLMs) are often deployed on text-only inputs, although they are trained with images. We find that removing the vision modality

researcharxiv-cs-ai
14 May 2026
Research

Debunking Grad-ECLIP: A Comprehensive Study on Its Incorrectness and Fundamental Principles for Model Interpretation

DGX agent

arXiv:2605.12952v1 Announce Type: new Abstract: Grad-ECLIP is published at ICML 2024 and represents a new Transformer interpretation technical route (intermediate features-based). First, this paper de

researcharxiv-cs-cv
14 May 2026
Model Releases

Few-Shot Physics-Informed Neural Network for Shape Reconstruction of Concentric-Tube Robots

DGX agent

arXiv:2605.12790v1 Announce Type: new Abstract: Modeling concentric tube robots (CTRs) involves complex nonlinear continuum mechanics, and despite recent progress, physics-based models often lack an a

model-releasesarxiv-cs-ro
14 May 2026
Research

SubspaceAD: Training-Free Few-Shot Anomaly Detection via Subspace Modeling

DGX agent

arXiv:2602.23013v3 Announce Type: replace Abstract: Detecting visual anomalies in industrial inspection often requires training with only a few normal images per category. Recent few-shot methods achi

researcharxiv-cs-cv
14 May 2026
Research

Unifying Physically-Informed Weather Priors in A Single Model for Image Restoration Across Multiple Adverse Weather Conditions

DGX agent

arXiv:2605.13158v1 Announce Type: new Abstract: Image restoration under multiple adverse weather conditions aims to develop a single model to recover the underlying scene with high visibility. Weather

researcharxiv-cs-cv
14 May 2026
Research

A Comparative Study of Model Selection Criteria for Symbolic Regression

DGX agent

arXiv:2605.11233v1 Announce Type: new Abstract: Effective model selection is critical in symbolic regression (SR) to identify mathematical expressions that balance accuracy and complexity, and have lo

researcharxiv-cs-lg
13 May 2026
Research

Action Hallucination in Generative Vision-Language-Action Models

DGX agent

arXiv:2602.06339v2 Announce Type: replace Abstract: Robot Foundation Models, such as VLAs, promise end-to-end generative robot policies with broad generalization. Yet it remains unclear whether they f

researcharxiv-cs-ro
13 May 2026
Local Ai

ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models

DGX agent

arXiv:2605.11222v1 Announce Type: new Abstract: Quantization is an effective strategy to reduce the storage and computation footprint of large language models (LLMs). Post-training quantization (PTQ)

local-aiarxiv-cs-lg
13 May 2026
Safety

Couple to Control: Joint Initial Noise Design in Diffusion Models

DGX agent

arXiv:2605.11311v1 Announce Type: cross Abstract: Diffusion models typically generate image batches from independent Gaussian initial noises. We argue that this independence assumption is only one cho

safetyarxiv-cs-cv
13 May 2026
Research

Design Your Ad: Personalized Advertising Image and Text Generation with Unified Autoregressive Models

DGX agent

arXiv:2605.12138v1 Announce Type: cross Abstract: Generating realistic and user-preferred advertisements is a key challenge in e-commerce. Existing approaches utilize multiple independent models drive

researcharxiv-cs-cl
13 May 2026
Applications

ECHO: Continuous Hierarchical Memory for Vision-Language-Action Models

DGX agent

arXiv:2605.10993v1 Announce Type: new Abstract: Memory capacity is a critical factor determining the performance of Vision-Language-Action (VLA) models in long-horizon manipulation tasks. Existing mem

applicationsarxiv-cs-ro
13 May 2026
Research

Efficient Adjoint Matching for Fine-tuning Diffusion Models

DGX agent

arXiv:2605.11480v1 Announce Type: new Abstract: Reward fine-tuning has become a common approach for aligning pretrained diffusion and flow models with human preferences in text-to-image generation. Am

researcharxiv-cs-lg
13 May 2026
Tutorials

FERMI: Exploiting Relations for Membership Inference Against Tabular Diffusion Models

DGX agent

arXiv:2605.11527v1 Announce Type: new Abstract: Diffusion models are the leading approach for tabular data synthesis and are increasingly used to share sensitive records. Whether they actually protect

tutorialsarxiv-cs-lg
13 May 2026
Research

G^2TR: Generation-Guided Visual Token Reduction for Separate-Encoder Unified Multimodal Models

DGX agent

arXiv:2605.12309v1 Announce Type: new Abstract: The development of separate-encoder Unified multimodal models (UMMs) comes with a rapidly growing inference cost due to dense visual token processing. I

researcharxiv-cs-cv
13 May 2026
Research

HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation

DGX agent

arXiv:2605.11596v1 Announce Type: new Abstract: Closed-loop driving simulation requires real-time interaction beyond short offline clips, pushing current driving world models toward autoregressive (AR

researcharxiv-cs-cv
13 May 2026
Local Ai

Instruction Lens Score: Your Instruction Contributes a Powerful Object Hallucination Detector for Multimodal Large Language Models

DGX agent

arXiv:2605.12258v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have achieved remarkable progress, yet the object hallucination remains a critical challenge for reliable deplo

local-aiarxiv-cs-lg
13 May 2026
Research

Interpretable rainfall modelling reveals rapid reorganisation of Amazonian rainfall under vegetation loss

DGX agent

arXiv:2605.10948v1 Announce Type: cross Abstract: Understanding how vegetation loss alters rainfall remains a major challenge in climate and hydrological science, as deforestation modifies precipitati

researcharxiv-cs-lg
13 May 2026
Local Ai

Local and Mixing-Based Algorithms for Gaussian Graphical Model Selection from Glauber Dynamics

DGX agent

arXiv:2412.18594v3 Announce Type: replace Abstract: Gaussian graphical model selection is usually studied under independent sampling, but in many applications observations arise from dependent dynamic

local-aiarxiv-cs-lg
13 May 2026
Applications

Make It Long, Keep It Fast: End-to-End 10k-Sequence Modeling at Billion Scale on Douyin Recommendation

DGX agent

arXiv:2511.06077v2 Announce Type: replace Abstract: Short-video recommenders such as Douyin must exploit extremely long user histories without breaking latency or cost budgets. We present an end-to-en

applicationsarxiv-cs-lg
13 May 2026
Applications

Measuring Accuracy and Energy-to-Solution of Quantum Fine-Tuning of Foundational AI Models

DGX agent

arXiv:2605.02798v1 Announce Type: cross Abstract: We present an experimental study of energy-to-solution (ETS) of hybrid quantum-classical applications, enabled by direct instrumentation of power cons

applicationsarxiv-cs-lg
13 May 2026
Model Releases

Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference

DGX agent

arXiv:2510.05497v5 Announce Type: replace-cross Abstract: Large-scale Mixture of Experts (MoE) Large Language Models (LLMs) have recently become the frontier open-weight models, achieving remarkable m

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Procedural-skill SFT across capacity tiers: A W-Shaped pre-SFT Trajectory and Regime-Asymmetric Mechanism on 0.8B-4B Qwen3.5 Models

DGX agent

arXiv:2605.11907v1 Announce Type: new Abstract: We measure procedural-skill SFT contribution across three Qwen3.5 dense scales (0.8B, 2B, 4B) on a 200-task / 40-skill holdout, with Claude Haiku 4.5 as

model-releasesarxiv-cs-lg
13 May 2026
Safety

Question Difficulty Estimation for Large Language Models via Answer Plausibility Scoring

DGX agent

arXiv:2605.12398v1 Announce Type: new Abstract: Estimating question difficulty is a critical component in evaluating and improving large language models (LLMs) for question answering (QA). Existing ap

safetyarxiv-cs-cl
13 May 2026
Applications

Reconstruction of Personally Identifiable Information from Supervised Finetuned Models

DGX agent

arXiv:2605.12264v1 Announce Type: cross Abstract: Supervised Finetuning (SFT) has become one of the primary methods for adapting a large language model (LLM) with extensive pre-trained knowledge to do

applicationsarxiv-cs-cl
13 May 2026
Safety

The DAWN of World-Action Interactive Models

DGX agent

arXiv:2605.11550v1 Announce Type: new Abstract: A plausible scene evolution depends on the maneuver being considered, while a good maneuver depends on how the scene may evolve. Existing World Action M

safetyarxiv-cs-cv
13 May 2026
Applications

A Breast Vision Pathology Foundation Model for Real-world Clinical Utility

DGX agent

arXiv:2605.08207v1 Announce Type: new Abstract: Pathology foundation models have shown strong retrospective performance, but whether such systems can support clinically relevant use remains unclear. T

applicationsarxiv-cs-cv
12 May 2026
Research

A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models

DGX agent

arXiv:2605.08504v1 Announce Type: new Abstract: We investigate the origins of massive activations in large language models (LLMs) and identify a specific layer named the extbf{Massive Emergence Layer

researcharxiv-cs-cl
12 May 2026
← Previous
1…188189190191192…1030
Next →