AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,935 results
Model Releases

Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination

DGX agent

arXiv:2605.19250v1 Announce Type: new Abstract: Modality-conflict hallucination occurs when multimodal large language models (MLLMs) prioritize erroneous textual premises over contradictory visual evi

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

DGX agent

arXiv:2605.18808v1 Announce Type: cross Abstract: We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

DGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Depth2Pose: A Pose-Based Benchmark for Monocular Depth Estimation without Ground-Truth Depth

DGX agent

arXiv:2605.19797v1 Announce Type: new Abstract: Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted d

model-releasesarxiv-cs-cv
20 May 2026
Local Ai

Descriptive versus Regulatory Uncertainty in Bounded Predictive Systems

DGX agent

arXiv:2605.18909v1 Announce Type: new Abstract: Any system that models the world under finite representational capacity must compress; any compression entails a prior; and the prior is the system's bi

local-aiarxiv-cs-lg
20 May 2026
Model Releases

Differential-Integral Neural Operator for Long-Term Turbulence Forecasting

DGX agent

arXiv:2509.21196v3 Announce Type: replace-cross Abstract: Accurately forecasting the long-term evolution of turbulence represents a grand challenge in scientific computing and is crucial for applicati

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

DMN: A Compositional Framework for Jailbreaking Multimodal LLMs with Multi-Image Inputs

DGX agent

arXiv:2605.18915v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are vulnerable to jailbreak attacks, which can elicit harmful responses from MLLMs. Many MLLMs support multi-

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

DocQT: Improving Document Forgery Localization Robustness via Diverse JPEG Quantization Tables

DGX agent

arXiv:2605.19688v1 Announce Type: new Abstract: Document manipulation localization models achieve strong performance on public benchmarks yet fail to generalize to operational document workflows. We i

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

EventPrune: Cascaded Event-Assisted Token Pruning for Efficient First-Person Dynamic Spatial Reasoning

DGX agent

arXiv:2605.19506v1 Announce Type: new Abstract: First-person dynamic spatial reasoning requires models to track continuous motion and precise geometric structure, but the quadratic attention cost of T

model-releasesarxiv-cs-cv
20 May 2026
Safety

Feature-Space Smoothing: Certified Robustness of Deep Representations

DGX agent

arXiv:2601.16200v3 Announce Type: replace-cross Abstract: Modern deep learning models exhibit strong capabilities across diverse applications, yet remain vulnerable to malicious inputs that induce err

safetyarxiv-cs-cv
20 May 2026
Safety

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

DGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

safetyarxiv-cs-ai
20 May 2026
Model Releases

Generalization Bounds of Surrogate Policies for Combinatorial Optimization Problems

DGX agent

arXiv:2407.17200v3 Announce Type: replace-cross Abstract: Many real-world decision problems require solving, again and again, combinatorial optimization instances drawn from a common distribution. A r

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

GoTTA be Diverse: Rethinking Memory Policies for Test-Time Adaptation

DGX agent

arXiv:2605.19890v1 Announce Type: new Abstract: Test-time adaptation (TTA) enables a pre-trained model to adapt online to an unlabeled test stream under distribution shift. While most TTA research foc

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding

DGX agent

arXiv:2605.19223v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) exhibit strong performance on standard video tasks, their ability to faithfully summarize and reason over

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

DGX agent

arXiv:2605.18822v1 Announce Type: cross Abstract: Post-training has become essential for adapting large language models (LLMs) to complex downstream behaviors, including instruction following, prefere

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Less Back-and-Forth: A Comparative Study of Structured Prompting

DGX agent

arXiv:2605.20149v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for open-ended tasks, but underspecified prompts can lead to low-quality answers and additional interacti

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Library Hallucinations in LLM-Generated Code: A Risk Analysis Grounded in Developer Queries

DGX agent

arXiv:2509.22202v3 Announce Type: replace-cross Abstract: Large language models (LLMs) now play a central role in code generation, yet they continue to hallucinate, frequently inventing non-existent l

model-releasesarxiv-cs-cl
20 May 2026
Research

Lossless Anti-Distillation Sampling

DGX agent

arXiv:2605.18829v1 Announce Type: new Abstract: Frontier commercial generative models face a growing threat from distillation, whereby a distiller harvests generated responses and trains a competing m

researcharxiv-cs-lg
20 May 2026
Model Releases

MIRO: MultI-Reward cOnditioned pretraining improves T2I quality and efficiency

DGX agent

arXiv:2510.25897v2 Announce Type: replace Abstract: The default paradigm of post-training text-to-image generators includes post-hoc selection of generated images, and subsequent training with one rew

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs

DGX agent

arXiv:2511.14159v2 Announce Type: replace Abstract: Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-wo

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

DGX agent

arXiv:2605.19354v1 Announce Type: cross Abstract: MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes mo

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Model Releases

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

DGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

RECIPE: Procedural Planning via Grounding in Instructional Video

DGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Retrieval-Augmented Generation for Natural Language Processing: A Survey

DGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Smooth Piecewise Cutting for Neural Operator to Handle Discontinuities and Sharp Transitions

DGX agent

arXiv:2605.19823v1 Announce Type: cross Abstract: Neural operators have achieved strong performance in learning solution operators of partial differential equations (PDEs), but their inherently contin

model-releasesarxiv-cs-ai
20 May 2026
Applications

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

DGX agent

arXiv:2505.23747v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced performance on 2D visual tasks. However, improving

applicationsarxiv-cs-ai
20 May 2026
Model Releases

Tail Annealing for Heavy-Tailed Flow Matching

DGX agent

arXiv:2605.20068v1 Announce Type: cross Abstract: Standard generative models struggle with heavy-tailed data: Lipschitz architectures cannot produce power-law tails from Gaussian noise, and interpolat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

DGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

model-releasesarxiv-cs-cl
20 May 2026
Safety

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

DGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

safetyarxiv-cs-lg
20 May 2026
Model Releases

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

DGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

DGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

DGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Do Evolutionary Coding Agents Evolve?

DGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

DGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Feature-Driven Framework for Software Fault Prediction

DGX agent

arXiv:2605.17611v1 Announce Type: cross Abstract: Software fault prediction (SFP) is a critical task in software engineering, enabling early identification of faults in modules to improve software qua

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

A Machine With Human-Like Memory Systems

DGX agent

arXiv:2204.01611v3 Announce Type: replace Abstract: Inspired by the cognitive science theory, we explicitly model an agent with both semantic and episodic memory systems, and show that it is better th

model-releasesarxiv-cs-ai
19 May 2026
Research

A Theory of Training Profit-Optimal LLMs

DGX agent

arXiv:2605.16430v1 Announce Type: cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure

researcharxiv-cs-ai
19 May 2026
Local Ai

AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training

DGX agent

arXiv:2605.17923v1 Announce Type: cross Abstract: In video generation models, particularly world models, training large-scale video diffusion Transformers (such as DiT and MMDiT) poses significant com

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory

DGX agent

arXiv:2605.18733v1 Announce Type: new Abstract: Autoregressive video generation has improved rapidly in visual fidelity and interactivity, but it still suffers from long-term inconsistency and memory

model-releasesarxiv-cs-cv
19 May 2026
Agents

AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent

DGX agent

arXiv:2602.03955v2 Announce Type: replace Abstract: While large language model (LLM) multi-agent systems achieve superior reasoning performance through iterative debate, practical deployment is limite

agentsarxiv-cs-ai
19 May 2026
Model Releases

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

DGX agent

arXiv:2605.17583v1 Announce Type: new Abstract: While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

AgentWall: A Runtime Safety Layer for Local AI Agents

DGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing

DGX agent

arXiv:2603.23069v2 Announce Type: replace-cross Abstract: The task of authorship style transfer involves rewriting text in the style of a target author while preserving the meaning of the original tex

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity

DGX agent

arXiv:2510.00304v3 Announce Type: replace-cross Abstract: Deep learning models excel in stationary data but struggle in non-stationary environments due to a phenomenon known as loss of plasticity (LoP

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs

DGX agent

arXiv:2602.09805v2 Announce Type: replace-cross Abstract: As reasoning LLMs increasingly trade tokens for accuracy through deliberation, search, and self-correction, a single accuracy score can no lon

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…427428429430431…1082
Next →