AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
53,690 results
Model Releases

MVI-Bench: A Comprehensive Benchmark for Evaluating Robustness to Misleading Visual Inputs in LVLMs

DGX agent

arXiv:2511.14159v2 Announce Type: replace Abstract: Evaluating the robustness of Large Vision-Language Models (LVLMs) is essential for their continued development and responsible deployment in real-wo

model-releasesarxiv-cs-cv
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Next-Acceleration-Scale Prediction for Autoregressive MRI Reconstruction

DGX agent

arXiv:2605.19354v1 Announce Type: cross Abstract: MRI reconstruction is an inherently ill-posed inverse problem, since incomplete measurements admit many plausible solutions. This ambiguity becomes mo

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Model Releases

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling

DGX agent

arXiv:2605.20052v1 Announce Type: cross Abstract: Automatic report labeling facilitates the identification of clinical findings from unstructured text and enables large-scale annotation for medical im

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

RECIPE: Procedural Planning via Grounding in Instructional Video

DGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Retrieval-Augmented Generation for Natural Language Processing: A Survey

DGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Smooth Piecewise Cutting for Neural Operator to Handle Discontinuities and Sharp Transitions

DGX agent

arXiv:2605.19823v1 Announce Type: cross Abstract: Neural operators have achieved strong performance in learning solution operators of partial differential equations (PDEs), but their inherently contin

model-releasesarxiv-cs-ai
20 May 2026
Applications

Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

DGX agent

arXiv:2505.23747v2 Announce Type: replace-cross Abstract: Recent advancements in Multimodal Large Language Models (MLLMs) have significantly enhanced performance on 2D visual tasks. However, improving

applicationsarxiv-cs-ai
20 May 2026
Model Releases

Tail Annealing for Heavy-Tailed Flow Matching

DGX agent

arXiv:2605.20068v1 Announce Type: cross Abstract: Standard generative models struggle with heavy-tailed data: Lipschitz architectures cannot produce power-law tails from Gaussian noise, and interpolat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Taming the Thinker: Conditional Entropy Shaping for Adaptive LLM Reasoning

DGX agent

arXiv:2605.19358v1 Announce Type: new Abstract: Entropy-based deep reasoning has emerged as a promising direction for improving the reasoning capabilities of Large Language Models (LLMs), but existing

model-releasesarxiv-cs-cl
20 May 2026
Safety

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

DGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

safetyarxiv-cs-lg
20 May 2026
Model Releases

The Annotation Scarcity Paradox in Low-Resource NLP Evaluation: A Decade of Acceleration and Emerging Constraints

DGX agent

arXiv:2605.19066v1 Announce Type: new Abstract: Over the past decade, low-resource natural language processing (NLP) has experienced explosive growth, propelled by cross-lingual transfer, massively mu

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

DGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

DGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Do Evolutionary Coding Agents Evolve?

DGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

DGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Feature-Driven Framework for Software Fault Prediction

DGX agent

arXiv:2605.17611v1 Announce Type: cross Abstract: Software fault prediction (SFP) is a critical task in software engineering, enabling early identification of faults in modules to improve software qua

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

A Machine With Human-Like Memory Systems

DGX agent

arXiv:2204.01611v3 Announce Type: replace Abstract: Inspired by the cognitive science theory, we explicitly model an agent with both semantic and episodic memory systems, and show that it is better th

model-releasesarxiv-cs-ai
19 May 2026
Research

A Theory of Training Profit-Optimal LLMs

DGX agent

arXiv:2605.16430v1 Announce Type: cross Abstract: Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure

researcharxiv-cs-ai
19 May 2026
Local Ai

AdaptiveLoad: Towards Efficient Video Diffusion Transformer Training

DGX agent

arXiv:2605.17923v1 Announce Type: cross Abstract: In video generation models, particularly world models, training large-scale video diffusion Transformers (such as DiT and MMDiT) poses significant com

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Advancing Narrative Long Video Generation via Training-Free Identity-Aware Memory

DGX agent

arXiv:2605.18733v1 Announce Type: new Abstract: Autoregressive video generation has improved rapidly in visual fidelity and interactivity, but it still suffers from long-term inconsistency and memory

model-releasesarxiv-cs-cv
19 May 2026
Agents

AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent

DGX agent

arXiv:2602.03955v2 Announce Type: replace Abstract: While large language model (LLM) multi-agent systems achieve superior reasoning performance through iterative debate, practical deployment is limite

agentsarxiv-cs-ai
19 May 2026
Model Releases

AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech

DGX agent

arXiv:2605.17583v1 Announce Type: new Abstract: While existing text-to-speech (TTS) models exhibit high expressiveness, fine-grained control over composite instructions remains challenging due to the

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

AgentWall: A Runtime Safety Layer for Local AI Agents

DGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

AuthorMix: Modular Authorship Style Transfer via Layer-wise Adapter Mixing

DGX agent

arXiv:2603.23069v2 Announce Type: replace-cross Abstract: The task of authorship style transfer involves rewriting text in the style of a target author while preserving the meaning of the original tex

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Barriers for Learning in an Evolving World: Mathematical Understanding of Loss of Plasticity

DGX agent

arXiv:2510.00304v3 Announce Type: replace-cross Abstract: Deep learning models excel in stationary data but struggle in non-stationary environments due to a phenomenon known as loss of plasticity (LoP

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs

DGX agent

arXiv:2602.09805v2 Announce Type: replace-cross Abstract: As reasoning LLMs increasingly trade tokens for accuracy through deliberation, search, and self-correction, a single accuracy score can no lon

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

BioProAgent: Neuro-Symbolic Grounding for Constrained Scientific Planning

DGX agent

arXiv:2603.00876v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated significant reasoning capabilities in scientific discovery but struggle to bridge the gap to physical

model-releasesarxiv-cs-ai
19 May 2026
Research

CAB: Accelerating Flow and Diffusion Sampling via Rectification and Corrected Adams-Bashforth

DGX agent

arXiv:2605.16736v1 Announce Type: new Abstract: Flow and diffusion models achieve high-fidelity, high-resolution image synthesis, but often require many function evaluations (NFEs) at sampling time. E

researcharxiv-cs-cv
19 May 2026
Model Releases

CAREBench: Evaluating LLMs' Emotion Understanding by Assessing Cognitive Appraisal Reasoning

DGX agent

arXiv:2605.17176v1 Announce Type: new Abstract: Emotion understanding is a core capability for LLMs to interact effectively with humans, yet existing evaluation paradigms rely on discrete emotion labe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ClawArena: Benchmarking AI Agents in Evolving Information Environments

DGX agent

arXiv:2604.04202v2 Announce Type: replace-cross Abstract: AI agents deployed as persistent assistants must maintain correct beliefs as their information environment evolves. In practice, evidence is s

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selection

DGX agent

arXiv:2605.16839v1 Announce Type: new Abstract: Chunked prefill has become a widely adopted serving strategy for long-context large language models, but efficient attention computation in this regime

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

ContraFix: Agentic Vulnerability Repair via Differential Runtime Evidence and Skill Reuse

DGX agent

arXiv:2605.17450v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly used for automated vulnerability repair (AVR), where repository-level reasoning enables them to ins

model-releasesarxiv-cs-ai
19 May 2026
Agents

DECODE: Domain-aware Continual Domain Expansion for Motion Prediction

DGX agent

arXiv:2411.17917v2 Announce Type: replace Abstract: Motion prediction is critical for autonomous vehicles to effectively navigate complex environments and accurately anticipate the behaviors of other

agentsarxiv-cs-cv
19 May 2026
Local Ai

Diagnosing Korean-Language LLM Political Bias via Census-Grounded Agent Simulation

DGX agent

arXiv:2605.18395v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit systematic political biases in voter simulations, but their underlying mechanisms and cross-lingual generalizatio

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations

DGX agent

arXiv:2605.17107v1 Announce Type: cross Abstract: We introduce a novel framework for uncertainty quantification of solution operators associated with stochastic partial differential equations (SPDEs).

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

DisasterVQA: A Visual Question Answering Benchmark Dataset for Disaster Scenes

DGX agent

arXiv:2601.13839v2 Announce Type: replace Abstract: Social media imagery provides a low-latency source of situational information during natural and human-induced disasters, enabling rapid damage asse

model-releasesarxiv-cs-cv
19 May 2026
Safety

DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning

DGX agent

arXiv:2405.19189v3 Announce Type: replace Abstract: With the great success of diffusion models (DMs) in generating realistic synthetic vision data, many researchers have investigated their potential i

safetyarxiv-cs-lg
19 May 2026
Model Releases

Evaluating Cognitive Age Alignment in Interactive AI Agents

DGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

DGX agent

arXiv:2511.20857v2 Announce Type: replace-cross Abstract: Statefulness is essential for large language model (LLM) agents to perform long-term planning and problem-solving. This makes memory a critica

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

DGX agent

arXiv:2605.18421v1 Announce Type: cross Abstract: Recent benchmarks for Large Language Model (LLM) agents mainly evaluate reasoning, planning, and execution. However, memory is also essential for agen

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Experimentally validated quantum-secure federated learning over a multi-user quantum network

DGX agent

arXiv:2501.12709v2 Announce Type: replace-cross Abstract: Federated learning enables decentralized, privacy-preserving training but remains vulnerable to privacy leakage in the quantum era. Quantum fe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

exttt{SynC}: Synergistic Boosting of Structure and Representation for Deep Graph Clustering

DGX agent

arXiv:2406.15797v2 Announce Type: replace-cross Abstract: Employing graph neural networks (GNNs) for graph clustering has shown promising results in deep graph clustering. However, existing methods di

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Fine-grained List-wise Alignment for Generative Medication Recommendation

DGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

model-releasesarxiv-cs-lg
19 May 2026
Research

Flowing with Confidence

DGX agent

arXiv:2605.18472v1 Announce Type: cross Abstract: Generative models can produce nonsensical text, unrealistic images, and unstable materials faster than simulation or human review can absorb; without

researcharxiv-cs-ai
19 May 2026
Applications

From Documents to Segments: A Contextual Reformulation for Topic Assignment

DGX agent

arXiv:2605.17714v1 Announce Type: new Abstract: Traditional topic modeling assigns a single topic to each document. In practice, however, many real-world documents, such as product reviews or open-end

applicationsarxiv-cs-cl
19 May 2026
← Previous
1…443444445446447…1119
Next →