AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

Towards Apples to Apples for AI Evaluations: From Real-World Use Cases to Evaluation Scenarios

DGX agent

arXiv:2605.07986v1 Announce Type: cross Abstract: AI measurement science has a wide variety of methodologies and measurements for comparing AI systems, resulting in what often appear to be 'apples-to-

applicationsarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Towards Autonomous Business Intelligence via Data-to-Insight Discovery Agent

DGX agent

arXiv:2605.07202v1 Announce Type: new Abstract: Transforming fragmented enterprise data into actionable insights remains a significant challenge for LLMs, constrained by complex database schemas, limi

agentsarxiv-cs-ai
11 May 2026
Hardware

Towards Billion-scale Multi-modal Biometric Search

DGX agent

arXiv:2605.07655v1 Announce Type: cross Abstract: Searching a multi-biometric database of a billion records for a country-level identity system requires pushing the limits of all aspects of a biometri

hardwarearxiv-cs-ai
11 May 2026
Safety

Towards Differentially Private Reinforcement Learning with General Function Approximation

DGX agent

arXiv:2605.07049v1 Announce Type: cross Abstract: We present the first theoretical guarantees for differentially private online reinforcement learning (RL) with general function approximation, extendi

safetyarxiv-cs-ai
11 May 2026
Agents

Towards Security-Auditable LLM Agents: A Unified Graph Representation

DGX agent

arXiv:2605.06812v1 Announce Type: new Abstract: LLM-based agentic systems are rapidly evolving to perform complex autonomous tasks through dynamic tool invocation, stateful memory management, and mult

agentsarxiv-cs-ai
11 May 2026
Research

TRACE: Tourism Recommendation with Accountable Citation Evidence

DGX agent

arXiv:2605.07677v1 Announce Type: cross Abstract: Tourism is a high-stakes setting for conversational recommender systems (CRS): a plausible-sounding suggestion can waste real money and trip time once

researcharxiv-cs-ai
11 May 2026
Agents

TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples

DGX agent

arXiv:2605.07935v1 Announce Type: new Abstract: We present TraceFix, a verification-first pipeline for Large Language Model (LLM) multi-agent coordination. An agent synthesizes a protocol topology as

agentsarxiv-cs-ai
11 May 2026
Model Releases

Tracing Uncertainty in Language Model 'Reasoning'

DGX agent

arXiv:2605.07776v1 Announce Type: cross Abstract: Language model (LM) 'reasoning', commonly described as Chain-of-Thought or test-time scaling, often improves benchmark performance, but the dynamics u

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Tracking Large-scale Shared Bikes with Inertial Motion Learning in GNSS Blocked Environments

DGX agent

arXiv:2605.07412v1 Announce Type: cross Abstract: Although Global Navigation Satellite Systems (GNSS) provide a general solution for bike tracking outdoors, there still exist complex riding environmen

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation

DGX agent

arXiv:2605.07924v1 Announce Type: cross Abstract: Discrete flow matching generates text by iteratively transforming noise tokens into coherent language, but may require hundreds of forward passes. Dis

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TSRBench: A Comprehensive Multi-task Multi-modal Time Series Reasoning Benchmark for Generalist Models

DGX agent

arXiv:2601.18744v2 Announce Type: replace Abstract: Time series are ubiquitous in real-world scenarios and crucial for applications ranging from energy management to traffic control. Consequently, the

model-releasesarxiv-cs-ai
11 May 2026
Research

TTF: Temporal Token Fusion for Efficient Video-Language Model

DGX agent

arXiv:2605.07355v1 Announce Type: cross Abstract: Video-language models (VLMs) face rapid inference costs as visual token counts scale with video length. For example, 32 frames at 448{imes}448 resolut

researcharxiv-cs-ai
11 May 2026
Applications

Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors

DGX agent

arXiv:2505.11325v4 Announce Type: replace-cross Abstract: Prior-data fitted networks (PFNs) have emerged as promising foundation models for prediction from tabular datasets, achieving state-of-the-art

applicationsarxiv-cs-ai
11 May 2026
Model Releases

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios

DGX agent

arXiv:2410.06355v3 Announce Type: replace-cross Abstract: This paper presents UNCOM, a novel hybrid framework for interpreting natural human commands in tabletop scenarios. The system integrates multi

model-releasesarxiv-cs-ai
11 May 2026
Research

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

DGX agent

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

researcharxiv-cs-ai
11 May 2026
Model Releases

Uneven Evolution of Cognition Across Generations of Generative AI Models

DGX agent

arXiv:2605.06815v1 Announce Type: new Abstract: The pursuit of artificial general intelligence necessitates robust methods for evaluating the cognitive capabilities of models beyond narrow task perfor

model-releasesarxiv-cs-ai
11 May 2026
Local Ai

Unlocking High-Fidelity Molecular Generation from Mass Spectra via Dual-Stream Line Graph Diffusion

DGX agent

arXiv:2605.07048v1 Announce Type: cross Abstract: De novo molecular generation from tandem mass spectra is a challenging inverse problem whose core difficulty lies in the circular dependency between a

local-aiarxiv-cs-ai
11 May 2026
Model Releases

Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts

DGX agent

arXiv:2605.07395v1 Announce Type: cross Abstract: Efficient routing across multiple LLMs enables cost-quality tradeoffs by directing queries to the cheapest capable model. Prior work attributes routin

model-releasesarxiv-cs-ai
11 May 2026
Applications

Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs

DGX agent

arXiv:2605.07481v1 Announce Type: cross Abstract: In this paper, we investigate the recent state-of-the-art schemes for watermarking large language models (LLMs) outputs. These techniques are claimed

applicationsarxiv-cs-ai
11 May 2026
Agents

VDCook:DIY video data cook your MLLMs

DGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

agentsarxiv-cs-ai
11 May 2026
Research

VecCISC: Improving Confidence-Informed Self-Consistency with Reasoning Trace Clustering and Candidate Answer Selection

DGX agent

arXiv:2605.08070v1 Announce Type: new Abstract: A standard technique for scaling inference-time reasoning is Self-Consistency, whereby multiple candidate answers are sampled from an LLM and the most c

researcharxiv-cs-ai
11 May 2026
Safety

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

DGX agent

arXiv:2602.10693v3 Announce Type: replace-cross Abstract: Off-policy updates are inevitable in reinforcement learning (RL) for large language models (LLMs) due to rollout staleness from asynchronous t

safetyarxiv-cs-ai
11 May 2026
Applications

Vibe coding before the trend

DGX agent

arXiv:2605.07751v1 Announce Type: cross Abstract: Early 2025 we ran a series of vibe coding challenges across four different student cohorts. The cohorts included 54 ICT students, 24 digital marketing

applicationsarxiv-cs-ai
11 May 2026
Agents

VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents

DGX agent

arXiv:2506.21582v5 Announce Type: replace-cross Abstract: Text analytics has traditionally required specialized knowledge in Natural Language Processing (NLP) or text analysis, which presents a barrie

agentsarxiv-cs-ai
11 May 2026
Model Releases

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models

DGX agent

arXiv:2605.07872v1 Announce Type: cross Abstract: Multimodal reward models have advanced substantially in text and image domains, yet progress in video understanding reward modeling remains severely l

model-releasesarxiv-cs-ai
11 May 2026
Safety

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

DGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

safetyarxiv-cs-ai
11 May 2026
Safety

VISD: Enhancing Video Reasoning via Structured Self-Distillation

DGX agent

arXiv:2605.06094v2 Announce Type: replace-cross Abstract: Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignme

safetyarxiv-cs-ai
11 May 2026
Model Releases

Visual Text Compression as Measure Transport

DGX agent

arXiv:2605.06708v1 Announce Type: cross Abstract: Visual text compression (VTC) promises efficient long-context processing by rendering text into an image and re-encoding it with a vision-language mod

model-releasesarxiv-cs-ai
11 May 2026
Research

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

DGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

researcharxiv-cs-ai
11 May 2026
Agents

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

DGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

agentsarxiv-cs-ai
11 May 2026
Applications

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

DGX agent

arXiv:2605.06761v1 Announce Type: new Abstract: The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection at

applicationsarxiv-cs-ai
11 May 2026
Applications

What if AI systems weren't chatbots?

DGX agent

arXiv:2605.07896v1 Announce Type: cross Abstract: The rapid convergence of artificial intelligence (AI) toward conversational chatbot interfaces marks a critical moment for the industry. This paper ar

applicationsarxiv-cs-ai
11 May 2026
Research

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

DGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

researcharxiv-cs-ai
11 May 2026
Model Releases

When Does Critique Improve AI-Assisted Theoretical Physics? SCALAR: Structured Critic--Actor Loop for Agentic Reasoning

DGX agent

arXiv:2605.06772v1 Announce Type: new Abstract: As large language models (LLMs) show increasing promise on research-level physics reasoning tasks and agentic AI becomes more common, a practical questi

model-releasesarxiv-cs-ai
11 May 2026
Research

When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining

DGX agent

arXiv:2605.07756v1 Announce Type: cross Abstract: Modern deep models are often pretrained on large-scale data with missing labels using composite objectives, where the relative weights of multiple los

researcharxiv-cs-ai
11 May 2026
Agents

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

DGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

agentsarxiv-cs-ai
11 May 2026
Model Releases

Where's the Plan? Locating Latent Planning in Language Models with Lightweight Mechanistic Interventions

DGX agent

arXiv:2605.07984v1 Announce Type: cross Abstract: We study planning site formation in language models -- where internal representations of structurally-constrained future tokens form during the forwar

model-releasesarxiv-cs-ai
11 May 2026
Safety

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

DGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

safetyarxiv-cs-ai
11 May 2026
Tutorials

Why DDIM Hallucinates More than DDPM: A Theoretical Analysis of Reverse Dynamics

DGX agent

arXiv:2605.06831v1 Announce Type: cross Abstract: We theoretically study the hallucination phenomena in two canonical diffusion samplers: the stochastic Denoising Diffusion Probabilistic Model (DDPM)

tutorialsarxiv-cs-ai
11 May 2026
Model Releases

Why Self-Inconsistency Arises in GNN Explanations and How to Exploit It

DGX agent

arXiv:2605.07527v1 Announce Type: cross Abstract: Recent work has observed that explanations produced by Self-Interpretable Graph Neural Networks (SI-GNNs) can be self-inconsistent: when the model is

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems

DGX agent

arXiv:2605.07068v1 Announce Type: cross Abstract: The LLM Wiki pattern, to compile and provide domain knowledge into a persistent artifact and serve it to LLMs via KV cache inference, promises context

model-releasesarxiv-cs-ai
11 May 2026
Hardware

XiYOLO: Energy-Aware Object Detection via Iterative Architecture Search and Scaling

DGX agent

arXiv:2605.06927v1 Announce Type: cross Abstract: Object detection on heterogeneous edge devices must satisfy strict energy, latency, and memory constraints while still providing reliable perception f

hardwarearxiv-cs-ai
11 May 2026
Model Releases

Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

DGX agent

arXiv:2605.07579v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) for Large Reasoning Models hinges on baseline estimation for variance reduction, but existing ap

model-releasesarxiv-cs-ai
11 May 2026
Research

A Fast Model Counting Algorithm for Two-Variable Logic with Counting and Modulo Counting Quantifiers

DGX agent

arXiv:2605.03391v1 Announce Type: cross Abstract: Weighted first-order model counting (WFOMC) is a central task in lifted probabilistic inference: It asks for the weighted sum of all models of a first

researcharxiv-cs-ai
7 May 2026
Safety

A Skill-Based AI Agentic Pipeline for Library of Congress Subject Indexing

DGX agent

arXiv:2605.03537v1 Announce Type: cross Abstract: This paper presents a modular AI agentic skill pipeline for automating subject indexing with Library of Congress Subject Headings (LCSH). Subject inde

safetyarxiv-cs-ai
7 May 2026
Research

A Universal Space of Brain Dynamics for Unveiling Cognitive Transitions and Individual Differences

DGX agent

arXiv:2605.02936v1 Announce Type: cross Abstract: Representing dynamical systems through data-driven universal spaces has proven effective; however, achieving this universality for human brain activit

researcharxiv-cs-ai
7 May 2026
Local Ai

A Workflow-Oriented Framework for Asynchronous Human-AI Collaboration in Hybrid and Compute-Intensive HPC Environments

DGX agent

arXiv:2605.03743v1 Announce Type: cross Abstract: Human involvement is critical in training and deploying AI systems in high-stakes defence and security contexts. However, real-time interaction is imp

local-aiarxiv-cs-ai
7 May 2026
Research

AdapShot: Adaptive Many-Shot In-Context Learning with Semantic-Aware KV Cache Reuse

DGX agent

arXiv:2605.03644v1 Announce Type: new Abstract: Many-Shot In-Context Learning (ICL) has emerged as a promising paradigm, leveraging extensive examples to unlock the reasoning potential of Large Langua

researcharxiv-cs-ai
7 May 2026
← Previous
1…355356357358359…448
Next →