AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
11 May 2026

Uncertainty Quantification for Prior-Data Fitted Networks using Martingale Posteriors

ApplicationsDGX agent

arXiv:2505.11325v4 Announce Type: replace-cross Abstract: Prior-data fitted networks (PFNs) have emerged as promising foundation models for prediction from tabular datasets, achieving state-of-the-art

UNCOM: Zero-shot Context-Aware Command Understanding for Tabletop Scenarios

Model ReleasesDGX agent

arXiv:2410.06355v3 Announce Type: replace-cross Abstract: This paper presents UNCOM, a novel hybrid framework for interpreting natural human commands in tabletop scenarios. The system integrates multi

Understanding Performance Collapse in Layer-Pruned Large Language Models via Decision Representation Transitions

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.07271v1 Announce Type: cross Abstract: Layer pruning efficiently reduces Large Language Model (LLM) computational costs but often triggers sudden performance collapse. Existing representati

Uneven Evolution of Cognition Across Generations of Generative AI Models

Model ReleasesDGX agent

arXiv:2605.06815v1 Announce Type: new Abstract: The pursuit of artificial general intelligence necessitates robust methods for evaluating the cognitive capabilities of models beyond narrow task perfor

Unlocking High-Fidelity Molecular Generation from Mass Spectra via Dual-Stream Line Graph Diffusion

Local AiDGX agent

arXiv:2605.07048v1 Announce Type: cross Abstract: De novo molecular generation from tandem mass spectra is a challenging inverse problem whose core difficulty lies in the circular dependency between a

Unsolvability Ceiling in Multi-LLM Routing: An Empirical Study of Evaluation Artifacts

Model ReleasesDGX agent

arXiv:2605.07395v1 Announce Type: cross Abstract: Efficient routing across multiple LLMs enables cost-quality tradeoffs by directing queries to the cheapest capable model. Prior work attributes routin

Vaporizer: Breaking Watermarking Schemes for Large Language Model Outputs

ApplicationsDGX agent

arXiv:2605.07481v1 Announce Type: cross Abstract: In this paper, we investigate the recent state-of-the-art schemes for watermarking large language models (LLMs) outputs. These techniques are claimed

VDCook:DIY video data cook your MLLMs

AgentsDGX agent

arXiv:2603.05539v2 Announce Type: replace-cross Abstract: We introduce VDCook: a self-evolving video data operating system, a configurable video data construction platform for researchers and vertical

VecCISC: Improving Confidence-Informed Self-Consistency with Reasoning Trace Clustering and Candidate Answer Selection

ResearchDGX agent

arXiv:2605.08070v1 Announce Type: new Abstract: A standard technique for scaling inference-time reasoning is Self-Consistency, whereby multiple candidate answers are sampled from an LLM and the most c

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

SafetyDGX agent

arXiv:2602.10693v3 Announce Type: replace-cross Abstract: Off-policy updates are inevitable in reinforcement learning (RL) for large language models (LLMs) due to rollout staleness from asynchronous t

Vibe coding before the trend

ApplicationsDGX agent

arXiv:2605.07751v1 Announce Type: cross Abstract: Early 2025 we ran a series of vibe coding challenges across four different student cohorts. The cohorts included 54 ICT students, 24 digital marketing

VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents

AgentsDGX agent

arXiv:2506.21582v5 Announce Type: replace-cross Abstract: Text analytics has traditionally required specialized knowledge in Natural Language Processing (NLP) or text analysis, which presents a barrie

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models

Model ReleasesDGX agent

arXiv:2605.07872v1 Announce Type: cross Abstract: Multimodal reward models have advanced substantially in text and image domains, yet progress in video understanding reward modeling remains severely l

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

SafetyDGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

VISD: Enhancing Video Reasoning via Structured Self-Distillation

SafetyDGX agent

arXiv:2605.06094v2 Announce Type: replace-cross Abstract: Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignme

Visual Text Compression as Measure Transport

Model ReleasesDGX agent

arXiv:2605.06708v1 Announce Type: cross Abstract: Visual text compression (VTC) promises efficient long-context processing by rendering text into an image and re-encoding it with a vision-language mod

VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing

ResearchDGX agent

arXiv:2605.06765v1 Announce Type: cross Abstract: Human speech conveys expressiveness beyond linguistic content, including personality, mood, or performance elements, such as a comforting tone or humm

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

AgentsDGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

ApplicationsDGX agent

arXiv:2605.06761v1 Announce Type: new Abstract: The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection at

What if AI systems weren't chatbots?

ApplicationsDGX agent

arXiv:2605.07896v1 Announce Type: cross Abstract: The rapid convergence of artificial intelligence (AI) toward conversational chatbot interfaces marks a critical moment for the industry. This paper ar

When Does a Language Model Commit? A Finite-Answer Theory of Pre-Verbalization Commitment

ResearchDGX agent

arXiv:2605.06723v1 Announce Type: new Abstract: Language models often generate reasoning before giving a final answer, but the visible answer does not reveal when the model's answer preference became

When Does Critique Improve AI-Assisted Theoretical Physics? SCALAR: Structured Critic--Actor Loop for Agentic Reasoning

Model ReleasesDGX agent

arXiv:2605.06772v1 Announce Type: new Abstract: As large language models (LLMs) show increasing promise on research-level physics reasoning tasks and agentic AI becomes more common, a practical questi

When Losses Align: Gradient-Based Composite Loss Weighting for Efficient Pretraining

ResearchDGX agent

arXiv:2605.07756v1 Announce Type: cross Abstract: Modern deep models are often pretrained on large-scale data with missing labels using composite objectives, where the relative weights of multiple los

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

AgentsDGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

Where's the Plan? Locating Latent Planning in Language Models with Lightweight Mechanistic Interventions

Model ReleasesDGX agent

arXiv:2605.07984v1 Announce Type: cross Abstract: We study planning site formation in language models -- where internal representations of structurally-constrained future tokens form during the forwar

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

SafetyDGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

Why DDIM Hallucinates More than DDPM: A Theoretical Analysis of Reverse Dynamics

TutorialsDGX agent

arXiv:2605.06831v1 Announce Type: cross Abstract: We theoretically study the hallucination phenomena in two canonical diffusion samplers: the stochastic Denoising Diffusion Probabilistic Model (DDPM)

Why Self-Inconsistency Arises in GNN Explanations and How to Exploit It

Model ReleasesDGX agent

arXiv:2605.07527v1 Announce Type: cross Abstract: Recent work has observed that explanations produced by Self-Interpretable Graph Neural Networks (SI-GNNs) can be self-inconsistent: when the model is

WiCER: Wiki-memory Compile, Evaluate, Refine Iterative Knowledge Compilation for LLM Wiki Systems

Model ReleasesDGX agent

arXiv:2605.07068v1 Announce Type: cross Abstract: The LLM Wiki pattern, to compile and provide domain knowledge into a persistent artifact and serve it to LLMs via KV cache inference, promises context

XiYOLO: Energy-Aware Object Detection via Iterative Architecture Search and Scaling

HardwareDGX agent

arXiv:2605.06927v1 Announce Type: cross Abstract: Object detection on heterogeneous edge devices must satisfy strict energy, latency, and memory constraints while still providing reliable perception f

Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States

Model ReleasesDGX agent

arXiv:2605.07579v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) for Large Reasoning Models hinges on baseline estimation for variance reduction, but existing ap

7 May 2026

A Fast Model Counting Algorithm for Two-Variable Logic with Counting and Modulo Counting Quantifiers

ResearchDGX agent

arXiv:2605.03391v1 Announce Type: cross Abstract: Weighted first-order model counting (WFOMC) is a central task in lifted probabilistic inference: It asks for the weighted sum of all models of a first

A Skill-Based AI Agentic Pipeline for Library of Congress Subject Indexing

SafetyDGX agent

arXiv:2605.03537v1 Announce Type: cross Abstract: This paper presents a modular AI agentic skill pipeline for automating subject indexing with Library of Congress Subject Headings (LCSH). Subject inde

A Universal Space of Brain Dynamics for Unveiling Cognitive Transitions and Individual Differences

ResearchDGX agent

arXiv:2605.02936v1 Announce Type: cross Abstract: Representing dynamical systems through data-driven universal spaces has proven effective; however, achieving this universality for human brain activit

A Workflow-Oriented Framework for Asynchronous Human-AI Collaboration in Hybrid and Compute-Intensive HPC Environments

Local AiDGX agent

arXiv:2605.03743v1 Announce Type: cross Abstract: Human involvement is critical in training and deploying AI systems in high-stakes defence and security contexts. However, real-time interaction is imp

AdapShot: Adaptive Many-Shot In-Context Learning with Semantic-Aware KV Cache Reuse

ResearchDGX agent

arXiv:2605.03644v1 Announce Type: new Abstract: Many-Shot In-Context Learning (ICL) has emerged as a promising paradigm, leveraging extensive examples to unlock the reasoning potential of Large Langua

Adaptive Dual-Path Framework for Covert Semantic Communication

SafetyDGX agent

arXiv:2605.03423v1 Announce Type: new Abstract: This paper proposes a novel adaptive dual-path framework for covert semantic communication (SemCom), which integrates covert information transmission wi

Adaptive Long-term Embedding with Denoising and Augmentation for Recommendation

TutorialsDGX agent

arXiv:2504.13614v2 Announce Type: replace-cross Abstract: The rapid growth of the internet has made personalized recommendation systems indispensable. Graph-based sequential recommendation systems, po

Adaptive Reorganization of Neural Pathways for Continual Learning with Spiking Neural Networks

ResearchDGX agent

arXiv:2309.09550v4 Announce Type: replace-cross Abstract: The human brain can self-organize rich and diverse sparse neural pathways to incrementally master hundreds of cognitive tasks. However, most e

Agent-Based Modeling of Low-Emission Fertilizer Adoption for Dairy Farm Decarbonisation using Empirical Farm Data

SafetyDGX agent

arXiv:2605.03648v1 Announce Type: new Abstract: To understand complex system dynamics in dairy farming, it is essential to use modeling tools that capture farm heterogeneity, social interactions, and

Agentic publications: redesigning scientific publishing in the age of thinking large language models

AgentsDGX agent

arXiv:2505.13246v2 Announce Type: replace Abstract: Purpose: This paper introduces the concept of 'Agentic Publication,' a novel LLM-driven framework designed to complement traditional scientific publ

AI Advocate: Educational Path to Transform Squads to the Future

ApplicationsDGX agent

arXiv:2605.03800v1 Announce Type: cross Abstract: This paper analyzes the strategic education process aimed at transitioning traditional software development squads into hybrid structures centered on

AI4EOSC: a Federated Cloud Platform for Artificial Intelligence in Scientific Research

ApplicationsDGX agent

arXiv:2512.16455v3 Announce Type: replace-cross Abstract: The rapid growth of Artificial Intelligence and Machine Learning in scientific research has highlighted a gap between industry-standard MLOps

An Agent-Oriented Pluggable Experience-RAG Skill for Experience-Driven Retrieval Strategy Orchestration

AgentsDGX agent

arXiv:2605.03989v1 Announce Type: new Abstract: Retrieval-augmented generation systems often assume that one fixed retrieval pipeline is sufficient across heterogeneous tasks, yet factoid question ans

Are you with me? A Framework for Detecting Mental Model Discrepancies in Task-Based Team Dialogues

ResearchDGX agent

arXiv:2605.03149v1 Announce Type: new Abstract: Humans typically use natural language to update teammates on task states. Since not all updates are communicated, discrepancies arise between the team m

ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.03042v1 Announce Type: cross Abstract: This report describes ARIS (Auto-Research-in-sleep), an open-source research harness for autonomous research, including its architecture, assurance me

ARISE: A Repository-level Graph Representation and Toolset for Agentic Fault Localization and Program Repair

Local AiDGX agent

arXiv:2605.03117v1 Announce Type: cross Abstract: Repository-level fault localization (FL) and automated program repair (APR) require an agent to identify the relevant code units across files, follow

Automated Large-scale CVRP Solver Design via LLM-assisted Flexible MCTS

Model ReleasesDGX agent

arXiv:2605.03339v1 Announce Type: new Abstract: Solving large-scale CVRP (LSCVRP) with hundreds to thousands of nodes remains difficult for even state-of-the-art solvers. Divide-and-conquer can scale

Brainrot: Deskilling and Addiction are Overlooked AI Risks

SafetyDGX agent

arXiv:2605.03512v1 Announce Type: cross Abstract: The scope of AI safety and alignment work in generative artificial intelligence (GenAI) has so far mostly been limited to harms related to: (a) discri

Can AI Help You Get Over Your Breakup? One Session with a Belief-Reframing Chatbot Shows Sustained Distress Reduction

ResearchDGX agent

arXiv:2605.03261v1 Announce Type: cross Abstract: Romantic breakups are among the most common and intense sources of psychological distress. We evaluated *overit*, a single-session AI chatbot that use

Can LLMs Make (Personalized) Access Control Decisions?

AgentsDGX agent

arXiv:2511.20284v2 Announce Type: replace-cross Abstract: Precise access control decisions are crucial for the security of both traditional applications and emerging agent-based systems. Typically, th

Closed-Loop Vision-Language Planning for Multi-Agent Coordination

Model ReleasesDGX agent

arXiv:2502.10148v3 Announce Type: replace Abstract: Cooperative multi-agent reinforcement learning (MARL) struggles with sample efficiency, interpretability, and generalization. While Large Language M

Computing Thiele Rules on Interval Elections and their Generalizations

ResearchDGX agent

arXiv:2605.03067v1 Announce Type: new Abstract: Approval-based committee voting has received significant attention in the social choice community. Among the studied rules, Thiele rules, and especially

Contextual Multi-Objective Optimization: Rethinking Objectives in Frontier AI Systems

Local AiDGX agent

arXiv:2605.03900v1 Announce Type: new Abstract: Frontier AI systems perform best in settings with clear, stable, and verifiable objectives, such as code generation, mathematical reasoning, games, and

Continuum: Efficient and Robust Multi-Turn LLM Agent Scheduling with KV Cache Time-to-Live

Model ReleasesDGX agent

arXiv:2511.02230v4 Announce Type: replace-cross Abstract: KV cache management is essential for efficient LLM inference. To maximize utilization, existing inference engines evict finished requests' KV

Copula-Based Endogeneity Correction for Doubly Robust Estimation of Treatment Effect

SafetyDGX agent

arXiv:2605.03278v2 Announce Type: cross Abstract: Doubly Robust (DR) estimation of treatment effect relies on an untestable assumption that is the absence of unobserved confounding. This assumption is

cotomi Act: Learning to Automate Work by Watching You

AgentsDGX agent

arXiv:2605.03231v1 Announce Type: new Abstract: What if a browser agent could learn your work simply by watching you do it? We present cotomi Act, a browser-based computer-using agent that combines re

Cryptographic Registry Provenance: Structural Defense Against Dependency Confusion in AI Package Ecosystems

ApplicationsDGX agent

arXiv:2605.03309v1 Announce Type: cross Abstract: Dependency confusion attacks exploit a structural gap in software distribution: once a package is installed, there is no cryptographic proof of which

Deco: Extending Personal Physical Objects into Pervasive AI Companion through a Dual-Embodiment Framework

ResearchDGX agent

arXiv:2605.03882v1 Announce Type: cross Abstract: Individuals frequently form deep attachments to physical objects (e.g., plush toys) that usually cannot sense or respond to their emotions. While AI c

Decompose to Understand, Fuse to Detect: Frequency-Decoupled Anomaly Detection for Encrypted Network Traffic

SafetyDGX agent

arXiv:2605.02970v1 Announce Type: cross Abstract: Network traffic anomaly detection represents a critical cybersecurity task, yet widespread encryption makes this task increasingly challenging. In res

← Previous
1…284285286287288…358
Next →