AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Agents

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems

DGX agent

arXiv:2605.10555v1 Announce Type: new Abstract: As AI agents transition from research prototypes to enterprise production systems, the tool interfaces they consume remain rooted in human-oriented CRUD

agentsarxiv-cs-ai
12 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Agent-Omit: Adaptive Context Omission for Efficient LLM Agents

DGX agent

arXiv:2602.04284v2 Announce Type: replace Abstract: Managing agent context (e.g., thought and observation) during multi-turn agent-environment interactions is an emerging strategy to improve agent eff

safetyarxiv-cs-ai
12 May 2026
Safety

Agent-Sentry: Bounding LLM Agents via Execution Provenance

DGX agent

arXiv:2603.22868v2 Announce Type: replace-cross Abstract: Agentic computing systems, while immensely capable, raise serious security, privacy, and safety concerns. A key issue is that the full set of

safetyarxiv-cs-ai
12 May 2026
Model Releases

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

DGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

Agent-X: Full Pipeline Acceleration of On-device AI Agents

DGX agent

arXiv:2605.10380v1 Announce Type: new Abstract: LLM-based agents deliver state-of-the-art performance across tasks but incur high end-to-end latency on edge devices. We introduce Agent-X, a software-o

local-aiarxiv-cs-ai
12 May 2026
Model Releases

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

DGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

DGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

model-releasesarxiv-cs-ai
12 May 2026
Agents

Agentic AI Scientists Are Not Built For Autonomous Scientific Discovery

DGX agent

arXiv:2605.08956v1 Announce Type: new Abstract: A growing body of work pursues AI scientists capable of end-to-end autonomous scientific discovery. This position paper argues that although they alread

agentsarxiv-cs-ai
12 May 2026
Model Releases

Agentic MIP Research: Accelerated Constraint Handler Generation

DGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Agentic Performance at the Edge: Insights from Benchmarking

DGX agent

arXiv:2605.10384v1 Announce Type: new Abstract: Agentic artificial intelligence (AI) is a natural fit for Internet of Things (IoT) and edge systems, but edge deployments are often constrained to model

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization

DGX agent

arXiv:2605.08704v1 Announce Type: new Abstract: Multi-agent reasoning has shown promise for improving the problem-solving ability of large language models by allowing multiple agents to explore divers

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

DGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

DGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

model-releasesarxiv-cs-ai
12 May 2026
Safety

AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care

DGX agent

arXiv:2605.08480v1 Announce Type: new Abstract: Individuals with Alzheimer's disease (AD) and Alzheimer's disease-related dementia (ADRD) experience memory and thinking changes that impact their abili

safetyarxiv-cs-ai
12 May 2026
Applications

AI Native Asset Intelligence

DGX agent

arXiv:2605.09115v1 Announce Type: cross Abstract: Modern security environments generate fragmented signals across cloud resources, identities, configurations, and third-party security tools. Although

applicationsarxiv-cs-ai
12 May 2026
Safety

AIPO: : Learning to Reason from Active Interaction

DGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

safetyarxiv-cs-ai
12 May 2026
Safety

ALAM: Algebraically Consistent Latent Transitions for Vision-Language-Action Models

DGX agent

arXiv:2605.10819v1 Announce Type: cross Abstract: Vision-language-action (VLA) models remain constrained by the scarcity of action-labeled robot data, whereas action-free videos provide abundant evide

safetyarxiv-cs-ai
12 May 2026
Safety

Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification

DGX agent

arXiv:2605.09476v1 Announce Type: cross Abstract: Text simplification plays a crucial role in improving the accessibility and comprehensibility of written information for diverse audiences, including

safetyarxiv-cs-ai
12 May 2026
Model Releases

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

DGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

model-releasesarxiv-cs-ai
12 May 2026
Safety

Alignment as Jurisprudence

DGX agent

arXiv:2605.08416v1 Announce Type: new Abstract: Jurisprudence, the study of how judges should properly decide cases, and alignment, the science of getting AI models to conform to human values, share a

safetyarxiv-cs-ai
12 May 2026
Research

AllocMV: Optimal Resource Allocation for Music Video Generation via Structured Persistent State

DGX agent

arXiv:2605.10723v1 Announce Type: cross Abstract: Generating long-horizon music videos (MVs) is frequently constrained by prohibitive computational costs and difficulty maintaining cross-shot consiste

researcharxiv-cs-ai
12 May 2026
Model Releases

AlpsBench: An LLM Personalization Benchmark for Real-Dialogue Memorization and Preference Alignment

DGX agent

arXiv:2603.26680v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into lifelong AI assistants, LLM personalization has become a critical frontier. However, progress is c

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

DGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

model-releasesarxiv-cs-ai
12 May 2026
Agents

An agentic framework for gravitational-wave counterpart association in the multi-messenger era

DGX agent

arXiv:2605.10584v1 Announce Type: cross Abstract: With the detection of gravitational waves (GWs), multi-messenger astronomy has opened a new window for advancing our understanding of astrophysics, de

agentsarxiv-cs-ai
12 May 2026
Model Releases

An Empirical Study of Multi-Agent Collaboration for Automated Research

DGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

model-releasesarxiv-cs-ai
12 May 2026
Research

An Explainable Unsupervised-to-Supervised Machine Learning Framework for Dietary Pattern Discovery Using UK National Dietary Survey Data

DGX agent

arXiv:2605.08242v1 Announce Type: cross Abstract: Clinical dietary assessment can generate detailed but high-dimensional nutrient and food-group information that is difficult to translate quickly into

researcharxiv-cs-ai
12 May 2026
Local Ai

An Uncertainty-Aware Resilience Micro-Agent for Causal Observability in the Computing Continuum

DGX agent

arXiv:2605.10718v1 Announce Type: cross Abstract: Grey failures in the computing continuum produce ambiguous overlapping symptoms that existing approaches fail to diagnose reliably, either due to a la

local-aiarxiv-cs-ai
12 May 2026
Model Releases

AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation

DGX agent

arXiv:2605.10397v1 Announce Type: cross Abstract: Visual anomaly detection (VAD) is crucial in many real-world fields, such as industrial inspection, medical imaging, infrastructure monitoring, and re

model-releasesarxiv-cs-ai
12 May 2026
Research

Any2Any 3D Diffusion Models with Knowledge Transfer: A Radiotherapy Planning Study

DGX agent

arXiv:2605.09622v1 Announce Type: cross Abstract: Voxel-wise dose prediction is a critical yet challenging task in practical radiotherapy (RT) planning, as bespoke models trained from scratch often st

researcharxiv-cs-ai
12 May 2026
Tutorials

APCD: Adaptive Path-Contrastive Decoding for Reliable Large Language Model Generation

DGX agent

arXiv:2605.09492v1 Announce Type: cross Abstract: Large language models (LLMs) often suffer from hallucinations due to error accumulation in autoregressive decoding, where suboptimal early token choic

tutorialsarxiv-cs-ai
12 May 2026
Research

AR-VLA: True Autoregressive Action Expert for Vision-Language-Action Models

DGX agent

arXiv:2603.10126v2 Announce Type: replace-cross Abstract: We propose a standalone autoregressive (AR) Action Expert that generates actions as a continuous causal sequence while conditioning on refresh

researcharxiv-cs-ai
12 May 2026
Model Releases

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

DGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

model-releasesarxiv-cs-ai
12 May 2026
Research

ArchRAG: Attributed Community-based Hierarchical Retrieval-Augmented Generation

DGX agent

arXiv:2502.09891v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has proven effective in integrating external knowledge into large language models (LLMs) for solving ques

researcharxiv-cs-ai
12 May 2026
Model Releases

Artificial Intelligence in Number Theory: LLMs for Algorithm Generation and Ensemble Methods for Conjecture Verification

DGX agent

arXiv:2504.19451v3 Announce Type: cross Abstract: This paper presents two concrete applications of Artificial Intelligence to algorithmic and analytic number theory. Recent benchmarks of large languag

model-releasesarxiv-cs-ai
12 May 2026
Agents

ASIA: an Autonomous System Identification Agent

DGX agent

arXiv:2605.10480v1 Announce Type: new Abstract: Over the years, research in system identification has provided a rich set of methods for learning dynamical models, together with well-established theor

agentsarxiv-cs-ai
12 May 2026
Model Releases

AssayBench: An Assay-Level Virtual Cell Benchmark for LLMs and Agents

DGX agent

arXiv:2605.10876v1 Announce Type: cross Abstract: Recent advances in machine learning and large-scale biological data collections have revived the prospect of building a virtual cell, a computational

model-releasesarxiv-cs-ai
12 May 2026
Safety

Assessing Trustworthiness of AI Training Dataset using Subjective Logic -- A Use Case on Bias

DGX agent

arXiv:2508.13813v2 Announce Type: replace-cross Abstract: As AI systems increasingly rely on training data, assessing dataset trustworthiness has become critical, particularly for properties like fair

safetyarxiv-cs-ai
12 May 2026
Applications

Assessment of RAG and Fine-Tuning for Industrial Question-Answering-Applications

DGX agent

arXiv:2605.09533v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly employed in enterprise question-answering (QA) systems, requiring adaptation to domain-specific knowledg

applicationsarxiv-cs-ai
12 May 2026
Agents

AtteConDA: Attention-Based Conflict Suppression in Multi-Condition Diffusion Models and Synthetic Data Augmentation

DGX agent

arXiv:2605.09425v1 Announce Type: cross Abstract: Recent conditional image generation methods can improve controllability by generating images that are faithful to conditions such as sketches, human p

agentsarxiv-cs-ai
12 May 2026
Tutorials

Attention-based graph neural networks: a survey

DGX agent

arXiv:2605.08679v1 Announce Type: cross Abstract: Graph neural networks (GNNs) aim to learn well-trained representations in a lower-dimension space for downstream tasks while preserving the topologica

tutorialsarxiv-cs-ai
12 May 2026
Tutorials

Attention Drift: What Autoregressive Speculative Decoding Models Learn

DGX agent

arXiv:2605.09992v1 Announce Type: cross Abstract: Speculative decoding accelerates LLM inference by drafting future tokens with a small model, but drafter models degrade sharply under template perturb

tutorialsarxiv-cs-ai
12 May 2026
Safety

Attention-Mamba: A Mamba-Enhanced Multi-Scale Parallel Inference Network for Medical Image Segmentation

DGX agent

arXiv:2402.02286v4 Announce Type: replace-cross Abstract: U-shaped architectures have long dominated the field of medical image segmentation, while Transformers are widely employed for modeling long-r

safetyarxiv-cs-ai
12 May 2026
Research

Attractor-Vascular Coupling Theory: Formal Grounding and Empirical Validation for AAMI-Standard Cuffless Blood Pressure Estimation from Smartphone Photoplethysmography

DGX agent

arXiv:2605.10871v1 Announce Type: cross Abstract: This work proposes Attractor-Vascular Coupling Theory (AVCT), a mathematical framework showing that cardiac attractor geometry encodes blood pressure

researcharxiv-cs-ai
12 May 2026
Research

Attribution-based Explanations for Markov Decision Processes

DGX agent

arXiv:2605.09780v1 Announce Type: new Abstract: Attribution techniques explain the outcome of an AI model by assigning a numerical score to its inputs. So far, these techniques have mainly focused on

researcharxiv-cs-ai
12 May 2026
Research

AU-Harness: An Open-Source Toolkit for Holistic Evaluation of Audio LLMs

DGX agent

arXiv:2509.08031v3 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) are rapidly advancing, but evaluating them remains challenging due to inefficient and non-standardized too

researcharxiv-cs-ai
12 May 2026
Safety

Auditing Data Membership in Reinforcement Learning With Verifiable Rewards

DGX agent

arXiv:2511.14045v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a core training stage in recent large language models (LLMs). Its reliance on

safetyarxiv-cs-ai
12 May 2026
Safety

Auto-Rubric as Reward: From Implicit Preferences to Explicit Multimodal Generative Criteria

DGX agent

arXiv:2605.08354v1 Announce Type: new Abstract: Aligning multimodal generative models with human preferences demands reward signals that respect the compositional, multi-dimensional structure of human

safetyarxiv-cs-ai
12 May 2026
Model Releases

Automated Approach for Solving Infinite-state Polynomial Reachability Games

DGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…328329330331332…448
Next →