AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation

DGX agent

arXiv:2509.26600v2 Announce Type: replace-cross Abstract: As LLMs rapidly saturate existing benchmarks, automated benchmark creation using LLMs (LLM-as-a-benchmark) -- where a model generates test inp

model-releasesarxiv-cs-ai
27 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Where Code Meets Natural Language: Taxonomy-Driven Information Flow Analysis for LLM-Integrated Applications

DGX agent

arXiv:2603.28345v2 Announce Type: replace-cross Abstract: LLM API calls are becoming a ubiquitous program construct, yet they create a boundary that no existing program analysis can cross: runtime val

applicationsarxiv-cs-ai
27 May 2026
Safety

Which Changes Matter? Towards Trustworthy Legal AI via Relevance-Sensitive Evaluation and Solver-Grounded Reasoning

DGX agent

arXiv:2605.26530v1 Announce Type: new Abstract: Legal reasoning requires distinguishing changes that matter from those that do not. Legal AI should remain stable under legally irrelevant perturbations

safetyarxiv-cs-ai
27 May 2026
Research

Why LLMs Hallucinate on Structured Knowledge: A Mechanistic Analysis of Reasoning over Linearized Representations

DGX agent

arXiv:2605.26362v1 Announce Type: cross Abstract: In many reasoning tasks, large language models (LLMs) rely on structured external knowledge, such as graphs and tables, which is typically linearized

researcharxiv-cs-ai
27 May 2026
Model Releases

Workflow Closure Is Not Scientific Closure in Auto-Research Systems

DGX agent

arXiv:2605.26200v1 Announce Type: cross Abstract: This paper argues that workflow closure is not scientific closure in auto-research systems. Current systems can increasingly complete research-like lo

model-releasesarxiv-cs-ai
27 May 2026
Hardware

Xe-Forge: Multi-Stage LLM-Powered Kernel Optimization for Intel GPU

DGX agent

arXiv:2605.26118v1 Announce Type: cross Abstract: Porting deep learning algorithms to new hardware accelerators requires developers to repeatedly apply the same low-level optimizations -- quantization

hardwarearxiv-cs-ai
27 May 2026
Agents

XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs

DGX agent

arXiv:2601.04426v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on dynamic structured generation, such as tool calling and response protocols. Unlike traditional structured gen

agentsarxiv-cs-ai
27 May 2026
Research

Yes, Q-learning Helps Offline In-Context RL

DGX agent

arXiv:2502.17666v4 Announce Type: replace-cross Abstract: Existing offline in-context reinforcement learning (ICRL) methods have predominantly relied on supervised training objectives, which are known

researcharxiv-cs-ai
27 May 2026
Model Releases

Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems

DGX agent

arXiv:2605.26302v1 Announce Type: new Abstract: Long-lived AI agents are increasingly deployed as persistent operational systems, yet they are still evaluated like freshly initialized models. Day-one

model-releasesarxiv-cs-ai
27 May 2026
Research

A Comprehensive Dataset for Human vs. AI Generated Image Detection

DGX agent

arXiv:2601.00553v2 Announce Type: replace-cross Abstract: Multimodal generative AI systems like Stable Diffusion, DALL-E, and MidJourney have fundamentally changed how synthetic images are created. Th

researcharxiv-cs-ai
26 May 2026
Model Releases

A Controlled Synthetic Benchmark for Educational Aspect-Based Sentiment Analysis

DGX agent

arXiv:2605.25502v1 Announce Type: cross Abstract: Educational aspect-based sentiment analysis (ABSA) can support course improvement, but public aspect-labeled student feedback remains scarce because e

model-releasesarxiv-cs-ai
26 May 2026
Research

A Deep Dive into Axiomatic Design -- Part I: Problem Formulation

DGX agent

arXiv:2605.25735v1 Announce Type: new Abstract: Problem formulation translating customer needs and constraints into a minimum set of independent first-level functional requirements, is arguably the mo

researcharxiv-cs-ai
26 May 2026
Research

A Dynamical Framework for Cognitive Processes Based on Transformations and Semantic Equivalence

DGX agent

arXiv:2605.23942v1 Announce Type: new Abstract: This paper proposes a structural and dynamical framework for modeling cognitive processes within a cybernetic perspective. Cognitive states are represen

researcharxiv-cs-ai
26 May 2026
Research

A general tensor-structured compression scheme for efficient large language models

DGX agent

arXiv:2605.25344v1 Announce Type: cross Abstract: Large language models (LLMs) are dominated by dense linear transformations, whose storage, memory and computational overheads hinder efficient adaptat

researcharxiv-cs-ai
26 May 2026
Safety

A governance horizon for ethical-use constraints in open-weight AI models

DGX agent

arXiv:2605.24383v1 Announce Type: new Abstract: Ethical constraints on open-weight AI models are both a reflection of societal concerns and a foundation for AI governance policy. They are expected to

safetyarxiv-cs-ai
26 May 2026
Model Releases

A Large-Scale Dataset and Benchmark: Do Protein-Ligand Models Learn Binding Sites or Just Binding Likelihood?

DGX agent

arXiv:2605.24045v1 Announce Type: cross Abstract: Protein-ligand modeling underpins computational drug discovery and molecular design. Existing protein-ligand benchmarks typically evaluate whether a p

model-releasesarxiv-cs-ai
26 May 2026
Agents

A Multi-Agent LLM Framework for Rating the Quality of Surgical Feedback

DGX agent

arXiv:2605.25440v1 Announce Type: cross Abstract: Verbal feedback delivered by attending surgeons in the operating room plays a critical formative role in resident trainee skill acquisition. Yet, asse

agentsarxiv-cs-ai
26 May 2026
Safety

A Multimodal 3D Foundation Model for Light Sheet Fluorescence Microscopy Enables Few-Shot Segmentation, Classification, and Deblurring

DGX agent

arXiv:2605.26026v1 Announce Type: cross Abstract: Light sheet fluorescence microscopy (LSM) enables high-resolution, three-dimensional (3D) imaging of biological specimens, providing rich volumetric d

safetyarxiv-cs-ai
26 May 2026
Research

A Signal-Language Foundation Model for Broad-Spectrum Cardiovascular Assessment from Routine Electrocardiography

DGX agent

arXiv:2605.25446v1 Announce Type: new Abstract: Electrocardiography (ECG) is central to cardiovascular care, but conventional AI models are often restricted to common arrhythmias and may generalize po

researcharxiv-cs-ai
26 May 2026
Safety

A Sober Look at Agentic Misalignment in Automated Workflows

DGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

safetyarxiv-cs-ai
26 May 2026
Safety

A Tertiary Review of Large Language Model-Based Code Generating Tasks: Trends, Challenges, and Future Directions

DGX agent

arXiv:2605.25536v1 Announce Type: cross Abstract: Context. Large language models (LLMs) are increasingly applied to code-generating tasks (CGTs) in software engineering. While reported results are pro

safetyarxiv-cs-ai
26 May 2026
Agents

A Token/KV-Cache Communication Media Selection and Resource Allocation Strategy for Multi-Agent Collaboration

DGX agent

arXiv:2605.25422v1 Announce Type: cross Abstract: The convergence of large language models (LLMs) with 6G networks is fostering a paradigm of autonomous multi-agent cooperation, which in turn is expec

agentsarxiv-cs-ai
26 May 2026
Model Releases

A World Model of Radiologist Reading for Medical Image Representation Learning

DGX agent

arXiv:2605.23992v1 Announce Type: cross Abstract: Radiologist eye-tracking data provide a rich record of how experts search, compare, and accumulate evidence during image reading; yet, existing method

model-releasesarxiv-cs-ai
26 May 2026
Research

Abduction-Deduction Entanglement: Domain Generalization via Representation Transplants

DGX agent

arXiv:2605.25156v1 Announce Type: cross Abstract: Prediction models trained under the source distribution do not generalize well to a different target distribution. A valid inference about an unseen d

researcharxiv-cs-ai
26 May 2026
Research

Accelerating Long-Tail Generation in Synchronous RLHF Training via Adaptive Tensor Parallelism

DGX agent

arXiv:2605.23945v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a key post-training paradigm for improving model quality. However, the synchronous three-st

researcharxiv-cs-ai
26 May 2026
Model Releases

Acting on the Unseen: Communication-Free Collaborative Filtering for Decentralized Multi-Robot Task Allocation

DGX agent

arXiv:2605.25584v1 Announce Type: cross Abstract: Multi-robot task allocation usually assumes some combination of communication, known task models, or a coordinator. We study the opposite extreme, a r

model-releasesarxiv-cs-ai
26 May 2026
Applications

Actionable and diverse counterfactual explanations incorporating domain knowledge and plausibility constraints

DGX agent

arXiv:2511.20236v3 Announce Type: replace Abstract: Counterfactual explanations improve the actionable interpretability of machine learning models by identifying minimal changes required to achieve a

applicationsarxiv-cs-ai
26 May 2026
Model Releases

ActQuant: Sub-4-bit Action-Guided Quantization for Vision-Language-Action Models

DGX agent

arXiv:2605.24011v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models exhibit remarkable action generation for embodied intelligence, but their heavy compute make deployment on edge pl

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Adaptive Graph Refinement and Label Propagation with LLMs for Cost-Effective Entity Resolution

DGX agent

arXiv:2605.25814v1 Announce Type: cross Abstract: Dirty entity resolution (ER), which identifies records referring to the same real-world entity from a single, messy dataset, is a fundamental task in

model-releasesarxiv-cs-ai
26 May 2026
Safety

Adaptive Human-AI Coordination via Hierarchical Action Disentanglement

DGX agent

arXiv:2605.24343v1 Announce Type: new Abstract: Human-AI collaboration requires agents that can adapt to diverse partner behaviors and skill levels while remaining robust to unseen partners. Existing

safetyarxiv-cs-ai
26 May 2026
Agents

Adaptive Punishment for Cooperation in Mixed-Motive Games

DGX agent

arXiv:2605.24516v1 Announce Type: cross Abstract: Mixed-motive scenarios are ubiquitous in real-world multi-agent interactions, where self-interested agents often defect for immediate rewards, overloo

agentsarxiv-cs-ai
26 May 2026
Applications

ADMFormer: An Adaptive-Decomposition Transformer with Time-Varying Masked Spatial Attention for Traffic Forecasting

DGX agent

arXiv:2605.25543v1 Announce Type: new Abstract: Accurate traffic forecasting is essential for intelligent transportation systems, supporting a wide range of real-world applications. However, it remain

applicationsarxiv-cs-ai
26 May 2026
Model Releases

Advancing Graph Few-Shot Learning via In-Context Learning

DGX agent

arXiv:2605.24410v1 Announce Type: new Abstract: Graph few-shot learning, which aims to classify nodes from novel classes with only a few labeled examples, is a widely studied problem in graph learning

model-releasesarxiv-cs-ai
26 May 2026
Safety

AdvantageFlow: Advantage-Weighted Least Squares for RL in Flow Models

DGX agent

arXiv:2605.26013v1 Announce Type: cross Abstract: We introduce AdvantageFlow, a forward-process reinforcement learning algorithm for rectified flow models. Unlike Flow-GRPO, which optimizes the revers

safetyarxiv-cs-ai
26 May 2026
Safety

Adversarial Error Correction for Visual Autoregressive Generation

DGX agent

arXiv:2605.24843v1 Announce Type: cross Abstract: Visual Autoregressive (VAR) models have emerged as a powerful paradigm for image synthesis by performing hierarchical next-scale prediction. However,

safetyarxiv-cs-ai
26 May 2026
Research

Adversarial Network Imagination: Causal LLMs and Digital Twins for Proactive Telecom Mitigation

DGX agent

arXiv:2602.13203v2 Announce Type: replace-cross Abstract: Telecommunication networks experience complex failures such as fiber cuts, traffic overloads, and cascading outages. Existing monitoring and d

researcharxiv-cs-ai
26 May 2026
Research

Adversarial Orthogonal Disentanglement for LVLM Hallucination Mitigation

DGX agent

arXiv:2605.25377v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have advanced multimodal understanding, yet their reliability is limited by hallucination, where generated conten

researcharxiv-cs-ai
26 May 2026
Agents

Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis

DGX agent

arXiv:2605.24600v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for qualitative data analysis (QDA), yet their outputs often miss the depth and nuance of human analy

agentsarxiv-cs-ai
26 May 2026
Safety

Agent-Centric Social Trajectory Prediction: A Free Energy Principle Perspective

DGX agent

arXiv:2605.25748v1 Announce Type: new Abstract: Trajectory prediction methods have demonstrated remarkable capabilities in capturing complex motion patterns. However, existing methods rely on global s

safetyarxiv-cs-ai
26 May 2026
Safety

Agent-Facing Information Design in LLM Tool Registries

DGX agent

arXiv:2605.23916v1 Announce Type: cross Abstract: LLM tool registries function as unregulated advertising platforms: providers write free-text descriptions that agents use for selection, yet no measur

safetyarxiv-cs-ai
26 May 2026
Safety

Agent Learning via Early Experience

DGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

safetyarxiv-cs-ai
26 May 2026
Agents

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

DGX agent

arXiv:2605.24823v1 Announce Type: new Abstract: Manufacturing has passed through four widely recognized paradigms - mechanization, electrification, programmable automation, and Smart Manufacturing - e

agentsarxiv-cs-ai
26 May 2026
Agents

Agent Primitives: Reusable Latent Building Blocks for Multi-Agent Systems

DGX agent

arXiv:2602.03695v2 Announce Type: replace-cross Abstract: While existing multi-agent systems (MAS) can handle complex problems by enabling collaboration among multiple agents, they are often highly ta

agentsarxiv-cs-ai
26 May 2026
Safety

Agent-ToM: Learning to Monitor Autonomous LLM Agents via Theory-of-Mind Reasoning

DGX agent

arXiv:2605.24216v1 Announce Type: cross Abstract: Monitoring autonomous large language model (LLM) agents for covert malicious behavior is challenging due to delayed, context-dependent, and long-horiz

safetyarxiv-cs-ai
26 May 2026
Model Releases

Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning

DGX agent

arXiv:2602.10090v3 Announce Type: replace Abstract: Recent advances in large language model (LLM) have empowered autonomous agents to perform multi-turn interactions with tools and environments. Howev

model-releasesarxiv-cs-ai
26 May 2026
Agents

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

DGX agent

arXiv:2605.24486v1 Announce Type: new Abstract: Recent progress on long-horizon agentic tasks has been driven largely by scaling up individual agents through stronger models, better tools, and more ef

agentsarxiv-cs-ai
26 May 2026
Model Releases

AgentHijack: Benchmarking Computer Use Agent Robustness to Common Environment Corruptions

DGX agent

arXiv:2605.25707v1 Announce Type: new Abstract: Autonomous computer use agents that powered by multimodal large language models (MLLMs) are emerging as capable assistants for completing complex digita

model-releasesarxiv-cs-ai
26 May 2026
Agents

AGI Requires a Coordination Layer on Top of Pattern Repositories

DGX agent

arXiv:2512.05765v2 Announce Type: replace Abstract: In this paper we argue that influential critiques dismissing Large Language Models (LLMs) as a dead end for AGI misidentify the bottleneck: they con

agentsarxiv-cs-ai
26 May 2026
← Previous
1…262263264265266…448
Next →