AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies

DGX agent

arXiv:2508.01049v2 Announce Type: replace Abstract: Independent on-policy policy gradient algorithms are widely used for multi-agent reinforcement learning (MARL) in cooperative and no-conflict games,

safetyarxiv-cs-lg
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation

DGX agent

arXiv:2605.12857v1 Announce Type: cross Abstract: Existing API-based agentic systems for RTL code generation are fundamentally misaligned with industrial practice: they assume a golden testbench is av

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack

DGX agent

arXiv:2605.12673v1 Announce Type: new Abstract: Agent benchmarks have become the de facto measure of frontier AI competence, guiding model selection, investment, and deployment. However, reward hackin

model-releasesarxiv-cs-ai
14 May 2026
Safety

Quantitative Certification of Agentic Tool Selection

DGX agent

arXiv:2510.03992v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in agentic systems, where a fundamental task is mapping user intents to relevant extern

safetyarxiv-cs-ai
14 May 2026
Model Releases

Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs

DGX agent

arXiv:2505.11556v4 Announce Type: replace-cross Abstract: Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet syst

model-releasesarxiv-cs-ai
14 May 2026
Agents

ToolMol: Evolutionary Agentic Framework for Multi-objective Drug Discovery

DGX agent

arXiv:2605.12784v1 Announce Type: new Abstract: Advances in large language models (LLMs) have recently opened new and promising avenues for small-molecule drug discovery. Yet existing LLM-based approa

agentsarxiv-cs-lg
14 May 2026
Model Releases

An Empirical Study of Automating Agent Evaluation

DGX agent

arXiv:2605.11378v1 Announce Type: new Abstract: Agent evaluation requires assessing complex multi-step behaviors involving tool use and intermediate reasoning, making it costly and expertise-intensive

model-releasesarxiv-cs-cl
13 May 2026
Agents

Deep Reasoning in General Purpose Agents via Structured Meta-Cognition

DGX agent

arXiv:2605.11388v1 Announce Type: new Abstract: Humans intuitively solve complex problems by flexibly shifting among reasoning modes: they plan, execute, revise intermediate goals, resolve ambiguity t

agentsarxiv-cs-cl
13 May 2026
Agents

DORA: Dynamic Online Reinforcement Agent for Token Merging in Vision Transformers

DGX agent

arXiv:2605.11683v1 Announce Type: new Abstract: Vision Transformers (ViTs) incur significant computational overhead due to the quadratic complexity of self-attention relative to the token sequence len

agentsarxiv-cs-cv
13 May 2026
Agents

Dynamic Full-body Motion Agent with Object Interaction via Blending Pre-trained Modular Controllers

DGX agent

arXiv:2605.11369v1 Announce Type: new Abstract: Generating physically plausible dynamic motions of human-object interaction (HOI) remains challenging, mainly due to existing HOI datasets limited to st

agentsarxiv-cs-cv
13 May 2026
Model Releases

ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks?

DGX agent

arXiv:2605.11086v1 Announce Type: cross Abstract: AI agents are rapidly gaining capabilities that could significantly reshape cybersecurity, making rigorous evaluation urgent. A critical capability is

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

No More, No Less: Task Alignment in Terminal Agents

DGX agent

arXiv:2605.12233v1 Announce Type: new Abstract: Terminal agents are increasingly capable of executing complex, long-horizon tasks autonomously from a single user prompt. To do so, they must interpret

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

SkillSafetyBench: Evaluating Agent Safety under Skill-Facing Attack Surfaces

DGX agent

arXiv:2605.12015v1 Announce Type: cross Abstract: Reusable skills are becoming a common interface for extending large language model agents, packaging procedural guidance with access to files, tools,

model-releasesarxiv-cs-cl
13 May 2026
Local Ai

AdaSwitch: Adaptive Switching between Small and Large Agents for Effective Cloud-Local Collaborative Learning

DGX agent

arXiv:2410.13181v2 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have been remarkable. Users face a choice between using cloud-based LLMs for generation quality

local-aiarxiv-cs-cl
12 May 2026
Hardware

CellDX AI Autopilot: Agent-Guided Training and Deployment of Pathology Classifiers

DGX agent

arXiv:2605.10362v1 Announce Type: new Abstract: Training AI models for computational pathology currently requires access to expensive whole-slide-image datasets, GPU infrastructure, deep expertise in

hardwarearxiv-cs-cv
12 May 2026
Model Releases

Collective Alignment in LLM Multi-Agent Systems: Disentangling Bias from Cooperation via Statistical Physics

DGX agent

arXiv:2605.10528v1 Announce Type: cross Abstract: We investigate the emergent collective dynamics of LLM-based multi-agent systems on a 2D square lattice and present a model-agnostic statistical-physi

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents

DGX agent

arXiv:2605.09826v1 Announce Type: new Abstract: Theory of Mind (ToM), the ability to track others epistemic state, makes humans efficient collaborators. AI agents need the same capacity in multi agent

model-releasesarxiv-cs-ai
12 May 2026
Agents

Enhancing Consistency Models for Multi-Agent Trajectory Prediction

DGX agent

arXiv:2605.08572v1 Announce Type: new Abstract: Diffusion models for multi-agent trajectory prediction are limited by iterative denoising, which causes inference latency that hinders their use in time

agentsarxiv-cs-cv
12 May 2026
Model Releases

M2A: Synergizing Mathematical and Agentic Reasoning in Large Language Models

DGX agent

arXiv:2605.09879v1 Announce Type: new Abstract: While reasoning has become a central capability of large language models (LLMs), the reasoning patterns required for different scenarios are often misal

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

MAGS-SLAM: Monocular Multi-Agent Gaussian Splatting SLAM for Geometrically and Photometrically Consistent Reconstruction

DGX agent

arXiv:2605.10760v1 Announce Type: new Abstract: Collaborative photorealistic 3D reconstruction from multiple agents enables rapid large-scale scene capture for virtual production and cooperative multi

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

MDrive: Benchmarking Closed-Loop Cooperative Driving for End-to-End Multi-agent Systems

DGX agent

arXiv:2605.10904v1 Announce Type: new Abstract: Vehicle-to-Everything (V2X) communication has emerged as a promising paradigm for autonomous driving, enabling connected agents to share complementary p

model-releasesarxiv-cs-ro
12 May 2026
Safety

Mem-W: Latent Memory-Native GUI Agents

DGX agent

arXiv:2605.09317v1 Announce Type: new Abstract: GUI agents are beginning to operate the web, mobile, and desktop as interactive worlds, where successful control depends on carrying forward visual, pro

safetyarxiv-cs-cl
12 May 2026
Agents

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning

DGX agent

arXiv:2605.09287v1 Announce Type: new Abstract: Large Language Model (LLM)-based search agents trained with reinforcement learning (RL) have significantly improved the performance of knowledge-intensi

agentsarxiv-cs-ai
12 May 2026
Agents

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

DGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

agentsarxiv-cs-lg
12 May 2026
Local Ai

Scaling Mobile Agent Systems: From Capability Density to Collective Intelligence

DGX agent

arXiv:2605.08124v1 Announce Type: cross Abstract: Mobile agent systems are emerging as a key paradigm for enabling intelligent applications on edge devices and in AIoT ecosystems. However, their scala

local-aiarxiv-cs-cl
12 May 2026
Agents

SkillRAE: Agent Skill-Based Context Compilation for Retrieval-Augmented Execution

DGX agent

arXiv:2605.10114v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents (e.g., OpenClaw) increasingly rely on reusable skill libraries to solve artifact-rich tasks such as document-cen

agentsarxiv-cs-cl
12 May 2026
Agents

TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

DGX agent

arXiv:2605.10344v1 Announce Type: new Abstract: Test-time scaling has become an effective paradigm for improving the reasoning ability of large language models by allocating additional computation dur

agentsarxiv-cs-ai
12 May 2026
Agents

ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering

DGX agent

arXiv:2510.20036v2 Announce Type: replace Abstract: Large language model (LLM) agents rely on external tools to solve complex tasks, but real-world toolsets often contain redundant tools with overlapp

agentsarxiv-cs-cl
12 May 2026
Agents

When Independent Sampling Outperforms Agentic Reasoning

DGX agent

arXiv:2605.08478v1 Announce Type: new Abstract: We study how to allocate inference-time compute for competitive programming under fixed budgets. Evaluating 216 Codeforces problems across Divisions 1-3

agentsarxiv-cs-lg
12 May 2026
Agents

A Multi-Memory Segment System for Generating High-Quality Long-Term Memory Content in Agents

DGX agent

arXiv:2508.15294v4 Announce Type: replace Abstract: In the current field of agent memory, extensive explorations have been conducted in the area of memory retrieval, yet few studies have focused on ex

agentsarxiv-cs-ai
11 May 2026
Local Ai

ADKO: Agentic Decentralized Knowledge Optimization

DGX agent

arXiv:2605.07863v1 Announce Type: new Abstract: We present Agentic Decentralized Knowledge Optimization (ADKO), a framework for collaborative black-box optimization across autonomous agents that achie

local-aiarxiv-cs-lg
11 May 2026
Safety

Agentic Coding Needs Proactivity, Not Just Autonomy

DGX agent

arXiv:2605.06717v1 Announce Type: cross Abstract: Coding agents are rapidly changing the landscape of software development, moving from inline completion to autonomous systems that edit repositories,

safetyarxiv-cs-ai
11 May 2026
Model Releases

Beyond the Black Box: Interpretability of Agentic AI Tool Use

DGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

model-releasesarxiv-cs-ai
11 May 2026
Agents

Exponential Sample Complexity Separation between Flat and Hierarchical Agentic Theorem Provers

DGX agent

arXiv:2602.10512v2 Announce Type: replace Abstract: Agentic theorem provers often introduce intermediate lemmas, proof sketches, or subgoal decompositions before returning to tactic-level search. This

agentsarxiv-cs-lg
11 May 2026
Agents

HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion

DGX agent

arXiv:2605.07472v1 Announce Type: cross Abstract: Insider threat detection assumes that an adaptive insider leaves behavioral residue distinguishing them from legitimate users. We test this assumption

agentsarxiv-cs-ai
11 May 2026
Local Ai

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning

DGX agent

arXiv:2605.07505v1 Announce Type: new Abstract: Developing lightweight, on-device vision-language GUI agents is essential for efficient cross-platform automated interaction. However, current on-device

local-aiarxiv-cs-ai
11 May 2026
Model Releases

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

DGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Region4Web: Rethinking Observation Space Granularity for Web Agents

DGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

model-releasesarxiv-cs-ai
11 May 2026
Research

Rethinking Experience Utilization in Self-Evolving Language Model Agents

DGX agent

arXiv:2605.07164v1 Announce Type: new Abstract: Self-evolving agents improve by accumulating and reusing experience from past interactions. Existing work has largely focused on how experience is const

researcharxiv-cs-cl
11 May 2026
Safety

Self-Programmed Execution for Language-Model Agents

DGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

safetyarxiv-cs-ai
11 May 2026
Safety

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

DGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

safetyarxiv-cs-ro
11 May 2026
Agents

Adaptivity Under Realizability Constraints: Comparing In-Context and Agentic Learning

DGX agent

arXiv:2605.04995v1 Announce Type: new Abstract: We compare in-context learning with fixed queries and agentic learning with adaptive queries for uniform approximation of task families. We consider two

agentsarxiv-cs-lg
7 May 2026
Agents

GEM: Graph-Enhanced Mixture-of-Experts with ReAct Agents for Dialogue State Tracking

DGX agent

arXiv:2605.04449v1 Announce Type: new Abstract: Dialogue State Tracking (DST) requires precise extraction of structured information from multi-domain conversations, a task where Large Language Models

agentsarxiv-cs-cl
7 May 2026
Agents

Learning Correct Behavior from Examples: Validating Sequential Execution in Autonomous Agents

DGX agent

arXiv:2605.03159v1 Announce Type: new Abstract: As autonomous agents become increasingly sophisticated, validating their sequential behavior presents a significant challenge. Traditional testing appro

agentsarxiv-cs-ai
7 May 2026
Model Releases

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

DGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

model-releasesarxiv-cs-ai
7 May 2026
Model Releases

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

DGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne

model-releasesarxiv-cs-ai
7 May 2026
Agents

Agentic Multi-Source Grounding for Enhanced Query Intent Understanding: A DoorDash Case Study

DGX agent

arXiv:2603.01486v2 Announce Type: replace Abstract: Accurately mapping user queries to business categories is a fundamental Information Retrieval challenge for multi-category marketplaces, where conte

agentsarxiv-cs-ai
6 May 2026
Model Releases

Multi-Agent Reasoning Improves Compute Efficiency: Pareto-Optimal Test-Time Scaling

DGX agent

arXiv:2605.01566v1 Announce Type: new Abstract: Advances in inference methods have enabled language models to improve their predictions without additional training. These methods often prioritize raw

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…6061626364…233
Next →