AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

Self-Evolving World Models for LLM Agent Planning

DGX agent

arXiv:2606.30639v1 Announce Type: new Abstract: World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unrel

agentsarxiv-cs-ai
30 Jun 2026
Agents

Agentic Publication Protocol: An Attempt to Modernize Scientific Publication

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.27386v1 Announce Type: cross Abstract: Scientific publication is still organized primarily around static manuscripts, even though much of scientific progress depends on tacit know-how: how

agentsarxiv-cs-ai
29 Jun 2026
Agents

Delayed Verification Destabilizes Multi-Agent LLM Belief: Instability Thresholds and Optimal Corrector Placement

DGX agent

arXiv:2606.27409v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems often rely on verifier and critic agents to suppress hallucinations, but verification is delayed. Durin

agentsarxiv-cs-cl
29 Jun 2026
Model Releases

Govern the Repository, Not the Agent: Measuring Ecosystem-Level Risk in AI-Native Software

DGX agent

arXiv:2606.28235v1 Announce Type: cross Abstract: Autonomous coding agents now open and merge pull requests in shared repositories at scale, and the field evaluates them the way it has always evaluate

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

A Process Harness for Uplifting Legacy Workflows to Agentic BPM: Design and Realization in CUGA FLO

DGX agent

arXiv:2606.27188v1 Announce Type: new Abstract: We introduce the process harness, a new mechanism for uplifting legacy workflows into Agentic Business Process Management (Agentic BPM) without replacin

safetyarxiv-cs-ai
26 Jun 2026
Agents

Agentic System as Compressor: Quantifying System Intelligence in Bits

DGX agent

arXiv:2606.25960v1 Announce Type: new Abstract: Large language models are turning from isolated predictors into agentic systems: they call tools, retrieve evidence, obey environment constraints, use v

agentsarxiv-cs-ai
25 Jun 2026
Agents

Autodata: An agentic data scientist to create high quality synthetic data

DGX agent

arXiv:2606.25996v1 Announce Type: cross Abstract: We introduce Autodata, a general method that enables AI agents to act as data scientists who build high quality training and evaluation data. We show

agentsarxiv-cs-cl
25 Jun 2026
Agents

Domain-Specific Agents for Cherenkov Telescope Array Control Software and Gamma-Ray Data Analysis

DGX agent

arXiv:2510.01299v3 Announce Type: replace-cross Abstract: We present domain-adapted large language model agents designed to support Cherenkov Telescope Array operation and data analysis. The agents co

agentsarxiv-cs-ai
25 Jun 2026
Agents

Evaluating AGENTS.md: Are Repository-Level Context Files Helpful for Coding Agents?

DGX agent

arXiv:2602.11988v2 Announce Type: replace-cross Abstract: A widespread practice in software development is to tailor coding agents to repositories using context files, such as AGENTS.md. Although this

agentsarxiv-cs-ai
25 Jun 2026
Safety

Low Variance Trust Region Optimization with Independent Actors and Sequential Updates in Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.25526v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning assumes each agent shares the same reward function and can be trained effectively using the Trust Region

safetyarxiv-cs-lg
25 Jun 2026
Agents

RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

DGX agent

arXiv:2606.23927v1 Announce Type: new Abstract: Agentic AI systems powered by large language models (LLMs) are rapidly evolving into autonomous decision-making systems, exposing attack vectors beyond

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents

DGX agent

arXiv:2506.04018v3 Announce Type: replace-cross Abstract: As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents'

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

DGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

safetyarxiv-cs-lg
23 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

DGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations

DGX agent

arXiv:2606.11218v1 Announce Type: cross Abstract: Ethical deliberation is often misunderstood as a search for single right or wrong answers, creating difficulties for non-ethically trained personnel w

safetyarxiv-cs-ai
11 Jun 2026
Agents

FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse

DGX agent

arXiv:2606.11290v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems are increasingly powerful, but current agentic workflow optimization paradigms make an unsatisfyi

agentsarxiv-cs-ai
11 Jun 2026
Agents

Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security

DGX agent

arXiv:2606.11671v1 Announce Type: cross Abstract: Agent skills let LLM agents reuse instructions, resources, tools, and workflows, but they also create a new place for malicious behavior to hide. A sk

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

DGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

DGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

safetyarxiv-cs-lg
9 Jun 2026
Agents

DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

DGX agent

arXiv:2605.22781v2 Announce Type: replace-cross Abstract: LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid chec

agentsarxiv-cs-ai
9 Jun 2026
Agents

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

DGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

DGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

DGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

agentsarxiv-cs-ai
8 Jun 2026
Agents

When Does Multi-Agent Collaboration Help? An Entropy Perspective

DGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

agentsarxiv-cs-ai
8 Jun 2026
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

SentinelBench: A Benchmark for Long-Running Monitoring Agents

DGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

MARDoc: A Memory-Aware Refinement Agent Framework for Multimodal Long Document QA

DGX agent

arXiv:2606.05749v1 Announce Type: new Abstract: Iterative retrieval-reasoning agents have recently shown promise for multimodal long-document question answering. However, most existing systems maintai

agentsarxiv-cs-cl
5 Jun 2026
Agents

SkillComposer: Learning to Evolve Agent Skills for Specification and Generalization

DGX agent

arXiv:2606.06079v1 Announce Type: new Abstract: Agent skills, which consist of reusable strategies that guide agent reasoning and action, have shown strong potential for improving model capability at

agentsarxiv-cs-cl
5 Jun 2026
Agents

Cascading Hallucination in Agentic RAG: The CHARM Framework for Detection and Mitigation

DGX agent

arXiv:2606.04435v1 Announce Type: new Abstract: Multi-step agentic retrieval-augmented generation (RAG) pipelines have demonstrated significant capability for complex reasoning tasks, yet remain vulne

agentsarxiv-cs-ai
4 Jun 2026
Safety

Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling

DGX agent

arXiv:2606.05021v1 Announce Type: new Abstract: We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algor

safetyarxiv-cs-lg
4 Jun 2026
Agents

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

DGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

agentsarxiv-cs-ai
4 Jun 2026
Agents

Adaptive Latent Agentic Reasoning

DGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

eMEM: A Hybrid Spatio-Temporal Memory System For Embodied Agents

DGX agent

arXiv:2606.03374v1 Announce Type: new Abstract: We present eMEM (Embodied Memory), a hybrid graph-based memory system for embodied agents operating in physical environments. Current agent memory archi

model-releasesarxiv-cs-ro
3 Jun 2026
Agents

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

DGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

agentsarxiv-cs-ai
3 Jun 2026
Agents

OpenAgenet/OAN: Technical Architecture for Trust-Governed Agent Identity and Discovery

DGX agent

arXiv:2606.03163v1 Announce Type: cross Abstract: This paper describes the technical architecture of OpenAgenet / OAN. OAN is a protocol-neutral trust layer for open Agent interconnection. It specifie

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

The DeepSpeak-Agentic Dataset

DGX agent

arXiv:2606.03686v1 Announce Type: new Abstract: We present DeepSpeak-Agentic, a dataset of videos comprising over 37 hours of semi-structured conversations between a human and an embodied AI agent. We

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection

DGX agent

arXiv:2606.02812v1 Announce Type: new Abstract: Modeling patient trajectories from longitudinal electronic health records (EHRs) requires reasoning over sparse, noisy, and long-context multimodal sequ

agentsarxiv-cs-ai
3 Jun 2026
Safety

ANDES: Agent Native Data Evolving Synthesis Tool for Autonomous Instruction Alignment

DGX agent

arXiv:2606.01279v1 Announce Type: new Abstract: AI agents are increasingly being tasked with automating AI research itself, particularly the critical post-training phase that transforms base LLMs into

safetyarxiv-cs-ai
2 Jun 2026
Agents

CAREAgent: Clinical Agent with Structured Reasoning and Tool-Integrated for Order Generation

DGX agent

arXiv:2606.01094v1 Announce Type: new Abstract: Clinical order generation serves as a critical bridge between clinical decision-making and real-world practice, translating medical decisions into concr

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

DGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Dynamic Trust-Aware Sparse Communication Topology for LLM-Based Multi-Agent Consensus

DGX agent

arXiv:2606.01828v1 Announce Type: cross Abstract: Large language model-driven multi-agent systems enhance the reliability of complex reasoning tasks through multi-round deliberation, role specializati

agentsarxiv-cs-ai
2 Jun 2026
Agents

MACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment

DGX agent

arXiv:2312.03644v3 Announce Type: replace Abstract: Offline Multi-agent Reinforcement Learning (MARL) is valuable in scenarios where online interaction is impractical or risky. While independent learn

agentsarxiv-cs-lg
2 Jun 2026
Agents

Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

DGX agent

arXiv:2606.00408v1 Announce Type: cross Abstract: Long-horizon search agents accumulate large amounts of retrieved content across many tool calls, making context-budget efficiency increasingly importa

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

DGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Scaling Multi-Agent Environment Co-Design with Diffusion Models

DGX agent

arXiv:2511.03100v2 Announce Type: replace-cross Abstract: The agent-environment co-design paradigm jointly optimises agent policies and environment configurations in search of improved system performa

safetyarxiv-cs-ai
1 Jun 2026
Agents

Skill Reuse as Compression in Agentic RL

DGX agent

arXiv:2605.31509v1 Announce Type: cross Abstract: Large language model agents trained with reinforcement learning (RL) often learn brittle, task-specific shortcuts. We hypothesize that agents generali

agentsarxiv-cs-ai
1 Jun 2026
Agents

Estimating the Empowerment of Language Model Agents

DGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

agentsarxiv-cs-ai
29 May 2026
← Previous
1…2627282930…233
Next →