AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

PACE: A Proxy for Agentic Capability Evaluation

DGX agent

arXiv:2607.02032v1 Announce Type: new Abstract: Evaluating LLM agents on benchmarks like SWE-Bench and GAIA can be expensive, time-consuming, and requires complex infrastructure. A single evaluation c

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

DGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

model-releasesarxiv-cs-lg
1 Jul 2026
Agents

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2507.23604v2 Announce Type: replace Abstract: Decentralized Multi-Agent Reinforcement Learning (MARL) methods allow for learning scalable multi-agent policies, but suffer from partial observabil

agentsarxiv-cs-lg
1 Jul 2026
Agents

HyphaeDB: A Living Knowledge Topology for Agent-First Memory

DGX agent

arXiv:2606.28781v1 Announce Type: new Abstract: Every existing vector database and agent memory framework treats memory as passive storage that agents query explicitly. No system propagates knowledge

agentsarxiv-cs-ai
30 Jun 2026
Agents

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

DGX agent

arXiv:2606.30555v1 Announce Type: new Abstract: The rapid integration of Large Language Models (LLMs) has driven the evolution of Multi-Agent Systems (MAS), where specialized agents collaborate to exe

agentsarxiv-cs-ai
30 Jun 2026
Agents

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

DGX agent

arXiv:2606.23664v1 Announce Type: new Abstract: Multi-agent systems (MAS) offer a scalable path forward for agentic AI, comprising multiple LLM-based agents, each assigned a system prompt and a positi

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

DGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Agent Economics: An Entropy-Controlled Pluralistic Alignment Framework for Preventing Artificial Hivemind in Autonomous Agents

DGX agent

arXiv:2606.09039v1 Announce Type: new Abstract: This study proposes the Behavioral Protocol Framework (BPF), an entropy-controlled pluralistic alignment framework designed to address two critical chal

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

DGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Detecting Perspective Shifts in Multi-agent Systems

DGX agent

arXiv:2512.05013v2 Announce Type: replace Abstract: Generative models augmented with external tools and update mechanisms (or extit{agents}) have demonstrated capabilities beyond intelligent prompting

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

Do More Agents Help? Controlled and Protocol-Aligned Evaluation of LLM Agent Workflows

DGX agent

arXiv:2606.05670v1 Announce Type: new Abstract: Does adding more agents help an LLM workflow once compared systems share the same benchmark loader, tool access, answer contract, usage accounting, and

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

DGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

agentsarxiv-cs-ai
3 Jun 2026
Agents

Agentic Clustering: Controllable Text Taxonomies via Multi-Agent Refinement

DGX agent

arXiv:2606.01255v1 Announce Type: new Abstract: Recent text-clustering methods use large language models to propose a cluster taxonomy from a corpus and then assign each text to it. These pipelines ar

agentsarxiv-cs-cl
2 Jun 2026
Agents

LRAgent: Efficient KV Cache Sharing for Multi-LoRA LLM Agents

DGX agent

arXiv:2602.01053v2 Announce Type: replace Abstract: Role specialization in multi-LLM agent systems is often realized via multi-LoRA, where agents share a pretrained backbone and differ only by lightwe

agentsarxiv-cs-lg
2 Jun 2026
Agents

Modeling Distinct Human Interaction in Web Agents

DGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

agentsarxiv-cs-cl
2 Jun 2026
Model Releases

How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

DGX agent

arXiv:2605.29442v1 Announce Type: cross Abstract: AI coding agents increasingly act directly within software environments, yet existing analyses of their failures rely on benchmark trajectories that m

model-releasesarxiv-cs-ai
29 May 2026
Agents

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

DGX agent

arXiv:2605.24486v1 Announce Type: new Abstract: Recent progress on long-horizon agentic tasks has been driven largely by scaling up individual agents through stronger models, better tools, and more ef

agentsarxiv-cs-ai
26 May 2026
Model Releases

Push Your Agent: Measuring and Enforcing Quantitative Goal Persistence in Long-Horizon LLM Agents

DGX agent

arXiv:2605.23574v1 Announce Type: new Abstract: Long-horizon language agents can make many plausible local tool calls yet fail to persist until a requested count is actually complete. We study this ga

model-releasesarxiv-cs-lg
25 May 2026
Agents

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

DGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

agentsarxiv-cs-ai
22 May 2026
Safety

Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics

DGX agent

arXiv:2605.16300v1 Announce Type: cross Abstract: Robotic systems are moving from isolated platforms to interconnected multi-agent ecosystems that operate in human environments. This shift raises a go

safetyarxiv-cs-ai
19 May 2026
Agents

MAC: Masked Agent Collaboration Boosts Large Language Model Medical Decision-Making

DGX agent

arXiv:2507.21159v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have proven effective in artificial intelligence, where the multi-agent system (MAS) holds considerable promise f

agentsarxiv-cs-lg
13 May 2026
Agents

When Child Inherits: Modeling and Exploiting Subagent Spawn in Multi-Agent Networks

DGX agent

arXiv:2605.08460v1 Announce Type: cross Abstract: Since the official release of ChatGPT in 2022, large language models (LLMs) have rapidly evolved from chatbot-style interfaces into agentic systems th

agentsarxiv-cs-ai
12 May 2026
Safety

MAGIQ: A Post-Quantum Multi-Agentic AI Governance System with Provable Security

DGX agent

arXiv:2605.06933v1 Announce Type: new Abstract: Our computing ecosystem is being transformed by two emerging paradigms: the increased deployment of agentic AI systems and advancements in quantum compu

safetyarxiv-cs-lg
11 May 2026
Model Releases

6G Needs Agents: Toward Agentic AI-Native Networks for Autonomous Intelligence

DGX agent

arXiv:2605.01546v1 Announce Type: cross Abstract: Sixth-generation (6G) networks are increasingly envisioned as AI-native infrastructures integrating communication, sensing, and computing into a unifi

model-releasesarxiv-cs-ai
6 May 2026
Agents

Near-Optimal Privacy-Preserving Learning for Max-Min Fair Multi-Agent Bandits

DGX agent

arXiv:2306.04498v3 Announce Type: replace Abstract: We study fair multi-agent multi-armed bandit learning under collision-only coordination. Agents cannot communicate explicitly during learning and ob

agentsarxiv-cs-lg
5 May 2026
Safety

Agentic Memory: Learning Unified Long-Term and Short-Term Memory Management for Large Language Model Agents

DGX agent

arXiv:2601.01885v2 Announce Type: replace Abstract: Large language model (LLM) agents face fundamental limitations in long-horizon reasoning due to finite context windows, making effective memory mana

safetyarxiv-cs-cl
1 May 2026
Model Releases

ORFS-agent: Tool-Using Agents for Chip Design Optimization

DGX agent

arXiv:2506.08332v3 Announce Type: replace Abstract: Machine learning has been widely used to optimize complex engineering workflows across numerous domains. In integrated circuit design, modern flows

model-releasesarxiv-cs-ai
1 May 2026
Local Ai

Emergent Coordination in Multi-Agent Language Models

DGX agent

arXiv:2510.05174v4 Announce Type: replace-cross Abstract: When are multi-agent LLM systems merely a collection of individual agents versus an integrated collective with higher-order structure? We intr

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses

DGX agent

arXiv:2604.25850v1 Announce Type: new Abstract: Harnesses have become a central determinant of coding-agent performance, shaping how models interact with repositories, tools, and execution environment

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation

DGX agent

arXiv:2604.25318v1 Announce Type: cross Abstract: Cutscenes are carefully choreographed cinematic sequences embedded in video games and interactive media, serving as the primary vehicle for narrative

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

Food4All: A Multi-Agent Framework for Real-time Free Food Discovery with Integrated Nutritional Metadata

DGX agent

arXiv:2510.18289v2 Announce Type: replace Abstract: Food insecurity remains a persistent public health emergency in the United States, tightly interwoven with chronic disease, mental illness, and opio

agentsarxiv-cs-cl
28 Apr 2026
Model Releases

GraphPlanner: Graph Memory-Augmented Agentic Routing for Multi-Agent LLMs

DGX agent

arXiv:2604.23626v1 Announce Type: new Abstract: LLM routing has achieved promising results in integrating the strengths of diverse models while balancing efficiency and performance. However, to suppor

model-releasesarxiv-cs-cl
28 Apr 2026
Agents

Poster: ClawdGo: Endogenous Security Awareness Training for Autonomous AI Agents

DGX agent

arXiv:2604.24020v1 Announce Type: cross Abstract: Autonomous AI agents deployed on platforms such as OpenClaw face prompt injection, memory poisoning, supply-chain attacks, and social engineering, yet

agentsarxiv-cs-ai
28 Apr 2026
Agents

WinkTPG: An Execution Framework for Multi-Agent Path Finding Using Temporal Reasoning

DGX agent

arXiv:2508.01495v2 Announce Type: replace Abstract: Planning collision-free paths for a large group of agents is a challenging problem in many real-world applications. While recent advances in Multi-A

agentsarxiv-cs-ai
28 Apr 2026
Agents

Sound Agentic Science Requires Adversarial Experiments

DGX agent

arXiv:2604.22080v1 Announce Type: new Abstract: LLM-based agents are rapidly being adopted for scientific data analysis, automating tasks once limited by human time and expertise. This capability is o

agentsarxiv-cs-ai
27 Apr 2026
Agents

LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

DGX agent

arXiv:2411.10109v2 Announce Type: replace Abstract: Machine learning can predict human behavior well when substantial structured data and well-defined outcomes are available, but these models are typi

agentsarxiv-cs-ai
23 Apr 2026
Safety

LLM Agents Predict Social Media Reactions but Do Not Outperform Text Classifiers: Benchmarking Simulation Accuracy Using 120K+ Personas of 1511 Humans

DGX agent

arXiv:2604.19787v1 Announce Type: cross Abstract: Social media platforms mediate how billions form opinions and engage with public discourse. As autonomous AI agents increasingly participate in these

safetyarxiv-cs-ai
23 Apr 2026
Agents

Memory-Augmented LLM-based Multi-Agent System for Automated Feature Generation on Tabular Data

DGX agent

arXiv:2604.20261v1 Announce Type: new Abstract: Automated feature generation extracts informative features from raw tabular data without manual intervention and is crucial for accurate, generalizable

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models

DGX agent

arXiv:2604.18612v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, while recent prompting strategies such as Chain-of-Thou

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

DGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

COSEARCH: Joint Training of Reasoning and Document Ranking via Reinforcement Learning for Agentic Search

DGX agent

arXiv:2604.17555v1 Announce Type: cross Abstract: Agentic search -- the task of training agents that iteratively reason, issue queries, and synthesize retrieved information to answer complex questions

safetyarxiv-cs-cl
21 Apr 2026
Agents

Personalizing Student-Agent Interactions Using Log-Contextualized Retrieval-Augmented Generation (RAG)

DGX agent

arXiv:2505.17238v3 Announce Type: replace Abstract: Collaborative dialogue offers rich insights into students' learning and critical thinking, which is essential for personalizing pedagogical agent in

agentsarxiv-cs-cl
21 Apr 2026
Agents

Capture the Flags: Family-Based Evaluation of Agentic LLMs via Semantics-Preserving Transformations

DGX agent

arXiv:2602.05523v2 Announce Type: replace-cross Abstract: Agentic large language models (LLMs) are increasingly evaluated on cybersecurity tasks using capture-the-flag (CTF) benchmarks, yet existing p

agentsarxiv-cs-ai
20 Apr 2026
Safety

Symbolic Guardrails for Domain-Specific Agents: Stronger Safety and Security Guarantees Without Sacrificing Utility

DGX agent

arXiv:2604.15579v1 Announce Type: cross Abstract: AI agents that interact with their environments through tools enable powerful applications, but in high-stakes business settings, unintended actions c

safetyarxiv-cs-ai
20 Apr 2026
Agents

AgentForge: Execution-Grounded Multi-Agent LLM Framework for Autonomous Software Engineering

DGX agent

arXiv:2604.13120v1 Announce Type: cross Abstract: Large language models generate plausible code but cannot verify correctness. Existing multi-agent systems simulate execution or leave verification opt

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

Theory of Mind in Action: The Instruction Inference Task in Dynamic Human-Agent Collaboration

DGX agent

arXiv:2507.02935v2 Announce Type: replace Abstract: Successful human-agent teaming relies on an agent being able to understand instructions given by a (human) principal. In many cases, an instruction

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ExpSeek: Self-Triggered Experience Seeking for Web Agents

DGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

pi-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data

DGX agent

arXiv:2604.14054v1 Announce Type: cross Abstract: Deep search agents have emerged as a promising paradigm for addressing complex information-seeking tasks, but their training remains challenging due t

agentsarxiv-cs-cl
16 Apr 2026
← Previous
1…1011121314…230
Next →