AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,736 results
Agents

Enhancing Multi-Agent Communication through Attention Steering with Context Relevance

DGX agent

arXiv:2605.30136v1 Announce Type: new Abstract: LLM-based multi-agent systems have demonstrated remarkable performance on complex tasks through collaborative reasoning. However, these systems tend to

agentsarxiv-cs-ai
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs

DGX agent

arXiv:2605.29512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for social and strategic reasoning over extended intera

agentsarxiv-cs-ai
29 May 2026
Safety

A Policy-Driven Runtime Layer for Agentic LLM Serving

DGX agent

arXiv:2605.27744v1 Announce Type: new Abstract: Multi-agent LLM systems have become the dominant production workload, but the serving stack was not built for them. The agent framework above knows agen

safetyarxiv-cs-ai
28 May 2026
Agents

Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet Named

DGX agent

arXiv:2605.27396v1 Announce Type: cross Abstract: Autonomous AI agents now plan, decide, and act on behalf of users across healthcare, financial services, and workplace contexts, often without step-by

agentsarxiv-cs-ai
28 May 2026
Agents

Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

DGX agent

arXiv:2605.27954v1 Announce Type: new Abstract: Agentic large language models are increasingly used to solve real-world tasks by reasoning over goals, invoking tools, and interacting with external env

agentsarxiv-cs-lg
28 May 2026
Agents

Defending LLM-based Multi-Agent Systems Against Cooperative Attacks with Sentence-Level Rectification

DGX agent

arXiv:2605.28104v1 Announce Type: new Abstract: Recent years have witnessed the rapid development of Large Language Model-based Multi-Agent Systems (MAS), which excel at collaborative decision-making

agentsarxiv-cs-ai
28 May 2026
Safety

LACUNA: Safe Agents as Recursive Program Holes

DGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

safetyarxiv-cs-ai
28 May 2026
Agents

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

DGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

agentsarxiv-cs-ai
28 May 2026
Agents

CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly

DGX agent

arXiv:2605.26195v1 Announce Type: cross Abstract: LLM-based agents are increasingly used for cybersecurity tasks, but most existing systems rely on fixed, human-designed scaffolds that struggle to ada

agentsarxiv-cs-ai
27 May 2026
Agents

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments

DGX agent

arXiv:2605.27209v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have facilitated the widespread deployment of LLMs as interactive agents capable of reasoning, planning,

agentsarxiv-cs-ai
27 May 2026
Agents

The Necessity of a Unified Framework for LLM-Based Agent Evaluation

DGX agent

arXiv:2602.03238v2 Announce Type: replace Abstract: With the advent of Large Language Models (LLMs), general-purpose agents have seen fundamental advancements. However, evaluating these agents present

agentsarxiv-cs-ai
27 May 2026
Model Releases

AMA-Bench: Evaluating Long-Horizon Memory for Agentic Applications

DGX agent

arXiv:2602.22769v3 Announce Type: replace Abstract: Large Language Models (LLMs) are deployed as autonomous agents in increasingly complex applications, where enabling long-horizon memory is critical

model-releasesarxiv-cs-ai
26 May 2026
Safety

Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?

DGX agent

arXiv:2605.25929v1 Announce Type: cross Abstract: The effectiveness of multi-agent LLM deliberation depends not only on the agents' individual predictions, but also on how they communicate and collabo

safetyarxiv-cs-lg
26 May 2026
Safety

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

DGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

safetyarxiv-cs-ai
26 May 2026
Agents

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEn…

DGX agent

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEngineer Singapore last week that comprehensively traces through

agentsjerry-liu--x
25 May 2026
Safety

MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems

DGX agent

arXiv:2602.04431v2 Announce Type: replace Abstract: LLM-based multi-agent systems have demonstrated impressive capabilities, but they also introduce significant safety risks when individual agents fai

safetyarxiv-cs-lg
25 May 2026
Agents

LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems

DGX agent

arXiv:2605.22786v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on intermediate communication to coordinate complex tasks. While most existing

agentsarxiv-cs-lg
23 May 2026
Agents

Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development

DGX agent

arXiv:2605.20456v1 Announce Type: cross Abstract: Agentic AI coding systems can inspect repositories, plan implementation steps, edit files, call tools, run tests, and submit pull requests. These capa

agentsarxiv-cs-ai
22 May 2026
Agents

babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs…

DGX agent

babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems' https://arxiv.org/a

agentsyohei-nakajima--x
22 May 2026
Agents

Observability for any agent, anywhere: Production-ready tracing with OpenTelemetry & Unity Catalog on Databricks

DGX agent

Databricks provides production-ready observability solutions for AI agents using OpenTelemetry integration and Unity Catalog, enabling comprehensive tracing and monitoring across distributed agent dep

agentsdatabricks
22 May 2026
Agents

SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations

DGX agent

arXiv:2605.22564v1 Announce Type: new Abstract: Today, tool-calling agents are commonly evaluated or tested on static datasets of execution traces, including input commands, agent responses, and assoc

agentsarxiv-cs-cl
22 May 2026
Agents

Learning Incentive Structures for Cooperative Resilience in Multi-Agent Systems under Social Dilemmas

DGX agent

arXiv:2601.22292v2 Announce Type: replace-cross Abstract: Multi-agent social dilemmas, such as the tragedy of the commons, capture settings where individual incentives conflict with collective well-be

agentsarxiv-cs-lg
21 May 2026
Agents

Agent Security is a Systems Problem

DGX agent

arXiv:2605.18991v1 Announce Type: cross Abstract: We take the position that agent security must be approached as a systems problem: the AI model powering the agent must be treated as an untrusted comp

agentsarxiv-cs-ai
20 May 2026
Agents

Toward Training Superintelligent Software Agents through Self-Play SWE-RL

DGX agent

arXiv:2512.18552v2 Announce Type: replace-cross Abstract: While current software agents powered by large language models (LLMs) and agentic reinforcement learning (RL) can boost programmer productivit

agentsarxiv-cs-ai
20 May 2026
Agents

CORAL: Towards Autonomous Multi-Agent Evolution for Open-Ended Discovery

DGX agent

arXiv:2604.01658v2 Announce Type: replace Abstract: Large language model (LLM)-based evolution is a promising approach for open-ended discovery, where progress requires sustained search and knowledge

agentsarxiv-cs-ai
19 May 2026
Model Releases

FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics

DGX agent

arXiv:2605.17373v1 Announce Type: cross Abstract: AI research agents accelerate ML research by automating hypothesis generation, experimentation, and empirical refinement. Existing agent strategies ra

model-releasesarxiv-cs-ai
19 May 2026
Safety

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

DGX agent

arXiv:2605.16303v1 Announce Type: cross Abstract: Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demog

safetyarxiv-cs-ai
19 May 2026
Agents

S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination

DGX agent

arXiv:2605.17076v1 Announce Type: cross Abstract: Concurrent LLM agents sharing mutable natural-language state produce Structural Race Conditions (SRCs): write-write and cross-shard stale-read conflic

agentsarxiv-cs-ai
19 May 2026
Agents

Skim: Speculative Execution for Fast and Efficient Web Agents

DGX agent

arXiv:2605.16565v1 Announce Type: new Abstract: Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is

agentsarxiv-cs-ai
19 May 2026
Agents

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

DGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

agentsarxiv-cs-ai
18 May 2026
Agents

Recently @pinecone introduced Nexus – a new knowledge-engine layer for AI agents that reduces token use by up to 90%. It’s built on top of a…

DGX agent

Recently @pinecone introduced Nexus – a new knowledge-engine layer for AI agents that reduces token use by up to 90%. It’s built on top of a vector database, but shifts reasoning earlier in the pipeli

agentspinecone--x
17 May 2026
Agents

xAI has expanded access to X Premium+ subscribers in Hermes Agent. Enjoy!

DGX agent

xAI has expanded access to X Premium+ subscribers in Hermes Agent. Enjoy! You can now use X Premium subscriptions in Hermes Agent, and Hermes Agent can now search X posts. https://x.ai/news/grok-herme

agentsnous-research--x
16 May 2026
Model Releases

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

DGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

model-releasesarxiv-cs-ai
15 May 2026
Agents

MetaAgent-X : Breaking the Ceiling of Automatic Multi-Agent Systems via End-to-End Reinforcement Learning

DGX agent

arXiv:2605.14212v1 Announce Type: new Abstract: Automatic multi-agent systems aim to instantiate agent workflows without relying on manually designed or fixed orchestration. However, existing automati

agentsarxiv-cs-ai
15 May 2026
Agents

Harnessing Agentic Evolution

DGX agent

arXiv:2605.13821v1 Announce Type: new Abstract: Agentic evolution has emerged as a powerful paradigm for improving programs, workflows, and scientific solutions by iteratively generating candidates, e

agentsarxiv-cs-ai
14 May 2026
Agents

if you want to mainline hermes agent memory architecture real quick...

DGX agent

if you want to mainline hermes agent memory architecture real quick... the three-tier memory of Hermes agent. AI agents forgets everything when your session ends. Hermes doesn't. it has three memory l

agentsnous-research--x
14 May 2026
Agents

LangSmith Engine feels like the missing piece for reliable agents. Clustering failures and suggesting targeted fixes turns raw traces into a…

DGX agent

LangSmith Engine feels like the missing piece for reliable agents. Clustering failures and suggesting targeted fixes turns raw traces into actionable improvements. This is how we scale agent developme

agentsharrison-chase--x
14 May 2026
Agents

which was your favorite launch? SmithDB (database purpose built for agent trace data): https://www.langchain.com/blog/introducing-smithdb La…

DGX agent

which was your favorite launch? SmithDB (database purpose built for agent trace data): https://www.langchain.com/blog/introducing-smithdb LangSmith Engine (agent for improving your agents based on tra

agentsharrison-chase--x
14 May 2026
Safety

Events as Triggers for Behavioral Diversity in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.12388v1 Announce Type: cross Abstract: Effective multi-agent cooperation requires agents to adopt diverse behaviors as task conditions evolve-and to do so at the right moment. Yet, current

safetyarxiv-cs-lg
13 May 2026
Agents

So many agent builders at Langchain Interrupt 2026 😍😍😍

DGX agent

At LangChain Interrupt 2026, multiple agent builder tools and frameworks were showcased, reflecting the growing ecosystem of agentic AI development. This post highlights the proliferation of agent-bui

agentsharrison-chase--x
13 May 2026
Agents

Spend less time on triaging Ship fixes faster Catch regressions earlier Introducing LangSmith Engine: an agent that works autonomously to fi…

DGX agent

Spend less time on triaging Ship fixes faster Catch regressions earlier Introducing LangSmith Engine: an agent that works autonomously to find patterns in your agent's failures 💻 Accelerate the agent

agentsharrison-chase--x
13 May 2026
Agents

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

DGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

agentsarxiv-cs-ai
12 May 2026
Model Releases

An Empirical Study of Multi-Agent Collaboration for Automated Research

DGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

model-releasesarxiv-cs-ai
12 May 2026
Agents

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability

DGX agent

arXiv:2605.10516v1 Announce Type: new Abstract: This paper establishes a rigorous measurement science for AI agent reliability, providing a foundational framework for quantifying consistency under sem

agentsarxiv-cs-ai
12 May 2026
Agents

Generalization Bounds of Emergent Communications for Agentic AI Networking

DGX agent

arXiv:2605.08613v1 Announce Type: new Abstract: The evolution of 6G networking toward agentic AI networking (AgentNet) systems requires a shift from traditional data pipelines to task-aware, agentic A

agentsarxiv-cs-ai
12 May 2026
Model Releases

Pairwise is Not Enough: Hypergraph Neural Networks for Multi-Agent Pathfinding

DGX agent

arXiv:2602.06733v2 Announce Type: replace-cross Abstract: Multi-Agent Path Finding (MAPF) is a representative multi-agent coordination problem, where multiple agents are required to navigate to their

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

SAP SAPPHIRE 2026: Google Cloud unveils unified agentic vision and massive compute scaling

DGX agent

In today's hyper-connected market, an enterprise's most valuable asset — mission-critical data — often remains trapped in legacy silos. For years, leadership teams have navigated a data pipeline dilem

model-releasesgoogle-cloud-ai
12 May 2026
Agents

Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments

DGX agent

arXiv:2605.09721v1 Announce Type: cross Abstract: Tool-enabled AI agents are increasingly deployed in cloud-hosted environments and offered as services, where they perform side-effecting operations th

agentsarxiv-cs-ai
12 May 2026
← Previous
1…3940414243…370
Next →