AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
7 Jul 2026

MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution

Local AiDGX agent

arXiv:2607.05297v1 Announce Type: new Abstract: Recent LLM agents tackle increasingly long-horizon, open-ended tasks, and external skills, reusable procedural knowledge supplied to the agent, further

Shutdownable Agents through POST-Agency

AgentsDGX agent

arXiv:2505.20203v4 Announce Type: replace Abstract: Many fear that future artificial agents will resist shutdown. I present an idea - the POST-Agents Proposal - for ensuring that doesn't happen. I pro

Towards Mitigation of Hallucination for LLM-empowered Agents: Progressive Generalization Bound Exploration and Watchdog Monitor

AgentsDGX agent

arXiv:2507.15903v2 Announce Type: replace-cross Abstract: Empowered by large language models (LLMs), intelligent agents have become a popular paradigm for interacting with open environments to facilit

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
3 Jul 2026

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/shea…

Model ReleasesDGX agent

Excited to share our paper, “Learning Multi-Agent Coordination via Sheaf-ADMM” to be presented at #ICML2026 Blog: https://pub.sakana.ai/sheaf-admm/ Most AI models process information as one giant, mon

2 Jul 2026

@john_my07 I ran the same prompt through @imagine Agent mode and the results were pretty amazing!

AgentsDGX agent

This post discusses running a prompt through an AI agent mode feature, comparing results between standard and agent modes, with the author noting positive outcomes from the agent mode implementation.

llm-coding-agent 0.1a0

Model ReleasesDGX agent

Release: llm-coding-agent 0.1a0 Another Fable 5 experiment. Now that my LLM library has evolved into more of an agent framework it's time to see what a simple coding agent would look like built on it.

1 Jul 2026

Agentic AI Enhances Physician Trust in Clinical Decision Making

AgentsDGX agent

arXiv:2606.30658v1 Announce Type: cross Abstract: Medical AI has shifted from reasoning to agentic AI, a new paradigm that autonomously invokes external tools during reasoning, rendering intermediate

RigorBench: Benchmarking Engineering Process Discipline in Autonomous AI Coding Agents

Model ReleasesDGX agent

arXiv:2606.22678v2 Announce Type: replace-cross Abstract: Agentic coding harnesses - such as Agent-Skills, Superpowers, and Agent-Rigor - are increasingly deployed to augment underlying LLMs for real-

30 Jun 2026

DAIN: Dynamic Agent-Based Interaction Network for Efficient and Collaborative Multimodal Reasoning

AgentsDGX agent

arXiv:2606.30189v1 Announce Type: new Abstract: Current multimodal fusion approaches, particularly those based on static Mixture-of-Experts (MoE) architectures, often struggle to provide the adaptive

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

Model ReleasesDGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

Memory as an Attack Surface in LLM Agents: A Study on Multiple-Choice Question Answering

AgentsDGX agent

arXiv:2606.29030v1 Announce Type: new Abstract: AI agents extend conventional large language model (LLM) applications by integrating language understanding with task execution, external tool use, and

Preventing Error Propagation in Multi-Agent AI through Runtime Monitoring

AgentsDGX agent

arXiv:2606.29026v1 Announce Type: new Abstract: Multi-agent AI systems can improve answer selection by allowing different language models to exchange reasoning traces, revise initial predictions, and

Vorlon debuts Guardian to block risky AI agent actions before they complete

AgentsDGX agent

Agentic ecosystem security startup Vorlon Inc. today launched Guardian, a real-time enforcement gateway that aims to block risky actions by artificial intelligence agents before a transaction complete

29 Jun 2026

Beyond In-Memory Graphs: Query, Share, and Scale Your Agent Graph Projection with FalkorDB and ActiveGraph https://www.falkordb.com/blog/bey…

AgentsDGX agent

FalkorDB and ActiveGraph enable querying, sharing, and scaling of agent graph projections beyond in-memory limitations, allowing developers to persist and manage complex agent architectures in a graph

CoreWeave debuts ARIA agent to automate AI research in Weights & Biases

AgentsDGX agent

Artificial intelligence cloud operator CoreWeave Inc. today launched ARIA, an AI research agent built into the Weights & Biases platform. The agent reads experiment data and surfaces insights research

Straiker lands $64M to defend enterprise AI agents from attack

AgentsDGX agent

Agentic security company Straiker Inc. today revealed that it has raised 64 million in new funding to expand a platform built that secures the artificial intelligence agents now spreading across enter

When Multi-Robot Systems Meet Agentic AI:Towards Embodied Collective Intelligence

AgentsDGX agent

arXiv:2606.27929v1 Announce Type: new Abstract: Embodied AI is increasingly becoming agentic, shifting robots from perception--control pipelines towards closed-loop systems that can retrieve context,

27 Jun 2026

langchain is one of the few teams that actually explains the engineering behind agent concepts instead of just naming them. loop engineering…

AgentsDGX agent

langchain is one of the few teams that actually explains the engineering behind agent concepts instead of just naming them. loop engineering, managed agents, memory systems. every time a new concept d

25 Jun 2026

AI Snitches Get Glitches: Towards Evading Agentic Surveillance

AgentsDGX agent

arXiv:2606.25836v1 Announce Type: new Abstract: To better assist users with completing challenging tasks, AI agents mediate communications, access data, and interact with different APIs. Many employer

24 Jun 2026

Critique of Agent Model

SafetyDGX agent

arXiv:2606.23991v1 Announce Type: new Abstract: What is an agent? What constitutes agency? With the rise of Large Language Model (LLM) systems marketed as ``coding agents'', ``AI co-scientists'', and

ReM-MoA: Reasoning Memory Sustains Mixture-of-Agents Scaling

AgentsDGX agent

arXiv:2606.24437v1 Announce Type: new Abstract: Mixture-of-Agents (MoA) architectures improve inference-time scaling by organizing multiple LLM agents into layered reasoning pipelines. However, existi

23 Jun 2026

From Discrete Plans to Real-World Execution: A World-Model-Driven Framework for Execution-Aware Multi-Agent Path Finding

AgentsDGX agent

arXiv:2511.21886v2 Announce Type: replace Abstract: Multi-Agent Path Finding (MAPF) studies how to coordinate multiple agents to reach their goals without collisions and underpins a range of large-sca

PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate

AgentsDGX agent

arXiv:2606.20621v1 Announce Type: cross Abstract: Multi-agent debate improves the reliability of large language models (LLMs) through iterative peer critiques. However, fixed topologies often introduc

19 Jun 2026

Hermes Agent v0.17.0 - The Reach Release Changelog below:

AgentsDGX agent

Hermes Agent v0.17.0, released by Nous Research, is a major update referred to as 'The Reach Release' that introduces new capabilities and improvements to the Hermes Agent framework. The changelog det

11 Jun 2026

Preregistration for Experiments with AI Agents

AgentsDGX agent

arXiv:2606.11217v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and autonomous AI agents has given rise to a rapidly growing methodological paradigm: 'in silico' be

Search Discipline for Long-Horizon Research Agents

AgentsDGX agent

arXiv:2606.11522v1 Announce Type: new Abstract: Autoresearch agents now propose, evaluate, and select scientific candidates against a metric, and that metric is usually an aggregate reduced over a het

10 Jun 2026

Harnessing the Collective Intelligence of AI Agents in the Wild for New Discoveries

AgentsDGX agent

arXiv:2606.10402v1 Announce Type: cross Abstract: Scientific discovery is often a collective process: researchers share partial results, inspect failed attempts, and build on each other's ideas over l

Less Context, Better Agents: Efficient Context Engineering for Long-Horizon Tool-Using LLM Agents

Model ReleasesDGX agent

arXiv:2606.10209v1 Announce Type: new Abstract: Large language models deployed as autonomous agents for enterprise workflows face a key challenge: verbose tool responses from enterprise systems can ca

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

AgentsDGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

What makes a harness a harness: necessary and sufficient conditions for an agent harness

Model ReleasesDGX agent

arXiv:2606.10106v1 Announce Type: cross Abstract: The term agent harness now circulates widely in software engineering with generative artificial intelligence. It names the layer that wraps a language

9 Jun 2026

A Survey on Large Language Model-Based Game Agents

AgentsDGX agent

arXiv:2404.02039v5 Announce Type: replace Abstract: Game environments provide rich, controllable settings that stimulate many aspects of real-world complexity. As such, game agents offer a valuable te

Anything2Skill: Compiling External Knowledge into Reusable Skills for Agents

AgentsDGX agent

arXiv:2606.09316v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enables agents to access external knowledge at inference time, but it primarily retrieves fragmented declarative ev

Collaborative Human-Agent Protocol (CHAP)

AgentsDGX agent

arXiv:2606.09751v1 Announce Type: new Abstract: Foundation models are moving from response generation into operational roles. They plan across steps, call tools, request human input, coordinate with o

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

Model ReleasesDGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

Prisma-World: Camera-Controllable Multi-Agent Video World Model

SafetyDGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey

AgentsDGX agent

arXiv:2606.08500v1 Announce Type: cross Abstract: Software engineering agents (SWE agents) increasingly work through tool-mediated trajectories in real repositories, yet their behavior remains difficu

RAILS: Verification-Native Clearing For Agentic Commerce

AgentsDGX agent

arXiv:2606.08790v1 Announce Type: new Abstract: Autonomous agents negotiate, purchase, deploy code, and move funds, but no neutral mechanism determines whether they met their delegated obligation, who

Voting Protocols as Coordination Mechanisms for Role-Constrained Multi-Agent Tutoring Systems

AgentsDGX agent

arXiv:2606.08030v1 Announce Type: cross Abstract: Agentic tutoring systems introduce a coordination challenge: multiple agents may propose different but reasonable interventions, yet only one response

8 Jun 2026

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

AgentsDGX agent

arXiv:2606.06525v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text genera

Agentopia: Long-Term Life Simulation and Learning in Agent Societies

AgentsDGX agent

arXiv:2606.07513v1 Announce Type: new Abstract: Humans learn from social life. Simulating this process with LLM-powered agents represents a promising research direction, raising a natural question: wh

Dual Latent Memory for Visual Multi-agent System

AgentsDGX agent

arXiv:2602.00471v2 Announce Type: replace Abstract: While Visual Multi-Agent Systems (VMAS) promise to enhance comprehensive abilities through inter-agent collaboration, empirical evidence reveals a c

Skill-3D: Evolving Scene-Aware Skills for Agentic 3D Spatial Reasoning

Model ReleasesDGX agent

arXiv:2606.07436v1 Announce Type: new Abstract: This paper explores agentic 3D spatial understanding, i.e., MLLM agents performing 3D reasoning through tool use. Existing methods often misuse tools an

7 Jun 2026

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1…

AgentsDGX agent

imo there’s a pretty solid default recipe that everyone should use to optimize a system of Agent = Model + Harness you should “train” both 1. Build v1 agent using a sensible base harness and some task

6 Jun 2026

CaMeLs Can Use Computers Too: System-level Security for Computer Use Agents

AgentsDGX agent

arXiv:2601.09923v3 Announce Type: replace Abstract: AI agents are vulnerable to prompt injection attacks, where malicious content hijacks agent behavior. Among proposed defenses, architectural isolati

5 Jun 2026

Thousand Token Wood: shipping a multi-agent economy on a 3B model

AgentsDGX agent

Thousand Token Wood is a multi-agent economy simulation built on a 3 billion parameter language model, demonstrating how small models can power complex interactive systems with multiple agents. The pr

4 Jun 2026

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline

AgentsDGX agent

arXiv:2606.04315v1 Announce Type: new Abstract: LLM agents accumulate histories that outgrow their context windows, motivating a growing literature on memory systems. Yet most existing designs are tun

3 Jun 2026

Bring your favorite agent into Devin Desktop using ACP: https://docs.devin.ai/desktop/acp

AgentsDGX agent

Devin Desktop now supports Agent Control Protocol (ACP), allowing users to integrate their preferred AI agents into the platform. This feature enables users to bring custom or third-party agents into

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

AgentsDGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

2 Jun 2026

Agent Skills Should Go Beyond Text: The Case for Visual Skills

AgentsDGX agent

arXiv:2606.01414v1 Announce Type: new Abstract: Reusable skills are a key mechanism for extending agent capabilities, allowing agents to accumulate experience and solve increasingly complex tasks. Yet

Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations

Model ReleasesDGX agent

arXiv:2606.00832v1 Announce Type: new Abstract: Recent advances in agentic AI have enabled agents to complete complex tasks through tool use, reasoning, and multi-step planning. Yet existing benchmark

SS-ZKR: Spatial-Semantic Zero-Knowledge Routing for Privacy-Preserving Multi-Agent Collaboration

SafetyDGX agent

arXiv:2606.00962v1 Announce Type: cross Abstract: Foundational agent interoperability standards, notably the Agent-to-Agent (A2A) protocol and the Model Context Protocol (MCP), have advanced multi-age

Tracking the Behavioral Trajectories of Adapting Agents

AgentsDGX agent

arXiv:2606.02536v1 Announce Type: new Abstract: Text files such as skill files, memory files, and behavioral configuration files play a central role in defining how modern agents act. Through edits by

1 Jun 2026

Dreaming Of Others: Latent Teammate Modeling In World Models For Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.31361v1 Announce Type: cross Abstract: In cooperative multi-agent reinforcement learning (MARL), agents must coordinate with partners whose internal policies and intentions are not directly

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration

AgentsDGX agent

arXiv:2605.31365v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to promising progress in web agents. However, existing web agents often rely on han

29 May 2026

agent builder!

AgentsDGX agent

agent builder! Start creating agents using everyday language with LangSmith Fleet. Learn how to build no-code agents for real work. Take our free LangChain Academy course today: https://academy.langch

E-valuator: Reliable Agent Verifiers with Sequential Hypothesis Testing

AgentsDGX agent

arXiv:2512.03109v2 Announce Type: replace-cross Abstract: Agentic AI systems execute a sequence of actions, such as reasoning steps or tool calls, in response to a user prompt. To evaluate the success

Enhancing Multi-Agent Communication through Attention Steering with Context Relevance

AgentsDGX agent

arXiv:2605.30136v1 Announce Type: new Abstract: LLM-based multi-agent systems have demonstrated remarkable performance on complex tasks through collaborative reasoning. However, these systems tend to

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs

AgentsDGX agent

arXiv:2605.29512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for social and strategic reasoning over extended intera

28 May 2026

A Policy-Driven Runtime Layer for Agentic LLM Serving

SafetyDGX agent

arXiv:2605.27744v1 Announce Type: new Abstract: Multi-agent LLM systems have become the dominant production workload, but the serving stack was not built for them. The agent framework above knows agen

Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet Named

AgentsDGX agent

arXiv:2605.27396v1 Announce Type: cross Abstract: Autonomous AI agents now plan, decide, and act on behalf of users across healthcare, financial services, and workplace contexts, often without step-by

← Previous
1…3031323334…296
Next →