AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
2 Jun 2026

FALAT: Tracing Failures in LLM Agent Trajectories via Dependency-Guided Search

Model ReleasesDGX agent

arXiv:2606.00765v1 Announce Type: new Abstract: LLM-based agents increasingly solve complex tasks through long trajectories involving reasoning steps, tool calls, and inter-agent communication. Howeve

Multi-Agent Computer Use

Model ReleasesDGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

AgentsDGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
29 May 2026

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.29790v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) have emerged as an effective paradigm for complex and long-horizon tasks. However, in real-world tasks, MAS often

ValueFlow: Measuring the Propagation of Value Perturbations in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2602.08567v2 Announce Type: replace-cross Abstract: Multi-agent large language model (LLM) systems increasingly consist of agents that observe and respond to one another's outputs. While value a

27 May 2026

AgentSociety: Incentivizing Agentic Social Intelligence

Model ReleasesDGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

Learning to Orchestrate Agents under Uncertainty

SafetyDGX agent

arXiv:2605.27073v1 Announce Type: new Abstract: Adaptive orchestration of heterogeneous agents requires making sequential delegation decisions under uncertain and evolving agent behaviour, e.g., coord

UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.26646v1 Announce Type: new Abstract: LLM-based multi-agent systems decompose complex tasks into interacting roles, but most remain manually orchestrated by prompts, tools, and control rules

26 May 2026

Market Regime Council for Dynamic Credit Assignment in Multi-Agent LLM Decision Systems

AgentsDGX agent

arXiv:2605.24490v1 Announce Type: new Abstract: Multi-agent LLM decision systems for portfolio management still lack a principled way to assign credit across specialist agents, remain vulnerable to co

24 May 2026

datasette-agent 0.1a4

AgentsDGX agent

Release: datasette-agent 0.1a4 Taking advantage of the new makeJumpSections() JavaScript plugin hook added in Datasette 1.0a30, datasette-agent now presents this 'Start a new agent chat' interface as

21 May 2026

datasette-agent-charts 0.1a2

AgentsDGX agent

datasette-agent-charts adds charts to Datasette Agent, powered by Observable Plot. The plugin extends Datasette Agent, which provides a conversational interface for asking questions about data in Data

19 May 2026

Code as Agent Harness

SafetyDGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

EXG: Self-Evolving Agents with Experience Graphs

AgentsDGX agent

arXiv:2605.17721v1 Announce Type: new Abstract: Large language model (LLM)-based agents have demonstrated strong capabilities in complex reasoning and problem solving through multi-step interactions,

Great new paper to read: Code as Agent Harness (bookmark it)

AgentsDGX agent

Great new paper to read: Code as Agent Harness (bookmark it) // Code as Agent Harness // 100+ page report on all things related to agent harnesses. (bookmark it) In particular, the survey summarizes m

PPAI: Enabling Personalized LLM Agent Interoperability for Collaborative Edge Intelligence

Local AiDGX agent

arXiv:2605.18067v1 Announce Type: new Abstract: Deploying large language model (LLM) on edge device enables personalized LLM agents for various users. The growing availability of diverse personalized

PULSE: Agentic Investigation with Passive Sensing for Proactive Intervention in Cancer Survivorship

AgentsDGX agent

arXiv:2605.17679v1 Announce Type: cross Abstract: Cancer survivors face elevated rates of depression, anxiety, and general emotional distress, yet the precise moments they most need support are often

The agentic era: Architecting the blueprint for mission impact across the public sector

Model ReleasesDGX agent

This is a new era — the agentic era – and the question is no longer, “what’s possible?” but rather, “what creates impact?” Today, organizations across industries around the world are swiftly moving fr

15 May 2026

GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration

Model ReleasesDGX agent

arXiv:2605.13848v1 Announce Type: new Abstract: Agentic LLM frameworks that rely on prompted orchestration, where the model itself determines workflow transitions, often suffer from hallucinated routi

14 May 2026

MMSkills: Towards Multimodal Skills for General Visual Agents

AgentsDGX agent

arXiv:2605.13527v1 Announce Type: new Abstract: Reusable skills have become a core substrate for improving agent capabilities, yet most existing skill packages encode reusable behavior primarily as te

Reinforced Collaboration in Multi-Agent Flow Networks

AgentsDGX agent

arXiv:2605.12943v1 Announce Type: new Abstract: Multi-agent systems provide a powerful way to extend large language models (LLMs) by decomposing a complex task into specialized subtasks handled by dif

12 May 2026

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

SafetyDGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation

Model ReleasesDGX agent

arXiv:2605.09315v1 Announce Type: new Abstract: Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and mai

Engineering Robustness into Personal Agents with the AI Workflow Store

AgentsDGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

Formal Policy Enforcement for Real-World Agentic Systems

SafetyDGX agent

arXiv:2602.16708v3 Announce Type: replace-cross Abstract: Security policy enforcement in contemporary agentic systems predominantly consists of embedding natural-language policies within an agent's sy

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier f…

Model ReleasesDGX agent

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier for teams building AI agents on LangChain to run those worklo

PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

Local AiDGX agent

arXiv:2605.08646v1 Announce Type: cross Abstract: Large language model (LLM) agents face a structural tension: cloud agents provide strong reasoning but expose user data, while on-device agents preser

ShadowMerge: A Novel Poisoning Attack on Graph-Based Agent Memory via Relation-Channel Conflicts

AgentsDGX agent

arXiv:2605.09033v1 Announce Type: cross Abstract: Graph-based agent memory is increasingly used in LLM agents to support structured long-term recall and multi-hop reasoning, but it also creates a new

11 May 2026

Agentic AI and the Industrialization of Cyber Offense: Forecast, Consequences, and Defensive Priorities for Enterprises and the Mittelstand

AgentsDGX agent

arXiv:2605.06713v1 Announce Type: cross Abstract: Agentic AI systems can plan, call tools, inspect code, interact with web applications, and coordinate multi-step workflows. These same capabilities ch

if you've ever wanted a cofounder that would manage all your agents for you... try http://cofounder.co by @intelligenceco

AgentsDGX agent

Cofounder.co is a platform by Intelligence Co that provides AI agent management, allowing users to delegate agent oversight and coordination to an automated system rather than managing multiple agents

Learning CLI Agents with Structured Action Credit under Selective Observation

AgentsDGX agent

arXiv:2605.08013v1 Announce Type: new Abstract: Command line interface (CLI) agents are emerging as a practical paradigm for agent-computer interaction over evolving filesystems, executable command li

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

AgentsDGX agent

arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories

7 May 2026

What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity

AgentsDGX agent

arXiv:2605.03782v1 Announce Type: new Abstract: To navigate partially observable visual environments, recent VLM agents increasingly internalize world modeling capabilities into their policies via exp

4 May 2026

fleet is built on top of deepagents, a model agnostic base harness! try building multi-model agents in fleet today, or build your own deep a…

AgentsDGX agent

fleet is built on top of deepagents, a model agnostic base harness! try building multi-model agents in fleet today, or build your own deep agent with the underlying sdk! Not every step in an agent wor

1 May 2026

AID: Agent Intent from Diffusion for Multi-Agent Informative Path Planning

SafetyDGX agent

arXiv:2512.02535v2 Announce Type: replace Abstract: Information gathering in large-scale or time-critical scenarios (e.g., environmental monitoring, search and rescue) requires broad coverage within l

many sensitive Agent Workloads today require some sort of human feedback LangGraph supplies the runtime primitives for LangChain + Deep Agen…

AgentsDGX agent

many sensitive Agent Workloads today require some sort of human feedback LangGraph supplies the runtime primitives for LangChain + Deep Agents and makes it easy to durably pause, resume, and replay ag

most of the time, you want an agent loop to run uninterrupted. that's where the utility comes from! but some decisions shouldn't be delegate…

AgentsDGX agent

most of the time, you want an agent loop to run uninterrupted. that's where the utility comes from! but some decisions shouldn't be delegated to the agent. two situations come up consistently: 1/ befo

29 Apr 2026

Hermes Agent is the go-to for complex creative workflows

AgentsDGX agent

Hermes Agent is the go-to for complex creative workflows i just spent $10,000 on AI credits building out my project for @AIPSummit and i can't be thankful enough to @NousResearch's Hermes harness and

Leverage Laws: A Per-Task Framework for Human-Agent Collaboration

AgentsDGX agent

arXiv:2604.25040v1 Announce Type: cross Abstract: We propose a per-task leverage ratio for human-agent collaboration: human work displaced by an agent, divided by the human time required to specify th

28 Apr 2026

Agentic Adversarial Rewriting Exposes Architectural Vulnerabilities in Black-Box NLP Pipelines

AgentsDGX agent

arXiv:2604.23483v1 Announce Type: new Abstract: Multi-component natural language processing (NLP) pipelines are increasingly deployed for high-stakes decisions, yet no existing adversarial method can

At @ListenLabs, agents analyze data from conversations by building their own tables — adding columns like 'user sentiment,' filling in value…

AgentsDGX agent

At @ListenLabs, agents analyze data from conversations by building their own tables — adding columns like 'user sentiment,' filling in values automatically, then charting the results. Co-Founder & CTO

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

SafetyDGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean

AgentsDGX agent

arXiv:2604.23002v1 Announce Type: new Abstract: Formalising informal mathematical reasoning into formally verifiable code is a significant challenge for large language models. In scientific fields suc

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

Model ReleasesDGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

Phenom adds Plum psychometric science to its agentic AI hiring stack

AgentsDGX agent

Artificial intelligence-based human resources company Phenom People Inc. announced today that it has acquired Plum.io Inc., a psychometric-based talent assessments company that measures the durable sk

The Last Human-Written Paper: Agent-Native Research Artifacts

AgentsDGX agent

arXiv:2604.24658v1 Announce Type: new Abstract: Scientific publication compresses a branching, iterative research process into a linear narrative, discarding the majority of what was discovered along

The most powerful real-time visual tool in creative coding also has the steepest learning curve Now your Hermes agent can just run TouchDesi…

AgentsDGX agent

The most powerful real-time visual tool in creative coding also has the steepest learning curve Now your Hermes agent can just run TouchDesigner for you. Video credit: made by @macbethAI, a talented A

27 Apr 2026

How real-time data pipelines are giving AI agents something worth acting on

AgentsDGX agent

As enterprises race to wire AI into their operations, the infrastructure bottleneck has shifted from model capability to data access — and the enterprises winning the race are those treating real-time

ml-intern from @huggingface is such a cool repository, an 🤖 agentic harness that allows you to wield the larger🤗ecosystem to run all the t…

AgentsDGX agent

ml-intern from @huggingface is such a cool repository, an 🤖 agentic harness that allows you to wield the larger🤗ecosystem to run all the tasks a ml researcher would. it's all one can ask for as a side

25 Apr 2026

it's literally 2 lines to upload your second brain to a private @huggingface bucket just ask your agent to us hf cli to create a private buc…

AgentsDGX agent

This post describes a simple two-line process for uploading a personal knowledge base or 'second brain' to a private Hugging Face bucket using the Hugging Face CLI, with the suggestion that an AI agen

24 Apr 2026

Breaking MCP with Function Hijacking Attacks: Novel Threats for Function Calling and Agentic Models

AgentsDGX agent

arXiv:2604.20994v1 Announce Type: cross Abstract: The growth of agentic AI has drawn significant attention to function calling Large Language Models (LLMs), which are designed to extend the capabiliti

Hermes Agent v0.11.0 - “The Interface Release” Full changelog below ↓

AgentsDGX agent

Hermes Agent v0.11.0, released by Nous Research, is a significant update focused on interface improvements and enhancements. This release, dubbed 'The Interface Release,' includes various updates and

Multi-Agent Empowerment and Emergence of Complex Behavior in Groups

AgentsDGX agent

arXiv:2604.21155v1 Announce Type: new Abstract: Intrinsic motivations are receiving increasing attention, i.e. behavioral incentives that are not engineered, but emerge from the interaction of an agen

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning

AgentsDGX agent

arXiv:2604.21190v1 Announce Type: new Abstract: Understanding visual scenes requires not only recognizing objects but also reasoning about their spatial relationships. Unlike general vision-language t

What is an agent harness?

AgentsDGX agent

A version of this article originally appeared on X. Someone asked me at a hacker event last week: “Can anyone actually tell me what a harness really is?” It was... The post What is an agent harness? a

23 Apr 2026

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

Model ReleasesDGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

If you're waiting for a sign... that might not be it! Mitigating Trust Boundary Confusion from Visual Injections on Vision-Language Agentic Systems

AgentsDGX agent

arXiv:2604.19844v1 Announce Type: cross Abstract: Recent advances in embodied Vision-Language Agentic Systems (VLAS), powered by large vision-language models (LVLMs), enable AI systems to perceive and

Salesforce and Google partner on agentic cross-platform collaboration

AgentsDGX agent

Salesforce Inc. said today it’s trying to break down the siloes that separate enterprise’s customer relationship management data from their productivity tools, specifically for artificial intelligence

22 Apr 2026

AgentDynEx: Nudging the Mechanics and Dynamics of Multi-Agent Simulations

AgentsDGX agent

arXiv:2504.09662v3 Announce Type: replace-cross Abstract: Multi-agent large language model simulations have the potential to model complex human behaviors and interactions. If the mechanics are set up

Best Agent Identification for General Game Playing

AgentsDGX agent

arXiv:2507.00451v2 Announce Type: replace-cross Abstract: We present an efficient and generalised procedure to accurately identify the best (or near best) performing algorithm for each sub-task in a m

Google announces innovations in mega-scale networking for the agentic era

AgentsDGX agent

Rising to meet demands for increasing latency and scale, Google LLC today introduced a mega-scale datacenter network fabric and cross-cloud infrastructure aimed at agentic artificial intelligence deli

← Previous
1…2728293031…296
Next →