AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
Model Releases

FALAT: Tracing Failures in LLM Agent Trajectories via Dependency-Guided Search

DGX agent

arXiv:2606.00765v1 Announce Type: new Abstract: LLM-based agents increasingly solve complex tasks through long trajectories involving reasoning steps, tool calls, and inter-agent communication. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Multi-Agent Computer Use

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

DGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

agentsarxiv-cs-ai
2 Jun 2026
Agents

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.29790v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) have emerged as an effective paradigm for complex and long-horizon tasks. However, in real-world tasks, MAS often

agentsarxiv-cs-ai
29 May 2026
Safety

ValueFlow: Measuring the Propagation of Value Perturbations in Multi-Agent LLM Systems

DGX agent

arXiv:2602.08567v2 Announce Type: replace-cross Abstract: Multi-agent large language model (LLM) systems increasingly consist of agents that observe and respond to one another's outputs. While value a

safetyarxiv-cs-cl
29 May 2026
Model Releases

AgentSociety: Incentivizing Agentic Social Intelligence

DGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

model-releasesarxiv-cs-ai
27 May 2026
Safety

Learning to Orchestrate Agents under Uncertainty

DGX agent

arXiv:2605.27073v1 Announce Type: new Abstract: Adaptive orchestration of heterogeneous agents requires making sequential delegation decisions under uncertain and evolving agent behaviour, e.g., coord

safetyarxiv-cs-lg
27 May 2026
Model Releases

UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.26646v1 Announce Type: new Abstract: LLM-based multi-agent systems decompose complex tasks into interacting roles, but most remain manually orchestrated by prompts, tools, and control rules

model-releasesarxiv-cs-ai
27 May 2026
Agents

Market Regime Council for Dynamic Credit Assignment in Multi-Agent LLM Decision Systems

DGX agent

arXiv:2605.24490v1 Announce Type: new Abstract: Multi-agent LLM decision systems for portfolio management still lack a principled way to assign credit across specialist agents, remain vulnerable to co

agentsarxiv-cs-ai
26 May 2026
Agents

datasette-agent 0.1a4

DGX agent

Release: datasette-agent 0.1a4 Taking advantage of the new makeJumpSections() JavaScript plugin hook added in Datasette 1.0a30, datasette-agent now presents this 'Start a new agent chat' interface as

agentssimon-willison
24 May 2026
Agents

datasette-agent-charts 0.1a2

DGX agent

datasette-agent-charts adds charts to Datasette Agent, powered by Observable Plot. The plugin extends Datasette Agent, which provides a conversational interface for asking questions about data in Data

agentssimon-willison
21 May 2026
Safety

Code as Agent Harness

DGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

safetyarxiv-cs-ai
19 May 2026
Agents

EXG: Self-Evolving Agents with Experience Graphs

DGX agent

arXiv:2605.17721v1 Announce Type: new Abstract: Large language model (LLM)-based agents have demonstrated strong capabilities in complex reasoning and problem solving through multi-step interactions,

agentsarxiv-cs-ai
19 May 2026
Agents

Great new paper to read: Code as Agent Harness (bookmark it)

DGX agent

Great new paper to read: Code as Agent Harness (bookmark it) // Code as Agent Harness // 100+ page report on all things related to agent harnesses. (bookmark it) In particular, the survey summarizes m

agentsdair-ai--x
19 May 2026
Local Ai

PPAI: Enabling Personalized LLM Agent Interoperability for Collaborative Edge Intelligence

DGX agent

arXiv:2605.18067v1 Announce Type: new Abstract: Deploying large language model (LLM) on edge device enables personalized LLM agents for various users. The growing availability of diverse personalized

local-aiarxiv-cs-cl
19 May 2026
Agents

PULSE: Agentic Investigation with Passive Sensing for Proactive Intervention in Cancer Survivorship

DGX agent

arXiv:2605.17679v1 Announce Type: cross Abstract: Cancer survivors face elevated rates of depression, anxiety, and general emotional distress, yet the precise moments they most need support are often

agentsarxiv-cs-ai
19 May 2026
Model Releases

The agentic era: Architecting the blueprint for mission impact across the public sector

DGX agent

This is a new era — the agentic era – and the question is no longer, “what’s possible?” but rather, “what creates impact?” Today, organizations across industries around the world are swiftly moving fr

model-releasesgoogle-cloud-ai
19 May 2026
Model Releases

GraphBit: A Graph-based Agentic Framework for Non-Linear Agent Orchestration

DGX agent

arXiv:2605.13848v1 Announce Type: new Abstract: Agentic LLM frameworks that rely on prompted orchestration, where the model itself determines workflow transitions, often suffer from hallucinated routi

model-releasesarxiv-cs-ai
15 May 2026
Agents

MMSkills: Towards Multimodal Skills for General Visual Agents

DGX agent

arXiv:2605.13527v1 Announce Type: new Abstract: Reusable skills have become a core substrate for improving agent capabilities, yet most existing skill packages encode reusable behavior primarily as te

agentsarxiv-cs-ai
14 May 2026
Agents

Reinforced Collaboration in Multi-Agent Flow Networks

DGX agent

arXiv:2605.12943v1 Announce Type: new Abstract: Multi-agent systems provide a powerful way to extend large language models (LLMs) by decomposing a complex task into specialized subtasks handled by dif

agentsarxiv-cs-lg
14 May 2026
Safety

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

DGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

safetyarxiv-cs-ai
12 May 2026
Model Releases

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation

DGX agent

arXiv:2605.09315v1 Announce Type: new Abstract: Recent advances in LLM agents enable systems that autonomously refine workflows, accumulate reusable skills, self-train their underlying models, and mai

model-releasesarxiv-cs-ai
12 May 2026
Agents

Engineering Robustness into Personal Agents with the AI Workflow Store

DGX agent

arXiv:2605.10907v1 Announce Type: cross Abstract: The dominant paradigm for AI agents is an 'on-the-fly' loop in which agents synthesize plans and execute actions within seconds or minutes in response

agentsarxiv-cs-ai
12 May 2026
Safety

Formal Policy Enforcement for Real-World Agentic Systems

DGX agent

arXiv:2602.16708v3 Announce Type: replace-cross Abstract: Security policy enforcement in contemporary agentic systems predominantly consists of embedding natural-language policies within an agent's sy

safetyarxiv-cs-ai
12 May 2026
Model Releases

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier f…

DGX agent

$NBIS announced a partnership with LangChain to integrate Nebius Token Factory with LangChain’s Deep Agents. The goal is to make it easier for teams building AI agents on LangChain to run those worklo

model-releasesharrison-chase--x
12 May 2026
Local Ai

PAAC: Privacy-Aware Agentic Device-Cloud Collaboration

DGX agent

arXiv:2605.08646v1 Announce Type: cross Abstract: Large language model (LLM) agents face a structural tension: cloud agents provide strong reasoning but expose user data, while on-device agents preser

local-aiarxiv-cs-cl
12 May 2026
Agents

ShadowMerge: A Novel Poisoning Attack on Graph-Based Agent Memory via Relation-Channel Conflicts

DGX agent

arXiv:2605.09033v1 Announce Type: cross Abstract: Graph-based agent memory is increasingly used in LLM agents to support structured long-term recall and multi-hop reasoning, but it also creates a new

agentsarxiv-cs-ai
12 May 2026
Agents

Agentic AI and the Industrialization of Cyber Offense: Forecast, Consequences, and Defensive Priorities for Enterprises and the Mittelstand

DGX agent

arXiv:2605.06713v1 Announce Type: cross Abstract: Agentic AI systems can plan, call tools, inspect code, interact with web applications, and coordinate multi-step workflows. These same capabilities ch

agentsarxiv-cs-ai
11 May 2026
Agents

if you've ever wanted a cofounder that would manage all your agents for you... try http://cofounder.co by @intelligenceco

DGX agent

Cofounder.co is a platform by Intelligence Co that provides AI agent management, allowing users to delegate agent oversight and coordination to an automated system rather than managing multiple agents

agentsyohei-nakajima--x
11 May 2026
Agents

Learning CLI Agents with Structured Action Credit under Selective Observation

DGX agent

arXiv:2605.08013v1 Announce Type: new Abstract: Command line interface (CLI) agents are emerging as a practical paradigm for agent-computer interaction over evolving filesystems, executable command li

agentsarxiv-cs-ai
11 May 2026
Agents

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

DGX agent

arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories

agentsarxiv-cs-ai
11 May 2026
Agents

What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic Curiosity

DGX agent

arXiv:2605.03782v1 Announce Type: new Abstract: To navigate partially observable visual environments, recent VLM agents increasingly internalize world modeling capabilities into their policies via exp

agentsarxiv-cs-ai
7 May 2026
Agents

fleet is built on top of deepagents, a model agnostic base harness! try building multi-model agents in fleet today, or build your own deep a…

DGX agent

fleet is built on top of deepagents, a model agnostic base harness! try building multi-model agents in fleet today, or build your own deep agent with the underlying sdk! Not every step in an agent wor

agentsharrison-chase--x
4 May 2026
Safety

AID: Agent Intent from Diffusion for Multi-Agent Informative Path Planning

DGX agent

arXiv:2512.02535v2 Announce Type: replace Abstract: Information gathering in large-scale or time-critical scenarios (e.g., environmental monitoring, search and rescue) requires broad coverage within l

safetyarxiv-cs-ro
1 May 2026
Agents

many sensitive Agent Workloads today require some sort of human feedback LangGraph supplies the runtime primitives for LangChain + Deep Agen…

DGX agent

many sensitive Agent Workloads today require some sort of human feedback LangGraph supplies the runtime primitives for LangChain + Deep Agents and makes it easy to durably pause, resume, and replay ag

agentsharrison-chase--x
1 May 2026
Agents

most of the time, you want an agent loop to run uninterrupted. that's where the utility comes from! but some decisions shouldn't be delegate…

DGX agent

most of the time, you want an agent loop to run uninterrupted. that's where the utility comes from! but some decisions shouldn't be delegated to the agent. two situations come up consistently: 1/ befo

agentsharrison-chase--x
1 May 2026
Agents

Hermes Agent is the go-to for complex creative workflows

DGX agent

Hermes Agent is the go-to for complex creative workflows i just spent $10,000 on AI credits building out my project for @AIPSummit and i can't be thankful enough to @NousResearch's Hermes harness and

agentsnous-research--x
29 Apr 2026
Agents

Leverage Laws: A Per-Task Framework for Human-Agent Collaboration

DGX agent

arXiv:2604.25040v1 Announce Type: cross Abstract: We propose a per-task leverage ratio for human-agent collaboration: human work displaced by an agent, divided by the human time required to specify th

agentsarxiv-cs-cl
29 Apr 2026
Agents

Agentic Adversarial Rewriting Exposes Architectural Vulnerabilities in Black-Box NLP Pipelines

DGX agent

arXiv:2604.23483v1 Announce Type: new Abstract: Multi-component natural language processing (NLP) pipelines are increasingly deployed for high-stakes decisions, yet no existing adversarial method can

agentsarxiv-cs-ai
28 Apr 2026
Agents

At @ListenLabs, agents analyze data from conversations by building their own tables — adding columns like 'user sentiment,' filling in value…

DGX agent

At @ListenLabs, agents analyze data from conversations by building their own tables — adding columns like 'user sentiment,' filling in values automatically, then charting the results. Co-Founder & CTO

agentsharrison-chase--x
28 Apr 2026
Safety

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

DGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

safetyarxiv-cs-lg
28 Apr 2026
Agents

FormalScience: Scalable Human-in-the-Loop Autoformalisation of Science with Agentic Code Generation in Lean

DGX agent

arXiv:2604.23002v1 Announce Type: new Abstract: Formalising informal mathematical reasoning into formally verifiable code is a significant challenge for large language models. In scientific fields suc

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Mind the Gap: Evaluating Model- and Agentic-Level Vulnerabilities in LLMs with Action Graphs

DGX agent

arXiv:2509.04802v3 Announce Type: replace Abstract: As large language models increasingly deployed into agentic systems, existing methods face critical gaps in observing, assessing, and mitigating dep

model-releasesarxiv-cs-cl
28 Apr 2026
Agents

Phenom adds Plum psychometric science to its agentic AI hiring stack

DGX agent

Artificial intelligence-based human resources company Phenom People Inc. announced today that it has acquired Plum.io Inc., a psychometric-based talent assessments company that measures the durable sk

agentssiliconangle
28 Apr 2026
Agents

The Last Human-Written Paper: Agent-Native Research Artifacts

DGX agent

arXiv:2604.24658v1 Announce Type: new Abstract: Scientific publication compresses a branching, iterative research process into a linear narrative, discarding the majority of what was discovered along

agentsarxiv-cs-lg
28 Apr 2026
Agents

The most powerful real-time visual tool in creative coding also has the steepest learning curve Now your Hermes agent can just run TouchDesi…

DGX agent

The most powerful real-time visual tool in creative coding also has the steepest learning curve Now your Hermes agent can just run TouchDesigner for you. Video credit: made by @macbethAI, a talented A

agentsnous-research--x
28 Apr 2026
Agents

How real-time data pipelines are giving AI agents something worth acting on

DGX agent

As enterprises race to wire AI into their operations, the infrastructure bottleneck has shifted from model capability to data access — and the enterprises winning the race are those treating real-time

agentssiliconangle
27 Apr 2026
Agents

ml-intern from @huggingface is such a cool repository, an 🤖 agentic harness that allows you to wield the larger🤗ecosystem to run all the t…

DGX agent

ml-intern from @huggingface is such a cool repository, an 🤖 agentic harness that allows you to wield the larger🤗ecosystem to run all the tasks a ml researcher would. it's all one can ask for as a side

agentsclem-delangue--x
27 Apr 2026
← Previous
1…3435363738…370
Next →