AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
Agents

Group of Skills: Group-Structured Skill Retrieval for Agent Skill Libraries

DGX agent

arXiv:2605.06978v1 Announce Type: cross Abstract: Skill-augmented agents increasingly rely on large reusable skill libraries, but retrieving relevant skills is not the same as presenting usable contex

agentsarxiv-cs-ai
11 May 2026
Local Ai
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

DGX agent

arXiv:2605.07637v1 Announce Type: new Abstract: Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning problems, where multiple homogeneous agents move simulta

local-aiarxiv-cs-ai
11 May 2026
Model Releases

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

DGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

model-releasesarxiv-cs-ai
11 May 2026
Agents

TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples

DGX agent

arXiv:2605.07935v1 Announce Type: new Abstract: We present TraceFix, a verification-first pipeline for Large Language Model (LLM) multi-agent coordination. An agent synthesizes a protocol topology as

agentsarxiv-cs-ai
11 May 2026
Agents

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

DGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

agentsarxiv-cs-ai
11 May 2026
Safety

Why Does Agentic Safety Fail to Generalize Across Tasks?

DGX agent

arXiv:2605.06992v1 Announce Type: new Abstract: AI agents are increasingly deployed in multi-task settings, where the task to perform is specified at test time, and the agent must generalize to unseen

safetyarxiv-cs-lg
11 May 2026
Agents

Join the Nous Research team for another Hermes Agent Jam in our Discord This one will be an interactive session, so come prepared to discuss…

DGX agent

Nous Research is hosting an interactive Hermes Agent Jam session in their Discord community where participants can discuss and collaborate on Hermes agent development. The event invites community memb

agentsnous-research--x
10 May 2026
Agents

'OncoAgent: A Dual-Tier Multi-Agent Framework for Privacy-Preserving Oncology Clinical Decision Support'

DGX agent

OncoAgent is a dual-tier multi-agent framework designed to provide clinical decision support in oncology while maintaining patient privacy. The system likely leverages multiple specialized AI agents w

agentshugging-face
9 May 2026
Agents

Signals V2: an LLM-free analyzer to scores live agent trajectories as OpenTelemetry spans.

DGX agent

Signals V2 is an LLM-free analyzer tool that evaluates live agent execution trajectories by scoring them as OpenTelemetry spans, enabling developers to monitor and trace AI agent behavior without rely

agentsr-ollama
9 May 2026
Agents

The future of Math is mathematicians and AI agents working together. Very pleased to introduce @GoogleDeepMind's AI co-mathematician: a mult…

DGX agent

The future of Math is mathematicians and AI agents working together. Very pleased to introduce @GoogleDeepMind's AI co-mathematician: a multi-agent system designed to actively collaborate with human e

agentsgary-marcus--x
8 May 2026
Agents

Improving token efficiency in GitHub Agentic Workflows

DGX agent

Agentic workflows that run on every pull request can quietly accumulate large API bills. Here's how we instrumented our own production workflows, found the inefficiencies, and built agents to fix them

agentsgithub-ai-blog
7 May 2026
Agents

New course: Build agents that respond to users with not only plaintext, but custom UIs like charts, forms, and whiteboards, generated on dem…

DGX agent

New course: Build agents that respond to users with not only plaintext, but custom UIs like charts, forms, and whiteboards, generated on demand and displayed right in the chat. This short course is bu

agentsandrew-ng--x
7 May 2026
Agents

Pact: A Choreographic Language for Agentic Ecosystems

DGX agent

arXiv:2605.03143v1 Announce Type: cross Abstract: Recent advances in large language models have led to the rise of software systems (i.e. agents) that execute with increasing autonomy on behalf of use

agentsarxiv-cs-ai
7 May 2026
Agents

HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness

DGX agent

arXiv:2605.02396v1 Announce Type: new Abstract: Recent advances in agentic harness with orchestration frameworks that coordinate multiple agents with memory, skills, and tool use have achieved remarka

agentsarxiv-cs-ai
6 May 2026
Agents

Hybrid Inspection and Task-Based Access Control in Zero-Trust Agentic AI

DGX agent

arXiv:2605.02682v1 Announce Type: new Abstract: Authorizing Large Language Model (LLM)-driven agents to dynamically invoke tools and access protected resources introduces significant security risks, a

agentsarxiv-cs-ai
6 May 2026
Agents

Less Interaction But More Explanation: A Communication Perspective on Agentic AI Interfaces

DGX agent

arXiv:2605.01610v1 Announce Type: cross Abstract: AI systems have long been expected to interact with users, answering questions, generating content, and continuing (social) conversations. Agentic AI,

agentsarxiv-cs-ai
6 May 2026
Agents

Lifting Traces to Logic: Programmatic Skill Induction with Neuro-Symbolic Learning for Long-Horizon Agentic Tasks

DGX agent

arXiv:2605.01293v1 Announce Type: new Abstract: Foundation model-driven agents often struggle with long-horizon planning due to the transient nature of purely prompting-based reasoning. While existing

agentsarxiv-cs-ai
6 May 2026
Agents

OpenSeeker-v2: Pushing the Limits of Search Agents with Informative and High-Difficulty Trajectories

DGX agent

arXiv:2605.04036v1 Announce Type: cross Abstract: Deep search capabilities have become an indispensable competency for frontier Large Language Model (LLM) agents, yet their development remains dominat

agentsarxiv-cs-cl
6 May 2026
Agents

according to google trends, babyagi is what made people start caring about ai agents 😎

DGX agent

BabyAGI, an AI agent project, generated significant public interest and search attention according to Google Trends data, potentially catalyzing broader awareness of AI agents as a category. The obser

agentsyohei-nakajima--x
5 May 2026
Safety

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

DGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

safetyarxiv-cs-ai
5 May 2026
Agents

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguabl…

DGX agent

open-weight LLMs have come a long way on agent tasks! but the harness you wrap them in matters just as much as the model itself, and arguably the interface you use to drive that harness matters even m

agentsharrison-chase--x
5 May 2026
Agents

Optimized and kinematically feasible multi-agent motion planning

DGX agent

arXiv:2605.01996v1 Announce Type: new Abstract: Multi-agent motion planning (MAMP) is an important problem for autonomous systems with multiple agents. In this work we propose a two-step method for fi

agentsarxiv-cs-ro
5 May 2026
Model Releases

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning

DGX agent

arXiv:2605.01489v1 Announce Type: cross Abstract: Frontier scientific reasoning is rapidly emerging as a key foundation for advancing AI agents in automated scientific discovery. Deep research agents

model-releasesarxiv-cs-cl
5 May 2026
Agents

Agent-guided workflows to accelerate model customization in Amazon SageMaker AI

DGX agent

Amazon SageMaker AI now offers an agentic experience that changes this. Developers describe their use case using natural language, and the AI coding agent streamlines the entire journey, from use case

agentsaws-ml-blog
4 May 2026
Agents

LangGraph is battle tested against production weirdness. For example you can update an agent while it is running without breaking active thr…

DGX agent

LangGraph is battle tested against production weirdness. For example you can update an agent while it is running without breaking active threads Deep Agents inherits this resilience because it runs on

agentsharrison-chase--x
4 May 2026
Agents

Position: agentic AI orchestration should be Bayes-consistent

DGX agent

arXiv:2605.00742v1 Announce Type: cross Abstract: LLMs excel at predictive tasks and complex reasoning tasks, but many high-value deployments rely on decisions under uncertainty, for example, which to

agentsarxiv-cs-lg
4 May 2026
Agents

this is the coolest demo I've seen a while, a must-watch for anyone interested in human-agent interaction!

DGX agent

this is the coolest demo I've seen a while, a must-watch for anyone interested in human-agent interaction! What if chat is the wrong interface for managing agents? In this talk, @steveruizok shows wha

agentsswyx--x
4 May 2026
Agents

Features - Multi-agent: each task runs on a specialized profile, with its own tools, skills, and personality. - Linked tasks: parent → child…

DGX agent

Features - Multi-agent: each task runs on a specialized profile, with its own tools, skills, and personality. - Linked tasks: parent → child dependencies. Fan out work, gather results, continue. - Sha

agentsnous-research--x
3 May 2026
Agents

@SakanaAILabs Sakana Fugu: A Multi-Agent Orchestration System as a Foundation Model https://sakana.ai/fugu-beta/

DGX agent

Sakana AI Labs has developed Fugu, a multi-agent orchestration system designed to function as a foundation model. The system likely enables coordinated interaction between multiple AI agents to improv

agentsdavid-ha--x
3 May 2026
Agents

great intro to agent harnesses and harness engineering!

DGX agent

great intro to agent harnesses and harness engineering! Is agent harness engineering the new prompt engineering? @LangChain co-founder and CEO, @hwchase17, joined us at #GoogleCloudNext to explain why

agentsharrison-chase--x
2 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Agents

here's a deep dive on how middleware lets you customize your agent harness, excellent writeup by @Vtrivedy10 !! deepagents offers a powerful…

DGX agent

here's a deep dive on how middleware lets you customize your agent harness, excellent writeup by @Vtrivedy10 !! deepagents offers a powerful base harness that you can customize for your use case! crea

agentsharrison-chase--x
1 May 2026
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
Agents

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to …

DGX agent

It is honestly shocking how little code can get you so far within this base primitive. create_agent is one of the most fun things for me to show people who are trying to get started here - because it

agentsharrison-chase--x
1 May 2026
Agents

Microsoft wants lawyers to trust its new AI agent in Word documents

DGX agent

Microsoft is launching a new AI agent inside Word that's specifically designed for legal teams. Legal Agent handles document edits, negotiation history, and complex documents to help legal teams handl

agentsthe-verge-ai
1 May 2026
Agents

Modeling Clinical Concern Trajectories in Language Model Agents

DGX agent

arXiv:2604.27872v1 Announce Type: new Abstract: Large language model (LLM) agents deployed in clinical settings often exhibit abrupt, threshold-driven behavior, offering little visibility into accumul

agentsarxiv-cs-ai
1 May 2026
Agents

Rethinking Agentic Reinforcement Learning In Large Language Models

DGX agent

arXiv:2604.27859v1 Announce Type: new Abstract: Reinforcement Learning (RL) has traditionally focused on training specialized agents to optimize predefined reward functions within narrowly defined env

agentsarxiv-cs-ai
1 May 2026
Safety

Safe Bilevel Delegation (SBD): A Formal Framework for Runtime Delegation Safety in Multi-Agent Systems

DGX agent

arXiv:2604.27358v1 Announce Type: new Abstract: As large language model (LLM) agents are deployed in high-stakes environments, the question of how safely to delegate subtasks to specialized sub-agents

safetyarxiv-cs-ai
1 May 2026
Agents

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

DGX agent

arXiv:2604.27464v1 Announce Type: cross Abstract: Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, in

agentsarxiv-cs-ai
1 May 2026
Agents

Self-Evolving Software Agents

DGX agent

arXiv:2604.27264v1 Announce Type: cross Abstract: Autonomous agents can adapt their behaviour to changing environments, but remain bound to requirements, goals, and capabilities fixed at design time,

agentsarxiv-cs-ai
1 May 2026
Agents

A Deep Dive into Deep Agents

DGX agent

A Deep Dive into Deep Agents 🚀DeepAgents deploy is a simple, configuration driven way to get an agent harness deployed to the cloud deepagents.toml is the file that configures it. It has four sections

agentsharrison-chase--x
30 Apr 2026
Agents

A Survey of Multi-Agent Deep Reinforcement Learning with Graph Neural Network-Based Communication

DGX agent

arXiv:2604.25972v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), the integration of a communication mechanism, allowing agents to better learn to coordinate their action

agentsarxiv-cs-ai
30 Apr 2026
Agents

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

DGX agent

arXiv:2604.26752v1 Announce Type: new Abstract: We present GLM-5V-Turbo, a step toward native foundation models for multimodal agents. As foundation models are increasingly deployed in real environmen

agentsarxiv-cs-cv
30 Apr 2026
Model Releases

LATTICE: Evaluating Decision Support Utility of Crypto Agents

DGX agent

arXiv:2604.26235v1 Announce Type: cross Abstract: We introduce LATTICE, a benchmark for evaluating the decision support utility of crypto agents in realistic user-facing scenarios. Prior crypto agent

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior…

DGX agent

resharing this note, find it helpful given all the great open evals work + teams building vertical agents Evals are a proxy for the behavior we want our agent to exhibit in production Model+Harness pu

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

FAMA: Failure-Aware Meta-Agentic Framework for Open-Source LLMs in Interactive Tool Use Environments

DGX agent

arXiv:2604.25135v1 Announce Type: new Abstract: Large Language Models are being increasingly deployed as the decision-making core of autonomous agents capable of effecting change in external environme

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

Interacting with the agent in the studio is the easiest way to get a user friendly UI! I can share it with others in my LangSmith org, and t…

DGX agent

The post discusses LangSmith's agent interaction capabilities in the studio environment, highlighting how users can access a user-friendly interface for agent testing and collaboration. It emphasizes

agentsharrison-chase--x
29 Apr 2026
Agents

Your hermes agent can make its own little live todo list now 😊 @NousResearch @Teknium

DGX agent

Nous Research has introduced a feature enabling Hermes agents to dynamically create and manage their own live to-do lists during task execution. This capability allows agents to better organize and tr

agentsnous-research--x
29 Apr 2026
← Previous
1…5556575859…371
Next →