AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
18 May 2026

flymyai is a simple cloud agent with tons of APIs/models baked in under one api key once you have that working, you can drop that into somet…

AgentsDGX agent

flymyai is a simple cloud agent with tons of APIs/models baked in under one api key once you have that working, you can drop that into something like lovable In 90 seconds I shipped the agent that tur

ICYMI: we shipped Deep Agents v0.6 last week, our biggest release yet!

AgentsDGX agent

LangChain released Deep Agents v0.6, described as their largest release to date. The announcement was made by Harrison Chase on X (formerly Twitter), indicating a significant update to the Deep Agents

some good discussions and experiments around stateful agents in the replies, but seems like we’re not quite there yet, as in we’re starting …

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

some good discussions and experiments around stateful agents in the replies, but seems like we’re not quite there yet, as in we’re starting to track memory and traces, but not quite agent capability a

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination

AgentsDGX agent

arXiv:2605.15207v1 Announce Type: new Abstract: Multi-agent LLM systems have shown promise for complex reasoning, yet recent evaluations reveal they often underperform single-model baselines. We ident

Toward Natural and Companionable Virtual Agents via Cross-Temporal Emotional Modeling

AgentsDGX agent

arXiv:2605.15812v1 Announce Type: cross Abstract: Recent advances in foundation models have enabled conversational agents that aim for sustained companionship rather than mere task completion. Yet mos

17 May 2026

Hermes Agent v0.14.0 - “The Foundation Release” Changelog below

AgentsDGX agent

Hermes Agent v0.14.0, released by Nous Research as 'The Foundation Release,' represents a significant update to their Hermes agent framework. This release likely includes core infrastructure improveme

Many AI agents in finance rely on extremely high quality context engineering from documents 📑 They can be roughly divided into two categori…

AgentsDGX agent

Many AI agents in finance rely on extremely high quality context engineering from documents 📑 They can be roughly divided into two categories: 1️⃣ Repetitive, operational work common in back-office us

This is one of the first real continual learning systems for agents in production. Not just monitoring. Actually getting better over time.

AgentsDGX agent

This is one of the first real continual learning systems for agents in production. Not just monitoring. Actually getting better over time. LangSmith Engine is how we’re spinning the always-on, self-im

15 May 2026

APWA: A Distributed Architecture for Parallelizable Agentic Workflows

AgentsDGX agent

arXiv:2605.15132v1 Announce Type: new Abstract: Autonomous multi-agent systems based on large language models (LLMs) have demonstrated remarkable abilities in independently solving complex tasks in a

Auditing Agent Harness Safety

Model ReleasesDGX agent

arXiv:2605.14271v1 Announce Type: new Abstract: LLM agents increasingly run inside execution harnesses that dispatch tools, allocate resources, and route messages between specialized components. Howev

Had the honor of sharing the stage with the one and only @sydneyrunkle at Interrupt 2026 to talk about Deep Agents. Sign-up for the Managed …

AgentsDGX agent

Had the honor of sharing the stage with the one and only @sydneyrunkle at Interrupt 2026 to talk about Deep Agents. Sign-up for the Managed Deep Agents waitlist today to get early access! https://www.

ICYMI: 1️⃣ LangSmith Engine 2️⃣ SmithDB 3️⃣ Managed Deep Agents 4️⃣ LangSmith Sandboxes: Now Generally Available 5️⃣ Context Hub 6️⃣ LangSmi…

AgentsDGX agent

ICYMI: 1️⃣ LangSmith Engine 2️⃣ SmithDB 3️⃣ Managed Deep Agents 4️⃣ LangSmith Sandboxes: Now Generally Available 5️⃣ Context Hub 6️⃣ LangSmith LLM Gateway 7️⃣ Sandboxes, Prebuilt agents, + free model

it almost never makes sense to use real api's for your evals. with how good coding agents have become, i will pretty much always opt to crea…

AgentsDGX agent

it almost never makes sense to use real api's for your evals. with how good coding agents have become, i will pretty much always opt to create a fake mock server for my agent to hit. the workflow is u

MIMIC-D: Multi-modal Imitation for MultI-agent Coordination with Decentralized Diffusion Policies

AgentsDGX agent

arXiv:2509.14159v3 Announce Type: replace Abstract: As robots become more integrated in society, their ability to coordinate with other robots and humans on multi-modal tasks (those with multiple vali

Quantum Advantage in Multi Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.14235v1 Announce Type: new Abstract: We present an empirical evaluation of quantum entanglement in agent coordination within quantum multi agent reinforcement learning (QMARL). While QMARL

TeachAnything: A Multimodal Crowdsourcing Platform for Training Embodied AI Agents in Symmetrical Reality

AgentsDGX agent

arXiv:2605.14556v1 Announce Type: new Abstract: Symmetrical Reality (SR) is emerging as a future trend for human-agent coexistence, placing higher demands on agents to acquire human-like intelligence.

WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections

AgentsDGX agent

arXiv:2605.15030v1 Announce Type: cross Abstract: Web agents can autonomously complete online tasks by interacting with websites, but their exposure to open web environments makes them vulnerable to p

14 May 2026

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

SafetyDGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents

AgentsDGX agent

arXiv:2605.13357v1 Announce Type: cross Abstract: Foundation models have transformed automated code generation, yet autonomous software-engineering agents remain unreliable in realistic development se

Autodesk taps Permiso Security to monitor AI agents across its cloud and workforce

AgentsDGX agent

Unified identity security platform provider Permiso Security Inc. today launched AI agent runtime security capabilities that give security teams continuous visibility into agent activity across cloud

Boomi and AWS built the guardrails for agents before anyone was asking for them

AgentsDGX agent

Agent governance has gone from niche concern to boardroom prerequisite — and the enterprises that saw it coming are now pulling ahead. The rise of agentic AI across enterprise operations has made gove

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

Model ReleasesDGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

Neurodata Without Boredom: Benchmarking Agentic AI for Data Reuse

AgentsDGX agent

arXiv:2605.12808v1 Announce Type: new Abstract: Neuroscience data are highly fragmented across labs, formats, and experimental paradigms, and reuse often requires substantial manual effort. A persiste

Okta extends AI agent security to Amazon Bedrock, opens platform to rival identity providers

AgentsDGX agent

Identity and access management company Okta Inc. today expanded its Okta for AI Agents platform to cover any agent ecosystem, any enterprise resource and any identity provider, including a new integra

PersonalAlign: Hierarchical Implicit Intent Alignment for Personalized GUI Agent with Long-Term User-Centric Records

Model ReleasesDGX agent

arXiv:2601.09636v2 Announce Type: replace Abstract: While GUI agents have shown strong performance under explicit and completion instructions, real-world deployment requires aligning with users' more

13 May 2026

Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

AgentsDGX agent

arXiv:2605.12411v1 Announce Type: cross Abstract: AI agents negotiate and transact in natural language with unfamiliar counterparts: a buyer bot facing an unknown seller, or a procurement assistant ne

SkillGen: Verified Inference-Time Agent Skill Synthesis

AgentsDGX agent

arXiv:2605.10999v1 Announce Type: new Abstract: Skills are a promising way to improve LLM agent capabilities without retraining, while keeping the added procedure reusable and controllable. However, h

We just shipped tons of new products to accelerate the full agent development lifecycle: https://www.langchain.com/blog TLDR: ✅ LangSmith En…

AgentsDGX agent

We just shipped tons of new products to accelerate the full agent development lifecycle: https://www.langchain.com/blog TLDR: ✅ LangSmith Engine ✅ SmithDB ✅ Sandboxes ✅ Managed Deep Agents ✅ LLM Gatew

12 May 2026

AgentSlimming: Towards Efficient and Cost-Aware Multi-Agent Systems

AgentsDGX agent

arXiv:2605.08813v1 Announce Type: new Abstract: Large Language Model-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in complex tasks. However, manually designing optimal com

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

Model ReleasesDGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

Autonomous Continual Learning for Environment Adaptation of Computer-Use Agents

AgentsDGX agent

arXiv:2602.10356v2 Announce Type: replace Abstract: Real-world digital environments are highly diverse and dynamic. These characteristics cause agents to frequently encounter unseen environments and d

Decentralized Contingency MPC based on Safe Sets for Nonlinear Multi-agent Collision Avoidance

AgentsDGX agent

arXiv:2605.10738v1 Announce Type: cross Abstract: Decentralized collision avoidance remains challenging, particularly when agents do not communicate any information related to planned trajectories. Mo

DeepRefine: Agent-Compiled Knowledge Refinement via Reinforcement Learning

AgentsDGX agent

arXiv:2605.10488v1 Announce Type: cross Abstract: Agent-compiled knowledge bases provide persistent external knowledge for large language model (LLM) agents in open-ended, knowledge-intensive downstre

EquiMem: Calibrating Shared Memory in Multi-Agent Debate via Game-Theoretic Equilibrium

AgentsDGX agent

arXiv:2605.09278v1 Announce Type: new Abstract: Multi-agent debate (MAD) systems increasingly rely on shared memory to support long-horizon reasoning, but this convenience opens a critical vulnerabili

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

Local AiDGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

Honeycomb introduces agent observability features to keep an eye on production

AgentsDGX agent

Full-stack observability startup Hound Technology Inc., which does business as Honeycomb, introduced a number of new platform updates aimed at investigating artificial intelligence agent activity in p

Join us for a live build session exploring Hermes Agent integrated with ComfyUI. If you're already comfortable navigating ComfyUI workflows …

AgentsDGX agent

Join us for a live build session exploring Hermes Agent integrated with ComfyUI. If you're already comfortable navigating ComfyUI workflows and want to understand what an agent layer brings to the tab

LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments

Model ReleasesDGX agent

arXiv:2605.10779v1 Announce Type: cross Abstract: The rapid proliferation of LLM-based autonomous agents in real operating system environments introduces a new category of safety risk beyond content s

Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks

SafetyDGX agent

arXiv:2605.08257v1 Announce Type: cross Abstract: Motivated by the challenge to improve the adversarial robustness, security, and trust of medical decision making intelligent agents, this study develo

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

Model ReleasesDGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

Thanks to everyone who showcased their projects in today's Hermes Agent Jam!

AgentsDGX agent

Nous Research held a 'Hermes Agent Jam' event where developers showcased projects built using or related to their Hermes agent framework. The event appears to have been a community-driven showcase hig

Thanks to the @huggingface team for adding Hermes Agent to local apps and shipping a native Hermes traces viewer!

Local AiDGX agent

Thanks to the @huggingface team for adding Hermes Agent to local apps and shipping a native Hermes traces viewer! 🆕 Hugging Face 🤝 Hermes Agent 🔥 > we added Hermes Agent to local apps: run it locally

Tomorrow @JulienAIArt and I are going to dive into Hermes Agent with the ComfyUI Skill and see what it can do! Bring your questions and tips…

AgentsDGX agent

Tomorrow @JulienAIArt and I are going to dive into Hermes Agent with the ComfyUI Skill and see what it can do! Bring your questions and tips! Join us for a live build session exploring Hermes Agent in

11 May 2026

A Self-Healing Framework for Reliable LLM-Based Autonomous Agents

AgentsDGX agent

arXiv:2605.06737v1 Announce Type: cross Abstract: Autonomous agents based on Large Language Models (LLMs) are increasingly being utilized in complex software systems. However, reliability remains a si

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management

AgentsDGX agent

arXiv:2512.10371v2 Announce Type: replace Abstract: The rapid development of mobile GUI agents has stimulated growing research interest in long-horizon task automation. However, building agents for th

agents make for a surprisingly great product

AgentsDGX agent

# Agents as Products This post likely explores how AI agents can be effectively deployed as standalone products, discussing their practical applications and advantages over traditional software soluti

Belief Memory: Agent Memory Under Partial Observability

AgentsDGX agent

arXiv:2605.05583v2 Announce Type: replace Abstract: LLM agents that operate over long context depend on external memory to accumulate knowledge over time. However, existing methods typically store eac

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

Model ReleasesDGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

Group of Skills: Group-Structured Skill Retrieval for Agent Skill Libraries

AgentsDGX agent

arXiv:2605.06978v1 Announce Type: cross Abstract: Skill-augmented agents increasingly rely on large reusable skill libraries, but retrieving relevant skills is not the same as presenting usable contex

Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

Local AiDGX agent

arXiv:2605.07637v1 Announce Type: new Abstract: Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning problems, where multiple homogeneous agents move simulta

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

Model ReleasesDGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples

AgentsDGX agent

arXiv:2605.07935v1 Announce Type: new Abstract: We present TraceFix, a verification-first pipeline for Large Language Model (LLM) multi-agent coordination. An agent synthesizes a protocol topology as

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

AgentsDGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

Why Does Agentic Safety Fail to Generalize Across Tasks?

SafetyDGX agent

arXiv:2605.06992v1 Announce Type: new Abstract: AI agents are increasingly deployed in multi-task settings, where the task to perform is specified at test time, and the agent must generalize to unseen

10 May 2026

Join the Nous Research team for another Hermes Agent Jam in our Discord This one will be an interactive session, so come prepared to discuss…

AgentsDGX agent

Nous Research is hosting an interactive Hermes Agent Jam session in their Discord community where participants can discuss and collaborate on Hermes agent development. The event invites community memb

9 May 2026

'OncoAgent: A Dual-Tier Multi-Agent Framework for Privacy-Preserving Oncology Clinical Decision Support'

AgentsDGX agent

OncoAgent is a dual-tier multi-agent framework designed to provide clinical decision support in oncology while maintaining patient privacy. The system likely leverages multiple specialized AI agents w

Signals V2: an LLM-free analyzer to scores live agent trajectories as OpenTelemetry spans.

AgentsDGX agent

Signals V2 is an LLM-free analyzer tool that evaluates live agent execution trajectories by scoring them as OpenTelemetry spans, enabling developers to monitor and trace AI agent behavior without rely

8 May 2026

The future of Math is mathematicians and AI agents working together. Very pleased to introduce @GoogleDeepMind's AI co-mathematician: a mult…

AgentsDGX agent

The future of Math is mathematicians and AI agents working together. Very pleased to introduce @GoogleDeepMind's AI co-mathematician: a multi-agent system designed to actively collaborate with human e

7 May 2026

Improving token efficiency in GitHub Agentic Workflows

AgentsDGX agent

Agentic workflows that run on every pull request can quietly accumulate large API bills. Here's how we instrumented our own production workflows, found the inefficiencies, and built agents to fix them

New course: Build agents that respond to users with not only plaintext, but custom UIs like charts, forms, and whiteboards, generated on dem…

AgentsDGX agent

New course: Build agents that respond to users with not only plaintext, but custom UIs like charts, forms, and whiteboards, generated on demand and displayed right in the chat. This short course is bu

← Previous
1…4344454647…297
Next →