AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,737 results
Agents

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

DGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

agentsarxiv-cs-ai
27 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

DGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

model-releasesarxiv-cs-ai
27 May 2026
Safety

Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation

DGX agent

arXiv:2510.19420v2 Announce Type: replace-cross Abstract: Multi-Agent Systems (MAS) have become a prevalent paradigm for Large Language Model (LLM) applications. However, the complex multi-agent desig

safetyarxiv-cs-ai
27 May 2026
Agents

DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations

DGX agent

arXiv:2605.24539v1 Announce Type: new Abstract: Agent harness evolution improves frozen language-model agents by modifying the executable structures around them. We study this paradigm as a form of sa

agentsarxiv-cs-ai
26 May 2026
Tutorials

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from lon…

DGX agent

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from long-horizon agents. Here is the problem with agents today: Att

tutorialsdair-ai--x
26 May 2026
Agents

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

DGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

agentsarxiv-cs-ai
26 May 2026
Agents

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection

DGX agent

arXiv:2605.23723v1 Announce Type: new Abstract: Large language model agents increasingly rely on persistent memory to store past interactions, retrieve relevant demonstrations, and improve long-horizo

agentsarxiv-cs-ai
25 May 2026
Agents

When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.23414v1 Announce Type: new Abstract: LLM-based multi-agent systems can fail even when planned actions are executed correctly because agents may misjudge their knowledge when evaluating plan

agentsarxiv-cs-ai
25 May 2026
Agents

Every 'self-evolving agent' paper this year has mutated text: prompts, skill files, workflow graphs, memory schemas. MOSS from USTC & HKUST …

DGX agent

Every 'self-evolving agent' paper this year has mutated text: prompts, skill files, workflow graphs, memory schemas. MOSS from USTC & HKUST argues this is the wrong layer. The thing that actually brea

agentsyohei-nakajima--x
23 May 2026
Agents

Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs

DGX agent

arXiv:2605.20315v1 Announce Type: new Abstract: LLM agents have recently emerged as a powerful paradigm for solving complex tasks through planning, tool use, memory retrieval, and multi-step interacti

agentsarxiv-cs-cl
21 May 2026
Agents

CANTANTE: Optimizing Agentic Systems via Contrastive Credit Attribution [R]

DGX agent

CANTANTE addresses the challenge of optimizing LLM-based multi-agent systems where system-level performance scores are available but individual agent parameters cannot be directly optimized. The frame

agentsr-machinelearning
20 May 2026
Agents

Railway: The Agent-Native Cloud — Jake Cooper

DGX agent

Railway is a cloud platform that provides agent-native infrastructure, enabling AI agents to be deployed and executed natively within the cloud environment. The discussion with Jake Cooper likely cove

agentslatent-space
20 May 2026
Agents

When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity

DGX agent

arXiv:2605.20023v1 Announce Type: new Abstract: Agent Skills, structured packages of procedural knowledge loaded into an LLM agent at inference time, are widely reported to improve task pass rates by

agentsarxiv-cs-ai
20 May 2026
Agents

WisdomAI’s new analytics agents go beyond insights, automating business work through autonomous action

DGX agent

WisdomAI Inc., creator of an artificial intelligence-native business intelligence platform, is jumping on the agentic AI bandwagon with the latest update to its flagship Federated Agentic Intelligence

agentssiliconangle
20 May 2026
Agents

Before enterprises can run with agentic AI, they need to learn to walk with their data

DGX agent

Multi-agent orchestration is the destination, but for most enterprises, the road is blocked long before the first agent gets deployed by the quality of the data feeding those systems. As organizations

agentssiliconangle
19 May 2026
Agents

Can LLM Agents Be CFOs? Benchmarking Long-Horizon Resource Allocation in an Uncertain Enterprise Environment

DGX agent

arXiv:2603.23638v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly tested on complex tasks, but their ability to allocate scarce resources over long horizons remain

agentsarxiv-cs-ai
19 May 2026
Agents

GRASP: Graph Agentic Search over Propositions for Multi-hop Question Answering

DGX agent

arXiv:2605.16598v1 Announce Type: cross Abstract: Agentic retrieval improves multi-hop question answering by giving language models autonomy to iteratively gather evidence. Recent work augments these

agentsarxiv-cs-ai
19 May 2026
Safety

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

DGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

safetyarxiv-cs-cl
19 May 2026
Model Releases

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

DGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

NeuroMAS: Multi-Agent Systems as Neural Networks with Joint Reinforcement Learning

DGX agent

arXiv:2605.16757v1 Announce Type: new Abstract: Multi-agent language systems are often built as hand-designed workflows, where agents are assigned semantic roles and communication protocols are specif

model-releasesarxiv-cs-ai
19 May 2026
Agents

Stop rogue AI: How Unity Catalog secures your agent actions

DGX agent

Unity Catalog is a Databricks governance solution that helps secure AI agent actions by managing access controls and permissions across agent operations. It prevents unauthorized or 'rogue' AI behavio

agentsdatabricks
19 May 2026
Model Releases

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

DGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

model-releasesarxiv-cs-ai
19 May 2026
Agents

Differentiable Mixture-of-Agents Incentivizes Swarm Intelligence of Large Language Models

DGX agent

arXiv:2605.15706v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have catalyzed the development of multi-agent systems (MAS) for complex reasoning tasks. However, existi

agentsarxiv-cs-lg
18 May 2026
Agents

Every time I ask my 10-year-old to use coding agents, he gets extremely disappointed. It turns out that all he wants is to build his own roc…

DGX agent

Every time I ask my 10-year-old to use coding agents, he gets extremely disappointed. It turns out that all he wants is to build his own rocket simulator. No amount of context engineering helps. No mo

agentsdair-ai--x
18 May 2026
Agents

H-Mem: A Novel Memory Mechanism for Evolving and Retrieving Agent Memory via a Hybrid Structure

DGX agent

arXiv:2605.15701v1 Announce Type: cross Abstract: Memory data are ubiquitous in Large Language Model (LLM)-based agents (e.g., OpenClaw and Manus). A few recent works have attempted to exploit agents'

agentsarxiv-cs-ai
18 May 2026
Agents

Look Before You Leap: Autonomous Exploration for LLM Agents

DGX agent

arXiv:2605.16143v1 Announce Type: new Abstract: Large language model based agents often fail in unfamiliar environments due to premature exploitation: a tendency to act on prior knowledge before acqui

agentsarxiv-cs-ai
18 May 2026
Agents

The Open Agent Leaderboard

DGX agent

The Open Agent Leaderboard is a benchmarking system hosted on Hugging Face that evaluates and ranks AI agents based on their performance across various tasks and capabilities. It provides a standardiz

agentshugging-face
18 May 2026
Agents

when evaluating long running agents, all of your evals don't need to be end to end. i'm working on a proper blog about this, but in our eval…

DGX agent

when evaluating long running agents, all of your evals don't need to be end to end. i'm working on a proper blog about this, but in our evals for our agents that run for 30-60 minutes, we have two set

agentsharrison-chase--x
18 May 2026
Agents

Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces

DGX agent

arXiv:2605.14786v1 Announce Type: cross Abstract: As LLM-based agents increasingly browse the web on users' behalf, a natural question arises: can websites passively identify which underlying model po

agentsarxiv-cs-ai
15 May 2026
Agents

SuperGrok now in Hermes Agent

DGX agent

SuperGrok has been integrated into the Hermes Agent, representing an advancement in Nous Research's AI agent capabilities. This integration likely combines SuperGrok's reasoning or processing features

agentsnous-research--x
15 May 2026
Model Releases

Web Agents Should Adopt the Plan-Then-Execute Paradigm

DGX agent

arXiv:2605.14290v1 Announce Type: cross Abstract: ReAct has become the default architecture across LLM agents, and many existing web agents follow this paradigm. We argue that it is the wrong default

model-releasesarxiv-cs-ai
15 May 2026
Agents

You can now use your @grok subscription inside @NousResearch Hermes Agent. http://x.ai/news/grok-hermes

DGX agent

Nous Research has integrated Grok, xAI's large language model, into the Hermes Agent framework, allowing users with active @grok subscriptions to leverage Grok's capabilities within the Hermes Agent e

agentsnous-research--x
15 May 2026
Agents

AgenticFlict: A Large-Scale Dataset of Merge Conflicts in AI Coding Agent Pull Requests on GitHub

DGX agent

arXiv:2604.03551v2 Announce Type: replace-cross Abstract: Software Engineering 3.0 marks a paradigm shift in software development, in which AI coding agents are no longer just assistive tools but acti

agentsarxiv-cs-ai
14 May 2026
Agents

CHAL: Council of Hierarchical Agentic Language

DGX agent

arXiv:2605.12718v1 Announce Type: new Abstract: Multi-agent debate has emerged as a promising approach for improving LLM reasoning on ground-truth tasks, yet current methodologies face certain structu

agentsarxiv-cs-ai
14 May 2026
Model Releases

Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue

DGX agent

arXiv:2605.12920v1 Announce Type: cross Abstract: Effective collaboration between embodied agents requires more than acting in a shared environment; it demands communication grounded in each agent's e

model-releasesarxiv-cs-ai
14 May 2026
Agents

Excited to see SWE-ZERO trending on HF alongside awesome agentic trace datasets like AgentTrove!

DGX agent

SWE-ZERO and AgentTrove are trending datasets on Hugging Face related to software engineering and agentic AI systems. The post indicates growing interest in trace datasets that capture agent behaviors

agentsclem-delangue--x
14 May 2026
Agents

Freshworks unveils Freddy AI Agent Studio and MCP Gateway for Freshservice

DGX agent

Freshworks Inc. today unveiled an expanded set of agentic capabilities in its Freshservice information technology service management platform led by a new no-code Freddy AI Agent Studio that lets ente

agentssiliconangle
14 May 2026
Agents

Introducing the Anyscale Agent Skill for LLM Post-Training

DGX agent

Anyscale has introduced a new Agent Skill designed to enhance LLM post-training capabilities, likely enabling developers to build and train agentic systems more effectively using Ray's distributed com

agentsanyscale-ray
14 May 2026
Agents

OpenAaaS: An Open Agent-as-a-Service Framework for Distributed Materials-Informatics Research

DGX agent

arXiv:2605.13618v1 Announce Type: cross Abstract: The Materials Genome Initiative catalyzed the proliferation of centralized platforms--SaaS, PaaS, and IaaS--that aggregate computational and experimen

agentsarxiv-cs-ai
14 May 2026
Safety

The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration

DGX agent

arXiv:2602.01453v3 Announce Type: replace Abstract: We study cooperative multi-agent reinforcement learning in the setting of reward-free exploration, where multiple agents jointly explore an unknown

safetyarxiv-cs-lg
14 May 2026
Agents

Focusing Influence Mechanism for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2506.19417v2 Announce Type: replace Abstract: Cooperative multi-agent reinforcement learning (MARL) under sparse rewards remains fundamentally challenging because agents often fail to concentrat

agentsarxiv-cs-lg
13 May 2026
Agents

just tried this out and it one-shotted* this video: 'before the agent does anything' *i generated the narrative using chatgpt and used that …

DGX agent

just tried this out and it one-shotted* this video: 'before the agent does anything' *i generated the narrative using chatgpt and used that as a prompt. featuring: @e2b @runanywhereai @composio @mem0a

agentsyohei-nakajima--x
13 May 2026
Safety

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

DGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

safetyarxiv-cs-ai
12 May 2026
Model Releases

GAMBIT: A Three-Mode Benchmark for Adversarial Robustness in Multi-Agent LLM Collectives

DGX agent

arXiv:2605.09027v1 Announce Type: new Abstract: In multi-agent systems (MAS), a single deceptive agent can nullify all gains of an agentic AI collective and evade deployed defenses. However, existing

model-releasesarxiv-cs-cl
12 May 2026
Agents

MIND-Skill: Quality-Guaranteed Skill Generation via Multi-Agent Induction and Deduction

DGX agent

arXiv:2605.08670v1 Announce Type: new Abstract: Large language model (LLM) powered AI agents have emerged as a promising paradigm for autonomous problem-solving, yet they continue to struggle with com

agentsarxiv-cs-ai
12 May 2026
Agents

Robust Multi-Agent LLMs under Byzantine Faults

DGX agent

arXiv:2605.09076v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly collaborate over peer-to-peer networks to improve their reliability. However, these same interactions c

agentsarxiv-cs-ai
12 May 2026
Agents

WebTrap: Stealthy Mid-Task Hijacking of Browser Agents During Navigation

DGX agent

arXiv:2605.08310v1 Announce Type: cross Abstract: Browser agents are increasingly deployed in long-horizon tasks, which require executing extended action chains to accomplish user goals. However, this

agentsarxiv-cs-ai
12 May 2026
Agents

Willful Disobedience: Automatically Detecting Failures in Agentic Traces

DGX agent

arXiv:2603.23806v2 Announce Type: replace-cross Abstract: AI agents are increasingly embedded in real software systems, where they execute multi-step workflows through multi-turn dialogue, tool invoca

agentsarxiv-cs-ai
12 May 2026
← Previous
1…4546474849…370
Next →