AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
28 Jul 2026

‼️ Hugging Face built an interactive replay of the OpenAI agent that breached them. It includes 17,613 logged attacker actions across the 4.…

AgentsDGX agent

‼️ Hugging Face built an interactive replay of the OpenAI agent that breached them. It includes 17,613 logged attacker actions across the 4.5-day campaign, with the live command stream and more. https

ReasFlow: Assisting Reasoning-Centric Scientific Discovery in Applied Mathematics via a Knowledge-Based Multi-Agent System

AgentsDGX agent

arXiv:2607.14178v2 Announce Type: replace Abstract: Recent advances in Large Language Models have fueled autonomous AI agents capable of tackling complex scientific tasks, yet existing automated resea

SF-AMS: Strategic Forgetting for Structured Memory in LLM Agent

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.22562v1 Announce Type: new Abstract: Managing long-context dependencies remains a primary bottleneck in LLM agents, as redundant and irrelevant information can degrade multi-step reasoning.

Snowflake debuts Cortex AI Gateway to govern and monitor enterprise AI agents

AgentsDGX agent

Snowflake Inc. today introduced Cortex AI Gateway, a centralized control layer that lets enterprises connect, govern and monitor artificial intelligence agents as they reach into models, tools, Model

27 Jul 2026

Are agent skills always worth using? The answer is no? This paper provides some important insights to understand this more. (bookmark it) Pa…

AgentsDGX agent

Are agent skills always worth using? The answer is no? This paper provides some important insights to understand this more. (bookmark it) Paper summary: Adding procedural skills to an agent is usually

Building the enterprise environment for agentic AI

SafetyDGX agent

For the enterprise, the promise of agentic AI is much more than just a better chatbot. It is software agents that execute business tasks end-to-end across people, business workflows, data, and systems

here’s a video on my approach at applying this to agents w @activegraphai:

AgentsDGX agent

here’s a video on my approach at applying this to agents w @activegraphai: 🆕 ActiveGraph: The Log is the Agent my talk from AI Engineer is live!!! 😆 https://www.youtube.com/watch?v=khVX_BUnEwU it's ab

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

Model ReleasesDGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

24 Jul 2026

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers

AgentsDGX agent

arXiv:2607.21594v1 Announce Type: new Abstract: Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evo

16 Jul 2026

A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Multi-Agent Systems

AgentsDGX agent

arXiv:2507.19593v3 Announce Type: replace Abstract: Classical game-theoretic models typically assume rational agents, complete information, and common knowledge of payoffs - assumptions that are often

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

Model ReleasesDGX agent

arXiv:2607.14004v1 Announce Type: new Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is repor

15 Jul 2026

AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

AgentsDGX agent

arXiv:2607.11098v2 Announce Type: replace-cross Abstract: Tool-using LLM agents are mostly evaluated assuming all tools work. When a tool times out, returns a week-stale value, or has its description

Kiro CLI observability: trace and evaluate agent changes with Arize Skills

AgentsDGX agent

Use Arize Skills with Kiro CLI to trace coding-agent changes, build datasets from failures, run experiments, and validate prompts before shipping. The post Kiro CLI observability: trace and evaluate a

14 Jul 2026

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eva…

AgentsDGX agent

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eval signal that reveals whether an agent fails, and where. The

13 Jul 2026

GPUs for all 🤗 Creating ZeroGPU demos or apps is now available to ALL @huggingface users tell your agent: 'Build a HF ZeroGPU demo for this…

AgentsDGX agent

GPUs for all 🤗 Creating ZeroGPU demos or apps is now available to ALL @huggingface users tell your agent: 'Build a HF ZeroGPU demo for this model' New Space: https://huggingface.co/new-space Agent ski

Pinecone has introduced Nexus, a knowledge engine built specifically for AI agents. Instead of repeatedly searching documents with tradition…

AgentsDGX agent

Pinecone has introduced Nexus, a knowledge engine built specifically for AI agents. Instead of repeatedly searching documents with traditional RAG, Nexus compiles knowledge once from sources like data

8 Jul 2026

Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development

AgentsDGX agent

arXiv:2607.06101v1 Announce Type: cross Abstract: AI coding agents are rapidly reshaping how software is built, with developers increasingly delegating substantial coding tasks to autonomous agents in

Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with @NVIDIA ✅ A fully op…

Model ReleasesDGX agent

Introducing the NemoClaw Deep Agents Blueprint, a reference architecture for building open agent systems developed with @NVIDIA ✅ A fully open stack enterprises can own and customize ✅ Benchmark-leadi

7 Jul 2026

Explicit Credit Assignment through Local Rewards and Dependence Graphs in Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2601.21523v2 Announce Type: replace Abstract: To promote cooperation in Multi-Agent Reinforcement Learning, the reward signals of all agents can be aggregated together, forming global rewards th

How a local-first personal AI agent uses orchestration, memory, tools, LangGraph, background workflows, and child agents

Local AiDGX agent

This post describes architectural patterns for building a local-first personal AI agent, covering key components including orchestration frameworks, memory systems, tool integration, LangGraph for wor

No Time Like the Present: Agentic Test-Time Training for LLM Agents

SafetyDGX agent

arXiv:2607.03441v1 Announce Type: cross Abstract: LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previ

3 Jul 2026

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've f…

AgentsDGX agent

Multimodal prompting is clearly the future. I love experimenting with new ways to interact with agents. As a researcher and engineer, I've found that the richer the inputs to the agent and the richer

PACE: A Proxy for Agentic Capability Evaluation

Model ReleasesDGX agent

arXiv:2607.02032v1 Announce Type: new Abstract: Evaluating LLM agents on benchmarks like SWE-Bench and GAIA can be expensive, time-consuming, and requires complex infrastructure. A single evaluation c

1 Jul 2026

A Systematic Approach to Multi-Agent AI from Advanced Regulatory Control Theory: Safe and Auditable LLM Operator Agents for Process Control

Model ReleasesDGX agent

arXiv:2606.30877v1 Announce Type: cross Abstract: Recent literature shows that large language models (LLMs) are useful for general-purpose tasks yet perform poorly on specific domain ones. One reason

Claude Fable 5 is available in Devin. You can use Fable 5 in Devin Cloud’s Ultra agent, our smartest and most capable agent, which excels at…

Model ReleasesDGX agent

Claude Fable 5 is available in Devin. You can use Fable 5 in Devin Cloud’s Ultra agent, our smartest and most capable agent, which excels at long-horizon tasks and debugging. Claude Fable 5 is also av

Hierarchical Message-Passing Policies for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2507.23604v2 Announce Type: replace Abstract: Decentralized Multi-Agent Reinforcement Learning (MARL) methods allow for learning scalable multi-agent policies, but suffer from partial observabil

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And…

Model ReleasesDGX agent

Some of LangSmith's biggest and most engaged customers trace with: - @aisdk - Claude Agents SDK - OpenAI Agent SDK - No framework at all And I'm sure we'll see tons of http://pi.dev in the near future

30 Jun 2026

HyphaeDB: A Living Knowledge Topology for Agent-First Memory

AgentsDGX agent

arXiv:2606.28781v1 Announce Type: new Abstract: Every existing vector database and agent memory framework treats memory as passive storage that agents query explicitly. No system propagates knowledge

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

AgentsDGX agent

arXiv:2606.30555v1 Announce Type: new Abstract: The rapid integration of Large Language Models (LLMs) has driven the evolution of Multi-Agent Systems (MAS), where specialized agents collaborate to exe

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Fir…

Model ReleasesDGX agent

SEO Agent. Security Agent. Design in Claude. Replit Desktop. Shopify on Replit. Skills + Custom Instructions. 450+ integrations. Package Firewall. And more! ✨ All shipped in June! 🤯 And there's a good

29 Jun 2026

Introducing Cursor for iOS. Build from anywhere by launching always-on cloud agents. Or remotely control agents running on your computer fro…

ToolsDGX agent

Introducing Cursor for iOS. Build from anywhere by launching always-on cloud agents. Or remotely control agents running on your computer from the app. Composer 2.5 is 75% off in the app now through Ju

Supercharging the agentic era with Spanner’s multi-model architecture

Model ReleasesDGX agent

In the agentic era, the role of the database has fundamentally changed. It is no longer a passive repository; it’s a critical context engine designed to ground generative AI apps, models and power aut

we just released dynamic subagents, which let your agent programmatically orchestrate subagents in a code interpreter. this lets agents do w…

Model ReleasesDGX agent

we just released dynamic subagents, which let your agent programmatically orchestrate subagents in a code interpreter. this lets agents do work at scale that tool calls can't reliably handle, like pro

27 Jun 2026

🧑‍🏫3hr Long Deep Agents Course Great course from a member of the community on Deep Agents Covers task planning, file systems for context m…

TutorialsDGX agent

🧑‍🏫3hr Long Deep Agents Course Great course from a member of the community on Deep Agents Covers task planning, file systems for context management, subagent-spawning, and long-term memory https://www

25 Jun 2026

Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents:…

ApplicationsDGX agent

Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents: full-stack examples with streaming UI, subagents, thread hi

23 Jun 2026

MAS-PromptBench: When Does Prompt Optimization Improve Multi-Agent LLM Systems?

AgentsDGX agent

arXiv:2606.23664v1 Announce Type: new Abstract: Multi-agent systems (MAS) offer a scalable path forward for agentic AI, comprising multiple LLM-based agents, each assigned a system prompt and a positi

10 Jun 2026

Bittensor Agent Arenas as a Trajectory Primitive: Distilling a Shopping Agent from ShoppingBench Subnet Traces

Model ReleasesDGX agent

arXiv:2606.10064v1 Announce Type: cross Abstract: Small-model agentic post-training is bottlenecked less by the algorithm than by the trajectory substrate it consumes. Leading recipes (RLVR, group-rel

9 Jun 2026

Agent Economics: An Entropy-Controlled Pluralistic Alignment Framework for Preventing Artificial Hivemind in Autonomous Agents

SafetyDGX agent

arXiv:2606.09039v1 Announce Type: new Abstract: This study proposes the Behavioral Protocol Framework (BPF), an entropy-controlled pluralistic alignment framework designed to address two critical chal

Causal Agent Replay: Counterfactual Attribution for LLM-Agent Failures

Model ReleasesDGX agent

arXiv:2606.08275v1 Announce Type: cross Abstract: When an LLM agent fails -- issues a refund it should not have, calls the wrong tool, leaks data -- existing tooling answers what happened (observabili

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage cont…

Model ReleasesDGX agent

Just landed nested subagent support in Claude Code Starting to experiment more with agents kicking off agents as a way to better manage context. Capped at depth=5 to start, going out in today’s releas

Rubrik turns its platform into an AI agent and ships Agent Cloud for Claude

Model ReleasesDGX agent

Rubrik Inc. today turned its data security platform into an autonomous agent and made its control layer for Anthropic PBC’s Claude generally available, the headline items in a wave of announcements at

Use Ollama with Hermes Desktop by @NousResearch. Hermes Desktop brings the same agent (its multi-agent engine, self-improving skills, and me…

Local AiDGX agent

Use Ollama with Hermes Desktop by @NousResearch. Hermes Desktop brings the same agent (its multi-agent engine, self-improving skills, and messaging integrations) into a desktop app on macOS, Windows,

6 Jun 2026

Detecting Perspective Shifts in Multi-agent Systems

AgentsDGX agent

arXiv:2512.05013v2 Announce Type: replace Abstract: Generative models augmented with external tools and update mechanisms (or extit{agents}) have demonstrated capabilities beyond intelligent prompting

Do More Agents Help? Controlled and Protocol-Aligned Evaluation of LLM Agent Workflows

Model ReleasesDGX agent

arXiv:2606.05670v1 Announce Type: new Abstract: Does adding more agents help an LLM workflow once compared systems share the same benchmark loader, tool access, answer contract, usage accounting, and

3 Jun 2026

Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks

AgentsDGX agent

arXiv:2606.02875v1 Announce Type: new Abstract: Coding-agent benchmarks evaluate whether a single uninterrupted agent can resolve a repository issue. Real software work is messier: tasks are interrupt

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + …

Model ReleasesDGX agent

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + two tool sets (global company data and firm-specific context

2 Jun 2026

Agentic Clustering: Controllable Text Taxonomies via Multi-Agent Refinement

AgentsDGX agent

arXiv:2606.01255v1 Announce Type: new Abstract: Recent text-clustering methods use large language models to propose a cluster taxonomy from a corpus and then assign each text to it. These pipelines ar

LRAgent: Efficient KV Cache Sharing for Multi-LoRA LLM Agents

AgentsDGX agent

arXiv:2602.01053v2 Announce Type: replace Abstract: Role specialization in multi-LLM agent systems is often realized via multi-LoRA, where agents share a pretrained backbone and differ only by lightwe

Modeling Distinct Human Interaction in Web Agents

AgentsDGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

new in deepagents: agent rubrics! you define a rubric, and the agent self-evaluates and iterates until it satisfies every rubric criterion. …

Model ReleasesDGX agent

new in deepagents: agent rubrics! you define a rubric, and the agent self-evaluates and iterates until it satisfies every rubric criterion. this is similar to /goal in claude code or codex, but more f

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVID…

HardwareDGX agent

Teach an agent a workflow once. Have it remember after every rebuild. This tutorial shows how to deploy @nousresearch Hermes Agent with NVIDIA NemoClaw and OpenShell, connect it to Slack, Outlook, Git

1 Jun 2026

We have worked with @nvidia to integrate their official Agent Skills catalog into the Hermes Skills Hub. These skills teach your agent how t…

HardwareDGX agent

We have worked with @nvidia to integrate their official Agent Skills catalog into the Hermes Skills Hub. These skills teach your agent how to use CUDA-X libraries, Omniverse and Physical AI workflows,

29 May 2026

How Coding Agents Fail Their Users: A Large-Scale Analysis of Developer-Agent Misalignment in 20,574 Real-World Sessions

Model ReleasesDGX agent

arXiv:2605.29442v1 Announce Type: cross Abstract: AI coding agents increasingly act directly within software environments, yet existing analyses of their failures rely on benchmark trajectories that m

26 May 2026

AgentFugue: Agent Scaling for Long-Horizon Tasks through Collective Reasoning

AgentsDGX agent

arXiv:2605.24486v1 Announce Type: new Abstract: Recent progress on long-horizon agentic tasks has been driven largely by scaling up individual agents through stronger models, better tools, and more ef

25 May 2026

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official…

Model ReleasesDGX agent

Hermes Agent now can orchestrate the @OpenHandsDev agents with a new optional skill! `hermes update` then do `hermes skills install official/autonomous-ai-agents/openhands` Reminder: You can already d

Push Your Agent: Measuring and Enforcing Quantitative Goal Persistence in Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2605.23574v1 Announce Type: new Abstract: Long-horizon language agents can make many plausible local tool calls yet fail to persist until a requested count is actually complete. We study this ga

23 May 2026

Most agents die after a few seconds. @AnthropicAI's workshop shows how to build agents that run for hours. full 75-min session with Ash Prab…

TutorialsDGX agent

Most agents die after a few seconds. @AnthropicAI's workshop shows how to build agents that run for hours. full 75-min session with Ash Prabaker & Andrew Wilson. https://www.youtube.com/watch?v=mR-WAv

22 May 2026

Enabling Regulatory Multi-Agent Collaboration: Architecture, Challenges, and Solutions

AgentsDGX agent

arXiv:2509.09215v2 Announce Type: replace Abstract: Large language models (LLMs)-empowered autonomous agents are transforming both digital and physical environments by enabling adaptive, multi-agent c

19 May 2026

Consent Chain Degradation in Embodied Multi-Agent Systems: Bridging the Gap Between AI Agent Governance and Robot Ethics

SafetyDGX agent

arXiv:2605.16300v1 Announce Type: cross Abstract: Robotic systems are moving from isolated platforms to interconnected multi-agent ecosystems that operate in human environments. This shift raises a go

14 May 2026

Real-time voice agents with Stream Vision Agents and Amazon Nova 2 Sonic

TutorialsDGX agent

In this post, you learn how to combine Stream's Vision Agents open-source framework with Amazon Bedrock and Amazon Nova 2 Sonic to build real-time voice agents that can be production-ready in minutes.

← Previous
1…2021222324…294
Next →