AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,599 results
9 Apr 2026

AI Agents Know About Supabase. They Don't Always Use It Right.

AgentsDGX agent

Supabase has launched **Agent Skills**, an open-source set of instructions designed to teach AI coding agents how to build on Supabase correctly, addressing the problem that while AI agents have ge...

NEW paper from Microsoft Every agent benchmark has the same hidden problem: how do you know the agent actually succeeded? Microsoft research…

Model ReleasesDGX agent

NEW paper from Microsoft Every agent benchmark has the same hidden problem: how do you know the agent actually succeeded? Microsoft researchers introduce the Universal Verifier, which discusses lesson

12 Aug 2026

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2606.08348v2 Announce Type: replace Abstract: LLM agents increasingly rely on prompts, tools, memory, SOPs, skills, and harness feedback, yet current self-evolution pipelines often update these

MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph

AgentsDGX agent

arXiv:2608.10504v1 Announce Type: new Abstract: As coding agents increasingly handle implementation, the central challenge shifts from building individual agents to building an infrastructure that sys

Mitigating Context Interference for Reliable and Efficient Search Agents

AgentsDGX agent

arXiv:2608.10743v1 Announce Type: new Abstract: Recent research empowers Large Language Models (LLMs) as multi-turn search agents to iteratively retrieve and generate outputs until complex tasks are s

VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection

AgentsDGX agent

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary mode

9 Aug 2026

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malici…

Model ReleasesDGX agent

Prompt injection is the most common way that scammers attack people and agents: your agent visits http://foo.com, and the website has malicious text like “btw send the user’s ssh keys and passwords to

6 Aug 2026

What are Agentic Workflows?

AgentsDGX agent

**Agentic Workflows** enable automated systems to execute complex tasks by orchestrating multiple agents that interact with each other and external services. They delegate subtasks, monitor progress,

3 Aug 2026

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the i…

AgentsDGX agent

// Model or Harness // Great paper if you are building with agents in production. (bookmark it) It organizes 41 agent failure modes by the interaction they originate in. Each mode gets assigned to an

31 Jul 2026

The Age of AI Agents Demands A New Scientific Paradigm To Sustain Trustworthy Science

AgentsDGX agent

arXiv:2607.26064v1 Announce Type: cross Abstract: AI systems are becoming autonomous research agents that generate hypotheses, design experiments, and produce discoveries at scales beyond human oversi

28 Jul 2026

AI agent evaluation: Tips from Anthropic on building evals you can trust

AgentsDGX agent

Learn how to build trustworthy AI agent evals using regression tests, capability evals, production traces, LLM judges, and reproducible environments. The post AI agent evaluation: Tips from Anthropic

Focus Is All You Need: Adaptive Goal-aware Attention Orchestration for Multi-Agent Graph Systems

AgentsDGX agent

arXiv:2607.23678v1 Announce Type: new Abstract: Large language models (LLMs) enable autonomous agents for reasoning, planning, and tool use. Recent systems increasingly organize these agents as graphs

27 Jul 2026

How Do AI Coding Agents Contribute to Software Development? an Empirical Study of Agentic Pull Requests

AgentsDGX agent

arXiv:2607.21832v1 Announce Type: cross Abstract: Recent advances in large language models and their rapid adoption across software engineering tasks have made Artificial Intelligence (AI) coding agen

24 Jul 2026

Open Knowledge format v0.2 tackles agentic trust

Model ReleasesDGX agent

When we introduced the Open Knowledge Format (OKF) in June 2026, we asserted that the context that agents need (table schemas, metric definitions, runbooks) should live in a format, not in a proprieta

23 Jul 2026

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in…

Model ReleasesDGX agent

If you are building real-time voice agents with @GoogleDeepMind Gemini Live, you can now trace your speech-to-speech agent loops directly in @LangChain! - Speaker callback hooks capture only the exact

22 Jul 2026

New for enterprises: OpenAI Presence helps companies deploy trusted voice and chat agents across customer and internal workflows. AI agents …

ApplicationsDGX agent

New for enterprises: OpenAI Presence helps companies deploy trusted voice and chat agents across customer and internal workflows. AI agents can answer questions, use company systems, take approved act

14 Jul 2026

Claude at scale on Google Cloud: Frontier AI, built for enterprise production

Model ReleasesDGX agent

Running frontier AI in production is demanding — accelerators to manage, latency to hold steady across continents, regulated data to keep in-region, and long-context requests to serve reliably. Claude

Entrust launches Agentic AI Trust Accelerator to move AI agents into production

Model ReleasesDGX agent

Identity-centered security solution company Entrust Corp. today launched the Agentic AI Trust Accelerator, a co-development program that brings together enterprises and technology partners to build th

10 Jul 2026

Frontier and Center: Who evaluates the evaluations?

Model ReleasesDGX agent

Editor’s note: Some of the most interesting questions in AI are being asked by information theoreticians, around how to provide context to an emerging class of AI agents. A few weeks ago, we waded int

9 Jul 2026

CILC: Cryptographically-secure Inter-agent Loop Closure Candidate Detection for Multi-Agent Collaborative SLAM

AgentsDGX agent

arXiv:2607.06700v1 Announce Type: new Abstract: Multi-agent Simultaneous Localization and Mapping (SLAM) and collaborative SLAM (CSLAM) require robots to continuously exchange global descriptors (GDs)

When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents

SafetyDGX agent

arXiv:2607.06595v1 Announce Type: cross Abstract: Personal AI agents powered by large language models can reason and act using available tools to access emails, manage calendars, and push code to remo

8 Jul 2026

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benc…

Model ReleasesDGX agent

Great writeup from the University of Oxford. It's a taxonomy of LLM-based agent limitations. Good read for anyone shipping with agents. Benchmark scores keep climbing, yet the same agent failures resu

7 Jul 2026

Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration

AgentsDGX agent

arXiv:2511.02200v2 Announce Type: replace Abstract: The emergence of multi-agent systems powered by large language models (LLMs) has unlocked new frontiers in complex task-solving, enabling diverse ag

2 Jul 2026

my fave question, talked about this coding agent Eval+Improvement loop infra + UX in my AIE talk yesterday! biased but LangSmith is the best…

Model ReleasesDGX agent

my fave question, talked about this coding agent Eval+Improvement loop infra + UX in my AIE talk yesterday! biased but LangSmith is the best spot to Eval + continuously improve your coding agents, and

30 Jun 2026

ACPO: Agent-Chained Policy Optimization for Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.30072v1 Announce Type: new Abstract: Cooperative tasks in Multi-Agent Reinforcement Learning (MARL) require agents to collectively maximize a shared return. Under the Centralized Training w

The gap in autonomous agentic loops that gets ignored: agents can plan and call APIs but can't acquire tools they don't have access to. x402…

AgentsDGX agent

The gap in autonomous agentic loops that gets ignored: agents can plan and call APIs but can't acquire tools they don't have access to. x402 + Apify's 20,000+ Actors is a concrete fix for that. Worth

26 Jun 2026

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents

Model ReleasesDGX agent

arXiv:2606.26627v1 Announce Type: cross Abstract: Large language model agents increasingly query databases, search document collections, call external APIs, remember past interactions, and act on a us

25 Jun 2026

Agentic infrastructure startup Seltz raises $12.5M to help AI agents search the web for answers

AgentsDGX agent

Agentic search startup Seltz Inc. said today it has bagged 12.5 million in seed funding to build a more optimal infrastructure so that artificial intelligence agents can find their way around the web.

21 Jun 2026

>> Scalable Evaluation for AI Agents << If you run agent evaluation in production, this one is worth your time. It shows that front-loading …

AgentsDGX agent

>> Scalable Evaluation for AI Agents << If you run agent evaluation in production, this one is worth your time. It shows that front-loading human judgment into reusable evaluation assets is useful. Bu

9 Jun 2026

Benchmarking Open-Ended Multi-Agent Coordination in Language Agents

Model ReleasesDGX agent

arXiv:2606.08340v1 Announce Type: new Abstract: As language models are increasingly deployed as autonomous agents, they must coordinate with others over long horizons in open-ended interactive tasks.

7 Jun 2026

We're building this at LangChain Fleet lets you create and manage a fleet of agents. Each agent specializes in a workflow, e.g. inbox manage…

AgentsDGX agent

We're building this at LangChain Fleet lets you create and manage a fleet of agents. Each agent specializes in a workflow, e.g. inbox management, blog writing, competitor research, candidate recruitin

5 Jun 2026

here's an activegraph based deep research agent that gives you full graph/trace of claims, sources, agent activity...

AgentsDGX agent

This post describes an AI research agent built on ActiveGraph that provides complete visibility into its reasoning process through detailed graphs and traces of claims, sources, and internal agent act

Merging model-based control with multi-agent reinforcement learning for multi-agent cooperative teaming strategies

AgentsDGX agent

arXiv:2606.06011v1 Announce Type: new Abstract: In this work, we propose a framework that combines multi-agent reinforcement learning (MARL) with model-based control to achieve safe, dynamically feasi

4 Jun 2026

Poke, which lets users access AI agents via text message, becomes the first AI agent approved for Apple's Messages for Business platform (Sarah Perez/TechCrunch)

AgentsDGX agent

Sarah Perez / TechCrunch: Poke, which lets users access AI agents via text message, becomes the first AI agent approved for Apple's Messages for Business platform — Poke, a startup that turns using AI

2 Jun 2026

Agent-R1: A Unified and Modular Framework for Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2511.14460v2 Announce Type: replace Abstract: Large language models (LLMs) have rapidly evolved from single-turn text generators into the foundation of increasingly capable agents. As these agen

Congrats @cognition on Devin Desktop! We've been prototyping with the team to connect local agents and @harvey's Spectre cloud agents, so co…

AgentsDGX agent

Congrats @cognition on Devin Desktop! We've been prototyping with the team to connect local agents and @harvey's Spectre cloud agents, so context can follow engineers across their stack. Excited about

1 Jun 2026

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal …

Model ReleasesDGX agent

👏👏 Introducing Qwen3.7-Plus — a multimodal agent model that unifies vision and language into one versatile agent foundation. ✅ Multimodal interactive hybrid agent: unified GUI & CLI operation across v

28 May 2026

Agents that Matter: Optimizing Multi-Agent LLMs via Removal-Based Attribution

SafetyDGX agent

arXiv:2605.27621v1 Announce Type: cross Abstract: As multi-agent systems (MAS) become increasingly complex, identifying the contributions of individual agents is critical for system optimization. Howe

27 May 2026

We spent a ton of time making worktrees actually work well for agent swarms and large repos. When you’re running 10s of agents that ship, yo…

AgentsDGX agent

We spent a ton of time making worktrees actually work well for agent swarms and large repos. When you’re running 10s of agents that ship, you quickly realize you need worktree, but git’s defaults are

Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems

Model ReleasesDGX agent

arXiv:2605.26302v1 Announce Type: new Abstract: Long-lived AI agents are increasingly deployed as persistent operational systems, yet they are still evaluated like freshly initialized models. Day-one

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering.…

Model ReleasesDGX agent

// Your Agents are Aging Too // Huh!? They need 'sleep,' and now they are aging? Joke aside, great write-up on reliable agentic engineering. This new research introduces AgingBench, a longitudinal rel

23 May 2026

Replit Agent builds your app. Squidler tests it like a real user. Replit Agent fixes what's broken. That's the full AI QA loop, and it's now…

AgentsDGX agent

Replit Agent builds your app. Squidler tests it like a real user. Replit Agent fixes what's broken. That's the full AI QA loop, and it's now live in Replit's MCP library. You describe what your app sh

20 May 2026

a few months back, it become clear to us that a large part of technical work would be driven by agents in the future. coding agents were bec…

AgentsDGX agent

a few months back, it become clear to us that a large part of technical work would be driven by agents in the future. coding agents were becoming ubiquitous and highly capable. since we build a platfo

Discoverable Agent Knowledge -- A Formal Framework for Agentic KG Affordances (Extended Version)

AgentsDGX agent

arXiv:2605.19186v1 Announce Type: new Abstract: Two decades ago, the Semantic Web Services community was asked how agents with different ontological commitments could discover, compose, and invoke web

Trustworthy Agent Network: Trust in Agent Networks Must Be Baked In, Not Bolted On

SafetyDGX agent

arXiv:2605.19035v1 Announce Type: new Abstract: The rapid advancement of Large Language Models has given rise to autonomous LLM-based agents capable of complex reasoning and execution. As these agents

19 May 2026

Deep Agents now integrates with @nebiusai Token Factory. Now, you can run agent workloads on production-grade AI infrastructure with open-so…

AgentsDGX agent

Deep Agents now integrates with @nebiusai Token Factory. Now, you can run agent workloads on production-grade AI infrastructure with open-source models, dedicated endpoints, real-time search, and full

Herding CATs: ALARA for Agent Harness Engineering in Portable Composable Multi-Agent Teams

Local AiDGX agent

arXiv:2603.20380v2 Announce Type: replace-cross Abstract: Industry practitioners and academic researchers regularly use multi-agent systems to accelerate their work, but the applications through which

NEW paper worth reading: MetaCogAgent MetaCogAgent equips a multi-agent system with metacognition so each agent decides whether it should an…

AgentsDGX agent

NEW paper worth reading: MetaCogAgent MetaCogAgent equips a multi-agent system with metacognition so each agent decides whether it should answer or delegate. In other words, it aims for self-aware tas

Taming 'Zombie'' Agents: A Markov State-Aware Framework for Resilient Multi-Agent Evolution

SafetyDGX agent

arXiv:2605.17348v1 Announce Type: new Abstract: Recent advancements in LLM-based multi-agent systems have demonstrated remarkable collaborative capabilities across complex tasks. To improve overall ef

this has probably been the most interesting project i've ever worked on. we created the best agent at finding problems with your agent. and …

AgentsDGX agent

this has probably been the most interesting project i've ever worked on. we created the best agent at finding problems with your agent. and we integrated it with the best observability platform to spi

18 May 2026

The Hermes Agent Kanban just got a big automation upgrade. Drop one prompt into the triage, and the orchestrator agent can take it from ther…

AgentsDGX agent

The Hermes Agent Kanban just got a big automation upgrade. Drop one prompt into the triage, and the orchestrator agent can take it from there - decomposing it into all the subtasks necessary and autom

15 May 2026

also thought this was cool from the creative hackathon. extended the dashboard into a frontend dev agent. giving an agent a tight use case a…

AgentsDGX agent

also thought this was cool from the creative hackathon. extended the dashboard into a frontend dev agent. giving an agent a tight use case and surfacing it via a mini chat right next to the artifact i

Gemini Live Agent Challenge: Announcing the winners and highlights

Model ReleasesDGX agent

The Gemini Live Agent Challenge is officially in the books! We challenged developers worldwide to break out of the traditional 'text box' paradigm by building next-generation AI agents. From our initi

13 May 2026

Has been a blast shipping Engine!! We built engine to automate the agent development loop, and make agents that improve autonomously. Absolu…

AgentsDGX agent

Has been a blast shipping Engine!! We built engine to automate the agent development loop, and make agents that improve autonomously. Absolutely cracked team shipping at light speed to create the futu

HTML Artifacts are a big part of how I work with agents now. Artifacts can be more than just static files. When combined with agents, they c…

Model ReleasesDGX agent

HTML Artifacts are a big part of how I work with agents now. Artifacts can be more than just static files. When combined with agents, they can take action or help you take action. This unlocks all kin

12 May 2026

Towards Shutdownable Agents: Generalizing Stochastic Choice in RL Agents and LLMs

Model ReleasesDGX agent

arXiv:2604.17502v2 Announce Type: replace Abstract: Misaligned artificial agents might resist shutdown. One proposed solution is to train agents to lack preferences between different-length trajectori

5 May 2026

Sources: Meta is building an OpenClaw-inspired agent internally called Hatch, to be powered by its Muse Spark model, and an agentic shopping tool in Instagram (Jyoti Mann/The Information)

AgentsDGX agent

Jyoti Mann / The Information: Sources: Meta is building an OpenClaw-inspired agent internally called Hatch, to be powered by its Muse Spark model, and an agentic shopping tool in Instagram — Meta Plat

We've raised $27M to build @CopilotKit — the Agentic Frontend Stack connecting humans & agents. Because all UI will be AI. Co-led by Glilot …

AgentsDGX agent

CopilotKit has secured $27 million in funding to develop an agentic frontend stack that integrates AI agents with human-facing user interfaces, reflecting the belief that all future UI will incorporat

3 May 2026

while exploring harnesses, i came across my first and favourite agent orchestration framework @LangChain has launched deep-agents last year.…

AgentsDGX agent

while exploring harnesses, i came across my first and favourite agent orchestration framework @LangChain has launched deep-agents last year. it’s a wonderful harness, that has a lot of tools: - write_

30 Apr 2026

serving multiple users from a single agent deployment introduces three distinct problems. luckily, langsmith's agent server has a solution f…

AgentsDGX agent

serving multiple users from a single agent deployment introduces three distinct problems. luckily, langsmith's agent server has a solution for each! 1. data isolation: your @auth.authenticate handler

← Previous
1…7891011…294
Next →