AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
1 May 2026

Build agents with LangChain + @browserbase. Give your Deep Agents search, fetch, and browser subagents to access the full web. All with full…

AgentsDGX agent

This post announces the ability to build AI agents using LangChain integrated with Browserbase, enabling agents to perform web searches, retrieve content, and control browser subagents to access and i

30 Apr 2026

AgentTrove: new agentic dataset with 1.7M samples Thanks to OpenThoughts for this great work The @huggingface Hub needs more agentic dataset…

AgentsDGX agent

AgentTrove is a newly released agentic dataset containing 1.7 million samples, designed to address the shortage of agentic datasets available on Hugging Face Hub. The dataset aims to support training

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale

ResearchDGX agent

Safe agents don’t guarantee a safe ecosystem of interconnected agents. Microsoft Research examines what breaks when AI agents interact and why network-level risks require new approaches. The post Red-

29 Apr 2026

AI Agent Compliance & Governance Layer Autonomous compliance agents evolve from tools into always-on governance infrastructure. As regulatio…

AgentsDGX agent

AI Agent Compliance & Governance Layer Autonomous compliance agents evolve from tools into always-on governance infrastructure. As regulation accelerates (AI, ESG, cross-border data, tax), the bottlen

Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, and multimodal reasoning. AI na…

AgentsDGX agent

Introducing Qwen3.6-Plus from @Alibaba_Qwen, a 1M-context model built for real-world agents, agentic coding, and multimodal reasoning. AI natives can now use Qwen3.6-Plus on Together AI and benefit fr

16 Apr 2026

Need to secure your production agents? @Cisco AI Defense Runtime Protection can integrate into your LangChain agents through middleware and …

ApplicationsDGX agent

Need to secure your production agents? @Cisco AI Defense Runtime Protection can integrate into your LangChain agents through middleware and produce a consistent runtime contract. 📄 What you need to kn

14 Apr 2026

Agent-Dice: Disentangling Knowledge Updates via Geometric Consensus for Agent Continual Learning

Model ReleasesDGX agent

arXiv:2601.03641v4 Announce Type: replace Abstract: Large Language Model (LLM)-based agents significantly extend the utility of LLMs by interacting with dynamic environments. However, enabling agents

12 Aug 2026

Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

AgentsDGX agent

arXiv:2608.10299v1 Announce Type: new Abstract: Agentic systems are increasingly expected to improve after deployment, yet single-entity self-evolution is often bounded by a static learning context, s

11 Aug 2026

El Agente Grafico: A Semantic Execution Runtime for Scientific Agents

AgentsDGX agent

arXiv:2602.17902v2 Announce Type: replace Abstract: Large language models (LLMs) can plan scientific workflows and generate code, but these capabilities do not specify how scientific state is validate

We built the Agentic World Cup - LLMs that compete in 1v1 Soccer. [P]

AgentsDGX agent

Hey everyone - we've been building something particularly relevant to ML at large - The Agentic World Cup - a platform where Agents compete in sports. As you know, today's Agents can code, do math, an

10 Aug 2026

Agentic AI: User Empowerment or Enclosure?

AgentsDGX agent

arXiv:2608.06510v1 Announce Type: cross Abstract: Agentic AI promises a more flexible form of digital agency: systems that can act on users' behalf, from filtering content to negotiating prices to sel

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

SafetyDGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

Kimi K2.5: Visual Agentic Intelligence

AgentsDGX agent

arXiv:2602.02276v2 Announce Type: replace-cross Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint op

7 Aug 2026

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

SafetyDGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

3 Aug 2026

Code Is the Body: Agent-Owned Software Bodies for Recursive Evolution and Descent

AgentsDGX agent

arXiv:2607.28691v1 Announce Type: cross Abstract: Personalized AI agents are often configurable without giving users control over the artifacts that determine their future behavior. We present OurArk,

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

Model ReleasesDGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up …

Model ReleasesDGX agent

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up with 27.3% of the money a human makes. MerchantBench is a 36

29 Jul 2026

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power…

AgentsDGX agent

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power and capability of agents, and the need to harden systems an

24 Jul 2026

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

AgentsDGX agent

arXiv:2607.21503v1 Announce Type: new Abstract: Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning co

15 Jul 2026

Data-Native AI Agents: Why Agents Must Move to Your Data

ApplicationsDGX agent

Enterprise AI agents lose effectiveness when they must access data outside of protected, governed systems. The article argues that truly reliable agents need to be *data‑native*—run directly inside th

When and Why Does Multi-Agent Debate Fail and Does It Really Underperform?

AgentsDGX agent

arXiv:2510.20963v2 Announce Type: replace Abstract: Multi-agent debate (MAD) was proposed as a promising approach for ensembling the wisdom of multiple large language models (LLMs) to improve reasonin

14 Jul 2026

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk …

Model ReleasesDGX agent

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk to anything with a phone number. Just one prompt. 15 seconds

11 Jul 2026

Ramp for Agents. Replit Agent can now incorporate your company, apply for Ramp, and get your business ready to spend, pay bills, and manage …

AgentsDGX agent

Ramp for Agents. Replit Agent can now incorporate your company, apply for Ramp, and get your business ready to spend, pay bills, and manage money. Every company used to start with paperwork. The next

10 Jul 2026

agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the…

AgentsDGX agent

agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the ground running . what we're building is really exciting! tr

7 Jul 2026

Hermes Agent can now export your agent sessions, or sets of sessions, into a variety of formats and places. Get full conversations out in HT…

AgentsDGX agent

Hermes Agent can now export your agent sessions, or sets of sessions, into a variety of formats and places. Get full conversations out in HTML, Markdown, JSON and more, or upload entire datasets of yo

Introducing Flint Ads Agent: we optimize your Google Ads spend for you. Our agent figures out what’s costing you conversions, using our comp…

AgentsDGX agent

Introducing Flint Ads Agent: we optimize your Google Ads spend for you. Our agent figures out what’s costing you conversions, using our comprehensive data and optimization engine. Then, it executes th

2 Jul 2026

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepage…

AgentsDGX agent

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepagents We added the ability to programatically call subagents.

When AI Agents Compete for Jobs: Strategic Capabilities and Economic Dynamics of AI Labour Markets

AgentsDGX agent

arXiv:2512.04988v2 Announce Type: replace-cross Abstract: Emerging agentic marketplaces provide the economic infrastructure for matching and coordinating the large amounts of AI agents used in agentic

1 Jul 2026

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents…

Model ReleasesDGX agent

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents need to autonomously navigate large, evolving knowledge bas

get started with deep agents code, our open source coding agent with first class open model support

AgentsDGX agent

get started with deep agents code, our open source coding agent with first class open model support You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. …

Model ReleasesDGX agent

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. The defaults are to expose the agent to the whole skill coll

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents

Model ReleasesDGX agent

arXiv:2606.31179v1 Announce Type: new Abstract: As AI agents become increasingly capable of complex, long-horizon reasoning, rigorous and holistic evaluation is essential for measuring progress toward

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, …

AgentsDGX agent

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, proposed by @a1zhang from MIT, help: instead of working mode

30 Jun 2026

Giving a talk on agent-to-agent and AI network effects at @swyx 's AI Engineer World Fair today at 1:30p in Room 2010. Come say hi! I think …

AgentsDGX agent

Giving a talk on agent-to-agent and AI network effects at @swyx 's AI Engineer World Fair today at 1:30p in Room 2010. Come say hi! I think this talk will be a good one if I may say so myself. https:/

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a…

AgentsDGX agent

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a real, reproducible, isolated environment, many times in par

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundan…

AgentsDGX agent

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundant processing steps; large pages are saved locally and paged

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instea…

AgentsDGX agent

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instead of chained tool calls Code interpreters don't need a sandb

29 Jun 2026

Deep Agents now supports dynamic subagents. Instead of invoking subagents with tool calls, the main agent writes orchestration code to coord…

AgentsDGX agent

Deep Agents now supports dynamic subagents. Instead of invoking subagents with tool calls, the main agent writes orchestration code to coordinate work at scale. This enables workflows like processing

25 Jun 2026

Agentic coding forces you to design clean interfaces and document them well. An agent cannot read the implicit mental model shared by your e…

AgentsDGX agent

Agentic coding forces you to design clean interfaces and document them well. An agent cannot read the implicit mental model shared by your engineering team, it can only read your API contracts and doc

Agentic Knowledge Tracing: A Multi-Agent LLM Architecture for Stealth Assessment of Financial Literacy in Serious Games

AgentsDGX agent

arXiv:2606.25358v1 Announce Type: new Abstract: Assessing financial literacy during gameplay without disrupting the learning experience remains a key challenge in serious games for education. We prese

Building Agent Skills is great, but testing them clearly and across agents is difficult. The Pinecone DevRel team has released Cultivar, a C…

Model ReleasesDGX agent

Building Agent Skills is great, but testing them clearly and across agents is difficult. The Pinecone DevRel team has released Cultivar, a CLI tool and agent skill designed to solve this problem. With

24 Jun 2026

memory is the key to continual learning for agents!! memory helps your agent improve over time, check out this guide on how to automate memo…

AgentsDGX agent

memory is the key to continual learning for agents!! memory helps your agent improve over time, check out this guide on how to automate memory development! i've been talking a lot about loops recently

Your Hermes Agent can now adopt an animated pet: a small sprite that reacts to what the agent is doing (idle, running a tool, thinking, wait…

AgentsDGX agent

Your Hermes Agent can now adopt an animated pet: a small sprite that reacts to what the agent is doing (idle, running a tool, thinking, waiting, finishing, failing) in the GUI or TUI. You have nearly

22 Jun 2026

Deep Agents v0.6 feature spotlight: a code interpreter. Agents can now call tools from inside a runtime, keep intermediate results out of mo…

AgentsDGX agent

Deep Agents v0.6 feature spotlight: a code interpreter. Agents can now call tools from inside a runtime, keep intermediate results out of model context, and only pass the relevant output back. Fewer r

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report…

SafetyDGX agent

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report builds a five-dimensional taxonomy (counterparty, payload,

If you are building agents, Max Agency with @hwchase17 is THE best behind the scenes of how the top agents are built at companies like @Sier…

AgentsDGX agent

If you are building agents, Max Agency with @hwchase17 is THE best behind the scenes of how the top agents are built at companies like @SierraPlatform @RampLabs @benchling @ListenLabs @_hex_tech @coge

21 Jun 2026

people want to build agents, and they want it to be easy even more important, it should be easy for your agents to improve over time that's …

AgentsDGX agent

people want to build agents, and they want it to be easy even more important, it should be easy for your agents to improve over time that's what the 4th loop ('hill climbing') enables. automating the

10 Jun 2026

How can we assess human-agent interactions? Case studies in software agent design

Model ReleasesDGX agent

arXiv:2510.09801v3 Announce Type: replace Abstract: While benchmarks measure the accuracy of LLM-powered agents, they mostly assume full automation, failing to represent the collaborative nature of re

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

SafetyDGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

9 Jun 2026

And @DevangSachdev announced the @nebiusai Agents Blueprint and consists of - @LangChain Deep Agents - @tavilyai - @pinecone - @guardrails_a…

ToolsDGX agent

Nebius AI announced the Agents Blueprint, a framework integrating multiple tools including LangChain Deep Agents, Tavily AI, Pinecone for vector storage, and Guardrails for building AI agents. This bl

8 Jun 2026

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers pro…

AgentsDGX agent

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers production-ready output in PPTX, Word, PDF, and Excel, straight

6 Jun 2026

What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.05304v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models are typically organized around roles, pipelines, and turn schedules, while the content that age

Will the Agent Recuse Itself? Measuring LLM-Agent Compliance with In-Band Access-Deny Signals

Model ReleasesDGX agent

arXiv:2606.06460v1 Announce Type: cross Abstract: As autonomous LLM agents increasingly hold real credentials and operate infrastructure without a human in the loop, operators have no standard way to

5 Jun 2026

Sandboxes are already helping teams move from agents that answer questions to agents that can do work safely. At @mondaydotcom, that means g…

AgentsDGX agent

Sandboxes are already helping teams move from agents that answer questions to agents that can do work safely. At @mondaydotcom, that means giving Sidekick a secure environment to write and run code fo

4 Jun 2026

1/5 Our latest Labs in Front piece: Agent pipeline order matters. By reversing a common agent recipe - scale first, enrich second - we reach…

AgentsDGX agent

AI21 Labs discusses how the order of operations in agent pipelines affects performance, presenting findings that reversing the typical 'scale first, enrich second' approach by instead enriching agent

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not posses…

AgentsDGX agent

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not possession of the credential. Control plane outside runtime. Fail

3 Jun 2026

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

Model ReleasesDGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

2 Jun 2026

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criter…

Model ReleasesDGX agent

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criteria: they're used both in agent benchmarking and as reward si

1 Jun 2026

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

Model ReleasesDGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

← Previous
1…1011121314…294
Next →