AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
Safety

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

DGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

safetyarxiv-cs-ai
10 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

Kimi K2.5: Visual Agentic Intelligence

DGX agent

arXiv:2602.02276v2 Announce Type: replace-cross Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint op

agentsarxiv-cs-ai
10 Aug 2026
Safety

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

DGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

safetyarxiv-cs-ai
7 Aug 2026
Agents

Code Is the Body: Agent-Owned Software Bodies for Recursive Evolution and Descent

DGX agent

arXiv:2607.28691v1 Announce Type: cross Abstract: Personalized AI agents are often configurable without giving users control over the artifacts that determine their future behavior. We present OurArk,

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up …

DGX agent

Nice benchmark to measure agentic e-commerce capabilities. They ran an agent for one simulated year of e-commerce operations and it ends up with 27.3% of the money a human makes. MerchantBench is a 36

model-releasesdair-ai--x
3 Aug 2026
Agents

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power…

DGX agent

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power and capability of agents, and the need to harden systems an

agentsgary-marcus--x
29 Jul 2026
Agents

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

DGX agent

arXiv:2607.21503v1 Announce Type: new Abstract: Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning co

agentsarxiv-cs-ai
24 Jul 2026
Applications

Data-Native AI Agents: Why Agents Must Move to Your Data

DGX agent

Enterprise AI agents lose effectiveness when they must access data outside of protected, governed systems. The article argues that truly reliable agents need to be *data‑native*—run directly inside th

applicationsdatabricks
15 Jul 2026
Agents

When and Why Does Multi-Agent Debate Fail and Does It Really Underperform?

DGX agent

arXiv:2510.20963v2 Announce Type: replace Abstract: Multi-agent debate (MAD) was proposed as a promising approach for ensembling the wisdom of multiple large language models (LLMs) to improve reasonin

agentsarxiv-cs-lg
15 Jul 2026
Model Releases

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk …

DGX agent

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk to anything with a phone number. Just one prompt. 15 seconds

model-releasesyohei-nakajima--x
14 Jul 2026
Agents

Ramp for Agents. Replit Agent can now incorporate your company, apply for Ramp, and get your business ready to spend, pay bills, and manage …

DGX agent

Ramp for Agents. Replit Agent can now incorporate your company, apply for Ramp, and get your business ready to spend, pay bills, and manage money. Every company used to start with paperwork. The next

agentsreplit--x
11 Jul 2026
Agents

agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the…

DGX agent

agent memory has always been reactive. OpenWiki makes it proactive. connect to sources, tell it what you care about, and your agent hits the ground running . what we're building is really exciting! tr

agentsharrison-chase--x
10 Jul 2026
Agents

Hermes Agent can now export your agent sessions, or sets of sessions, into a variety of formats and places. Get full conversations out in HT…

DGX agent

Hermes Agent can now export your agent sessions, or sets of sessions, into a variety of formats and places. Get full conversations out in HTML, Markdown, JSON and more, or upload entire datasets of yo

agentsclem-delangue--x
7 Jul 2026
Agents

Introducing Flint Ads Agent: we optimize your Google Ads spend for you. Our agent figures out what’s costing you conversions, using our comp…

DGX agent

Introducing Flint Ads Agent: we optimize your Google Ads spend for you. Our agent figures out what’s costing you conversions, using our comprehensive data and optimization engine. Then, it executes th

agentsharrison-chase--x
7 Jul 2026
Agents

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepage…

DGX agent

4/ programatic subagents in deepagents Deep Agents is our open source, model agnostic agent harness: https://github.com/langchain-ai/deepagents We added the ability to programatically call subagents.

agentsharrison-chase--x
2 Jul 2026
Agents

When AI Agents Compete for Jobs: Strategic Capabilities and Economic Dynamics of AI Labour Markets

DGX agent

arXiv:2512.04988v2 Announce Type: replace-cross Abstract: Emerging agentic marketplaces provide the economic infrastructure for matching and coordinating the large amounts of AI agents used in agentic

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents…

DGX agent

Agentic retrieval is changing the way retrieval-augmented applications are built, especially in domains like legal and fintech, where agents need to autonomously navigate large, evolving knowledge bas

model-releasesllamaindex--x
1 Jul 2026
Agents

get started with deep agents code, our open source coding agent with first class open model support

DGX agent

get started with deep agents code, our open source coding agent with first class open model support You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3

agentsharrison-chase--x
1 Jul 2026
Model Releases

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. …

DGX agent

Great paper on managing agent skills. Skill libraries keep growing, and picking the right skills has become a bottleneck for coding agents. The defaults are to expose the agent to the whole skill coll

model-releasesdair-ai--x
1 Jul 2026
Model Releases

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents

DGX agent

arXiv:2606.31179v1 Announce Type: new Abstract: As AI agents become increasingly capable of complex, long-horizon reasoning, rigorous and holistic evaluation is essential for measuring progress toward

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, …

DGX agent

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, proposed by @a1zhang from MIT, help: instead of working mode

agentsharrison-chase--x
1 Jul 2026
Agents

Giving a talk on agent-to-agent and AI network effects at @swyx 's AI Engineer World Fair today at 1:30p in Room 2010. Come say hi! I think …

DGX agent

Giving a talk on agent-to-agent and AI network effects at @swyx 's AI Engineer World Fair today at 1:30p in Room 2010. Come say hi! I think this talk will be a good one if I may say so myself. https:/

agentsswyx--x
30 Jun 2026
Agents

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a…

DGX agent

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a real, reproducible, isolated environment, many times in par

agentsharrison-chase--x
30 Jun 2026
Agents

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundan…

DGX agent

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundant processing steps; large pages are saved locally and paged

agentsnous-research--x
30 Jun 2026
Agents

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instea…

DGX agent

we launched code interpreters for deep agents last month. Basic idea is to let agents plan, delegate, and organize context using code instead of chained tool calls Code interpreters don't need a sandb

agentsharrison-chase--x
30 Jun 2026
Agents

Deep Agents now supports dynamic subagents. Instead of invoking subagents with tool calls, the main agent writes orchestration code to coord…

DGX agent

Deep Agents now supports dynamic subagents. Instead of invoking subagents with tool calls, the main agent writes orchestration code to coordinate work at scale. This enables workflows like processing

agentsharrison-chase--x
29 Jun 2026
Agents

Agentic coding forces you to design clean interfaces and document them well. An agent cannot read the implicit mental model shared by your e…

DGX agent

Agentic coding forces you to design clean interfaces and document them well. An agent cannot read the implicit mental model shared by your engineering team, it can only read your API contracts and doc

agentsfrancois-chollet--x
25 Jun 2026
Agents

Agentic Knowledge Tracing: A Multi-Agent LLM Architecture for Stealth Assessment of Financial Literacy in Serious Games

DGX agent

arXiv:2606.25358v1 Announce Type: new Abstract: Assessing financial literacy during gameplay without disrupting the learning experience remains a key challenge in serious games for education. We prese

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

Building Agent Skills is great, but testing them clearly and across agents is difficult. The Pinecone DevRel team has released Cultivar, a C…

DGX agent

Building Agent Skills is great, but testing them clearly and across agents is difficult. The Pinecone DevRel team has released Cultivar, a CLI tool and agent skill designed to solve this problem. With

model-releasespinecone--x
25 Jun 2026
Agents

memory is the key to continual learning for agents!! memory helps your agent improve over time, check out this guide on how to automate memo…

DGX agent

memory is the key to continual learning for agents!! memory helps your agent improve over time, check out this guide on how to automate memory development! i've been talking a lot about loops recently

agentsharrison-chase--x
24 Jun 2026
Agents

Your Hermes Agent can now adopt an animated pet: a small sprite that reacts to what the agent is doing (idle, running a tool, thinking, wait…

DGX agent

Your Hermes Agent can now adopt an animated pet: a small sprite that reacts to what the agent is doing (idle, running a tool, thinking, waiting, finishing, failing) in the GUI or TUI. You have nearly

agentsnous-research--x
24 Jun 2026
Agents

Deep Agents v0.6 feature spotlight: a code interpreter. Agents can now call tools from inside a runtime, keep intermediate results out of mo…

DGX agent

Deep Agents v0.6 feature spotlight: a code interpreter. Agents can now call tools from inside a runtime, keep intermediate results out of model context, and only pass the relevant output back. Fewer r

agentsharrison-chase--x
22 Jun 2026
Safety

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report…

DGX agent

Great report on LLM agent communication protocols. Communication is a huge bottleneck in multi-agent systems. (worth bookmarking) The report builds a five-dimensional taxonomy (counterparty, payload,

safetydair-ai--x
22 Jun 2026
Agents

If you are building agents, Max Agency with @hwchase17 is THE best behind the scenes of how the top agents are built at companies like @Sier…

DGX agent

If you are building agents, Max Agency with @hwchase17 is THE best behind the scenes of how the top agents are built at companies like @SierraPlatform @RampLabs @benchling @ListenLabs @_hex_tech @coge

agentsharrison-chase--x
22 Jun 2026
Agents

people want to build agents, and they want it to be easy even more important, it should be easy for your agents to improve over time that's …

DGX agent

people want to build agents, and they want it to be easy even more important, it should be easy for your agents to improve over time that's what the 4th loop ('hill climbing') enables. automating the

agentsharrison-chase--x
21 Jun 2026
Model Releases

How can we assess human-agent interactions? Case studies in software agent design

DGX agent

arXiv:2510.09801v3 Announce Type: replace Abstract: While benchmarks measure the accuracy of LLM-powered agents, they mostly assume full automation, failing to represent the collaborative nature of re

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

DGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

safetyarxiv-cs-ai
10 Jun 2026
Tools

And @DevangSachdev announced the @nebiusai Agents Blueprint and consists of - @LangChain Deep Agents - @tavilyai - @pinecone - @guardrails_a…

DGX agent

Nebius AI announced the Agents Blueprint, a framework integrating multiple tools including LangChain Deep Agents, Tavily AI, Pinecone for vector storage, and Guardrails for building AI agents. This bl

toolspinecone--x
9 Jun 2026
Agents

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers pro…

DGX agent

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers production-ready output in PPTX, Word, PDF, and Excel, straight

agentskimi-moonshot--x
8 Jun 2026
Model Releases

What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

DGX agent

arXiv:2606.05304v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models are typically organized around roles, pipelines, and turn schedules, while the content that age

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Will the Agent Recuse Itself? Measuring LLM-Agent Compliance with In-Band Access-Deny Signals

DGX agent

arXiv:2606.06460v1 Announce Type: cross Abstract: As autonomous LLM agents increasingly hold real credentials and operate infrastructure without a human in the loop, operators have no standard way to

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

Sandboxes are already helping teams move from agents that answer questions to agents that can do work safely. At @mondaydotcom, that means g…

DGX agent

Sandboxes are already helping teams move from agents that answer questions to agents that can do work safely. At @mondaydotcom, that means giving Sidekick a secure environment to write and run code fo

agentsharrison-chase--x
5 Jun 2026
Agents

1/5 Our latest Labs in Front piece: Agent pipeline order matters. By reversing a common agent recipe - scale first, enrich second - we reach…

DGX agent

AI21 Labs discusses how the order of operations in agent pipelines affects performance, presenting findings that reversing the typical 'scale first, enrich second' approach by instead enriching agent

agentsai21-labs--x
4 Jun 2026
Agents

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not posses…

DGX agent

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not possession of the credential. Control plane outside runtime. Fail

agentsharrison-chase--x
4 Jun 2026
Model Releases

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

DGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

model-releasesarxiv-cs-lg
3 Jun 2026
Model Releases

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criter…

DGX agent

Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM judges that check an agent’s work against rubric criteria: they're used both in agent benchmarking and as reward si

model-releasesharrison-chase--x
2 Jun 2026
Model Releases

The fully-managed Remote MCP Server for AlloyDB is now Generally Available

DGX agent

AI agents possess incredible reasoning capabilities and can perform increasingly complex actions. But the reliability of agentic outcomes depends entirely on the quality of the context they can access

model-releasesgoogle-cloud-ai
1 Jun 2026
← Previous
1…1314151617…367
Next →