AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,770 results
Agents

Codex CLI now runs inside Devin Desktop. You can pair your Codex subscription with Devin Desktop to have native multi-agent orchestration. L…

DGX agent

Codex CLI now runs inside Devin Desktop. You can pair your Codex subscription with Devin Desktop to have native multi-agent orchestration. Learn more about using your favorite agents in Devin Desktop:

agentswindsurf--x
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Critic-R: Improving Agentic Search using Instruction-tuned Retrievers with Natural Language Introspective Feedback

DGX agent

arXiv:2606.00590v1 Announce Type: cross Abstract: Agentic search systems iteratively interact with retrieval models to answer complex queries. Despite substantial progress, optimizing retrievers for a

agentsarxiv-cs-ai
2 Jun 2026
Agents

Defenses & Enablers For Skill Injection Attacks on Terminal Based Agents

DGX agent

arXiv:2606.01567v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly rely on reusable skills i.e. documents describing task-specific procedures. However, this introduces a

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains

DGX agent

arXiv:2606.02357v1 Announce Type: cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that agents have learned to use tools. We argue that this interp

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Ev-Trust: An Evolutionarily Stable Trust Mechanism for Decentralized LLM-Based Multi-Agent Service Economies

DGX agent

arXiv:2512.16167v3 Announce Type: replace-cross Abstract: Decentralized LLM-based multi-agent service economies face three vulnerabilities that undermine traditional trust mechanisms: reduced cost of

agentsarxiv-cs-ai
2 Jun 2026
Agents

Introducing Devin Desktop: the next generation of Windsurf Manage fleets of local and cloud agents from one surface Support for any ACP-comp…

DGX agent

Introducing Devin Desktop: the next generation of Windsurf Manage fleets of local and cloud agents from one surface Support for any ACP-compatible agent With a full IDE for when you need to jump into

agentswindsurf--x
2 Jun 2026
Agents

Microsoft’s Project Solara is an OS for AI agent gadgets

DGX agent

Microsoft just announced 'Project Solara,' a new OS designed for gadgets that run AI agents, at Build 2026. The company is calling it 'a new platform built from the ground up to power agent-driven exp

agentsthe-verge-ai
2 Jun 2026
Model Releases

MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

DGX agent

arXiv:2606.01993v1 Announce Type: cross Abstract: Abundant procedural knowledge on the Web holds great potential for helping agents solve long-horizon tasks. However, such knowledge is often multimoda

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

DGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

TechGraphRAG: An Agentic Graph-Augmented RAG Framework for Technical Literature Reasoning

DGX agent

arXiv:2606.01613v1 Announce Type: cross Abstract: This paper presents an agentic retrieval-augmented generation (RAG) framework for domain-specific technical reasoning support, instantiated over a cur

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

DGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Agents

Beyond LLMs: Why Scalable Enterprise AI Adoption Depends on Agent Logic

DGX agent

Enterprise AI adoption at scale requires moving beyond large language models to implement agent logic systems that can handle complex reasoning, planning, and decision-making autonomously. Agent-based

agentshugging-face
1 Jun 2026
Agents

Connecting agents directly to Snowflake, Slack, or BigQuery via MCP/CLIs is a multi-token disaster. They default to brute-force exploration,…

DGX agent

Connecting agents directly to Snowflake, Slack, or BigQuery via MCP/CLIs is a multi-token disaster. They default to brute-force exploration, firing 20-30 tool calls just to rediscover context on every

agentspinecone--x
1 Jun 2026
Model Releases

Counterfactual Trace Auditing of LLM Agent Skills

DGX agent

arXiv:2605.11946v2 Announce Type: replace Abstract: Large Language Model agents are increasingly augmented with agent skills. Current evaluation methods for skills remain limited. Most deployed benchm

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval

DGX agent

arXiv:2605.31377v1 Announce Type: cross Abstract: Agentic Retrieval-Augmented Generation improves retrieval by integrating planning, tool use, and iterative reasoning, but existing agentic RAG methods

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

if you build any agent on activegraph, the trace is automatic and first-class, not bolted on

DGX agent

if you build any agent on activegraph, the trace is automatic and first-class, not bolted on a parallel experiment building a coding agent on top of @activegraphai. you can see everything flattened do

agentsyohei-nakajima--x
1 Jun 2026
Agents

LLM Anonymization Against Agentic Re-Identificatio

DGX agent

arXiv:2605.30848v1 Announce Type: cross Abstract: Agentic LLMs with web search change the threat model for text anonymization: weak contextual cues can become cross-referenceable evidence for re-ident

agentsarxiv-cs-cl
1 Jun 2026
Agents

More info about Search as Code in the Perplexity Agent API docs: https://docs.perplexity.ai/docs/agent-api/tools/sandbox

DGX agent

The Perplexity Agent API documentation includes a 'Search as Code' feature accessible through the sandbox tools section, enabling developers to integrate search functionality programmatically within a

agentsperplexity--x
1 Jun 2026
Agents

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Cover…

DGX agent

🧑‍⚖️Evaluating Deep Agents with LangSmith on AWS Great deep dive blog with our friends at AWS on evaluating DeepAgents with LangSmith Covers datapoint and evaluator design for longer horizon agents ht

agentsharrison-chase--x
31 May 2026
Agents

Learn more about the latest from @james_y_zou and our Frontier Agents Research team!

DGX agent

Learn more about the latest from @james_y_zou and our Frontier Agents Research team! To evaluate frontier AI agents, we need more complex tasks. But such tasks are also more prone to have design mista

agentstogether-ai--x
30 May 2026
Agents

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, an…

DGX agent

Step 3.7 Flash is now free for 30 days via Nous Portal It is a new MoE vision-language model focused on agent efficiency, coding, search, and multimodal workflows — and Hermes Agent users have been lo

agentsnous-research--x
30 May 2026
Agents

AgentSchool: An LLM-Powered Multi-Agent Simulation for Education

DGX agent

arXiv:2605.30144v1 Announce Type: new Abstract: Despite the rapid deployment of LLMs into classrooms, validating educational AI remains uniquely intractable: interventions act on developing learners w

agentsarxiv-cs-ai
29 May 2026
Agents

Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents

DGX agent

arXiv:2605.29927v1 Announce Type: cross Abstract: Despite recent advances, LLM-based web agents still struggle with limited exploration, omission of critical steps, and sensitivity to task constraints

agentsarxiv-cs-ai
29 May 2026
Agents

Governing Technical Debt in Agentic AI Systems

DGX agent

arXiv:2605.29129v1 Announce Type: new Abstract: Agentic AI systems are increasingly being explored as production infrastructure: they reason over multiple steps, call tools, act through workflows, and

agentsarxiv-cs-ai
29 May 2026
Agents

Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction

DGX agent

arXiv:2605.29960v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly leverage long term memory to support persistent and autonomous task execution. However, this capability

agentsarxiv-cs-ai
29 May 2026
Agents

make something agents want

DGX agent

make something agents want studying ActiveGraph by @yoheinakajima and apart from the concept the implementation itself is brilliant on the site it shares a prompt that will make my agent study docs, i

agentsyohei-nakajima--x
29 May 2026
Model Releases

OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

DGX agent

arXiv:2605.29253v1 Announce Type: new Abstract: Task success can hide process anomalies in real-world agent executions. An agent may pass the final task oracle while still accumulating unresolved ambi

model-releasesarxiv-cs-ai
29 May 2026
Agents

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

DGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

agentsarxiv-cs-ai
29 May 2026
Agents

a hot (cold at this point?) take that lead us to build this: every agent in the future will need a sandbox to connect to writing/executing c…

DGX agent

a hot (cold at this point?) take that lead us to build this: every agent in the future will need a sandbox to connect to writing/executing code is not just for coding agents! is useful for all sorts o

agentsharrison-chase--x
28 May 2026
Model Releases

A Unified Framework for the Evaluation of LLM Agentic Capabilities

DGX agent

arXiv:2605.27898v1 Announce Type: new Abstract: As LLMs are increasingly deployed as agents, reliable assessment of their agentic capabilities has become essential. However, reported benchmark scores

model-releasesarxiv-cs-ai
28 May 2026
Agents

Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents

DGX agent

arXiv:2605.22166v2 Announce Type: replace Abstract: LLM agents are shaped not only by their language models, but also by the runtime harness that mediates observation, tool use, action execution, feed

agentsarxiv-cs-ai
28 May 2026
Model Releases

Ask Now, Use Later: Benchmarking the Proactivity Gap in Long-Lived LLM Agents

DGX agent

arXiv:2605.28108v1 Announce Type: new Abstract: A long-lived LLM agent, such as OpenClaw, earns its value by acting on a user's preferences and constraints across sessions, not just the current reques

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

DGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

model-releasesarxiv-cs-ai
28 May 2026
Agents

From Instructor to Collaborator: What a 90-Participant Study Reveals about Human-Agent Collaboration in a Mobile Serious Game

DGX agent

arXiv:2605.27384v1 Announce Type: cross Abstract: This position paper reflects empirical data collected during my PhD from a large-scale within-subjects study (N = 90). The study compared a highly hum

agentsarxiv-cs-ai
28 May 2026
Agents

GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection

DGX agent

arXiv:2605.28534v1 Announce Type: new Abstract: Despite the rapid progress of multimodal large language models in building Graphical User Interface (GUI) agents, their real-world task completion is fu

agentsarxiv-cs-cl
28 May 2026
Agents

How Endava builds an agentic organization with Codex

DGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

agentsopenai
28 May 2026
Agents

Is grep 𝘳𝘦𝘢𝘭𝘭𝘺 all your AI agent needs for search? For a small codebase or a docs folder, the answer might be yes, but in most enterpr…

DGX agent

Is grep 𝘳𝘦𝘢𝘭𝘭𝘺 all your AI agent needs for search? For a small codebase or a docs folder, the answer might be yes, but in most enterprise environments, agents face millions of PDFs, spreadsheets, and

agentsjerry-liu--x
28 May 2026
Agents

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture'

DGX agent

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture' babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is t

agentsyohei-nakajima--x
28 May 2026
Agents

Orchid Security targets AI agent sprawl with new identity governance tools

DGX agent

Orchid Security Inc. today extended its Identity Control Plane with a set of capabilities aimed at governing artificial intelligence agents, saying existing identity and access management models canno

agentssiliconangle
28 May 2026
Agents

Adaptation-Free Heterogeneous Collaborative Perception with Unseen Agent Configurations

DGX agent

arXiv:2605.26642v1 Announce Type: new Abstract: Collaborative perception improves 3D object detection by enabling agents to share complementary observations, but most existing methods assume fixed or

agentsarxiv-cs-cv
27 May 2026
Model Releases

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

DGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

model-releasesarxiv-cs-ai
27 May 2026
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Agents

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

DGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

agentsarxiv-cs-ai
27 May 2026
Agents

i first wrote down the agent labs thesis last year https://x.com/swyx/status/1990886806250782876?s=46

DGX agent

i first wrote down the agent labs thesis last year https://x.com/swyx/status/1990886806250782876?s=46 New @latentspacepod Essay: why Agent Labs are clearly emerging in 2025 as a complement to Model La

agentsswyx--x
27 May 2026
Agents

It's crazy that this is even possible today. It inspired me to build my own self-improving coding agent with simple read, write, bash,... I …

DGX agent

It's crazy that this is even possible today. It inspired me to build my own self-improving coding agent with simple read, write, bash,... I already used the coding agent to build an entire production-

agentsdair-ai--x
27 May 2026
Agents

Lessons from Penetration Tests on Large-Scale Agent Systems

DGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of

agentsarxiv-cs-ai
27 May 2026
Agents

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

DGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

agentsarxiv-cs-ai
27 May 2026
Agents

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proac…

DGX agent

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proactive agent impl ive seen https://x.com/russelljkaplan/status

agentsswyx--x
27 May 2026
← Previous
1…6263646566…371
Next →