AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
28 May 2026

How Endava builds an agentic organization with Codex

AgentsDGX agent

Endava, a software services company, leverages OpenAI's Codex to transform its organizational operations into an agentic model where AI agents autonomously handle tasks and decision-making. The approa

Is grep 𝘳𝘦𝘢𝘭𝘭𝘺 all your AI agent needs for search? For a small codebase or a docs folder, the answer might be yes, but in most enterpr…

AgentsDGX agent

Is grep 𝘳𝘦𝘢𝘭𝘭𝘺 all your AI agent needs for search? For a small codebase or a docs folder, the answer might be yes, but in most enterprise environments, agents face millions of PDFs, spreadsheets, and

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture'

AgentsDGX agent

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture' babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is t

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Orchid Security targets AI agent sprawl with new identity governance tools

AgentsDGX agent

Orchid Security Inc. today extended its Identity Control Plane with a set of capabilities aimed at governing artificial intelligence agents, saying existing identity and access management models canno

27 May 2026

Adaptation-Free Heterogeneous Collaborative Perception with Unseen Agent Configurations

AgentsDGX agent

arXiv:2605.26642v1 Announce Type: new Abstract: Collaborative perception improves 3D object detection by enabling agents to share complementary observations, but most existing methods assume fixed or

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

Model ReleasesDGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

SafetyDGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

AgentsDGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

i first wrote down the agent labs thesis last year https://x.com/swyx/status/1990886806250782876?s=46

AgentsDGX agent

i first wrote down the agent labs thesis last year https://x.com/swyx/status/1990886806250782876?s=46 New @latentspacepod Essay: why Agent Labs are clearly emerging in 2025 as a complement to Model La

It's crazy that this is even possible today. It inspired me to build my own self-improving coding agent with simple read, write, bash,... I …

AgentsDGX agent

It's crazy that this is even possible today. It inspired me to build my own self-improving coding agent with simple read, write, bash,... I already used the coding agent to build an entire production-

Lessons from Penetration Tests on Large-Scale Agent Systems

AgentsDGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proac…

AgentsDGX agent

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proactive agent impl ive seen https://x.com/russelljkaplan/status

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

AgentsDGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

Xage extends zero trust to autonomous AI agents across cloud, SaaS and edge

AgentsDGX agent

Zero-trust cybersecurity company Xage Security Inc. today unveiled new capabilities in its platform designed to give enterprises deterministic visibility into autonomous artificial intelligence agents

26 May 2026

A Token/KV-Cache Communication Media Selection and Resource Allocation Strategy for Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.25422v1 Announce Type: cross Abstract: The convergence of large language models (LLMs) with 6G networks is fostering a paradigm of autonomous multi-agent cooperation, which in turn is expec

Build highly scalable serverless LangGraph multi-agent systems in AWS with Amazon Bedrock AgentCore

AgentsDGX agent

In this post, we provide a solution to build highly scalable, serverless multi-agent generative AI systems on AWS using LangGraph Agents as orchestrators integrated with Amazon Bedrock AgentCore Memor

CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures

AgentsDGX agent

arXiv:2605.25338v1 Announce Type: cross Abstract: Large language model (LLM) agents frequently fail on multi-step tasks involving reasoning, tool use, and environment interaction. While such failures

DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

Model ReleasesDGX agent

arXiv:2605.25188v1 Announce Type: new Abstract: Multi-agent LLM systems improve reasoning by combining outputs from multiple agents, but interaction-heavy methods can introduce error propagation and h

HyLaT: Efficient Multi-Agent Communication via Hybrid Latent-Text Protocol

AgentsDGX agent

arXiv:2605.25421v1 Announce Type: new Abstract: Communication protocol design is a central challenge in large language model-based multi-agent systems. Existing single-channel approaches face an inher

New agent skill to convert YouTube videos to slides and notes.

AgentsDGX agent

New agent skill to convert YouTube videos to slides and notes. Just built an insane new agent skill. It can perfectly extract slides from YT videos, then write notes, images, transcripts, and slides i

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

AgentsDGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

Proper Scoring Rules for Agentic Uncertainty Quantification

AgentsDGX agent

arXiv:2605.24756v1 Announce Type: new Abstract: Language-model agents increasingly emit uncertainty signals throughout a trajectory, but existing agentic UQ evaluations often conflate ranking usefulne

Rethinking organizational design in the age of agentic AI

AgentsDGX agent

Amid rapidly growing adoption of enterprise-level AI agents, there’s a disconnect emerging between ambition and execution. Although 85% of organizations say they want to be agentic within the next thr

25 May 2026

ARMS: Automatic Reward Shaping for Sparse-Reward Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.23562v1 Announce Type: cross Abstract: Sparse rewards are a major bottleneck in multi-agent reinforcement learning (MARL), where simultaneous learning induces non-stationarity and makes rew

Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents

AgentsDGX agent

arXiv:2605.23590v1 Announce Type: new Abstract: ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which r

From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.23382v1 Announce Type: new Abstract: Agentic reinforcement learning (Agentic RL) has achieved strong progress in tasks with clear success signals. However, many real-world agent application

/goal is really insane! It's how you can get the most out of coding agents today. For efficiency, I find it works best when you do planning …

AgentsDGX agent

/goal is really insane! It's how you can get the most out of coding agents today. For efficiency, I find it works best when you do planning before /goal. This ensures the agent has the right context a

Harness, Scaffold, and the AI Agent Terms Worth Getting Right

AgentsDGX agent

This article defines and clarifies key terminology related to AI agents, including the concepts of 'harness' and 'scaffold,' which are important architectural and operational components in building an

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

Model ReleasesDGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

Redrawing the AI Map: A Theory of Accountability Boundaries in Agentic Ecosystems

AgentsDGX agent

arXiv:2605.23179v1 Announce Type: new Abstract: Agentic AI orchestrators reduce the interface and assembly costs of composing information systems capabilities across organizational boundaries, seeming

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2602.01665v2 Announce Type: replace-cross Abstract: The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (

23 May 2026

Agents shouldn’t have direct visibility into env vars or credentials that can expose sensitive systems and data. Keeping secrets outside the…

AgentsDGX agent

Agents shouldn’t have direct visibility into env vars or credentials that can expose sensitive systems and data. Keeping secrets outside the agent’s context helps secure the env while still allowing a

22 May 2026

Agents are too good at sounding right. And terrible at proving what actually happened, especially on complex, long-running tasks. So I built…

AgentsDGX agent

Agents are too good at sounding right. And terrible at proving what actually happened, especially on complex, long-running tasks. So I built a fully traceable and forkable research agent with @yoheina

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

Model ReleasesDGX agent

arXiv:2605.22138v1 Announce Type: cross Abstract: How should an agent decide when and how to plan? A dominant approach builds agents as reactive policies with adaptive computation (e.g., chain-of-thou

From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)

AgentsDGX agent

arXiv:2605.20608v1 Announce Type: new Abstract: Realizing Level 4/5 Autonomous Networks (AN) demands a shift from static automation to agent-native intelligence. Current operations, reliant on rigid s

OPERA: An Agent for Image Restoration with End-to-End Joint Planning-Execution Optimization

AgentsDGX agent

arXiv:2605.22104v1 Announce Type: new Abstract: Real-world image restoration is challenging due to complex and interacting mixed degradations. Recent agent-based approaches address this problem by com

We're hosting an agentic AI meetup in LA on May 28th — 5–7pm at Gulp in Playa Vista. Builders, founders, engineers. Drinks, no fluff. RAG sy…

AgentsDGX agent

We're hosting an agentic AI meetup in LA on May 28th — 5–7pm at Gulp in Playa Vista. Builders, founders, engineers. Drinks, no fluff. RAG systems, agentic workflows, or just starting out — all welcome

21 May 2026

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents

AgentsDGX agent

arXiv:2605.20616v1 Announce Type: new Abstract: Language agents increasingly operate over streams of related tasks, yet existing memory systems struggle to convert accumulated experience into reusable

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.…

Model ReleasesDGX agent

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.7 and GPT-5.5 variants above it. This release puts Composer

datasette-agent 0.1a3

AgentsDGX agent

Release: datasette-agent 0.1a3 'View SQL query' buttons for both visible tables and collapsed SQL result tool calls. Don't display empty reasoning chunks Improved handling of truncated responses - tab

🆕Daytona’s Agent-Native Compute: 60ms sandboxes, 50K startups in 75 sec, 850K daily runs, RL/evals, CLI > MCP, & the end of localhost https…

AgentsDGX agent

🆕Daytona’s Agent-Native Compute: 60ms sandboxes, 50K startups in 75 sec, 850K daily runs, RL/evals, CLI > MCP, & the end of localhost https://latent.space/p/daytona @daytonaio CEO @ivanburazin explain

@Replit longer form video on Active Graph [7 min 22 sec] - flip the agent architecture - 1970s blackboard system - rollback, fork, diff agen…

AgentsDGX agent

@Replit longer form video on Active Graph [7 min 22 sec] - flip the agent architecture - 1970s blackboard system - rollback, fork, diff agent runs - experience + behaviors + beliefs = you - behaviors

20 May 2026

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

AgentsDGX agent

arXiv:2605.18770v1 Announce Type: cross Abstract: We present a collaborative agentic GraphRAG framework for expert analysis of commercial registry data. Public registries are often formally accessible

and the article before talked even more broadly about the state of stateful agents, which inspired this project https://x.com/yoheinakajima/…

AgentsDGX agent

This post references a previous article discussing stateful agents broadly, which served as inspiration for Yohei Nakajima's project. The post likely documents how earlier discourse on agent architect

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

Model ReleasesDGX agent

arXiv:2605.19418v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM ag

Exclusive: Juicebox autonomous recruiting agents help source candidates proactively

AgentsDGX agent

Juicebox App Inc. today introduced a suite of autonomous recruiting agents that help hiring teams proactively identify and engage candidates across multiple open roles simultaneously, recommending pro

I highly recommend this. The Agentic Review is a new podcast from @QodoAI hosted by Itamar Friedman and Nnenna Ndukwe, and it's a great AI c…

AgentsDGX agent

I highly recommend this. The Agentic Review is a new podcast from @QodoAI hosted by Itamar Friedman and Nnenna Ndukwe, and it's a great AI coding show that's neither hype nor doom. It's honest convers

ICYMI: LangSmith Sandboxes are GA ✅ Agents get a real filesystem, shell, and package manager. Isolated from your infra. ✅ Works with Deep Ag…

AgentsDGX agent

ICYMI: LangSmith Sandboxes are GA ✅ Agents get a real filesystem, shell, and package manager. Isolated from your infra. ✅ Works with Deep Agents, Open SWE, or your own code. ✅ Auth with the same API k

PAVE: A Cognitive Architecture for Legitimate Violation in Generative Agent Societies

AgentsDGX agent

arXiv:2605.19351v1 Announce Type: cross Abstract: Generative agents based on large language models reproduce believable human behavior in cooperative settings, but how they should reason in situations

we've been building agents around the llm, starting with conversations, adding tools, giving rules, logging everything, and storing a form o…

AgentsDGX agent

we've been building agents around the llm, starting with conversations, adding tools, giving rules, logging everything, and storing a form of it as retrievable state Active Graphs flips this. what if

19 May 2026

AgentWall: A Runtime Safety Layer for Local AI Agents

Model ReleasesDGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

DocOS: Towards Proactive Document-Guided Actions in GUI Agents

Model ReleasesDGX agent

arXiv:2605.18048v1 Announce Type: new Abstract: While Graphical User Interface (GUI) agents have shown promising performance in automated device interaction, they primarily depend on static parametric

extsc{MasFACT}: Continual Multi-Agent Topology Learning via Geometry-Aware Posterior Transfer

AgentsDGX agent

arXiv:2605.17361v1 Announce Type: cross Abstract: Multi-agent systems (MAS) powered by large language models (LLMs) have emerged as a powerful paradigm for complex problem solving, where performance c

MemRepair: Hierarchical Memory for Agentic Repository-Level Vulnerability Repair

AgentsDGX agent

arXiv:2605.17444v1 Announce Type: cross Abstract: Modern software ecosystems face a rapidly growing number of disclosed vulnerabilities, increasing the need for automated repair techniques that can op

OProver: A Unified Framework for Agentic Formal Theorem Proving

AgentsDGX agent

arXiv:2605.17283v1 Announce Type: cross Abstract: Recent progress in formal theorem proving has benefited from large-scale proof generation and verifier-aware training, but agentic proving is rarely i

Read more on what we've learned while developing long-horizon agent evals:

AgentsDGX agent

Read more on what we've learned while developing long-horizon agent evals: Curious finding while creating evals and benchmarks for long-horizon (100+ turn) agents While it’s generally thought that a d

Reversa: A Reverse Documentation Engineering Framework for Converting Legacy Software into Operational Specifications for AI Agents

AgentsDGX agent

arXiv:2605.18684v1 Announce Type: cross Abstract: Legacy systems concentrate business rules, architectural decisions, and operational exceptions that often remain implicit in code, data, configuration

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

Model ReleasesDGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

The two most used branding phrases at Google IO: 1) “agent-first” 2) “at the speed of voice” #Google

AgentsDGX agent

At Google I/O, two dominant branding phrases emerged: 'agent-first,' emphasizing AI agents as a primary computing paradigm, and 'at the speed of voice,' highlighting voice-based interaction speed and

← Previous
1…5051525354…297
Next →