AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,608 results
7 Aug 2026

Agentic self-driving microscopy benchmarks support qualification but do not necessarily generalize to unseen tasks

Model ReleasesDGX agent

arXiv:2608.05266v1 Announce Type: new Abstract: Large language model agents are increasingly being developed to control a wide range of scientific characterization tools including microscopes and sync

When Do Prompt-Side Agent Playbooks Transfer? Accuracy, Cost, and Runtime Shift in Agent Deployment

AgentsDGX agent

arXiv:2608.05778v1 Announce Type: new Abstract: Prompt-side playbooks can improve tool-using language agents without retraining, but their portability beyond the source setting is unclear. We study fr

6 Aug 2026

AI agent observability: Why production systems need a reasoning layer

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
AgentsDGX agent

Traditional APM can collect every span and still leave developers guessing about intent, causality, and drift. As agents multiply, the observability stack must learn to interpret the systems it watche

Binding Biometrics with AI Agent Identifiers for Delegation of Authority

AgentsDGX agent

arXiv:2608.04292v1 Announce Type: new Abstract: The proliferation of agentic artificial intelligence (AI) systems has raised serious questions about the accountability for tasks performed by AI agents

From ranking to recommended: get your site ready to thrive in the age of AI agents

AgentsDGX agent

More than half of requests now come from machines, not people. Agent Readiness shows how well agents can discover and read your site, while Answer Engine Optimization tracks how often AI assistants re

Introducing Kitesurf: The agent-first browser that runs in V8 isolates on Cloudflare Workers

AgentsDGX agent

We should be giving all agents tools that excel at what’s important for an AI model. Kitesurf is Cloudflare’s new stateless, highly scalable, and cost-effective web browser that runs entirely on top o

31 Jul 2026

LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger

AgentsDGX agent

arXiv:2607.28374v1 Announce Type: new Abstract: Multimodal agents for visual question answering increasingly operate as multi-step trajectories that interleave perception, retrieval, and reasoning, ye

Multi-Agent Debate Strategies: Survey, Taxonomy, and Challenges

AgentsDGX agent

arXiv:2607.26212v1 Announce Type: cross Abstract: Multi-Agent Debate (MAD) is a promising paradigm for improving the accuracy and robustness of Large Language Model (LLM)-based agentic systems. It ena

30 Jul 2026

SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response

Model ReleasesDGX agent

arXiv:2607.26791v1 Announce Type: cross Abstract: Large Language Model (LLM) agents are increasingly adopted in real-world security operations with access to host artifacts and command-line interfaces

29 Jul 2026

Agent Retrieval Bench: Evaluating Repository Context Retrieval for Coding Agents

Model ReleasesDGX agent

arXiv:2607.24882v1 Announce Type: cross Abstract: Modern coding agents are usually evaluated by whether they eventually produce a correct patch, but patch generation depends on an earlier context-acqu

Beyond Memory: A Templated Substrate for Heterogeneous Collaborative Knowledge Work with LLM Agents

AgentsDGX agent

arXiv:2607.24759v1 Announce Type: new Abstract: Research projects, educational efforts, and adjacent knowledge work accumulate findings, decisions, and reasoning that future collaborators rarely recov

28 Jul 2026

ACM: Agentic Context Management for Long Horizon Tasks

AgentsDGX agent

arXiv:2607.23809v1 Announce Type: new Abstract: Agentic tasks are inherently long-horizon and multi-turn, constantly accumulating context through interactions with the environment. Existing context co

An Agentic Orchestration of Atomistic Simulations

AgentsDGX agent

arXiv:2607.22596v1 Announce Type: new Abstract: Atomistic simulations are central to materials design, but their execution involves complex, multi-step workflows that require significant human experti

AutoWorld: Learning Multi-Agent Traffic Simulation with Self-Supervised World Models

AgentsDGX agent

arXiv:2603.28963v2 Announce Type: replace-cross Abstract: Simulation with realistic traffic agents is essential for validating autonomous driving systems. Existing data-driven simulators learn agent b

Market surveillance agent with LangGraph and Strands on AgentCore

AgentsDGX agent

Learn how to architect and deploy a production-ready multi-agent AI system using LangGraph for workflow orchestration and Strands for agent reasoning on Amazon Bedrock AgentCore. This post walks throu

MemTX: Transactional Belief Commit for Stateful Agent Memory

SafetyDGX agent

arXiv:2607.23929v1 Announce Type: new Abstract: LLM agents increasingly coordinate through persistent shared memory: one agent's write becomes another agent's premise, and eventually a tool call with

26 Jul 2026

“The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!”

AgentsDGX agent

“The first autonomous agent cyberattack is an unprecedented event. It deserves an unprecedented response!” In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s rel

20 Jul 2026

Highly recommended. I've often claimed there's huge alpha in building agent harnesses. Turns out harnesses are compositional generalizers. T…

AgentsDGX agent

Highly recommended. I've often claimed there's huge alpha in building agent harnesses. Turns out harnesses are compositional generalizers. The RLM harness is an instance of this. This could lead to in

15 Jul 2026

Isolation as a First-Class Principle for LLM-Agent System Safety: Concepts, Taxonomy, Challenges and Future Directions

SafetyDGX agent

arXiv:2607.12406v1 Announce Type: new Abstract: The capability of LLM agents to function as the ``brain'' of a system fundamentally expands the scope of analysis beyond a standalone model. Consequentl

PalmClaw: A Native On-Device Agent Framework for Mobile Phones

Local AiDGX agent

arXiv:2607.13027v1 Announce Type: cross Abstract: Large Language Model (LLM) agents have moved beyond generating responses to executing multi-step tasks by calling tools, observing the results, and it

PM-Bench: Evaluating Prospective Memory in LLM Agents

Model ReleasesDGX agent

arXiv:2607.12385v1 Announce Type: new Abstract: A significant challenge in agentic AI is prospective memory: the ability to execute an intention at a specific future cue or state while other activitie

14 Jul 2026

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan o…

AgentsDGX agent

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan of Fireworks - how open-source agents can use a frontier mode

13 Jul 2026

Language models and coding agents are great, but there is more to life, and more to AI, than just LLM agents.

ResearchDGX agent

Hardmaru posted a tweet on July 13, 2026 stating that while language models and coding agents are valuable, “there is more to life, and more to AI, than just LLM agents.” The tweet received 61.1 k vie

9 Jul 2026

AGAPI-Agents: An Open-Access Agentic AI Platform for Accelerated Materials Design on AtomGPT.org

SafetyDGX agent

arXiv:2512.11935v2 Announce Type: replace Abstract: Agentic AI systems increasingly connect large language models (LLMs) to external scientific tools, yet whether and when tool access improves predict

TIR-Agent: Training an Explorative and Efficient Agent for Image Restoration

SafetyDGX agent

arXiv:2603.27742v2 Announce Type: replace Abstract: Vision-language agents that orchestrate specialized tools for image restoration (IR) have emerged as a promising method, yet most existing framework

7 Jul 2026

🎓 New course launch from LangChain Academy: Introduction to Deep Agents ✅ Learn what a harness is, and why agents need one ✅ Understand the…

TutorialsDGX agent

🎓 New course launch from LangChain Academy: Introduction to Deep Agents ✅ Learn what a harness is, and why agents need one ✅ Understand the 4 core capabilities of a harness ✅ Start building with Deep

6 Jul 2026

'Ghost memory' is a real problem with agents. You might have seen the issue where a long-running agent still confidently repeats a user fact…

Model ReleasesDGX agent

'Ghost memory' is a real problem with agents. You might have seen the issue where a long-running agent still confidently repeats a user fact that stopped being true weeks ago? New research names the f

3 Jul 2026

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates

SafetyDGX agent

arXiv:2607.02507v1 Announce Type: new Abstract: LLM agents will increasingly act in socially structured settings where role, audience, and relational context can shape what is advantageous or costly t

2 Jul 2026

Planning over MAPF Agent Dependencies via Multi-Dependency PIBT

AgentsDGX agent

arXiv:2603.23405v2 Announce Type: replace-cross Abstract: Modern Multi-Agent Path Finding (MAPF) algorithms must plan for hundreds to thousands of agents in congested environments within a second, req

30 Jun 2026

Whose Side Is Your Agent On? Multi-Party Principal Loyalty in LLM Agents

Model ReleasesDGX agent

arXiv:2606.30383v1 Announce Type: new Abstract: A rapidly growing class of LLM agents is multi-party: the agent acts for a principal (who briefs it, sends follow-ups, and receives results) while also

29 Jun 2026

Cloud CISO Perspectives: How Google Cloud Security uses AI internally

Model ReleasesDGX agent

Welcome to the second Cloud CISO Perspectives for June 2026. Today, we’re discussing how we use AI to chart a path to autonomous software development lifecycle security.As with all Cloud CISO Perspect

25 Jun 2026

Probabilistic Agents in Deterministic Audits: Evaluating Multi-Agent Systems for Automated Audits Based on the German IT-Grundschutz

AgentsDGX agent

arXiv:2606.25622v1 Announce Type: cross Abstract: The NIS-2 Directive mandates robust Risk Management from thousands of small and medium enterprises. To ensure compliance, companies rely on establishe

The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

SafetyDGX agent

arXiv:2606.24937v1 Announce Type: cross Abstract: The Hitchhiker's Guide to Agentic AI is a comprehensive practitioner's reference for building autonomous AI systems. The book covers the full stack fr

22 Jun 2026

Hermes Agent has reached 200,000 GitHub stars Thank you to our contributors, supporters, users, and agents!

AgentsDGX agent

Nous Research announced that their Hermes Agent project has achieved 200,000 GitHub stars, acknowledging the contributions of developers, supporters, and users who helped reach this milestone. The ann

11 Jun 2026

Multi-agent rendezvous in fluid flows via reinforcement learning

AgentsDGX agent

arXiv:2606.11274v1 Announce Type: cross Abstract: Rendezvous is a critical task for multi-agent systems, requiring agents to coordinate to meet at an unspecified location. However, achieving this in f

9 Jun 2026

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

SafetyDGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

You can try Claude Fable 5 as part of Devin Cloud’s Ultra agent. Devin Ultra is our smartest and most capable agent, which excels at long-ho…

Model ReleasesDGX agent

You can try Claude Fable 5 as part of Devin Cloud’s Ultra agent. Devin Ultra is our smartest and most capable agent, which excels at long-horizon tasks and debugging. We tuned the harness so Ultra cos

8 Jun 2026

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

SafetyDGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

Model ReleasesDGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

7 Jun 2026

The first wave of AI-native applications is wrapping tokens and providing in-app agents. As agent usage centralizes around core apps (e.g. C…

Model ReleasesDGX agent

The first wave of AI-native applications is wrapping tokens and providing in-app agents. As agent usage centralizes around core apps (e.g. Claude Code, Codex), there's this emerging wave of building s

6 Jun 2026

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

SafetyDGX agent

arXiv:2606.06223v1 Announce Type: new Abstract: Language-model agents act through repeated cycles of observation, reasoning, and action selection, making safety monitoring depend on both internal mode

5 Jun 2026

// Agents' Last Exam // Agents' Last Exam is a living benchmark of over 1,000 economically valuable tasks, built with 250+ industry experts …

Model ReleasesDGX agent

// Agents' Last Exam // Agents' Last Exam is a living benchmark of over 1,000 economically valuable tasks, built with 250+ industry experts and mapped to the U.S. federal occupational taxonomy. The ha

Seeing is Believing? Evaluating Vision-Language Model Susceptibility in Agent-to-Agent Multimodal Persuasion

Model ReleasesDGX agent

arXiv:2510.22768v2 Announce Type: replace Abstract: As autonomous agents increasingly interact, they inevitably attempt to influence one another. While prior work in text-only settings has explored th

4 Jun 2026

A couple of weeks ago we started rebuilding the 𝚑𝚏 CLI with AI agents in mind. it now detects when an agent is using it and gives clean, t…

Model ReleasesDGX agent

A couple of weeks ago we started rebuilding the 𝚑𝚏 CLI with AI agents in mind. it now detects when an agent is using it and gives clean, token-efficient output, next-command hints, and more, all desig

2 Jun 2026

Agent Operating Systems (AOS): Integrating Agentic Control Planes into, and Beyond, Traditional Operating Systems

SafetyDGX agent

arXiv:2606.01508v1 Announce Type: cross Abstract: Traditional operating systems were designed around deterministic programs, explicit control flow, and human initiated workflows. Their core abstractio

Build and run agents at scale with Microsoft Foundry at Build 2026

AgentsDGX agent

Learn how Microsoft Foundry helps developers build, deploy, and operate production-ready agents with Agent Framework, Toolboxes, hosted agents, Microsoft 365 distribution, observability, and agent opt

MARFT: Multi-Agent Reinforcement Fine-Tuning

AgentsDGX agent

arXiv:2504.16129v5 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (LaMAS) have demonstrated strong capabilities on complex agentic tasks requiring multifac

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02282v1 Announce Type: new Abstract: Orchestrating Large Language Models into Multi-Agent Systems (LLM-MAS) has unlocked remarkable reasoning capabilities, yet emergent failures and halluci

1 Jun 2026

Learning Agent-Compatible Context Management for Long-Horizon Tasks

AgentsDGX agent

arXiv:2605.30785v1 Announce Type: new Abstract: LLM agents increasingly face long-horizon tasks such as web search and deep research in real-world applications, where accumulated context can cause lon

30 May 2026

This recent @latentspacepod pod was a good one (they usually all are). One specific piece resonated: They were talking about AI agents writi…

AgentsDGX agent

This recent @latentspacepod pod was a good one (they usually all are). One specific piece resonated: They were talking about AI agents writing code, and the line was basically that without caution / d

29 May 2026

Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms

AgentsDGX agent

arXiv:2605.30169v1 Announce Type: cross Abstract: As autonomous language model agents proliferate, forming an emerging agentic web with real-world consequences, what credibility signals can you use to

28 May 2026

Almost everyone is building agent harness systems the wrong way. The default move: pick LangChain or LangGraph or the OpenAI Agents SDK, acc…

SafetyDGX agent

Almost everyone is building agent harness systems the wrong way. The default move: pick LangChain or LangGraph or the OpenAI Agents SDK, accept the loop, the tools, the memory, the orchestration, the

Got a Secret? LLM Agents Can't Keep It: Evaluating Privacy in Multi-Agent Systems

SafetyDGX agent

arXiv:2605.27766v1 Announce Type: new Abstract: LLM safety evaluations predominantly test models in isolation, yet deployed AI agents increasingly operate within persistent social environments alongsi

27 May 2026

Krea is now built in to Hermes Agent as an image generation API provider, allowing your agent to use Krea 2: a new foundation model trained …

Model ReleasesDGX agent

Krea is now built in to Hermes Agent as an image generation API provider, allowing your agent to use Krea 2: a new foundation model trained from scratch to balance aesthetic quality and fine control,

26 May 2026

Architecting Agentic Communities using Design Patterns

AgentsDGX agent

arXiv:2601.03624v3 Announce Type: replace Abstract: The rapid evolution of Large Language Models (LLM) and subsequent Agentic AI technologies requires systematic architectural guidance for building so

Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures

AgentsDGX agent

arXiv:2605.25435v1 Announce Type: new Abstract: The rapid evolution of large language model (LLM)-driven autonomous agents has given rise to OpenClaw, a new class of open-source agent frameworks that

22 May 2026

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic c…

Model ReleasesDGX agent

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic coding, reasoning, and long-horizon autonomy. AI natives can

21 May 2026

Managed Deep Agents is now in Private Beta ICYMI: It’s managed, model-agnostic infra for deep agents you can deploy with a single line of co…

ApplicationsDGX agent

Managed Deep Agents is a new LangChain feature in private beta that provides managed infrastructure for deep agents, designed to be model-agnostic and deployable with minimal code. The service aims to

20 May 2026

Informatica expands agentic AI strategy with headless data services and unified agent governance

AgentsDGX agent

In its first major announcement since being acquired by Salesforce Inc., data management vendor Informatica today unveiled a “headless” version of its flagship Intelligent Data Management Cloud, posit

19 May 2026

Calibrate-Then-Act: Cost-Aware Exploration in LLM Agents

AgentsDGX agent

arXiv:2602.16699v3 Announce Type: replace-cross Abstract: LLM agents are deployed in environments where they must interact to acquire information. In these scenarios, the agent must reason about inher

← Previous
1…1718192021…294
Next →