AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
26 May 2026

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

SafetyDGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents

Model ReleasesDGX agent

arXiv:2605.24069v1 Announce Type: cross Abstract: The rise of tool-using Large Language Model (LLM) agents, standardized by protocols like the Model Context Protocol (MCP), has unlocked unprecedented

25 May 2026

CultivAgents: Cultivating Relationship-Centered Multi-Agent Systems for Personalized Gardening

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local AiDGX agent

arXiv:2605.23193v1 Announce Type: cross Abstract: Gardening is critical to support well-being, cultural continuity, and food autonomy, yet existing digital tools often provide generic advice that over

RMA: an Agentic System for Research-Level Mathematical Problems

Model ReleasesDGX agent

arXiv:2605.22875v1 Announce Type: new Abstract: We present extbf{Research Math Agents (RMA)}, an agentic framework for automated reasoning on research-level mathematical problems. Unlike prior studies

24 May 2026

The Top AI Papers of the Week (May 18 - 24): - AIRA - MetaCogAgent - Memory as a Model - Code as Agent Harness - Weak-Model Critic-Comparato…

AgentsDGX agent

The Top AI Papers of the Week (May 18 - 24): - AIRA - MetaCogAgent - Memory as a Model - Code as Agent Harness - Weak-Model Critic-Comparator - OpenAI Disproves the Unit Distance Conjecture - Producti

23 May 2026

Compiling Agentic Workflows into LLM Weights: Near-Frontier Quality at Two Orders of Magnitude Less Cost

Model ReleasesDGX agent

arXiv:2605.22502v1 Announce Type: cross Abstract: Agent orchestration frameworks have proliferated, collectively exceeding 290,000 GitHub stars across LangGraph, CrewAI, Google ADK, OpenAI Agents SDK,

22 May 2026

An Application-Layer Multi-Modal Covert-Channel Reference Monitor for LLM Agent Egress

AgentsDGX agent

arXiv:2605.20734v1 Announce Type: cross Abstract: A large language model (LLM) agent that sends messages can leak data inside them. Destination allowlists and content scanners do not police whether an

An Entity Linking Agent for Question Answering

AgentsDGX agent

arXiv:2508.03865v4 Announce Type: replace Abstract: Some Question Answering (QA) systems rely on knowledge bases (KBs) to provide accurate answers. Entity Linking (EL) plays a critical role in linking

Causal Past Logic for Runtime Verification of Distributed LLM Agent Workflows

AgentsDGX agent

arXiv:2605.20923v1 Announce Type: cross Abstract: Distributed LLM agent workflows should not be monitored as if they produced a single sequential log. In an asynchronous execution, a decision can only

COAgents: Multi-Agent Framework to Learn and Navigate Routing Problems Search Space

AgentsDGX agent

arXiv:2605.20618v1 Announce Type: new Abstract: Although Vehicle Routing Problems (VRP) are essential to many real-world systems, they remain computationally intractable at scale due to their combinat

Evaluating multimodal emotion recognition in proactive conversational agents: A user study

AgentsDGX agent

arXiv:2605.20200v1 Announce Type: cross Abstract: This article presents a multimodal emotion recognition module integrated into a proactive Socially Interactive Agent (SIA) powered by generative artif

Governance by Construction for Generalist Agents

SafetyDGX agent

arXiv:2605.20874v1 Announce Type: new Abstract: Enterprise agents are increasingly expected to operate autonomously across tools and interfaces, yet production deployments require governance by constr

Most vibe-coded apps forget every user who opens them. With one prompt to Replit Agent fixes it! 💻 Try adding authentication to your app to…

AgentsDGX agent

Replit Agent can automatically add authentication features to web applications with a single prompt, solving the common problem of apps losing user sessions. This tool helps developers quickly impleme

software security is a hot topic right now timely podcast with @cogent_security on building agents for cyber defense

AgentsDGX agent

software security is a hot topic right now timely podcast with @cogent_security on building agents for cyber defense Great conversation on the Max Agency podcast with @cogent_security Co-Founder + CTO

Thank you @NousResearch team! Here's to Security for All, agents included. 😎🤖

AgentsDGX agent

Bitwarden acknowledges Nous Research for collaboration on security initiatives that extend to AI agents, emphasizing inclusive security practices. The post suggests a partnership or joint effort betwe

This is one of the things I dislike about managed agents. Is it the best DX? yes. Is it now much, much more usable because it's bring your o…

AgentsDGX agent

This is one of the things I dislike about managed agents. Is it the best DX? yes. Is it now much, much more usable because it's bring your own sandbox? yes (most startups now have Sandbox credits and

21 May 2026

Agent streaming has outgrown token deltas. Real apps need to render tools, state, subagents, media, interrupts, and reconnects without parsi…

AgentsDGX agent

Agent streaming has outgrown token deltas. Real apps need to render tools, state, subagents, media, interrupts, and reconnects without parsing a firehose of raw events. The new @LangChain streaming pr

Anatomy of Agentic Memory: Taxonomy and Empirical Analysis of Evaluation and System Limitations

Model ReleasesDGX agent

arXiv:2602.19320v2 Announce Type: replace Abstract: Agentic memory systems enable large language model (LLM) agents to maintain state across long interactions, supporting long-horizon reasoning and pe

Dell’s agentic inflection point puts the rack at the center of enterprise transformation

AgentsDGX agent

Enterprises have reached an agentic inflection point where the shift from AI-first experimentation to AI-native operations is forcing a fundamental redesign of infrastructure and business processes. T

MemGym: a Long-Horizon Memory Environment for LLM Agents

Model ReleasesDGX agent

arXiv:2605.20833v1 Announce Type: new Abstract: Memory is a central capability for LLM agents operating across long-horizon tasks. Existing memory benchmarks predominantly evaluate retention of person

Pivot pulls in $40M to push agentic AI deeper into enterprise procurement

AgentsDGX agent

Artificial-intelligence-powered procurement platform startup Pivot Technologies SAS revealed today that it has raised 40 million in new funding to expand its agentic AI capabilities and push deeper in

RT @huntlovell: re: interpreters (since we''ve gotten asked this a few times) The proof will be in the evals/task-set for your agent, but…

AgentsDGX agent

This tweet by Hunt Lovell, shared by Harrison Chase, addresses questions about interpreters in the context of AI agent evaluation. It suggests that the effectiveness of interpreters should be validate

This is real agent security alpha 🔐

AgentsDGX agent

This post likely discusses early-stage security features or improvements for AI agents, possibly from LangChain or a related AI framework that Harrison Chase is involved with. The 'alpha' designation

Versa applies zero-trust controls to AI agent actions with new MCP architecture

AgentsDGX agent

Secure access service edge firm Versa Networks Inc. today introduced a zero-trust architecture for the Model Context Protocol that validates every action an artificial intelligence agent takes inside

We are quite short of compute, and that is going to result in compute becoming very expensive for complex agentic workflows even as single-t…

AgentsDGX agent

We are quite short of compute, and that is going to result in compute becoming very expensive for complex agentic workflows even as single-turn chatbots get cheaper. So the richest companies & most pr

'You actually don't have too many PRs. You have too many bad PRs.' @dexhorthy on how agents write fast, but the intellectual burden of judgi…

AgentsDGX agent

'You actually don't have too many PRs. You have too many bad PRs.' @dexhorthy on how agents write fast, but the intellectual burden of judging the code still lands on a human reviewer. A perfect PR is

20 May 2026

Both LlamaParse and LiteParse can be used with your favorite agent through minimal MCP/skill setup: ✅ Setup LlamaParse MCP for high-quality …

AgentsDGX agent

Both LlamaParse and LiteParse can be used with your favorite agent through minimal MCP/skill setup: ✅ Setup LlamaParse MCP for high-quality document processing and extraction: https://www.llamaindex.a

Buckle up: Google is set to remake search with agentic AI in 2026

AgentsDGX agent

At Google I/O 2026, Google announced major AI-backed upgrades to Search representing the biggest overhaul in over 25 years, featuring a redesigned intelligent search box and new agentic AI capabilitie

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

Model ReleasesDGX agent

arXiv:2605.20176v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has a

Distribution-Free Uncertainty Quantification for Continuous AI Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.19779v1 Announce Type: new Abstract: We adapt split conformal prediction and adaptive conformal inference (ACI) to continuous AI agent evaluation, providing distribution-free coverage guara

GPUs changed the equation of enterprise compute. AMD and Dell say agentic AI is flipping the math once again

AgentsDGX agent

As enterprises graduate from AI experimentation to production-scale agentic deployments, the infrastructure assumptions of the chatbot era are rapidly giving way to a more distributed and cost-conscio

If you design production agent systems, this matters. Most devs accidentally let their framework defaults make critical architecture decisio…

AgentsDGX agent

If you design production agent systems, this matters. Most devs accidentally let their framework defaults make critical architecture decisions without thinking it through. This paper shows you how to

I've been wanting to do this for a long time... The Agentic Review Podcast is live 🔥 Apple, Spotify, YouTube. The gap I keep seeing when I …

AgentsDGX agent

I've been wanting to do this for a long time... The Agentic Review Podcast is live 🔥 Apple, Spotify, YouTube. The gap I keep seeing when I talk to engineering leaders is the one between 'AI wrote the

NanoCo raises $12M to accelerate NanoClaw, a secure, enterprise-grade agentic AI assistant for every office worker

AgentsDGX agent

NanoCo, the startup behind NanoClaw, a fast-growing open-source artificial intelligent agent that’s secure by design, said today it has raised 12 million in seed funding. The oversubscribed round was

OEP: Poisoning Self-Evolving LLM Agents via Locally Correct but Non-Transferable Experiences

Local AiDGX agent

arXiv:2605.18930v1 Announce Type: cross Abstract: Memory-augmented large language model (LLM) agents use iterative reflection and self-evolution to solve complex tasks, but these mechanisms introduce

PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents

SafetyDGX agent

arXiv:2605.19932v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate over long and recurring external contexts, like document corpora and code repositories. Across in

Tenable adds multistep reasoning and MCP support to Hexa AI agent

AgentsDGX agent

Exposure management company Tenable Holdings Inc. today unveiled new capabilities for Hexa AI, the agentic engine inside its Tenable One Exposure Management Platform, pushing the tool into general ava

Toward User Comprehension Supports for LLM Agent Skill Specifications

AgentsDGX agent

arXiv:2605.19362v1 Announce Type: cross Abstract: Users often interpret and select agent skills through their exttt{SKILL.md} specifications. To protect users, existing audits mainly focus on maliciou

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

Model ReleasesDGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

We raised 250M in Series C funding at a 2.2B valuation, led by a16z. Exa is a search lab organizing the web's data for agents.

AgentsDGX agent

Exa, a search lab focused on organizing web data for AI agents, announced a Series C funding round of 250 million at a 2.2 billion valuation, led by venture capital firm Andreessen Horowitz (a16z). Th

What and When to Distill: Selective Hindsight Distillation for Multi-Turn Agents

AgentsDGX agent

arXiv:2605.19447v1 Announce Type: new Abstract: Reinforcement learning can train LLM agents from sparse task rewards, but long-horizon credit assignment remains challenging: a single success-or-failur

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

Model ReleasesDGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

19 May 2026

Agentic AI Governance and Lifecycle Management in Healthcare

Model ReleasesDGX agent

arXiv:2601.15630v2 Announce Type: replace Abstract: Healthcare organizations are beginning to embed agentic AI into routine workflows, including clinical documentation support and early-warning monito

Agentic success has a prerequisite — building the systems most enterprises left undone

AgentsDGX agent

Enterprises racing toward assured autonomy in agentic AI are running headlong into a decades-old problem: Most of their infrastructure was never designed to be connected — let alone governed at the sp

AMATA: Adaptive Multi-Agent Trajectory Alignment for Knowledge-Intensive Question Answering

SafetyDGX agent

arXiv:2605.17352v1 Announce Type: new Abstract: Despite substantial advances in large language models (LLMs), generating factually consistent responses for knowledge-intensive question answering remai

Building a self-improving agent on a context graph of human disagreement

AgentsDGX agent

You can build a measurably better agent from data you already have, without retraining a thing. The data is what your experienced humans do when they correct the AI. Capture... The post Building a sel

Dynamic Generation of Multi-LLM Agents Communication Topologies with Graph Diffusion Models

AgentsDGX agent

arXiv:2510.07799v2 Announce Type: replace-cross Abstract: The efficiency of multi-agent systems driven by large language models (LLMs) largely hinges on their communication topology. However, designin

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

Model ReleasesDGX agent

arXiv:2605.17554v1 Announce Type: new Abstract: Frontier deep research agents (DRAs) plan a research task, synthesize across documents, and return a structured deliverable on demand. They are being de

HINT-SD: Targeted Hindsight Self-Distillation for Long-Horizon Agents

AgentsDGX agent

arXiv:2605.17873v1 Announce Type: cross Abstract: Training long-horizon LLM agents with reinforcement learning is challenging because sparse outcome rewards reveal whether a task succeeds, but not whi

How do you AI engineer an agent to do AI engineering? Turns out this is how 💯

AgentsDGX agent

This post likely discusses techniques and methodologies for building AI agents capable of performing AI engineering tasks, such as code generation, model optimization, or system design. It probably co

How GoPerfect Built an Agentic Recruiting Workforce with Qdrant Cloud

AgentsDGX agent

GoPerfect mission is to use an AI recruiting workforce that replaces the manual, low-leverage parts of recruiting. Instead, an agent decomposes recruiter intent and runs the work end to end to find to

I didn’t really understand what people meant by stateful agents so I started exploring my current take is that we’re not there yet, which le…

AgentsDGX agent

Yohei Nakajima discusses his exploration of stateful agents and concludes that current implementations have not yet achieved what the concept truly entails, suggesting the field is still in early stag

Learning Transferable Topology Priors for Multi-Agent LLM Collaboration Across Domains

SafetyDGX agent

arXiv:2605.17359v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential for complex reasoning by coordinating specialized agents through struct

SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering

Model ReleasesDGX agent

arXiv:2605.17526v1 Announce Type: cross Abstract: As autonomous coding agents become capable of handling increasingly long-horizon tasks, they have gradually demonstrated the potential to complete end

SkillGenBench: Benchmarking Skill Generation Pipelines for LLM Agents

Model ReleasesDGX agent

arXiv:2605.18693v1 Announce Type: new Abstract: As LLM agents are increasingly built around reusable skills, a central challenge is no longer only whether agents can use provided skills, but whether t

Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception

AgentsDGX agent

arXiv:2601.09413v2 Announce Type: replace-cross Abstract: We introduce a voice-agentic framework that learns one critical omni-understanding skill: knowing when to trust itself versus when to consult

This release is absolutely stacked. I want to call out a sleeper feature: Interpreters Imagine your cloud agent gets a 10k-row CSV of suppor…

AgentsDGX agent

This release is absolutely stacked. I want to call out a sleeper feature: Interpreters Imagine your cloud agent gets a 10k-row CSV of support tickets. Without a code runtime, the model is stuck reason

Throughput-Optimal Scheduling Algorithms for LLM Inference and AI Agents

AgentsDGX agent

arXiv:2504.07347v3 Announce Type: replace-cross Abstract: As demand for Large Language Models (LLMs) and AI agents grows rapidly, optimizing systems for efficient LLM inference becomes critical. While

what if we mapped older distributed systems patterns (like actor models or reactive, state-driven blackboards) to LLM agents?

AgentsDGX agent

This post explores conceptual parallels between classical distributed systems design patterns—such as actor models and reactive blackboard architectures—and their potential application to LLM agent de

With expanded Antigravity platform, Google accelerates agent-native software development

Model ReleasesDGX agent

Google Cloud is enhancing its “agent-first” coding platform for developers with the launch of Antigravity 2.0, a new standalone desktop application that enables a full “agent-optimized” user experienc

← Previous
1…7172737475…299
Next →