AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
13 May 2026

AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

AgentsDGX agent

arXiv:2605.11026v1 Announce Type: cross Abstract: Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks

@ankush_gola11 absolutely cooking on stage!! Excited for our launch of SmithDB, a database purpose built for agent observability with as muc…

AgentsDGX agent

SmithDB is a new database system designed specifically for agent observability, announced by Harrison Chase and the LangChain team. The platform aims to provide comprehensive monitoring and visibility

Emergent Communication between Heterogeneous Visual Agents through Decentralized Learning

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.11695v1 Announce Type: new Abstract: Symbols are shared, but perception is private. We study emergent communication between heterogeneous visual agents through decentralized learning, askin

Hermes Unlocks Self-Improving AI Agents, Powered by NVIDIA RTX PCs and DGX Spark

Local AiDGX agent

Agentic AI is changing the way users get work done. Following the success of OpenClaw, the community is embracing new open source agentic frameworks. The latest is Hermes Agent, which crossed 140,000

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues

Model ReleasesDGX agent

arXiv:2605.12493v1 Announce Type: new Abstract: Long-term memory is crucial for agents in specialized web environments, where success depends on recalling interface affordances, state dynamics, workfl

MCPShield: Content-Aware Attack Detection for LLM Agent Tool-Call Traffic

AgentsDGX agent

arXiv:2605.11053v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become a widely adopted interface for LLM agents to invoke external tools, yet learned monitoring of MCP tool-cal

One of the biggest challenges with agents is that they produce way more than just tokens. Tool calls. State updates. Intermediate reasoning.…

AgentsDGX agent

One of the biggest challenges with agents is that they produce way more than just tokens. Tool calls. State updates. Intermediate reasoning. Subagent progress. File changes. Custom events. Structured

Sign up below for a Hermes Agent meetup in Montreal!

AgentsDGX agent

Sign up below for a Hermes Agent meetup in Montreal! Save the date for May 29 at 6:30pm. Be here in Montreal or be square. You are cordially invited to a cozy apartment MEETUP hosted by myself and @vn

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send m…

AgentsDGX agent

We’re training models wrong and it’s due to chatGPT. Even the modern coding agents used daily still use message-based exchanges: They send messages to users, to themselves (CoT) and to tools, and rece

12 May 2026

Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution

AgentsDGX agent

arXiv:2605.08887v1 Announce Type: new Abstract: Self-evolving agents present a promising path toward continual adaptation by distilling task interactions into reusable knowledge artifacts. In practice

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

Model ReleasesDGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

Bridging the Cognitive Gap: A Unified Memory Paradigm for 6G Agentic AI-RAN

AgentsDGX agent

arXiv:2605.10036v1 Announce Type: cross Abstract: As 6G evolves, the radio access network must transcend traditional automation to embrace agentic AI capable of perception, reasoning, and evolution. A

DSGBench: A Diverse Strategic Game Benchmark for Evaluating LLM-based Agents in Complex Decision-Making Environments

Model ReleasesDGX agent

arXiv:2503.06047v2 Announce Type: replace Abstract: Large language model (LLM)-based agents are increasingly applied to complex strategic environments that demand long-horizon reasoning, multi-agent i

Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling

SafetyDGX agent

arXiv:2603.03759v2 Announce Type: replace-cross Abstract: Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents und

OpenClaw-RL: Train Any Agent Simply by Talking

SafetyDGX agent

arXiv:2603.10165v2 Announce Type: replace-cross Abstract: Every agent interaction generates a next-state signal, namely the user reply, tool output, terminal or GUI state change that follows each acti

PECMAN: Perception-enabled Collaborative Multi-Agent Navigation in Unknown Environments

AgentsDGX agent

arXiv:2605.09344v1 Announce Type: new Abstract: Most path planners assume fully known, static environments, assumptions that fail when robots navigate in dynamic and partially observable environments.

Rethinking Ratio-Based Trust Regions for Policy Optimization in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.09212v1 Announce Type: new Abstract: Centralized training with decentralized execution (CTDE) is a standard framework for cooperative multi-agent policy-gradient reinforcement learning, all

SkillLens: Adaptive Multi-Granularity Skill Reuse for Cost-Efficient LLM Agents

AgentsDGX agent

arXiv:2605.08386v1 Announce Type: new Abstract: Skill libraries have become a practical way for LLM agents to reuse procedural experience across tasks. However, existing systems typically treat skills

SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System

AgentsDGX agent

arXiv:2605.09341v1 Announce Type: cross Abstract: Large language model (LLM) agent systems are increasingly expected to improve after deployment, but existing work often decouples two adaptation targe

Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding

AgentsDGX agent

arXiv:2507.11198v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) enable new possibilities for qualitative research at scale, including annotation and qualitative coding of educat

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent

AgentsDGX agent

arXiv:2605.09443v1 Announce Type: cross Abstract: The advancement of Multimodal Large Language Models (MLLMs) has expanded Role-Playing Agents (RPAs) into visually grounded environments. However, huma

Towards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.09366v1 Announce Type: new Abstract: Transforming neuroimaging data into clinically actionable biomarkers is a knowledge-intensive and labor-intensive process. Standardized workflows such a

11 May 2026

Active Learning for Communication Structure Optimization in LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.05703v2 Announce Type: replace-cross Abstract: Optimizing the communication structure of large language model based multi-agent systems (LLM-MAS) has been shown to improve downstream perfor

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run you…

Model ReleasesDGX agent

Agents + file sandboxes are all in the range in 2026 🤖🗃️ This is a nifty reference implementation by @itsclelia showing you how to run your agent over a collection of docs (PDFs, images, Office) with

Alternating Target-Path Planning for Scalable Multi-Agent Coordination

AgentsDGX agent

arXiv:2605.07744v1 Announce Type: new Abstract: The concurrent target assignment and pathfinding (TAPF) problem extends multi-agent pathfinding (MAPF) by asking planners to allocate distinct targets a

Conformal Agent Error Attribution

AgentsDGX agent

arXiv:2605.06788v1 Announce Type: new Abstract: When multi-agent systems (MAS) fail, identifying where the decisive error occurred is the first step for automated recovery to an earlier state. Error a

CyBiasBench: Benchmarking Bias in LLM Agents for Cyber-Attack Scenarios

Model ReleasesDGX agent

arXiv:2605.07830v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in offensive cybersecurity. In this paper, we reveal an interesting phenom

DiffeoMorph: Learning to Morph 3D Shapes Using Differentiable Agent-Based Simulations

SafetyDGX agent

arXiv:2512.17129v2 Announce Type: replace Abstract: Biological systems can form complex three-dimensional structures through the collective behavior of agents that share a common update rule and opera

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

SafetyDGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms

AgentsDGX agent

arXiv:2605.06716v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents have fundamentally reshaped artificial intelligence by integrating external tools and planning capabilities. Whi

MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents

AgentsDGX agent

arXiv:2605.07594v1 Announce Type: new Abstract: Existing memory systems for embodied agents typically inject retrieved memory as static context at episode start, a paradigm we term Ahead-of-time Monol

On Time, Within Budget: Constraint-Driven Online Resource Allocation for Agentic Workflows

AgentsDGX agent

arXiv:2605.06110v2 Announce Type: replace Abstract: Agentic systems increasingly solve complex user requests by executing orchestrated workflows, where subtasks are assigned to specialized models or t

Our internal AI agent was a brute-force nightmare. 🛑 It burned 40,000 tokens, took 2 minutes, and only hit 68% accuracy just to answer a si…

AgentsDGX agent

Our internal AI agent was a brute-force nightmare. 🛑 It burned 40,000 tokens, took 2 minutes, and only hit 68% accuracy just to answer a single question. Why? Traditional databases are built for human

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

Model ReleasesDGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

Switchcraft: AI Model Router for Agentic Tool Calling

AgentsDGX agent

arXiv:2605.07112v1 Announce Type: new Abstract: Agentic AI systems that invoke external tools are powerful but costly, leading developers to default to large models and overspend inference budgets. Mo

9 May 2026

Agentic coding is a form of machine learning. Generated code is best treated as a blackbox artifact whose behavior and generalization should…

AgentsDGX agent

Agentic coding is a form of machine learning. Generated code is best treated as a blackbox artifact whose behavior and generalization should be managed via empirical evaluation, like with any ML model

viv always says it better than me lots of talk recently of thinking of agents as systems to measure and iteratively improve but - thats not …

AgentsDGX agent

viv always says it better than me lots of talk recently of thinking of agents as systems to measure and iteratively improve but - thats not JUST a technical thing. its also a human & team thing my fav

8 May 2026

AWS adds agentic payment features to Amazon Bedrock AgentCore

AgentsDGX agent

Amazon Web Services Inc. today debuted a set of features that artificial intelligence agents can use to make purchases. The capabilities are available in preview through the cloud giant’s Amazon Bedro

MCP Marketplace Brings Real-Time Intelligence to Agentic Applications

AgentsDGX agent

The MCP Marketplace enables agentic applications to access real-time intelligence and data through a centralized platform of Model Context Protocol integrations. Databricks has launched this marketpla

7 May 2026

ARISE: A Repository-level Graph Representation and Toolset for Agentic Fault Localization and Program Repair

Local AiDGX agent

arXiv:2605.03117v1 Announce Type: cross Abstract: Repository-level fault localization (FL) and automated program repair (APR) require an agent to identify the relevant code units across files, follow

GLM-5V-Turbo Tech Report: Toward a Native Foundation Model for Multimodal Agents This report summarizes the main improvements behind GLM-5V-…

AgentsDGX agent

GLM-5V-Turbo Tech Report: Toward a Native Foundation Model for Multimodal Agents This report summarizes the main improvements behind GLM-5V-Turbo across model design, multimodal training, reinforcemen

Open Models Make Agentic Batch Processing Economically Viable A lot of world’s work looks like “Do X for EVERY Y” - read every trace - respo…

AgentsDGX agent

Open Models Make Agentic Batch Processing Economically Viable A lot of world’s work looks like “Do X for EVERY Y” - read every trace - respond to every email - deep dive into every document - enrich e

6 May 2026

Agentic AI Systems Should Be Designed as Marginal Token Allocators

AgentsDGX agent

arXiv:2605.01214v1 Announce Type: new Abstract: This position paper argues that agentic AI systems should be designed and evaluated as marginal token allocation economies rather than as text generator

AI agents push infrastructure beyond human-centric design

AgentsDGX agent

Autonomous agents are rapidly redefining how enterprise systems operate, exposing new security gaps as machine-driven activity begins to outpace the infrastructure designed for human users. Systems bu

Beyond Isolated Investor: Predicting Startup Success via Roleplay-Based Collective Agents

AgentsDGX agent

arXiv:2512.22608v3 Announce Type: replace Abstract: Due to the high value and high failure rates of startups, predicting their success is a critical challenge. Existing approaches typically model star

Collibra’s new AI Command Center promises to combat agentic hallucinations

AgentsDGX agent

Big data management startup Collibra BV is trying to position itself as the nerve center for artificial intelligence agents with the launch of its new AI Command Center offering. It’s designed to give

From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework

Model ReleasesDGX agent

arXiv:2605.02939v1 Announce Type: new Abstract: Multimodal controversy detection (MCD) identifies controversial content in videos and their associated user comments, to support risk management for soc

Google has quietly shut down Project Mariner, its Chrome-browsing AI agent for completing tasks on users' behalf, after highlighting it onstage at I/O 2025 (Max Zeff/@zeffmax)

AgentsDGX agent

Max Zeff / @zeffmax: Google has quietly shut down Project Mariner, its Chrome-browsing AI agent for completing tasks on users' behalf, after highlighting it onstage at I/O 2025 — NEW: Google quietly s

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic

AgentsDGX agent

arXiv:2601.21972v4 Announce Type: replace Abstract: Recent work has explored optimizing LLM collaboration through Multi-Agent Reinforcement Learning (MARL). However, most MARL fine-tuning approaches r

MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory

SafetyDGX agent

arXiv:2605.03228v1 Announce Type: cross Abstract: As large language model (LLM)-powered agents are increasingly deployed to perform complex, real-world tasks, they face a growing class of attacks that

Vibe coding and agentic engineering are getting closer than I'd like

Model ReleasesDGX agent

I recently talked with Joseph Ruscio about AI coding tools for Heavybit's High Leverage podcast: Ep. #9, The AI Coding Paradigm Shift with Simon Willison. Here are some of my highlights, including my

5 May 2026

Cisco buys Astrix Security to strengthen AI agent discovery and governance

AgentsDGX agent

Cisco Systems Inc. today revealed it will buy the Israeli cybersecurity firm Astrix Security Ltd. in an effort to bolster the defenses of artificial intelligence agents. The startup has developed a pl

Cloud Engineer’s AI Toolkit: Sign up Now for a Developer Workshop Near You!

Model ReleasesDGX agent

The world of AI is rapidly shifting from experimental Large Language Models to an era of Agentic AI. In the agentic era, autonomous software agents act on behalf of employees and consumers—driving a f

Deepgram STT is now natively available on @togethercompute. One platform, full voice-agent stack: Deepgram transcription, Together-hosted LL…

AgentsDGX agent

Deepgram STT is now natively available on @togethercompute. One platform, full voice-agent stack: Deepgram transcription, Together-hosted LLMs, Aura-2 TTS. Sub-3-second round trip on reference builds.

Kuo: OpenAI appears to be fast-tracking its AI agent phone with two NPUs and a custom MediaTek Dimensity 9600 SoC, targeting mass production as early as H1 2027 (@mingchikuo)

AgentsDGX agent

@mingchikuo: Kuo: OpenAI appears to be fast-tracking its AI agent phone with two NPUs and a custom MediaTek Dimensity 9600 SoC, targeting mass production as early as H1 2027 — 【Industry Check Update】O

Markets with Heterogeneous Agents: Dynamics and Survival of Bayesian vs. No-Regret Learners

AgentsDGX agent

arXiv:2502.08597v3 Announce Type: replace-cross Abstract: We analyze the performance of heterogeneous learning agents in asset markets with stochastic payoffs. Our main focus is on comparing Bayesian

// Skills as Verifiable Artifacts // Pay attention to this one, AI devs. If you ship agent skills, your runtime is treating signed-and-clear…

AgentsDGX agent

// Skills as Verifiable Artifacts // Pay attention to this one, AI devs. If you ship agent skills, your runtime is treating signed-and-cleared skills as trusted by default. This paper argues a skill i

4 May 2026

Better Models Won’t Save Your Agent

AgentsDGX agent

This article likely argues that improving the underlying language models is insufficient for building effective AI agents, and that other critical components—such as knowledge retrieval, context manag

Can Coding Agents Reproduce Findings in Computational Materials Science?

Model ReleasesDGX agent

arXiv:2605.00803v1 Announce Type: cross Abstract: Large language models are increasingly deployed as autonomous coding agents and have achieved remarkably strong performance on software engineering be

← Previous
1…6061626364…297
Next →