AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
Agents

GUI Agents for Continual Game Generation

DGX agent

arXiv:2605.28258v1 Announce Type: cross Abstract: Generating a game is not the same as making one that can be played. Despite advances in code generation, existing approaches treat game generation as

agentsarxiv-cs-ai
28 May 2026
Model Releases

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2605.28158v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used to assist with operations research (OR) modeling, yet existing OR-oriented benchmarks often redu

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

DGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
Agents

AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito

DGX agent

arXiv:2601.18381v2 Announce Type: replace Abstract: To facilitate the transformation of legacy finite difference implementations into the Devito environment, this study develops an integrated AI agent

agentsarxiv-cs-ai
27 May 2026
Agents

Harmonia: Enhancing Data Placement and Migration in Hybrid Storage Systems via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2503.20507v4 Announce Type: replace-cross Abstract: Modern high-performance computing (HPC) environments rely on hybrid storage systems (HSS) that combine multiple storage devices with diverse l

agentsarxiv-cs-lg
27 May 2026
Agents

Hermes Agent now has a built-in MCP Catalog

DGX agent

Nous Research announced that their Hermes Agent now includes an integrated MCP (Model Context Protocol) Catalog, enabling users to discover and access available Model Context Protocol tools and integr

agentsnous-research--x
27 May 2026
Agents

LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning

DGX agent

arXiv:2412.20505v2 Announce Type: replace Abstract: Participatory Urban Planning (PUP) is increasingly supported by LLM-based agents, yet existing methods largely rely on static preference elicitation

agentsarxiv-cs-ai
27 May 2026
Agents

Stateful Inference for Low-Latency Multi-Agent Tool Calling

DGX agent

arXiv:2605.26289v1 Announce Type: new Abstract: Multi-agent tool calling is becoming the dominant interaction pattern for LLM-based systems, yet existing inference frameworks treat each tool call as a

agentsarxiv-cs-lg
27 May 2026
Agents

.@AdamRLucek on how we use traces to build evals for production agents.

DGX agent

.@AdamRLucek on how we use traces to build evals for production agents. Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating e

agentsharrison-chase--x
26 May 2026
Safety

Agent Learning via Early Experience

DGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

safetyarxiv-cs-ai
26 May 2026
Agents

APT-Agent: Automated Penetration Testing using Large Language Models

DGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

agentsarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Agents

Exclusive: Canyon Code closes on $5M pre-seed round to bring visibility to agentic application workflows

DGX agent

Agentic artificial intelligence startup Canyon Code stepped out of the shadows today after closing on a 5 million pre-seed funding round. Its mission is to give enterprises the granular controls requi

agentssiliconangle
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Agents

Harnessing AtomisticSkills for Agentic Atomistic Research

DGX agent

arXiv:2605.24002v1 Announce Type: cross Abstract: Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) hav

agentsarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Agents

More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries

DGX agent

arXiv:2605.24050v1 Announce Type: cross Abstract: Skill libraries allow LLM agents to load task-specific instructions on demand, letting non-expert users solve domain-specific tasks through natural la

agentsarxiv-cs-ai
26 May 2026
Safety

MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection

DGX agent

arXiv:2505.23803v2 Announce Type: replace-cross Abstract: Phishing email detection faces significant challenges due to evolving adversarial tactics and heterogeneous attack patterns. Traditional appro

safetyarxiv-cs-ai
26 May 2026
Agents

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

DGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

agentsarxiv-cs-cl
26 May 2026
Safety

SEAL: Synergistic Co-Evolution of Agents and Learning Environments

DGX agent

arXiv:2605.24426v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly improved through interaction, yet most self-evolution methods adapt either the policy or the learning

safetyarxiv-cs-cl
26 May 2026
Agents

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

DGX agent

arXiv:2605.25310v1 Announce Type: new Abstract: Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether t

agentsarxiv-cs-cl
26 May 2026
Model Releases

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

DGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

model-releasesarxiv-cs-ai
25 May 2026
Agents

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

DGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an

agentsarxiv-cs-ai
25 May 2026
Agents

Join the team on Wednesday for another Hermes Agent Jam!

DGX agent

Nous Research is inviting developers and enthusiasts to participate in a 'Hermes Agent Jam' event scheduled for Wednesday, likely a hackathon or collaborative workshop focused on building or improving

agentsnous-research--x
25 May 2026
Agents

SFG-ROS: A Resource-Aware Framework for Dense Multi-Agent Perception

DGX agent

arXiv:2605.23832v1 Announce Type: new Abstract: Deploying heterogeneous multi-agent robot fleets for collaborative perception requires robust data exchange and scalable software architectures. However

agentsarxiv-cs-ro
25 May 2026
Model Releases

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

DGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

model-releasesarxiv-cs-ai
25 May 2026
Agents

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-1…

DGX agent

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-19b0-4534-9c90-465b8e29ad73/unit-distance-proof.pdf` and come

agentselon-musk--x
24 May 2026
Safety

Heterogeneous Agent Collaborative Reinforcement Learning

DGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

safetyarxiv-cs-lg
23 May 2026
Agents

AWS User Group Malaysia Talk: Experience on Building AI Agent Harness https://www.youtube.com/watch?v=u145ea-fNOw

DGX agent

This talk from an AWS User Group Malaysia event covers practical experience and insights on building an AI Agent Harness, likely discussing architecture patterns, implementation challenges, and best p

agentsharrison-chase--x
22 May 2026
Model Releases

CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking

DGX agent

arXiv:2602.08023v3 Announce Type: replace-cross Abstract: Existing benchmarks for LLM-based offensive security agents use isolated, single-target setups with a known vulnerable service and fixed objec

model-releasesarxiv-cs-ai
22 May 2026
Safety

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

DGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

safetyarxiv-cs-cl
22 May 2026
Agents

General Agentic Planning Through Simulative Reasoning with World Models

DGX agent

arXiv:2507.23773v3 Announce Type: replace-cross Abstract: What does it mean to plan? Current agentic systems, whether scaffolded workflows or end-to-end policies, rely on reactive decision-making: sel

agentsarxiv-cs-cl
22 May 2026
Agents

Highlights: 👉 Long-horizon autonomy: maintained coherent execution across a 35-hour autonomous kernel optimization run 👉 Agentic coding: l…

DGX agent

Highlights: 👉 Long-horizon autonomy: maintained coherent execution across a 35-hour autonomous kernel optimization run 👉 Agentic coding: leading Terminal-Bench 2.0-Terminus performance for terminal-ba

agentstogether-ai--x
22 May 2026
Safety

Learning to Configure Agentic AI Systems

DGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

safetyarxiv-cs-ai
22 May 2026
Agents

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while pres…

DGX agent

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while preserving near-frontier task quality. The workflow includes mul

agentsdair-ai--x
22 May 2026
Agents

Reflecti-Mate: A Conversational Agent for Adaptive Decision-Making Support Through System 1 and System 2 Thinking

DGX agent

arXiv:2605.22509v1 Announce Type: cross Abstract: Making high-stakes personal decisions involves cognitive, emotional, and intuitive processes, and individuals differ in how they allocate attention ac

agentsarxiv-cs-cl
22 May 2026
Safety

Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.22748v1 Announce Type: new Abstract: Autonomous systems have achieved superhuman performance in isolation or simulation, yet they remain brittle in shared, dynamic real-world spaces. This f

safetyarxiv-cs-ro
22 May 2026
Agents

these guys built a research agent with activegraph (using their @monid_ai tool) and found that every claim was traced to a source - which is…

DGX agent

these guys built a research agent with activegraph (using their @monid_ai tool) and found that every claim was traced to a source - which is not prompted, but natively baked in to the approach (they a

agentsyohei-nakajima--x
22 May 2026
Agents

Giving Agents Computers — Ivan Burazin, Daytona

DGX agent

Latent Space episode featuring Ivan Burazin discussing Daytona, a platform or framework that enables AI agents to use computers and interact with software systems autonomously. The discussion likely c

agentslatent-space
21 May 2026
Agents

Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents

DGX agent

arXiv:2605.21347v1 Announce Type: cross Abstract: Diagnosing failures in LLM agents remains largely manual. Practitioners inspect a small subset of execution traces, form ad-hoc hypotheses, and iterat

agentsarxiv-cs-lg
21 May 2026
Agents

Introducing the sandbox Auth Proxy: A way to control the boundary between agent-generated behavior and the rest of the world. An explainer f…

DGX agent

The sandbox Auth Proxy is a security mechanism designed to control and limit the interactions between AI agents and external systems by establishing a controlled boundary. This explainer from Harrison

agentsharrison-chase--x
21 May 2026
Agents

Reimagining ML Operations with Agent Skills: a new maturity model for on-call

DGX agent

This article presents a maturity model for ML operations that leverages agent skills to improve on-call practices and incident response workflows. It likely discusses how organizations can evolve thei

agentsanyscale-ray
21 May 2026
Agents

streaming from modern agents is pretty complex! especially with - parallel tools / subagents - multimodal content - human in the loop events…

DGX agent

streaming from modern agents is pretty complex! especially with - parallel tools / subagents - multimodal content - human in the loop events our new streaming primitives make all of this ergonomic! Ag

agentsharrison-chase--x
21 May 2026
Agents

AQuaUI: Visual Token Reduction for GUI Agents with Adaptive Quadtrees

DGX agent

arXiv:2605.19260v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently emerged as promising backbones for GUI-agent models, where high-resolution GUI screenshots are introduced t

agentsarxiv-cs-ai
20 May 2026
Agents

datasette-agent-charts 0.1a1

DGX agent

Release: datasette-agent-charts 0.1a1 More color! Bar and waffle charts without a color column are shaded by magnitude with a sequential color scheme; color columns holding text values use the observa

agentssimon-willison
20 May 2026
Agents

From performance reviews to pink slips, managing AI agents looks a lot like managing people

DGX agent

The workforce is no longer purely human — and closing the gap between how companies manage people and how they govern AI agents has become one of the defining operational challenges of the digital wor

agentssiliconangle
20 May 2026
Agents

i'm excited to open source Active Graph: an event-sourced reactive graph runtime for long-running, agents 🔄🧠 events/logs projects a graph.…

DGX agent

i'm excited to open source Active Graph: an event-sourced reactive graph runtime for long-running, agents 🔄🧠 events/logs projects a graph. reactive behaviors react and affect the graph. fork-and-diff

agentsyohei-nakajima--x
20 May 2026
← Previous
1…7475767778…374
Next →