AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

SERA: Soft-Verified Efficient Repository Agents

DGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

model-releasesarxiv-cs-cl
1 Jun 2026
Agents

Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.30862v1 Announce Type: cross Abstract: Text2SQL agents powered by LLMs translate natural language intent into SQL by exploring the data system through tool calls before formulating the quer

agentsarxiv-cs-ai
1 Jun 2026
Safety

Task-Focused Memorization for Multimodal Agents

DGX agent

arXiv:2605.31075v1 Announce Type: new Abstract: Long-term memory is essential for multimodal agents to build coherent experience, accumulate world knowledge, and achieve continual learning. However, c

safetyarxiv-cs-cv
1 Jun 2026
Safety

Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information

DGX agent

arXiv:2605.31445v1 Announce Type: cross Abstract: In this work we study agents in simulated bargaining scenarios, where a buyer and a seller communicate through a text channel and attempt to negotiate

safetyarxiv-cs-ai
1 Jun 2026
Agents

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

DGX agent

arXiv:2605.29354v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly participate in software development workflows by generating code, selecting dependencies, and producing package

agentsarxiv-cs-lg
29 May 2026
Agents

PatchBoard: Schema-Grounded State Mutation for Reliable and Auditable LLM Multi-Agent Collaboration

DGX agent

arXiv:2605.29313v1 Announce Type: new Abstract: LLM multi-agent systems often coordinate through natural-language dialogue or loosely structured shared memory, making intermediate state difficult to v

agentsarxiv-cs-cl
29 May 2026
Agents

Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework

DGX agent

arXiv:2605.29397v1 Announce Type: new Abstract: HTML observations in LLM-based web agents are extremely long, and while many reduction methods have been proposed, it remains unclear which methods redu

agentsarxiv-cs-cl
29 May 2026
Agents

RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models

DGX agent

arXiv:2603.18859v2 Announce Type: replace Abstract: Reinforcement learning (RL) shows promise for enhancing LLM agentic reasoning, yet sparse terminal rewards hinder fine-grained optimization. Process

agentsarxiv-cs-ai
29 May 2026
Agents

SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents

DGX agent

arXiv:2605.29440v1 Announce Type: cross Abstract: Retrieval-augmented LLM agents increasingly rely on curated skill banks: collections of reusable textual principles that guide decision making on comp

agentsarxiv-cs-ai
29 May 2026
Model Releases

A Query Engine for the Agents

DGX agent

arXiv:2605.27785v1 Announce Type: new Abstract: The fastest-growing data in production today is unstructured text: agent traces, chat logs, reasoning chains, model outputs. People want to analyze it,

model-releasesarxiv-cs-ai
28 May 2026
Agents

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

DGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

agentsarxiv-cs-ai
28 May 2026
Agents

GUI Agents for Continual Game Generation

DGX agent

arXiv:2605.28258v1 Announce Type: cross Abstract: Generating a game is not the same as making one that can be played. Despite advances in code generation, existing approaches treat game generation as

agentsarxiv-cs-ai
28 May 2026
Model Releases

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

DGX agent

arXiv:2605.28158v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used to assist with operations research (OR) modeling, yet existing OR-oriented benchmarks often redu

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

DGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

DGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

model-releasesarxiv-cs-ai
28 May 2026
Agents

AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito

DGX agent

arXiv:2601.18381v2 Announce Type: replace Abstract: To facilitate the transformation of legacy finite difference implementations into the Devito environment, this study develops an integrated AI agent

agentsarxiv-cs-ai
27 May 2026
Agents

Harmonia: Enhancing Data Placement and Migration in Hybrid Storage Systems via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2503.20507v4 Announce Type: replace-cross Abstract: Modern high-performance computing (HPC) environments rely on hybrid storage systems (HSS) that combine multiple storage devices with diverse l

agentsarxiv-cs-lg
27 May 2026
Agents

LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning

DGX agent

arXiv:2412.20505v2 Announce Type: replace Abstract: Participatory Urban Planning (PUP) is increasingly supported by LLM-based agents, yet existing methods largely rely on static preference elicitation

agentsarxiv-cs-ai
27 May 2026
Agents

Stateful Inference for Low-Latency Multi-Agent Tool Calling

DGX agent

arXiv:2605.26289v1 Announce Type: new Abstract: Multi-agent tool calling is becoming the dominant interaction pattern for LLM-based systems, yet existing inference frameworks treat each tool call as a

agentsarxiv-cs-lg
27 May 2026
Safety

Agent Learning via Early Experience

DGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

safetyarxiv-cs-ai
26 May 2026
Agents

APT-Agent: Automated Penetration Testing using Large Language Models

DGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

agentsarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Model Releases

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

DGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

model-releasesarxiv-cs-ai
26 May 2026
Agents

Harnessing AtomisticSkills for Agentic Atomistic Research

DGX agent

arXiv:2605.24002v1 Announce Type: cross Abstract: Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) hav

agentsarxiv-cs-ai
26 May 2026
Model Releases

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

DGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

model-releasesarxiv-cs-ai
26 May 2026
Agents

More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries

DGX agent

arXiv:2605.24050v1 Announce Type: cross Abstract: Skill libraries allow LLM agents to load task-specific instructions on demand, letting non-expert users solve domain-specific tasks through natural la

agentsarxiv-cs-ai
26 May 2026
Safety

MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection

DGX agent

arXiv:2505.23803v2 Announce Type: replace-cross Abstract: Phishing email detection faces significant challenges due to evolving adversarial tactics and heterogeneous attack patterns. Traditional appro

safetyarxiv-cs-ai
26 May 2026
Agents

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

DGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

agentsarxiv-cs-cl
26 May 2026
Safety

SEAL: Synergistic Co-Evolution of Agents and Learning Environments

DGX agent

arXiv:2605.24426v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly improved through interaction, yet most self-evolution methods adapt either the policy or the learning

safetyarxiv-cs-cl
26 May 2026
Agents

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

DGX agent

arXiv:2605.25310v1 Announce Type: new Abstract: Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether t

agentsarxiv-cs-cl
26 May 2026
Model Releases

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

DGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

model-releasesarxiv-cs-ai
25 May 2026
Agents

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

DGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an

agentsarxiv-cs-ai
25 May 2026
Agents

SFG-ROS: A Resource-Aware Framework for Dense Multi-Agent Perception

DGX agent

arXiv:2605.23832v1 Announce Type: new Abstract: Deploying heterogeneous multi-agent robot fleets for collaborative perception requires robust data exchange and scalable software architectures. However

agentsarxiv-cs-ro
25 May 2026
Model Releases

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

DGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

model-releasesarxiv-cs-ai
25 May 2026
Safety

Heterogeneous Agent Collaborative Reinforcement Learning

DGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

safetyarxiv-cs-lg
23 May 2026
Model Releases

CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking

DGX agent

arXiv:2602.08023v3 Announce Type: replace-cross Abstract: Existing benchmarks for LLM-based offensive security agents use isolated, single-target setups with a known vulnerable service and fixed objec

model-releasesarxiv-cs-ai
22 May 2026
Safety

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

DGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

safetyarxiv-cs-cl
22 May 2026
Agents

General Agentic Planning Through Simulative Reasoning with World Models

DGX agent

arXiv:2507.23773v3 Announce Type: replace-cross Abstract: What does it mean to plan? Current agentic systems, whether scaffolded workflows or end-to-end policies, rely on reactive decision-making: sel

agentsarxiv-cs-cl
22 May 2026
Safety

Learning to Configure Agentic AI Systems

DGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

safetyarxiv-cs-ai
22 May 2026
Agents

Reflecti-Mate: A Conversational Agent for Adaptive Decision-Making Support Through System 1 and System 2 Thinking

DGX agent

arXiv:2605.22509v1 Announce Type: cross Abstract: Making high-stakes personal decisions involves cognitive, emotional, and intuitive processes, and individuals differ in how they allocate attention ac

agentsarxiv-cs-cl
22 May 2026
Safety

Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.22748v1 Announce Type: new Abstract: Autonomous systems have achieved superhuman performance in isolation or simulation, yet they remain brittle in shared, dynamic real-world spaces. This f

safetyarxiv-cs-ro
22 May 2026
Agents

Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents

DGX agent

arXiv:2605.21347v1 Announce Type: cross Abstract: Diagnosing failures in LLM agents remains largely manual. Practitioners inspect a small subset of execution traces, form ad-hoc hypotheses, and iterat

agentsarxiv-cs-lg
21 May 2026
Agents

AQuaUI: Visual Token Reduction for GUI Agents with Adaptive Quadtrees

DGX agent

arXiv:2605.19260v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have recently emerged as promising backbones for GUI-agent models, where high-resolution GUI screenshots are introduced t

agentsarxiv-cs-ai
20 May 2026
Safety

Metric-Gradient Projection for Stable Multi-Agent Policy Learning

DGX agent

arXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent's policy update changes the optimization landscape fa

safetyarxiv-cs-ai
20 May 2026
Model Releases

Search Self-play: Pushing the Frontier of Agent Capability without Supervision

DGX agent

arXiv:2510.18821v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become the mainstream technique for training LLM agents. However, RLVR highly depends on w

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

DGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs

DGX agent

arXiv:2605.19798v1 Announce Type: new Abstract: As Socially Interactive Agents (SIAs) become increasingly integrated into daily life, the ability to calibrate user trust to an agent's actual capabilit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Vision Harnessing Agent for Open Ad-hoc Segmentation

DGX agent

arXiv:2605.19410v1 Announce Type: new Abstract: Segmentation has become easy when the concept is known, requiring retrieval of a learned visual grounding from text. It remains hard for open ad-hoc con

model-releasesarxiv-cs-cv
20 May 2026
← Previous
1…4041424344…233
Next →