AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

POINTS-Seeker: Towards Training a Multimodal Agentic Search Model from Scratch

DGX agent

arXiv:2604.14029v1 Announce Type: new Abstract: While Large Multimodal Models (LMMs) demonstrate impressive visual perception, they remain epistemically constrained by their static parametric knowledg

agentsarxiv-cs-cv
16 Apr 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Agentic Discovery with Active Hypothesis Exploration for Visual Recognition

DGX agent

arXiv:2604.12999v1 Announce Type: new Abstract: We introduce HypoExplore, an agentic framework that formulates neural architecture discovery for visual recognition as a hypothesis-driven scientific in

agentsarxiv-cs-cv
15 Apr 2026
Agents

M2HRI: An LLM-Driven Multimodal Multi-Agent Framework for Personalized Human-Robot Interaction

DGX agent

arXiv:2604.11975v1 Announce Type: new Abstract: Multi-robot systems hold significant promise for social environments such as homes and hospitals, yet existing multi-robot works treat robots as functio

agentsarxiv-cs-ro
15 Apr 2026
Agents

Beyond RAG for Cyber Threat Intelligence: A Systematic Evaluation of Graph-Based and Agentic Retrieval

DGX agent

arXiv:2604.11419v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) analysts must answer complex questions over large collections of narrative security reports. Retrieval-augmented generat

agentsarxiv-cs-ai
14 Apr 2026
Agents

ChatInject: Abusing Chat Templates for Prompt Injection in LLM Agents

DGX agent

arXiv:2509.22830v3 Announce Type: replace Abstract: The growing deployment of large language model (LLM) based agents that interact with external environments has created new attack surfaces for adver

agentsarxiv-cs-cl
14 Apr 2026
Agents

Verify Before You Fix: Agentic Execution Grounding for Trustworthy Cross-Language Code Analysis

DGX agent

arXiv:2604.10800v1 Announce Type: cross Abstract: Learned classifiers deployed in agentic pipelines face a fundamental reliability problem: predictions are probabilistic inferences, not verified concl

agentsarxiv-cs-ai
14 Apr 2026
Local Ai

Bayesian Ego-graph Inference for Networked Multi-Agent Reinforcement Learning

DGX agent

arXiv:2509.16606v5 Announce Type: replace-cross Abstract: In networked multi-agent reinforcement learning (Networked-MARL), decentralized agents must act under local observability and constrained comm

local-aiarxiv-cs-lg
13 Apr 2026
Model Releases

Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models

DGX agent

arXiv:2604.08970v1 Announce Type: cross Abstract: We study predictive multilingual evaluation: estimating how well a model will perform on a task in a target language when direct benchmark results are

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

Multi-agent Adaptive Mechanism Design

DGX agent

arXiv:2512.21794v3 Announce Type: replace-cross Abstract: We study a sequential mechanism design problem in which a principal seeks to elicit truthful reports from multiple rational agents while start

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent

DGX agent

arXiv:2604.06296v1 Announce Type: cross Abstract: AI agents are increasingly deployed in real-world applications, including systems such as Manus, OpenClaw, and coding agents. Existing research has pr

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Learning Without Losing Identity: Capability Evolution for Embodied Agents

DGX agent

arXiv:2604.07799v1 Announce Type: new Abstract: Embodied agents are expected to operate persistently in dynamic physical environments, continuously acquiring new capabilities over time. Existing appro

safetyarxiv-cs-ro
10 Apr 2026
Safety

Qualixar OS: A Universal Operating System for AI Agent Orchestration

DGX agent

arXiv:2604.06392v1 Announce Type: new Abstract: We present Qualixar OS, the first application-layer operating system for universal AI agent orchestration. Unlike kernel-level approaches (AIOS) or sing

safetyarxiv-cs-ai
10 Apr 2026
Agents

Reasoning Graphs: Deterministic Agent Accuracy through Evidence-Centric Chain-of-Thought Feedback

DGX agent

arXiv:2604.07595v1 Announce Type: cross Abstract: Language model agents reason from scratch on every query: each time an agent retrieves evidence and deliberates, the chain of thought is discarded and

agentsarxiv-cs-cl
10 Apr 2026
Agents

Bandwidth-Efficient Multi-Agent Communication through Information Bottleneck and Vector Quantization

DGX agent

arXiv:2602.02035v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning systems deployed in real-world robotics applications face severe communication constraints that significant

agentsarxiv-cs-ai
12 Aug 2026
Agents

On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models

DGX agent

arXiv:2608.10530v1 Announce Type: cross Abstract: Large Language Models (LLMs) have undergone a shift from stateless conversational interfaces to autonomous agents capable of multi-step planning, tool

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

DGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Coupling Planning with Episodic Memory in LLM Agents for Software Issue Resolution

DGX agent

arXiv:2608.06811v1 Announce Type: cross Abstract: Resolving a real software issue with a large language model (LLM) agent is a long repair episode, often tens to hundreds of steps spanning exploration

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

HyperAgent: Planning and Acting over Tool-Schema Hypergraphs for Tool-Use LLM Agents

DGX agent

arXiv:2608.02650v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly rely on external tools to complete complex real-world tasks. However, reliable tool-use planning remains

agentsarxiv-cs-ai
5 Aug 2026
Safety

The Agent Operating System (AOS): A Reference Operating Architecture for Distributed Agentic Systems

DGX agent

arXiv:2608.03214v1 Announce Type: new Abstract: Large language models have transformed artificial intelligence from isolated prediction services into components of long-running, distributed systems th

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

DGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Bayesian and Motivated Reasoning in AI Agents

DGX agent

arXiv:2608.00339v1 Announce Type: cross Abstract: AI agents increasingly perform open-ended tasks in settings where their conclusions can guide consequential decisions. We provide evidence that AI age

agentsarxiv-cs-cl
4 Aug 2026
Agents

Progressive Agent Skill Generation via Reinforcement Learning

DGX agent

arXiv:2608.01678v1 Announce Type: cross Abstract: Existing skill generation methods largely rely on heuristics or pipeline-style consolidation, which must be specially designed for different evidence

agentsarxiv-cs-cl
4 Aug 2026
Agents

RoMeRL: Balancing Feedback Coverage and the Memory-Reward Trap in Self-Evolving Agent Memory via Reduced-Order Utility States

DGX agent

arXiv:2608.02508v1 Announce Type: cross Abstract: Learning-based memory systems for self-evolving LLM agents face two tightly coupled challenges. First, trajectory-indexed utilities grow with the inte

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step

DGX agent

arXiv:2608.02358v1 Announce Type: new Abstract: To operate robustly in open-world environments, autonomous agents should be able to infer the behavior of unfamiliar systems through interaction alone,

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

SWE-Touch: Benchmarking Coding Agents When Users Touch the Code

DGX agent

arXiv:2608.02499v1 Announce Type: cross Abstract: Real-world software development requires coding agents to operate in shared workspaces where users may inspect and modify code during an ongoing task,

agentsarxiv-cs-cl
4 Aug 2026
Agents

V-Mem: Modality-Routed Retrieval for Long-Term Multimodal Agentic Memory

DGX agent

arXiv:2608.01543v1 Announce Type: cross Abstract: Interaction between users and LLM agents is increasingly multimodal: conversations interleave text with images, and a later question may target either

agentsarxiv-cs-cl
4 Aug 2026
Agents

AgenticRepair: Multi-Faceted Program Context Engineering for Agentic Vulnerability Repair

DGX agent

arXiv:2607.29422v1 Announce Type: cross Abstract: Automated vulnerability repair aims to reduce the time and effort required to patch security flaws from a vulnerability triage report. Recent agentic

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

EMBL AI Librarian: Life-Sciences Knowledge Layer for AI Agents

DGX agent

arXiv:2607.28229v1 Announce Type: new Abstract: The web is increasingly accessed by AI agents rather than humans. Every agent needs knowledge, especially in the life-sciences, where agentic pipelines

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Living-Harness Is an Interactive-Agent Evolver

DGX agent

arXiv:2607.26598v1 Announce Type: cross Abstract: Large language model (LLM) agents may recover from a failure within an episode or after a retry, yet the same execution failure can recur in later tas

agentsarxiv-cs-cl
30 Jul 2026
Agents

SkillCAT: Contrastive, Assessment-Augmented and Topology-AwareSkill Self-Evolution for LLM Agents

DGX agent

arXiv:2606.13317v2 Announce Type: replace Abstract: Skill self-evolution methods for LLM agents aim to turn execution trajectories into reusable skill documents. However, current pipelines typically d

agentsarxiv-cs-cl
30 Jul 2026
Safety

PLATO: Pointer Learner for Agent and Task Openness

DGX agent

arXiv:2607.25082v1 Announce Type: new Abstract: Open agent systems (OASYS) are increasingly prevalent in real-world domains where the sets of agents and tasks change unpredictably over time. Such open

safetyarxiv-cs-ai
29 Jul 2026
Agents

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents

DGX agent

arXiv:2607.23588v1 Announce Type: new Abstract: Creative AI is moving from single-step asset generation toward long-horizon multimodal production. Although recent generative models can synthesize high

agentsarxiv-cs-cv
28 Jul 2026
Model Releases

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows

DGX agent

arXiv:2510.24411v3 Announce Type: replace Abstract: Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobi

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Where Is the Cost of Third-Party API Routers in Agentic Software Development?

DGX agent

arXiv:2607.23624v1 Announce Type: cross Abstract: Third-party API routers have become a common layer that unifies access across increasingly diverse LLM providers. In coding-agent workflows, high-auto

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

InteractComp: Evaluating Search Agents With Ambiguous Queries

DGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

model-releasesarxiv-cs-cl
27 Jul 2026
Agents

AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics

DGX agent

arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments. Thi

agentsarxiv-cs-ai
24 Jul 2026
Agents

Critic Experience Bank: Self-Evolving Step-Level Confidence Estimation for LLM Agents

DGX agent

arXiv:2607.12397v1 Announce Type: new Abstract: LLM agents act in external environments where each action changes the state that later decisions condition on, and where a single wrong step can waste i

agentsarxiv-cs-ai
15 Jul 2026
Agents

Tracing Agentic Failure from the Flow of Success

DGX agent

arXiv:2607.12747v1 Announce Type: new Abstract: Failure attribution for LLM-based agentic systems, i.e., identifying which steps in a failure trajectory caused the task to fail, is critical for debugg

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

Think Big, Search Small: Where Capacity Matters in Hierarchical Search Agents?

DGX agent

arXiv:2607.07548v1 Announce Type: new Abstract: Large language model based search agents increasingly adopt multi-agent architectures in which a main agent decomposes a complex question into sub-queri

model-releasesarxiv-cs-cl
9 Jul 2026
Agents

Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale

DGX agent

arXiv:2607.06233v1 Announce Type: new Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize

agentsarxiv-cs-ai
8 Jul 2026
Agents

Conflict-Based Search for Multi-Agent Path Finding with Elevators

DGX agent

arXiv:2602.20512v2 Announce Type: replace Abstract: This paper investigates a problem called Multi-Agent Path Finding with Elevators (MAPF-E), which seeks conflict-free paths for multiple agents from

agentsarxiv-cs-ro
7 Jul 2026
Agents

SkillFab: An Agent-Native Skill Production Platform

DGX agent

arXiv:2607.03780v1 Announce Type: cross Abstract: SkillFab is an agent-native platform for turning missing capabilities into reviewed, reusable Agent Skills. At runtime, agents first search for reusab

agentsarxiv-cs-ai
7 Jul 2026
Agents

Investigating Multi-Agent Deliberation in Law

DGX agent

arXiv:2606.30906v1 Announce Type: new Abstract: Artificial Intelligence is increasingly applied to the field of law, and has the potential to increase access to justice. One particular movement that i

agentsarxiv-cs-ai
1 Jul 2026
Safety

Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols

DGX agent

arXiv:2606.26203v1 Announce Type: new Abstract: As AI agent protocols proliferate, the governance structures shaping their interoperability standards remain empirically underexamined. We introduce an

safetyarxiv-cs-ai
26 Jun 2026
Agents

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents

DGX agent

arXiv:2606.24893v1 Announce Type: new Abstract: For agents to learn continuously from interaction with the world at test time, they must be able to explore effectively, acquire new world knowledge and

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

CORE-Bench: Fostering the Credibility of Published Research Through a Computational Reproducibility Agent Benchmark

DGX agent

arXiv:2409.11363v2 Announce Type: replace-cross Abstract: AI agents have the potential to aid users on a variety of consequential tasks, including conducting scientific research. To spur the developme

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History

DGX agent

arXiv:2606.08671v1 Announce Type: new Abstract: Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually

agentsarxiv-cs-lg
9 Jun 2026
Agents

Parthenon Law: A Self-Evolving Legal-Agent Framework

DGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

agentsarxiv-cs-ai
4 Jun 2026
← Previous
1…1112131415…230
Next →