AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Towards LLM-Assisted Architecture Recovery for Real-World ROS~2 Systems: An Agent-Based Multi-Level Approach to Hierarchical Structural Architecture Reconstruction

DGX agent

arXiv:2605.20055v1 Announce Type: cross Abstract: Explicit software architecture models are essential artifacts for communicating, analyzing, and evolving complex software-intensive systems. In ROS~2-

agentsarxiv-cs-ai
20 May 2026
Safety

TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents

DGX agent

arXiv:2602.11767v3 Announce Type: replace Abstract: Advances in large language models (LLMs) are driving a shift toward using reinforcement learning (RL) to train agents from iterative, multi-turn int

safetyarxiv-cs-ai
20 May 2026
Model Releases

Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap

DGX agent

arXiv:2605.17903v1 Announce Type: new Abstract: We automatically generate feedback causal fuzzy cognitive maps (FCMs) from text by teaching large-language-model agents to break the text into overlappi

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning

DGX agent

arXiv:2605.16309v1 Announce Type: new Abstract: LLM-based agents can recover from individual execution errors, yet they repeatedly fail on the same fault when the underlying process knowledge--operato

local-aiarxiv-cs-ai
19 May 2026
Safety

Body-Grounded Perspective Formation and Conative Attunement in Artificial Agents

DGX agent

arXiv:2605.16728v1 Announce Type: new Abstract: This paper proposes a minimal architecture for body-grounded perspective formation in artificial agents. Extending prior work, the model introduces an i

safetyarxiv-cs-ai
19 May 2026
Model Releases

Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis

DGX agent

arXiv:2605.18451v1 Announce Type: new Abstract: Designing realistic and functional 3D indoor rooms is essential for a wide range of applications, including interior design, virtual reality, gaming, an

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

ContractBench: Can LLM Agents Preserve Observation Contracts?

DGX agent

arXiv:2605.17281v1 Announce Type: cross Abstract: Tool-augmented LLM agents call APIs whose intermediate outputs, such as presigned URLs, session tokens, and OAuth state parameters, are observation co

model-releasesarxiv-cs-ai
19 May 2026
Safety

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

DGX agent

arXiv:2605.17909v1 Announce Type: new Abstract: As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency p

safetyarxiv-cs-ai
19 May 2026
Model Releases

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

DGX agent

arXiv:2507.21035v3 Announce Type: replace Abstract: Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injectio

DGX agent

arXiv:2605.17986v1 Announce Type: cross Abstract: AI agents such as OpenClaw are increasingly deployed in local workflows with access to external tools. This creates indirect prompt-injection (IPI) ri

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Multi-agent AI systems outperform human teams in creativity

DGX agent

arXiv:2605.17885v1 Announce Type: cross Abstract: Although artificial intelligence (AI) now matches or exceeds human performance across numerous cognitive tasks, creativity remains a highly contested

local-aiarxiv-cs-ai
19 May 2026
Model Releases

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

DGX agent

arXiv:2605.17596v1 Announce Type: new Abstract: We present NeuSymMS, an adaptive memory system that enables large language model (LLM) agents to learn, remember, and reason about users across sessions

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments

DGX agent

arXiv:2603.23231v2 Announce Type: replace Abstract: Empowering large language models with long-term memory is crucial for building agents that adapt to users' evolving needs. Existing evaluations of t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

The Scaling Laws of Skills in LLM Agent Systems

DGX agent

arXiv:2605.16508v1 Announce Type: cross Abstract: As agent systems scale, skills accumulate into large reusable libraries, yet their scaling laws remain poorly understood. Across 15 frontier LLMs, 1,1

model-releasesarxiv-cs-ai
19 May 2026
Hardware

TriAxialKV: Toward Extreme Low-Precision KV-Cache Quantization for Agentic Inference Tasks

DGX agent

arXiv:2605.17170v1 Announce Type: new Abstract: Agentic workloads have emerged as a major workload for LLM inference. They differ significantly from chat-only workloads, requiring long-context process

hardwarearxiv-cs-lg
19 May 2026
Model Releases

Watching, Reasoning, and Searching: A Video Deep Research Benchmark on Open Web for Agentic Video Reasoning

DGX agent

arXiv:2601.06943v2 Announce Type: replace-cross Abstract: In real-world video question answering scenarios, videos often provide only localized visual cues, while verifiable answers are distributed ac

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation

DGX agent

arXiv:2605.15343v1 Announce Type: new Abstract: LLM-based agents are increasingly used to simulate deliberative interactions such as negotiation, conflict resolution, and multi-turn opinion exchange.

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory

DGX agent

arXiv:2605.15759v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term memory to leverage information from past interactions. However, existing memory systems often face a

model-releasesarxiv-cs-cl
18 May 2026
Safety

From I/O to Code with Discovery Agent

DGX agent

arXiv:2605.15334v1 Announce Type: cross Abstract: The automatic synthesis of a program from any form of specification is regarded as a holy grail of computer science. Fueled by LLMs, NL2Code has achie

safetyarxiv-cs-ai
18 May 2026
Model Releases

PBT-Bench: Benchmarking AI Agents on Property-Based Testing

DGX agent

arXiv:2605.15229v1 Announce Type: cross Abstract: Existing code benchmarks measure whether an agent can produce any test that reproduces a known bug, or whether it can produce a patch that fixes a des

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

DGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

DGX agent

arXiv:2605.15204v1 Announce Type: new Abstract: Multi-agent orchestration frameworks such as LangChain, LangGraph, and CrewAI route tasks through graph-based pipelines but do not enforce the stage con

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

A Minimal Agent for Automated Theorem Proving

DGX agent

arXiv:2602.24273v3 Announce Type: replace Abstract: We propose a minimal agentic baseline that enables systematic comparison across different AI-based theorem prover architectures. This design impleme

model-releasesarxiv-cs-ai
15 May 2026
Safety

A Security Analysis of the OpenClaw AI Agent Framework

DGX agent

arXiv:2603.27517v3 Announce Type: replace-cross Abstract: AI agent frameworks connecting large language model (LLM) reasoning to host execution surfaces -- shell, filesystem, containers, and messaging

safetyarxiv-cs-ai
15 May 2026
Safety

CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG

DGX agent

arXiv:2605.11611v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for training agentic retrieval-augmented generation (RAG)

safetyarxiv-cs-ai
15 May 2026
Model Releases

Do Coding Agents Understand Least-Privilege Authorization?

DGX agent

arXiv:2605.14859v1 Announce Type: cross Abstract: As coding agents gain access to shells, repositories, and user files, least-privilege authorization becomes a prerequisite for safe deployment: an age

model-releasesarxiv-cs-ai
15 May 2026
Safety

Dual Hierarchical Dialogue Policy Learning for Legal Inquisitive Conversational Agents

DGX agent

arXiv:2605.14057v1 Announce Type: new Abstract: Most existing dialogue systems are user-driven, primarily designed to fulfill user requests. However, in many critical real-world scenarios, a conversat

safetyarxiv-cs-cl
15 May 2026
Model Releases

From Descriptive to Prescriptive: Uncover the Social Value Alignment of LLM-based Agents

DGX agent

arXiv:2605.14034v1 Announce Type: new Abstract: Wide applications of LLM-based agents require strong alignment with human social values. However, current works still exhibit deficiencies in self-cogni

model-releasesarxiv-cs-ai
15 May 2026
Local Ai

GraphFlow: An Architecture for Formally Verifiable Visual Workflows Enabling Reliable Agentic AI Automation

DGX agent

arXiv:2605.14968v1 Announce Type: new Abstract: GraphFlow is a visual workflow system designed to improve the reliability of agentic AI automation in multi-step, mission-critical processes. In these w

local-aiarxiv-cs-ai
15 May 2026
Model Releases

Optimizing PyTorch Inference with LLM-Based Multi-Agent Systems

DGX agent

arXiv:2511.16964v2 Announce Type: replace-cross Abstract: Maximizing performance on available GPU hardware is an ongoing challenge for modern AI inference systems. Traditional approaches include writi

model-releasesarxiv-cs-ai
15 May 2026
Agents

ProtoMedAgent: Multimodal Clinical Interpretability via Privacy-Aware Agentic Workflows

DGX agent

arXiv:2605.14113v1 Announce Type: cross Abstract: While interpretable prototype networks offer compelling case-based reasoning for clinical diagnostics, their raw continuous outputs lack the semantic

agentsarxiv-cs-ai
15 May 2026
Safety

SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration

DGX agent

arXiv:2605.14089v1 Announce Type: new Abstract: In recent years, a variety of powerful LLM-based agentic systems have been applied to automate complex tasks through task orchestration. However, existi

safetyarxiv-cs-ai
15 May 2026
Model Releases

EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents

DGX agent

arXiv:2605.13841v1 Announce Type: cross Abstract: Voice agents, artificial intelligence systems that conduct spoken conversations to complete tasks, are increasingly deployed across enterprise applica

model-releasesarxiv-cs-ai
14 May 2026
Research

GAAMA: Graph Augmented Associative Memory for Agents

DGX agent

arXiv:2603.27910v2 Announce Type: replace Abstract: AI agents that interact with users across multiple sessions require persistent long-term memory to maintain coherent, personalized behavior. Current

researcharxiv-cs-ai
14 May 2026
Safety

Integration of an Agent Model into an Open Simulation Architecture for Scenario-Based Testing of Automated Vehicles

DGX agent

arXiv:2605.13539v1 Announce Type: new Abstract: Simulative and scenario-based testing are crucial methods in the safety assurance for automated driving systems. To ensure that simulation results are r

safetyarxiv-cs-ro
14 May 2026
Safety

Large Language Models for Agentic NetOps and AIOps: Architectures, Evaluation, and Safety

DGX agent

arXiv:2605.12729v1 Announce Type: cross Abstract: Large language models are increasingly being used to support network operations (NetOps) and artificial intelligence for IT operations (AIOps), includ

safetyarxiv-cs-ai
14 May 2026
Safety

Macro-Action Based Multi-Agent Instruction Following through Value Cancellation

DGX agent

arXiv:2605.12655v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) in real-world use cases may need to adapt to external natural language instructions that interrupt ongoing beh

safetyarxiv-cs-ai
14 May 2026
Model Releases

ReTool-Video: Recursive Tool-Using Video Agents with Meta-Augmented Tool Grounding

DGX agent

arXiv:2605.13228v1 Announce Type: cross Abstract: Video understanding requires active evidence seeking, motivating tool-augmented video agents for temporal reasoning, cross-modal understanding, and co

model-releasesarxiv-cs-ai
14 May 2026
Agents

Beyond Manual Curation: Augmenting Targeted Protein Degradation Databases via Agentic Literature Extraction Workflows

DGX agent

arXiv:2605.11221v1 Announce Type: cross Abstract: Predictive models in biomedicine depend on structured assay data locked in the text, tables, and supplements of primary publications. This bottleneck

agentsarxiv-cs-lg
13 May 2026
Model Releases

Correcting Selection Bias in Sparse User Feedback for Large Language Model Quality Estimation: A Multi-Agent Hierarchical Bayesian Approach

DGX agent

arXiv:2605.12177v1 Announce Type: new Abstract: [Abridged] Production LLM deployments receive feedback from a non-random fraction of users: thumbs sit mostly in the tails of the satisfaction distribut

model-releasesarxiv-cs-cl
13 May 2026
Safety

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation

DGX agent

arXiv:2605.11853v1 Announce Type: cross Abstract: Reinforcement learning has become a widely used post-training approach for LLM agents, where training commonly relies on outcome-level rewards that pr

safetyarxiv-cs-cl
13 May 2026
Safety

JACoP: Joint Alignment for Compliant Multi-Agent Prediction

DGX agent

arXiv:2605.11385v1 Announce Type: new Abstract: Stochastic Human Trajectory Prediction (HTP) using generative modeling has emerged as a significant area of research. Although state-of-the-art models e

safetyarxiv-cs-cv
13 May 2026
Research

ReVision: Scaling Computer-Use Agents via Temporal Visual Redundancy Reduction

DGX agent

arXiv:2605.11212v1 Announce Type: new Abstract: Computer-use agents~(CUAs) rely on visual observations of graphical user interfaces, where each screenshot is encoded into a large number of visual toke

researcharxiv-cs-cl
13 May 2026
Model Releases

Agentic MIP Research: Accelerated Constraint Handler Generation

DGX agent

arXiv:2605.09186v1 Announce Type: new Abstract: Mixed-integer programming (MIP) research is both mathematically sophisticated and engineering-intensive: testing an algorithmic hypothesis within a bran

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Aligning Agents via Planning: A Benchmark for Trajectory-Level Reward Modeling

DGX agent

arXiv:2604.08178v2 Announce Type: replace Abstract: In classical Reinforcement Learning from Human Feedback (RLHF), Reward Models (RMs) serve as the fundamental signal provider for model alignment. As

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Bridging Modalities, Spanning Time: Structured Memory for Ultra-Long Agentic Video Reasoning

DGX agent

arXiv:2605.08271v1 Announce Type: cross Abstract: Understanding ultra-long videos such as egocentric recordings, live streams, or surveillance footage spanning days to weeks, remains a challenge. For

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…9192939495…236
Next →