AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

Agentic SABRE: An Uncertainty-Aware Neuro-Symbolic Multi-Agent Framework for Adaptive Ransomware Detection

DGX agent

arXiv:2607.04292v1 Announce Type: new Abstract: Ransomware has evolved into a complex, adaptive, and fast-moving adversary category in which static signatures and monolithic classifiers fail to genera

agentsarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

NKI-Agent: Domain-Specific Fine-Tuning and Agentic Tool Use for Neuron Kernel Generation

DGX agent

arXiv:2607.04395v1 Announce Type: new Abstract: Recent agentic approaches to LLM-based kernel generation have achieved impressive results on CUDA. For emerging AI accelerators such as AWS Trainium and

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Power Systems Agent Benchmark: Executable Evaluation of AI Agents in Electric Power Engineering

DGX agent

arXiv:2606.20950v2 Announce Type: replace Abstract: Executable evaluation -- checking the consequences of an agent's actions with a program rather than grading its prose -- has become a prominent way

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Position: Collaborative Agentic AI Needs Interoperability Across Ecosystems

DGX agent

arXiv:2505.21550v2 Announce Type: replace-cross Abstract: Collaborative agentic AI is projected to transform entire industries by enabling AI-powered agents to autonomously perceive, plan, and act wit

agentsarxiv-cs-ai
1 Jul 2026
Agents

Is Lying an Emergent Behaviour in LLMs? Evidence from Gaslighting AI agents in a Sustainability Game

DGX agent

arXiv:2606.28456v1 Announce Type: cross Abstract: LLMs agents are increasingly used in multi-agent settings, yet their behaviour in sustainability games remains largely unexplored. This work investiga

agentsarxiv-cs-ai
30 Jun 2026
Agents

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

DGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Why Trust Your Agent? Empirical Security Gains from TRiSM-Guided Agentic Workflows in Healthcare

DGX agent

arXiv:2606.28666v1 Announce Type: cross Abstract: Agent-based AI has enabled the automation of tasks by exposing application tools and resources to large language models (LLMs). However, to improve sc

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

IPO Finance Agent: Evaluation of LLM Financial Analysts beyond Finance Agent v2, with Automated Rubric Generation -- the Case of the SpaceX (SPCX) IPO

DGX agent

arXiv:2606.23032v2 Announce Type: replace Abstract: Finance Agent v2 (by Vals AI) has emerged as the reference benchmark for evaluating both Anthropic Claude and OpenAI ChatGPT frontier language model

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

MoCA-Agent: A Market-of-Claims Code Agent for Financial and Numerical Reasoning

DGX agent

arXiv:2606.11537v1 Announce Type: new Abstract: Financial and tabular question answering requires more than fluent reasoning: answers must be grounded in the exact facts, formulas, units, signs, and s

agentsarxiv-cs-ai
11 Jun 2026
Local Ai

Agent libOS: A Library-OS-Inspired Runtime for Long-Running, Capability-Controlled LLM Agents

DGX agent

arXiv:2606.03895v1 Announce Type: cross Abstract: Large language model (LLM) agents are evolving from request-response assistants into long-running software actors: they maintain state across model ca

local-aiarxiv-cs-ai
3 Jun 2026
Agents

SAGE: A Quantitative Evaluation of Socialized Evolution in Agent Ecosystems

DGX agent

arXiv:2606.03544v1 Announce Type: new Abstract: Self-improving language agents are typically evaluated in isolation: an agent attempts a task, receives feedback, and iteratively refines its own behavi

agentsarxiv-cs-ai
3 Jun 2026
Agents

Distributed Non-Uniform Scaling Control of Multi-Agent Formation with Dynamic Agent Joining

DGX agent

arXiv:2605.29191v1 Announce Type: cross Abstract: Non-uniform scaling control of formation enables multi-agent systems to adjust their shape by scaling with different ratios along different coordinate

agentsarxiv-cs-ro
29 May 2026
Safety

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

DGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

safetyarxiv-cs-ai
29 May 2026
Model Releases

Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

DGX agent

arXiv:2605.28816v1 Announce Type: new Abstract: World models for interactive video generation have largely focused on single-agent settings, where future observations are generated from a single contr

model-releasesarxiv-cs-cv
28 May 2026
Agents

Multi-agent Collaboration with State Management

DGX agent

arXiv:2605.20563v1 Announce Type: cross Abstract: Recent advances in multi-agent systems have shown great potential for solving complex tasks. However, when multiple agents edit a shared codebase conc

agentsarxiv-cs-cl
21 May 2026
Agents

ARM: Discovering Agentic Reasoning Modules for Generalizable Multi-Agent Systems

DGX agent

arXiv:2510.05746v2 Announce Type: replace Abstract: Large Language Model (LLM)-powered Multi-agent systems (MAS) have achieved state-of-the-art results on various complex reasoning tasks. Recent works

agentsarxiv-cs-ai
20 May 2026
Model Releases

ADR: An Agentic Detection System for Enterprise Agentic AI Security

DGX agent

arXiv:2605.17380v1 Announce Type: new Abstract: We present the Agentic AI Detection and Response (ADR) system, the first large-scale, production-proven enterprise framework for securing AI agents oper

model-releasesarxiv-cs-ai
19 May 2026
Safety

Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces

DGX agent

arXiv:2605.17698v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) as autonomous economic agents introduces systemic risks that extend beyond individual capability failures

safetyarxiv-cs-lg
19 May 2026
Local Ai

MementoGUI: Learning Agentic Multimodal Memory Control for Long-Horizon GUI Agents

DGX agent

arXiv:2605.18652v1 Announce Type: new Abstract: Recent GUI agents have made substantial progress in visual grounding and action prediction, yet they remain brittle in long-horizon tasks that require m

local-aiarxiv-cs-cv
19 May 2026
Agents

SkillMaster: Toward Autonomous Skill Mastery in LLM Agents

DGX agent

arXiv:2605.08693v1 Announce Type: new Abstract: Skills provide an effective mechanism for improving LLM agents on complex tasks, yet in existing agent frameworks, their creation, refinement, and selec

agentsarxiv-cs-ai
12 May 2026
Agents

Material Database Agent: A Multimodal Agentic Framework for Scientific Literature Mining

DGX agent

arXiv:2605.04278v1 Announce Type: new Abstract: Materials science workflows rely on structured and unstructured data from the vast body of available scientific literature. However, most of the experim

agentsarxiv-cs-cl
7 May 2026
Agents

A Study of Belief Revision Postulates in Multi-Agent Systems (Extended Version)

DGX agent

arXiv:2605.02249v1 Announce Type: new Abstract: We investigate the belief revision problem in epistemic planning, i.e., what will be the beliefs of all agents in a multi-agent system after an agent ga

agentsarxiv-cs-ai
6 May 2026
Model Releases

Agentic-imodels: Evolving agentic interpretability tools via autoresearch

DGX agent

arXiv:2605.03808v1 Announce Type: cross Abstract: Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Agent Factories for High Level Synthesis: How Far Can General-Purpose Coding Agents Go in Hardware Optimization?

DGX agent

arXiv:2603.25719v2 Announce Type: replace-cross Abstract: We present an empirical study of how far general-purpose coding agents -- without hardware-specific training -- can optimize hardware designs

model-releasesarxiv-cs-lg
4 May 2026
Agents

Agentic AI platforms for autonomous training and rule induction of human-human and virus-human protein-protein interactions

DGX agent

arXiv:2604.23924v1 Announce Type: new Abstract: We instruct an AI agent to construct two separate agentic AI platforms: one for autonomous training of predictive ML models for human-human and virus-hu

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

Complete Cyclic Subtask Graphs for Tool-Using LLM Agents: Flexibility, Cost, and Bottlenecks in Multi-Agent Workflows

DGX agent

arXiv:2604.22820v1 Announce Type: cross Abstract: Long-horizon tool-using tasks sometimes benefit from revisiting earlier subtasks for recovery and exploration, but added multi-agent workflow flexibil

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Cost-Effective Communication: An Auction-based Method for Language Agent Interaction

DGX agent

arXiv:2511.13193v2 Announce Type: replace Abstract: Multi-agent systems (MAS) built on large language models (LLMs) often suffer from inefficient 'free-for-all' communication, leading to exponential t

agentsarxiv-cs-ai
27 Apr 2026
Safety

Survey on Evaluation of LLM-based Agents

DGX agent

arXiv:2503.16416v2 Announce Type: replace Abstract: LLM-based agents represent a paradigm shift in AI, enabling autonomous systems to plan, reason, and use tools while interacting with dynamic environ

safetyarxiv-cs-ai
24 Apr 2026
Agents

Forage V2: Knowledge Evolution and Transfer in Autonomous Agent Organizations

DGX agent

arXiv:2604.19837v1 Announce Type: new Abstract: Autonomous agents operating in open-world tasks -- where the completion boundary is not given in advance -- face denominator blindness: they systematica

agentsarxiv-cs-ai
23 Apr 2026
Agents

DR-MMSearchAgent: Deepening Reasoning in Multimodal Search Agents

DGX agent

arXiv:2604.19264v1 Announce Type: new Abstract: Agentic multimodal models have garnered significant attention for their ability to leverage external tools to tackle complex tasks. However, it is obser

agentsarxiv-cs-cv
22 Apr 2026
Agents

Explicit Trait Inference for Multi-Agent Coordination

DGX agent

arXiv:2604.19278v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) show promise on complex tasks but remain prone to coordination failures such as goal drift, error cascades, and misa

agentsarxiv-cs-ai
22 Apr 2026
Agents

OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration

DGX agent

arXiv:2505.11765v3 Announce Type: replace-cross Abstract: Agents powered by advanced large language models (LLMs) have demonstrated impressive capabilities across diverse complex applications. Recentl

agentsarxiv-cs-ai
22 Apr 2026
Agents

Rethinking Scale: Deployment Trade-offs of Small Language Models under Agent Paradigms

DGX agent

arXiv:2604.19299v1 Announce Type: cross Abstract: Despite the impressive capabilities of large language models, their substantial computational costs, latency, and privacy risks hinder their widesprea

agentsarxiv-cs-ai
22 Apr 2026
Agents

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0)

DGX agent

arXiv:2604.17091v1 Announce Type: new Abstract: Long-horizon large language model (LLM) agents are fundamentally limited by context. As interactions become longer, tool descriptions, retrieved memorie

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

StepPO: Step-Aligned Policy Optimization for Agentic Reinforcement Learning

DGX agent

arXiv:2604.18401v1 Announce Type: new Abstract: General agents have given rise to phenomenal applications such as OpenClaw and Claude Code. As these agent systems (a.k.a. Harnesses) strive for bolder

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Bilevel Optimization of Agent Skills via Monte Carlo Tree Search

DGX agent

arXiv:2604.15709v1 Announce Type: new Abstract: Agent exttt{skills} are structured collections of instructions, tools, and supporting resources that help large language model (LLM) agents perform part

agentsarxiv-cs-ai
20 Apr 2026
Agents

MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation

DGX agent

arXiv:2604.16175v1 Announce Type: new Abstract: Automated 3D radiology report generation often suffers from clinical hallucinations and a lack of the iterative verification found in human practice. Wh

agentsarxiv-cs-ai
20 Apr 2026
Agents

AgentSPEX: An Agent SPecification and EXecution Language

DGX agent

arXiv:2604.13346v1 Announce Type: new Abstract: Language-model agent systems commonly rely on reactive prompting, in which a single instruction guides the model through an open-ended sequence of reaso

agentsarxiv-cs-cl
16 Apr 2026
Agents

Form Without Function: Agent Social Behavior in the Moltbook Network

DGX agent

arXiv:2604.13052v1 Announce Type: cross Abstract: Moltbook is a social network where every participant is an AI agent. We analyze 1,312,238 posts, 6.7~million comments, and over 120,000 agent profiles

agentsarxiv-cs-cl
16 Apr 2026
Model Releases

Modality-Native Routing in Agent-to-Agent Networks: A Multimodal A2A Protocol Extension

DGX agent

arXiv:2604.12213v1 Announce Type: new Abstract: Preserving multimodal signals across agent boundaries is necessary for accurate cross-modal reasoning, but it is not sufficient. We show that modality-n

model-releasesarxiv-cs-ai
15 Apr 2026
Agents

Credit-Budgeted ICPC-Style Coding: When Agents Must Pay for Every Decision

DGX agent

arXiv:2604.10182v1 Announce Type: new Abstract: Current evaluations of autonomous coding agents assume an unrealistic, infinite-resource environment. However, real-world software engineering is a reso

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Do Agent Rules Shape or Distort? Guardrails Beat Guidance in Coding Agents

DGX agent

arXiv:2604.11088v1 Announce Type: new Abstract: Developers increasingly guide AI coding agents through natural language instruction files (e.g., CLAUDE.md, .cursorrules), yet no controlled study has m

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

From Agent Loops to Structured Graphs:A Scheduler-Theoretic Framework for LLM Agent Execution

DGX agent

arXiv:2604.11378v1 Announce Type: new Abstract: The dominant paradigm for building LLM based agents is the Agent Loop, an iterative cycle where a single language model decides what to do next by readi

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Mem^2Evolve: Towards Self-Evolving Agents via Co-Evolutionary Capability Expansion and Experience Distillation

DGX agent

arXiv:2604.10923v1 Announce Type: cross Abstract: While large language model--powered agents can self-evolve by accumulating experience or by dynamically creating new assets (i.e., tools or expert age

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

PAC-BENCH: Evaluating Multi-Agent Collaboration under Privacy Constraints

DGX agent

arXiv:2604.11523v1 Announce Type: new Abstract: We are entering an era in which individuals and organizations increasingly deploy dedicated AI agents that interact and collaborate with other agents. H

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets

DGX agent

arXiv:2604.02460v2 Announce Type: replace Abstract: Recent work reports strong performance from multi-agent LLM systems (MAS), but these gains are often confounded by increased test-time computation.

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

Group-Aware Coordination Graph for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2404.10976v4 Announce Type: replace-cross Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) necessitates seamless collaboration among agents, often represented by an underlying rel

agentsarxiv-cs-ai
13 Apr 2026
Agents

Computer Environments Elicit General Agentic Intelligence in LLMs

DGX agent

arXiv:2601.16206v3 Announce Type: replace-cross Abstract: Agentic intelligence in large language models (LLMs) requires not only model intrinsic capabilities but also interactions with external enviro

agentsarxiv-cs-ai
10 Apr 2026
← Previous
1…56789…230
Next →