AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Prompt, Plan, Extract: Zero-Shot Agentic LLMs Workflows for Lung Pathology Extraction from Clinical Narratives

DGX agent

arXiv:2606.19852v2 Announce Type: replace Abstract: Information extraction from pathology reports is essential for cancer staging, tumor registry population. Yet key data remains embedded in narrative

agentsarxiv-cs-cl
26 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Socratic agents for autonomous scientific discovery in high-dimensional physical systems

DGX agent

arXiv:2606.26722v1 Announce Type: new Abstract: The automation of scientific discovery has reached an inflection point. While AI systems now operate instruments, optimize parameters and generate hypot

agentsarxiv-cs-ai
26 Jun 2026
Safety

Beyond Next-Observation Prediction: Agent-Authored World Modeling for Sequential Decision Making

DGX agent

arXiv:2606.25421v1 Announce Type: new Abstract: Recent studies on world modeling for Large Language Model (LLM) agents typically formulate the learning objective as next-observation prediction. Howeve

safetyarxiv-cs-cl
25 Jun 2026
Safety

GCT-MARL: Graph-Based Contrastive Transfer for Sample-Efficient Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.25073v1 Announce Type: new Abstract: In cooperative multi-agent reinforcement learning (MARL), from a deployment perspective, it is challenging and expensive to train agents from scratch fo

safetyarxiv-cs-lg
25 Jun 2026
Agents

MANGO: Automated Multi-Agent Test Oracle Generation for Vision-Language-Action Models

DGX agent

arXiv:2606.24815v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are emerging robotic control systems that integrate perception, language understanding, and action generation in a

agentsarxiv-cs-ro
25 Jun 2026
Agents

Memory Makes the Difference: Evaluating How Different Memory Roles Shape Conversational Agents

DGX agent

arXiv:2606.25361v1 Announce Type: new Abstract: Prior research on memory mechanism in RAG-based conversational system has emphasized how memory is stored and retrieved. However, far less is known abou

agentsarxiv-cs-cl
25 Jun 2026
Safety

Offline Multi-agent Continual Cooperation via Skill Partition and Reuse

DGX agent

arXiv:2606.25389v1 Announce Type: new Abstract: Extracting skills from multi-agent offline dataset improves learning efficiency via sharing task-invariant coordination skills among tasks. In settings

safetyarxiv-cs-ai
25 Jun 2026
Safety

Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems

DGX agent

arXiv:2606.24416v1 Announce Type: new Abstract: Network operators' changing policies, service requirements, and stringent real-time constraints render existing methods designed with fixed objectives a

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.24601v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) addresses the problem of training multiple agents that pursue collaborative, competitive, or mixed objectives.

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent

DGX agent

arXiv:2606.17698v2 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the que

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Safe and Generalizable Hierarchical Multi-Agent RL via Constraint Manifold Control

DGX agent

arXiv:2606.24010v1 Announce Type: new Abstract: Multi-agent systems are widely used in safety-critical applications that require coordinated behavior under strict safety constraints. Existing approach

safetyarxiv-cs-ai
24 Jun 2026
Agents

Varying Bundle Size Reactive Multi-Task Assignment using Selective Cost Estimation for Multi-Agent Systems

DGX agent

arXiv:2606.24462v1 Announce Type: new Abstract: This paper presents a scalable framework for multi-robot task allocation in complex environments where estimating task execution costs is computationall

agentsarxiv-cs-ro
24 Jun 2026
Agents

Hypothesis-Disciplined Multi-Agent Automated Formalization of Asymptotic Statistical Theory

DGX agent

arXiv:2606.20642v1 Announce Type: cross Abstract: Asymptotic statistical theory is a challenging domain for AI-assisted formalization: its central results mix convergence statements, asymptotic expans

agentsarxiv-cs-lg
23 Jun 2026
Agents

RAPID: A Reproducible Multi-Agent Pipeline for Interpretable Disaster Damage Assessment from Satellite and Street-View Imagery

DGX agent

arXiv:2606.21819v1 Announce Type: new Abstract: Due to the increasing frequency and intensity of extreme climate events, there is a clear demand for intelligent, scalable, and autonomous approaches to

agentsarxiv-cs-cv
23 Jun 2026
Agents

SPARC: A Multi-Agent System for Electrical Circuit Question Answering

DGX agent

arXiv:2606.20643v1 Announce Type: cross Abstract: Electrical circuit diagram QA tasks require complex mathematical reasoning, which remains challenging for multimodal LLMs. We present SPARC, a multi-a

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

DGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

WebCryptoAgent: Agentic Crypto Trading with Web Informatics

DGX agent

arXiv:2601.04687v2 Announce Type: replace Abstract: Cryptocurrency trading increasingly depends on timely integration of heterogeneous web information and market microstructure signals to support shor

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

DGX agent

arXiv:2606.12320v1 Announce Type: new Abstract: Enterprise security was built to govern data boundaries: the protected surface was data at rest and in transit, and the controls -- access control, data

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

DGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

model-releasesarxiv-cs-cl
11 Jun 2026
Local Ai

Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment

DGX agent

arXiv:2606.11672v1 Announce Type: cross Abstract: This paper explores the value of agentic AI tools for cybersecurity purposes. We evaluate the efficacy of a general-purpose GenAI Large Language Model

local-aiarxiv-cs-ai
11 Jun 2026
Model Releases

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

DGX agent

arXiv:2606.12344v1 Announce Type: cross Abstract: General-purpose agents such as OpenClaw are increasingly used as autonomous tool users, but their coding ability is difficult to measure under SWE-ben

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies

DGX agent

arXiv:2602.18291v2 Announce Type: replace Abstract: Online Multi-Agent Reinforcement Learning (MARL) is a prominent framework for efficient agent coordination. Crucially, enhancing policy expressivene

safetyarxiv-cs-ai
11 Jun 2026
Agents

DrivingAgent: Design and Scheduling Agents for Autonomous Driving Systems

DGX agent

arXiv:2606.12236v1 Announce Type: cross Abstract: Many autonomous driving systems are increasingly incorporating foundation models to improve generalization and handle long-tail scenarios. However, th

agentsarxiv-cs-cv
11 Jun 2026
Agents

Fanar-Sadiq: A Multi-Agent Architecture for Grounded Islamic QA

DGX agent

arXiv:2603.08501v3 Announce Type: replace Abstract: Large language models (LLMs) can answer religious knowledge queries fluently, yet they often hallucinate and misattribute sources, which is especial

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

DGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

model-releasesarxiv-cs-ai
11 Jun 2026
Hardware

INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration

DGX agent

arXiv:2606.11440v1 Announce Type: new Abstract: Existing multi-agent LLM orchestration methods, ranging from brute-force ensembles to learned routers, select models and topologies based on task and mo

hardwarearxiv-cs-ai
11 Jun 2026
Model Releases

MARIC: Multi-Agent Reasoning for Image Classification

DGX agent

arXiv:2509.14860v2 Announce Type: replace-cross Abstract: Image classification has traditionally relied on parameter-intensive model training, requiring large-scale annotated datasets and extensive fi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MedCTA: A Benchmark for Clinical Tool Agents

DGX agent

arXiv:2606.11702v1 Announce Type: cross Abstract: To make clinically grounded decisions, medical AI agents are expected to go beyond simple recognition and be capable of tool retrieval, evidence acqui

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

DGX agent

arXiv:2606.12329v1 Announce Type: new Abstract: AI coding assistants now support a growing share of software work, from quick scripts to production applications. Yet these agents remain largely statel

local-aiarxiv-cs-ai
11 Jun 2026
Agents

Sustainability assessment using multimodal AI agents

DGX agent

arXiv:2507.17012v2 Announce Type: replace Abstract: Reducing the rapidly growing environmental impact of the computing industry requires assessing the emissions of electronics at scale. However, a tra

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

DGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Assessing Automated Prompt Injection Attacks in Agentic Environments

DGX agent

arXiv:2606.10525v1 Announce Type: cross Abstract: Indirect prompt injection poses a critical threat to LLM agents that interact with untrusted external data, yet automated attack methods--proven effec

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

au-Rec: A Verifiable Benchmark for Agentic Recommender Systems

DGX agent

arXiv:2606.10156v1 Announce Type: cross Abstract: As recommender systems transition toward agentic, multi-turn conversational interfaces, evaluation paradigms have struggled to keep pace. Current benc

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering

DGX agent

arXiv:2510.04514v3 Announce Type: replace Abstract: Recent multimodal LLMs have shown promise in chart-based visual question answering, but their performance declines sharply on unannotated charts-tho

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

Deployment-Time Memorization in Foundation-Model Agents

DGX agent

arXiv:2606.10062v1 Announce Type: new Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time fun

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

Agentic Search for Counterfactual Recourse under Fixed LLM Budgets

DGX agent

arXiv:2606.08696v1 Announce Type: cross Abstract: Counterfactual recourse aims to provide actionable feature changes that would alter an unfavorable decision made by a predictive model. In practice, a

agentsarxiv-cs-ai
9 Jun 2026
Local Ai

AGENTSERVESIM: A Hardware-aware Simulator for Multi-Turn LLM Agent Serving

DGX agent

arXiv:2606.09613v1 Announce Type: cross Abstract: Multi-turn LLM agents interleave model calls with external tool invocations, shifting serving from stateless request processing to stateful program ex

local-aiarxiv-cs-ai
9 Jun 2026
Agents

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

DGX agent

arXiv:2606.08952v1 Announce Type: new Abstract: Multimodal Foundation Models (MFMs) have made substantial progress, yet remain fragile in spatial reasoning over the physical world. A key bottleneck li

agentsarxiv-cs-ai
9 Jun 2026
Agents

An AI Security Agent for University ACMIS: Multi-Vector Threat Detection and Automated Response

DGX agent

arXiv:2606.08270v1 Announce Type: cross Abstract: University Academic Management Information Systems (ACMIS) are high-value targets for a wide spectrum of security threats including brute-force login

agentsarxiv-cs-ai
9 Jun 2026
Safety

Autonomous Incident Resolution at Hyperscale: An Agentic AI Architecture for Network Operations

DGX agent

arXiv:2606.09122v1 Announce Type: cross Abstract: Cloud network infrastructure at hyperscale presents unique operational challenges where traditional human-driven incident response cannot keep pace wi

safetyarxiv-cs-ai
9 Jun 2026
Hardware

BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in Mobile Networks

DGX agent

arXiv:2602.14033v1 Announce Type: cross Abstract: Future sixth-generation (6G) mobile networks will demand artificial intelligence (AI) agents that are not only autonomous and efficient, but also capa

hardwarearxiv-cs-ai
9 Jun 2026
Safety

Brain-Prompt Injection: A Route-Safety Audit for BCI-LLM Agents

DGX agent

arXiv:2606.09315v1 Announce Type: cross Abstract: BCI-to-agent pipelines turn decoded neural activity into an authorization channel for tool-use agents, exposing a new attack surface we call brain-pro

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory

DGX agent

arXiv:2606.09365v1 Announce Type: new Abstract: Medical agent systems are increasingly expected to support interactive clinical decision making rather than only static question answering. In such sett

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

From Human Guidance to Autonomy: Agent Skill System for End-to-End LLM Deployment on Spatial NPUs

DGX agent

arXiv:2606.07586v1 Announce Type: cross Abstract: Spatial neural processing units (NPUs) provide an energy-efficient platform for edge LLM inference, but efficiently deploying an LLM end-to-end on suc

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

DGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

MAR:Multi-Agent Reflexion Improves Reasoning Abilities in LLMs

DGX agent

arXiv:2512.20845v2 Announce Type: replace Abstract: LLMs have shown the capacity to improve their performance on reasoning tasks through reflecting on their mistakes, and acting with these reflections

agentsarxiv-cs-ai
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…6768697071…236
Next →