AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

NaviGNN: Multi-Agent Reinforcement Learning and Graph Neural Network for Sustainable Mobility in Futuristic Smart Cities

DGX agent

arXiv:2507.15143v3 Announce Type: replace Abstract: This paper investigates the feasibility of human mobility in extreme urban morphologies characterized by high-density vertical structures and linear

agentsarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PhysicianBench: Evaluating LLM Agents in Real-World EHR Environments

DGX agent

arXiv:2605.02240v1 Announce Type: new Abstract: We introduce PhysicianBench, a benchmark for evaluating LLM agents on physician tasks grounded in real clinical setting within electronic health record

model-releasesarxiv-cs-ai
6 May 2026
Agents

When LLM Agents Meet Graph Optimization: An Automated Data Quality Improvement Approach

DGX agent

arXiv:2510.08952v4 Announce Type: replace Abstract: Text-attributed graphs (TAGs) have become a key form of graph-structured data in modern data management and analytics, combining structural relation

agentsarxiv-cs-lg
6 May 2026
Agents

Are Tools All We Need? Unveiling the Tool-Use Tax in LLM Agents

DGX agent

arXiv:2605.00136v1 Announce Type: new Abstract: Tool-augmented reasoning has become a popular direction for LLM-based agents, and it is widely assumed to improve reasoning and reliability. However, we

agentsarxiv-cs-ai
5 May 2026
Safety

MAD-OPD: Breaking the Ceiling in On-Policy Distillation via Multi-Agent Debate

DGX agent

arXiv:2605.01347v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own trajectories under token-level teacher supervision, but existing methods are capped by a single

safetyarxiv-cs-cl
5 May 2026
Agents

Optimistic {epsilon}-Greedy Exploration for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2502.03506v2 Announce Type: replace-cross Abstract: The Centralized Training with Decentralized Execution (CTDE) paradigm is widely used in cooperative multi-agent reinforcement learning. Howeve

agentsarxiv-cs-lg
5 May 2026
Agents

Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectives

DGX agent

arXiv:2604.06091v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly acting as human delegates in multi-agent environments, where a representative agent integrates di

agentsarxiv-cs-cl
5 May 2026
Agents

Synergistic Perception and Generative Recomposition: A Multi-Agent Orchestration for Expert-Level Building Inspection

DGX agent

arXiv:2603.20143v2 Announce Type: replace Abstract: Building facade defect inspection is fundamental to structural health monitoring and sustainable urban maintenance, yet it remains a formidable chal

agentsarxiv-cs-cv
5 May 2026
Safety

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy

DGX agent

arXiv:2605.01101v1 Announce Type: cross Abstract: This paper develops Virtual Speech Therapist (VST), an intelligent agent-based platform that streamlines stuttering assessment and delivers customized

safetyarxiv-cs-cl
5 May 2026
Local Ai

NonZero: Interaction-Guided Exploration for Multi-Agent Monte Carlo Tree Search

DGX agent

arXiv:2605.00751v1 Announce Type: new Abstract: Monte Carlo Tree Search (MCTS) scales poorly in cooperative multi-agent domains because expansion must consider an exponentially large set of joint acti

local-aiarxiv-cs-lg
4 May 2026
Safety

SAGA: Workflow-Atomic Scheduling for AI Agent Inference on GPU Clusters

DGX agent

arXiv:2605.00528v1 Announce Type: cross Abstract: AI agents execute tens to hundreds of chained LLM calls per task, yet GPU schedulers treat each call as independent, discarding gigabytes of intermedi

safetyarxiv-cs-lg
4 May 2026
Agents

A Grid-Aware Agent-Based Model for Analyzing Electric Vehicle Charging Systems

DGX agent

arXiv:2604.27849v1 Announce Type: new Abstract: This paper presents a configurable, grid-aware Agent-Based Model (ABM) for the systematic analysis of electric vehicle (EV) charging systems under confi

agentsarxiv-cs-ai
1 May 2026
Agents

Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making

DGX agent

arXiv:2604.27753v1 Announce Type: new Abstract: This article outlines a new framework of traffic light optimization through a digital twin of the transport infrastructure, managed by agentic AI to ens

agentsarxiv-cs-ai
1 May 2026
Agents

Chronology of Multi-Agent Interactions for Provenance of Evolving Information

DGX agent

arXiv:2504.12612v2 Announce Type: replace Abstract: Provenance is the chronological history of things, resonating with the fundamental pursuit to uncover origins, trace connections, and situate entiti

agentsarxiv-cs-ai
1 May 2026
Model Releases

ObjectGraph: From Document Injection to Knowledge Traversal -- A Native File Format for the Agentic Era

DGX agent

arXiv:2604.27820v1 Announce Type: new Abstract: Every document format in existence was designed for a human reader moving linearly through text. Autonomous LLM agents do not read - they retrieve. This

model-releasesarxiv-cs-ai
1 May 2026
Safety

OpAgent: Operator Agent for Web Navigation

DGX agent

arXiv:2602.13559v2 Announce Type: replace Abstract: To fulfill user instructions, autonomous web agents must contend with the inherent complexity and volatile nature of real-world websites. Convention

safetyarxiv-cs-ai
1 May 2026
Model Releases

WindowsWorld: A Process-Centric Benchmark of Autonomous GUI Agents in Professional Cross-Application Environments

DGX agent

arXiv:2604.27776v1 Announce Type: new Abstract: While GUI agents have shown impressive capabilities in common computer-use tasks such as OSWorld, current benchmarks mainly focus on isolated and single

model-releasesarxiv-cs-ai
1 May 2026
Agents

DreamProver: Evolving Transferable Lemma Libraries via a Wake-Sleep Theorem-Proving Agent

DGX agent

arXiv:2604.26311v1 Announce Type: new Abstract: We introduce DreamProver, an agentic framework that leverages a 'wake-sleep' program induction paradigm to discover reusable lemmas for formal theorem p

agentsarxiv-cs-ai
30 Apr 2026
Model Releases

SWE-Edit: Rethinking Code Editing for Efficient SWE-Agent

DGX agent

arXiv:2604.26102v1 Announce Type: cross Abstract: Large language model agents have achieved remarkable progress on software engineering tasks, yet current approaches suffer from a fundamental context

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

WebAggregator: Enhancing Compositional Reasoning Capabilities of Deep Research Agent Foundation Models

DGX agent

arXiv:2510.14438v2 Announce Type: replace Abstract: The hallmark of Deep Research agents lies in compositional reasoning, the capacity to aggregate distributed, heterogeneous information into coherent

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Benchmarking and Improving GUI Agents in High-Dynamic Environments

DGX agent

arXiv:2604.25380v1 Announce Type: new Abstract: Recent advancements in Graphical User Interface (GUI) agents have predominantly focused on training paradigms like supervised fine-tuning (SFT) and rein

model-releasesarxiv-cs-cv
29 Apr 2026
Model Releases

Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

DGX agent

arXiv:2604.24964v1 Announce Type: cross Abstract: Existing web agent benchmarks have largely converged on short, single-site tasks that frontier models are approaching saturation on. However, real wor

model-releasesarxiv-cs-cl
29 Apr 2026
Research

Why Do LLM-based Web Agents Fail? A Hierarchical Planning Perspective

DGX agent

arXiv:2603.14248v2 Announce Type: replace-cross Abstract: Large language model (LLM) web agents are increasingly used for web navigation but remain far from human reliability on realistic, long-horizo

researcharxiv-cs-cl
29 Apr 2026
Hardware

Agentic Fusion of Large Atomic and Language Models to Accelerate Materials Discovery

DGX agent

arXiv:2604.23758v1 Announce Type: new Abstract: The discovery of novel materials is critical for global energy and quantum technology transitions. While deep learning has fundamentally reshaped this l

hardwarearxiv-cs-lg
28 Apr 2026
Model Releases

AgentPulse: A Continuous Multi-Signal Framework for Evaluating AI Agents in Deployment

DGX agent

arXiv:2604.24038v1 Announce Type: new Abstract: Static benchmarks measure what AI agents can do at a fixed point in time but not how they are adopted, maintained, or experienced in deployment. We intr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents

DGX agent

arXiv:2604.23781v1 Announce Type: new Abstract: Language-model agents are increasingly used as persistent coworkers that assist users across multiple working days. During such workflows, the surroundi

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ClawTrace: Cost-Aware Tracing for LLM Agent Skill Distillation

DGX agent

arXiv:2604.23853v1 Announce Type: new Abstract: Skill-distillation pipelines learn reusable rules from LLM agent trajectories, but they lack a key signal: how much each step costs. Without per-step co

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Cooperative Informative Sensing for Monitoring Dynamic Indoor Environments via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.23179v1 Announce Type: cross Abstract: Monitoring human activity in indoor environments is important for applications such as facility management, safety assessment, and space utilization a

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems

DGX agent

arXiv:2604.24477v1 Announce Type: cross Abstract: The rapid integration of Large Language Models (LLMs) into Multi-Agent Systems (MAS) has significantly enhanced their collaborative problem-solving ca

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Scaling Coding Agents via Atomic Skills

DGX agent

arXiv:2604.05013v2 Announce Type: replace-cross Abstract: Current LLM coding agents are predominantly trained on composite benchmarks (e.g., bug fixing), which often leads to task-specific overfitting

researcharxiv-cs-ai
28 Apr 2026
Model Releases

Recent Advances in Multi-Agent Human Trajectory Prediction: A Comprehensive Review

DGX agent

arXiv:2506.14831v3 Announce Type: replace Abstract: With the emergence of powerful data-driven methods in human trajectory prediction (HTP), gaining a finer understanding of multi-agent interactions l

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

From Recall to Forgetting: Benchmarking Long-Term Memory for Personalized Agents

DGX agent

arXiv:2604.20006v1 Announce Type: new Abstract: Personalized agents that interact with users over long periods must maintain persistent memory across sessions and update it as circumstances change. Ho

model-releasesarxiv-cs-cl
23 Apr 2026
Safety

MOA: Multi-Objective Alignment for Role-Playing Agents

DGX agent

arXiv:2512.09756v2 Announce Type: replace Abstract: Role-playing agents (RPAs) require balancing multiple objectives, such as instruction following, persona consistency, and stylistic fidelity, which

safetyarxiv-cs-cl
23 Apr 2026
Agents

More Is Different: Toward a Theory of Emergence in AI-Native Software Ecosystems

DGX agent

arXiv:2604.19827v1 Announce Type: cross Abstract: Software engineering faces a fundamental challenge: multi-agent AI systems fail in ways that defy explanation by traditional theories. While individua

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

From Experience to Skill: Multi-Agent Generative Engine Optimization via Reusable Strategy Learning

DGX agent

arXiv:2604.19516v1 Announce Type: new Abstract: Generative engines (GEs) are reshaping information access by replacing ranked links with citation-grounded answers, yet current Generative Engine Optimi

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

MATA: Multi-Agent Framework for Reliable and Flexible Table Question Answering

DGX agent

arXiv:2602.09642v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLMs) have significantly improved table understanding tasks such as Table Question Answering (TableQ

agentsarxiv-cs-ai
22 Apr 2026
Agents

Remote Rowhammer Attack using Adversarial Observations on Federated Learning Clients

DGX agent

arXiv:2505.06335v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has the potential for simultaneous global learning amongst a large number of parallel agents, enabling emerging AI suc

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

StepFly: Agentic Troubleshooting Guide Automation for Incident Diagnosis

DGX agent

arXiv:2510.10074v2 Announce Type: replace Abstract: Effective incident management in large-scale IT systems relies on troubleshooting guides (TSGs), but their manual execution is slow and error-prone.

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Agentic Large Language Models for Training-Free Neuro-Radiological Image Analysis

DGX agent

arXiv:2604.16729v1 Announce Type: new Abstract: State-of-the-art large language models (LLMs) show high performance in general visual question answering. However, a fundamental limitation remains: cur

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

Bolzano: Case Studies in LLM-Assisted Mathematical Research

DGX agent

arXiv:2604.16989v1 Announce Type: new Abstract: We report new results on six problems in mathematics and theoretical computer science, produced with the assistance of Bolzano, an open-source multi-age

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

Do LLM-derived graph priors improve multi-agent coordination?

DGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

From Clinical Intent to Clinical Model: An Autonomous Coding-Agent Framework for Clinician-driven AI Development

DGX agent

arXiv:2604.17110v1 Announce Type: new Abstract: Clinical AI development has traditionally followed a collaborative paradigm that depends on close interaction between clinicians and specialized AI team

agentsarxiv-cs-cv
21 Apr 2026
Safety

IceBreaker for Conversational Agents: Breaking the First-Message Barrier with Personalized Starters

DGX agent

arXiv:2604.18375v1 Announce Type: new Abstract: Conversational agents, such as ChatGPT and Doubao, have become essential daily assistants for billions of users. To further enhance engagement, these sy

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

PersonalHomeBench: Evaluating Agents in Personalized Smart Homes

DGX agent

arXiv:2604.16813v1 Announce Type: cross Abstract: Agentic AI systems are rapidly advancing toward real-world applications, yet their readiness in complex and personalized environments remains insuffic

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

Satellite Chasers: Divergent Adversarial Reinforcement Learning to Engage Intelligent Adversaries on Orbit

DGX agent

arXiv:2409.17443v2 Announce Type: replace Abstract: As space becomes increasingly crowded and contested, robust autonomous capabilities for multi-agent environments are gaining critical importance. Cu

agentsarxiv-cs-ro
21 Apr 2026
Model Releases

Scaling External Knowledge Input Beyond Context Windows of LLMs via Multi-Agent Collaboration

DGX agent

arXiv:2505.21471v2 Announce Type: replace Abstract: With the rapid advancement of post-training techniques for reasoning and information seeking, large language models (LLMs) can incorporate a large q

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Scaling Test-Time Compute for Agentic Coding

DGX agent

arXiv:2604.16529v1 Announce Type: cross Abstract: Test-time scaling has become a powerful way to improve large language models. However, existing methods are best suited to short, bounded outputs that

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows

DGX agent

arXiv:2505.19897v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have extended their impact beyond Natural Language Processing, substantially fostering the development of interdi

model-releasesarxiv-cs-cl
21 Apr 2026
← Previous
1…6162636465…233
Next →