AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

GAMED.AI: A Hierarchical Multi-Agent Framework for Automated Educational Game Generation

DGX agent

arXiv:2604.23947v1 Announce Type: new Abstract: We introduce GameDAI, a hierarchical multi-agent framework that transforms instructor-provided questions into fully playable, pedagogically grounded edu

safetyarxiv-cs-ai
28 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AgentBound: Securing Execution Boundaries of AI Agents

DGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

safetyarxiv-cs-ai
27 Apr 2026
Safety

AgentGL: Towards Agentic Graph Learning with LLMs via Reinforcement Learning

DGX agent

arXiv:2604.05846v2 Announce Type: replace Abstract: Large Language Models (LLMs) increasingly rely on agentic capabilities-iterative retrieval, tool use, and decision-making-to overcome the limits of

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

When Agents Look the Same: Quantifying Distillation-Induced Similarity in Tool-Use Behaviors

DGX agent

arXiv:2604.21255v1 Announce Type: new Abstract: Model distillation is a primary driver behind the rapid progress of LLM agents, yet it often leads to behavioral homogenization. Many emerging agents sh

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

Why Do Language Model Agents Whistleblow?

DGX agent

arXiv:2511.17085v3 Announce Type: replace-cross Abstract: The deployment of Large Language Models (LLMs) as tool-using agents causes their alignment training to manifest in new ways. Recent work finds

safetyarxiv-cs-ai
24 Apr 2026
Safety

Ask Only When Needed: Proactive Retrieval from Memory and Skills for Experience-Driven Lifelong Agents

DGX agent

arXiv:2604.20572v1 Announce Type: new Abstract: Online lifelong learning enables agents to accumulate experience across interactions and continually improve on long-horizon tasks. However, existing me

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

Learning to Evolve: A Self-Improving Framework for Multi-Agent Systems via Textual Parameter Graph Optimization

DGX agent

arXiv:2604.20714v1 Announce Type: new Abstract: Designing and optimizing multi-agent systems (MAS) is a complex, labor-intensive process of 'Agent Engineering.' Existing automatic optimization methods

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

Shift-Up: A Framework for Software Engineering Guardrails in AI-native Software Development -- Initial Findings

DGX agent

arXiv:2604.20436v1 Announce Type: cross Abstract: Generative AI (GenAI) is reshaping software engineering by shifting development from manual coding toward agent-driven implementation. While vibe codi

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Soft-Label Governance for Distributional Safety in Multi-Agent Systems

DGX agent

arXiv:2604.19752v1 Announce Type: cross Abstract: Multi-agent AI systems exhibit emergent risks that no single agent produces in isolation. Existing safety frameworks rely on binary classifications of

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Integrating Anomaly Detection into Agentic AI for Proactive Risk Management in Human Activity

DGX agent

arXiv:2604.19538v1 Announce Type: new Abstract: Agentic AI, with goal-directed, proactive, and autonomous decision-making capabilities, offers a compelling opportunity to address movement-related risk

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Mango: Multi-Agent Web Navigation via Global-View Optimization

DGX agent

arXiv:2604.18779v1 Announce Type: new Abstract: Existing web agents typically initiate exploration from the root URL, which is inefficient for complex websites with deep hierarchical structures. Witho

model-releasesarxiv-cs-cl
22 Apr 2026
Agents

Revac: A Social Deduction Reasoning Agent

DGX agent

arXiv:2604.19523v1 Announce Type: new Abstract: Social deduction games such as Mafia present a unique AI challenge: players must reason under uncertainty, interpret incomplete and intentionally mislea

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

Agentic Risk-Aware Set-Based Engineering Design

DGX agent

arXiv:2604.16687v1 Announce Type: cross Abstract: This paper introduces a multi-agent framework guided by Large Language Models (LLMs) to assist in the early stages of engineering design, a phase ofte

model-releasesarxiv-cs-lg
21 Apr 2026
Agents

ARMove: Learning to Predict Human Mobility through Agentic Reasoning

DGX agent

arXiv:2604.17419v1 Announce Type: cross Abstract: Human mobility prediction is a critical task but remains challenging due to its complexity and variability across populations and regions. Recently, l

agentsarxiv-cs-lg
21 Apr 2026
Agents

Is Agentic RAG worth it? An experimental comparison of RAG approaches

DGX agent

arXiv:2601.07711v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems are usually defined by the combination of a generator and a retrieval component that extracts textual c

agentsarxiv-cs-cl
21 Apr 2026
Safety

Multimodal Policy Internalization for Conversational Agents

DGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

safetyarxiv-cs-cl
21 Apr 2026
Agents

Unified Ultrasound Intelligence Toward an End-to-End Agentic System

DGX agent

arXiv:2604.16914v1 Announce Type: new Abstract: Clinical ultrasound analysis demands models that generalize across heterogeneous organs, views, and devices, while supporting interpretable workflow-lev

agentsarxiv-cs-cv
21 Apr 2026
Agents

DeepER-Med: Advancing Deep Evidence-Based Research in Medicine Through Agentic AI

DGX agent

arXiv:2604.15456v1 Announce Type: new Abstract: Trustworthiness and transparency are essential for the clinical adoption of artificial intelligence (AI) in healthcare and biomedical research. Recent d

agentsarxiv-cs-ai
20 Apr 2026
Agents

Integrating Graphs, Large Language Models, and Agents: Reasoning and Retrieval

DGX agent

arXiv:2604.15951v1 Announce Type: new Abstract: Generative AI, particularly Large Language Models, increasingly integrates graph-based representations to enhance reasoning, retrieval, and structured d

agentsarxiv-cs-ai
20 Apr 2026
Agents

SocialWise: LLM-Agentic Conversation Therapy for Individuals with Autism Spectrum Disorder to Enhance Communication Skills

DGX agent

arXiv:2604.15347v1 Announce Type: cross Abstract: Autism Spectrum Disorder (ASD) affects more than 75 million people worldwide. However, scalable support for practicing everyday conversation is scarce

agentsarxiv-cs-ai
20 Apr 2026
Safety

The Price of Paranoia: Robust Risk-Sensitive Cooperation in Non-Stationary Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.15695v1 Announce Type: cross Abstract: Cooperative equilibria are fragile. When agents learn alongside each other rather than in a fixed environment, the process of learning destabilizes th

safetyarxiv-cs-ai
20 Apr 2026
Safety

CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas

DGX agent

arXiv:2604.15267v1 Announce Type: cross Abstract: It is increasingly important that LLM agents interact effectively and safely with other goal-pursuing agents, yet, recent works report the opposite tr

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games

DGX agent

arXiv:2506.03610v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are reshaping the game industry, by enabling more intelligent and human-preferable characters. Yet, current game b

model-releasesarxiv-cs-ai
17 Apr 2026
Agents

Agentic LLM Reasoning in a Self-Driving Laboratory for Air-Sensitive Lithium Halide Spinel Conductors

DGX agent

arXiv:2604.11957v1 Announce Type: cross Abstract: Self-driving laboratories promise to accelerate materials discovery. Yet current automated solid-state synthesis platforms are limited to ambient cond

agentsarxiv-cs-lg
15 Apr 2026
Agents

CascadeDebate: Multi-Agent Deliberation for Cost-Aware LLM Cascades

DGX agent

arXiv:2604.12262v1 Announce Type: cross Abstract: Cascaded LLM systems coordinate models of varying sizes with human experts to balance accuracy, cost, and abstention under uncertainty. However, singl

agentsarxiv-cs-ai
15 Apr 2026
Safety

CIA: Inferring the Communication Topology from LLM-based Multi-Agent Systems

DGX agent

arXiv:2604.12461v1 Announce Type: new Abstract: LLM-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in solving complex tasks. Central to MAS is the communication topology whi

safetyarxiv-cs-ai
15 Apr 2026
Agents

Development, Evaluation, and Deployment of a Multi-Agent System for Thoracic Tumor Board

DGX agent

arXiv:2604.12161v1 Announce Type: new Abstract: Tumor boards are multidisciplinary conferences dedicated to producing actionable patient care recommendations with live review of primary radiology and

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

CocoaBench: Evaluating Unified Digital Agents in the Wild

DGX agent

arXiv:2604.11201v1 Announce Type: cross Abstract: LLM agents now perform strongly in software engineering, deep research, GUI automation, and various other applications, while recent agent scaffolds a

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Data Mixing Agent: Learning to Re-weight Domains for Continual Pre-training

DGX agent

arXiv:2507.15640v2 Announce Type: replace-cross Abstract: Continual pre-training on small-scale task-specific data is an effective method for improving large language models in new target fields, yet

agentsarxiv-cs-ai
14 Apr 2026
Safety

DeepFleet: Multi-Agent Foundation Models for Mobile Robots

DGX agent

arXiv:2508.08574v3 Announce Type: replace Abstract: We introduce DeepFleet, a suite of foundation models designed to support coordination and planning for large-scale mobile robot fleets. These models

safetyarxiv-cs-ro
14 Apr 2026
Agents

From Query to Counsel: Structured Reasoning with a Multi-Agent Framework and Dataset for Legal Consultation

DGX agent

arXiv:2604.10470v1 Announce Type: cross Abstract: Legal consultation question answering (Legal CQA) presents unique challenges compared to traditional legal QA tasks, including the scarcity of high-qu

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

HiPRAG: Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation

DGX agent

arXiv:2510.07794v2 Announce Type: replace-cross Abstract: Agentic RAG is a powerful technique for incorporating external information that LLMs lack, enabling better problem solving and question answer

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

LOLGORITHM: Funny Comment Generation Agent For Short Videos

DGX agent

arXiv:2604.09729v1 Announce Type: cross Abstract: Short-form video platforms have become central to multimedia information dissemination, where comments play a critical role in driving engagement, pro

agentsarxiv-cs-ai
14 Apr 2026
Safety

OOM-RL: Out-of-Money Reinforcement Learning Market-Driven Alignment for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2604.11477v1 Announce Type: new Abstract: The alignment of Multi-Agent Systems (MAS) for autonomous software engineering is constrained by evaluator epistemic uncertainty. Current paradigms, suc

safetyarxiv-cs-ai
14 Apr 2026
Safety

The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents

DGX agent

arXiv:2601.11496v2 Announce Type: replace-cross Abstract: The integration of AI agents into economic markets fundamentally alters the landscape of strategic interaction. We investigate the economic im

safetyarxiv-cs-ai
14 Apr 2026
Safety

YIELD: A Large-Scale Dataset and Evaluation Framework for Information Elicitation Agents

DGX agent

arXiv:2604.10968v1 Announce Type: new Abstract: Most conversational agents (CAs) are designed to satisfy user needs through user-driven interactions. However, many real-world settings, such as academi

safetyarxiv-cs-cl
14 Apr 2026
Safety

Artifacts as Memory Beyond the Agent Boundary

DGX agent

arXiv:2604.08756v1 Announce Type: new Abstract: The situated view of cognition holds that intelligent behavior depends not only on internal memory, but on an agent's active use of environmental resour

safetyarxiv-cs-ai
13 Apr 2026
Agents

Automated Standardization of Legacy Biomedical Metadata Using an Ontology-Constrained LLM Agent

DGX agent

arXiv:2604.08552v1 Announce Type: cross Abstract: Scientific metadata are often incomplete and noncompliant with community standards, limiting dataset findability, interoperability, and reuse. When re

agentsarxiv-cs-ai
13 Apr 2026
Agents

Multi-User Large Language Model Agents

DGX agent

arXiv:2604.08567v1 Announce Type: new Abstract: Large language models (LLMs) and LLM-based agents are increasingly deployed as assistants in planning and decision making, yet most existing systems are

agentsarxiv-cs-cl
13 Apr 2026
Model Releases

ClawsBench: Evaluating Capability and Safety of LLM Productivity Agents in Simulated Workspaces

DGX agent

arXiv:2604.05172v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly deployed to automate productivity tasks (e.g., email, scheduling, document management), but evalu

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

DGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems

DGX agent

arXiv:2604.06811v1 Announce Type: cross Abstract: Skill-based agent systems tackle complex tasks by composing reusable skills, improving modularity and scalability while introducing a largely unexamin

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Causal Agent based on Large Language Model

DGX agent

arXiv:2408.06849v3 Announce Type: replace Abstract: The large language model (LLM) has achieved significant success across various domains. However, the inherent complexity of causal problems and caus

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

ExRole: From Team Trajectories to Executable Roles in Multi-Agent Language Models

DGX agent

arXiv:2608.11949v1 Announce Type: new Abstract: Roles provide an interpretable interface for organizing language-model agents, yet most multi-agent systems treat them as hand-written prompt labels dis

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Harness-IF: Evaluating Instruction Following Across Instruction Surfaces in Coding Agents

DGX agent

arXiv:2608.11727v1 Announce Type: new Abstract: When a coding agent obeys a rule, it may simply have been going to do that anyway. Existing instruction-following benchmarks cannot tell the difference:

agentsarxiv-cs-ai
13 Aug 2026
Agents

Agentic Instruction Data Selection: Let DataMaster Interpret Your Intent

DGX agent

arXiv:2608.10579v1 Announce Type: new Abstract: Although existing instruction data selection methods have introduced various metrics, the inherent complexity of real-world datasets makes it impractica

agentsarxiv-cs-ai
12 Aug 2026
Safety

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

DGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

safetyarxiv-cs-lg
12 Aug 2026
Model Releases

Navigation Alone Is Not Enough: Evaluating Explanatory Assistive UI Agents

DGX agent

arXiv:2608.09944v1 Announce Type: cross Abstract: Modern web interfaces are increasingly difficult to use with screen readers, particularly when pages update dynamically or hide important structure be

model-releasesarxiv-cs-ai
12 Aug 2026
← Previous
1…3536373839…233
Next →