AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Joint Agent Memory and Exploration Learning via Novelty Signals

DGX agent

arXiv:2606.01528v1 Announce Type: new Abstract: In open-ended environments, exploration is fundamental for autonomous agents, yet current language model agents struggle with this. Effective exploratio

safetyarxiv-cs-ai
2 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation

DGX agent

arXiv:2606.00610v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become an essential method for mitigating hallucinations in Large Language Models (LLMs) by leveraging extern

agentsarxiv-cs-ai
2 Jun 2026
Agents

Multi-Agent Conformal Prediction with Personalized Statistical Validity

DGX agent

arXiv:2606.00717v1 Announce Type: cross Abstract: Uncertainty quantification is essential in high-stakes machine learning tasks. However, one of the principled solutions, conformal prediction, faces c

agentsarxiv-cs-ai
2 Jun 2026
Agents

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

DGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

agentsarxiv-cs-ai
2 Jun 2026
Agents

RocketSmith: An Agentic System for High-Powered Rocket Design and Manufacturing

DGX agent

arXiv:2606.00097v1 Announce Type: new Abstract: This work presents RocketSmith, an agentic system capable of the design, manufacturing, and optimization processes in high powered rocket development. T

agentsarxiv-cs-ro
2 Jun 2026
Agents

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents

DGX agent

arXiv:2602.12984v2 Announce Type: replace Abstract: Scientific reasoning inherently demands integrating sophisticated toolkits to navigate domain-specific knowledge. Yet, current benchmarks largely ov

agentsarxiv-cs-cl
2 Jun 2026
Agents

SkillSmith: Co-Evolving Skills and Tools for Self-Improving Agent Systems

DGX agent

arXiv:2606.01314v1 Announce Type: new Abstract: Recent self-evolving agents have shown that skills can be discovered, refined, and accumulated through execution. However, existing skill-evolution fram

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

When Safe Skills Collide: Measuring Compositional Risk in Agent Skill Ecosystems

DGX agent

arXiv:2606.00448v1 Announce Type: cross Abstract: LLM agents increasingly rely on community-contributed skills that expand an agent's operational capability set. We study a core safety problem in agen

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

AutoSci: A Memory-Centric Agentic System for the Full Scientific Research Lifecycle

DGX agent

arXiv:2605.31468v1 Announce Type: new Abstract: Scientific research has traditionally been human-intensive, requiring researchers to coordinate literature, ideas, experiments, manuscripts, and review

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

EMBGuard: Constructing Hazard-Aware Guardrails for Safe Planning in Embodied Agents

DGX agent

arXiv:2605.30924v1 Announce Type: new Abstract: MLLM-powered embodied agents deployed in real-world environments encounter physical hazards. However, existing approaches lack explicit mechanisms for i

model-releasesarxiv-cs-cl
1 Jun 2026
Model Releases

Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion

DGX agent

arXiv:2605.31170v1 Announce Type: cross Abstract: Monitoring autonomous language model agents currently relies mostly on surface behavior. But what happens when agent populations invent new languages

model-releasesarxiv-cs-ai
1 Jun 2026
Research

Enhancing Human-Likeness in Reinforcement Learning Agents via Hierarchical Macro Action Quantization

DGX agent

arXiv:2605.30928v1 Announce Type: new Abstract: Human-like agents are a long-standing goal of artificial intelligence. Despite strong performance, most reinforcement learning (RL) agents remain reward

researcharxiv-cs-ro
1 Jun 2026
Model Releases

Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship

DGX agent

arXiv:2605.30947v1 Announce Type: new Abstract: LLM-based research agents have advanced rapidly in science and engineering, where research is organized around executable experiments, code, and quantit

model-releasesarxiv-cs-cl
1 Jun 2026
Agents

NEMO: Execution-Aware Optimization Modeling via Autonomous Coding Agents

DGX agent

arXiv:2601.21372v2 Announce Type: replace Abstract: We present NEMO, a system that translates Natural-language descriptions of decision problems into formal Executable Mathematical Optimization implem

agentsarxiv-cs-ai
1 Jun 2026
Hardware

PithTrain: A Compact and Agent-Native MoE Training System

DGX agent

arXiv:2605.31463v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become the dominant architecture for frontier language models. To meet this demand, production frameworks have built opti

hardwarearxiv-cs-ai
1 Jun 2026
Safety

Seeing Before Agreeing: Aligning Multi-Agent Consensus with Visual Evidence

DGX agent

arXiv:2605.30698v1 Announce Type: cross Abstract: Vision-language models (VLMs) have achieved strong performance on visual question answering (VQA). To mitigate individual hallucinations and blind spo

safetyarxiv-cs-ai
1 Jun 2026
Agents

Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Education Systems

DGX agent

arXiv:2501.10332v2 Announce Type: replace-cross Abstract: Personalized learning represents a promising educational strategy within intelligent educational systems, aiming to enhance learners' practice

agentsarxiv-cs-ai
29 May 2026
Model Releases

AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-or-Reject Pruning

DGX agent

arXiv:2602.23258v2 Announce Type: replace Abstract: While Multi-Agent Systems (MAS) excel in complex reasoning, they suffer from the cascading impact of erroneous information from individual agents. C

model-releasesarxiv-cs-ai
29 May 2026
Agents

AnomalyAgent: Training-Free Agentic Models for Zero-/Few-Shot Anomaly Detection

DGX agent

arXiv:2605.30140v1 Announce Type: new Abstract: Benefiting from generalizability of vision-language models (VLMs) such as CLIP, many zero-/few-shot anomaly detection (AD) approaches have achieved impr

agentsarxiv-cs-cv
29 May 2026
Agents

Compass: Navigating Global Marine Lead Data Integration through Expert-Guided LLM Agent

DGX agent

arXiv:2605.29966v1 Announce Type: new Abstract: Marine lead (Pb) and its isotopes are critical tracers for ocean circulation and anthropogenic pollution, yet in-situ observations remain costly and spa

agentsarxiv-cs-ai
29 May 2026
Safety

Graph-Enhanced Policy Optimization in LLM Agent Training

DGX agent

arXiv:2510.26270v2 Announce Type: replace Abstract: Multi-step LLM agents in interactive environments represent a crucial step toward long-horizon decision-making. To train such agents, group-based re

safetyarxiv-cs-ai
29 May 2026
Research

Honest Lying: Understanding Memory Confabulation in Reflexive Agents

DGX agent

arXiv:2605.29463v1 Announce Type: cross Abstract: Reflexion-style agents rely on self-generated reflections as memory, implicitly assuming that agents can accurately diagnose their own failures.We sho

researcharxiv-cs-ai
29 May 2026
Safety

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

DGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

safetyarxiv-cs-ai
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Model Releases

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

DGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code

DGX agent

arXiv:2603.24631v2 Announce Type: replace-cross Abstract: Code agents resolve 65-70% of SWE-bench Verified issues, but Pass@1 cannot tell us why the rest fail, and, as we show, capable-model failures

model-releasesarxiv-cs-ai
28 May 2026
Agents

Discovery Agents for Real-Time Analytics: Toward Proactive Insight Systems

DGX agent

arXiv:2605.27571v1 Announce Type: new Abstract: Modern analytics systems are fundamentally reactive, requiring users to define queries over increasingly complex and continuously evolving data. In real

agentsarxiv-cs-ai
28 May 2026
Agents

Hierarchical Prompt-Domain Control and Learning for Resource-Constrained Agentic Language Models

DGX agent

arXiv:2605.27703v1 Announce Type: new Abstract: Large Language Models are increasingly deployed inside agentic systems, where they must follow structured protocols, adapt to evolving states, and opera

agentsarxiv-cs-ai
28 May 2026
Agents

Roles with Rails: Contract-Preserving Role Evolution in Multi-Agent Structured Reasoning

DGX agent

arXiv:2605.28433v1 Announce Type: new Abstract: Role-based LLM multi-agent systems need adaptive role pools, yet adapting such systems is not merely a matter of prompt optimization: roles often carry

agentsarxiv-cs-cl
28 May 2026
Safety

Structured Agent Distillation for Large Language Model

DGX agent

arXiv:2505.13820v5 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong capabilities as decision-making agents by interleaving reasoning and actions, as seen in ReAct-sty

safetyarxiv-cs-ai
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

BeyondSWE: Can Current Code Agent Survive Beyond Single-Repo Bug Fixing?

DGX agent

arXiv:2603.03194v2 Announce Type: replace Abstract: Current code-agent benchmarks primarily evaluate localized issue resolution within a single target repository, leaving under-tested many software en

model-releasesarxiv-cs-cl
27 May 2026
Agents

Communication Gain and Delay Cost Under Cross-Timestep Delays in Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.03785v2 Announce Type: replace Abstract: Communication is essential for coordination in cooperative multi-agent reinforcement learning under partial observability, yet cross-timestep delays

agentsarxiv-cs-ai
27 May 2026
Safety

Counterfactual Credit Policy Optimization for Multi-Agent Collaboration

DGX agent

arXiv:2603.21563v2 Announce Type: replace Abstract: Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such s

safetyarxiv-cs-ai
27 May 2026
Agents

Interactive Agents: Simulating Counselor-Client Psychological Counseling via Role-Playing LLM-to-LLM Interactions

DGX agent

arXiv:2408.15787v2 Announce Type: replace Abstract: Creating effective dialogue systems for mental health support requires high-quality multi-turn counseling dialogue data, yet collecting real counsel

agentsarxiv-cs-cl
27 May 2026
Model Releases

JobBench: Aligning Agent Work With Human Will

DGX agent

arXiv:2605.26329v1 Announce Type: new Abstract: Current benchmarks for occupational AI agents are scoped primarily by economic values, telling a replacement story. We introduce JobBench, which evaluat

model-releasesarxiv-cs-ai
27 May 2026
Safety

RICE-PO: Turning Retrieval Interactions into Credit Signals for Reasoning Agents

DGX agent

arXiv:2605.26352v1 Announce Type: new Abstract: Retrieval is increasingly moving from one-shot matching toward interactive reasoning, where language agents iteratively inspect evidence, reformulate qu

safetyarxiv-cs-cl
27 May 2026
Agents

SPEAR: Code-Augmented Agentic Prompt Optimization

DGX agent

arXiv:2605.26275v1 Announce Type: new Abstract: Automatic prompt engineering (APE) rewrites prompts to improve downstream task performance, but existing APE loops treat the optimizer itself as a fixed

agentsarxiv-cs-cl
27 May 2026
Model Releases

SWE-Adept: An LLM-Based Agentic Framework for Deep Codebase Analysis and Structured Issue Resolution

DGX agent

arXiv:2603.01327v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong performance on self-contained programming tasks. However, they still struggle with repository-leve

model-releasesarxiv-cs-cl
27 May 2026
Safety

Think Twice Before You Act: Enhancing Agent Behavioral Safety with Thought Correction

DGX agent

arXiv:2505.11063v3 Announce Type: replace Abstract: LLM-based agents solve complex tasks through iterative reasoning, tool use, and environment interaction, where each intermediate thought directly sh

safetyarxiv-cs-ai
27 May 2026
Model Releases

Verus-SpecGym: An Agentic Environment for Evaluating Specification Autoformalization

DGX agent

arXiv:2605.26457v1 Announce Type: cross Abstract: AI coding agents are increasingly used to write real-world software, but ensuring that their outputs are correct remains a fundamental challenge. Form

model-releasesarxiv-cs-ai
27 May 2026
Agents

Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models

DGX agent

arXiv:2605.24755v1 Announce Type: new Abstract: Speech monologues recorded in naturalistic settings provide opportunities to characterize mental illness phenomenology and detect symptom exacerbation.

agentsarxiv-cs-ai
26 May 2026
Agents

Inference-Time Backdoors via Chat Templates: From LLM Supply Chains to Agentic System Compromise

DGX agent

arXiv:2602.04653v4 Announce Type: replace-cross Abstract: Open-weight language models are increasingly used in production settings, raising new security challenges. One prominent threat is backdoor at

agentsarxiv-cs-lg
26 May 2026
Agents

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

DGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

agentsarxiv-cs-ai
26 May 2026
Agents

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

DGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

agentsarxiv-cs-ai
26 May 2026
Agents

Why Agentic Theorem Prover Works: A Statistical Provability Theory of Mathematical Reasoning Models

DGX agent

arXiv:2602.10538v3 Announce Type: replace-cross Abstract: Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components

agentsarxiv-cs-lg
26 May 2026
Agents

Why Your Deep Research Agent Fails? On Hallucination Evaluation in Full Research Trajectory

DGX agent

arXiv:2601.22984v2 Announce Type: replace Abstract: Diagnosing failure patterns in Deep Research Agents (DRAs) remains a critical challenge. Existing benchmarks predominantly rely on end-to-end evalua

agentsarxiv-cs-ai
26 May 2026
Agents

A Proactive Multi-Agent Dialogue Framework for Assessing Social Language Disorder Traits in Autism

DGX agent

arXiv:2605.22993v1 Announce Type: cross Abstract: Characteristic linguistic behaviors associated with Social Language Disorder (SLD) in autism spectrum disorder, including echoic repetition, pronoun d

agentsarxiv-cs-ai
25 May 2026
← Previous
1…5859606162…233
Next →