AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Multi-Agent DRL for QoS and Energy Optimization in RIS-Enabled Open-RAN Industrial 6G TN/NTN Networks

DGX agent

arXiv:2606.28339v1 Announce Type: cross Abstract: Industrial 6G networks require ultra-reliable, low-latency, and energy-efficient connectivity in dynamic and blockage-prone environments, where conven

agentsarxiv-cs-ai
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Recursive Self-Evolving Agents via Held-Out Selection

DGX agent

arXiv:2606.28374v1 Announce Type: new Abstract: LLM agents are increasingly improved without weight updates by evolving a natural-language artifact, such as reflections, workflows, playbooks, cheatshe

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

The Crowded Embedding Space: A Mean-Field Mechanism for Emergent Marginalization in Retrieval-Augmented Agents

DGX agent

arXiv:2606.28343v1 Announce Type: cross Abstract: Retrieval-augmented generative agents rely on retrieval for grounding, yet are typically evaluated on a query-by-query basis. This isolates interactio

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Towards Continual Motion-Language Agents: LoRA Variants for Incremental Motion Understanding and Generation

DGX agent

arXiv:2606.30266v1 Announce Type: cross Abstract: Motion-language agents must possess the bidirectional capability to both understand human movement (motion-to-text, M2T) and generate it from natural

model-releasesarxiv-cs-ai
30 Jun 2026
Local Ai

UCOB: Learning to Utilize and Evolve Agentic Skills via Credit-Aware On-Policy Bidirectional Self-Distillation

DGX agent

arXiv:2606.29502v1 Announce Type: new Abstract: Skill memories can improve agentic reinforcement learning by reusing past experience as textual guidance, but retrieved skills are not oracular: they ma

local-aiarxiv-cs-ai
30 Jun 2026
Agents

When LLMs Develop Languages: Symbolic Communication for Efficient Multi-Agent Reasoning

DGX agent

arXiv:2606.29354v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves large language models (LLMs) on difficult reasoning tasks, but it often incurs long natural-language rationales that are

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Ko-WideSearch: A Korean Breadth-Search Benchmark for Exhaustive Set Enumeration by Web Agents

DGX agent

arXiv:2606.27595v1 Announce Type: new Abstract: Web-agent benchmarks overwhelmingly measure depth -- pinning one obscure answer behind a chain of constraints -- while breadth, exhaustively enumerating

model-releasesarxiv-cs-cl
29 Jun 2026
Agents

SEA-TS: Self-Evolving Agent for Autonomous Code Generation of Time Series Forecasting Algorithms

DGX agent

arXiv:2603.04873v3 Announce Type: replace Abstract: Accurate time series forecasting underpins decision-making in many domains, yetconventional ML development often faces data scarcity, distribution s

agentsarxiv-cs-ai
29 Jun 2026
Model Releases

Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

DGX agent

arXiv:2606.26479v1 Announce Type: cross Abstract: Recent work (2024 to 2026) has converged on a strategy for defending tool-using LLM agents against indirect prompt injection: rather than training the

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

AgentX: Towards Agent-Driven Self-Iteration of Industrial Recommender Systems

DGX agent

arXiv:2606.26859v1 Announce Type: new Abstract: Recommendation algorithm iteration is moving from an artisanal, engineer-bound process toward an industrialized research loop, but this transition remai

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

auto-psych: Automating the science of mind using agent-driven theory discovery and experimentation

DGX agent

arXiv:2606.26460v1 Announce Type: new Abstract: AI-based scientific automation is increasingly possible by using agents to generate hypotheses, design experiments, and analyze data. Data collection is

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Governing Actions, Not Agents: Institutional Attestation as a Governance Model for Autonomous AI Systems

DGX agent

arXiv:2606.26298v1 Announce Type: new Abstract: Autonomous AI agents may begin to perform consequential, irreversible actions such as clinical prescribing and production software deployment. This pape

safetyarxiv-cs-ai
26 Jun 2026
Local Ai

Joint Learning of Experiential Rules and Policies for Large Language Model Agents

DGX agent

arXiv:2606.27136v1 Announce Type: new Abstract: For LLM agents in multi-step interactive environments, a key challenge is to make effective use of accumulated interaction experience. Existing work has

local-aiarxiv-cs-ai
26 Jun 2026
Safety

The Verification Horizon: No Silver Bullet for Coding Agent Rewards

DGX agent

arXiv:2606.26300v1 Announce Type: new Abstract: A classical intuition holds that verifying a solution is easier than producing one. For today's coding agents, this intuition is being inverted: as foun

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

DGX agent

arXiv:2606.25819v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that solve tasks by interacting with external tool environments. Although recent tool-use benc

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

Exploring Information Seeking Agent Consolidation

DGX agent

arXiv:2602.00585v2 Announce Type: replace Abstract: Information-seeking agents have emerged as a powerful paradigm for knowledge-intensive tasks, yet today's systems remain specialized for the open we

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

GUI agent: Guided Exploration of User-Sensitive Screens

DGX agent

arXiv:2606.25705v1 Announce Type: new Abstract: LLM agents are increasingly being used to automate tasks for users within an open GUI environment. They inevitably encounter screens containing user-sen

safetyarxiv-cs-ai
25 Jun 2026
Research

The Interplay of Harness Design and Post-Training in LLM Agents

DGX agent

arXiv:2606.25447v1 Announce Type: cross Abstract: Tool-integrated LLM agents are often wrapped within a harness: the scaffolding that determines which tools are exposed, how they are described, and wh

researcharxiv-cs-cl
25 Jun 2026
Safety

AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming

DGX agent

arXiv:2606.24245v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly automate complex tasks by integrating language models with external tools and environments. However, th

safetyarxiv-cs-ai
24 Jun 2026
Safety

Governed Shared Memory for Multi-Agent LLM Systems

DGX agent

arXiv:2606.24535v1 Announce Type: new Abstract: Multi-agent LLM environments require robust mechanisms for shared knowledge management. This paper formalizes the fleet-memory problem and identifies fo

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

How Much Can We Trust LLM Search Agents? Measuring Endorsement Vulnerability to Web Content Manipulation

DGX agent

arXiv:2606.16821v2 Announce Type: replace Abstract: Large language model (LLM)-based search agents synthesize open-web content into actionable recommendations on behalf of users, creating a risk that

model-releasesarxiv-cs-cl
24 Jun 2026
Agents

OmniPath: A Multi-Modal Agentic Framework for Auditing Wheelchair Accessibility

DGX agent

arXiv:2606.24129v1 Announce Type: new Abstract: For a wheelchair user, a standard blue line on a map is often a broken promise. While platforms like OpenStreetMap (OSM) successfully capture where a pa

agentsarxiv-cs-ai
24 Jun 2026
Safety

Reinforcement Learning for Computer-Use Agents with Autonomous Evaluation

DGX agent

arXiv:2606.24515v1 Announce Type: new Abstract: Computer-Use Agents (CUAs) execute high-level user goals by perceiving and acting directly within graphical user interfaces. However, reinforcement lear

safetyarxiv-cs-ai
24 Jun 2026
Local Ai

Subjective-Graph LLM Agents for Simulating Uncertainty in Classroom Social Perception

DGX agent

arXiv:2603.20750v2 Announce Type: replace Abstract: Social actors do not observe a common social world: each individual forms judgments from a partial and potentially distorted view of the surrounding

local-aiarxiv-cs-ai
24 Jun 2026
Agents

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent

DGX agent

arXiv:2606.24094v1 Announce Type: new Abstract: Unifying image clustering across different clustering scenarios remains challenging due to fundamental gaps among tasks. We introduce a Guideline-Driven

agentsarxiv-cs-cv
24 Jun 2026
Model Releases

VisCritic: Visual State Comparison as Process Reward for GUI Agents

DGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

model-releasesarxiv-cs-cv
24 Jun 2026
Model Releases

CFAgentBench: A Reproducible Environment and Benchmark for Autonomous Construction-Finance Agents

DGX agent

arXiv:2606.22000v1 Announce Type: cross Abstract: We introduce CFAgentBench, a reproducible, self-hostable environment and benchmark for autonomous construction-finance agents: a CFO/controller-class

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

Democratizing and accelerating AI-driven pathology research through agentic intelligence

DGX agent

arXiv:2606.20677v1 Announce Type: cross Abstract: Computational pathology has advanced rapidly with the emergence of foundation models, yet widespread adoption remains limited by substantial technical

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

From Knowing to Acting: Benchmarking Self-Awareness Capability of LLM Agents

DGX agent

arXiv:2606.20661v1 Announce Type: cross Abstract: The integration of external tools has transitioned LLM agents from passive responders to autonomous systems. However, current benchmarks prioritize ex

model-releasesarxiv-cs-lg
23 Jun 2026
Local Ai

MIRAGE: Stealthy Visual Prompt Injection for Vulnerability Detection in Web Agents

DGX agent

arXiv:2606.20717v1 Announce Type: new Abstract: Multimodal Large Language Model (MLLM)-based web agents provide practical, high-precision solutions for visual browser automation; however, they inheren

local-aiarxiv-cs-cv
23 Jun 2026
Safety

Spark: Strategic Policy-Aware Exploration via Dynamic Branching for Long-Horizon Agentic Learning

DGX agent

arXiv:2601.20209v2 Announce Type: replace Abstract: Reinforcement learning has empowered large language models to act as intelligent agents, yet training them for long-horizon tasks remains challengin

safetyarxiv-cs-lg
23 Jun 2026
Agents

Tell Me: An LLM-powered Mental Well-being Assistant with RAG, Synthetic Dialogue Generation, and Agentic Planning

DGX agent

arXiv:2511.14445v2 Announce Type: replace-cross Abstract: We present Tell Me, a mental well-being system that leverages advances in large language models to provide accessible, context-aware support f

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Grounding Computer Use Agents on Human Demonstrations

DGX agent

arXiv:2511.07332v2 Announce Type: replace-cross Abstract: Building reliable computer-use agents requires grounding: accurately connecting natural language instructions to the correct on-screen element

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios

DGX agent

arXiv:2602.22638v2 Announce Type: replace Abstract: Route-planning agents powered by large language models (LLMs) have emerged as a promising paradigm for supporting everyday human mobility through na

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

SIL: Symbiotic Interactive Learning for Language-Conditioned Human-Agent Co-Adaptation

DGX agent

arXiv:2511.05203v3 Announce Type: replace Abstract: Today's autonomous agents, largely driven by foundation models (FMs), can understand natural language instructions and solve long-horizon tasks with

safetyarxiv-cs-ro
11 Jun 2026
Agents

Dmsh: A Multi-Agent Reinforcement Learning Framework for All-Quad Mesh Generation

DGX agent

arXiv:2606.10601v1 Announce Type: cross Abstract: Generating high-quality meshes for arbitrary geometries remains a fundamental bottleneck in computational engineering, often demanding heuristic tunin

agentsarxiv-cs-ai
10 Jun 2026
Safety

Early-Token Confidence Predicts Reasoning Quality in Multi-Agent LLM Debate

DGX agent

arXiv:2606.10307v1 Announce Type: new Abstract: Evaluating reasoning quality in multi-agent LLM systems is challenging, especially for open-ended tasks without reference answers. We investigate whethe

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Fact-Augmented Lookahead Planning for LLM Agents

DGX agent

arXiv:2506.09171v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly capable, but LLM agents still struggle to plan effectively in interactive, partially observable,

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Failure Modes of Deep Multi-Agent RL in Asynchronous Pricing: Reproducible Triggers, Trace Diagnostics, and a Partial Fix

DGX agent

arXiv:2606.09884v1 Announce Type: cross Abstract: We study two reproducible failure modes of deep multi-agent reinforcement learning in continuous-time pricing markets: (i) tacit cartel formation betw

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

GUI-AC: Enhancing Continual Learning in GUI Agents

DGX agent

arXiv:2606.10522v1 Announce Type: new Abstract: Graphical User Interfaces (GUIs) serve as the dominant medium for human-computer interaction, yet building GUI agents that generalize across the vast di

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

MIRAGE: A Polarity-Flipping Encoding Subspace in LLM Agents

DGX agent

arXiv:2606.10304v1 Announce Type: new Abstract: When LLM agents are coerced into covertly encoding sensitive data (Base64, ROT13, acrostic, synonym chains, and beyond), the resulting outputs evade out

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

The Interlocutor Effect: Why LLMs Leak More Personal Data to Agents Than Humans

DGX agent

arXiv:2606.09844v1 Announce Type: cross Abstract: Large Language Models (LLMs) alter their privacy behavior based on the perceived identity of their interlocutor. While safety mechanisms typically pre

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

TRACE: A Unified Rollout Budget Allocation Framework for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.11119v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a promising approach for enhancing reasoning and agentic behavior in large language models. H

safetyarxiv-cs-ai
10 Jun 2026
Safety

What Should a Skill Remember? Quality--Cost Trade-offs in Cost-Aware Skill Rewriting for Language Model Agents

DGX agent

arXiv:2606.09421v2 Announce Type: replace Abstract: Large language model agents increasingly rely on skills: reusable procedural documents encoding workflows, tool use, implementation patterns, valida

safetyarxiv-cs-cl
10 Jun 2026
Agents

Agentic Neuro-Symbolic Planning and Commissioning for Human-in-the-Loop Industrial Robotics with Digital Twins

DGX agent

arXiv:2606.08214v1 Announce Type: new Abstract: Flexible robotic automation requires systems that interpret operator intent, verify physical feasibility, and recover from execution failures across bot

agentsarxiv-cs-ro
9 Jun 2026
Model Releases

Bidirectional Semantic Complementary Tool Retrieval for Remote Sensing Agents

DGX agent

arXiv:2606.07538v1 Announce Type: cross Abstract: Large language model (LLM)-based agents provide a novel paradigm for the automated processing of remote sensing(RS) data. Their success in complex RS

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Decision-Aware Memory Cards: Counterfactual-Inspired Context Selection and Compression for Tool-Using LLM Agents

DGX agent

arXiv:2606.08151v1 Announce Type: new Abstract: Tool-using LLM agents often fail not because relevant text is absent, but because decisive evidence is not selected, compressed, or surfaced at action t

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

DICE: Entropy-Regularized Equilibrium Selection for Stable Multi-Agent LLM Coordination

DGX agent

arXiv:2606.08068v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems often fail to reliably outperform a single strong model equipped with best-of-N sampling. We argue that a

model-releasesarxiv-cs-lg
9 Jun 2026
← Previous
1…7778798081…236
Next →