AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

GPTNT: Benchmarking Real-Time Collaboration Between Multimodal Agents on Keep Talking And Nobody Explodes

DGX agent

arXiv:2606.28514v1 Announce Type: new Abstract: Multimodal models are increasingly deployed to solve tasks collaboratively with humans or other artificial agents. Existing benchmarks show that these m

model-releasesarxiv-cs-ai
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents

DGX agent

arXiv:2606.29459v1 Announce Type: cross Abstract: Inverse design of metal-organic frameworks (MOFs) requires searching a combinatorially vast space where property labels are expensive and most machine

agentsarxiv-cs-ai
30 Jun 2026
Agents

KbSD: Knowledge Boundary aware Self-Distillation for Behavioral Calibration in Agentic Search

DGX agent

arXiv:2606.29863v1 Announce Type: new Abstract: Agentic search equips large language models with dynamic retrieval abilities, but existing reinforcement learning methods remain limited by reward spars

agentsarxiv-cs-cl
30 Jun 2026
Model Releases

MESA: Prioritizing Vulnerable Communication Channels for Securing Multi-Agent Systems

DGX agent

arXiv:2606.30602v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly used to automate complex, distributed workflows. However, their inter-agent communication channels introduc

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

ReasonRec: A Reasoning-Augmented Multimodal Agent for Unified Recommendation

DGX agent

arXiv:2606.28357v1 Announce Type: cross Abstract: Recent advances in multimodal recommenders excel at feature fusion but remain opaque and inefficient decision-makers, lacking explicit reasoning and s

agentsarxiv-cs-ai
30 Jun 2026
Agents

SafeGEO: Understanding Generative Engine Optimization Risks in Recommendation Agents

DGX agent

arXiv:2606.28356v1 Announce Type: cross Abstract: Generative Engine Optimization (GEO) lets content owners rewrite web content to increase their visibility in generative systems. In recommendation age

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

SEVA: Self-Evolving Verification Agent with Process Reward for Fact Attribution

DGX agent

arXiv:2606.29713v1 Announce Type: cross Abstract: Hallucination is the reliability bottleneck for LLM-based agents, and fact attribution verifiers are the last line of defense -- yet today's verifiers

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

TUA-Bench: A Benchmark for General-Purpose Terminal-Use Agents

DGX agent

arXiv:2606.28480v1 Announce Type: cross Abstract: As large language models and harness frameworks continue to advance, agents operating in terminals are increasingly capable of performing a broader ra

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

COOPA: A Modular LLM Agent Architecture for Operations Research Problems

DGX agent

arXiv:2606.27611v1 Announce Type: new Abstract: Operations Research (OR) provides a rigorous framework for high-stakes decision-making, but effective OR modeling requires substantial domain knowledge,

agentsarxiv-cs-lg
29 Jun 2026
Model Releases

DMV-Bench: Diagnosing Long-Horizon Multimodal Agents' Visual Memory with Incidental Cue Injection

DGX agent

arXiv:2606.27499v1 Announce Type: cross Abstract: Research on agent memory has matured rapidly, but almost entirely on the text side: few existing benchmarks ask, in an interactive environment, when a

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning

DGX agent

arXiv:2606.27483v1 Announce Type: new Abstract: Large language model (LLM) agents have demonstrated strong capability in sequential decision-making, yet they remains fundamentally reactive in long-hor

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

SidConArena: An Environment Evaluating Agents in Open-Ended,Positive-Sum Bargaining Game

DGX agent

arXiv:2606.27397v1 Announce Type: cross Abstract: Evaluating LLM agents requires dynamic environments that go beyond static reasoning and zero-sum games. Real-world economic interaction is often open-

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

ToolPrivacyBench: Benchmarking Purpose-Bound Privacy in Tool-Using LLM Agents

DGX agent

arXiv:2606.28061v1 Announce Type: cross Abstract: Large language models (LLMs) have increasingly moved from standalone text generation systems to agents that invoke external tools, access environments

model-releasesarxiv-cs-ai
29 Jun 2026
Agents

When Does Personality Composition Matter for Multi-Agent LLM Teams?

DGX agent

arXiv:2606.27443v1 Announce Type: new Abstract: Personality prompting shapes how large language models communicate, yet whether these behavioral shifts affect objective task outcomes remains under-exp

agentsarxiv-cs-ai
29 Jun 2026
Agents

A Guideline-Aware AI Agent for Zero-Shot Target Volume Auto-Delineation

DGX agent

arXiv:2603.09448v2 Announce Type: replace-cross Abstract: Delineating the clinical target volume (CTV) in radiotherapy involves complex margins constrained by tumor location and anatomical barriers. W

agentsarxiv-cs-ai
26 Jun 2026
Agents

Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning

DGX agent

arXiv:2606.27330v1 Announce Type: cross Abstract: Multimodal web agents can assist humans in operating repetitive GUI tasks, where effective task planning is essential for decomposing complex tasks in

agentsarxiv-cs-ai
26 Jun 2026
Safety

IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control

DGX agent

arXiv:2606.26575v1 Announce Type: cross Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, motivating the adoption of learning-based

safetyarxiv-cs-ai
26 Jun 2026
Agents

Privacy-Aware Agent Collaboration for Dynamic VR Slice Management in 6G SD-RAN

DGX agent

arXiv:2606.26123v1 Announce Type: cross Abstract: Ultra-low latency and high throughput are required for Virtual Reality (VR) services in 6G networks, which presents critical challenges for Software-D

agentsarxiv-cs-ai
26 Jun 2026
Model Releases

Qwen-Image-Agent: Bridging the Context Gap in Real-World Image Generation

DGX agent

arXiv:2606.26907v1 Announce Type: new Abstract: While text-to-image (T2I) models have achieved remarkable progress, they struggle with real-world requests that are often underspecified, implicit, or d

model-releasesarxiv-cs-cv
26 Jun 2026
Agents

Skill-MAS: Evolving Meta-Skill for Automatic Multi-Agent Systems

DGX agent

arXiv:2606.18837v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based automatic Multi-Agent Systems (MAS) generation has become a crucial frontier for tackling complex tasks. Howe

agentsarxiv-cs-lg
25 Jun 2026
Agents

Stagnant Neuron: Towards Understanding the Plasticity Loss in Multi-Agent Reinforcement Learning Value Factorization Methods

DGX agent

arXiv:2606.25335v1 Announce Type: new Abstract: Multi-Agent Reinforcement Learning (MARL) value factorization methods can suffer from a loss of plasticity, gradually failing to adapt when transferring

agentsarxiv-cs-lg
25 Jun 2026
Safety

The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapable AI Systems

DGX agent

arXiv:2606.26057v1 Announce Type: cross Abstract: AI agents are granted access to tools, APIs, and other infrastructure, making them active principals in those systems. The dominant approach places co

safetyarxiv-cs-lg
25 Jun 2026
Model Releases

Uncertainty Quantification for Computer-Use Agents: A Benchmark across Vision-Language Models and GUI Grounding Datasets

DGX agent

arXiv:2606.25760v1 Announce Type: cross Abstract: Computer-use agents turn vision-language model (VLM) predictions into executable GUI clicks, so reliable uncertainty estimates are essential for rejec

model-releasesarxiv-cs-cl
25 Jun 2026
Agents

ATRIA: Adaptive Traceable ECG Reporting with Iterative Agents

DGX agent

arXiv:2606.24392v1 Announce Type: new Abstract: Existing ECG report generation is tightly coupled -- interpretation and reporting fused end-to-end, so errors propagate without stage-level recourse --

agentsarxiv-cs-ai
24 Jun 2026
Agents

Bayesian control for coding agents

DGX agent

arXiv:2606.24453v1 Announce Type: new Abstract: Modern coding agents pair LLM generators with various tools, including cheap diagnostics and expensive verifiers. The tool-use decisions are typically g

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

BioMedArena: An Open-source Toolkit for Building and Evaluating Biomedical Deep Research Agents

DGX agent

arXiv:2605.06177v2 Announce Type: replace Abstract: Reproducing and comparing deep research agents today is hard: the same backbone evaluated on the same benchmark can report different accuracies acro

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Metis: Bridging Text and Code Memory for Self-Evolving Agents

DGX agent

arXiv:2606.24151v1 Announce Type: cross Abstract: Self-evolving agents improve over time by distilling experience from past executions and reusing it in future tasks. Existing systems represent such e

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

Qwen-AgentWorld: Language World Models for General Agents

DGX agent

arXiv:2606.24597v1 Announce Type: new Abstract: A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning.

model-releasesarxiv-cs-cl
24 Jun 2026
Local Ai

SHERLOC: Structured Diagnostic Localization for Code Repair Agents

DGX agent

arXiv:2606.24820v1 Announce Type: new Abstract: LLM agents solve repository-level coding tasks through multi-turn tool use, but utilize half their budget on locating faults before editing. Dedicated l

local-aiarxiv-cs-cl
24 Jun 2026
Agents

The Latent Bridge: A Continuous Slow-Fast Channel for Real-Time Game Agents

DGX agent

arXiv:2606.24470v1 Announce Type: new Abstract: A real-time agent for general computer use - with games as the most demanding case - must act within tens of milliseconds while still planning over seco

agentsarxiv-cs-ai
24 Jun 2026
Local Ai

World Models in Pieces: Structural Certification for General Agents

DGX agent

arXiv:2606.24842v1 Announce Type: new Abstract: In the big-world regime, agents cannot be universally capable and their ability is inevitably specialized across a world model in pieces. Consequently,

local-aiarxiv-cs-ai
24 Jun 2026
Safety

Agentic Time Machine as an Infrastructure for Future-Event Forecasting

DGX agent

arXiv:2606.21013v1 Announce Type: cross Abstract: Forecasting future events is a critical challenge for large language model (LLM) agents, spanning domains from elections and monetary policy to financ

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

AI Agents Can Already Autonomously Perform Experimental High Energy Physics

DGX agent

arXiv:2603.20179v3 Announce Type: replace-cross Abstract: Large language model-based AI agents are now able to autonomously execute substantial portions of a high energy physics (HEP) analysis pipelin

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures

DGX agent

arXiv:2606.19380v2 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly

safetyarxiv-cs-lg
23 Jun 2026
Agents

Drowning in Routine: Signal Dilution in Multi-Turn Agent Training

DGX agent

arXiv:2606.22164v1 Announce Type: new Abstract: Multi-turn agents interleave consequential decisions with routine execution: some actions change the downstream return distribution, while others are ne

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Beyond Compaction: Structured Context Eviction for Long-Horizon Agents

DGX agent

arXiv:2606.11213v1 Announce Type: new Abstract: We present Context Window Lifecycle (CWL), a context-management scheme that gives long-horizon LLM agents an effectively unbounded working horizon. As a

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Bootstrapped Monitoring: Leveraging Transparent Reasoning to Oversee Stronger AI Agents

DGX agent

arXiv:2606.11998v1 Announce Type: new Abstract: Trusted monitoring is a cornerstone of AI control. However, as frontier models grow more capable, the increasing capabilities gap between trusted and un

agentsarxiv-cs-lg
11 Jun 2026
Model Releases

Agentic Hybrid RAG for Evidence-Grounded Muon Collider Analysis

DGX agent

arXiv:2606.10381v1 Announce Type: cross Abstract: Muon collider research spans accelerator physics, detector instrumentation, and high-energy phenomenology, with relevant evidence scattered across a r

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

AutoPDE: Reliable Agentic PDE Solving via Explicitly Represented Solver Strategies

DGX agent

arXiv:2606.10752v1 Announce Type: new Abstract: Numerical solvers for partial differential equations (PDEs) are core computational tools in science and engineering. Building reliable PDE solvers requi

agentsarxiv-cs-ai
10 Jun 2026
Local Ai

Decentralized Multi-Agent Systems with Shared Context

DGX agent

arXiv:2606.10662v1 Announce Type: cross Abstract: Multi-agent systems (MAS) can scale large language model reasoning at test time by decomposing complex problems into parallel subtasks. However, most

local-aiarxiv-cs-ai
10 Jun 2026
Model Releases

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

DGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

TabClaw: An Interactive and Self-Evolving Agent for Spreadsheet Manipulation and Table Reasoning

DGX agent

arXiv:2606.10316v1 Announce Type: new Abstract: Spreadsheets and tables are widely used representations for structured data analysis, but effective analysis still requires substantial manual effort an

agentsarxiv-cs-cl
10 Jun 2026
Agents

What Spatial Memory Must Store: Occlusion as the Test for Language-Agent Memory

DGX agent

arXiv:2606.10299v1 Announce Type: new Abstract: Language-agent 'memory palace' systems anchor each memory to a world coordinate, on the intuition that geometry adds something text cannot. We make that

agentsarxiv-cs-ai
10 Jun 2026
Agents

A multi-agent system for spine MRI report generation from multi-sequence imaging

DGX agent

arXiv:2606.08897v1 Announce Type: cross Abstract: Spinal pathology is a leading cause of pain and disability worldwide. Spine MRI is central to clinical evaluation, yet its interpretation remains comp

agentsarxiv-cs-ai
9 Jun 2026
Agents

Agentic multi-fidelity learning of quasiparticle and excitonic properties

DGX agent

arXiv:2606.07836v1 Announce Type: cross Abstract: Many-body GW-Bethe-Salpeter equation calculations are essential for accurate simulations of electronic structure and optical properties in modern low-

agentsarxiv-cs-ai
9 Jun 2026
Agents

AgentTrust: A Self-Improving Trust Layer for AI-Agent Actions

DGX agent

arXiv:2606.08539v1 Announce Type: new Abstract: AI agents increasingly take consequential actions -- shell commands, cloud operations, and arbitrary tool-calls -- so a trust layer must decide, per act

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

DGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…5657585960…233
Next →