AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

To Isolate or to Score? Model-Adaptive Assessment for Cost-Efficient Multi-Agent RAG

DGX agent

arXiv:2606.25191v1 Announce Type: cross Abstract: Multi-agent document assessment for retrieval-augmented generation is computationally expensive, driving practitioners toward smaller, deployable mode

agentsarxiv-cs-cl
25 Jun 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Are We Ready For An Agent-Native Memory System?

DGX agent

arXiv:2606.24775v1 Announce Type: new Abstract: Memory for large language model (LLM) agents has rapidly evolved from simple retrieval-augmented mechanisms into a data management system that supports

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

Detecting AI Coding Agents in Open Source: A Validated Multi-Method Census of 180 Million Repositories

DGX agent

arXiv:2606.24429v1 Announce Type: cross Abstract: Generative AI coding agents are entering the open-source supply chain, yet their diverse and often invisible traces leave their prevalence poorly unde

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Emergent Relational Order in LLM Agent Societies: From Collective Affect to Authority Stratification

DGX agent

arXiv:2606.23764v1 Announce Type: cross Abstract: Fei Xiaotong's Differential Order Pattern characterizes rural society as egocentric and relationally graded, with cooperation attenuating over social

agentsarxiv-cs-ai
24 Jun 2026
Agents

Paying to Know: Micro-Transaction Markets for Verified Product Information in Agentic E-Commerce

DGX agent

arXiv:2606.24783v1 Announce Type: cross Abstract: Commercial NLP treats the shopping chatbot as a recommender or a conversion tool: its job is to match a user to a catalogue entry and close a sale. We

agentsarxiv-cs-ai
24 Jun 2026
Safety

Red-Teaming the Agentic Red-Team

DGX agent

arXiv:2606.24496v1 Announce Type: cross Abstract: The use of agentic systems to perform offensive security operations has moved from a theoretical possibility to a commoditized capability. However, wh

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

ReMMD: Realistic Multilingual Multi-Image Agentic Verification for Multimodal Misinformation Detection

DGX agent

arXiv:2606.24112v1 Announce Type: new Abstract: Multimodal misinformation detection is increasingly important because viral posts now combine long multilingual narratives, several images, mixed proven

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents

DGX agent

arXiv:2606.20625v1 Announce Type: cross Abstract: LLM agents are promising for alpha mining via combining financial priors, symbolic reasoning, executable factor generation, and feedback-driven refine

agentsarxiv-cs-lg
23 Jun 2026
Agents

Causal Discovery in the Era of Agents

DGX agent

arXiv:2606.23608v1 Announce Type: cross Abstract: Recent attempts to combine large language models (LLMs) with causal discovery ask models to infer pairwise directions, propose graph structures, or in

agentsarxiv-cs-lg
23 Jun 2026
Safety

Darwin Mobile Agent: A Roadmap for Self-Evolution

DGX agent

arXiv:2606.20622v1 Announce Type: cross Abstract: The goal of artificial intelligence is to create agents capable of general, adaptive behaviour in open-ended environments. Guided by the 'Bitter Lesso

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Distilling Collaborative Dynamics into Latent Space for Implicit Coordination in Decentralized Multi-Agent Manipulation

DGX agent

arXiv:2606.22982v1 Announce Type: new Abstract: Multi-arm manipulation demands precise spatiotemporal coordination, yet many centralized approaches scale poorly as team size increases. To address this

model-releasesarxiv-cs-ro
23 Jun 2026
Agents

Dynamic multi-agent deep reinforcement learning-based pricing and incentivization approach in multimodal transportation networks

DGX agent

arXiv:2606.23257v1 Announce Type: new Abstract: In multimodal transportation systems, shared mobility services (SMSs) are promoted for their potential to enhance flexibility and reduce congestion. How

agentsarxiv-cs-lg
23 Jun 2026
Safety

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

DGX agent

arXiv:2606.22995v1 Announce Type: new Abstract: Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy up

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

Next-Gen CAPTCHAs: Leveraging the Cognitive Gap for Scalable and Diverse GUI-Agent Defense

DGX agent

arXiv:2602.09012v2 Announce Type: replace Abstract: The rapid evolution of GUI-enabled agents has rendered traditional CAPTCHAs obsolete. While previous benchmarks like OpenCaptchaWorld established a

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Probe-and-Refine Tuning of Repository Guidance for Coding Agents

DGX agent

arXiv:2606.20512v2 Announce Type: replace-cross Abstract: LLM-based coding agents need higher-level operational knowledge about a repository (which files house which subsystems, how to run the test su

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Skill Coverage: A Test Adequacy Metric for Agent Skills

DGX agent

arXiv:2606.20659v1 Announce Type: cross Abstract: Agent skills encode reusable procedural knowledge that guides large language model agents across tasks and execution contexts. Existing evaluations pr

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

Towards Adaptive Categories: Dimensional Governance for Agentic AI

DGX agent

arXiv:2505.11579v3 Announce Type: replace-cross Abstract: As AI systems evolve from static tools to dynamic agents, traditional categorical governance frameworks -- based on fixed risk tiers, levels o

agentsarxiv-cs-lg
23 Jun 2026
Model Releases

Can AI Agents Synthesize Scientific Conclusions?

DGX agent

arXiv:2606.11337v1 Announce Type: new Abstract: Scientific AI agents increasingly retrieve evidence, reason across sources, and synthesize conclusions used in consequential decisions. Yet, their abili

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

Counterexample Guided Learning in the Large using Reasoning Agents

DGX agent

arXiv:2606.11521v1 Announce Type: new Abstract: LLMs and LLM agents should improve when given feedback, but identifying when they are able to do so is difficult: feedback is heterogeneous, domain-spec

agentsarxiv-cs-lg
11 Jun 2026
Agents

Goal-Autopilot: A Verifiable Anti-Fabrication Firewall for Unattended Long-Horizon Agents

DGX agent

arXiv:2606.11688v1 Announce Type: cross Abstract: Long-horizon LLM agents are not trusted to run unattended: with no human watching, they confidently report success they never verified. We treat hones

agentsarxiv-cs-ai
11 Jun 2026
Safety

Improving Generalization and Data Efficiency with Diffusion in Offline Multi-agent RL

DGX agent

arXiv:2307.01472v2 Announce Type: replace Abstract: We present a novel Diffusion Offline Multi-agent Model (DOM2) for offline Multi-Agent Reinforcement Learning (MARL). Different from existing algorit

safetyarxiv-cs-ai
11 Jun 2026
Local Ai

Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness

DGX agent

arXiv:2606.11686v1 Announce Type: cross Abstract: End-to-end task-success is the dominant way to evaluate LLM agents, but one aggregate number tells you that an agent regressed, not where. We present

local-aiarxiv-cs-ai
11 Jun 2026
Agents

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

DGX agent

arXiv:2606.11897v1 Announce Type: new Abstract: Scientific discovery workflows usually contain and rely heavily on lab notes, where researchers record observations, interpret uncertain results, and pl

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

DGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

WebChallenger: A Reliable and Efficient Generalist Web Agent

DGX agent

arXiv:2606.10423v1 Announce Type: new Abstract: Autonomous web navigation remains challenging for LLM agents, and the strongest generalist systems rely on proprietary reasoning models whose inference

model-releasesarxiv-cs-cl
10 Jun 2026
Agents

A Multi-Agent System for IPMSM Design Optimization via an FEA-AI Hybrid Approach

DGX agent

arXiv:2606.09037v1 Announce Type: new Abstract: Interior permanent magnet synchronous motor (IPMSM) design requires balancing conflicting objectives and multi-physics constraints, while modern optimiz

agentsarxiv-cs-ai
9 Jun 2026
Agents

Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading Systems

DGX agent

arXiv:2606.08285v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems are increasingly proposed for financial trading, yet their reported performance remains difficult to co

agentsarxiv-cs-ai
9 Jun 2026
Safety

Data Agents Under Attack: Vulnerabilities in LLM-Driven Analytical Systems

DGX agent

arXiv:2606.08661v1 Announce Type: cross Abstract: Data agents integrate LLM-driven reasoning with relational data access, executable analytical tools, and multi-step workflow orchestration, making the

safetyarxiv-cs-ai
9 Jun 2026
Agents

From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG

DGX agent

arXiv:2603.03292v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) exhibit high reasoning capacity in medical question-answering, but their tendency to produce hallucinations and o

agentsarxiv-cs-ai
9 Jun 2026
Agents

MAVIS: Multi-Agent Video Retrieval via Structured Video Understanding

DGX agent

arXiv:2606.09641v1 Announce Type: new Abstract: The dominant paradigm in video retrieval relies on embedding-based full-corpus scanning, which suffers from inherent computational inefficiency and the

agentsarxiv-cs-cv
9 Jun 2026
Agents

Observability for Delegated Execution in Agentic AI Systems

DGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

agentsarxiv-cs-ai
9 Jun 2026
Safety

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

DGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

safetyarxiv-cs-ai
9 Jun 2026
Agents

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

DGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

agentsarxiv-cs-ai
9 Jun 2026
Agents

SKILL.nb: Selective Formalization and Gated Execution for Durable Agent Workflows

DGX agent

arXiv:2606.08049v1 Announce Type: new Abstract: AI agents increasingly turn past experience into reusable artifacts such as code, workflows, and procedural memories. Reuse can improve efficiency, but

agentsarxiv-cs-ai
9 Jun 2026
Safety

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

DGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

safetyarxiv-cs-ai
9 Jun 2026
Safety

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

DGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

safetyarxiv-cs-ai
9 Jun 2026
Agents

ViMax: Agentic Video Generation

DGX agent

arXiv:2606.07649v1 Announce Type: cross Abstract: Long-form video generation requires systematic narrative planning and visual consistency that current short-clip methods cannot provide. Existing meth

agentsarxiv-cs-ai
9 Jun 2026
Safety

Visual Para-Thinker++: A Single-Policy Multi-Agent Framework for Visual Reasoning

DGX agent

arXiv:2606.09290v1 Announce Type: new Abstract: Visual reasoning requires integrating evidence distributed across regions, attributes, and relations, making single-chain reasoning prone to early perce

safetyarxiv-cs-cv
9 Jun 2026
Model Releases

EvoClaw: Evaluating AI Agents on Continuous Software Evolution

DGX agent

arXiv:2603.13428v2 Announce Type: replace-cross Abstract: With AI agents increasingly deployed as long-running systems, it becomes essential to autonomously construct and continuously evolve customize

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning

DGX agent

arXiv:2606.06532v1 Announce Type: new Abstract: Despite significant progress in agentic long video understanding, existing methods still lack detailed motion comprehension coupled with an efficient me

agentsarxiv-cs-cv
8 Jun 2026
Agents

Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA

DGX agent

arXiv:2603.24481v2 Announce Type: replace Abstract: Miscalibrated confidence scores are a practical obstacle to deploying AI in clinical settings. A model that is always overconfident offers no useful

agentsarxiv-cs-ai
8 Jun 2026
Safety

Self-evolving LLM agents with in-distribution Optimization

DGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

safetyarxiv-cs-lg
8 Jun 2026
Agents

Signal-Driven Observation for Long-Horizon Web Agents

DGX agent

arXiv:2606.06708v1 Announce Type: new Abstract: Web agents operating over long horizons ingest raw DOM and accessibility trees -- routinely tens of thousands of tokens -- at every action step, causing

agentsarxiv-cs-cl
8 Jun 2026
Model Releases

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

DGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

model-releasesarxiv-cs-cl
8 Jun 2026
Agents

TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

DGX agent

arXiv:2606.07054v1 Announce Type: cross Abstract: Autonomous LLM agents can pursue hidden malicious objectives through sequences of individually benign actions, making sabotage difficult to detect usi

agentsarxiv-cs-ai
8 Jun 2026
Agents

2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support

DGX agent

arXiv:2602.21889v2 Announce Type: replace Abstract: Predictions from ML models support human decision making in several fields, including high-stakes ones such as healthcare and the judiciary. Yet, we

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

DGX agent

arXiv:2606.06448v1 Announce Type: new Abstract: LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale re

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

SciVisAgentSkills: Design and Evaluation of Agent Skills for Scientific Data Analysis and Visualization

DGX agent

arXiv:2606.05525v1 Announce Type: new Abstract: Recent advances in agentic visualization have enabled the translation of natural language into executable scientific visualization (SciVis) workflows. W

model-releasesarxiv-cs-ai
6 Jun 2026
← Previous
1…4748495051…233
Next →