AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,215 results
26 May 2026

MedBeads: An Agent-Native, Immutable Data Substrate for Trustworthy Medical AI

AgentsDGX agent

arXiv:2602.01086v2 Announce Type: replace Abstract: Background: As of 2026, Large Language Models (LLMs) demonstrate expert-level medical knowledge. However, deploying them as autonomous 'Clinical Age

MemForest: An Efficient Agent Memory System with Hierarchical Temporal Indexing

AgentsDGX agent

arXiv:2605.23986v1 Announce Type: cross Abstract: Memory is a fundamental component for enabling long-context LLM agents, supporting persistent state across interactions through a continuous serve-and

Meta-Agent: From Task Descriptions to Verified Multi-Agent Systems

AgentsDGX agent

arXiv:2605.25233v1 Announce Type: new Abstract: AI agents are increasingly used to solve complex, multi-step tasks, but existing multi-agent frameworks remain brittle as workflows grow in scale and de


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report

AgentsDGX agent

arXiv:2605.25665v1 Announce Type: cross Abstract: AI-native software development is often evaluated at the level of individual models, prompts, or generated artifacts. This framing is insufficient for

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

AgentsDGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

Microsoft Copilot Cowork Exfiltrates Files

AgentsDGX agent

Microsoft Copilot Cowork Exfiltrates Files The biggest challenge in designing agentic systems continues to be preventing them from enabling attackers to exfiltrate data. In this case Microsoft Copilot

MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research

AgentsDGX agent

arXiv:2605.26114v1 Announce Type: new Abstract: We present MobileGym, a browser-hosted, lightweight, fully controllable environment for everyday mobile use, targeting interaction fidelity without repl

More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries

AgentsDGX agent

arXiv:2605.24050v1 Announce Type: cross Abstract: Skill libraries allow LLM agents to load task-specific instructions on demand, letting non-expert users solve domain-specific tasks through natural la

Multi-Agent Specification-based Metamorphic Testing of FMU-Based Simulations

AgentsDGX agent

arXiv:2605.25101v1 Announce Type: cross Abstract: In many industrial domains, the Functional Mock-up Interface (FMI) is used to exchange simulation models as Functional Mock-up Units (FMUs) across dif

Multi-Level Strategic Classification: Incentivizing Improvement through Promotion and Relegation Dynamics

AgentsDGX agent

arXiv:2602.11439v2 Announce Type: replace Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outc

Multi-market value-stacking: Battery control for combined imbalance participation and non-uniform FCR bidding

AgentsDGX agent

arXiv:2605.23964v1 Announce Type: cross Abstract: The growing share of Renewable Energy Sources (RES) in modern power systems increases both grid imbalances and frequency deviations, reinforcing the n

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

AgentsDGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

NASA has just launched a new website for its Moon Base missions, which aims to build a permanent $20 billion U.S. base on the Moon. @SpaceX'…

AgentsDGX agent

NASA has just launched a new website for its Moon Base missions, which aims to build a permanent $20 billion U.S. base on the Moon. @SpaceX's Starship rocket will play a big role in these missions. 'T

Neuronal Stochastic Attention Circuit (NSAC) for Probabilistic Representation Learning

AgentsDGX agent

arXiv:2605.26061v1 Announce Type: cross Abstract: Reliable quantification of uncertainty estimates in continuous-time (CT) representation learning remains nascent, particularly within CT attention arc

New agent skill to convert YouTube videos to slides and notes.

AgentsDGX agent

New agent skill to convert YouTube videos to slides and notes. Just built an insane new agent skill. It can perfectly extract slides from YT videos, then write notes, images, transcripts, and slides i

OPAL: Omnidirectional Path-efficient Aerial 3D expLoration

AgentsDGX agent

arXiv:2605.25423v1 Announce Type: new Abstract: Autonomous exploration is critical for robot mapping unknown environments. Desirable characteristics of exploration algorithms include compute efficienc

Optimal Design for Multinomial Logit Model with Applications to Best Assortment Identification

AgentsDGX agent

arXiv:2605.25592v1 Announce Type: cross Abstract: We study optimal experimental design for multinomial logit (MNL) bandits, where an agent repeatedly selects a subset of K items from a ground set of s

PANDO: Efficient Multimodal AI Agents via Online Skill Distillation

AgentsDGX agent

arXiv:2605.24785v1 Announce Type: new Abstract: Recent advances in multimodal web agents often rely on increased inference-time computation, including rollout search, verifier passes, offline skill di

Passivity-based Semi-autonomous Rotational Motion Navigation for Rigid-body Networks: Stability and Human Passivity Analysis

AgentsDGX agent

arXiv:2605.24731v1 Announce Type: cross Abstract: This paper presents a novel passivity-based semi-autonomous attitude control framework, with a particular focus on attitude kinematics defined on the

Path Following Control System of Line-of-Sight Guidance for Robotic Dolphin with Multi-Link Mechanism in Underwater Simulator

AgentsDGX agent

arXiv:2605.25401v1 Announce Type: new Abstract: Biomimetic autonomous underwater vehicle (BAUV) with multi-link mechanism is widely used in aquatic life observation and environmental surveys due to it

PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning

AgentsDGX agent

arXiv:2502.10906v2 Announce Type: replace Abstract: Reward design plays a pivotal role in the training of game AIs, requiring substantial domain-specific knowledge and human effort. In recent years, s

Performance Comparison of Classical and Neural Sampling Algorithms for Robotic Navigation

AgentsDGX agent

arXiv:2605.25010v1 Announce Type: cross Abstract: Integrating artificial intelligence (AI) into sampling-based motion planning provides new possibilities for improving autonomous navigation efficiency

Persuasion Should be Double-Blind: A Multi-Domain Dialogue Dataset With Faithfulness Based on Causal Theory of Mind

AgentsDGX agent

arXiv:2502.21297v2 Announce Type: replace Abstract: Persuasive dialogue is central to human communication, yet existing datasets often rely on a single language model generating both roles, producing

Physical AI takes off: How real-time data keeps Fraport’s airports running on time

AgentsDGX agent

Physical AI is no longer a concept confined to factory floors and autonomous vehicles — it is reshaping the complex, time-sensitive operations of some of the world’s busiest airports. The operators ma

Practical Quantum CIM Empowerment via All-Domestic-Core Agentic Large Model

AgentsDGX agent

arXiv:2605.23934v1 Announce Type: new Abstract: Quantum computing devices are recognized as powerful tools for solving NP-complete problems. However, the intricacy of their modeling presents notable b

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

AgentsDGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

Proper Scoring Rules for Agentic Uncertainty Quantification

AgentsDGX agent

arXiv:2605.24756v1 Announce Type: new Abstract: Language-model agents increasingly emit uncertainty signals throughout a trajectory, but existing agentic UQ evaluations often conflate ranking usefulne

Rethinking organizational design in the age of agentic AI

AgentsDGX agent

Amid rapidly growing adoption of enterprise-level AI agents, there’s a disconnect emerging between ambition and execution. Although 85% of organizations say they want to be agentic within the next thr

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

AgentsDGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

Reward Shaping and Action Masking for Compositional Tasks using Behavior Trees and LLMs

AgentsDGX agent

arXiv:2605.05795v2 Announce Type: replace Abstract: Decomposing complex tasks into a sequence of simpler subtasks can improve learning efficiency for an autonomous agent. Reinforcement learning (RL) c

SAM: State-Adaptive Memory for Long-Horizon Reasoning Agent

AgentsDGX agent

arXiv:2605.24468v1 Announce Type: new Abstract: Long-horizon agentic reasoning requires large language models to act over long interaction histories containing thoughts, tool calls, observations, and

Scaling up Energy-Aware Multi-Agent Reinforcement Learning for Mission-Oriented Drone Networks with Individual Reward

AgentsDGX agent

arXiv:2605.24992v1 Announce Type: cross Abstract: Multi-agent reinforcement learning (MARL) has shown wide applicability in collaborative systems such as autonomous driving and smart cities for its ab

Security of OpenClaw Agents: Fundamentals, Attacks, and Countermeasures

AgentsDGX agent

arXiv:2605.25435v1 Announce Type: new Abstract: The rapid evolution of large language model (LLM)-driven autonomous agents has given rise to OpenClaw, a new class of open-source agent frameworks that

SoK: DARPA's AI Cyber Challenge (AIxCC): Competition Design, Architectures, and Lessons Learned

AgentsDGX agent

arXiv:2602.07666v3 Announce Type: replace-cross Abstract: DARPA's AI Cyber Challenge (AIxCC, 2023--2025) is the largest competition to date for building fully autonomous cyber reasoning systems (CRSs)

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

AgentsDGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

STORM: Internalized Modeling for Spatial-Temporal Reasoning in Video-Language Models

AgentsDGX agent

arXiv:2605.26014v1 Announce Type: cross Abstract: Many video reasoning tasks require tracking motion, temporal order, and evolving visual states across frames. Existing methods built on large vision-l

Technical deep dive: AgentCore payments and innovation in agentic commerce

AgentsDGX agent

Amazon Bedrock AgentCore payments is now available in preview, it provides instant payments to paid external services with no manual billing setup per provider, stablecoin support for cost-effective m

Test-Time Deep Thinking to Explore Implicit Rules

AgentsDGX agent

arXiv:2605.24828v1 Announce Type: new Abstract: With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in

the importance of traces!

AgentsDGX agent

the importance of traces! Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective agents requires shipping early, obser

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

AgentsDGX agent

arXiv:2605.25310v1 Announce Type: new Abstract: Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether t

Toward Enactive Artificial Intelligence

AgentsDGX agent

arXiv:2605.24238v1 Announce Type: new Abstract: In this paper, we advocate for incorporating enactive approaches to perception and cognition into artificial intelligence (AI). Enactive approaches view

Towards Multi-Turn Dialog Systems for Industrial Asset Operations and Maintenance

AgentsDGX agent

arXiv:2605.24953v1 Announce Type: new Abstract: Industrial asset operations and maintenance question answering is inherently multi-turn, iterative, and highly dependent on external tool invocation. Ho

Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective age…

AgentsDGX agent

Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating effective agents requires shipping early, observing behavior, and iterati

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

AgentsDGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

Why Agentic Theorem Prover Works: A Statistical Provability Theory of Mathematical Reasoning Models

AgentsDGX agent

arXiv:2602.10538v3 Announce Type: replace-cross Abstract: Agentic theorem provers combine a reasoning model, retrieval, search, and a proof assistant verifier, yet it remains unclear which components

Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform

AgentsDGX agent

arXiv:2605.23972v1 Announce Type: new Abstract: Large language models achieve strong performance in language generation and knowledge-intensive tasks, yet remain limited in settings requiring causal r

Why Your Deep Research Agent Fails? On Hallucination Evaluation in Full Research Trajectory

AgentsDGX agent

arXiv:2601.22984v2 Announce Type: replace Abstract: Diagnosing failure patterns in Deep Research Agents (DRAs) remains a critical challenge. Existing benchmarks predominantly rely on end-to-end evalua

You can’t build an autonomous agent on a static RAG pipeline. Because goals mutate dynamically, agents require an advanced knowledge infrast…

AgentsDGX agent

You can’t build an autonomous agent on a static RAG pipeline. Because goals mutate dynamically, agents require an advanced knowledge infrastructure, rather than forcing the model to hunt through a raw

your bank knows what you did. it has no idea why. that context, the why behind every transaction, is the most valuable data in banking. and …

AgentsDGX agent

your bank knows what you did. it has no idea why. that context, the why behind every transaction, is the most valuable data in banking. and nobody's capturing it. this changes with agentic banking. Me

25 May 2026

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

AgentsDGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEn…

AgentsDGX agent

A full tour through RAG, document context, and AI agents - from 2023 to 2026 🌎🤖 @hexapode gave a comprehensive 90-min workshop at @aiDotEngineer Singapore last week that comprehensively traces through

A Proactive Multi-Agent Dialogue Framework for Assessing Social Language Disorder Traits in Autism

AgentsDGX agent

arXiv:2605.22993v1 Announce Type: cross Abstract: Characteristic linguistic behaviors associated with Social Language Disorder (SLD) in autism spectrum disorder, including echoic repetition, pronoun d

Agentivism: a learning theory for the age of artificial intelligence

AgentsDGX agent

arXiv:2604.07813v2 Announce Type: replace Abstract: Learning theories have historically changed when the conditions of learning evolved. Generative and agentic AI create a new condition by allowing le

AI Assurance: A Comprehensive Testing Strategy for Enterprise AI Systems

AgentsDGX agent

arXiv:2605.23459v1 Announce Type: cross Abstract: Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software

As token costs shoot up, Dell doubles down on desktop AI

AgentsDGX agent

As enterprises confront spiraling cloud inference bills, on-prem AI computing is emerging as the decisive lever for making agentic AI economically viable at scale. Dell Technologies Inc.’s announcemen

Autonomous Frontier-Based Exploration with VLM Guidance

AgentsDGX agent

arXiv:2605.23165v1 Announce Type: cross Abstract: Autonomous robotic exploration of unknown and hazardous environments, a long-standing challenge, can be significantly improved by leveraging the advan

AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery

AgentsDGX agent

arXiv:2605.23204v1 Announce Type: new Abstract: Scientific research is being reshaped by AI systems that move beyond isolated assistance toward longer-horizon workflows spanning literature grounding,

Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents

AgentsDGX agent

arXiv:2605.23590v1 Announce Type: new Abstract: ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which r

Curriculum reinforcement learning with measurable task representation learning

AgentsDGX agent

arXiv:2605.23372v1 Announce Type: cross Abstract: In curriculum reinforcement learning (CRL), an agent incrementally accumulates knowledge over a sequence of tasks (i.e., a curriculum), and the learni

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

AgentsDGX agent

arXiv:2605.23271v1 Announce Type: cross Abstract: The rapid evolution of generative video foundation models has propelled the field toward professional-grade cinematic synthesis. To achieve such deman

← Previous
1…6162636465…121
Next →