AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,963 results
Research

DRIVE: Modeling Skills at the Reasoning and Interaction Levels for Web Agents under Continual Learning

DGX agent

arXiv:2605.23939v1 Announce Type: new Abstract: Web agents require both high-level reasoning (for task decomposition) and low-level interactions (for page elements manipulation) to conduct different t

researcharxiv-cs-ai
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Dynamic Dual-Granularity Skill Bank for Agentic RL

DGX agent

arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often l

safetyarxiv-cs-ai
26 May 2026
Safety

ECHO: Terminal Agents Learn World Models for Free

DGX agent

arXiv:2605.24517v1 Announce Type: cross Abstract: CLI agents are the closest thing language models have to an embodied setting: the model emits commands, the terminal executes them, and the returned s

safetyarxiv-cs-cl
26 May 2026
Model Releases

Memory-Induced Tool-Drift in LLM Agents

DGX agent

arXiv:2605.24941v1 Announce Type: cross Abstract: Modern LLM agents combine long-term memory for personalization with tool-calling interfaces for taking actions in the world -- a combination underpinn

model-releasesarxiv-cs-lg
26 May 2026
Research

Mitigating Provenance-Role Collapse in Long-Term Agents via Typed Memory Representation

DGX agent

arXiv:2605.25869v1 Announce Type: new Abstract: Long-term memory is essential for persistent LLM agents, yet prevailing architectures store historical interactions as unstructured, flat text. This unc

researcharxiv-cs-cl
26 May 2026
Agents

MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research

DGX agent

arXiv:2605.26114v1 Announce Type: new Abstract: We present MobileGym, a browser-hosted, lightweight, fully controllable environment for everyday mobile use, targeting interaction fidelity without repl

agentsarxiv-cs-ai
26 May 2026
Hardware

Rambus targets agentic AI workloads with faster client memory chipset

DGX agent

Rambus Inc. today announced a complete DDR5 9600 client memory module chipset designed to push PC memory speeds to 9,600 megatransfers per second, targeting the bandwidth and capacity demands of agent

hardwaresiliconangle
26 May 2026
Safety

The OpenAI insider @thsottiaux has a warning for everyone offloading their thinking to agents.

DGX agent

An OpenAI insider (@thsottiaux) raises concerns about the risks of over-relying on AI agents to handle cognitive tasks, warning against wholesale delegation of thinking to autonomous systems. The warn

safetygary-marcus--x
26 May 2026
Model Releases

VeriTrace: Evolving Mental Models for Deep Research Agents

DGX agent

arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing systems explore what evolving intermediate representatio

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

When Do LLM Agents Treat Surface Noise Differently from Semantic Noise? A 68-Cell Measurement Study with a Held-Out Trace-Level Validation

DGX agent

arXiv:2605.25981v1 Announce Type: new Abstract: We document an empirical phenomenon in chain-of-thought and ReAct agents driven by ten large language models from seven architecture families: meaning-b

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

Ax-Prover: A Deep Reasoning Agentic Framework for Theorem Proving in Mathematics and Quantum Physics

DGX agent

arXiv:2510.12787v4 Announce Type: replace Abstract: We present Ax-Prover, a multi-agent system for automated theorem proving in Lean that can solve problems across diverse scientific domains and opera

model-releasesarxiv-cs-ai
25 May 2026
Safety

Goal-Conditioned Agents that Learn Everything All at Once

DGX agent

arXiv:2605.23551v1 Announce Type: cross Abstract: A goal-conditioned reinforcement learning agent exploring an environment will see a wealth of information throughout a trajectory, most of which is di

safetyarxiv-cs-ai
25 May 2026
Local Ai

HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation

DGX agent

arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e

local-aiarxiv-cs-cl
25 May 2026
Model Releases

LLM-driven design of physics-constrained constitutive models: two agents are better than one

DGX agent

arXiv:2605.23754v1 Announce Type: new Abstract: Developing constitutive models that capture how materials deform under load traditionally requires years of specialized expertise in continuum mechanics

model-releasesarxiv-cs-lg
25 May 2026
Model Releases

What Training Data Teaches RL Memory Agents: An Empirical Study of Curriculum Effects in Memory-Augmented QA

DGX agent

arXiv:2605.23067v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a viable recipe for training LLM agents to reason over external memory banks in multi-session dialogue. Exist

model-releasesarxiv-cs-cl
25 May 2026
Research

WMAttack: Automated Attack Search for Adversarial Evaluation of World-Model Agents

DGX agent

arXiv:2605.23220v1 Announce Type: new Abstract: Despite the growing use of world models as decision-making agents, their adversarial robustness remains underexplored due to the lack of dedicated autom

researcharxiv-cs-lg
25 May 2026
Model Releases

Blind Spots in the Guard: How Domain-Camouflaged Injection Attacks Evade Detection in Multi-Agent LLM Systems

DGX agent

arXiv:2605.22001v1 Announce Type: cross Abstract: Injection detectors deployed to protect LLM agents are calibrated on static, template-based payloads that announce themselves as override directives.

model-releasesarxiv-cs-cl
22 May 2026
Model Releases

Diverse Yet Consistent: Context-Guided Diffusion with Energy-Based Joint Refinement for Multi-Agent Motion Prediction

DGX agent

arXiv:2605.22017v1 Announce Type: new Abstract: Deepgenerative models havebecomeapromisingapproach for human motion prediction due to their ability to capture multimodal distributions and represent di

model-releasesarxiv-cs-cv
22 May 2026
Agents

SpecHop: Continuous Speculation for Accelerating Multi-Hop Retrieval Agents

DGX agent

arXiv:2605.21965v1 Announce Type: new Abstract: Large language models increasingly use external tools such as web search and document retrieval to solve information-intensive tasks. However, multi-hop

agentsarxiv-cs-cl
22 May 2026
Local Ai

VBFDD-Agent for Electric Vehicle Battery Fault Detection and Diagnosis: Descriptive Text Modeling of Battery Digital Signals

DGX agent

arXiv:2605.20742v1 Announce Type: new Abstract: With the rapid proliferation of electric vehicles, the safety and reliability of lithium-ion batteries have become critical concerns. Effective anomaly

local-aiarxiv-cs-ai
22 May 2026
Model Releases

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure…

DGX agent

Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure out what comes next: 1:46 Omni: 'Nano Banana for video' 4:5

model-releasesrowan-cheung--x
21 May 2026
Model Releases

Lean Refactor: Multi-Objective Controllable Proof Optimization via Agentic Strategy Search

DGX agent

arXiv:2605.20244v1 Announce Type: cross Abstract: We present Lean Refactor, a plug-and-play retrieval-augmented agentic framework for multi-objective, controllable, and version-robust refactoring of L

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory

DGX agent

arXiv:2602.06025v2 Announce Type: replace Abstract: Memory is increasingly central to Large Language Model (LLM) agents operating beyond a single context window, yet most existing systems rely on offl

model-releasesarxiv-cs-cl
21 May 2026
Agents

Spotify Studio’s AI agent creates a daily podcast just for you

DGX agent

Studio by Spotify Labs is a new standalone AI app that generates a daily briefing, podcasts, and playlists on your PC using chatbot prompts. The AI-generated content draws from your Spotify listening

agentsthe-verge-ai
21 May 2026
Model Releases

TRAM: Test-Time Risk Adaptation with Mixture of Agents

DGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

[AINews] Google I/O 2026: Gemini 3.5 Flash, Omni (NanoBanana for Video), Spark (background agents), and Antigravity 2.0

DGX agent

Google I/O 2026 featured several new AI model releases including Gemini 3.5 Flash, an Omni model codenamed NanoBanana for video processing, Spark for background agent tasks, and Antigravity 2.0. These

model-releaseslatent-space
20 May 2026
Safety

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

DGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

safetyarxiv-cs-ai
20 May 2026
Agents

KadiAssistant: A conversational AI Agent for information retrieval in Kadi4Mat

DGX agent

arXiv:2605.18850v1 Announce Type: cross Abstract: We introduce KadiAssistant, a privacy-by-design AI assistant integrated into the Kadi research data ecosystem, enabling researchers to efficiently acc

agentsarxiv-cs-ai
20 May 2026
Safety

LLM agents & memory systems operate in continuously updated environments (Git repos, evolving docs). They must process long contexts, recove…

DGX agent

LLM agents & memory systems operate in continuously updated environments (Git repos, evolving docs). They must process long contexts, recover earlier information, and reason over many updates that cre

safetyjeremy-howard--x
20 May 2026
Research

OpenComputer: Verifiable Software Worlds for Computer-Use Agents

DGX agent

arXiv:2605.19769v1 Announce Type: new Abstract: We present OpenComputer, a verifier-grounded framework for constructing verifiable software worlds for computer-use agents. OpenComputer integrates four

researcharxiv-cs-ai
20 May 2026
Model Releases

Physics-in-the-Loop: A Hybrid Agentic Architecture for Validated CAD Engineering Design

DGX agent

arXiv:2605.19717v1 Announce Type: new Abstract: Large Language Models (LLMs) can generate Computer-Aided Design (CAD), yet lack physical comprehension required for reliable engineering design. Instead

model-releasesarxiv-cs-cv
20 May 2026
Safety

Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

DGX agent

arXiv:2605.20061v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) is a promising paradigm for improving large language model (LLM) agents on long-horizon interactiv

safetyarxiv-cs-cl
20 May 2026
Agents

Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking

DGX agent

arXiv:2605.18852v1 Announce Type: cross Abstract: Checkpoint selection for multimodal large language models (MLLMs) presents significant challenges when performance differentials are marginal and eval

agentsarxiv-cs-ai
20 May 2026
Model Releases

STAR-PolyaMath: Multi-Agent Reasoning under Persistent Meta-Strategic Supervision

DGX agent

arXiv:2605.19338v1 Announce Type: cross Abstract: Frontier AI models and multi-agent systems have led to significant improvements in mathematical reasoning. However, for problems requiring extended, l

model-releasesarxiv-cs-ai
20 May 2026
Agents

Towards LLM-Assisted Architecture Recovery for Real-World ROS~2 Systems: An Agent-Based Multi-Level Approach to Hierarchical Structural Architecture Reconstruction

DGX agent

arXiv:2605.20055v1 Announce Type: cross Abstract: Explicit software architecture models are essential artifacts for communicating, analyzing, and evolving complex software-intensive systems. In ROS~2-

agentsarxiv-cs-ai
20 May 2026
Safety

TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents

DGX agent

arXiv:2602.11767v3 Announce Type: replace Abstract: Advances in large language models (LLMs) are driving a shift toward using reinforcement learning (RL) to train agents from iterative, multi-turn int

safetyarxiv-cs-ai
20 May 2026
Model Releases

Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap

DGX agent

arXiv:2605.17903v1 Announce Type: new Abstract: We automatically generate feedback causal fuzzy cognitive maps (FCMs) from text by teaching large-language-model agents to break the text into overlappi

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning

DGX agent

arXiv:2605.16309v1 Announce Type: new Abstract: LLM-based agents can recover from individual execution errors, yet they repeatedly fail on the same fault when the underlying process knowledge--operato

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Benchmarking inference at scale: coding agents

DGX agent

This article presents benchmarking results for AI coding agents evaluated at scale, likely comparing performance metrics such as code generation accuracy, execution success rates, and inference effici

model-releasestogether-ai-blog
19 May 2026
Safety

Body-Grounded Perspective Formation and Conative Attunement in Artificial Agents

DGX agent

arXiv:2605.16728v1 Announce Type: new Abstract: This paper proposes a minimal architecture for body-grounded perspective formation in artificial agents. Extending prior work, the model introduces an i

safetyarxiv-cs-ai
19 May 2026
Industry

Can I get my agents on the phone?

DGX agent

This article likely discusses the feasibility and methods of contacting AI agents or customer service representatives by phone, exploring whether voice communication is available as an interaction opt

industryben-s-bites
19 May 2026
Model Releases

Code-as-Room: Generating 3D Rooms from Top-Down View Images via Agentic Code Synthesis

DGX agent

arXiv:2605.18451v1 Announce Type: new Abstract: Designing realistic and functional 3D indoor rooms is essential for a wide range of applications, including interior design, virtual reality, gaming, an

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

ContractBench: Can LLM Agents Preserve Observation Contracts?

DGX agent

arXiv:2605.17281v1 Announce Type: cross Abstract: Tool-augmented LLM agents call APIs whose intermediate outputs, such as presigned URLs, session tokens, and OAuth state parameters, are observation co

model-releasesarxiv-cs-ai
19 May 2026
Safety

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

DGX agent

arXiv:2605.17909v1 Announce Type: new Abstract: As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency p

safetyarxiv-cs-ai
19 May 2026
Model Releases

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

DGX agent

arXiv:2507.21035v3 Announce Type: replace Abstract: Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Google just announced agentic coding in search. Based on your search, Gemini processes your ask, decides whether it should build a new inter…

DGX agent

Google just announced agentic coding in search. Based on your search, Gemini processes your ask, decides whether it should build a new interface, reasons through the steps, and creates a personalized

model-releasesallie-k--miller--x
19 May 2026
Model Releases

LivePI: More Realistic Benchmarking of Agents Against Indirect Prompt Injectio

DGX agent

arXiv:2605.17986v1 Announce Type: cross Abstract: AI agents such as OpenClaw are increasingly deployed in local workflows with access to external tools. This creates indirect prompt-injection (IPI) ri

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Multi-agent AI systems outperform human teams in creativity

DGX agent

arXiv:2605.17885v1 Announce Type: cross Abstract: Although artificial intelligence (AI) now matches or exceeds human performance across numerous cognitive tasks, creativity remains a highly contested

local-aiarxiv-cs-ai
19 May 2026
← Previous
1…148149150151152…375
Next →