AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,724 results
29 Jul 2026

PLATO: Pointer Learner for Agent and Task Openness

SafetyDGX agent

arXiv:2607.25082v1 Announce Type: new Abstract: Open agent systems (OASYS) are increasingly prevalent in real-world domains where the sets of agents and tasks change unpredictably over time. Such open

28 Jul 2026

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents

AgentsDGX agent

arXiv:2607.23588v1 Announce Type: new Abstract: Creative AI is moving from single-step asset generation toward long-horizon multimodal production. Although recent generative models can synthesize high

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2510.24411v3 Announce Type: replace Abstract: Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobi

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interacti…

AgentsDGX agent

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interactive replay (AND how they defended against the attack)! Here i

Where Is the Cost of Third-Party API Routers in Agentic Software Development?

SafetyDGX agent

arXiv:2607.23624v1 Announce Type: cross Abstract: Third-party API routers have become a common layer that unifies access across increasingly diverse LLM providers. In coding-agent workflows, high-auto

27 Jul 2026

InteractComp: Evaluating Search Agents With Ambiguous Queries

Model ReleasesDGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

24 Jul 2026

AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics

AgentsDGX agent

arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments. Thi

Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates (Ina Fried/Axios)

AgentsDGX agent

Ina Fried / Axios: Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates — Meta is giving its AI a

23 Jul 2026

Detecting silent agent failures with Amazon Bedrock AgentCore optimization

AgentsDGX agent

Amazon Bedrock AgentCore optimization surfaces silent behavioral failures in production AI agents: the ones that pass every health check but still deliver wrong outcomes. Learn how insights discovers,

22 Jul 2026

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dens…

AgentsDGX agent

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dense tables, footnoted adjustments, and the details buried in t

15 Jul 2026

Critic Experience Bank: Self-Evolving Step-Level Confidence Estimation for LLM Agents

AgentsDGX agent

arXiv:2607.12397v1 Announce Type: new Abstract: LLM agents act in external environments where each action changes the state that later decisions condition on, and where a single wrong step can waste i

How to Analyze and Govern Gemini Enterprise App Usage at Scale with BigQuery

Model ReleasesDGX agent

Deploying the Gemini Enterprise app across an organization marks a transformative leap forward in workforce productivity, providing employees with an amazing, high-performance suite of agentic AI tool

Tracing Agentic Failure from the Flow of Success

AgentsDGX agent

arXiv:2607.12747v1 Announce Type: new Abstract: Failure attribution for LLM-based agentic systems, i.e., identifying which steps in a failure trajectory caused the task to fail, is critical for debugg

14 Jul 2026

How Retail Finance teams are using Agentic AI to protect omni-channel margins

AgentsDGX agent

Retail Finance teams are increasingly deploying agentic AI to safeguard omni‑channel margins. These systems automatically monitor sales performance, forecast demand, and adjust pricing or inventory co

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

AgentsDGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

10 Jul 2026

Hermes Agent now comes preinstalled on rabbitOS

AgentsDGX agent

Hermes Agent now comes preinstalled on rabbitOS rabbitOS 2.3 is here, with hermes agent 🥕🪽 a fresh OTA is rolling out to r1 now, and this one’s packed: hermes agent, proactive rabbit, openclaw v4, cre

9 Jul 2026

Think Big, Search Small: Where Capacity Matters in Hierarchical Search Agents?

Model ReleasesDGX agent

arXiv:2607.07548v1 Announce Type: new Abstract: Large language model based search agents increasingly adopt multi-agent architectures in which a main agent decomposes a complex question into sub-queri

8 Jul 2026

Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale

AgentsDGX agent

arXiv:2607.06233v1 Announce Type: new Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize

7 Jul 2026

Conflict-Based Search for Multi-Agent Path Finding with Elevators

AgentsDGX agent

arXiv:2602.20512v2 Announce Type: replace Abstract: This paper investigates a problem called Multi-Agent Path Finding with Elevators (MAPF-E), which seeks conflict-free paths for multiple agents from

SkillFab: An Agent-Native Skill Production Platform

AgentsDGX agent

arXiv:2607.03780v1 Announce Type: cross Abstract: SkillFab is an agent-native platform for turning missing capabilities into reviewed, reusable Agent Skills. At runtime, agents first search for reusab

1 Jul 2026

Investigating Multi-Agent Deliberation in Law

AgentsDGX agent

arXiv:2606.30906v1 Announce Type: new Abstract: Artificial Intelligence is increasingly applied to the field of law, and has the potential to increase access to justice. One particular movement that i

26 Jun 2026

Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols

SafetyDGX agent

arXiv:2606.26203v1 Announce Type: new Abstract: As AI agent protocols proliferate, the governance structures shaping their interoperability standards remain empirically underexamined. We introduce an

25 Jun 2026

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents

AgentsDGX agent

arXiv:2606.24893v1 Announce Type: new Abstract: For agents to learn continuously from interaction with the world at test time, they must be able to explore effectively, acquire new world knowledge and

Building agentic AI applications with a modern data mesh strategy on AWS

AgentsDGX agent

This AWS ML Blog article explores how Data Product Agent Mesh makes data mesh practical by solving operational complexity through intelligent automation, grounding agentic AI capabilities in well-defi

24 Jun 2026

CORE-Bench: Fostering the Credibility of Published Research Through a Computational Reproducibility Agent Benchmark

Model ReleasesDGX agent

arXiv:2409.11363v2 Announce Type: replace-cross Abstract: AI agents have the potential to aid users on a variety of consequential tasks, including conducting scientific research. To spur the developme

9 Jun 2026

SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History

AgentsDGX agent

arXiv:2606.08671v1 Announce Type: new Abstract: Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually

5 Jun 2026

Unlocking dependable responses with Gemini Enterprise Agent Platform’s Agentic RAG

Model ReleasesDGX agent

Google's RAG Engine securely connects private enterprise data to LLMs to improve answer accuracy and reduce hallucinations , making it a key component of the Gemini Enterprise Agent Platform for build

4 Jun 2026

Parthenon Law: A Self-Evolving Legal-Agent Framework

AgentsDGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

AgentsDGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

AgentsDGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

2 Jun 2026

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

AgentsDGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

Learning to Construct Practical Agentic Systems

AgentsDGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

1 Jun 2026

Counterfactual Graph for Multi-Agent LLM Calibration

AgentsDGX agent

arXiv:2605.30653v1 Announce Type: new Abstract: Multi-agent LLM systems often treat agreement as evidence: when many agents in a panel give the same answer, that answer is assumed to be more reliable.

Merge launches Agent Handler for Employees as an IT gatekeeper for workplace AI agents

Model ReleasesDGX agent

Merge API Inc., a platform provider delivering connective infrastructure for artificial intelligence to business data and tools, launched Agent Handler for Employees, easing the strain on information

29 May 2026

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

Model ReleasesDGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

AgentsDGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

SafetyDGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

28 May 2026

Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

SafetyDGX agent

arXiv:2602.15198v2 Announce Type: replace-cross Abstract: Multi-agent systems, where LLM agents communicate through free-form language, enable sophisticated coordination for solving complex cooperativ

27 May 2026

Is Agent Memory a Database? Rethinking Data Foundations for Long-Term AI Agent Memory

Local AiDGX agent

arXiv:2605.26252v1 Announce Type: new Abstract: Long-running AI agents need persistent memory. Memory supports learning across sessions, reduces repeated context injection, and enables auditing of pas

Robinhood opens its platform to AI agents for trading and credit card spending

AgentsDGX agent

Robinhood Markets Inc. today opened its trading and banking platforms to artificial intelligence agents, launching a beta Agentic Trading product and a virtual Agentic Credit Card that can place stock

26 May 2026

A Sober Look at Agentic Misalignment in Automated Workflows

SafetyDGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

25 May 2026

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

AgentsDGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

AgentsDGX agent

arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education

Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness

SafetyDGX agent

arXiv:2605.23146v1 Announce Type: cross Abstract: Classical reinforcement learning assumes the agent interacts with a fixed environment whose behavior does not depend on the agent's policy. This assum

21 May 2026

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

Model ReleasesDGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

20 May 2026

Automating what already exists isn’t enough — agents demand a fundamental rethink of the enterprise

AgentsDGX agent

Enterprises have figured out how to stand up AI agents, but agent management is another problem entirely. The agentic moment is forcing organizations to interrogate everything, including the operating

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and …

Model ReleasesDGX agent

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and encounter. They’re not really random benchmark tasks, they r

19 May 2026

VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.17467v1 Announce Type: new Abstract: Large language model-driven multi-agent systems (LLM-MAS) excel at complex tasks, yet unreliable agents remain a key bottleneck to system-level reliabil

18 May 2026

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

Model ReleasesDGX agent

arXiv:2605.15218v1 Announce Type: new Abstract: Large language models deployed for MAPDL finite-element simulation face practical reliability challenges: without structured execution control, tool enc

17 May 2026

Eval engineering: The missing piece of agentic AI governance

AgentsDGX agent

As artificial intelligence agents become more powerful, agentic AI governance becomes increasingly important – and yet, today’s governance solutions struggle to keep AI agents from going off the rails

14 May 2026

Language-Based Agent Control

AgentsDGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

12 May 2026

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

Model ReleasesDGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

Evolutionary Ensemble of Agents

AgentsDGX agent

arXiv:2605.09018v1 Announce Type: cross Abstract: We introduce Evolutionary Ensemble (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving sys

General Agent Evaluation

Model ReleasesDGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents

AgentsDGX agent

arXiv:2605.08580v1 Announce Type: cross Abstract: To cope with the large contexts that long-horizon LLM agents produce, modern frameworks increasingly rely on compaction -- invoking an LLM to rewrite

Token Economics for LLM Agents: A Dual-View Study from Computing and Economics

AgentsDGX agent

arXiv:2605.09104v1 Announce Type: new Abstract: As LLM agents evolve, tokens have emerged as the core economic primitives of Agentic AI. However, their exponential consumption introduces severe comput

11 May 2026

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

Model ReleasesDGX agent

arXiv:2605.06869v1 Announce Type: new Abstract: AI agent research spans a wide spectrum: from RL agents that learn from scratch to foundation model agents that leverage pre-trained knowledge, yet no u

Architecting a resilient, scalable and secure foundation for the agentic era

Model ReleasesDGX agent

Across the public sector, the conversation has shifted; we are no longer just talking about the potential of AI, we are already seeing the impact. While visionary leadership and cultural buy-in are cr

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

Model ReleasesDGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

7 May 2026

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

AgentsDGX agent

arXiv:2506.00886v3 Announce Type: replace Abstract: As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Exi

← Previous
1…2324252627…296
Next →