AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,723 results
Agents

JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents

DGX agent

arXiv:2607.23588v1 Announce Type: new Abstract: Creative AI is moving from single-step asset generation toward long-horizon multimodal production. Although recent generative models can synthesize high

agentsarxiv-cs-cv
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OS-Sentinel: Towards Safety-Enhanced Mobile GUI Agents via Hybrid Validation in Realistic Workflows

DGX agent

arXiv:2510.24411v3 Announce Type: replace Abstract: Computer-using agents powered by Vision-Language Models (VLMs) have demonstrated human-like capabilities in operating digital environments like mobi

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interacti…

DGX agent

Super impressed with @huggingface's breakdown of the AI agent autonomous cyber attack from OpenAI, including technical timeline, & interactive replay (AND how they defended against the attack)! Here i

agentsclem-delangue--x
28 Jul 2026
Safety

Where Is the Cost of Third-Party API Routers in Agentic Software Development?

DGX agent

arXiv:2607.23624v1 Announce Type: cross Abstract: Third-party API routers have become a common layer that unifies access across increasingly diverse LLM providers. In coding-agent workflows, high-auto

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

InteractComp: Evaluating Search Agents With Ambiguous Queries

DGX agent

arXiv:2510.24668v2 Announce Type: replace Abstract: Language agents have demonstrated remarkable potential in web search and information retrieval. However, many search-agent benchmarks assume that us

model-releasesarxiv-cs-cl
27 Jul 2026
Agents

AINTMA: Agentic AI Architecture for Autonomous Test Management with Generative Intelligence, Secure Cloud Communication and Adaptive Quality Analytics

DGX agent

arXiv:2607.20452v1 Announce Type: new Abstract: Modern software quality assurance demands intelligent, autonomous systems capable of adaptive decision-making across distributed cloud environments. Thi

agentsarxiv-cs-ai
24 Jul 2026
Agents

Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates (Ina Fried/Axios)

DGX agent

Ina Fried / Axios: Meta updates Meta AI with Muse Spark 1.1-powered agentic capabilities, connecting to Gmail and Google Calendar to perform tasks like creating daily updates — Meta is giving its AI a

agentstechmeme
24 Jul 2026
Agents

Detecting silent agent failures with Amazon Bedrock AgentCore optimization

DGX agent

Amazon Bedrock AgentCore optimization surfaces silent behavioral failures in production AI agents: the ones that pass every health check but still deliver wrong outcomes. Learn how insights discovers,

agentsaws-ml-blog
23 Jul 2026
Agents

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dens…

DGX agent

Your agentic workflows are only as good as the context you feed them. In finance, that context is stuck inside your messiest documents: dense tables, footnoted adjustments, and the details buried in t

agentsjerry-liu--x
22 Jul 2026
Agents

Critic Experience Bank: Self-Evolving Step-Level Confidence Estimation for LLM Agents

DGX agent

arXiv:2607.12397v1 Announce Type: new Abstract: LLM agents act in external environments where each action changes the state that later decisions condition on, and where a single wrong step can waste i

agentsarxiv-cs-ai
15 Jul 2026
Model Releases

How to Analyze and Govern Gemini Enterprise App Usage at Scale with BigQuery

DGX agent

Deploying the Gemini Enterprise app across an organization marks a transformative leap forward in workforce productivity, providing employees with an amazing, high-performance suite of agentic AI tool

model-releasesgoogle-cloud-ai
15 Jul 2026
Agents

Tracing Agentic Failure from the Flow of Success

DGX agent

arXiv:2607.12747v1 Announce Type: new Abstract: Failure attribution for LLM-based agentic systems, i.e., identifying which steps in a failure trajectory caused the task to fail, is critical for debugg

agentsarxiv-cs-ai
15 Jul 2026
Agents

How Retail Finance teams are using Agentic AI to protect omni-channel margins

DGX agent

Retail Finance teams are increasingly deploying agentic AI to safeguard omni‑channel margins. These systems automatically monitor sales performance, forecast demand, and adjust pricing or inventory co

agentsdatabricks
14 Jul 2026
Agents

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

DGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

agentsfrancois-chollet--x
14 Jul 2026
Agents

Hermes Agent now comes preinstalled on rabbitOS

DGX agent

Hermes Agent now comes preinstalled on rabbitOS rabbitOS 2.3 is here, with hermes agent 🥕🪽 a fresh OTA is rolling out to r1 now, and this one’s packed: hermes agent, proactive rabbit, openclaw v4, cre

agentsnous-research--x
10 Jul 2026
Model Releases

Think Big, Search Small: Where Capacity Matters in Hierarchical Search Agents?

DGX agent

arXiv:2607.07548v1 Announce Type: new Abstract: Large language model based search agents increasingly adopt multi-agent architectures in which a main agent decomposes a complex question into sub-queri

model-releasesarxiv-cs-cl
9 Jul 2026
Agents

Demonstrating TOFFEE: A Learned System for Synthesizing Data Agent Trajectories at Scale

DGX agent

arXiv:2607.06233v1 Announce Type: new Abstract: LLM-powered data agents are playing an increasingly important role in data-driven decision making. However, existing data agents struggle to generalize

agentsarxiv-cs-ai
8 Jul 2026
Agents

Conflict-Based Search for Multi-Agent Path Finding with Elevators

DGX agent

arXiv:2602.20512v2 Announce Type: replace Abstract: This paper investigates a problem called Multi-Agent Path Finding with Elevators (MAPF-E), which seeks conflict-free paths for multiple agents from

agentsarxiv-cs-ro
7 Jul 2026
Agents

SkillFab: An Agent-Native Skill Production Platform

DGX agent

arXiv:2607.03780v1 Announce Type: cross Abstract: SkillFab is an agent-native platform for turning missing capabilities into reviewed, reusable Agent Skills. At runtime, agents first search for reusab

agentsarxiv-cs-ai
7 Jul 2026
Agents

Investigating Multi-Agent Deliberation in Law

DGX agent

arXiv:2606.30906v1 Announce Type: new Abstract: Artificial Intelligence is increasingly applied to the field of law, and has the potential to increase access to justice. One particular movement that i

agentsarxiv-cs-ai
1 Jul 2026
Safety

Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols

DGX agent

arXiv:2606.26203v1 Announce Type: new Abstract: As AI agent protocols proliferate, the governance structures shaping their interoperability standards remain empirically underexamined. We introduce an

safetyarxiv-cs-ai
26 Jun 2026
Agents

AgentOdyssey: Open-Ended Long-Horizon Text Game Generation for Test-Time Continual Learning Agents

DGX agent

arXiv:2606.24893v1 Announce Type: new Abstract: For agents to learn continuously from interaction with the world at test time, they must be able to explore effectively, acquire new world knowledge and

agentsarxiv-cs-cl
25 Jun 2026
Agents

Building agentic AI applications with a modern data mesh strategy on AWS

DGX agent

This AWS ML Blog article explores how Data Product Agent Mesh makes data mesh practical by solving operational complexity through intelligent automation, grounding agentic AI capabilities in well-defi

agentsaws-ml-blog
25 Jun 2026
Model Releases

CORE-Bench: Fostering the Credibility of Published Research Through a Computational Reproducibility Agent Benchmark

DGX agent

arXiv:2409.11363v2 Announce Type: replace-cross Abstract: AI agents have the potential to aid users on a variety of consequential tasks, including conducting scientific research. To spur the developme

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History

DGX agent

arXiv:2606.08671v1 Announce Type: new Abstract: Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually

agentsarxiv-cs-lg
9 Jun 2026
Model Releases

Unlocking dependable responses with Gemini Enterprise Agent Platform’s Agentic RAG

DGX agent

Google's RAG Engine securely connects private enterprise data to LLMs to improve answer accuracy and reduce hallucinations , making it a key component of the Gemini Enterprise Agent Platform for build

model-releasesgoogle-research
5 Jun 2026
Agents

Parthenon Law: A Self-Evolving Legal-Agent Framework

DGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

agentsarxiv-cs-ai
4 Jun 2026
Agents

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

DGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

agentsarxiv-cs-ai
4 Jun 2026
Agents

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

DGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

agentsarxiv-cs-ai
4 Jun 2026
Agents

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

DGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

agentsarxiv-cs-ro
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Agents

Counterfactual Graph for Multi-Agent LLM Calibration

DGX agent

arXiv:2605.30653v1 Announce Type: new Abstract: Multi-agent LLM systems often treat agreement as evidence: when many agents in a panel give the same answer, that answer is assumed to be more reliable.

agentsarxiv-cs-cl
1 Jun 2026
Model Releases

Merge launches Agent Handler for Employees as an IT gatekeeper for workplace AI agents

DGX agent

Merge API Inc., a platform provider delivering connective infrastructure for artificial intelligence to business data and tools, launched Agent Handler for Employees, easing the strain on information

model-releasessiliconangle
1 Jun 2026
Model Releases

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

DGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

model-releasesarxiv-cs-ai
29 May 2026
Agents

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

agentsarxiv-cs-cl
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Safety

Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

DGX agent

arXiv:2602.15198v2 Announce Type: replace-cross Abstract: Multi-agent systems, where LLM agents communicate through free-form language, enable sophisticated coordination for solving complex cooperativ

safetyarxiv-cs-ai
28 May 2026
Local Ai

Is Agent Memory a Database? Rethinking Data Foundations for Long-Term AI Agent Memory

DGX agent

arXiv:2605.26252v1 Announce Type: new Abstract: Long-running AI agents need persistent memory. Memory supports learning across sessions, reduces repeated context injection, and enables auditing of pas

local-aiarxiv-cs-ai
27 May 2026
Agents

Robinhood opens its platform to AI agents for trading and credit card spending

DGX agent

Robinhood Markets Inc. today opened its trading and banking platforms to artificial intelligence agents, launching a beta Agentic Trading product and a virtual Agentic Credit Card that can place stock

agentssiliconangle
27 May 2026
Safety

A Sober Look at Agentic Misalignment in Automated Workflows

DGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

safetyarxiv-cs-ai
26 May 2026
Agents

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

DGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

agentsarxiv-cs-ai
25 May 2026
Agents

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

DGX agent

arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education

agentsarxiv-cs-ai
25 May 2026
Safety

Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness

DGX agent

arXiv:2605.23146v1 Announce Type: cross Abstract: Classical reinforcement learning assumes the agent interacts with a fixed environment whose behavior does not depend on the agent's policy. This assum

safetyarxiv-cs-ai
25 May 2026
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Agents

Automating what already exists isn’t enough — agents demand a fundamental rethink of the enterprise

DGX agent

Enterprises have figured out how to stand up AI agents, but agent management is another problem entirely. The agentic moment is forcing organizations to interrogate everything, including the operating

agentssiliconangle
20 May 2026
Model Releases

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and …

DGX agent

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and encounter. They’re not really random benchmark tasks, they r

model-releasesharrison-chase--x
20 May 2026
Model Releases

VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

DGX agent

arXiv:2605.17467v1 Announce Type: new Abstract: Large language model-driven multi-agent systems (LLM-MAS) excel at complex tasks, yet unreliable agents remain a key bottleneck to system-level reliabil

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

DGX agent

arXiv:2605.15218v1 Announce Type: new Abstract: Large language models deployed for MAPDL finite-element simulation face practical reliability challenges: without structured execution control, tool enc

model-releasesarxiv-cs-ai
18 May 2026
← Previous
1…2930313233…370
Next →