AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Safety

Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Systems

DGX agent

arXiv:2605.27628v1 Announce Type: new Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remai

safetyarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

LegalGraphRAG: Multi-Agent Graph Retrieval-Augmented Generation for Reliable Legal Reasoning

DGX agent

arXiv:2605.28120v1 Announce Type: cross Abstract: Graph-based Retrieval-Augmented Generation (GraphRAG) advances flat document retrieval by structuring knowledge as relational graphs, enabling more co

agentsarxiv-cs-ai
28 May 2026
Model Releases

MemCog: From Memory-as-Tool to Memory-as-Cognition in Conversational Agents

DGX agent

arXiv:2605.28046v1 Announce Type: new Abstract: Existing agent memory systems universally follow what we term a Memory-as-Tool paradigm where a single query triggers one-shot retrieval of flat passage

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

StoryMI: Steerable Multi-Agent Therapeutic Dialogue Generation

DGX agent

arXiv:2605.27393v1 Announce Type: cross Abstract: Large language models (LLMs) can generate fluent dialogue, but prior works lack situational grounding, dynamic strategy control, and evaluation aligne

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context en…

DGX agent

Took some inspiration from @vboykis and converted my first ever talk into a blog post. I talk about the role of agentic search in context engineering. Together we build an intuition on the strengths a

model-releasesswyx--x
28 May 2026
Agents

AD-H: Language-guided Autonomous Driving with Hierarchical Agents

DGX agent

arXiv:2406.03474v2 Announce Type: replace Abstract: Language-guided autonomous driving requires bridging a large abstraction gap between high-level natural-language instructions and low-level vehicle

agentsarxiv-cs-cv
27 May 2026
Model Releases

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

DGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

Experiments in Agentic AI for Science

DGX agent

arXiv:2605.26305v1 Announce Type: new Abstract: This paper details two novel frameworks for developing autonomous, agentic AI in scientific workflows. Both systems leverage a hybrid Local Body, Remote

local-aiarxiv-cs-ai
27 May 2026
Research

GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL

DGX agent

arXiv:2602.22190v2 Announce Type: replace-cross Abstract: Open-source native GUI agents still lag behind closed-source systems on long-horizon navigation tasks. This gap stems from two limitations: a

researcharxiv-cs-ai
27 May 2026
Model Releases

Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Multi-Agent LLMs

DGX agent

arXiv:2605.26835v1 Announce Type: new Abstract: LLM-based multi-agent systems have been widely adopted for knowledge retrieval and report generation, synthesizing known information through web search

model-releasesarxiv-cs-ai
27 May 2026
Tools

How Conductor moved parallel coding agents from the laptop to the cloud with Vercel Sandbox

DGX agent

Conductor migrated its parallel coding agents infrastructure from local laptops to the cloud using Vercel Sandbox, enabling improved scalability and distributed execution of AI-driven code generation

toolsvercel-blog
27 May 2026
Model Releases

I think Anthropic and OpenAI have found product-market fit

DGX agent

Anthropic are strongly rumored to be about to have their first profitable quarter. Stories are circulating of companies surprised at how expensive their LLM bills are becoming from usage by their staf

model-releasessimon-willison
27 May 2026
Agents

Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction

DGX agent

arXiv:2602.00959v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can be seen as compressed knowledge bases, but it remains unclear what knowledge they truly contain and how far t

agentsarxiv-cs-cl
27 May 2026
Research

RepoMirage: Probing Repository Context Reasoning in Code Agents with Perturbations

DGX agent

arXiv:2605.26177v1 Announce Type: cross Abstract: Code agents are currently having skillful performance on repository-level software engineering benchmarks, but it remains unclear whether success on e

researcharxiv-cs-ai
27 May 2026
Safety

StepOPSD: Step-Aware Online Preference Distillation for Agent Reinforcement Learning

DGX agent

arXiv:2605.27140v1 Announce Type: new Abstract: Reinforcement learning for multi-turn agents suffers from a credit-assignment mismatch: rewards are sparse and trajectory-level, while success often hin

safetyarxiv-cs-ai
27 May 2026
Model Releases

7AI launches PLAID ELITE fully managed agentic security operations service

DGX agent

Agentic artificial intelligence security startup 7AI Inc. today announced the launch of PLAID ELITE, a fully managed AI-native security operations service. The new service combines autonomous investig

model-releasessiliconangle
26 May 2026
Agents

A Multi-Agent LLM Framework for Rating the Quality of Surgical Feedback

DGX agent

arXiv:2605.25440v1 Announce Type: cross Abstract: Verbal feedback delivered by attending surgeons in the operating room plays a critical formative role in resident trainee skill acquisition. Yet, asse

agentsarxiv-cs-ai
26 May 2026
Agents

AgentWatch: Proactive AWS monitoring with ambient agents

DGX agent

In this post, we demonstrate the capabilities of AgentWatch through practical implementation. You will see how the solution performs infrastructure checks every 15 minutes, summarizing CloudWatch metr

agentsaws-ml-blog
26 May 2026
Hardware

As agentic AI surges, CPUs and air-cooled infrastructure move to the fore

DGX agent

Agents are turning up the heat for enterprises, turning air-cooled AI infrastructure into a boardroom priority. While GPUs have largely dominated the AI conversation, CPUs are increasingly coming into

hardwaresiliconangle
26 May 2026
Model Releases

AvalancheBench: Evaluating Enterprise Data Agents Through Latent World Recovery

DGX agent

arXiv:2605.24183v1 Announce Type: cross Abstract: We introduce AvalancheBench, a benchmark for evaluating enterprise data agents through latent world recovery. AvalancheBench improves on existing benc

model-releasesarxiv-cs-ai
26 May 2026
Hardware

Build high-performance generative AI systems with Strands Agents, NVIDIA NIM, and Amazon Bedrock AgentCore

DGX agent

In this post you'll learn how to build a multi-agent campaign review system that demonstrates parallel reasoning, context persistence, and traceable execution paths using an integrated architecture th

hardwareaws-ml-blog
26 May 2026
Model Releases

Can LLMs Time Travel? Enhancing Temporal Consistency in Legal Agentic Search through Reinforcement Learning

DGX agent

arXiv:2605.25920v1 Announce Type: cross Abstract: While large language models (LLMs) augmented with agentic search capabilities show promise for legal reasoning, they overlook a fundamental constraint

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

CUA-Gym: Scaling Verifiable Training Environments and Tasks for Computer-Use Agents

DGX agent

arXiv:2605.25624v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has driven breakthroughs in domains such as math, tool-use, and software engineering, yet its exte

model-releasesarxiv-cs-ai
26 May 2026
Safety

Hide-and-Shill: A Reinforcement Learning Framework for Market Manipulation Detection in Symphony-a Decentralized Multi-Agent System

DGX agent

arXiv:2507.09179v3 Announce Type: replace Abstract: Decentralized finance (DeFi) has introduced a new era of permissionless financial innovation but also led to unprecedented market manipulation. With

safetyarxiv-cs-ai
26 May 2026
Model Releases

Insuring Every Action: An Authority Frontier Framework for Runtime Actuarial Control of Autonomous AI Agents

DGX agent

arXiv:2605.25632v1 Announce Type: new Abstract: Autonomous AI agents increasingly issue side-effect-bearing actions: database mutations, refunds, payments, external commitments. We propose the Actuari

model-releasesarxiv-cs-ai
26 May 2026
Safety

MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents

DGX agent

arXiv:2602.02474v2 Announce Type: replace-cross Abstract: Most Large Language Model (LLM) agent memory systems rely on a small set of static, hand-designed operations for extracting memory. These fixe

safetyarxiv-cs-ai
26 May 2026
Model Releases

MimirRAG: A Multi-Agent RAG Framework for Financial Data Retrieval with Metadata Integration

DGX agent

arXiv:2605.25030v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems offer a promising approach to reduce hallucinations and improve answer accuracy in large language models (L

model-releasesarxiv-cs-lg
26 May 2026
Safety

PolyGnosis 2.0: Enhancing LLM Reasoning via Agentic Harness Engineering for Polymarket and OSINT Insight Extraction

DGX agent

arXiv:2605.25958v1 Announce Type: new Abstract: This paper introduces PolyGnosis 2.0, a pioneering multi-agent architecture designed to extract predictive intelligence by synthesizing Polymarket anoma

safetyarxiv-cs-cl
26 May 2026
Model Releases

Qwen 3.7 Max is now supported in Hermes Agent

DGX agent

Nous Research has added support for Qwen 3.7 Max, a large language model, within their Hermes Agent framework. This integration enables users to leverage Qwen 3.7 Max's capabilities when building or d

model-releasesnous-research--x
26 May 2026
Model Releases

SODE: Analyzing Social Dynamics in LLM Agents

DGX agent

arXiv:2605.23949v1 Announce Type: cross Abstract: As Large Language Models (LLMs) evolve into interactive agents, understanding their behavioral alignment within human social dynamics becomes essentia

model-releasesarxiv-cs-ai
26 May 2026
Safety

Stop Comparing LLM Agents Without Disclosing the Harness

DGX agent

arXiv:2605.23950v1 Announce Type: new Abstract: This position paper argues that, for long-horizon tasks evaluated across models with comparable frontier capability, the agent execution harness, namely

safetyarxiv-cs-ai
26 May 2026
Agents

Technical deep dive: AgentCore payments and innovation in agentic commerce

DGX agent

Amazon Bedrock AgentCore payments is now available in preview, it provides instant payments to paid external services with no manual billing setup per provider, stablecoin support for cost-effective m

agentsaws-ml-blog
26 May 2026
Model Releases

Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development

DGX agent

arXiv:2605.24138v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to software engineering (SE), yet their potential for autonomous, role-oriented collaboration re

model-releasesarxiv-cs-ai
26 May 2026
Model Releases

Agentic-VLA: Efficient Online Adaptation for Vision-Language-Action Models

DGX agent

arXiv:2605.22896v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for robotic manipulation by leveraging pre-trained vision-language representa

model-releasesarxiv-cs-ai
25 May 2026
Tools

Do you have your coding agents include automated tests for the code that they write?

DGX agent

Simon Willison discusses whether coding agents should automatically generate tests alongside the code they produce, addressing a key quality assurance consideration in AI-assisted development. This li

toolssimon-willison--x
25 May 2026
Safety

Energy per Successful Goal: Goal-Level Energy Accounting for Agentic AI Systems

DGX agent

arXiv:2605.22883v1 Announce Type: new Abstract: Current AI energy benchmarks measure consumption at the granularity of a single model invocation or training run. For classical single-turn workloads th

safetyarxiv-cs-ai
25 May 2026
Model Releases

Evaluating Memory Structure in LLM Agents

DGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

model-releasesarxiv-cs-cl
25 May 2026
Safety

IntentScore: Intent-Conditioned Action Evaluation for Computer-Use Agents

DGX agent

arXiv:2604.05157v2 Announce Type: replace Abstract: Computer-Use Agents (CUAs) leverage large language models to execute GUI operations on desktop environments, yet they generate actions without evalu

safetyarxiv-cs-ai
25 May 2026
Agents

KPI2KVI: A Multi Agent Workflow for Calculating Key Value Indicators from Service Descriptions

DGX agent

arXiv:2605.22825v1 Announce Type: cross Abstract: Key Value Indicators (KVIs) provide a decision oriented view of a service by summarizing how operational performance translates into stakeholder value

agentsarxiv-cs-ai
25 May 2026
Agents

NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

DGX agent

arXiv:2602.13473v2 Announce Type: replace Abstract: Although foundation models have demonstrated remarkable success in general domains, the application of these models to electroencephalography (EEG)

agentsarxiv-cs-ai
25 May 2026
Model Releases

PrefBench: Evaluating Zero-Shot LLM Agents in Hidden-Preference Personalized Pricing Negotiations

DGX agent

arXiv:2605.22855v1 Announce Type: cross Abstract: Personalized pricing negotiations are a challenging testbed for LLM agents because successful interaction does not guarantee profitable decision makin

model-releasesarxiv-cs-ai
25 May 2026
Safety

SafeHarbor: Hierarchical Memory-Augmented Guardrail for LLM Agent Safety

DGX agent

arXiv:2605.05704v2 Announce Type: replace-cross Abstract: Recent advances in foundation models have transformed LLMs from passive conversational systems into autonomous agents capable of reasoning and

safetyarxiv-cs-ai
25 May 2026
Research

SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering Agents

DGX agent

arXiv:2602.11210v4 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a key paradigm for training software engineering (SWE) agents, but existing pipelines typically rely on

researcharxiv-cs-lg
23 May 2026
Agents

Amid tool sprawl and agentic fragmentation, the need for an AI operating system has grown critical

DGX agent

As enterprises move from AI experimentation to full-scale production, the absence of a unified enterprise AI operating system is emerging as the single most consequential bottleneck in realizing retur

agentssiliconangle
22 May 2026
Safety

Governance by Design: Architecting Agentic AI for Organizational Learning and Scalable Autonomy

DGX agent

arXiv:2605.20210v1 Announce Type: cross Abstract: Agentic AI systems - systems that can pursue goals through multi-step planning and tool-mediated action with limited direct supervision - are moving f

safetyarxiv-cs-ai
22 May 2026
Agents

ImProver: Agent-Based Automated Proof Optimization

DGX agent

arXiv:2410.04753v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been used to generate formal proofs of mathematical theorems in proofs assistants such as Lean. However, we

agentsarxiv-cs-cl
22 May 2026
Applications

OpenAI named a Leader in enterprise coding agents by Gartner

DGX agent

OpenAI has been recognized as a Leader in Gartner's 2026 evaluation of enterprise coding agents, highlighting its competitive position in AI-powered code generation and software development tools. Thi

applicationsopenai
22 May 2026
Model Releases

Ratchet: A Minimal Hygiene Recipe for Self-Evolving LLM Agents

DGX agent

arXiv:2605.22148v1 Announce Type: cross Abstract: Self-evolving skill libraries, pioneered by Voyager, let frozen LLM agents accumulate reusable knowledge without weight updates, yet recent evaluation

model-releasesarxiv-cs-cl
22 May 2026
← Previous
1…133134135136137…375
Next →