AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,762 results
23 Jun 2026

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now.

AgentsDGX agent

Learn to use the new eve agentic framework from Vercel. Go try out the hands-on labs now. I'm digging the eve agentic framework from Vercel. I like that everything is files, from the tools to the skil

you need docs built for agents

AgentsDGX agent

you need docs built for agents Docs are the eyes and ears of Agents. But we're moving so fast that they are always outdated. So hence an outdated doc also confuses the models. But it's also hard keep

22 Jun 2026

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management syst…

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

context engineering docs for agentic engineering - plans, research, etc SHOULD NOT be stored in version control: A good docs management system keeps them: > outside your repo > accesible to agent via

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

AgentsDGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

11 Jun 2026

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

AgentsDGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

Model ReleasesDGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations

SafetyDGX agent

arXiv:2606.11218v1 Announce Type: cross Abstract: Ethical deliberation is often misunderstood as a search for single right or wrong answers, creating difficulties for non-ethically trained personnel w

FlowBank: Query-Adaptive Agentic Workflows Optimization through Precompute-and-Reuse

AgentsDGX agent

arXiv:2606.11290v1 Announce Type: cross Abstract: Large Language Model (LLM)-based multi-agent systems are increasingly powerful, but current agentic workflow optimization paradigms make an unsatisfyi

Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security

AgentsDGX agent

arXiv:2606.11671v1 Announce Type: cross Abstract: Agent skills let LLM agents reuse instructions, resources, tools, and workflows, but they also create a new place for malicious behavior to hide. A sk

10 Jun 2026

STAGE-Claw: Automated State-based Agent Benchmarking for Realistic Scenarios

Model ReleasesDGX agent

arXiv:2606.10394v1 Announce Type: new Abstract: Large language models are increasingly used to power personal agents for everyday applications, but evaluating these agents remains a challenge. Existin

Stop hand-tuning kernels: How Neuron Agentic Development accelerates AWS Trainium optimizations

AgentsDGX agent

Today, we’re announcing the Neuron Agentic Development capabilities: a collection of AI agents and skills that make this possible for developers building on AWS Trainium and AWS Inferentia. In this po

Visa partners with OpenAI to let AI agents make payments for users

AgentsDGX agent

Visa Inc. has struck a deal with OpenAI Group PBC to let artificial intelligence agents make payments for users, bringing one of the world’s largest payment networks into ChatGPT’s push toward agentic

9 Jun 2026

Claw-R1: A Step-Level Data Middleware System for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.09138v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has become an important post-training paradigm for turning LLMs from static chatbots into interactive agents, giving

DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

AgentsDGX agent

arXiv:2605.22781v2 Announce Type: replace-cross Abstract: LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid chec

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

AgentsDGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

8 Jun 2026

Act As a Real Researcher: A Suite of Benchmarks Evaluating Frontier LLMs and Agentic Harnesses in Research Lifecycle

Model ReleasesDGX agent

arXiv:2606.07462v1 Announce Type: new Abstract: As foundation models advance and agent scaffolding becomes increasingly sophisticated, agents have demonstrated remarkable proficiency in complex, long-

Deep Agents explained in <90 seconds by @sydneyrunkle

AgentsDGX agent

Deep Agents are an AI concept that Sydney Runkle explains concisely in under 90 seconds, likely covering how agents can be designed to operate with deeper reasoning and decision-making capabilities. T

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://try…

AgentsDGX agent

Now you can text Hermes on iMessage 💬 It's just that simple. Photon brings the BEST agents to everyone where they already are -> http://tryphoton.ai Your Hermes Agent now lives in iMessage via @photon

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

AgentsDGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

When Does Multi-Agent Collaboration Help? An Entropy Perspective

AgentsDGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

7 Jun 2026

Snowflake, Databricks and the model makers: The battle for the agentic client and AI back end

AgentsDGX agent

Agentic artificial intelligence is being misread as a set of separate battles – for example, Snowflake Inc. versus Databricks Inc., copilots versus agents, model makers versus application vendors. We

6 Jun 2026

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

SafetyDGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

Hermes Agent is for the Artists

AgentsDGX agent

Hermes Agent, developed by Nous Research, is an AI agent designed specifically for creative professionals and artists. The tool likely focuses on automating tasks, enhancing workflows, or providing sp

Hermes Agent v0.16.0 - “The Surface Release” Changelog below:

AgentsDGX agent

Hermes Agent v0.16.0, released by Nous Research and dubbed 'The Surface Release,' represents an update to their Hermes Agent framework. This version likely includes bug fixes, performance improvements

SentinelBench: A Benchmark for Long-Running Monitoring Agents

Model ReleasesDGX agent

arXiv:2606.05342v1 Announce Type: new Abstract: AI agents are increasingly asked to carry out work that spans minutes, hours, or longer. Yet the default model of agent behavior is continuous action: i

The good news: agentic is leading to lots of new apps! The bad news: ain’t nobody adopting them. Slop FTL [for the loss]

AgentsDGX agent

Gary Marcus discusses a paradox in the agentic AI market: while developers are creating numerous new applications powered by agentic AI systems, these applications are failing to achieve meaningful us

5 Jun 2026

MARDoc: A Memory-Aware Refinement Agent Framework for Multimodal Long Document QA

AgentsDGX agent

arXiv:2606.05749v1 Announce Type: new Abstract: Iterative retrieval-reasoning agents have recently shown promise for multimodal long-document question answering. However, most existing systems maintai

SkillComposer: Learning to Evolve Agent Skills for Specification and Generalization

AgentsDGX agent

arXiv:2606.06079v1 Announce Type: new Abstract: Agent skills, which consist of reusable strategies that guide agent reasoning and action, have shown strong potential for improving model capability at

4 Jun 2026

Agentic commerce puts data quality at the center of retail’s next era

AgentsDGX agent

The rise of agentic commerce is forcing retailers to rethink their entire data strategy — not as a future concern, but as the immediate prerequisite for competing in a world where AI agents make purch

Call your Hermes Agent

AgentsDGX agent

This post likely announces or describes how to invoke or utilize a Hermes Agent, possibly related to function calling, tool use, or agentic capabilities within the Hermes model framework developed by

Cascading Hallucination in Agentic RAG: The CHARM Framework for Detection and Mitigation

AgentsDGX agent

arXiv:2606.04435v1 Announce Type: new Abstract: Multi-step agentic retrieval-augmented generation (RAG) pipelines have demonstrated significant capability for complex reasoning tasks, yet remain vulne

Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling

SafetyDGX agent

arXiv:2606.05021v1 Announce Type: new Abstract: We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algor

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

AgentsDGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

Why the semantic layer is becoming the foundation for trusted agentic AI

AgentsDGX agent

Agentic AI is making the semantic layer an essential enterprise priority because headless agents asking thousands of questions simultaneously have zero tolerance for inconsistent data definitions. Thi

3 Jun 2026

Adaptive Latent Agentic Reasoning

AgentsDGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

Build 2026: From observability to ROI for AI agents on any framework

AgentsDGX agent

9 min read · June 3, 2026 · Sebastian Kohlmeier Shipping an AI agent is the easy part. Keeping it accurate, safe, and accountable in production is where teams get stuck. Agents are non-deterministic.

eMEM: A Hybrid Spatio-Temporal Memory System For Embodied Agents

Model ReleasesDGX agent

arXiv:2606.03374v1 Announce Type: new Abstract: We present eMEM (Embodied Memory), a hybrid graph-based memory system for embodied agents operating in physical environments. Current agent memory archi

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also …

AgentsDGX agent

I'm finally launching this agent tomorrow! Will be free, open source, and powered exclusively by open models on @togethercompute. Will also drop a full guide on how it works! Building an agent that ca

New research from Google. Just shows the impressive results you can get from custom agent harnesses. LEAP wraps a general-purpose LLM in an …

AgentsDGX agent

New research from Google. Just shows the impressive results you can get from custom agent harnesses. LEAP wraps a general-purpose LLM in an agentic scaffold that grounds every step in the Lean compile

OpenAgenet/OAN: Open Infrastructure for Trusted Agent Interconnection

AgentsDGX agent

arXiv:2606.03161v1 Announce Type: cross Abstract: OpenAgenet, abbreviated as OAN, is an open infrastructure project for trusted Agent interconnection. It addresses a problem that becomes visible when

OpenAgenet/OAN: Technical Architecture for Trust-Governed Agent Identity and Discovery

AgentsDGX agent

arXiv:2606.03163v1 Announce Type: cross Abstract: This paper describes the technical architecture of OpenAgenet / OAN. OAN is a protocol-neutral trust layer for open Agent interconnection. It specifie

The DeepSpeak-Agentic Dataset

Model ReleasesDGX agent

arXiv:2606.03686v1 Announce Type: new Abstract: We present DeepSpeak-Agentic, a dataset of videos comprising over 37 hours of semi-structured conversations between a human and an embodied AI agent. We

Traj-Evolve: A Self-Evolving Multi-Agent System for Patient Trajectory Modeling in Lung Cancer Early Detection

AgentsDGX agent

arXiv:2606.02812v1 Announce Type: new Abstract: Modeling patient trajectories from longitudinal electronic health records (EHRs) requires reasoning over sparse, noisy, and long-context multimodal sequ

2 Jun 2026

ANDES: Agent Native Data Evolving Synthesis Tool for Autonomous Instruction Alignment

SafetyDGX agent

arXiv:2606.01279v1 Announce Type: new Abstract: AI agents are increasingly being tasked with automating AI research itself, particularly the critical post-training phase that transforms base LLMs into

CAREAgent: Clinical Agent with Structured Reasoning and Tool-Integrated for Order Generation

AgentsDGX agent

arXiv:2606.01094v1 Announce Type: new Abstract: Clinical order generation serves as a critical bridge between clinical decision-making and real-world practice, translating medical decisions into concr

CoMIC: Collaborative Memory and Insights Circulation for Long-Horizon LLM Agents in Cloud-Edge Systems

Model ReleasesDGX agent

arXiv:2606.00756v1 Announce Type: new Abstract: Deploying lightweight Large Language Model (LLM) agents on edge servers can reduce latency and move agentic services closer to users, but resource-const

Dynamic Trust-Aware Sparse Communication Topology for LLM-Based Multi-Agent Consensus

AgentsDGX agent

arXiv:2606.01828v1 Announce Type: cross Abstract: Large language model-driven multi-agent systems enhance the reliability of complex reasoning tasks through multi-round deliberation, role specializati

Fleet offers the most secure computer use experience, ensuing your can let your agent access private resources, without worrying about it mi…

AgentsDGX agent

Fleet offers the most secure computer use experience, ensuing your can let your agent access private resources, without worrying about it mishandling or leaking secrets Giving agents access to sensiti

Foundry IQ: Build smarter agents faster with unified knowledge and serverless retrieval

AgentsDGX agent

Learn how Foundry IQ helps developers ground agents with unified enterprise knowledge, serverless retrieval, improved agentic retrieval quality, and production-ready security. The post Foundry IQ: Bui

MACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment

AgentsDGX agent

arXiv:2312.03644v3 Announce Type: replace Abstract: Offline Multi-agent Reinforcement Learning (MARL) is valuable in scenarios where online interaction is impractical or risky. While independent learn

Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

AgentsDGX agent

arXiv:2606.00408v1 Announce Type: cross Abstract: Long-horizon search agents accumulate large amounts of retrieved content across many tool calls, making context-budget efficiency increasingly importa

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

Model ReleasesDGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

1 Jun 2026

Build → Test → Deploy → Monitor @hwchase17 on the agent development lifecycle: https://www.langchain.com/blog/the-agent-development-lifecycl…

AgentsDGX agent

This post outlines the agent development lifecycle, consisting of four key phases: Build, Test, Deploy, and Monitor. The content covers best practices and methodologies for developing AI agents using

Scaling Multi-Agent Environment Co-Design with Diffusion Models

SafetyDGX agent

arXiv:2511.03100v2 Announce Type: replace-cross Abstract: The agent-environment co-design paradigm jointly optimises agent policies and environment configurations in search of improved system performa

Skill Reuse as Compression in Agentic RL

AgentsDGX agent

arXiv:2605.31509v1 Announce Type: cross Abstract: Large language model agents trained with reinforcement learning (RL) often learn brittle, task-specific shortcuts. We hypothesize that agents generali

29 May 2026

Estimating the Empowerment of Language Model Agents

AgentsDGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

GenClaw: Code-Driven Agentic Image Generation

AgentsDGX agent

arXiv:2605.30248v1 Announce Type: new Abstract: Image generation models have evolved from text-conditioned pixel synthesis toward multimodal agents endowed with visual comprehension and tool invocatio

How Consistent Are LLM Agents? Measuring Behavioral Reproducibility in Multi-Step Tool-Calling Pipelines

AgentsDGX agent

arXiv:2605.28840v1 Announce Type: cross Abstract: Large language model (LLM) agents with tool-calling capabilities are increasingly deployed in production systems, yet a fundamental reliability questi

How to build a better agent harness with traces and evals

AgentsDGX agent

Agents are easy to prototype and hard to improve. A repeatable loop of traces, evals, failed-span inspection, and targeted harness changes makes agent behavior easier to debug and improve. The post Ho

28 May 2026

Adaptive Multimodal Agents-Based Framework for Automatic Workflow Execution

AgentsDGX agent

arXiv:2605.28607v1 Announce Type: new Abstract: Modern information systems require autonomous agents capable of navigating complex workflows, yet current methodologies often struggle with the transiti

← Previous
1…4142434445…297
Next →