AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,951 results
12 May 2026

MAGS-SLAM: Monocular Multi-Agent Gaussian Splatting SLAM for Geometrically and Photometrically Consistent Reconstruction

Model ReleasesDGX agent

arXiv:2605.10760v1 Announce Type: new Abstract: Collaborative photorealistic 3D reconstruction from multiple agents enables rapid large-scale scene capture for virtual production and cooperative multi

MDrive: Benchmarking Closed-Loop Cooperative Driving for End-to-End Multi-agent Systems

Model ReleasesDGX agent

arXiv:2605.10904v1 Announce Type: new Abstract: Vehicle-to-Everything (V2X) communication has emerged as a promising paradigm for autonomous driving, enabling connected agents to share complementary p

Mem-W: Latent Memory-Native GUI Agents

SafetyDGX agent

arXiv:2605.09317v1 Announce Type: new Abstract: GUI agents are beginning to operate the web, mobile, and desktop as interactive worlds, where successful control depends on carrying forward visual, pro

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PiCA: Pivot-Based Credit Assignment for Search Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2605.09287v1 Announce Type: new Abstract: Large Language Model (LLM)-based search agents trained with reinforcement learning (RL) have significantly improved the performance of knowledge-intensi

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

AgentsDGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

Scaling Mobile Agent Systems: From Capability Density to Collective Intelligence

Local AiDGX agent

arXiv:2605.08124v1 Announce Type: cross Abstract: Mobile agent systems are emerging as a key paradigm for enabling intelligent applications on edge devices and in AIoT ecosystems. However, their scala

SkillRAE: Agent Skill-Based Context Compilation for Retrieval-Augmented Execution

AgentsDGX agent

arXiv:2605.10114v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents (e.g., OpenClaw) increasingly rely on reusable skill libraries to solve artifact-rich tasks such as document-cen

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn…

AgentsDGX agent

The agent also reimplemented the “Blues Improvisation” experiment by @douglas_eck and @SchmidhuberAI in 2002 which show that LSTMs can learn temporal structure in music. Finding temporal structure in

the Mini Shai-Hulud attack is scary because it attacks new AI coding workflows like CI, editor hooks, agent configs, etc

AgentsDGX agent

The Mini Shai-Hulud attack targets emerging AI-assisted development workflows by compromising multiple integration points including continuous integration systems, code editor hooks, and AI agent conf

The scale of the infra on HF is insane. If you're still hosting models, datasets, agent memory,... in S3 or R2, talk to use and we can help …

AgentsDGX agent

Hugging Face offers substantial infrastructure capabilities for hosting machine learning models, datasets, and agent memory systems. The statement suggests that organizations currently using alternati

TMAS: Scaling Test-Time Compute via Multi-Agent Synergy

AgentsDGX agent

arXiv:2605.10344v1 Announce Type: new Abstract: Test-time scaling has become an effective paradigm for improving the reasoning ability of large language models by allocating additional computation dur

🚨 Today: OpenMed Agent ships in preview. Built on @huggingface: → HF endpoints power clinical extraction + terminology → MCP for your own s…

AgentsDGX agent

🚨 Today: OpenMed Agent ships in preview. Built on @huggingface: → HF endpoints power clinical extraction + terminology → MCP for your own services → Every tool call, every plan, fully visible 1,000+ O

ToolScope: Enhancing LLM Agent Tool Use through Tool Merging and Context-Aware Filtering

AgentsDGX agent

arXiv:2510.20036v2 Announce Type: replace Abstract: Large language model (LLM) agents rely on external tools to solve complex tasks, but real-world toolsets often contain redundant tools with overlapp

we just shipped delta channels in langgraph 1.2. as agents run longer and use more context, full-state checkpointing doesn't scale, but delt…

AgentsDGX agent

we just shipped delta channels in langgraph 1.2. as agents run longer and use more context, full-state checkpointing doesn't scale, but delta channel snapshots do. this new algorithm is now powering m

When Independent Sampling Outperforms Agentic Reasoning

AgentsDGX agent

arXiv:2605.08478v1 Announce Type: new Abstract: We study how to allocate inference-time compute for competitive programming under fixed budgets. Evaluating 216 Codeforces problems across Divisions 1-3

11 May 2026

A Multi-Memory Segment System for Generating High-Quality Long-Term Memory Content in Agents

AgentsDGX agent

arXiv:2508.15294v4 Announce Type: replace Abstract: In the current field of agent memory, extensive explorations have been conducted in the area of memory retrieval, yet few studies have focused on ex

ADKO: Agentic Decentralized Knowledge Optimization

Local AiDGX agent

arXiv:2605.07863v1 Announce Type: new Abstract: We present Agentic Decentralized Knowledge Optimization (ADKO), a framework for collaborative black-box optimization across autonomous agents that achie

Agentic Coding Needs Proactivity, Not Just Autonomy

SafetyDGX agent

arXiv:2605.06717v1 Announce Type: cross Abstract: Coding agents are rapidly changing the landscape of software development, moving from inline completion to autonomous systems that edit repositories,

Agentic inference is set to be different than today's inference, and will change compute infrastructure because speed won't matter when humans aren't involved (Ben Thompson/Stratechery)

AgentsDGX agent

Ben Thompson / Stratechery: Agentic inference is set to be different than today's inference, and will change compute infrastructure because speed won't matter when humans aren't involved — Subscribe t

Beyond the Black Box: Interpretability of Agentic AI Tool Use

Model ReleasesDGX agent

arXiv:2605.06890v1 Announce Type: new Abstract: AI agents are promising for high-stakes enterprise workflows, but dependable deployment remains limited because tool-use failures are difficult to diagn

Cursor is now available in Microsoft Teams. Mention @​Cursor in any channel to delegate tasks to an agent or pull information from Cursor in…

AgentsDGX agent

Cursor, an AI coding assistant, has been integrated into Microsoft Teams, allowing users to mention @Cursor in channels to delegate tasks to an AI agent or retrieve information. This integration enabl

Dealing with quirks introduced by switching models doesn't have to be hard -- we recently introduced a 'harness profile' API in Deep Agents …

AgentsDGX agent

Dealing with quirks introduced by switching models doesn't have to be hard -- we recently introduced a 'harness profile' API in Deep Agents as a solution. Profiles adjust system prompts, tool descript

Exponential Sample Complexity Separation between Flat and Hierarchical Agentic Theorem Provers

AgentsDGX agent

arXiv:2602.10512v2 Announce Type: replace Abstract: Agentic theorem provers often introduce intermediate lemmas, proof sketches, or subgoal decompositions before returning to tactic-level search. This

HBEE: Human Behavioral Entropy Engine -- Pre-Registered Multi-Agent LLM Simulation of Peer-Suspicion-Based Detection Inversion

AgentsDGX agent

arXiv:2605.07472v1 Announce Type: cross Abstract: Insider threat detection assumes that an adaptive insider leaves behavioral residue distinguishing them from legitimate users. We test this assumption

LiteGUI: Distilling Compact GUI Agents with Reinforcement Learning

Local AiDGX agent

arXiv:2605.07505v1 Announce Type: new Abstract: Developing lightweight, on-device vision-language GUI agents is essential for efficient cross-platform automated interaction. However, current on-device

LLM-Based Agents for Competitive Landscape Mapping in Drug Asset Due Diligence

Model ReleasesDGX agent

arXiv:2508.16571v4 Announce Type: replace Abstract: In this paper, we describe and benchmark a competitor-discovery component used within an agentic AI system for fast drug asset due diligence. A comp

Q&A with Scott Wu, CEO of Cognition, which hit a $445M revenue run rate in its first 18 months, on his math competition roots, the Devin AI coding agent, more (Jeremy Stern/Colossus)

AgentsDGX agent

Jeremy Stern / Colossus: Q&A with Scott Wu, CEO of Cognition, which hit a $445M revenue run rate in its first 18 months, on his math competition roots, the Devin AI coding agent, more — An afternoon w

Region4Web: Rethinking Observation Space Granularity for Web Agents

Model ReleasesDGX agent

arXiv:2605.07134v1 Announce Type: cross Abstract: Web agents perceive web pages through an observation space, yet its granularity has remained an underexamined design choice. Existing work treats obse

Rethinking Experience Utilization in Self-Evolving Language Model Agents

ResearchDGX agent

arXiv:2605.07164v1 Announce Type: new Abstract: Self-evolving agents improve by accumulating and reusing experience from past interactions. Existing work has largely focused on how experience is const

Self-Programmed Execution for Language-Model Agents

SafetyDGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

SafetyDGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

Try Parallel Agents today at http://Replit.com

ToolsDGX agent

Replit announced the availability of Parallel Agents, a feature that allows developers to run multiple AI agents simultaneously on their platform. This capability enables more efficient and complex ta

10 May 2026

Experian says 40% of the 5,000 data breaches it serviced in 2025 were AI-powered, and predicts agentic AI will be the leading cause of data breaches in 2026 (Jennah Haque/Bloomberg)

AgentsDGX agent

Jennah Haque / Bloomberg: Experian says 40% of the 5,000 data breaches it serviced in 2025 were AI-powered, and predicts agentic AI will be the leading cause of data breaches in 2026 — A few months ag

How Superset built the IDE for AI agents on Vercel

ToolsDGX agent

Superset developed an integrated development environment (IDE) for AI agents deployed on Vercel's platform, enabling developers to build, test, and manage AI agent applications more efficiently. The s

Over the weekend I set up the Hermes agent, and basically force fed it every resource I could find on X to upgrade it. I then asked it to ra…

AgentsDGX agent

Over the weekend I set up the Hermes agent, and basically force fed it every resource I could find on X to upgrade it. I then asked it to rank each resource and provide a simple explanation: So yea, h

Shopify's River agent system lives in Slack and can only be used in public so that other employees can learn from what you do with it Remind…

AgentsDGX agent

Shopify's River agent system lives in Slack and can only be used in public so that other employees can learn from what you do with it Reminds me of how Midjourney's Discord-only launch helped people f

9 May 2026

Gall’s law: “a complex system that works is invariably found to have evolved from a simple system that worked.” also very relevant to agent …

AgentsDGX agent

Gall’s law: “a complex system that works is invariably found to have evolved from a simple system that worked.” also very relevant to agent systems. most teams are trying to jump straight to autonomou

Hermes Agent is now #1 on the Global @OpenRouter token rankings. While our journey together has just begun, we'd like to take this opportuni…

AgentsDGX agent

Hermes Agent is now #1 on the Global @OpenRouter token rankings. While our journey together has just begun, we'd like to take this opportunity to thank our contributors, supporters, and users for all

This means that agentic coding isn't exactly a replacement for software engineering. It is a fundamentally different way of producing softwa…

AgentsDGX agent

This means that agentic coding isn't exactly a replacement for software engineering. It is a fundamentally different way of producing software, with different best practices and different use cases. J

Tip of the day: You can use Hermes Agent's Credential Pools to login or add multiple API keys for the same LLM Provider and it will rotate a…

AgentsDGX agent

Tip of the day: You can use Hermes Agent's Credential Pools to login or add multiple API keys for the same LLM Provider and it will rotate across them to ensure stability of your operations. Learn mor

8 May 2026

i keep coming back to this post -- it enumerates a ton of requirements you need to think about when taking an agent to production, and how t…

AgentsDGX agent

i keep coming back to this post -- it enumerates a ton of requirements you need to think about when taking an agent to production, and how the langgraph runtime is built to address these needs! https:

integrating good system design and thinking, not just into software, but into all agentic ai interactions is a massive design space/opportun…

AgentsDGX agent

integrating good system design and thinking, not just into software, but into all agentic ai interactions is a massive design space/opportunity right now @mattpocockuk with @swyx, @latentspacepod Medi

Pushing the Frontier for Data Agents with Genie

AgentsDGX agent

Databricks' Genie is an AI-powered data agent designed to enable natural language interactions with data, allowing users to query, analyze, and generate insights without requiring SQL or coding expert

'Tis the year of open source LLMs in agents!

AgentsDGX agent

'Tis the year of open source LLMs in agents! .@BraceSproul changed our org's internal model in Fleet from Sonnet 4.6 to Kimi K2.6 and I didn't even notice. Open models are already good enough for most

We've published our internal manual for building agent skills. Skills require a new way of thinking for developers. https://research.perplex…

AgentsDGX agent

We've published our internal manual for building agent skills. Skills require a new way of thinking for developers. https://research.perplexity.ai/articles/designing-refining-and-maintaining-agent-ski

7 May 2026

A few major use cases for agentic coding for me: 1. Adhoc data visualizations. Anytime I have a question that can be answered quantitatively…

AgentsDGX agent

A few major use cases for agentic coding for me: 1. Adhoc data visualizations. Anytime I have a question that can be answered quantitatively, I generate some code to make a plot. 2. Adhoc data annotat

Adaptivity Under Realizability Constraints: Comparing In-Context and Agentic Learning

AgentsDGX agent

arXiv:2605.04995v1 Announce Type: new Abstract: We compare in-context learning with fixed queries and agentic learning with adaptive queries for uniform approximation of task families. We consider two

Announcing the official poster of Hermes Agent Meetup @ Seoul by http://Instruct.KR & Team Attention - Sponsored by @sionic_ai @hashed_offic…

AgentsDGX agent

Announcing the official poster of Hermes Agent Meetup @ Seoul by http://Instruct.KR & Team Attention - Sponsored by @sionic_ai @hashed_official @NousResearch @Teknium thanks for all the support! It sh

Anthropic just shipped sleep into agents. When you sleep, your hippocampus replays the day's neural sequences to the cortex during 150-220 H…

Model ReleasesDGX agent

Anthropic just shipped sleep into agents. When you sleep, your hippocampus replays the day's neural sequences to the cortex during 150-220 Hz bursts called sharp-wave ripples. The replay runs about 20

GEM: Graph-Enhanced Mixture-of-Experts with ReAct Agents for Dialogue State Tracking

AgentsDGX agent

arXiv:2605.04449v1 Announce Type: new Abstract: Dialogue State Tracking (DST) requires precise extraction of structured information from multi-domain conversations, a task where Large Language Models

Hermes Agent can use Autobrowse to make better browser skills. In this HN example, 2 iterations take us from: 102 seconds -> 35 23 turns -> …

AgentsDGX agent

Hermes Agent can use Autobrowse to make better browser skills. In this HN example, 2 iterations take us from: 102 seconds -> 35 23 turns -> 8 1.46 -> 0.28 Instead of clicking step by step it decides t

Learning Correct Behavior from Examples: Validating Sequential Execution in Autonomous Agents

AgentsDGX agent

arXiv:2605.03159v1 Announce Type: new Abstract: As autonomous agents become increasingly sophisticated, validating their sequential behavior presents a significant challenge. Traditional testing appro

Lightpanda is now a browser backend in Hermes by @NousResearch. Open source autonomous agent. Open source browser built for machines. It had…

AgentsDGX agent

Lightpanda is now a browser backend in Hermes by @NousResearch. Open source autonomous agent. Open source browser built for machines. It had to be done. Set Lightpanda as default with automatic Chrome

Nova Intelligence, which is building agentic AI for SAP ahead of a 2030 migration, raised a 31.5M Series A led by Chemistry, taking its total funding to 40M+ (Lily Mae Lazarus/Fortune)

AgentsDGX agent

Lily Mae Lazarus / Fortune: Nova Intelligence, which is building agentic AI for SAP ahead of a 2030 migration, raised a 31.5M Series A led by Chemistry, taking its total funding to 40M+ — An estimated

SkCC: Portable and Secure Skill Compilation for Cross-Framework LLM Agents

Model ReleasesDGX agent

arXiv:2605.03353v1 Announce Type: cross Abstract: LLM-Agents have evolved into autonomous systems for complex task execution, with the SKILL.md specification emerging as a de facto standard for encaps

This builds on our existing research on multi-agent systems, with a key addition: verifiers. Planners spawn workers that write code and veri…

AgentsDGX agent

This builds on our existing research on multi-agent systems, with a key addition: verifiers. Planners spawn workers that write code and verifiers that run it. If verification fails, the planner spawns

What Happens Inside Agent Memory? Circuit Analysis from Emergence to Diagnosis

Model ReleasesDGX agent

arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne

6 May 2026

Agentic Multi-Source Grounding for Enhanced Query Intent Understanding: A DoorDash Case Study

AgentsDGX agent

arXiv:2603.01486v2 Announce Type: replace Abstract: Accurately mapping user queries to business categories is a fundamental Information Retrieval challenge for multi-category marketplaces, where conte

As agents generate more code, review is becoming the bottleneck. You can write a PR in minutes, but understanding whether it's correct, safe…

AgentsDGX agent

As agents generate more code, review is becoming the bottleneck. You can write a PR in minutes, but understanding whether it's correct, safe, and ready to merge still takes longer than writing it. Dev

.@Clay uses LangSmith to manage 300M agent runs a month, with an average 10-30 steps each. @hwchase17’s conversation with Clay’s Head of AI …

AgentsDGX agent

.@Clay uses LangSmith to manage 300M agent runs a month, with an average 10-30 steps each. @hwchase17’s conversation with Clay’s Head of AI @jeffbarg on how they run this at scale → https://youtu.be/c

← Previous
1…8687888990…300
Next →