AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
21 Jun 2026

We just passed 1500 Contributors to Hermes Agent's repo! Thank you to all the contributors and developers!

AgentsDGX agent

Nous Research announced that the Hermes Agent repository has reached 1,500 contributors, celebrating the collaborative effort of developers who have contributed to the project. This milestone reflects

11 Jun 2026

Can AI Agents Synthesize Scientific Conclusions?

Model ReleasesDGX agent

arXiv:2606.11337v1 Announce Type: new Abstract: Scientific AI agents increasingly retrieve evidence, reason across sources, and synthesize conclusions used in consequential decisions. Yet, their abili

Counterexample Guided Learning in the Large using Reasoning Agents

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.11521v1 Announce Type: new Abstract: LLMs and LLM agents should improve when given feedback, but identifying when they are able to do so is difficult: feedback is heterogeneous, domain-spec

Goal-Autopilot: A Verifiable Anti-Fabrication Firewall for Unattended Long-Horizon Agents

AgentsDGX agent

arXiv:2606.11688v1 Announce Type: cross Abstract: Long-horizon LLM agents are not trusted to run unattended: with no human watching, they confidently report success they never verified. We treat hones

Improving Generalization and Data Efficiency with Diffusion in Offline Multi-agent RL

SafetyDGX agent

arXiv:2307.01472v2 Announce Type: replace Abstract: We present a novel Diffusion Offline Multi-agent Model (DOM2) for offline Multi-Agent Reinforcement Learning (MARL). Different from existing algorit

Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness

Local AiDGX agent

arXiv:2606.11686v1 Announce Type: cross Abstract: End-to-end task-success is the dominant way to evaluate LLM agents, but one aggregate number tells you that an agent regressed, not where. We present

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

AgentsDGX agent

arXiv:2606.11897v1 Announce Type: new Abstract: Scientific discovery workflows usually contain and rely heavily on lab notes, where researchers record observations, interpret uncertain results, and pl

Reminder that profiles are still one of the most powerful features in Hermes Agent that not everyone seems to have fully grasped. It's one o…

AgentsDGX agent

Reminder that profiles are still one of the most powerful features in Hermes Agent that not everyone seems to have fully grasped. It's one of those things that's obvious in hindsight, but yeah - tryin

WorldReasoner: Evaluating Whether Language Model Agents Forecast Events with Valid Reasoning

Model ReleasesDGX agent

arXiv:2606.11816v1 Announce Type: cross Abstract: Forecasting real-world events requires language-model agents to reason under uncertainty from incomplete, time-bounded information. Yet evaluating whe

10 Jun 2026

AI agents have become surprisingly good at talking to servers 🤖 The next challenge is getting them to interact with the environment users a…

AgentsDGX agent

AI agents have become surprisingly good at talking to servers 🤖 The next challenge is getting them to interact with the environment users actually live in. 🧑‍💻 Browsers. 🖥️ Apps. 📱 Devices. 🌐 Local st

Exclusive: MotherDuck adds agentic data ingestion to its cloud analytics service

AgentsDGX agent

MotherDuck Corp., the maker of a cloud-native data warehouse based on the open-source DuckDB analytical engine, is betting that artificial intelligence agents will reshape how data pipelines are built

Introducing @PoeticHQ: a new AI system that executes complex multi-hour tasks with 99%+ accuracy and 10x fewer tokens than agents. We raised…

AgentsDGX agent

Introducing @PoeticHQ: a new AI system that executes complex multi-hour tasks with 99%+ accuracy and 10x fewer tokens than agents. We raised 50M at 500M from Kleiner Perkins, Founders Fund, First Harm

Introducing Write Gate in Hermes Agent. Now you have the capability to be able to approve/deny memory updates, skill updates, and skill crea…

AgentsDGX agent

Introducing Write Gate in Hermes Agent. Now you have the capability to be able to approve/deny memory updates, skill updates, and skill creation with the same familiar mechanisms as approving dangerou

Most AI agents reset every session. @JenovaAIAgent's don't. Longest session on their platform: 16M tokens. All of it retrievable in <10ms vi…

AgentsDGX agent

Most AI agents reset every session. @JenovaAIAgent's don't. Longest session on their platform: 16M tokens. All of it retrievable in <10ms via Pinecone vector retrieval. Result: Fast ramp to $1M+ ARR,

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now…

AgentsDGX agent

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now. TL;DR from @latentspacepod: • Model Labs compete on capabi

WebChallenger: A Reliable and Efficient Generalist Web Agent

Model ReleasesDGX agent

arXiv:2606.10423v1 Announce Type: new Abstract: Autonomous web navigation remains challenging for LLM agents, and the strongest generalist systems rely on proprietary reasoning models whose inference

9 Jun 2026

A Multi-Agent System for IPMSM Design Optimization via an FEA-AI Hybrid Approach

AgentsDGX agent

arXiv:2606.09037v1 Announce Type: new Abstract: Interior permanent magnet synchronous motor (IPMSM) design requires balancing conflicting objectives and multi-physics constraints, while modern optimiz

Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading Systems

AgentsDGX agent

arXiv:2606.08285v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems are increasingly proposed for financial trading, yet their reported performance remains difficult to co

Build an agentic incident triage assistant with Amazon Quick and New Relic

AgentsDGX agent

This post shows engineering teams how to apply that principle to one of the most time-sensitive workflows in engineering: incident triage. You will build a custom incident triage assistant agent using

Data Agents Under Attack: Vulnerabilities in LLM-Driven Analytical Systems

SafetyDGX agent

arXiv:2606.08661v1 Announce Type: cross Abstract: Data agents integrate LLM-driven reasoning with relational data access, executable analytical tools, and multi-step workflow orchestration, making the

From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG

AgentsDGX agent

arXiv:2603.03292v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) exhibit high reasoning capacity in medical question-answering, but their tendency to produce hallucinations and o

MAVIS: Multi-Agent Video Retrieval via Structured Video Understanding

AgentsDGX agent

arXiv:2606.09641v1 Announce Type: new Abstract: The dominant paradigm in video retrieval relies on embedding-based full-corpus scanning, which suffers from inherent computational inefficiency and the

Observability for Delegated Execution in Agentic AI Systems

AgentsDGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

SafetyDGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

AgentsDGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

SKILL.nb: Selective Formalization and Gated Execution for Durable Agent Workflows

AgentsDGX agent

arXiv:2606.08049v1 Announce Type: new Abstract: AI agents increasingly turn past experience into reusable artifacts such as code, workflows, and procedural memories. Reuse can improve efficiency, but

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

SafetyDGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

VESTA: A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents

SafetyDGX agent

arXiv:2606.08531v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly evolving from simple text-based interaction systems into LLM agents that can maintain memory, use tools, a

ViMax: Agentic Video Generation

AgentsDGX agent

arXiv:2606.07649v1 Announce Type: cross Abstract: Long-form video generation requires systematic narrative planning and visual consistency that current short-clip methods cannot provide. Existing meth

Visual Para-Thinker++: A Single-Policy Multi-Agent Framework for Visual Reasoning

SafetyDGX agent

arXiv:2606.09290v1 Announce Type: new Abstract: Visual reasoning requires integrating evidence distributed across regions, attributes, and relations, making single-chain reasoning prone to early perce

8 Jun 2026

A year ago the closest thing we had to an AI agent was o3.

AgentsDGX agent

One year prior to this post, o3 represented the most advanced AI agent available, marking a significant milestone in AI development. The statement reflects how rapidly AI capabilities have evolved, wi

Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows (Hartley Charlton/MacRumors)

AgentsDGX agent

Hartley Charlton / MacRumors: Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows — Apple today

EvoClaw: Evaluating AI Agents on Continuous Software Evolution

Model ReleasesDGX agent

arXiv:2603.13428v2 Announce Type: replace-cross Abstract: With AI agents increasingly deployed as long-running systems, it becomes essential to autonomously construct and continuously evolve customize

GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning

AgentsDGX agent

arXiv:2606.06532v1 Announce Type: new Abstract: Despite significant progress in agentic long video understanding, existing methods still lack detailed motion comprehension coupled with an efficient me

Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA

AgentsDGX agent

arXiv:2603.24481v2 Announce Type: replace Abstract: Miscalibrated confidence scores are a practical obstacle to deploying AI in clinical settings. A model that is always overconfident offers no useful

Self-evolving LLM agents with in-distribution Optimization

SafetyDGX agent

arXiv:2606.07367v1 Announce Type: new Abstract: Large Language Models (LLMs) have recently emerged as powerful controllers for interactive agents in complex environments, yet training them to perform

Signal-Driven Observation for Long-Horizon Web Agents

AgentsDGX agent

arXiv:2606.06708v1 Announce Type: new Abstract: Web agents operating over long horizons ingest raw DOM and accessibility trees -- routinely tens of thousands of tokens -- at every action step, causing

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

Model ReleasesDGX agent

arXiv:2606.07297v1 Announce Type: cross Abstract: Repository-level coding benchmarks such as SWE-bench have driven a rapid surge in the capabilities of coding agents. Yet they usually treat coding tas

TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

AgentsDGX agent

arXiv:2606.07054v1 Announce Type: cross Abstract: Autonomous LLM agents can pursue hidden malicious objectives through sequences of individually benign actions, making sabotage difficult to detect usi

6 Jun 2026

2-Step Agent: A Framework for the Interaction of a Decision Maker with AI Decision Support

AgentsDGX agent

arXiv:2602.21889v2 Announce Type: replace Abstract: Predictions from ML models support human decision making in several fields, including high-stakes ones such as healthcare and the judiciary. Yet, we

A super useful feature I like that we have in Grok Build: we load your .envrc and pass it straight into the agent’s shell environment. Same …

AgentsDGX agent

A super useful feature I like that we have in Grok Build: we load your .envrc and pass it straight into the agent’s shell environment. Same API keys, PATH, sccache, remote caches, .etc — everything yo

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads

Model ReleasesDGX agent

arXiv:2606.06448v1 Announce Type: new Abstract: LLM agents are increasingly deployed on long-horizon tasks requiring sustained reasoning over extended interaction histories. Realizing this at scale re

and here's what a coding agent on activegraph looks like basically you always get a trace and a graph automatically

AgentsDGX agent

This post describes the automatic features of a coding agent built on ActiveGraph, highlighting that developers receive both an execution trace and a visual graph representation without additional con

Open-source AI Agent Security resource , fully on the @LangChain 🦜 stack. It shows real attacks (prompt injection, Indirect prompt injectio…

AgentsDGX agent

Open-source AI Agent Security resource , fully on the @LangChain 🦜 stack. It shows real attacks (prompt injection, Indirect prompt injection, tenant data exfiltration, memory poisoning) and the securi

SciVisAgentSkills: Design and Evaluation of Agent Skills for Scientific Data Analysis and Visualization

Model ReleasesDGX agent

arXiv:2606.05525v1 Announce Type: new Abstract: Recent advances in agentic visualization have enabled the translation of natural language into executable scientific visualization (SciVis) workflows. W

Updated the skills hub integration search in the Hermes Agent Dashboard to be way more comprehensive and give you the info you need.

AgentsDGX agent

The Hermes Agent Dashboard's skills hub integration search functionality has been enhanced to provide more comprehensive results and improved information delivery to users. This update, shared by Nous

5 Jun 2026

Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding

AgentsDGX agent

arXiv:2512.05774v2 Announce Type: replace-cross Abstract: Long video understanding (LVU) is challenging because answering real-world queries often depends on sparse, temporally dispersed cues buried i

AI agent web traffic has surpassed that of humans, lending weight to the ‘dead internet’ theory

AgentsDGX agent

Cloudflare Inc. co-founder and Chief Executive Matthew Prince says that artificial intelligence agents now drive the bulk of the internet’s traffic, having surpassed human web activity for the first t

Dynamic Multi-Agent Pickup and Delivery in Robotic Cellular Warehousing Systems

AgentsDGX agent

arXiv:2606.05669v1 Announce Type: new Abstract: Robotic Cellular Warehousing Systems (RCWS) give rise to multi-agent pickup and delivery (MAPD) processes in which robots sequentially collect multiple

Give Cursor visual prompts to shrink the gap between what you see and what the agent understands. Learn more: http://cursor.com/blog/design-…

AgentsDGX agent

Cursor's visual prompts feature helps bridge the understanding gap between user interface elements and AI agent perception by providing clearer visual cues and contextual information. This enhancement

Massive output uptick due to agentic AI. Complete flat adoption.

AgentsDGX agent

Jeremy Howard discusses a significant increase in output capabilities driven by agentic AI systems, while noting that adoption rates remain completely flat across the board. The observation suggests a

NitroGen just won CVPR Best Paper Honorable Mention!! We are making strides towards general-purpose embodied agents that master not only the…

AgentsDGX agent

NitroGen just won CVPR Best Paper Honorable Mention!! We are making strides towards general-purpose embodied agents that master not only the real world physics, but also all possible physics across a

4 Jun 2026

Can Generalist Agents Automate Data Curation?

Model ReleasesDGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

Episodic Memory Temporal Consistency for Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2606.04492v1 Announce Type: new Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) frequently suffers from severe reward sparsity and exploration bottlenecks. While episodic memory

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

SafetyDGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

Model ReleasesDGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

SafetyDGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

Great summy of the future of agents from @hwchase17 and @LangChain 🧵

AgentsDGX agent

Harrison Chase, co-founder of LangChain, shares insights on the future development and evolution of AI agents in this Twitter thread. The post likely covers emerging trends, technical capabilities, an

London- and NY-based Airspeed, which aims to use AI agents to replace sales software like traditional CRM dashboards, raised a $20M Series A led by DN Capital (Mike Butcher/Pathfounders)

AgentsDGX agent

Mike Butcher / Pathfounders: London- and NY-based Airspeed, which aims to use AI agents to replace sales software like traditional CRM dashboards, raised a 20M Series A led by DN Capital — Airspeed, t

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running cod…

Model ReleasesDGX agent

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running coding agents. Today we're shipping Nemotron 3 Ultra. A 550B Mo

← Previous
1…6869707172…299
Next →