AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,957 results
Hardware

Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit

DGX agent

The NVIDIA BioNeMo Agent Toolkit enables AI scientists to use structure prediction, molecular generation, docking, sequence analysis, design, and genomics as callable tools , allowing AI agents to exe

hardwarenvidia-developer
23 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures

DGX agent

arXiv:2606.19380v2 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly

safetyarxiv-cs-lg
23 Jun 2026
Agents

Drowning in Routine: Signal Dilution in Multi-Turn Agent Training

DGX agent

arXiv:2606.22164v1 Announce Type: new Abstract: Multi-turn agents interleave consequential decisions with routine execution: some actions change the downstream return distribution, while others are ne

agentsarxiv-cs-lg
23 Jun 2026
Agents

GLM 5.2 is already top 10 on OpenRouter this week, sitting right next to Opus 4.8. it's also been really strong in Deep Agents across long-h…

DGX agent

GLM 5.2 is already top 10 on OpenRouter this week, sitting right next to Opus 4.8. it's also been really strong in Deep Agents across long-horizon coding, big context, tool use, and verifying work bef

agentsharrison-chase--x
23 Jun 2026
Agents

Model council as a tool to push performance definitely works, I think the interesting next frontier is trying to scale this at the agent lev…

DGX agent

Model council as a tool to push performance definitely works, I think the interesting next frontier is trying to scale this at the agent level have been thinking a bunch about model routing and relate

agentsharrison-chase--x
23 Jun 2026
Agents

🧠Self-Harness: Harnesses that improve themselves New paper on agents shaping their own harnesses to improve over time. Not from LangChain, …

DGX agent

🧠Self-Harness: Harnesses that improve themselves New paper on agents shaping their own harnesses to improve over time. Not from LangChain, but builds on top of DeepAgents! Three key steps: 1/ Weakness

agentsharrison-chase--x
23 Jun 2026
Agents

self-harness: harnesses that improve themselves spoiler alert, it's all loop engineering! 1. weakness mining - run agent and observe failure…

DGX agent

self-harness: harnesses that improve themselves spoiler alert, it's all loop engineering! 1. weakness mining - run agent and observe failures 2. propose improvements to the harness 3. confirm improvem

agentsharrison-chase--x
23 Jun 2026
Agents

Want to get into Hermes Agent but don't know where to start? Perhaps you're not ready to invest multiple hours of listening to my voice (unt…

DGX agent

Want to get into Hermes Agent but don't know where to start? Perhaps you're not ready to invest multiple hours of listening to my voice (unthinkable!), so I made a quick 10 minute video to get you up

agentsnous-research--x
23 Jun 2026
Agents

Bring your hot takes and your drop shots.🔥 Next week, we're co-hosting The Agent Open: an afternoon of pickleball, food, drinks, and indust…

DGX agent

Bring your hot takes and your drop shots.🔥 Next week, we're co-hosting The Agent Open: an afternoon of pickleball, food, drinks, and industry-leading speakers who ace their code and commit to their se

agentsjerry-liu--x
22 Jun 2026
Agents

Cool side fact about GardenCam: it was my first time driving a project mainly via an agent. Between the launch of the site and the conclusio…

DGX agent

Cool side fact about GardenCam: it was my first time driving a project mainly via an agent. Between the launch of the site and the conclusion last week, I had 4 talks across both coasts and basically

agentsnous-research--x
22 Jun 2026
Agents

Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the perfo…

DGX agent

Introducing Sakana Fugu: A full multi-agent orchestration system accessible via a single model API. Our ‘Fugu Ultra’ model matches the performance of Fable and Mythos, delivering frontier capability w

agentsdavid-ha--x
22 Jun 2026
Agents

Join us Tuesday, June 30th, for the inaugural Agent Open pickleball tournament! 🏟️ Our co-founder, @travers00, will join leaders from @Prag…

DGX agent

Join us Tuesday, June 30th, for the inaugural Agent Open pickleball tournament! 🏟️ Our co-founder, @travers00, will join leaders from @Pragmatic_Eng, @braintrust, @llama_index, @turbopuffer, and @moda

agentsjerry-liu--x
22 Jun 2026
Agents

Want to chat with your Hermes Agent on iMessage without a Mac? The latest v0.17.0 update made this possible (and super easy)! I made a quick…

DGX agent

Want to chat with your Hermes Agent on iMessage without a Mac? The latest v0.17.0 update made this possible (and super easy)! I made a quick video showing how to get this all set up in less than three

agentsnous-research--x
22 Jun 2026
Agents

You don't need to be AI Dependent to build State of the Art software. @FactoryAI and Fireworks use GLM to build the most advanced agentic AI…

DGX agent

You don't need to be AI Dependent to build State of the Art software. @FactoryAI and Fireworks use GLM to build the most advanced agentic AI you can train and own today. GLM 5.2 is available in Droid,

agentsfireworks-ai--x
22 Jun 2026
Agents

One of the better agentic AI courses I've seen Nearly 10 hours of great content. Covers LangChain, LangGraph, RAG, deepagents, guardrails, a…

DGX agent

One of the better agentic AI courses I've seen Nearly 10 hours of great content. Covers LangChain, LangGraph, RAG, deepagents, guardrails, and more Any other good Lang* resources out there for folks w

agentsharrison-chase--x
20 Jun 2026
Agents

[AINews] Open Models, Model Labs vs Agent Labs, and What's Untrainable — Sarah Guo

DGX agent

This episode discusses the growing landscape of open-source AI models, contrasts the approaches and philosophies of Model Labs versus Agent Labs in AI development, and explores fundamental limitations

agentslatent-space
11 Jun 2026
Model Releases

Beyond Compaction: Structured Context Eviction for Long-Horizon Agents

DGX agent

arXiv:2606.11213v1 Announce Type: new Abstract: We present Context Window Lifecycle (CWL), a context-management scheme that gives long-horizon LLM agents an effectively unbounded working horizon. As a

model-releasesarxiv-cs-cl
11 Jun 2026
Agents

Bootstrapped Monitoring: Leveraging Transparent Reasoning to Oversee Stronger AI Agents

DGX agent

arXiv:2606.11998v1 Announce Type: new Abstract: Trusted monitoring is a cornerstone of AI control. However, as frontier models grow more capable, the increasing capabilities gap between trusted and un

agentsarxiv-cs-lg
11 Jun 2026
Model Releases

Agentic Hybrid RAG for Evidence-Grounded Muon Collider Analysis

DGX agent

arXiv:2606.10381v1 Announce Type: cross Abstract: Muon collider research spans accelerator physics, detector instrumentation, and high-energy phenomenology, with relevant evidence scattered across a r

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

AutoPDE: Reliable Agentic PDE Solving via Explicitly Represented Solver Strategies

DGX agent

arXiv:2606.10752v1 Announce Type: new Abstract: Numerical solvers for partial differential equations (PDEs) are core computational tools in science and engineering. Building reliable PDE solvers requi

agentsarxiv-cs-ai
10 Jun 2026
Agents

Cursor’s code review agent is now over 3x faster, 22% cheaper, and finds 10% more bugs. You can also use /review to run Bugbot locally to ca…

DGX agent

Cursor has released performance improvements to its code review agent, achieving 3x faster speed, 22% cost reduction, and 10% increased bug detection rates. The update introduces a /review command tha

agentscursor--x
10 Jun 2026
Model Releases

datasette-agent 0.2a0

DGX agent

Release: datasette-agent 0.2a0 Highlights from the release notes: Tools can now ask the user questions mid-execution. Tools that declare a context parameter receive a ToolContext object, and await con

model-releasessimon-willison
10 Jun 2026
Local Ai

Decentralized Multi-Agent Systems with Shared Context

DGX agent

arXiv:2606.10662v1 Announce Type: cross Abstract: Multi-agent systems (MAS) can scale large language model reasoning at test time by decomposing complex problems into parallel subtasks. However, most

local-aiarxiv-cs-ai
10 Jun 2026
Agents

How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latency o…

DGX agent

How do you support full-text search JSON filtering over agent traces that span up to hundreds of MBs, while keeping a median (P50) latency of 400ms? Here’s an inside look at how we built a custom inve

agentsharrison-chase--x
10 Jun 2026
Agents

in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: …

DGX agent

in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, Held-Out Gated Improvement Loop Demonstrated o

agentsyohei-nakajima--x
10 Jun 2026
Agents

Introducing the Hermes Agent Profile Builder You can now build a complete profile in the dashboard with full control over identity/name/desc…

DGX agent

Introducing the Hermes Agent Profile Builder You can now build a complete profile in the dashboard with full control over identity/name/description, model/provider, built-in + optional skills, skills-

agentsnous-research--x
10 Jun 2026
Agents

less novel, but still very interesting impo is the gated approach to self-modification the agent basically forks itself, propose a patch, ru…

DGX agent

less novel, but still very interesting impo is the gated approach to self-modification the agent basically forks itself, propose a patch, run through multiple tests (static/sandbox/diff), and somethin

agentsyohei-nakajima--x
10 Jun 2026
Model Releases

MemVenom: Triggered Poisoning of Multimodal Memories in Web Agents

DGX agent

arXiv:2606.10742v1 Announce Type: cross Abstract: External memory has become a core component of modern web agents, enabling long-horizon reasoning through the retrieval of past experiences. However,

model-releasesarxiv-cs-lg
10 Jun 2026
Model Releases

T1-Bench: Benchmarking Multi-Scenario Agents in Real-World Domains

DGX agent

arXiv:2606.11070v1 Announce Type: cross Abstract: Recent advances in reasoning and tool-calling capabilities of large language models (LLMs) have enabled increasingly capable agentic systems. However,

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

TabClaw: An Interactive and Self-Evolving Agent for Spreadsheet Manipulation and Table Reasoning

DGX agent

arXiv:2606.10316v1 Announce Type: new Abstract: Spreadsheets and tables are widely used representations for structured data analysis, but effective analysis still requires substantial manual effort an

agentsarxiv-cs-cl
10 Jun 2026
Agents

two fun surprises from using activegraph: - the coding agent i was using would query the trace db to debug instead of looking at the logs li…

DGX agent

two fun surprises from using activegraph: - the coding agent i was using would query the trace db to debug instead of looking at the logs like they normally would (i didn't ask it to) - when long eval

agentsyohei-nakajima--x
10 Jun 2026
Agents

What Spatial Memory Must Store: Occlusion as the Test for Language-Agent Memory

DGX agent

arXiv:2606.10299v1 Announce Type: new Abstract: Language-agent 'memory palace' systems anchor each memory to a world coordinate, on the intuition that geometry adds something text cannot. We make that

agentsarxiv-cs-ai
10 Jun 2026
Agents

A multi-agent system for spine MRI report generation from multi-sequence imaging

DGX agent

arXiv:2606.08897v1 Announce Type: cross Abstract: Spinal pathology is a leading cause of pain and disability worldwide. Spine MRI is central to clinical evaluation, yet its interpretation remains comp

agentsarxiv-cs-ai
9 Jun 2026
Agents

Agentic multi-fidelity learning of quasiparticle and excitonic properties

DGX agent

arXiv:2606.07836v1 Announce Type: cross Abstract: Many-body GW-Bethe-Salpeter equation calculations are essential for accurate simulations of electronic structure and optical properties in modern low-

agentsarxiv-cs-ai
9 Jun 2026
Agents

AgentTrust: A Self-Improving Trust Layer for AI-Agent Actions

DGX agent

arXiv:2606.08539v1 Announce Type: new Abstract: AI agents increasingly take consequential actions -- shell commands, cloud operations, and arbitrary tool-calls -- so a trust layer must decide, per act

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Beyond Goodhart's Law: A Dynamic Benchmark for Evaluating Compliance in Multi-Agent Systems

DGX agent

arXiv:2606.07805v1 Announce Type: new Abstract: The rapid evolution of Large Language Models (LLMs) from passive assistants to autonomous, execution-capable agents has introduced critical operational

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Context-Fractured Decomposition Attacks on Tool-Using LLM Agents: Exploiting Artifact Provenance Gaps

DGX agent

arXiv:2606.09084v1 Announce Type: cross Abstract: Tool-using LLM agents interact with the world through actions that persist state in artifacts (e.g., workspace files or logs). Consequently, jailbreak

safetyarxiv-cs-ai
9 Jun 2026
Agents

Cost-Aware Speculative Execution for LLM-Agent Workflows: An Integrated Five-Dimension Method

DGX agent

arXiv:2606.07846v1 Announce Type: cross Abstract: LLM-agent workflows chain model calls and tool invocations, and spend most of their wall-clock time waiting on upstream operations before downstream o

agentsarxiv-cs-ai
9 Jun 2026
Agents

Introducing Cohere's first open-source coding model: North Mini Code Small & efficient, designed for agentic performance and built for commu…

DGX agent

Cohere released North Mini Code, an open-source coding model designed to be small and efficient while optimizing for agentic performance and community use. The model represents Cohere's initial offeri

agentsclem-delangue--x
9 Jun 2026
Agents

man its kinda wild to write “agent lab” on my random blog and a few months later its adopted by people like @breeves08 and @ScottWu46 🫡 htt…

DGX agent

man its kinda wild to write “agent lab” on my random blog and a few months later its adopted by people like @breeves08 and @ScottWu46 🫡 https://x.com/cognition/status/2062923088778367314?s=46 More tha

agentsswyx--x
9 Jun 2026
Model Releases

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

DGX agent

arXiv:2605.22664v2 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet ente

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

DGX agent

arXiv:2606.08919v1 Announce Type: new Abstract: As LLM agents begin to take real, irreversible actions (shell commands, file edits, deploys), the standard safety pattern is a human-in-the-loop approva

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

POISE: Position-Aware Undetectable Skill Injection on LLM Agents

DGX agent

arXiv:2606.07943v1 Announce Type: cross Abstract: Agent skills provide a lightweight mechanism for extending general-purpose agents, but their open format exposes them to skill-poisoning attacks. A pr

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

DGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

agentsarxiv-cs-ai
9 Jun 2026
Agents

Reminder: The only place to download the Hermes Agent Desktop App is https://hermes-agent.nousresearch.com/desktop Any other website or sour…

DGX agent

Reminder: The only place to download the Hermes Agent Desktop App is https://hermes-agent.nousresearch.com/desktop Any other website or source is dangerous and could contain old, or worse, dangerous a

agentsnous-research--x
9 Jun 2026
Agents

SAGE: An LLM-driven Self Reflective Agentic Framework for Fraud Detection

DGX agent

arXiv:2606.08146v1 Announce Type: new Abstract: Fraud detection in payment, e-commerce, and telecommunications systems requires accuracy at the individual level, robustness under severe class imbalanc

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

WeaveBench: A Long-Horizon, Real-World Benchmark for Computer-Use Agents with Hybrid Interfaces

DGX agent

arXiv:2606.09426v1 Announce Type: new Abstract: Computer-use agents (CUAs) increasingly operate in runtimes that combine visual desktop control, command-line execution, code editing, browsers, and ext

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Do Coding Agents Deceive Us? Detecting and Preventing Cheating via Capped Evaluation with Randomized Tests

DGX agent

arXiv:2606.07379v1 Announce Type: cross Abstract: A growing failure mode in agent evaluation and training is that models can achieve high evaluation scores by exploiting shortcuts instead of solving t

agentsarxiv-cs-ai
8 Jun 2026
← Previous
1…101102103104105…375
Next →