AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,975 results
Model Releases

From Facts to Insights: A Persona-Driven Dual Memory Framework and Dataset for Role-Playing Agents

DGX agent

arXiv:2605.25693v1 Announce Type: new Abstract: While role-playing agents excel in short-term interactions, long-term conversations overwhelm context windows, motivating external memory frameworks. Cu

model-releasesarxiv-cs-cl
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

DGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

model-releasesclem-delangue--x
26 May 2026
Local Ai

LLM Agent Based Renewable Energy Forecasting Using Edge and IoT Data A Review of Solar Wind Weather and Grid Aware Decision Support

DGX agent

arXiv:2605.25141v1 Announce Type: cross Abstract: Reliable forecasting of renewable energy generation is a foundational requirement for grid stability energy trading battery scheduling and carbon awar

local-aiarxiv-cs-ai
26 May 2026
Safety

Micro-Swarm Locomotion Optimization in Dynamic Flow using Multi-Objective Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.25025v1 Announce Type: new Abstract: Coordinating micro-robotic swarms in physiologically realistic, time-dependent fluid environments remains an unsolved challenge for biomedical and envir

safetyarxiv-cs-ro
26 May 2026
Model Releases

Neural Router: Semantic Content Matching for Agentic AI

DGX agent

arXiv:2605.25701v1 Announce Type: cross Abstract: Large language models (LLMs) can serve as the semantic-matching engine of a content-based publish/subscribe broker for agentic AI across the edge-clou

model-releasesarxiv-cs-cl
26 May 2026
Model Releases

New on the Engineering Blog: The access and permissions we grant agents should evolve with their capabilities. In our own products, we set t…

DGX agent

New on the Engineering Blog: The access and permissions we grant agents should evolve with their capabilities. In our own products, we set these parameters through sandboxing, which limits the scope o

model-releasesboris-cherny--x
26 May 2026
Safety

Operationalizing Reconstructive Authority: Runtime Construction, Dependency Resolution, and Execution Gating in Autonomous Agent Systems

DGX agent

arXiv:2605.23935v1 Announce Type: new Abstract: Autonomous agent systems fail not only due to incorrect decisions, but due to executing decisions whose authority no longer holds at runtime. Prior work

safetyarxiv-cs-ai
26 May 2026
Model Releases

Our strategy lead @yeahfortommy was just on stage with CTO of @alibaba_cloud discussing Hermes Agent at the Qwen Conference, check it out: h…

DGX agent

Our strategy lead @yeahfortommy was just on stage with CTO of @alibaba_cloud discussing Hermes Agent at the Qwen Conference, check it out: https://www.youtube.com/live/r99c3sfgkmc?si=cnuR07ofhO_V69l-&

model-releasesnous-research--x
26 May 2026
Safety

ProActor: Timing-Aware Reinforcement Learning for Proactive Task Scheduling Agents

DGX agent

arXiv:2605.24900v1 Announce Type: new Abstract: Proactive task-oriented agents must autonomously anticipate user needs, identify actionable opportunities, and trigger software actions at appropriate m

safetyarxiv-cs-ai
26 May 2026
Industry

Q&A with Sundar Pichai about reshaping the information ecosystem with Search changes, putting AI agents in everything, when AI will replace him as CEO, and more (Nilay Patel/The Verge)

DGX agent

Nilay Patel / The Verge: Q&A with Sundar Pichai about reshaping the information ecosystem with Search changes, putting AI agents in everything, when AI will replace him as CEO, and more — Today, I'm t

industrytechmeme
26 May 2026
Industry

Sources: Qualcomm reached a deal with ByteDance to supply millions of ASICs for AI data centers to support AI agents in the Doubao chatbot; QCOM jumps 5%+ (Ian King/Bloomberg)

DGX agent

Ian King / Bloomberg: Sources: Qualcomm reached a deal with ByteDance to supply millions of ASICs for AI data centers to support AI agents in the Doubao chatbot; QCOM jumps 5%+ — Qualcomm Inc. reached

industrytechmeme
26 May 2026
Local Ai

DART: Semantic Recoverability for Structured Tool Agents

DGX agent

arXiv:2605.23311v1 Announce Type: new Abstract: When a structured tool agent fails mid-execution, the runtime faces a dilemma: replaying the entire task is safe but wasteful, while restoring from a lo

local-aiarxiv-cs-ai
25 May 2026
Model Releases

DeepSeek V4 Flash IS BACK on Nous Portal for FREE for use in Hermes Agent! Check it out at https://portal.nousresearch.com/manage-subscripti…

DGX agent

DeepSeek V4 Flash model has been made available again on the Nous Research portal at no cost for use with Hermes Agent applications. Users can access and utilize this model through the Nous portal's s

model-releasesnous-research--x
25 May 2026
Model Releases

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up o…

DGX agent

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up on complex tasks, and it has actual personality. Few rough ed

model-releaseselon-musk--x
25 May 2026
Local Ai

PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA

DGX agent

arXiv:2605.23559v1 Announce Type: cross Abstract: Whole-slide image visual question answering (WSI-VQA) frames pathology as an extreme-context search problem: to answer a free-form clinical query, a s

local-aiarxiv-cs-ai
25 May 2026
Safety

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over

DGX agent

> I’m now in the LeCun/Marcus camp on LLMs > real programming agents will need world models > not some RLVR shit it’s over The Eternal Sloptember https://geohot.github.io//blog/jekyll/update/2026/05/2

safetygary-marcus--x
24 May 2026
Industry

Q: How are job postings for software engineers rising rapidly despite AI agents automating coding? A: Because there’s far more code to manag…

DGX agent

Q: How are job postings for software engineers rising rapidly despite AI agents automating coding? A: Because there’s far more code to manage than ever before. We’re already seeing a 14x YoY increase

industryclem-delangue--x
24 May 2026
Research

Dynamic Mixture of Latent Memories for Self-Evolving Agents

DGX agent

arXiv:2605.21951v1 Announce Type: new Abstract: Achieving self-evolution in intelligent agents requires the continual accumulation of new knowledge across changing task sequences without forgetting pr

researcharxiv-cs-lg
23 May 2026
Model Releases

MoralityGym: A Benchmark for Evaluating Hierarchical Moral Alignment in Sequential Decision-Making Agents

DGX agent

arXiv:2602.13372v2 Announce Type: replace-cross Abstract: Evaluating moral alignment in agents navigating conflicting, hierarchically structured human norms is a critical challenge at the intersection

model-releasesarxiv-cs-lg
23 May 2026
Safety

NEW: Google claimed this week that a team of agents had built an entire operating system, based on a single prompt, costing only about $900 …

DGX agent

NEW: Google claimed this week that a team of agents had built an entire operating system, based on a single prompt, costing only about $900 in tokens. We fact check this claim and analyze what it mean

safetygary-marcus--x
23 May 2026
Safety

Q&A with Sundar Pichai on the future of Google Search, Google's place in the AI race, public skepticism toward AI, AI agents, AI safety, TPUs, and more (New York Times)

DGX agent

New York Times: Q&A with Sundar Pichai on the future of Google Search, Google's place in the AI race, public skepticism toward AI, AI agents, AI safety, TPUs, and more — After a busy Google I/O, the c

safetytechmeme
23 May 2026
Model Releases

Big fan of teaching more people the basics of using Claude Code in an accessible way. So much of the world has not yet used agents. There's …

DGX agent

Big fan of teaching more people the basics of using Claude Code in an accessible way. So much of the world has not yet used agents. There's a lot of opportunity to level the playing field and expand a

model-releasesboris-cherny--x
22 May 2026
Model Releases

Evaluating Temporal Semantic Caching and Workflow Optimization in Agentic Plan-Execute Pipelines

DGX agent

arXiv:2605.20630v1 Announce Type: new Abstract: Industrial asset operations workflows are latency-sensitive because a single user query may require coordination over sensor data, work orders, failure

model-releasesarxiv-cs-ai
22 May 2026
Tutorials

Personality Engineering with AI Agents: A New Methodology for Negotiation Research

DGX agent

arXiv:2605.20554v1 Announce Type: new Abstract: According to canonical negotiation theory, people's success in a negotiation depends on how well they balance competing demands--empathizing and asserti

tutorialsarxiv-cs-ai
22 May 2026
Model Releases

RAGCap-Bench: Benchmarking Capabilities of LLMs in Agentic Retrieval Augmented Generation Systems

DGX agent

arXiv:2510.13910v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) mitigates key limitations of Large Language Models (LLMs)-such as factual errors, outdated knowledge, and hallu

model-releasesarxiv-cs-cl
22 May 2026
Safety

Retaining Suboptimal Actions to Follow Shifting Optima in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2602.17062v2 Announce Type: replace Abstract: Value decomposition is a core approach for cooperative multi-agent reinforcement learning (MARL). However, existing methods still rely on a single o

safetyarxiv-cs-ai
22 May 2026
Tools

With the Cursor SDK, you can build your own agents with Composer 2.5. It's now available in Python and TypeScript. This long weekend, Compos…

DGX agent

With the Cursor SDK, you can build your own agents with Composer 2.5. It's now available in Python and TypeScript. This long weekend, Composer usage is 90% off in the SDK. We're excited to see what yo

toolscursor--x
22 May 2026
Applications

Everyone on earth will get really good chatbots for free, so thats good for democratization of AI, but really good agents that can do comple…

DGX agent

Everyone on earth will get really good chatbots for free, so thats good for democratization of AI, but really good agents that can do complex work burn thousands of times more tokens, and will be rese

applicationsethan-mollick--x
21 May 2026
Model Releases

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite d…

DGX agent

I released the first alpha of Datasette Agent - a conversational AI assistant for Datasette that can answer questions about data in SQLite databases, and can be extended with plugins to add extra tool

model-releasessimon-willison--x
21 May 2026
Model Releases

In the next version of Claude Code: run /usage to see a breakdown of which Skills, Agents, MCPs, and Plugins are using your tokens CLI today…

DGX agent

The next version of Claude Code will introduce a `/usage` command that provides a detailed breakdown of token consumption across different components including Skills, Agents, MCPs (Model Context Prot

model-releasesboris-cherny--x
21 May 2026
Safety

Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty

DGX agent

arXiv:2605.20255v1 Announce Type: new Abstract: Simulation-based testing of self-driving cars (SDCs) typically relies on scripted or simplified pedestrian models that do not capture the heterogeneity

safetyarxiv-cs-lg
21 May 2026
Safety

STEAM: A Training-Free Congestion-Aware Enhancement Framework for Decentralized Multi-Agent Path Finding

DGX agent

arXiv:2605.20929v1 Announce Type: new Abstract: We propose STEAM (Spatial, Temporal, and Emergent congestion Awareness for MAPF), a training-free test-time enhancement framework for learning-based dec

safetyarxiv-cs-ro
21 May 2026
Model Releases

The Yes-Man Syndrome: Benchmarking Abstention in Embodied Robotic Agents

DGX agent

arXiv:2605.20544v1 Announce Type: cross Abstract: Vision-language models (VLMs) are used as high-level planners for embodied agents, translating natural language instructions and visual observations i

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million par…

DGX agent

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million parameters and runs nearly 2,000 tokens per second prefill on a

model-releasesswyx--x
21 May 2026
Research

4 big upgrades to Hermes Agents speed today `hermes update` to get moving faster now

DGX agent

Nous Research announced four major performance upgrades to their Hermes Agents framework, aimed at improving speed and efficiency. The update, referred to as the `hermes update`, enables faster execut

researchnous-research--x
20 May 2026
Safety

ContextFlow: Hierarchical Task-State Alignment for Long-Horizon Embodied Agents

DGX agent

arXiv:2605.19314v1 Announce Type: cross Abstract: Long-horizon embodied agents increasingly delegate navigation, search, approach, and manipulation to specialist executors. As these executors become s

safetyarxiv-cs-ai
20 May 2026
Model Releases

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

DGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

model-releasesarxiv-cs-ai
20 May 2026
Tutorials

Evaluating Memory Condensation Strategies for Coding Agents in Data-Driven Scientific Discovery

DGX agent

arXiv:2605.18854v1 Announce Type: new Abstract: Coding agents accumulate extensive context during long-running tasks, yet fixed context windows force practitioners to choose between truncation and tas

tutorialsarxiv-cs-lg
20 May 2026
Research

I'm switching to Hermes.... I've been using it for a month.....and I'm sold...moving all of my @openclaw agents to Hermes (@NousResearch) Wh…

DGX agent

I'm switching to Hermes.... I've been using it for a month.....and I'm sold...moving all of my @openclaw agents to Hermes (@NousResearch) Why? -----> https://youtu.be/QQEgIo4Juxg Thank you to @Hosting

researchnous-research--x
20 May 2026
Safety

Memory-Augmented Reinforcement Learning Agent for CAD Generation

DGX agent

arXiv:2605.19748v1 Announce Type: new Abstract: Automatic generation of computer-aided design (CAD) models is a core technology for enabling intelligence in advanced manufacturing. Existing generation

safetyarxiv-cs-ai
20 May 2026
Tools

MiniMax Speech 2.8 Turbo is built for voice agents that need natural delivery, not just clean audio. → Sound Tags for laughter, breathing, s…

DGX agent

MiniMax Speech 2.8 Turbo is built for voice agents that need natural delivery, not just clean audio. → Sound Tags for laughter, breathing, sighs, gasps, and other vocal cues → 60% prosody improvement

toolstogether-ai--x
20 May 2026
Safety

Phase-Aware Mixture of Experts for Agentic Reinforcement Learning

DGX agent

arXiv:2602.17038v3 Announce Type: replace Abstract: Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a single

safetyarxiv-cs-ai
20 May 2026
Model Releases

Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory

DGX agent

arXiv:2605.19952v1 Announce Type: new Abstract: To enable reliable long-term interaction, LLM agents require a memory system that can faithfully store, efficiently retrieve, and deeply reason over acc

model-releasesarxiv-cs-cl
20 May 2026
Safety

SimGym: A Framework for A/B Test Simulation in E-Commerce with Traffic-Grounded VLM Agents

DGX agent

arXiv:2605.19219v1 Announce Type: new Abstract: A/B testing remains the gold standard for evaluating modifications to e-commerce storefronts, yet it diverts traffic, requires weeks to reach statistica

safetyarxiv-cs-ai
20 May 2026
Model Releases

To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents

DGX agent

arXiv:2605.18882v1 Announce Type: cross Abstract: LLM agents exhibit a consistent tendency to over-call, invoking tools even in situations where none is needed. On the When2Call benchmark, six models

model-releasesarxiv-cs-ai
20 May 2026
Applications

Albert Einstein + ElevenLabs. AI agents can make education more accessible - a teacher for every student in every field. A classroom size of…

DGX agent

Albert Einstein + ElevenLabs. AI agents can make education more accessible - a teacher for every student in every field. A classroom size of one learning from icons who shaped the world Today with his

applicationsemad-mostaque--x
19 May 2026
Local Ai

BLAgent: Agentic RAG for File-Level Bug Localization

DGX agent

arXiv:2605.17965v1 Announce Type: cross Abstract: Bug localization remains a key bottleneck in downstream software maintenance tasks, including root cause analysis, triage, and automated program repai

local-aiarxiv-cs-ai
19 May 2026
Model Releases

Causal Intervention-Based Memory Selection for Long-Horizon LLM Agents

DGX agent

arXiv:2605.17641v1 Announce Type: new Abstract: Long-horizon LLM agents rely on persistent memory to support interactions across sessions, yet existing memory systems often retrieve context using sema

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…165166167168169…375
Next →