AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,975 results
Model Releases

Computer use turns Claude into an agent that can operate real UIs. New blog post on making it reliable in production: getting click accuracy…

DGX agent

Computer use turns Claude into an agent that can operate real UIs. New blog post on making it reliable in production: getting click accuracy right, choosing thinking effort levels, keeping long sessio

model-releasesboris-cherny--x
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Enhancing Cloud Network Resilience via a Robust LLM-Empowered Multi-Agent Reinforcement Learning Framework

DGX agent

arXiv:2601.07122v2 Announce Type: replace-cross Abstract: While virtualization and resource pooling empower cloud networks with structural flexibility and elastic scalability, they inevitably expand t

local-aiarxiv-cs-ai
19 May 2026
Safety

Equilibrium Selection in Multi-Agent Policy Gradients via Opponent-Aware Basin Entry

DGX agent

arXiv:2605.18078v1 Announce Type: new Abstract: Multi-agent policy-gradient methods have been shown to converge locally near stable Nash equilibria. Local convergence, however, does not determine whic

safetyarxiv-cs-lg
19 May 2026
Model Releases

extsc{PrivScope}: Task-scoped Disclosure Control for Hybrid Agentic Systems

DGX agent

arXiv:2605.16630v1 Announce Type: cross Abstract: Hybrid local--cloud agents enrich user requests with context from persistent working state before delegating capability-intensive subtasks to a cloud

model-releasesarxiv-cs-ai
19 May 2026
Safety

Generation Navigator: A State-Aware Agentic Framework for Image Generation

DGX agent

arXiv:2605.17969v1 Announce Type: new Abstract: Despite rapid advances in text-to-image generation, faithfully realizing user intent remains challenging, often requiring manual multi-turn trial and er

safetyarxiv-cs-cv
19 May 2026
Model Releases

Introducing Claude Managed Agents with Modal Sandboxes

DGX agent

This article announces a collaboration between Anthropic's Claude and Modal that enables Claude to operate as managed agents within Modal's sandboxed environments. The integration allows Claude to saf

model-releasesmodal-blog
19 May 2026
Model Releases

LaunchDarkly launches runtime control layer for the agentic AI era

DGX agent

LaunchDarkly, a feature control platform that helps developers and software engineers launch and manage products, today announced the launch of AgentControl, a new solution providing real-time managem

model-releasessiliconangle
19 May 2026
Model Releases

MAVEN A Multi-Agent Framework for Multicultural Text-to-Video Generation

DGX agent

arXiv:2605.16716v1 Announce Type: cross Abstract: Text-to-video (T2V) generation has rapidly progressed in visual fidelity, yet its ability to faithfully represent multiple cultures within a single pr

model-releasesarxiv-cs-ai
19 May 2026
Safety

Mitigating Conversational Inertia in Multi-Turn Agents

DGX agent

arXiv:2602.03664v3 Announce Type: replace Abstract: Large language models excel as few-shot learners when provided with appropriate demonstrations, yet this strength becomes problematic in multiturn a

safetyarxiv-cs-ai
19 May 2026
Local Ai

Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction

DGX agent

arXiv:2605.18729v1 Announce Type: cross Abstract: The ability to navigate and interact with complex environments is central to real-world embodied agents, yet navigation in unseen environments remains

local-aiarxiv-cs-cv
19 May 2026
Model Releases

Skills on the Fly: Test-Time Adaptive Skill Synthesis for LLM Agents

DGX agent

arXiv:2605.16986v1 Announce Type: cross Abstract: LLM agents benefit from reusable skills, yet test-time tasks often require guidance more specific than a static skill library can provide. We propose

model-releasesarxiv-cs-ai
19 May 2026
Research

SPIKE: An Adaptive Dual Controller Framework for Cost-Efficient Long-Horizon Game Agents

DGX agent

arXiv:2605.18636v1 Announce Type: new Abstract: Long-horizon multimodal agents in open-world games must stay goal-directed across many low-level interactions under tight token and latency budgets. Exi

researcharxiv-cs-cv
19 May 2026
Model Releases

TusoAI: Agentic Optimization for Scientific Methods

DGX agent

arXiv:2509.23986v2 Announce Type: replace Abstract: Scientific discovery is often slowed by the manual development of computational tools needed to analyze complex experimental data. Building such too

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

WEBSERV: A Full-Stack and RL-Ready Web Environment for Training Web Agents at Scale

DGX agent

arXiv:2510.16252v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) for web agents demands environments that are both effective for evaluation and efficient enough for large-scale on

model-releasesarxiv-cs-cl
19 May 2026
Research

A mental model for working with coding agents is that they're blind squirrels running into a maze and bumping into walls. You must place the…

DGX agent

A mental model for working with coding agents is that they're blind squirrels running into a maze and bumping into walls. You must place the walls (verifiable constraints) strategically so that they e

researchfrancois-chollet--x
18 May 2026
Industry

Companies running bug bounty programs are tightening background checks and building AI agents to triage a flood of low-quality reports generated by AI (Jamie John/Financial Times)

DGX agent

Jamie John / Financial Times: Companies running bug bounty programs are tightening background checks and building AI agents to triage a flood of low-quality reports generated by AI — ‘Bug bounty’ prog

industrytechmeme
18 May 2026
Safety

TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices

DGX agent

arXiv:2605.15611v1 Announce Type: new Abstract: Root cause analysis (RCA) in microservices is challenging due to (i) noisy and heterogeneous multimodal observability (metrics, logs, traces), (ii) casc

safetyarxiv-cs-ai
18 May 2026
Model Releases

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the env…

DGX agent

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the environment itself is the verifier. 🏆Among frontier models, @An

model-releaseszhipu-ai--x
17 May 2026
Safety

LLM-powered AI agents are gonna be great! You should totally trust them!

DGX agent

Gary Marcus expresses optimism about the potential of LLM-powered AI agents in a post on X (formerly Twitter). The post advocates for confidence in these AI systems, though without access to the full

safetygary-marcus--x
16 May 2026
Model Releases

AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

DGX agent

arXiv:2509.26100v2 Announce Type: replace Abstract: The rapid integration of Large Language Models (LLMs) into high-stakes domains necessitates reliable safety and compliance evaluation. However, exis

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents

DGX agent

arXiv:2605.13941v1 Announce Type: cross Abstract: Long-term memory is essential for LLM agents that operate across multiple sessions, yet existing memory systems treat retrieval infrastructure as fixe

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

From Text to Voice: A Reproducible and Verifiable Framework for Evaluating Tool Calling LLM Agents

DGX agent

arXiv:2605.15104v1 Announce Type: new Abstract: Voice agents increasingly require reliable tool use from speech, whereas prominent tool-calling benchmarks remain text-based. We study whether verified

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Latency-Quality Routing for Functionally Equivalent Tools in LLM Agents

DGX agent

arXiv:2605.14241v1 Announce Type: new Abstract: Tool-augmented LLM agents increasingly access the same tool type through multiple functionally equivalent providers, such as web-search APIs, retrievers

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

MemReranker: Reasoning-Aware Reranking for Agent Memory Retrieval

DGX agent

arXiv:2605.06132v2 Announce Type: replace Abstract: In agent memory systems, the reranking model serves as the critical bridge connecting user queries with long-term memory. Most systems adopt the 're

model-releasesarxiv-cs-cl
15 May 2026
Safety

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.14758v1 Announce Type: new Abstract: History-dependent policies induced by recurrent neural networks (RNNs) rely on latent hidden state dynamics, making verification in partially observable

safetyarxiv-cs-ai
15 May 2026
Model Releases

Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR)

DGX agent

arXiv:2605.14126v1 Announce Type: cross Abstract: Fast Healthcare Interoperability Resources (FHIR) is the dominant standard for interoperable exchange of healthcare data. In FHIR, electronic health r

model-releasesarxiv-cs-ai
15 May 2026
Safety

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy

DGX agent

arXiv:2605.14558v1 Announce Type: cross Abstract: Agentic reinforcement learning trains large language models using multi-turn trajectories that interleave long reasoning traces with short environment

safetyarxiv-cs-ai
15 May 2026
Local Ai

Towards In-Depth Root Cause Localization for Microservices with Multi-Agent Recursion-of-Thought

DGX agent

arXiv:2605.14866v1 Announce Type: cross Abstract: As modern microservice systems grow increasingly complex due to dynamic interactions and evolving runtime environments, they experience failures with

local-aiarxiv-cs-ai
15 May 2026
Local Ai

Agentic Interpretation: Lattice-Structured Evidence for LLM-Based Program Analysis

DGX agent

arXiv:2605.12694v1 Announce Type: cross Abstract: Large language models can consult information that fixed static analyzers cannot, such as documentation, current security advisories, version-specific

local-aiarxiv-cs-ai
14 May 2026
Model Releases

Another banger of a model, free for Hermes agent users via Nous Portal!

DGX agent

Nous Research announced the release of a new model available for free to Hermes agent users through the Nous Portal. The post suggests this is a significant model release from Nous Research, their org

model-releasesnous-research--x
14 May 2026
Safety

Control where your AI agents can browse with Chrome enterprise policies on Amazon Bedrock AgentCore

DGX agent

In this post, you will configure Chrome enterprise policies to restrict a browser agent to a specific website, observe the policy enforcement through session recording, and demonstrate custom root CA

safetyaws-ml-blog
14 May 2026
Safety

Data Agent: Learning to Select Data via End-to-End Dynamic Optimization

DGX agent

arXiv:2603.07433v2 Announce Type: replace-cross Abstract: Dynamic Data selection aims to accelerate training by prioritizing informative samples during online training. However, existing methods typic

safetyarxiv-cs-cv
14 May 2026
Hardware

How the NVIDIA Vera Rubin Platform is Solving Agentic AI’s Scale-Up Problem

DGX agent

The NVIDIA Vera Rubin Platform is a rack-scale AI supercomputer designed to power agentic AI and reasoning models at scale by eliminating bottlenecks in communication and memory movement for efficient

hardwarenvidia-developer
14 May 2026
Model Releases

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询…

DGX agent

LangChain 在 Interrupt 大会上发布了底层数据库 SmithDB 和自动化排障引擎 LangSmith Engine。 Agent 运行会产生海量 trace(执行轨迹),把旧数据库撑到了瓶颈。新底座 SmithDB 放弃了本地磁盘,全面转向对象存储,将核心查询速度拉高了 15 倍。 底座换新后,LangSmith Engine 顺势接管了查 Bug 的体力活。它在后台持续监控生

model-releasesharrison-chase--x
14 May 2026
Model Releases

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete…

DGX agent

Meet Kimi Web Bridge - Kimi's browser extension. Agent can now interact with websites like a human: search, scroll, click, type and complete tasks. Supports Kimi Code CLI, Claude Code, Cursor, Codex,

model-releaseskimi-moonshot--x
14 May 2026
Safety

Not Just RLHF: Why Alignment Alone Won't Fix Multi-Agent Sycophancy

DGX agent

arXiv:2605.12991v1 Announce Type: cross Abstract: LLM-based multi-agent pipelines flip from correct to incorrect answers under simulated peer disagreement at rates we term yield, a vulnerability widel

safetyarxiv-cs-ai
14 May 2026
Research

two creative hackathon entries that caught my eye: @Liftaris1's herm puts hermes primitives (sessions, skills, cron, agents) as first class …

DGX agent

two creative hackathon entries that caught my eye: @Liftaris1's herm puts hermes primitives (sessions, skills, cron, agents) as first class citizens in the tui. @arm64le's theia turns your sessions hi

researchnous-research--x
14 May 2026
Safety

VideoSEAL: Mitigating Evidence Misalignment in Agentic Long Video Understanding by Decoupling Answer Authority

DGX agent

arXiv:2605.12571v1 Announce Type: cross Abstract: Long video question answering requires locating sparse, time-scattered visual evidence within highly redundant content. Although current MLLMs perform

safetyarxiv-cs-ai
14 May 2026
Applications

We are excited to be partnering with @LangChain for deploying self-improving agents. Continual learning in your production environment unloc…

DGX agent

We are excited to be partnering with @LangChain for deploying self-improving agents. Continual learning in your production environment unlocks compounding capability gains for model-product optimizati

applicationsharrison-chase--x
14 May 2026
Applications

“Whimsey attacks” that seem absurd (“I cannot pay that much because of the Geneva Convention”) work against AI agents as guardrails are weak…

DGX agent

“Whimsey attacks” that seem absurd (“I cannot pay that much because of the Geneva Convention”) work against AI agents as guardrails are weak against out-of-distribution arguments. Smaller models fall

applicationsethan-mollick--x
14 May 2026
Safety

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2605.11880v1 Announce Type: new Abstract: TD(lambda) in value-based MARL algorithms or the Temporal Difference critic learning in Actor-Critic-based (AC-based) algorithms synergistically integra

safetyarxiv-cs-lg
13 May 2026
Model Releases

Agent-Based Post-Hoc Correction of Agricultural Yield Forecasts

DGX agent

arXiv:2605.12375v1 Announce Type: new Abstract: Accurate crop yield forecasting in commercial soft fruit production is constrained by the data available in typical commercial farm records, which lack

model-releasesarxiv-cs-lg
13 May 2026
Tools

Computer is secure by default. Every task runs in its own hardware-isolated sandbox with VPC-level storage and compute separation. Agents ar…

DGX agent

Computer is secure by default. Every task runs in its own hardware-isolated sandbox with VPC-level storage and compute separation. Agents are authenticated with short-lived proxy tokens instead of raw

toolsperplexity--x
13 May 2026
Model Releases

Courtroom-Style Multi-Agent Debate with Progressive RAG and Role-Switching for Controversial Claim Verification

DGX agent

arXiv:2603.28488v2 Announce Type: replace Abstract: Large language models (LLMs) remain unreliable for high-stakes claim verification due to hallucinations and shallow reasoning. While retrieval-augme

model-releasesarxiv-cs-cl
13 May 2026
Tools

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted…

DGX agent

External content is scanned in parallel by ML classifiers and the BrowseSafe model before agents act on it. File connector data is encrypted in transit and at rest, uploaded files automatically delete

toolsperplexity--x
13 May 2026
Safety

Missing Old Logits in Asynchronous Agentic RL: Semantic Mismatch and Repair Methods for Off-Policy Correction

DGX agent

arXiv:2605.12070v1 Announce Type: new Abstract: Asynchronous reinforcement learning improves rollout throughput for large language model agents by decoupling sample generation from policy optimization

safetyarxiv-cs-lg
13 May 2026
Model Releases

PRISM: Pareto-Efficient Retrieval over Intent-Aware Structured Memory for Long-Horizon Agents

DGX agent

arXiv:2605.12260v1 Announce Type: new Abstract: Long-horizon language agents accumulate conversation history far faster than any fixed context window can hold, making memory management critical to bot

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch

DGX agent

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal i

model-releasesqwen--x
13 May 2026
← Previous
1…166167168169170…375
Next →