AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,976 results
Model Releases

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch

DGX agent

🚀Qwen3.6-Plus is on Nous Portal now and FREE for a limited time. Hermes Agent, here we go!! ⚡️ @NousResearch Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal i

model-releasesqwen--x
13 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tools

Starting today, you can run cloud agents inside fully configured development environments. Set them up the same way you'd set up a laptop fo…

DGX agent

Starting today, you can run cloud agents inside fully configured development environments. Set them up the same way you'd set up a laptop for an engineer: cloned repos, installed dependencies, and too

toolscursor--x
13 May 2026
Model Releases

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK…

DGX agent

This is misleading. This policy redefines the term 'interactive' to mean 'using an Anthropic front-end'. If you use `claude -p` or Agent SDK to do something interactively, it now uses credits, not you

model-releasesjeremy-howard--x
13 May 2026
Model Releases

A Communication-Theoretic Framework for LLM Agents: Cost-Aware Adaptive Reliability

DGX agent

arXiv:2605.09121v1 Announce Type: cross Abstract: Agents built on large language models (LLMs) rely on a range of reliability techniques, including retry, majority voting, and self-consistency, that h

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AnomalyClaw: A Universal Visual Anomaly Detection Agent via Tool-Grounded Refutation

DGX agent

arXiv:2605.10397v1 Announce Type: cross Abstract: Visual anomaly detection (VAD) is crucial in many real-world fields, such as industrial inspection, medical imaging, infrastructure monitoring, and re

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AssemPlanner: A Multi-Agent Based Task Planning Framework for Flexible Assembly System

DGX agent

arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio

model-releasesarxiv-cs-ro
12 May 2026
Model Releases

CIVeX: Causal Intervention Verification for Language Agents

DGX agent

arXiv:2605.09168v1 Announce Type: new Abstract: A valid tool call is not necessarily a valid intervention. Tool-using language agents are guarded by schema validators, policy filters, provenance check

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Combining Mechanical and Agentic Specification Inference for Move

DGX agent

arXiv:2605.10005v1 Announce Type: cross Abstract: In this paper, we describe early work on a specification inference tool for the Move Prover that combines a weakest-precondition (WP) analysis over Mo

model-releasesarxiv-cs-ai
12 May 2026
Safety

Conformity Generates Collective Misalignment in AI Agents Societies

DGX agent

arXiv:2605.10721v1 Announce Type: cross Abstract: Artificial intelligence safety research focuses on aligning individual language models with human values, yet deployed AI systems increasingly operate

safetyarxiv-cs-cl
12 May 2026
Research

Defense effectiveness across architectural layers: a mechanistic evaluation of persistent memory attacks on stateful LLM agents

DGX agent

arXiv:2605.08442v1 Announce Type: cross Abstract: Persistent memory attacks against LLM agents achieve high attack success rates against open-source models. In these attacks, malicious instructions in

researcharxiv-cs-ai
12 May 2026
Tutorials

Do Agents Need to Plan Step-by-Step? Rethinking Planning Horizon in Data-Centric Tool Calling

DGX agent

arXiv:2605.08477v1 Announce Type: new Abstract: Explicit planning is a critical capability for LLM-based agents solving complex data-centric tasks, which require precise tool calling over external dat

tutorialsarxiv-cs-cl
12 May 2026
Research

Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents

DGX agent

arXiv:2605.10663v1 Announce Type: new Abstract: Experience-driven self-evolving agents aim to overcome the static nature of large language models by distilling reusable experience from past interactio

researcharxiv-cs-ai
12 May 2026
Industry

Exaforce, which uses AI agents to detect and thwart cyberattacks, raised a 125M Series B at a 725M valuation, bringing its total funding to $200M (Marina Temkin/TechCrunch)

DGX agent

Marina Temkin / TechCrunch: Exaforce, which uses AI agents to detect and thwart cyberattacks, raised a 125M Series B at a 725M valuation, bringing its total funding to $200M — As bad actors weaponize

industrytechmeme
12 May 2026
Local Ai

FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration

DGX agent

arXiv:2605.08520v1 Announce Type: new Abstract: LLM-based evolution has emerged as a promising way to improve agents by refining non-parametric artifacts, but its wall-clock cost remains a major bottl

local-aiarxiv-cs-lg
12 May 2026
Model Releases

Human-Inspired Memory Architecture for LLM Agents

DGX agent

arXiv:2605.08538v1 Announce Type: new Abstract: Current LLM agents lack principled mechanisms for managing persistent memory across long interaction horizons. We present a biologically-grounded memory

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pa…

DGX agent

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pages with complex text layouts and tables, and it will extrac

model-releasesjerry-liu--x
12 May 2026
Model Releases

LLM Agents Already Know When to Call Tools -- Even Without Reasoning

DGX agent

arXiv:2605.09252v1 Announce Type: new Abstract: Tool-augmented LLM agents tend to call tools indiscriminately, even when the model can answer directly. Each unnecessary call wastes API fees and latenc

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

MemQ: Integrating Q-Learning into Self-Evolving Memory Agents over Provenance DAGs

DGX agent

arXiv:2605.08374v1 Announce Type: new Abstract: Episodic memory allows LLM agents to accumulate and retrieve experience, but current methods treat each memory independently, i.e., evaluating retrieval

model-releasesarxiv-cs-ai
12 May 2026
Safety

Position: Academic Conferences are Potentially Facing Denominator Gaming Caused by Fully Automated Scientific Agents

DGX agent

arXiv:2605.09915v1 Announce Type: cross Abstract: The implicit policy of maintaining relatively stable acceptance rates at top AI conferences, despite exponentially growing submissions, introduces a c

safetyarxiv-cs-ai
12 May 2026
Local Ai

PYTHALAB-MERA: Validation-Grounded Memory, Retrieval, and Acceptance Control for Frozen-LLM Coding Agents

DGX agent

arXiv:2605.08468v1 Announce Type: cross Abstract: Local LLM-based coding agents increasingly work in settings where correctness is earned through execution feedback, persistent state, and bounded repa

local-aiarxiv-cs-ai
12 May 2026
Safety

Skill-R1: Agent Skill Evolution via Reinforcement Learning

DGX agent

arXiv:2605.09359v1 Announce Type: cross Abstract: Agentic large language models often rely on skills, reusable natural language procedures that guide planning, action, and tool use. In practice, skill

safetyarxiv-cs-ai
12 May 2026
Research

Symbolic learning is not a replacement for coding agents, it's a replacement for gradient descent & NNs: a low-level, completely general, ex…

DGX agent

Francois Chollet argues that symbolic learning represents a fundamental alternative to gradient descent and neural networks rather than a replacement for coding agents, offering a low-level, general-p

researchfrancois-chollet--x
12 May 2026
Model Releases

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

DGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

An Anthropic engineer argues HTML is a better output format for AI agents than Markdown, citing information density, ease of sharing, and two-way interaction (@trq212)

DGX agent

@trq212: An Anthropic engineer argues HTML is a better output format for AI agents than Markdown, citing information density, ease of sharing, and two-way interaction — Using Claude Code: The Unreason

model-releasestechmeme
11 May 2026
Model Releases

End-to-end PDDL Planning with Hardcoded and Dynamic Agents

DGX agent

arXiv:2512.09629v2 Announce Type: replace Abstract: We present an end-to-end framework for planning supported by verifiers. An orchestrator receives a human specification written in natural language a

model-releasesarxiv-cs-ai
11 May 2026
Tools

For browser-use AI agents, every task is dozens of model calls in a tight loop. The inference layer isn’t background infrastructure. It’s wh…

DGX agent

For browser-use AI agents, every task is dozens of model calls in a tight loop. The inference layer isn’t background infrastructure. It’s what the product runs on. @yutori_ai runs Scouts, Delegate, an

toolstogether-ai--x
11 May 2026
Model Releases

MAS-Algorithm: A Workflow for Solving Algorithmic Programming Problems with a Multi-Agent System

DGX agent

arXiv:2605.05949v2 Announce Type: replace Abstract: Algorithmic problem solving serves as a rigorous testbed for evaluating structured reasoning in AI coding systems, as it directly reflects a model's

model-releasesarxiv-cs-ai
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Safety

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

DGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

safetyarxiv-cs-lg
11 May 2026
Model Releases

Start 'claude agents' in a high level directory with all your repos in it (for me thats ~/Projects). It keeps track of which sessions need y…

DGX agent

Start 'claude agents' in a high level directory with all your repos in it (for me thats ~/Projects). It keeps track of which sessions need your input and makes it really easy to resume and pick up whe

model-releasesthariq--x
11 May 2026
Model Releases

Swap models & view their capabilities! Try out in Deep Agents CLI: https://docs.langchain.com/oss/python/deepagents/cli/

DGX agent

Swap models & view their capabilities! Try out in Deep Agents CLI: https://docs.langchain.com/oss/python/deepagents/cli/ here's model profile details look like in practice, using @NVIDIAAIDev's Nemotr

model-releasesharrison-chase--x
11 May 2026
Applications

Weblica: Scalable and Reproducible Training Environments for Visual Web Agents

DGX agent

arXiv:2605.06761v1 Announce Type: new Abstract: The web is complex, open-ended, and constantly changing, making it challenging to scale training data for visual web agents. Existing data collection at

applicationsarxiv-cs-ai
11 May 2026
Industry

Q&A with Qualcomm CEO Cristiano Amon on 2026 as 'year of agents', the end of the smartphone-centric world, 6G turning humans into 'walking cameras', and more (Fortune)

DGX agent

Fortune: Q&A with Qualcomm CEO Cristiano Amon on 2026 as “year of agents”, the end of the smartphone-centric world, 6G turning humans into “walking cameras”, and more — Cristiano Amon's pitch is that

industrytechmeme
10 May 2026
Model Releases

Downloading now... 1M token context window with supposedly usable coding agent capability all on a 128GB Macbook Pro is 🤯

DGX agent

Downloading now... 1M token context window with supposedly usable coding agent capability all on a 128GB Macbook Pro is 🤯 🚨 OPEN SOURCE AI IS LITERALLY UNSTOPPABLE 🚨 The legendary founder of Redis (An

model-releasesclem-delangue--x
9 May 2026
Model Releases

Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misalign…

DGX agent

Chain of thought monitors are a key layer of defense against AI agent misalignment. To preserve monitorability, we avoid penalizing misaligned reasoning during RL. We found a limited amount of acciden

model-releasessam-altman--x
8 May 2026
Model Releases

Last week we shipped 50+ Claude Code reliability fixes. This week it's 60+ more. Smoother long-running sessions, a more efficient agent loop…

DGX agent

Last week we shipped 50+ Claude Code reliability fixes. This week it's 60+ more. Smoother long-running sessions, a more efficient agent loop, auth that works in more environments, and terminal fixes:

model-releasesboris-cherny--x
8 May 2026
Industry

New research: long-running agents often fail by stopping too early, not because the model can't make progress. We tested 5 harness designs a…

DGX agent

New research: long-running agents often fail by stopping too early, not because the model can't make progress. We tested 5 harness designs across 8 long-horizon coding tasks. Our new orchestration har

industryemad-mostaque--x
8 May 2026
Industry

Agents that transact: Introducing Amazon Bedrock AgentCore payments, built with Coinbase and Stripe

DGX agent

Today, we're announcing a preview of Amazon Bedrock AgentCore Payments, a new set of features in Amazon Bedrock AgentCore that enables AI agents to instantly access and pay for what they use. AgentCor

industryaws-ml-blog
7 May 2026
Tools

Introducing /orchestrate, a skill that recursively spawns agents to tackle your most ambitious tasks with the Cursor SDK. We’ve used it to: …

DGX agent

Introducing /orchestrate, a skill that recursively spawns agents to tackle your most ambitious tasks with the Cursor SDK. We’ve used it to: - Autoresearch our internal skills, cutting token use by 20%

toolscursor--x
7 May 2026
Tutorials

Most teams can build agents, but far fewer have the infrastructure or know-how to run them reliably in production. Join us for an evening in…

DGX agent

Most teams can build agents, but far fewer have the infrastructure or know-how to run them reliably in production. Join us for an evening in San Francisco on May 19th, with Victor Moreira + @Vtrivedy1

tutorialsharrison-chase--x
7 May 2026
Applications

Our new voice models are now available in the Realtime API: 🎙️ GPT-Realtime-2: Build production-ready voice agents that can think harder, t…

DGX agent

Our new voice models are now available in the Realtime API: 🎙️ GPT-Realtime-2: Build production-ready voice agents that can think harder, take action, handle interruptions, and keep conversations flow

applicationsopenai--x
7 May 2026
Safety

Scalable Multi Agent Diffusion Policies for Coverage Control

DGX agent

arXiv:2509.17244v2 Announce Type: replace Abstract: We propose MADP, a novel diffusion-model-based approach for collaboration in decentralized robot swarms. MADP leverages diffusion models to generate

safetyarxiv-cs-ro
7 May 2026
Model Releases

Storage Is Not Memory: A Retrieval-Centered Architecture for Agent Recall

DGX agent

arXiv:2605.04897v1 Announce Type: new Abstract: Extraction at ingestion is the wrong primitive for agent memory: content discarded before the query is known cannot be recovered at retrieval time. We p

model-releasesarxiv-cs-cl
7 May 2026
Safety

Agentic AI-Based Joint Computing and Networking via Mixture of Experts and Large Language Models

DGX agent

arXiv:2605.02911v1 Announce Type: new Abstract: Future sixth-generation (6G) mobile networks are envisioned to be equipped with a diverse set of powerful, yet highly specialized, optimization experts.

safetyarxiv-cs-lg
6 May 2026
Research

AI agents often struggle to plan movements because their internal representations of the physical world can be overly tangled. CDS PhD stude…

DGX agent

AI agents often struggle to plan movements because their internal representations of the physical world can be overly tangled. CDS PhD student Ying Wang (@yingwww_) shows how straightening these pathw

researchyann-lecun--x
6 May 2026
Tools

As agents generate more code, review is becoming the bottleneck. You can write a PR in minutes, but understanding whether it's correct, safe…

DGX agent

As agents generate more code, review is becoming the bottleneck. You can write a PR in minutes, but understanding whether it's correct, safe, and ready to merge still takes longer than writing it. Dev

toolswindsurf--x
6 May 2026
Model Releases

Continuous-video agents (computer use, robotics, static scenes) burn compute re-ingesting pixels that didn't move. VLMaxxing teaches a froze…

DGX agent

Continuous-video agents (computer use, robotics, static scenes) burn compute re-ingesting pixels that didn't move. VLMaxxing teaches a frozen video VLM to skip the reruns. 54 fps perception on Gemma 4

model-releasesswyx--x
6 May 2026
Model Releases

DocSync: Agentic Documentation Maintenance via Critic-Guided Reflexion

DGX agent

arXiv:2605.02163v1 Announce Type: cross Abstract: Software documentation frequently drifts from executable logic as codebases evolve, creating technical debt that degrades maintainability and causes d

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…167168169170171…375
Next →