AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,950 results
Model Releases

Agentic Coding in the Wild: Characterizing GitHub Copilot Traces at Production Scale

DGX agent

arXiv:2608.00101v1 Announce Type: cross Abstract: AI coding agents like GitHub Copilot, Claude Code, and Codex interleave multi-step LLM inference with tool execution, creating a workload different fr

model-releasesarxiv-cs-lg
4 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency

DGX agent

Members of the Open Secure AI Alliance — now more than 120 organizations strong — are developing new guidelines to strengthen agentic AI cybersecurity as the annual Black Hat conference begins in Las

agentsnvidia-blog
4 Aug 2026
Safety

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent

DGX agent

arXiv:2608.01321v1 Announce Type: new Abstract: Multi-step search is a fundamental capability for search agents, enabling them to iteratively acquire, refine, and integrate external evidence for compl

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Global Optimization and Inference-Time Region Grafting for Agentic Workflows

DGX agent

arXiv:2608.02353v1 Announce Type: new Abstract: Recent advances in agentic workflow optimization automate workflow design through task-specific workflow search or input-conditioned architecture select

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

HopRefusalBench: Diagnosing Refusal Failures in Search-Augmented Agents for Multi-Hop Reasoning

DGX agent

arXiv:2608.01358v1 Announce Type: new Abstract: Search-augmented large language model agents are increasingly capable of solving knowledge-intensive tasks, but their behavior when a multi-hop question

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now mo…

DGX agent

Picking the right agent harness is now a crucial skill for any AI engineer. Imagine using the same model, same task, and same prompt. Now move it between two agent harnesses and the cost per success c

model-releasesdair-ai--x
4 Aug 2026
Agents

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing inter…

DGX agent

Finally a good paper testing whether agent memory needs an LLM at all. Production memory stacks spend extra model calls on summarizing interactions, writing records, and reranking retrievals. Every on

agentsdair-ai--x
3 Aug 2026
Agents

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman…

DGX agent

GOP AGs warn OpenAI's Altman to preserve records in AI agent hacking probe https://www.foxbusiness.com/technology/gop-ags-warn-openai-altman-preserve-records-ai-agent-hacking-probe?%3Fintcmp=tw_fbn&ta

agentsgary-marcus--x
3 Aug 2026
Agents

Self-Supervised Skill Optimization

DGX agent

arXiv:2607.28777v1 Announce Type: new Abstract: Agent skills provide frozen large language model (LLM) agents with reusable procedural guidance, and recent work shows that such skills can be optimized

agentsarxiv-cs-cl
3 Aug 2026
Agents

Transcript-Managed Transformers: Monotone Multi-Agent Collapse and Universality with Two Pop-Enabled Transcripts

DGX agent

arXiv:2607.29496v1 Announce Type: new Abstract: We study transcript management for fixed, finite-precision causal Transformers. A transcript is partitioned into channels of bounded blocks. Each transi

agentsarxiv-cs-lg
3 Aug 2026
Agents

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models…

DGX agent

Try Qwen3.8-Max on Hermes Agent and you will have to doubt on how much these open frontier models have caught up with frontier closed models. These new open models are insanely good. Meet Qwen3.8-Max:

agentsdair-ai--x
3 Aug 2026
Agents

When an agent starts with shared context, more people can get reliable answers. When truth is shared, AI becomes infrastucture. Read about h…

DGX agent

When an agent starts with shared context, more people can get reliable answers. When truth is shared, AI becomes infrastucture. Read about how our internal truth layer drives our teams at Replit: http

agentsreplit--x
3 Aug 2026
Agents

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skil…

DGX agent

I wish I had found this sooner. Nous Research launched a FREE Hermes agent Skills Hub. 90,000+ community skills across 200+ categories. Skills from OpenAI, Anthropic, HuggingFace & more. Thank me late

agentsnous-research--x
2 Aug 2026
Model Releases

LayerRAG-Bench: A Cross-Layer Reliability Benchmark for Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2607.27353v1 Announce Type: new Abstract: Agentic retrieval-augmented generation systems can produce answers that appear grounded while failing at the evidence, tool-contract, authorization, or

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Partner Capability Estimation for Task-Agnostic Adaptation in Ad-Hoc Teamwork

DGX agent

arXiv:2607.27177v1 Announce Type: new Abstract: Effective collaboration with novel and diverse partners is a crucial skill for autonomous agents. Most current ad-hoc teamwork (AHT) approaches assume t

agentsarxiv-cs-ai
31 Jul 2026
Safety

Procedural Fairness in Multi-Agent Bandits

DGX agent

arXiv:2601.10600v2 Announce Type: replace-cross Abstract: In the context of multi-agent multi-armed bandits (MA-MAB), fairness is often reduced to outcomes: maximizing welfare, reducing inequality, or

safetyarxiv-cs-lg
31 Jul 2026
Model Releases

Safety-Gated Agentic Supervisory Control on a Coupled Distillation Benchmark: Regime Map, Auditable Gate, and Co-Design Findings

DGX agent

arXiv:2607.27849v1 Announce Type: cross Abstract: An open-weight LLM can write composition setpoints every five minutes. What a plant still needs is a hard check: named constraints, logged margins, an

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/207051816739969…

DGX agent

Just for reference, from what I observed in my Using Local Coding Agents blog article last month: https://x.com/rasbt/status/2070518167399698490?s=20 'I tried to analyze why Claude Code uses more toke

model-releasessebastian-raschka--x
30 Jul 2026
Safety

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution

DGX agent

arXiv:2607.26784v1 Announce Type: new Abstract: Large language model agents often encounter related yet distinct tasks that share reusable solution patterns. Yet standard agentic reinforcement learnin

safetyarxiv-cs-lg
30 Jul 2026
Agents

Software Engineers: Do you honestly get anything useful out of LLMs?

DGX agent

For 6 months now I've been trying to make agentic coding work for me, using Pi and a handful 30-120B models (Qwens, Nemotrons, Leguna...etc). I'm not greedy either, I stick to decent quants, never qua

agentsr-localllama
30 Jul 2026
Safety

WikiLoop: Jointly Learning to Build and Navigate Agent-Native Wikis with Downstream Feedback

DGX agent

arXiv:2607.26604v1 Announce Type: new Abstract: Knowledge-base construction and querying are typically optimized in isolation: retrieval-augmented agents operate over a fixed, externally maintained in

safetyarxiv-cs-cl
30 Jul 2026
Model Releases

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models a…

DGX agent

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models are like that. So what changes? The way to interact with Opus

model-releasesdair-ai--x
29 Jul 2026
Agents

Agentic inference wastes GPUs on KV cache thrashing. ThunderAgent fixes it at the scheduler level: 2.5x higher single-node throughput and ~1…

DGX agent

Agentic inference wastes GPUs on KV cache thrashing. ThunderAgent fixes it at the scheduler level: 2.5x higher single-node throughput and ~10x lower P50 latency at high concurrency. ThunderAgent was a

agentstogether-ai--x
29 Jul 2026
Agents

Cardiologent: Multi-Agent Clinical Decision Support for Patient-Level Arrhythmia Assessment, Urgency, and Management

DGX agent

arXiv:2607.25340v1 Announce Type: new Abstract: The same episode of atrial fibrillation is a minor finding in a healthy adult and grounds for anticoagulation in an elderly patient with hypertension: i

agentsarxiv-cs-ai
29 Jul 2026
Safety

ContractHIL-HLS: Contract-Aligned Multi-Agent Workflow with Hardware-in-the-Loop Feedback for HLS Design

DGX agent

arXiv:2607.25283v1 Announce Type: new Abstract: This paper presents ContractHIL-HLS, a contract-aligned multi-agent workflow for practical high-level synthesis (HLS) engineering. The workflow makes th

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

COVENANT: Natural-Language Workflow Compilation for Aligned Agent Execution

DGX agent

arXiv:2607.25400v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly entrusted with natural-language workflow instructions (e.g., retail-payment policies) that specify no

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

From Naive RAG to Deep Agentic Retrieval: An Evolving Context Engineering Pipeline for Regulatory Compliance

DGX agent

arXiv:2607.24791v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) is the dominant paradigm for applying large language models (LLMs) to enterprise document corpora, yet naive impl

agentsarxiv-cs-ai
29 Jul 2026
Agents

How agentic AI can help telecom finance teams protect the margin when every moment matters

DGX agent

Agentic AI empowers telecom finance teams to preserve margins by rapidly detecting and preventing revenue leakage across billing, provisioning, and cost‑allocation processes. The technology automates

agentsdatabricks
29 Jul 2026
Model Releases

Interactive Reward Agent: GUI Task Evaluation via Environment-State Verification

DGX agent

arXiv:2607.25904v1 Announce Type: new Abstract: Graphical user interface task evaluation aims to determine whether a GUI agent has successfully completed a user instruction. Automated GUI task evaluat

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Messier: A High-Resolution Corpus for Cross-Benchmark Agent Evaluation

DGX agent

arXiv:2607.25891v1 Announce Type: new Abstract: Evaluating AI agents in interactive environments is hindered by fragmented tasks, scaffolds, verifiers, and scoring rules. Existing efforts focus on nar

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

OpenAI’s rogue AI agent didn’t stop at hacking Hugging Face

DGX agent

The AI agent that escaped from OpenAI and hacked developer platform Hugging Face attacked other companies as well, OpenAI revealed on Tuesday. The update substantially widens the scope of an already c

agentsthe-verge-ai
29 Jul 2026
Safety

SearchArt: Training Long-Horizon Search Agent with Scalable Synthetic and Verified Task

DGX agent

arXiv:2607.24850v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled search agents to autonomously tackle complex tasks across extended search and reasoning h

safetyarxiv-cs-lg
29 Jul 2026
Agents

The Future of Agentic AI Depends on Cloud Cost Optimization

DGX agent

Capitalizing on agentic AI depends on the right infrastructure investments, yet enterprises struggle to reduce cloud costs. Vultr VX1™ Cloud Compute presents a way to free up budget for CPUs and GPUs

agentsvultr
29 Jul 2026
Model Releases

WorkSurface-Bench: Benchmarking Enterprise Agents on Multi-Surface Knowledge Routing

DGX agent

arXiv:2607.25765v1 Announce Type: new Abstract: Enterprise agents often need to integrate heterogeneous knowledge sources: documents for narrative facts, tables for computation, and dependency graphs

model-releasesarxiv-cs-cl
29 Jul 2026
Model Releases

AlloBench: Measuring Online Tool Allocation Capability in LLM Agents

DGX agent

arXiv:2607.23332v1 Announce Type: new Abstract: Creating a reusable tool is an investment: an agent pays a fixed cost now in exchange for the potential of future reuse. Therefore, a user should prefer

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

ConsistencyGate: Preventing Memory Contamination in LLM Agents via Self-Consistency Admission Control

DGX agent

arXiv:2607.22962v1 Announce Type: new Abstract: LLM agents that operate over many turns accumulate facts in an external memory store and reuse them as premises for downstream reasoning. A hallucinated

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

CoopReflect: Towards Natural Language Communication for Cooperative Autonomous Driving via Multi-Agent Learning

DGX agent

arXiv:2505.18334v2 Announce Type: replace-cross Abstract: Past work has demonstrated that autonomous vehicles can drive more safely if they communicate with each other. However, this communication is

safetyarxiv-cs-ai
28 Jul 2026
Agents

Knowledge-Centric Agents for Workflow Generation in ComfyUI

DGX agent

arXiv:2607.15845v2 Announce Type: replace Abstract: Workflow generation in visual creation systems such as ComfyUI demands not only syntactic accuracy but also expert-level reasoning over modular comp

agentsarxiv-cs-ai
28 Jul 2026
Local Ai

Let AI Agents Translate Networks, Not Reason About Them

DGX agent

arXiv:2607.22947v1 Announce Type: new Abstract: A formal model enables verifying reachability, localizing an outage, or anticipating the blast radius of a change. Yet, virtually no production network

local-aiarxiv-cs-ai
28 Jul 2026
Model Releases

MulRobBench: A Decision-Level Benchmark for Safe and Security-Policy-Compliant Multimodal UAV Agents

DGX agent

arXiv:2607.23870v1 Announce Type: cross Abstract: Smart-city airspace is transforming Uncrewed Aerial Vehicles (UAVs) from passive sensing platforms into cyber-physical decision makers that must follo

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Multi-Agent Privacy Game in Federated Learning: A Unified Mean-Field View

DGX agent

arXiv:2607.23029v1 Announce Type: cross Abstract: Federated learning enables collaborative model training across distributed clients without centralising their data, yet privacy remains a persistent c

agentsarxiv-cs-ai
28 Jul 2026
Safety

The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation

DGX agent

arXiv:2607.24720v1 Announce Type: cross Abstract: Multi-turn long-horizon planning is critical for foundation model agents, yet how to fundamentally improve it remains unclear. Existing models are tra

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

UAV-ON: A Benchmark for Open-World Object Goal Navigation with Aerial Agents

DGX agent

arXiv:2508.00288v5 Announce Type: replace-cross Abstract: Aerial navigation is a fundamental yet underexplored capability in embodied intelligence, enabling agents to operate in large-scale, unstructu

model-releasesarxiv-cs-cv
28 Jul 2026
Safety

What Can Be Enforced? A Theory of Certified Runtime Safety for Tool-Using Agents

DGX agent

arXiv:2607.22868v1 Announce Type: new Abstract: Runtime guardrails act before irreversible tool calls, but their guarantees depend on what policy state is representable, what a judge observes, and whe

safetyarxiv-cs-ai
28 Jul 2026
Agents

Together gives developers an efficient, high-throughput production path for K3’s long, tool-heavy agent workloads, hosted on Together AI’s U…

DGX agent

Together gives developers an efficient, high-throughput production path for K3’s long, tool-heavy agent workloads, hosted on Together AI’s US-based infrastructure with zero data retention. Start build

agentstogether-ai--x
27 Jul 2026
Safety

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy…

DGX agent

Very cool paper from Microsoft. The idea is to train agents on replayed teacher trajectories instead of live environment rollouts. On-policy distillation for agentic tasks is expensive because every u

safetydair-ai--x
27 Jul 2026
Agents

web archive for agent testing

DGX agent

web archive for agent testing 🔍 Introducing BackSearch. LLMs are increasingly asked to predict the future, but a good backtest requires a snapshot of the internet at a point in time. BackSearch allows

agentsyohei-nakajima--x
25 Jul 2026
Model Releases

AppWorld-UL: Benchmarking Diverse Agent-User Interactions for Tool-Use

DGX agent

arXiv:2607.20536v1 Announce Type: new Abstract: Tool-use agents that address day-to-day digital tasks such as ordering groceries must not only operate applications, but also interact with the user, e.

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…96979899100…374
Next →