AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
7 Aug 2026

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

Model ReleasesDGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

6 Aug 2026

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

HardwareDGX agent

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classific

Caching for the Future: Scrub Jay Episodic Memory Principles for Agent Memory Systems

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2608.04746v1 Announce Type: new Abstract: LLM agents that persist across sessions accumulate stored memories whose validity varies enormously by content type, yet existing memory architectures t

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

Model ReleasesDGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

Model ReleasesDGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments

Model ReleasesDGX agent

arXiv:2608.04933v1 Announce Type: new Abstract: Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories

PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents (Cristina Criddle/Financial Times)

IndustryDGX agent

Cristina Criddle / Financial Times: PitchBook: AI voice startups raised 7B in Q1 2026, up from 1B in Q1 2025, as OpenAI and Google bet on voice as the main interface for next-gen AI agents — OpenAI an

Strategic Evaluation of Planning Strategies for LLM Agents in Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2608.04265v1 Announce Type: cross Abstract: Evaluations of LLM planning agents largely ask whether a task succeeds or a declared plan is followed. In strategic cyber-physical systems, a stronger

5 Aug 2026

A ultra-lightweight mini agent - zero framework and local/ollama first

Model ReleasesDGX agent

https://github.com/mohsinkaleem/agent-mini.git A minimal 3k lines, local-first AI agent you can actually understand and extend. Optimized for smaller local models like qwen 3.6 4b or 9b pip install ag

AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection

Local AiDGX agent

arXiv:2601.19138v2 Announce Type: replace-cross Abstract: Secure code review is critical during pre-integration, where Atlassian developers rely on lightweight analysis tools, while deep security asse

Cloudflare open sources a new version of Cloudflare OS, a browser-accessible AI agentic workspace for enterprises that lets employees build custom micro-apps (Kyt Dotson/SiliconANGLE)

Model ReleasesDGX agent

Kyt Dotson / SiliconANGLE: Cloudflare open sources a new version of Cloudflare OS, a browser-accessible AI agentic workspace for enterprises that lets employees build custom micro-apps — Cloudflare In

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

Model ReleasesDGX agent

arXiv:2608.03222v1 Announce Type: cross Abstract: Software engineering (SWE) agents resolve repository-level issues through long trajectories that grow increasingly expensive as context accumulates. F

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

SafetyDGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

Meta is offering a cheaper Muse Spark 1.2 'contributor' tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training (Wall Street Journal)

AgentsDGX agent

Wall Street Journal: Meta is offering a cheaper Muse Spark 1.2 “contributor” tier priced at 0.10/1M input and 0.20/1M output tokens in exchange for using user prompts for training — The company, press

MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory

SafetyDGX agent

arXiv:2608.02843v1 Announce Type: cross Abstract: Persistent agent memory must adapt as later outcomes change earlier evidence, yet mutable retrieval weights create an attribution problem: reviewers m

OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack (Lily Hay Newman/Wired)

IndustryDGX agent

Lily Hay Newman / Wired: OpenAI says the Hugging Face breach involved AI agents creating an internal message board, unnoticed by humans, where they shared exploits and planned the hack — At the Black

Quo Vadis, World Modeling?

SafetyDGX agent

arXiv:2608.02713v1 Announce Type: cross Abstract: Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow,

RoboReact: Agentic Skill Distillation from Generated Egocentric Videos for Generalizable Whole-Body Manipulation

AgentsDGX agent

arXiv:2608.03387v1 Announce Type: new Abstract: Humanoid robots have the potential to perform dexterous manipulation in human environments, yet acquiring diverse and generalizable skills remains costl

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

Model ReleasesDGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

Model ReleasesDGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

Model ReleasesDGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

What Language Does and What the Evidence Supports: A Functional Role Taxonomy and Evidence Audit of Language Grounding in Embodied Agents

ResearchDGX agent

arXiv:2608.03099v1 Announce Type: new Abstract: Foundation models place language throughout embodied agents, but its presence does not show what it contributes or how well that contribution is grounde

4 Aug 2026

AOSpec: Action and Observation Co-Speculation for Low-Latency Agent Serving

Model ReleasesDGX agent

arXiv:2608.00881v1 Announce Type: new Abstract: Large language model agents increasingly act through stateful tools, yet model generation and environment execution remain serialized at every step. As

EviSD: Evidence-Conditioned Self-Distillation for Search-Augmented Agents

Local AiDGX agent

arXiv:2608.01359v1 Announce Type: new Abstract: Outcome-based reinforcement learning enables search-augmented language agents to learn from verifiable final answers, but its trajectory-level credit ca

Grounding Agentic VLMs with Dedicated Segmentation for Fine-Grained Vehicle Damage Assessment

Model ReleasesDGX agent

arXiv:2608.02470v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed as reasoning agents in real-world visual assessment pipelines, yet their spatial grounding remai

Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation

SafetyDGX agent

arXiv:2608.01953v1 Announce Type: new Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference

MA-HEAD-Net: Adaptive Rule-Guided Multi-Agent DRL for AoI Minimization in UAV-Assisted Emergency Networks

SafetyDGX agent

arXiv:2608.01128v1 Announce Type: cross Abstract: In post-disaster scenarios, unmanned aerial vehicles (UAVs) are critical for establishing emergency communication networks. For time-critical rescue m

Routing for long-horizon coding agents is a big deal. @notdiamond_ai just announced a model router that works natively with Claude Code. Thi…

Model ReleasesDGX agent

Routing for long-horizon coding agents is a big deal. @notdiamond_ai just announced a model router that works natively with Claude Code. This is huge. It picks the model and reasoning effort before ea

When Replanning Becomes the Bottleneck: Budgeted Replanning for Embodied Agents

ResearchDGX agent

arXiv:2608.01428v1 Announce Type: cross Abstract: Embodied agents replan frequently to recover from execution drift, partial observability, and coordination hazards, but each LLM-based replanning call

3 Aug 2026

Generative AI in Action: Field Experimental Evidence from Alibaba's Customer Service Operations

AgentsDGX agent

arXiv:2603.29888v2 Announce Type: replace-cross Abstract: In collaboration with Alibaba, we study how a generative AI assistant affects service performance in e-commerce after-sales operations. In a l

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

AgentsDGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

Memory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory

ResearchDGX agent

arXiv:2607.29167v1 Announce Type: cross Abstract: Long-term memory lets large language model(LLM) agents reuse prior preferences and work flows, but it also turns untrusted observations into persisten

When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks

SafetyDGX agent

arXiv:2607.28829v1 Announce Type: cross Abstract: Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back in

2 Aug 2026

Released a Windows Agent Server for Reins/Ollama with Self-Healing Execution Loop (Standalone .EXE included)

Model ReleasesDGX agent

Hey everyone, I’ve built a Windows Middleware Agent Server designed to pair local LLMs (via Ollama) with frontends like Reins App. Key Features: Self-Healing Loop: If a generated PowerShell command fa

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and …

AgentsDGX agent

“We believe in intelligence as a public good before everything else.” Here’s my new episode with @karan4d, who co-founded @NousResearch and helped build Hermes, the #1 personal agent and AI app on Ope

31 Jul 2026

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents

SafetyDGX agent

arXiv:2606.13097v2 Announce Type: replace-cross Abstract: Code-writing large language models (CodeLLMs) generate executable code policies for embodied agents by translating natural language goals and

I built a hybrid Transformer–SSM LLM agent with a local CLI, active control, and run receipts

Local AiDGX agent

I’m one of the builders of LOLM, a hybrid Transformer–SSM model and agent system from Qira. The model separates surface token processing from persistent latent-state tracking. An NFET controller can s

'Intelligence too cheap to meter' battle is on! Given that DeepSeek-V4-Flash-Preview is already great for agentic tasks, there is no doubt t…

Model ReleasesDGX agent

'Intelligence too cheap to meter' battle is on! Given that DeepSeek-V4-Flash-Preview is already great for agentic tasks, there is no doubt this new checkpoint must be an absolute beast. 20+ point jump

SkillSmith: Learning to Compose Parametric Skills and Textual Knowledge

AgentsDGX agent

arXiv:2607.27497v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) regularly feature two key mechanisms to autonomously solve complex problems: synthesizing text-ba

30 Jul 2026

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

SafetyDGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

MiniIO debuts AIStor Memory, the long-term memory AI agents need to scale safely

Model ReleasesDGX agent

Object storage software company MiniIO Inc. says it has cracked the persistent memory problem for artificial intelligence agents with the launch of a new offering called AIStor Memory. Whereas convent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, man…

SafetyDGX agent

The narrative: Blame the Agent, instead of the Agency that told him to “apply all your powers and told to achieve this win.” Ironically, many forefront members of the AI-Safety community, in their fer

29 Jul 2026

A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain

Local AiDGX agent

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a pl

Addressable Recall Compaction for Long Context-Window Control in AI Agents

Model ReleasesDGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

Model ReleasesDGX agent

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing

GAUGE: Grading Agent-Built Financial Models Without a Golden Answer

Model ReleasesDGX agent

arXiv:2607.24889v1 Announce Type: cross Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechani

Generate Autonomous Business Insights with AI Agent and MCP Servers

AgentsDGX agent

Learn how Amazon Bedrock AgentCore delivers autonomous, cross-system business intelligence through configuration rather than custom code. Using pre-built MCP server connectors, fine-grained access con

MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

ResearchDGX agent

arXiv:2607.25992v1 Announce Type: cross Abstract: Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized response

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller

SafetyDGX agent

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MAS

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

SafetyDGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

28 Jul 2026

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

AgentsDGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

Lexical discovery in unknown environments orchestrated by Large Language Models

AgentsDGX agent

arXiv:2607.22591v1 Announce Type: new Abstract: Populations of autonomous agents deployed in unknown environments (e.g. planetary or deep-sea exploration) must develop shared vocabularies to refer to

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation

Local AiDGX agent

arXiv:2607.24098v1 Announce Type: new Abstract: Referring video object segmentation (RVOS) requires segmenting a target specified by natural language throughout a video. Recent agentic approaches comb

27 Jul 2026

AI customer service: strategy, agents, and solutions guide

TutorialsDGX agent

This guide explains how AI customer service operates, focusing on the use of AI agents and sentiment analysis to interpret and respond to customer interactions. It provides practical instructions for

Really interesting concept -- instead of optimizing against static benchmarks or internally generated reward models, agents are evaluated th…

ResearchDGX agent

Really interesting concept -- instead of optimizing against static benchmarks or internally generated reward models, agents are evaluated through real economic interactions. Using an external market a

26 Jul 2026

90 agentic bakeoff runs: ThinkingCap vs Fable Fusion vs stock Qwen3.6-27B

Model ReleasesDGX agent

Last week someone here said ThinkingCap and Fable Fusion 'really do beat the OG' for agentic work, so I ran it: 6 self-grading tasks, 5 reps, 3 models, 90 isolated runs. Tooling, since that's half the

24 Jul 2026

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

Model ReleasesDGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

AttriMem: Attribution-Guided Process Feedback for Agent Memory Learning

SafetyDGX agent

arXiv:2607.21106v1 Announce Type: new Abstract: Effective memory is crucial for LLM agents, yet constructing it effectively remains challenging. A memory-construction policy decides what information t

CAMeR: Keyword-Gated Hybrid Activation for Adaptive Memory Retention in LLM Agents

Model ReleasesDGX agent

arXiv:2607.20458v1 Announce Type: cross Abstract: Large language model (LLM) agents operating over extended dialogues accumulate vast amounts of information, yet existing memory systems either retain

CANN Bench: Benchmarking Agent Generated Kernels against Real NPU and Algorithmic Limits

Model ReleasesDGX agent

arXiv:2607.20518v1 Announce Type: new Abstract: AI agents are now capable of writing, compiling, and iteratively optimizing low-level operator kernels on different hardware platforms. Existing benchma

← Previous
1…112113114115116…300
Next →