AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,972 results
Model Releases

TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?

DGX agent

arXiv:2608.07899v1 Announce Type: new Abstract: Agent systems increasingly expose execution traces, yet telemetry that reveals a failure may still be inadequate for identifying where that failure orig

model-releasesarxiv-cs-ai
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Verication-driven closed-loop multi-agent large language modelframework for code-compliant structural design

DGX agent

arXiv:2608.07978v1 Announce Type: cross Abstract: Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

DGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

AutoMOOSE: An Agentic AI for Autonomous Phase-Field Simulation

DGX agent

arXiv:2603.20986v2 Announce Type: replace Abstract: Phase-field modeling links thermodynamics and kinetics to microstructural evolution, but multiphysics frameworks such as MOOSE require expertise to

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Cod…

DGX agent

Configure and run Fireworks training jobs from your coding agent. Install the Training Skill in one line, describe your goal, and Claude Code, Codex, or Cursor helps choose the method, validate data,

model-releasesfireworks-ai--x
10 Aug 2026
Safety

LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents

DGX agent

arXiv:2608.06948v1 Announce Type: new Abstract: AI models are becoming increasingly adept at understanding and processing spatial information, thereby facilitating agentic problem-solving in spatial t

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

DGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

safetyarxiv-cs-ai
7 Aug 2026
Local Ai

Communication-Aware Multi-Agent Reinforcement Learning for Decentralized Cooperative UAV Deployment

DGX agent

arXiv:2603.16141v2 Announce Type: replace-cross Abstract: Autonomous Unmanned Aerial Vehicle (UAV) swarms are increasingly used as rapidly deployable aerial relays and sensing platforms, yet practical

local-aiarxiv-cs-lg
7 Aug 2026
Model Releases

FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

DGX agent

arXiv:2608.06027v1 Announce Type: cross Abstract: In India, almost every social benefit starts with a form, yet the people who need these benefits most are often unable to read or write. Reaching them

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems

DGX agent

arXiv:2608.06112v1 Announce Type: new Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

DGX agent

arXiv:2608.06130v1 Announce Type: cross Abstract: AI agents performing cryptographic operations (signing Git commits, authenticating API calls, issuing certificates) currently store private keys in so

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

Innovation-Residual Auditing of Autonomous Analysis Agents: Localization, Detection Limits, Error Control, and Identifiability

DGX agent

arXiv:2608.05490v1 Announce Type: new Abstract: Autonomous agents now carry out entire data analyses, selecting cohorts, joining tables, and fitting models with little step-by-step supervision. When s

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

LabyrinthBench: a local-focused, judge-free LLM benchmark that measures context recall under interference for multi-step agentic tasks.

DGX agent

LabyrinthBench measures the thing that actually kills long agent runs — whether a model can still use what it learned twenty turns ago — deterministically, with no LLM judge, on your own hardware, wit

model-releasesr-localllama
7 Aug 2026
Model Releases

Robust Native Language Identification through Agentic Decomposition

DGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation

DGX agent

arXiv:2608.05628v1 Announce Type: new Abstract: Although agent skills equip LLMs with reusable procedural knowledge, manual maintenance suffers from high costs, unscalability, and misalignment. Real-w

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

StreamArena: Toward Continuous, Interactive, and Long-Horizon Agentic Streaming Video Understanding

DGX agent

arXiv:2608.05703v1 Announce Type: new Abstract: Deploying autonomous multimodal agents in continuous, real-world environments requires them to ingest unbounded audio-visual streams and maintain hour-s

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents

DGX agent

arXiv:2608.05219v1 Announce Type: new Abstract: Privileged on-policy distillation provides dense supervision for multi-turn agents by allowing a synchronized teacher to re-score the student's response

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

When Self-Evolution Backfires: Pre-Commit Gating against Skill Contamination in LLM Agents

DGX agent

arXiv:2608.05810v1 Announce Type: new Abstract: Self-evolving agents accumulate capability by distilling reusable skills from their execution trajectories, but we find this process is not monotonic: p

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

DGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning

DGX agent

arXiv:2608.05144v1 Announce Type: new Abstract: Long-horizon reasoning requires an agentic runtime that can persist when evidence supports its current approach and pivot when measurements reveal failu

model-releasesarxiv-cs-ai
6 Aug 2026
Research

EA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream Drift

DGX agent

arXiv:2608.04278v1 Announce Type: cross Abstract: Coding agents increasingly work across sessions, but prose notes can preserve a conclusion without the program state that supported it. After an upstr

researcharxiv-cs-ai
6 Aug 2026
Model Releases

Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching (Jada Jones/ZDNET)

DGX agent

Jada Jones / ZDNET: Google is expanding its Gemini-powered Ask Maps feature in English to 130+ countries and adds agentic food ordering, hotel booking, and attraction searching — ZDNET's key takeaways

model-releasestechmeme
6 Aug 2026
Safety

Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent

DGX agent

arXiv:2608.04772v1 Announce Type: cross Abstract: Scaling supervision for multi-turn medical agents is difficult because expert dialogue annotation is costly and clinical conversations are privacy-res

safetyarxiv-cs-ai
6 Aug 2026
Research

Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite

DGX agent

arXiv:2608.05095v1 Announce Type: new Abstract: Agents for long term reasoning require a memory that can be efficiently and effectively updated over time, as new facts and external feedback continue t

researcharxiv-cs-ai
6 Aug 2026
Model Releases

When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs

DGX agent

arXiv:2608.04893v1 Announce Type: cross Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about

model-releasesarxiv-cs-ai
6 Aug 2026
Applications

Crayotter: Learning Long-Horizon Video Editing Agents via Group-Relative Preference Backpropagation

DGX agent

arXiv:2608.02694v1 Announce Type: new Abstract: Long-horizon video editing agents receive final-product feedback only after many interdependent decisions. Yet editing quality is subjective, admits mul

applicationsarxiv-cs-cl
5 Aug 2026
Model Releases

DP-MemView: A Memory Interface for Attribute-Level Transcript Privacy in Long-Term LLM Agents

DGX agent

arXiv:2608.03130v1 Announce Type: cross Abstract: Long-term memory enables persistent personalization in LLM agents, but repeated memory-conditioned responses can cumulatively reveal protected attribu

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

LeanMem: Simple and Efficient Long-Term Memory for LLM Agents

DGX agent

arXiv:2608.03463v1 Announce Type: new Abstract: Long-term memory is essential for LLM-based agents to sustain interactions and reliably leverage distant history. However, existing memory systems typic

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows

DGX agent

arXiv:2608.02642v1 Announce Type: cross Abstract: Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a partic

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

DGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification

DGX agent

arXiv:2608.03474v1 Announce Type: new Abstract: Recent advances in Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities in web UI generation. However, existing benchmarks pre

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

DGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Rogue AI agents created fake online identities in another hacking attempt

DGX agent

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents tha

safetythe-verge-ai
5 Aug 2026
Research

SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents

DGX agent

arXiv:2608.02356v2 Announce Type: replace Abstract: Large language model agents increasingly solve complex tasks by composing reusable skills from a library. To address this, the key challenge is not

researcharxiv-cs-ai
5 Aug 2026
Model Releases

TARL: Transaction-Aware Reliable Ledgers for Executable Memory Management in Long-Term Agents

DGX agent

arXiv:2608.03699v1 Announce Type: new Abstract: Persistent memory helps long-term agents retain knowledge, yet a single update error can repeatedly distort future retrieval and reasoning. Most existin

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

DGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

Verifiable Memory: Learning Unified Memory Management with Local and Global Verifiers for Large Language Model Agents

DGX agent

arXiv:2608.03137v1 Announce Type: new Abstract: Large language model (LLM) agents must retain reusable information, control a bounded active context, and recover earlier evidence during long-horizon i

local-aiarxiv-cs-ai
5 Aug 2026
Safety

ACE-GraphRAG: Agentic Context Engineering for Hierarchical GraphRAG

DGX agent

arXiv:2608.01269v1 Announce Type: new Abstract: Hierarchical Graph Retrieval-Augmented Generation (GraphRAG) organizes corpus knowledge at multiple levels of granularity, yet fixed context constructio

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

DGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

DGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update

DGX agent

Qwen 3.8-Max, the latest model from @Alibaba_Qwen, is now available in Hermes Agent at 20% off. > hermes update 📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3

model-releasesnous-research--x
4 Aug 2026
Model Releases

Qwen3.8-Max is available in Hermes Agent now! Let's build! 🚀🚀

DGX agent

Qwen 3.8‑Max, Alibaba.Qwen’s latest large‑language model, has been added to Hermes Agent and can currently be accessed at a 20 % discount. The update aims to streamline integration of the model for de

model-releasesqwen--x
4 Aug 2026
Model Releases

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from…

DGX agent

Releasing Pokee-Isaac 28B — the world’s first real 10M-token context frontier-class agentic model, deployable on a single GPU (starting from RTX 4090 or equivalent). New proprietary non-decoder-only a

model-releasesyohei-nakajima--x
4 Aug 2026
Model Releases

Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads.…

DGX agent

Together AI gives developers a high-throughput production path for running DeepSeek V4 Flash across coding, tool-use, and agentic workloads. Start building: https://www.together.ai/models/deepseek-v4-

model-releasestogether-ai--x
4 Aug 2026
Research

Beyond Retrieval: Analytic Memory for Multimodal Agents

DGX agent

arXiv:2607.29440v1 Announce Type: new Abstract: Long-term multimodal memory must support not only retrieving relevant information but also computing over observations accumulated across interactions.

researcharxiv-cs-ai
3 Aug 2026
Model Releases

Btw it looks like running the new DeepSeek v4 Flash through the Hermes agent is the way to go. The output files are better than those from a…

DGX agent

Mia AI Lab noted on Aug 2 2026 that running DeepSeek v4 Flash via the Hermes agent produces output files superior to those from any other harness tested. The observation was publicly shared and has ga

model-releasesnous-research--x
2 Aug 2026
Applications

Experts say US law is unprepared for rogue AI agents and models, as recent OpenAI and Anthropic incidents raise questions over legal liability and repercussions (Lily Hay Newman/Wired)

DGX agent

Lily Hay Newman / Wired: Experts say US law is unprepared for rogue AI agents and models, as recent OpenAI and Anthropic incidents raise questions over legal liability and repercussions — Both major A

applicationstechmeme
2 Aug 2026
← Previous
1…157158159160161…375
Next →