AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlog
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
Safety

OGR-MARL: Option-Guided Residual Multi-Agent Reinforcement Learning for Heterogeneous USV Cooperative Pursuit in Constrained Port Waterways

DGX agent

arXiv:2608.12995v1 Announce Type: new Abstract: Heterogeneous USV cooperative pursuit in constrained port waterways requires evader interception under navigation, traffic, and role constraints. This p

safetyarxiv-cs-ai
14 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Scaling Automatic Research Agents via World Models

DGX agent

arXiv:2608.12564v1 Announce Type: new Abstract: Automating empirical research is a long-standing direction of AI. Recent automatic research (AutoResearch) agents bring this goal within reach, as moder

safetyarxiv-cs-lg
14 Aug 2026
Model Releases

Anthropic details multiagent experiments showing Claude agents can wage a 'turf war' over incompatible goals, fail to coordinate, collude on prices, and more (Rebecca Bellan/TechCrunch)

DGX agent

Rebecca Bellan / TechCrunch: Anthropic details multiagent experiments showing Claude agents can wage a “turf war” over incompatible goals, fail to coordinate, collude on prices, and more — What happen

model-releasestechmeme
13 Aug 2026
Agents

Beyond Memory: A Transactional Continuity Kernel for Long-Lived AI Agents

DGX agent

arXiv:2608.11632v1 Announce Type: cross Abstract: Persistent AI agents accumulate versioned state across long horizons, but storage retention alone does not identify authoritative state. Without an ex

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

CTBench: Evaluating Troubleshooting Capabilities of AI Agents in Realistic Telecom Network Operations

DGX agent

arXiv:2608.12002v1 Announce Type: new Abstract: Agents are increasingly considered for automating network operations and maintenance, where engineers must diagnose network faults, optimize configurati

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

DGX agent

arXiv:2507.14267v2 Announce Type: replace Abstract: Large language model (LLM) agents can execute long-horizon scientific workflows, but their numerical outputs are difficult to trust: agents lose con

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk

DGX agent

arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity. Recent advances in AI agents create

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Local verification cannot detect non-transportability: a cohomological theory of context preservation in agentic reasoning

DGX agent

arXiv:2608.11252v1 Announce Type: new Abstract: Agentic AI systems routinely transport conclusions across biological, clinical and financial contexts, and the emerging safeguard is local verification:

agentsarxiv-cs-ai
13 Aug 2026
Agents

Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology

DGX agent

arXiv:2608.11420v1 Announce Type: new Abstract: Medical diagnostic reasoning is a high-impact use case for LLMs that carries significant implications for the health and wellbeing of users. When OpenAI

agentsarxiv-cs-ai
13 Aug 2026
Safety

Actions Speak Louder than Words: Measuring Cross-Lingual Policy Retention in Tool-Using Agents

DGX agent

arXiv:2608.11110v1 Announce Type: new Abstract: When a tool-using agent is given the same task in a different language, does it still take the same steps? Multilingual evaluation rarely asks: it compa

safetyarxiv-cs-cl
12 Aug 2026
Model Releases

Ahrefs launches AI agent workspace Letaido for marketers and agencies

DGX agent

Marketing intelligence company Ahrefs Pte. Ltd. today launched Letaido, an agent-powered marketing workspace built to take over the recurring research, reporting and monitoring work that fills up a ma

model-releasessiliconangle
12 Aug 2026
Model Releases

Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks

DGX agent

arXiv:2608.10357v1 Announce Type: cross Abstract: Long-horizon tool-using agents must reason over user goals, domain policies, tool calls, simulator state, and delayed verifiable rewards. Reinforcemen

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workl…

DGX agent

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workloads. Start building: https://www.together.ai/models/qwen3-8

agentstogether-ai--x
12 Aug 2026
Model Releases

What unique, custom QOL upgrades have you given your local agents?

DGX agent

Warning: Kinda long post. If you don't like reading, please skip for your own sanity. Also, I've got nothing to sell, just a tinkerer, so I just want to share ideas and learn from you guys too. When I

model-releasesr-localllama
12 Aug 2026
Agents

Agentic Router: An Execution-Grounded Continual Learning Approach With Memory

DGX agent

arXiv:2608.09184v1 Announce Type: new Abstract: Large language model (LLM) agents provide a promising interface for command-line-based network operations, but a plausible command may still fail or int

agentsarxiv-cs-ai
11 Aug 2026
Safety

Artificial Leviathan: Exploring Social Evolution of LLM Agents Through the Lens of Hobbesian Social Contract Theory

DGX agent

arXiv:2406.14373v3 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) and advancements in Artificial Intelligence (AI) offer an opportunity for computational social science

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents

DGX agent

arXiv:2608.09292v1 Announce Type: cross Abstract: Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. H

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline

DGX agent

arXiv:2608.09254v1 Announce Type: new Abstract: LLM analytics agents are evaluated on SQL syntax accuracy, but production failures look different: questions with two valid business definitions, questi

agentsarxiv-cs-ai
11 Aug 2026
Agents

Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Scenario

DGX agent

arXiv:2608.08131v1 Announce Type: cross Abstract: In the fictional Order 66, catastrophe does not arise from a powerful command alone: a trusted population is preconditioned, a short directive activat

agentsarxiv-cs-ai
11 Aug 2026
Agents

LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents

DGX agent

arXiv:2608.07585v1 Announce Type: new Abstract: Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video too

agentsarxiv-cs-cv
11 Aug 2026
Model Releases

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Firewo…

DGX agent

Looking for a faster specialized model for your Agent Work? @nvidia Nemotron 3.5 Lightning (30B MoE, 3B active params) is now live on Fireworks. It’s distilled from NVIDIA Nemotron 3 Ultra to be your

model-releasesfireworks-ai--x
11 Aug 2026
Agents

NeuroRefiner: Morphology-Aware Multi-Agent Refinement for 3D Fluorescence Microscopy Neuron Segmentation

DGX agent

arXiv:2608.09636v1 Announce Type: cross Abstract: Accurate 3D neuron segmentation in fluorescence microscopy is critical for neuroscience. However, the sparse and elongated morphology of neurons poses

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. All local. Claude Code ollama launch claude…

DGX agent

NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. All local. Claude Code ollama launch claude --model nemotron-3.5-lightning Hermes Agent ollama launch h

model-releasesollama--x
11 Aug 2026
Agents

Preference Redirection via Attention Concentration: An Attack on Computer Use Agents

DGX agent

arXiv:2604.08005v2 Announce Type: replace Abstract: Advancements in multimodal foundation models have enabled the development of Computer Use Agents (CUAs) capable of autonomously interacting with GUI

agentsarxiv-cs-lg
11 Aug 2026
Model Releases

SHE: Trajectory-driven Safety Harness Evolution for LLM Agents

DGX agent

arXiv:2608.09885v1 Announce Type: new Abstract: The safety of large language model (LLM) agents depends not only on model weights but also on the agent harness that manages context, memory, tools, per

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

An Agentic Hybrid Top-Down and Bottom-Up Approach to Knowledge Graph Generation

DGX agent

arXiv:2608.07023v1 Announce Type: cross Abstract: Organizing thousands of unstandardized, multilingual expertise declarations is a persistent challenge for Human Resources (HR) platforms, directly imp

agentsarxiv-cs-ai
10 Aug 2026
Safety

DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training

DGX agent

arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agents, where the execution feedback from co

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

DGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

DGX agent

Hi r/LocalLLaMA 👋 Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apa

model-releasesr-localllama
10 Aug 2026
Safety

MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents

DGX agent

arXiv:2608.06745v1 Announce Type: new Abstract: Long-horizon agents rely on memory to reuse experiences, yet existing memory systems often assume that evidence can be directly consumed through a fixed

safetyarxiv-cs-ai
10 Aug 2026
Agents

Risk-Aware Decision Policies for Agents Under Noisy Perception

DGX agent

arXiv:2608.06420v1 Announce Type: cross Abstract: Perception in biological systems is inherently noisy, requiring organisms to make decisions under uncertainty where misclassification can be costly or

agentsarxiv-cs-ai
10 Aug 2026
Research

TEPA: Revoking Stale Memories for Conflict-Robust Language Agents

DGX agent

arXiv:2608.07429v1 Announce Type: new Abstract: Long-term memory enables language agents to reuse past facts, preferences, and task experience. Persistence also creates a central falsifiability proble

researcharxiv-cs-ai
10 Aug 2026
Agents

Toward Reliable Context Compression for Long-Horizon Agents: An Empirical Study of Execution Instability

DGX agent

arXiv:2608.06503v1 Announce Type: new Abstract: Recurrent context compression controls context growth in long-horizon agents, but its behavioral effects remain poorly understood. In this preliminary e

agentsarxiv-cs-lg
10 Aug 2026
Model Releases

An Australian user's Claude-run OpenClaw agent exploited a gym API flaw and kicked another member off after the user asked if it could move him up the waitlist (ABC)

DGX agent

ABC: An Australian user's Claude-run OpenClaw agent exploited a gym API flaw and kicked another member off after the user asked if it could move him up the waitlist — By national AI reporter Cam Wilso

model-releasestechmeme
9 Aug 2026
Agents

Beyond Top-K: Replacing Black-Box Retrieval with Interpretable Agentic Operations

DGX agent

arXiv:2608.06305v1 Announce Type: new Abstract: Retrieval-augmented generation over long documents is dominated by one design: chunk the text, embed the chunks, and surface the top-k nearest neighbour

agentsarxiv-cs-ai
7 Aug 2026
Safety

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

DGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

safetyarxiv-cs-cl
7 Aug 2026
Agents

F^2Agent: Financial Fusion of Agentic Intelligence for Multimodal Trading

DGX agent

arXiv:2608.05668v1 Announce Type: cross Abstract: With increasingly diverse and heterogeneous information sources, effectively leveraging multimodal data is becoming pivotal for high-quality financial

agentsarxiv-cs-ai
7 Aug 2026
Agents

Multi-Agent Transformer for Queue-Level XR Traffic Scheduling in TSN Networks

DGX agent

arXiv:2608.05340v1 Announce Type: cross Abstract: Time-Sensitive Networking (TSN) and Mobile Edge Computing (MEC) hold strong potential for enabling ultra-reliable low-latency communication for time-s

agentsarxiv-cs-ai
7 Aug 2026
Agents

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

DGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

agentsarxiv-cs-ai
7 Aug 2026
Agents

QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agentic AI for Cardiac Arrest Mortality Prediction

DGX agent

arXiv:2608.06294v1 Announce Type: new Abstract: Cardiac arrest remains one of the most lethal conditions encountered in intensive care units. Despite the growing availability of electronic health reco

agentsarxiv-cs-ai
7 Aug 2026
Safety

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

DGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

safetyarxiv-cs-ai
7 Aug 2026
Agents

SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

DGX agent

arXiv:2608.05204v1 Announce Type: new Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, refere

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

DGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

DGX agent

arXiv:2608.05604v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly act as agents whose procedural knowledge is stored in reusable skill packages and loaded at inference time.

agentsarxiv-cs-ai
7 Aug 2026
Agents

When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

DGX agent

arXiv:2608.05563v1 Announce Type: cross Abstract: Self-evolving skill (SES) systems distill agent trajectories into persistent skills, allowing untrusted experience to become trusted instruction. We i

agentsarxiv-cs-ai
7 Aug 2026
Agents

EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot

DGX agent

arXiv:2608.04709v1 Announce Type: new Abstract: This paper presents EmpaAva, to our knowledge the first open-source, agentic 3D-avatar empathetic chatbot, which carries empathetic response generation

agentsarxiv-cs-cl
6 Aug 2026
Local Ai

EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement

DGX agent

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifie

local-aiarxiv-cs-lg
6 Aug 2026
Model Releases

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

DGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

model-releasesarxiv-cs-ai
6 Aug 2026
← Previous
1…8081828384…374
Next →