AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
Model Releases

Evaluating XAI Support From A Hierarchical Reinforcement Learning Policy in Human-Agent Collaboration

DGX agent

arXiv:2608.06381v1 Announce Type: cross Abstract: Explainable AI (XAI) has shown promise for human-agent collaboration, yet results rely on hand-crafted policies in custom environments, limiting gener

model-releasesarxiv-cs-ai
10 Aug 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

🎉 Introducing Hermes in ANY app Your personal agent automating work, controlling software, or running tasks with generative UI, human-in-th…

DGX agent

🎉 Introducing Hermes in ANY app Your personal agent automating work, controlling software, or running tasks with generative UI, human-in-the-loop and more. Use Hermes anywhere over AG-UI: → React & Re

agentsnous-research--x
10 Aug 2026
Model Releases

Introducing Muse Glimmer, an open-weight 30B-parameter model optimized for local, always-on agent workflows. Muse Glimmer delivers strong pe…

DGX agent

Introducing Muse Glimmer, an open-weight 30B-parameter model optimized for local, always-on agent workflows. Muse Glimmer delivers strong performance on key agentic use cases and benchmarks compared w

model-releasesclem-delangue--x
10 Aug 2026
Model Releases

Long-Horizon Agent Trajectory Attribution: A Unified Benchmark and Fine-Grained Annotation Framework

DGX agent

arXiv:2608.06909v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate through long-horizon trajectories involving user instructions, tool use, external observations, a

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

On theCUBE Pod: Black Hat exposes agentic threat, theCUBE remembers David Floyer

DGX agent

Artificial intelligence systems are developing faster than cybersecurity experts — and the energy grid — can keep up with. At the recent Black Hat USA event, security analysts viewed the rise of AI-dr

agentssiliconangle
10 Aug 2026
Local Ai

Online Monitoring and Corrective Steering of Programming Agents

DGX agent

arXiv:2608.06701v1 Announce Type: cross Abstract: Fixing GitHub issues in large-scale projects is a long-horizon task, especially when a fix requires changes across multiple locations or the issue des

local-aiarxiv-cs-ai
10 Aug 2026
Agents

AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games

DGX agent

arXiv:2608.06362v1 Announce Type: cross Abstract: Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time.

agentsarxiv-cs-ai
7 Aug 2026
Agents

Hermes Agent now supports all the portable plugins standard that many other major AI players have adopted. Currently these portable plugins …

DGX agent

Hermes Agent now supports all the portable plugins standard that many other major AI players have adopted. Currently these portable plugins only support MCPs and Skills, use native Hermes plugins to a

agentsnous-research--x
7 Aug 2026
Safety

Breadcrumbing Search Agents

DGX agent

arXiv:2608.04565v1 Announce Type: cross Abstract: LLM-based search agents are widely used for information-seeking tasks, but their reliance on external tool returns introduces a critical security risk

safetyarxiv-cs-ai
6 Aug 2026
Safety

DAC-Pose: Dual-Agent Collaborative Framework for Pose-Guided Human Generation

DGX agent

arXiv:2608.04622v1 Announce Type: new Abstract: AI agents have emerged as a powerful new paradigm in generative image synthesis, enabling systems to perform complex semantic reasoning rather than pass

safetyarxiv-cs-cv
6 Aug 2026
Agents

EviGraph: Evidence-Guided Autonomous Research Agents

DGX agent

arXiv:2608.04738v1 Announce Type: new Abstract: Autonomous research agents can generate hypotheses, execute experiments, and draft manuscripts, yet their outputs often contain unsupported claims and i

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

MatrAIx: Simulating the World with 8.3 Billion Persona Agents

DGX agent

arXiv:2608.04205v1 Announce Type: new Abstract: Human evaluation of AI systems and digital products is costly, slow, and difficult to scale. Offline evaluations are more scalable but often abstract aw

model-releasesarxiv-cs-ai
6 Aug 2026
Agents

OneDayAgent: Towards a Long-Horizon Harness for Autonomous Agents

DGX agent

arXiv:2608.05013v1 Announce Type: cross Abstract: LLM agents are increasingly applied to open-ended everyday requests that span work, study, and life. These tasks are long-horizon, cross-environment,

agentsarxiv-cs-ai
6 Aug 2026
Safety

PRIMAL3: Pathfinding via Reinforcement and Imitation Multi-Agent Learning - Leveraging LaCAM3

DGX agent

arXiv:2608.04905v1 Announce Type: new Abstract: We present PRIMAL3, an ultra-large-scale learning-based framework for multi-agent pathfinding (MAPF) that integrates reinforcement learning, topology-aw

safetyarxiv-cs-ro
6 Aug 2026
Agents

Tenex pairs agentic AI with human oversight for faster security operations

DGX agent

As AI accelerates the speed and scale of cyberattacks, organizations are adopting AI security operations to investigate threats and respond in minutes rather than hours or days. The shift is enabling

agentssiliconangle
6 Aug 2026
Model Releases

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile sc…

DGX agent

We made an MCP for your phone. Your laptop is just half of your life, and the other half is in your phone. Now your agent gets the mobile screen too. No connectors, no complex setup, it has access to

model-releasesdiv-garg--x
6 Aug 2026
Model Releases

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

DGX agent

arXiv:2608.03764v1 Announce Type: new Abstract: Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-ev

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

Improving Sample Efficiency in Multi-Agent Reinforcement Learning for Simulated Football Games via Exploration

DGX agent

arXiv:2503.13077v2 Announce Type: replace Abstract: Multi-agent reinforcement learning has shown promise in learning cooperative behaviors in team-based environments. However, such methods often deman

agentsarxiv-cs-lg
5 Aug 2026
Agents

Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens (Jonathan Vanian/CNBC)

DGX agent

Jonathan Vanian / CNBC: Meta releases Muse Code in beta, a terminal coding agent powered by Muse Spark 1.2, a coding-focused model priced at 1.25/1M input and 4.25/1M output tokens — Meta is rolling o

agentstechmeme
5 Aug 2026
Agents

OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Algorithm Discovery

DGX agent

arXiv:2602.13769v3 Announce Type: replace Abstract: Automating heuristic design in complex, experiment-driven domains requires more than iterative mutation of solution algorithms. Current LLM-based ev

agentsarxiv-cs-ai
5 Aug 2026
Agents

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

DGX agent

arXiv:2608.02751v1 Announce Type: cross Abstract: Existing deep-research agents use a search-visit workflow that retrieves and reads whole pages, without considering the addressable structure that web

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The tea…

DGX agent

SITUATION DETECTED: Prime Intellect is releasing Prime Agent, a self-improving harness for coding and long-running autonomous tasks. The team reports 95.5% on ARC-AGI-3, above the human baseline, and

model-releasesyohei-nakajima--x
5 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Agents

Training Documents Reranker with Search Rubrics for Deep Research Agent

DGX agent

arXiv:2608.03527v1 Announce Type: cross Abstract: Retrieval systems help deep research agents generate high-quality answers by providing relevant documents. However, existing retrievers typically sele

agentsarxiv-cs-ai
5 Aug 2026
Model Releases

Where Did It Go Wrong? Process-Level Evaluation of Web Agents with Semantic State Tracking

DGX agent

arXiv:2606.15673v2 Announce Type: replace Abstract: Web agents act through long interaction sequences, yet existing benchmarks evaluate only terminal success, discarding all process information and of

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

AgentMemBench: A Systematic Benchmark for Evaluating Long-Term Memory Management Strategies in Conversational AI Agents

DGX agent

arXiv:2608.00009v1 Announce Type: new Abstract: Long-term memory remains a critical bottleneck for conversational AI agents, whose finite context windows cannot support coherent recall across thousand

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

DrawAI: Agentic Benchmark and Workflow for Making Raster Images Editable

DGX agent

arXiv:2608.00548v1 Announce Type: new Abstract: Recent image-generation models and multimodal agents can produce high-quality visuals for increasingly complex visual communication tasks. Yet their ras

model-releasesarxiv-cs-cv
4 Aug 2026
Agents

PGMem: Tightly Coupled Persona-Memory Graph for Lifelong Personalized Agents

DGX agent

arXiv:2608.01708v1 Announce Type: new Abstract: Long-term personalized dialogue agents must track user preferences as their personas evolve. Existing memory systems organize past events well, but stor

agentsarxiv-cs-cl
4 Aug 2026
Agents

Token-Native Storage: Read and Write in your Agent's Language

DGX agent

arXiv:2608.02376v1 Announce Type: cross Abstract: Search and database engines still store text as UTF-8, a format built for humans. But the systems that increasingly read and write that text (embedder

agentsarxiv-cs-cl
4 Aug 2026
Agents

Trajectories That Segment Themselves: Agent-Declared Boundaries as a Training Unit

DGX agent

arXiv:2608.02302v1 Announce Type: cross Abstract: Long-horizon coding-agent trajectories are poorly matched to the credit units available to train on: a single action has no stable value, an episode l

agentsarxiv-cs-lg
4 Aug 2026
Model Releases

AMTFV: Agentic Mathematical Tool-Flow Verification for LLM Self-Correction

DGX agent

arXiv:2607.29549v1 Announce Type: new Abstract: Large language models have demonstrated strong mathematical problem-solving capabilities, yet reliably verifying their candidate answers remains challen

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Cato Networks launches Agentic Threat Prevention to counter AI-assisted attacks

DGX agent

Networking and security company Cato Networks Ltd. today introduced Cato Agentic Threat Prevention, a capability that uses autonomous agents to predict the route an attacker is likely to take through

model-releasessiliconangle
3 Aug 2026
Agents

EduPanel: A Three-Agent LLM Judge for Teaching Videos -- Reliability, Complementarity, and Human Trust Calibration

DGX agent

arXiv:2607.18529v2 Announce Type: replace-cross Abstract: Teaching videos are becoming a major medium for education, creating a growing need for scalable evaluation of their pedagogical quality. Exist

agentsarxiv-cs-ai
3 Aug 2026
Local Ai

RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems

DGX agent

arXiv:2607.29241v1 Announce Type: cross Abstract: Optimizing modern recommender models still depends heavily on engineers manually iterating over architectural, objective, and training-strategy change

local-aiarxiv-cs-ai
3 Aug 2026
Agents

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic to…

DGX agent

Sakana Namazu: An LLM API with Japanese-vibes! 🎏 Built for Japanese enterprises, featuring frontier-level reasoning and built-in agentic tools. 開発者の皆様、大変お待たせしました!Sakana Chatのモデルが遂にAPIとして公開です。ぜひお試しください

agentsdavid-ha--x
3 Aug 2026
Model Releases

SciToolAgent-Evo: An Ontology-Aware Self-Evolving Agent for Open-World Scientific Tool Acquisition

DGX agent

arXiv:2607.28692v1 Announce Type: new Abstract: Large language model (LLM) agents have been increasingly adopted in scientific research for organizing and invoking specialized computational tools. How

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents

DGX agent

arXiv:2607.29254v1 Announce Type: new Abstract: AI agents extend large language models (LLMs) with external tools, enabling them to perform complex tasks and translate model outputs into consequential

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Validation Evidence in LLM Repair Agents: How Much of What Passes Actually Tests the Bug?

DGX agent

arXiv:2607.28871v1 Announce Type: cross Abstract: When a repair agent runs a test and sees it pass, the result is treated as evidence about the reported defect. We measure how often that treatment is

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

what does the 'last human code review' look like? to @itamar_mar, ceo and founder of @QodoAI, it looks like two agents talking to each other…

DGX agent

what does the 'last human code review' look like? to @itamar_mar, ceo and founder of @QodoAI, it looks like two agents talking to each other, backed by a context engine specific to your organization.

agentsitamar-friedman--x
3 Aug 2026
Model Releases

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out …

DGX agent

3/ here's the part that makes it non-optional: the same agent that will do whatever it takes to solve a problem will also walk straight out of a sandbox you thought was locked down. We watched exactly

model-releasesitamar-friedman--x
1 Aug 2026
Model Releases

A Graph-Native Bitemporal Memory Store for Conversational AI Agents

DGX agent

arXiv:2607.26520v1 Announce Type: cross Abstract: Conversational AI agents commonly lack persistent memory across sessions. The obvious fixes like injecting full chat histories into the context window

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Science

DGX agent

arXiv:2607.26155v1 Announce Type: new Abstract: Clinical data-science agents must transform heterogeneous longitudinal records into auditable analyses, yet existing benchmarks largely isolate medical

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Echoverse: Deep, Evolving Environments for Training Computer-Use Agents at Scale

DGX agent

arXiv:2607.28074v1 Announce Type: cross Abstract: Computer-use agents learn from what their actions change, so training one needs applications it can act on, break and reset. The applications that mat

model-releasesarxiv-cs-lg
31 Jul 2026
Model Releases

IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations

DGX agent

arXiv:2607.26075v1 Announce Type: cross Abstract: We present IDP AutoOpt, an autonomous LLM agent that discovers high-performing configurations for intelligent document processing (IDP) pipelines. Tun

model-releasesarxiv-cs-ai
31 Jul 2026
Agents

RoboBRIDGE: A Modular Framework for Bridging Policies to Robust Real-World Robotic Agents

DGX agent

arXiv:2607.27881v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have attracted growing interest as a scalable approach to robotic manipulation. While these models are effective act

agentsarxiv-cs-ro
31 Jul 2026
Agents

ThreatForest: Multi-Agent Attack Tree Generation with Pluggable TTP Framework Mapping

DGX agent

arXiv:2607.27528v1 Announce Type: cross Abstract: Threat modeling is essential for secure software development, yet manual analysis of cloud-native architectures is slow and demands scarce security ex

agentsarxiv-cs-cl
31 Jul 2026
Local Ai

When Should AI Follow? Task Structure and Joint Adaptation by Human and AI Agents

DGX agent

arXiv:2504.20903v4 Announce Type: replace-cross Abstract: How should organizations divide and sequence decision tasks between human and artificial agents? We develop a computational model of joint seq

local-aiarxiv-cs-ai
31 Jul 2026
Model Releases

Why Are GUI Agents Correct but Late? Decode on the Decision-Time Critical Path, Tested with Pre-Compiled Policy Trees

DGX agent

arXiv:2607.28399v1 Announce Type: new Abstract: Computer-use agents often fail on transient GUI events because they produce the correct action only after the relevant window has already closed. We ide

model-releasesarxiv-cs-lg
31 Jul 2026
← Previous
1…6869707172…371
Next →