AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,911 results
Agents

pretty sick that i get to work with Jake every day on making continual learning + memory accessible at scale for every single agent one comm…

DGX agent

pretty sick that i get to work with Jake every day on making continual learning + memory accessible at scale for every single agent one common thread here is...the Trace a large part of Continual Lear

agentsharrison-chase--x
24 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

DGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Skills for the future software profession: beyond agentic AI!

DGX agent

arXiv:2606.21894v2 Announce Type: replace-cross Abstract: As coding agents are rapidly changing software engineering, a natural question is: what are the core skills needed by future software engineer

agentsarxiv-cs-ai
24 Jun 2026
Agents

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

DGX agent

arXiv:2606.23743v1 Announce Type: cross Abstract: Modern video diffusion models achieve higher generation quality through scaling, but this also increases inference cost. Although many acceleration me

agentsarxiv-cs-ai
24 Jun 2026
Agents

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

DGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

agentsarxiv-cs-cl
24 Jun 2026
Agents

A trillion tokens a day and 200k GitHub stars Very proud of the Hermes Agent team and what we've built at @NousResearch 'Better today than y…

DGX agent

Nous Research announced that their Hermes Agent has achieved a trillion tokens processed daily and reached 200,000 GitHub stars, reflecting significant adoption and performance milestones for their AI

agentsnous-research--x
23 Jun 2026
Agents

GRADE: Graph Representation of LLM Agent Dependency and Execution

DGX agent

arXiv:2606.22741v1 Announce Type: new Abstract: Can one graph represent every kind of LLM agent's run? A trace records what each step did, never what it relied on, the state it read, and the results i

agentsarxiv-cs-lg
23 Jun 2026
Agents

Hermes Agent can now /learn from anything: feed it directories of any source material (code, API docs, manuals, PDFs, configs) and it distil…

DGX agent

Hermes Agent has been updated to accept diverse source materials including code repositories, API documentation, manuals, PDFs, and configuration files as input, with the ability to distill and learn

agentsnous-research--x
23 Jun 2026
Agents

PolicyGuard: Towards Test-time and Step-level Adversary (Backdoor) Defense for Reinforcement Learning Agent

DGX agent

arXiv:2606.12896v2 Announce Type: replace Abstract: While real-world applications of reinforcement learning (RL) are becoming increasingly popular, the security of RL systems deserve more attention an

agentsarxiv-cs-lg
23 Jun 2026
Agents

UltraQuant: 4-bit KV Caching for Context-Heavy Agents

DGX agent

arXiv:2606.20474v2 Announce Type: replace Abstract: Context-heavy agents place unusual pressure on the key-value (KV) cache: long prefixes are reused across many short turns, while concurrency determi

agentsarxiv-cs-lg
23 Jun 2026
Agents

Hermes Agent now supports computer use via @trycua on Windows and Linux in addition to existing macOS support

DGX agent

Hermes Agent, developed by Nous Research, has expanded its computer use capabilities to support Windows and Linux operating systems in addition to its existing macOS functionality. This enhancement, i

agentsnous-research--x
22 Jun 2026
Safety

APPO: Agentic Procedural Policy Optimization

DGX agent

arXiv:2606.12384v1 Announce Type: cross Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents

safetyarxiv-cs-ai
11 Jun 2026
Agents

ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories

DGX agent

arXiv:2606.11520v1 Announce Type: cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution

agentsarxiv-cs-ai
11 Jun 2026
Agents

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

DGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

agentsarxiv-cs-ai
11 Jun 2026
Agents

Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents

DGX agent

arXiv:2606.10315v1 Announce Type: cross Abstract: LLM-as-judge is the default instrument for evaluating conversational agents, yet its reliability is almost always reported as agreement with human rat

agentsarxiv-cs-ai
10 Jun 2026
Agents

Exclusive: Relai raises $6.9M to enable verifiable and continuous learning for AI agents

DGX agent

Artificial intelligence infrastructure startup Relai Inc. said today it has closed on 6.9 million in funding as it bids to ensure the reliability of autonomous AI agents for enterprises. The company a

agentssiliconangle
10 Jun 2026
Agents

From Confident Closing to Silent Failure: Characterizing False Success in LLM Agents

DGX agent

arXiv:2606.09863v1 Announce Type: new Abstract: LLM agents can fail silently by asserting task completion when the environment state shows otherwise. We study this failure mode, false success, across

agentsarxiv-cs-lg
10 Jun 2026
Agents

Infini Memory: Maintainable Topic Documents for Long-Term LLM Agent Memory

DGX agent

arXiv:2606.10677v1 Announce Type: new Abstract: Long-term LLM agents need persistent memory that can track changing facts and provide relevant evidence across sessions. Existing memory systems often s

agentsarxiv-cs-ai
10 Jun 2026
Agents

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp …

DGX agent

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp & team • DSPy (arXiv:2310.03714) — @lateinteraction + @hazyr

agentsyohei-nakajima--x
10 Jun 2026
Agents

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not …

DGX agent

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not too far now! Excited to get this to work with my @dair_ai co

agentsdair-ai--x
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo!

DGX agent

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo! Excited to launch a new way to upskill with AI agents. This is how we are making it possible for anyone to learn to build with cod

model-releasesdair-ai--x
9 Jun 2026
Agents

ConMem: Structured Memory-Guided Adaptation in Training-Free Multi-Agent Systems

DGX agent

arXiv:2606.08702v1 Announce Type: new Abstract: Recent advances have improved the adaptive capabilities of LLM-based multi-agent systems (MAS) through memory-, skill-, and learning-based approaches, y

agentsarxiv-cs-ai
9 Jun 2026
Agents

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents

DGX agent

arXiv:2606.07904v1 Announce Type: new Abstract: Tool-augmented large language model agents increasingly rely on external APIs, but standard tool schemas describe how to call a tool, not when the tool

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

DGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems

DGX agent

arXiv:2605.12213v2 Announce Type: replace Abstract: LLM-based conversational AI agents struggle to maintain coherent behavior over long horizons due to limited context. While RAG-based approaches are

agentsarxiv-cs-ai
9 Jun 2026
Agents

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and re…

DGX agent

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and remain frozen or mostly unchanged. The harness, like the skill

agentsdair-ai--x
9 Jun 2026
Agents

SIGA: Self-Evolving Coding-Agent Adapters for Scientific Simulation

DGX agent

arXiv:2606.09774v1 Announce Type: new Abstract: Advanced scientific simulators expose specialized input languages that turn simulation goals into executable configurations, but learning them can cost

agentsarxiv-cs-ai
9 Jun 2026
Agents

Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing

DGX agent

arXiv:2606.08256v1 Announce Type: new Abstract: Verifiability, attribution, and reproducibility are foundational requirements of scientific knowledge, yet current publishing infrastructure does not en

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation

DGX agent

arXiv:2606.08091v1 Announce Type: new Abstract: Recent agent frameworks such as Claude Code, Codex, and OpenClaw are strong at tool use and orchestration, but whether they can handle long video genera

model-releasesarxiv-cs-cv
9 Jun 2026
Agents

Autonomous computational catalysis through an agentic research system

DGX agent

arXiv:2601.13508v4 Announce Type: replace-cross Abstract: Autonomous agents are beginning to transform scientific research from tool-assisted workflows toward self-sustaining discovery processes. Comp

agentsarxiv-cs-ai
8 Jun 2026
Agents

Exploring Agentic Tool-Calling Decisions via Uncertainty-Aligned Reinforcement Learning

DGX agent

arXiv:2606.06976v1 Announce Type: new Abstract: Large language model (LLM)-based agents often make suboptimal tool-use decisions, including unsupported tool invocation and hallucinated direct response

agentsarxiv-cs-ai
8 Jun 2026
Agents

How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope

DGX agent

arXiv:2606.07489v1 Announce Type: new Abstract: Frontier AI systems are bridging the gap between intelligence and utility by shifting from conversational assistants to autonomous agents that execute t

agentsarxiv-cs-ai
8 Jun 2026
Local Ai

Off-Policy Evaluation with Strategic Agents via Local Disclosure

DGX agent

arXiv:2606.07308v1 Announce Type: new Abstract: We study off-policy evaluation (OPE) under strategic behavior where decision subjects (or agents) respond to a decision maker's policy by strategically

local-aiarxiv-cs-ai
8 Jun 2026
Agents

OpenSkill: Open-World Self-Evolution for LLM Agents

DGX agent

arXiv:2606.06741v1 Announce Type: new Abstract: Self-evolving agents requires adaptation after deployment, but existing approaches assume a usable learning loop, such as curated skills, successful tra

agentsarxiv-cs-ai
8 Jun 2026
Agents

Pega expands AI platform with agent orchestration, development tools and new pricing model

DGX agent

Workflow automation vendor Pegasystems Inc. today unveiled a broad set of artificial intelligence enhancements aimed at helping enterprises deploy AI agents in mission-critical business processes whil

agentssiliconangle
8 Jun 2026
Agents

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built…

DGX agent

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built out courts. Stadium seating. Exhibition matches by AI leader

agentsjerry-liu--x
8 Jun 2026
Agents

The Open Source Community is backing OpenEnv for Agentic RL

DGX agent

OpenEnv is an open-source framework backed by the community for training and developing agentic reinforcement learning systems. The project represents collaborative efforts within the open-source ecos

agentshugging-face
8 Jun 2026
Model Releases

datasette-agent-edit 0.1a0

DGX agent

Release: datasette-agent-edit 0.1a0 I'm planning several plugins for Datasette Agent which can make edits to existing pieces of text - things like collaborative Markdown editing, updating large SQL qu

model-releasessimon-willison
7 Jun 2026
Industry

good dev tools are cached intelligence for agents!

DGX agent

good dev tools are cached intelligence for agents! Token costs are why there will be no saas apocalypse / good dev tools are cached intelligence for agents! The popular theory goes: agents can write c

industryclem-delangue--x
6 Jun 2026
Agents

Unsupervised Skill Discovery for Agentic Data Analysis

DGX agent

arXiv:2606.06416v1 Announce Type: cross Abstract: Inference-time skill augmentation provides a lightweight way to improve data-analytic agents by injecting reusable procedural knowledge without updati

agentsarxiv-cs-cl
5 Jun 2026
Hardware

AgentJet: A Flexible Swarm Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2606.04484v1 Announce Type: new Abstract: We present AgentJet, a distributed swarm training framework for large language model (LLM) agent reinforcement learning. Unlike centralized frameworks t

hardwarearxiv-cs-ai
4 Jun 2026
Agents

From Prompt to Process: a Process Taxonomy and Comparative Assessment of Frameworks Supporting AI Software Development Agents

DGX agent

arXiv:2606.04967v1 Announce Type: cross Abstract: AI tools for programming are no longer just autocomplete or chat assistants: they organize themselves as development frameworks, with process, roles,

agentsarxiv-cs-ai
4 Jun 2026
Agents

Radiant Logic extends identity visibility platform to enterprise AI agents with real-time risk scoring

DGX agent

Radiant Logic Inc., a platform that provides identity visibility and intelligence, today announced it’s extending its services to agentic artificial intelligence to help companies control and govern t

agentssiliconangle
4 Jun 2026
Safety

RUBAS: Rubric-Based Reinforcement Learning for Agent Safety

DGX agent

arXiv:2606.04051v1 Announce Type: cross Abstract: The evolution of LLMs into tool-enabled agents creates a new class of safety challenges associated with real-world execution rather than simple text g

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Streaming Communication in Multi-Agent Reasoning

DGX agent

arXiv:2606.05158v1 Announce Type: cross Abstract: Multi-agent reasoning systems adopt a 'generate-then-transfer' paradigm that forces end-to-end latency to scale linearly with pipeline depth. We intro

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

We have a Fleet agent called @docs_plz in our Slack that's made a very noticeable impact on our velocity of docs changes. In the chart below…

DGX agent

We have a Fleet agent called @docs_plz in our Slack that's made a very noticeable impact on our velocity of docs changes. In the chart below, you can clearly see that after it was added, the amount of

agentsharrison-chase--x
4 Jun 2026
Model Releases

Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward

DGX agent

arXiv:2602.12430v4 Announce Type: replace-cross Abstract: The transition from monolithic language models to modular, skill-equipped agents marks a defining shift in how large language models (LLMs) ar

model-releasesarxiv-cs-ai
3 Jun 2026
← Previous
1…7273747576…374
Next →