AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,770 results
Applications

deployment cookbook for langchain agents!

DGX agent

deployment cookbook for langchain agents! Agents are easy to demo locally. The hard part is shipping them inside a real app. We published a deployment cookbook for @LangChain agents: full-stack exampl

applicationsharrison-chase--x
25 Jun 2026
Agents

Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2601.03785v3 Announce Type: replace Abstract: Long-term human-agent dialogues are organized by topic continuity: adjacent turns often develop the same goal, plan, problem, or event, while relate

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and Factorized Branch-and-Bound

DGX agent

arXiv:2606.25978v1 Announce Type: cross Abstract: Multi-agent goal recognition asks an observer to jointly infer which agents act together and what each team is trying to achieve, so the hypothesis sp

model-releasesarxiv-cs-lg
25 Jun 2026
Agents

Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows

DGX agent

arXiv:2604.25345v2 Announce Type: replace Abstract: Agentic AI systems are increasingly being integrated into scientific workflows, yet their behavior under realistic conditions remains insufficiently

agentsarxiv-cs-ai
25 Jun 2026
Safety

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

DGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

safetyarxiv-cs-cl
24 Jun 2026
Agents

Grading the Grader: Lessons from Evaluating an Agentic Data Analysis System

DGX agent

arXiv:2606.24839v1 Announce Type: new Abstract: Agentic data analysis systems produce rich outputs, including code, numerical results, and verbal diagnostics. This makes them more challenging to evalu

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

DGX agent

arXiv:2606.24551v1 Announce Type: new Abstract: Computer-use agents can execute software tasks through either graphical interfaces or programmatic command interfaces, but existing evaluations confound

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

MEMPROBE: Probing Long-Term Agent Memory via Hidden User-State Recovery

DGX agent

arXiv:2606.24595v1 Announce Type: new Abstract: Long-term memory promises LLM agents that grow more capable across sessions, maintaining an accurate, evolving understanding of the user that interactio

model-releasesarxiv-cs-cl
24 Jun 2026
Agents

'Most agents don't learn, they just leave traces.' In 12 minutes, @jakebroekhuizen breaks down how to actually close the loop. Surface issue…

DGX agent

'Most agents don't learn, they just leave traces.' In 12 minutes, @jakebroekhuizen breaks down how to actually close the loop. Surface issues with LangSmith Engine Write memory updates back to Context

agentsharrison-chase--x
24 Jun 2026
Safety

Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

DGX agent

arXiv:2509.18371v2 Announce Type: replace-cross Abstract: Multi-agent games in dynamic nonlinear settings are challenging due to the time-varying interactions among the agents and the non-stationarity

safetyarxiv-cs-ro
24 Jun 2026
Agents

pretty sick that i get to work with Jake every day on making continual learning + memory accessible at scale for every single agent one comm…

DGX agent

pretty sick that i get to work with Jake every day on making continual learning + memory accessible at scale for every single agent one common thread here is...the Trace a large part of Continual Lear

agentsharrison-chase--x
24 Jun 2026
Model Releases

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

DGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Skills for the future software profession: beyond agentic AI!

DGX agent

arXiv:2606.21894v2 Announce Type: replace-cross Abstract: As coding agents are rapidly changing software engineering, a natural question is: what are the core skills needed by future software engineer

agentsarxiv-cs-ai
24 Jun 2026
Agents

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

DGX agent

arXiv:2606.23743v1 Announce Type: cross Abstract: Modern video diffusion models achieve higher generation quality through scaling, but this also increases inference cost. Although many acceleration me

agentsarxiv-cs-ai
24 Jun 2026
Agents

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

DGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

agentsarxiv-cs-cl
24 Jun 2026
Agents

A trillion tokens a day and 200k GitHub stars Very proud of the Hermes Agent team and what we've built at @NousResearch 'Better today than y…

DGX agent

Nous Research announced that their Hermes Agent has achieved a trillion tokens processed daily and reached 200,000 GitHub stars, reflecting significant adoption and performance milestones for their AI

agentsnous-research--x
23 Jun 2026
Agents

GRADE: Graph Representation of LLM Agent Dependency and Execution

DGX agent

arXiv:2606.22741v1 Announce Type: new Abstract: Can one graph represent every kind of LLM agent's run? A trace records what each step did, never what it relied on, the state it read, and the results i

agentsarxiv-cs-lg
23 Jun 2026
Agents

Hermes Agent can now /learn from anything: feed it directories of any source material (code, API docs, manuals, PDFs, configs) and it distil…

DGX agent

Hermes Agent has been updated to accept diverse source materials including code repositories, API documentation, manuals, PDFs, and configuration files as input, with the ability to distill and learn

agentsnous-research--x
23 Jun 2026
Agents

PolicyGuard: Towards Test-time and Step-level Adversary (Backdoor) Defense for Reinforcement Learning Agent

DGX agent

arXiv:2606.12896v2 Announce Type: replace Abstract: While real-world applications of reinforcement learning (RL) are becoming increasingly popular, the security of RL systems deserve more attention an

agentsarxiv-cs-lg
23 Jun 2026
Agents

UltraQuant: 4-bit KV Caching for Context-Heavy Agents

DGX agent

arXiv:2606.20474v2 Announce Type: replace Abstract: Context-heavy agents place unusual pressure on the key-value (KV) cache: long prefixes are reused across many short turns, while concurrency determi

agentsarxiv-cs-lg
23 Jun 2026
Agents

Hermes Agent now supports computer use via @trycua on Windows and Linux in addition to existing macOS support

DGX agent

Hermes Agent, developed by Nous Research, has expanded its computer use capabilities to support Windows and Linux operating systems in addition to its existing macOS functionality. This enhancement, i

agentsnous-research--x
22 Jun 2026
Safety

APPO: Agentic Procedural Policy Optimization

DGX agent

arXiv:2606.12384v1 Announce Type: cross Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents

safetyarxiv-cs-ai
11 Jun 2026
Agents

ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories

DGX agent

arXiv:2606.11520v1 Announce Type: cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution

agentsarxiv-cs-ai
11 Jun 2026
Agents

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

DGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

agentsarxiv-cs-ai
11 Jun 2026
Agents

Catching One in Five: LLM-as-Judge Blind Spots in Production Multi-Turn Transaction Agents

DGX agent

arXiv:2606.10315v1 Announce Type: cross Abstract: LLM-as-judge is the default instrument for evaluating conversational agents, yet its reliability is almost always reported as agreement with human rat

agentsarxiv-cs-ai
10 Jun 2026
Agents

Exclusive: Relai raises $6.9M to enable verifiable and continuous learning for AI agents

DGX agent

Artificial intelligence infrastructure startup Relai Inc. said today it has closed on 6.9 million in funding as it bids to ensure the reliability of autonomous AI agents for enterprises. The company a

agentssiliconangle
10 Jun 2026
Agents

From Confident Closing to Silent Failure: Characterizing False Success in LLM Agents

DGX agent

arXiv:2606.09863v1 Announce Type: new Abstract: LLM agents can fail silently by asserting task completion when the environment state shows otherwise. We study this failure mode, false success, across

agentsarxiv-cs-lg
10 Jun 2026
Agents

Infini Memory: Maintainable Topic Documents for Long-Term LLM Agent Memory

DGX agent

arXiv:2606.10677v1 Announce Type: new Abstract: Long-term LLM agents need persistent memory that can track changing facts and provide relevant evidence across sessions. Existing memory systems often s

agentsarxiv-cs-ai
10 Jun 2026
Agents

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp …

DGX agent

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp & team • DSPy (arXiv:2310.03714) — @lateinteraction + @hazyr

agentsyohei-nakajima--x
10 Jun 2026
Agents

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not …

DGX agent

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not too far now! Excited to get this to work with my @dair_ai co

agentsdair-ai--x
10 Jun 2026
Model Releases

Workflow-GYM: Towards Long-Horizon Evaluation of Computer-use Agentic tasks in Real-World Professional Fields

DGX agent

arXiv:2606.11042v1 Announce Type: new Abstract: Recent years have witnessed the rapid evolution of AI agents toward handling increasingly complex, real-world tasks. However, existing benchmarks rarely

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo!

DGX agent

Also, I found that Hermes Agent + Nemotron 3 Ultra is a mighty combo! Excited to launch a new way to upskill with AI agents. This is how we are making it possible for anyone to learn to build with cod

model-releasesdair-ai--x
9 Jun 2026
Agents

ConMem: Structured Memory-Guided Adaptation in Training-Free Multi-Agent Systems

DGX agent

arXiv:2606.08702v1 Announce Type: new Abstract: Recent advances have improved the adaptive capabilities of LLM-based multi-agent systems (MAS) through memory-, skill-, and learning-based approaches, y

agentsarxiv-cs-ai
9 Jun 2026
Agents

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents

DGX agent

arXiv:2606.07904v1 Announce Type: new Abstract: Tool-augmented large language model agents increasingly rely on external APIs, but standard tool schemas describe how to call a tool, not when the tool

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Emergence World: A Platform for Evaluating Long-Horizon Multi-Agent Autonomy

DGX agent

arXiv:2606.08367v1 Announce Type: cross Abstract: Most evaluations of LLM agents look like exams: a discrete task, a clean environment, a score in minutes or hours. We argue that this approach is mism

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems

DGX agent

arXiv:2605.12213v2 Announce Type: replace Abstract: LLM-based conversational AI agents struggle to maintain coherent behavior over long horizons due to limited context. While RAG-based approaches are

agentsarxiv-cs-ai
9 Jun 2026
Agents

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and re…

DGX agent

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and remain frozen or mostly unchanged. The harness, like the skill

agentsdair-ai--x
9 Jun 2026
Agents

SIGA: Self-Evolving Coding-Agent Adapters for Scientific Simulation

DGX agent

arXiv:2606.09774v1 Announce Type: new Abstract: Advanced scientific simulators expose specialized input languages that turn simulation goals into executable configurations, but learning them can cost

agentsarxiv-cs-ai
9 Jun 2026
Agents

Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing

DGX agent

arXiv:2606.08256v1 Announce Type: new Abstract: Verifiability, attribution, and reproducibility are foundational requirements of scientific knowledge, yet current publishing infrastructure does not en

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation

DGX agent

arXiv:2606.08091v1 Announce Type: new Abstract: Recent agent frameworks such as Claude Code, Codex, and OpenClaw are strong at tool use and orchestration, but whether they can handle long video genera

model-releasesarxiv-cs-cv
9 Jun 2026
Agents

Autonomous computational catalysis through an agentic research system

DGX agent

arXiv:2601.13508v4 Announce Type: replace-cross Abstract: Autonomous agents are beginning to transform scientific research from tool-assisted workflows toward self-sustaining discovery processes. Comp

agentsarxiv-cs-ai
8 Jun 2026
Agents

Exploring Agentic Tool-Calling Decisions via Uncertainty-Aligned Reinforcement Learning

DGX agent

arXiv:2606.06976v1 Announce Type: new Abstract: Large language model (LLM)-based agents often make suboptimal tool-use decisions, including unsupported tool invocation and hallucinated direct response

agentsarxiv-cs-ai
8 Jun 2026
Agents

How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope

DGX agent

arXiv:2606.07489v1 Announce Type: new Abstract: Frontier AI systems are bridging the gap between intelligence and utility by shifting from conversational assistants to autonomous agents that execute t

agentsarxiv-cs-ai
8 Jun 2026
Local Ai

Off-Policy Evaluation with Strategic Agents via Local Disclosure

DGX agent

arXiv:2606.07308v1 Announce Type: new Abstract: We study off-policy evaluation (OPE) under strategic behavior where decision subjects (or agents) respond to a decision maker's policy by strategically

local-aiarxiv-cs-ai
8 Jun 2026
Agents

OpenSkill: Open-World Self-Evolution for LLM Agents

DGX agent

arXiv:2606.06741v1 Announce Type: new Abstract: Self-evolving agents requires adaptation after deployment, but existing approaches assume a usable learning loop, such as curated skills, successful tra

agentsarxiv-cs-ai
8 Jun 2026
Agents

Pega expands AI platform with agent orchestration, development tools and new pricing model

DGX agent

Workflow automation vendor Pegasystems Inc. today unveiled a broad set of artificial intelligence enhancements aimed at helping enterprises deploy AI agents in mission-critical business processes whil

agentssiliconangle
8 Jun 2026
Agents

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built…

DGX agent

The Agent Open: AI's Pickleball Tournament 🏓 Come put your code and backhand to the test and embrace the full Open experience. Custom built out courts. Stadium seating. Exhibition matches by AI leader

agentsjerry-liu--x
8 Jun 2026
Agents

The Open Source Community is backing OpenEnv for Agentic RL

DGX agent

OpenEnv is an open-source framework backed by the community for training and developing agentic reinforcement learning systems. The project represents collaborative efforts within the open-source ecos

agentshugging-face
8 Jun 2026
← Previous
1…7172737475…371
Next →