AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,771 results
1 Jun 2026

ForecastCompass: Guiding Agentic Forecasting with Adaptive Factor Memory

Model ReleasesDGX agent

arXiv:2605.30858v1 Announce Type: new Abstract: Agentic forecasting is important for decision-making in dynamic environments, but it remains challenging because agents must reason from incomplete, tim

Generalized Intention Modeling in Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.31318v1 Announce Type: new Abstract: Modeling an opponent's intent is critical for effective decision-making in non-cooperative, competitive, and general-sum multi-agent reinforcement learn

Hermes Agent is now natively supported on @Windows

AgentsDGX agent

Nous Research announced native support for Hermes Agent on Windows, expanding the availability of their Hermes model to Windows-based systems. This development enables Windows users to run Hermes Agen

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How Hermes implements an open source agent harness architecture

AgentsDGX agent

Hermes from NousResearch is a strong open-source agent harness. This post examines how its runtime loop, context management, tool scoping, session infrastructure, and orchestration patterns map to a m

HypoAgent: An Agentic Framework for Interactive Abductive Hypothesis Generation over Knowledge Graphs

AgentsDGX agent

arXiv:2605.31370v1 Announce Type: new Abstract: Abductive reasoning over knowledge graphs aims to generate logical hypotheses that explain observed entities or facts. Existing controllable hypothesis

Let engine help you build better agents

AgentsDGX agent

This post likely discusses how the LangChain framework (which Harrison Chase co-founded) can assist developers in constructing more effective AI agents by providing tools, abstractions, and patterns f

MosaicLeaks:Privacy Risks in Querying-in-the-Open for Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.30727v1 Announce Type: new Abstract: Deep research agents increasingly combine private local documents with external tools like web retrieval, creating a privacy risk: an agent's external q

Multi-Agent Teams Hold Experts Back

SafetyDGX agent

arXiv:2602.01011v4 Announce Type: replace-cross Abstract: Multi-agent LLM systems are increasingly deployed as autonomous collaborators, where agents interact freely rather than execute fixed, pre-spe

SERA: Soft-Verified Efficient Repository Agents

Model ReleasesDGX agent

arXiv:2601.20789v3 Announce Type: replace Abstract: Open-weight coding agents should hold a fundamental advantage over closed-source systems because they can specialize to private codebases, encoding

Sophrosyne: Agentic Exploration of Relational Data Systems Needs Moderation

AgentsDGX agent

arXiv:2605.30862v1 Announce Type: cross Abstract: Text2SQL agents powered by LLMs translate natural language intent into SQL by exploring the data system through tool calls before formulating the quer

Task-Focused Memorization for Multimodal Agents

SafetyDGX agent

arXiv:2605.31075v1 Announce Type: new Abstract: Long-term memory is essential for multimodal agents to build coherent experience, accumulate world knowledge, and achieve continual learning. However, c

Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information

SafetyDGX agent

arXiv:2605.31445v1 Announce Type: cross Abstract: In this work we study agents in simulated bargaining scenarios, where a buyer and a seller communicate through a text channel and attempt to negotiate

30 May 2026

In a few months, people will start to realize how fundamentally important MCP for agents is. It's not even about connecting tools. There are…

AgentsDGX agent

In a few months, people will start to realize how fundamentally important MCP for agents is. It's not even about connecting tools. There are many ways to do that. It's about the types of abstraction i

Step 3.7 Flash, free for 30 days for Hermes Agent users. What could possibly go wrong? 🍿 Thanks @NousResearch for making it happen. Can’t w…

AgentsDGX agent

Step 3.7 Flash, free for 30 days for Hermes Agent users. What could possibly go wrong? 🍿 Thanks @NousResearch for making it happen. Can’t wait to see what Hermes users build! Step 3.7 Flash is now fre

29 May 2026

Harmless Yet Harmful: Neutral Prompting Attacks for Stealthy Hallucination Steering in Agent Skills

AgentsDGX agent

arXiv:2605.29354v1 Announce Type: cross Abstract: LLM-powered coding agents increasingly participate in software development workflows by generating code, selecting dependencies, and producing package

PatchBoard: Schema-Grounded State Mutation for Reliable and Auditable LLM Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.29313v1 Announce Type: new Abstract: LLM multi-agent systems often coordinate through natural-language dialogue or loosely structured shared memory, making intermediate state difficult to v

Revisiting Observation Reduction for Web Agents: Comprehensive Evaluation with a Lightweight Framework

AgentsDGX agent

arXiv:2605.29397v1 Announce Type: new Abstract: HTML observations in LLM-based web agents are extremely long, and while many reduction methods have been proposed, it remains unclear which methods redu

RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large Language Models

AgentsDGX agent

arXiv:2603.18859v2 Announce Type: replace Abstract: Reinforcement learning (RL) shows promise for enhancing LLM agentic reasoning, yet sparse terminal rewards hinder fine-grained optimization. Process

SkillBrew: Multi-Objective Curation of Skill Banks for LLM Agents

AgentsDGX agent

arXiv:2605.29440v1 Announce Type: cross Abstract: Retrieval-augmented LLM agents increasingly rely on curated skill banks: collections of reusable textual principles that guide decision making on comp

The team at @llama_index built an awesome template using LlamaParse and the new Managed Agents in the Gemini API. See how they built an agen…

Model ReleasesDGX agent

The team at @llama_index built an awesome template using LlamaParse and the new Managed Agents in the Gemini API. See how they built an agent that can tackle unstructured documents. 📄↓ 🚀 The team at @

28 May 2026

A Query Engine for the Agents

Model ReleasesDGX agent

arXiv:2605.27785v1 Announce Type: new Abstract: The fastest-growing data in production today is unstructured text: agent traces, chat logs, reasoning chains, model outputs. People want to analyze it,

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

AgentsDGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

CoreWeave introduces autonomous improvement capabilities for AI agents

AgentsDGX agent

Artificial intelligence cloud operator CoreWeave Inc. today announced the launch of a new offering that gives enterprise outfits the ability to deploy AI agents that can learn and improve themselves a

GUI Agents for Continual Game Generation

AgentsDGX agent

arXiv:2605.28258v1 Announce Type: cross Abstract: Generating a game is not the same as making one that can be played. Despite advances in code generation, existing approaches treat game generation as

OR-Space: A Full-Lifecycle Workspace Benchmark for Industrial Optimization Agents

Model ReleasesDGX agent

arXiv:2605.28158v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used to assist with operations research (OR) modeling, yet existing OR-oriented benchmarks often redu

PEAM: Parametric Embodied Agent Memory through Contrastive Internalization of Experience in Minecraft

Model ReleasesDGX agent

arXiv:2605.27762v1 Announce Type: new Abstract: We present PEAM, a Parametric Embodied Agent Memory framework in Minecraft that transforms agent memory from inference-time retrieval into parameter-res

Plant, Persist, Trigger: Sleeper Attack on Large Language Model Agents

Model ReleasesDGX agent

arXiv:2605.28201v1 Announce Type: new Abstract: Large Language Model (LLM) agents remain vulnerable to safety threats from the external environment, where attackers inject adversarial content into ext

27 May 2026

AI Agent for Reverse-Engineering Legacy Finite-Difference Code and Translating to Devito

AgentsDGX agent

arXiv:2601.18381v2 Announce Type: replace Abstract: To facilitate the transformation of legacy finite difference implementations into the Devito environment, this study develops an integrated AI agent

Harmonia: Enhancing Data Placement and Migration in Hybrid Storage Systems via Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2503.20507v4 Announce Type: replace-cross Abstract: Modern high-performance computing (HPC) environments rely on hybrid storage systems (HSS) that combine multiple storage devices with diverse l

Hermes Agent now has a built-in MCP Catalog

AgentsDGX agent

Nous Research announced that their Hermes Agent now includes an integrated MCP (Model Context Protocol) Catalog, enabling users to discover and access available Model Context Protocol tools and integr

LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning

AgentsDGX agent

arXiv:2412.20505v2 Announce Type: replace Abstract: Participatory Urban Planning (PUP) is increasingly supported by LLM-based agents, yet existing methods largely rely on static preference elicitation

Stateful Inference for Low-Latency Multi-Agent Tool Calling

AgentsDGX agent

arXiv:2605.26289v1 Announce Type: new Abstract: Multi-agent tool calling is becoming the dominant interaction pattern for LLM-based systems, yet existing inference frameworks treat each tool call as a

26 May 2026

.@AdamRLucek on how we use traces to build evals for production agents.

AgentsDGX agent

.@AdamRLucek on how we use traces to build evals for production agents. Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating e

Agent Learning via Early Experience

SafetyDGX agent

arXiv:2510.08558v3 Announce Type: replace Abstract: A long-term goal of language agents is to learn and improve through their own experience, ultimately outperforming humans in complex, real-world tas

APT-Agent: Automated Penetration Testing using Large Language Models

AgentsDGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

CODESKILL: Learning Self-Evolving Skills for Coding Agents

SafetyDGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

Exclusive: Canyon Code closes on $5M pre-seed round to bring visibility to agentic application workflows

AgentsDGX agent

Agentic artificial intelligence startup Canyon Code stepped out of the shadows today after closing on a 5 million pre-seed funding round. Its mission is to give enterprises the granular controls requi

From Model Scaling to System Scaling: Scaling the Harness in Agentic AI

Model ReleasesDGX agent

arXiv:2605.26112v1 Announce Type: new Abstract: This paper studies the next major bottleneck in agentic AI as system scaling, not only model scaling: the design of auditable, persistent, modular, and

Harnessing AtomisticSkills for Agentic Atomistic Research

AgentsDGX agent

arXiv:2605.24002v1 Announce Type: cross Abstract: Computational materials science and chemistry span vast knowledge domains and fractured software ecosystems. Although large language models (LLMs) hav

How Many Tools Should an LLM Agent See? A Chance-Corrected Answer

Model ReleasesDGX agent

arXiv:2605.24660v1 Announce Type: cross Abstract: Before an LLM agent can use a tool, a retrieval system must decide which candidate tools to show to the agent. How long should that shortlist be? Show

More Skills, Worse Agents? Skill Shadowing Degrades Performance When Expanding Skill Libraries

AgentsDGX agent

arXiv:2605.24050v1 Announce Type: cross Abstract: Skill libraries allow LLM agents to load task-specific instructions on demand, letting non-expert users solve domain-specific tasks through natural la

MultiPhishGuard: An Explainable and Adaptive Multi-Agent LLM System for Phishing Email Detection

SafetyDGX agent

arXiv:2505.23803v2 Announce Type: replace-cross Abstract: Phishing email detection faces significant challenges due to evolving adversarial tactics and heterogeneous attack patterns. Traditional appro

Retrieval as Reasoning: Self-Evolving Agent-Native Retrieval via LLM-Wiki

AgentsDGX agent

arXiv:2605.25480v1 Announce Type: new Abstract: LLM agents require retrieval to behave less like one-shot context fetching and more like reasoning: searching, reading, traversing, and deciding when ev

SEAL: Synergistic Co-Evolution of Agents and Learning Environments

SafetyDGX agent

arXiv:2605.24426v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly improved through interaction, yet most self-evolution methods adapt either the policy or the learning

Tool-Call Dependency Structure is Linearly Decodable in LLM Agent Residual Streams

AgentsDGX agent

arXiv:2605.25310v1 Announce Type: new Abstract: Tool-using LLM agents produce trajectories whose calls form a directed dependency graph: earlier tool outputs supply arguments to later calls. Whether t

25 May 2026

A measurement substrate for agentic Kubernetes operations: Methodology and a case study in retrieval-compounding falsification

Model ReleasesDGX agent

arXiv:2605.23058v1 Announce Type: cross Abstract: Empirical claims about autonomous Kubernetes operations agents are largely unfalsifiable. Published work reports observational results without control

From Raw Experience to Skill Consumption: A Systematic Study of Model-Generated Agent Skills

AgentsDGX agent

arXiv:2605.23899v1 Announce Type: new Abstract: Language agents increasingly improve by reusing skills -- structured procedural artifacts distilled from past experience. In particular, domain-level an

Join the team on Wednesday for another Hermes Agent Jam!

AgentsDGX agent

Nous Research is inviting developers and enthusiasts to participate in a 'Hermes Agent Jam' event scheduled for Wednesday, likely a hackathon or collaborative workshop focused on building or improving

SFG-ROS: A Resource-Aware Framework for Dense Multi-Agent Perception

AgentsDGX agent

arXiv:2605.23832v1 Announce Type: new Abstract: Deploying heterogeneous multi-agent robot fleets for collaborative perception requires robust data exchange and scalable software architectures. However

The Misattribution Gap: When Memory Poisoning Looks Like Model Failure in Agentic AI Systems

Model ReleasesDGX agent

arXiv:2605.22842v1 Announce Type: cross Abstract: Multi-agent AI pipelines typically assume that agent misconduct originates from model misalignment. We identify a structural failure in this assumptio

24 May 2026

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-1…

AgentsDGX agent

Grok Build sub-agent swarm weekend fun. You can reuse the prompt for your projects: Read the proof of `https://cdn.openai.com/pdf/74c24085-19b0-4534-9c90-465b8e29ad73/unit-distance-proof.pdf` and come

23 May 2026

Heterogeneous Agent Collaborative Reinforcement Learning

SafetyDGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

22 May 2026

AWS User Group Malaysia Talk: Experience on Building AI Agent Harness https://www.youtube.com/watch?v=u145ea-fNOw

AgentsDGX agent

This talk from an AWS User Group Malaysia event covers practical experience and insights on building an AI Agent Harness, likely discussing architecture patterns, implementation challenges, and best p

CTFExplorer: Evaluating LLM Offensive Agents Through Multi-Target Web CTF Benchmarking

Model ReleasesDGX agent

arXiv:2602.08023v3 Announce Type: replace-cross Abstract: Existing benchmarks for LLM-based offensive security agents use isolated, single-target setups with a known vulnerable service and fixed objec

FlyRoute: Self-Evolving Agent Profiling via Data Flywheel for Adaptive Task Routing

SafetyDGX agent

arXiv:2605.22057v1 Announce Type: new Abstract: Enterprise routers assign queries to expert agents, yet deployed profiles stay static while agents evolve (prompts, tools, models), and developers rarel

General Agentic Planning Through Simulative Reasoning with World Models

AgentsDGX agent

arXiv:2507.23773v3 Announce Type: replace-cross Abstract: What does it mean to plan? Current agentic systems, whether scaffolded workflows or end-to-end policies, rely on reactive decision-making: sel

Highlights: 👉 Long-horizon autonomy: maintained coherent execution across a 35-hour autonomous kernel optimization run 👉 Agentic coding: l…

AgentsDGX agent

Highlights: 👉 Long-horizon autonomy: maintained coherent execution across a 35-hour autonomous kernel optimization run 👉 Agentic coding: leading Terminal-Bench 2.0-Terminus performance for terminal-ba

Learning to Configure Agentic AI Systems

SafetyDGX agent

arXiv:2602.11574v3 Announce Type: replace Abstract: Configuring LLM-based agent systems involves choosing workflows, tools, token budgets, and prompts from a large combinatorial design space, and is t

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while pres…

AgentsDGX agent

NEW paper worth reading. A full agentic workflow can be distilled into model weights and run at roughly 100x lower inference cost while preserving near-frontier task quality. The workflow includes mul

Reflecti-Mate: A Conversational Agent for Adaptive Decision-Making Support Through System 1 and System 2 Thinking

AgentsDGX agent

arXiv:2605.22509v1 Announce Type: cross Abstract: Making high-stakes personal decisions involves cognitive, emotional, and intuitive processes, and individuals differ in how they allocate attention ac

← Previous
1…5859606162…297
Next →