AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,964 results
19 May 2026

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

SafetyDGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map

AgentsDGX agent

arXiv:2605.16899v1 Announce Type: new Abstract: A fundamental challenge in embodied AI is verifying if agents build internal models of spatial structure or merely learn to mimic task-specific expert t

Qumus: Realization of An Embodied AI Quantum Material Experimentalist

AgentsDGX agent

arXiv:2605.18407v1 Announce Type: cross Abstract: While modern Large Language Models (LLMs) and agentic artificial intelligence (AI) have demonstrated transformative capabilities in digital domains, t

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
18 May 2026

DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

SafetyDGX agent

arXiv:2511.19399v3 Announce Type: replace-cross Abstract: Deep research agents perform multi-step research to produce long-form, well-attributed answers. However, most open deep research agents are tr

Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach

AgentsDGX agent

arXiv:2501.19128v5 Announce Type: replace-cross Abstract: In many real-world scenarios, reward signal for agents are exceedingly sparse, making it challenging to learn an effective reward function for

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation

AgentsDGX agent

arXiv:2605.15964v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires agents to follow natural-language instructions through closed-loop perception and action in 3D enviro

X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention

AgentsDGX agent

arXiv:2605.15505v1 Announce Type: new Abstract: In enterprise operations, the context required for an AI agent task is scattered across systems of record, static information stores, and communication

15 May 2026

Closing the Gap on the Sample Complexity of 1-Identification

AgentsDGX agent

arXiv:2601.15620v2 Announce Type: replace Abstract: The 1-identification problem is a fundamental pure-exploration problem in multi-armed bandits. An agent aims to determine whether there exists an ar

Heuristic Pathologies and Further Variance Reduction via Uncertainty Propagation in the AIVAT Family of Techniques

AgentsDGX agent

arXiv:2605.14261v1 Announce Type: new Abstract: How should an agent's performance in a multiagent environment be evaluated when there is a limited sample size or a high cost of running a trial? The AI

In-IDE Toolkit for Developers of AI-Based Features

AgentsDGX agent

arXiv:2605.14612v1 Announce Type: cross Abstract: AI-enabled features built on LLMs and agentic workflows are difficult to test, debug, and reproduce, especially for product-focused software engineers

14 May 2026

ERPPO: Entropy Regularization-based Proximal Policy Optimization

SafetyDGX agent

arXiv:2605.13131v1 Announce Type: new Abstract: Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent

LangSmith Engine, 6 months later

AgentsDGX agent

LangSmith Engine, 6 months later imo, reading and analyzing long agent traces is a huge bottleneck and obstacle for progress (especially as these systems grow in time horizon). Imagine reading a langs

Learning POMDP World Models from Observations with Language-Model Priors

AgentsDGX agent

arXiv:2605.13740v1 Announce Type: new Abstract: Whether navigating a building, operating a robot, or playing a game, an agent that acts effectively in an environment must first learn an internal model

this is an early beta. we still have much to improve, which is why its so exciting for people to start trying. the goal is for Grok Build to…

AgentsDGX agent

this is an early beta. we still have much to improve, which is why its so exciting for people to start trying. the goal is for Grok Build to be the best coding agent when it comes to solving real and

What Limits Vision-and-Language Navigation ?

AgentsDGX agent

arXiv:2605.13328v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) is a cornerstone of embodied intelligence. However, current agents often suffer from significant performance degr

13 May 2026

PRISM: : Planning and Reasoning with Intent in Simulated Embodied Environments

Model ReleasesDGX agent

arXiv:2605.11534v1 Announce Type: new Abstract: When an LLM-based embodied agent fails at a household task, the culprit could be misidentified objects, forgotten sub-goals, or poor action sequencing -

so cool to see another blazing fast database built on vortex!

AgentsDGX agent

so cool to see another blazing fast database built on vortex! Just announced at Interrupt! SmithDB. Agent traces have outgrown the databases built to hold them. That’s why we built SmithDB, a purpose-

12 May 2026

AIPO: : Learning to Reason from Active Interaction

SafetyDGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

HULK: Large-scale Hierarchical Coordination under Continual and Uncertain Temporal Tasks

AgentsDGX agent

arXiv:2605.08722v1 Announce Type: new Abstract: Multi-agent systems can be extremely efficient when working concurrently and collaboratively, e.g., for delivery, surveillance, search and rescue. Coord

MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation

AgentsDGX agent

arXiv:2502.13451v5 Announce Type: replace Abstract: Vision-and-language navigation (VLN) is a key task in Embodied AI, requiring agents to navigate diverse and unseen environments while following natu

Policy Gradient Methods for Non-Markovian Reinforcement Learning

SafetyDGX agent

arXiv:2605.10816v1 Announce Type: cross Abstract: We study policy gradient methods for reinforcement learning in non-Markovian decision processes (NMDPs), where observations and rewards depend on the

Preventing Prompt Injection with Type-Directed Privilege Separation

AgentsDGX agent

arXiv:2509.25926v2 Announce Type: replace-cross Abstract: Modern language models have enabled the development of agentic systems that achieve strong performance on reasoning-intensive tasks. Unfortuna

SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation

AgentsDGX agent

arXiv:2602.04712v2 Announce Type: replace-cross Abstract: We present a visual-context image-retrieval-augmented generation (ImageRAG)- assisted AI agent for automatic target recognition (ATR) of synth

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

AgentsDGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

11 May 2026

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

AgentsDGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation

Model ReleasesDGX agent

arXiv:2605.07247v1 Announce Type: new Abstract: Scalable AI agents training relies on interactive environments that faithfully simulate the consequences of agent actions. Manually crafted environments

From Assistance to Agency: Rethinking Autonomy and Control in CI/CD Pipelines

SafetyDGX agent

arXiv:2605.07062v1 Announce Type: cross Abstract: AI agents are assuming active roles in Continuous Integration and Continuous Deployment (CI/CD) workflows, yet the research community lacks a shared v

GameGen-Verifier: Parallel Keypoint-Based Verification for LLM-Generated Games via Runtime State Injection

AgentsDGX agent

arXiv:2605.07442v1 Announce Type: new Abstract: LLM-based game generation promises to turn natural-language specifications into executable games, but progress is limited by the lack of reliable automa

Learning on the Shop floor

AgentsDGX agent

Learning on the Shop floor Tobias Lütke describes Shopify's internal coding agent tool, River, which operates entirely in public on their Slack: River does not respond to direct messages. She politely

@trycua https://hermes-agent.nousresearch.com/docs/user-guide/features/computer-use

AgentsDGX agent

This documentation page covers the computer-use feature in Hermes Agent, detailing how the AI system can interact with computer interfaces and perform automated tasks through screen interaction and in

9 May 2026

state management, observability, retries, permissioning, recovery paths, eval drift, human escalation. the model is only one component now.

AgentsDGX agent

This post from Harrison Chase discusses how AI agents have evolved beyond just the underlying model, emphasizing critical infrastructure components including state management, observability, retry mec

We just hit number one globally across all AI apps on OpenRouter. Super grateful to the nearly 1000 contributors who've helped make Hermes A…

AgentsDGX agent

We just hit number one globally across all AI apps on OpenRouter. Super grateful to the nearly 1000 contributors who've helped make Hermes Agent great, thank you! What do you want to see next? Hermes

7 May 2026

one of the features i'm most excited about in our upcoming langgraph release is delta channels! the langgraph runtime lets you 'checkpoint' …

AgentsDGX agent

one of the features i'm most excited about in our upcoming langgraph release is delta channels! the langgraph runtime lets you 'checkpoint' agent progress at every step (model call, tool call, hooks).

Personal Computer is built for the Mac ecosystem, so you can manage it from anywhere. Start tasks from your iPhone with local files on your …

AgentsDGX agent

Personal Computer is built for the Mac ecosystem, so you can manage it from anywhere. Start tasks from your iPhone with local files on your Mac. Add a Mac mini for full continuity and always-on agents

6 May 2026

CyberAId: AI-Driven Cybersecurity for Financial Service Providers

AgentsDGX agent

arXiv:2605.01892v1 Announce Type: new Abstract: European financial institutions face mounting regulatory pressure while their security operations centres remain constrained not by data or staffing but

“Every enterprise needs a claw strategy.” How did @LangChain go from a weekend project to 1B+ downloads in 3 years? We sat down with CEO and…

AgentsDGX agent

“Every enterprise needs a claw strategy.” How did @LangChain go from a weekend project to 1B+ downloads in 3 years? We sat down with CEO and co-founder Harrison Chase (@hwchase17) to talk deep agents,

NEW paper from Microsoft Research. (bookmark it) The entire interpretability literature is built around human readers. As more analysis gets…

Model ReleasesDGX agent

NEW paper from Microsoft Research. (bookmark it) The entire interpretability literature is built around human readers. As more analysis gets delegated to agents, the right target of interpretability s

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

SafetyDGX agent

arXiv:2605.02900v1 Announce Type: cross Abstract: Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, saf

Taming the Curses of Multiagency in Robust Markov Games with Large State Space through Linear Function Approximation

AgentsDGX agent

arXiv:2605.03125v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) holds great potential but faces robustness challenges due to environmental uncertainty. To address this, distr

5 May 2026

$ hermes skills install hyperframes http://github.com/NousResearch/hermes-agent/blob/main/optional-skills/creative/hyperframes/SKILL.md

AgentsDGX agent

The Hermes agent framework includes an optional 'hyperframes' skill within its creative module that enables enhanced multi-dimensional reasoning and response generation capabilities. This skill instal

I’m excited to announce that @llama_index is on the @CBInsights AI 100 list for 2026 🔥 We’re on a mission to parse all of the world’s PDFs,…

AgentsDGX agent

I’m excited to announce that @llama_index is on the @CBInsights AI 100 list for 2026 🔥 We’re on a mission to parse all of the world’s PDFs, and make them accessible to both humans and AI agents. List:

4 May 2026

Great conversation on harnesses from @hwchase17 at #GoogleCloudNext

AgentsDGX agent

Great conversation on harnesses from @hwchase17 at #GoogleCloudNext Is agent harness engineering the new prompt engineering? @LangChain co-founder and CEO, @hwchase17, joined us at #GoogleCloudNext to

had a couple ppl ask what made me want to invest (i am lucky they welcomed me), here’s what i answered

AgentsDGX agent

had a couple ppl ask what made me want to invest (i am lucky they welcomed me), here’s what i answered Announcing Cofounder 2: Run an entire company with agents. It's the infrastructure for the one pe

The latest AI news we announced in April 2026

Model ReleasesDGX agent

Google announced major AI infrastructure and platform advances at Google Cloud Next '26 in April 2026, including the Gemini Enterprise Agent Platform with new Agent Designer capabilities, and 8th Gene

this is a great GTM strategy for dev tools

AgentsDGX agent

this is a great GTM strategy for dev tools It's easy to talk agents... it's harder to know what the heck to build. Introducing 'Build with AgentMail' - a series of curated templates for the most power

This is my favorite launch video i've seen and luckily the product matches the craft (and is getting better at a very very fast rate).

AgentsDGX agent

This is my favorite launch video i've seen and luckily the product matches the craft (and is getting better at a very very fast rate). Announcing Cofounder 2: Run an entire company with agents. It's t

1 May 2026

End-to-end autonomous scientific discovery on a real optical platform

AgentsDGX agent

arXiv:2604.27092v1 Announce Type: new Abstract: Scientific research has long been human-led, driving new knowledge and transformative technologies through the continual revision of questions, methods

30 Apr 2026

EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses

AgentsDGX agent

arXiv:2604.26417v1 Announce Type: new Abstract: Emotion perception and adaptive expression are fundamental capabilities in human-agent interaction. While recent advances in speech emotion captioning (

Hackathon info https://x.com/NousResearch/status/2045225469088326039

AgentsDGX agent

Hackathon info https://x.com/NousResearch/status/2045225469088326039 The Hermes Agent Creative Hackathon starts now 16 Days, $25k in Prizes Presented by @Kimi_Moonshot & @NousResearch For the tinkerer

the folks @MadrigalPharma are some of my favorites to work with, they're always at the cutting edge! check out this case study on how they'r…

AgentsDGX agent

the folks @MadrigalPharma are some of my favorites to work with, they're always at the cutting edge! check out this case study on how they're building a multi-agent deep research system, powered by de

We shipped a LlamaParse MCP server to let you parse, classify, split, and generally operate over your hardest documents with your favorite A…

AgentsDGX agent

We shipped a LlamaParse MCP server to let you parse, classify, split, and generally operate over your hardest documents with your favorite AI agent 📄🤖 Check out the MCP server: http://mcp.llamaindex.a

29 Apr 2026

Appian and PwC make the case that governed AI is the only kind that scales

AgentsDGX agent

Agentic transformation is reshaping how enterprises operate, moving beyond basic automation to AI systems that can manage entire business processes on their own. But unlocking the full potential of th

ComfyUI is the most flexible open-source creation tool out there but the hard part has been finding and assembling the right pipeline for th…

AgentsDGX agent

ComfyUI is the most flexible open-source creation tool out there but the hard part has been finding and assembling the right pipeline for the job. We are excited about the new ComfyUI skill in @NousRe

SODA-CitrON: Static Object Data Association by Clustering Multi-Modal Sensor Detections Online

AgentsDGX agent

arXiv:2602.22243v2 Announce Type: replace Abstract: The online fusion and tracking of static objects from heterogeneous sensor detections is a fundamental problem in robotics, autonomous systems, and

28 Apr 2026

Accelerating Reinforcement Learning for Wind Farm Control via Expert Demonstrations

AgentsDGX agent

arXiv:2604.22794v1 Announce Type: cross Abstract: Reinforcement learning (RL) offers a promising approach for adaptive wind farm flow control, yet its practical deployment is hindered by slow training

Benchmarking Emergent Coordination in Large-Scale LLM Populations: An Evaluation Framework on the MoltBook Archive

Model ReleasesDGX agent

arXiv:2603.03555v2 Announce Type: replace-cross Abstract: As multi-agent Large Language Model (LLM) systems scale, evaluating their emergent coordination dynamics becomes increasingly critical. Howeve

Reconstructive Authority Model: Runtime Execution Validity Under Partial Observability

AgentsDGX agent

arXiv:2604.22898v1 Announce Type: cross Abstract: Autonomous systems increasingly operate under partial observability where execution-relevant state is never fully accessible. Existing governance mech

Think Anywhere in Code Generation

AgentsDGX agent

arXiv:2603.29957v3 Announce Type: replace-cross Abstract: Recent advances in reasoning Large Language Models (LLMs) have primarily relied on upfront thinking, where reasoning occurs before final answe

ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems

SafetyDGX agent

arXiv:2604.23878v1 Announce Type: new Abstract: Despite a century of empirical memory research, existing AI agent memory systems rely on system-engineering metaphors (virtual-memory paging, flat LLM s

27 Apr 2026

Devin for Terminal is available for all Windsurf users, and can even be used directly from Windsurf Next.

AgentsDGX agent

Devin for Terminal is available for all Windsurf users, and can even be used directly from Windsurf Next. The terminal hasn’t changed much since the 1970s. What you do with it has. Introducing Devin f

← Previous
1…162163164165166…300
Next →