AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,214 results
18 May 2026

Introducing Devin Auto-Triage: Your AI first-responder with long-term memory. Devin can monitor incoming bugs, alerts, and incidents, invest…

AgentsDGX agent

Introducing Devin Auto-Triage: Your AI first-responder with long-term memory. Devin can monitor incoming bugs, alerts, and incidents, investigate them, and come back with context, next steps, or a PR.

Introducing http://Browse.sh, the largest open-source catalog of skills to reliably perform any task on the internet. We've researched hundr…

AgentsDGX agent

Introducing http://Browse.sh, the largest open-source catalog of skills to reliably perform any task on the internet. We've researched hundreds of sites to give your agents the playbook they need to n

Lamarckian Inheritance in Dynamic Environments: How Key Variables Affect Evolutionary Dynamics

Agents

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.15769v1 Announce Type: cross Abstract: The co-optimization of a robot's body and brain presents a coupled challenge: the morphology constrains which control strategies are effective, while

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say…

AgentsDGX agent

LangSmith Engine automates the full agent fix loop — detecting failures, diagnosing causes and drafting PRs. But multi-model enterprises say a neutral observability layer still wins. http://venturebea

last week, a friend was telling me how running down a steep trail is great cognitive exercise due to the vast amount of quick processing you…

AgentsDGX agent

last week, a friend was telling me how running down a steep trail is great cognitive exercise due to the vast amount of quick processing you do from vision and tactile input to micro adjustments requi

LEAP: Trajectory-Level Evaluation of LLMs in Iterative Scientific Design

AgentsDGX agent

arXiv:2605.15341v1 Announce Type: cross Abstract: LLMs are increasingly deployed in autonomous laboratories, under the assumption that their domain priors and reasoning over iterative feedback let the

LLM apps fail in ways normal logs can’t explain. Same input, different outputs. Subtle drift. Silent hallucinations. That’s the problem spac…

AgentsDGX agent

LLM apps fail in ways normal logs can’t explain. Same input, different outputs. Subtle drift. Silent hallucinations. That’s the problem space LangSmith by @LangChain was built for: observability + eva

Look Before You Leap: Autonomous Exploration for LLM Agents

AgentsDGX agent

arXiv:2605.16143v1 Announce Type: new Abstract: Large language model based agents often fail in unfamiliar environments due to premature exploitation: a tendency to act on prior knowledge before acqui

lots of good things in 0.6 release of deepagents! great write up by sydney

AgentsDGX agent

The post discusses positive features and improvements included in version 0.6 of DeepAgents, with Harrison Chase praising a write-up by Sydney that explains these updates. This likely covers new capab

@masondrxy and @Vtrivedy10 are doing an awesome push on harness profiles @huntlovell doing cool stuff w code interpreter / REPL @bromann lea…

AgentsDGX agent

@masondrxy and @Vtrivedy10 are doing an awesome push on harness profiles @huntlovell doing cool stuff w code interpreter / REPL @bromann leading the charge on streaming our OSS team is stacked! lots o

maybe you can’t just tack statefulness onto an agent, you have to figure out how to represent the agent as a state

AgentsDGX agent

This post explores the architectural challenge of implementing state management in AI agents, arguing that statefulness cannot simply be added as an afterthought but requires fundamental redesign of h

Mecha-nudges for Machines

AgentsDGX agent

arXiv:2603.23433v2 Announce Type: replace Abstract: AI agents are becoming active decision-makers on the Internet. As they make decisions in the same environments as humans, the environments themselve

Multiplayer AI startup Dust raises $40M to help enterprises move beyond isolated AI assistants

AgentsDGX agent

Dust, an agentic artificial intelligence startup that’s trying to push enterprise workers away from isolated chatbots into a more collaborative, multiplayer ecosystem, said today it has raised 40 mill

Nebius and @LangChain have partnered to integrate Nebius Token Factory with LangChain's Deep Agents. The integration, combined with LangChai…

AgentsDGX agent

Nebius and @LangChain have partnered to integrate Nebius Token Factory with LangChain's Deep Agents. The integration, combined with LangChain's existing Tavily integration, gives teams building on Lan

NIMO Controller: a self-driving laboratory orchestrator based on the Model Context Protocol

AgentsDGX agent

arXiv:2605.15227v1 Announce Type: new Abstract: Self-driving laboratories (SDLs) have attracted increasing attention as a means of accelerating scientific discovery; however, developing SDL software r

now i wanna see @drjimfan train robots in a sim doing this https://x.com/i/status/1795746905336725673

AgentsDGX agent

now i wanna see @drjimfan train robots in a sim doing this https://x.com/i/status/1795746905336725673 VIDEO: Competitors chase wheels of cheese down a steep hill, few remaining upright, in the annual

On the Convergence Rates of Federated Q-Learning across Heterogeneous Environments

AgentsDGX agent

arXiv:2409.03897v3 Announce Type: replace Abstract: Large-scale multi-agent systems are often deployed across wide geographic areas, where agents interact with heterogeneous environments. There is an

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

AgentsDGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

paper.json: A Coordination Convention for LLM-Agent-Actionable Papers

AgentsDGX agent

arXiv:2605.16194v1 Announce Type: cross Abstract: LLM agents routinely serve as first (and sometimes only) readers of academic papers, skimming for sub-claims, extracting reproducibility steps, and ge

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

AgentsDGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

Prospective multi-pathogen disease forecasting using autonomous LLM-guided tree search

AgentsDGX agent

arXiv:2605.16238v1 Announce Type: new Abstract: Probabilistic forecasting of infectious diseases is crucial for public health but relies on labor-intensive manual model curation by expert modeling tea

Real question: what is the actual latest state-of-the-art for file search and retrieval? - Actual grep over filesystem - Virtualized grep / …

AgentsDGX agent

Real question: what is the actual latest state-of-the-art for file search and retrieval? - Actual grep over filesystem - Virtualized grep / BM25 over a db (what @mintlify did) - Vector search over a d

RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents

AgentsDGX agent

arXiv:2605.16045v1 Announce Type: cross Abstract: Memory systems often organize user-agent interactions as retrievable external memory and are crucial for long-running agents by overcoming the limited

Runtime-Structured Task Decomposition for Agentic Coding Systems

AgentsDGX agent

arXiv:2605.15425v1 Announce Type: cross Abstract: Agentic coding systems increasingly use large language models (LLMs) for software engineering tasks such as debugging, root cause analysis, and code r

Scalable neuromorphic computing from autonomous spiking dynamics in a clockless reconfigurable chip

AgentsDGX agent

arXiv:2605.16114v1 Announce Type: cross Abstract: We propose a scalable neuromorphic architecture based on spiking dynamics emerging from the autonomous time-continuous evolution of clockless (asynchr

Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach

AgentsDGX agent

arXiv:2501.19128v5 Announce Type: replace-cross Abstract: In many real-world scenarios, reward signal for agents are exceedingly sparse, making it challenging to learn an effective reward function for

Sigma Computing seals $80M funding round as it pivots toward ‘agentic analytics’

AgentsDGX agent

Cloud-native data analytics startup Sigma Computing Inc. has closed on an 80 million Series E funding round that doubles its valuation to 3 billion, almost a year to the day after its previous Series

Simulation is a core part of how we build and evaluate robots. I spoke about simulation, robotics, and the role of interactive worlds at AI …

AgentsDGX agent

Simulation is a core part of how we build and evaluate robots. I spoke about simulation, robotics, and the role of interactive worlds at AI Engineer Singapore this past weekend! Full talk : https://ww

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

AgentsDGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

some good discussions and experiments around stateful agents in the replies, but seems like we’re not quite there yet, as in we’re starting …

AgentsDGX agent

some good discussions and experiments around stateful agents in the replies, but seems like we’re not quite there yet, as in we’re starting to track memory and traces, but not quite agent capability a

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data

AgentsDGX agent

arXiv:2509.21465v3 Announce Type: replace Abstract: Tabular foundation models are becoming increasingly popular for low-resource tabular problems. These models make up for small training datasets by p

Teams like @modal are already using Devin Auto-Triage for incidents on their inference team. “Devin Automations feels like a step forward fr…

AgentsDGX agent

Teams like @modal are already using Devin Auto-Triage for incidents on their inference team. “Devin Automations feels like a step forward from other auto-triage tools we’ve tried. It monitors our chan

TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination

AgentsDGX agent

arXiv:2605.15207v1 Announce Type: new Abstract: Multi-agent LLM systems have shown promise for complex reasoning, yet recent evaluations reveal they often underperform single-model baselines. We ident

The Hermes Agent Kanban just got a big automation upgrade. Drop one prompt into the triage, and the orchestrator agent can take it from ther…

AgentsDGX agent

The Hermes Agent Kanban just got a big automation upgrade. Drop one prompt into the triage, and the orchestrator agent can take it from there - decomposing it into all the subtasks necessary and autom

The Open Agent Leaderboard

AgentsDGX agent

The Open Agent Leaderboard is a benchmarking system hosted on Hugging Face that evaluates and ranks AI agents based on their performance across various tasks and capabilities. It provides a standardiz

there was a crow being mean to the ducklings and the parents were stressing out and I pointed it out to the kids who shoo’d the crow away an…

AgentsDGX agent

A person observed a crow harassing ducklings while their parents were distressed, and alerted nearby children who successfully drove the crow away. The post describes a brief wildlife intervention mom

this would add some pretty granular feedback that doesn’t interrupt conversation flow, using a UI ppl are familiar with

AgentsDGX agent

This post likely discusses implementing non-intrusive feedback mechanisms in conversational AI systems that leverage familiar UI patterns to provide detailed user input without disrupting the natural

Toward Natural and Companionable Virtual Agents via Cross-Temporal Emotional Modeling

AgentsDGX agent

arXiv:2605.15812v1 Announce Type: cross Abstract: Recent advances in foundation models have enabled conversational agents that aim for sustained companionship rather than mere task completion. Yet mos

Training on Documents About Monitoring Leads to CoT Obfuscation

AgentsDGX agent

arXiv:2605.15257v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is one of the most promising tools we have for detecting model misbehavior, but its effectiveness depends on models fa

Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

AgentsDGX agent

arXiv:2510.10454v2 Announce Type: replace Abstract: Large language models (LLMs) offer a generalizable approach for modeling patient trajectories, but suffer from the long and noisy nature of electron

Try it out … Improvements are landing every few days!

AgentsDGX agent

Try it out … Improvements are landing every few days! Grok Build CLI Beta can now be installed directly from Grok Web with a single terminal command. The agentic coding and workflow tool is currently

very happy for you do we have to maintain our SDKs manually again

AgentsDGX agent

very happy for you do we have to maintain our SDKs manually again We're thrilled to announce that Stainless is joining @AnthropicAI! Stainless was founded to make software better for everyone, and we'

watching super mario galaxy over the weekend and seeing all the references, i kept realizing all the mario games my kids haven't played yet

AgentsDGX agent

A parent reflects on playing Super Mario Galaxy with their children and becoming aware of the numerous references to other Mario games in the series that their kids have yet to experience. The observa

We are currently hiring full stack engineers to work on managed services, nous portal, UX/UI for hermes agent and applications around it, an…

AgentsDGX agent

We are currently hiring full stack engineers to work on managed services, nous portal, UX/UI for hermes agent and applications around it, and solving technical challenges cross-domain. If you're inter

We have hit 1000 contributors on the repo, a nice milestone to start out the week. Thank you to all of the contributors who make the Hermes …

AgentsDGX agent

Nous Research announced that their Hermes repository has reached 1,000 contributors, marking a significant milestone for the project. The post expresses gratitude to all contributors who have particip

we’re evolving into better vibe coders

AgentsDGX agent

we’re evolving into better vibe coders Here's an example of ongoing human physiological change: some people have a third artery in their arm. Some don't. ~10% of people born in the 1880s had the third

when evaluating long running agents, all of your evals don't need to be end to end. i'm working on a proper blog about this, but in our eval…

AgentsDGX agent

when evaluating long running agents, all of your evals don't need to be end to end. i'm working on a proper blog about this, but in our evals for our agents that run for 30-60 minutes, we have two set

Where to Perch in a Tree: Vision-Guidance for Tree-Grasping Drones

AgentsDGX agent

arXiv:2605.15430v1 Announce Type: cross Abstract: This study demonstrates a method to locate an ideal perch location on a tree for vision-guided autonomous tree-perching drones. Various image processi

Who Owns This Agent? Tracing AI Agents Back to Their Owners

AgentsDGX agent

arXiv:2605.16035v1 Announce Type: cross Abstract: AI agents are increasingly deployed to act autonomously in the world, yet there is still no reliable way to trace a harmful agent back to the account

Woke up to see that Grok Build finished my feature build from last night. But what's the most interesting to me, is that it has a set of sug…

AgentsDGX agent

Woke up to see that Grok Build finished my feature build from last night. But what's the most interesting to me, is that it has a set of suggestions for what to do to really make sure everything is do

WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes

AgentsDGX agent

arXiv:2605.15843v1 Announce Type: new Abstract: Recent 3D world modeling systems based on generative scene synthesis, such as Marble, can create coherent and explorable 3D environments, yet their outp

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation

AgentsDGX agent

arXiv:2605.15964v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires agents to follow natural-language instructions through closed-loop perception and action in 3D enviro

X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention

AgentsDGX agent

arXiv:2605.15505v1 Announce Type: new Abstract: In enterprise operations, the context required for an AI agent task is scattered across systems of record, static information stores, and communication

17 May 2026

build v1 of agent ship it (dogfooding counts) ⭐️ collect tracing data ⭐️ ⭐️⭐️ point agentic compute at data ⭐️⭐️ understand failures at scal…

AgentsDGX agent

This post outlines a development roadmap for an AI agent system, prioritizing initial version building with internal testing (dogfooding), implementing tracing and observability for data collection, a

Eval engineering: The missing piece of agentic AI governance

AgentsDGX agent

As artificial intelligence agents become more powerful, agentic AI governance becomes increasingly important – and yet, today’s governance solutions struggle to keep AI agents from going off the rails

good day for super mario galaxy movie and din tai w the kiddos

AgentsDGX agent

Yohei Nakajima shared a personal post about having a good day watching the Super Mario Galaxy movie and dining at Din Tai Fung restaurant with children. The post appears to be a casual social media up

Hermes Agent v0.14.0 - “The Foundation Release” Changelog below

AgentsDGX agent

Hermes Agent v0.14.0, released by Nous Research as 'The Foundation Release,' represents a significant update to their Hermes agent framework. This release likely includes core infrastructure improveme

is a heart beat just a cron job?

AgentsDGX agent

This post explores conceptual parallels between biological heartbeats and cron jobs (scheduled automated tasks in computing), likely examining how both function as reliable, recurring processes that m

LangSmith Engine feels like the missing CI/CD loop for AI agents , automatically detecting failures, clustering issues, proposing fixes, and…

AgentsDGX agent

LangSmith Engine feels like the missing CI/CD loop for AI agents , automatically detecting failures, clustering issues, proposing fixes, and generating evals from production traces. Agent engineering

LangSmith Engine is how we’re spinning the always-on, self-improvement loop for every agent - Tracing is on for every single agent - Purpose…

AgentsDGX agent

LangSmith Engine is how we’re spinning the always-on, self-improvement loop for every agent - Tracing is on for every single agent - Purpose built infra with SmithDB to handle data at agent scale (mor

← Previous
1…7172737475…121
Next →