AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,195 results
30 Jun 2026

Had a lot of fun at @aiDotEngineer talking about BM25!

AgentsDGX agent

Had a lot of fun at @aiDotEngineer talking about BM25! Cool Search paradigm by @jobergum: “BM25 + Grep is all you need”: 1. Use BM25 to narrow down a large universe of documents to a few candidates 2.

harbor is a great framework for running evals for long running, stateful agents its becoming industry standard, powering benchmarks like ter…

AgentsDGX agent

harbor is a great framework for running evals for long running, stateful agents its becoming industry standard, powering benchmarks like terminal bench 2 we've integrated deeply with harbor: across La

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a…

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

.@harborframework can now integrate directly with Deep Agents, LangSmith Sandboxes, and LangSmith Observability. You need to run agents in a real, reproducible, isolated environment, many times in par

He3-Seeker: Robotic Information Planning for Lunar Helium-3 Distribution Mapping

AgentsDGX agent

arXiv:2606.28746v1 Announce Type: new Abstract: Lunar helium-3 is a highly valuable strategic resource, pivotal to the advancement of both deep-space exploration and space mining. Existing lunar heliu

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundan…

AgentsDGX agent

Hermes Agent now reads the web up to 60x faster and 49x cheaper. Scraping backends pass clean content straight to the agent without redundant processing steps; large pages are saved locally and paged

HMARS: A Hierarchical Multi-Agent Memory System for Long-Context Reasoning

AgentsDGX agent

arXiv:2606.28349v1 Announce Type: cross Abstract: Long-context reasoning requires models to access, retrieve, and integrate evidence scattered across documents, dialogues, and accumulated interaction

how do you run agent code without a full blown sandbox? we did a lot of work to harden the code interpreter runtime we use

AgentsDGX agent

Harrison Chase discusses techniques for executing agent code safely without requiring a complete sandbox environment, highlighting hardening measures implemented in their code interpreter runtime. The

human in the loop is very nice for these wiki memory systems

AgentsDGX agent

human in the loop is very nice for these wiki memory systems @hwchase17 My rule is that nothing gets added to the knowledge system without me reviewing it first, which has served me well so far, along

@hwchase17 very spot-on! for our (http://qontext.ai) customers, the journey is usually: local .md files -> shared github repo (-> gbrain) be…

AgentsDGX agent

@hwchase17 very spot-on! for our (http://qontext.ai) customers, the journey is usually: local .md files -> shared github repo (-> gbrain) before these approaches break down at a) multiplayer (incl acc

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents

AgentsDGX agent

arXiv:2606.29648v1 Announce Type: cross Abstract: Different retrievers, including lexical, semantic, and multimodal approaches, provide highly complementary strengths for multimodal document understan

HyphaeDB: A Living Knowledge Topology for Agent-First Memory

AgentsDGX agent

arXiv:2606.28781v1 Announce Type: new Abstract: Every existing vector database and agent memory framework treats memory as passive storage that agents query explicitly. No system propagates knowledge

I love CLI loading spinners so much. there's something so satisfying about this UI

AgentsDGX agent

This post expresses enthusiasm for command-line interface (CLI) loading spinners, highlighting the aesthetic and user experience satisfaction they provide during software operations. The author apprec

If you build with MCPs, this one is worth reading. (bookmark it) The paper covers five recurring MCP server patterns across fifteen independ…

AgentsDGX agent

If you build with MCPs, this one is worth reading. (bookmark it) The paper covers five recurring MCP server patterns across fifteen independently developed servers. That taxonomy is useful because I s

Improved Multi-Dimensional Forecasting for Swap Regret

AgentsDGX agent

arXiv:2606.29533v1 Announce Type: cross Abstract: We study the problem of forecasting for an arbitrary number of downstream agents with unknown objectives, each of whom best responds to the forecaster

Interpretable Clustering: A Survey

AgentsDGX agent

arXiv:2409.00743v4 Announce Type: replace-cross Abstract: In recent years, much of the research on clustering algorithms has primarily focused on enhancing their accuracy and efficiency, frequently at

Interpretable Inverse Design of Metal-Organic Frameworks with Large Language Model Agents

AgentsDGX agent

arXiv:2606.29459v1 Announce Type: cross Abstract: Inverse design of metal-organic frameworks (MOFs) requires searching a combinatorially vast space where property labels are expensive and most machine

Is Lying an Emergent Behaviour in LLMs? Evidence from Gaslighting AI agents in a Sustainability Game

AgentsDGX agent

arXiv:2606.28456v1 Announce Type: cross Abstract: LLMs agents are increasingly used in multi-agent settings, yet their behaviour in sustainability games remains largely unexplored. This work investiga

I've added video support to my 'shot-scraper' browser automation tool - you (or your coding agent) can now create a storyboard YAML file and…

AgentsDGX agent

I've added video support to my 'shot-scraper' browser automation tool - you (or your coding agent) can now create a storyboard YAML file and use that to record a video demo of new web application feat

joined @LangChain almost three months ago and have spent a ton of time thinking about evals! i think that harbor is the future of eval frame…

AgentsDGX agent

joined @LangChain almost three months ago and have spent a ton of time thinking about evals! i think that harbor is the future of eval frameworks so really enjoyed being able to bring it into tighter

KbSD: Knowledge Boundary aware Self-Distillation for Behavioral Calibration in Agentic Search

AgentsDGX agent

arXiv:2606.29863v1 Announce Type: new Abstract: Agentic search equips large language models with dynamic retrieval abilities, but existing reinforcement learning methods remain limited by reward spars

LAMP: Lean-based Agentic framework with MCP and Proof Repair

AgentsDGX agent

arXiv:2606.28841v1 Announce Type: cross Abstract: Large language models are increasingly capable of mathematical reasoning, but the proofs they generate are often unreliable and hard to verify. Intera

LangChain 今天连发两记,直接把 agent reliability 往前推了一步:一套统一评估栈,一个不用完整沙箱的代码执行方案。 先说评估。长运行、有状态的 agent 怎么测?Harbor + LangSmith 的组合给出一个 runner,能挂载 sandbox…

AgentsDGX agent

LangChain 今天连发两记,直接把 agent reliability 往前推了一步:一套统一评估栈,一个不用完整沙箱的代码执行方案。 先说评估。长运行、有状态的 agent 怎么测?Harbor + LangSmith 的组合给出一个 runner,能挂载 sandbox、接入 observability,终于不是靠单元测试糊弄了。 再看代码执行。Deep Agents 需要跑不受信代码,

Legible Shared Autonomy: Implicit Communication of Robot Belief through Motion

AgentsDGX agent

arXiv:2606.29846v1 Announce Type: new Abstract: Shared autonomy systems combine user input with autonomous assistance to help users with motor impairments control robot arms to perform everyday manipu

Linguistic Firewall: Geometry as Defense in Multi-Agent Systems Routing

AgentsDGX agent

arXiv:2606.30555v1 Announce Type: new Abstract: The rapid integration of Large Language Models (LLMs) has driven the evolution of Multi-Agent Systems (MAS), where specialized agents collaborate to exe

LLM agents security duality: a comprehensive survey of self-security and empowered cybersecurity

AgentsDGX agent

arXiv:2606.28450v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly being integrated into real-world systems. Their autonomy and tool-use capabilities generate substantial

Mandol: An Agglomerative Agent Memory System for Long-Term Conversations

AgentsDGX agent

arXiv:2606.29778v1 Announce Type: cross Abstract: Long-term conversational agents need to remember and query cross-session, multi-typed information with complex correlations. Existing agent memory sys

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

AgentsDGX agent

arXiv:2606.30296v1 Announce Type: new Abstract: Multi-round reflection lets agents built on large language models recover from failures within a single task, but each task remains an isolated episode:

Manufactured Confidence: How Memory Consolidation Turns Hearsay into Confident Facts

AgentsDGX agent

arXiv:2606.29279v1 Announce Type: cross Abstract: LLM agents carry conclusions across steps and sessions in compressed memory, and memory products (e.g., mem0, LangMem) rewrite conversation into store

MAVIN: Multi-Shot Audio-Visual Generation with Narrative Control

AgentsDGX agent

arXiv:2606.29473v1 Announce Type: new Abstract: While recent generative models produce high-fidelity videos, they struggle with the complex narrative control required for coherent multi-shot audio-vis

Memory as an Attack Surface in LLM Agents: A Study on Multiple-Choice Question Answering

AgentsDGX agent

arXiv:2606.29030v1 Announce Type: new Abstract: AI agents extend conventional large language model (LLM) applications by integrating language understanding with task execution, external tool use, and

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsol…

AgentsDGX agent

Memory has somehow consistently been the most exciting area of agent development over the last 3 years (imo), and it's still a largely unsolved problem!! Wiki's are the biggest advancement I've seen i

Modeling Earth-Scale Human-Like Societies with One Billion Agents

AgentsDGX agent

arXiv:2506.12078v2 Announce Type: replace-cross Abstract: Understanding the dynamic evolution of complex social phenomena requires both high-fidelity modeling of human behavior and large-scale simulat

MonoSR: Open-Vocabulary Spatial Reasoning from Monocular Images

AgentsDGX agent

arXiv:2511.19119v2 Announce Type: replace Abstract: Spatial reasoning (SR), the ability to infer 3D spatial information from 2D inputs, is essential for real-world applications such as embodied AI and

Monte Carlo Query Search: Active Capability Assessment of AI Agents

AgentsDGX agent

arXiv:2512.16733v3 Announce Type: replace Abstract: Black-box AI (BBAI) systems, including foundation-model agents, are increasingly used for sequential decision making. Safe deployment requires metho

More on Hermes Agent web search & extraction: https://hermes-agent.nousresearch.com/docs/user-guide/features/web-search

AgentsDGX agent

This documentation page covers advanced features of Hermes Agent's web search and information extraction capabilities, explaining how users can leverage the tool to search the internet and extract rel

More time to build with Step 3.7 Flash: in partnership with @StepFun_ai, we’re extending the free usage period in Nous Portal by an addition…

AgentsDGX agent

More time to build with Step 3.7 Flash: in partnership with @StepFun_ai, we’re extending the free usage period in Nous Portal by an additional 15 days! Media Step 3.7 Flash is now free for 30 days via

Multi-Agent DRL for QoS and Energy Optimization in RIS-Enabled Open-RAN Industrial 6G TN/NTN Networks

AgentsDGX agent

arXiv:2606.28339v1 Announce Type: cross Abstract: Industrial 6G networks require ultra-reliable, low-latency, and energy-efficient connectivity in dynamic and blockage-prone environments, where conven

NaLA: A 3D Native LLM Layout Agent for High-quality 3D Scene Generation

AgentsDGX agent

arXiv:2606.29395v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have emerged as promising layout agents for 3D scene generation. Existing layout agents still suffer from implaus

Neural Procedural Memory: Empowering LLM Agents with Implicit Activation Steering

AgentsDGX agent

arXiv:2606.29824v1 Announce Type: cross Abstract: While Large Language Models (LLMs) excel as static solvers, transforming them into autonomous agents remains challenging. This transition requires con

Neural Stereo Video Compression with Hybrid Disparity Compensation

AgentsDGX agent

arXiv:2504.20383v3 Announce Type: replace Abstract: Disparity compensation represents the primary strategy in stereo video compression (SVC) for exploiting cross-view redundancy. These mechanisms can

NeuralMUSIC: A Hybrid Neural-Subspace Framework for Robot Sound Source Localization

AgentsDGX agent

arXiv:2606.18664v2 Announce Type: replace-cross Abstract: Reliable sound source localization is fundamental to robot audition, enabling autonomous robots to perceive spatial cues and operate effective

Normalizing Flow-Enhanced Message Passing for Multirobot Collaborative Localization

AgentsDGX agent

arXiv:2606.29868v1 Announce Type: new Abstract: Accurate, robust, and adaptive localization is essential for various robotic operations. This paper proposes a new message passing (MP) algorithm for re

On the Necessity of a Liquid Substrate for Mesh Intelligence

AgentsDGX agent

arXiv:2606.28413v1 Announce Type: cross Abstract: A mesh of sovereign agents has no center: no shared clock, no shared model, and no coordinator to gather data or retrain. Its competence rests on each

PIXELRAG: Web Screenshots Beat Text for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2606.28344v1 Announce Type: cross Abstract: Augmenting large language models (LLMs) with retrieved web text has become a dominant paradigm, yet the web is not natively textual: existing systems

PLOT: Pseudo-Labeling via Object Tracking for Monocular 3D Object Detection

AgentsDGX agent

arXiv:2507.02393v2 Announce Type: replace Abstract: Monocular 3D object detection is crucial for scalable perception across fields like autonomous driving, robotics, and surveillance. However, progres

Preventing Error Propagation in Multi-Agent AI through Runtime Monitoring

AgentsDGX agent

arXiv:2606.29026v1 Announce Type: new Abstract: Multi-agent AI systems can improve answer selection by allowing different language models to exchange reasoning traces, revise initial predictions, and

Queue raises $12.6M to launch ‘fully robotic pharmacy’ kiosk to make picking up meds more convenient

AgentsDGX agent

Queue, a company building an autonomous “robotic” pharmacy kiosk that dispenses medication, today announced that it has raised 12.6 million in seed funding led by AlleyCorp. The company is launching t

Real-Time Underwater Image Enhancement via Frequency-Guided Dual-Path Attention

AgentsDGX agent

arXiv:2606.30314v1 Announce Type: new Abstract: Real-time underwater image enhancement (UIE) is crucial for mobile underwater photography and autonomous robotic systems, where practical deployment typ

ReasonRec: A Reasoning-Augmented Multimodal Agent for Unified Recommendation

AgentsDGX agent

arXiv:2606.28357v1 Announce Type: cross Abstract: Recent advances in multimodal recommenders excel at feature fusion but remain opaque and inefficient decision-makers, lacking explicit reasoning and s

Reinforcement Learning for Software Vulnerability Analysis: A Systematic Review with Emphasis on C/C++ Source Code and Static Analysis

AgentsDGX agent

arXiv:2606.28403v1 Announce Type: cross Abstract: Vulnerability detection in C/C++ software remains a major security challenge due to code complexity, manual memory management, and the limitations of

Reinforcement Learning in Super Mario Bros: Curriculum, Pedagogy, and Optimal Level Design in World 1-1

AgentsDGX agent

arXiv:2606.29511v1 Announce Type: new Abstract: World 1-1 of Super Mario Bros is widely celebrated as a masterclass in game design: its progressive structure is credited with teaching players core mec

RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources

AgentsDGX agent

arXiv:2606.29538v1 Announce Type: cross Abstract: Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill librari

Rethinking Role-Playing Evaluation: Anonymous Benchmarking and a Systematic Study of Personality Effects

AgentsDGX agent

arXiv:2603.03915v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have shown remarkable potential in developing role-playing agents (RPAs). However, current evaluation frameworks

Row-Bot's approach to memory is hybrid. We have taken the best of both worlds. A knowledge graph for the agent + bi-directionaly synced Wiki…

AgentsDGX agent

Row-Bot's approach to memory is hybrid. We have taken the best of both worlds. A knowledge graph for the agent + bi-directionaly synced Wiki for human readability/audit. With in-app graph visualizatio

SafeGEO: Understanding Generative Engine Optimization Risks in Recommendation Agents

AgentsDGX agent

arXiv:2606.28356v1 Announce Type: cross Abstract: Generative Engine Optimization (GEO) lets content owners rewrite web content to increase their visibility in generative systems. In recommendation age

SAGA: Scene-Aware, Goal-Evolving Agents for Long-Horizon CivRealm Strategy Planning

AgentsDGX agent

arXiv:2606.29932v1 Announce Type: new Abstract: Long-horizon strategic planning in complex strategy games demands concurrent reasoning across multiple decision domains under imperfect information and

Scene-aware Prediction of Diverse Human Movement Goals

AgentsDGX agent

arXiv:2606.29942v1 Announce Type: new Abstract: Anticipation of human behaviours facilitates autonomous systems in proactive planning. Human behaviour could be stochastic due to varying goals. Human g

SCREP: Scene Coordinate Regression and Evidential Learning-based Perception-Aware Trajectory Generation

AgentsDGX agent

arXiv:2507.07467v3 Announce Type: replace Abstract: Autonomous flight in GPS-denied indoor spaces requires trajectories that keep visual-localization error tightly bounded across varied missions. Map-

Self-Evolving Agentic Image Restoration via Deliberate Planning and Intuitive Execution

AgentsDGX agent

arXiv:2606.28971v1 Announce Type: new Abstract: Real-world image restoration (IR) remains challenging due to complex and coupled degradations. While recent agentic IR frameworks leverage Large Languag

Self-Evolving World Models for LLM Agent Planning

AgentsDGX agent

arXiv:2606.30639v1 Announce Type: new Abstract: World models offer a principled way to equip long-horizon LLM agents with foresight: predictions of action consequences before execution. However, unrel

← Previous
1…3031323334…120
Next →