AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,201 results
9 Jun 2026

DeltaBox: Scaling Stateful AI Agents with Millisecond-Level Sandbox Checkpoint/Rollback

AgentsDGX agent

arXiv:2605.22781v2 Announce Type: replace-cross Abstract: LLM-powered AI agents require high-frequency state exploration (e.g., test-time tree search and reinforcement learning), relying on rapid chec

Differentiable Weightless Controllers: Learning Logic Circuits for Continuous Control

AgentsDGX agent

arXiv:2512.01467v2 Announce Type: replace Abstract: Controlling autonomous systems under real-world conditions often requires policies that can be evaluated with low latency and minimal energy consump

Disturbance-Aware Aerial Robotics for Ethical Wildlife Monitoring

AgentsDGX agent

arXiv:2606.08249v1 Announce Type: cross Abstract: Reliable wildlife monitoring is essential for ecology and conservation, yet many existing methods, such as tagging, capture, and close-range observati


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

DTEX adds AI Risk Management to track how agents and employees use AI

AgentsDGX agent

Behavioral intelligence security company DTEX Systems Inc. today introduced an expanded AI Risk Management product that reads the intent behind how employees and autonomous artificial intelligence age

Earlytrade raises $10M to bring agentic AI to construction payments

AgentsDGX agent

Earlytrade Pty. Ltd., a company solving payments flow for contractors in the construction industry, today announced it has raised about 10 million in new funding. Today’s capital infusion brings the t

EditSSC: Toward Editable Semantic Occupancy Scenes with Unconditional Diffusion Models

AgentsDGX agent

arXiv:2606.09273v1 Announce Type: new Abstract: 3D semantic scene generation is crucial for autonomous driving applications, yet most methods rely on complex 3D-specific architectures such as triplane

Engagement Process: Rethinking the Temporal Interface of Action and Observation

AgentsDGX agent

arXiv:2605.11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over

Excited to launch a new way to upskill with AI agents. This is how we are making it possible for anyone to learn to build with coding agents…

AgentsDGX agent

Excited to launch a new way to upskill with AI agents. This is how we are making it possible for anyone to learn to build with coding agents. To start, we are launching 4 new hands-on labs on the foll

Eyes All Around: Design and Analysis of 360-Degree LiDAR Perception Using Equivariant Feature Learning in Unstructured Traffic

AgentsDGX agent

arXiv:2606.07626v1 Announce Type: cross Abstract: Perception in dense, unstructured urban traffic remains a major challenge for autonomous driving because of the wide variety of road users, frequent o

FASE: Fast Adaptive Semantic Entropy for Code Quality

AgentsDGX agent

arXiv:2606.09800v1 Announce Type: cross Abstract: Multi-agent code generation offers a promising paradigm for autonomous software development by simulating the human software engineering lifecycle. Ho

From 0-to-1 to 1-to-N: Reproducible Engineering Evidence for MetaAI Recursive Self-Design

AgentsDGX agent

arXiv:2606.09663v1 Announce Type: new Abstract: Recursive self-design refers to AI-assisted modification of the mechanisms by which an AI system is built, evaluated, and improved. This paper treats Me

From Conflict to Consensus: Boosting Medical Reasoning via Multi-Round Agentic RAG

AgentsDGX agent

arXiv:2603.03292v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) exhibit high reasoning capacity in medical question-answering, but their tendency to produce hallucinations and o

Geometry-Aware Fisheye-LiDAR Fusion for Robust 3D Object Detection in Low-Overlap Setups

AgentsDGX agent

arXiv:2606.08844v1 Announce Type: new Abstract: As autonomous systems expand from capital-intensive robotaxis to cost-sensitive logistics, sensor configurations are increasingly optimized for coverage

GIFT: LLM-Guided State-Reward Interface for Financial Reinforcement Learning

AgentsDGX agent

arXiv:2606.08450v1 Announce Type: new Abstract: Financial portfolio trading is naturally formulated as a reinforcement learning problem, where an agent sequentially rebalances assets under changing ma

Goal-Oriented Reasoning for RAG-based Memory in Conversational Agentic LLM Systems

AgentsDGX agent

arXiv:2605.12213v2 Announce Type: replace Abstract: LLM-based conversational AI agents struggle to maintain coherent behavior over long horizons due to limited context. While RAG-based approaches are

Governance Controls for AI-Generated Test Artifacts in Autonomous Software Testing

AgentsDGX agent

arXiv:2606.08806v1 Announce Type: cross Abstract: Artificial Intelligence (AI) and Large Language Models (LLMs) are increasingly used in autonomous software testing; however, AI-generated test artifac

GPT-Micro: A large language paradigm for accelerated, inexpensive, and thermodynamics-consistent discovery of constitutive models in manufacturing

AgentsDGX agent

arXiv:2606.08238v1 Announce Type: new Abstract: Constitutive modeling of the relationship between process-imposed material states and fundamental material properties is critical to control of material

GraphER: An Efficient Graph-Based Enrichment and Reranking Method for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2603.24925v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) systems that rely on semantic search often fail to retrieve the complete set of evidence for complex queries, p

Happy to present that terminal tool calls will be in prettified markdown codeblocks now if you have display tool calls enabled in Hermes Age…

AgentsDGX agent

Nous Research announced an enhancement to their Hermes AI model where terminal tool calls will now be displayed in formatted markdown codeblocks when the display tool calls feature is enabled. This im

How an Agent Built a 3D Paris Gallery by Chaining Two Hugging Face Spaces

AgentsDGX agent

This article describes how an AI agent was used to create a 3D virtual gallery of Paris by chaining together two Hugging Face Spaces applications. It demonstrates a practical example of using agents t

How to detect credential theft in AI agent harness traces

AgentsDGX agent

In May 2026, a malicious version of a popular VS Code extension spent 18 minutes in the marketplace before anyone caught it. In that time it ran on roughly 6,000... The post How to detect credential t

If people only knew how much OpenMed runs on HF Stack like Buckets, Datasets, and Spaces, From datasets, to agent traces, to medical intelli…

AgentsDGX agent

If people only knew how much OpenMed runs on HF Stack like Buckets, Datasets, and Spaces, From datasets, to agent traces, to medical intelligent MCPs, to binary builds for OpenMed Agent, @huggingface

In-Context Reinforcement Learning via Communicative World Models

AgentsDGX agent

arXiv:2508.06659v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) agents often struggle to generalize to new tasks and contexts without updating their parameters, mainly because th

Introducing Cohere's first open-source coding model: North Mini Code Small & efficient, designed for agentic performance and built for commu…

AgentsDGX agent

Cohere released North Mini Code, an open-source coding model designed to be small and efficient while optimizing for agentic performance and community use. The model represents Cohere's initial offeri

loops are great because its just ai running in the background 'triggers' are the way we kick those off in fleet

AgentsDGX agent

loops are great because its just ai running in the background 'triggers' are the way we kick those off in fleet What are loops, and how do you build one? A 'loop' is the repeated process where some ev

man its kinda wild to write “agent lab” on my random blog and a few months later its adopted by people like @breeves08 and @ScottWu46 🫡 htt…

AgentsDGX agent

man its kinda wild to write “agent lab” on my random blog and a few months later its adopted by people like @breeves08 and @ScottWu46 🫡 https://x.com/cognition/status/2062923088778367314?s=46 More tha

MAR:Multi-Agent Reflexion Improves Reasoning Abilities in LLMs

AgentsDGX agent

arXiv:2512.20845v2 Announce Type: replace Abstract: LLMs have shown the capacity to improve their performance on reasoning tasks through reflecting on their mistakes, and acting with these reflections

MASS: Deep Research for Social Sciences with Memory-Augmented Social Simulation

AgentsDGX agent

arXiv:2606.09198v1 Announce Type: new Abstract: Deep Research agents powered by Large Language Models (LLMs) have exhibited extraordinary potential in automated paper writing tasks. However, existing

MAVIS: Multi-Agent Video Retrieval via Structured Video Understanding

AgentsDGX agent

arXiv:2606.09641v1 Announce Type: new Abstract: The dominant paradigm in video retrieval relies on embedding-based full-corpus scanning, which suffers from inherent computational inefficiency and the

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

AgentsDGX agent

arXiv:2606.07909v1 Announce Type: new Abstract: Modern large language model (LLM) agents can use external tools to help users solve complex tasks. However, for problems that require learning from long

MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

AgentsDGX agent

arXiv:2606.07603v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning capabilities, yet most LLM-based agents are statically deployed and unable to improve through ta

MinNav: Minimalist Navigation Using Optical Flow For Active Tiny Aerial Robots

AgentsDGX agent

arXiv:2606.07813v1 Announce Type: cross Abstract: Navigation using a monocular camera is pivotal for autonomous operation on tiny aerial robots due to their perfect balance of versatility, cost and ac

Motion planning for hundreds of floating robots

AgentsDGX agent

arXiv:2606.09620v1 Announce Type: new Abstract: Planning collision-free motion for large robot fleets is difficult because collision avoidance induces strong inter-agent coupling that grows rapidly wi

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

AgentsDGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

Observability for Delegated Execution in Agentic AI Systems

AgentsDGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

AgentsDGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

Online Agent-as-a-Judge: Situation-Generating Evaluation for Interactive Agents

AgentsDGX agent

arXiv:2606.08200v1 Announce Type: new Abstract: Evaluating LLM-powered interactive social agents is challenging because socially relevant behaviors depend not only on isolated outputs, but also on pri

PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents

AgentsDGX agent

arXiv:2606.08106v1 Announce Type: new Abstract: Self-evolving agents improve by repeatedly proposing changes to their own prompts, skills, or workflows and keeping those that score higher on a small h

Parsing a document accurately is one thing. Proving where every value came from is another. When a compliance team reviews an AI extraction,…

AgentsDGX agent

Parsing a document accurately is one thing. Proving where every value came from is another. When a compliance team reviews an AI extraction, or an auditor needs to sign off on a figure pulled from a f

Perturbative Contrastive Physical Learning

AgentsDGX agent

arXiv:2606.09756v1 Announce Type: new Abstract: Responses to perturbations are key to understanding physical systems. The ability to contrast such responses by comparing how a system reacts under slig

PRISM: Recovering Instruction Sets from Language Model Activations

AgentsDGX agent

arXiv:2606.09563v1 Announce Type: new Abstract: As LLMs are deployed as agents, reliable monitoring requires knowing not only what they output, but which instructions are steering their behavior. This

Projecting the Emerging Mindset of SWE Agent by Launching a Wild Code Understanding Journey

AgentsDGX agent

arXiv:2606.08500v1 Announce Type: cross Abstract: Software engineering agents (SWE agents) increasingly work through tool-mediated trajectories in real repositories, yet their behavior remains difficu

puffy plays pickleball june 30 in SF, co-hosted with good friends http://luma.com/the-agent-open

AgentsDGX agent

Puffy participated in a pickleball event on June 30 in San Francisco called 'The Agent Open,' which was co-hosted with Good Friends. The event appears to have been organized or promoted through Luma,

QueryWeaver: Reliable Multi-Tool Query Execution Planning via LLM-Based Graph Generation

AgentsDGX agent

arXiv:2606.08300v1 Announce Type: new Abstract: Many real-world queries over personal data span multiple applications and require structured planning, as individual tools expose only partial informati

RAILS: Verification-Native Clearing For Agentic Commerce

AgentsDGX agent

arXiv:2606.08790v1 Announce Type: new Abstract: Autonomous agents negotiate, purchase, deploy code, and move funds, but no neutral mechanism determines whether they met their delegated obligation, who

REFLECT: Intervention-Supported Error Attribution for Silent Failures in LLM Agent Traces

AgentsDGX agent

arXiv:2606.09071v1 Announce Type: new Abstract: Large language model (LLM) agents now solve complex tasks through long plan-and-execution traces, yet the ability to locate errors in a completed traces

Reflection in the Dark: Exposing and Escaping the Black Box in Reflective Prompt Optimization

AgentsDGX agent

arXiv:2603.18388v2 Announce Type: replace Abstract: Automatic prompt optimization (APO) has emerged as a powerful paradigm for improving LLM performance without manual prompt engineering. Reflective A

Reminder: The only place to download the Hermes Agent Desktop App is https://hermes-agent.nousresearch.com/desktop Any other website or sour…

AgentsDGX agent

Reminder: The only place to download the Hermes Agent Desktop App is https://hermes-agent.nousresearch.com/desktop Any other website or source is dangerous and could contain old, or worse, dangerous a

RepoLaunch: Automating Build and Management of Code Repositories across Languages and Platforms

AgentsDGX agent

arXiv:2603.05026v2 Announce Type: replace-cross Abstract: Language model (LM) agents have driven substantial progress in automated software engineering (SWE), yet building and testing software reposit

Reward Evolution with Graph-of-Thoughts: A Bi-Level Language Model Framework for Reinforcement Learning

AgentsDGX agent

arXiv:2509.16136v5 Announce Type: replace Abstract: Designing effective reward functions remains a major challenge in reinforcement learning (RL), often requiring considerable human expertise and iter

SAGE: An LLM-driven Self Reflective Agentic Framework for Fraud Detection

AgentsDGX agent

arXiv:2606.08146v1 Announce Type: new Abstract: Fraud detection in payment, e-commerce, and telecommunications systems requires accuracy at the individual level, robustness under severe class imbalanc

SearchSwarm: Towards Delegation Intelligence in Agentic LLMs for Long-Horizon Deep Research

AgentsDGX agent

arXiv:2606.09730v1 Announce Type: new Abstract: Large language models are increasingly expected to handle complex, long-horizon real-world tasks whose context demands can grow without bound, yet model

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and re…

AgentsDGX agent

// Self-Harness: Harnesses That Improve Themselves // (bookmark this one) Most of the agent scaffolds we rely on today are built once and remain frozen or mostly unchanged. The harness, like the skill

Self-Paced Curriculum Reinforcement Learning for Autonomous Superbike Racing in Simulation

AgentsDGX agent

arXiv:2606.09236v1 Announce Type: cross Abstract: Autonomous Racing has seen remarkable progress through deep Reinforcement Learning (RL), primarily for four-wheeled vehicles. However, motorbikes intr

Shape Formation for the Cooperative Transportation of Arbitrary Objects Using Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2606.09610v1 Announce Type: cross Abstract: Cooperative object transportation is essential in numerous domains, including industrial to domestic services. A popular transportation strategy is to

SIGA: Self-Evolving Coding-Agent Adapters for Scientific Simulation

AgentsDGX agent

arXiv:2606.09774v1 Announce Type: new Abstract: Advanced scientific simulators expose specialized input languages that turn simulation goals into executable configurations, but learning them can cost

SkillHone: A Harness for Continual Agent Skill Evolution Through Persistent Decision History

AgentsDGX agent

arXiv:2606.08671v1 Announce Type: new Abstract: Agent skills extend language-model agents with task-specific procedures, scripts, and references, but the tasks and environments they target continually

SKILL.nb: Selective Formalization and Gated Execution for Durable Agent Workflows

AgentsDGX agent

arXiv:2606.08049v1 Announce Type: new Abstract: AI agents increasingly turn past experience into reusable artifacts such as code, workflows, and procedural memories. Reuse can improve efficiency, but

Structuring agentic AI for HPC code modernization

AgentsDGX agent

arXiv:2606.08710v1 Announce Type: cross Abstract: Modernization of legacy scientific codes is often necessary to keep up with the ever-evolving changes in the compute resource ecosystem. Parallelizati

Syll: Open-Source Personal Automation with Cross-Surface Execution

AgentsDGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

← Previous
1…4243444546…121
Next →