AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,214 results
2 Jun 2026

LeARN: Learnable and Adaptive Representations for Nonlinear Dynamics in System Identification

AgentsDGX agent

arXiv:2412.12036v2 Announce Type: replace Abstract: System identification, the process of deriving mathematical models of dynamical systems from observed input-output data, has undergone a paradigm sh

Learning Query-Specific Rubrics from Human Preferences for DeepResearch Report Generation

AgentsDGX agent

arXiv:2602.03619v2 Announce Type: replace Abstract: Nowadays, developing reliable DeepResearch-style long-form report generation remains challenging, as training and evaluation lack verifiable reward

Learning to Construct Practical Agentic Systems

AgentsDGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning-To-Measure: In-Context Active Feature Acquisition

AgentsDGX agent

arXiv:2510.12624v2 Announce Type: replace-cross Abstract: Active feature acquisition (AFA) is a sequential decision-making problem where the goal is to improve model performance for test instances by

LRAgent: Efficient KV Cache Sharing for Multi-LoRA LLM Agents

AgentsDGX agent

arXiv:2602.01053v2 Announce Type: replace Abstract: Role specialization in multi-LLM agent systems is often realized via multi-LoRA, where agents share a pretrained backbone and differ only by lightwe

MACCA: Offline Multi-agent Reinforcement Learning with Causal Credit Assignment

AgentsDGX agent

arXiv:2312.03644v3 Announce Type: replace Abstract: Offline Multi-agent Reinforcement Learning (MARL) is valuable in scenarios where online interaction is impractical or risky. While independent learn

MARFT: Multi-Agent Reinforcement Fine-Tuning

AgentsDGX agent

arXiv:2504.16129v5 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based Multi-Agent Systems (LaMAS) have demonstrated strong capabilities on complex agentic tasks requiring multifac

Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

AgentsDGX agent

arXiv:2606.00408v1 Announce Type: cross Abstract: Long-horizon search agents accumulate large amounts of retrieved content across many tool calls, making context-budget efficiency increasingly importa

MemGraphRAG: Memory-based Multi-Agent System for Graph Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2606.00610v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has become an essential method for mitigating hallucinations in Large Language Models (LLMs) by leveraging extern

MemPro: Agentic Memory Systems as Evolvable Programs

AgentsDGX agent

arXiv:2606.00619v1 Announce Type: cross Abstract: Long-horizon autonomous agents require memory systems to retain historical information, track evolving states, and reuse relevant knowledge beyond fin

micropython-wasm 0.1a1

AgentsDGX agent

micropython-wasm 0.1a1 is a Python library for running a MicroPython sandbox using WebAssembly . This alpha release includes fixes for limitations discovered while building datasette-agent-micropython

Microsoft announces the Agent Control Specification, an open-source standard that gives developers a granular, consistent way to control what AI agents can do (Ram Iyer/TechCrunch)

AgentsDGX agent

Ram Iyer / TechCrunch: Microsoft announces the Agent Control Specification, an open-source standard that gives developers a granular, consistent way to control what AI agents can do — As AI agents gro

Microsoft Build 2026: All the news about Windows, AI, RTX Spark, and more

AgentsDGX agent

Microsoft’s annual developer conference is kicking off on June 2nd in San Francisco with the keynote presentation streaming live at 12:30PM ET / 9:30AM PT, and we will be following along here with eve

Microsoft debuts an expansion of its model families and agentic AI intelligence for developers

AgentsDGX agent

Microsoft Corp. announced an expansion to its artificial intelligence models and agentic AI infrastructure today that brings more data and context into the hands of developers and business users as th

Microsoft unveils Project Solara, an Android-based platform for agent-first devices, with concept hardware and pilots planned at Best Buy, Target, and others (Todd Bishop/GeekWire)

AgentsDGX agent

Todd Bishop / GeekWire: Microsoft unveils Project Solara, an Android-based platform for agent-first devices, with concept hardware and pilots planned at Best Buy, Target, and others — [Editor's Note:

Microsoft’s Project Solara is an OS for AI agent gadgets

AgentsDGX agent

Microsoft just announced 'Project Solara,' a new OS designed for gadgets that run AI agents, at Build 2026. The company is calling it 'a new platform built from the ground up to power agent-driven exp

MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention

AgentsDGX agent

arXiv:2606.01063v1 Announce Type: new Abstract: Theory of Mind (ToM) enables an agent to reason about another actor's beliefs, goals, and intentions, which is essential for human-centered embodied ass

MIRROR: A Multi-Agent Framework with Iterative Adaptive Revision and Hierarchical Retrieval for Optimization Modeling in Operations Research

AgentsDGX agent

arXiv:2602.03318v3 Announce Type: replace Abstract: Operations Research (OR) relies on expert-driven modeling-a slow and fragile process ill-suited to novel scenarios. While large language models (LLM

Modeling Distinct Human Interaction in Web Agents

AgentsDGX agent

arXiv:2602.17588v3 Announce Type: replace Abstract: Despite rapid progress in autonomous web agents, human involvement remains essential for shaping preferences and correcting agent behavior as tasks

Monitoring Agentic Systems Before They're Reliable

AgentsDGX agent

arXiv:2606.02494v1 Announce Type: cross Abstract: Agentic systems entering production typically operate as partially integrated assemblies where structural defects, not task-level errors, dominate the

Multi-Agent Conformal Prediction with Personalized Statistical Validity

AgentsDGX agent

arXiv:2606.00717v1 Announce Type: cross Abstract: Uncertainty quantification is essential in high-stakes machine learning tasks. However, one of the principled solutions, conformal prediction, faces c

MUSCLE-NET: Predicted-Multiscale-Aware Network for Pedestrian Trajectory Forecasting

AgentsDGX agent

arXiv:2606.00471v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction is essential for safe navigation in autonomous driving and intelligent transportation systems. Despite substan

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

AgentsDGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

NestRL: A Nested Training Regime for Mutual Adaptation in Human-AI Teaming

AgentsDGX agent

arXiv:2602.17737v2 Announce Type: replace-cross Abstract: Mutual adaptation is a central challenge in human-AI teaming, as humans naturally adjust their strategies in response to an AI agent's behavio

New in Deep Agents: Agent Rubrics! Attach a rubric to your agent invocation, and a grader evaluates and self-corrects output until it satisf…

AgentsDGX agent

New in Deep Agents: Agent Rubrics! Attach a rubric to your agent invocation, and a grader evaluates and self-corrects output until it satisfies all requirements. This is helpful for long/complex tasks

Not All Flips Are Conformity: Decomposing Stance Convergence in Multi-Agent LLM Debate

AgentsDGX agent

arXiv:2606.00820v1 Announce Type: new Abstract: Multi-agent debate (MAD) is a promising strategy for improving LLM reasoning, but when agents converge on a shared answer, it is unclear whether that co

OctoT2I: A Self-Evolving Agentic Text-to-Image Router

AgentsDGX agent

arXiv:2606.01803v1 Announce Type: new Abstract: The explosive growth of Text-to-Image (T2I) models, from large-scale versions to lightweight, real-time ones, now faces diminishing marginal returns fro

ODTQA-FoRe: An Open-Domain Tabular Question Answering Dataset for Future Data Forecasting and Reasoning

AgentsDGX agent

arXiv:2606.02433v1 Announce Type: cross Abstract: The rapid development of LLMs has significantly advanced tabular question answering, but most systems cannot perform future-oriented numerical predict

On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents

AgentsDGX agent

arXiv:2603.12109v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a de facto paradigm for building LLM-based agents that act, interact, and reason over extended task horizons.

OpenAI unveils new Codex plugins for tasks related to public equity investment, banking and sales, and other roles, and plans to integrate Codex into ChatGPT (Shirin Ghaffary/Bloomberg)

AgentsDGX agent

Shirin Ghaffary / Bloomberg: OpenAI unveils new Codex plugins for tasks related to public equity investment, banking and sales, and other roles, and plans to integrate Codex into ChatGPT — OpenAI is e

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

AgentsDGX agent

arXiv:2603.14771v3 Announce Type: replace Abstract: Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting th

PairedGTA: Generating Driving Datasets for Controlled Photometric Shift Analysis

AgentsDGX agent

arXiv:2606.01192v1 Announce Type: new Abstract: Evaluating the performance of visual perception systems for autonomous driving is essential to ensure reliable operation across diverse environmental sc

PillarDETR: YOLO-Backbone and RT-DETR Head for Real-Time 3D Object Detection

AgentsDGX agent

arXiv:2606.01757v1 Announce Type: new Abstract: Real-time 3D object detection is a critical component for the safe operation of autonomous driving systems and robotics. While LiDAR point clouds provid

PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation

AgentsDGX agent

arXiv:2602.01662v4 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have enabled increasing progress in real-world robot manipulation. However, long-horizon manipulati

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps

AgentsDGX agent

arXiv:2606.01788v1 Announce Type: new Abstract: Embodied visual navigation, where an agent perceives a complex environment and acts to reach a goal from raw sensory input, underpins a wide range of ap

Quick explainer of our managed deepagents offering

AgentsDGX agent

This post likely provides a brief explanation of a managed deepagents offering, describing what deepagents are, how the managed service works, and its key benefits or use cases. As a post from Harriso

RadAgent: A tool-using AI agent for stepwise interpretation of chest computed tomography

AgentsDGX agent

arXiv:2604.15231v2 Announce Type: replace Abstract: Vision-language models (VLM) have markedly advanced AI-driven interpretation and reporting of complex medical imaging, such as computed tomography (

Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory

AgentsDGX agent

arXiv:2604.03588v3 Announce Type: replace Abstract: AI agents operating over extended time horizons accumulate experiences that serve multiple concurrent goals, and must often maintain conflicting int

RCM-ACT: Imitation Learning with Dynamic RCM Calibration for Autonomous Intraocular Foreign Body Removal

AgentsDGX agent

arXiv:2508.19191v3 Announce Type: replace Abstract: Intraocular foreign body removal demands millimeter-level precision in confined intraocular spaces, yet existing robotic systems predominantly rely

RDA: Reward Design Agent for Reinforcement Learning

AgentsDGX agent

arXiv:2606.01672v1 Announce Type: new Abstract: Reinforcement learning has enabled the acquisition of impressive robotic skills, but typically requires hand-crafted reward functions that are slow to d

Read more about hybrid agentic inference in Perplexity Computer: https://www.perplexity.ai/hub/blog/the-data-center-moves-to-your-machine

AgentsDGX agent

Perplexity discusses hybrid agentic inference in Perplexity Computer, which likely describes a computational approach that combines processing between local machines and data centers. The concept sugg

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

AgentsDGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

Rehumanizing global health care with agentic AI

AgentsDGX agent

The global health care sector is under increasing strain. Decades of chronic underinvestment and constraints in recruitment have coincided with a surge in demand for services for aging populations. Ga

RelationalAI beefs up its reasoning capabilities to enhance AI agent decision-making

AgentsDGX agent

Enterprise decision intelligence startup RelationalAI Inc. is advancing its capabilities for Snowflake Inc.’s AI Data Cloud platform. At Snowflake Summit 2026 today, it announced a series of updates t

RocketSmith: An Agentic System for High-Powered Rocket Design and Manufacturing

AgentsDGX agent

arXiv:2606.00097v1 Announce Type: new Abstract: This work presents RocketSmith, an agentic system capable of the design, manufacturing, and optimization processes in high powered rocket development. T

Scaling Agentic Capabilities via Grounded Interaction Synthesis

AgentsDGX agent

arXiv:2606.02001v1 Announce Type: new Abstract: General agentic intelligence hinges on the ability to interact with diverse real-world tools to complete complex tasks, a capability fundamentally tied

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

AgentsDGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's poss…

AgentsDGX agent

// Scaling Behavior of Single LLM-Driven Multi-Agent Systems // Does adding more agents actually make a multi-agent system better? It's possible that collective intelligence emerges from interaction d

SciAgentGym: Benchmarking Multi-Step Scientific Tool-use in LLM Agents

AgentsDGX agent

arXiv:2602.12984v2 Announce Type: replace Abstract: Scientific reasoning inherently demands integrating sophisticated toolkits to navigate domain-specific knowledge. Yet, current benchmarks largely ov

Self-Revising Discovery Systems for Science: A Categorical Framework for Agentic Artificial Intelligence

AgentsDGX agent

arXiv:2606.01444v1 Announce Type: new Abstract: Scientific discovery is not only answer generation but revision of the representational regime in which evidence, artifacts, operations, and verifiers a

Seq-DeepIPC: Sequential Sensing for End-to-End Control in Legged Robot Navigation

AgentsDGX agent

arXiv:2510.23057v2 Announce Type: replace-cross Abstract: We present Seq-DeepIPC, a sequential end-to-end perception-to-control model for legged robot navigation in real-world environments. Seq-DeepIP

Site4Drug: Predicting Drug-Binding Target Sites with an AI Agent

AgentsDGX agent

arXiv:2606.01816v1 Announce Type: cross Abstract: Selecting where to intervene on a protein (i.e., choosing a targetable site) is often a more ambiguous and failure-prone bottleneck than selecting wha

Situation-Aware Interactive MPC Switching for Autonomous Driving

AgentsDGX agent

arXiv:2512.06182v2 Announce Type: replace Abstract: Autonomous driving in interactive traffic scenarios remains challenging because of the mutual influence among vehicles and the inherent uncertainty

'Skill issues'': data-centric optimization of lakehouse agents

AgentsDGX agent

arXiv:2606.01185v1 Announce Type: new Abstract: Coding agents are becoming users of data infrastructure, but their success depends not only on model quality: it also depends on the skills and environm

SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision

AgentsDGX agent

arXiv:2606.01139v1 Announce Type: new Abstract: Agent skills are procedural artifacts that enable LLM agents to execute workflows, verify constraints, and recover from failures. Existing self-evolving

SkillSmith: Co-Evolving Skills and Tools for Self-Improving Agent Systems

AgentsDGX agent

arXiv:2606.01314v1 Announce Type: new Abstract: Recent self-evolving agents have shown that skills can be discovered, refined, and accumulated through execution. However, existing skill-evolution fram

SortingHat: Redefining Operating Systems Education with a Tailored Digital Teaching Assistant

AgentsDGX agent

arXiv:2606.00015v1 Announce Type: cross Abstract: Operating Systems (OS) courses are among the most challenging in computer science education due to the complexity of internal structures and the diver

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL

AgentsDGX agent

arXiv:2512.04069v2 Announce Type: replace Abstract: Vision Language Models (VLMs) demonstrate strong qualitative visual understanding, but struggle with metrically precise spatial reasoning required f

STEM: Semantic Target Search and Exploration using MAVs in Cluttered Environments

AgentsDGX agent

arXiv:2606.00762v1 Announce Type: new Abstract: Autonomous target search is crucial for deploying Micro Aerial Vehicles (MAVs) in emergency response and rescue missions. Existing approaches either foc

Stop Wandering, Find the Keys: LLMs Discriminate Key States for Efficient Multi-Agent Exploration

AgentsDGX agent

arXiv:2410.02511v2 Announce Type: replace Abstract: With expansive state-action spaces, efficient multi-agent exploration remains a longstanding challenge in reinforcement learning. Although pursuing

← Previous
1…5152535455…121
Next →