AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
9 Jun 2026

Baichuan-M4: A Clinical-Grade Medical Agent System for Continuous Care

SafetyDGX agent

arXiv:2606.08982v1 Announce Type: new Abstract: Baichuan-M4 is Baichuan Intelligence's clinical-grade medical large model, designed for continuous care rather than single-turn medical question answeri

Cooperative Long Rope Skipping via Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.08064v1 Announce Type: new Abstract: Humans exhibit remarkable motor agility, enabling a wide range of dynamic skills such as running and jumping, which highlights the great potential of hu

HARBOR: A Harness Framework for Agentic Robot Reinforcement Learning

SafetyDGX agent

arXiv:2606.08610v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a powerful paradigm for robot learning, particularly in sim-to-real settings, but its broader adoption remains

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

HDRAgent: An Agentic Framework for Multi-Exposure HDR Imaging

SafetyDGX agent

arXiv:2606.09110v1 Announce Type: new Abstract: Most existing multi-exposure HDR methods follow a fixed feed-forward reconstruction paradigm, making them prone to ghosting artifacts in complex dynamic

Principled Agent Debate: Adversarial Arbitration for Sycophancy Reduction in Large Language Models

Model ReleasesDGX agent

arXiv:2606.07532v1 Announce Type: cross Abstract: RLHF-trained models are systematically biased toward agreement over accuracy, a structural property of the training process. We present Principled Age

Syll: Open-Source Personal Automation with Cross-Surface Execution

AgentsDGX agent

arXiv:2606.07594v1 Announce Type: new Abstract: Personal AI agents must increasingly operate across APIs, shells, web surfaces, and desktop GUIs, yet many systems remain tuned to a single interface an

8 Jun 2026

MADE: Beyond Scoring via a Multilingual Agentic Diagnosing Engine for Fine-Grained Evaluation Insights

Model ReleasesDGX agent

arXiv:2606.07020v1 Announce Type: new Abstract: Multilingual and multicultural benchmarks now cover dozens of languages and model families, but the resulting score landscapes remain metric-rich and in

6 Jun 2026

Agent-Orchestrated Adaptive RAG: A Comparative Study on Structured and Multi-Hop Retrieval

Model ReleasesDGX agent

arXiv:2606.05658v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding their responses in external knowledge, but conventional pipeli

CuTeGen: An LLM-Based Agentic Framework for Generation and Optimization of High-Performance GPU Kernels using CuTe

HardwareDGX agent

arXiv:2604.01489v2 Announce Type: replace-cross Abstract: High-performance GPU kernels are critical to modern machine learning systems, yet developing them remains a manual, expert-driven process. Rec

Vortex: Efficient and Programmable Sparse Attention Serving for AI Agents

Model ReleasesDGX agent

arXiv:2606.06453v1 Announce Type: new Abstract: Sparse attention is becoming increasingly important for serving large language models (LLMs) as generation lengths continue to grow. However, deploying

5 Jun 2026

Agents' Last Exam

Model ReleasesDGX agent

arXiv:2606.05405v1 Announce Type: cross Abstract: Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deploym

Thinking with Imagination: Agentic Visual Spatial Reasoning with World Simulators

Model ReleasesDGX agent

arXiv:2606.06476v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have shown strong visual reasoning capabilities, their spatial reasoning abilities remain largely constrained to the

4 Jun 2026

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities

Model ReleasesDGX agent

arXiv:2606.04460v1 Announce Type: cross Abstract: AI has the potential to transform cybersecurity by enabling systems that can autonomously detect, analyze, and remediate software vulnerabilities. How

SocialCoach: Personalized Social Skill Learning with RL-based Agentic Tutoring and Practice

SafetyDGX agent

arXiv:2606.04155v1 Announce Type: cross Abstract: Social skills such as negotiation and leadership are crucial for personal and professional success in today's interconnected world. However, scalable

Solving Zebra Puzzles Using Constraint-Guided Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2407.03956v3 Announce Type: replace-cross Abstract: Prior research has enhanced the ability of Large Language Models (LLMs) to solve logic puzzles using techniques such as chain-of-thought promp

What's new for Managed Service for Apache Spark clusters

Model ReleasesDGX agent

At Google Cloud, our goal is to let you run large-scale analytical and data science workloads with maximum efficiency so you can process big data pipelines, machine learning, and ETL tasks. We recentl

3 Jun 2026

A Scoping Review of the Ethical Perspectives on Anthropomorphising Large Language Model-Based Conversational Agents

ResearchDGX agent

arXiv:2601.09869v2 Announce Type: replace Abstract: Anthropomorphisation -- the phenomenon whereby non-human entities are ascribed human-like qualities -- has become increasingly salient with the rise

Discovery to Execution: Scaling Agents with Toolboxes and Routines in Microsoft Foundry

ApplicationsDGX agent

Tooling doesn’t break at a small scale—it breaks when teams move to production. AI adoption accelerates, so does the number of tools available to them. Discovering, managing and securing the right too

From Prompt to Service: An SLM-Based Agent Orchestration Gateway for AI-Driven Virtual Worlds

Model ReleasesDGX agent

arXiv:2606.03557v1 Announce Type: new Abstract: As generative AI capabilities expand, AI-driven virtual worlds face a growing architectural challenge. Users interact through in-world interfaces in mul

Libra: Efficient Resource Management for Agentic RL Post-Training

SafetyDGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

Validation-Gated Multi-Agent Governance for Online Adaptation of Thermal-Hydraulic Surrogate Models under Operating-Regime Shift

SafetyDGX agent

arXiv:2606.03321v1 Announce Type: new Abstract: Artificial-intelligence surrogates can support second-by-second thermal-hydraulic forecasting, but models selected and frozen offline may become conditi

VLESA: Vision-Language Embodied Safety Agent for Human Activity Monitoring

Model ReleasesDGX agent

arXiv:2606.03954v1 Announce Type: new Abstract: As AI systems increasingly assist humans in physical tasks, ensuring safety becomes paramount -- physical actions carry immediate and irreversible conse

2 Jun 2026

3rd Place at CVPR 2026 CASTLE Challenge: Agentic Multi-View Long-Context Video Understanding via Hierarchical Knowledge Graph Retrieval

Model ReleasesDGX agent

arXiv:2606.01933v1 Announce Type: new Abstract: This paper presents our winning methodology for the CASTLE 2026 Challenge at the CVPR 2026 EgoVis Workshop, where our team secured third place globally.

Bridging the Last Mile of Time Series Forecasting with LLM Agents

SafetyDGX agent

arXiv:2606.02497v1 Announce Type: new Abstract: Time series forecasting has advanced rapidly, especially with the emergence of foundation models that show strong zero-shot performance on numerical ext

CodeCytos: AI-assisted spatial molecular imaging analysis via code-augmented agent action space

Model ReleasesDGX agent

arXiv:2606.00472v1 Announce Type: cross Abstract: Conventional tissue image analysis software provides foundational capabilities for cellular analysis, including segmentation, basic morphological feat

HALO: Learning Human-Robot Collaboration via Heterogeneous-Agent Lyapunov Policy Optimization

Model ReleasesDGX agent

arXiv:2603.03741v2 Announce Type: replace-cross Abstract: To improve generalization and resilience in human-robot collaboration (HRC), robots must contend with diverse combinations of human behaviors

IMAC-AgriVLN: Can Agricultural Vision-and-Language Navigation Agents be Aware of Instruction Mistakes?

Model ReleasesDGX agent

arXiv:2606.02519v1 Announce Type: new Abstract: Agricultural robots are serving as powerful assistants across a wide range of agricultural tasks, nevertheless, still heavily relying on manual operatio

MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment Simulation

Model ReleasesDGX agent

arXiv:2606.02470v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has emerged as a transformative standard for connecting large language models (LLMs) with external data sources and too

RadioMaster: Multi-Agent System for Autonomous Radio Signal Generation

Model ReleasesDGX agent

arXiv:2606.01862v1 Announce Type: cross Abstract: Translating user intents into physical radio signals represents the critical yet notoriously tedious final step in wireless prototyping, as it require

Sandboxed Coding Agents are Competitive Omni-modal Task Solvers

Model ReleasesDGX agent

arXiv:2606.00579v1 Announce Type: new Abstract: As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not al

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

SafetyDGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

1 Jun 2026

An Organization-Scoped LLM Agent Runtime Architecture for Regulated Cybersecurity Operations

Local AiDGX agent

arXiv:2605.30604v1 Announce Type: cross Abstract: Regulated cybersecurity workflows lack a runtime substrate that enforces organization-level scope across retrieval, tool calls, memory, findings, repo

Crafter: A Multi-Agent Harness for Editable Scientific Figure Generation from Diverse Inputs

Model ReleasesDGX agent

arXiv:2605.30611v1 Announce Type: cross Abstract: Scientific figures are among the most effective means of communicating complex research ideas, yet producing publication-quality illustrations remains

More on the Hermes Skills Hub: https://hermes-agent.nousresearch.com/docs/guides/work-with-skills#the-skills-hub

AgentsDGX agent

The Hermes Skills Hub is a feature that allows users to discover, manage, and integrate skills within the Hermes agent framework, enabling extended functionality and customization of agent capabilitie

31 May 2026

Great article on harness engineering. https://www.langchain.com/blog/the-anatomy-of-an-agent-harness

AgentsDGX agent

This article from LangChain explores the architectural components and design principles of agent harnesses, which are systems that manage the execution and behavior of AI agents. The piece likely cove

30 May 2026

Personal agents light the fuse as Snowflake and Databricks move up the AI stack

IndustryDGX agent

The artificial intelligence wave is starting to look a bit like the personal computer era – with some obvious differences. The first similarity is personal productivity. Individuals are taking control

29 May 2026

feel very aligned with our vision & Ronak + the awesome Trajectory team’s on practically tackling Continual Learning at scale 🚀 there’s a v…

AgentsDGX agent

feel very aligned with our vision & Ronak + the awesome Trajectory team’s on practically tackling Continual Learning at scale 🚀 there’s a very good reason why teams are partly building “Observability

Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Evaluation

SafetyDGX agent

arXiv:2605.29430v1 Announce Type: new Abstract: Automatic speech recognition (ASR) is a core component of human--computer interaction and an increasingly important front-end for LLM-based assistants a

Train the Agent, Not the Expert: Learning to Harness Heterogeneous Experts for Multi-Turn Visual Reasoning

SafetyDGX agent

arXiv:2605.29894v1 Announce Type: new Abstract: Recent progress in computer vision has produced a wide range of powerful specialized models for detection, segmentation, counting, and other visual task

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

Model ReleasesDGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

28 May 2026

Chrome Enterprise rolls out AI agents and automation to streamline security management

Model ReleasesDGX agent

Google LLC today launched new enhancements to Chrome Enterprise, the company’s enterprise version of its Chrome browser, designed to provide greater administrative and security control to information

MACReD: A Multi-Agent Collaborative Reasoning Framework for Reaction Diagram Parsing

Model ReleasesDGX agent

arXiv:2605.28077v1 Announce Type: new Abstract: Parsing chemical reaction diagrams from scientific literature is challenging due to heterogeneous layouts, intertwined visual elements, and the difficul

MetaboT: An LLM-based Multi-Agent Frameworkfor Interactive Analysis of Mass SpectrometryMetabolomics Knowledge Graphs

Model ReleasesDGX agent

arXiv:2510.01724v2 Announce Type: replace Abstract: Mass spectrometry-based metabolomics generates complex, high-dimensional data that holds vast potential for biological discovery but remains difficu

Segment to Focus: Guiding Latent Action Models in the Presence of Distractors

AgentsDGX agent

arXiv:2602.02259v2 Announce Type: replace-cross Abstract: Latent action models (LAMs) offer a promising path to pre-training embodied agents on large amounts of action-free video. They infer latent ac

SwarmHarness: Skill-Based Task Routing via Decentralized Incentive-Aligned AI Agent Networks

HardwareDGX agent

arXiv:2605.28764v1 Announce Type: new Abstract: Vast quantities of compute (GPU cycles on personal workstations, idle inference servers, and edge devices between jobs) go unused because no incentive-a

27 May 2026

Adversarial Training for Robust Coverage Network under Worst-case Facility Losses

AgentsDGX agent

arXiv:2605.26763v1 Announce Type: cross Abstract: The Maximal Covering Location-Interdiction Problem (MCLIP) is a classic bi-level optimization problem, which is fundamental to resilient infrastructur

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

Model ReleasesDGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

MechRL: Reinforcement Learning Agents Perform Circuit Discovery for Mechanistic Interpretability

SafetyDGX agent

arXiv:2605.26343v1 Announce Type: new Abstract: Mechanistic interpretability has identified small sets of attention heads that implement specific behaviours in transformer language models, but recover

Scaling, Benchmarking, and Reasoning of Vision-Language Agents for Mobile GUI Navigation

Model ReleasesDGX agent

arXiv:2605.27134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown rapid progress in mobile GUI navigation. This paper presents a systematic study of data scaling, benchmarking,

SEAL: Self-Evolving Agentic Learning for Conversational Question Answering over Knowledge Graphs

Model ReleasesDGX agent

arXiv:2512.04868v2 Announce Type: replace-cross Abstract: Knowledge-based conversational question answering (KBCQA) confronts persistent challenges in resolving coreference, modeling contextual depend

sqlite AGENTS.md

AgentsDGX agent

sqlite AGENTS.md SQLite gained an AGENTS.md file five days ago - but it's not intended for their own development, it's presumably aimed at people who are pointing agents at the SQLite codebase. It inc

Synergetic Empowerment: Wireless Communications Meets Embodied Intelligence

AgentsDGX agent

arXiv:2509.10481v2 Announce Type: replace-cross Abstract: Wireless communication is evolving into an agent era, where large-scale agents with inherent embodied intelligence are not just users but acti

TADDLE: A Tool-Augmented Agent for Detecting Deficient LLM-Generated Peer Reviews

Model ReleasesDGX agent

arXiv:2605.26911v1 Announce Type: new Abstract: LLM-generated peer reviews are increasingly common at major venues, yet their deficiencies are hard to detect because they are uniformly fluent and well

26 May 2026

Agent-Centric Social Trajectory Prediction: A Free Energy Principle Perspective

SafetyDGX agent

arXiv:2605.25748v1 Announce Type: new Abstract: Trajectory prediction methods have demonstrated remarkable capabilities in capturing complex motion patterns. However, existing methods rely on global s

Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat

SafetyDGX agent

arXiv:2605.25091v1 Announce Type: new Abstract: As modern air combat evolves toward beyond-visual-range (BVR) multi-aircraft cooperative engagements, autonomous decision-making for unmanned combat aer

Explore Before You Solve: The Speed--Depth Trade-off in Epistemic Agents for ARC-AGI-3

Model ReleasesDGX agent

arXiv:2605.25931v1 Announce Type: new Abstract: We systematically investigate all 25 public ARC-AGI-3 games and find that every one is reachable through non-intelligent strategies: 10 in a single blin

LipoAgent: Coordinating Fine-Tuned LLM Agents for Safer Lipid Design

SafetyDGX agent

arXiv:2605.25250v1 Announce Type: new Abstract: Lipid nanoparticles (LNPs) are among the most clinically mature platforms for nucleic acid delivery, yet designing lipids that are both effective and bi

MMUEChange: A Generalized LLM Agent Framework for Intelligent Multi-Modal Urban Environment Change Analysis

SafetyDGX agent

arXiv:2601.05483v2 Announce Type: replace Abstract: Understanding urban environment change is essential for sustainable development. However, current approaches, particularly remote sensing change det

PrivFusion: A Privacy-preserving Multi-Agent Framework for Harmonizing Distributed Datasets

SafetyDGX agent

arXiv:2605.24249v1 Announce Type: new Abstract: The growing availability of clinical data has increased the use of machine learning, yet centralized data aggregation is often infeasible for sensitive

Spectral Retrieval: Multi-Scale Sinc Convolution over Token Embeddings for Localized Retrieval in LLM Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.24764v1 Announce Type: cross Abstract: [Abridged] - Spectral Retrieval is a plug-in re-ranking stage that interpolates between per-token MaxSim and mean-pool retrieval through a multi-scale

← Previous
1…139140141142143…300
Next →