AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
26 May 2026

Test-Time Deep Thinking to Explore Implicit Rules

AgentsDGX agent

arXiv:2605.24828v1 Announce Type: new Abstract: With the continuous advancement of Large Language Models (LLMs), intelligent agents are becoming increasingly vital. However, these agents often fail in

25 May 2026

Efficient and Transferable Agentic Knowledge Graph RAG via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.26383v5 Announce Type: replace-cross Abstract: Knowledge-graph retrieval-augmented generation (KG-RAG) couples large language models (LLMs) with structured, verifiable knowledge graphs (KGs

One Policy, Infinite NPCs: Persona-Traceable Shared RL Policies for Scalable Game Agents

Model Releases
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.23652v1 Announce Type: new Abstract: On a 300-persona life-simulation benchmark, pcsp achieves compositional zero-shot persona identification up to 17x above chance, Spearman rho approx 0.7

VideoTemp-o3: Harmonizing Temporal Grounding and Video Understanding in Agentic Thinking-with-Videos

Model ReleasesDGX agent

arXiv:2602.07801v4 Announce Type: replace-cross Abstract: In long-video understanding, conventional uniform frame sampling often fails to capture key visual evidence, leading to degraded performance a

23 May 2026

AutoMCU: Feasibility-First MCU Neural Network Customization via LLM-based Multi-Agent Systems

HardwareDGX agent

arXiv:2605.21560v1 Announce Type: new Abstract: Deploying neural networks on microcontroller units (MCUs) is critical for edge intelligence but remains challenging due to tight memory, storage, and co

22 May 2026

AgroTools: A Benchmark for Tool-Augmented Multimodal Agents in Agriculture

Model ReleasesDGX agent

arXiv:2605.22366v1 Announce Type: new Abstract: Agricultural decision-making increasingly requires multimodal systems that can transform visual observations into reliable, executable actions. However,

Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety

Model ReleasesDGX agent

arXiv:2605.22643v1 Announce Type: new Abstract: Background. Traditional safety benchmarks for language models evaluate generated text: whether a model outputs toxic language, reproduces bias, or follo

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2605.22144v1 Announce Type: new Abstract: Existing approaches for digital short-drama production typically rely on one-shot LLM generated scripts and loosely coupled pipelines, which fail to sat

21 May 2026

Beyond Text-to-SQL: An Agentic LLM System for Governed Enterprise Analytics APIs

SafetyDGX agent

arXiv:2605.21027v1 Announce Type: new Abstract: Enterprise analytics aims to make organizational data accessible for decision-making, yet non-technical users still face barriers when using traditional

IndusAgent: Reinforcing Open-Vocabulary Industrial Anomaly Detection with Agentic Tools

Local AiDGX agent

arXiv:2605.20682v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown remarkable capability in bridging visual perception and textual reasoning, enabling zero-shot unders

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

Model ReleasesDGX agent

arXiv:2605.20189v1 Announce Type: cross Abstract: Despite the remarkable success of large language models (LLMs), they still face bottlenecks while deploying in dynamic, real-world settings with prima

20 May 2026

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

Model ReleasesDGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization

HardwareDGX agent

arXiv:2605.19782v1 Announce Type: new Abstract: LLM discovery and optimization systems are increasingly applied across domains, implementing a common propose-evaluate-revise loop. Such optimization or

What Do Evolutionary Coding Agents Evolve?

Model ReleasesDGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

19 May 2026

AscendOptimizer: Episodic Agent for Ascend NPU Operator Optimization

Model ReleasesDGX agent

arXiv:2603.23566v2 Announce Type: replace-cross Abstract: Optimizing AscendC (Ascend C) operators for Ascend NPUs is difficult for two reasons. First, unlike CUDA, the ecosystem offers few public kern

EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis

Model ReleasesDGX agent

arXiv:2508.07292v3 Announce Type: replace Abstract: Endoscopic diagnosis is an iterative process in which clinicians progressively acquire, compare, and verify local visual evidence before reaching a

Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

Model ReleasesDGX agent

arXiv:2605.18144v1 Announce Type: new Abstract: Nanomedicine research spans delivery chemistry, immunology, imaging, biomaterials, and disease-specific translational science, yet its conceptual design

LARGER: Lexically Anchored Repository Graph Exploration and Retrieval

AgentsDGX agent

arXiv:2605.16352v1 Announce Type: cross Abstract: Repository-level coding agents must first localize the files and symbols relevant to a task; failures at this stage can cascade across downstream obje

LLM Agents Are the Antidote to Walled Gardens

ApplicationsDGX agent

arXiv:2506.23978v3 Announce Type: replace-cross Abstract: While the Internet's core infrastructure was designed to be open and universal, today's application layer is dominated by closed, proprietary

MADP: A Multi-Agent Pipeline for Sustainable Document Processing with Human-in-the-Loop

Model ReleasesDGX agent

arXiv:2605.17159v1 Announce Type: new Abstract: Document processing automation remains a critical challenge in enterprise environments, where traditional manual approaches are labor-intensive and erro

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

Model ReleasesDGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

Model ReleasesDGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

SafetyDGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback

Model ReleasesDGX agent

arXiv:2605.17448v1 Announce Type: cross Abstract: Computer-aided design (CAD) is the backbone of modern industrial design, yet learned CAD generators still fall short of real engineering pipelines: th

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.17423v1 Announce Type: new Abstract: We study series-level cinematic remaking, a long-horizon video-to-video generation problem that localizes full episodes or films via stylization or acto

18 May 2026

Dell targets enterprise AI execution gap with local agentic AI systems and integrated AI infrastructure

Local AiDGX agent

Dell Technologies Inc. today is kicking off its Dell Technologies World conference by expanding its artificial intelligence portfolio with enhancements aimed at helping enterprises move AI projects fr

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

Model ReleasesDGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

Task-Semantic Graph-Driven Distributed Agent Networking for Underwater Target Tracking

SafetyDGX agent

arXiv:2605.15528v1 Announce Type: new Abstract: Autonomous underwater vehicle (AUV) swarms are emerging as intelligent underwater networks, where each node must sense, communicate, process local data,

15 May 2026

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

Model ReleasesDGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

ASH: Agents that Self-Hone via Embodied Learning

SafetyDGX agent

arXiv:2605.14211v1 Announce Type: new Abstract: Long-horizon embodied tasks remain a fundamental challenge in AI, as current methods rely on hand-engineered rewards or action-labeled demonstrations, n

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution

Model ReleasesDGX agent

arXiv:2605.14504v1 Announce Type: new Abstract: Long-horizon household tasks demand robust high-level planning and sustained reasoning capabilities, which are largely overlooked by existing embodied A

14 May 2026

1/5 Caching carries a deterministic assumption baked in: same input, same output. That breaks down with LLMs, and especially with agents run…

Model ReleasesDGX agent

This post discusses how traditional caching mechanisms assume deterministic behavior (identical inputs producing identical outputs), an assumption that breaks down with large language models and espec

An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing

Local AiDGX agent

arXiv:2605.13221v1 Announce Type: new Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms

EvoGround: Self-Evolving Video Agents for Video Temporal Grounding

TutorialsDGX agent

arXiv:2605.13803v1 Announce Type: new Abstract: Video temporal grounding (VTG) takes an untrimmed video and a natural-language query as input and localizes the temporal moment that best matches the qu

Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation

Model ReleasesDGX agent

arXiv:2605.12953v1 Announce Type: cross Abstract: Language-guided segmentation transcends the scope limitations of traditional semantic segmentation, enabling models to segment arbitrary target region

13 May 2026

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

HardwareDGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

Distributed Quantum Gaussian Processes for Multi-Agent Systems

Local AiDGX agent

arXiv:2602.15006v2 Announce Type: replace-cross Abstract: Gaussian Processes (GPs) are a powerful tool for probabilistic modeling, but their performance is often constrained in complex, large-scale re

12 May 2026

An Uncertainty-Aware Resilience Micro-Agent for Causal Observability in the Computing Continuum

Local AiDGX agent

arXiv:2605.10718v1 Announce Type: cross Abstract: Grey failures in the computing continuum produce ambiguous overlapping symptoms that existing approaches fail to diagnose reliably, either due to a la

Breaking the Impasse: Dual-Scale Evolutionary Policy Training for Social Language Agents

SafetyDGX agent

arXiv:2605.08721v1 Announce Type: new Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for closed-ended tasks, extending it to open-ended social language game

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

Model ReleasesDGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

Mirror, Mirror on the Wall: Can VLM Agents Tell Who They Are at All?

Model ReleasesDGX agent

arXiv:2605.08816v1 Announce Type: new Abstract: In the animal kingdom, mirror self-recognition is a canonical probe of higher-order cognition, emerging only in some species. We ask whether an analogou

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces

Model ReleasesDGX agent

arXiv:2605.08904v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and tool use. However, the fundamental cognitive faculties essential

RewardHarness: Self-Evolving Agentic Post-Training

Model ReleasesDGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

TRACER: Verifiable Generative Provenance for Multimodal Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.09934v1 Announce Type: new Abstract: Multimodal large language models increasingly solve vision-centric tasks by calling external tools for visual inspection, OCR, retrieval, calculation, a

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

Model ReleasesDGX agent

arXiv:2605.09769v1 Announce Type: new Abstract: This paper describes our system for classifying psychological defense mechanisms in emotional support dialogues using the Defense Mechanism Rating Scale

11 May 2026

Cognitive Agent Compilation for Explicit Problem Solver Modeling

SafetyDGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

Many-to-Many Multi-Agent Pickup and Delivery

SafetyDGX agent

arXiv:2605.07835v1 Announce Type: new Abstract: Multi-robot systems in automated warehouses must manage continuous streams of pickup-and-delivery tasks while ensuring efficiency and safety. Prior work

SOD: Step-wise On-policy Distillation for Small Language Model Agents

SafetyDGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent

Model ReleasesDGX agent

arXiv:2601.18700v2 Announce Type: replace Abstract: Emotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance. Howeve

Tools as Continuous Flow for Evolving Agentic Reasoning

Model ReleasesDGX agent

arXiv:2605.07339v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in orchestrating tools for reasoning tasks. However, existing methods rely on a s

7 May 2026

cotomi Act: Learning to Automate Work by Watching You

AgentsDGX agent

arXiv:2605.03231v1 Announce Type: new Abstract: What if a browser agent could learn your work simply by watching you do it? We present cotomi Act, a browser-based computer-using agent that combines re

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents

Model ReleasesDGX agent

arXiv:2604.01496v2 Announce Type: replace-cross Abstract: We introduce SWE-ZERO to SWE-HERO, a two-stage SFT recipe that achieves state-of-the-art results on SWE-bench by distilling open-weight fronti

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents

Model ReleasesDGX agent

arXiv:2509.24943v2 Announce Type: replace Abstract: Long videos, characterized by temporal complexity and sparse task-relevant information, pose significant reasoning challenges for AI systems. Althou

6 May 2026

AcademiClaw: When Students Set Challenges for AI Agents

Model ReleasesDGX agent

arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw

AI Agents for Inventory Control: Human-LLM-OR Complementarity

Model ReleasesDGX agent

arXiv:2602.12631v2 Announce Type: replace-cross Abstract: Inventory control is a fundamental operations problem in which ordering decisions are traditionally guided by theoretically grounded operation

EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents

Model ReleasesDGX agent

arXiv:2605.01250v1 Announce Type: new Abstract: Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

SafetyDGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

Model ReleasesDGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

SafetyDGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

5 May 2026

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

SafetyDGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

← Previous
1…140141142143144…300
Next →