AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

T2T-LA: A Topology-to-Topology LLM Agent for Graph Learning with Neither Feature Access nor Task Knowledge

DGX agent

arXiv:2512.08964v4 Announce Type: replace Abstract: Graph learning aims to convert data into graph representations, which are fundamental to many problems in machine learning for CAD, where circuits,

model-releasesarxiv-cs-lg
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Task-Semantic Graph-Driven Distributed Agent Networking for Underwater Target Tracking

DGX agent

arXiv:2605.15528v1 Announce Type: new Abstract: Autonomous underwater vehicle (AUV) swarms are emerging as intelligent underwater networks, where each node must sense, communicate, process local data,

safetyarxiv-cs-ro
18 May 2026
Model Releases

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

DGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

model-releasesarxiv-cs-ai
15 May 2026
Safety

ASH: Agents that Self-Hone via Embodied Learning

DGX agent

arXiv:2605.14211v1 Announce Type: new Abstract: Long-horizon embodied tasks remain a fundamental challenge in AI, as current methods rely on hand-engineered rewards or action-labeled demonstrations, n

safetyarxiv-cs-ai
15 May 2026
Model Releases

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution

DGX agent

arXiv:2605.14504v1 Announce Type: new Abstract: Long-horizon household tasks demand robust high-level planning and sustained reasoning capabilities, which are largely overlooked by existing embodied A

model-releasesarxiv-cs-ai
15 May 2026
Local Ai

An Agentic AI Framework with Large Language Models and Chain-of-Thought for UAV-Assisted Logistics Scheduling with Mobile Edge Computing

DGX agent

arXiv:2605.13221v1 Announce Type: new Abstract: In cloud manufacturing, unmanned aerial vehicles (UAVs) can support both product collection and mobile edge computing (MEC). This joint operation forms

local-aiarxiv-cs-ai
14 May 2026
Tutorials

EvoGround: Self-Evolving Video Agents for Video Temporal Grounding

DGX agent

arXiv:2605.13803v1 Announce Type: new Abstract: Video temporal grounding (VTG) takes an untrimmed video and a natural-language query as input and localizes the temporal moment that best matches the qu

tutorialsarxiv-cs-cv
14 May 2026
Model Releases

Seg-Agent: Test-Time Multimodal Reasoning for Training-Free Language-Guided Segmentation

DGX agent

arXiv:2605.12953v1 Announce Type: cross Abstract: Language-guided segmentation transcends the scope limitations of traditional semantic segmentation, enabling models to segment arbitrary target region

model-releasesarxiv-cs-ai
14 May 2026
Hardware

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

DGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

hardwarearxiv-cs-cl
13 May 2026
Local Ai

Distributed Quantum Gaussian Processes for Multi-Agent Systems

DGX agent

arXiv:2602.15006v2 Announce Type: replace-cross Abstract: Gaussian Processes (GPs) are a powerful tool for probabilistic modeling, but their performance is often constrained in complex, large-scale re

local-aiarxiv-cs-lg
13 May 2026
Local Ai

An Uncertainty-Aware Resilience Micro-Agent for Causal Observability in the Computing Continuum

DGX agent

arXiv:2605.10718v1 Announce Type: cross Abstract: Grey failures in the computing continuum produce ambiguous overlapping symptoms that existing approaches fail to diagnose reliably, either due to a la

local-aiarxiv-cs-ai
12 May 2026
Safety

Breaking the Impasse: Dual-Scale Evolutionary Policy Training for Social Language Agents

DGX agent

arXiv:2605.08721v1 Announce Type: new Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for closed-ended tasks, extending it to open-ended social language game

safetyarxiv-cs-cl
12 May 2026
Model Releases

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

DGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Mirror, Mirror on the Wall: Can VLM Agents Tell Who They Are at All?

DGX agent

arXiv:2605.08816v1 Announce Type: new Abstract: In the animal kingdom, mirror self-recognition is a canonical probe of higher-order cognition, emerging only in some species. We ask whether an analogou

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces

DGX agent

arXiv:2605.08904v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and tool use. However, the fundamental cognitive faculties essential

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RewardHarness: Self-Evolving Agentic Post-Training

DGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

TRACER: Verifiable Generative Provenance for Multimodal Tool-Using Agents

DGX agent

arXiv:2605.09934v1 Announce Type: new Abstract: Multimodal large language models increasingly solve vision-centric tasks by calling external tools for visual inspection, OCR, retrieval, calculation, a

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

DGX agent

arXiv:2605.09769v1 Announce Type: new Abstract: This paper describes our system for classifying psychological defense mechanisms in emotional support dialogues using the Defense Mechanism Rating Scale

model-releasesarxiv-cs-ai
12 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Safety

Many-to-Many Multi-Agent Pickup and Delivery

DGX agent

arXiv:2605.07835v1 Announce Type: new Abstract: Multi-robot systems in automated warehouses must manage continuous streams of pickup-and-delivery tasks while ensuring efficiency and safety. Prior work

safetyarxiv-cs-ro
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Model Releases

TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent

DGX agent

arXiv:2601.18700v2 Announce Type: replace Abstract: Emotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance. Howeve

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Tools as Continuous Flow for Evolving Agentic Reasoning

DGX agent

arXiv:2605.07339v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in orchestrating tools for reasoning tasks. However, existing methods rely on a s

model-releasesarxiv-cs-ai
11 May 2026
Agents

cotomi Act: Learning to Automate Work by Watching You

DGX agent

arXiv:2605.03231v1 Announce Type: new Abstract: What if a browser agent could learn your work simply by watching you do it? We present cotomi Act, a browser-based computer-using agent that combines re

agentsarxiv-cs-ai
7 May 2026
Model Releases

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents

DGX agent

arXiv:2604.01496v2 Announce Type: replace-cross Abstract: We introduce SWE-ZERO to SWE-HERO, a two-stage SFT recipe that achieves state-of-the-art results on SWE-bench by distilling open-weight fronti

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents

DGX agent

arXiv:2509.24943v2 Announce Type: replace Abstract: Long videos, characterized by temporal complexity and sparse task-relevant information, pose significant reasoning challenges for AI systems. Althou

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

AcademiClaw: When Students Set Challenges for AI Agents

DGX agent

arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

AI Agents for Inventory Control: Human-LLM-OR Complementarity

DGX agent

arXiv:2602.12631v2 Announce Type: replace-cross Abstract: Inventory control is a fundamental operations problem in which ordering decisions are traditionally guided by theoretically grounded operation

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents

DGX agent

arXiv:2605.01250v1 Announce Type: new Abstract: Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical

model-releasesarxiv-cs-ai
6 May 2026
Safety

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

DGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

safetyarxiv-cs-ai
6 May 2026
Model Releases

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

DGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

model-releasesarxiv-cs-ai
6 May 2026
Safety

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

DGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

safetyarxiv-cs-lg
6 May 2026
Safety

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

DGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

safetyarxiv-cs-lg
5 May 2026
Safety

Causal Foundations of Collective Agency

DGX agent

arXiv:2605.00248v1 Announce Type: new Abstract: A key challenge for the safety of advanced AI systems is the possibility that multiple simpler agents might inadvertently form a collective agent with c

safetyarxiv-cs-ai
5 May 2026
Model Releases

Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization

DGX agent

arXiv:2605.02011v1 Announce Type: new Abstract: Automating the drafting of judgment documents is pivotal to judicial efficiency, yet it remains challenging due to the dual requirements of comprehensiv

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Training-Free Time Series Classification via In-Context Reasoning with LLM Agents

DGX agent

arXiv:2510.05950v2 Announce Type: replace Abstract: Time series classification (TSC) spans diverse application scenarios, yet labeled data are often scarce, making task-specific training costly and in

model-releasesarxiv-cs-ai
5 May 2026
Safety

Agent-Agnostic Evaluation of SQL Accuracy in Production Text-to-SQL Systems

DGX agent

arXiv:2604.28049v1 Announce Type: new Abstract: Text-to-SQL (T2SQL) evaluation in production environments poses fundamental challenges that existing benchmarks do not address. Current evaluation metho

safetyarxiv-cs-ai
1 May 2026
Model Releases

Agentic Education: Using Claude Code to Teach Claude Code

DGX agent

arXiv:2604.17460v2 Announce Type: replace-cross Abstract: AI coding assistants have proliferated rapidly, yet structured pedagogical frameworks for learning these tools remain scarce. Developers face

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Can Large Language Models Implement Agent-Based Models? An ODD-based Replication Study

DGX agent

arXiv:2602.10140v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now synthesize non-trivial executable code from textual descriptions, raising an important question: can LLMs

model-releasesarxiv-cs-ai
1 May 2026
Local Ai

Echo-{alpha}: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation

DGX agent

arXiv:2604.28011v1 Announce Type: new Abstract: Ultrasound interpretation requires both precise lesion localization and holistic clinical reasoning, yet existing methods typically excel at only one of

local-aiarxiv-cs-cv
1 May 2026
Safety

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

DGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

safetyarxiv-cs-ai
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft

DGX agent

arXiv:2604.24697v1 Announce Type: new Abstract: Discovering causal regularities and applying them to build functional systems--the discovery-to-application loop--is a hallmark of general intelligence,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

IntrAgent: An LLM Agent for Content-Grounded Information Retrieval through Literature Review

DGX agent

arXiv:2604.22861v1 Announce Type: cross Abstract: Scientific research relies on accurate information retrieval from literature to support analytical decisions. In this work, we introduce a new task, I

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning

DGX agent

arXiv:2604.22770v1 Announce Type: cross Abstract: Most digital language learning curricula rely on discrete-item quizzes that test recall rather than applied conversational proficiency. When progressi

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Measuring Successful Cooperation in Human-AI Teamwork: Development and Validation of the Perceived Cooperativity and Teaming Perception Scales

DGX agent

arXiv:2604.24461v1 Announce Type: cross Abstract: As human-AI cooperation becomes increasingly prevalent, reliable instruments for assessing the subjective quality of cooperative human-AI interaction

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Zero-to-CAD: Agentic Synthesis of Interpretable CAD Programs at Million-Scale Without Real Data

DGX agent

arXiv:2604.24479v1 Announce Type: new Abstract: Computer-Aided Design (CAD) models are defined by their construction history: a parametric recipe that encodes design intent. However, existing large-sc

model-releasesarxiv-cs-cv
28 Apr 2026
← Previous
1…108109110111112…236
Next →