AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

LessonBench-V1: A Benchmark Dataset for Evaluating AI Lesson Generation Agents

DGX agent

arXiv:2607.13041v1 Announce Type: cross Abstract: Large Language Model (LLM) based AI educational content generation systems are increasingly being developed, yet no standardised benchmark exists to s

model-releasesarxiv-cs-ai
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

PhysClaw-0: A Symbiotic Agentic System for Robot Autonomy via Language Corrections

DGX agent

arXiv:2607.14047v1 Announce Type: new Abstract: Autonomous data collection governs the volume and quality of real-world trajectories for manipulation policy learning. Existing pipelines reduce human e

safetyarxiv-cs-ro
16 Jul 2026
Agents

Survival Dynamics of Neural and Programmatic Policies in Evolutionary Reinforcement Learning

DGX agent

arXiv:2601.04365v2 Announce Type: replace Abstract: In evolutionary reinforcement learning tasks (ERL), agent policies are often encoded as small artificial neural networks (NERL). Such representation

agentsarxiv-cs-lg
16 Jul 2026
Local Ai

A Multi-Agent System for Autonomous, Fine-Tuning-Free Clinical Symptom Detection: Development and Validation Study

DGX agent

arXiv:2607.12886v1 Announce Type: new Abstract: Clinical notes contain many of the signs and symptoms that bring patients to care, yet this information rarely reaches structured fields. Existing extra

local-aiarxiv-cs-ai
15 Jul 2026
Agents

Amplitude-Only FFN Intervention for Tool-Structured LLM Inference Method: Gated Evaluation Protocol, and Cross-Model Empirical Results

DGX agent

arXiv:2607.11183v2 Announce Type: replace Abstract: Large language models increasingly operate as tool-using agents, where small format, argument, or function-call errors can invalidate otherwise plau

agentsarxiv-cs-cl
15 Jul 2026
Model Releases

Metric-Guided Synthetic Image Data Rendering for Deep Learning compatible with Agentic AI

DGX agent

arXiv:2607.12874v1 Announce Type: new Abstract: Deep learning computer vision for scientific applications requires collecting and annotating large datasets in a laborious, expensive and error-prone pr

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

DGX agent

arXiv:2607.13028v1 Announce Type: cross Abstract: Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground beh

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

Win by Silence: Deletion Non-Monotonicity, Autonomous Exploitation, and Typed-State Gating in LLM Plan Evaluation

DGX agent

arXiv:2607.12986v1 Announce Type: new Abstract: Plan evaluators can reward a strategic plan for becoming less explicit. This paper studies that failure in a staged expected-value scorer for LLM-genera

agentsarxiv-cs-ai
15 Jul 2026
Safety

From Prompts to Contracts: Harness Engineering for Auditable Enterprise LLM Agents

DGX agent

arXiv:2607.08028v1 Announce Type: new Abstract: Enterprise large language model (LLM) applications often begin as prototypes whose behavior is carried by prompts and retrieval context. Productization

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

ArtisanCAD: An Industrial-Level CAD Agent with Expert-Grounded Knowledge Distillation

DGX agent

arXiv:2607.05750v1 Announce Type: new Abstract: Computer-aided design (CAD) for industrial components requires long-horizon procedural modeling, robust feature dependencies, editable parametric geomet

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

KernelEvolve: Scaling Agentic Kernel Coding for Heterogeneous AI Accelerators at Meta

DGX agent

arXiv:2512.23236v4 Announce Type: replace-cross Abstract: Making deep learning recommendation model (DLRM) training and inference fast and efficient is important. However, this presents three key syst

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

Agentic Retrieval-Augmented Generation for Financial Document Question Answering

DGX agent

arXiv:2605.05409v2 Announce Type: replace Abstract: Financial document question answering (QA) demands complex multi-step numerical reasoning over heterogeneous evidence--structured tables, textual na

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Integrated Altruistic and Fairness Preference Induces Advanced Mutual Cooperation in Sequential Social Dilemmas

DGX agent

arXiv:2607.04710v1 Announce Type: new Abstract: Inducing cooperation among distributed agents is still a difficult problem in the field of multi-agent reinforcement learning (MARL), particularly in so

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

MedCalc-Pro: Solving Complex Medical Calculations with LLM Agents

DGX agent

arXiv:2607.02879v1 Announce Type: new Abstract: Current benchmarks for evaluating large language models (LLMs) in medical calculation are largely based on simplified settings, where each patient case

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

ParEVO: Synthesizing Code for Irregular Data: High-Performance Parallelism through Agentic Evolution

DGX agent

arXiv:2603.02510v2 Announce Type: replace Abstract: The transition from sequential to parallel computing is essential for modern high-performance applications but is hindered by the steep learning cur

model-releasesarxiv-cs-lg
7 Jul 2026
Safety

RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents

DGX agent

arXiv:2607.04713v1 Announce Type: cross Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

SafeGuard: A Multi-Agent Perception-Reasoning Framework for Social-Risk AI-Generated Video Detection

DGX agent

arXiv:2607.03069v1 Announce Type: new Abstract: As video generation paradigms evolve from localized manipulation to full-scene synthesis, AI-generated video detection becomes increasingly challenging,

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

Taming I2V models for Image HOI Editing: A Cognitive Benchmark and Agentic Self-Correcting Framework

DGX agent

arXiv:2606.19073v2 Announce Type: replace Abstract: Current image editing methods excel at static attributes but fail at complex Human-Object Interactions (HOI), a critical challenge unaddressed by ex

model-releasesarxiv-cs-cv
7 Jul 2026
Safety

Training Verifiably Robust Agents Using Set-Based Reinforcement Learning

DGX agent

arXiv:2408.09112v2 Announce Type: replace Abstract: Reinforcement learning policies parametrized by deep neural networks have achieved strong performance for continuous control, yet even small input p

safetyarxiv-cs-lg
7 Jul 2026
Safety

ASPIRE: Agentic /Skills Discovery for Robotics

DGX agent

arXiv:2607.00272v1 Announce Type: cross Abstract: Traditional robot programming is challenging: it requires orchestrating multimodal perception, managing physical contact dynamics, and handling divers

safetyarxiv-cs-ai
2 Jul 2026
Safety

Certified Speculative Execution for Untrusted AI Agents

DGX agent

arXiv:2606.31023v1 Announce Type: cross Abstract: Hard-constrained sequential decision systems have no certified way to spend the test-time compute of modern AI: executing the multi-step drafts of a l

safetyarxiv-cs-lg
1 Jul 2026
Safety

DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation

DGX agent

arXiv:2606.31537v1 Announce Type: new Abstract: Text-rich image generation is one of the most challenging settings in image generation, since models must simultaneously produce visually realistic imag

safetyarxiv-cs-cv
1 Jul 2026
Model Releases

HyPOLE: Hyperproperty-Guided Multi-Agent Reinforcement Learning under Partial Observation

DGX agent

arXiv:2606.30966v1 Announce Type: new Abstract: Formal specification is a powerful tool to guide the learning process and provides significant advantages over reward shaping: (1) mathematical rigor; (

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.12634v2 Announce Type: replace-cross Abstract: Long-horizon tool-use reinforcement learning learns from outcome verification, but trajectory-level advantages are broadcast over reasoning, A

safetyarxiv-cs-ai
1 Jul 2026
Safety

Paper2Rebuttal: A Multi-Agent Framework for Transparent Author Response Assistance

DGX agent

arXiv:2601.14171v2 Announce Type: replace Abstract: Writing effective rebuttals is a high-stakes task that demands more than linguistic fluency, as it requires precise alignment between reviewer inten

safetyarxiv-cs-ai
1 Jul 2026
Safety

ReGRPO: Reflection-Augmented Policy Optimization for Tool-Using Agents

DGX agent

arXiv:2606.31392v1 Announce Type: new Abstract: Tool-augmented vision-language models (VLMs) can solve multimodal, multi-step tasks by calling external tools, yet they remain fragile in practice. Exis

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

REMSA: Foundation Model Selection for Remote Sensing via a Constraint-Aware Agent

DGX agent

arXiv:2511.17442v3 Announce Type: replace-cross Abstract: Foundation Models (FMs) are increasingly integrated into remote sensing (RS) pipelines. These models include unimodal vision encoders and mult

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Think While You Map: Asynchronous Vision-Language Agents for Incremental 3D Scene Graphs

DGX agent

arXiv:2606.31471v1 Announce Type: new Abstract: Open-vocabulary 3D scene graph methods typically operate in two stages: first reconstruct, then enrich with vision-language models, leaving the graph un

researcharxiv-cs-cv
1 Jul 2026
Model Releases

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

DGX agent

arXiv:2606.31307v1 Announce Type: new Abstract: Large language models used in task-oriented dialogue often produce fluent but unsafe responses when backend database calls fail, return empty results, o

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

ADEPT: An Entropy-Driven Dual-Strategy Agent for Interactive Video Retrieval

DGX agent

arXiv:2606.28326v1 Announce Type: cross Abstract: This research aims to solve the challenge of video retrieval from massive datasets, caused by ambiguous user queries. Prevailing single-round retrieva

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Building Multi-Task Agentic LLMs via Two-Phase Distillation

DGX agent

arXiv:2606.30044v1 Announce Type: new Abstract: A key step toward artificial general intelligence is to train models that can perform multiple tasks. In this paper, we study how to build such models b

safetyarxiv-cs-lg
30 Jun 2026
Safety

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval

DGX agent

arXiv:2606.28365v1 Announce Type: cross Abstract: RAG ingestion pipelines frequently augment search corpus index with semantic enrichment indices (e.g., synthetic queries or summaries generated from c

safetyarxiv-cs-ai
30 Jun 2026
Safety

Carolina Guide: A Multi-Agent RAG System with Institutional Guardrails for Academic Policy Assistance

DGX agent

arXiv:2606.28360v1 Announce Type: cross Abstract: University students often struggle to navigate complex academic policies, leading to advising bottlenecks and delayed access to critical information.

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Dynamo: Dynamic Skill-Tool Evolution for Vision-Language Agents

DGX agent

arXiv:2606.30185v1 Announce Type: new Abstract: Improving vision-language models (VLMs) on visual reasoning typically requires retraining or hand-designed prompts and tools. We present Dynamo, a train

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Efficient Visual Pointing for Embodied AI:Agent-Driven Data Synthesis, Cross-Block Attention, and Iterative Correction

DGX agent

arXiv:2606.29850v1 Announce Type: new Abstract: Visual pointing maps a language instruction to pixel co ordinates, a core skill for embodied AI. We describe our PointArena 2026 solution, which achieve

model-releasesarxiv-cs-cv
30 Jun 2026
Safety

Latent Actions from Factorized Transition Effects under Agent Ambiguity

DGX agent

arXiv:2606.30544v1 Announce Type: new Abstract: Latent Action Models (LAMs) learn action-like proxies from observation transitions. However, in multi-object or distractor-rich scenes, these visual eff

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Multi-Agentic System Leveraging Open-Source LLMs to Mitigate Disinformation Threats

DGX agent

arXiv:2606.30259v1 Announce Type: new Abstract: In contemporary societies, the threat of disinformation has reached alarming levels, exacerbated by the proliferation of electronic communication, socia

model-releasesarxiv-cs-cl
30 Jun 2026
Safety

CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association

DGX agent

arXiv:2606.28179v1 Announce Type: cross Abstract: Identifying robust associations between cardiac imaging phenotypes and clinical diseases is fundamental to population-scale cardiovascular research an

safetyarxiv-cs-ai
29 Jun 2026
Model Releases

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration

DGX agent

arXiv:2606.28215v1 Announce Type: cross Abstract: Extracting dynamic 4D object interactions from massive, in-the-wild monocular videos offers a highly efficient data collection pathway for scaling Emb

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking

DGX agent

arXiv:2606.26205v1 Announce Type: new Abstract: Patients increasingly seek medication information online, yet safety knowledge for psychiatric drugs is split between regulatory adverse-event records,

safetyarxiv-cs-ai
26 Jun 2026
Local Ai

ProfileFoundry: A Synthetic Person-Object Substrate for Privacy, Memory, and Tool-Use Evaluation in LLM Agent

DGX agent

arXiv:2606.26403v1 Announce Type: new Abstract: Foundation-model research increasingly needs data about people: user state, personal histories, relationships, contact-like fields, documents, and longi

local-aiarxiv-cs-cl
26 Jun 2026
Model Releases

Colon-Bench: An Agentic Workflow for Scalable Dense Lesion Annotation in Full-Procedure Colonoscopy Videos

DGX agent

arXiv:2603.25645v2 Announce Type: replace-cross Abstract: Early screening via colonoscopy is critical for colon cancer prevention, yet developing robust AI systems for this domain is hindered by the l

model-releasesarxiv-cs-cv
25 Jun 2026
Local Ai

Explainable Control Framework (XCF) based on Fuzzy Model-Agnostic Explanation and LLM Agent-Supported Interface

DGX agent

arXiv:2606.25941v1 Announce Type: cross Abstract: Increasing demand for precise and reliable control in complex scenarios has led to the development of increasingly sophisticated controllers, includin

local-aiarxiv-cs-ai
25 Jun 2026
Model Releases

DataClaw0: Agentic Tailoring Multimodal Data from Raw Streams

DGX agent

arXiv:2606.21337v1 Announce Type: new Abstract: Massive unstructured multimodal streams suffer from high 'data entropy,' impeding both efficient human knowledge acquisition and high-quality AI post-tr

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

One Image is All You Need: Agentic One-Shot Image Generation via Text-Based World Models for Long-Tail Spatial Perception

DGX agent

arXiv:2606.20764v1 Announce Type: new Abstract: Reliable spatial decision automation, such as autonomous driving and maritime surveillance, critically depends on robust visual perception. However, rea

safetyarxiv-cs-cv
23 Jun 2026
Safety

A Survey of Reasoning and Agentic Systems in Time Series with Large Language Models

DGX agent

arXiv:2509.11575v3 Announce Type: replace Abstract: Time series reasoning treats time as a first-class axis and incorporates intermediate evidence directly into the answer. This survey defines the pro

safetyarxiv-cs-ai
11 Jun 2026
Agents

PRInTS: Reward Modeling for Long-Horizon Information Seeking

DGX agent

arXiv:2511.19314v2 Announce Type: replace Abstract: Information-seeking is a core capability for AI agents, requiring them to gather and reason over tool-generated information across long trajectories

agentsarxiv-cs-ai
11 Jun 2026
Agents

The Impossibility of Eliciting Latent Knowledge

DGX agent

arXiv:2606.12268v1 Announce Type: new Abstract: Advanced AI systems have extensive knowledge of their environments; in fact, their knowledge may (far) exceed that of their developers or users. Consequ

agentsarxiv-cs-ai
11 Jun 2026
← Previous
1…111112113114115…236
Next →