AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

PilotBench: A Benchmark for General Aviation Agents with Safety Constraints

DGX agent

arXiv:2604.08987v1 Announce Type: new Abstract: As Large Language Models (LLMs) advance toward embodied AI agents operating in physical environments, a fundamental question emerges: can models trained

model-releasesarxiv-cs-ai
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Process Reward Agents for Steering Knowledge-Intensive Reasoning

DGX agent

arXiv:2604.09482v1 Announce Type: new Abstract: Reasoning in knowledge-intensive domains remains challenging as intermediate steps are often not locally verifiable: unlike math or code, evaluating ste

local-aiarxiv-cs-ai
13 Apr 2026
Model Releases

On Emotion-Sensitive Decision Making of Small Language Model Agents

DGX agent

arXiv:2604.06562v1 Announce Type: new Abstract: Small language models (SLM) are increasingly used as interactive decision-making agents, yet most decision-oriented evaluations ignore emotion as a caus

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

Auditable agentic AI for evidence-grounded thyroid ultrasound diagnosis and reporting

DGX agent

arXiv:2608.12590v1 Announce Type: new Abstract: Thyroid ultrasound diagnosis requires coordinated lesion localization, measurement, risk stratification and reporting, yet most AI systems address these

local-aiarxiv-cs-ai
14 Aug 2026
Safety

Static analysis-guided agentic AI translation enables Rust as a full stack bioinformatics language

DGX agent

arXiv:2608.13029v1 Announce Type: cross Abstract: The field of bioinformatics struggles with legacy code - old code that is commonly used but may no longer have a maintainer, or may be written in an n

safetyarxiv-cs-ai
14 Aug 2026
Model Releases

Advancing MLLM-based UAV Image Understanding and Reasoning: A Benchmark and a Training-Free Multi-Agent System

DGX agent

arXiv:2608.11738v1 Announce Type: cross Abstract: Multimodal Large Language Model (MLLM)-based UAV aerial image understanding and reasoning is essential for aerial intelligence yet poses distinct chal

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

AgenticTwin: An Agentic LLM Framework Integrated with Digital Twin for Anomaly Detection

DGX agent

arXiv:2608.11679v1 Announce Type: new Abstract: Digital twins are increasingly used to monitor and simulate the behavior of cyber-physical systems. Even with skilled operators, interpreting anomalies

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Beyond Trial-and-Error: Agentic Optimization for Image-to-Video Adherence

DGX agent

arXiv:2608.12290v1 Announce Type: cross Abstract: Modern black-box Image-to-Video (I2V) models offer powerful capabilities in automated content creation, yet their lack of fine-grained control and rel

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

Glance, Scrutinize, and Think: Advancing Video Anomaly Detection from Training-Free to Agentic Reasoning

DGX agent

arXiv:2608.11260v1 Announce Type: new Abstract: Video Anomaly Detection (VAD) aims to identify anomalous events and localize their temporal intervals. Existing approaches exhibit a 'when-what' dissoci

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem

DGX agent

arXiv:2608.10760v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become the de-facto interface for connecting LLM agents to enterprise tools, and adoption has been explosive: wit

agentsarxiv-cs-ai
12 Aug 2026
Safety

Detecting an Effect Is Not Learning to Act on It: A Reward-SNR Floor for LLM Acquisition Agents

DGX agent

arXiv:2608.10441v1 Announce Type: cross Abstract: Many pipelines can pay a per-example cost to acquire an auxiliary, model-derived observation -- an LLM's structured reasoning, a slow oracle, an expen

safetyarxiv-cs-cl
12 Aug 2026
Agents

Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization

DGX agent

arXiv:2608.10694v1 Announce Type: cross Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answerin

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

Agentic AI for Clustering, Relationship Discovery, and Semantic Trading in Prediction Markets

DGX agent

arXiv:2512.02436v2 Announce Type: replace Abstract: Prediction markets allow users to trade on outcomes of real-world events, but are prone to fragmentation with overlapping questions, implicit equiva

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ATLAS: Agentic Taxonomy of Large-Scale Software Ecosystems

DGX agent

arXiv:2606.21597v2 Announce Type: replace-cross Abstract: The open-source ecosystem on GitHub lacks a systematic hierarchical taxonomy of software repositories. GitHub Topics, the dominant organizatio

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

DGX agent

arXiv:2608.08160v1 Announce Type: cross Abstract: The rapid advancement of Large Language Models (LLMs) is revolutionizing AI for Games by enabling open-ended and fluid interactive storytelling. Howev

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Diagnosing as Cardiologists Do: ECG Agents with Doctor-Grounded Priors for Clinical Reasoning Across Diseases and Populations

DGX agent

arXiv:2608.09053v1 Announce Type: cross Abstract: Cardiologists interpret electrocardiograms by localizing waveform components, measuring rhythm and interval patterns, and translating these structured

model-releasesarxiv-cs-cv
11 Aug 2026
Agents

Query-Only Backdoor Attacks on Self-Evolving Skills via Trajectory Poisoning

DGX agent

arXiv:2608.08303v1 Announce Type: new Abstract: Agentic skills improve large language model (LLM) agents by encoding reusable procedures for complex tasks. However, manually authored skills often adap

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation

DGX agent

arXiv:2503.22122v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly for long-horizon tasks that require

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Politician, the Liar, and the Obedient Worker: Emerging Behavior of LLM Agents in Hierarchical Games

DGX agent

arXiv:2608.09574v1 Announce Type: new Abstract: LLMs are rapidly embedding themselves into daily life: drafting our emails, managing our schedules, and making decisions on our behalf. As they move fro

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems

DGX agent

arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-driven interactions that produce emergent

safetyarxiv-cs-ai
10 Aug 2026
Safety

Shape Your Feed: An LLM-based Agentic System for Conversational Recommendation

DGX agent

arXiv:2608.06632v1 Announce Type: new Abstract: Industrial recommendation systems predominantly adopt a passive ranking paradigm that infers user preferences from implicit behavioral signals (e.g., cl

safetyarxiv-cs-ai
10 Aug 2026
Safety

AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning

DGX agent

arXiv:2608.05987v1 Announce Type: new Abstract: Reinforcement learning (RL) with verifiable rewards constructs trajectory-level advantage estimates, yet it often fails to credit the few pivotal decisi

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Evaluating Investment Logic in Large Language Models: A Real-World Benchmark Towards Personalzied Financial Agents

DGX agent

arXiv:2608.06108v1 Announce Type: new Abstract: Investment competence is inherently personalized: the same market evidence can justify different actions for investors with different goals, horizons, p

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

A-SR: Self-Evolving Agentic LLMs for Symbolic Regression via Hierarchical Coordination

DGX agent

arXiv:2608.04872v1 Announce Type: cross Abstract: Symbolic regression aims to discover closed-form equations from data, but existing LLM-guided methods often rely on a unified proposal loop that compr

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation

DGX agent

arXiv:2511.07322v3 Announce Type: replace-cross Abstract: While LLMs have shown great success in financial tasks like stock prediction and question answering, their application in fully automating Equ

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

PICopilot: An LLM-based Agentic Framework for Assisting Photonic Integrated Circuit Design via Script Generation

DGX agent

arXiv:2608.01791v2 Announce Type: replace-cross Abstract: The rapid development of photonic integrated circuits (PICs) is shifting the design flow from traditional graphical user interface (GUI)-based

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

SimMOF: AI agent for Automated MOF Simulations

DGX agent

arXiv:2603.29152v2 Announce Type: replace Abstract: Metal-organic frameworks (MOFs) offer a vast design space, and as such, computational simulations play a critical role in predicting their structura

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Beyond the Single Camera: Agentic Multi-View Reasoning in Sports Video Understanding

DGX agent

arXiv:2607.11844v2 Announce Type: replace Abstract: Recent Multimodal Large Language Models (MLLMs) achieve strong performance on single-view video understanding benchmarks. However, sports videos inv

model-releasesarxiv-cs-cv
5 Aug 2026
Agents

Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details

DGX agent

arXiv:2608.03644v1 Announce Type: new Abstract: AI agents deployed in real-world settings must be capable of coordinating with humans and other AI agents they have not encountered before. Zero-shot co

agentsarxiv-cs-ai
5 Aug 2026
Safety

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

DGX agent

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

safetyarxiv-cs-ai
5 Aug 2026
Model Releases

MemArena: An Ego-Centric Benchmark for On-Device Agentic Personal Memory Assistants at Scale

DGX agent

arXiv:2608.02613v1 Announce Type: cross Abstract: Edge-deployed personal memory assistants must handle private interpersonal conversations on-device with open-weight models. Yet, existing memory bench

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

PhysAgent: A Multi-Agent Framework for Reliable Remote Heart Rate Estimation

DGX agent

arXiv:2608.00066v1 Announce Type: new Abstract: Remote photoplethysmography (rPPG) enables non-contact heart-rate estimation from facial videos, but its weak physiological signal is easily corrupted b

model-releasesarxiv-cs-cv
4 Aug 2026
Agents

AuEmoChat: Authentic Emotion Understanding and Rendering for Conversational Speech Synthesis

DGX agent

arXiv:2607.15755v2 Announce Type: replace-cross Abstract: Conversational Speech Synthesis (CSS) aims to synthesize speech with human-like emotional expression and contextual consistency in user-agent

agentsarxiv-cs-ai
3 Aug 2026
Agents

SERUM: State Extraction and Refinement for User Modeling

DGX agent

arXiv:2607.29181v1 Announce Type: cross Abstract: Agentic assistants capable of proactive, personalized interactions require structured models of user intent and workflow. However, building these mode

agentsarxiv-cs-ai
3 Aug 2026
Agents

Shaping Scientific Explanations to Expert Perspectives with Persona-Conditioned Reinforcement Learning

DGX agent

arXiv:2603.21846v2 Announce Type: replace Abstract: Explainable AI is increasingly important to scientific discovery. However, existing methods largely ignore that explanation quality is not universal

agentsarxiv-cs-ai
3 Aug 2026
Agents

The Formalism Trap: Are LLM-as-a-Judge Evaluators Blinded by Consensus Mimicry under Social Load?

DGX agent

arXiv:2607.28641v1 Announce Type: cross Abstract: We introduce the extit{Agentic Formalism Trap} and the Evaluative Dissonance Index (D_E), quantifying how LLM-as-a-Judge systems conflate structural p

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

A Physics-Informed Framework for PID Tuning of Chemical Processes Using Large Language Model Agents

DGX agent

arXiv:2607.26594v1 Announce Type: cross Abstract: PID tuning for chemical processes commonly relies on identified process models, whereas plant engineers often retune loops iteratively by observing re

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

Fantastic Adaptive Taxonomies and How to Use Them

DGX agent

arXiv:2607.16387v2 Announce Type: replace-cross Abstract: An agent system's execution traces record how it fails, and procedures that improve such a system without changing model weights (trajectory s

model-releasesarxiv-cs-ai
31 Jul 2026
Agents

Pramana: A Composable, Domain-Specific Backend for Empirical Networking Research

DGX agent

arXiv:2607.26352v1 Announce Type: cross Abstract: Networking research advances by turning hypotheses into empirical evidence, so accelerating it means reducing the lag between ideation (synthesizing a

agentsarxiv-cs-ai
31 Jul 2026
Safety

Learning Implicit Causal World Models from Multi-Agent Demonstrations

DGX agent

arXiv:2607.26336v1 Announce Type: new Abstract: In model-based reinforcement learning, world models exist as internal simulators, but their training often conflates statistical correlations with causa

safetyarxiv-cs-lg
30 Jul 2026
Local Ai

MARS: Multi-Agent Re-ranking for Repeat-Order Food Delivery Recommendation

DGX agent

arXiv:2607.25420v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in recommender systems, but it is often unclear how much performance can be obtained from strong pr

local-aiarxiv-cs-ai
29 Jul 2026
Agents

Coherent Without Grounding, Grounded Without Success: Observability and Epistemic Failure

DGX agent

arXiv:2603.28371v2 Announce Type: replace-cross Abstract: When an agent can articulate why something works, we typically take this as evidence of genuine understanding. This presupposes that effective

agentsarxiv-cs-ai
28 Jul 2026
Agents

WCM: World-Cognition Model for Generalizable Human-Robot Interaction

DGX agent

arXiv:2607.22999v1 Announce Type: cross Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable interaction to physical tasks. Current

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Encoding Invisible Causation for Bridge Diagnostic Agents: Triple-Guided Retrieval-Augmented Fine-Tuning with QLoRA

DGX agent

arXiv:2607.21680v1 Announce Type: new Abstract: Bridge infrastructure deteriorates gradually, yet its root causes---salt intrusion, freezing, fatigue cracking, and others---remain invisible to the nak

model-releasesarxiv-cs-lg
27 Jul 2026
Safety

TRACE-ROUTER: Task-Consistent and Adaptive Online Routing for Agentic AI

DGX agent

arXiv:2607.22465v1 Announce Type: cross Abstract: Routing to select large language models (LLMs) with different cost-quality trade-offs has become a fundamental deployment feature of enterprise AI. Ex

safetyarxiv-cs-lg
27 Jul 2026
Local Ai

Conflict Resolution under Degraded Surveillance in Air Corridors Using Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.20547v1 Announce Type: new Abstract: Safe Advanced Air Mobility operations require aircraft to maintain separation when surveillance information is noisy, delayed, incomplete, or temporaril

local-aiarxiv-cs-lg
24 Jul 2026
Safety

SAGE: A Socially-Aware Generative Engine for Heterogeneous Multi-Agent Navigation

DGX agent

arXiv:2607.16619v2 Announce Type: replace Abstract: Safe and socially compliant navigation in open human-robot environments requires robots to reason about heterogeneous participants with different dy

safetyarxiv-cs-ro
24 Jul 2026
Model Releases

Verifier-First Evaluation of Agentic LLMs for Infrastructure-as-Code Generation

DGX agent

arXiv:2607.20478v1 Announce Type: cross Abstract: Infrastructure-as-Code (IaC) generation from natural language requires satisfying provider schemas, dependency planning, and organizational policy con

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…104105106107108…236
Next →