AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,942 results
Model Releases

(How) Do Large Language Models Understand High-Level Message Sequence Charts?

DGX agent

arXiv:2605.13773v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being employed widely to automate tasks across the software development life-cycle. It is, however, unclear whether t

model-releasesarxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Is a Picture Worth a Thousand Words? Adaptive Multimodal Fact-Checking with Visual Evidence Necessity

DGX agent

arXiv:2604.04692v2 Announce Type: replace-cross Abstract: Automated fact-checking is a crucial task that supports a responsible information ecosystem. While recent research has progressed from text-on

researcharxiv-cs-ai
14 May 2026
Research

JANUS: Anatomy-Conditioned Gating for Robust CT Triage Under Distribution Shift

DGX agent

arXiv:2605.13813v1 Announce Type: new Abstract: Automated CT triage requires models that are simultaneously accurate across diverse pathologies and reliable under institutional shift. While Vision Tra

researcharxiv-cs-cv
14 May 2026
Model Releases

LLMs as annotators of credibility assessment in Danish asylum decisions: evaluating classification performance and errors beyond aggregated metrics

DGX agent

arXiv:2605.13412v1 Announce Type: cross Abstract: Off-the-shelf large language models (LLMs) are increasingly used to automate text annotation, yet their effectiveness remains underexplored for underr

model-releasesarxiv-cs-ai
14 May 2026
Safety

MoCCA: A Movable Circle Probability of Collision Approximation

DGX agent

arXiv:2605.13125v1 Announce Type: new Abstract: In automated driving, crash mitigation is crucial to ensure passenger safety. Accurate avoidance requires precise knowledge of the object's position and

safetyarxiv-cs-ro
14 May 2026
Model Releases

Pattern-Enhanced RT-DETR for Multi-Class Battery Detection

DGX agent

arXiv:2605.13670v1 Announce Type: new Abstract: Accurate and efficient battery detection is increasingly important for applications in electronic waste recycling, industrial quality control, and autom

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

PreFIQs: Face Image Quality Is What Survives Pruning

DGX agent

arXiv:2605.13396v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) evaluates the utility of a face image for automated face recognition (FR) systems. In this work, we propose PreFIQs

model-releasesarxiv-cs-cv
14 May 2026
Research

PRISM: Perinuclear Ring-based Image Segmentation Method for Acute Lymphoblastic Leukemia Classification

DGX agent

arXiv:2605.12851v1 Announce Type: cross Abstract: Automated analysis of peripheral blood smears for Acute Lymphoblastic Leukemia (ALL) is hindered by low contrast and substantial variability in cytopl

researcharxiv-cs-ai
14 May 2026
Research

AOI-SSL: Self-Supervised Framework for Efficient Segmentation of Wire-bonded Semiconductors In Optical Inspection

DGX agent

arXiv:2605.12430v1 Announce Type: new Abstract: Segmentation models in automated optical inspection of wire-bonded semiconductors are typically device-specific and must be re-trained when new devices

researcharxiv-cs-cv
13 May 2026
Model Releases

ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder

DGX agent

arXiv:2605.11091v1 Announce Type: new Abstract: Automated ASD screening tools remain limited by single-architecture evaluations, axis-restricted assessment, and near-exclusive focus on adult cohorts,

model-releasesarxiv-cs-lg
13 May 2026
Safety

Causal Bias Detection in Generative Artifical Intelligence

DGX agent

arXiv:2605.11365v1 Announce Type: cross Abstract: Automated systems built on artificial intelligence (AI) are increasingly deployed across high-stakes domains, raising critical concerns about fairness

safetyarxiv-cs-lg
13 May 2026
Model Releases

Covering Human Action Space for Computer Use: Data Synthesis and Benchmark

DGX agent

arXiv:2605.12501v1 Announce Type: new Abstract: Computer-use agents (CUAs) automate on-screen work, as illustrated by GPT-5.4 and Claude. Yet their reliability on complex, low-frequency interactions i

model-releasesarxiv-cs-cv
13 May 2026
Safety

Epistemic Uncertainty for Test-Time Discovery

DGX agent

arXiv:2605.11328v1 Announce Type: new Abstract: Automated scientific discovery using large language models relies on identifying genuinely novel solutions. Standard reinforcement learning penalizes hi

safetyarxiv-cs-lg
13 May 2026
Research

GATA2Floor: Graph attention for floor counting in street-view facades

DGX agent

arXiv:2605.11863v1 Announce Type: new Abstract: Automated analysis of building facades from street-level imagery has great potential for urban analytics, energy assessment, and emergency planning. How

researcharxiv-cs-cv
13 May 2026
Model Releases

Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation

DGX agent

arXiv:2605.11208v1 Announce Type: new Abstract: Automated, clinician-grade assessment reports for surgical procedures could reduce documentation burden and provide objective feedback, yet remain chall

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function

DGX agent

arXiv:2602.02406v2 Announce Type: replace-cross Abstract: Data-driven algorithm design automates hyperparameter tuning, but its statistical foundations remain limited because model performance can dep

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

VERDI: Single-Call Confidence Estimation for Verification-Based LLM Judges via Decomposed Inference

DGX agent

arXiv:2605.11334v1 Announce Type: cross Abstract: LLM-as-Judge systems are widely deployed for automated evaluation, yet practitioners lack reliable methods to know when a judge's verdict should be tr

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Arcane: An Assertion Reduction Framework through Semantic Clustering and MCTS-Guided Rule Exploring

DGX agent

arXiv:2605.10107v1 Announce Type: new Abstract: Assertion-based Verification (ABV) is essential for ensuring that hardware designs conform to their intended specifications. However, existing automated

model-releasesarxiv-cs-ai
12 May 2026
Agents

Bridging the Cognitive Gap: A Unified Memory Paradigm for 6G Agentic AI-RAN

DGX agent

arXiv:2605.10036v1 Announce Type: cross Abstract: As 6G evolves, the radio access network must transcend traditional automation to embrace agentic AI capable of perception, reasoning, and evolution. A

agentsarxiv-cs-ai
12 May 2026
Model Releases

Can Deep Research Agents Retrieve and Organize? Evaluating the Synthesis Gap with Expert Taxonomies

DGX agent

arXiv:2601.12369v3 Announce Type: replace Abstract: Deep Research Agents increasingly automate survey generation, yet whether they match human experts at retrieving essential papers and organizing the

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

ComplexMCP: Evaluation of LLM Agents in Dynamic, Interdependent, and Large-Scale Tool Sandbox

DGX agent

arXiv:2605.10787v1 Announce Type: new Abstract: Current LLM agents are proficient at calling isolated APIs but struggle with the 'last mile' of commercial software automation. In real-world scenarios,

model-releasesarxiv-cs-ai
12 May 2026
Research

Cross-Modal Semantic-Enhanced Diffusion Framework for Diabetic Retinopathy Grading

DGX agent

arXiv:2605.09242v1 Announce Type: cross Abstract: Automated grading of diabetic retinopathy (DR) faces several critical challenges: subtle inter-grade visual distinctions in fine-grained lesion patter

researcharxiv-cs-cv
12 May 2026
Safety

Crowding Out The Noise: Algorithmic Collective Action Under Differential Privacy

DGX agent

arXiv:2505.05707v2 Announce Type: replace Abstract: The integration of AI into daily life has generated considerable attention and excitement, while also raising concerns about automating algorithmic

safetyarxiv-cs-lg
12 May 2026
Model Releases

Explicit Reasoning Makes Better Judges: A Systematic Study on Accuracy, Efficiency, and Robustness

DGX agent

arXiv:2509.13332v2 Announce Type: replace Abstract: As Large Language Models (LLMs) are increasingly adopted as automated judges in benchmarking and reward modeling, ensuring their reliability, effici

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Geometrically Constrained Stenosis Editing in Coronary Angiography via Entropic Optimal Transport

DGX agent

arXiv:2605.08851v1 Announce Type: cross Abstract: The scarcity of high-quality imaging data for coronary angiography (CAG) stenosis limits the clinical translation of automated stenosis detection. Syn

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Hierarchical Reinforced Trader (HRT): A Bi-Level Approach for Optimizing Stock Selection and Execution

DGX agent

arXiv:2410.14927v2 Announce Type: replace-cross Abstract: Automated equity trading requires converting noisy market and news signals into executable portfolio decisions under risk, turnover, and trans

model-releasesarxiv-cs-lg
12 May 2026
Local Ai

iPay: Integrated Payment Action Recognition via Multimodal Networks and Adaptive Spatial Prior Learning

DGX agent

arXiv:2605.10732v1 Announce Type: cross Abstract: Automated transit payment analysis is vital for scalable fare auditing and passenger analytics, yet practice still relies on limited manual inspection

local-aiarxiv-cs-ai
12 May 2026
Model Releases

Metis: Learning to Jailbreak LLMs via Self-Evolving Metacognitive Policy Optimization

DGX agent

arXiv:2605.10067v1 Announce Type: cross Abstract: Red teaming is critical for uncovering vulnerabilities in Large Language Models (LLMs). While automated methods have improved scalability, existing ap

model-releasesarxiv-cs-ai
12 May 2026
Safety

Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs

DGX agent

arXiv:2605.10877v1 Announce Type: new Abstract: Automated question answering (QA) over electronic health records (EHRs) demands precise evidence retrieval, faithful answer generation, and explicit gro

safetyarxiv-cs-cl
12 May 2026
Model Releases

Optimized Culprit Identification Using Mobilenet and Attention Mechanisms

DGX agent

arXiv:2605.08169v1 Announce Type: cross Abstract: Automated culprit identification in surveillance systems is a critical task that requires high accuracy along with computational efficiency for real-t

model-releasesarxiv-cs-ai
12 May 2026
Safety

PhysEDA: Physics-Aware Learning Framework for Efficient EDA With Manhattan Distance Decay

DGX agent

arXiv:2605.10547v1 Announce Type: new Abstract: Electronic design automation (EDA) addresses placement, routing, timing analysis, and power-integrity verification for integrated circuits. Learning met

safetyarxiv-cs-lg
12 May 2026
Safety

Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge

DGX agent

arXiv:2605.10805v1 Announce Type: new Abstract: Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings r

safetyarxiv-cs-ai
12 May 2026
Research

Robust Building Damage Detection in Cross-Disaster Settings Using Domain Adaptation

DGX agent

arXiv:2603.14694v2 Announce Type: replace-cross Abstract: Rapid structural damage assessment from remote sensing imagery is essential for timely disaster response. Within human-machine systems (HMS) f

researcharxiv-cs-ai
12 May 2026
Research

ShifaMind: A Multiplicative Concept Bottleneck for Interpretable ICD-10 Coding

DGX agent

arXiv:2605.08482v1 Announce Type: cross Abstract: Automated ICD-10 coding from clinical discharge summaries requires models that are both accurate on long-tailed multi-label classification tasks and i

researcharxiv-cs-cl
12 May 2026
Research

Spatial Priming Outperforms Semantic Prompting: A Grid-Based Approach to Improving LLM Accuracy on Chart Data Extraction

DGX agent

arXiv:2605.08220v1 Announce Type: new Abstract: The automated extraction of data from scientific charts is a critical task for large-scale literature analysis. While multimodal Large Language Models (

researcharxiv-cs-ai
12 May 2026
Model Releases

SpectraLLM: Uncovering the Ability of LLMs for Molecular Structure Elucidation from Multi-Spectral Data

DGX agent

arXiv:2508.08441v3 Announce Type: replace-cross Abstract: Automated molecular structure elucidation remains challenging, as existing approaches often depend on pre-compiled databases or restrict thems

model-releasesarxiv-cs-lg
12 May 2026
Local Ai

UPA: Unsupervised Prompt Agent via Tree-Based Search and Selection

DGX agent

arXiv:2601.23273v2 Announce Type: replace Abstract: Prompt agents have recently emerged as a promising paradigm for automated prompt optimization, framing prompt discovery as a sequential decision-mak

local-aiarxiv-cs-cl
12 May 2026
Tutorials

VulTriage: Triple-Path Context Augmentation for LLM-Based Vulnerability Detection

DGX agent

arXiv:2605.09461v1 Announce Type: new Abstract: Automated vulnerability detection is a fundamental task in software security, yet existing learning-based methods still struggle to capture the structur

tutorialsarxiv-cs-ai
12 May 2026
Agents

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management

DGX agent

arXiv:2512.10371v2 Announce Type: replace Abstract: The rapid development of mobile GUI agents has stimulated growing research interest in long-horizon task automation. However, building agents for th

agentsarxiv-cs-ai
11 May 2026
Model Releases

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

DGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing

DGX agent

arXiv:2605.06936v1 Announce Type: cross Abstract: LLM-based agents are increasingly applied to the 'last mile' of Electronic Design Automation (EDA): repairing residual sign-off Design Rule Check (DRC

model-releasesarxiv-cs-ai
11 May 2026
Agents

Conformal Agent Error Attribution

DGX agent

arXiv:2605.06788v1 Announce Type: new Abstract: When multi-agent systems (MAS) fail, identifying where the decisive error occurred is the first step for automated recovery to an earlier state. Error a

agentsarxiv-cs-lg
11 May 2026
Local Ai

HMACE: Heterogeneous Multi-Agent Collaborative Evolution for Combinatorial Optimization

DGX agent

arXiv:2605.07214v1 Announce Type: new Abstract: Large Language Models have recently emerged as a promising paradigm for automated heuristic design for NP-hard combinatorial optimization problems. Desp

local-aiarxiv-cs-ai
11 May 2026
Research

Inference of Qualitative Models from Steady-State Data via Weighted MaxSMT

DGX agent

arXiv:2605.07433v1 Announce Type: cross Abstract: Qualitative models provide crucial instruments for modelling complex biological systems. While advances in automated reasoning and symbolic encodings

researcharxiv-cs-lg
11 May 2026
Model Releases

Is Your Prompt Poisoning Code? Defect Induction Rates and Security Mitigation Strategies

DGX agent

arXiv:2510.22944v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have become indispensable for automated code generation, yet the quality and security of their outputs remain a c

model-releasesarxiv-cs-ai
11 May 2026
Safety

Many-to-Many Multi-Agent Pickup and Delivery

DGX agent

arXiv:2605.07835v1 Announce Type: new Abstract: Multi-robot systems in automated warehouses must manage continuous streams of pickup-and-delivery tasks while ensuring efficiency and safety. Prior work

safetyarxiv-cs-ro
11 May 2026
Research

Minerva: Reinforcement Learning with Verifiable Rewards for Cyber Threat Intelligence LLMs

DGX agent

arXiv:2602.00513v3 Announce Type: replace Abstract: Cyber threat intelligence (CTI) analysts routinely convert noisy, unstructured security artifacts into standardized, automation-ready representation

researcharxiv-cs-lg
11 May 2026
Research

Multi-Dimensional Evaluation of LLMs for Grammatical Error Correction

DGX agent

arXiv:2605.07635v1 Announce Type: new Abstract: Automated assistants for Grammatical Error Correction are now embedded in educational platforms serving millions of learners, yet three critical gaps re

researcharxiv-cs-cl
11 May 2026
← Previous
1…2728293031…83
Next →