AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Local Ai

Efficient Asynchronous Federated Evaluation with Strategy Similarity Awareness for Intent-Based Networking in Industrial Internet of Things

DGX agent

arXiv:2512.20627v2 Announce Type: replace-cross Abstract: Intent-Based Networking (IBN) offers a promising paradigm for intelligent and automated network control in Industrial Internet of Things (IIoT

local-aiarxiv-cs-ai
6 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Learning Adaptive Parallel Execution for Efficient Code Localization

DGX agent

arXiv:2601.19568v2 Announce Type: replace Abstract: Code localization constitutes a key bottleneck in automated software development pipelines. While concurrent tool execution can enhance discovery sp

researcharxiv-cs-ai
6 Jun 2026
Research

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation

DGX agent

arXiv:2601.12983v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are increasingly used to automate chart generation from data tables, improving analysis and reporting effic

researcharxiv-cs-cl
5 Jun 2026
Research

Deep Learning-assisted AMD Staging based on OCT and OCT Angiography

DGX agent

arXiv:2606.05379v1 Announce Type: new Abstract: To develop and evaluate deep learning models for automated grading of age-related macular degeneration (AMD) severity using optical coherence tomography

researcharxiv-cs-cv
5 Jun 2026
Research

From Scoring to Explanations: Evaluating SHAP and LLM Rationales for Rubric-based Teaching Quality Assessment

DGX agent

arXiv:2606.05180v1 Announce Type: new Abstract: Automated scoring models are increasingly used to assign rubric-based quality ratings to complex language performances, including classroom transcripts,

researcharxiv-cs-cl
5 Jun 2026
Agents

Meridian: Metric-Semantic Primitive Matching for Cross-View Geo-Localization Beyond Urban Environments

DGX agent

arXiv:2606.06312v1 Announce Type: new Abstract: Successful robot automation requires accurate global localization to support repeatability, task planning, goal specification, and safe operation. Howev

agentsarxiv-cs-ro
5 Jun 2026
Safety

Real-Time Threat Detection from Surveillance Cameras using Machine Learning

DGX agent

arXiv:2606.05708v1 Announce Type: new Abstract: Ensuring public safety in densely populated urban environments remains a critical challenge, necessitating the deployment of intelligent and automated v

safetyarxiv-cs-cv
5 Jun 2026
Model Releases

Stability vs. Manipulability: Evaluating Robustness Under Post-Decision Interaction in LLM Judges

DGX agent

arXiv:2606.05384v1 Announce Type: cross Abstract: LLM-as-judge evaluation is widely used in benchmarking pipelines, where model outputs are compared and ranked using automated evaluators. These pipeli

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Abduction Prover in Isabelle/HOL

DGX agent

arXiv:2606.04877v1 Announce Type: cross Abstract: Proof assistants based on expressive logics suffer limited automation for proof search, raising the cost of formal verification based on proof assista

researcharxiv-cs-ai
4 Jun 2026
Model Releases

An Open-Source Two-Stage Computer Vision Pipeline for Fine-Grained Vehicle Classification using Vision Transformers

DGX agent

arXiv:2606.05149v1 Announce Type: new Abstract: Vehicle body type is a significant determinant of cyclist injury severity in overtaking crashes, yet automated tools for classifying vehicles into injur

model-releasesarxiv-cs-cv
4 Jun 2026
Safety

Be Fair! Can Machine Learning Engineering Agents Adhere to Fairness Constraints?

DGX agent

arXiv:2606.04971v1 Announce Type: new Abstract: Machine learning engineering (MLE) agents promise to automate end-to-end ML pipeline development from raw data and natural language instructions, potent

safetyarxiv-cs-lg
4 Jun 2026
Agents

Beyond Prompt-Based Planning: MCP-Native Graph Planning-based Biomedical Agent System

DGX agent

arXiv:2606.04494v1 Announce Type: new Abstract: Biomedical agents promise to automate complex biological workflows, yet current systems face two fundamental bottlenecks: bioinformatics tools are highl

agentsarxiv-cs-ai
4 Jun 2026
Research

FoeGlass: Simple In-Context Learning Is Enough for Red Teaming Audio Deepfake Detectors

DGX agent

arXiv:2606.05101v1 Announce Type: cross Abstract: Audio deepfake detection (ADD) models are critical for countering the malicious use of text-to-speech (TTS) models. Evaluating and strengthening ADD m

researcharxiv-cs-lg
4 Jun 2026
Tutorials

J-RAS: Mutual Adaptation for Medical Image Segmentation via Contrastive Retrieval-Augmented Joint Optimization

DGX agent

arXiv:2510.09953v3 Announce Type: replace Abstract: Manual medical image segmentation by clinicians, though accurate, is time-consuming and variable across experts, whereas AI-based models automate th

tutorialsarxiv-cs-cv
4 Jun 2026
Model Releases

LCSHBench: A Multilingual, Consensus-Grounded Benchmark for Library of Congress Subject Heading Assignment

DGX agent

arXiv:2606.04382v1 Announce Type: cross Abstract: Automated subject cataloging assigns controlledvocabulary headings to bibliographic records, but LCSH has no standard public benchmark. We introduce L

model-releasesarxiv-cs-ai
4 Jun 2026
Research

NLLog: Lightweight, Explainable SOC Anomaly Detection via Log-to-Language Rewriting

DGX agent

arXiv:2606.04957v1 Announce Type: cross Abstract: System-generated logs underpin security monitoring, yet their rigid template-based format hinders both automated analysis and human comprehension. We

researcharxiv-cs-lg
4 Jun 2026
Model Releases

OckBench: Measuring the Efficiency of LLM Reasoning

DGX agent

arXiv:2511.05722v3 Announce Type: replace-cross Abstract: Large language models (LLMs) such as GPT-5 and Gemini 3 have pushed the frontier of automated reasoning and code generation. Yet current bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Plan First, Judge Later, Run Better: A DMAIC-Inspired Agentic System for Industrial Anomaly Detection

DGX agent

arXiv:2606.04599v1 Announce Type: new Abstract: Large language model (LLM) agents have shown promise in automating complex data-analysis workflows, but their reliable deployment remains challenging in

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

DGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

model-releasesarxiv-cs-lg
4 Jun 2026
Model Releases

Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models

DGX agent

arXiv:2606.04739v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong potential for automated software vulnerability detection, particularly in retrieval-augmented generatio

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Aligning Data-Driven Predictors with Allocation: A Decision-Focused Approach to Survival Analysis

DGX agent

arXiv:2606.02671v1 Announce Type: cross Abstract: Machine learning predictors have become essential tools for guiding automated decision making. However, a major misalignment persists: predictive mode

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

AnyAudio-Judge: A Dynamic Rubric-Based Benchmark and Evaluator for Audio Instruction Following

DGX agent

arXiv:2606.03116v1 Announce Type: cross Abstract: The rapid advancement of instruction-guided audio generation has highlighted the critical need for robust alignment evaluation. Current automated eval

model-releasesarxiv-cs-ai
3 Jun 2026
Research

CoughSense: Five-Class Respiratory Disease Classification via Whisper Encoder Fine-Tuning and Dual-Encoder Cross-Attention Fusion with Balanced Contrastive Learning

DGX agent

arXiv:2606.02998v1 Announce Type: new Abstract: Automated cough analysis offers a path to low-cost respiratory screening, but most existing work stops at binary COVID-19 detection. A practical tool ne

researcharxiv-cs-lg
3 Jun 2026
Research

Don't Gamble, GAMBLe: An Analytical Framework for AI-Driven Research Systems

DGX agent

arXiv:2606.02863v1 Announce Type: new Abstract: AI-Driven Research Systems (ADRS) -- systems coupling LLMs with automated evaluation to discover algorithms, proofs, and designs -- are being optimized

researcharxiv-cs-ai
3 Jun 2026
Agents

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

DGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

agentsarxiv-cs-ai
3 Jun 2026
Safety

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

DGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

MedCUA-Bench: A Screenshot-Only Benchmark for Clinical Computer-Use Agents

DGX agent

arXiv:2606.03203v1 Announce Type: new Abstract: Computer-use agents could automate repetitive screen-based clinical work, but their reliability in medical graphical user interfaces remains largely unv

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Structure-Guided Mixed Masked Pretraining and Spatial Continuity Regularization for Printed Circuit Board Defect Detection

DGX agent

arXiv:2606.03508v1 Announce Type: new Abstract: Printed circuit board (PCB) defect detection is an essential part of automated optical inspection (AOI); yet it remains challenging in practice because

researcharxiv-cs-cv
3 Jun 2026
Agents

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

DGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

A Machine-to-Machine Knowledge-Guided LLM Agent for Generalizable Radiotherapy Treatment Planning

DGX agent

arXiv:2606.00922v1 Announce Type: cross Abstract: In this work, we propose a prototype machine-to-machine (M2M) knowledge-guided Large Language Model (LLM) framework for automated radiotherapy treatme

model-releasesarxiv-cs-ro
2 Jun 2026
Research

A Unified Framework for Probabilistic Dynamic-, Trajectory- and Vision-based Virtual Fixtures

DGX agent

arXiv:2506.10239v3 Announce Type: replace Abstract: Probabilistic Virtual Fixtures (VFs) enable the adaptive selection of the most suitable haptic feedback for each phase of a task, based on learned o

researcharxiv-cs-ro
2 Jun 2026
Model Releases

ADRA-Bank: A Modular Benchmark for Academic Deep Research Agents

DGX agent

arXiv:2512.00986v3 Announce Type: replace Abstract: A surge in academic publications calls for automated deep research (DR) systems, but accurately evaluating them is still an open problem. First, exi

model-releasesarxiv-cs-cl
2 Jun 2026
Research

Cognitive-Linguistic Indicators of Depression in Online Communities: Analysed by DistilBERT and Holographic Reduced Representation

DGX agent

arXiv:2606.00026v1 Announce Type: new Abstract: This paper investigates whether combining cognitively grounded linguistic features with transformer-based embeddings improves automated detection of dep

researcharxiv-cs-cl
2 Jun 2026
Research

ELF: A Family of Encoder-Free ECG-Language Models

DGX agent

arXiv:2601.18798v2 Announce Type: replace-cross Abstract: ECG-Language Models (ELMs) extend recent advances in Multimodal Large Language Models (MLLMs) to automated ECG interpretation. However, most e

researcharxiv-cs-ai
2 Jun 2026
Research

GloResNet: A lightweight 3D CNN with global topological features for preterm brain injury prediction

DGX agent

arXiv:2606.02498v1 Announce Type: new Abstract: This study introduces an automated deep learning framework for predicting brain injury (BI) in preterm infants from T2-weighted MRI (dHCP dataset). We p

researcharxiv-cs-cv
2 Jun 2026
Agents

iML: Executable, Problem-Grounded, and Broadly Exploratory Code-Driven AutoML

DGX agent

arXiv:2602.13937v2 Announce Type: replace Abstract: Automated Machine Learning (AutoML) has improved access to machine learning, yet existing techniques often remain limited in flexibility, transparen

agentsarxiv-cs-lg
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Safety

Mitigating Perceptual Judgment Bias in Multimodal LLM-as-a-Judge via Perceptual Perturbation and Reward Modeling

DGX agent

arXiv:2606.02578v1 Announce Type: cross Abstract: Recent multimodal large language models have demonstrated strong reasoning ability, yet their reliability as automated evaluators remains limited by a

safetyarxiv-cs-ai
2 Jun 2026
Hardware

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

DGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

hardwarearxiv-cs-ai
2 Jun 2026
Model Releases

Truth, Trust, and Trouble: Medical AI on the Edge

DGX agent

arXiv:2507.02983v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) hold significant promise for transforming digital health by enabling automated medical question answering. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

What to Format and How: A Benchmark and Workflow Approach for Document Formatting

DGX agent

arXiv:2606.01936v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have opened up new possibilities for automated document formatting. However, real-world formatting often

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Automatically Attacking Software Reverse Engineering AI Agents

DGX agent

arXiv:2605.30667v1 Announce Type: cross Abstract: Software tools for reverse engineering executable binary files, such as Ghidra, enable malware analysts to safely conduct robust static analysis witho

agentsarxiv-cs-ai
1 Jun 2026
Research

Beyond Tokens: Enhancing RTL Quality Estimation via Structural Graph Learning

DGX agent

arXiv:2508.18730v2 Announce Type: replace Abstract: Estimating the quality of register transfer level (RTL) designs is crucial in the electronic design automation (EDA) workflow, as it enables instant

researcharxiv-cs-lg
1 Jun 2026
Research

Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models

DGX agent

arXiv:2605.30631v1 Announce Type: cross Abstract: While automated diagnosis systems have achieved remarkable success in computed tomography (CT)-based lung cancer screening, their development remains

researcharxiv-cs-ai
1 Jun 2026
Safety

Cost-aware Stopping for Bayesian Optimization

DGX agent

arXiv:2507.12453v5 Announce Type: replace Abstract: In automated machine learning, scientific discovery, and other applications of Bayesian optimization, deciding when to stop evaluating expensive bla

safetyarxiv-cs-lg
1 Jun 2026
Safety

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

DGX agent

arXiv:2602.00521v2 Announce Type: replace Abstract: While LLM-as-a-Judge is widely used in automated evaluation, existing validation practices primarily operate at the level of observed outputs, offer

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

DGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

model-releasesarxiv-cs-cl
1 Jun 2026
← Previous
1…2324252627…82
Next →