AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Model Releases

False Confidence: Automated Labels Confound Fairness Audits in Cervical Spine Segmentation

DGX agent

arXiv:2607.07852v1 Announce Type: cross Abstract: Automated segmentation of cervical-spine MRI is increasingly used in clinical workflows, yet no fairness audit exists for this anatomy. We show that a

model-releasesarxiv-cs-cv
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Designing Maintainable Hybrid Generative Systems: A Quantum-Inspired Approach to Automated Music Harmony Generation

DGX agent

arXiv:2607.06296v1 Announce Type: cross Abstract: This paper presents the design and evaluation of a maintainable hybrid generative architecture for automated music harmony generation from melody. The

researcharxiv-cs-ai
8 Jul 2026
Agents

Automated Data Readiness for Scientific AI

DGX agent

arXiv:2607.02771v1 Announce Type: new Abstract: Leadership computing facilities steward large-scale scientific datasets that routinely require substantial transformation before serving as AI training

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

QEDBENCH: Quantifying the Alignment Gap in Automated Evaluation of University-Level Mathematical Proofs

DGX agent

arXiv:2602.20629v3 Announce Type: replace Abstract: As Large Language Models (LLMs) saturate elementary benchmarks, the research frontier has shifted from generation to the reliability of automated ev

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Structured Prompting and Automated Evaluation in Fixed Synthetic Japanese-Language Counseling Dialogues

DGX agent

arXiv:2507.02950v3 Announce Type: replace-cross Abstract: Large language models (LLMs) may support counseling training, yet evidence from Japanese-language interactions and automated quality ratings r

model-releasesarxiv-cs-ai
7 Jul 2026
Applications

RAISE: LLM-based Automated Heuristic Design with Robust Adversary Instance Search

DGX agent

arXiv:2606.31801v1 Announce Type: new Abstract: Automated Heuristic Design (AHD) with Large Language Models (LLMs) has shown remarkable progress in discovering high-quality heuristics. However, existi

applicationsarxiv-cs-ai
1 Jul 2026
Model Releases

Proteus: Automated Adversarial Robustness Testing for Audio Deepfake Detectors

DGX agent

arXiv:2606.29544v1 Announce Type: cross Abstract: We present Proteus, a framework developed at Resemble AI for automated robustness testing of our audio deepfake detection system. Given a detector, Pr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Symmetry-Aware Transformer Training for Automated Planning

DGX agent

arXiv:2508.07743v2 Announce Type: replace Abstract: While transformers excel in many settings, their application in the field of automated planning is limited. Prior work like PlanGPT, a state-of-the-

safetyarxiv-cs-ai
29 Jun 2026
Agents

Closing the Loop to Discover Psychological Theories with an Automated Cognitive Scientist

DGX agent

arXiv:2606.26448v1 Announce Type: cross Abstract: Across the sciences, autonomous systems are increasingly being used in closed-loop discovery, proposing new theories and designing and running experim

agentsarxiv-cs-ai
26 Jun 2026
Research

An Approach for a Supporting Multi-LLM System for Automated Certification Based on the German IT-Grundschutz

DGX agent

arXiv:2606.25608v1 Announce Type: cross Abstract: This paper presents a novel approach to perform semi-automated BSI IT-Grundschutz certification using a MultiLarge Language Model system (MLS) with Hy

researcharxiv-cs-ai
25 Jun 2026
Research

Female-RHINO: A Real-Time Scanner-Integrated Framework for Automated Quantitative Uterine MRI Analysis and Structured Reporting

DGX agent

arXiv:2606.24390v1 Announce Type: cross Abstract: Standardized assessment of uterine MRI remains challenging due to anatomical variability, observer dependence, and the lack of workflow-integrated aut

researcharxiv-cs-ai
24 Jun 2026
Model Releases

Systematic Exploration of 4-Expert Heterogeneous Mixture-of-Experts via Automated Pipeline Search

DGX agent

arXiv:2606.23739v1 Announce Type: cross Abstract: We present an automated large-scale search pipeline for heterogeneous 4-Expert Mixture-of-Experts (MoE4) architectures within the LEMUR neural network

model-releasesarxiv-cs-cv
24 Jun 2026
Safety

A UAV-Based Multi-Modal Vision System for Automated Sideslope Deformation Monitoring and Hazard Detection

DGX agent

arXiv:2606.20681v1 Announce Type: new Abstract: Slope hazards constitute a major safety threat to expressway infrastructure, and their evolution is typically manifested as slow surface deformation. Co

safetyarxiv-cs-cv
23 Jun 2026
Research

An Effective Strategy for Modeling Score Ordinality and Non-uniform Intervals in Automated Speaking Assessment

DGX agent

arXiv:2509.03372v3 Announce Type: replace-cross Abstract: A recent line of research on automated speaking assessment (ASA) has benefited from self-supervised learning (SSL) representations, which capt

researcharxiv-cs-lg
23 Jun 2026
Research

Causal Reward World Models: Zero-shot Reward Design for Automated Skill Generation

DGX agent

arXiv:2606.23280v1 Announce Type: new Abstract: Automated Reward Design (ARD) aims to replace manual reward engineering in reinforcement learning with language-driven reward function synthesis. Howeve

researcharxiv-cs-ro
23 Jun 2026
Research

FetSelect: Task-Specific Architectures and Self-Supervised Learning for Automated Fetal Ultrasound Frame Selection

DGX agent

arXiv:2606.22487v1 Announce Type: new Abstract: Automated frame selection for fetal biometry remains under addressed, with most prior work targeting generic quality assessment or downstream measuremen

researcharxiv-cs-cv
23 Jun 2026
Research

Multi-Target Maneuver Coordinations: Unlocking Coordination Opportunities in Connected Automated Driving

DGX agent

arXiv:2606.22055v1 Announce Type: cross Abstract: Maneuver coordination is a key enabler of connected and automated driving, allowing vehicles to negotiate and execute maneuvers that would otherwise b

researcharxiv-cs-ro
23 Jun 2026
Model Releases

Reinforcement learning to improve large language model-based automated code compliance systems

DGX agent

arXiv:2606.22402v1 Announce Type: cross Abstract: Large language model (LLM)-based approaches for automated code compliance (ACC) of building regulations are prone to generating incorrect and hallucin

model-releasesarxiv-cs-lg
23 Jun 2026
Applications

LaQual: An Automated Framework for LLM App Quality Evaluation

DGX agent

arXiv:2508.18636v2 Announce Type: replace-cross Abstract: Representing a new paradigm in software distribution, LLM app stores are rapidly emerging, offering users diverse choices for content generati

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

ProbeLLM: Automating Principled Diagnosis of LLM Failures

DGX agent

arXiv:2602.12966v2 Announce Type: replace Abstract: Understanding how and why large language models (LLMs) fail is becoming a central challenge as models rapidly evolve and static evaluations fall beh

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

The Order Matters: Sequential Fine-Tuning of LLaMA for Coherent Automated Essay Scoring

DGX agent

arXiv:2606.10327v1 Announce Type: new Abstract: Automated Essay Scoring (AES) systems must judge interdependent discourse elements (e.g., lead, claim, evidence, conclusion), yet most approaches treat

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

DGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

AblationBench: Evaluating Automated Planning of Ablations in Empirical AI Research

DGX agent

arXiv:2507.08038v3 Announce Type: replace-cross Abstract: Language model agents are increasingly used to automate scientific research, yet evaluating their scientific contributions remains a challenge

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Automated Essay Scoring and Language Certification: Assessing Generalizability, Agreement and Validity for French

DGX agent

arXiv:2606.02009v1 Announce Type: new Abstract: In Automated Essay Scoring (AES), benchmarking practices have fostered minimalist evaluation practices, in contrast with the broader-view recommendation

safetyarxiv-cs-cl
2 Jun 2026
Applications

From Capability Models to Automated Planning: An AAS-Native Approach for Automatic PDDL Generation

DGX agent

arXiv:2606.02167v1 Announce Type: new Abstract: Engineers designing production systems need to verify that a given layout supports all required production sequences. Automated planning techniques can

applicationsarxiv-cs-ai
2 Jun 2026
Safety

LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving

DGX agent

arXiv:2606.00372v1 Announce Type: new Abstract: Reliable object detection is critical for automated driving, yet even state-of-the-art detectors inevitably make errors that can compromise safety. Intr

safetyarxiv-cs-cv
2 Jun 2026
Safety

LLM Trainer: Automated Robotic Data Generation via Demonstration Augmentation using LLMs

DGX agent

arXiv:2509.20070v2 Announce Type: replace Abstract: We present LLM Trainer, a fully automated pipeline that leverages the world knowledge of Large Language Models (LLMs) to transform a small number of

safetyarxiv-cs-ro
2 Jun 2026
Model Releases

Automating Formal Verification with Reinforcement Learning and Recursive Inference

DGX agent

arXiv:2605.30914v1 Announce Type: new Abstract: Automated formal verification remains challenging for large language models because data for proof assistants and verification-aware languages is scarce

model-releasesarxiv-cs-lg
1 Jun 2026
Safety

DeepSurvey: Enhancing Analytical Depth and Citation Reliability in Automated Survey Generation

DGX agent

arXiv:2605.29522v1 Announce Type: new Abstract: As scientific literature grows rapidly, automated survey generation has become a key capability for AI scientists and human researchers. However, existi

safetyarxiv-cs-ai
29 May 2026
Model Releases

Gram: Assessing sabotage propensities via automated alignment auditing

DGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

Learnable Assessment Skills for LLM-based Automated Scoring: Rubric Construction via Iterative Optimization

DGX agent

arXiv:2605.29274v1 Announce Type: new Abstract: LLM-based automated scoring approaches near-human performance, but scaling to new tasks remains bottlenecked by the per-item human configuration of upst

tutorialsarxiv-cs-cl
29 May 2026
Safety

Towards automated data analysis: A guided framework for LLM-based risk estimation

DGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

safetyarxiv-cs-ai
28 May 2026
Agents

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

DGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

agentsarxiv-cs-ai
28 May 2026
Model Releases

A Hybrid Vision-Language Architecture for Automated Defect Reasoning and Report Generation in Industrial Inspection

DGX agent

arXiv:2605.26533v1 Announce Type: cross Abstract: Automated industrial inspection requires both precise defect localization and structured maintenance report generation; in current practice these task

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation

DGX agent

arXiv:2509.26600v2 Announce Type: replace-cross Abstract: As LLMs rapidly saturate existing benchmarks, automated benchmark creation using LLMs (LLM-as-a-benchmark) -- where a model generates test inp

model-releasesarxiv-cs-ai
27 May 2026
Agents

APT-Agent: Automated Penetration Testing using Large Language Models

DGX agent

arXiv:2605.24949v1 Announce Type: cross Abstract: Penetration testing is essential to securing modern web infrastructures, yet traditional manual methods struggle to keep pace with their scale and com

agentsarxiv-cs-ai
26 May 2026
Agents

Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models

DGX agent

arXiv:2605.24755v1 Announce Type: new Abstract: Speech monologues recorded in naturalistic settings provide opportunities to characterize mental illness phenomenology and detect symptom exacerbation.

agentsarxiv-cs-ai
26 May 2026
Agents

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

DGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

agentsarxiv-cs-cl
26 May 2026
Research

Does Continued Pretraining on a Learner Corpus Improve Automated Essay Scoring on English Proficiency Tests? Evidence from EFCAMDAT

DGX agent

arXiv:2605.25924v1 Announce Type: new Abstract: Recent automated essay scoring (AES) studies increasingly use pretrained transformer models, but these models are usually pretrained on general-domain E

researcharxiv-cs-cl
26 May 2026
Research

Eye Gaze-Informed and Context-Aware Pedestrian Trajectory Prediction in Shared Spaces with Automated Shuttles: A Virtual Reality Study

DGX agent

arXiv:2603.19812v2 Announce Type: replace Abstract: To address this gap, we conduct a Virtual Reality experiment in which pedestrians interact with automated shuttles under varying approach angles (45

researcharxiv-cs-lg
25 May 2026
Tutorials

Exploring the Effectiveness of Using LLMs for Automated Assessment of Student Self Explanations in Programming Education

DGX agent

arXiv:2605.21614v1 Announce Type: cross Abstract: Worked examples are step-by-step solutions to problems in a specific domain, offered to students to acquire domain-specific problem-solving skills. Th

tutorialsarxiv-cs-lg
23 May 2026
Agents

From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)

DGX agent

arXiv:2605.20608v1 Announce Type: new Abstract: Realizing Level 4/5 Autonomous Networks (AN) demands a shift from static automation to agent-native intelligence. Current operations, reliant on rigid s

agentsarxiv-cs-ai
22 May 2026
Agents

MARS: Modular Agent with Reflective Search for Automated AI Research

DGX agent

arXiv:2602.02660v3 Announce Type: replace Abstract: A critical bottleneck in automating AI research is the execution of complex machine learning engineering (MLE) tasks. MLE differs from general softw

agentsarxiv-cs-ai
22 May 2026
Research

Automated Grading of Handwritten Mathematics Using Vision-Capable LLMs

DGX agent

arXiv:2605.19043v1 Announce Type: cross Abstract: Automated grading systems have enabled scalable assessment for many response types, but handwritten mathematics remains a barrier due to the complexit

researcharxiv-cs-ai
20 May 2026
Applications

Closed-Loop Hybrid Digital Twin Platform for Connected and Automated Vehicle Validation

DGX agent

arXiv:2605.19490v1 Announce Type: cross Abstract: Comprehensive and efficient validation of connected and automated vehicles (CAVs) is critical prior to real-world deployment. While simulation-based t

applicationsarxiv-cs-cv
20 May 2026
Model Releases

Fine-tuning Large Language Model for Automated Algorithm Design

DGX agent

arXiv:2507.10614v2 Announce Type: replace-cross Abstract: The integration of large language models (LLMs) into automated algorithm design has shown promising potential. A prevalent approach embeds LLM

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

XNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception

DGX agent

arXiv:2603.22453v2 Announce Type: replace Abstract: Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

BacktestBench: Benchmarking Large Language Models for Automated Quantitative Strategy Backtesting

DGX agent

arXiv:2605.17937v1 Announce Type: cross Abstract: Quantitative backtesting is essential for evaluating trading strategies but remains hampered by high technical barriers and limited scalability. While

model-releasesarxiv-cs-ai
19 May 2026
← Previous
123456…82
Next →