AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Model Releases

Can Generalist Agents Automate Data Curation?

DGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Ekka: Automated Diagnosis of Silent Errors in LLM Inference

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.04594v1 Announce Type: cross Abstract: LLM serving frameworks are quickly evolving with a complex software stack and a vast number of optimizations. The rapid development process can introd

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Automated Report-Derived Oncology VQA Benchmark for Evaluating Vision-Language Models on 3D Medical Imaging

DGX agent

arXiv:2606.02809v1 Announce Type: new Abstract: Evaluating vision-language models (VLMs) on medical images requires benchmarks that are clinically grounded, scalable, and controlled for evaluation con

model-releasesarxiv-cs-cv
3 Jun 2026
Agents

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

DGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

ROGLE: Robust Global-Local Alignment with Automated Region Supervision for Text-Based Person Search

DGX agent

arXiv:2606.01825v1 Announce Type: new Abstract: Text-Based Person Search (TBPS) aims to retrieve pedestrian images using natural language queries. However, existing TBPS models, especially those based

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

SkillHarm: Lifecycle-Aware Skill-Based Attacks via Automated Construction

DGX agent

arXiv:2606.02540v1 Announce Type: new Abstract: Agent skills occupy a privileged position in the agent workflow, as agents are expected to implicitly follow and execute them, rendering third-party ski

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

COLLEAGUE.SKILL: Automated AI Skill Generation via Expert Knowledge Distillation

DGX agent

arXiv:2605.31264v1 Announce Type: new Abstract: LLM agents are increasingly expected not only to complete isolated tasks, but also to carry bounded representations of human expertise, judgment, and in

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents

DGX agent

arXiv:2509.23694v5 Announce Type: replace Abstract: Search agents connect LLMs to the Internet, enabling them to access broader and more up-to-date information. However, this also introduces a new thr

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Automating Formal Verification with Agent-Guided Tree Search

DGX agent

arXiv:2605.27485v1 Announce Type: cross Abstract: Formal verification offers a path to provably correct software, but writing verified code remains expensive enough that the technique is rarely used i

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

CktGen: Automated Analog Circuit Design with Generative Artificial Intelligence

DGX agent

arXiv:2410.00995v3 Announce Type: replace Abstract: The automatic synthesis of analog circuits presents significant challenges. Most existing approaches formulate the problem as a single-objective opt

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

PashtoTTS-Bench: automated screening for low-resource non-Latin-script text-to-speech

DGX agent

arXiv:2605.26978v1 Announce Type: new Abstract: Text-to-speech (TTS) evaluation for low-resource non-Latin-script languages can fail when it relies on a single ASR round-trip word error rate (WER). A

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

Automated Benchmark Auditing for AI Agents and Large Language Models

DGX agent

arXiv:2605.26079v1 Announce Type: new Abstract: Modern AI benchmarks operate at a complexity that outpaces traditional verification methods. Tasks authored by domain experts often contain implicit ass

model-releasesarxiv-cs-cl
26 May 2026
Applications

Automated regime classification in multidimensional time series data using sliced Wasserstein k-means clustering

DGX agent

arXiv:2310.01285v2 Announce Type: replace-cross Abstract: Recent work has proposed Wasserstein k-means (Wk-means) clustering as a powerful method to classify regimes in time series data, and one-dimen

applicationsarxiv-cs-lg
26 May 2026
Safety

From Automation to Collaboration: Human-in-the-Loop Methods for Safe and Trustworthy NLP

DGX agent

arXiv:2605.25226v1 Announce Type: new Abstract: Large language models are widely deployed in high-stakes NLP tasks, yet risks such as bias, hallucination, adversarial vulnerability and unreliable gene

safetyarxiv-cs-cl
26 May 2026
Agents

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

DGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

agentsarxiv-cs-cl
26 May 2026
Research

Scaling Natural-Language Graph-Based Test Time Compute for Automated Theorem Proving

DGX agent

arXiv:2503.11657v3 Announce Type: replace Abstract: Large language models have demonstrated remarkable capabilities in natural language processing tasks requiring multi-step logical reasoning capabili

researcharxiv-cs-cl
26 May 2026
Applications

ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning

DGX agent

arXiv:2605.23454v1 Announce Type: new Abstract: Rubric-based rewards offer a promising way to extend reinforcement learning (RL) for large language models beyond tasks with automatically verifiable an

applicationsarxiv-cs-cl
25 May 2026
Applications

Automated Random Embedding for Practical Bayesian Optimization with Unknown Effective Dimension

DGX agent

arXiv:2605.23473v1 Announce Type: cross Abstract: Bayesian optimization is widely employed for optimizing complex black-box functions but struggles with the curse of dimensionality. Random embedding,

applicationsarxiv-cs-ai
25 May 2026
Agents

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

DGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

agentsarxiv-cs-ai
22 May 2026
Agents

N3P: Accelerated Automated Parking via a Learning-Based Naturalistic Three-Stage Scheme

DGX agent

arXiv:2605.22722v1 Announce Type: new Abstract: Autonomous parking requires efficient path planning that ensures kinematic feasibility and collision avoidance in constrained environments. Hybrid A* is

agentsarxiv-cs-ro
22 May 2026
Model Releases

Automated ICD Classification of Psychiatric Diagnoses: From Classical NLP to Large Language Models

DGX agent

arXiv:2605.21154v1 Announce Type: new Abstract: Mental health has become a global priority, leading to a massive administrative burden in the coding of clinical diagnoses. This study proposes the auto

model-releasesarxiv-cs-cl
21 May 2026
Research

Automated Kernel Discovery Towards Understanding High-dimensional Bayesian Optimization

DGX agent

arXiv:2605.20249v1 Announce Type: new Abstract: Gaussian Process (GP) kernels are central to Bayesian optimization (BO), yet designing effective kernels for high-dimensional problems still relies on e

researcharxiv-cs-lg
21 May 2026
Applications

GradeLegal: Automated Grading for German Legal Cases

DGX agent

arXiv:2605.21076v1 Announce Type: new Abstract: Grading German legal exam solutions faces growing volumes and a shortage of qualified graders, delaying feedback and creating a bottleneck. At the same

applicationsarxiv-cs-cl
21 May 2026
Tutorials

BCI-sift: An automated feature selection toolbox for Brain Computer Interface applications

DGX agent

arXiv:2605.19646v1 Announce Type: cross Abstract: Advancements in clinical Brain-Computer Interfaces (BCIs) depend on precise and reliable signal interpretation. However, the high-dimensional and nois

tutorialsarxiv-cs-lg
20 May 2026
Research

Automated Coding of Communication Data Using ChatGPT: Consistency Across Subgroups

DGX agent

arXiv:2510.20584v3 Announce Type: replace-cross Abstract: Assessing communication and collaboration at scale depends on a labor-intensive task of coding communication data into categories according to

researcharxiv-cs-ai
19 May 2026
Tutorials

Automated Knowledge Component Generation for Interpretable Knowledge Tracing in Coding Problems

DGX agent

arXiv:2502.18632v4 Announce Type: replace Abstract: Knowledge components (KCs) mapped to problems help model student learning, tracking their mastery levels on fine-grained skills thereby facilitating

tutorialsarxiv-cs-ai
19 May 2026
Safety

Beyond Safety Filtering: Control Barrier Function-Informed Reinforcement Learning for Connected and Automated Vehicles

DGX agent

arXiv:2605.16894v1 Announce Type: new Abstract: Reinforcement Learning (RL) uses rewards to guide learning, yet reward design is typically hand-crafted using heuristics that can be difficult to tune.

safetyarxiv-cs-ro
19 May 2026
Local Ai

CheckSupport: A Local LLM-Powered Tool for Automated Manuscript Submission Checklist Selection and Completion

DGX agent

arXiv:2605.16377v1 Announce Type: cross Abstract: Transparent and standardized reporting is essential for reproducible scientific research, yet adherence to reporting guidelines remains inconsistent b

local-aiarxiv-cs-ai
19 May 2026
Agents

SkillJect: Effectively Automating Skill-Based Prompt Injection for Skill-Enabled Agents

DGX agent

arXiv:2602.14211v2 Announce Type: replace-cross Abstract: Agent skills are increasingly used to extend LLM agents with task-specific instructions, executable scripts, and auxiliary resources. While im

agentsarxiv-cs-ai
19 May 2026
Research

Vidya: An AI-Driven Modular Pipeline for Archival Automation and Semantic Metadata Enrichment

DGX agent

arXiv:2605.16338v1 Announce Type: cross Abstract: The large-scale digitization of historical archives has created a paradox: 'dark data'-digital objects lacking metadata for retrieval. Manual archival

researcharxiv-cs-cl
19 May 2026
Model Releases

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

DGX agent

arXiv:2605.15218v1 Announce Type: new Abstract: Large language models deployed for MAPDL finite-element simulation face practical reliability challenges: without structured execution control, tool enc

model-releasesarxiv-cs-ai
18 May 2026
Agents

BOOST: A Data-Driven Framework for the Automated Joint Selection of Kernel and Acquisition Functions in Bayesian Optimization

DGX agent

arXiv:2508.02332v3 Announce Type: replace Abstract: The performance of Bayesian optimization (BO), a highly sample-efficient method for expensive black-box problems, is critically governed by the sele

agentsarxiv-cs-lg
15 May 2026
Model Releases

CUICurate: A GraphRAG-based Framework for Automated Clinical Concept Curation for NLP applications

DGX agent

arXiv:2602.17949v2 Announce Type: replace-cross Abstract: Background: Clinical named entity recognition tools commonly map free text to Unified Medical Language System (UMLS) Concept Unique Identifier

model-releasesarxiv-cs-ai
15 May 2026
Safety

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

DGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

safetyarxiv-cs-ai
14 May 2026
Model Releases

SAGE: Scalable Automated Robustness Augmentation for LLM Knowledge Evaluation

DGX agent

arXiv:2605.12022v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong performance on standard knowledge evaluation benchmarks, yet recent work shows that their knowledge capabili

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

Automated Approach for Solving Infinite-state Polynomial Reachability Games

DGX agent

arXiv:2605.10169v1 Announce Type: new Abstract: Reachability games are two-player games played on a graph, where the objective of exttt{REACH} player is to reach the target set whereas the objective o

model-releasesarxiv-cs-ai
12 May 2026
Research

Automated high-frequency quantification of fish communities and biomass using computer vision

DGX agent

arXiv:2605.10449v1 Announce Type: new Abstract: Quantifying fish community structure is essential for understanding biodiversity and ecosystem responses in a changing environment, yet existing survey

researcharxiv-cs-cv
12 May 2026
Research

Automated Robotic Moisture Monitoring in Agricultural Fields

DGX agent

arXiv:2605.09050v1 Announce Type: cross Abstract: Monitoring moisture level of land in a large-scale plantation is tedious. The main objective of this project is to use a robotic kit in collaboration

researcharxiv-cs-cv
12 May 2026
Hardware

Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes

DGX agent

arXiv:2505.02184v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used for generating parallel scientific codes, with a primary focus on generating functionally correct

hardwarearxiv-cs-ai
12 May 2026
Model Releases

MonitoringBench: Semi-Automated Red-Teaming for Agent Monitoring

DGX agent

arXiv:2605.09684v1 Announce Type: cross Abstract: We introduce a red-teaming methodology that exposes harder-to-catch attacks for coding-agent monitors, suggesting that current practices may under-eli

model-releasesarxiv-cs-ai
12 May 2026
Safety

Position: Academic Conferences are Potentially Facing Denominator Gaming Caused by Fully Automated Scientific Agents

DGX agent

arXiv:2605.09915v1 Announce Type: cross Abstract: The implicit policy of maintaining relatively stable acceptance rates at top AI conferences, despite exponentially growing submissions, introduces a c

safetyarxiv-cs-ai
12 May 2026
Agents

Direction for Detection: A Survey of Automated Vulnerability Detection and all of its Pain Points

DGX agent

arXiv:2412.11194v2 Announce Type: replace-cross Abstract: Security vulnerabilities in software can have severe consequences; however, manual vulnerability detection is costly and does not scale, espec

agentsarxiv-cs-ai
11 May 2026
Safety

AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation

DGX agent

arXiv:2507.12768v2 Announce Type: replace Abstract: Learning generalizable manipulation policies hinges on data, yet robot manipulation data is scarce and often entangled with specific embodiments, ma

safetyarxiv-cs-cv
7 May 2026
Safety

Confronting Label Indeterminacy in Automated Bail Decisions

DGX agent

arXiv:2605.04073v1 Announce Type: new Abstract: Bail decisions present a fundamental challenge for data-driven decision support systems. When bail is denied, the counterfactual outcome of whether the

safetyarxiv-cs-lg
7 May 2026
Safety

Heterogeneous Graph Importance Scoring and Clustering with Automated LLM-based Interpretation

DGX agent

arXiv:2605.02919v1 Announce Type: new Abstract: Urban bridge networks are critical infrastructure whose disruption can cascade into severe impacts on transportation, emergency services, and economic a

safetyarxiv-cs-lg
6 May 2026
Agents

When LLM Agents Meet Graph Optimization: An Automated Data Quality Improvement Approach

DGX agent

arXiv:2510.08952v4 Announce Type: replace Abstract: Text-attributed graphs (TAGs) have become a key form of graph-structured data in modern data management and analytics, combining structural relation

agentsarxiv-cs-lg
6 May 2026
Research

Automated In-the-Wild Data Collection for Continual AI Generated Image Detection

DGX agent

arXiv:2605.02567v1 Announce Type: new Abstract: The rapid advancement of generative Artificial Intelligence (AI) has introduced significant challenges for reliable AI-generated image detection. Existi

researcharxiv-cs-cv
5 May 2026
Agents

Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation

DGX agent

arXiv:2604.09718v2 Announce Type: cross Abstract: LLM-driven web agents operating through continuous inference loops -- repeatedly querying a model to evaluate browser state and select actions -- exhi

agentsarxiv-cs-ai
1 May 2026
← Previous
1…7891011…82
Next →