AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,914 results
Agents

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents

DGX agent

arXiv:2606.31229v1 Announce Type: new Abstract: Ideation plays a pivotal role in scientific discovery. Recent LLM, especially AI Scientist systems, show promising potential for automated ideation. How

agentsarxiv-cs-ai
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
Agents

DigitalCoach: Communication and Grounding Gaps in Human and Agentic Computer Use Coaching

DGX agent

arXiv:2606.31980v1 Announce Type: new Abstract: Agents are increasingly capable of automating software tasks, but can they teach humans how to use software themselves? We introduce DigitalCoach, a mul

agentsarxiv-cs-cl
1 Jul 2026
Model Releases

DiscoGen: Procedural Generation of Algorithm Discovery Tasks in Machine Learning

DGX agent

arXiv:2603.17863v2 Announce Type: replace-cross Abstract: Automating the development of machine learning algorithms has the potential to unlock new breakthroughs. However, our ability to improve and e

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Optimizing Image Preparation and Compression for Face Recognition within 1024 Bytes

DGX agent

arXiv:2606.30321v1 Announce Type: new Abstract: ICAO-compliant machine readable travel documents enable automated biometric face verification. The biometric reference is stored on an RFID chip include

researcharxiv-cs-cv
30 Jun 2026
Agents

Socratic agents for autonomous scientific discovery in high-dimensional physical systems

DGX agent

arXiv:2606.26722v1 Announce Type: new Abstract: The automation of scientific discovery has reached an inflection point. While AI systems now operate instruments, optimize parameters and generate hypot

agentsarxiv-cs-ai
26 Jun 2026
Agents

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

DGX agent

arXiv:2606.25332v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise for automated penetration testing, yet existing end-to-end black-box evaluations are highly susceptibl

agentsarxiv-cs-ai
25 Jun 2026
Research

Externalizing Research Synthesis and Validation in AI Scientists through a Research Harness

DGX agent

arXiv:2606.18874v2 Announce Type: replace Abstract: AI systems can increasingly automate scientific workflows, but the reasoning that links prior evidence, generated ideas, experiments and final claim

researcharxiv-cs-ai
25 Jun 2026
Model Releases

Type Checking Project Haystack Grids using JSON Schema and Pydantic

DGX agent

arXiv:2606.24891v1 Announce Type: cross Abstract: Ontologies enable scalable energy services in buildings by supporting interoperability and automation. Project Haystack is a building ontology that is

model-releasesarxiv-cs-ai
25 Jun 2026
Model Releases

When Multi-Sensor Fusion Fails to Generalize: Cattle Posture Classification Under Animal-Level and Temporal Distribution Shift

DGX agent

arXiv:2606.24986v1 Announce Type: new Abstract: Automated cattle posture-classification systems frequently report near-perfect accuracy, yet their robustness under realistic deployment conditions rema

model-releasesarxiv-cs-lg
25 Jun 2026
Research

Beyond Logprobs: A Multi-Signal Confidence Engine for LLM-Based Document Field Extraction

DGX agent

arXiv:2606.24420v1 Announce Type: new Abstract: In high-stakes document processing pipelines, including financial reconciliation, compliance verification, and procurement automation, an LLM extraction

researcharxiv-cs-cl
24 Jun 2026
Safety

Overcoming Imperfect Kinematics in Surgical Robotics Through Sim-to-Real Visuomotor Learning

DGX agent

arXiv:2606.21396v1 Announce Type: new Abstract: Robot-Assisted Surgery is integral to modern minimally invasive procedures, with automation emerging as the next frontier to enhance precision and reduc

safetyarxiv-cs-ro
23 Jun 2026
Research

Structured Spectral Graph Representation Learning for Multi-label Abnormality Analysis from 3D CT Scans

DGX agent

arXiv:2510.10779v5 Announce Type: replace Abstract: With the growing volume of CT examinations, there is an increasing demand for automated tools such as organ segmentation, abnormality detection, and

researcharxiv-cs-cv
23 Jun 2026
Model Releases

VideoAgent: All-in-One Framework for Video Understanding and Editing

DGX agent

arXiv:2606.23327v1 Announce Type: new Abstract: Video editing has become essential in digital media creation, yet existing automated systems are restricted to short segment processing and domain-speci

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

Benchmarking Cross-Domain Audio-Visual Deception Detection

DGX agent

arXiv:2405.06995v4 Announce Type: replace-cross Abstract: Automated deception detection is crucial for assisting humans in accurately assessing truthfulness and identifying deceptive behavior. Convent

model-releasesarxiv-cs-cv
11 Jun 2026
Agents

A case study of evaluating AI agents on a neuroscience data-to-discovery pipeline

DGX agent

arXiv:2606.07718v1 Announce Type: new Abstract: Agentic AI tools offer a promising path to automating software development bottlenecks in scientific research pipelines, particularly for stages that ta

agentsarxiv-cs-ai
9 Jun 2026
Research

LATTEArena: An Evaluation Framework for LLM-powered Tabular Feature Engineering (Extended Version)

DGX agent

arXiv:2606.09004v1 Announce Type: new Abstract: Feature engineering remains essential for tabular data analysis, and Large Language Models (LLMs) have emerged as a promising paradigm for automating th

researcharxiv-cs-ai
9 Jun 2026
Model Releases

Real-Time Industrial Defect Detection on Edge Hardware Using Fine-Tuned YOLOv8: A Systematic Benchmark on the NEU Surface Defect Database and MVTec AD with Automotive & Battery Manufacturing Extensions

DGX agent

arXiv:2606.07659v1 Announce Type: new Abstract: Automated surface defect detection is critical for ensuring rigorous quality control in high-speed manufacturing environments. While deep learning model

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents

DGX agent

arXiv:2602.08235v2 Announce Type: replace-cross Abstract: Although computer-use agents (CUAs) hold significant potential to automate increasingly complex OS workflows, they can demonstrate unsafe unin

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Research

SleepVLM: Explainable and Rule-Grounded Sleep Staging via a Vision-Language Model

DGX agent

arXiv:2603.26738v3 Announce Type: replace-cross Abstract: While automated sleep staging has achieved expert-level accuracy, its clinical adoption is hindered by a lack of auditable reasoning. We intro

researcharxiv-cs-ai
3 Jun 2026
Safety

ANDES: Agent Native Data Evolving Synthesis Tool for Autonomous Instruction Alignment

DGX agent

arXiv:2606.01279v1 Announce Type: new Abstract: AI agents are increasingly being tasked with automating AI research itself, particularly the critical post-training phase that transforms base LLMs into

safetyarxiv-cs-ai
2 Jun 2026
Applications

MiCU: End-to-End Smart Home Command Understanding with Large Language Model

DGX agent

arXiv:2606.01099v1 Announce Type: cross Abstract: Command understanding systems in smart home ecosystems can automate device control and substantially improve user experience. However, while they perf

applicationsarxiv-cs-ai
2 Jun 2026
Model Releases

VESTA: Visual Exploration with Statistical Tool Agents

DGX agent

arXiv:2606.00384v1 Announce Type: new Abstract: Fitting quantitative models to data is a central step in scientific workflows, yet it remains one of the least automated. Recent agent-based systems lev

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

LLM Judges Inconsistently Disagree Across Safety Criteria and Harm Categories

DGX agent

arXiv:2605.31381v1 Announce Type: new Abstract: We evaluate the consistency of automated judges in conducting a multi-dimensional safety evaluation in a reference-free setup. Our results indicate that

safetyarxiv-cs-cl
1 Jun 2026
Agents

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

DGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

agentsarxiv-cs-ai
28 May 2026
Model Releases

Benchmarking Convolutional, Transformer, Hybrid, and Vision Language Models for Multi Disease Retinal Screening

DGX agent

arXiv:2605.26283v1 Announce Type: new Abstract: Modern deep learning offers powerful tools for automated retinal screening, but it remains unclear how different visual model families compare in realis

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

PRISM: A Multi-Dimensional Benchmark for Evaluating LLM Peer Reviewers

DGX agent

arXiv:2605.26730v1 Announce Type: new Abstract: The rapid growth in submissions to machine learning venues has strained the scientific peer-review system and intensified interest in LLM-based automate

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

AutoSG: LLM-Driven Solver Generation Solely from Task Prompts for Expensive Optimization

DGX agent

arXiv:2605.25658v1 Announce Type: cross Abstract: Expensive optimization tasks are ubiquitous in real-world applications, demanding highly specialized solvers. While LLM-driven automated solver genera

local-aiarxiv-cs-ai
26 May 2026
Local Ai

Catching MRI outliers: unsupervised detection and localization of MRI artefacts and clinical anomalies using deep learning

DGX agent

arXiv:2605.24609v1 Announce Type: cross Abstract: Artificial intelligence is increasingly integrated into radiotherapy workflows, yet such pipelines remain vulnerable to out-of-distribution image data

local-aiarxiv-cs-ai
26 May 2026
Research

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence

DGX agent

arXiv:2509.23573v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to help security analysts manage the surge of cyber threats, automating tasks from vulnerab

researcharxiv-cs-ai
26 May 2026
Model Releases

OSS: Open Suturing Skills Vision-Based Assessment Challenge 2024-2025

DGX agent

arXiv:2605.22200v1 Announce Type: new Abstract: Achieving high levels of surgical skill through effective training is essential for optimal patient outcomes. Automated, data-driven skill assessment ho

model-releasesarxiv-cs-cv
22 May 2026
Applications

Geo-Data-Driven HD Map Generation Workflow with Integrated Reference-Free Constraint-Based Verification

DGX agent

arXiv:2605.18921v1 Announce Type: new Abstract: High-definition (HD) maps are core artifacts for automated driving systems, but their generation commonly relies on sensor-intensive mobile mapping camp

applicationsarxiv-cs-ro
20 May 2026
Model Releases

AI for Auto-Research: Roadmap & User Guide

DGX agent

arXiv:2605.18661v1 Announce Type: new Abstract: AI-assisted research is crossing a threshold: fully automated systems can now generate research papers for as little as $15, while long-horizon agents c

model-releasesarxiv-cs-ai
19 May 2026
Applications

Beyond Imperfect Alternatives with Rulemapping: A Neuro-Symbolic Case Study on Online Hate Speech

DGX agent

arXiv:2605.16280v1 Announce Type: cross Abstract: Automating legal reasoning forces a choice between imperfect alternatives: symbolic systems offer transparency but struggle with ambiguity, whereas ne

applicationsarxiv-cs-ai
19 May 2026
Research

Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs

DGX agent

arXiv:2508.15878v2 Announce Type: replace-cross Abstract: Formal theorem proving (FTP) has emerged as a critical foundation for evaluating the reasoning capabilities of large language models, enabling

researcharxiv-cs-ai
19 May 2026
Agents

MemRepair: Hierarchical Memory for Agentic Repository-Level Vulnerability Repair

DGX agent

arXiv:2605.17444v1 Announce Type: cross Abstract: Modern software ecosystems face a rapidly growing number of disclosed vulnerabilities, increasing the need for automated repair techniques that can op

agentsarxiv-cs-ai
19 May 2026
Model Releases

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

DGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

TriALS: Triphasic-Aided Liver Lesion Segmentation Benchmark in Non-Contrast CT

DGX agent

arXiv:2605.16572v1 Announce Type: new Abstract: Automated segmentation of liver lesions on non-contrast computed tomography (NCCT) is clinically important but fundamentally challenging, particularly i

model-releasesarxiv-cs-cv
19 May 2026
Local Ai

Surrogate Neural Architecture Codesign Package (SNAC-Pack)

DGX agent

arXiv:2605.16138v1 Announce Type: cross Abstract: Neural architecture search (NAS) is a powerful approach for automating model design, but existing methods often optimize for accuracy alone or rely on

local-aiarxiv-cs-ai
18 May 2026
Model Releases

CausalReasoningBenchmark: A Real-World Benchmark for Disentangled Evaluation of Causal Identification and Estimation

DGX agent

arXiv:2602.20571v2 Announce Type: replace Abstract: Many benchmarks for automated causal inference evaluate a system's performance based on a single numerical output, such as an Average Treatment Effe

model-releasesarxiv-cs-ai
15 May 2026
Safety

SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration

DGX agent

arXiv:2605.14089v1 Announce Type: new Abstract: In recent years, a variety of powerful LLM-based agentic systems have been applied to automate complex tasks through task orchestration. However, existi

safetyarxiv-cs-ai
15 May 2026
Model Releases

Formal Conjectures: An Open and Evolving Benchmark for Verified Discovery in Mathematics

DGX agent

arXiv:2605.13171v1 Announce Type: new Abstract: As automated reasoning systems advance rapidly, there is a growing need for research-level formal mathematical problems to accurately evaluate their cap

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

VERA-MH Concept Paper

DGX agent

arXiv:2510.15297v4 Announce Type: replace-cross Abstract: We introduce VERA-MH (Validation of Ethical and Responsible AI in Mental Health), an automated evaluation of the safety of AI chatbots used in

model-releasesarxiv-cs-ai
14 May 2026
Safety

Persona-Conditioned Adversarial Prompting: Multi-Identity Red-Teaming for Adversarial Discovery and Mitigation

DGX agent

arXiv:2605.11730v1 Announce Type: new Abstract: Automated red-teaming for LLMs often discovers narrow attack slices, missing diverse real-world threats, and yielding insufficient data for safety fine-

safetyarxiv-cs-lg
13 May 2026
Research

Read, Extract, Classify: A Tool for Smarter Requirements Engineering

DGX agent

arXiv:2605.11045v1 Announce Type: cross Abstract: This paper presents the ReXCL tool, which automates the extraction and classification processes in requirements engineering, enhancing the software de

researcharxiv-cs-lg
13 May 2026
Model Releases

CUDAHercules: Benchmarking Hardware-Aware Expert-level CUDA Optimization for LLMs

DGX agent

arXiv:2605.08467v1 Announce Type: new Abstract: Large language models show promise for automated CUDA programming, however even the strongest coding models (e.g., Claude-Opus-4.6) may still fall short

model-releasesarxiv-cs-lg
12 May 2026
Hardware

LLMs for Secure Hardware Design and Related Problems: Opportunities and Challenges

DGX agent

arXiv:2605.10807v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into Electronic Design Automation (EDA) and hardware security is rapidly reshaping the semiconductor i

hardwarearxiv-cs-lg
12 May 2026
← Previous
1…1415161718…82
Next →