AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
3,942 results
Research

Self-Supervised Learning of Plant Image Representations

DGX agent

arXiv:2604.27538v1 Announce Type: new Abstract: Automated plant recognition plays a crucial role in biodiversity monitoring and conservation, yet current approaches rely heavily on supervised learning

researcharxiv-cs-cv
1 May 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Skills-Coach: A Self-Evolving Skill Optimizer via Training-Free GRPO

DGX agent

arXiv:2604.27488v1 Announce Type: new Abstract: We introduce Skills-Coach, a novel automated framework designed to significantly enhance the self-evolution of skills within Large Language Model (LLM)-

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

Step-level Optimization for Efficient Computer-use Agents

DGX agent

arXiv:2604.27151v1 Announce Type: new Abstract: Computer-use agents provide a promising path toward general software automation because they can interact directly with arbitrary graphical user interfa

model-releasesarxiv-cs-ai
1 May 2026
Safety

Taxon: Hierarchical Tax Code Prediction with Semantically Aligned LLM Expert Guidance

DGX agent

arXiv:2601.08418v2 Announce Type: replace-cross Abstract: Tax code prediction is a crucial yet underexplored task in automating invoicing and compliance management for large-scale e-commerce platforms

safetyarxiv-cs-ai
1 May 2026
Model Releases

The Epistemic Planning Domain Definition Language: Official Guideline

DGX agent

arXiv:2601.20969v3 Announce Type: replace Abstract: Epistemic planning extends (multi-agent) automated planning by making agents' knowledge and beliefs first-class aspects of the planning formalism. O

model-releasesarxiv-cs-ai
1 May 2026
Agents

Think it, Run it: Autonomous ML pipeline generation via self-healing multi-agent AI

DGX agent

arXiv:2604.27096v1 Announce Type: new Abstract: The purpose of our paper is to develop a unified multi-agent architecture that automates end-to-end machine learning (ML) pipeline generation from datas

agentsarxiv-cs-ai
1 May 2026
Model Releases

WebMall -- A Multi-Shop Benchmark for Evaluating Web Agents

DGX agent

arXiv:2508.13024v3 Announce Type: replace Abstract: LLM-based web agents have the potential to automate long-running web tasks, such as searching for products in multiple e-shops and subsequently orde

model-releasesarxiv-cs-cl
1 May 2026
Research

A Self-Calibrating Framework for Analog Circuit Sizing Using LLM-Derived Analytical Equations

DGX agent

arXiv:2604.07387v2 Announce Type: replace-cross Abstract: We present a design automation framework for analog circuit sizing that produces calibrated, topology-specific analytical equations from raw c

researcharxiv-cs-ai
30 Apr 2026
Model Releases

EvoDev: An Iterative Feature-Driven Framework for End-to-End Software Development with LLM-based Agents

DGX agent

arXiv:2511.02399v2 Announce Type: replace-cross Abstract: Recent advances in large language model agents offer the promise of automating end-to-end software development from natural language requireme

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Inference-Time Scaling of Verification: Self-Evolving Deep Research Agents via Test-Time Rubric-Guided Verification

DGX agent

arXiv:2601.15808v2 Announce Type: replace Abstract: Recent advances in Deep Research Agents (DRAs) are transforming automated knowledge discovery and problem-solving. While the majority of existing ef

safetyarxiv-cs-ai
30 Apr 2026
Research

Motion-Driven Multi-Object Tracking of Model Organisms in Space Science Experiments

DGX agent

arXiv:2604.26321v1 Announce Type: new Abstract: Automated animal behavior analysis relies on long-term, interpretable individual trajectories; however, multi-animal tracking in space science experimen

researcharxiv-cs-cv
30 Apr 2026
Model Releases

OMEGA: Optimizing Machine Learning by Evaluating Generated Algorithms

DGX agent

arXiv:2604.26211v1 Announce Type: new Abstract: In order to automate AI research we introduce a full, end-to-end framework, OMEGA: Optimizing Machine learning by Evaluating Generated Algorithms, that

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

A Blueprint for AI-Driven Software Quality: Integrating LLMs with Established Standards

DGX agent

arXiv:2505.13766v5 Announce Type: replace-cross Abstract: Software Quality Assurance (SQA) is critical for delivering reliable, secure, and efficient software products. The Software Quality Assurance

safetyarxiv-cs-cl
29 Apr 2026
Research

Bye Bye Perspective API: Lessons for Measurement Infrastructure in NLP, CSS and LLM Evaluation

DGX agent

arXiv:2604.25580v1 Announce Type: new Abstract: The closure of Perspective API at the end of 2026 discards what has functioned as the de facto standard for automated toxicity measurement in NLP, CSS,

researcharxiv-cs-cl
29 Apr 2026
Agents

FGDM: Reasoning Aware Multi-Agentic Framework for Software Bug Detection using Chain of Thought and Tree of Thought Prompting

DGX agent

arXiv:2604.24831v1 Announce Type: cross Abstract: Deep Learning methods are becoming prominent in automated software bug detection; however, they lack the global understanding of the given code. Conse

agentsarxiv-cs-lg
29 Apr 2026
Model Releases

VLM Judges Can Rank but Cannot Score: Task-Dependent Uncertainty in Multimodal Evaluation

DGX agent

arXiv:2604.25235v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used as automated judges for multimodal systems, yet their scores provide no indication of reliability.

model-releasesarxiv-cs-cl
29 Apr 2026
Model Releases

Why Search When You Can Transfer? Amortized Agentic Workflow Design from Structural Priors

DGX agent

arXiv:2604.25012v1 Announce Type: new Abstract: Automated agentic workflow design currently relies on per-task iterative search, which is computationally prohibitive and fails to reuse structural know

model-releasesarxiv-cs-lg
29 Apr 2026
Applications

A Hierarchical Ensemble Inference Pipeline for Robust White Blood Cell Classification Under Domain Shifts

DGX agent

arXiv:2604.23271v1 Announce Type: new Abstract: Automated white blood cell (WBC) classification is essential for scalable leukaemia screening. However, real-world deployment is challenged by domain sh

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

AI Security Beyond Core Domains: Resume Screening as a Case Study of Adversarial Vulnerabilities in Specialized LLM Applications

DGX agent

arXiv:2512.20164v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) excel at text comprehension and generation, making them ideal for automated tasks like code review and content mo

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Anonymization-Enhanced Privacy Protection for Mobile GUI Agents: Available but Invisible

DGX agent

arXiv:2602.10139v3 Announce Type: replace-cross Abstract: Mobile Graphical User Interface (GUI) agents have demonstrated strong capabilities in automating complex smartphone tasks by leveraging multim

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

AutoRISE: Agent-Driven Strategy Evolution for Red-Teaming Large Language Models

DGX agent

arXiv:2604.22871v1 Announce Type: cross Abstract: Automated red-teaming methods for large language models typically optimize attack prompts within a fixed, human-designed strategy, leaving the attack

model-releasesarxiv-cs-ai
28 Apr 2026
Research

H-SemiS: Hierarchical Fusion of Semi and Self-Supervised Learning for Knee Osteoarthritis Severity Grading

DGX agent

arXiv:2604.23335v1 Announce Type: new Abstract: Knee osteoarthritis (KOA) is a degenerative joint disease that can lead to chronic pain, reduced mobility, and long-term disability. Automated severity

researcharxiv-cs-cv
28 Apr 2026
Model Releases

HalalBench: A Multilingual OCR Benchmark for Food Packaging Ingredient Extraction

DGX agent

arXiv:2604.22754v1 Announce Type: cross Abstract: No standardized benchmark exists for evaluating OCR on food packaging, despite its critical role in automated halal food verification. Existing benchm

model-releasesarxiv-cs-cl
28 Apr 2026
Applications

IoT-Enhanced CNN-Based Labelled Crack Detection for Additive Manufacturing Image Annotation in Industry 4.0

DGX agent

arXiv:2604.22857v1 Announce Type: new Abstract: This paper presents an IoT-enhanced deep learning framework for automated crack detection in Additive Manufacturing (AM) surfaces using convolutional ne

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systems

DGX agent

arXiv:2604.23478v1 Announce Type: new Abstract: Large language models are increasingly deployed as automated judges for evaluating other models, yet the stability of their verdicts under semantically

model-releasesarxiv-cs-cl
28 Apr 2026
Research

K-SENSE: A Knowledge-Guided Self-Augmented Encoder for Neuro-Semantic Evaluation of Mental Health Conditions on Social Media

DGX agent

arXiv:2604.23493v1 Announce Type: cross Abstract: Early detection of mental health conditions, particularly stress and depression, from social media text remains a challenging open problem in computat

researcharxiv-cs-ai
28 Apr 2026
Model Releases

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

DGX agent

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Scalable Production Scheduling: Linear Complexity via Unified Homogeneous Graphs

DGX agent

arXiv:2604.23841v1 Announce Type: cross Abstract: Efficiently solving the Job Shop Scheduling Problem in real-world industrial applications requires policies that are both computationally lean and top

safetyarxiv-cs-ai
28 Apr 2026
Agents

SeaEvo: Advancing Algorithm Discovery with Strategy Space Evolution

DGX agent

arXiv:2604.24372v1 Announce Type: cross Abstract: LLM-guided evolutionary search has emerged as a promising paradigm for automated algorithm discovery, yet most systems track search progress primarily

agentsarxiv-cs-ai
28 Apr 2026
Research

STELLAR-E: a Synthetic, Tailored, End-to-end LLM Application Rigorous Evaluator

DGX agent

arXiv:2604.24544v1 Announce Type: new Abstract: The increasing reliance on Large Language Models (LLMs) across diverse sectors highlights the need for robust domain-specific and language-specific eval

researcharxiv-cs-ai
28 Apr 2026
Safety

Towards Fair and Robust Volumetric CT Classification via KL-Regularised Group Distributionally Robust Optimisation

DGX agent

arXiv:2603.15941v2 Announce Type: replace Abstract: Automated diagnosis from chest computed tomography (CT) scans faces two persistent challenges in clinical deployment: distribution shift across acqu

safetyarxiv-cs-cv
28 Apr 2026
Safety

Voxify3D: Pixel Art Meets Volumetric Rendering

DGX agent

arXiv:2512.07834v2 Announce Type: replace Abstract: Voxel art is a distinctive stylization widely used in games and digital media, yet automated generation from 3D meshes remains challenging due to co

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

DGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification

DGX agent

arXiv:2604.22601v1 Announce Type: cross Abstract: Large Language Models (LLMs) show promise in automated software engineering, yet their guarantee of correctness is frequently undermined by erroneous

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Railway Artificial Intelligence Learning Benchmark (RAIL-BENCH): A Benchmark Suite for Perception in the Railway Domain

DGX agent

arXiv:2604.22507v1 Announce Type: new Abstract: Automated train operation on existing railway infrastructure requires robust camera-based perception, yet the railway domain lacks public benchmark suit

model-releasesarxiv-cs-cv
27 Apr 2026
Research

Removing Sandbagging in LLMs by Training with Weak Supervision

DGX agent

arXiv:2604.22082v1 Announce Type: cross Abstract: As AI systems begin to automate complex tasks, supervision increasingly relies on weaker models or limited human oversight that cannot fully verify ou

researcharxiv-cs-ai
27 Apr 2026
Research

SAMIDARE: Advanced Tracking-by-Segmentation for Dense Scenarios

DGX agent

arXiv:2604.22162v1 Announce Type: new Abstract: Automated sports analysis demands robust multi-object tracking (MOT), yet segmentation-based methods often struggle with mask errors and ID switches in

researcharxiv-cs-cv
27 Apr 2026
Agents

Sound Agentic Science Requires Adversarial Experiments

DGX agent

arXiv:2604.22080v1 Announce Type: new Abstract: LLM-based agents are rapidly being adopted for scientific data analysis, automating tasks once limited by human time and expertise. This capability is o

agentsarxiv-cs-ai
27 Apr 2026
Model Releases

A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair

DGX agent

arXiv:2604.21579v1 Announce Type: cross Abstract: LLM-based automated program repair (APR) techniques have shown promising results in reducing debugging costs. However, prior results can be affected b

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Fairness under uncertainty in sequential decisions

DGX agent

arXiv:2604.21711v1 Announce Type: cross Abstract: Fair machine learning (ML) methods help identify and mitigate the risk that algorithms encode or automate social injustices. Algorithmic approaches al

safetyarxiv-cs-ai
24 Apr 2026
Agents

PosterForest: Hierarchical Multi-Agent Collaboration for Scientific Poster Generation

DGX agent

arXiv:2508.21720v2 Announce Type: replace Abstract: Automating scientific poster generation requires hierarchical document understanding and coherent content-layout planning. Existing methods often re

agentsarxiv-cs-ai
24 Apr 2026
Model Releases

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

DGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

model-releasesarxiv-cs-lg
24 Apr 2026
Safety

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

DGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

safetyarxiv-cs-lg
24 Apr 2026
Safety

Tumor-anchored deep feature random forests for out-of-distribution detection in lung cancer segmentation

DGX agent

arXiv:2512.08216v3 Announce Type: replace-cross Abstract: Accurate segmentation of lung tumors from 3D computed tomography (CT) scans is essential for automated treatment planning and response assessm

safetyarxiv-cs-cv
24 Apr 2026
Model Releases

ActuBench: A Multi-Agent LLM Pipeline for Generation and Evaluation of Actuarial Reasoning Tasks

DGX agent

arXiv:2604.20273v1 Announce Type: new Abstract: We present ActuBench, a multi-agent LLM pipeline for the automated generation and evaluation of advanced actuarial assessment items aligned with the Int

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

AgentLens: Adaptive Visual Modalities for Human-Agent Interaction in Mobile GUI Agents

DGX agent

arXiv:2604.20279v1 Announce Type: cross Abstract: Mobile GUI agents can automate smartphone tasks by interacting directly with app interfaces, but how they should communicate with users during executi

agentsarxiv-cs-ai
23 Apr 2026
Agents

AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite

DGX agent

arXiv:2510.21652v2 Announce Type: replace Abstract: AI agents hold the potential to revolutionize scientific productivity by automating literature reviews, replicating experiments, analyzing data, and

agentsarxiv-cs-ai
23 Apr 2026
Research

Auto-Unrolled Proximal Gradient Descent: An AutoML Approach to Interpretable Waveform Optimization

DGX agent

arXiv:2603.17478v2 Announce Type: replace-cross Abstract: This study explores the combination of automated machine learning (AutoML) with model-based deep unfolding (DU) for optimizing wireless beamfo

researcharxiv-cs-ai
23 Apr 2026
← Previous
1…2930313233…83
Next →