AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
5 Aug 2026

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

Model ReleasesDGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

HyperFL: Query-Adaptive Representation Learning for Software Fault Localization

Model ReleasesDGX agent

arXiv:2608.02967v1 Announce Type: cross Abstract: Software fault localization identifies the code locations responsible for reported issues and is a fundamental step toward automated debugging and pro

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2608.02611v1 Announce Type: cross Abstract: Automating GPU kernel optimization remains difficult in practice: generated variants can violate correctness constraints, runtime measurements are noi

MuEvo: LLM-Driven Evolution of Multi-Heuristic Ensemble

ResearchDGX agent

arXiv:2608.03636v1 Announce Type: cross Abstract: Large language model-based automated heuristic design (LLM-AHD) has shown strong potential in discovering effective heuristics for combinatorial optim

Poisoning Prompt-Guided Sampling in Video Large Language Models

SafetyDGX agent

arXiv:2509.20851v2 Announce Type: replace Abstract: Video Large Language Models (VideoLLMs) are increasingly deployed as automated moderators on user-generated video platforms, where a few unwatched s

Reddit is introducing a new moderator: AI

IndustryDGX agent

Reddit is enlisting AI to help moderate new subreddits - and eventually the rest of site. The company is introducing automated moderation tools that rely on LLMs to help mods manage their communities,

Towards Reliable and Reproducible Fetal Brain Biometry: A Deep Learning Approach Using MRI

ResearchDGX agent

arXiv:2608.03724v1 Announce Type: new Abstract: Fetal brain biometry is essential for quantitative assessment of brain development, supporting gestational age estimation, developmental monitoring, and

Unequal Verdicts: Investigating Gender Bias in LLM-Based Fake News Detection

Model ReleasesDGX agent

arXiv:2608.03627v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for automated fact-checking, yet their susceptibility to gender bias in this context remains underexp

VeriTrace: Human-Like Temporal Exploration Completes Agentic Action Space

Model ReleasesDGX agent

arXiv:2608.02878v1 Announce Type: new Abstract: Large language models have shown promise for automated Verilog RTL generation, yet state-of-the-art multi-agent systems plateau at ~95% accuracy on stan

4 Aug 2026

ARMOR: A Robust Self-Supervised Framework for Root Cause Analysis in Microservices under Missing Modality

SafetyDGX agent

arXiv:2603.25538v3 Announce Type: replace Abstract: Automated incident management is critical for microservice reliability. While recent unified frameworks leverage multimodal data for joint optimizat

DeBERTa-Sentinel: Toward Transparent and Trustworthy Detection of AI-Generated Text

Model ReleasesDGX agent

arXiv:2608.01046v1 Announce Type: new Abstract: The rapid spread of large language models (LLMs) across the web raises concerns about misinformation, academic integrity, automated content manipulation

Deep Learning for Retinal Degeneration Assessment: A Comprehensive Analysis of the MARIO Challenge

Model ReleasesDGX agent

arXiv:2506.02976v4 Announce Type: replace Abstract: The MARIO challenge, held at MICCAI 2024, focused on advancing the automated detection and monitoring of age-related macular degeneration (AMD) thro

DeepSurvey-Bench: Evaluating Academic Value of Automatically Generated Scientific Surveys

Model ReleasesDGX agent

arXiv:2601.15307v2 Announce Type: replace-cross Abstract: The rapid development of automated survey generation technology has made it increasingly important to establish a comprehensive benchmark to e

DS@GT ARC at MEDIQA-CORE-Task-1 2026: Trimodal Model Fusion with Task-Specific Gates for Brain Tumor Subtype Classification

Model ReleasesDGX agent

arXiv:2608.00086v1 Announce Type: new Abstract: Brain tumor diagnosis is a time-sensitive process in which patients may wait weeks for a finalized pathology report. This problem motivates automated sy

Fairness Auditing: Lower Bounds on Company Manipulation

SafetyDGX agent

arXiv:2608.00568v1 Announce Type: new Abstract: Fairness audits are increasingly mandated in high-stakes applications such as hiring, lending, and automated decision-making. Recent work has establishe

HappyRobot raises 150M at 1.2B valuation to bring AI agents to critical enterprise work

ApplicationsDGX agent

HappyRobot Inc., a San Francisco-based artificial intelligence startup that automates enterprise operations, said today it raised 150 million, led by Prysm Capital and co-led by Eurazeo, bringing the

How Deutsche Bank unlocked agility with an API-ready ecosystem

SafetyDGX agent

When people think about digital transformation in banking, they often focus on the visible results: mobile apps and new digital services. But there's an invisible infrastructure making all these servi

MDWD: A Street-Level Dataset for Municipal Solid Waste Detection in Dense Urban Environments

Model ReleasesDGX agent

arXiv:2608.00257v1 Announce Type: new Abstract: Automated visual monitoring of urban environments is a growing Computer Vision research area, but municipal solid waste detection remains under-represen

PackingGPT: 3D Packing Agent for Real Furniture in Last-Mile Delivery

Model ReleasesDGX agent

arXiv:2608.01427v1 Announce Type: new Abstract: 3D bin packing rectangular items into standardised containers to maximise space utilisation under geometric shipping automation. Loading a furniture pur

Quick on the Uptake: Eliciting Implicit Intents from Human Demonstrations for Personalized Mobile-Use Agents

SafetyDGX agent

arXiv:2508.08645v3 Announce Type: replace Abstract: As multimodal large language models advance rapidly, the automation of mobile tasks has become increasingly feasible through the use of mobile-use a

ReMiX-MAE: Learning Missing-Channel Cross-Modal Representations from RGB-Only Clinical Facial Videos for Sympathetic-Mediated Pain Assessment

ApplicationsDGX agent

arXiv:2608.02561v1 Announce Type: new Abstract: Automated pain assessment in real clinics is limited by scarce clinically grounded facial video data with weak labels (often sequence-level self-report)

Rethinking PPG-based Sleep Staging: Datasets, Metrics, and Benchmarks

ResearchDGX agent

arXiv:2608.00943v1 Announce Type: cross Abstract: Automated sleep staging assigns discrete stage labels to successive time epochs throughout an overnight recording; conventionally each window spans at

Unleashing the Potential of Large Language Models: A Blueprint for Real-Time, Enterprise-Ready Deployments

SafetyDGX agent

arXiv:2608.00419v1 Announce Type: cross Abstract: Large language models deployed in real-time, regulated settings face knowledge staleness, catastrophic forgetting, hallucination, and weak feedback lo

3 Aug 2026

CalibratedRubric: Task-Adaptive Rubric Banks for Open-Ended LLM Evaluation

ApplicationsDGX agent

arXiv:2607.29252v1 Announce Type: cross Abstract: Reliable evaluation of open-ended LLM outputs requires fine-grained rubrics, yet expert curation is costly and difficult to scale. Existing automated

Can Large Language Models Derive New Knowledge? A Dynamic Benchmark for Biological Knowledge Discovery

Model ReleasesDGX agent

arXiv:2603.03322v2 Announce Type: replace-cross Abstract: Recent advancements in Large Language Model (LLM) agents have demonstrated remarkable potential in automatic knowledge discovery. However, rig

Real-world mainframe modernization with AI: A safe, scalable path from mainframe to cloud

Model ReleasesDGX agent

For too long, enterprises with legacy mainframe estates have been faced with a high-stakes dilemma: continue maintaining their mainframes, essentially kicking the modernization can down the road (they

Unanticipated Effects of Generative AI on Expertise Pathways and Performance Perception in System Administration

SafetyDGX agent

arXiv:2607.28650v1 Announce Type: cross Abstract: While industry discourse often emphasizes immediate productivity gains and frames GenAI primarily as a tool for automation, the integration of GenAI i

31 Jul 2026

AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining

Local AiDGX agent

arXiv:2607.26642v1 Announce Type: new Abstract: Automated alpha mining has increasingly adopted large language model (LLM) agents for factor generation and iterative discovery. However, existing LLM-b

Cybersecurity Detection Classification with Reasoning-enabled Language Models

ResearchDGX agent

arXiv:2607.28460v1 Announce Type: new Abstract: A major issue in Security Operations Centers (SOCs) is alert fatigue, as the number of detections reported is more than staff can triage in a given day.

Efficient LLMs with AMP: Attention Heads and MLP Pruning

Model ReleasesDGX agent

arXiv:2504.21174v2 Announce Type: replace Abstract: Deep learning drives a new wave in computing systems and triggers the automation of increasingly complex problems. In particular, Large Language Mod

How does downsampling affect needle electromyography signals? A generalisable workflow for understanding downsampling effects on high-frequency time series

ResearchDGX agent

arXiv:2601.10191v2 Announce Type: replace Abstract: Automated analysis of needle electromyography (nEMG) signals is emerging as a tool to support the detection of neuromuscular diseases (NMDs), yet th

Model-Driven Requirements Configuration with Three-Valued Uncertainty Scoring

AgentsDGX agent

arXiv:2607.26220v1 Announce Type: cross Abstract: Context: Large Language Models (LLMs) offer natural-language flexibility for automated requirements elicitation but frequently generate structurally i

The MADRS Pipeline: Supporting Depression Assessment in Clinical Trials

ResearchDGX agent

arXiv:2607.28190v1 Announce Type: new Abstract: Depression is a major mental disorder for which diagnosis relies primarily on clinical assessments. Automated methods to support its detection via the p

Toward Multi-Modal Deep Learning for Pulmonary Disease Classification: A Texture-Based Machine Learning Pilot Study on Public Chest X-Ray Data

ResearchDGX agent

arXiv:2607.27286v1 Announce Type: cross Abstract: Automated classification of pulmonary disease from chest radiographs is a widely studied application of machine learning in medical imaging. This pape

UrbanDS: A Graph-Guided LLM Multi-Agent System for Data-Intensive Urban Tasks

Model ReleasesDGX agent

arXiv:2607.26724v1 Announce Type: new Abstract: Large language model (LLM) agents have been widely applied in automating data science tasks. However, existing methods typically rely on a limited set o

What Makes Deep Learning Work for Traditional Chinese Medicine Tongue Diagnosis? A Comprehensive Ablation Study

Model ReleasesDGX agent

arXiv:2607.28148v1 Announce Type: new Abstract: Deep learning has shown promise for automated tongue diagnosis in traditional Chinese medicine (TCM), yet the design space remains underexplored. We con

30 Jul 2026

Deep Expert Injection for Anchoring Retinal VLMs with Domain-Specific Knowledge

ResearchDGX agent

arXiv:2603.07131v4 Announce Type: replace Abstract: Large Vision Language Models (LVLMs) show immense potential for automated ophthalmic diagnosis. However, their clinical deployment is severely hinde

Do Methods Support the Claims? Intra-Paper Verification for Peer Review

SafetyDGX agent

arXiv:2607.26066v1 Announce Type: new Abstract: The growing volume of scientific submissions has motivated interest in using large language models (LLMs) to assist peer review. Existing automated nove

Large-Scale ChatBot Validation Through Customer Digital Twin Simulations

SafetyDGX agent

arXiv:2607.26060v1 Announce Type: new Abstract: LLM-based chatbots are transforming customer service in regulated domains such as banking, but scalable and cost-effective validation remains a critical

Rethinking Clinical Relevance in Chest X-ray Machine Learning: How Evaluation References Define Performance

SafetyDGX agent

arXiv:2607.26333v1 Announce Type: cross Abstract: Chest X-ray (CXR) machine learning relies heavily on automated evaluation using reference standards that aim to approximate clinical judgment. However

ScratchSim: A Procedural Synthetic Data Pipeline for Surface Scratch Detection

Local AiDGX agent

arXiv:2607.27065v1 Announce Type: new Abstract: While automated defect detection such as the detection of surface scratched is an important aspect in industrial quality control, the scarcity of annota

The Rise of AI in Weather and Climate Information and its Impact on Global Inequality

ApplicationsDGX agent

arXiv:2603.05710v2 Announce Type: replace-cross Abstract: AI development's current trajectory risks automating and amplifying the North-South divide in the global climate information system. Frontier

29 Jul 2026

A Machine-Learning-Based Gas Lift Optimization Workflow for Unconventional Fields

ApplicationsDGX agent

arXiv:2607.25885v1 Announce Type: cross Abstract: In this paper, we present an automated data-driven workflow using Machine Learning (ML) for gas lift optimization in unconventional fields. This workf

AlphaCrafter: Harnessing Multi-Agent Workflows for Cross-Sectional Quantitative Trading

SafetyDGX agent

arXiv:2605.05580v2 Announce Type: replace Abstract: Quantitative trading agents have demonstrated substantial promise in automating factor discovery, signal aggregation, and portfolio execution. Howev

Decompose and Reorganize: Planning with Primitives and Visuomotor Policies Learned from Demonstrations

ApplicationsDGX agent

arXiv:2607.25397v1 Announce Type: new Abstract: Successfully automating dexterous, long-horizon robotic manipulation requires frameworks capable of both high-level reasoning and fine-grained execution

Diff2DGS: Reliable Reconstruction of Occluded Surgical Scenes via 2D Gaussian Splatting

ResearchDGX agent

arXiv:2602.18314v2 Announce Type: replace Abstract: Real-time reconstruction of deformable surgical scenes is vital for advancing robotic surgery, improving intraoperative guidance, and enabling autom

DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Multilingual Numerical Claim Verification

ResearchDGX agent

arXiv:2607.25069v1 Announce Type: cross Abstract: Automated verification of numerical claims is a challenging problem, as it requires both language understanding and quantitative reasoning. This paper

Faces of Fairness: Examining Bias in Facial Expression Recognition Datasets and Models

SafetyDGX agent

arXiv:2502.11049v3 Announce Type: replace Abstract: Automated Facial Expression Recognition (FER), involves two critical aspects: data and model design. Both significantly influence bias and fairness

How agentic AI can help telecom finance teams protect the margin when every moment matters

AgentsDGX agent

Agentic AI empowers telecom finance teams to preserve margins by rapidly detecting and preventing revenue leakage across billing, provisioning, and cost‑allocation processes. The technology automates

Interactive Reward Agent: GUI Task Evaluation via Environment-State Verification

Model ReleasesDGX agent

arXiv:2607.25904v1 Announce Type: new Abstract: Graphical user interface task evaluation aims to determine whether a GUI agent has successfully completed a user instruction. Automated GUI task evaluat

RIDGE: An Autonomous Framework for Validation and Method Discovery in LLM-Generated Option Pricing

Model ReleasesDGX agent

arXiv:2607.25199v1 Announce Type: cross Abstract: Automated code generation is becoming an important tool in quantitative finance, where large language models can generate option pricing implementatio

Track-Leakage-Free Hold-Out Self-Validation for Photogrammetric Reconstruction: Protocol, Sensitivity, and Limits

Model ReleasesDGX agent

arXiv:2607.24852v1 Announce Type: new Abstract: Automated photogrammetric inspection emits metric measurements from a 3D reconstruction whose own correctness is normally unknown without an external su

Understanding User Experiences of Computer Use Agents: Design Space and Opportunities for Building Agent UX Prototypes

AgentsDGX agent

arXiv:2510.04452v3 Announce Type: replace-cross Abstract: Computer use agents (or 'agents') are generative AI that automates actions within user interfaces from user commands. Current research focuses

What’s new in Gemini Enterprise Agent Platform

Model ReleasesDGX agent

Since we launched Gemini Enterprise Agent Platform a few months ago, we’ve seen inspiring progress from businesses and builders alike. To stir up development, we’ve also shared 13 demos that can walk

28 Jul 2026

A Modern ConvNet for Solar Filament Detection

ResearchDGX agent

arXiv:2607.24525v1 Announce Type: cross Abstract: Automated solar filament detection using deep learning faces several challenges. Semantic segmentation of solar filaments is a complicated multiscale

Bringing Conversational Analytics to your entire data ecosystem

Model ReleasesDGX agent

Increasing the adoption of generative AI across the enterprise requires you to do more than deploy a generic chatbot with a custom wrapper. Interacting with business-critical databases demands absolut

Co-Harness: Co-Evolving Harnesses and Model Weights for LLM Agents

Local AiDGX agent

arXiv:2607.22688v1 Announce Type: new Abstract: Post-training agents for automated AI research requires optimizing not only model parameters, but also the runtime harness that shapes how research traj

CodeEvo: Interaction-Driven Synthesis of Code-centric Data through Hybrid and Iterative Feedback

AgentsDGX agent

arXiv:2507.22080v2 Announce Type: replace-cross Abstract: Acquiring high-quality instruction-code pairs is essential for training Large Language Models for code generation. While automated synthesis h

Codifying the Judge: Scalable Evaluation via Program Distillation

ResearchDGX agent

arXiv:2607.22561v1 Announce Type: new Abstract: LLM-as-a-judge has become the standard for automated evaluation, but it suffers from high cost, significant latency, and opaque decisions -- limitations

Cost-Aware Recovery-Pathway Identification and Bayesian Optimization for Autonomous Materials Discovery

Model ReleasesDGX agent

arXiv:2607.23896v1 Announce Type: new Abstract: Autonomous laboratories automate experimental execution, but a campaign must also decide which recovery pathway merits optimization. We formulate this a

← Previous
1…1920212223…83
Next →