AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,979 results
Model Releases

What to Format and How: A Benchmark and Workflow Approach for Document Formatting

DGX agent

arXiv:2606.01936v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have opened up new possibilities for automated document formatting. However, real-world formatting often

model-releasesarxiv-cs-cl
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Automatically Attacking Software Reverse Engineering AI Agents

DGX agent

arXiv:2605.30667v1 Announce Type: cross Abstract: Software tools for reverse engineering executable binary files, such as Ghidra, enable malware analysts to safely conduct robust static analysis witho

agentsarxiv-cs-ai
1 Jun 2026
Agents

@bentannyhill @Zach_Kamran Langsmith Engine is the 'Full Self-Driving' moment for AI Engineering

DGX agent

Harrison Chase compares LangSmith Engine to Tesla's 'Full Self-Driving' moment, suggesting it represents a significant advancement in AI engineering capabilities and automation. The post indicates Lan

agentsharrison-chase--x
1 Jun 2026
Research

Beyond Tokens: Enhancing RTL Quality Estimation via Structural Graph Learning

DGX agent

arXiv:2508.18730v2 Announce Type: replace Abstract: Estimating the quality of register transfer level (RTL) designs is crucial in the electronic design automation (EDA) workflow, as it enables instant

researcharxiv-cs-lg
1 Jun 2026
Research

Controllable Lung Nodule Synthesis via Histogram-Regularized Latent Diffusion Models

DGX agent

arXiv:2605.30631v1 Announce Type: cross Abstract: While automated diagnosis systems have achieved remarkable success in computed tomography (CT)-based lung cancer screening, their development remains

researcharxiv-cs-ai
1 Jun 2026
Safety

Cost-aware Stopping for Bayesian Optimization

DGX agent

arXiv:2507.12453v5 Announce Type: replace Abstract: In automated machine learning, scientific discovery, and other applications of Bayesian optimization, deciding when to stop evaluating expensive bla

safetyarxiv-cs-lg
1 Jun 2026
Safety

Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory

DGX agent

arXiv:2602.00521v2 Announce Type: replace Abstract: While LLM-as-a-Judge is widely used in automated evaluation, existing validation practices primarily operate at the level of observed outputs, offer

safetyarxiv-cs-ai
1 Jun 2026
Agents

have manually read 1000s traces since joining LangChain! it’s a great way to learn and understand your agent but completely infeasible to do…

DGX agent

have manually read 1000s traces since joining LangChain! it’s a great way to learn and understand your agent but completely infeasible to do at agent scale 🙃 engine helps us automate that process so h

agentsharrison-chase--x
1 Jun 2026
Model Releases

Knowledge Boundary Probing and Demand-Guided Intervention for LLM-Based Power System Code Generation

DGX agent

arXiv:2605.31478v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to automate power-system analysis, but many utilities and energy-research labs require on-premise s

model-releasesarxiv-cs-cl
1 Jun 2026
Research

Multilingual and Cross-Lingual Citation Needed Detection on Wikipedia for Lower-Resource Languages

DGX agent

arXiv:2605.31136v1 Announce Type: new Abstract: In automated fact-checking (AFC), check-worthiness detection identifies claims requiring verification based on domain-specific criteria. On Wikipedia, t

researcharxiv-cs-cl
1 Jun 2026
Hardware

NVIDIA Factory Operations Blueprint Gives Factories a New AI Brain

DGX agent

As factories move from isolated automation to plant-wide intelligence, manufacturers need AI systems that can connect live machine signals, quality systems, work instructions and operational alerts in

hardwarenvidia-blog
1 Jun 2026
Safety

REAL: Regression-Aware Reinforcement Learning for LLM-as-a-Judge

DGX agent

arXiv:2603.17145v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as automated evaluators that assign numeric scores to model outputs, a paradigm known a

safetyarxiv-cs-ai
1 Jun 2026
Safety

Reinterpreting Safety Thresholds as Neuron Spiking Thresholds

DGX agent

arXiv:2605.30368v1 Announce Type: cross Abstract: Surrogate Safety Measures (SSMs) are extensively utilised in the evaluation of traffic risk in automated driving contexts. However, the majority of SS

safetyarxiv-cs-ai
1 Jun 2026
Agents

Stop manually triaging agent failures. Let LangSmith Engine fix it.

DGX agent

LangSmith Engine is a tool designed to automatically diagnose and resolve agent failures, eliminating the need for manual troubleshooting and triage. The feature appears to leverage automated analysis

agentsharrison-chase--x
1 Jun 2026
Agents

Agentic AI success helps UiPath swing to a profit, but investors weren’t impressed

DGX agent

Business automation software company UiPath Inc. delivered mixed results in its latest quarter, posting a solid revenue beat but falling short on earnings — but it did at least manage to return to pro

agentssiliconangle
29 May 2026
Agents

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

DGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

agentsarxiv-cs-ai
29 May 2026
Safety

Inferring Code Correctness from Specification

DGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

safetyarxiv-cs-ai
29 May 2026
Model Releases

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

DGX agent

arXiv:2605.29523v1 Announce Type: new Abstract: Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barri

model-releasesarxiv-cs-lg
29 May 2026
Safety

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

DGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

safetyarxiv-cs-ai
29 May 2026
Model Releases

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions

DGX agent

arXiv:2605.29862v1 Announce Type: cross Abstract: AI-driven respiratory sound classification (RSC) is promising for automated pulmonary disease detection, yet multi-site deployment is hindered by inte

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring

DGX agent

arXiv:2605.29852v1 Announce Type: new Abstract: Histological scoring is essential for diagnosing Non-Alcoholic Fatty Liver Disease (NAFLD), yet its automation remains challenging due to the high annot

model-releasesarxiv-cs-cv
29 May 2026
Safety

PRO-CUA: Process-Reward Optimization for Computer Use Agents

DGX agent

arXiv:2605.29119v1 Announce Type: new Abstract: Computer use agents (CUAs) have shown strong potential for automating complex digital workflows, yet their training remains constrained by costly live e

safetyarxiv-cs-ai
29 May 2026
Model Releases

SoundnessBench: Can Your AI Scientist Really Tell Good Research Ideas from Bad Ones?

DGX agent

arXiv:2605.30329v1 Announce Type: new Abstract: Autonomous AI research agents aim to accelerate scientific discovery by automating the research pipeline, from hypothesis generation to peer review. How

model-releasesarxiv-cs-lg
29 May 2026
Local Ai

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

DGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

local-aiarxiv-cs-ai
29 May 2026
Research

Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bioacoustic Data

DGX agent

arXiv:2502.20838v3 Announce Type: replace-cross Abstract: Passive acoustic monitoring (PAM) systems generate continuous recordings spanning months, yet automated bioacoustic analysis of whale calls re

researcharxiv-cs-ai
29 May 2026
Applications

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

DGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

applicationsarxiv-cs-cl
29 May 2026
Model Releases

Adaptive Cost-Efficient Evaluation for Reliable Patent Claim Generation

DGX agent

arXiv:2604.04295v3 Announce Type: replace Abstract: Automated patent claim validation demands low error tolerance. However, existing approaches face a rigidity-resource dilemma: lightweight encoders c

model-releasesarxiv-cs-cl
28 May 2026
Agents

AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation

DGX agent

arXiv:2605.28655v1 Announce Type: new Abstract: Scientific research proceeds through iterative cycles of hypothesis generation, experiment design, execution, and revision. AI agents can automate parts

agentsarxiv-cs-ai
28 May 2026
Agents

CircuitLM: A Multi-Agent LLM-Aided Design Framework for Generating Circuit Schematics from Natural Language Prompts

DGX agent

arXiv:2601.04505v3 Announce Type: replace Abstract: Generating accurate circuit schematics from high-level natural language descriptions remains a persistent challenge in electronic design automation

agentsarxiv-cs-ai
28 May 2026
Safety

Mining Multi-Modality Spatio-Temporal Cues for Video Important Person Identification

DGX agent

arXiv:2605.28604v1 Announce Type: cross Abstract: Identifying key individuals in video scenes is essential for applications such as automated video editing and intelligent surveillance. Current method

safetyarxiv-cs-ai
28 May 2026
Model Releases

MolLingo: Molecule-Native Representations for LLM-Powered Scientific Agents

DGX agent

arXiv:2605.27853v1 Announce Type: new Abstract: We present MolLingo, a multi-agent system that emulates the reasoning process of a chemist to automate molecular design. Existing LLM-based approaches e

model-releasesarxiv-cs-ai
28 May 2026
Agents

One of the best ways to learn what LangSmith Engine is capable of is to talk to the team that built it. Join @bentannyhill for a live sessio…

DGX agent

One of the best ways to learn what LangSmith Engine is capable of is to talk to the team that built it. Join @bentannyhill for a live session on June 11th and see how your team can automate the agent

agentsharrison-chase--x
28 May 2026
Research

Poison with Style: A Practical Poisoning Attack on Code Large Language Models

DGX agent

arXiv:2605.27631v1 Announce Type: cross Abstract: Code Large Language Models (CLLMs) serve as the core of modern code agents, enabling developers to automate complex software development tasks. In thi

researcharxiv-cs-lg
28 May 2026
Agents

RAG-Coding: Enhancing LLM Medical Coding with Structured External Knowledge

DGX agent

arXiv:2605.27377v1 Announce Type: cross Abstract: We present RAG-Coding, an agentic method for automated ICD-10-CM coding. RAG-Coding orchestrates four large language model (LLM) agents and grounds th

agentsarxiv-cs-ai
28 May 2026
Research

Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis

DGX agent

arXiv:2512.14561v2 Announce Type: replace Abstract: Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compare

researcharxiv-cs-cl
27 May 2026
Applications

Cisco and OpenAI redefine enterprise engineering with Codex

DGX agent

Cisco and OpenAI partnered to integrate OpenAI's Codex AI model into Cisco's enterprise engineering tools to enhance software development and automation capabilities. The collaboration aims to improve

applicationsopenai
27 May 2026
Agents

Datasets for Lane Detection in Autonomous Driving: A Comprehensive Review

DGX agent

arXiv:2504.08540v2 Announce Type: replace Abstract: Accurate lane detection is essential for automated driving, enabling safe and reliable vehicle navigation across a variety of road scenarios. Numero

agentsarxiv-cs-cv
27 May 2026
Agents

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLM…

DGX agent

For future-proof, build AI that's composable. Regardless of what you use, all these should be composable, iterative, and customizable: - LLMs - Evals - Automations - MCP/CLI tools - Skills/Memory/Cont

agentsdair-ai--x
27 May 2026
Model Releases

SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks?

DGX agent

arXiv:2605.26548v1 Announce Type: cross Abstract: Large language models (LLMs) now support automated software security tasks, including vulnerability discovery and proof-of-concept (PoC) generation. E

model-releasesarxiv-cs-lg
27 May 2026
Research

Sleep-stage efficient classification using a lightweight self-supervised model

DGX agent

arXiv:2605.26295v1 Announce Type: new Abstract: Accurate classification of sleep stages is crucial for diagnosing sleep disorders and automating this process can significantly enhance clinical assessm

researcharxiv-cs-cv
27 May 2026
Model Releases

Why Prompt Optimization Works, and Why It Sometimes Doesn't: A Causal-Inspired Edit-Level Analysis

DGX agent

arXiv:2605.26655v1 Announce Type: new Abstract: Automated prompt optimization methods (e.g., DSpy, TextGrad) can substantially improve the performance of large language model (LLM), however, their gen

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

YOLO26-RipeLoc Lite: A lightweight architecture for tomato ripeness detection and picking point localization in greenhouse robotic harvesting

DGX agent

arXiv:2605.27129v1 Announce Type: new Abstract: In greenhouse tomato production, automated harvesting requires accurate detection of ripe tomatoes, ripeness classification, and precise picking-point l

local-aiarxiv-cs-cv
27 May 2026
Agents

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

DGX agent

arXiv:2605.24823v1 Announce Type: new Abstract: Manufacturing has passed through four widely recognized paradigms - mechanization, electrification, programmable automation, and Smart Manufacturing - e

agentsarxiv-cs-ai
26 May 2026
Tutorials

An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods

DGX agent

arXiv:2605.24298v1 Announce Type: cross Abstract: The growing use of Large Language Models (LLMs) for automated code generation has enhanced software development efficiency, but often at the cost of s

tutorialsarxiv-cs-ai
26 May 2026
Applications

By Their Fruits You Will Know Them: Comparing Formalizations of Law by the Decisions They Encode

DGX agent

arXiv:2605.25186v1 Announce Type: cross Abstract: Formalizing legal provisions promises machine-accessible law and automated legal reasoning, and recent LLMs make it tempting to generate such formaliz

applicationsarxiv-cs-ai
26 May 2026
Research

Catching The Correct Answer Trap: Characterising AI Tutor Blind Spots When Analysing Student Reasoning

DGX agent

arXiv:2605.23925v1 Announce Type: cross Abstract: Intelligent tutoring systems increasingly provide automated feedback on student work, but robust feedback requires assessing reasoning, not only final

researcharxiv-cs-ai
26 May 2026
Applications

Generating Legal Commentaries from Case Databases via Retrieval, Clustering, and Generation

DGX agent

arXiv:2605.24534v1 Announce Type: new Abstract: We present a fully automated pipeline that transforms large collections of court decisions into legal commentaries for statutes - without providing any

applicationsarxiv-cs-cl
26 May 2026
Research

Generative AI impacts on intra-urban inequality and skill premium in Beijing

DGX agent

arXiv:2605.25505v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI) is the first automation wave to reach high-cognitive tasks at scale, yet its effects on intra-urban inequal

researcharxiv-cs-ai
26 May 2026
← Previous
1…3233343536…104
Next →