AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,980 results
19 May 2026

CVE-Factory: Scaling Expert-Level Agentic Tasks for Code Security Vulnerability

Model ReleasesDGX agent

arXiv:2602.03012v2 Announce Type: replace-cross Abstract: Evaluating and improving the security capabilities of code agents requires high-quality, executable vulnerability tasks. However, existing wor

Data-driven and distributed governance of building facilities management using decentralized autonomous organization, digital twin, and large language models

AgentsDGX agent

arXiv:2605.16298v1 Announce Type: cross Abstract: While traditional AI and data-driven facilities management approaches have improved building operational efficiency, they remain constrained by centra

Employing Vision-Language Models for Face Image Quality Assessment

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

Enhancing Metacognitive AI: Knowledge-Graph Population with Graph-Theoretic LLM Enrichment

Model ReleasesDGX agent

arXiv:2605.16676v1 Announce Type: new Abstract: Metacognition-the ability to monitor one's own knowledge state, spot gaps, and autonomously fill them--remains largely absent from modern AI. Here, we p

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

SafetyDGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets

Model ReleasesDGX agent

arXiv:2605.18475v1 Announce Type: cross Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. Ho

Generation Navigator: A State-Aware Agentic Framework for Image Generation

SafetyDGX agent

arXiv:2605.17969v1 Announce Type: new Abstract: Despite rapid advances in text-to-image generation, faithfully realizing user intent remains challenging, often requiring manual multi-turn trial and er

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

Model ReleasesDGX agent

arXiv:2507.21035v3 Announce Type: replace Abstract: Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due t

GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning

ResearchDGX agent

arXiv:2605.16371v1 Announce Type: cross Abstract: Large Multimodal Models (LMMs) often struggle with geometric reasoning due to visual hallucinations and a lack of mathematically precise Chain-of-Thou

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

SafetyDGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

HPC-LLM: Practical Domain Adaptation and Retrieval-Augmented Generation for HPC Support

Model ReleasesDGX agent

arXiv:2605.16347v1 Announce Type: new Abstract: Modern scientific research increasingly depends on High-Performance Computing (HPC) infrastructures, yet many researchers face significant operational b

Human-Certified Module Repositories for the AI Age

ResearchDGX agent

arXiv:2603.02512v4 Announce Type: replace-cross Abstract: Human-Certified Module Repositories (HCMRs) are introduced in this work as a new architectural model for constructing trustworthy software in

LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters

SafetyDGX agent

arXiv:2605.12012v2 Announce Type: replace Abstract: Public-sector legal departments in the Netherlands face acute staff shortages, increased case volumes, and increased pressure to meet regulatory com

Leveraging Multimodal Self-Consistency Reasoning in Coding Motivational Interviewing for Alcohol Use Reduction

ResearchDGX agent

arXiv:2605.12987v2 Announce Type: replace Abstract: BACKGROUND: Coding Motivational Interviewing (MI) sessions is essential for understanding client behaviors and predicting outcomes, but it requires

Leveraging Unsupervised Learning for Cost-Effective Visual Anomaly Detection

ResearchDGX agent

arXiv:2409.15980v2 Announce Type: replace-cross Abstract: Traditional machine learning-based visual inspection systems require extensive data collection and repetitive model training to improve accura

LinAlg-Bench: A Forensic Benchmark Revealing Structural Failure Modes in LLM Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2605.16675v1 Announce Type: new Abstract: We introduce LinAlg-Bench, a diagnostic benchmark evaluating 10 frontier large language models on structured linear algebra computation across a strict

LLM-Safety Evaluations Lack Robustness

SafetyDGX agent

arXiv:2503.02574v2 Announce Type: replace-cross Abstract: In this paper, we argue that current safety alignment research efforts for large language models are hindered by many intertwined sources of n

LLMs for automatic annotation of Mandarin narrative transcripts

Local AiDGX agent

arXiv:2605.17205v1 Announce Type: new Abstract: Linguistic annotation of transcribed speech is essential for research in language acquisition, language disorders, and sociolinguistics, yet remains lab

Machine Learning Enabled Graph Analysis of Particulate Composites: Application to Solid-state Battery Cathodes

TutorialsDGX agent

arXiv:2512.16085v2 Announce Type: replace-cross Abstract: Particulate composites underpin many solid-state chemical and electrochemical systems, where microstructural features such as multiphase bound

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

Model ReleasesDGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

Model ReleasesDGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis

ResearchDGX agent

arXiv:2605.17539v1 Announce Type: new Abstract: Combinatorial optimization (CO) underlies decision-making from logistics to chip design, where infeasible solutions are operationally unusable and small

MHMamba: Multi-Head Mamba for 3D Brain Tumor Segmentation

ResearchDGX agent

arXiv:2605.16464v1 Announce Type: cross Abstract: Brain tumors exhibit high heterogeneity in morphology and multimodal contrast, making manual slice-by-slice de lineation time-consuming and experience

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

Model ReleasesDGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your …

Model ReleasesDGX agent

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your behalf, under your direction. It seamlessly integrates with

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

Model ReleasesDGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

Perovskite-R1: a domain-specialized large language model for intelligent discovery of precursor additives and experimental design

ApplicationsDGX agent

arXiv:2507.16307v2 Announce Type: replace-cross Abstract: Perovskite solar cells (PSCs) have rapidly emerged as a leading contender in next-generation photovoltaic technologies, owing to their excepti

PopPy: Opportunistically Exploiting Parallelism in Python Compound AI Applications

ApplicationsDGX agent

arXiv:2605.18697v1 Announce Type: cross Abstract: Compound AI applications, which compose calls to ML models using a general-purpose programming language like Python, are widely used for a variety of

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

Model ReleasesDGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

Principal Component Analysis for Lunar Crater Detection

ResearchDGX agent

arXiv:2605.17125v1 Announce Type: new Abstract: Optical navigation is a critical component for lunar orbiter and lander missions. Image-based crater identification has emerged as a promising technolog

PromptDecipher: Supporting AI Tutor Authoring Through Editable Simulated Interactions

TutorialsDGX agent

arXiv:2605.16605v1 Announce Type: cross Abstract: Chatbots have long been explored as tools to support learning, and recent advances in large language models have significantly expanded the availabili

QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI

SafetyDGX agent

arXiv:2605.17382v1 Announce Type: new Abstract: The rapid progress of generative artificial intelligence has exposed fundamental limitations in existing evaluation methodologies, particularly for open

RadGame: An AI-Powered Platform for Radiology Education

Local AiDGX agent

arXiv:2509.13270v2 Announce Type: replace-cross Abstract: We introduce RadGame, an AI-powered gamified platform for radiology education that targets two core skills: localizing findings and generating

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

SafetyDGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

Robust and Resilient Soft Robotic Object Insertion with Compliance-Enabled Contact Formation and Failure Recovery

ResearchDGX agent

arXiv:2509.17666v2 Announce Type: replace Abstract: Object insertion tasks are prone to failure under pose uncertainty and environmental variation, often requiring manual fine-tuning or controller ret

StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting

Model ReleasesDGX agent

arXiv:2605.17309v1 Announce Type: cross Abstract: We present StyleText, a large-scale dataset and benchmark for localized scene-text inpainting with style preservation. StyleText contains 28,518 image

Sustainability via LLM Right-sizing

Model ReleasesDGX agent

arXiv:2504.13217v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have become increasingly embedded in organizational workflows. This has raised concerns over their energy consump

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

Towards Robust Argumentative Essay Understanding via TIDE: An Interactive Framework with Trial and Debate

ResearchDGX agent

arXiv:2605.17247v1 Announce Type: new Abstract: Argumentative essays serve as a vital medium for assessing critical thinking and reasoning skills, yet there is limited works on accurately understandin

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

Model ReleasesDGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

Understanding the modern cybercrime landscape

ApplicationsDGX agent

Throughout 2025, HPE observed significant changes in how cybercriminals operate. Analyzing real-world threats, our HPE Threat Labs highlighted an industrialization of the cyber criminals’ methods in i

VeriHGN: Heterogeneous Graph-Based Congestion Prediction for Chip Layout Verification

ResearchDGX agent

arXiv:2603.11075v2 Announce Type: replace-cross Abstract: As Very Large Scale Integration (VLSI) designs continue to scale in size and complexity, layout verification has become a central challenge in

WavFlow: Audio Generation in Waveform Space

Model ReleasesDGX agent

arXiv:2605.18749v1 Announce Type: cross Abstract: Modern audio generation predominantly relies on latent-space compression, introducing additional complexity and potential information loss. In this wo

YawDD+: Frame-level Annotations for Accurate Yawn Prediction

Local AiDGX agent

arXiv:2512.11446v3 Announce Type: replace Abstract: Driver fatigue remains a leading cause of road accidents, responsible for 24% of crashes. While yawning serves as an early behavioral indicator of f

You can ask Cursor to fix bugs, add features, update tests, or investigate a task described in the work item. Learn more: http://cursor.com/…

TutorialsDGX agent

Cursor is an AI tool that can assist with software development tasks including bug fixes, feature additions, test updates, and task investigation based on work item descriptions. The platform enables

18 May 2026

A Cascaded Generative Approach for e-Commerce Recommendations

ApplicationsDGX agent

arXiv:2605.11118v2 Announce Type: replace Abstract: Personalized storefronts in large e-commerce marketplaces are often assembled from many independent components: static themes per page section ('pla

Adesua: Development and Feasibility Study of an AI WhatsApp Bot for Science Learning in West Africa

ApplicationsDGX agent

arXiv:2605.15376v1 Announce Type: new Abstract: Sub-Saharan Africa faces persistently high student-teacher ratios and shortages of qualified teachers, limiting students' access to personalized learnin

AgriMind: An Ensemble Deep Learning Framework for Multi-Class Plant Disease Classification

HardwareDGX agent

arXiv:2605.16076v1 Announce Type: cross Abstract: Plant disease detection is still largely manual in Bangladesh, where extension workers eyeball leaf samples across millions of smallholdings. We built

Benchmark of Benchmarks: Unpacking Influence and Code Repository Quality in LLM Safety Benchmarks

Model ReleasesDGX agent

arXiv:2603.04459v3 Announce Type: replace-cross Abstract: The rapid expansion of research in LLM safety presents challenges in tracking advancements, making benchmarks important evaluation infrastruct

CitePrism: Human-in-the-Loop AI for Citation Auditing and Editorial Integrity

Local AiDGX agent

arXiv:2605.16000v1 Announce Type: cross Abstract: Editors and reviewers are expected to ensure that manuscripts cite relevant, accurate, current, and ethically appropriate literature, yet manuscript-l

Context-aware Entity-Relation Extraction for Threat Intelligence Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.15904v1 Announce Type: new Abstract: Cybersecurity Knowledge Graphs (CKGs) unify diverse Cyber Threat Intelligence (CTI) sources into structured, queryable formats, offering scalable soluti

End-to-end plaque counting and virus titration from laboratory plate images with deep learning

Model ReleasesDGX agent

arXiv:2605.16008v1 Announce Type: new Abstract: Plaque assays remain the gold standard readout of virus infectivity; however, plaque counting from plate images is labor-intensive and prone to inter-op

Evaluating Design Video Generation: Metrics for Compositional Fidelity

ResearchDGX agent

arXiv:2605.16223v1 Announce Type: cross Abstract: Generative video models are increasingly used in design animation tasks, yet no standardized evaluation framework exists for this domain. Unlike natur

FINESSE-Bench: A Hierarchical Benchmark Suite for Financial Domain Knowledge and Technical Analysis in Large Language Models

Model ReleasesDGX agent

arXiv:2605.15482v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being applied to financial analysis, reporting, investment decision support, risk management, compliance,

From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery

Model ReleasesDGX agent

arXiv:2605.15412v1 Announce Type: cross Abstract: Modern quantitative trading increasingly relies on systematic models to extract predictive signals from large-scale financial data, where alpha factor

Genome-Factory: A Library for Tuning, Deploying, and Interpreting Genomic Foundation Models

Model ReleasesDGX agent

arXiv:2509.12266v2 Announce Type: replace-cross Abstract: We introduce Genome-Factory, the first integrated Python library for tuning, deploying, and interpreting genomic foundation models. Our core c

Is Agentic AI Ready for Real-World Hardware Engineering? A Deep Dive with Phoenix-bench

Local AiDGX agent

arXiv:2605.15226v1 Announce Type: cross Abstract: We ask whether agentic AI systems built for software engineering transfer to realistic hardware engineering. Existing hardware LLM benchmarks isolate

Learning Structured Robot Policies from Vision-Language Models via Synthetic Neuro-Symbolic Supervision

Model ReleasesDGX agent

arXiv:2604.02812v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have recently demonstrated strong capabilities in mapping multimodal observations to robot behaviors. However, most cu

LUIVITON: Learned Universal Interoperable VIrtual Try-ON

SafetyDGX agent

arXiv:2509.05030v2 Announce Type: replace Abstract: To enable large-scale reuse of real-world 3D assets, where garments and characters rarely share skeletons, templates, or dense correspondences, we p

Prospective multi-pathogen disease forecasting using autonomous LLM-guided tree search

AgentsDGX agent

arXiv:2605.16238v1 Announce Type: new Abstract: Probabilistic forecasting of infectious diseases is crucial for public health but relies on labor-intensive manual model curation by expert modeling tea

← Previous
1…6667686970…83
Next →