AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “automated”

GridTimelineEvolution
4,951 results
Research

XFlowMap: Cross-Scale Generalization and Mapping of Massive Origin-Destination Data

DGX agent

arXiv:2605.18777v1 Announce Type: cross Abstract: Mapping large origin-destination (OD) datasets remains challenging because flow maps become cluttered, meaningful patterns occur at multiple spatial s

researcharxiv-cs-cv
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

A2RBench: An Automatic Paradigm for Formally Verifiable Abstract Reasoning Benchmark Generation

DGX agent

arXiv:2605.17278v1 Announce Type: new Abstract: Abstract reasoning ability reflects the intelligence and generalization capacity of LLMs to extract and apply abstract rules. However, accurately measur

model-releasesarxiv-cs-ai
19 May 2026
Safety

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making

DGX agent

arXiv:2605.17228v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in high-stakes domains such as clinical decision support and medical documentation. However, the

safetyarxiv-cs-cl
19 May 2026
Safety

Beyond Transcripts: Iterative Peer-Editing with Audio Unlocks High-Quality Human Summaries of Conversational Speech

DGX agent

arXiv:2605.17652v1 Announce Type: new Abstract: There are not enough established benchmarks for the task fo speech summarization. Creating new benchmarks demands human annotation, as LLMs could embed

safetyarxiv-cs-cl
19 May 2026
Research

Bridging the Version Gap: Multi-version Training Improves ICD Code Prediction, Especially for Rare Codes

DGX agent

arXiv:2605.17755v1 Announce Type: cross Abstract: Clinical coding maps clinical documentation to standardized medical codes, an essential yet time-consuming administrative task that could benefit from

researcharxiv-cs-ai
19 May 2026
Model Releases

Can LLMs Generate and Solve Linguistic Olympiad Puzzles?

DGX agent

arXiv:2509.21820v2 Announce Type: replace Abstract: In this paper, we introduce a combination of novel and exciting tasks: the solution and generation of linguistic puzzles. We focus on puzzles used i

model-releasesarxiv-cs-cl
19 May 2026
Safety

Code as Agent Harness

DGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

safetyarxiv-cs-ai
19 May 2026
Local Ai

Cross-Source Supervision for Bone Infection Segmentation in Dual-Modality PET-CT

DGX agent

arXiv:2605.16373v1 Announce Type: cross Abstract: Early and accurate diagnosis and lesion localization of bone infections are crucial for clinical treatment. PET-CT integrates anatomical information f

local-aiarxiv-cs-ai
19 May 2026
Model Releases

CVE-Factory: Scaling Expert-Level Agentic Tasks for Code Security Vulnerability

DGX agent

arXiv:2602.03012v2 Announce Type: replace-cross Abstract: Evaluating and improving the security capabilities of code agents requires high-quality, executable vulnerability tasks. However, existing wor

model-releasesarxiv-cs-ai
19 May 2026
Agents

Data-driven and distributed governance of building facilities management using decentralized autonomous organization, digital twin, and large language models

DGX agent

arXiv:2605.16298v1 Announce Type: cross Abstract: While traditional AI and data-driven facilities management approaches have improved building operational efficiency, they remain constrained by centra

agentsarxiv-cs-ai
19 May 2026
Model Releases

Employing Vision-Language Models for Face Image Quality Assessment

DGX agent

arXiv:2605.17489v1 Announce Type: new Abstract: Face Image Quality Assessment (FIQA) is a crucial control step in biometric pipelines. It ensures only reliable samples are processed to maintain system

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Enhancing Metacognitive AI: Knowledge-Graph Population with Graph-Theoretic LLM Enrichment

DGX agent

arXiv:2605.16676v1 Announce Type: new Abstract: Metacognition-the ability to monitor one's own knowledge state, spot gaps, and autonomously fill them--remains largely absent from modern AI. Here, we p

model-releasesarxiv-cs-ai
19 May 2026
Safety

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

DGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

safetyarxiv-cs-ai
19 May 2026
Model Releases

GAMMA: Global Bit Allocation for Mixed-Precision Models under Arbitrary Budgets

DGX agent

arXiv:2605.18475v1 Announce Type: cross Abstract: Mixed-precision quantization improves the budget--accuracy trade-off for large language models (LLMs) by allocating more bits to sensitive modules. Ho

model-releasesarxiv-cs-ai
19 May 2026
Safety

Generation Navigator: A State-Aware Agentic Framework for Image Generation

DGX agent

arXiv:2605.17969v1 Announce Type: new Abstract: Despite rapid advances in text-to-image generation, faithfully realizing user intent remains challenging, often requiring manual multi-turn trial and er

safetyarxiv-cs-cv
19 May 2026
Model Releases

GenoMAS: A Multi-Agent Framework for Scientific Discovery via Code-Driven Gene Expression Analysis

DGX agent

arXiv:2507.21035v3 Announce Type: replace Abstract: Gene expression analysis holds the key to many biomedical discoveries, yet extracting insights from raw transcriptomic data remains formidable due t

model-releasesarxiv-cs-ai
19 May 2026
Research

GeoSym127K: Scalable Symbolically-verifiable Synthesis for Multimodal Geometric Reasoning

DGX agent

arXiv:2605.16371v1 Announce Type: cross Abstract: Large Multimodal Models (LMMs) often struggle with geometric reasoning due to visual hallucinations and a lack of mathematically precise Chain-of-Thou

researcharxiv-cs-ai
19 May 2026
Safety

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

DGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

safetyarxiv-cs-cl
19 May 2026
Model Releases

HPC-LLM: Practical Domain Adaptation and Retrieval-Augmented Generation for HPC Support

DGX agent

arXiv:2605.16347v1 Announce Type: new Abstract: Modern scientific research increasingly depends on High-Performance Computing (HPC) infrastructures, yet many researchers face significant operational b

model-releasesarxiv-cs-lg
19 May 2026
Research

Human-Certified Module Repositories for the AI Age

DGX agent

arXiv:2603.02512v4 Announce Type: replace-cross Abstract: Human-Certified Module Repositories (HCMRs) are introduced in this work as a new architectural model for constructing trustworthy software in

researcharxiv-cs-ai
19 May 2026
Safety

LegalCheck: Retrieval- and Context-Augmented Generation for Drafting Municipal Legal Advice Letters

DGX agent

arXiv:2605.12012v2 Announce Type: replace Abstract: Public-sector legal departments in the Netherlands face acute staff shortages, increased case volumes, and increased pressure to meet regulatory com

safetyarxiv-cs-ai
19 May 2026
Research

Leveraging Multimodal Self-Consistency Reasoning in Coding Motivational Interviewing for Alcohol Use Reduction

DGX agent

arXiv:2605.12987v2 Announce Type: replace Abstract: BACKGROUND: Coding Motivational Interviewing (MI) sessions is essential for understanding client behaviors and predicting outcomes, but it requires

researcharxiv-cs-cl
19 May 2026
Research

Leveraging Unsupervised Learning for Cost-Effective Visual Anomaly Detection

DGX agent

arXiv:2409.15980v2 Announce Type: replace-cross Abstract: Traditional machine learning-based visual inspection systems require extensive data collection and repetitive model training to improve accura

researcharxiv-cs-ai
19 May 2026
Model Releases

LinAlg-Bench: A Forensic Benchmark Revealing Structural Failure Modes in LLM Mathematical Reasoning

DGX agent

arXiv:2605.16675v1 Announce Type: new Abstract: We introduce LinAlg-Bench, a diagnostic benchmark evaluating 10 frontier large language models on structured linear algebra computation across a strict

model-releasesarxiv-cs-ai
19 May 2026
Safety

LLM-Safety Evaluations Lack Robustness

DGX agent

arXiv:2503.02574v2 Announce Type: replace-cross Abstract: In this paper, we argue that current safety alignment research efforts for large language models are hindered by many intertwined sources of n

safetyarxiv-cs-ai
19 May 2026
Local Ai

LLMs for automatic annotation of Mandarin narrative transcripts

DGX agent

arXiv:2605.17205v1 Announce Type: new Abstract: Linguistic annotation of transcribed speech is essential for research in language acquisition, language disorders, and sociolinguistics, yet remains lab

local-aiarxiv-cs-cl
19 May 2026
Tutorials

Machine Learning Enabled Graph Analysis of Particulate Composites: Application to Solid-state Battery Cathodes

DGX agent

arXiv:2512.16085v2 Announce Type: replace-cross Abstract: Particulate composites underpin many solid-state chemical and electrochemical systems, where microstructural features such as multiphase bound

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

ManiSoft: Towards Vision-Language Manipulation for Soft Continuum Robotics

DGX agent

arXiv:2605.18617v1 Announce Type: cross Abstract: Most existing vision-language manipulation research targets rigid robotic arms, whose fixed morphology limits adaptability in cluttered or confined sp

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution

DGX agent

arXiv:2603.05308v2 Announce Type: replace-cross Abstract: Assessing whether an article supports an assertion is essential for hallucination detection and claim verification. While large language model

model-releasesarxiv-cs-ai
19 May 2026
Research

Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis

DGX agent

arXiv:2605.17539v1 Announce Type: new Abstract: Combinatorial optimization (CO) underlies decision-making from logistics to chip design, where infeasible solutions are operationally unusable and small

researcharxiv-cs-ai
19 May 2026
Research

MHMamba: Multi-Head Mamba for 3D Brain Tumor Segmentation

DGX agent

arXiv:2605.16464v1 Announce Type: cross Abstract: Brain tumors exhibit high heterogeneity in morphology and multimodal contrast, making manual slice-by-slice de lineation time-consuming and experience

researcharxiv-cs-ai
19 May 2026
Model Releases

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

DGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your …

DGX agent

New upgrades to the @GeminiApp are you helping you get more done: ✨Gemini Spark is your 24/7 personal AI agent that can take action on your behalf, under your direction. It seamlessly integrates with

model-releasesgoogle-ai--x
19 May 2026
Model Releases

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

DGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

model-releasesarxiv-cs-cl
19 May 2026
Applications

Perovskite-R1: a domain-specialized large language model for intelligent discovery of precursor additives and experimental design

DGX agent

arXiv:2507.16307v2 Announce Type: replace-cross Abstract: Perovskite solar cells (PSCs) have rapidly emerged as a leading contender in next-generation photovoltaic technologies, owing to their excepti

applicationsarxiv-cs-ai
19 May 2026
Applications

PopPy: Opportunistically Exploiting Parallelism in Python Compound AI Applications

DGX agent

arXiv:2605.18697v1 Announce Type: cross Abstract: Compound AI applications, which compose calls to ML models using a general-purpose programming language like Python, are widely used for a variety of

applicationsarxiv-cs-ai
19 May 2026
Model Releases

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency

DGX agent

arXiv:2605.18732v1 Announce Type: cross Abstract: While scaling laws govern aggregate large language model performance, no scaling law has linked factual recall to both model size and training-data co

model-releasesarxiv-cs-ai
19 May 2026
Research

Principal Component Analysis for Lunar Crater Detection

DGX agent

arXiv:2605.17125v1 Announce Type: new Abstract: Optical navigation is a critical component for lunar orbiter and lander missions. Image-based crater identification has emerged as a promising technolog

researcharxiv-cs-cv
19 May 2026
Tutorials

PromptDecipher: Supporting AI Tutor Authoring Through Editable Simulated Interactions

DGX agent

arXiv:2605.16605v1 Announce Type: cross Abstract: Chatbots have long been explored as tools to support learning, and recent advances in large language models have significantly expanded the availabili

tutorialsarxiv-cs-ai
19 May 2026
Safety

QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI

DGX agent

arXiv:2605.17382v1 Announce Type: new Abstract: The rapid progress of generative artificial intelligence has exposed fundamental limitations in existing evaluation methodologies, particularly for open

safetyarxiv-cs-ai
19 May 2026
Local Ai

RadGame: An AI-Powered Platform for Radiology Education

DGX agent

arXiv:2509.13270v2 Announce Type: replace-cross Abstract: We introduce RadGame, an AI-powered gamified platform for radiology education that targets two core skills: localizing findings and generating

local-aiarxiv-cs-ai
19 May 2026
Safety

Rethinking Code Review in the Age of AI: A Vision for Agentic Code Review

DGX agent

arXiv:2605.17548v1 Announce Type: cross Abstract: Code review has evolved for decades, from informal peer checking to today's pull request (PR) workflows, yet it remains a largely manual, uneven, and

safetyarxiv-cs-ai
19 May 2026
Research

Robust and Resilient Soft Robotic Object Insertion with Compliance-Enabled Contact Formation and Failure Recovery

DGX agent

arXiv:2509.17666v2 Announce Type: replace Abstract: Object insertion tasks are prone to failure under pose uncertainty and environmental variation, often requiring manual fine-tuning or controller ret

researcharxiv-cs-ro
19 May 2026
Model Releases

StyleText: A Large-Scale Dataset and Benchmark for Stylized Scene Text Inpainting

DGX agent

arXiv:2605.17309v1 Announce Type: cross Abstract: We present StyleText, a large-scale dataset and benchmark for localized scene-text inpainting with style preservation. StyleText contains 28,518 image

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Sustainability via LLM Right-sizing

DGX agent

arXiv:2504.13217v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have become increasingly embedded in organizational workflows. This has raised concerns over their energy consump

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

TOBench: A Task-Oriented Omni-Modal Benchmark for Real-World Tool-Using Agents

DGX agent

arXiv:2605.16909v1 Announce Type: new Abstract: Tool-using agents are increasingly expected to operate across realistic professional workflows, where they must interpret multimodal inputs, coordinate

model-releasesarxiv-cs-ai
19 May 2026
Research

Towards Robust Argumentative Essay Understanding via TIDE: An Interactive Framework with Trial and Debate

DGX agent

arXiv:2605.17247v1 Announce Type: new Abstract: Argumentative essays serve as a vital medium for assessing critical thinking and reasoning skills, yet there is limited works on accurately understandin

researcharxiv-cs-ai
19 May 2026
Model Releases

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

DGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…8283848586…104
Next →