AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
7 Aug 2026

How Much Reconstruction Does Quantum Machine Learning Need? Late Fusion of Independently Trained Quantum Subcircuits

Model ReleasesDGX agent

arXiv:2608.05595v1 Announce Type: cross Abstract: Circuit cutting lets a large quantum neural network (QNN) run as independent subcircuits on small devices, but rebuilding its outputs by reconstructio

Hybrid-Adaptive Thread Tuning to Mitigate Simulation Execution Bottlenecks in High-Performance Reinforcement Learning Inference

TutorialsDGX agent

arXiv:2608.06025v1 Announce Type: new Abstract: In simulation-in-the-loop decision-making systems, reinforcement learning (RL) inference is often constrained by simulator-side execution overhead, wher

Hybrid Probabilistic Zonotopes for Identifiable and Refinable Predictive Uncertainty

Research

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.05454v1 Announce Type: new Abstract: Probabilistic prediction heads in neural networks typically output either a Gaussian mixture or a single conformal region. Neither separates the distinc

Hypothesis Testing with Conditional Queries: Learnability and the Value of Interaction

SafetyDGX agent

arXiv:2608.06262v1 Announce Type: new Abstract: Model evaluations may fix all tests before observing any responses or select later tests using earlier responses. We study this choice in a conditional-

IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games

TutorialsDGX agent

arXiv:2608.05422v1 Announce Type: new Abstract: While many algorithms blend reinforcement learning (RL) with counterfactual regret (CFR) methods to leverage tradeoffs in computational speed and perfor

Kastor: An efficient fine-tuning strategy for generative emulation of PDE simulations

Model ReleasesDGX agent

arXiv:2608.06107v1 Announce Type: new Abstract: Machine learning offers a promising avenue to accelerate physical simulations by replacing computationally expensive traditional Partial Differential Eq

KV-Skill: Forging Expertise in the Model's Native Language

Model ReleasesDGX agent

arXiv:2608.05475v1 Announce Type: new Abstract: Task knowledge is commonly stored either as text in the prompt or as an update to model weights. Text is modular but must be interpreted on every use, w

Latent Utility Q-Learning for Preference-Adaptive Dynamic Treatment Regimes

SafetyDGX agent

arXiv:2307.12022v3 Announce Type: replace-cross Abstract: Optimizing individualized treatment sequences for patients who weigh multiple, competing outcomes differently poses a challenge for dynamic tr

LC-Implicit-QAOA: Active-Workspace-Capped Exact Objective-and-Gradient Evaluation for Training over Bounded QUBO Light Cones

ResearchDGX agent

arXiv:2608.05610v1 Announce Type: cross Abstract: QAOA training repeatedly queries an objective and all shared gradients, making exact evaluation a feasibility bottleneck even when QUBO terms have bou

Learning to Rank Tensor Network Contraction Plans for GPU-Accelerated Quantum Circuit Simulation

HardwareDGX agent

arXiv:2608.05819v1 Announce Type: new Abstract: Classical simulation remains essential for developing and validating quantum algorithms, but its cost grows rapidly with circuit size. Tensor-network co

LILAC: An Idempotent Neural Speech Codec

ResearchDGX agent

arXiv:2608.05727v1 Announce Type: cross Abstract: Neural Audio Codecs are widely adopted in speech generation and editing. However, existing neural audio codecs are not idempotent: across the paper's

LLM Inference Under Bursty Workload Distribution: Modifying the WAIT Algorithm

Model ReleasesDGX agent

arXiv:2608.06135v1 Announce Type: new Abstract: Large Language Models (LLMs) such as ChatGPT and Claude are widely used for information retrieval and problem-solving. Recent work has focused on improv

MetaboLLM: a metabolomics-specialized large language model for biochemical knowledge integration and predictive metabolite graph construction

Model ReleasesDGX agent

arXiv:2608.06253v1 Announce Type: new Abstract: Metabolomics knowledge is distributed across heterogeneous resources and remains difficult to translate into predictive representations. We developed Me

Minimax Optimal Early-Stopped Gradient Descent for Gaussian Mixture Classification

SafetyDGX agent

arXiv:2608.06250v1 Announce Type: cross Abstract: In overparameterised classification, training data can be linearly separable even when the underlying distribution is not. In this setting, gradient d

ML-for-ML

ResearchDGX agent

arXiv:2608.06046v1 Announce Type: cross Abstract: AI training workloads are growing rapidly, making their time, energy, and infrastructure costs increasingly important. In shared cloud clusters, train

MS-MLB: An Open Machine Learning Benchmark for Blood-Based MS Classification

Model ReleasesDGX agent

arXiv:2608.05196v1 Announce Type: new Abstract: Multiple sclerosis (MS) is diagnosed through clinical assessment, magnetic resonance imaging, laboratory evidence when appropriate, and exclusion of bet

Muon on the Stiefel Manifold Admits an Exact Closed-Form Update

ResearchDGX agent

arXiv:2608.06218v1 Announce Type: cross Abstract: We study Muon, a recently proposed matrix-aware optimization method, in the context of the Stiefel manifold. This manifold consists of matrices with o

Neuro-Symbolic Closed-Loop Control of Laser Powder Bed Fusion with an In-Loop Ontology

Model ReleasesDGX agent

arXiv:2608.05773v1 Announce Type: new Abstract: A geometry-conditioned, neuro-symbolic closed-loop architecture is proposed for laser powder bed fusion, in which a standards-aligned ontology operates

Observation-Grounded Self-Predictive Reinforcement Learning for Visual Continuous Control

SafetyDGX agent

arXiv:2608.05989v1 Announce Type: new Abstract: Sample-efficient policy learning from pixels is a long-standing challenge in reinforcement learning (RL). Recent dynamics-based representation learning

On-Policy Self-Distillation without Any Supervision

SafetyDGX agent

arXiv:2608.06296v1 Announce Type: new Abstract: On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still re

On Same-Sample and Independent-Sample Stochastic Extragradient for Monotone Variational Inequalities

ResearchDGX agent

arXiv:2608.06182v1 Announce Type: cross Abstract: We study stochastic extragradient (SEG) methods for solving monotone variational inequality problems (VIPs) over a feasible set. Although extragradien

On the Anisotropy of Score-Based Generative Models

ResearchDGX agent

arXiv:2510.22899v2 Announce Type: replace Abstract: We investigate the role of network architecture in shaping the inductive biases of modern score-based generative models. To this end, we introduce t

Operating Multi-Node Full Fine-Tuning on NVIDIA B300: A Field Report on Telemetry-Based Triage, Negative Results, and Operational Hardening

Model ReleasesDGX agent

arXiv:2608.05944v1 Announce Type: cross Abstract: We report operational experience full-fine-tuning a 32.76B-parameter dense model (Qwen3-32B) on 16 x NVIDIA B300 (two nodes, FSDP / ZeRO-3) -- among t

Optimal or Greedy Decision Trees? Revisiting their Objectives, Tuning, and Performance

ResearchDGX agent

arXiv:2409.12788v3 Announce Type: replace Abstract: Recently there has been a surge of interest in optimal decision tree (ODT) methods that globally optimize accuracy directly, in contrast to traditio

Optimal Rates for Learning with Monotone Adversaries

ResearchDGX agent

arXiv:2608.06337v1 Announce Type: cross Abstract: A monotone adversary observes an i.i.d. labeled sample and appends a finite number of further examples of its choice, every one of them labeled correc

Perfect reconstruction of sparse signals using nonconvexity control and one-step RSB message passing

Model ReleasesDGX agent

arXiv:2512.17426v2 Announce Type: replace-cross Abstract: We consider sparse signal reconstruction via minimization of the smoothly clipped absolute deviation (SCAD) penalty, and develop one-step repl

Physics-Based Molecular Fingerprints from Spectral Graph Theory Provide Efficient Geometry-Aware Measures of Chemical Similarity

ResearchDGX agent

arXiv:2608.05336v1 Announce Type: cross Abstract: Molecular representations are essential for the evaluation of molecular similarity and the development of structure-property relationships. Despite th

Potential Matching Optimal Transport: Continuous Normalizing Flows for Exact p-Wasserstein Dynamics

ResearchDGX agent

arXiv:2608.05666v1 Announce Type: new Abstract: We introduce Potential Matching Optimal Transport (PMOT), a potential-flow framework for general p-cost optimal transport with c_p(x,y)=|x-y|^p. PMOT pa

PPDL: LLM-Based Flows as Probabilistic Programs

AgentsDGX agent

arXiv:2608.05234v1 Announce Type: new Abstract: Building reliable applications that leverage large language models (LLMs) remains a significant challenge. While LLMs offer impressive capabilities acro

Provably Efficient Self-Calibrating Quantum Fault Tolerance

ResearchDGX agent

arXiv:2608.05686v1 Announce Type: cross Abstract: Quantum error correction protects logical information only when every physical operation remains below the fault-tolerance threshold, a condition that

Quantum-Structured World Models (QSWMs) for Predictive Latent Dynamics

TutorialsDGX agent

arXiv:2608.05371v1 Announce Type: new Abstract: World models learn latent states that summarize interaction histories, evolve over time, and support prediction, simulation, or planning. Most existing

RASP-QAOA: Resource-Aware Per-Instance Selection for Exact QAOA Simulation

SafetyDGX agent

arXiv:2608.05646v1 Announce Type: cross Abstract: Exact QAOA simulation spans several computational representations whose useful regions differ sharply across graph structure, circuit depth, precision

Rectifying Geometric Misalignment: Online Source-Free Adaptation for Class-Imbalanced EEG

Model ReleasesDGX agent

arXiv:2608.05315v1 Announce Type: new Abstract: Electroencephalography (EEG) based Brain-Computer Interfaces (BCIs) often require unsupervised domain adaptation (UDA) to generalize across subjects and

Robot Learning from Human Demonstrations: Handwritten Alphabet Trajectories and Human-Likeness Evaluation

Model ReleasesDGX agent

arXiv:2608.06221v1 Announce Type: cross Abstract: Learning from demonstration (LfD) provides a developmental framework through which robots can develop motor skills by observing and imitating human dy

Robust Context-Aware Detection of Malicious Instructions in Text

AgentsDGX agent

arXiv:2608.05430v1 Announce Type: cross Abstract: The remarkable instruction-following ability of modern LLMs has enabled their practical use as the minds of agents that can autonomously complete incr

RxnCLF: Contrastive Transformation-Aware Reaction Foundation Model for Improved Reactivity Prediction

TutorialsDGX agent

arXiv:2608.06259v1 Announce Type: new Abstract: Reaction yield prediction remains challenging because labeled data are scarce and reaction space is both combinatorially large and sparsely populated, l

SAGA: Score-Weighted Adaptive Generation Alignment for Low-Resource Nordic Language Models

SafetyDGX agent

arXiv:2608.06179v1 Announce Type: new Abstract: Preference optimisation has proven effective for improving large language models but typically relies on costly human preference annotations. Extending

Scalable estimation of VARMA models

SafetyDGX agent

arXiv:2608.06340v1 Announce Type: cross Abstract: Vector autoregressive moving-average (VARMA) models have long been considered impractical beyond moderate dimensions: the likelihood is non-convex, th

Scientific Machine Learning of Chaotic Systems Learns Reduced-Order Equations for Neural Populations

Model ReleasesDGX agent

arXiv:2507.03631v4 Announce Type: replace Abstract: Extracting interpretable mathematical models from complex dynamical systems is difficult, especially for chaotic dynamics observed with noisy experi

SEAM: Global consistency beyond local accuracy in scientific machine learning

Model ReleasesDGX agent

arXiv:2608.05702v1 Announce Type: new Abstract: Scientific machine learning commonly validates models at the level of a subdomain, a benchmark split, or an explanation for one prediction. Yet such loc

SkillTFM: Gated Skill Evolution for Training-Free Adaptation of Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2608.06137v1 Announce Type: new Abstract: Tabular data are ubiquitous in real-world applications and are crucial for data-driven prediction and decision-making across science, industry, finance,

Spectral Distillation: From Nonlinear Dynamics to Linear State-Space Models

TutorialsDGX agent

arXiv:2608.05416v1 Announce Type: new Abstract: Can nonlinear dynamical systems be learned through a compact linear state-space representation, without directly solving a non-convex system-identificat

Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning

TutorialsDGX agent

arXiv:2608.06276v1 Announce Type: cross Abstract: Persistence diagrams (PDs) provide stable and interpretable summaries of multiscale topological structure. While substantial progress has been made in

Surv-IPTB: An Attention-Based Model for Estimating Individual Probability of Treatment Benefit with Survival Data

ResearchDGX agent

arXiv:2608.06288v1 Announce Type: new Abstract: This work presents a novel attention-based framework for estimating the Individual Probability of Treatment Benefit (IPTB) in survival analysis contexts

THBKG: A Temporal Biomedical Knowledge Graph for Decision-Aligned Clinical Advancement Prediction

Model ReleasesDGX agent

arXiv:2608.05982v1 Announce Type: new Abstract: Inadequate target--disease linkage accounts for 40--50% of Phase~II efficacy failures, so anticipating which programmes will advance would let sponsors

The Tamed Subgradient Unadjusted Langevin Algorithm beyond Convexity

ResearchDGX agent

arXiv:2608.06283v1 Announce Type: new Abstract: We study the problem of sampling from target distributions whose potentials are simultaneously non-smooth, subject to superlinear gradient growth, and n

Threshold-Based Early Stopping of Accumulations in Neural Networks with Binary Activation

Model ReleasesDGX agent

arXiv:2608.06177v1 Announce Type: new Abstract: Binary neural networks are very attractive for constrained deployment, enabling small footprint and low-power inference. For binary activations, the dot

Timestep-Conditioned Transformers for Global Weather Forecasting

ResearchDGX agent

arXiv:2608.06241v1 Announce Type: new Abstract: Existing machine-learning weather forecasting models rely on predetermined and fixed autoregressive timesteps. The choice of model timestep involves a f

Velocity- and Regime-Aware Detection of Intraday Options Market Manipulation, with Explainable Attribution

ResearchDGX agent

arXiv:2608.05373v1 Announce Type: cross Abstract: Intraday market manipulation is hard to detect because its footprint is brief, buried in millions of quotes, and statistically similar to ordinary vol

Verifiable Regularity Criterion for Conditional Expectation Operators and Conditional Mean Embeddings with Applications to Nonparametric Regression, Bayesian Inverse Problems, and Koopman Operators

ResearchDGX agent

arXiv:2608.06155v1 Announce Type: cross Abstract: Conditional expectation operators (CEOs) and their associated conditional mean embeddings (CMEs) play a central role across applied mathematics and ma

When Do Corrective Features Help? An Agent for Corrective Feature Discovery on Black-Box Forecasters

AgentsDGX agent

arXiv:2608.05207v1 Announce Type: new Abstract: Frozen pretrained forecasters often fail in structured, recurring ways that are costly to repair through fine-tuning. We study corrective feature discov

When Does Consensus Mean Correctness? Measuring the Agreement-Accuracy Coupling with Semantics-Preserving Re-Rendering

Local AiDGX agent

arXiv:2608.05670v1 Announce Type: new Abstract: A model's agreement across perturbed inputs is used both as a label-free reliability signal and as a self-training target, on the premise that agreement

Worst-Case Distance-Aware Error Bounds for Neural Networks

SafetyDGX agent

arXiv:2510.22021v3 Announce Type: replace Abstract: Safety-critical applications of machine learning require uncertainty estimates that support reliable worst-case analysis. Neural networks (NNs) prov

6 Aug 2026

A Comparative Study of Feature Selection Methods for EHR Diagnosis Codes in Opioid Use Disorder Prediction

ResearchDGX agent

arXiv:2608.04180v1 Announce Type: new Abstract: Feature selection is a critical step in electronic health record (EHR)-based predictive modeling, where input variables are often high-dimensional, spar

A Counterexample to Fourier Alignment in Single-Neuron Modular Addition

Model ReleasesDGX agent

arXiv:2608.04451v1 Announce Type: cross Abstract: We give a negative solution to MAIS-O60. We first construct an example in which an initially active ReLU neuron becomes completely inactive in finite

A geometry-based deep equilibrium model for image restoration under multiplicative Gamma noise

ResearchDGX agent

arXiv:2608.04944v1 Announce Type: cross Abstract: We propose a deep learning framework for image restoration from images degraded by both multiplicative Gamma noise and blur. Unlike conventional deep

A Mechanistic Analysis of Transformers for Dynamical Systems

ResearchDGX agent

arXiv:2512.21113v2 Announce Type: replace Abstract: Transformers are increasingly adopted for modeling and forecasting time-series, yet their internal mechanisms remain poorly understood from a dynami

A Multi-Cohort Validation of Censoring-Aware Conformal Lower Predictive Bounds for Pathology Survival Models

ResearchDGX agent

arXiv:2608.04025v1 Announce Type: cross Abstract: Whole-slide survival models commonly provide risk rankings without calibrated statements about individual event times. We evaluate fixed-cutoff drcosa

Above-ground Biomass Estimation with Geospatial Foundation Models

Model ReleasesDGX agent

arXiv:2608.04792v1 Announce Type: new Abstract: Accurate estimation of Above-Ground Biomass (AGB) from satellite imagery is essential for the large-scale monitoring of carbon stocks, yet it remains a

Active Learning Guided Design Space Refinement for Scalable Multi-Objective Bayesian Optimization in Materials Discovery

AgentsDGX agent

arXiv:2608.04651v1 Announce Type: new Abstract: Advanced materials discovery increasingly relies on machine learning and Bayesian optimization to explore large discrete design spaces under limited eva

← Previous
1…7891011…239
Next →