AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
5 Aug 2026

Automatic Patient-Specific Microwave Ablation Planning Accelerated by a Physics-Guided Deep Learning Model

ResearchDGX agent

arXiv:2608.03086v1 Announce Type: cross Abstract: Microwave ablation (MWA) is a promising minimally invasive treatment for liver tumors, but its therapeutic outcome strongly depends on patient-specifi

Bayesian Data Reweighting Improves Multimodal Retrieval for Knowledge-Based Visual Question Answering

ResearchDGX agent

arXiv:2608.02907v1 Announce Type: new Abstract: Multimodal retrievers are essential for knowledge-based visual question answering, where they retrieve external evidence for image-question pairs. Howev

Benign interpolation and Occam's razor

ResearchDGX agent

arXiv:2608.03386v1 Announce Type: new Abstract: Contemporary deep learning methods generalize well even when they fit their training data perfectly, a phenomenon known as benign interpolation. This ph


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond Solving: Prescriptive Probing for Neural Routing Solvers

Local AiDGX agent

arXiv:2602.07216v2 Announce Type: replace Abstract: Neural combinatorial optimization (NCO) trains fast heuristics for routing problems, but planners often need more than a single solve: they ask whic

Beyond the Gegenbauer Paradigm: q-Orthogonal Kernels for Machine Learning

Model ReleasesDGX agent

arXiv:2608.03482v1 Announce Type: new Abstract: The performance of Support Vector Machines (SVMs) critically depends on the kernel function choice, which enables implicit mapping of data into high-dim

Bi-Lipschitz Ansatz for Anti-Symmetric Functions

ResearchDGX agent

arXiv:2503.04263v2 Announce Type: replace Abstract: Motivated by applications to the simulation of quantum many-body systems by neural networks, researchers have suggested several models which are ant

Bi-semantic Chemical Embedder for Joint Representation Learning of SMILES and Natural Language

ResearchDGX agent

arXiv:2608.03855v1 Announce Type: new Abstract: Transformer models have revolutionized natural language processing (NLP), and text-based molecular representations like SMILES have successfully extende

Bridging Prediction and Attribution: Identifying Forward and Backward Causal Influence Ranges Using Assimilative Causal Inference

SafetyDGX agent

arXiv:2510.21889v2 Announce Type: replace-cross Abstract: Causal inference identifies cause-and-effect relationships between variables. While traditional approaches rely on data to reveal causal links

Causal Inference with Unstructured Outcomes

TutorialsDGX agent

arXiv:2608.03085v1 Announce Type: cross Abstract: Causal inference has traditionally centered on scalar outcomes: whether a patient recovers, how much a worker earns, or how many visits a website rece

CausalOPD: First-Wrong-Step Supervision for Distilling Causal Chain Reasoning

SafetyDGX agent

arXiv:2608.03673v1 Announce Type: new Abstract: Many critical reasoning tasks, including clinical diagnosis, legal judgment, and industrial fault diagnosis, require step-dependent causal chains in whi

Conditionally Identifiable Latent-Environment Modeling for Out-of-Distribution Recommendation

ResearchDGX agent

arXiv:2608.03647v1 Announce Type: cross Abstract: Out-of-distribution (OOD) recommendation is vulnerable to preference shifts induced by a latent environment. Existing methods can infer latent states

Conformal risk control for model-form uncertainty in parametric non-intrusive reduced-order models

ApplicationsDGX agent

arXiv:2608.03360v1 Announce Type: cross Abstract: Non-intrusive reduced-order models (NIROMs) have become a standard tool for approximating parametric partial differential equations from computer desi

ConformalShift: Targeted Event Reordering Against Adaptive ECG Monitoring

ApplicationsDGX agent

arXiv:2608.03628v1 Announce Type: new Abstract: Adaptive conformal prediction can recover clinically important heartbeat classes missed by a point classifier, but delayed feedback makes its decisions

Contrast-invariant deep ptychography neural networks

ApplicationsDGX agent

arXiv:2608.02869v1 Announce Type: new Abstract: Ptychography neural networks suffer from scaling inconsistencies when generalizing out of distribution, limiting their real world viability. We address

Convergence analysis of controlled particle systems arising in deep learning: from finite to infinite sample size

ResearchDGX agent

arXiv:2404.05185v4 Announce Type: replace-cross Abstract: This paper deals with a class of neural SDEs and studies the limiting behavior of the associated sampled optimal control problems as the sampl

Cross-Country Learning for National Infectious Disease Forecasting Using European Data

ApplicationsDGX agent

arXiv:2601.20771v2 Announce Type: replace-cross Abstract: Accurate forecasting of infectious disease incidence is critical for public health planning and timely intervention. While most data-driven fo

Cross-Model KV Cache Transfer in LLM Families: A Closed-Form Linear Mapping for Prefill Reuse

ApplicationsDGX agent

arXiv:2608.03893v1 Announce Type: new Abstract: Production deployments often swap between different-sized models in a family for cost-quality cascading, mid-conversation switching, and routing, and ea

CRS-Triage: Confidence- and Reliability-Aware Selective Triage under Incomplete Clinical Evidence

ResearchDGX agent

arXiv:2608.03862v1 Announce Type: new Abstract: Emergency triage requires reliable decisions within a short time period. However, the available electronic health record (EHR) data, including structure

DAIF: A Data-Driven Intermediate Fusion Framework for Multimodal Supervised Learning via Approximate Message Passing

Model ReleasesDGX agent

arXiv:2608.02769v1 Announce Type: cross Abstract: Multimodal supervised learning seeks to leverage multiple heterogeneous data sources to improve predictive performance. A central challenge is determi

Design Criteria for SGD Preconditioners: Local Conditioning, Noise Floors, and Basin Stability

ResearchDGX agent

arXiv:2511.19716v3 Announce Type: replace-cross Abstract: Stochastic Gradient Descent (SGD) often slows in the late stage of training due to anisotropic curvature and gradient noise. We analyze precon

Design-Time Optimization of Deep Neural Networks for Intermittent Learning on Microcontrollers

Local AiDGX agent

arXiv:2608.03589v1 Announce Type: new Abstract: We present a method for designing deep neural networks (DNNs) for intermittent, energy-autonomous, on-device learning on microcontroller units (MCUs). I

Detecting high-frequency brain disorder signals using dynamic mode decomposition from EEG

ResearchDGX agent

arXiv:2608.02804v1 Announce Type: cross Abstract: Recent studies have reported clearly identifiable dynamical changes in the high-frequency range of EEG signals recorded during specific stimuli, such

DiagLoop: A Counterfactual Data Flywheel with Stage-Localized Reinforcement for Diagnostic LLMs

Local AiDGX agent

arXiv:2608.03674v1 Announce Type: new Abstract: Causal diagnostic models must explain how conclusions follow from evidence because diagnoses guide repairs and treatments. Yet serious cases are scarce,

Divide-and-Conquer: Towards Generalizable Amortized Bayesian Inference for the Drift Diffusion Model

TutorialsDGX agent

arXiv:2608.03566v1 Announce Type: cross Abstract: The drift diffusion model (DDM) is a cornerstone of cognitive decision-making research. Although numerous estimation methods exist, researchers contin

Double Descent in Gradient Boosting Decision Trees via Split-Candidate Scaling

Model ReleasesDGX agent

arXiv:2608.03111v1 Announce Type: new Abstract: Double descent is commonly studied by scaling an explicit capacity parameter, such as neural-network width. For gradient boosting decision trees (GBDTs)

ED-DiT: Physics-Guided Diffusion Pretraining for Transferable Molecular Representations from Electron Density

TutorialsDGX agent

arXiv:2608.03260v1 Announce Type: new Abstract: Pretraining has shown strong potential for learning transferable representations, yet it remains underexplored for electron-density-based molecular lear

Efficient quantum-enhanced classical simulation for patches of quantum landscapes

ResearchDGX agent

arXiv:2411.19896v2 Announce Type: replace-cross Abstract: Understanding the capabilities of classical simulation methods is key to identifying where quantum computers are advantageous. Not only does t

Enhancing Q-Value Updates in Deep Q-Learning via Successor-State Prediction

SafetyDGX agent

arXiv:2511.03836v2 Announce Type: replace Abstract: Deep Q-Networks (DQNs) estimate future returns by learning from transitions sampled from a replay buffer. However, the target updates in DQN often r

Evaluation Blindness: How Silent Measurement Failures Corrupt AI Systems from Training to Deployment

Model ReleasesDGX agent

arXiv:2608.02786v1 Announce Type: new Abstract: AI systems can fail silently. The failure propagates through training loops, evaluation pipelines, and production monitoring stacks until downstream har

Exploiting Separability in Multi-Scale Grey-Box Bayesian Optimization

Model ReleasesDGX agent

arXiv:2608.03045v1 Announce Type: new Abstract: We consider grey-box optimization problems where the decision variables naturally partition into black-box variables (as arguments to an expensive black

FedCARE: A Multi-Objective Personalised Federated Learning Framework for Smart Healthcare

TutorialsDGX agent

arXiv:2608.03498v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training across distributed healthcare institutions without centralising sensitive patient data. How

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

Model ReleasesDGX agent

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource c

Federated generative event models for tokenized electronic health records

ResearchDGX agent

arXiv:2608.02939v1 Announce Type: new Abstract: Electronic health record foundation models are limited by institutionally siloed data and substantial performance degradation under cross-site transfer.

FedRings: A Scalable and Topology-Aware Federated Learning Framework for LEO Satellite Constellations

ResearchDGX agent

arXiv:2608.03436v1 Announce Type: cross Abstract: Federated learning over low Earth orbit (LEO) satellite networks is limited by frequent link changes, short contact times, and a highly dynamic topolo

Field Aware Agent Skill Retrieval

AgentsDGX agent

arXiv:2608.02880v1 Announce Type: cross Abstract: As lifelong learning agents accumulate lifelong growing skill banks, retrieving the correct skill becomes an increasingly important bottleneck. Most c

Forecasting Revenue with its Customer-Base Drivers: When and Why Coordination Helps

Model ReleasesDGX agent

arXiv:2608.02911v1 Announce Type: new Abstract: Revenue forecasts guide acquisition budgets, demand planning, and customer-based valuations, yet an aggregate forecast does not show whether change refl

Fretiq: Browser-Native Electric Guitar String Classification via Engineered Spectral Features and Held-Out Free-Play Evaluation

ResearchDGX agent

arXiv:2607.18303v2 Announce Type: replace-cross Abstract: Identifying which string produces a given pitch in monophonic electric guitar audio is a classification challenge: a single pitch can often be

GeoID-PINN: Identifiability-Aware Regional Epidemic Inference with Geographic Coupling

Local AiDGX agent

arXiv:2608.02633v1 Announce Type: new Abstract: Regional surveillance data reflect local transmission, reporting, seeding, and external infection pressure, which are difficult to identify separately.

GLOBE: Trajectory-Aligned Gradient Matching with Structured SparseOptimization for Coreset Selection

Local AiDGX agent

arXiv:2608.02690v1 Announce Type: new Abstract: On-device training of deep neural networks is fundamentally constrained by the computational and memory costs of large-scale datasets. Coreset selection

GoT-CD: Graph-of-Thoughts Causal Discovery and the Fragility of Post-hoc Path-Specific Fairness Audits

Model ReleasesDGX agent

arXiv:2608.02877v1 Announce Type: new Abstract: Causal discovery recovers directed structure from observational data and is increasingly used in clinical settings to support mechanism reasoning and fa

HAPEns: Hardware-Aware Post-Hoc Ensembling for Tabular Data

ResearchDGX agent

arXiv:2603.10582v2 Announce Type: replace Abstract: Ensembling is commonly used in machine learning on tabular data to boost predictive performance and robustness, but larger ensembles often lead to i

Heteroscedasticity of Denoising Score Matching with Generalised Smooth Noise

ResearchDGX agent

arXiv:2508.01597v2 Announce Type: replace Abstract: Score Matching (SM) is a powerful framework for estimating the log-density derivatives of a distribution without calculating its normalizing constan

Improving Sample Efficiency in Multi-Agent Reinforcement Learning for Simulated Football Games via Exploration

AgentsDGX agent

arXiv:2503.13077v2 Announce Type: replace Abstract: Multi-agent reinforcement learning has shown promise in learning cooperative behaviors in team-based environments. However, such methods often deman

Information-Geometric Forward Policy Training in GFlowNets

Local AiDGX agent

arXiv:2608.03967v1 Announce Type: cross Abstract: Generative Flow Networks (GFlowNets) have emerged as a flexible framework for amortised inference over discrete and mixed discrete-continuous objects,

Inverted Detection and Control in Steering Vectors

Model ReleasesDGX agent

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption un

Joint Affine Spectral Shaping: Coupling Weight and Bias Updates Beyond Weight-Only Muon

SafetyDGX agent

arXiv:2608.02991v1 Announce Type: new Abstract: Matrix spectral optimizers reshape weight-update spectra but usually delegate vector-valued biases to a separate optimizer. We study whether this separa

Kernel weighted importance sampling for off-policy evaluation in contextual bandits

SafetyDGX agent

arXiv:2607.15067v2 Announce Type: replace Abstract: This article presents a novel estimator for performing off-policy evaluation using only offline data for contextual bandits. The proposed estimator,

LAEF: A Lead-Agnostic ECG Foundation Model Towards Point-of-Care Diagnostics

Model ReleasesDGX agent

arXiv:2608.03690v1 Announce Type: new Abstract: Point-of-care cardiac devices such as smartwatches and handheld ECG recorders typically capture 1--2 leads, yet existing ECG foundation models are archi

Learning and Clustering on Temporal Graphs: Principles, Primitives, and Pooling

HardwareDGX agent

arXiv:2608.03696v1 Announce Type: new Abstract: This work focuses on the problem of learning on temporal graphs, with particular emphasis on the task of clustering: obtaining coarse-grained representa

LLM-Derived Priors for Thompson Sampling in Cold-Start Comment Recommendation

SafetyDGX agent

arXiv:2608.03382v1 Announce Type: cross Abstract: Multi-armed bandit algorithms, especially Thompson sampling, are widely used in online recommendation. Despite their ability to adapt from online feed

LLMs Can Annotate Attribution Graphs

ResearchDGX agent

arXiv:2608.02632v1 Announce Type: new Abstract: Circuit tracing is an exciting technique for revealing the internal computation of language models, but it requires a time-intensive manual step of grou

LoBoost: Fast Model-Native Local Conformal Prediction for Gradient-Boosted Trees

Local AiDGX agent

arXiv:2602.22432v2 Announce Type: replace-cross Abstract: Gradient-boosted decision trees are among the strongest off-the-shelf predictors for tabular regression, but point predictions alone do not qu

Maglev: Sliding Recurrent Memory

Model ReleasesDGX agent

arXiv:2608.02870v1 Announce Type: new Abstract: We introduce ours{}, a recurrent Transformer architecture with fixed-size memory that generalizes sliding-window attention while remaining parallelizabl

Muon Meets Mamba: Spectral Optimization for State Space Models

ResearchDGX agent

arXiv:2608.03941v1 Announce Type: new Abstract: Muon is a recent optimizer that orthogonalizes the update to each weight matrix with a Newton-Schulz iteration, which performs steepest descent under th

Neural network realization of binary refinement iterates via a two-chart atlas selector

ResearchDGX agent

arXiv:2608.02624v1 Announce Type: cross Abstract: Refinement operators generate many functions used in wavelet constructions, subdivision schemes, and geometric modeling. Their finite iterates can dev

Neural Networks with Local Converging Inputs for Efficient Options Pricing Models

Model ReleasesDGX agent

arXiv:2608.02778v1 Announce Type: new Abstract: We present a novel application of Neural Networks with Local Converging Inputs (NNLCI) to improve the efficiency of existing numerical methods for prici

Noise-Aware Shrinkage for Differentially Private Zeroth-Order Fine-Tuning of Large Language Models

ResearchDGX agent

arXiv:2608.03277v1 Announce Type: new Abstract: Differentially private zeroth-order optimization (DP-ZO) enables memory-efficient private fine-tuning of large language models using only forward evalua

NOMADD: Numerical Optimization of Models Adapting to Data Drift

Model ReleasesDGX agent

arXiv:2608.02845v1 Announce Type: new Abstract: Tabular model performance degrades when feature distributions change over time or the relationship between features and outcome variables change over ti

Omega-S: A Functional Resilience Index for LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2608.03887v1 Announce Type: new Abstract: Fine-tuning a large language model on new data degrades what it previously learned. We present Omega-S, a drop-in penalty computed from the weight matri

On the Implicit Flatness Bias of Sharpness-Aware Minimization: A Linear Stability Analysis with Quantitative Hyperparameter Bounds

Model ReleasesDGX agent

arXiv:2608.03197v1 Announce Type: new Abstract: Sharpness-Aware Minimization (SAM) improves generalization by seeking parameters whose loss is robust to local adversarial perturbations, but the quanti

← Previous
1…1011121314…239
Next →