AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
18 May 2026

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

Model ReleasesDGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

Harnessing Unimodality in Semiparametric Contextual Pricing via Oracle Price Map Learning

Model ReleasesDGX agent

arXiv:2605.15411v1 Announce Type: cross Abstract: We study contextual dynamic pricing in a semiparametric scalar-index valuation model where the latent value is v_t=mu_ast(mathsf c_t)+xi_t, with an un

Heuristic-Based Merging of HPC Traces to Extend Hardware Counter Coverage

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.15832v1 Announce Type: cross Abstract: This work extends a framework for predicting the performance of High-Performance Computing (HPC) workloads using Machine Learning (ML). A common limit

High-accuracy log-concave sampling with stochastic queries

ResearchDGX agent

arXiv:2602.14342v2 Announce Type: replace-cross Abstract: We show that high-accuracy guarantees for log-concave sampling -- that is, iteration and query complexities which scale as polylog(1/elta), wh

How Data Augmentation Shapes Neural Representations

ResearchDGX agent

arXiv:2605.15306v1 Announce Type: new Abstract: Data augmentation is widely recognized for improving generalization in deep networks, yet its impact on the geometry of learned representations remains

Hypothesis-driven construction of mesoscopic dynamics

TutorialsDGX agent

arXiv:2605.16211v1 Announce Type: new Abstract: Traditional scientific modeling typically begins with fixed, instance-wise effective equations and then carries out equation-specific analysis and compu

Imitation learning for clinical decision support in pediatric ECMO

TutorialsDGX agent

arXiv:2605.16175v1 Announce Type: new Abstract: Pediatric critical care is a dynamic, high-stakes process involving constant monitoring and adjustments in life-saving treatments. Modeling these interv

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

Model ReleasesDGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

Model ReleasesDGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

Information-Preserving Domain Transfer with Unlabeled Data in Misspecified Simulation-Based Inference

Model ReleasesDGX agent

arXiv:2605.05652v2 Announce Type: replace Abstract: Simulation-based inference (SBI) provides amortized Bayesian parameter inference from simulator-generated data without requiring explicit likelihood

Intrinsic Wasserstein Rates for Score-Based Generative Models on Smooth Manifolds

ResearchDGX agent

arXiv:2605.15822v1 Announce Type: new Abstract: Score-based generative models are trained in high-dimensional ambient spaces, yet many data distributions are supported on low-dimensional nonlinear str

IO-SVD: Input-Output Whitened SVD for Adaptive-Rank LLM Compression

ResearchDGX agent

arXiv:2605.15626v1 Announce Type: new Abstract: Large language models deliver strong performance across language and reasoning tasks, but their storage and compute costs remain major barriers to deplo

ITGPT: Generative Pretraining on Irregular Timeseries

ApplicationsDGX agent

arXiv:2605.16069v1 Announce Type: new Abstract: Timeseries regression models often struggle to leverage large volumes of labeled multimodal data, particularly when the data are irregularly sampled or

Lagrangian Flow Matching: A Least-Action Framework for Principled Path Design

ResearchDGX agent

arXiv:2605.15419v1 Announce Type: new Abstract: Flow matching trains a neural velocity field by regression against a target velocity associated with a prescribed probability path connecting a simple i

LAtte: Hyperbolic Lorentz Attention for Cross-Subject EEG Classification

ResearchDGX agent

arXiv:2603.10881v2 Announce Type: replace Abstract: Electroencephalogram (EEG) classification plays a key role in medical diagnosis and brain-computer interfaces, but remains challenging due to low si

Layer-wise Derivative Controlled Networks

ApplicationsDGX agent

arXiv:2605.15463v1 Announce Type: new Abstract: As machine learning models grow in complexity, they increasingly struggle with three conflicting demands: the need for high accuracy, the requirement fo

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking

SafetyDGX agent

arXiv:2605.16154v1 Announce Type: new Abstract: Reinforcement learning (RL) allows vision-language-action (VLA) policies to generalize beyond their training distribution by optimizing directly for tas

Learning Context-conditioned Gaussian Overbounds for Convolution-Based Uncertainty Propagation

SafetyDGX agent

arXiv:2605.15789v1 Announce Type: new Abstract: Uncertainty quantification is essential in safety-critical settings--from autonomous driving to aviation, finance, and health--where decisions must rely

Learning in Structured Stackelberg Games

SafetyDGX agent

arXiv:2504.09006v4 Announce Type: replace-cross Abstract: We initiate the study of structured Stackelberg games, a novel form of strategic interaction between a leader and a follower where contextual

Learning Where It Matters: Geometric Anchoring for Robust Preference Alignment

Local AiDGX agent

arXiv:2602.04909v3 Announce Type: replace Abstract: Direct Preference Optimization (DPO) and related methods align large language models from pairwise preferences by regularizing updates against a fix

Logical Grammar Induction via Graph Kolmogorov Complexity: A Neuro-Symbolic Framework for Self-Healing Clinical Data Integrity

ApplicationsDGX agent

arXiv:2605.15242v1 Announce Type: new Abstract: The reliability of Healthcare Information Systems (HIS) is frequently compromised by human-induced data entry errors, which existing statistical anomaly

LPDS: Evaluating LLM Robustness Through Logic-Preserving Difficulty Scaling

ResearchDGX agent

arXiv:2605.15393v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed to perform tasks with minimal human oversight, it is crucial that these models operate robustl

Martingale Neural Operators: Learning Stochastic Marginals via Doob-Meyer Factorization

ResearchDGX agent

arXiv:2605.15806v1 Announce Type: new Abstract: Neural operators excel as deterministic surrogates, but inevitably collapse to the conditional mean when applied to stochastic PDEs, discarding the vari

MaxSketch: Robust Distinct Counting in Streams via Random Projections

ResearchDGX agent

arXiv:2605.15571v1 Announce Type: cross Abstract: Estimating the number of distinct elements in a data stream is well understood when repeated elements are identical. In modern settings, however, obse

Metropolis-Scale Road Network Datasets for Fine-Grained Urban Traffic Modeling

ApplicationsDGX agent

arXiv:2510.02278v2 Announce Type: replace Abstract: Modeling traffic dynamics is a critical challenge for urban computing, with applications from real-time traffic management to infrastructure plannin

Mind Dreamer: Untethering Imagination via Active Latent Intervention on Latent Manifolds

SafetyDGX agent

arXiv:2605.16030v1 Announce Type: new Abstract: Model-Based Reinforcement Learning (MBRL) leverages latent imagination for sample efficiency, yet remains constrained by Historical Tethering: imaginati

Missing-Data-Induced Phase Transitions in Spectral PLS for Multimodal Learning

SafetyDGX agent

arXiv:2601.21294v2 Announce Type: replace Abstract: Partial Least Squares (PLS) learns shared structure from paired data via the top singular vectors of the empirical cross-covariance (PLS-SVD), but m

Multi-Fidelity Flow Matching: Cascaded Refinement of PDE Solutions

Model ReleasesDGX agent

arXiv:2605.16118v1 Announce Type: new Abstract: The source distribution in conditional flow matching is a design parameter that can be calibrated to data, not a default isotropic prior. We exploit thi

Multi-Probe Zero Collision Hash (MPZCH): Mitigating Embedding Collisions and Enhancing Model Freshness in Large-Scale Recommenders

Model ReleasesDGX agent

arXiv:2602.17050v3 Announce Type: replace Abstract: Embedding tables are critical components of large-scale recommendation systems, facilitating the efficient mapping of high-cardinality categorical f

MuteBench: Modality Unavailability Tolerance Evaluation for Incomplete Multimodal Fusion

Model ReleasesDGX agent

arXiv:2605.15235v1 Announce Type: new Abstract: Multimodal physiological data powers clinical AI systems from intensive care units to wearable devices, but sensors routinely fail in practice. Two fail

Neural Backward Filtering Forward Guiding

TutorialsDGX agent

arXiv:2601.23030v2 Announce Type: replace-cross Abstract: Inference in nonlinear continuous stochastic processes on trees is challenging, particularly when observations are sparse and the topology is

Njord: A Probabilistic Graph Neural Network for Ensemble Ocean Forecasting

Model ReleasesDGX agent

arXiv:2605.15470v1 Announce Type: new Abstract: Ocean dynamics are inherently chaotic, yet existing machine learning ocean models produce only deterministic forecasts. We introduce Njord, a probabilis

OgBench: A Framework for Evaluating Graph Neural Networks on Omics Data

Model ReleasesDGX agent

arXiv:2605.15511v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become the dominant framework for inductive graph-level learning. Yet most benchmarks focus on the regime n gg p, wher

On Kernel Eigen-alignments of KRR: Reconstruction and Generalization

SafetyDGX agent

arXiv:2605.15240v1 Announce Type: cross Abstract: This paper investigates the critical role of eigenalignments between the kernel matrix and learning targets in achieving robust generalization in lear

On the Convergence Rates of Federated Q-Learning across Heterogeneous Environments

AgentsDGX agent

arXiv:2409.03897v3 Announce Type: replace Abstract: Large-scale multi-agent systems are often deployed across wide geographic areas, where agents interact with heterogeneous environments. There is an

On the Power of Adaptivity for arepsilon-Best Arm Identification in Linear Bandits

ResearchDGX agent

arXiv:2605.15663v1 Announce Type: new Abstract: We study the minimax sample complexity of arepsilon-best arm identification in linear bandits. Given a compact action set X that spans R^d and an unknow

On the Stability of Growth in Structural Plasticity

ResearchDGX agent

arXiv:2605.15435v1 Announce Type: new Abstract: Standard deep-learning pipelines usually choose the network architecture before training and keep it fixed throughout optimization. In contrast, a model

Online Algorithms for Repeated Optimal Stopping: Balancing Baseline Guarantees and Regret

ResearchDGX agent

arXiv:2511.04484v2 Announce Type: replace-cross Abstract: We study the repeated optimal stopping problem, in which the same optimal stopping instance with an unknown distribution is solved repeatedly

Online Vector Quantized Attention

ResearchDGX agent

arXiv:2602.03922v3 Announce Type: replace Abstract: Standard sequence mixing layers used in language models struggle to balance efficiency and performance. Self-attention performs well on long context

Overfitting has a limitation: a model-independent generalization gap bound based on Renyi entropy

ResearchDGX agent

arXiv:2506.00182v3 Announce Type: replace-cross Abstract: Will further scaling up of machine learning models continue to bring success? A significant challenge in answering this question lies in under

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

SafetyDGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

Perforated Neural Networks for Keyword Spotting

Model ReleasesDGX agent

arXiv:2605.15647v1 Announce Type: new Abstract: Edge machine learning presents a unique set of constraints not encountered in cloud-scale model deployment: strict memory budgets, limited compute, and

Pessimistic Risk-Aware Policy Learning in Contextual Bandits

SafetyDGX agent

arXiv:2605.15620v1 Announce Type: cross Abstract: We study risk-aware offline policy learning, aiming to learn a decision rule from logged data that is optimal under general risk criteria. This proble

phi-Balancing for Mixture-of-Experts Training

SafetyDGX agent

arXiv:2605.15403v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models rely on balanced expert utilization to fully realize their scalability. However, existing load-balancing methods are lar

Position: Ideas Should be the Center of Machine Learning Research

Model ReleasesDGX agent

arXiv:2605.15253v1 Announce Type: new Abstract: Machine learning research increasingly bifurcates into two disconnected modes: benchmark-driven engineering that prioritizes metrics over understanding,

Position: Zeroth-Order Optimization in Deep Learning Is Underexplored, Not Underpowered

ResearchDGX agent

arXiv:2605.15622v1 Announce Type: new Abstract: Zeroth-order (ZO) optimization, learning from finite differences of function evaluations without backpropagation, has recently regained attention in dee

Practical Validity Conditions for Byzantine-Tolerant Federated Learning

ResearchDGX agent

arXiv:2605.15887v1 Announce Type: new Abstract: Robust aggregation is the core operation in Byzantine-tolerant federated learning. To ensure the quality of aggregation independently of data distributi

PRB-RUPFormer: A Recursive Unified Probabilistic Transformer for Residual PRB Forecasting

ResearchDGX agent

arXiv:2605.15363v1 Announce Type: new Abstract: Accurate forecasting of residual Physical Resource Blocks (PRBs) is critical for proactive network slice provisioning, energy-efficient operation, and s

Preconditioned Regularized Wasserstein Proximal Sampling

SafetyDGX agent

arXiv:2509.01685v2 Announce Type: replace-cross Abstract: We consider sampling from a Gibbs distribution by evolving finitely many particles. We propose a preconditioned version of a recently proposed

Privacy Evaluation of Generative Models for Trajectory Generation

ResearchDGX agent

arXiv:2605.15246v1 Announce Type: new Abstract: Trajectory data is fundamental to modern urban intelligence, yet its sensitivity raises significant privacy concerns. Generative models such as Generati

Process-Informed Forecasting of Complex Thermal Dynamics in Pharmaceutical Manufacturing

ApplicationsDGX agent

arXiv:2509.20349v3 Announce Type: replace Abstract: Accurate time-series forecasting for complex physical systems is the backbone of modern industrial monitoring and control, yet deep learning models

Quantum Feature Pyramid Gating for Seismic Image Segmentation

Model ReleasesDGX agent

arXiv:2605.15370v1 Announce Type: cross Abstract: Accurate salt-body delineation is essential for seismic interpretation because salt structures distort wave propagation, complicate velocity-model bui

RanSOM: Second-Order Momentum with Randomized Scaling for Constrained and Unconstrained Optimization

SafetyDGX agent

arXiv:2602.06824v2 Announce Type: replace-cross Abstract: Momentum methods, such as Polyak's Heavy Ball, are the standard for training deep networks but suffer from curvature-induced bias in stochasti

RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach

SafetyDGX agent

arXiv:2603.18396v3 Announce Type: replace Abstract: Bus holding control is challenging due to stochastic traffic and passenger demand. While deep reinforcement learning (DRL) shows promise, standard a

Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2605.15239v1 Announce Type: new Abstract: Safety alignment often improves robustness to harmful queries at the cost of reasoning ability, a tradeoff known as the safety tax. A common cause is di

Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains

ResearchDGX agent

arXiv:2602.16274v2 Announce Type: replace Abstract: We present the first regret bound for classical online Q-learning in infinite-horizon discounted Markov decision processes (MDPs), without relying o

Reinforcement learning for adaptive interior point methods in convex quadratic programming

Model ReleasesDGX agent

arXiv:2509.07404v2 Announce Type: replace-cross Abstract: Quadratic programming is a workhorse of modern nonlinear optimization, control, and data science. Although regularized methods offer convergen

Rethinking Neural Network Learning Rates: A Stackelberg Perspective

ResearchDGX agent

arXiv:2605.15530v1 Announce Type: new Abstract: Neural networks are typically trained with a single learning rate across all layers. While recent empirical evidence suggests that assigning layer-speci

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers

Local AiDGX agent

arXiv:2505.12601v2 Announce Type: replace Abstract: As large language models (LLMs) grow in scale and specialization, routing--selecting the best model for a given input--has become essential for effi

Reweighting free energy profiles between universal machine learning interatomic potentials for fast consensus building

ApplicationsDGX agent

arXiv:2605.15630v1 Announce Type: cross Abstract: Free energy profiles serve as a fundamental bridge between microscopic atomic fluctuations and macroscopic thermodynamic observables. Estimating the f

← Previous
1…152153154155156…243
Next →