AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
19 May 2026

Can Adaptive Gradient Methods Converge under Heavy-Tailed Noise? A Case Study of AdaGrad

ApplicationsDGX agent

arXiv:2605.18694v1 Announce Type: cross Abstract: Many tasks in modern machine learning are observed to involve heavy-tailed gradient noise during the optimization process. To manage this realistic an

Can machine learning for quantum-gas experiments be explainable?

ResearchDGX agent

arXiv:2605.18689v1 Announce Type: cross Abstract: Virtually all aspects of many-body atomic physics are challenging: experiments are technically demanding, datasets have become enormous, and the memor

Canonical Regularisation of Wide Feature-Learning Neural Networks

SafetyDGX agent

arXiv:2605.18180v1 Announce Type: cross Abstract: Wide neural networks in the feature-learning regime drive modern deep learning, and yet they remain far less studied than their kernel-regime counterp


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CAST: Causal Anchored Simplex Transport for Distribution-Valued Time Series

ResearchDGX agent

arXiv:2605.16919v1 Announce Type: cross Abstract: Many decision-facing stochastic systems are observed through aggregate distributions rather than scalar trajectories: queue occupancies, mobility shar

Causal Anomaly Detection for Lithium-Ion Battery Degradation

Model ReleasesDGX agent

arXiv:2605.17334v1 Announce Type: cross Abstract: Reliable early detection of lithium-ion battery degradation requires health indicators that are physically interpretable and computable from routine c

Causal Influences over Social Learning Networks

AgentsDGX agent

arXiv:2307.09575v2 Announce Type: replace-cross Abstract: This paper investigates causal influences between agents linked by a social graph and interacting over time. In particular, the work examines

CayleyPy RL: Pathfinding and Reinforcement Learning on Cayley Graphs

Model ReleasesDGX agent

arXiv:2502.18663v3 Announce Type: replace Abstract: This paper is the second in a series of studies on developing efficient artificial intelligence-based approaches to pathfinding on extremely large g

ClaHF: A Human Feedback-inspired Reinforcement Learning Framework for Improving Classification Tasks

SafetyDGX agent

arXiv:2605.17458v1 Announce Type: new Abstract: Text classification models are typically trained via supervised fine-tuning (SFT). However, SFT essentially performs behavior cloning from instance-wise

Clipped Gradient Methods for Nonsmooth Convex Optimization under Heavy-Tailed Noise: A Refined Analysis

ResearchDGX agent

arXiv:2512.23178v3 Announce Type: replace-cross Abstract: Optimization under heavy-tailed noise has become popular recently, since it better fits many modern machine learning tasks, as captured by emp

Compass: SLO-aware Query Planner for Compound AI Serving at Scale

AgentsDGX agent

arXiv:2504.16397v2 Announce Type: replace-cross Abstract: The rise of compound AI serving that integrates multiple operators in a pipeline enables end-user applications such as generative AI-powered m

Consistency of Learned Sparse Grid Quadrature Rules using NeuralODEs

TutorialsDGX agent

arXiv:2507.01533v2 Announce Type: replace-cross Abstract: We prove consistency of a recently proposed scheme that evaluates expected values by composing a learned transport map with Clenshaw--Curtis s

Constrained Policy Optimization via Sampling-Based Weight-Space Projection

Model ReleasesDGX agent

arXiv:2512.13788v2 Announce Type: replace Abstract: Safety-critical learning requires policies that improve performance without leaving the safe operating regime. We study constrained policy learning

Convex Dataset Valuation for Post-Training

SafetyDGX agent

arXiv:2605.16704v1 Announce Type: new Abstract: Improving LLM performance on downstream tasks sometimes requires leveraging auxiliary datasets during post-training. In practice, however, developers fa

Coordinate Heterogeneity Governs Binary Quantization: From InfoNCE to Recall

Model ReleasesDGX agent

arXiv:2605.17524v1 Announce Type: new Abstract: Binary quantization (BQ) compresses high-dimensional embeddings into one or two bits per coordinate, enabling nearest neighbor search at extreme speed.

Corruptions of Supervised Learning Problems: Typology and Mitigations

ResearchDGX agent

arXiv:2307.08643v4 Announce Type: replace Abstract: Corruption is notoriously widespread in data collection. Despite extensive research, the existing literature predominantly focuses on specific setti

Cost-aware Duration Prediction for Software Upgrades in Datacenters

Model ReleasesDGX agent

arXiv:2212.05155v2 Announce Type: replace-cross Abstract: Software upgrades are critical to maintaining server reliability in datacenters. While job duration prediction and scheduling have been extens

Could Large Language Models work as Post-hoc Explainability Tools in Credit Risk Models?

Model ReleasesDGX agent

arXiv:2602.18895v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in translating model-based explanations into human-readable narratives. This study evaluates w

Counterfactual Explanations Under Concept Drift

ResearchDGX agent

arXiv:2605.17651v1 Announce Type: new Abstract: Counterfactual explanations (CFEs) provide actionable recourse, but most methods assume a static framework with fixed data and a trained classifier. Thi

CoX-MoE: Coalesced Expert Execution for High-Throughput MoE Inference with AMX-Enabled CPU-GPU Co-Execution

Model ReleasesDGX agent

arXiv:2605.17889v1 Announce Type: new Abstract: The Mixture-of-Experts (MoE) architecture improves computational efficiency via sparse expert activation, but throughput-oriented inference faces substa

CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning

Model ReleasesDGX agent

arXiv:2602.02979v2 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated strong potential in complex reasoning, yet their progress remains fundamentally constrained by reliance

DAD4TS: Data-Augmentation-Oriented Diffusion Model for Time-Series Forecasting with Small-Scale Data

ApplicationsDGX agent

arXiv:2605.17866v1 Announce Type: new Abstract: Small-scale data is a critical problem in time-series forecasting tasks. Data augmentation is an effective strategy for this task, but it has a limitati

Decision-Aware Proximal Bridge Learning for Optimal Treatment Selection

SafetyDGX agent

arXiv:2605.16989v1 Announce Type: new Abstract: Individualized treatment selection with continuous actions requires accurate causal response estimation in decision-relevant regions, rather than unifor

Decouple then Converge: Handling Unknown Unlabeled Distributions in Long-Tailed Semi-Supervised Learning

ApplicationsDGX agent

arXiv:2406.13187v2 Announce Type: replace Abstract: While long-tailed semi-supervised learning (LTSSL) has attracted growing attention in many real-world classification tasks, existing LTSSL algorithm

Decoupled Conformal Optimisation: Efficient Prediction Sets via Independent Tuning and Calibration

Model ReleasesDGX agent

arXiv:2605.18354v1 Announce Type: new Abstract: Bayesian conformal optimisation methods often use the same held-out data both to search for efficient prediction sets and to certify coverage or risk. T

Deep Learning-Based Channel Extrapolation for Dual-Band Massive MIMO Systems

ResearchDGX agent

arXiv:2601.06858v2 Announce Type: replace-cross Abstract: Future wireless communication systems will increasingly rely on the integration of millimeter wave (mmWave) and sub-6 GHz bands to meet hetero

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets

SafetyDGX agent

arXiv:2510.01479v2 Announce Type: replace Abstract: Offline reinforcement learning (RL) enables policy optimization from fixed datasets, making it suitable for safety-critical applications where onlin

Differentiable Optimization Layers for Guaranteed Fairness in Deep Learning

SafetyDGX agent

arXiv:2605.17118v1 Announce Type: new Abstract: Differentiable optimization layers are traditionally integrated in predict-then-optimize frameworks where a neural model estimates parameters that subse

Diffusion-Based Stochastic Operator Networks for Uncertainty Quantification in Stochastic Partial Differential Equations

Model ReleasesDGX agent

arXiv:2605.17107v1 Announce Type: cross Abstract: We introduce a novel framework for uncertainty quantification of solution operators associated with stochastic partial differential equations (SPDEs).

Dimension-Free Convergence of Discrete Diffusion Models: Adjoint Equations Induce the Right Space

ResearchDGX agent

arXiv:2605.17232v1 Announce Type: new Abstract: Discrete diffusion has become a leading framework for generative modeling in various applications including language, vision, and biology. Existing conv

Dimension-Uniform Discretization Analysis of Preconditioned Annealed Langevin Dynamics for Multimodal Gaussian Mixtures

ResearchDGX agent

arXiv:2605.16473v1 Announce Type: cross Abstract: Obtaining stable diffusion-based samplers in high- and infinite-dimensional settings is challenging because errors can accumulate across high-frequenc

Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs

Model ReleasesDGX agent

arXiv:2603.12676v2 Announce Type: replace Abstract: Generalizing neural surrogate models across different PDE parameters remains difficult because changes in PDE coefficients often make learning harde

Distributed Perceptron under Bounded Staleness, Partial Participation, and Noisy Communication

Model ReleasesDGX agent

arXiv:2601.10705v3 Announce Type: replace Abstract: We study a semi-asynchronous client-server perceptron trained via iterative parameter mixing (IPM-style averaging): clients run local perceptron upd

Distribution Transformers: Fast Approximate Bayesian Inference With On-The-Fly Prior Adaptation

Model ReleasesDGX agent

arXiv:2502.02463v3 Announce Type: replace-cross Abstract: While Bayesian inference provides a principled framework for reasoning under uncertainty, its widespread adoption is limited by the intractabi

Does Weight Decay Enhance Training Stability?

Model ReleasesDGX agent

arXiv:2605.16622v1 Announce Type: new Abstract: In modern deep learning, weight decay is often credited with 'stabilizing' training dynamics, diverging from its classical role as a static regularizati

DP-SelFT: Differentially Private Selective Fine-Tuning for Large Language Models

Model ReleasesDGX agent

arXiv:2605.17432v1 Announce Type: new Abstract: Large language models (LLMs) are commonly adapted to downstream tasks through fine-tuning, but fine-tuning data often contains sensitive information tha

DyDiff: Long-Horizon Rollout via Dynamics Diffusion for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2405.19189v3 Announce Type: replace Abstract: With the great success of diffusion models (DMs) in generating realistic synthetic vision data, many researchers have investigated their potential i

DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization

SafetyDGX agent

arXiv:2605.17486v1 Announce Type: cross Abstract: Recent progress in Reinforcement Learning (RL) provides a principled approach to optimizing Vision-Language-Action (VLA) models, facilitating a shift

Dynamic Elliptical Graph Factor Models via Riemannian Optimization with Geodesic Temporal Regularization

Model ReleasesDGX agent

arXiv:2605.18316v1 Announce Type: new Abstract: Inferring time-varying graph structures from high-dimensional nodal observations is a fundamental problem arising in neuroscience, finance, climatology,

Dynamic robotic cloth folding with efficient Koopman operator-based model predictive control

ResearchDGX agent

arXiv:2605.18373v1 Announce Type: cross Abstract: Robotic cloth folding is a challenging task, particularly when considering dynamic folding tasks, which aim at folding cloth by fast motions that leve

Efficient and Noise-Tolerant PAC Learning of Multiclass Linear Classifiers

ResearchDGX agent

arXiv:2605.18662v1 Announce Type: new Abstract: Noise-tolerant PAC learning of linear models has been of central interests in machine learning community since the last century. In recent years, many c

Elastic-dLLM: Position Preserving Context Compression and Augmentation of Diffusion LLMs

ResearchDGX agent

arXiv:2605.18165v1 Announce Type: new Abstract: Unlike autoregressive models, which generate one token at a time, dLLMs denoise a chunk of [MASK] tokens jointly and sample one or more tokens per step;

Empirical evaluation of Time Series Foundation Models for Day-ahead and Imbalance Electricity Price Forecasting in Belgium

ResearchDGX agent

arXiv:2605.17045v1 Announce Type: cross Abstract: Recent advances in Time Series Foundation Models (TSFMs) promise zero-shot forecasting capabilities with minimal task-specific training. While these m

Emulating the Forced Response of Climate Models with Flow Matching

ResearchDGX agent

arXiv:2605.16929v1 Announce Type: new Abstract: Global climate models are essential tools to simulate past and potential future pathways of climate change, as well as associated climate impacts. Share

Enhancing the Code Reasoning Capabilities of LLMs via Consistency-based Reinforcement Learning

ResearchDGX agent

arXiv:2605.17958v1 Announce Type: new Abstract: Code reasoning refers to the task of predicting the output of a program given its source code and specific inputs. It can measure the reasoning capabili

Equilibrium Selection in Multi-Agent Policy Gradients via Opponent-Aware Basin Entry

SafetyDGX agent

arXiv:2605.18078v1 Announce Type: new Abstract: Multi-agent policy-gradient methods have been shown to converge locally near stable Nash equilibria. Local convergence, however, does not determine whic

Evaluating Inter-Column Logical Relationships in Synthetic Tabular Data Generation

ApplicationsDGX agent

arXiv:2502.04055v2 Announce Type: replace Abstract: Current evaluations of synthetic tabular data mainly focus on how well joint distributions are modeled, often overlooking the assessment of their ef

EvilGenie: A Reward Hacking Benchmark

Model ReleasesDGX agent

arXiv:2511.21654v2 Announce Type: replace Abstract: We introduce EvilGenie, a benchmark for reward hacking in programming settings. We source problems from LiveCodeBench and create an environment in w

Exact Convex Reformulations of Linear Neural Networks via Completely Positive Lifting

ResearchDGX agent

arXiv:2605.17692v1 Announce Type: new Abstract: We show that the training problem of a deep linear neural network under the squared loss admits an exact convex reformulation in a lifted space over a g

Factored Causal Representation Learning for Robust Reward Modeling in RLHF

SafetyDGX agent

arXiv:2601.21350v2 Announce Type: replace Abstract: A reliable reward model is essential for aligning large language models with human preferences through reinforcement learning from human feedback. H

Fast Rates for Nonstationary Weighted Risk Minimization

ResearchDGX agent

arXiv:2602.05742v2 Announce Type: replace-cross Abstract: Weighted empirical risk minimization is a common approach to prediction under distribution drift. This article studies its out-of-sample predi

Feature Learning in Linear-Width Two-Layer Networks: Two vs. One Step of Gradient Descent

SafetyDGX agent

arXiv:2605.17767v1 Announce Type: cross Abstract: We study feature learning in two-layer neural networks within the linear-width regime, where the number of hidden neurons, sample size, and input dime

Federated Distillation on Edge Devices: Efficient Client-Side Filtering for Non-IID Data

ApplicationsDGX agent

arXiv:2508.14769v2 Announce Type: replace Abstract: Federated distillation has emerged as a promising collaborative machine learning approach, offering enhanced privacy protection and reduced communic

Federated Learning by Utility-Constrained Stochastic Aggregation for Improving Rational Participation

Local AiDGX agent

arXiv:2605.18020v1 Announce Type: new Abstract: Federated Learning (FL) algorithms implicitly assume that clients passively comply with server-side orchestration by sharing local model updates upon se

Federated Martingale Posterior Samping

Model ReleasesDGX agent

arXiv:2605.18554v1 Announce Type: new Abstract: Federated Bayesian neural networks require fixing a prior on the model parameters together with a likelihood. Eliciting meaningful priors on the weight

FEG-Pro: Forecast-Error Growth Profiling for Finite-Horizon Instability Analysis of Nonlinear Time Series

ResearchDGX agent

arXiv:2605.17282v1 Announce Type: cross Abstract: Estimating the largest Lyapunov exponent from a scalar time series is difficult when the governing equations, tangent dynamics, and full state vector

Filter-then-Verify: A Multiphase GNN and ModernBERT Framework for Social Engineering Detection in Email Networks

ResearchDGX agent

arXiv:2605.17201v1 Announce Type: cross Abstract: Social engineering attacks exploit human trust rather than software vulnerabilities, making them difficult to detect using conventional filters. We pr

Fine-grained List-wise Alignment for Generative Medication Recommendation

Model ReleasesDGX agent

arXiv:2505.20218v2 Announce Type: replace Abstract: Accurate and safe medication recommendations are critical for effective clinical decision-making, especially in multimorbidity cases. However, exist

Finite-Particle Rates for Regularized Stein Variational Gradient Descent

Model ReleasesDGX agent

arXiv:2602.05172v2 Announce Type: replace-cross Abstract: We derive finite-particle rates for the regularized Stein variational gradient descent (R-SVGD) algorithm introduced by He et al. (2024) that

FLEX-MoE: Federated Mixture-of-Experts with Load-balanced Expert Assignment for Edge Computing

ResearchDGX agent

arXiv:2512.23070v2 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models enable scalable neural networks through conditional computation, offering enhanced effectiveness and efficiency for

FlowMixer: A Depth-Agnostic Neural Architecture for Interpretable Spatiotemporal Forecasting

ResearchDGX agent

arXiv:2505.16786v2 Announce Type: replace Abstract: We introduce FlowMixer, a single-layer neural architecture that leverages constrained matrix operations to model structured spatiotemporal patterns

← Previous
1…146147148149150…243
Next →