AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
21 May 2026

Neural Estimation of Pairwise Mutual Information in Masked Discrete Sequence Models

ResearchDGX agent

arXiv:2605.20187v1 Announce Type: new Abstract: Understanding dependencies between variables is critical for interpretability and efficient generation in masked diffusion models (MDMs), yet these mode

Neural Negative Binomial Regression for Weekly Seismicity Forecasting: Per-Cell Dispersion Estimation and Tail Risk Assessment

Model ReleasesDGX agent

arXiv:2605.21437v1 Announce Type: cross Abstract: Standard approaches to forecasting the weekly number of earthquakes on a spatial grid rely on the Poisson distribution with a single global dispersion

Nonlocal operator learning for fMRI encoding and decoding tasks

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.20389v1 Announce Type: new Abstract: Functional MRI data exhibit high-dimensional spatiotemporal structure, making both prediction and decoding challenging. In this work, we investigate neu

Nonparametric Learning and Earning with One-Point Feedback under Nonstationarity

Model ReleasesDGX agent

arXiv:2605.21263v1 Announce Type: new Abstract: Firms increasingly rely on dynamic pricing to respond to evolving customer demand, yet in many applications they observe only the revenue generated by a

OCTOPUS: Optimized KV Cache for Transformers via Octahedral Parametrization Under optimal Squared error quantization

ResearchDGX agent

arXiv:2605.21226v1 Announce Type: new Abstract: The key-value (KV) cache dominates memory bandwidth and footprint in long-context autoregressive inference. Recent rotation-preconditioned codecs (Turbo

OmniISR: A Unified Framework for Centralized and Federated Learning via Intermediate Supervision and Regularization

Local AiDGX agent

arXiv:2605.20276v1 Announce Type: new Abstract: The global deployment of edge intelligence operates across heterogeneous legal frameworks. While some regions permit centralized learning (CL) via cloud

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective

ResearchDGX agent

arXiv:2605.21260v1 Announce Type: new Abstract: We develop a learning-theoretic framework for understanding Chain of Thought (CoT). We model CoT as the interaction between an answer map and a chain ru

On the Regularity and Generalization of One-Step Wasserstein-guided Generative Models for PDE-Induced Measures

TutorialsDGX agent

arXiv:2605.21388v1 Announce Type: new Abstract: Despite the remarkable empirical success of generative models, the available theory on their statistical accuracy in scientific computing remains largel

On the Suboptimality of GP-UCB under Polynomial Effective Optimism

Model ReleasesDGX agent

arXiv:2312.01386v2 Announce Type: replace Abstract: Gaussian process upper confidence bound (GP-UCB) is widely used for sequential optimization of expensive black-box functions. Although many upper bo

One Operator to Rule Them All? On Boundary-Indexed Operator Families in Neural PDE Solvers

ResearchDGX agent

arXiv:2603.01406v2 Announce Type: replace Abstract: Neural PDE solvers are often described as learning solution operators that map problem data to PDE solutions. In this work, we argue that this inter

Online Conformal Prediction with Corrupted Feedback

ApplicationsDGX agent

arXiv:2605.20515v1 Announce Type: new Abstract: Modern artificial intelligence systems require calibrated uncertainty estimates that remain reliable in sequential and non-stationary environments. Onli

OpenSeisML: Open Large-Scale Real Seismic and well-log Dataset for Generative AI

ResearchDGX agent

arXiv:2605.20539v1 Announce Type: new Abstract: The advent of machine learning (ML) and computer vision has significantly accelerated seismic inversion workflows by reducing the computational cost of

Optimization Hyper-parameter Laws for Large Language Models

Model ReleasesDGX agent

arXiv:2409.04777v4 Announce Type: replace Abstract: Large Language Models have driven significant AI advancements, yet their training is resource-intensive and highly sensitive to hyper-parameter sele

Optimized Federated Knowledge Distillation with Distributed Neural Architecture Search

Local AiDGX agent

arXiv:2605.21322v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative model training without centralizing data. However, real-world deployments must simultaneously address stat

PACD-Net: Pseudo-Augmented Contrastive Distillation for Glycemic Control Estimation from SMBG

TutorialsDGX agent

arXiv:2605.20751v1 Announce Type: new Abstract: Effective diabetes management requires continuous monitoring of glycemic levels. Clinically, glycemic control is assessed using metrics such as Time in

Personalized Weight Loss Management through Wearable Devices and Artificial Intelligence

ApplicationsDGX agent

arXiv:2409.08700v2 Announce Type: replace Abstract: Early detection of chronic and Non-Communicable Diseases (NCDs) is crucial for effective treatment during the initial stages. This study explores th

Physics-informed convolutional neural networks for fluid flow through porous media

ResearchDGX agent

arXiv:2605.20250v1 Announce Type: new Abstract: Accurate simulation of fluid flow in porous media is challenging due to complex pore-space geometries and the computational cost of solving the Navier-S

PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Models

Model ReleasesDGX agent

arXiv:2605.20873v1 Announce Type: cross Abstract: Planning is a fundamental capability for large language models (LLMs) because such complex tasks require models to coordinate goals, constraints, reso

PlexRL: Cluster-Level Orchestration of Serviceized LLM Execution for RLVR

HardwareDGX agent

arXiv:2605.20863v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has recently unlocked strong reasoning capabilities in large language models (LLMs), triggering

Plug-and-Play Spiking Operators: Breaking the Nonlinearity Bottleneck in Spiking Transformers

ResearchDGX agent

arXiv:2605.20289v1 Announce Type: new Abstract: ANN-to-SNN conversion offers a practical, training-free route to spiking large language models. However, current pipelines primarily focus on spike-driv

Point Cloud Sequence Encoding for Material-conditioned Graph Network Simulators

Model ReleasesDGX agent

arXiv:2605.20978v1 Announce Type: new Abstract: Graph Network Simulators (GNSs) have emerged as powerful surrogates for complex physics-based simulation, offering inherent differentiability and orders

Polynomial-Time Robust Multiclass Linear Classification under Gaussian Marginals

ResearchDGX agent

arXiv:2605.21428v1 Announce Type: new Abstract: We study the task of agnostic learning of multiclass linear classifiers under the Gaussian distribution. Given labeled examples (x, y) from a distributi

Praxium: Diagnosing Cloud Anomalies with AI-based Telemetry and Dependency Analysis

ResearchDGX agent

arXiv:2603.23890v2 Announce Type: replace-cross Abstract: As the modern microservice architecture for cloud applications grows in popularity, cloud services are becoming increasingly complex and more

Preference-aware Influence-function-based Data Selection Method for Efficient Fine-Tuning

SafetyDGX agent

arXiv:2605.21422v1 Announce Type: new Abstract: As LLMs continue to scale, improving training efficiency increasingly depends on using data more effectively. Data selection addresses this problem by a

PREFINE: Preference-Based Implicit Reward and Cost Fine-Tuning for Safety Alignment

SafetyDGX agent

arXiv:2605.21225v1 Announce Type: new Abstract: We address the problem of making a pre-trained reinforcement learning (RL) policy safety-aware by incorporating cost constraints without retraining it f

PrefixWall: Mitigating Prefix Caching Side Channels in Shared LLM Systems

ResearchDGX agent

arXiv:2603.10726v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) rely on optimizations like Automatic Prefix Caching (APC) to accelerate inference. APC works by reusing previousl

Prism: Structural Symmetry Scanning via Duality-Constrained Laplacian Projection

ResearchDGX agent

arXiv:2605.20245v1 Announce Type: cross Abstract: We introduce extbf{Prism}, a framework for structural symmetry diagnosis in complex networks. Given a graph Laplacian L and a duality operator P (a sy

Provably Learning Diffusion Models under the Manifold Hypothesis: Collapse and Refine

ResearchDGX agent

arXiv:2605.20235v1 Announce Type: new Abstract: Diffusion models generate high-dimensional data with remarkable quality, yet how their training efficiently learns the score function, bypassing the cur

Proximal State Nudging: Reducing Skill Atrophy from AI Assistance

SafetyDGX agent

arXiv:2605.20355v1 Announce Type: cross Abstract: Skill atrophy, the gradual decline of human capability under AI assistance, poses a safety risk in shared-control of semi-autonomous systems, where op

Pseudo-Formalization for Automatic Proof Verification

Model ReleasesDGX agent

arXiv:2605.20531v1 Announce Type: cross Abstract: Reliable verification of proofs remains a bottleneck for training and evaluating AI systems on hard mathematical reasoning. Fully formal proofs, in la

Q-Net: Queue Length Estimation via Kalman-based Neural Networks

TutorialsDGX agent

arXiv:2509.24725v3 Announce Type: replace Abstract: Estimating queue lengths at signalized intersections is a long-standing challenge in traffic management. Partial observability of vehicle flows comp

Q-SYNTH: Hybrid Quantum-Classical Adversarial Augmentation for Imbalanced Fraud Detection

ResearchDGX agent

arXiv:2605.21164v1 Announce Type: new Abstract: Credit card fraud detection is fundamentally challenged by extreme class imbalance, where fraudulent transactions are rare yet operationally critical. T

Quadratic Characterizations for Reachability Analysis of Neural Networks

SafetyDGX agent

arXiv:2605.20482v1 Announce Type: new Abstract: Quadratic constraints (QCs) are widely used to characterize nonlinearities and uncertainties, but generic analytical characterizations can be conservati

Quantifying Hyperparameter Transfer and the Importance of Embedding Layer Learning Rate

Model ReleasesDGX agent

arXiv:2605.21486v1 Announce Type: new Abstract: Hyperparameter transfer allows extrapolating optimal optimization hyperparameters from small to large scales, making it critical for training large lang

Quant.npu: Enabling Efficient Mobile NPU Inference for on-device LLMs via Fully Static Quantization

Local AiDGX agent

arXiv:2605.20295v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed on mobile devices, where Neural Processing Units (NPUs) necessitate fully static quantization for

Quantum End-to-End Learning for Contextual Combinatorial Optimization

SafetyDGX agent

arXiv:2605.20222v1 Announce Type: cross Abstract: Contextual combinatorial optimization (CCO) plays a critical role in decision-making under uncertainty, yet remains a significant challenge. We presen

Quantum reservoir computing in Jaynes-Cummings models: Nonlinear memory and time-series prediction

Model ReleasesDGX agent

arXiv:2510.00171v2 Announce Type: replace-cross Abstract: We investigate quantum reservoir computing (QRC) using a hybrid qubit-boson system described by the Jaynes-Cummings (JC) Hamiltonian and its d

Reasoning-Trace Collapse: Evaluating the Loss of Explicit Reasoning During Fine-Tuning

ResearchDGX agent

arXiv:2605.21127v1 Announce Type: new Abstract: Explicit reasoning models are trained to produce intermediate reasoning traces before final answers, but downstream fine-tuning is often performed on or

REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak

SafetyDGX agent

arXiv:2605.20654v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable capabilities, they remain susceptible to sophisticated, multi-step jailbreak attacks that circ

Reinforcement Learning-based Control via Y-wise Affine Neural Networks: Comparative Case Studies for Chemical Processes

ResearchDGX agent

arXiv:2605.21211v1 Announce Type: cross Abstract: In this work we present an efficient and practically implementable approach for the application of reinforcement learning (RL)-based control in chemic

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces

SafetyDGX agent

arXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov

rePIRL: Learn PRM with Inverse RL for LLM Reasoning

SafetyDGX agent

arXiv:2602.07832v2 Announce Type: replace Abstract: Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In

Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing

Model ReleasesDGX agent

arXiv:2605.20262v1 Announce Type: new Abstract: We study selective refusal editing as a three-way control problem: induce non-refusal on designated edit prompts while preserving benign behavior and ha

ReversedQ: Opportunities for Faster Q-Learning in Episodic Online Reinforcement Learning

ResearchDGX agent

arXiv:2605.20592v1 Announce Type: new Abstract: We study model-free Q-learning in finite-horizon episodic Markov Decision Processes (MDPs) with stationary dynamics across episodes. We identify a centr

Reviving Error Correction in Modern Deep Time-Series Forecasting

ResearchDGX agent

arXiv:2605.21088v1 Announce Type: new Abstract: Modern deep-learning models have achieved remarkable success in time-series forecasting. Yet, their performance degrades in long-term prediction due to

Riemannian MeanFlow for One-Step Generation on Manifolds

ResearchDGX agent

arXiv:2603.10718v2 Announce Type: replace Abstract: Flow Matching enables simulation-free training of generative models on Riemannian manifolds, yet sampling typically still relies on numerically inte

Robust Personalized Recommendation under Hidden Confounding in MNAR

Model ReleasesDGX agent

arXiv:2605.21066v1 Announce Type: new Abstract: Recommender systems often rely on observational user--item interaction data, which is prone to selection bias due to users' selective interactions with

Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework

SafetyDGX agent

arXiv:2605.20721v1 Announce Type: new Abstract: Learning from implicit feedback in recommender systems is fundamentally challenged by pervasive label noise. While conventional denoising approaches oft

Robust Subspace-Constrained Quadratic Models for Low-Dimensional Structure Learning

ResearchDGX agent

arXiv:2605.20300v1 Announce Type: new Abstract: In this paper, we propose a robust subspace-constrained quadratic model (SCQM) for learning low-dimensional structure from high-dimensional data. Buildi

roto 2.0: The Robot Tactile Olympiad

Model ReleasesDGX agent

arXiv:2605.21429v1 Announce Type: cross Abstract: Tactile-based reinforcement learning (RL) is currently hindered by fragmented research and a focus on over-saturated orientation tasks. We introduce v

Runtime-Certified Bounded-Error Quantized Attention

Model ReleasesDGX agent

arXiv:2605.20868v1 Announce Type: new Abstract: KV cache quantization reduces the memory cost of long-context LLM inference, but introduces approximation error that is typically validated only empiric

Same Target, Different Basins: Hard vs. Soft Labels for Annotator Distributions

ResearchDGX agent

arXiv:2605.20642v1 Announce Type: new Abstract: When annotators disagree, that disagreement can reflect epistemic uncertainty rather than simple label noise. We study hard-label delivery as an alterna

Sample Complexity of Transfer Learning: An Optimal Transport Approach

ResearchDGX agent

arXiv:2605.20545v1 Announce Type: cross Abstract: Transfer learning is an essential technique for many machine learning/AI models of complex structures such as large language models and generative AI.

Scale-Calibrated Median-of-Means for Robust Distributed Principal Component Analysis

Local AiDGX agent

arXiv:2605.20681v1 Announce Type: cross Abstract: Distributed principal component analysis (PCA) produces node-level estimates of both a mean vector and a principal subspace. Robustly aggregating thes

Score-Based Causal Discovery of Latent Variable Causal Models

ResearchDGX agent

arXiv:2605.20396v1 Announce Type: new Abstract: Identifying latent variables and the causal structure involving them is essential across various scientific fields. While many existing works fall under

Secure, Verifiable, and Scalable Multi-Client Data Sharing via Consensus-Based Privacy-Preserving Data Distribution

SafetyDGX agent

arXiv:2601.00418v2 Announce Type: replace-cross Abstract: We propose the Consensus-Based Privacy-Preserving Data Distribution (CPPDD) framework, a lightweight and post-setup autonomous protocol for se

Self-Improving Skill Learning for Robust Skill-based Meta-Reinforcement Learning

ResearchDGX agent

arXiv:2502.03752v5 Announce Type: replace Abstract: Meta-reinforcement learning (Meta-RL) facilitates rapid adaptation to unseen tasks but faces challenges in long-horizon environments. Skill-based ap

Semiparametric Efficient Bilevel Gradient Estimation

Model ReleasesDGX agent

arXiv:2605.21341v1 Announce Type: cross Abstract: Functional bilevel methods estimate a lower-level function and plug it into a hypergradient, but this plug-in gradient can retain first-order bias whe

Sequential Data Augmentation for Generative Recommendation

Model ReleasesDGX agent

arXiv:2509.13648v3 Announce Type: replace Abstract: Generative recommendation plays a crucial role in personalized systems, predicting users' future interactions from their historical behavior sequenc

ShapeBench: A Scalable Benchmark and Diagnostic Suite for Standardized Evaluation in Aerodynamic Shape Optimization

Model ReleasesDGX agent

arXiv:2605.20763v1 Announce Type: new Abstract: Rapid progress in aerodynamic shape optimization (ASO) has outpaced currently-available standardized evaluation frameworks. Fair comparison requires a u

← Previous
1…140141142143144…243
Next →