AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
19 May 2026

Proximal basin hopping: global optimization with guarantees

ResearchDGX agent

arXiv:2605.18364v1 Announce Type: new Abstract: Global optimization is a challenging problem, with plenty of algorithms displaying empirical success, but scarce theoretical backing. In this work, we p

Proximal-IMH: Proximal Posterior Proposals for Independent Metropolis-Hastings with Approximate Operators

Local AiDGX agent

arXiv:2602.21426v2 Announce Type: replace Abstract: We consider the problem of sampling from a posterior distribution arising in Bayesian inverse problems in science, engineering, and imaging. Our met

Prune, Update and Trim: Robust Structured Pruning for Large Language Models

ResearchDGX agent

arXiv:2605.18331v1 Announce Type: new Abstract: Large Language Models (LLMs) have experienced significant growth and development in recent years. However, performing inference on LLMs remains costly,


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Public-Decay Homomorphic State Space Models for Private Sequence Inference

Local AiDGX agent

arXiv:2605.16647v1 Announce Type: cross Abstract: Fully homomorphic encryption (FHE) changes sequence-model design because rotations, encrypted products, ciphertext materialization, multiplicative dep

PULSE: Generative Phase Evolution for Non-Stationary Time Series Forecasting

SafetyDGX agent

arXiv:2605.16793v1 Announce Type: new Abstract: Time series forecasting under non-stationarity faces a fundamental tension between capturing stable representations and adapting to distribution shifts.

pyforce-1.0.0: Python Framework for data-driven model Order Reduction of multi-physiCs problEms

ResearchDGX agent

arXiv:2605.18082v1 Announce Type: new Abstract: pyforce is a Python package implementing Data-Driven Reduced Order Modelling techniques for applications to multi-physics problems, mainly set in the Nu

Q-LocalAdam: Memory-Efficient Client-Side Adaptive Optimization for Edge Federated Learning

Local AiDGX agent

arXiv:2605.17552v1 Announce Type: new Abstract: Federated learning on edge devices must cope with non-IID client data and tight memory budgets. Adaptive optimizers like Adam stabilize training under d

QLIF-CAST: Quantum Leaky-Integrate-and-Fire for Time-Series Weather Forecasting

Model ReleasesDGX agent

arXiv:2605.18333v1 Announce Type: cross Abstract: Accurate and efficient time-series forecasting remains a challenging problem for both classical and quantum neural architectures, particularly in mult

QuadraSHAP: Stable and Scalable Shapley Values for Product Games via Gauss-Legendre Quadrature

ResearchDGX agent

arXiv:2605.05870v2 Announce Type: replace Abstract: We study the efficient computation of Shapley values for product games -- cooperative games in which the coalition value factorizes as a product of

Quantitative Linear Logic for Neuro-Symbolic Learning and Verification

ResearchDGX agent

arXiv:2605.13845v2 Announce Type: cross Abstract: Differentiable Logics are deployed in neuro-symbolic learning tasks as a way of embedding logical constraints in the training objective of neural netw

QuChaTeR: A Hybrid Quantum-Chaotic Temporal Framework for Earthquake Prediction

ApplicationsDGX agent

arXiv:2605.16454v1 Announce Type: new Abstract: Seismic prediction remains challenging due to the highly nonlinear and chaotic dynamics of earthquake signals. While classical deep learning models such

Queue Length Regret Bounds for Contextual Queueing Bandits

Model ReleasesDGX agent

arXiv:2601.19300v2 Announce Type: replace Abstract: We introduce contextual queueing bandits, a new context-aware framework for scheduling while simultaneously learning unknown service rates. Individu

R2V Agent: Teaching SLMs When to Ask for Help

Local AiDGX agent

arXiv:2605.16604v1 Announce Type: new Abstract: Efficient agentic systems should incur expensive frontier-model costs only on decisions where a cheaper local model is likely to fail. Existing LLM casc

Ranking-Aware Calibration for Reliable Multimodal Reinforcement Learning

SafetyDGX agent

arXiv:2605.16999v1 Announce Type: new Abstract: Reinforcement learning post-training has substantially improved the reasoning accuracy of vision-language models, yet the resulting policies remain poor

Re-analysis of the Human Transcription Factor Atlas Recovers TF-Specific Signatures from Pooled Single-Cell Screens with Missing Controls

ResearchDGX agent

arXiv:2604.02511v2 Announce Type: replace Abstract: Public pooled single-cell perturbation atlases are valuable resources for studying transcription factor (TF) function, but downstream re-analysis ca

Real-time Multi-instrument Autonomous Discovery of Novel Phase-change Memory Materials

AgentsDGX agent

arXiv:2605.18033v1 Announce Type: cross Abstract: Autonomous labs enable the integration of automated experiment execution, data analysis and decision making. The main challenge remains the integratio

Reasoning as Compression: Unifying Budget Forcing via the Conditional Information Bottleneck

ResearchDGX agent

arXiv:2603.08462v2 Announce Type: replace Abstract: ac{CoT} prompting improves LLM accuracy on complex tasks but often increases token usage and inference cost. Existing ``Budget Forcing'' methods red

Revisiting the Adam-SGD Gap in LLM Pre-Training: The Role of Large Effective Learning Rates

Model ReleasesDGX agent

arXiv:2605.17787v1 Announce Type: new Abstract: It is widely believed that stochastic gradient descent (SGD) performs significantly worse than adaptive optimizers such as Adam in pre-training Large La

RIE-Greedy: Regularization-Induced Exploration for Contextual Bandits

Model ReleasesDGX agent

arXiv:2603.11276v2 Announce Type: replace-cross Abstract: Real-world contextual bandit problems with complex reward models are often tackled with iteratively trained models, such as boosting trees. Ho

Ringmaster LMO: Asynchronous Linear Minimization Oracle Momentum Method

Model ReleasesDGX agent

arXiv:2605.18174v1 Announce Type: new Abstract: Muon has recently emerged as a strong alternative to AdamW for training neural networks, with encouraging large-scale pretraining results and growing ev

RL4RLA: Teaching ML to Discover Randomized Linear Algebra Algorithms Through Curriculum Design and Graph-Based Search

SafetyDGX agent

arXiv:2605.18004v1 Announce Type: new Abstract: Randomized linear algebra (RLA) algorithms are a modern class of numerical linear algebra techniques that play an essential role in scientific computing

Robust Player-Conditional Champion Ranking for League of Legends: Style Similarity, Mastery Priors, and Archetype-Constrained Discovery

ApplicationsDGX agent

arXiv:2605.18338v1 Announce Type: cross Abstract: Champion recommendation in multiplayer online battle arena games is usually framed informally as a problem of metagame strength, personal comfort, or

S2Aligner: Pair-Efficient and Transferable Pre-Training for Sparse Text-Attributed Graphs

SafetyDGX agent

arXiv:2605.18579v1 Announce Type: new Abstract: Pre-training on text-attributed graphs (TAGs) is central to building transferable graph foundation models, where LLM-as-Aligner methods align graph and

Sample efficient inductive matrix completion with noise and inexact side information

ApplicationsDGX agent

arXiv:2605.17189v1 Announce Type: cross Abstract: Low-rank matrix completion is a widely studied problem with many variants. Inductive matrix completion (IMC) incorporates row and column side informat

SC3D: Dynamic and Differentiable Causal Discovery for Temporal and Instantaneous Graphs

ApplicationsDGX agent

arXiv:2602.02830v3 Announce Type: replace Abstract: Discovering causal structures from multivariate time series is a key problem because interactions span across multiple lags and possibly involve ins

Scalable Bi-causal Optimal Transport via KL Relaxation and Policy Gradients

SafetyDGX agent

arXiv:2605.17271v1 Announce Type: cross Abstract: Bi-causal optimal transport (OT) is a natural framework for comparing and coupling stochastic processes under nonanticipative information constraints,

Scalable Decision-Focused Learning through Cost-Sensitive Regression

ApplicationsDGX agent

arXiv:2605.18005v1 Announce Type: new Abstract: Many real-world combinatorial problems involve uncertain parameters, which can be predicted given contextual features and historical data. These `predic

Scalable Knowledge Editing for Mixture-of-Experts LLMs via Tensor-Structured Updates

Model ReleasesDGX agent

arXiv:2605.16686v1 Announce Type: new Abstract: Knowledge editing (KE) provides a lightweight alternative to repeated fine-tuning of LLMs. However, most existing KE methods target dense feed-forward l

Scale-Equivariant Generative Forecasting: Weight-Tied Dilated Convolutions, Wavelet Scattering Inputs, and Spectral-Consistency Training for Self-Similar Time Series

Model ReleasesDGX agent

arXiv:2605.17582v1 Announce Type: new Abstract: Many natural and engineered time series -- equity returns, climate anomalies, turbulent velocities, neural recordings, packet-level network traffic -- a

Scale-Invariant Neural Network Optimization: Norm Geometry and Heavy-Tailed Noise

ResearchDGX agent

arXiv:2605.18528v1 Announce Type: cross Abstract: A growing lesson from neural network optimization is that optimizer design should respect how the model is parametrized. Scale-invariant methods becom

scHelix: Asymmetric Dual-Stream Integration via Explicit Gene-Level Disentanglement

TutorialsDGX agent

arXiv:2605.18576v1 Announce Type: new Abstract: A critical challenge in single-cell RNA sequencing (scRNA-seq) integration is resolving the tension between eliminating batch effects and maintaining bi

SCOUT: Cyclic Causal Discovery Under Soft Interventions with Unknown Targets

ApplicationsDGX agent

arXiv:2605.16620v1 Announce Type: new Abstract: Learning causal relationships between variables from data is a fundamental research area with many applications across disciplines. Most existing causal

SE-GA: Memory-Augmented Self-Evolution for GUI Agents

Model ReleasesDGX agent

arXiv:2605.16883v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents often struggle with multi-step tasks due to constrained context windows and static policies that fail t

Self-Distillation is Optimal Among Spectral Shrinkage Estimators in Spiked Covariance Models

ResearchDGX agent

arXiv:2605.17778v1 Announce Type: cross Abstract: Self-distillation has emerged as a promising technique for improving model performance in modern machine learning systems. We develop the statistical

Self-Supervised Learning for Sparse Matrix Reordering

ResearchDGX agent

arXiv:2605.17403v1 Announce Type: new Abstract: Rearranging the rows or columns of a sparse matrix using an appropriate ordering can significantly reduce fill-ins, i.e., new nonzeros introduced during

Self-supervised local learning rules learn the hidden hierarchical structure of high-dimensional data

TutorialsDGX agent

arXiv:2605.18557v1 Announce Type: new Abstract: The brain learns abstract representations of high-dimensional sensory input, but the plasticity rules that enable such learning are unknown. We study bi

Self-Supervised On-Policy Distillation for Reasoning Language Models

Model ReleasesDGX agent

arXiv:2605.17497v1 Announce Type: new Abstract: GRPO-style RLVR trains reasoning models from multiple on-policy attempts per prompt, but typically uses these attempts only through terminal rewards. We

Sequential Structure in Intraday Futures Data: LSTM vs Gradient Boosting on MNQ

ApplicationsDGX agent

arXiv:2605.17724v1 Announce Type: cross Abstract: This paper compares gradient boosting and long short-term memory (LSTM) architectures for intraday directional prediction in Micro E-Mini Nasdaq 100 f

Shallow ReLU^s Networks in L^p-Type and Sobolev Spaces: Approximation and Path-Norm Controlled Generalization

Model ReleasesDGX agent

arXiv:2605.18468v1 Announce Type: cross Abstract: We study approximation by shallow ReLU^s networks, sigma_s(t)=max{0,t}^s, and the generalization behavior of such networks under ell_1 path-norm contr

Shift Detection and Adaptation for Network Intrusion Detection

ResearchDGX agent

arXiv:2508.15100v2 Announce Type: replace-cross Abstract: Distribution shift, a change in the statistical properties of data over time, poses a critical challenge for deep learning anomaly detection s

SignMuon: Communication-Efficient Distributed Muon Optimization

HardwareDGX agent

arXiv:2605.16311v1 Announce Type: new Abstract: Distributed training of large neural networks is bottlenecked by full-precision gradient communication and by coordinatewise optimizers that ignore the

SNLP: Layer-Parallel Inference via Structured Newton Corrections

SafetyDGX agent

arXiv:2605.17842v1 Announce Type: new Abstract: Autoregressive language models execute Transformer layers sequentially, creating a latency bottleneck that is not removed by conventional tensor or pipe

Sparse Deep Additive Model with Interactions: Enhancing Interpretability and Predictability

TutorialsDGX agent

arXiv:2509.23068v2 Announce Type: replace-cross Abstract: Recent advances in deep learning highlight the need for personalized models that can learn from small samples, handle high-dimensional feature

Sparse Mamba Decoder for Quantum Error Correction: Efficient Defect-Centric Processing of Surface Code Syndromes

HardwareDGX agent

arXiv:2605.17156v1 Announce Type: cross Abstract: Quantum error correction (QEC) is essential for building fault-tolerant quantum computers, requiring decoders that are simultaneously accurate, fast,

Sparse Training of Neural Networks based on Multilevel Mirror Descent

Model ReleasesDGX agent

arXiv:2602.03535v2 Announce Type: replace Abstract: We introduce a dynamic sparse training algorithm based on linearized Bregman iterations / mirror descent that exploits the naturally incurred sparsi

Spherical Harmonic Optimal Transport: Application to Climate Models Comparisons

HardwareDGX agent

arXiv:2605.18389v1 Announce Type: new Abstract: Optimal transport provides a powerful framework for comparing measures while respecting the geometry of their support, but comes with an expensive compu

ST-BCP: Tightening Coverage Bound for Backward Conformal Prediction via Non-Conformity Score Transformation

ResearchDGX agent

arXiv:2602.01733v2 Announce Type: replace-cross Abstract: Conformal Prediction (CP) provides a statistical framework for uncertainty quantification that constructs prediction sets with coverage guaran

StAD: Stein Amortized Divergence for Fast Likelihoods with Diffusion and Flow

TutorialsDGX agent

arXiv:2605.16486v1 Announce Type: cross Abstract: Diffusion and flow-based models are ubiquitously used for generative modelling and density estimation. They admit a deterministic probability flow ord

Statistical Unlearning of Distributions: A Hypothesis Testing Approach

ResearchDGX agent

arXiv:2605.16645v1 Announce Type: cross Abstract: Machine learning systems increasingly face requirements to forget not only individual data points, but entire domains of information, such as toxic la

StatQAT: Statistical Quantizer Optimization for Deep Networks

ResearchDGX agent

arXiv:2605.17745v1 Announce Type: cross Abstract: Quantization is essential for reducing the computational cost and memory usage of deep neural networks, enabling efficient inference on low-precision

Stein Diffusion Guidance: Training-Free Posterior Correction for Sampling Beyond High-Density Regions

ResearchDGX agent

arXiv:2507.05482v3 Announce Type: replace Abstract: Training-free diffusion guidance offers a flexible framework for leveraging off-the-shelf classifiers without additional training. Yet, current appr

Step-wise Rubric Rewards for LLM Reasoning

ResearchDGX agent

arXiv:2605.17291v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is widely used to improve reasoning in large language models, but rewards only final-answer correc

Stochastic Regret Guarantees for Online Zeroth- and First-Order Bilevel Optimization

ResearchDGX agent

arXiv:2511.01126v2 Announce Type: replace Abstract: Online bilevel optimization (OBO) is a powerful framework for machine learning problems where both outer and inner objectives evolve over time, requ

Stress-Testing Neural Network Verifiers with Provably Robust Instances

ResearchDGX agent

arXiv:2605.17153v1 Announce Type: new Abstract: Neural network verifiers aim to provide formal guarantees on model behavior, but existing verification benchmarks are fundamentally limited by their lac

Structure-Aware Masking for Protein Representation Learning

SafetyDGX agent

arXiv:2605.16581v1 Announce Type: new Abstract: Masked language modeling (MLM) is the standard objective for training protein language models, typically implemented by randomly masking individual resi

Structured Neural Marked Point Processes for Interpretable Event Interaction Modeling

Model ReleasesDGX agent

arXiv:2605.17568v1 Announce Type: new Abstract: Multi-class event streams arise in numerous real-world applications, where uncovering structured, interpretable inter-event relationships, together with

Subject-Specific Analysis of Self-Initiated Attention Shifts from EEG with Controlled Internal and External Attention Conditions

ResearchDGX agent

arXiv:2605.18251v1 Announce Type: cross Abstract: Self-initiated attention shifts play a critical role in voluntary behavior but are difficult to study due to the absence of explicit temporal markers.

SURGE: Approximation-free Training Free Particle Filter for Diffusion Surrogate

SafetyDGX agent

arXiv:2605.18745v1 Announce Type: cross Abstract: Diffusion-based generative models increasingly rely on inference-time guidance, adding a drift term or reweighting mixture of experts, to improve samp

SWING: Unlocking Implicit Graph Representations for Graph Random Features

ResearchDGX agent

arXiv:2602.12703v2 Announce Type: replace Abstract: We propose SWING: Space Walks for Implicit Network Graphs, a new class of algorithms for computations involving Graph Random Features on graphs give

Synthesis and Verification of Transformer Programs (Technical Report)

ResearchDGX agent

arXiv:2602.16473v2 Announce Type: replace Abstract: C-RASP is a simple programming language that was recently shown to capture concepts expressible by transformers. In this paper, we develop new algor

← Previous
1…149150151152153…243
Next →