AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
2 Jun 2026

Quartet II: Accurate LLM Pre-Training in NVFP4 by Improved Unbiased Gradient Estimation

HardwareDGX agent

arXiv:2601.22813v2 Announce Type: replace Abstract: The NVFP4 lower-precision format, supported in hardware by NVIDIA Blackwell GPUs, promises to allow, for the first time, end-to-end fully-quantized

Query-Limited Community Recovery in Stochastic Block Models

Model ReleasesDGX agent

arXiv:2606.02055v1 Announce Type: cross Abstract: We study exact community recovery in the two-community stochastic block model on n vertices under limited and noisy access to network data. The learne

RADE: Random Add-Drop Edge as a Regularizer

SafetyDGX agent

arXiv:2606.00757v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) suffer from overfitting and over-squashing of long-range information. Stochastic graph augmentations (e.g., edge deletion)


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Randomized Least Squares Value Iteration itself is Joint Differentially Private

ResearchDGX agent

arXiv:2606.01952v1 Announce Type: new Abstract: As reinforcement learning (RL) increasingly applies to sensitive domains, such as health care and recommendation systems, privacy-preserving techniques

RDA: Reward Design Agent for Reinforcement Learning

AgentsDGX agent

arXiv:2606.01672v1 Announce Type: new Abstract: Reinforcement learning has enabled the acquisition of impressive robotic skills, but typically requires hand-crafted reward functions that are slow to d

React to Surprises: Stable-by-Design Neural Feedback Control and the Youla-REN

Model ReleasesDGX agent

arXiv:2506.01226v3 Announce Type: replace-cross Abstract: We study parameterizations of stabilizing nonlinear policies for learning-based control. We propose a structure based on a nonlinear version o

Real-Time Sensing of Inaccessible Physical Fields via an Edge-Deployable Hardware-Portable Graph Neural Operator

SafetyDGX agent

arXiv:2604.01802v2 Announce Type: replace Abstract: Real-time inference of inaccessible interior physical fields from sparse boundary observations is a fundamental but unresolved problem in scientific

Realistic noise synthesis reduces bias and improves tissue microstructure estimation with supervised machine learning

Model ReleasesDGX agent

arXiv:2606.02044v1 Announce Type: new Abstract: Diffusion MRI enables non-invasive probing of tissue microstructure, but accurate parameter estimation is challenged by noise-related effects. In superv

Reconstructing Content via Collaborative Attention to Improve Multimodal Embedding Quality

ResearchDGX agent

arXiv:2603.01471v2 Announce Type: replace-cross Abstract: Multimodal embedding models, rooted in multimodal large language models (MLLMs), have yielded significant performance improvements across dive

ReFLEX: Length-Generalizable CSI Denoising for MIMO-OFDM via Relative-Frequency Bias

SafetyDGX agent

arXiv:2606.00263v1 Announce Type: cross Abstract: This letter studies CSI denoising for MIMO--OFDM with variable NR resource block (RB) allocations. ReFLEX is a length-generalizable Transformer whose

RefLoRA: Refactored Low-Rank Adaptation for Efficient Fine-Tuning of Large Models

Model ReleasesDGX agent

arXiv:2505.18877v4 Announce Type: replace Abstract: Low-Rank Adaptation (LoRA) lowers the computational and memory overhead of fine-tuning large models by updating a low-dimensional subspace of the pr

Regularized Large Neighborhood Search

ResearchDGX agent

arXiv:2606.02294v1 Announce Type: new Abstract: Operations research practitioners typically tackle NP-hard combinatorial problems using large neighborhood search (LNS), a scalable heuristic that itera

Reinforcement Learning for Optimal Experiment Design in Parameter Identification of Mechatronic Systems

Model ReleasesDGX agent

arXiv:2606.00059v1 Announce Type: cross Abstract: Informative excitation signals are critical for accurate system identification of mechatronic systems, yet classical system identification (SI) approa

Rethinking Bregman Divergences in Kronecker-Factored Optimizers

ResearchDGX agent

arXiv:2606.00542v1 Announce Type: new Abstract: Shampoo-style optimizers approximate gradient covariance matrices using Kronecker-factored structures. Recent work~ite{lin2026understanding} showed that

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete

Model ReleasesDGX agent

arXiv:2606.01532v1 Announce Type: new Abstract: Positional encoding (PE) is widely viewed as necessary for transformers to process ordered sequences: without them, the next-token map appears permutati

Revisiting Neural Processes via Fourier Transform and Volterra Series

ApplicationsDGX agent

arXiv:2606.01172v1 Announce Type: new Abstract: Modeling unknown latent functions from finite, irregularly sampled measurements is a recurring challenge across science and engineering. Neural processe

Riemannian Gradient Descent for Low-Rank Architectures

ResearchDGX agent

arXiv:2606.02328v1 Announce Type: new Abstract: We explore Riemannian optimization techniques for rank-factored matrix parameters, targeting contemporary deep learning applications. We examine ten poi

Riemannian Optimization for Hadamard Products of Low-Rank Matrices

Model ReleasesDGX agent

arXiv:2606.01216v1 Announce Type: new Abstract: The elementwise Hadamard product of two low-rank matrices provides a parameter-efficient model for data with multiplicative structure, but its modeling

Riemannian Stochastic Optimization for Sufficient Dimension Reduction

ResearchDGX agent

arXiv:2606.00413v1 Announce Type: cross Abstract: Sufficient dimension reduction (SDR) makes high-dimensional regression tractable by projecting the covariates onto a low-dimensional subspace that pre

Robust Learning of a Group DRO Neuron

Model ReleasesDGX agent

arXiv:2601.18115v2 Announce Type: replace Abstract: We study the problem of learning a single neuron under standard squared loss in the presence of arbitrary label noise and group-level distributional

Robust Predictive Uncertainty and Double Descent in Contaminated Bayesian Random Features

ResearchDGX agent

arXiv:2602.19126v2 Announce Type: replace Abstract: We propose a robust Bayesian formulation of random feature (RF) regression that accounts explicitly for prior and likelihood misspecification via Hu

RobustModelMaker: Coupling Bootstrap Stability Selection with Leakage-Safe Nested Cross-Validation for Scientific Machine Learning

ResearchDGX agent

arXiv:2606.01566v1 Announce Type: new Abstract: Small-to-medium scientific datasets place machine learning pipelines under two compounding pressures. Single-run feature selection produces feature sets

Safe-Subspace Pseudo-Label Refinement for Source-Free Graph Domain Adaptation

ApplicationsDGX agent

arXiv:2606.00808v1 Announce Type: new Abstract: Source-free graph domain adaptation (SF-GDA) aims to adapt source-trained graph models to unlabeled target graphs when source graphs are no longer acces

Safeguarded Stochastic Polyak Step Sizes for Non-smooth Optimization: Robust Performance Without Small (Sub)Gradients

ResearchDGX agent

arXiv:2512.02342v3 Announce Type: replace-cross Abstract: The stochastic Polyak step size (SPS) has proven to be a promising choice for stochastic gradient descent (SGD), delivering competitive perfor

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

SafetyDGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

Sample Complexity and Decision-Theoretic Guarantees for Bayesian Model Averaging over Decision Trees with Catalan-Exponential Priors

ResearchDGX agent

arXiv:2606.01340v1 Announce Type: new Abstract: We ask: when do Bayesian model averaging (BMA) weights over decision trees carry sufficient epistemic information to justify committed exploitation of t

Scalable Counterfactual Risk Estimation for Rare Events in Longitudinal Data

ResearchDGX agent

arXiv:2606.01539v1 Announce Type: cross Abstract: Estimating the causal effect of time-varying treatments on survival outcomes in large observational studies is computationally demanding, particularly

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

SafetyDGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

Scaling depth capacity via zero/one-layer model expansion

ResearchDGX agent

arXiv:2511.04981v2 Announce Type: replace Abstract: Model depth is a double-edged sword in deep learning: deeper models achieve higher accuracy but require higher computational cost. To efficiently tr

Score imes Decoder: A Unified View of Unsupervised Inference-Time Scaling for Hallucination Mitigation

ResearchDGX agent

arXiv:2606.00739v1 Announce Type: new Abstract: Large language models hallucinate even when the answer lies within their parameters. While inference-time scaling can surface this latent knowledge, the

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

SafetyDGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

Segment-driven Structural Induction and Semantic Alignment for Heterogeneous Tabular Representation

Local AiDGX agent

arXiv:2606.01890v1 Announce Type: new Abstract: Real-world domains often contain heterogeneous tables whose headers vary while their underlying attribute semantics are shared, making it difficult to i

Self-Regulating Annealing in Heavy-Tailed Diffusion Models

ResearchDGX agent

arXiv:2606.01645v1 Announce Type: cross Abstract: Diffusion models have emerged as a leading framework for deep generative modeling. While the standard Gaussian formulation is theoretically convenient

Semantic-Geometric Task Representations for Bimanual Manipulation from Human Demonstrations to Robot Action Planning

ResearchDGX agent

arXiv:2601.11460v2 Announce Type: replace-cross Abstract: Learning structured task representations from human demonstrations is essential for bimanual manipulation, where action ordering, object invol

Semantic Retrieval for Product Search in E-Commerce

SafetyDGX agent

arXiv:2606.01504v1 Announce Type: cross Abstract: Semantic retrieval in e-commerce must handle short, noisy, and colloquial queries over large product catalogs with fine-grained attribute distinctions

Semi-Supervised Hyperbolic Hierarchical Clustering with Set-Level Structural Priors

Model ReleasesDGX agent

arXiv:2606.01525v1 Announce Type: new Abstract: Semi-supervised hierarchical clustering aims to learn a tree structure consistent with data patterns and user-provided supervision. Supervision is usual

Semi-Supervised Learning with Noisy Proxy Covariates: Generalization Bounds and Distribution Regression

ResearchDGX agent

arXiv:2606.00512v1 Announce Type: new Abstract: In many modern machine learning pipelines, abundant pretrained representations serve as noisy proxy covariates, while task-specific labels remain scarce

Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise Domain

ResearchDGX agent

arXiv:2606.00558v1 Announce Type: new Abstract: Transfer learning aims to facilitate the learning of a target domain by transferring knowledge from a source domain. The source domain typically contain

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

ShaplEIG: Bayesian Experimental Design for Shapley Value Estimation

ResearchDGX agent

arXiv:2606.02247v1 Announce Type: cross Abstract: Shapley values are a principled attribution measure widely used in interpretable machine learning, but their exact computation scales exponentially wi

Sharpness-Aware Hybrid Model Learning for Architecture-Agnostic Parameter Estimation

Model ReleasesDGX agent

arXiv:2602.06837v2 Announce Type: replace Abstract: Hybrid modeling, the combination of machine learning models and scientific mathematical models, enables flexible and robust data-driven prediction w

Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning

Model ReleasesDGX agent

arXiv:2603.11653v2 Announce Type: replace Abstract: Continual Reinforcement Learning (CRL) for Vision-Language-Action (VLA) models is a promising direction toward self-improving embodied agents that c

Site4Drug: Predicting Drug-Binding Target Sites with an AI Agent

AgentsDGX agent

arXiv:2606.01816v1 Announce Type: cross Abstract: Selecting where to intervene on a protein (i.e., choosing a targetable site) is often a more ambiguous and failure-prone bottleneck than selecting wha

Sparse FEONet: A Low-Cost, Memory-Efficient Operator Network via Finite-Element Local Sparsity for Parametric PDEs

Model ReleasesDGX agent

arXiv:2601.00672v2 Announce Type: replace-cross Abstract: In this paper, we study the finite element operator network (FEONet), an operator-learning method for parametric problems, originally introduc

Spatially Distributed Task-Oriented Compression for Multi-Emitter Localization and Characterization with Spectral Overlap

Model ReleasesDGX agent

arXiv:2606.01446v1 Announce Type: cross Abstract: Radio frequency spectrum awareness requires the ability to detect, localize, and characterize emitters in dense and contested wireless environments. I

Spatiotemporal Multi-Task Graph Transformer for Trip-Level Transit Prediction

SafetyDGX agent

arXiv:2606.00572v1 Announce Type: new Abstract: Passenger count data from public transit systems reveals urban mobility patterns and is essential for planning, operation, and optimisation. However, no

Spectra-Guided Neural Tucker Factorization

Model ReleasesDGX agent

arXiv:2606.00584v1 Announce Type: cross Abstract: This paper proposes Spectra-Guided Neural Tucker Factorization (SG-NTF) for High-Dimensional and Incomplete (HDI) tensor completion. Circumventing dis

Spectral Audit of In-Context Operator Networks

Local AiDGX agent

arXiv:2606.02427v1 Announce Type: cross Abstract: Existing evaluations of neural operators and in-context operator learning rely primarily on prediction error, but accurate output prediction does not

Speculative Sampling For Faster Molecular Dynamics

ResearchDGX agent

arXiv:2606.02455v1 Announce Type: new Abstract: Molecular dynamics (MD) is a key tool for simulating the dynamical behavior of atomic systems. However, MD is inherently serial, which makes it difficul

Statistical Analysis of using the Shapley Value for Sensor Anomaly Localization with Accurate Classifiers

ResearchDGX agent

arXiv:2606.00867v1 Announce Type: cross Abstract: Recent publications have suggested using the Shap- ley value for sensor anomaly/attack localization. We study the performance of such an approach by u

Statistical Guarantees for Reasoning Probes on Looped Boolean Circuits

ResearchDGX agent

arXiv:2602.03970v3 Announce Type: replace-cross Abstract: We study the statistical behavior of reasoning probes in a stylized model of iterative computation inspired by neural algorithmic reasoning. T

Statistical Testing on Directed Graphs by Surrogate Data Generation

ResearchDGX agent

arXiv:2606.00758v1 Announce Type: cross Abstract: In recent years, graph signal processing has emerged as a powerful framework at the intersection of signal processing and graph theory, providing tool

Step-Level Sparse Autoencoder for Reasoning Process Interpretation

ResearchDGX agent

arXiv:2603.03031v2 Announce Type: replace Abstract: Large Language Models (LLMs) have achieved strong complex reasoning capabilities through Chain-of-Thought (CoT) reasoning. However, their reasoning

Stochastic Rounding Increases Small Singular Values

ResearchDGX agent

arXiv:2606.00312v1 Announce Type: cross Abstract: Over the past half-dozen years, stochastic rounding (SR) has regained significant attention as a quantization scheme for low-precision floating-point

Structure and Scale in Simplicial Sequence Modelling

ResearchDGX agent

arXiv:2606.01302v1 Announce Type: new Abstract: Modern large-scale deep learning exhibits two striking empirical phenomena: behavioural scaling laws (predictable performance gains with increasing scal

Symmetric Hermite quadrature-based balanced truncation for learning linear dynamical systems from derivative data

ResearchDGX agent

arXiv:2606.00298v1 Announce Type: cross Abstract: Data-driven reduced-order modeling is an essential component in the computer-aided design of control systems. In this work, we present a novel symmetr

Symmetries in PAC-Bayesian Learning

ApplicationsDGX agent

arXiv:2510.17303v2 Announce Type: replace Abstract: Symmetries are known to improve the empirical performance of machine learning models, yet theoretical guarantees explaining these gains remain limit

Synthesizing Neural Network Controllers with Closed-Loop Dissipativity Guarantees

ResearchDGX agent

arXiv:2404.07373v2 Announce Type: replace-cross Abstract: This paper presents a method to synthesize neural network controllers to maximize reward subject to the hard constraint that the feedback syst

Systematic Evaluation of Time-Frequency Features for Binaural Sound Source Localization

Local AiDGX agent

arXiv:2511.13487v3 Announce Type: replace-cross Abstract: This study presents a systematic evaluation of time-frequency feature design for binaural sound source localization (SSL), focusing on how fea

TabPrep: Closing the Feature Engineering Gap in Tabular Benchmarks

Model ReleasesDGX agent

arXiv:2606.02384v1 Announce Type: new Abstract: Progress in tabular machine learning has largely focused on increasingly sophisticated model architectures. At the same time, feature engineering remain

← Previous
1…108109110111112…243
Next →