AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
5 May 2026

Quantifying Multimodal Capabilities: Formal Generalization Guarantees in Pairwise Metric Learning

ApplicationsDGX agent

arXiv:2605.01424v1 Announce Type: new Abstract: Multimodal learning leverages the integration of diverse data modalities to enhance performance in complex tasks. Yet, it frequently encounters incomple

Quantum-inspired Techniques in Tensor Networks for Industrial Contexts

ResearchDGX agent

arXiv:2404.11277v2 Announce Type: replace-cross Abstract: In this paper we present a study of the applicability and feasibility of quantum-inspired algorithms and techniques in tensor networks for ind

RamanBench: A Large-Scale Benchmark for Machine Learning on Raman Spectroscopy

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.02003v1 Announce Type: new Abstract: Machine Learning (ML) has transformed many scientific fields, yet key applications still lack standardized benchmarks. Raman spectroscopy, a widely used

Random-Effects Algorithm for Random Objects in Metric Spaces

ResearchDGX agent

arXiv:2605.02693v1 Announce Type: cross Abstract: Across many scientific disciplines, multiple observations are collected from the same experimental units, and in modern datasets these observations of

RAST-MoE-RL: A Regime-Aware Spatio-Temporal MoE Framework for Deep Reinforcement Learning in Ride-Hailing

ApplicationsDGX agent

arXiv:2512.13727v2 Announce Type: replace Abstract: Ride-hailing platforms face the challenge of balancing passenger waiting times with overall system efficiency under highly uncertain supply-demand c

Rate-optimal Design for Anytime Best Arm Identification

ApplicationsDGX agent

arXiv:2510.23199v3 Announce Type: replace-cross Abstract: We consider the best arm identification problem, where the goal is to identify the arm with the highest mean reward from a set of K arms under

Rationality Measurement and Theory for Reinforcement Learning Agents

SafetyDGX agent

arXiv:2602.04737v2 Announce Type: replace Abstract: This paper proposes a suite of rationality measures and associated theory for reinforcement learning agents, a property increasingly critical yet ra

Real-Time Text Transmission via LLM-Based Entropy Coding over Fixed-Rate Channels

Model ReleasesDGX agent

arXiv:2605.01991v1 Announce Type: cross Abstract: Learning, prediction, and compression are intimately connected: a model that accurately predicts the next symbol in a sequence can be coupled with a s

Reconstructing conformal field theoretical compositions with Transformers

ResearchDGX agent

arXiv:2605.01072v1 Announce Type: cross Abstract: We study the use of transformers to reconstruct the compositions of tensor products of two-dimensional rational conformal field theories (RCFTs) based

Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability

Model ReleasesDGX agent

arXiv:2605.02552v1 Announce Type: new Abstract: Chemotherapy dose optimization can be formulated as a dynamic treatment regime, requiring sequential decisions under uncertainty that must balance tumor

Recurrent Graph Neural Networks and Arithmetic Circuits

ResearchDGX agent

arXiv:2603.05140v2 Announce Type: replace-cross Abstract: We characterise the computational power of recurrent graph neural networks (GNNs) in terms of arithmetic circuits over the real numbers. Our n

Reference-Sampled Boltzmann Projection for KL-Regularized RLVR: Target-Matched Weighted SFT, Finite One-Shot Gaps, and Policy Mirror Descent

Model ReleasesDGX agent

arXiv:2605.02469v1 Announce Type: new Abstract: Online reinforcement learning with verifiable rewards (RLVR) turns checkable outcomes into a scalable training signal, but it keeps rollout generation,

Remote Action Generation: Remote Control with Minimal Communication

SafetyDGX agent

arXiv:2605.01833v1 Announce Type: cross Abstract: We address the challenge of remote control where one or more actors, lacking direct reward access, are steered by a controller over a communication-co

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

Model ReleasesDGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

Retrieval with Multiple Query Vectors through Anomalous Pattern Detection

ResearchDGX agent

arXiv:2605.01965v1 Announce Type: new Abstract: A classical vector retrieval problem typically considers a single query embedding vector as input and retrieves the most similar embedding vectors from

Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI

ResearchDGX agent

arXiv:2605.01240v1 Announce Type: new Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of region-aware masking and hybrid sequence modeling remains under

Riemannian Generative Decoder

ResearchDGX agent

arXiv:2506.19133v3 Announce Type: replace Abstract: Euclidean representations distort data with intrinsic non-Euclidean structure. While Riemannian representation learning offers a solution by embeddi

Robust and Explainable Divide-and-Conquer Learning for Intrusion Detection

Local AiDGX agent

arXiv:2605.02015v1 Announce Type: new Abstract: Machine learning-based intrusion detection requires complex models to capture patterns in high-dimensional, noisy, and class-imbalanced raw network traf

Robust and Fast Training via Per-Sample Clipping

ResearchDGX agent

arXiv:2605.02701v1 Announce Type: cross Abstract: We propose a robust gradient estimator based on per-sample gradient clipping and analyze its properties both theoretically and empirically. We show th

Robust Conditional Conformal Prediction via Branched Normalizing Flow

ResearchDGX agent

arXiv:2605.01868v1 Announce Type: new Abstract: Conformal prediction (CP) constructs prediction sets with marginal coverage guarantees under the assumption that the calibration and test distributions

Robust Linear Dueling Bandits with Post-serving Context under Unknown Delays and Adversarial Corruptions

ResearchDGX agent

arXiv:2605.01752v1 Announce Type: new Abstract: We study linear dueling bandits in volatile environments characterized by the simultaneous presence of post-serving contexts, delayed feedback, and adve

Robust Parameter Learning for Uncertain MDPs

Model ReleasesDGX agent

arXiv:2605.01339v1 Announce Type: new Abstract: Learning-based approaches to verifying unknown Markov decision processes (MDPs) often employ uncertain MDPs. These models use, for example, confidence i

Robust volatility updates for Hierarchical Gaussian Filtering

Model ReleasesDGX agent

arXiv:2605.00966v1 Announce Type: new Abstract: Hierarchical Gaussian Filtering (HGF) networks allow for efficient updating of posterior distributions (beliefs) about hidden states of an agent's envir

Rule Extraction in Machine Learning: Chat Incremental Pattern Constructor

ResearchDGX agent

arXiv:2208.00335v4 Announce Type: replace Abstract: Rule extraction is a central problem in interpretable machine learning because it seeks to convert opaque predictive behavior into human-readable sy

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

AgentsDGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

SCALE-LoRA: Auditing Post-Retrieval LoRA Composition with Residual Merging and View Reliability

Model ReleasesDGX agent

arXiv:2605.01429v1 Announce Type: cross Abstract: Libraries of Low-Rank Adaptation (LoRA) adapters are becoming a practical by-product of parameter-efficient adaptation. Once such adapters accumulate,

Segment-Aligned Policy Optimization for Multi-Modal Reasoning

Model ReleasesDGX agent

arXiv:2605.01327v1 Announce Type: cross Abstract: Existing reinforcement learning approaches for Large Language Models typically perform policy optimization at the granularity of individual tokens or

Selective Prediction from Agreement: A Lipschitz-Consistent Version Space Approach

ResearchDGX agent

arXiv:2605.02611v1 Announce Type: new Abstract: We consider selective classification with abstention in the fixed-pool (or transductive) setting, where the unlabeled pool is given beforehand and only

Selector-Guided Autonomous Curriculum for One-Shot Reinforcement Learning from Verifiable Rewards

Model ReleasesDGX agent

arXiv:2605.01823v1 Announce Type: new Abstract: Recently, Reinforcement Learning from Verifiable Rewards (RLVR) has been established as a highly effective technique for augmenting the math reasoning s

Self-Normalized Martingales and Uniform Regret Bounds for Linear Regression

ResearchDGX agent

arXiv:2605.01628v1 Announce Type: cross Abstract: Self-normalized martingale inequalities lie at the heart of confidence ellipsoids for online least squares and, more broadly, many bandit and reinforc

Semi-Supervised Treatment Effect Estimation with Unlabeled Covariates for Prediction-Powered Causal Inference

ResearchDGX agent

arXiv:2511.08303v2 Announce Type: replace-cross Abstract: This study investigates treatment effect estimation in the semi-supervised setting, also can be interpreted as prediction-powered inference. I

Separation Assurance between Heterogeneous Fleets of Small Unmanned Aerial Systems via Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.01041v1 Announce Type: cross Abstract: In the envisioned future dense urban airspace, multiple companies will operate heterogeneous fleets of small unmanned aerial systems (sUASs), where ea

Sequential Learning and Catastrophic Forgetting in Differentiable Resistor Networks

ResearchDGX agent

arXiv:2605.01383v1 Announce Type: new Abstract: Differentiable physical networks provide a simple setting in which learning can be studied through the interaction between trainable parameters and phys

ShiftLIF: Efficient Multi-Level Spiking Neurons with Power-of-Two Quantization

ResearchDGX agent

arXiv:2605.01866v1 Announce Type: cross Abstract: Spiking neural networks (SNNs) are promising for edge sensing due to their event-driven computation and temporal filtering capability. However, standa

Short window attention enables long-term memorization

ResearchDGX agent

arXiv:2509.24552v3 Announce Type: replace Abstract: Recent works show that hybrid architectures combining local sliding window attention layers and global attention layers outperform either of these a

Singular Bayesian Neural Networks

SafetyDGX agent

arXiv:2602.00387v3 Announce Type: replace-cross Abstract: Bayesian neural networks promise calibrated uncertainty but require O(mn) parameters for standard mean-field Gaussian posteriors. We argue thi

Skipping the Zeros in Diffusion Models for Sparse Data Generation

ResearchDGX agent

arXiv:2605.01817v1 Announce Type: new Abstract: Diffusion models (DMs) excel on dense continuous data, but are not designed for sparse continuous data. They do not model exact zeros that represent the

SPAMoE: Spectrum-Aware Hybrid Operator Framework for Full-Waveform Inversion

ResearchDGX agent

arXiv:2604.07421v2 Announce Type: replace Abstract: Full-waveform inversion (FWI) is pivotal for reconstructing high-resolution subsurface velocity models but remains computationally intensive and ill

Sparse Regression under Correlation and Weak Signals: A Reproducible Benchmark of Classical and Bayesian Methods

Model ReleasesDGX agent

arXiv:2605.00835v1 Announce Type: new Abstract: Choosing between classical and Bayesian sparse regression methods involves a real trade-off: penalized estimators like Lasso run in milliseconds but giv

Spatial-Temporal Learning-Based Distributed Routing for Dynamic LEO Satellite Networks

Local AiDGX agent

arXiv:2605.02413v1 Announce Type: cross Abstract: In this paper, we propose a spatial-temporal learning-based distributed routing framework for dynamic Low Earth Orbit (LEO) satellite networks, where

SpecTM: Spectral Targeted Masking for Trustworthy Foundation Models

TutorialsDGX agent

arXiv:2603.22097v2 Announce Type: replace-cross Abstract: Foundation models are now increasingly being developed for Earth observation (EO), yet they often rely on stochastic masking that do not expli

Spectral Graph Sparsification Preserves Representation Geometry in Graph Neural Networks

ResearchDGX agent

arXiv:2605.01136v1 Announce Type: new Abstract: Spectral graph sparsification is a classical tool for reducing graph complexity while preserving Laplacian quadratic forms. In graph neural networks (GN

Spectral Model eXplainer: a chemically-grounded explainability framework for spectral-based machine learning models

Model ReleasesDGX agent

arXiv:2605.02684v1 Announce Type: new Abstract: Spectral-based machine learning models have been increasingly deployed in chemometrics and spectroscopy, where predictive accuracy is as important as ex

Split-on-Share: Mixture of Sparse Experts for Task-Agnostic Continual Learning

Model ReleasesDGX agent

arXiv:2601.17616v2 Announce Type: replace Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to

SplitZip: Ultra Fast Lossless KV Compression for Disaggregated LLM Serving

HardwareDGX agent

arXiv:2605.01708v1 Announce Type: cross Abstract: Contemporary systems serving large language models (LLMs) have adopted prefill-decode disaggregation to better load-balance between the compute-bound

Stability and Generalization for Decentralized Markov SGD

ResearchDGX agent

arXiv:2605.01701v1 Announce Type: new Abstract: Stochastic gradient methods are central to large-scale learning, yet their generalization theory typically relies on independent sampling assumptions. I

Stabilizing Private LASSO under Heterogeneous Covariates via Anisotropic Objective Perturbation

ResearchDGX agent

arXiv:2605.01492v1 Announce Type: cross Abstract: We study high-dimensional LASSO under differential privacy via objective perturbation with heterogeneous covariate scales. In practical scenarios, cov

Stable Blanket with Hidden Variables and Cycles

ResearchDGX agent

arXiv:2605.01856v1 Announce Type: cross Abstract: Stabilized regression aims to identify a set of predictors whose conditional relationship with a response variable remains invariant across different

Stable GFlowNets with Probabilistic Guarantees

TutorialsDGX agent

arXiv:2605.01729v1 Announce Type: new Abstract: Generative Flow Networks (GFlowNets) learn to sample states proportional to an unnormalized reward. Despite their theoretical promise, practical trainin

Stable Localized Conformal Prediction via Transduction

ResearchDGX agent

arXiv:2605.01452v1 Announce Type: cross Abstract: Existing evaluations of conformal prediction, such as prediction efficiency and test-conditional coverage, are defined in expectation over the calibra

STABLEVAL: Disagreement-Aware and Stable Evaluation of AI Systems

SafetyDGX agent

arXiv:2605.02122v1 Announce Type: new Abstract: Human evaluation remains the primary standard for assessing modern AI systems, yet annotator disagreement, bias, and variability make system rankings fr

Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

Model ReleasesDGX agent

arXiv:2605.02860v1 Announce Type: cross Abstract: Cross-language code clone detection (X-CCD) is challenging because semantically equivalent programs written in different languages often share little

STAR: Decode-Phase Rescheduling for LLM Inference

ResearchDGX agent

arXiv:2510.13668v2 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference has emerged as a fundamental paradigm, however, variations in output length cause severe workload imbalan

Statistical Consistency and Generalization of Contrastive Representation Learning

ResearchDGX agent

arXiv:2605.02116v1 Announce Type: new Abstract: Contrastive representation learning (CRL) underpins many modern foundation models. Despite recent theoretical progress, existing analyses suffer from se

Statistically-Lossless Quantization of Large Language Models

Model ReleasesDGX agent

arXiv:2605.02404v1 Announce Type: new Abstract: Model quantization has become essential for efficient large language model deployment, yet existing approaches involve clear trade-offs: methods such as

Stochastic Modeling of Human-Machine Authentication Channels under Partial Information Leakage

ApplicationsDGX agent

arXiv:2605.02102v1 Announce Type: cross Abstract: Reliable and secure human-machine communication is fundamental to IoT and cyber-physical ecosystems, where smartphones and wearables commonly serve as

Stochastic Sparse Attention for Memory-Bound Inference

HardwareDGX agent

arXiv:2605.01910v1 Announce Type: new Abstract: Autoregressive decoding becomes bandwidth-limited at long contexts, as generating each token requires reading all n_k key and value vectors from KV cach

StreamIndex: Memory-Bounded Compressed Sparse Attention via Streaming Top-k

Model ReleasesDGX agent

arXiv:2605.02568v1 Announce Type: new Abstract: DeepSeek-V3.2 and V4 introduce Compressed Sparse Attention (CSA): a lightning indexer (a learned scoring projection over compressed keys) scores them, t

Structured Analytic Coherent Point Drift for Non-Rigid Point Set Registration

ResearchDGX agent

arXiv:2605.00934v1 Announce Type: new Abstract: We introduce Analytic-CPD, a structured analytic variant of coherent point drift for non-rigid point set registration. The method retains the CPD poster

StyleShield: Exposing the Fragility of AIGC Detectors through Continuous Controllable Style Transfer

Model ReleasesDGX agent

arXiv:2605.00924v1 Announce Type: new Abstract: AI-generated content (AIGC) detectors are increasingly deployed in high-stakes settings such as academic integrity screening, yet their reliability rest

← Previous
1…191192193194195…241
Next →