AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

SAGE: Agentic Framework for Interpretable and Clinically Translatable Computational Pathology Biomarker Discovery

AgentsDGX agent

arXiv:2602.00953v2 Announce Type: replace Abstract: Engineered image-based biomarkers offer a clinically interpretable alternative to black-box AI in computational pathology, yet their discovery remai

Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift

SafetyDGX agent

arXiv:2605.10289v1 Announce Type: new Abstract: Offline-to-online learning aims to improve online decision-making by leveraging offline logged data. A central challenge in this setting is the distribu

Scalable Gaussian process inference via neural feature maps

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.10285v1 Announce Type: cross Abstract: We present a theoretically grounded Gaussian process framework that leverages neural feature maps to construct expressive kernels. We show that the le

Scalable Mamba-Based Message-Passing Neural Decoder for Error-Correcting Codes

ResearchDGX agent

arXiv:2605.10681v1 Announce Type: cross Abstract: Forward error correction is essential for reliable communication over noisy channels. Attention-based model-free neural decoders have shown strong per

Scaling the Memory of Balanced Adam

Model ReleasesDGX agent

arXiv:2605.10119v1 Announce Type: new Abstract: Recent evidence suggests that Adam performs robustly when its momentum parameters are tied, eta_1=eta_2, reducing the optimizer to a single remaining pa

SCOT: Multi-Source Cross-City Transfer with Optimal-Transport Soft-Correspondence Objective

SafetyDGX agent

arXiv:2604.07383v2 Announce Type: replace Abstract: Cross-city transfer improves prediction in label-scarce cities by leveraging labeled data from other cities, but it becomes challenging when cities

SeBA: Semi-supervised few-shot learning via Separated-at-Birth Alignment for tabular data

Model ReleasesDGX agent

arXiv:2605.08519v1 Announce Type: new Abstract: Learning from scarce labeled data with a larger pool of unlabeled samples, known as semi-supervised few-shot learning (SS-FSL), remains critical for app

Selection of the Best Policy under Fairness Constraints for Subpopulations

SafetyDGX agent

arXiv:2605.09945v1 Announce Type: new Abstract: Many high-stakes decisions in health care, public policy, and clinical development require committing to a single policy that will be applied uniformly

Selection Plateau and a Sparsity-Dependent Hierarchy of Pruning Features

SafetyDGX agent

arXiv:2605.09345v1 Announce Type: new Abstract: We identify a Selection Plateau phenomenon in one-shot neural network pruning: all rank-monotone weight scorers converge to identical accuracy at fixed

Self-Attention as a Covariance Readout: A Unified View of In-Context Learning and Repetition

ResearchDGX agent

arXiv:2605.10466v1 Announce Type: new Abstract: Large language models (LLMs) exhibit two striking and ostensibly unrelated behaviours: in-context learning (ICL) and repetitive generation. In both, the

SEMASIA: A Large-Scale Dataset of Semantically Structured Latent Representations

Model ReleasesDGX agent

arXiv:2605.09485v1 Announce Type: new Abstract: Latent representations learned by neural networks often exhibit semantic structure, where concept similarity is reflected by geometric proximity in embe

Sequential Causal Discovery with Noisy Language Model Priors

Model ReleasesDGX agent

arXiv:2506.16234v2 Announce Type: replace Abstract: Causal discovery from observational data typically assumes access to complete data and availability of perfect domain experts. In practice, data oft

Sequential Membership Inference Attacks

ResearchDGX agent

arXiv:2602.16596v2 Announce Type: replace Abstract: Modern AI models are not static. They go through multiple updates in their lifecycles. We propose to design Sequential Membership Inference (SeMI) a

Set Prediction for Next-Day Active Fire Forecasting

Model ReleasesDGX agent

arXiv:2605.10298v1 Announce Type: new Abstract: Accurate next-day active fire forecasts can support early warning, disaster response, forest risk assessment, and downstream estimation of fire-related

Sharp feature-learning transitions and Bayes-optimal neural scaling laws in extensive-width networks

ResearchDGX agent

arXiv:2605.10395v1 Announce Type: cross Abstract: We study the information-theoretic limits of learning a one-hidden-layer teacher network with hierarchical features from noisy queries, in the context

Signal from Structure: Exploiting Submodular Upper Bounds in Generative Flow Networks

TutorialsDGX agent

arXiv:2601.21061v2 Announce Type: replace Abstract: Generative Flow Networks (GFlowNets; GFNs) are a class of generative models that learn to sample compositional objects proportionally to their a pri

Signature Approach for Contextual Bandits with Nonlinear and Path-dependent Rewards

ResearchDGX agent

arXiv:2605.10313v1 Announce Type: new Abstract: We study contextual bandits with nonlinear and path-dependent rewards through a novel signature-transform-based approach. Leveraging the universal nonli

Single-Thread JPEG Decoder Benchmarks Mis-Evaluate ML Data Loaders

Model ReleasesDGX agent

arXiv:2605.08731v1 Announce Type: cross Abstract: JPEG decode is routine ML infrastructure, but Python decoder choices are often justified by single-process, single-thread microbenchmarks. We audit th

Sinkhorn Treatment Effects: A Causal Optimal Transport Measure

Model ReleasesDGX agent

arXiv:2605.08485v1 Announce Type: cross Abstract: We introduce the Sinkhorn treatment effect, an entropic optimal transport measure of divergence between counterfactual distributions. Unlike classical

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck

AgentsDGX agent

arXiv:2605.08526v1 Announce Type: new Abstract: While LLM-based agents excel at planning and executing long action sequences, their execution often remains inconsistent across trials, limiting reliabi

Sliced Inner Product Gromov-Wasserstein Distances

ResearchDGX agent

arXiv:2605.08546v1 Announce Type: cross Abstract: The Gromov-Wasserstein (GW) problem provides a framework for aligning heterogeneous datasets by matching their intrinsic geometry, but its statistical

Sliding Window Informative Canonical Correlation Analysis

ResearchDGX agent

arXiv:2507.17921v2 Announce Type: replace-cross Abstract: Canonical correlation analysis (CCA) is a technique for finding correlated sets of features between two datasets. In this paper, we propose a

SMIXAE: Towards Unsupervised Manifold Discovery in Language Models

Model ReleasesDGX agent

arXiv:2605.09224v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) have been used widely to decompose and interpret neural network activations, especially those of transformer language models.

SMOG: Scalable Meta-Learning for Multi-Objective Bayesian Optimization

ResearchDGX agent

arXiv:2601.22131v2 Announce Type: replace Abstract: Multi-objective optimization aims to solve problems with competing objectives. Evaluating such problems is often slow or expensive, limiting the bud

Social Determinants of Health and Fentanyl Overdose Mortality Across US Counties: An XGBoost and SHAP Analysis Identifying Silent Risk Counties and Treatment Deserts

ResearchDGX agent

arXiv:2605.08230v1 Announce Type: new Abstract: Background: Fentanyl overdose deaths are still increasing across the U.S. We do not fully understand which county-level social and structural conditions

SPDEBench: An Extensive Benchmark for Learning Stochastic PDEs

Model ReleasesDGX agent

arXiv:2505.18511v2 Announce Type: replace Abstract: Stochastic Partial Differential Equations (SPDEs) driven by random noise play a central role in modeling physical processes with rough spatio-tempor

Spectral Condition for muP under Width-Depth Scaling

ResearchDGX agent

arXiv:2603.00541v2 Announce Type: replace Abstract: Generative foundation models are increasingly scaled in both width and depth, posing significant challenges for stable feature learning and reliable

SpectraLLM: Uncovering the Ability of LLMs for Molecular Structure Elucidation from Multi-Spectral Data

Model ReleasesDGX agent

arXiv:2508.08441v3 Announce Type: replace-cross Abstract: Automated molecular structure elucidation remains challenging, as existing approaches often depend on pre-compiled databases or restrict thems

Spherical Boltzmann machines: a solvable theory of learning and generation in energy-based models

Model ReleasesDGX agent

arXiv:2605.09031v1 Announce Type: new Abstract: Energy-based models (EBMs) are flexible generative architectures inspired by statistical physics, but their learning and generative properties remain po

Split CNN Inference on Networked Microcontrollers

ResearchDGX agent

arXiv:2605.09357v1 Announce Type: cross Abstract: Running deep neural networks on microcontroller units (MCUs) is severely constrained by limited memory resources. While TinyML techniques reduce model

Stability of the Monge Map in Semi-Dual Optimal Transport

ResearchDGX agent

arXiv:2605.05569v2 Announce Type: replace-cross Abstract: This paper shows that the semi-dual formulation of the optimal transport problem has a degenerate saddle-point structure, and that its numeric

Stable Long-Horizon PDE Forecasting via Latent Structured Spectral Propagators

SafetyDGX agent

arXiv:2605.10154v1 Announce Type: new Abstract: Long-horizon forecasting of time-dependent partial differential equations (PDEs) is critical for characterizing the sustained evolution of physical syst

Statistical Inference and Quality Measures of KV Cache Quantisations Inspired by TurboQuant

ResearchDGX agent

arXiv:2605.08114v1 Announce Type: new Abstract: We analyse three KV cache quantization schemes under a fair bit budget: extbf{KV} (scalar MSE baseline), extbf{KQV} (WHT + MSE on K; WHT + MSE + QJL on

Steerable but Not Decodable: Function Vectors Operate Beyond the Logit Lens

Model ReleasesDGX agent

arXiv:2604.02608v2 Announce Type: replace Abstract: Activation steering presupposes that task-relevant behaviors correspond to linear directions in activation space -- directions that should both stee

Stellar Age Compression Reshapes Interpretations of the Milky Way Thick-Disk Formation History

ResearchDGX agent

arXiv:2605.10220v1 Announce Type: cross Abstract: The formation timescale of the Milky Way thick disk is one of the central debates in Galactic archaeology. The age-metallicity relation (AMR), formati

Streaming Sliced Optimal Transport

ResearchDGX agent

arXiv:2505.06835v4 Announce Type: replace Abstract: Sliced optimal transport (SOT), or sliced Wasserstein (SW) distance, is widely recognized for its statistical and computational scalability. In this

Structural Alignment Improves Graph Test-Time Adaptation

SafetyDGX agent

arXiv:2502.18334v5 Announce Type: replace Abstract: Graph-based learning excels at capturing interaction patterns in diverse domains like recommendation, fraud detection, and particle physics. However

Structure-Preserving Reconstruction of Convex Lipschitz Functionals on Hilbert Spaces from Finite Samples

Model ReleasesDGX agent

arXiv:2605.08559v1 Announce Type: cross Abstract: Convex functionals are ubiquitous in applied analysis, appearing as value functions, risk measures, super-hedging prices, and loss functionals in mach

Sub-Footprint Effect Correction in FW-LiDAR Point Clouds via Intra-Footprint Target Unmixing

ApplicationsDGX agent

arXiv:2605.09845v1 Announce Type: new Abstract: Sub-footprint target mixing within a laser footprint significantly increases LiDAR intensity uncertainty, especially in complex environments where heter

Sundial: A Family of Highly Capable Time Series Foundation Models

ApplicationsDGX agent

arXiv:2502.00816v4 Announce Type: replace Abstract: We introduce Sundial, a family of native, flexible, and scalable time series foundation models. To predict the next-patch's distribution, we propose

Supercharging Bayesian Inference with Reliable AI-Informed Priors

SafetyDGX agent

arXiv:2605.09834v1 Announce Type: cross Abstract: Modern predictive systems encode beliefs that can act as useful prior information for statistical inference in data-limited settings. Using them for p

Supervised Guidance Training for Infinite-Dimensional Diffusion Models

ResearchDGX agent

arXiv:2601.20756v2 Announce Type: replace Abstract: Score-based diffusion models have recently been extended to infinite-dimensional function spaces, with uses such as inverse problems arising from pa

Survey-aware Machine Learning: A Guideline for Valid Population Health Inference based on Scoping Review

SafetyDGX agent

arXiv:2605.08963v1 Announce Type: cross Abstract: Machine Learning (ML) models trained on complex health surveys such as the National Health and Nutrition Examination Survey (NHANES) often ignore prim

SWE Atlas: Benchmarking Coding Agents Beyond Issue Resolution

Model ReleasesDGX agent

arXiv:2605.08366v1 Announce Type: new Abstract: We introduce SWE Atlas, a benchmark suite for coding agents spanning three professional software engineering workflows: Codebase Q&A (124 tasks), Test W

SymTorch: Symbolic Distillation of Neural Networks

Local AiDGX agent

arXiv:2602.21307v2 Announce Type: replace Abstract: What mathematical functions do neural network components learn? Symbolic distillation addresses this question by expressing neural network component

Synergistic Simplex: Cooperative Runtime Assurance for Safety-Critical Autonomous Systems

SafetyDGX agent

arXiv:2605.08190v1 Announce Type: new Abstract: Autonomous systems increasingly rely on machine-learning (ML) components for safety-critical tasks such as perception and control in autonomous vehicles

Tabular Foundation Model for Generative Modelling

Model ReleasesDGX agent

arXiv:2605.09424v1 Announce Type: new Abstract: Generative modelling is a demanding test of foundation models, because it requires robust, holistic representation learning for a given data modality, r

TAH-QUANT: Effective Activation Quantization in Pipeline Parallelism over Slow Network

ResearchDGX agent

arXiv:2506.01352v2 Announce Type: replace Abstract: Decentralized training of large language models offers the opportunity to pool computational resources across geographically distributed participant

Targeted Synthetic Control Method

SafetyDGX agent

arXiv:2602.04611v2 Announce Type: replace-cross Abstract: The synthetic control method (SCM) estimates causal effects in panel data with a single-treated unit by constructing a counterfactual outcome

TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder Generation

SafetyDGX agent

arXiv:2605.09810v1 Announce Type: cross Abstract: Protein function is often controlled by ligands that bias the direction of state transitions, such as agonists and antagonists, rather than stabilizin

Teaching LLMs to See Graphs: Unifying Text and Structural Reasoning

Model ReleasesDGX agent

arXiv:2605.10247v1 Announce Type: new Abstract: Using Large Language Models (LLMs) to process graph-structured data is an active research area, yet current state-of-the-art approaches typically rely o

TeleResilienceBench: Quantifying Resilience for LLM Reasoning in Telecommunications

Model ReleasesDGX agent

arXiv:2605.09929v1 Announce Type: new Abstract: Deploying large language models in telecommunications requires more than task accuracy. In realistic workflows, a model may inherit partially completed

Temporal-Decay Shapley: A Time-Aware Data Valuation Framework for Time-Series Data

ResearchDGX agent

arXiv:2605.08153v1 Announce Type: new Abstract: With the rapid development of machine learning applications on time-series data, accurately assessing the value of training samples has become essential

Tensor Product Representation Probes Reveal Shared Structure Across Linear Directions

ResearchDGX agent

arXiv:2605.09967v1 Announce Type: new Abstract: While researchers are finding concepts represented as linear directions in language models, a bag of linear directions fails to capture relational struc

The Benefits of Temporal Correlations: SGD Learns k-Juntas from Random Walks Efficiently

ResearchDGX agent

arXiv:2605.10237v1 Announce Type: new Abstract: We study how temporal correlations in the data can make certain sparse learning problems efficiently learnable by gradient-based methods. Our focus is o

The Cancellation Hypothesis in Critic-Free RL: From Outcome Rewards to Token Credits

ResearchDGX agent

arXiv:2605.08666v1 Announce Type: new Abstract: A commonly accepted explanation of critic-free RL for LLMs, based on sequence-level rewards, is that it reinforces successful rollouts with a positive a

The Differences Between Direct Alignment Algorithms are a Blur

Model ReleasesDGX agent

arXiv:2502.01237v3 Announce Type: replace Abstract: Direct Alignment Algorithms (DAAs) simplify LLM alignment by directly optimizing policies, bypassing reward modeling and RL. While DAAs differ in th

The Ensemble Schr{odinger Bridge filter for Nonlinear Data Assimilation

ResearchDGX agent

arXiv:2512.18928v3 Announce Type: replace Abstract: This work introduces a novel nonlinear optimal filtering method, termed the Ensemble Schr{odinger Bridge nonlinear filter. The proposed filter combi

The finite expression method for turbulent dynamics with high-order moment recovery

TutorialsDGX agent

arXiv:2605.10687v1 Announce Type: new Abstract: Turbulent dynamical systems are characterized by nonlinear interactions and stochastic effects that generate coupled statistical quantities, such as non

The Geometric Structure of Models Learning Sparse Data

SafetyDGX agent

arXiv:2605.08464v1 Announce Type: new Abstract: The manifold hypothesis (MH) is often used to explain how machine learning can overcome the curse of dimensionality. However, the MH is only applicable

← Previous
1…174175176177178…243
Next →