AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,329 results
11 Aug 2026

Sparse corruption in low-rank matrix inference: the PCA benchmark

Model ReleasesDGX agent

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA

Spatial Heterogeneity-Aware Multi-Hazard Susceptibility and Risk Mapping at Regional Scale

ResearchDGX agent

arXiv:2608.08321v1 Announce Type: new Abstract: Floods and landslides often co-occur, but their relationships with environmental controls vary spatially. This study develops a spatial heterogeneity-aw

SPD Learn: A Geometric Deep Learning Python Library for Neural Decoding Through Trivialization

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.22895v2 Announce Type: replace-cross Abstract: Implementations of symmetric positive definite (SPD) matrix-based neural networks for neural decoding remain fragmented across research codeba

SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding

Model ReleasesDGX agent

arXiv:2608.07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Th

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

Model ReleasesDGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

Stochastic gradient descent with discontinuity across a manifold

ResearchDGX agent

arXiv:2608.07618v1 Announce Type: cross Abstract: Stochastic gradient descent for a loss function discontinuous across lower dimensional manifolds is analyzed by studying its differential equation lim

Support Selection Beyond Smooth DAG Exactness: Completion Geometry,Score Margins, and Selective Certificates

ResearchDGX agent

arXiv:2608.08103v1 Announce Type: new Abstract: Smooth acyclicity constraints answer whether a weighted support is a DAG, whereas structure learning asks which support change should be made. Existing

SwiftQK: Fast and Communication-Efficient Tensor Parallelism for Query-Key Normalization

HardwareDGX agent

arXiv:2608.09160v1 Announce Type: new Abstract: Query-Key Normalization (QK-Norm) improves the training stability and quality of modern Large Language Models (LLMs). However, under Tensor Parallelism

Targeted Label-Flipping and Oversampling Attacks on Federated Conditional GANs

Local AiDGX agent

arXiv:2608.09314v1 Announce Type: new Abstract: In a federated learning setup for GANs, several adversarial attacks are possible. One such attack is label flipping, in which malicious clients delibera

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing

Model ReleasesDGX agent

arXiv:2608.08528v1 Announce Type: new Abstract: Enterprise AI coding assistants incur substantial inference spend, and naive token-cost minimization often fails to reduce end-to-end cost once retries,

Test-time Generalization for Physics through Neural Operator Splitting

Model ReleasesDGX agent

arXiv:2602.00884v2 Announce Type: replace Abstract: Neural operators have shown promise in learning solution maps of partial differential equations (PDEs), but they often struggle to generalize when t

Test-Time Scaling for CAD Generation via Verifier-Free Consensus Selection

ResearchDGX agent

arXiv:2608.09706v1 Announce Type: cross Abstract: Large language models can write parametric CAD programs from a natural-language description (text-to-CAD generation), but a single sample is often wro

The Cost of Adaptivity: Matching Lower Bounds Across Learning Problems

Model ReleasesDGX agent

arXiv:2608.08826v1 Announce Type: new Abstract: Adaptive procedures must work without nuisance information an oracle may use, such as a gradient scale or smoothness index, and robust procedures may ha

The Neural Division of Labor: Biologically-Inspired Modular Architectures for Robust Neuromorphic Computing

ResearchDGX agent

arXiv:2608.08317v1 Announce Type: new Abstract: Biological neural systems achieve high efficiency and robustness through compartmentalized architectures. In contrast, modern artificial neural networks

The Sample Complexity of Policy Learning with Mu-Resets

SafetyDGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

The Spectral Neuron

Local AiDGX agent

arXiv:2608.08003v1 Announce Type: cross Abstract: As machine learned models increase in complexity and expressive power, features of simpler models, such as interpretability and control over the shape

Tracing sources of epistemic uncertainty in deep learning predictions: homo- and hetero-scedastic linearized estimators

ApplicationsDGX agent

arXiv:2608.07630v1 Announce Type: new Abstract: We adapt two classical statistical estimators for quantifying uncertainty to modern deep learning, in order to provide clearer insights into uncertainty

Tracking the Best Strategy in an Extensive-Form Game

Model ReleasesDGX agent

arXiv:2608.09501v1 Announce Type: new Abstract: We consider the extensive-form bandit problem where on each trial the learner plays an extensive-form game against an oblivious adversary. We focus on t

Training-Free Universal Approximation by Prompting Random Transformers

ResearchDGX agent

arXiv:2608.09558v1 Announce Type: new Abstract: How expressive is prompting a transformer? Answering this question is important for separating the roles of prompting, architecture, and pretraining in

Trajectory Design and Budgeted Querying for Digital Twin Calibration

Model ReleasesDGX agent

arXiv:2608.08631v1 Announce Type: new Abstract: Digital-twin calibration requires interaction data that is expensive to collect. We study two acquisition decisions: which trajectories to generate, and

Transfer Learning-Enabled Distortion Compensation for Amplitude-Phase-Time Block Modulation-Based Nonlinear Single-Carrier Wireless Communications

ResearchDGX agent

arXiv:2608.08554v1 Announce Type: cross Abstract: Power amplifier (PA) nonlinearity and memory effects significantly limit the spectral compliance, reliability, and energy efficiency of communication

Transformers for Multimodal Brain State Decoding: Integrating Functional Magnetic Resonance Imaging Data and Medical Metadata

ResearchDGX agent

arXiv:2512.08462v2 Announce Type: replace Abstract: Decoding brain states from functional magnetic resonance imaging (fMRI) data is vital for advancing neuroscience and clinical applications. While tr

TS-Mob: Social and Geographical-Aware Time Series Foundation-Model Framework for Human Mobility Prediction

TutorialsDGX agent

arXiv:2507.00945v2 Announce Type: replace Abstract: Short-term forecasting of aggregated human mobility flows supports urban planning, intelligent transportation systems, and emergency response, yet e

TSDS-Toolbox: A Toolbox for Measuring Time-Series Dataset Similarity

ResearchDGX agent

arXiv:2608.08119v1 Announce Type: new Abstract: The rapid advancement of artificial intelligence (AI) has significantly accelerated research in time-series analysis, particularly in forecasting, class

Twin Rollouts: Noise-Coupled Counterfactual Branching in Interactive Video World Models

ResearchDGX agent

arXiv:2608.08982v1 Announce Type: new Abstract: Interactive video world models generate rollouts autoregressively under an action stream, yet they are trained and evaluated almost exclusively on factu

Understanding Alternating Minimization for Matrix Completion

ResearchDGX agent

arXiv:1312.0925v4 Announce Type: replace Abstract: Alternating Minimization is a widely used and empirically successful heuristic for matrix completion and related low-rank optimization problems. Theo

Unimodality-Promoting Regularized Learning for Ordinal Regression

SafetyDGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control

ApplicationsDGX agent

arXiv:2608.07870v1 Announce Type: new Abstract: Improving sample efficiency remains a core challenge in reinforcement learning (RL), especially in real-world settings like robotics, where data collect

Variance reduction in lattice QCD observables via normalizing flows

ResearchDGX agent

arXiv:2603.02984v2 Announce Type: replace-cross Abstract: Normalizing flows can be used to construct unbiased, reduced-variance estimators for lattice field theory observables that are defined by a de

Walk-on-Spheres Monte Carlo and deep neural network approximations of elliptic PDEs with drift and killing

ResearchDGX agent

arXiv:2608.09494v1 Announce Type: cross Abstract: In this paper we provide Monte Carlo and deep neural network approximations for stochastic representations of solutions to linear elliptic partial dif

Weak Correlations as the Underlying Principle for Linearization of Gradient-Based Learning Systems

Model ReleasesDGX agent

arXiv:2401.04013v2 Announce Type: replace Abstract: Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of in

What Would Fix This RAG Failure? Auditing Counterfactual Response with Paired Evidence Interventions

Model ReleasesDGX agent

arXiv:2608.08944v1 Announce Type: cross Abstract: A failed retrieval-augmented generation (RAG) answer can be consistent with several unseen responses to evidence repair. We introduce Pair-ID, an offl

When Can Fraud Operations Authorize Automation? A Decision-Support Framework for Fresh Audit Evidence and Review Workload

ResearchDGX agent

arXiv:2608.08577v1 Announce Type: new Abstract: Fraud operations must allocate events among automatic approval, analyst review, and automatic blocking even though the labels needed to evaluate these a

When Counterbalancing Hides the Bias: Access-Conditioned Position Lock in Forced-Choice LLM Evaluation

Model ReleasesDGX agent

arXiv:2607.10202v2 Announce Type: replace Abstract: Forced-choice probes with counterbalanced orientations are a standard tool for measuring language-model 'value dispositions,' and a concentration/ex

When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition

Model ReleasesDGX agent

arXiv:2608.09490v1 Announce Type: new Abstract: Task arithmetic treats fine-tuning displacements as composable directions in weight space, yet it remains unclear when parameter addition reflects predi

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

SafetyDGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

When should we trust the annotation? Selective prediction for molecular structure retrieval from mass spectra

Model ReleasesDGX agent

arXiv:2603.10950v2 Announce Type: replace Abstract: Machine learning methods for identifying molecular structures from tandem mass spectra (MS/MS) have advanced rapidly, yet current approaches still e

When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs

Model ReleasesDGX agent

arXiv:2608.08542v1 Announce Type: new Abstract: Model merging has become the default way to give an aligned language model new skills without retraining: a practitioner folds task vectors from math, c

ZeroLock: Concurrent Memory-Efficient LLM Training via Modular Update Decoupling

Local AiDGX agent

arXiv:2608.07974v1 Announce Type: new Abstract: Large language model (LLM) fine-tuning at the edge adapts the model to scenario-specific data while preserving privacy. Although existing studies propos

10 Aug 2026

A foundation-model approach to pediatric headache classification from rs-fMRI

ResearchDGX agent

arXiv:2608.07287v1 Announce Type: new Abstract: Headache is the most common neurological disorder in children and substantially affects quality of life. We investigated whether resting-state functiona

A proximal subgradient method for nonconvex stochastic optimization under the Kurdyka-{L}ojasiewicz condition

ResearchDGX agent

arXiv:2608.05460v1 Announce Type: cross Abstract: This work introduces a proximal stochastic subgradient method for minimizing the sum of an expected cost, whose integrand is potentially nonsmooth and

A Rate Separation for Agnostic Direct Sums

ResearchDGX agent

arXiv:2608.06951v1 Announce Type: new Abstract: Hanneke, Moran, and Waknine ite{HannekeMoranWaknine2024} asked how the agnostic PAC learning curve of the direct sum C^r depends on the single-instance

A Transferable Autologistic Model for Predicting Rare Failures in Heterogeneous Equipment

ResearchDGX agent

arXiv:2608.06695v1 Announce Type: new Abstract: Predicting failures before they occur remains a major challenge in predictive maintenance, particularly when failures are rare, when equipment of the sa

Adversarial Causal Intervention Falsification

ResearchDGX agent

arXiv:2608.06427v1 Announce Type: new Abstract: Generative models can reproduce an observational distribution while encoding an incorrect causal structure. We study a sequential game in which a struct

ArchEGraph: A Large-Scale Graph Dataset for Geometry-Topology-Physics Aligned Building Energy Modeling

Model ReleasesDGX agent

arXiv:2608.06772v1 Announce Type: new Abstract: Accurate estimation of building energy use is essential for achieving carbon neutral and sustainable buildings. To better understand the influence of de

BDD2Seq: Enabling Scalable Reversible-Circuit Synthesis via Graph-to-Sequence Learning

ResearchDGX agent

arXiv:2511.08315v2 Announce Type: replace-cross Abstract: Binary Decision Diagrams (BDDs) are instrumental in many electronic design automation (EDA) tasks thanks to their compact representation of Bo

Beyond Attention: Signed Integrated Gradients Attribution in a BiomeGPT-Style Microbiome Transformer

ResearchDGX agent

arXiv:2608.06486v1 Announce Type: new Abstract: In a feature-tokenized transformer (arXiv:2106.11959) such as BiomeGPT (doi:10.64898/2026.01.05.697599), each input token is built by fusing a fixed ide

Beyond Co-Movement: Locality by Exposures Enables a Joint Factor-Graph Framework for Portfolio Diversification

ResearchDGX agent

arXiv:2608.06618v1 Announce Type: cross Abstract: Current portfolio construction methods are either agnostic to the effects of idiosyncratic shocks (standard factor models) or to the latent data struc

Beyond Myopic World Models: Long-Horizon End-to-End Training for Direct Future Prediction

Local AiDGX agent

arXiv:2608.07420v1 Announce Type: new Abstract: World models are expected to support imagination over extended temporal horizons, yet most are still trained through local few-step prediction objective

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration

SafetyDGX agent

arXiv:2608.07419v1 Announce Type: new Abstract: Preference alignment often makes large language models (LLMs) overconfident and poorly calibrated. Traditional post-hoc temperature scaling is inherentl

Bootstrap-Conditioned Action Selection with Tabular Foundation Models

SafetyDGX agent

arXiv:2608.06559v1 Announce Type: new Abstract: Contextual bandits offer a natural framework for sample-efficient personalization, but practical deployment remains difficult under sparse, biased inter

Capacity Confounds and Coverage Guarantees in Adaptive Sub-model Federated Learning

Model ReleasesDGX agent

arXiv:2608.07157v1 Announce Type: new Abstract: Sub-model federated learning lets resource-constrained clients train width-reduced versions of a global model, but existing methods allocate capacity by

Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving

SafetyDGX agent

arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges

Cascading Through the Hierarchy: Regularizer-Induced Feature Detection as Phase Transitions in Deep Linear Neural Networks

Model ReleasesDGX agent

arXiv:2608.06597v1 Announce Type: cross Abstract: A scientific theory of deep learning, comprising learning dynamics and statistical properties of learned models, is rapidly gaining attention. One of

Certified Feedforward Tracking for Unknown Nonlinear Systems via Invertible Neural Networks

ResearchDGX agent

arXiv:2608.06419v1 Announce Type: cross Abstract: In this paper, we address the certification of datadriven feedforward control for periodic tracking of unknown nonlinear systems under partial state m

Certified Interpolation Oversampling: Per-Instance Safety Guarantees for Imbalanced Learning

Model ReleasesDGX agent

arXiv:2501.15790v2 Announce Type: replace Abstract: Synthetic minority oversampling is typically designed and evaluated against a predictive objective, generating samples that improve downstream class

CHIME: A Case for Efficient Long-Context Attention-FC Disaggregated Inference with DIMM-PIM

SafetyDGX agent

arXiv:2504.17584v2 Announce Type: replace-cross Abstract: Attention-FC Disaggregated (AFD) LLM inference systems offload memory-bound Attention operations to memory-rich accelerators (e.g., CPUs, HBM-

Cloud-Boosted Low-Compute Multi-Channel Speech Enhancement

Local AiDGX agent

arXiv:2608.07423v1 Announce Type: cross Abstract: Low-latency, low-compute speech enhancement is essential for wearable devices with real-time communication requirements, but strict computational cons

Conditioning Protein Generation via Hopfield Pattern Multiplicity

ResearchDGX agent

arXiv:2603.20115v2 Announce Type: replace Abstract: Small protein-family alignments often contain a subset of interest but not enough labeled data to train a conditional generator. We condition a trai

Conformal Fusion Under Missing Modalities

ResearchDGX agent

arXiv:2608.07183v1 Announce Type: new Abstract: Multimodal fusion architectures typically assume all modalities are available at inference, yet sensor failures, acquisition variability, and cost const

← Previous
1…45678…239
Next →