AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
4 Jun 2026

(Mis)generalization of Helpful-only Fine-tuning

SafetyDGX agent

arXiv:2606.04413v1 Announce Type: new Abstract: Helpful-only models, that is, models that are trained to always follow user intent, are valuable for dangerous capability evaluations and other areas of

Modeling and Interpreting Teamwork Dynamics in Cancer Care Outcome Prediction

TutorialsDGX agent

arXiv:2606.04499v1 Announce Type: cross Abstract: Cancer care requires a longitudinal approach in which treatments are planned and delivered over time according to the needs of each individual patient

Near-Optimal Decentralized Stochastic Convex Optimization over Networks

ResearchDGX agent

arXiv:2606.04757v1 Announce Type: cross Abstract: We study decentralized stochastic smooth convex optimization, where M workers minimize an average objective using local stochastic gradients and neigh


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Neural Galerkin Normalizing Flows for Bayesian Inference of Diffusions with Inaccessible Boundaries

Model ReleasesDGX agent

arXiv:2606.04324v1 Announce Type: new Abstract: One of the primary challenges in Bayesian inference on the parameters of a diffusion model from discrete observations is the unavailability of an analyt

Neural Langevin Machine: a local asymmetric learning rule can be creative

ResearchDGX agent

arXiv:2506.23546v2 Announce Type: replace-cross Abstract: Fixed points of recurrent neural networks can be leveraged to store and generate information. These fixed points can be captured by the Boltzm

New Benchmarking Shows Limited Generalization Power of TCR Antigenic Epitope Prediction Models

Model ReleasesDGX agent

arXiv:2606.04994v1 Announce Type: new Abstract: Accurate computational prediction of T cell receptor (TCR) antigen specificity would transform the study of T cell biology and enable scalable immune en

NLLog: Lightweight, Explainable SOC Anomaly Detection via Log-to-Language Rewriting

ResearchDGX agent

arXiv:2606.04957v1 Announce Type: cross Abstract: System-generated logs underpin security monitoring, yet their rigid template-based format hinders both automated analysis and human comprehension. We

Nonlocal Mean Field Schrodinger Bridge with Learned Interactions

ResearchDGX agent

arXiv:2606.04265v1 Announce Type: cross Abstract: The Schrodinger Bridge Problem constructs a stochastic process that connects an initial distribution to a terminal distribution with minimum energy. T

Novel Aspects of IEEE SA P3109 Arithmetic Formats for Machine Learning

ResearchDGX agent

arXiv:2606.04028v1 Announce Type: new Abstract: The IEEE P3109 draft standard defines a parameterized family of binary floating-point formats and associated operations, with a focus on facilitating ma

Offline-to-Online Learning in Linear Bandits

ResearchDGX agent

arXiv:2606.04305v1 Announce Type: new Abstract: We study online learning with an additional offline dataset in the stochastic linear bandit setting. Although this problem arises frequently in practice

On Forgetting and Stability of Score-based Generative models

ResearchDGX agent

arXiv:2601.21868v2 Announce Type: replace-cross Abstract: Understanding the stability and long-time behavior of generative models is a fundamental problem in modern machine learning. This paper provid

On Out-of-sample Embedding in UMAP

ResearchDGX agent

arXiv:2606.04451v1 Announce Type: new Abstract: Neighbor embedding algorithms reveal correlations in high-dimensional data by constructing an equivalent graph representation in a lower-dimensional spa

On the Expressive Power of Permutation-Equivariant Weight-Space Networks

ResearchDGX agent

arXiv:2602.01083v2 Announce Type: replace Abstract: Weight-space learning studies neural architectures that operate directly on the parameters of other neural networks. Motivated by the growing availa

On the Relationship Between CoCoA and ADMM for Distributed Empirical Risk Minimization

Model ReleasesDGX agent

arXiv:2502.00470v3 Announce Type: replace-cross Abstract: Distributed empirical risk minimization (ERM) is often studied through two influential yet seemingly separate families of methods: CoCoA-type

Optimal Transport under Group Fairness Constraints

SafetyDGX agent

arXiv:2601.07144v3 Announce Type: replace-cross Abstract: Ensuring fairness in matching algorithms is a key challenge in allocating scarce resources and positions. Focusing on Optimal Transport (OT),

Orthogonal Learner for Estimating Heterogeneous Long-Term Treatment Effects

ApplicationsDGX agent

arXiv:2604.00915v2 Announce Type: replace Abstract: Estimation of heterogeneous long-term treatment effects (HLTEs) is relevant for personalized decision-making in marketing, economics, and medicine,

Overclocking Electrostatic Generative Models

ResearchDGX agent

arXiv:2509.22454v2 Announce Type: replace Abstract: Electrostatic generative models such as PFGM++ have recently emerged as a powerful framework, achieving competitive performance in image synthesis.

Path-conditioned training: a principled way to rescale ReLU neural networks

SafetyDGX agent

arXiv:2602.19799v2 Announce Type: replace-cross Abstract: Despite recent algorithmic advances, we still lack principled ways to leverage the well-documented rescaling symmetries in ReLU neural network

PE-MHL: Physics-Encoded Modular Hybrid Layers for Scalable Learning of Complex Systems

Model ReleasesDGX agent

arXiv:2606.04290v1 Announce Type: new Abstract: Hybrid models that combine physics-based and data-driven components have shown strong potential for achieving accuracy and interpretability in control a

Policy Gradient for Continuous-Time Robust Markov Decision Processes

SafetyDGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

Prediction Under Imperfect Compression: A Theory of Approximate MDL

ResearchDGX agent

arXiv:2606.04834v1 Announce Type: new Abstract: Minimum Description Length (MDL) formalizes the principle of Occam's razor by optimizing the total description length: L(model)+L(data | model). For seq

Preserving Data Privacy in Learning Causal Structure with Fully Homomorphic Encryption

ResearchDGX agent

arXiv:2606.05129v1 Announce Type: cross Abstract: Preserving data privacy is an important topic in structural data management and data mining. However, the issue of privacy leakage in distributed caus

Provably Reduced Sample Cost in Prior-Guided Hyperparameter Optimization

Model ReleasesDGX agent

arXiv:2606.04866v1 Announce Type: new Abstract: Large-scale hyperparameter optimization (HPO) in automated machine learning (AutoML) consumes substantial computational resources, raising growing conce

Pseudospectral Bounds for Transient Amplification in Coupled Gradient Descent

Model ReleasesDGX agent

arXiv:2606.04031v1 Announce Type: new Abstract: Coupled gradient descent--where the update of one parameter block depends on another--underlies bilevel optimization, two-time-scale stochastic approxim

QPredSGG: Hybrid Quantum Predicate Learning for Long-Tailed Scene Graph Generation

Model ReleasesDGX agent

arXiv:2606.04689v1 Announce Type: cross Abstract: Scene Graph Generation (SGG) requires relational reasoning over objects and their interactions, but performance is often limited by severe long-tail p

RAVQ-HoloNet: Rate-Adaptive Vector-Quantized Hologram Compression

Model ReleasesDGX agent

arXiv:2511.21035v2 Announce Type: replace Abstract: Holography offers significant potential for AR/VR applications. However, its adoption is limited by the high demand for data compression. Existing d

Reasoning Shift: How Context Silently Shortens LLM Reasoning

ResearchDGX agent

arXiv:2604.01161v2 Announce Type: replace Abstract: Large language models (LLMs) exhibiting test-time scaling behavior, such as extended reasoning traces and self-verification, have demonstrated remar

Reconciling Causality and Non-Equilibrium Thermodynamics with Hamiltonian Causal Models

ApplicationsDGX agent

arXiv:2606.04822v1 Announce Type: new Abstract: Causal modeling of physical temporal phenomena must handle interventions that act along trajectories, nonstationary induced laws, path-dependent effects

Reconstructing Unobservable Temperature Fields via Simulation-Aided Intelligent Sensing

ResearchDGX agent

arXiv:2606.04582v1 Announce Type: cross Abstract: Real-time monitoring of the temperature distribution within components and sub-structures is a challenging topic in many systems due to restrictions o

Reducing the Filtering Effect in Public School Admissions: A Bias-aware Analysis for Targeted Interventions

SafetyDGX agent

arXiv:2004.10846v5 Announce Type: replace-cross Abstract: Problem definition: Traditionally, New York City's top 8 public schools have selected candidates solely based on their scores in the Specializ

REGAIN: REconciliation GAIN-driven Auxiliary Direction Learning

SafetyDGX agent

arXiv:2606.04380v1 Announce Type: cross Abstract: Forecast reconciliation usually starts from a fixed measurement system and asks how forecasts should be projected onto a coherent space. We ask a diff

RePercENT: Scaling Disentangled Representation Learning Beyond Two Modalities

SafetyDGX agent

arXiv:2606.05109v1 Announce Type: new Abstract: To leverage the full potential of multimodal data, we need representations that go beyond the state-of-the-art alignment and fusion approaches and explo

Representation Matters in Randomized Smoothing for Audio Classification

SafetyDGX agent

arXiv:2606.04210v1 Announce Type: cross Abstract: Randomized smoothing (RS) certifies robustness in the vector space where Gaussian noise is added. In audio classification, this space is often not uni

ReSGA: A Large Tail Risk Model for Learning Value-at-Risk and Expected Shortfall

ResearchDGX agent

arXiv:2606.04576v1 Announce Type: cross Abstract: Learning Value-at-Risk (VaR) and Expected Shortfall (ES) is important for managing financial risks effectively. Existing approaches with limited param

Rethinking Incompleteness: Formalizing Protocol Divergence and Train-Once Learning for Robust IMVC

ResearchDGX agent

arXiv:2606.04857v1 Announce Type: new Abstract: Standard IMVC evaluation retrains separate models for different missing-data configurations. We show that this paradigm obscures a fundamental vulnerabi

Reusing Trajectories in Policy Gradients Enables Fast Convergence

SafetyDGX agent

arXiv:2506.06178v3 Announce Type: replace Abstract: Policy gradient (PG) methods are a class of effective reinforcement learning algorithms, particularly when dealing with continuous control problems.

Revisiting Privacy Amplification by Subsampling in Selective Release DPSGD

ResearchDGX agent

arXiv:2606.04384v1 Announce Type: new Abstract: Machine learning's reliance on sensitive data necessitates privacy-preserving techniques like Differentially Private Stochastic Gradient Descent (DPSGD)

RIDE: An Open Dataset and Benchmark for Train Delay Prediction

Model ReleasesDGX agent

arXiv:2606.05070v1 Announce Type: new Abstract: Train delay prediction is an important problem for both passengers and railway operators, yet progress in the field remains difficult to assess due to t

RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training

SafetyDGX agent

arXiv:2606.04272v1 Announce Type: new Abstract: The standard LLM training pipeline applies reinforcement learning (RL) only after pre-training and supervised fine-tuning (SFT). We question this status

SC-TauPath: A Structural Connectivity Attribution Framework for Mapping Tau Propagation Pathways in Alzheimer's Disease

ResearchDGX agent

arXiv:2606.04066v1 Announce Type: cross Abstract: Understanding how structural connections are associated with tau propagation in Alzheimer's disease (AD) remains a central open question, yet existing

Scaling Datasets for Multi-Sensor, Multi-Agent, and Multi-Domain Learning in Autonomous Systems

AgentsDGX agent

arXiv:2606.04444v1 Announce Type: cross Abstract: Existing datasets cannot support large-scale learning in multi-agent, multi-sensor, or multi-domain autonomy, where diversity and coordination are ess

Scheduling in Queueing Systems with Uncertain and Evolving Holding Costs

ResearchDGX agent

arXiv:2505.21331v2 Announce Type: replace-cross Abstract: In content moderation for social media platforms, the cost of delaying the review of a content is proportional to its view trajectory, which f

Self-Distilled Policy Gradient

SafetyDGX agent

arXiv:2606.04036v1 Announce Type: new Abstract: On-policy self-distillation, where a language model conditions on privileged context to supervise its own generations, is a promising source of dense su

Sequential Data Poisoning in LLM Post-Training

ResearchDGX agent

arXiv:2606.04929v1 Announce Type: new Abstract: LLM post-training proceeds through multiple stages, e.g., supervised fine-tuning (SFT) followed by reinforcement learning from human feedback (RLHF) or

SFMP: Fine-Grained, Hardware-Friendly and Search-Free Mixed-Precision Quantization for Large Language Models

ResearchDGX agent

arXiv:2602.01027v2 Announce Type: replace Abstract: Mixed-precision quantization is a promising approach for compressing large language models under tight memory budgets. However, existing mixed-preci

Shortcomings and capacities of real-constrained neural networks in complex spaces

ResearchDGX agent

arXiv:2606.04390v1 Announce Type: new Abstract: We find the asymptotic ratio between the storage capacities when enforcing real pre-activations in a complex hypothesis class as opposed to complex ones

Sparse Bayesian Deep Functional Learning with Structured Region Selection

ApplicationsDGX agent

arXiv:2602.20651v3 Announce Type: replace Abstract: In modern applications such as ECG monitoring, neuroimaging, wearable sensing, and industrial equipment diagnostics, complex and continuously struct

SpliceBind: Isoform-Aware Prediction of Binding Pocket Druggability

ResearchDGX agent

arXiv:2606.04020v1 Announce Type: cross Abstract: Splice-mediated drug resistance occurs in up to 40% of patients on targeted kinase inhibitors, yet state-of-the-art druggability tools operate on sing

SPLIT-PINN: Separable Probability Learning Technique via Physics-Informed Neural Networks for High-Dimensional Probabilistic Modeling

Model ReleasesDGX agent

arXiv:2606.04000v1 Announce Type: cross Abstract: We present a probabilistic modeling framework for incorporating small-scale spatial heterogeneity into macroscopic descriptions of material behavior f

STaR-Quant: State-Time Consistent Post-Training Quantization for Diffusion Large Language Models

ResearchDGX agent

arXiv:2606.04945v1 Announce Type: new Abstract: Diffusion large language models (DLLMs) have recently emerged as a promising alternative to autoregressive LLMs by generating text through iterative mas

Stationarity-Aware Retrieval-Augmented Time Series Forecasting

ApplicationsDGX agent

arXiv:2606.04135v1 Announce Type: new Abstract: Time series forecasting relies on historical patterns, but real-world series often exhibit non-stationarity and regime shifts that challenge fully param

Stein Kernelized Molecular Dynamics for Active Learning of Interatomic Potentials

TutorialsDGX agent

arXiv:2606.04100v1 Announce Type: new Abstract: Machine learning interatomic potentials (MLIPs) enable efficient and accurate atomistic simulations but depend critically on the quality and diversity o

Structure-Aware Prediction of PROTAC-Mediated Protein Degradability via Graph Neural Networks

Model ReleasesDGX agent

arXiv:2606.04021v1 Announce Type: cross Abstract: Proteolysis-targeting chimeras (PROTACs) can selectively degrade disease-causing proteins, yet predicting which targets are amenable to degradation re

SurvPFN: Towards Foundation Models for Survival Predictions

TutorialsDGX agent

arXiv:2606.04564v1 Announce Type: new Abstract: Tabular foundation models (TFMs) have made rapid progress in standard classification and regression, but time-to-event survival prediction tasks have re

Symbolic Regression for Shared Expressions: Introducing Partial Parameter Sharing

Model ReleasesDGX agent

arXiv:2601.04051v3 Announce Type: replace Abstract: Symbolic regression aims to find symbolic expressions that describe datasets. Due to its inherent interpretability, symbolic regression (SR) is a po

TANDEM: Bi-Level Data Mixture Optimization with Twin Networks

ResearchDGX agent

arXiv:2606.04401v1 Announce Type: new Abstract: The capabilities of large language models (LLMs) significantly depend on training data drawn from various domains. Optimizing domain-specific mixture ra

Test-Time Compute Scaling for ASR with Depth-Conditioned Looped Transformers

Model ReleasesDGX agent

arXiv:2606.04678v1 Announce Type: new Abstract: End-to-end ASR systems typically use fixed-depth acoustic encoders at inference, making it difficult to trade additional test-time computation for impro

Testing Neural Networks via Bayesian-Guided Exploration of Decision Landscapes

SafetyDGX agent

arXiv:2606.04314v1 Announce Type: new Abstract: As neural networks are increasingly deployed in safety-critical domains, testing is essential to evaluate and improve their reliability. Existing testin

The Cost of Learning Under Multiple Change Points

ApplicationsDGX agent

arXiv:2602.11406v2 Announce Type: replace-cross Abstract: We consider an online learning problem in environments with multiple change points. In contrast to the single change point problem that is wid

The price of multi-group transductive learning

ResearchDGX agent

arXiv:2606.04423v1 Announce Type: new Abstract: We show every multi-group learner in the transductive setting may incur a multiplicative penalty in its error rate on some group relative to the error r

← Previous
1…99100101102103…243
Next →