AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
29 May 2026

A Theoretical and Experimental Study of a Novel Adaptive Learning Algorithm

ApplicationsDGX agent

arXiv:2605.29273v1 Announce Type: new Abstract: A crucial component of machine learning algorithms is minimizing loss functions with less computational cost and less oscillations. While adaptive learn

A Training-Time Diagnostic for Generalization via the Log-Alignment Ratio

Model ReleasesDGX agent

arXiv:2605.28975v1 Announce Type: new Abstract: We study the log-alignment ratio (LAR), a measure of parameter-activation alignment, introduced in parameterization theory. We reformulate it as the ove

A Triple-Modal Contrastive Learning Framework with Sequence, Graph, and 3D Features for Drug-Target Interaction Prediction

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.29926v1 Announce Type: new Abstract: Accurate prediction of drug-target interactions (DTI) is critical for drug discovery. Existing methods often rely on single-modal representations (e.g.,

Access Sets Matter: Budgeting Expert Reads for Scalable Weight-Space Model Merging

Model ReleasesDGX agent

arXiv:2605.29489v1 Announce Type: new Abstract: Weight-space model merging is usually formulated as an algebraic operation on checkpoints, yet at LLM scale the limiting resource is often the set of ex

Achieving Linear Speedup for Composite Federated Learning

ResearchDGX agent

arXiv:2602.03357v2 Announce Type: replace Abstract: This paper proposes FedNMap, a normal map-based method for composite federated learning, where the objective consists of a smooth loss and a possibl

Active Continual Learning with Metaplastic Binary Bayesian Neural Networks

ResearchDGX agent

arXiv:2605.30198v1 Announce Type: new Abstract: Always-on edge systems must keep learning as conditions change under tight compute budgets and must detect unreliable predictions. Bayesian binary neura

Active Learning for Machine Learning Driven Molecular Dynamics

Model ReleasesDGX agent

arXiv:2509.17208v3 Announce Type: replace Abstract: Machine-learned coarse-grained (CG) potentials are fast, but degrade over time when simulations reach under-sampled bio-molecular conformations, and

Adapting Automotive Aerodynamics Surrogates to New Vehicle Families via Transfer Learning

Model ReleasesDGX agent

arXiv:2605.27968v1 Announce Type: cross Abstract: Deploying Scientific Machine Learning surrogates in industrial CFD workflows requires adapting pretrained models to new vehicle families without large

Adaptive Exponential Integration for Stable Gaussian Mixture Black-Box Variational Inference

ResearchDGX agent

arXiv:2601.14855v3 Announce Type: replace Abstract: Black-box variational inference (BBVI) with Gaussian mixture families offers a flexible approach for approximating complex posterior distributions w

Aggregate Models, Not Explanations: Improving Feature Importance Estimation

ResearchDGX agent

arXiv:2602.11760v2 Announce Type: replace-cross Abstract: Feature-importance methods show promise in transforming machine learning models from predictive engines into tools for scientific discovery. H

AMDP: Asynchronous Multi-Directional Pipeline Parallelism for Large-Scale Models Training

Model ReleasesDGX agent

arXiv:2605.29664v1 Announce Type: cross Abstract: Pipeline parallelism is essential for large-scale model training, but existing asynchronous approaches often degrade convergence due to parameter mism

An End-to-End PyTorch Interface for Differentiable PDE Solvers: A RANS Model-Correction Study

Model ReleasesDGX agent

arXiv:2605.28858v1 Announce Type: cross Abstract: This work presents an end-to-end strategy for solving inverse problems constrained by Partial Differential Equations within a fully differentiable Mac

Anti Mode-Collapse in Mean-Field Transformer via Auxiliary Variables

ResearchDGX agent

arXiv:2605.30229v1 Announce Type: new Abstract: We use a mean-field-based transformer model to theoretically investigate how auxiliary variables, such as positional encoding, prevent mode collapse of

Anytime-Valid Federated Conformal RAG for LLM Swarms

SafetyDGX agent

arXiv:2605.29139v1 Announce Type: cross Abstract: Federated Conformal RAG (FC-RAG) provides distribution-free coverage for a bandwidth-limited swarm of weak language models, but only at a fixed horizo

Apertus LLM Family Expansion via Distillation and Quantization

Model ReleasesDGX agent

arXiv:2605.29128v1 Announce Type: new Abstract: The wide adoption of LLMs has led to their use in great variety of applications and scenarios, such as chatbot assistants and data annotation, creating

AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference

Local AiDGX agent

arXiv:2605.29535v1 Announce Type: new Abstract: Vision-Language Models (VLMs) process thousands of visual tokens per image alongside comparatively few text tokens, yet existing compression methods tre

Attention as In-Context Empirical Bayes: A Two-Stage View via Particle Dynamics

ResearchDGX agent

arXiv:2605.29351v1 Announce Type: new Abstract: We study minimal attention-only transformers under all-token corruption and show they admit a two-stage empirical Bayes interpretation. A single attenti

Auditing Training Data in Generative Music Models via Black-Box Membership Inference

SafetyDGX agent

arXiv:2605.29202v1 Announce Type: new Abstract: Recent advances in text-to-music generation enable high-fidelity synthesis of structured musical audio, raising growing concerns about data provenance,

Bandit Algorithms for Deep Brain Stimulation

Model ReleasesDGX agent

arXiv:2601.12699v2 Announce Type: replace Abstract: Deep Brain Stimulation (DBS) is an effective treatment for Parkinson's disease, but conventional fixed-parameter stimulation can reduce battery life

Bastion: Budget-Aware Speculative Decoding with Tree-structured Block Diffusion Drafting

HardwareDGX agent

arXiv:2605.29727v1 Announce Type: new Abstract: Block-diffusion drafters have recently emerged as a powerful alternative for speculative decoding by predicting multiple future-token distributions in a

Bayesian model selection and misspecification testing in imaging inverse problems only from noisy and partial measurements

ResearchDGX agent

arXiv:2510.27663v3 Announce Type: replace-cross Abstract: Modern imaging techniques heavily rely on Bayesian statistical models to address difficult image reconstruction and restoration tasks. This pa

Bridging Chemists and AI: An Expert-Augmented Framework for Interpretable Route Evaluation

ResearchDGX agent

arXiv:2605.29108v1 Announce Type: new Abstract: Selecting efficient multi-step synthetic routes is a central challenge in organic synthesis, particularly in medicinal and process chemistry, where rout

Bridging Functional and Representational Similarity via Usable Information

ResearchDGX agent

arXiv:2601.21568v2 Announce Type: replace Abstract: We present a unified framework for quantifying the similarity between representations through the lens of extit{usable} information, offering a rigo

BuilDyn: Excitation-Driven Data Generation for Building Thermal Dynamics Modeling and Control

ApplicationsDGX agent

arXiv:2605.29849v1 Announce Type: cross Abstract: Machine learning (ML) is increasingly used for data-driven modeling of buildings to enable downstream tasks such as fault detection and diagnosis, and

Calibrating Generative Models to Distributional Constraints

ResearchDGX agent

arXiv:2510.10020v4 Announce Type: replace-cross Abstract: Generative models frequently suffer miscalibration, wherein statistics of the sampling distribution, such as the fraction of generations in a

Can AI Weather Models Predict Beyond Two Weeks? A Quantitative Benchmark and Analysis of Long Rollouts

Model ReleasesDGX agent

arXiv:2605.30184v1 Announce Type: new Abstract: While AI weather models excel at short-to-medium range forecasts (up to 15 days), they frequently suffer from ill-defined 'instabilities' when rolled ou

Causal Intelligence for Constraint-Aware Intervention Design to Induce State Transitions

TutorialsDGX agent

arXiv:2605.29008v1 Announce Type: new Abstract: Driving a system from one state to another through targeted interventions is a fundamental challenge in science, yet most predictive models offer limite

Certified Causal Defense with Generalizable Robustness

Model ReleasesDGX agent

arXiv:2408.15451v3 Announce Type: replace Abstract: While machine learning models have proven effective across various scenarios, it is widely acknowledged that many models are vulnerable to adversari

Chess-World-Model: A 10M-Game Benchmark for Exact State Tracking from Chess Move Sequences

Model ReleasesDGX agent

arXiv:2605.30100v1 Announce Type: new Abstract: World models require state tracking, which is the ability to maintain a correct latent state across action sequences. Existing benchmarks are often synt

CLUBench: A Clustering Benchmark

Model ReleasesDGX agent

arXiv:2605.29933v1 Announce Type: new Abstract: Clustering is a fundamental problem in data science with a long-standing research history, yielding numerous insightful algorithms. Despite this progres

Cluster-Level Attention-Guided Parallel Decoding for Masked Diffusion Language Models

ResearchDGX agent

arXiv:2605.29607v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) enable parallel decoding by predicting all masked positions at each denoising step, yet existing training-free

Coarse-Grained Boltzmann Generators

ResearchDGX agent

arXiv:2602.10637v2 Announce Type: replace Abstract: Sampling equilibrium molecular configurations from the Boltzmann distribution is a longstanding challenge. Boltzmann Generators (BGs) address this b

Collaborative Threshold Watermarking

ResearchDGX agent

arXiv:2602.10765v2 Announce Type: replace Abstract: In federated learning (FL), K clients jointly train a model without sharing raw data. Because each participant invests data and compute, clients nee

Comment on 'Spin-1/2 Kagome Heisenberg Antiferromagnet: Machine Learning Discovery of the Spinon Pair-Density-Wave Ground State'

ResearchDGX agent

arXiv:2605.28861v1 Announce Type: cross Abstract: A recent article [Phys. Rev. X 15, 011047 (2025)] utilizes group-equivariant convolutional neural networks to study the ground state of the kagome Hei

CompilerDream: Learning a Compiler World Model for General Code Optimization

AgentsDGX agent

arXiv:2404.16077v4 Announce Type: replace-cross Abstract: Effective code optimization in compilers is crucial for computer and software engineering. The success of these optimizations primarily depend

Computational Modeling of Antibody-Antigen Complexes: PLM-Based and MSA-Based Approaches

ResearchDGX agent

arXiv:2605.28886v1 Announce Type: cross Abstract: Antibodies play a central role in the immune response by specifically recognizing and neutralizing antigens, and therapeutic antibodies have become ma

Computationally Efficient Replicable Learning of Parities and Applications

ResearchDGX agent

arXiv:2602.09499v2 Announce Type: replace Abstract: We study the computational relationship between replicability (Impagliazzo et al. [STOC `22], Ghazi et al. [NeurIPS `21]) and other stability notion

Connecting Independently Trained Modes via Layer-Wise Connectivity

Model ReleasesDGX agent

arXiv:2505.02604v5 Announce Type: replace Abstract: Empirical studies have shown that continuous low-loss paths can be constructed between independently trained neural network models. This phenomenon,

Contrastive Representation Regularization for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2510.01711v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have shown strong capabilities in robot manipulation by leveraging rich representations from pre-trained V

Convergence Theory for Iterative LLM-Based Neural Architecture Search: A Parametric Cross-Entropy Framework with Closed-Form Proxy Reliability

ResearchDGX agent

arXiv:2605.30103v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as generators in iterative neural architecture search (NAS), yet no formal convergence theory exists

Convex Basins in Single-Index Model Loss Landscapes: Applications to Robust Recovery under Strong Adversarial Corruption

ResearchDGX agent

arXiv:2605.29497v1 Announce Type: new Abstract: We study the problem of robustly learning Gaussian Single Index Models (SIMs) in the presence of heavy-tailed noise and a constant fraction of adversari

Cooperative Variance Estimation and Bayesian Neural Networks for Disentangling Aleatoric and Epistemic Uncertainties

TutorialsDGX agent

arXiv:2505.02743v2 Announce Type: replace Abstract: Real-world data contains aleatoric uncertainty - irreducible noise arising from imperfect measurements or from incomplete knowledge about the data g

CRB-Guided Framework Design and Resource Allocation for Indoor mmWave ISCC Systems

ResearchDGX agent

arXiv:2605.29939v1 Announce Type: cross Abstract: Integrated sensing, communication, and computation (ISCC) provides a promising framework for indoor human-centric applications. In these applications,

Cross-Chirality Generalization by Axial Vectors for Hetero-Chiral Protein-Peptide Interaction Design

ResearchDGX agent

arXiv:2602.20176v2 Announce Type: replace-cross Abstract: D-peptide binders targeting L-proteins have promising therapeutic potential. Despite rapid advances in machine learning-based target-condition

Cycle-Space Informed Detection of Autoencoded Blind False Data Injection Attacks on Power Systems

ResearchDGX agent

arXiv:2605.28912v1 Announce Type: new Abstract: The rapid growth of AI-driven data centers and large-scale energy storage systems is increasing the reliance of power system operation on real-time meas

DCFO: Density-Based Counterfactuals for Outliers -- Additional Material

ResearchDGX agent

arXiv:2512.10659v3 Announce Type: replace Abstract: Outlier detection identifies data points that significantly deviate from the majority of the data distribution. Explaining outliers is crucial for u

Deep Adaptive Dimension Reduction for Bayesian Inference in Inverse Problems

Model ReleasesDGX agent

arXiv:2605.29373v1 Announce Type: new Abstract: Solving high-dimensional PDE-governed inverse problems is often challenging due to complex non-Gaussian posterior distributions, expensive forward model

Deep Optimal Individualized Treatment Rules for Bivariate Survival Outcomes via Adaptive Prediction-Powered Learning

ResearchDGX agent

arXiv:2605.29464v1 Announce Type: cross Abstract: In randomized trials involving multiple treatments, bivariate survival outcomes present significant analytical challenges for making decisions. This p

Designing Active Tether-Net Systems for Space Debris Capture with Graph-Learning-Aided Mixed-Combinatorial Optimization

AgentsDGX agent

arXiv:2605.29021v1 Announce Type: new Abstract: Active tether-net systems are a promising solution for capturing large non-cooperative targets, such as space debris, by deploying a flexible net manipu

Diffusion-based learning framework for Constrained Nonconvex Optimization with Weighted Bootstrapped Refinement

Model ReleasesDGX agent

arXiv:2502.10330v4 Announce Type: replace Abstract: Recent advances in diffusion models show promising potential to accelerate nonconvex problem solving by leveraging their multimodality. However, mos

Diffusion differentiable resampling

Model ReleasesDGX agent

arXiv:2512.10401v3 Announce Type: replace-cross Abstract: This paper is concerned with differentiable resampling in the context of sequential Monte Carlo (e.g., particle filtering). Drawing on reparam

Diffusion Models Are Statistically Optimal for Learning Low-Dimensional Multi-Modal Distributions

ResearchDGX agent

arXiv:2605.30153v1 Announce Type: cross Abstract: Score-based diffusion models have demonstrated remarkable empirical success in learning high-dimensional distributions, particularly those exhibiting

Digitally enriching a screening population for pancreatic cancer using routine blood-based measures and clinical histories

ResearchDGX agent

arXiv:2605.30275v1 Announce Type: new Abstract: Earlier detection of pancreatic cancer is key to enabling wider access to curative treatment and reducing cancer deaths; however, screening is presently

DiScoFormer: Plug-In Density and Score Estimation with Transformers

TutorialsDGX agent

arXiv:2511.05924v3 Announce Type: replace Abstract: Estimating probability density and its score from samples remains a core problem in generative modeling, Bayesian inference, and kinetic theory. Exi

Dissecting the Black Box: Circuit-Level Analysis of LLM Vulnerability Detection

Model ReleasesDGX agent

arXiv:2605.29901v1 Announce Type: cross Abstract: Large language models (LLMs) can detect software vulnerabilities, but how do they actually identify vulnerable code? We address this question using me

Distributionally Robust Set Representation Learning Under Inference-Time Element Corruption

ResearchDGX agent

arXiv:2605.30089v1 Announce Type: new Abstract: Standard Set Representation Learning methods typically excel on curated data but often overlook the challenge of inference-time element corruption. This

Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias

SafetyDGX agent

arXiv:2605.29152v1 Announce Type: new Abstract: Randomly initialized neural networks induce a prior over functions, but the predictor used in practice is produced only after training. We ask how much

DynaFLIP: Rethinking Robotics Perception via Tri-Modal-Dynamics Guided Representation

SafetyDGX agent

arXiv:2605.30350v1 Announce Type: cross Abstract: Robot manipulation critically depends on perception that preserves the action-relevant aspects of a scene. Yet most robot learning pipelines are built

Dynamic Mixture of Progressive Parameter-Efficient Expert Library for Lifelong Robot Learning

Model ReleasesDGX agent

arXiv:2506.05985v3 Announce Type: replace Abstract: A generalist agent must continuously learn and adapt throughout its lifetime, achieving efficient forward transfer while minimizing catastrophic for

Dynamics of Stochastic Momentum with Sparse Updates in High Dimensions

Model ReleasesDGX agent

arXiv:2605.28961v1 Announce Type: cross Abstract: Existing theory of momentum assumes that gradients arrive at every parameter at a roughly constant rate, an assumption violated in practice by heavy-t

← Previous
1…114115116117118…243
Next →