AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
24 Apr 2026

Even More Guarantees for Variational Inference in the Presence of Symmetries

ResearchDGX agent

arXiv:2604.21407v1 Announce Type: new Abstract: When approximating an intractable density via variational inference (VI) the variational family is typically chosen as a simple parametric family that v

FairyFuse: Multiplication-Free LLM Inference on CPUs via Fused Ternary Kernels

Model ReleasesDGX agent

arXiv:2604.20913v1 Announce Type: new Abstract: Large language models are increasingly deployed on CPU-only platforms where memory bandwidth is the primary bottleneck for autoregressive generation. We

Fine-Tuning Regimes Define Distinct Continual Learning Problems

Model ReleasesDGX agent

arXiv:2604.21927v1 Announce Type: new Abstract: Continual learning (CL) studies how models acquire tasks sequentially while retaining previously learned knowledge. Despite substantial progress in benc


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Forget, Then Recall: Learnable Compression and Selective Unfolding via Gist Sparse Attention

ResearchDGX agent

arXiv:2604.20920v1 Announce Type: new Abstract: Scaling large language models to long contexts is challenging due to the quadratic computational cost of full attention. Mitigation approaches include K

GARG-AML against Smurfing: A Scalable and Interpretable Graph-Based Framework for Anti-Money Laundering

ApplicationsDGX agent

arXiv:2506.04292v3 Announce Type: replace-cross Abstract: Purpose: We introduce GARG-AML, a fast and transparent graph-based method to catch `smurfing', a common money-laundering tactic. It assigns a

Geometric Characterisation and Structured Trajectory Surrogates for Clinical Dataset Condensation

Model ReleasesDGX agent

arXiv:2604.21638v1 Announce Type: new Abstract: Dataset condensation constructs compact synthetic datasets that retain the training utility of large real-world datasets, enabling efficient model devel

GFlowState: Visualizing the Training of Generative Flow Networks Beyond the Reward

SafetyDGX agent

arXiv:2604.21830v1 Announce Type: new Abstract: We present GFlowState, a visual analytics system designed to illuminate the training process of Generative Flow Networks (GFlowNets or GFNs). GFlowNets

Graph Neural Network-Informed Predictive Flows for Faster Ford-Fulkerson and PAC-Learnability

TutorialsDGX agent

arXiv:2604.21175v1 Announce Type: new Abstract: We propose a learning-augmented framework for accelerating max-flow computation and image segmentation by integrating Graph Neural Networks (GNNs) with

GSpaRC: Gaussian Splatting for Real-time Reconstruction of RF Channels

HardwareDGX agent

arXiv:2511.22793v2 Announce Type: replace Abstract: Channel state information (CSI) is essential for adaptive beamforming and maintaining robust links in wireless communication systems. However, acqui

H-EFT-VA: An Effective-Field-Theory Variational Ansatz with Provable Barren Plateau Avoidance

Local AiDGX agent

arXiv:2601.10479v2 Announce Type: replace-cross Abstract: Variational Quantum Algorithms (VQAs) are critically threatened by the Barren Plateau (BP) phenomenon. In this work, we introduce the H-EFT Va

Higher Order Approximation Rates for ReLU CNNs in Korobov Spaces

ResearchDGX agent

arXiv:2501.11275v2 Announce Type: replace Abstract: This paper investigates the L_p approximation error for higher order Korobov functions using deep convolutional neural networks (CNNs) with ReLU act

Hyperboloid GPLVM for Discovering Continuous Hierarchies via Nonparametric Estimation

ResearchDGX agent

arXiv:2410.16698v2 Announce Type: replace Abstract: Dimensionality reduction (DR) offers a useful representation of complex high-dimensional data. Recent DR methods focus on hyperbolic geometry to der

ICNN-enhanced 2SP: Leveraging input convex neural networks for solving two-stage stochastic programming

Model ReleasesDGX agent

arXiv:2505.05261v3 Announce Type: replace-cross Abstract: Two-stage stochastic programming (2SP) offers a basic framework for modelling decision-making under uncertainty, yet scalability remains a cha

ILDR: Geometric Early Detection of Grokking

Model ReleasesDGX agent

arXiv:2604.20923v1 Announce Type: new Abstract: Grokking describes a delayed generalization phenomenon in which a neural network achieves perfect training accuracy long before validation accuracy impr

Improving Performance in Classification Tasks with LCEN and the Weighted Focal Differentiable MCC Loss

ResearchDGX agent

arXiv:2604.21252v1 Announce Type: new Abstract: The LASSO-Clip-EN (LCEN) algorithm was previously introduced for nonlinear, interpretable feature selection and machine learning. However, its design an

Interpretable Quantile Regression by Optimal Decision Trees

ResearchDGX agent

arXiv:2604.21042v1 Announce Type: new Abstract: The field of machine learning is subject to an increasing interest in models that are not only accurate but also interpretable and robust, thus allowing

JEPAMatch: Geometric Representation Shaping for Semi-Supervised Learning

TutorialsDGX agent

arXiv:2604.21046v1 Announce Type: new Abstract: Semi-supervised learning has emerged as a powerful paradigm for leveraging large amounts of unlabeled data to improve the performance of machine learnin

KinetiDiff: Docking-Guided Diffusion for De Novo ACVR1 Inhibitor Design in Fibrodysplasia Ossificans Progressiva

ResearchDGX agent

arXiv:2604.20886v1 Announce Type: cross Abstract: We present KinetiDiff, a structure-based framework for de novo kinase inhibitor design that integrates a Geometry-Complete Diffusion Model with real-t

Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask

ResearchDGX agent

arXiv:2604.21645v1 Announce Type: new Abstract: Large-scale Nearest Neighbor (NN) search, though widely utilized in the similarity search field, remains challenged by the computational limitations inh

Learning Linear Regression with Low-Rank Tasks in-Context

TutorialsDGX agent

arXiv:2510.04548v2 Announce Type: replace-cross Abstract: In-context learning (ICL) is a key building block of modern large language models, yet its theoretical mechanisms remain poorly understood. It

Learning to Emulate Chaos: Adversarial Optimal Transport Regularization

TutorialsDGX agent

arXiv:2604.21097v1 Announce Type: cross Abstract: Chaos arises in many complex dynamical systems, from weather to power grids, but is difficult to accurately model using data-driven emulators, includi

Locating acts of mechanistic reasoning in student team conversations with mechanistic machine learning

SafetyDGX agent

arXiv:2604.21870v1 Announce Type: cross Abstract: STEM education researchers are often interested in identifying moments of students' mechanistic reasoning for deeper analysis, but have limited capaci

Low-Rank Adaptation Redux for Large Models

Model ReleasesDGX agent

arXiv:2604.21905v1 Announce Type: new Abstract: Low-rank adaptation (LoRA) has emerged as the de facto standard for parameter-efficient fine-tuning (PEFT) of foundation models, enabling the adaptation

MCAP: Deployment-Time Layer Profiling for Memory-Constrained LLM Inference

Model ReleasesDGX agent

arXiv:2604.21026v1 Announce Type: new Abstract: Deploying large language models to heterogeneous hardware is often constrained by memory, not compute. We introduce MCAP (Monte Carlo Activation Profili

Mind the Gap: Optimal and Equitable Encouragement Policies

Local AiDGX agent

arXiv:2309.07176v5 Announce Type: replace Abstract: In consequential domains, it is often impossible to compel individuals to take treatment, so that optimal policy rules are merely suggestions in the

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

Model ReleasesDGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

Neutron and X-ray Diffraction Reveal the Limits of Long-Range Machine Learning Potentials for Medium-Range Order in Silica Glass

Local AiDGX agent

arXiv:2604.21222v1 Announce Type: cross Abstract: Glassy silica is a foundational material in optics and electronics, yet accurately predicting its medium-range order (MRO) remains a major challenge f

Nonlinear Causal Discovery through a Sequential Edge Orientation Approach

ApplicationsDGX agent

arXiv:2506.05590v3 Announce Type: replace-cross Abstract: Recent advances have established the identifiability of a directed acyclic graph (DAG) under additive noise models (ANMs), spurring the develo

Not-a-Bandit: Provably No-Regret Drafter Selection in Speculative Decoding for LLMs

ResearchDGX agent

arXiv:2510.20064v2 Announce Type: replace Abstract: Speculative decoding is widely used in accelerating large language model (LLM) inference. In this work, we focus on the online draft model selection

On the algebra of Koopman eigenfunctions and on some of their infinities

ResearchDGX agent

arXiv:2604.21825v1 Announce Type: cross Abstract: For continuous-time dynamical systems with reversible trajectories, the nowhere-vanishing eigenfunctions of the Koopman operator of the system form a

Partially Lazy Gradient Descent for Smoothed Online Learning

ResearchDGX agent

arXiv:2601.15984v2 Announce Type: replace Abstract: We introduce extsc{k-lazyGD}, an online learning algorithm that bridges the gap between greedy Online Gradient Descent (OGD, for k{=}1) and lazy GD/

PDGMM-VAE: A Variational Autoencoder with Adaptive Per-Dimension Gaussian Mixture Model Priors for Nonlinear ICA

ResearchDGX agent

arXiv:2603.23547v2 Announce Type: replace-cross Abstract: Independent component analysis is a core framework within blind source separation for recovering latent source signals from observed mixtures

Post-Training Augmentation Invariance

ResearchDGX agent

arXiv:2505.11702v2 Announce Type: replace Abstract: This work develops a framework for post-training augmentation invariance, in which our goal is to add invariance properties to a pretrained network

Preconditioned DeltaNet: Curvature-aware Sequence Modeling for Linear Recurrences

TutorialsDGX agent

arXiv:2604.21100v1 Announce Type: new Abstract: To address the increasing long-context compute limitations of softmax attention, several subquadratic recurrent operators have been developed. This work

Principled Evaluation with Human Labels: One Rater at a Time and Rater Equivalence

Model ReleasesDGX agent

arXiv:2106.01254v3 Announce Type: replace Abstract: In many classification tasks, there is no definitive ground truth, only human judgments that may disagree. We address two challenges that arise in s

PrismaDV: Automated Task-Aware Data Unit Test Generation

TutorialsDGX agent

arXiv:2604.21765v1 Announce Type: new Abstract: Data is a central resource for modern enterprises, and data validation is essential for ensuring the reliability of downstream applications. However, ex

Product Quantization for Surface Soil Similarity

ResearchDGX agent

arXiv:2506.03374v2 Announce Type: replace Abstract: The use of machine learning (ML) techniques has allowed rapid advancements in many scientific and engineering fields. One of these problems is that

Refining Covariance Matrix Estimation in Stochastic Gradient Descent Through Bias Reduction

SafetyDGX agent

arXiv:2604.21203v1 Announce Type: cross Abstract: We study online inference and asymptotic covariance estimation for the stochastic gradient descent (SGD) algorithm. While classical methods (such as p

Relocation of compact sets in R^n by diffeomorphisms and linear separability of datasets in R^n

ResearchDGX agent

arXiv:2604.21393v1 Announce Type: new Abstract: Relocation of compact sets in an n-dimensional manifold by self-diffeomorphism is of its own interest as well as significant potential applications to d

Revealing Geography-Driven Signals in Zone-Level Claim Frequency Models: An Empirical Study using Environmental and Visual Predictors

Model ReleasesDGX agent

arXiv:2604.21893v1 Announce Type: cross Abstract: Geographic context is often consider relevant to motor insurance risk, yet public actuarial datasets provide limited location identifiers, constrainin

SCM: Sleep-Consolidated Memory with Algorithmic Forgetting for Large Language Models

Model ReleasesDGX agent

arXiv:2604.20943v1 Announce Type: new Abstract: We present SCM (Sleep-Consolidated Memory), a research preview of a memory architecture for large language models that draws on neuroscientific principl

SDNGuardStack: An Explainable Ensemble Learning Framework for High-Accuracy Intrusion Detection in Software-Defined Networks

ResearchDGX agent

arXiv:2604.20934v1 Announce Type: cross Abstract: Software-Defined Networking (SDN) is another technology that has been developing in the last few years as a relevant technique to improve network prog

Sink-Token-Aware Pruning for Fine-Grained Video Understanding in Efficient Video LLMs

ResearchDGX agent

arXiv:2604.20937v1 Announce Type: new Abstract: Video Large Language Models (Video LLMs) incur high inference latency due to a large number of visual tokens provided to LLMs. To address this, training

Spatio-temporal probabilistic forecast using MMAF-guided learning

ResearchDGX agent

arXiv:2603.15055v2 Announce Type: replace-cross Abstract: We present a theory-guided generalized Bayesian methodology for spatio-temporal raster data, which we use to train an ensemble of stochastic f

Spectral Embeddings Leak Graph Topology: Theory, Benchmark, and Adaptive Reconstruction

Model ReleasesDGX agent

arXiv:2604.21094v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) excel on relational data, but standard benchmarks unrealistically assume the graph is centrally available. In practice, set

Spectral Kernel Dynamics for Planetary Surface Graphs: Distinction Dynamics and Topological Conservation

ResearchDGX agent

arXiv:2604.20887v1 Announce Type: cross Abstract: The spectral kernel field equation R[k] = T[k] lacks a conservation-law analog. We prove (i) the fixed-point flow is strictly volume-expanding (tr DF

Strategic Heterogeneous Multi-Agent Architecture for Cost-Effective Code Vulnerability Detection

Model ReleasesDGX agent

arXiv:2604.21282v1 Announce Type: cross Abstract: Automated code vulnerability detection is critical for software security, yet existing approaches face a fundamental trade-off between detection accur

Tempered Sequential Monte Carlo for Trajectory and Policy Optimization with Differentiable Dynamics

SafetyDGX agent

arXiv:2604.21456v1 Announce Type: new Abstract: We propose a sampling-based framework for finite-horizon trajectory and policy optimization under differentiable dynamics by casting controller design a

Temporal Taskification in Streaming Continual Learning: A Source of Evaluation Instability

Model ReleasesDGX agent

arXiv:2604.21930v1 Announce Type: new Abstract: Streaming Continual Learning (CL) typically converts a continuous stream into a sequence of discrete tasks through temporal partitioning. We argue that

The Feedback Hamiltonian is the Score Function: A Diffusion-Model Framework for Quantum Trajectory Reversal

Model ReleasesDGX agent

arXiv:2604.21210v1 Announce Type: cross Abstract: In continuously monitored quantum systems, the feedback protocol of Garcia-Pintos, Liu, and Gorshkov reshapes the arrow of time: a Hamiltonian H_{meas

The Recurrent Transformer: Greater Effective Depth and Efficient Decoding

Model ReleasesDGX agent

arXiv:2604.21215v1 Announce Type: new Abstract: Transformers process tokens in parallel but are temporally shallow: at position t, each layer attends to key-value pairs computed based on the previous

The Sample Complexity of Multicalibration

ResearchDGX agent

arXiv:2604.21923v1 Announce Type: new Abstract: We study the minimax sample complexity of multicalibration in the batch setting. A learner observes n i.i.d. samples from an unknown distribution and mu

There Will Be a Scientific Theory of Deep Learning

ResearchDGX agent

arXiv:2604.21691v1 Announce Type: cross Abstract: In this paper, we make the case that a scientific theory of deep learning is emerging. By this we mean a theory which characterizes important properti

Toward a Multi-Layer ML-Based Security Framework for Industrial IoT

ApplicationsDGX agent

arXiv:2603.24111v3 Announce Type: replace-cross Abstract: The Industrial Internet of Things (IIoT) introduces significant security challenges as resource-constrained devices become increasingly integr

Toward Efficient Membership Inference Attacks against Federated Large Language Models: A Projection Residual Approach

Model ReleasesDGX agent

arXiv:2604.21197v1 Announce Type: new Abstract: Federated Large Language Models (FedLLMs) enable multiple parties to collaboratively fine-tune LLMs without sharing raw data, addressing challenges of l

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

SafetyDGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

Towards Universal Tabular Embeddings: A Benchmark Across Data Tasks

Model ReleasesDGX agent

arXiv:2604.21696v1 Announce Type: new Abstract: Tabular foundation models aim to learn universal representations of tabular data that transfer across tasks and domains, enabling applications such as t

Transfer Learning for Loan Recovery Prediction under Distribution Shifts with Heterogeneous Feature Spaces

ApplicationsDGX agent

arXiv:2604.02832v2 Announce Type: replace-cross Abstract: Accurate forecasting of recovery rates (RR) is central to credit risk management and regulatory capital determination. In many loan portfolios

Transferable Physics-Informed Representations via Closed-Form Head Adaptation

TutorialsDGX agent

arXiv:2604.21761v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have garnered significant interest for their potential in solving partial differential equations (PDEs) that go

Transferable SCF-Acceleration through Solver-Aligned Initialization Learning

ResearchDGX agent

arXiv:2604.21657v1 Announce Type: new Abstract: Machine learning methods that predict initial guesses from molecular geometry can reduce this cost, but matrix-prediction models fail when extrapolating

← Previous
1…208209210211212…241
Next →