AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
12 May 2026

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

SafetyDGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

Diagnosing and Mitigating Domain Shift in Permission-Based Android Malware Detection

ApplicationsDGX agent

arXiv:2605.09028v1 Announce Type: new Abstract: Machine learning-based Android malware detectors often fail in real-world deployment due to domain shift, where models trained on one data source perfor

DiffATS: Diffusion in Aligned Tensor Space

Model ReleasesDGX agent

arXiv:2605.09275v1 Announce Type: new Abstract: Direct diffusion modeling of high-resolution spatiotemporal fields is computationally challenging. Parameter-efficient primitives address this by repres


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Different Prompts, Different Ranks: Prompt-aware Dynamic Rank Selection for SVD-based LLM Compression

Model ReleasesDGX agent

arXiv:2605.08568v1 Announce Type: new Abstract: Large language models (LLMs) have rapidly grown in scale, creating substantial memory and computational costs that hinder efficient deployment. Singular

Differentially Private Hyperparameter Tuning using Local Bayesian Optimization

ResearchDGX agent

arXiv:2502.06044v3 Announce Type: replace-cross Abstract: Hyperparameter tuning is a key component of machine learning procedures, but when validation data contain sensitive user information, search m

Differentially Private Sampling from Distributions via Wasserstein Projection

ResearchDGX agent

arXiv:2605.10015v1 Announce Type: cross Abstract: In this paper, we study the problem of sampling from a distribution under the constraint of differential privacy (DP). Prior works measure the utility

Differentially Private Spectral Graph Clustering: Balancing Privacy, Accuracy, and Efficiency

ApplicationsDGX agent

arXiv:2510.07136v2 Announce Type: replace-cross Abstract: We study spectral graph clustering under edge differential privacy. We propose a matrix shuffling mechanism that combines randomized edge flip

Diffusion Models are Evolutionary Algorithms

Model ReleasesDGX agent

arXiv:2410.02543v3 Announce Type: replace-cross Abstract: In a convergence of machine learning and biology, we reveal that diffusion models are evolutionary algorithms. By considering evolution as a d

Dimension-Free Saddle-Point Escape in Muon

ResearchDGX agent

arXiv:2605.09331v1 Announce Type: new Abstract: Modern Large Language Model (LLM) training is fundamentally bottlenecked by pathologically flat saddle points in extreme high-dimensional landscapes. Mo

Direct Bethe Free Energy Minimization for Bayesian Neural Ne twork

Model ReleasesDGX agent

arXiv:2605.08446v1 Announce Type: new Abstract: We propose training Bayesian neural networks by directly minimizing the Bethe free energy rather than maximizing a variational lower bound. On tree-stru

Direct From Darwin: Deriving Advanced Optimizers From Evolutionary First Principles

ResearchDGX agent

arXiv:2605.05284v2 Announce Type: replace-cross Abstract: Evolutionary computation has long promised to deliver both high-performance optimization tools as well as rigorous scientific simulations of D

Discovery of Nonlinear Dynamics with Automated Basis Function Generation

ResearchDGX agent

arXiv:2605.09696v1 Announce Type: new Abstract: Discovering governing equations from observational data remains a fundamental challenge in scientific modeling, particularly when the underlying mathema

Discrete Double-Bracket Flows for Isotropic-Noise Invariant Eigendecomposition

ResearchDGX agent

arXiv:2602.13759v2 Announce Type: replace Abstract: We study eigendecomposition on SO(n) under streaming observations C_k = C_{sig} + sigma_k^2 I + E_k, where the isotropic background sigma_k^2 I may

Discrete Flow Matching: Convergence Guarantees Under Minimal Assumptions

ResearchDGX agent

arXiv:2605.08882v1 Announce Type: new Abstract: Flow Matching has recently emerged as a popular class of generative models for simulating a target distribution mu_1 from samples drawn from a source di

Disentangled Representation Learning via Flow Matching

SafetyDGX agent

arXiv:2602.05214v2 Announce Type: replace Abstract: Disentangled representation learning aims to capture the underlying explanatory factors of observed data, enabling a principled understanding of the

Dissecting Jet-Tagger Through Mechanistic Interpretability

ResearchDGX agent

arXiv:2605.09881v1 Announce Type: cross Abstract: Mechanistic interpretability seeks to reverse engineer a trained neural network by identifying the minimal subset of internal components. We perform a

Distributional Reinforcement Learning via the Cramer Distance

ResearchDGX agent

arXiv:2605.08104v1 Announce Type: new Abstract: This paper explores the application of the Soft Actor-Critic (SAC) algorithm within a Distributional Reinforcement Learning setting and introduces an im

Distributional Spectral Diagnostics for Localizing Grokking Transitions

Model ReleasesDGX agent

arXiv:2605.08237v1 Announce Type: new Abstract: In grokking, a model first fits the training data while test accuracy remains low, and only later begins to generalize. We ask whether this transition c

Domain-Adaptive Arrhythmia Classification Using a Hybrid Transformer on Wearable Heart Signals

ResearchDGX agent

arXiv:2605.08199v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of death globally, underscoring the need for effective, accessible monitoring solutions, particularly

Don't Fix the Basis -- Learn It: Spectral Representation with Adaptive Basis Learning for PDEs

TutorialsDGX agent

arXiv:2605.10451v1 Announce Type: new Abstract: Spectral neural operators achieve strong performance for PDE learning, but rely on fixed global bases that limit their ability to represent spatially he

Doubly Robust Proxy Causal Learning with Neural Mean Embeddings

ResearchDGX agent

arXiv:2605.09514v1 Announce Type: new Abstract: Unobserved confounding prevents standard covariate adjustment from identifying causal response functions in observational studies. Proxy causal learning

DRIFT: Drift-Resilient Invariant-Feature Transformer for DGA Detection

TutorialsDGX agent

arXiv:2605.10436v1 Announce Type: cross Abstract: Domain Generation Algorithms (DGAs) evolve continuously to evade botnet detection, posing a persistent challenge for dependable network defense. While

DynaMiCS: Fine-tuning LLMs with Performance Constraints using Dynamic Mixtures

SafetyDGX agent

arXiv:2605.10770v1 Announce Type: new Abstract: Multi-domain fine-tuning of large language models requires improving performance on target domains while preserving performance on constrained domains,

Dystruct: Dynamically Structured Diffusion Language Model Decoding via Bayesian Inference

Local AiDGX agent

arXiv:2605.09820v1 Announce Type: new Abstract: Diffusion language models (DLMs) have recently emerged as a promising alternative to autoregressive models, primarily due to their ability to enable par

EchoAlign: Bridging Generative and Discriminative Learning under Noisy Labels

Model ReleasesDGX agent

arXiv:2405.12969v3 Announce Type: replace Abstract: Noisy labels severely hinder the accuracy and generalization of machine learning models, especially when ambiguous instance features make reliable a

Edge-specific signal propagation on mature chromophore-region 3D mechanism graphs for fluorescent protein quantum-yield prediction

Model ReleasesDGX agent

arXiv:2605.06644v2 Announce Type: replace Abstract: Fluorescent protein quantum yield (QY) is governed by the mature chromophore and its three-dimensional microenvironment rather than sequence identit

Efficient Evaluation of LLM Performance with Statistical Guarantees

Model ReleasesDGX agent

arXiv:2601.20251v3 Announce Type: replace-cross Abstract: Exhaustively evaluating many large language models (LLMs) on a large suite of benchmarks is expensive. We cast benchmarking as finite-populati

Efficient Neural Architectures for Real-Time ECG Interpretation on Limited Hardware

Model ReleasesDGX agent

arXiv:2605.09848v1 Announce Type: new Abstract: Electrocardiogram (ECG) interpretation is essential for diagnosing a wide range of cardiac abnormalities. While deep learning has shown strong potential

Efficient Statistics With Unknown Truncation, Polynomial Time Algorithms, Beyond Gaussians

ResearchDGX agent

arXiv:2410.01656v2 Announce Type: replace-cross Abstract: We study the estimation of distributional parameters when samples are shown only if they fall in some unknown set S subseteq R^d. Kontonis, Tz

Elucidating Representation Degradation Problem in Diffusion Model Training

ResearchDGX agent

arXiv:2605.10790v1 Announce Type: new Abstract: Diffusion models have achieved remarkable success, yet their training remains inefficient due to a severe optimization bottleneck, which we term Represe

Embedding Dimension Lower Bounds for Universality of Deep Sets and Janossy Pooling

ResearchDGX agent

arXiv:2605.08377v1 Announce Type: new Abstract: In many practical applications it is important to build symmetries into neural network architectures. Consider the important case of permutation symmetr

Empirical Bayes 1-bit matrix completion

ResearchDGX agent

arXiv:2605.09509v1 Announce Type: cross Abstract: The problem of predicting unobserved entries in a binary matrix, known as 1-bit matrix completion, has found diverse applications in fields such as re

Enabling Structure-Only Initialization and Out-of-Distribution Generalization in GNN-based Molecular Dynamics Simulators

ResearchDGX agent

arXiv:2605.09495v1 Announce Type: cross Abstract: Machine learning-based simulators offer the potential to model the dynamics of complex systems more efficiently than classical approaches, while retai

End-to-End Keyword Spotting on FPGA Using Graph Neural Networks with a Neuromorphic Auditory Sensor

Local AiDGX agent

arXiv:2605.09570v1 Announce Type: new Abstract: With the rapid growth of mobile robotics and embedded intelligence, there is an increasing demand for efficient on-device data processing on edge platfo

Energy-based models for diagnostic reconstruction and analysis in a laboratory plasma device

ResearchDGX agent

arXiv:2605.08645v1 Announce Type: cross Abstract: Energy-based models (EBMs) provide a powerful and flexible way of learning a joint probability distribution over data by constructing an energy surfac

Enhancing Adversarial Robustness in Network Intrusion Detection: A Layer-wise Adaptive Regularization Approach

ResearchDGX agent

arXiv:2605.08910v1 Announce Type: cross Abstract: The new wave of adversarial attacks that utilize gradient-related vulnerabilities in neural network-based classifiers makes Network Intrusion Detectio

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models

Model ReleasesDGX agent

arXiv:2605.10410v1 Announce Type: new Abstract: Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

SafetyDGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

ERIS: Enhancing Privacy and Scalability in Federated Learning via Federated Shard Aggregation

Model ReleasesDGX agent

arXiv:2602.08617v2 Announce Type: replace Abstract: Scaling Federated Learning (FL) to billion-parameter models forces a challenging trade-off between privacy, scalability, and model utility. Existing

Estimating Heterogeneous Causal Effect on Networks via Orthogonal Learning

ResearchDGX agent

arXiv:2509.18484v2 Announce Type: replace-cross Abstract: Estimating causal effects on networks is challenging because treatments may affect both treated units and their neighbors, while network homop

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

SafetyDGX agent

arXiv:2601.21484v2 Announce Type: replace Abstract: Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complic

Evaluating Federated Learning approaches for mammography under breast density heterogeneity

Local AiDGX agent

arXiv:2605.09137v1 Announce Type: new Abstract: Breast density is a key factor that influences mammography interpretation and is a major source of heterogeneity in multicenter datasets. Such heterogen

Exact Fixed-Point Constraints in Neural-ODEs with Provable Universality

ResearchDGX agent

arXiv:2605.10613v1 Announce Type: cross Abstract: We introduce a technique that enables Neural-ODEs to approximate arbitrary velocity fields with a priori planted fixed-points. Specifically, a recipe

Exact Unlearning from Proxies Induces Closeness Guarantees on Approximate Unlearning

ResearchDGX agent

arXiv:2605.10680v1 Announce Type: new Abstract: This paper proposes a paradigm shift linking machine unlearning directly to the structure of the data distributions rather than a mere update of the neu

Exactness Matters for Physical Rule Enforcement

Model ReleasesDGX agent

arXiv:2605.08285v1 Announce Type: new Abstract: Autoregressive scientific forecasters often enforce physical or structural constraints by repairing each predicted state before feeding it back into the

ExecuTorch -- A Unified PyTorch Solution to Run AI Models On-Device

Local AiDGX agent

arXiv:2605.08195v1 Announce Type: new Abstract: Local execution of AI on edge devices is important for low latency and offline operation. However, deploying models on diverse hardware remains fragment

Explicit and Effectively Symmetric Schemes for Neural SDEs on Lie Groups

ResearchDGX agent

arXiv:2509.20599v2 Announce Type: replace Abstract: Backpropagation through (neural) SDE solvers is traditionally approached in two ways: discretise-then-optimise, which offers accurate gradients but

Exploration-Driven Optimization for Test-Time Large Language Model Reasoning

SafetyDGX agent

arXiv:2605.09853v1 Announce Type: new Abstract: Post-training techniques combined with inference-time scaling significantly enhance the reasoning and alignment capabilities of large language models (L

Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality

SafetyDGX agent

arXiv:2605.10206v1 Announce Type: cross Abstract: Distributional causal inference requires estimating not only average treatment effects but also interventional outcome distributions, including quanti

f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment

SafetyDGX agent

arXiv:2602.05946v3 Announce Type: replace Abstract: Recent work shows that preference alignment objectives can be interpreted as divergence estimators between aligned (preferred) & unaligned (less-pre

Factual recall in linear associative memories: sharp asymptotics and mechanistic insights

ResearchDGX agent

arXiv:2605.10795v1 Announce Type: cross Abstract: Large language models demonstrate remarkable ability in factual recall, yet the fundamental limits of storing and retrieving input--output association

Fast Training of Mixture-of-Experts for Time Series Forecasting via Expert Loss Integration

ResearchDGX agent

arXiv:2605.10330v1 Announce Type: cross Abstract: We propose a novel adaptive Mixture-of-Experts (MoE) framework for time series forecasting that enhances expert specialization by incorporating expert

Feature Augmentation of GNNs for ILPs: Local Uniqueness Suffices

ApplicationsDGX agent

arXiv:2509.21000v2 Announce Type: replace Abstract: Integer Linear Programs (ILPs) are central to real-world optimizations but notoriously difficult to solve. Learning to Optimize (L2O) has emerged as

FedAdaVR: Adaptive Variance Reduction for Robust Federated Learning under Limited Client Participation

ResearchDGX agent

arXiv:2601.22204v2 Announce Type: replace Abstract: Federated learning (FL) encounters substantial challenges due to heterogeneity, leading to gradient noise, client drift, and partial client particip

FedCIGAR: A Personalized Reconstruction Approach for Federated Graph-level Anomaly Detection

ResearchDGX agent

arXiv:2605.09428v1 Announce Type: new Abstract: Graph-level anomaly detection (GLAD) is crucial for ensuring the reliability of graph-driven applications by identifying abnormal graphs that deviate fr

Federated Concept-Based Models: Interpretable models with distributed supervision

ResearchDGX agent

arXiv:2602.04093v2 Announce Type: replace Abstract: Concept-based Models (CMs) enhance interpretability in deep learning by grounding predictions in human-understandable concepts. However, concept ann

FedGMI: Generative Model-Driven Federated Learning for Probabilistic Mixture Inference

Local AiDGX agent

arXiv:2605.08760v1 Announce Type: new Abstract: Federated Learning (FL) facilitates collaborative model training across decentralized clients while preserving data privacy by avoiding raw data exchang

FedVSSAM: Mitigating Flatness Incompatibility in Sharpness-Aware Federated Learning

Local AiDGX agent

arXiv:2605.09144v1 Announce Type: new Abstract: Sharpness-aware minimization (SAM) is an effective method for improving the generalization of federated learning (FL) by steering local training toward

Finding Connections: Membership Inference Attacks for the Multi-Table Synthetic Data Setting

ApplicationsDGX agent

arXiv:2602.07126v2 Announce Type: replace Abstract: Synthetic tabular data has gained attention for enabling privacy-preserving data sharing. While substantial progress has been made in single-table s

Finer is Better (with the Right Scaling)

ResearchDGX agent

arXiv:2605.08565v1 Announce Type: new Abstract: Microscaling is a critical technique for preserving the quality of Large Language Models (LLMs) quantized to ultra-low precision formats. Intuitively, f

← Previous
1…169170171172173…243
Next →