AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
5 May 2026

Gaussian Approximation and Multiplier Bootstrap for Stochastic Gradient Descent

ResearchDGX agent

arXiv:2502.06719v3 Announce Type: replace-cross Abstract: In this paper, we establish the non-asymptotic validity of the multiplier bootstrap procedure for constructing the confidence sets using the S

General Frameworks for Conditional Two-Sample Testing

SafetyDGX agent

arXiv:2410.16636v2 Announce Type: replace-cross Abstract: We study the problem of conditional two-sample testing, which aims to determine whether two populations have the same distribution after accou

Generalized Distributional Alignment Games for Unbiased Answer-Level Fine-Tuning

SafetyDGX agent

arXiv:2605.02435v1 Announce Type: new Abstract: The Distributional Alignment Game framework provides a powerful variational perspective on Answer-Level Fine-Tuning (ALFT). However, standard algorithms


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Geometric and Spectral Alignment for Deep Neural Network I

SafetyDGX agent

arXiv:2605.02108v1 Announce Type: new Abstract: Deep residual architectures are modeled as products of near-identity Jacobians. This paper proves deterministic quotient-geometric estimates for singula

Geometric and Spectral Alignment for Deep Neural Network II

SafetyDGX agent

arXiv:2605.02111v1 Announce Type: new Abstract: This paper develops the angular and static-channel component of Geometric and Spectral Alignment for residual Jacobian chains. Starting from Cartan-coor

Geospatial foundation-model embeddings improve population estimation unevenly across space and scale

Model ReleasesDGX agent

arXiv:2605.01650v1 Announce Type: new Abstract: Reliable subnational population estimates are essential for applications, yet remain difficult where censuses are sparse, outdated or spatially coarse.

GETA-3DGS: Automatic Joint Structured Pruning and Quantization for 3D Gaussian Splatting

SafetyDGX agent

arXiv:2605.02086v1 Announce Type: new Abstract: 3D Gaussian splatting (3DGS) is a state-of-the-art representation for real-time photorealistic novel-view synthesis, yet a single high-fidelity scene ty

Gradient Boosted Risk Scores

ResearchDGX agent

arXiv:2605.02593v1 Announce Type: new Abstract: Risk scores are an interpretable and actionable class of machine learning models with applications in medicine, insurance, and risk management. Unlike m

Gradient Boosting within a Single Attention Layer

Model ReleasesDGX agent

arXiv:2604.03190v2 Announce Type: replace Abstract: Transformer attention computes a single softmax-weighted average over values -- a one-pass estimate that cannot correct its own errors. We introduce

Gradient-Discrepancy Acquisition for Pool-Based Active Learning

ResearchDGX agent

arXiv:2605.02609v1 Announce Type: new Abstract: The effectiveness of active learning hinges on the choice of the acquisition criterion by which a learning algorithm selects potentially informative dat

Gradient-Gated DPO: Stabilizing Preference Optimization in Language Models

SafetyDGX agent

arXiv:2605.02626v1 Announce Type: new Abstract: Preference optimization has become a central paradigm for aligning large language models with human feedback. Direct Preference Optimization (DPO) simpl

Graph Federated Unlearning for Privacy Preservation

Local AiDGX agent

arXiv:2605.02297v1 Announce Type: new Abstract: Graph federated learning (GFL) facilitates decentralized training on distributed graph data while keeping sensitive user information local, aligning wit

GraphLand: Evaluating Graph Machine Learning Models on Diverse Industrial Data

Model ReleasesDGX agent

arXiv:2409.14500v5 Announce Type: replace Abstract: Although data that can be naturally represented as graphs is widespread in real-world applications across diverse industries, popular graph ML bench

GraphSculptor: Sculpting Pre-training Coreset for Graph Self-supervised Learning

ResearchDGX agent

arXiv:2605.01310v1 Announce Type: new Abstract: Graph self-supervised learning typically relies on large-scale unlabeled datasets, heavily inflating computational costs. However, empirical evidence su

Green Energy Management for Sustainable Data Centers Using Deep Reinforcement Learning

SafetyDGX agent

arXiv:2507.21153v2 Announce Type: replace Abstract: The exponential growth of digital services has positioned data centers among the most energy-intensive infrastructures in the modern economy, raisin

H3: A Healthcare Three-Hop Index for Physician Referral Network Prediction

ApplicationsDGX agent

arXiv:2605.02150v1 Announce Type: cross Abstract: Accurate prediction of physician referral links is essential for optimizing care coordination and reducing fragmentation in healthcare delivery. Howev

Hall-Like Transversal Stress and Sandpile Criticality on Real Production Networks

ApplicationsDGX agent

arXiv:2605.01561v1 Announce Type: cross Abstract: This paper develops a Hall-Sandpile model of economic instability that combines a Hall-like transversal stress mechanism with sandpile threshold dynam

HARMES: A Multi-Modal Dataset for Wearable Human Activity Recognition with Motion, Environmental Sensing and Sound

Model ReleasesDGX agent

arXiv:2605.02596v1 Announce Type: new Abstract: With each sensing modality exhibiting inherent strengths and limitations, multi-modal approaches for wearable Human Activity Recognition (HAR) are becom

Harnessing Reasoning Trajectories for Hallucination Detection via Answer-agreement Representation Shaping

ResearchDGX agent

arXiv:2601.17467v2 Announce Type: replace Abstract: Large reasoning models (LRMs) often generate long, seemingly coherent reasoning traces yet still produce incorrect answers, making hallucination det

Heavy-Tailed Principal Component Analysis

ResearchDGX agent

arXiv:2603.11308v2 Announce Type: replace Abstract: Principal Component Analysis (PCA) is a cornerstone of dimensionality reduction, yet its classical formulation relies critically on second-order mom

HELIX: Hybrid Encoding with Learnable Identity and Cross-dimensional Synthesis for Time Series Imputation

ResearchDGX agent

arXiv:2605.02278v1 Announce Type: new Abstract: Time series imputation benefits from leveraging cross-feature correlations, yet existing attention-based methods re-discover feature relationships at ea

Hierarchical Federated Learning for Networked AI: From Communication Saving to Architecture-Aware Design

Local AiDGX agent

arXiv:2605.00931v1 Announce Type: new Abstract: Federated learning (FL) is fundamentally a distributed optimization problem executed by communicating agents with local data, local computation, and par

High entropy leads to symmetry equivariant policies in Dec-POMDPs

SafetyDGX agent

arXiv:2511.22581v3 Announce Type: replace Abstract: We prove that in any Dec-POMDP, sufficiently high entropy regularization ensures that the policy gradient flow with tabular softmax parametrization

How Label Imbalance Shapes Geometry: A General Spectral Analysis of Multi-Label Neural Collapse

ResearchDGX agent

arXiv:2605.01897v1 Announce Type: new Abstract: This work investigates the phenomenon of Neural Collapse (NC) in multi-label classification, extending its conceptual framework from multi-class learnin

How Reasoning Evolves from Post-Training Data: An Empirical Study Using Chess

Model ReleasesDGX agent

arXiv:2604.05134v2 Announce Type: replace Abstract: We study how reasoning evolves in a language model -- from supervised fine-tuning (SFT) to reinforcement learning (RL) -- by analyzing how a set of

Hybrid Quantum Reinforcement Learning with QAOA for Improved Vehicle Routing Optimization

SafetyDGX agent

arXiv:2605.01574v1 Announce Type: new Abstract: Vehicle Routing Problem (VRP) is one of the most complex NP-hard combinatorial optimization problem in transportation and logistics that requires a dyna

Importance-Guided Basis Selection for Low-Rank Decomposition of Large Language Models

Model ReleasesDGX agent

arXiv:2605.01627v1 Announce Type: new Abstract: Low-rank decomposition is a compelling approach for compressing large language models, but its effectiveness hinges on selecting which singular-vector b

Inducing Permutation Invariant Priors in Bayesian Optimization for Carbon Capture and Storage Applications

TutorialsDGX agent

arXiv:2605.02409v1 Announce Type: new Abstract: Bayesian Optimization is an iterative method, tailored to optimizing expensive black box objective functions. Surrogate models like Gaussian Processes,

Interpretable experiential learning based on state history and global feedback

Model ReleasesDGX agent

arXiv:2605.00940v1 Announce Type: new Abstract: A new interpretable experiential learning model based on state history and global feedback is presented. It is capable of learning a behavioral model re

Is there 'Secret Sauce'' in Large Language Model Development?

Model ReleasesDGX agent

arXiv:2602.07238v2 Announce Type: replace-cross Abstract: Do leading LLM developers possess a proprietary ``secret sauce'', or is LLM performance driven by scaling up compute? Using training and bench

Isotropic Fourier Neural Operators

TutorialsDGX agent

arXiv:2605.02597v1 Announce Type: new Abstract: Fourier Neural Operators are deep learning models that learn mappings between function spaces and can be used to learn and solve partial differential eq

KANs need curvature: penalties for compositional smoothness

ResearchDGX agent

arXiv:2605.02190v1 Announce Type: new Abstract: Kolmogorov-Arnold networks (KANs) offer a potent combination of accuracy and interpretability, thanks to their compositions of learnable univariate acti

Kernel Treatment Effects with Adaptively Collected Data

ResearchDGX agent

arXiv:2510.10245v2 Announce Type: replace-cross Abstract: Adaptive experiments improve efficiency by adjusting treatment assignments based on past outcomes, but this adaptivity breaks the i.i.d. assum

Large margin classifier with graph-based adaptive regularization

ResearchDGX agent

arXiv:2605.02027v1 Announce Type: new Abstract: This paper introduces the use of per-class regularization hyperparameters in Gabriel graph-based binary classifiers. We demonstrate how the quality inde

Learning a Stochastic Differential Equation Model of Tropical Cyclone Intensification from Reanalysis and Observational Data

ResearchDGX agent

arXiv:2601.08116v2 Announce Type: replace Abstract: Tropical cyclones are dangerous natural hazards, but their hazard is challenging to quantify directly from historical datasets due to limited datase

Learning Discriminators for Resampling in the Ensemble Gaussian Mixture Filter through a Normalizing Flow Approach

ResearchDGX agent

arXiv:2605.01089v1 Announce Type: new Abstract: The ensemble Gaussian mixture filter (EnGMF) is a powerful, convergent particle filter capable of medium-to-high dimensional non-linear filtering. The E

Learning in the Fisher Subspace: A Guided Initialization for LoRA Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.01046v1 Announce Type: new Abstract: LoRA adapts large language models (LLMs) by restricting updates to low-rank subspaces of pre-trained weights. While this substantially reduces training

Learning Koopman operators for coupled systems via information on governing equations of subsystems

TutorialsDGX agent

arXiv:2605.01835v1 Announce Type: new Abstract: Nonlinear coupled systems are ubiquitous in science and engineering. The analysis and modeling of such systems is challenging due to their high dimensio

Learning to Race in Minutes: Infoprop Dyna on the Mini Wheelbot

TutorialsDGX agent

arXiv:2605.01096v1 Announce Type: new Abstract: Reinforcement Learning (RL) has the potential to enable robots with fast, nonlinear, and unstable dynamics to reach the limits of their performance. How

Leveraging Data Symmetries to Select an Optimal Subset of Training Data under Label Noise

ApplicationsDGX agent

arXiv:2605.01874v1 Announce Type: new Abstract: The performance of machine learning models often relies on large labeled datasets; however, data collected from diverse sources can contain label noise.

Leveraging Ensemble-Based Semi-Supervised Learning for Illicit Account Detection in Ethereum DeFi Transactions

ApplicationsDGX agent

arXiv:2412.02408v3 Announce Type: replace-cross Abstract: The advent of smart contracts has enabled the rapid rise of Decentralized Finance (DeFi) on the Ethereum blockchain, offering substantial rewa

Ligandformer: A Graph Neural Network for Predicting Compound Property with Robust Interpretation

ResearchDGX agent

arXiv:2202.10873v4 Announce Type: replace-cross Abstract: Robust and efficient interpretation of QSAR methods is quite useful to validate AI prediction rationales with subjective opinion (chemist or b

Linear-Readout Floors and Threshold Recovery in Computation in Superposition

ResearchDGX agent

arXiv:2605.01192v1 Announce Type: new Abstract: Two recent approaches to computation in superposition reach different recursive capacity regimes: Hanni et al. certify ilde{O}(d^{3/2}) computable featu

LittleBit-2: Maximizing the Spectral Energy Gain in Sub-1-Bit LLMs via Latent Geometry Alignment

Model ReleasesDGX agent

arXiv:2603.00042v2 Announce Type: replace Abstract: We identify the Spectral Energy Gain in extreme model compression, where low-rank binary approximations outperform tiny-rank floating-point baseline

LLM-VA: Resolving the Jailbreak-Overrefusal Trade-off via Vector Alignment

SafetyDGX agent

arXiv:2601.19487v2 Announce Type: replace Abstract: Safety-aligned LLMs suffer from two failure modes: jailbreak (answering harmful inputs) and over-refusal (declining benign queries). Existing vector

Local Hessian Spectral Filtering for Robust Intrinsic Dimension Estimation

ResearchDGX agent

arXiv:2605.01221v1 Announce Type: new Abstract: While diffusion models enable new approaches for estimating Local Intrinsic Dimension (LID), existing methods fail in high-dimensional spaces where nois

Low-rank surrogate modeling and stochastic zero-order optimization for training of neural networks with black-box layers

Local AiDGX agent

arXiv:2509.15113v2 Announce Type: replace Abstract: The growing demand for energy-efficient, high-performance AI systems has led to increased attention on alternative computing platforms (e.g., photon

LUMINA: A Grid Foundation Model for Benchmarking AC Optimal Power Flow Surrogate Learning

Model ReleasesDGX agent

arXiv:2605.02133v1 Announce Type: new Abstract: AC optimal power flow (ACOPF) is foundational yet computationally expensive in power grid operations, driving learning-based surrogates for large-scale

Machine Learning as Iterated Belief Change a la Darwiche and Pearl

ApplicationsDGX agent

arXiv:2506.13157v3 Announce Type: replace-cross Abstract: Artificial Neural Networks (ANNs) are powerful machine-learning models capable of capturing intricate non-linear relationships. They are widel

Machine Learning-Augmented Acceleration of Iterative Ptychographic Reconstruction

ApplicationsDGX agent

arXiv:2605.01122v1 Announce Type: new Abstract: Iterative ptychographic reconstruction algorithms are widely used for coherent diffractive imaging but can exhibit slow convergence under realistic expe

Machine Learning Enhanced Laser Spectroscopy for Multi-Species Gas Detection in Complex and Harsh Environments

SafetyDGX agent

arXiv:2605.01306v1 Announce Type: cross Abstract: Laser absorption spectroscopy (LAS) is a well-established technique for non-intrusive measurement of gas species in combustion and atmospheric environ

MAGIC: Multi-Step Advantage-Gated Causal Influence for Multi-agent Reinforcement Learning

AgentsDGX agent

arXiv:2605.01805v1 Announce Type: cross Abstract: A key challenge in multi-agent reinforcement learning (MARL) lies in designing learning signals that effectively promote coordination among agents. De

Manifold-Constrained Adversarial Training for Long-Tailed Robustness via Geometric Alignment

SafetyDGX agent

arXiv:2605.02183v1 Announce Type: new Abstract: Adversarial training is effective on balanced datasets, but its robustness degrades under longtailed class distributions, where tail classes suffer high

Mean Testing under Truncation beyond Gaussian

SafetyDGX agent

arXiv:2605.01335v1 Announce Type: cross Abstract: We characterize the fundamental limits of high-dimensional mean testing under arbitrary truncation, where samples are drawn from the conditional distr

Measuring Differences between Conditional Distributions using Kernel Embeddings

ResearchDGX agent

arXiv:2605.02260v1 Announce Type: cross Abstract: Comparing conditional distributions is a fundamental challenge in statistics and machine learning, with applications across a wide range of domains. W

Mesh Based Simulations with Spatial and Temporal awareness

ResearchDGX agent

arXiv:2605.01542v1 Announce Type: new Abstract: Machine Learning surrogates for Computational Fluid Dynamics (CFD), particularly Graph Neural Networks (GNNs) and Transformers, have become a new import

Meta-learning Structure-Preserving Dynamics

Model ReleasesDGX agent

arXiv:2508.11205v2 Announce Type: replace Abstract: Structure-preserving approaches to dynamics discovery have demonstrated great potential for modeling physical systems due to their use of strong ind

Metric-Normalized Posterior Leakage (mPL): Attacker-Aligned Privacy for Joint Consumption

Local AiDGX agent

arXiv:2605.01137v1 Announce Type: new Abstract: Metric differential privacy (mDP) strengthens local differential privacy (LDP) by scaling noise to semantic distance, but many machine learning (ML) sys

Middle-mile logistics through the lens of goal-conditioned reinforcement learning

ResearchDGX agent

arXiv:2605.02461v1 Announce Type: cross Abstract: Middle-mile logistics describes the problem of routing parcels through a network of hubs linked by trucks with finite capacity. We rephrase this as a

Minimizing Collateral Damage in Activation Steering

SafetyDGX agent

arXiv:2605.01167v1 Announce Type: new Abstract: Activation steering is a method for controlling Large Language Model (LLM) behavior by intervening in its internal representations to increase the align

← Previous
1…189190191192193…241
Next →