AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,557 results
27 May 2026

Extra-Merge: Tracing the Rank-1 Subspace of Model Merging in Language Model Pre-Training

Model ReleasesDGX agent

arXiv:2605.26484v1 Announce Type: new Abstract: Model merging has emerged as a lightweight paradigm for enhancing Large Language Models (LLMs), yet its underlying mechanisms remain poorly understood.

Flow Matching Policy Optimization with Mirror Descent and Entropy Constraints

SafetyDGX agent

arXiv:2603.17685v3 Announce Type: replace Abstract: Balancing policy expressiveness with the exploration-exploitation trade-off is a core challenge in online Reinforcement Learning (RL). While Stochas

FluxNet: Learning Capacity-Constrained Local Transport Operators for Conservative and Bounded PDE Surrogates

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.01941v2 Announce Type: replace-cross Abstract: Autoregressive learning of time-stepping operators provides an effective approach to data-driven partial differential equation (PDE) simulatio

FM-fMRI: Event Conditioned Flow Matching for Rest-to-Task fMRI Time-Series Synthesis

SafetyDGX agent

arXiv:2605.26423v1 Announce Type: new Abstract: Task-based fMRI provides a direct readout of task-evoked neural dynamics, but it is expensive and difficult to acquire at scale, motivating rest-to-task

Focal Reward: Balanced Reinforcement Learning under Rubric-Based Rewards

Model ReleasesDGX agent

arXiv:2605.26579v1 Announce Type: new Abstract: The open-ended generation in LLMs usually requires multi-dimensional rubrics to adequately assess quality and guide the improvement of reinforcement lea

From Privacy to Generalization: Linear Max-Information Bounds for DP-SGD

ResearchDGX agent

arXiv:2605.26222v1 Announce Type: new Abstract: Understanding the relationship between generalization and privacy remains a central challenge in modern machine learning theory, particularly for deep n

From Scores to Gibbs Correctors: Accelerating Uniform-Rate Discrete Diffusion Models

ResearchDGX agent

arXiv:2605.27352v1 Announce Type: new Abstract: Discrete diffusion models have achieved strong empirical performance in text and other symbolic domains, but, especially for uniform-rate models, they o

Function-Valued Causal Influence in Nonlinear Time Series

ApplicationsDGX agent

arXiv:2605.26408v1 Announce Type: new Abstract: Causal discovery in time series is increasingly performed using nonlinear machine-learning models, yet the resulting causal relationships are almost alw

Gaussian Process-based learning with new MCMC-based implementation of Wishart prior on correlation matrix

ApplicationsDGX agent

arXiv:2605.27093v1 Announce Type: cross Abstract: In probabilstic supervised learning of an input-output relationship - as a sample function of a Gaussian Process (GP) - priors are typically specified

Generalist Graph Anomaly Detection via Prototype-Based Distillation

SafetyDGX agent

arXiv:2605.26857v1 Announce Type: new Abstract: Driven by the pressing demand for graph anomaly detection (GAD) in high-stakes domains, the generalist GAD paradigm, which trains a single detector tran

Generating realistic global precipitation fields from modelled atmospheric circulation

ResearchDGX agent

arXiv:2504.00307v2 Announce Type: replace Abstract: Improving the representation of precipitation in Earth system models (ESMs) is critical for assessing the impacts of climate change and especially o

Greening AI Inference with Accuracy and Latency-aware User Incentives

ResearchDGX agent

arXiv:2605.27309v1 Announce Type: new Abstract: The widespread use of AI services has raised concerns for its environmental sustainability, towards which recent studies have identified carbon emission

Harmonia: Enhancing Data Placement and Migration in Hybrid Storage Systems via Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2503.20507v4 Announce Type: replace-cross Abstract: Modern high-performance computing (HPC) environments rely on hybrid storage systems (HSS) that combine multiple storage devices with diverse l

Image Feature Fusion-based Federated Client Unlearning (FCU)

ResearchDGX agent

arXiv:2605.26715v1 Announce Type: new Abstract: Major data protection regulations all mention the 'right to be forgotten,' and that's what pushed federated unlearning (FU) techniques forward. But one

Incremental Gauss-Newton Descent for Machine Learning

Local AiDGX agent

arXiv:2408.05560v2 Announce Type: replace Abstract: Stochastic gradient updates are widely used for their efficiency and scalability, but their effective step sizes can depend strongly on feature scal

Inferring Group Intent as a Cooperative Game. An NLP-based Framework for Trajectory Analysis

ResearchDGX agent

arXiv:2510.23905v2 Announce Type: replace-cross Abstract: This paper studies group target trajectory intent as the outcome of a cooperative game where the complex-spatio trajectories are modeled using

Information Theoretic Perspective on Representation Learning

ResearchDGX agent

arXiv:2601.11334v2 Announce Type: replace-cross Abstract: An information-theoretic framework is introduced to analyze last-layer embedding, focusing on learned representations for regression tasks. We

Interpretability and Generalization Bounds for Learning Spatial Physics

Model ReleasesDGX agent

arXiv:2506.15199v3 Announce Type: replace Abstract: While there are many applications of ML to scientific problems that look promising, visuals can be deceiving. Using numerical analysis techniques, w

Kan Extension Transformers: A Categorical Unification of Attention, Diffusion, and Predict-Detach Self-Conditioning

ResearchDGX agent

arXiv:2605.27259v1 Announce Type: new Abstract: We propose Kan Extension Transformers (KETs) as a unifying categorical framework for a diverse group of Transformer implementations. The core claim is t

Learnable Kernel Density Estimation for Graphs and Its Application to Graph-Level Anomaly Detection

Model ReleasesDGX agent

arXiv:2505.21285v4 Announce Type: replace Abstract: This work proposes a framework LGKDE that learns kernel density estimation for graphs. The key challenge in graph density estimation lies in effecti

LearnedCache: An eBPF-Integrated Perceptron-Based Eviction Policy for the Linux Page Cache

SafetyDGX agent

arXiv:2605.26168v1 Announce Type: cross Abstract: Linux is the foundation of the digital age, accounting for the majority of the cloud and mobile OS markets. Any device that runs Linux uses the Linux

Learning Dynamic Graph Representations through Timespan View Contrasts

SafetyDGX agent

arXiv:2605.27063v1 Announce Type: new Abstract: The rich information underlying graphs has inspired further investigation of unsupervised graph representation. Existing studies mainly depend on node f

Learning Energy-Based Models from Stochastic Interpolants using Spatiotemporal Differences

TutorialsDGX agent

arXiv:2605.26850v1 Announce Type: new Abstract: Learning an energy-based model from data samples is a central problem in machine learning. Many recent and popular methods, such as denoising score matc

Learning Nonlinear Factor Models with Unknown Monotone Links from Incomplete and Noisy Data

ResearchDGX agent

arXiv:2605.26271v1 Announce Type: cross Abstract: We study a nonlinear factor model in which observed responses depend on low-rank latent factors through an unknown monotone link function. This settin

Learning to Orchestrate Agents under Uncertainty

SafetyDGX agent

arXiv:2605.27073v1 Announce Type: new Abstract: Adaptive orchestration of heterogeneous agents requires making sequential delegation decisions under uncertain and evolving agent behaviour, e.g., coord

Learning to Reason Efficiently with Discounted Reinforcement Learning

SafetyDGX agent

arXiv:2510.23486v2 Announce Type: replace Abstract: Large reasoning models (LRMs) often consume excessive tokens, inflating computational cost and latency. More broadly, in goal reaching sequential de

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

Model ReleasesDGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

Localizing Memorized Regions in Diffusion Models via Coordinate-Wise Curvature Differences

Local AiDGX agent

arXiv:2605.26756v1 Announce Type: new Abstract: Diffusion models can unintentionally memorize training samples, raising concerns about privacy and copyright. While recent methods can detect memorizati

MATT-CTR: Unleashing a Model-Agnostic Test-Time Paradigm for CTR Prediction with Confidence-Guided Inference Paths

Model ReleasesDGX agent

arXiv:2510.08932v2 Announce Type: replace Abstract: Recently, a growing body of research has focused on either optimizing CTR model architectures to better model feature interactions or refining train

MechRL: Reinforcement Learning Agents Perform Circuit Discovery for Mechanistic Interpretability

SafetyDGX agent

arXiv:2605.26343v1 Announce Type: new Abstract: Mechanistic interpretability has identified small sets of attention heads that implement specific behaviours in transformer language models, but recover

Membership Inference Risks in Quantized Models: A Theoretical and Empirical Study

ApplicationsDGX agent

arXiv:2502.06567v2 Announce Type: replace-cross Abstract: Quantizing machine learning models has demonstrated its effectiveness in lowering memory and inference costs while maintaining performance lev

Mildly Overparameterized ReLU Networks on Orthogonal Data: Incremental Learning and Implicit Bias

SafetyDGX agent

arXiv:2605.27097v1 Announce Type: new Abstract: The successful training of neural networks hinges on the use of first order optimization methods, yet the theoretical characterization of these methods

Minimal surfaces, Knots, and Neural Networks

ResearchDGX agent

arXiv:2605.26234v1 Announce Type: cross Abstract: A recent conjecture by Joel Fine posits a relationship between the coefficients of the HOMFLY polynomial of a knot K in the 3-sphere S^3, and the sign

MolPIF: A Parameter Interpolation Flow Model for Molecule Generation

Model ReleasesDGX agent

arXiv:2507.13762v4 Announce Type: replace Abstract: Motivation: Structure-based drug design (SBDD) has advanced with deep generative models, but bridging the gap between continuous atomic coordinates

Morphling: Fast, Fused, and Flexible GNN Training at Scale

HardwareDGX agent

arXiv:2512.01678v5 Announce Type: replace Abstract: Graph Neural Networks (GNNs) present a fundamental hardware challenge by fusing irregular, memory-bound graph traversals with regular, compute-inten

MTL-FNO: A Lightweight Multi-Task Fourier Neural Operator for Sparse Field Reconstruction

Model ReleasesDGX agent

arXiv:2605.26718v1 Announce Type: new Abstract: Efficient onboard multi-field sparse reconstruction is essential for the autonomous operation of aerospace vehicles. While existing deep learning models

MuCon: Clipped Muon Updates for LLM Training

ResearchDGX agent

arXiv:2605.26459v1 Announce Type: new Abstract: Muon-style optimizers take a matrix-valued momentum or preconditioned update B = U operatorname{diag}(sigma_1,ldots,sigma_r) V^op and replace it with it

Near-Optimal Regret in Adversarial Kernel Bandits

Model ReleasesDGX agent

arXiv:2605.26585v1 Announce Type: new Abstract: We study the adversarial kernel bandit problem, in which the loss at each round is induced by an arbitrary bounded element of a reproducing kernel Hilbe

Neural Autoregressive Control Variates for the Quantum Monte Carlo Sign Problem

Model ReleasesDGX agent

arXiv:2605.26814v1 Announce Type: cross Abstract: We train a pair of autoregressive models to construct zero-mean control variates to mitigate the sign problem in quantum Monte Carlo simulations. The

Neural Bayesian Sequential Routing

AgentsDGX agent

arXiv:2605.26147v1 Announce Type: new Abstract: Human decision-making is sequential and uncertainty-aware, yet standard neural networks often rely on static, dense forward computation with limited vis

Neural Scaling Laws of Deep ReLU and Deep Operator Network: A Theoretical Study

ResearchDGX agent

arXiv:2410.00357v2 Announce Type: replace Abstract: Neural scaling laws play a pivotal role in the performance of deep neural networks and have been observed in a wide range of tasks. However, a compl

Nonlinear Data Integration via Kernel Methods for Data Collaboration Analysis

ResearchDGX agent

arXiv:2605.27219v1 Announce Type: new Abstract: Collaborative analysis of decentralized confidential datasets is important, but direct sharing of original datasets is often restricted by privacy and i

Normal Guidance is what Attention Needs

TutorialsDGX agent

arXiv:2605.27306v1 Announce Type: new Abstract: We consider training classifiers for 3D medical images using only one binary label for the entire volume rather than a label for each 2D slice. In such

Normalizing Flows on Quotient Manifolds via Boundary Quotients

ResearchDGX agent

arXiv:2511.22882v3 Announce Type: replace Abstract: We introduce boundary quotients and present a framework for learning densities on manifolds that arise as boundary quotients of simpler domains. We

Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation

Model ReleasesDGX agent

arXiv:2605.26844v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts with token-level teacher supervision. Recent selective OPD methods exploit the non-uni

On the Role of Inductive Bias in Time-Series Pretraining: A Case Study in Learning Generalizable Representations for Clinical Time Series

SafetyDGX agent

arXiv:2605.26194v1 Announce Type: new Abstract: Clinical time-series learning is routinely constrained by small, heterogeneous cohorts and protocol drift, while its downstream use spans both classific

Online Learning on Hidden-Convex Losses via Algorithmic Equivalence: Optimal Regret, Geometric Barrier, and Bandit Feedback

ResearchDGX agent

arXiv:2605.26373v1 Announce Type: new Abstract: We study adversarial online learning with hidden-convex losses, i.e., nonconvex losses that become convex after a nonlinear reparameterization. Ghai, Lu

Open-Weight LLM Fine-Tuning Defenses are Susceptible to Simple Attacks

SafetyDGX agent

arXiv:2605.26526v1 Announce Type: new Abstract: Recent defenses for safeguarding open-weight large language models (LLMs) are intended to prevent adversarial usage. Underlying these defenses is an ass

Optimal Rates for Feasible Payoff Set Estimation in Games

AgentsDGX agent

arXiv:2602.04397v2 Announce Type: replace-cross Abstract: We study a setting in which two players play a (possibly approximate) Nash equilibrium of a bimatrix game, while a learner observes only their

Over-Alignment vs Over-Fitting: The Role of Feature Learning Strength in Generalization

SafetyDGX agent

arXiv:2602.00827v2 Announce Type: replace Abstract: Feature learning strength (FLS), i.e., the inverse of the effective output scaling of a model, plays a critical role in shaping the optimization dyn

Parsimonious Learning-Augmented Online Metric Matching

ResearchDGX agent

arXiv:2605.26886v1 Announce Type: cross Abstract: Learning-augmented algorithms have received significant attention in recent years, particularly in the context of online optimization. Motivated by th

Particle-Lund Multimodality in Jet Taggers

ResearchDGX agent

arXiv:2605.26821v1 Announce Type: cross Abstract: The Lund plane offers a physics-motivated, hierarchical representation of QCD radiation within jets, while transformer-based taggers have reached stat

PATE-TabTransGAN: Differentially Private Synthetic Tabular Data Generation via Transformer-Based Student Discrimination

ResearchDGX agent

arXiv:2605.26802v1 Announce Type: new Abstract: Generating high-fidelity synthetic tabular data under formal differential privacy guarantees remains an open challenge. Methods that provide strong theo

PhyGHT: Physics-Guided HyperGraph Transformer for Signal Purification at the HL-LHC

Local AiDGX agent

arXiv:2602.20475v2 Announce Type: replace-cross Abstract: The High-Luminosity Large Hadron Collider (HL-LHC) at CERN will produce unprecedented datasets capable of revealing fundamental properties of

PIDM-DP: Physics-Informed Diffusion with Dormand-Prince Integration for Chaotic System Identification and State Reconstruction across Multiple Dynamical Regimes

Model ReleasesDGX agent

arXiv:2605.26619v1 Announce Type: new Abstract: Reconstructing continuous state trajectories of chaotic dynamical systems from sparse, noisy observations remains a fundamental open problem in nonlinea

PLAID: A Unified Data Model for Machine Learning on Heterogeneous Physics Simulations

ResearchDGX agent

arXiv:2505.02974v3 Announce Type: replace Abstract: Machine learning-based surrogate models have emerged as a powerful tool to accelerate simulation-driven scientific workflows, but their adoption is

Position: Machine Learning for Heart Transplant Allocation Policy Optimization Should Account for Incentives

SafetyDGX agent

arXiv:2602.04990v3 Announce Type: replace Abstract: The allocation of scarce donor organs constitutes one of the most consequential algorithmic challenges in healthcare. While the field is rapidly tra

Pretrained Approximators for Low-Thrust Trajectory Cost and Reachability

Model ReleasesDGX agent

arXiv:2605.26790v1 Announce Type: new Abstract: Low-thrust trajectory design relies heavily on repeated evaluations of fuel consumption and transfer feasibility, which require expensive optimal contro

PRISM: Position-encoded Regressive Inverse Spectral Model for Multilayer Thin-Film Design

Model ReleasesDGX agent

arXiv:2605.26502v1 Announce Type: new Abstract: The inverse problem of multilayer thin-film optical coatings design represents a complex combinatorial-continuous optimization challenge. We present PRI

Probabilistic Recurrent Intention Switching Model

ResearchDGX agent

arXiv:2605.26998v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) recovers reward functions from observed behavior, yet traditional methods assume a single stationary reward that ca

← Previous
1…122123124125126…243
Next →