AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
30 Jun 2026

Reinforcement Learning in Super Mario Bros: Curriculum, Pedagogy, and Optimal Level Design in World 1-1

AgentsDGX agent

arXiv:2606.29511v1 Announce Type: new Abstract: World 1-1 of Super Mario Bros is widely celebrated as a masterclass in game design: its progressive structure is credited with teaching players core mec

Reliability, Faithfulness, and the Limits of Post-hoc Explanations of Opaque Scientific Models

ResearchDGX agent

arXiv:2606.29346v1 Announce Type: new Abstract: Post-hoc explanation methods are routinely used to interpret scientific machine learning models, with the deliverable understood to be insight into the

Replica Symmetry Breaking and Algorithmic Thresholds in Empirical Risk Minimization under Multi-Index Model

TutorialsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.28573v1 Announce Type: new Abstract: Modern machine learning models are trained by optimizing high-dimensional non-convex empirical risk functions. Such cost functions can have a multitude

Reproducing FACTER: Fairness via Conformal Thresholding and Prompt Repair

SafetyDGX agent

arXiv:2606.28620v1 Announce Type: cross Abstract: Fayyazi et al. (2025) recently proposed FACTER, a model-agnostic framework designed to jointly enforce fairness and statistical coverage in LLM-based

Residual-Guided Dictionary Learning for Spectrally Accurate Koopman Approximation

Model ReleasesDGX agent

arXiv:2606.29083v1 Announce Type: cross Abstract: Koopman theory promises linear structure in nonlinear dynamics, but numerical Koopman spectra are easy to compute and hard to trust. A finite EDMD mat

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation

SafetyDGX agent

arXiv:2602.09305v2 Announce Type: replace Abstract: Large Language Models (LLMs) demonstrate transformative potential, yet their reasoning remains inconsistent and unreliable. Reinforcement learning (

Robust Strategic Classification under Decision-Dependent Cost Uncertainty

SafetyDGX agent

arXiv:2606.30136v1 Announce Type: new Abstract: Humans facing algorithmic decision systems have been found to ``game'' them by altering their input data (at a cost to them) in order to favorably chang

Robust Tangent Space Estimation via Laplacian Eigenvector Gradient Orthogonalization

TutorialsDGX agent

arXiv:2510.02308v2 Announce Type: replace Abstract: Estimating the tangent spaces of a data manifold is a fundamental problem in geometric data analysis. The standard approach, Local Principal Compone

Robustness and Structure Preservation in Flow-Based Generative Models via Wasserstein Path-Space Divergences

SafetyDGX agent

arXiv:2410.01244v2 Announce Type: replace-cross Abstract: We introduce a novel Wasserstein-1 (W_1) path-space divergence for stochastic and deterministic dynamics and establish a Wasserstein Uncertain

Sample Complexity of Scientific Discovery: PAC Learnability of Compositional Function Trees

ResearchDGX agent

arXiv:2606.29331v1 Announce Type: new Abstract: Scientific discovery via symbolic regression is often viewed as statistically and computationally intractable because the hypothesis space of expression

Scalar Representations of Neural Network Training Dynamics

Model ReleasesDGX agent

arXiv:2606.30384v1 Announce Type: new Abstract: Training in artificial neural networks can be viewed as a trajectory evolving through a high-dimensional loss landscape. However, the large number of tr

scKDGM: KAN-guided Dynamic Graph Masked Learning for Single-Cell RNA-seq Clustering

TutorialsDGX agent

arXiv:2606.28459v1 Announce Type: new Abstract: Single-cell RNA sequencing (scRNA-seq) clustering is essential for identifying cell types, but high dimensionality, sparsity, dropout, and technical noi

Self-Supervised Calibration of Scientific Instruments Using Physical Consistency Constraints

AgentsDGX agent

arXiv:2606.29466v1 Announce Type: new Abstract: Calibration remains one of the principal obstacles to the deployment of machine learning in scientific instrumentation because it typically relies on ex

Sequential Hiring of Contingent Workers Through Learning-Based Optimization

Model ReleasesDGX agent

arXiv:2606.18438v2 Announce Type: replace-cross Abstract: In this paper, we study a sequential workforce management problem in a contingent labor setting with uncertainty in both worker production and

SGD Provably Prioritizes a Shortcut Spurious Feature in the XOR Model

SafetyDGX agent

arXiv:2606.30444v1 Announce Type: cross Abstract: Neural networks are known to be susceptible to over-reliance on spurious correlations. However, the precise mechanism by which models exploit shortcut

Shoot from the HIP: Hessian Interatomic Potentials without derivatives

ResearchDGX agent

arXiv:2509.21624v3 Announce Type: replace Abstract: Fundamental tasks in computational chemistry, from transition state search to vibrational analysis, rely on molecular Hessians, which are the second

Simplifying Flow Matching Transformations with Low-Rank Mixture Models

TutorialsDGX agent

arXiv:2606.29724v1 Announce Type: new Abstract: Normalizing flows are powerful generative models that learn an invertible mapping between complex data distributions and simple latent distributions, ty

Singular Learning and Occam's Razor in Deep Monomial Networks

Model ReleasesDGX agent

arXiv:2606.28464v1 Announce Type: new Abstract: In the optimization of neural networks, gradient dynamics are influenced by critical points that arise from the model's architecture. These critical poi

SMaRT: Online Reusable Resource Assignment and an Application to Mediation in the Kenyan Judiciary

AgentsDGX agent

arXiv:2602.18431v2 Announce Type: replace-cross Abstract: Motivated by the problem of assigning mediators to cases in the Kenyan judicial system, we study an online resource allocation problem where i

Solver-Integrated Adversarial Attacking and Training of Neural Operators

ResearchDGX agent

arXiv:2510.18989v2 Announce Type: replace Abstract: Neural operators are commonly utilized as fast surrogates for numerical solvers in PDE problems, mapping input functions to solution functions. Howe

SP-CACW: Convergence-Aware Client Weighting for Selfish Personalized Learning

SafetyDGX agent

arXiv:2606.29322v1 Announce Type: new Abstract: Collaborative learning is sustainable only when it benefits each participant. Standard federated learning optimizes a global average objective, which ca

SparsePixels: Efficient Convolution for Sparse Data on FPGAs

ResearchDGX agent

arXiv:2512.06208v3 Announce Type: replace-cross Abstract: Inference of standard convolutional neural networks (CNNs) on FPGAs often incurs high latency and a long initiation interval due to the deep n

Spatial Deconfounder: Interference-Aware Deconfounding for Spatial Causal Inference

Model ReleasesDGX agent

arXiv:2510.08762v2 Announce Type: replace Abstract: Causal inference in spatial domains faces two intertwined challenges: (1) unmeasured spatial factors, such as weather, air pollution, or mobility, t

Spectral Embedding via Chebyshev Bases for Robust DeepONet Approximation

Model ReleasesDGX agent

arXiv:2512.09165v2 Announce Type: replace Abstract: Deep Operator Networks (DeepONets) have emerged as a powerful framework for data-driven operator learning, providing flexible surrogates for nonline

Spectral phase transitions and trainability in neural network learning dynamics

SafetyDGX agent

arXiv:2606.28486v1 Announce Type: cross Abstract: The emergence of low-dimensional structures in the spectra of neural network weight matrices is a common empirical feature of trained models, but the

Speculative Pre-Positioning: Decoding Stateful Sessions to the Next Decision Point Off the Critical Path

ResearchDGX agent

arXiv:2606.29565v1 Announce Type: new Abstract: A stateless inference server (vLLM, SGLang, TensorRT-LLM) idles between requests while the accelerator waits; a stateful session reclaims that idle time

Staged Hybridisation for Visual Quantum Reinforcement Learning via Knowledge Distillation

SafetyDGX agent

arXiv:2606.30520v1 Announce Type: cross Abstract: Visual environments are a demanding setting for quantum reinforcement learning (QRL): high-dimensional observations, unstable RL optimisation, and con

STEMGym: Benchmarking Sequential Decision-Making under Dose Budgets in Autonomous Electron Microscopy

Model ReleasesDGX agent

arXiv:2606.29592v1 Announce Type: new Abstract: A central premise of autonomous scientific imaging is that smarter navigation, whether Bayesian, RL-based, or otherwise adaptive, is the principal lever

Stochastic and Non-local Closure Modeling for Nonlinear Dynamical Systems via Latent Score-based Generative Models

Local AiDGX agent

arXiv:2506.20771v2 Announce Type: replace Abstract: We propose a latent score-based generative AI framework for learning stochastic, non-local closure models and constitutive laws in nonlinear dynamic

Structured Proper Loss Geometries for Multiclass Classification: Theory and Controlled Empirical Evaluation

ResearchDGX agent

arXiv:2606.29471v1 Announce Type: new Abstract: Strictly proper scoring rules identify the true conditional class distribution at population level, but their curvature can alter optimization and finit

Surrogate Modeling for Explainable Predictive Time Series Corrections

Local AiDGX agent

arXiv:2412.19897v3 Announce Type: replace-cross Abstract: We introduce a local surrogate approach for explainable time-series forecasting. An initially non-interpretable predictive model to improve th

SWE-INTERACT: Reimagining SWE Benchmarks as User-Driven Long-Horizon Coding Sessions

AgentsDGX agent

arXiv:2606.30573v1 Announce Type: new Abstract: We introduce SWE-Interact, a new testbed for evaluating coding agents on multi-turn, interactive, user-driven software engineering tasks. Existing front

Synthetic Interaction Data for Scalable Personalization in Large Language Models

AgentsDGX agent

arXiv:2602.12394v2 Announce Type: replace Abstract: Personalized prompting offers large opportunities for deploying large language models (LLMs) to diverse users, yet existing prompt optimization meth

t-STEP: An interpretable model for Total Electron Content predictions and irregularities estimations

ResearchDGX agent

arXiv:2606.29644v1 Announce Type: new Abstract: Earth system infrastructures relying on satellite-based technologies, such as Global Positioning System (GPS) communications, are affected by ionospheri

Temporal Posed and Spontaneous Gesture Recognition from Electromyography in the Rock-Paper-Scissors Game

ApplicationsDGX agent

arXiv:2606.29423v1 Announce Type: new Abstract: The importance of gesture recognition has been acknowledged in many domains requiring real-time recognition systems. Two requirements for these are fast

TextClusterLab: An Integrated Framework for Reliable Text Clustering Studies

Model ReleasesDGX agent

arXiv:2606.28328v1 Announce Type: cross Abstract: In recent years, text clustering has become a critical technique for applications including intent discovery, topic mining, and recommendation systems

The Contagion Tensor: A Framework for Measuring Output-Distribution Coupling in Multi-Agent LLM Systems -- and Auditing the Claims It Enables

Model ReleasesDGX agent

arXiv:2606.28839v1 Announce Type: new Abstract: We introduce the Contagion Tensor, a measurement framework for quantifying how large language model (LLM) output distributions couple across modalities,

The Forgetting-Retention Dilemma: Certified Unlearning Theory in Continual Learning

ResearchDGX agent

arXiv:2606.29832v1 Announce Type: new Abstract: Machine unlearning aims to eliminate the influence of specific data from trained models to safeguard privacy. However, this presents a significant chall

The Fundamental Limits of Valid Transport Map Estimation

ResearchDGX agent

arXiv:2606.30574v1 Announce Type: new Abstract: Many modern generative modeling methods, including diffusion models, normalizing flows, and flow matching, estimate transport maps or plans between dist

The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.29526v1 Announce Type: new Abstract: Reinforcement learning (RL) has gained growing attention in large language model (LLM) post-training, yet RL training remains fragile and can suffer fro

The Voronoi Bottleneck: Capacity-Aware Dense Retrieval for Product Search

ResearchDGX agent

arXiv:2606.28359v1 Announce Type: cross Abstract: Dense embedding retrieval compresses all relevance information into a single inner product, imposing a fundamental geometric limit -- the Voronoi Bott

Theory of Continual Learning Against Data Poisoning Attacks

SafetyDGX agent

arXiv:2606.29841v1 Announce Type: new Abstract: Continual learning (CL), where a model is trained on a sequence of data tasks, is increasingly being adopted across key fields such as large language mo

To Use or not to Use Muon: How Simplicity Bias in Optimizers Matters

SafetyDGX agent

arXiv:2603.00742v2 Announce Type: replace Abstract: While Adam has long been the ubiquitous default optimizer for deep neural networks, Muon has recently seen rapid adoption due to its superior traini

Toward an Energy-Optimized Operation of Data Centers Located in Wind Farms Using Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.30316v1 Announce Type: new Abstract: This paper studies Reinforcement Learning as an online controller for curtailment-aware workload shifting in wind-turbine-integrated high-performance co

Towards Complete Causal Explanation with Expert Knowledge

ResearchDGX agent

arXiv:2407.07338v4 Announce Type: replace-cross Abstract: We study the problem of restricting a Markov equivalence class of maximal ancestral graphs (MAGs) to only those MAGs that contain certain edge

Towards Evaluating Data Priors for Tabular Foundation Models

ResearchDGX agent

arXiv:2606.29241v1 Announce Type: new Abstract: Data-generating priors are a central component of tabular foundation models because they define the task distribution used during pretraining. However,

Towards Improved Anomaly Detection for Cloud Cybersecurity via Graph Neural Networks

ApplicationsDGX agent

arXiv:2606.28923v1 Announce Type: new Abstract: Detecting security threats in an organization's cloud computing environment has become necessary due to the increased reliance on cloud infrastructure.

Transformer-Based Active Learning for Data-Efficient Vaccine Epitope Selection in PRRS

SafetyDGX agent

arXiv:2606.28659v1 Announce Type: cross Abstract: High-fidelity molecular docking simulations can produce biologically relevant estimates of epitope-receptor binding affinity but are computationally e

Transolver-3: Scaling Up Transformer Solvers to Industrial-Scale Geometries

HardwareDGX agent

arXiv:2602.04940v2 Announce Type: replace Abstract: Deep learning has emerged as a transformative tool for the neural surrogate modeling of partial differential equations (PDEs), known as neural PDE s

Two kinds of robustness are not the same: disentangling fault tolerance and low-SNR robustness in multi-domain event detection on real data

Model ReleasesDGX agent

arXiv:2606.29339v1 Announce Type: cross Abstract: Reliable event detection underpins induced-seismicity monitoring for Carbon dioxide Capture and Storage (CCS) and geothermal operations, distributed a

Universality of empirical risk minimization

ResearchDGX agent

arXiv:2202.08832v3 Announce Type: replace-cross Abstract: We study a general class of optimization problems with decision variable oldsymbol{Theta} in R^{p imes k} and cost function which is the sum o

Variance Reduction for Stochastic Gradient Generalized Non-reversible Langevin Monte Carlo Algorithms

ResearchDGX agent

arXiv:2606.28808v1 Announce Type: cross Abstract: We study the leading-order fluctuation of stochastic gradient Euler-Maruyama estimators for generalized non-reversible Langevin dynamics. Under struct

Warm-Starting Iterative Gaussian Processes for Faster Sequential Inference

ResearchDGX agent

arXiv:2511.16340v2 Announce Type: replace Abstract: Efficient Gaussian process (GP) inference is critical for sequential decision-making tasks such as active learning, online prediction, and Bayesian

Wasserstein Distributionally Robust Regret Optimization

ResearchDGX agent

arXiv:2504.10796v4 Announce Type: replace-cross Abstract: Distributionally robust optimization (DRO) is widely used for decision-making under uncertainty, but its adversarial focus on worst-case loss

Weak Dominant Balance for Robust Identification of Dynamically Consistent Fluid Flow Structure

ResearchDGX agent

arXiv:2606.29047v1 Announce Type: cross Abstract: Extracting interpretable, localized physical mechanisms from complex spatiotemporal data is a foundational challenge across physics, biology, and engi

When Can Conformal Risk Control Certify LLM Outputs? Bounds, Impossibility, and Adaptation for Structured Generation

ResearchDGX agent

arXiv:2606.29054v1 Announce Type: new Abstract: Large language models (LLMs) deployed for structured generation (NER, JSON extraction, QA, and classification) lack formal reliability guarantees, and s

When Does Online Imitation Learning Help in LLM Post-Training? The Role of (Non-)Realizability Beyond Horizon

SafetyDGX agent

arXiv:2606.30445v1 Announce Type: new Abstract: Online imitation learning (IL), particularly on-policy distillation, has emerged as a strong LLM post-training approach, often outperforming offline sup

When Prices Double in a Week: Forecasting of Agricultural Volatility in Import-Isolated Markets

ResearchDGX agent

arXiv:2606.29248v1 Announce Type: new Abstract: Vegetable prices in Sri Lanka are highly volatile because the market is largely import-isolated, so supply disruptions quickly drive prices up. This stu

Why Do We Need Warm-up? A Theoretical Perspective

Model ReleasesDGX agent

arXiv:2510.03164v2 Announce Type: replace Abstract: Learning rate warm-up -- increasing the learning rate at the beginning of training -- has become a ubiquitous heuristic in modern deep learning, yet

Wireless Backdoor Attack and Defense for Semantic Communications over Multiple Access Channel

ResearchDGX agent

arXiv:2606.30595v1 Announce Type: cross Abstract: Semantic communication (SemCom) aims to preserve semantic meaning and task-oriented information beyond conventional message recovery over wireless cha

← Previous
1…6263646566…241
Next →