AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
23 Apr 2026

How Will My Business Process Unfold? Predicting Case Suffixes With Start and End Timestamps

ResearchDGX agent

arXiv:2509.14536v3 Announce Type: replace Abstract: Predictive process monitoring supports operational decision-making by forecasting future states of ongoing business cases. A key task is case suffix

Improved large-scale graph learning through ridge spectral sparsification

TutorialsDGX agent

arXiv:2604.20078v1 Announce Type: new Abstract: Graph-based techniques and spectral graph theory have enriched the field of machine learning with a variety of critical advances. A central object in th

Improving clinical interpretability of linear neuroimaging models through feature whitening

ResearchDGX agent

arXiv:2604.20675v1 Announce Type: new Abstract: Linear models are widely used in computational neuroimaging to identify biomarkers associated with brain pathologies. However, interpreting the learned


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Improving Large-Scale Recommender Systems with Auxiliary Learning

ApplicationsDGX agent

arXiv:2510.02215v3 Announce Type: replace Abstract: Training large-scale recommendation models under a single global objective implicitly assumes homogeneity across user populations. However, real-wor

Issues with Value-Based Multi-objective Reinforcement Learning: Value Function Interference and Overestimation Sensitivity

ResearchDGX agent

arXiv:2402.06266v2 Announce Type: replace Abstract: Multi-objective reinforcement learning (MORL) algorithms extend conventional reinforcement learning (RL) to the more general case of problems with m

Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning

SafetyDGX agent

arXiv:2505.07527v5 Announce Type: replace Abstract: The advantage function is a central concept in RL that helps reduce variance in policy gradient estimates. For language modeling, Group Relative Pol

KANMixer: a minimal KAN-centered mixer for long-term time series forecasting

Model ReleasesDGX agent

arXiv:2508.01575v2 Announce Type: replace Abstract: Long-term time series forecasting (LTSF) underpins critical applications from energy management to weather prediction, yet achieving reliable multi-

Latent Stochastic Interpolants

Model ReleasesDGX agent

arXiv:2506.02276v2 Announce Type: replace Abstract: Stochastic Interpolants (SI) is a powerful framework for generative modeling, capable of flexibly transforming between two probability distributions

Lever: Inference-Time Policy Reuse under Support Constraints

SafetyDGX agent

arXiv:2604.20174v1 Announce Type: new Abstract: Reinforcement learning (RL) policies are typically trained for fixed objectives, making reuse difficult when task requirements change. We study inferenc

Local Diffusion Models and Phases of Data Distributions

TutorialsDGX agent

arXiv:2508.06614v2 Announce Type: replace Abstract: As a class of generative artificial intelligence frameworks inspired by statistical physics, diffusion models have shown extraordinary performance i

Machine Learning for Two-Stage Graph Sparsification for the Travelling Salesman Problem

TutorialsDGX agent

arXiv:2604.20236v1 Announce Type: new Abstract: High-performance TSP solvers like LKH search within a sparsified candidate graph rather than over all possible edges. Graph sparsification is non-trivia

Machine learning moment closure models for the radiative transfer equation IV: enforcing symmetrizable hyperbolicity in two dimensions

ResearchDGX agent

arXiv:2604.20143v1 Announce Type: cross Abstract: This is our fourth work in the series on machine learning (ML) moment closure models for the radiative transfer equation (RTE). In the first three pap

MasconCube: Fast and Accurate Gravity Modeling with an Explicit Representation

ResearchDGX agent

arXiv:2509.08607v3 Announce Type: replace-cross Abstract: The geodesy of irregularly shaped small bodies presents fundamental challenges for gravitational field modeling, particularly as deep space ex

Maximum Entropy Semi-Supervised Inverse Reinforcement Learning

ResearchDGX agent

arXiv:2604.20074v1 Announce Type: new Abstract: A popular approach to apprenticeship learning (AL) is to formulate it as an inverse reinforcement learning (IRL) problem. The MaxEnt-IRL algorithm succe

Mechanistic Interpretability Tool for AI Weather Models

ResearchDGX agent

arXiv:2604.20467v1 Announce Type: cross Abstract: Artificial Intelligence (AI) weather models are improving rapidly, and their forecasts are already competitive with long-established traditional Numer

MGDA-Decoupled: Geometry-Aware Multi-Objective Optimisation for DPO-based LLM Alignment

SafetyDGX agent

arXiv:2604.20685v1 Announce Type: new Abstract: Aligning large language models (LLMs) to desirable human values requires balancing multiple, potentially conflicting objectives such as helpfulness, tru

MixLLM: LLM Quantization with Global Mixed-precision between Output-features and Highly-efficient System Design

Model ReleasesDGX agent

arXiv:2412.14590v2 Announce Type: replace Abstract: Quantization has become one of the most effective methodologies to compress LLMs into smaller size. However, the existing quantization solutions sti

Multi-Armed Bandits With Machine Learning-Generated Surrogate Rewards

SafetyDGX agent

arXiv:2506.16658v2 Announce Type: replace-cross Abstract: Multi-armed bandit (MAB) is a widely adopted framework for sequential decision-making under uncertainty. Traditional bandit algorithms rely so

Multi-Objective Reinforcement Learning for Generating Covalent Inhibitor Candidates

SafetyDGX agent

arXiv:2604.20019v1 Announce Type: new Abstract: Rational design of covalent inhibitors requires simultaneously optimizing multiple properties, such as binding affinity, target selectivity, or electrop

Near-Future Policy Optimization

SafetyDGX agent

arXiv:2604.20733v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a core post-training recipe. Introducing suitable off-policy trajectories into on-polic

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning

SafetyDGX agent

arXiv:2604.20627v1 Announce Type: new Abstract: The temporal lag between actions and their long-term consequences makes credit assignment a challenge when learning goal-directed behaviors from data. G

On Bayesian Softmax-Gated Mixture-of-Experts Models

Model ReleasesDGX agent

arXiv:2604.20551v1 Announce Type: cross Abstract: Mixture-of-experts models provide a flexible framework for learning complex probabilistic input-output relationships by combining multiple expert mode

On the definition and importance of interpretability in scientific machine learning

ResearchDGX agent

arXiv:2505.13510v3 Announce Type: replace Abstract: Though neural networks trained on large datasets have been successfully used to describe and predict many physical phenomena, there is a sense among

Online Survival Analysis: A Bandit Approach under Cox PH Model

ResearchDGX agent

arXiv:2604.20296v1 Announce Type: cross Abstract: Survival analysis is a widely used statistical framework for modeling time-to-event data under censoring. Classical methods, such as the Cox proportio

Optimal Single-Policy Sample Complexity and Transient Coverage for Average-Reward Offline RL

Model ReleasesDGX agent

arXiv:2506.20904v2 Announce Type: replace Abstract: We study offline reinforcement learning in average-reward MDPs, which presents increased challenges from the perspectives of distribution shift and

Option Pricing on Noisy Intermediate-Scale Quantum Computers: A Quantum Neural Network Approach

Model ReleasesDGX agent

arXiv:2604.19832v1 Announce Type: cross Abstract: In a global derivatives market with notional values in the hundreds of trillions of dollars, the accuracy and efficiency of pricing models are of fund

Overcoming the Modality Gap in Context-Aided Forecasting

ApplicationsDGX agent

arXiv:2603.12451v3 Announce Type: replace Abstract: Context-aided forecasting (CAF) holds promise for integrating domain knowledge and forward-looking information, enabling AI systems to surpass tradi

Personalized electric vehicle energy consumption estimation framework that integrates driver behavior with map data

ResearchDGX agent

arXiv:2604.20764v1 Announce Type: cross Abstract: This paper presents a personalized Battery Electric Vehicle (BEV) energy consumption estimation framework that integrates map-based contextual feature

Physics-Conditioned Synthesis of Internal Ice-Layer Thickness for Incomplete Layer Traces

TutorialsDGX agent

arXiv:2604.20783v1 Announce Type: new Abstract: Internal ice layers imaged by radar provide key evidence of snow accumulation and ice dynamics, but radar-derived layer boundary observations are often

Physics-Guided Dimension Reduction for Simulation-Free Operator Learning of Stiff Differential--Algebraic Systems

ResearchDGX agent

arXiv:2604.19930v1 Announce Type: new Abstract: Neural surrogates for stiff differential-algebraic equations (DAEs) face two key challenges: soft-constraint methods leave algebraic residuals that stif

Pre-Execution Query Slot-Time Prediction in Cloud Data Warehouses: A Feature-Scoped Machine Learning Approach

ResearchDGX agent

arXiv:2604.20145v1 Announce Type: cross Abstract: Cloud data warehouses bill compute based on slot-time consumed. In shared multi-tenant environments, query cost is highly variable and hard to estimat

Properties and limitations of geometric tempering for gradient flow dynamics

ResearchDGX agent

arXiv:2604.20301v1 Announce Type: cross Abstract: We consider the problem of sampling from a probability distribution pi. It is well known that this can be written as an optimisation problem over the

Quantum Adaptive Self-Attention for Quantum Transformer Models

ApplicationsDGX agent

arXiv:2504.05336v3 Announce Type: replace-cross Abstract: Integrating quantum computing into deep learning architectures is a promising but poorly understood endeavor: when does a quantum layer actual

R2IF: Aligning Reasoning with Decisions via Composite Rewards for Interpretable LLM Function Calling

ResearchDGX agent

arXiv:2604.20316v1 Announce Type: new Abstract: Function calling empowers large language models (LLMs) to interface with external tools, yet existing RL-based approaches suffer from misalignment betwe

Rashomon Sets and Model Multiplicity in Federated Learning

Model ReleasesDGX agent

arXiv:2602.09520v2 Announce Type: replace Abstract: The Rashomon set captures the collection of models that achieve near-identical empirical performance yet may differ substantially in their decision

Relative Entropy Estimation in Function Space: Theory and Applications to Trajectory Inference

Model ReleasesDGX agent

arXiv:2604.20775v1 Announce Type: new Abstract: Trajectory Inference (TI) seeks to recover latent dynamical processes from snapshot data, where only independent samples from time-indexed marginals are

Replicable Bandits with UCB based Exploration

ResearchDGX agent

arXiv:2604.20024v1 Announce Type: new Abstract: We study replicable algorithms for stochastic multi-armed bandits (MAB) and linear bandits with UCB (Upper Confidence Bound) based exploration. A bandit

Rethinking Intrinsic Dimension Estimation in Neural Representations

ResearchDGX agent

arXiv:2604.20276v1 Announce Type: new Abstract: The analysis of neural representation has become an integral part of research aiming to better understand the inner workings of neural networks. While t

Robust Out-of-Distribution Stochastic Optimization

TutorialsDGX agent

arXiv:2604.20147v1 Announce Type: cross Abstract: Data-driven decision-making under uncertainty typically presumes the collection of historical data from an unknown target probability distribution. Ho

Robustness of Spatio-temporal Graph Neural Networks for Fault Location in Partially Observable Distribution Grids

ResearchDGX agent

arXiv:2604.20403v1 Announce Type: new Abstract: Fault location in distribution grids is critical for reliability and minimizing outage durations. Yet, it remains challenging due to partial observabili

SAMix: Calibrated and Accurate Continual Learning via Sphere-Adaptive Mixup and Neural Collapse

SafetyDGX agent

arXiv:2510.15751v2 Announce Type: replace Abstract: While most continual learning methods focus on mitigating forgetting and improving accuracy, they often overlook the critical aspect of network cali

Scalable Quantum Reinforcement Learning on NISQ Devices with Dynamic-Circuit Qubit Reuse and Grover Optimization

AgentsDGX agent

arXiv:2509.16002v2 Announce Type: replace-cross Abstract: A scalable and resource-efficient quantum reinforcement learning framework is presented that eliminates the linear qubit-scaling barrier in mu

Scaling Self-Play with Self-Guidance

Model ReleasesDGX agent

arXiv:2604.20209v1 Announce Type: new Abstract: LLM self-play algorithms are notable in that, in principle, nothing bounds their learning: a Conjecturer model creates problems for a Solver, and both i

Sheaf Neural Networks on SPD Manifolds: Second-Order Geometric Representation Learning

ResearchDGX agent

arXiv:2604.20308v1 Announce Type: new Abstract: Graph neural networks face two fundamental challenges rooted in the linear structure of Euclidean vector spaces: (1) Current architectures represent geo

SMART: A Spectral Transfer Approach to Multi-Task Learning

ResearchDGX agent

arXiv:2604.20161v1 Announce Type: new Abstract: Multi-task learning is effective for related applications, but its performance can deteriorate when the target sample size is small. Transfer learning c

Spatio-temporal modelling of electric vehicle charging demand

Model ReleasesDGX agent

arXiv:2604.19841v1 Announce Type: cross Abstract: Accurate forecasting of electric vehicle (EV) charging demand is critical for grid management and infrastructure planning. Yet the field continues to

Stream-CQSA: Avoiding Out-of-Memory in Attention Computation via Flexible Workload Scheduling

HardwareDGX agent

arXiv:2604.20819v1 Announce Type: new Abstract: The scalability of long-context large language models is fundamentally limited by the quadratic memory cost of exact self-attention, which often leads t

Structure-Aware Variational Learning of a Class of Generalized Diffusions

ResearchDGX agent

arXiv:2604.20188v1 Announce Type: new Abstract: Learning the underlying potential energy of stochastic gradient systems from partial and noisy observations is a fundamental problem arising in physics,

Super Apriel: One Checkpoint, Many Speeds

Model ReleasesDGX agent

arXiv:2604.19877v1 Announce Type: new Abstract: We release Super Apriel, a 15B-parameter supernet in which every decoder layer provides four trained mixer choices -- Full Attention (FA), Sliding Windo

Surrogate Functionals for Machine-Learned Orbital-Free Density Functional Theory

ResearchDGX agent

arXiv:2604.20458v1 Announce Type: new Abstract: We introduce surrogate functionals: machine-learned energy functionals for orbital-free density functional theory (OF-DFT) which are defined not by univ

SwiftRepertoire: Few-Shot Immune-Signature Synthesis via Dynamic Kernel Codes

ResearchDGX agent

arXiv:2602.01051v4 Announce Type: replace Abstract: Repertoire-level analysis of T cell receptors offers a biologically grounded signal for disease detection and immune monitoring, yet practical deplo

Synthetic Flight Data Generation Using Generative Models

ApplicationsDGX agent

arXiv:2604.20293v1 Announce Type: new Abstract: The increasing adoption of synthetic data in aviation research offers a promising solution to data scarcity and confidentiality challenges. This study i

Temporal Difference Calibration in Sequential Tasks: Application to Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.20472v1 Announce Type: cross Abstract: Recent advances in vision-language-action (VLA) models for robotics have highlighted the importance of reliable uncertainty quantification in sequenti

Temporally Extended Mixture-of-Experts Models

HardwareDGX agent

arXiv:2604.20156v1 Announce Type: new Abstract: Mixture-of-Experts models, now popular for scaling capacity at fixed inference speed, switch experts at nearly every token. Once a model outgrows availa

The Costs of Pretending That There Are Data-Generating Probability Distributions in the Social World

TutorialsDGX agent

arXiv:2407.17395v5 Announce Type: replace Abstract: Machine Learning research, including work promoting fair or equitable algorithms, often relies on the concept of a data-generating probability distr

The effect of the number of parameters and the number of local feature patches on loss landscapes in distributed quantum neural networks

ResearchDGX agent

arXiv:2504.19239v2 Announce Type: replace-cross Abstract: Quantum neural networks hold promise for tackling computationally challenging tasks that are intractable for classical computers. However, the

The Optical and Infrared Are Connected

ResearchDGX agent

arXiv:2503.03816v2 Announce Type: replace-cross Abstract: Galaxies are often modelled as composites of separable components with distinct spectral signatures, implying that different wavelength ranges

The Origin of Edge of Stability

ResearchDGX agent

arXiv:2604.20446v1 Announce Type: new Abstract: Full-batch gradient descent on neural networks drives the largest Hessian eigenvalue to the threshold 2/eta, where eta is the learning rate. This phenom

Throat and acoustic paired speech dataset for deep learning-based speech enhancement

SafetyDGX agent

arXiv:2502.11478v3 Announce Type: replace-cross Abstract: In high-noise environments such as factories, subways, and busy streets, capturing clear speech is challenging. Throat microphones can offer a

Too Sharp, Too Sure: When Calibration Follows Curvature

ResearchDGX agent

arXiv:2604.20614v1 Announce Type: new Abstract: Modern neural networks can achieve high accuracy while remaining poorly calibrated, producing confidence estimates that do not match empirical correctne

← Previous
1…210211212213214…241
Next →