AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
29 Apr 2026

How Fast Should a Model Commit to Supervision? Training Reasoning Models on the Tsallis Loss Continuum

SafetyDGX agent

arXiv:2604.25907v1 Announce Type: new Abstract: Adapting reasoning models to new tasks during post-training with only output-level supervision stalls under reinforcement learning from verifiable rewar

Immediate Derivatives Suffice for Online Recurrent Adaptation

ResearchDGX agent

arXiv:2603.28750v3 Announce Type: replace Abstract: For three decades online recurrent learning has been assumed to require propagating a Jacobian tensor through the network's dynamics at O(n^4) per s

Incompressible Knowledge Probes: Estimating Black-Box LLM Parameter Counts via Factual Capacity

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.24827v1 Announce Type: new Abstract: Closed-source frontier labs do not disclose parameter counts, and the standard alternative -- inference economics -- carries 2imes+ uncertainty from har

Injecting Measurement Information Yields a Fast and Noise-Robust Diffusion-Based Inverse Problem Solver

Model ReleasesDGX agent

arXiv:2508.02964v3 Announce Type: replace Abstract: Diffusion models have been firmly established as principled zero-shot solvers for linear and nonlinear inverse problems, owing to their powerful ima

Interpretable Fuzzy Modeling Reveals Population-Level Representation Differences in P300 Brain Computer Interfaces Across Neurodivergent and Neurotypical Cohorts

Model ReleasesDGX agent

arXiv:2604.24765v1 Announce Type: cross Abstract: P300-based brain-computer interfaces (BCIs) are widely used for communication, but population heterogeneity may alter the neural patterns available fo

Investigation into In-Context Learning Capabilities of Transformers

Model ReleasesDGX agent

arXiv:2604.25858v1 Announce Type: new Abstract: Transformers have demonstrated a strong ability for in-context learning (ICL), enabling models to solve previously unseen tasks using only example input

JaGuard: Position Error Correction of GNSS Jamming with Deep Temporal Graphs

ApplicationsDGX agent

arXiv:2509.14000v4 Announce Type: replace Abstract: Global Navigation Satellite Systems (GNSS) face growing disruption from intentional jamming, undermining critical infrastructure where precise posit

Knowledge-Data Dually Driven Paradigm for Accurate Landslide Susceptibility Prediction under Data-Scarce Conditions Using Geomorphic Priors and Tabular Foundation Model

ResearchDGX agent

arXiv:2604.25196v1 Announce Type: new Abstract: Landslide susceptibility prediction is critical for geohazard risk assessment and mitigation. Conventional data-driven paradigm achieves high predictive

Knowledge Distillation Must Account for What It Loses

SafetyDGX agent

arXiv:2604.25110v1 Announce Type: new Abstract: This position paper argues that knowledge distillation must account for what it loses: student models should be judged not only by retained task scores,

Kohn-Sham Hamiltonian from Effective Field Theory: Quasiparticle Band Narrowing from Frozen Core Dynamics

AgentsDGX agent

arXiv:2604.25199v1 Announce Type: cross Abstract: Kohn-Sham (KS) eigenvalues are routinely compared with angle-resolved photoemission (ARPES) and used as input for many-body methods, yet density funct

Laplace-Bridged Randomized Smoothing for Fast Certified Robustness

HardwareDGX agent

arXiv:2604.24993v1 Announce Type: new Abstract: Randomized Smoothing (RS) offers formal ell_2 guarantees for arbitrary base classifiers but faces two key practical bottlenecks: (i) it often relies on

Learning biophysical models of gene regulation with probability flow matching

ResearchDGX agent

arXiv:2604.25062v1 Announce Type: cross Abstract: Cellular differentiation is governed by gene regulatory networks, the high-dimensional stochastic biochemical systems that determine the transcription

Learning Structure, Energy, and Dynamics: A Survey of Artificial Intelligence for Protein Dynamics

ResearchDGX agent

arXiv:2604.25244v1 Announce Type: cross Abstract: Protein dynamics underlie many biological functions, yet remain difficult to characterize due to the high computational cost of molecular dynamics sim

Learning with Embedded Linear Equality Constraints via Variational Bayesian Inference

ResearchDGX agent

arXiv:2604.24911v1 Announce Type: new Abstract: Machine Learning is becoming more prevalent in science and engineering, but many approaches do not provide meaningful uncertainty estimates and predicti

Liquid Neural Network Models for Natural Gas Spot Price Time-Series Forecasting

Model ReleasesDGX agent

arXiv:2604.24788v1 Announce Type: new Abstract: Natural gas is undoubtedly an essential component of the global energy system. Accurate short-term forecasting of natural gas price is challenging due t

Making AI-Assisted Grant Evaluation Auditable without Exposing the Model

ResearchDGX agent

arXiv:2604.25200v1 Announce Type: cross Abstract: Public agencies are beginning to consider large language models (LLMs) as decision-support tools for grant evaluation. This creates a practical govern

Measuring the Sensitivity of Classification Models with the Error Sensitivity Profile

ResearchDGX agent

arXiv:2604.25765v1 Announce Type: new Abstract: The quality of training data is critical to the performance of machine learning models. In this paper, the Error Sensitivity Profile (ESP) is proposed.

Measuring the stability and plasticity of recommender systems

ResearchDGX agent

arXiv:2508.03941v3 Announce Type: replace-cross Abstract: The typical offline protocol to evaluate recommendation algorithms is to collect a dataset of user-item interactions and then use a part of th

minAction.net: Energy-First Neural Architecture Design -- From Biological Principles to Systematic Validation

Model ReleasesDGX agent

arXiv:2604.24805v1 Announce Type: new Abstract: Modern machine learning optimizes for accuracy without explicitly accounting for internal computational cost, even though physical and biological system

Minimax Generalized Cross-Entropy

Model ReleasesDGX agent

arXiv:2603.19874v3 Announce Type: replace-cross Abstract: Loss functions play a central role in supervised classification. Cross-entropy (CE) is widely used, whereas the mean absolute error (MAE) loss

MobileLLM-Flash: Latency-Guided On-Device LLM Design for Industry Scale Deployment

Local AiDGX agent

arXiv:2603.15954v2 Announce Type: replace Abstract: Real-time AI experiences call for on-device large language models (OD-LLMs) optimized for efficient deployment on resource-constrained hardware. The

Monitoring exposure-length variations in submarine power cables using distributed fiber-optic sensing

ResearchDGX agent

arXiv:2604.24880v1 Announce Type: cross Abstract: This study proposes an anomaly-detection framework for monitoring exposure-length variations in submarine free-span cables using Distributed Acoustic

MotionBricks: Scalable Real-Time Motions with Modular Latent Generative Model and Smart Primitives

ApplicationsDGX agent

arXiv:2604.24833v1 Announce Type: cross Abstract: Despite transformative advances in generative motion synthesis, real-time interactive motion control remains dominated by traditional techniques. In t

Multi-layer Cross-Attention is Provably Optimal for Multi-modal In-context Learning

ResearchDGX agent

arXiv:2602.04872v2 Announce Type: replace-cross Abstract: Recent progress has rapidly advanced our understanding of the mechanisms underlying in-context learning in modern attention-based neural netwo

Nautile-370M: Spectral Memory Meets Attention in a Small Reasoning Model

Model ReleasesDGX agent

arXiv:2604.24809v1 Announce Type: new Abstract: We present Nautile-370M, a 371-million-parameter small language model designed for efficient reasoning under strict parameter and inference budgets. Nau

Near-Optimal Sample Complexities of Divergence-based S-rectangular Distributionally Robust Reinforcement Learning

ApplicationsDGX agent

arXiv:2505.12202v3 Announce Type: replace Abstract: Distributionally robust reinforcement learning (DR-RL) has recently gained significant attention as a principled approach that addresses discrepanci

Negative Ontology of True Target for Machine Learning: Towards Evaluation and Learning under Democratic Supervision

ApplicationsDGX agent

arXiv:2604.24824v1 Announce Type: new Abstract: This article philosophically examines how shifts in assumptions regarding the existence and non-existence of the true target (TT) give rise to new persp

NUBO: A Transparent Python Package for Bayesian Optimization

Model ReleasesDGX agent

arXiv:2305.06709v4 Announce Type: replace Abstract: NUBO, short for Newcastle University Bayesian Optimisation, is a Bayesian optimization framework for the optimization of expensive-to-evaluate black

Null Measurability at the Symmetrization Interface in VC Learning

ResearchDGX agent

arXiv:2604.25028v1 Announce Type: new Abstract: Recent work revisiting measurability in the fundamental theorem of statistical learning imposes Borel measurability of ghost-gap suprema. We show that,

On Halting vs Converging in Recurrent Graph Neural Networks

Local AiDGX agent

arXiv:2604.25551v1 Announce Type: new Abstract: Recurrent Graph Neural Networks (RGNNs) extend standard GNNs by iterating message-passing until some stopping condition is met. Various RGNN models have

On quantitative Laplace-type convergence results for some exponential probability measures, with two applications

ResearchDGX agent

arXiv:2110.12922v2 Announce Type: replace-cross Abstract: Laplace-type results characterize the limit of sequence of measures (pi_arepsilon)_{arepsilon >0} with density w.r.t the Lebesgue measure (d p

On the Trainability of Masked Diffusion Language Models via Blockwise Locality

Local AiDGX agent

arXiv:2604.24832v1 Announce Type: new Abstract: Masked diffusion language models (MDMs) have recently emerged as a promising alternative to standard autoregressive large language models (AR-LLMs), yet

Online combinatorial optimization with stochastic decision sets and adversarial losses

ResearchDGX agent

arXiv:2604.25269v1 Announce Type: new Abstract: Most work on sequential learning assumes a fixed set of actions that are available all the time. However, in practice, actions can consist of picking su

Online learning with Erdos-Renyi side-observation graphs

ResearchDGX agent

arXiv:2604.25271v1 Announce Type: cross Abstract: We consider adversarial multi-armed bandit problems where the learner is allowed to observe losses of a number of arms beside the arm that it actually

Optimization-Free Topological Sort for Causal Discovery via the Schur Complement of Score Jacobians

ResearchDGX agent

arXiv:2604.25295v1 Announce Type: new Abstract: Continuous causal discovery typically couples representation learning with structural optimization via non-convex acyclicity penalties, which subjects s

Origin-Destination Demand Prediction: An Urban Radiation and Attraction Perspective

Model ReleasesDGX agent

arXiv:2412.00167v2 Announce Type: replace Abstract: In recent years, origin-destination (OD) demand prediction has gained significant attention for its profound implications in urban development. Exis

Physics-Guided Tiny-Mamba Transformer for Reliability-Aware Early Fault Warning

Local AiDGX agent

arXiv:2601.21293v2 Announce Type: replace Abstract: Reliability-centered prognostics for rotating machinery requires early-warning signals that remain accurate under nonstationary operating conditions

Pimp My LLM: Leveraging Variability Modeling to Tune Inference Hyperparameters

ResearchDGX agent

arXiv:2602.17697v2 Announce Type: replace Abstract: Large Language Models (LLMs) are being increasingly used across a wide range of tasks. However, their substantial computational demands raise concer

PINNs in More General Geometry

ResearchDGX agent

arXiv:2604.25020v1 Announce Type: cross Abstract: Neural architectures trained with losses inspired by differential conditions are the basis for PINN models. Since many constructions in differential g

PLMGH: What Matters in PLM-GNN Hybrids for Code Classification and Vulnerability Detection

ResearchDGX agent

arXiv:2604.25599v1 Announce Type: cross Abstract: Code understanding models increasingly rely on pretrained language models (PLMs) and graph neural networks (GNNs), which capture complementary semanti

Policy Improvement Reinforcement Learning

SafetyDGX agent

arXiv:2604.00860v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a central post-training paradigm for improving the reasoning capabilities of large

Prior-Aligned Data Cleaning for Tabular Foundation Models

Model ReleasesDGX agent

arXiv:2604.25154v1 Announce Type: new Abstract: Tabular Foundation Models (TFMs) achieve state-of-the-art zero-shot accuracy on small tabular datasets by meta-learning over synthetic data-generating p

Provable Accelerated Bayesian Optimization with Knowledge Transfer

TutorialsDGX agent

arXiv:2511.03125v2 Announce Type: replace-cross Abstract: We study how to accelerate Bayesian optimization (BO) on a target task by transferring historical knowledge from related source tasks. Existin

QFlash: Bridging Quantization and Memory Efficiency in Vision Transformer Attention

ResearchDGX agent

arXiv:2604.25306v1 Announce Type: new Abstract: FlashAttention improves efficiency through tiling, but its online softmax still relies on floating-point arithmetic for numerical stability, making full

Quantum Dynamics via Score Matching on Bohmian Trajectories

ResearchDGX agent

arXiv:2604.25137v1 Announce Type: cross Abstract: We solve the time-dependent Schrodinger equation by learning the score function, the gradient of the log-probability density, on Bohmian trajectories.

Query-Efficient Quantum Approximate Optimization via Graph-Conditioned Trust Regions

Model ReleasesDGX agent

arXiv:2604.24803v1 Announce Type: new Abstract: In low-depth implementations of the Quantum Approximate Optimization Algorithm (QAOA), the dominant cost is often the number of objective evaluations ra

RCProb: Probabilistic Rule Extraction for Efficient Simplification of Tree Ensembles

Model ReleasesDGX agent

arXiv:2604.25304v1 Announce Type: new Abstract: Tree ensembles are widely used in industrial machine learning due to their strong predictive performance and efficient training procedures. However, as

Reinforcement Learning for Testing Interdependent Requirements in Autonomous Vehicles: An Empirical Study

SafetyDGX agent

arXiv:2502.15792v2 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) make driving decisions without humans, making dependability assurance critical. Scenario-based testing is widely use

Reinforcement Learning Using known Invariances

ApplicationsDGX agent

arXiv:2511.03473v2 Announce Type: replace Abstract: In many real-world reinforcement learning (RL) problems, the environment exhibits inherent symmetries that can be exploited to improve learning effi

Relational In-Context Learning via Synthetic Pre-training with Structural Prior

ApplicationsDGX agent

arXiv:2603.03805v2 Announce Type: replace Abstract: Relational Databases (RDBs) are the backbone of modern business, yet they lack foundation models comparable to those in text or vision. A key obstac

Residual-loss Anomaly Analysis of Physics-Informed Neural Networks: An Inverse Method for Change-point Detection in Nonlinear Dynamical Systems with Regime Switching

Model ReleasesDGX agent

arXiv:2604.25655v1 Announce Type: cross Abstract: Nonlinear dynamical systems with regime transitions are typically described by ordinary differential equations with jumping parameters parameters. Tra

Rethinking Efficiency in Neural Combinatorial Optimization: Batched Preference Optimization with Mamba

Local AiDGX agent

arXiv:2602.20730v2 Announce Type: replace Abstract: We study efficiency as a first-class objective in Neural Combinatorial Optimization (NCO) and present ECO, an efficient learning framework that comb

Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective

SafetyDGX agent

arXiv:2510.10150v3 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) serves as a cornerstone technique for enhancing the reasoning capabilities of Large Language M

Revisiting the Past: Data Unlearning with Model State History

ResearchDGX agent

arXiv:2506.20941v3 Announce Type: replace Abstract: Large language models are trained on massive corpora of web data, which may include private data, copyrighted material, factually inaccurate data, o

Safe-Support Q-Learning: Learning without Unsafe Exploration

SafetyDGX agent

arXiv:2604.25379v1 Announce Type: new Abstract: Ensuring safety during reinforcement learning (RL) training is critical in real-world applications where unsafe exploration can lead to devastating outc

Sharp Capacity Scaling of Spectral Optimizers in Learning Associative Memory

ResearchDGX agent

arXiv:2603.26554v2 Announce Type: replace Abstract: Spectral optimizers such as Muon have recently shown strong empirical performance in large-scale language model training, but the source and extent

Sharp Risk Bounds for Early-Stopping in Gaussian Linear Regression

ResearchDGX agent

arXiv:2503.03426v2 Announce Type: replace Abstract: We study early-stopped mirror descent (ESMD) for high-dimensional Gaussian linear regression over arbitrary convex bodies and design matrices, where

Shearlet Neural Operators for Anisotropic-Shock-Dominated and Multi-scale parametric partial differential equations

Model ReleasesDGX agent

arXiv:2604.25181v1 Announce Type: new Abstract: Neural operators have emerged as powerful data-driven surrogates for learning solution operators of parametric partial differential equations (PDEs). Ho

Spark Policy Toolkit: Semantic Contracts and Scalable Execution for Policy Learning in Spark

SafetyDGX agent

arXiv:2604.25061v1 Announce Type: cross Abstract: Custom policy-learning pipelines in Spark fail for two coupled systems reasons: rowwise Python execution makes inference impractical, and driver-side

Spectral bandits

SafetyDGX agent

arXiv:2604.25272v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this work, we study a bandit problem where the payoffs

← Previous
1…201202203204205…241
Next →