AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-lg”

GridTimelineEvolution
14,432 results
10 Apr 2026

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

SafetyDGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

LNN-PINN: A Unified Physics-Only Training Framework with Liquid Residual Blocks

Model ReleasesDGX agent

arXiv:2508.08935v4 Announce Type: replace Abstract: Physics-informed neural networks (PINNs) have attracted considerable attention for their ability to integrate partial differential equation priors i

LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2509.09926v5 Announce Type: replace Abstract: Long-tailed semi-supervised learning (LTSSL) presents a formidable challenge where models must overcome the scarcity of tail samples while mitigatin

Low-Rank Key Value Attention

ResearchDGX agent

arXiv:2601.11471v3 Announce Type: replace Abstract: The key-value (KV) cache is a primary memory bottleneck in Transformers. We propose Low-Rank Key-Value (LRKV) attention, which reduces KV cache memo

Lumbermark: Resistant Clustering by Chopping Up Mutual Reachability Minimum Spanning Trees

Model ReleasesDGX agent

arXiv:2604.07143v1 Announce Type: new Abstract: We introduce Lumbermark, a robust divisive clustering algorithm capable of detecting clusters of varying sizes, densities, and shapes. Lumbermark iterat

LUMINA: Foundation Models for Topology Transferable ACOPF

SafetyDGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

Matrix Profile for Time-Series Anomaly Detection: A Reproducible Open-Source Benchmark on TSB-AD

Model ReleasesDGX agent

arXiv:2604.02445v2 Announce Type: replace Abstract: Matrix Profile (MP) methods are an interpretable and scalable family of distance-based methods for time-series anomaly detection, but strong benchma

MDP modeling for multi-stage stochastic programs

SafetyDGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

Measurement of Generative AI Workload Power Profiles for Whole-Facility Data Center Infrastructure Planning

HardwareDGX agent

arXiv:2604.07345v1 Announce Type: cross Abstract: The rapid growth of generative artificial intelligence (AI) has introduced unprecedented computational demands, driving significant increases in the e

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

AgentsDGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

MENO: MeanFlow-Enhanced Neural Operators for Dynamical Systems

ResearchDGX agent

arXiv:2604.06881v1 Announce Type: new Abstract: Neural operators have emerged as powerful surrogates for dynamical systems due to their grid-invariant properties and computational efficiency. However,

MF-GLaM: A multifidelity stochastic emulator using generalized lambda models

Model ReleasesDGX agent

arXiv:2507.10303v2 Announce Type: replace-cross Abstract: Stochastic simulators exhibit intrinsic stochasticity due to unobservable, uncontrollable, or unmodeled input variables, resulting in random o

MICA: Multivariate Infini Compressive Attention for Time Series Forecasting

ResearchDGX agent

arXiv:2604.06473v1 Announce Type: new Abstract: Multivariate forecasting with Transformers faces a core scalability challenge: modeling cross-channel dependencies via attention compounds attention's q

Mining Electronic Health Records to Investigate Effectiveness of Ensemble Deep Clustering

ApplicationsDGX agent

arXiv:2604.07085v1 Announce Type: new Abstract: In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical

MoE Routing Testbed: Studying Expert Specialization and Routing Behavior at Small Scale

ResearchDGX agent

arXiv:2604.07030v1 Announce Type: new Abstract: Sparse Mixture-of-Experts (MoE) architectures are increasingly popular for frontier large language models (LLM) but they introduce training challenges d

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

SafetyDGX agent

arXiv:2604.07148v1 Announce Type: new Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) ad

NativeTernary: A Self-Delimiting Binary Encoding with Unary Run-Length Hierarchy Markers for Ternary Neural Network Weights, Structured Data, and General Computing Infrastructure

Local AiDGX agent

arXiv:2604.03336v2 Announce Type: replace Abstract: BitNet b1.58 (Ma et al., 2024) demonstrates that large language models can operate entirely on ternary weights {-1, 0, +1}, yet no native binary wir

Negative Binomial Variational Autoencoders for Overdispersed Latent Modeling

Model ReleasesDGX agent

arXiv:2508.05423v2 Announce Type: replace Abstract: Although artificial neural networks are often described as brain-inspired, their representations typically rely on continuous activations, such as t

NestPipe: Large-Scale Recommendation Training on 1,500+ Accelerators via Nested Pipelining

HardwareDGX agent

arXiv:2604.06956v1 Announce Type: cross Abstract: Modern recommendation models have increased to trillions of parameters. As cluster scales expand to O(1k), distributed training bottlenecks shift from

Neural parametric representations for thin-shell shape optimisation

Model ReleasesDGX agent

arXiv:2604.06612v1 Announce Type: cross Abstract: Shape optimisation of thin-shell structures requires a flexible, differentiable geometric representation suitable for gradient-based optimisation. We

Neural Two-Stage Stochastic Optimization for Solving Unit Commitment Problem

ResearchDGX agent

arXiv:2507.09503v2 Announce Type: replace-cross Abstract: This paper proposes a neural stochastic optimization method for efficiently solving the two-stage stochastic unit commitment (2S-SUC) problem

Non-Expansive Mappings in Two-Time-Scale Stochastic Approximation: Finite-Time Analysis

ResearchDGX agent

arXiv:2501.10806v4 Announce Type: replace-cross Abstract: Two-time-scale stochastic approximation algorithms are iterative methods used in applications such as optimization, reinforcement learning, an

Non-identifiability of Explanations from Model Behavior in Deep Networks of Image Authenticity Judgments

ResearchDGX agent

arXiv:2604.07254v1 Announce Type: cross Abstract: Deep neural networks can predict human judgments, but this does not imply that they rely on human-like information or reveal the cues underlying those

Nonparametric Instrumental Regression via Kernel Methods is Minimax Optimal

ResearchDGX agent

arXiv:2411.19653v2 Announce Type: replace-cross Abstract: We study the kernel instrumental variable (KIV) algorithm, a kernel-based two-stage least-squares method for nonparametric instrumental variab

ODE-free Neural Flow Matching for One-Step Generative Modeling

TutorialsDGX agent

arXiv:2604.06413v1 Announce Type: new Abstract: Diffusion and flow matching models generate samples by learning time-dependent vector fields whose integration transports noise to data, requiring tens

On the Price of Privacy for Language Identification and Generation

ResearchDGX agent

arXiv:2604.07238v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly trained on sensitive user data, understanding the fundamental cost of privacy in language learning beco

Operator Learning for Surrogate Modeling of Wave-Induced Forces from Sea Surface Waves

ResearchDGX agent

arXiv:2604.06433v1 Announce Type: cross Abstract: Wave setup plays a significant role in transferring wave-induced energy to currents and causing an increase in water elevation. This excess momentum f

Optimal Rates for Pure {arepsilon}-Differentially Private Stochastic Convex Optimization with Heavy Tails

Model ReleasesDGX agent

arXiv:2604.06492v1 Announce Type: new Abstract: We study stochastic convex optimization (SCO) with heavy-tailed gradients under pure epsilon-differential privacy (DP). Instead of assuming a bound on t

PAC-Bayesian Bounds on Constrained f-Entropic Risk Measures

ResearchDGX agent

arXiv:2510.11169v2 Announce Type: replace-cross Abstract: PAC generalization bounds on the risk, when expressed in terms of the expected loss, are often insufficient to capture imbalances between subg

PD-SOVNet: A Physics-Driven Second-Order Vibration Operator Network for Estimating Wheel Polygonal Roughness from Axle-Box Vibrations

ApplicationsDGX agent

arXiv:2604.06620v1 Announce Type: new Abstract: Quantitative estimation of wheel polygonal roughness from axle-box vibration signals is a challenging yet practically relevant problem for rail-vehicle

Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalization

Model ReleasesDGX agent

arXiv:2604.07343v1 Announce Type: cross Abstract: Pluralistic alignment has emerged as a critical frontier in the development of Large Language Models (LLMs), with reward models (RMs) serving as a cen

Physics-Informed Functional Link Constrained Framework with Domain Mapping for Solving Bending Analysis of an Exponentially Loaded Perforated Beam

Model ReleasesDGX agent

arXiv:2604.07025v1 Announce Type: cross Abstract: This article presents a novel and comprehensive approach for analyzing bending behavior of the tapered perforated beam under an exponential load. The

Physics-Informed Neural Networks for Joint Source and Parameter Estimation in Advection-Diffusion Equations

Model ReleasesDGX agent

arXiv:2512.07755v3 Announce Type: replace-cross Abstract: Recent studies have demonstrated the success of deep learning in solving forward and inverse problems in engineering and scientific computing

Predictive Representations for Skill Transfer in Reinforcement Learning

AgentsDGX agent

arXiv:2604.07016v1 Announce Type: new Abstract: A key challenge in scaling up Reinforcement Learning is generalizing learned behaviour. Without the ability to carry forward acquired knowledge an agent

Probabilistic Predictions of Process-Induced Deformation in Carbon/Epoxy Composites Using a Deep Operator Network

ApplicationsDGX agent

arXiv:2512.13746v4 Announce Type: replace-cross Abstract: Fiber reinforcement and polymer matrix respond differently to manufacturing conditions due to mismatch in coefficient of thermal expansion and

Production-Ready Automated ECU Calibration using Residual Reinforcement Learning

ApplicationsDGX agent

arXiv:2604.07059v1 Announce Type: new Abstract: Electronic Control Units (ECUs) have played a pivotal role in transforming motorcars of yore into the modern vehicles we see on our roads today. They ac

QNAS: A Neural Architecture Search Framework for Accurate and Efficient Quantum Neural Networks

Model ReleasesDGX agent

arXiv:2604.07013v1 Announce Type: cross Abstract: Designing quantum neural networks (QNNs) that are both accurate and deployable on NISQ hardware is challenging. Handcrafted ansatze must balance expre

Quality-preserving Model for Electronics Production Quality Tests Reduction

SafetyDGX agent

arXiv:2604.06451v1 Announce Type: new Abstract: Manufacturing test flows in high-volume electronics production are typically fixed during product development and executed unchanged on every unit, even

Quantum-Inspired Tensor Network Autoencoders for Anomaly Detection: A MERA-Based Approach

Model ReleasesDGX agent

arXiv:2604.06541v1 Announce Type: cross Abstract: We investigate whether a multiscale tensor-network architecture can provide a useful inductive bias for reconstruction-based anomaly detection in coll

RAGEN-2: Reasoning Collapse in Agentic RL

AgentsDGX agent

arXiv:2604.06268v1 Announce Type: new Abstract: RL training of multi-turn LLM agents is inherently unstable, and reasoning quality directly determines task performance. Entropy is widely used to track

ReDAct: Uncertainty-Aware Deferral for LLM Agents

AgentsDGX agent

arXiv:2604.07036v1 Announce Type: cross Abstract: Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However,

Replacing Tunable Parameters in Weather and Climate Models with State-Dependent Functions using Reinforcement Learning

Model ReleasesDGX agent

arXiv:2601.04268v2 Announce Type: replace Abstract: Weather and climate models rely on parametrisations to represent unresolved sub-grid processes. Traditional schemes rely on fixed coefficients that

Resistance Distance and Linearized Optimal Transport on Graphs

ResearchDGX agent

arXiv:2404.15261v4 Announce Type: replace-cross Abstract: We study the linearization of a discrete transportation distance between probability distributions on finite weighted graphs originally due to

Revisiting Fairness Impossibility with Endogenous Behavior

SafetyDGX agent

arXiv:2604.06378v1 Announce Type: cross Abstract: In many real-world settings, institutions can and do adjust the consequences attached to algorithmic classification decisions, such as the size of fin

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

HardwareDGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

Robust support vector model based on bounded asymmetric elastic net loss for binary classification

Model ReleasesDGX agent

arXiv:2603.06257v2 Announce Type: replace-cross Abstract: In this paper, we propose a novel bounded asymmetric elastic net (L_{baen}) loss function and combine it with the support vector machine (SV

SBBTS: A Unified Schrodinger-Bass Framework for Synthetic Financial Time Series

ResearchDGX agent

arXiv:2604.07159v1 Announce Type: new Abstract: We study the problem of generating synthetic time series that reproduce both marginal distributions and temporal dynamics, a central challenge in financ

Selective Neuron Amplification for Training-Free Task Enhancement

ResearchDGX agent

arXiv:2604.07098v1 Announce Type: new Abstract: Large language models often fail on tasks they seem to already understand. In our experiments, this appears to be less about missing knowledge and more

Self-Distilled RLVR

SafetyDGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

Shapes are not enough: CONSERVAttack and its use for finding vulnerabilities and uncertainties in machine learning applications

ResearchDGX agent

arXiv:2603.13970v2 Announce Type: replace Abstract: In High Energy Physics, as in many other fields of science, the application of machine learning techniques has been crucial in advancing our underst

SL-FAC: A Communication-Efficient Split Learning Framework with Frequency-Aware Compression

ResearchDGX agent

arXiv:2604.07316v1 Announce Type: new Abstract: The growing complexity of neural networks hinders the deployment of distributed machine learning on resource-constrained devices. Split learning (SL) of

Smart Commander: A Hierarchical Reinforcement Learning Framework for Fleet-Level PHM Decision Optimization

ResearchDGX agent

arXiv:2604.07171v1 Announce Type: new Abstract: Decision-making in military aviation Prognostics and Health Management (PHM) faces significant challenges due to the 'curse of dimensionality' in large-

Smoothing the Edges: Smooth Optimization for Sparse Regularization using Hadamard Overparametrization

ResearchDGX agent

arXiv:2307.03571v4 Announce Type: replace Abstract: We present a framework for smooth optimization of explicitly regularized objectives for (structured) sparsity. These non-smooth and possibly non-con

SMT-AD: a scalable quantum-inspired anomaly detection approach

ResearchDGX agent

arXiv:2604.06265v1 Announce Type: new Abstract: Quantum-inspired tensor networks algorithms have shown to be effective and efficient models for machine learning tasks, including anomaly detection. Her

Spatiotemporal Gaussian representation-based dynamic reconstruction and motion estimation framework for time-resolved volumetric MR imaging (DREME-GSMR)

ResearchDGX agent

arXiv:2604.06482v1 Announce Type: cross Abstract: Time-resolved volumetric MR imaging that reconstructs a 3D MRI within sub-seconds to resolve deformable motion is essential for motion-adaptive radiot

Spike-based alignment learning solves the weight transport problem

SafetyDGX agent

arXiv:2503.02642v3 Announce Type: replace-cross Abstract: In both machine learning and in computational neuroscience, plasticity in functional neural networks is frequently expressed as gradient desce

Splats under Pressure: Exploring Performance-Energy Trade-offs in Real-Time 3D Gaussian Splatting under Constrained GPU Budgets

HardwareDGX agent

arXiv:2604.07177v1 Announce Type: cross Abstract: We investigate the feasibility of real-time 3D Gaussian Splatting (3DGS) rasterisation on edge clients with varying Gaussian splat counts and GPU comp

SSPINNpose: A Self-Supervised PINN for Inertial Pose and Dynamics Estimation

ApplicationsDGX agent

arXiv:2506.11786v2 Announce Type: replace Abstract: Accurate real-time estimation of human movement dynamics, including internal joint moments and muscle forces, is essential for applications in clini

Stochastic Auto-conditioned Fast Gradient Methods with Optimal Rates

Model ReleasesDGX agent

arXiv:2604.06525v1 Announce Type: cross Abstract: Achieving optimal rates for stochastic composite convex optimization without prior knowledge of problem parameters remains a central challenge. In the

Stochastic Gradient Descent in the Saddle-to-Saddle Regime of Deep Linear Networks

ResearchDGX agent

arXiv:2604.06366v1 Announce Type: new Abstract: Deep linear networks (DLNs) are used as an analytically tractable model of the training dynamics of deep neural networks. While gradient descent in DLNs

← Previous
1…238239240241
Next →