Increasing Missingness to Reduce Bias: Richardson-SGD with Missing Data
DGX agentarXiv:2605.19641v1 Announce Type: cross Abstract: Stochastic gradient methods are central to modern large-scale learning, but their use with incomplete covariates remains delicate since imputation sch