GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization
arXiv:2605.31464v1 Announce Type: cross Abstract: GPU kernels are the workhorse of modern deep learning, and optimizing them (via evolutionary search or coding agents) usually requires repeated measur