Can Bayesian Optimization Efficiently Find a Strong Single Expert in Neural Thickets?
arXiv:2608.10867v1 Announce Type: new Abstract: Gradient-free post-training has emerged as a compelling alternative to gradient-based optimization for large language models (LLMs), but existing approa