DUET: Optimizing Training Data Mixtures via Feedback from Unseen Evaluation Tasks
DGX agentarXiv:2502.00270v3 Announce Type: replace-cross Abstract: The performance of an LLM depends heavily on the relevance of its training data to the downstream evaluation task. However, in practice, the d