Contrastive ESA: Human Evaluation of Multiple Translations at Once
DGX agentarXiv:2607.26640v1 Announce Type: new Abstract: Current human evaluation of machine translation typically assesses single outputs in isolation, a paradigm that suffers from high annotator noise and co