Jan 06, 2026
Contrastive ESA: Human Evaluation of Multiple Translations at Once
Current human evaluation of machine translation typically assesses single outputs in isolation, a paradigm that suffers from high annotator noise and cost. We introduce Contrastive Error Span Annotation (cESA), a protocol that presents multiple translations of the source input (text, video, audio, image).
Authors
Vilem Zouhar, Roman Grundkiewicz, Sara Rajaee, Parker Riley, Rachel Bawden, Philipp Koehn, Marine Carpuat and Tom Kocmi
Abstract
Related works









