Jun 06, 2026
Findings of Automated Translation Quality Evaluation
We present the findings of the WMT26 Shared Task on Automated Translation Quality Evaluation Systems, continuing last year’s unification of the earlier separate WMT Metrics and Quality Estimation shared tasks
Authors
Alon Lavie, Greg Hanneman, Stefano Perrella, Shuoyang Ding, Eleftherios Avramidis, Lorenzo Proietti, Chi-kiu Lo 羅致翹, Ammon Shurtz, Chrysoula Zerva, Archchana Sindhujan, Vilém Zouhar, Diptesh Kanojia, Frédéric Blain, Brian Thompson, Giorgos Filandrianos, Orfeas Menis Mastromichalakis, Tom Kocmi, Pranav Gupta
Abstract
Related works

Research
BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluation
Read

Research
Findings of the WMT26 General Machine Translation Shared Task: Contrastive Dynamic Human Evaluation at Scale
Read

Research
The IOL-AI Challenge: An Open Challenge towards Advancing Linguistic Reasoning
Read






