Téléchargement | - Voir la version finale : On the stability of system rankings at WMT (PDF, 5.9 Mio)
|
---|
Lien | https://aclanthology.org/2021.wmt-1.56/ |
---|
Auteur | Rechercher : Knowles, Rebecca1 |
---|
Affiliation | - Conseil national de recherches du Canada. Technologies numériques
|
---|
Format | Texte, Article |
---|
Conférence | Sixth Conference on Machine Translation, November 10-11, 2021, Punta Cana, Dominican Republic and Online |
---|
Résumé | The current approach to collecting human judgments of machine translation quality for the news translation task at WMT – segment rating with document context – is the most recent in a sequence of changes to WMT human annotation protocol. As these annotation protocols have changed over time, they have drifted away from some of the initial statistical assumptions underpinning them, with consequences that call the validity of WMT news task system rankings into question. In simulations based on real data, we show that the rankings can be influenced by the presence of outliers (high- or low-quality systems), resulting in different system rankings and clusterings. We also examine questions of annotation task composition and how ease or difficulty of translating different documents may influence system rankings. We provide discussion of ways to analyze these issues when considering future changes to annotation protocols. |
---|
Date de publication | 2021-11 |
---|
Maison d’édition | Association for Computational Linguistics |
---|
Licence | |
---|
Dans | |
---|
Langue | anglais |
---|
Publications évaluées par des pairs | Oui |
---|
Exporter la notice | Exporter en format RIS |
---|
Signaler une correction | Signaler une correction (s'ouvre dans un nouvel onglet) |
---|
Identificateur de l’enregistrement | 1e1c385a-a88d-4f15-b7c8-b2329a7d5513 |
---|
Enregistrement créé | 2022-01-17 |
---|
Enregistrement modifié | 2022-01-17 |
---|