Datasets › WMT 2021 - Multilingual Low-Resource Translation for Indo-European Languages
WMT 2021 - Multilingual Low-Resource Translation for Indo-European Languages
The Multilingual Low-Resource Translation task for Indo-European Languages, part of the EMNLP 2021 Conference, focused on improving machine translation in the cultural heritage domain for North-Germanic and Romance languages. It aimed to explore data transferability across related languages, prioritizing low-resource languages while allowing training in high-resource languages. The task had two subtasks: translating Europeana thesis abstracts and descriptions for North-Germanic languages, and translating Wikipedia cultural heritage articles for Romance languages. The task encouraged using diverse data sources and provided additional resources like lexicons and validation sets. Evaluation was based on translation quality, emphasizing multilinguality and resource efficiency in machine translation.
Benchmarks archive 2025-07-28
No leaderboard in the archive resolves to this dataset.
Papers archive 2025-07-28
No paper in the archive has a leaderboard row on this dataset; the archive counts 1 paper for it but never published that list.
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
No task tagged in the archive.
License archive 2025-07-28
No licence recorded in the archive. Absence here is not a statement about the dataset's terms.
Modalities archive 2025-07-28
No modality tagged.
Languages archive 2025-07-28
No language tagged.
Variants archive 2025-07-28
- WMT 2021 - Multilingual Low-Resource Translation for Indo-European Languages
1 variant name, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections