{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/jw300-a-wide-coverage-parallel-corpus-for-low","title":"JW300: A Wide-Coverage Parallel Corpus for Low-Resource Languages","arxiv_id":null,"date":"2019-07-01","proceeding":"ACL 2019 7","authors":["{\\v{Z}}eljko Agi{\\'c}","Ivan Vuli{\\'c}"],"abstract":"Viable cross-lingual transfer critically depends on the availability of parallel texts. Shortage of such resources imposes a development and evaluation bottleneck in multilingual processing. We introduce JW300, a parallel corpus of over 300 languages with around 100 thousand parallel sentences per language pair on average. In this paper, we present the resource and showcase its utility in experiments with cross-lingual word embedding induction and multi-source part-of-speech projection.","url_abs":"https://aclanthology.org/P19-1310","url_pdf":"https://aclanthology.org/P19-1310.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"cross-lingual-transfer","task_name":"Cross-Lingual Transfer"}],"methods":[],"datasets_introduced":[{"slug":"jw300","name":"JW300","full_name":""}],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}