{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/from-word-segmentation-to-pos-tagging-for","title":"From Word Segmentation to POS Tagging for Vietnamese","arxiv_id":"1711.04951","date":"2017-11-14","proceeding":"ALTA 2017 12","authors":["Dat Quoc Nguyen","Thanh Vu","Dai Quoc Nguyen","Mark Dras","Mark Johnson"],"abstract":"This paper presents an empirical comparison of two strategies for Vietnamese\nPart-of-Speech (POS) tagging from unsegmented text: (i) a pipeline strategy\nwhere we consider the output of a word segmenter as the input of a POS tagger,\nand (ii) a joint strategy where we predict a combined segmentation and POS tag\nfor each syllable. We also make a comparison between state-of-the-art (SOTA)\nfeature-based and neural network-based models. On the benchmark Vietnamese\ntreebank (Nguyen et al., 2009), experimental results show that the pipeline\nstrategy produces better scores of POS tagging from unsegmented text than the\njoint strategy, and the highest accuracy is obtained by using a feature-based\nmodel.","url_abs":"http://arxiv.org/abs/1711.04951v1","url_pdf":"http://arxiv.org/pdf/1711.04951v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"from-word-segmentation-to-pos-tagging-for","repo_url":"https://github.com/datquocnguyen/VnMarMoT","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[{"task_slug":"pos","task_name":"POS"},{"task_slug":"pos-tagging","task_name":"POS Tagging"},{"task_slug":"part-of-speech-tagging","task_name":"Part-Of-Speech Tagging"},{"task_slug":"tag","task_name":"TAG"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"syntology_url":"https://syntology.ai/paper/1711.04951","atlas_url":"https://app.syntology.ai/?focus=1711.04951","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}