{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/neural-lattice-language-models","title":"Neural Lattice Language Models","arxiv_id":"1803.05071","date":"2018-03-13","proceeding":"TACL 2018 1","authors":["Jacob Buckman","Graham Neubig"],"abstract":"In this work, we propose a new language modeling paradigm that has the\nability to perform both prediction and moderation of information flow at\nmultiple granularities: neural lattice language models. These models construct\na lattice of possible paths through a sentence and marginalize across this\nlattice to calculate sequence probabilities or optimize parameters. This\napproach allows us to seamlessly incorporate linguistic intuitions - including\npolysemy and existence of multi-word lexical items - into our language model.\nExperiments on multiple language modeling tasks show that English neural\nlattice language models that utilize polysemous embeddings are able to improve\nperplexity by 9.95% relative to a word-level baseline, and that a Chinese model\nthat handles multi-character tokens is able to improve perplexity by 20.94%\nrelative to a character-level baseline.","url_abs":"http://arxiv.org/abs/1803.05071v1","url_pdf":"http://arxiv.org/pdf/1803.05071v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"neural-lattice-language-models","repo_url":"https://github.com/jbuckman/neural-lattice-language-models","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":0,"framework":"none","reach":null}],"tasks":[{"task_slug":"language-modeling","task_name":"Language Modeling"},{"task_slug":"language-modelling","task_name":"Language Modelling"},{"task_slug":"sentence","task_name":"Sentence"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=1803.05071","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}