{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/what-did-you-mention-a-large-scale-mention","title":"What did you Mention? A Large Scale Mention Detection Benchmark for Spoken and Written Text","arxiv_id":"1801.07507","date":"2018-01-23","proceeding":null,"authors":["Yosi Mass","Lili Kotlerman","Shachar Mirkin","Elad Venezian","Gera Witzling","Noam Slonim"],"abstract":"We describe a large, high-quality benchmark for the evaluation of Mention\nDetection tools. The benchmark contains annotations of both named entities as\nwell as other types of entities, annotated on different types of text, ranging\nfrom clean text taken from Wikipedia, to noisy spoken data. The benchmark was\nbuilt through a highly controlled crowd sourcing process to ensure its quality.\nWe describe the benchmark, the process and the guidelines that were used to\nbuild it. We then demonstrate the results of a state-of-the-art system running\non that benchmark.","url_abs":"http://arxiv.org/abs/1801.07507v3","url_pdf":"http://arxiv.org/pdf/1801.07507v3.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[{"slug":"ibm-debater-mention-detection-benchmark","name":"IBM Debater Mention Detection Benchmark","full_name":""}],"methods_introduced":[],"results":[],"syntology":{"syntology_url":null,"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}