{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/a-novel-discourse-parser-based-on-support","title":"A Novel Discourse Parser Based on Support Vector Machine Classification","arxiv_id":null,"date":"2009-08-02","proceeding":null,"authors":["David duVerle","Helmut Prendinger"],"abstract":"This paper introduces a new algorithm to parse discourse within the framework of Rhetorical Structure Theory (RST). Our method is based on recent advances in the field of statistical machine learning (multivariate capabilities of Support Vector Machines) and a rich feature space. RST offers a formal framework for hierarchical text organization with strong applications in discourse analysis and text generation. We demonstrate automated annotation of a text with RST hierarchically organised relations, with results comparable to those achieved by specially trained human annotators. Using a rich set of shallow lexical, syntactic and structural features from the input text, our parser achieves, in linear time, 73.9% of professional annotators’ human agreement F-score. The parser is 5% to 12% more accurate than current state-of-the-art parsers.","url_abs":"https://aclanthology.org/P09-1075","url_pdf":"https://aclanthology.org/P09-1075.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"classification-1","task_name":"Classification"},{"task_slug":"discourse-parsing","task_name":"Discourse Parsing"},{"task_slug":"classification","task_name":"General Classification"},{"task_slug":"text-generation","task_name":"Text Generation"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/discourse-parsing-on-rst-dt","task":"Discourse Parsing","dataset":"RST-DT","model":"HILDA Parser","rank_in_archive_order":30,"of":40,"metrics":{"RST-Parseval (Full)":"54.8","RST-Parseval (Nuclearity)":"68.4","RST-Parseval (Relation)":"55.3","RST-Parseval (Span)":"83.0"},"uses_additional_data":false}],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}