Browse State-of-the-Art › Argument Mining
Argument Mining
98 papers with code · 1 benchmark · 7 datasets archive 2025-07-28
Argument Mining is a field of corpus-based discourse analysis that involves the automatic identification of argumentative structures in text.
Source: AMPERSAND: Argument Mining for PERSuAsive oNline Discussions
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| TACO -- Twitter Arguments from COnversations (1 row) | TACO | TACO -- Twitter Arguments from COnversations | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
7 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
6 subtasks in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 98 papers with code (284 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
14 Nov 2020 3 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Finally, we present a search engine for this dataset which is utilized extensively by members of the National Speech and Debate Association today.
-
17 Sep 2018 3 repositories listedWe tackle the tasks of automatically identifying comparative sentences and categorizing the intended preference (e.
-
5 Jun 2024 2 repositories listedIn our work, we introduce an argument mining dataset that captures the end-to-end process of preparing an argumentative essay for a debate, which covers the tasks of claim and evidence identification (Task 1 ED),…
-
22 Jun 2021 2 repositories listedOn DeSSE, which has a more even balance of complex sentence types, our model achieves higher accuracy on the number of atomic sentences than an encoder-decoder baseline.
-
10 Dec 2020 2 repositories listed Syntology ran 0 of 2 samples · 2 unverified · 2 pointer-only (licence)Peer reviewing is a central process in modern research and essential for ensuring high quality and reliability of published work.
-
1 Jul 2019 2 repositories listedWe address this task in an empirical manner by annotating 39 political debates from the last 50 years of US presidential campaigns, creating a new corpus of 29k argument components, labeled as premises and claims.
-
24 Jun 2019 2 repositories listedWe experiment with two recent contextualized word embedding methods (ELMo and BERT) in the context of open-domain argument search.
-
1 Nov 2018 2 repositories listedFact-checking is a journalistic practice that compares a claim made publicly against trusted sources of facts.
-
21 Jul 2025 1 repository listedIn this paper, we present our submission to the MM-ArgFallacy2025 shared task, which aims to advance research in multimodal argument mining, focusing on logical fallacies in political debates.
-
18 Dec 2024 1 repository listedTo overcome this issue, we develop a web application called "Find your Figure" that facilitates the identification and annotation of German rhetorical figures.
-
7 Oct 2024 1 repository listedIn this paper, we follow this direction, and we present, to the best of our knowledge, the first multilingual dataset for Medical Question Answering where correct and incorrect diagnoses for a clinical case are enriched…
-
29 Jul 2024 1 repository listedDialogical Argument Mining(DialAM) is an important branch of Argument Mining(AM).
-
4 Jul 2024 1 repository listedContrary to previous work, we show that for Argument Mining data transfer obtains better results than model-transfer and that fine-tuning outperforms few-shot methods.
-
10 Jun 2024 1 repository listedIn the training-free ICL setting, we show that GPT-4 is able to leverage relevant information from only a few demonstration examples and achieve very competitive classification accuracy on ATC.
-
2 May 2024 1 repository listedArgument structure learning~(ASL) entails predicting relations between arguments.
-
1 May 2024 1 repository listedFirst, we develop and release an Argument Detection model that can classify a piece of text as an argument with an F1 score between 79% and 86% on three different benchmark datasets.
-
17 Apr 2024 1 repository listed Syntology ran 8 of 13 samples · 5 unverified · 13 pointer-only (licence)Our objective is to train a generative model that can simultaneously provide a score indicating the presence of shared key point between a pair of arguments and generate the shared key point.
-
15 Apr 2024 1 repository listedIn this paper, we study the potential of using state-of-the-art large language models (LLMs) as proxies for argument quality annotators.
-
3 Apr 2024 1 repository listedWhen combined with automatic essay scoring, interactions of the argumentative structure and quality scores can be exploited for comprehensive writing support.
-
2 Apr 2024 1 repository listedWe publish LawInstruct as a resource for further study of instruction tuning in the legal domain.
-
30 Mar 2024 1 repository listedTwitter has emerged as a global hub for engaging in online conversations and as a research corpus for various disciplines that have recognized the significance of its user-generated content.
-
20 Jan 2024 1 repository listedRhetorical Structure Theory implies no single discourse interpretation of a text, and the limitations of RST parsers further exacerbate inconsistent parsing of similar structures.
-
20 Nov 2023 1 repository listedWith the increasing amount of problematic peer reviews in top AI conferences, the community is urgently in need of automatic quality control measures.
-
15 Nov 2023 1 repository listedAs large language models (LLMs) have demonstrated impressive capabilities in understanding context and generating natural language, it is worthwhile to evaluate the performance of LLMs on diverse computational…
-
14 Nov 2023 1 repository listedThe COVID-19 pandemic has sparked numerous discussions on social media platforms, with users sharing their views on topics such as mask-wearing and vaccination.
-
15 Oct 2023 1 repository listedThis paper presents an overview of the ImageArg shared task, the first multimodal Argument Mining shared task co-located with the 10th Workshop on Argument Mining at EMNLP 2023.
-
8 Oct 2023 1 repository listedA main goal of Argument Mining (AM) is to analyze an author's stance.
-
29 Sep 2023 1 repository listedWe evaluate two large language models (LLMs) ability to perform argumentative reasoning.
-
15 Sep 2023 1 repository listedOur findings challenge the previously asserted general superiority of in-context learning (ICL) for OOD.
-
9 Jul 2023 1 repository listedWe leverage the txtai semantic search and knowledge graph toolchain to produce and contribute 9 semantic knowledge graphs built on this dataset.
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections