Datasets › Word Sense Disambiguation: a Unified Evaluation Framework and Empirical Comparison
Word Sense Disambiguation: a Unified Evaluation Framework and Empirical Comparison
The Evaluation framework of Raganato et al. 2017 includes two training sets (SemCor-Miller et al., 1993- and OMSTI-Taghipour and Ng, 2015-) and five test sets from the Senseval/SemEval series (Edmonds and Cotton, 2001; Snyder and Palmer, 2004; Pradhan et al., 2007; Navigli et al., 2013; Moro and Navigli, 2015), standardized to the same format and sense inventory (i.e. WordNet 3.0).
Typically, there are two kinds of approach for WSD: supervised (which make use of sense-annotated training data) and knowledge-based (which make use of the properties of lexical resources).
Supervised: The most widely used training corpus used is SemCor, with 226,036 sense annotations from 352 documents manually annotated. All supervised systems in the evaluation table are trained on SemCor. Some supervised methods, particularly neural architectures, usually employ the SemEval 2007 dataset as development set (marked by *). The most usual baseline is the Most Frequent Sense (MFS) heuristic, which selects for each target word the most frequent sense in the training data.
Knowledge-based: Knowledge-based systems usually exploit WordNet or BabelNet as semantic network. The first sense given by the underlying sense inventory (i.e. WordNet 3.0) is included as a baseline.
Description from NLP Progress
Benchmarks archive 2025-07-28
All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.
| First row (archive order) | Paper | Code | ||||
|---|---|---|---|---|---|---|
| Word Sense Disambiguation | Supervised: | SANDWiCH Senseval 2 87.8 | SANDWiCH: Semantical Analysis of Neighbours for... | danielguzmanolivares/sandwich | 27 | Compare |
| Word Sense Disambiguation | Knowledge-based: | KEF All 68.0 | Word Sense Disambiguation: A comprehensive knowledge... | — | 6 | Compare |
| Quantization | Knowledge-based: | 3DCNN_VIVA_5 All 84809664 | Compressing 3DCNNs Based on Tensor Train Decomposition | — | 1 | Compare |
Papers archive 2025-07-28
22 shown of 22 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 109. The Syntology column is from Syntology's graph (read 2026-09-25), stated per sample; it is not part of any archive number.
Dataset loaders archive 2025-07-28
No loader listed in the archive.
Tasks archive 2025-07-28
License archive 2025-07-28
Unknown
Modalities archive 2025-07-28
Languages archive 2025-07-28
Variants archive 2025-07-28
- Knowledge-based:
- Supervised:
- Word Sense Disambiguation: a Unified Evaluation Framework and Empirical Comparison
3 variant names, as the archive lists them.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections