Datasets › WiC

WiC (Words in Context)

Introduced by Mohammad Taher Pilehvar et al. in WiC: the Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations archive 2025-07-28

WiC is a benchmark for the evaluation of context-sensitive word embeddings. WiC is framed as a binary classification task. Each instance in WiC has a target word w, either a verb or a noun, for which two contexts are provided. Each of these contexts triggers a specific meaning of w. The task is to identify if the occurrences of w in the two contexts correspond to the same meaning or not. In fact, the dataset can also be viewed as an application of Word Sense Disambiguation in practise.

Source: WiC: the Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations

Benchmarks archive 2025-07-28

All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Word Sense Disambiguation Words in Context COSINE + Transductive Learning Accuracy 85.3 Fine-Tuning Pre-trained Language Model with Weak... yueyu1030/COSINE 37 Compare
Classification WiC OPT-1.3B Test Accuracy 56.14% Achieving Dimension-Free Communication in Federated... ZidongLiu/DeComFL 2 Compare
Text Generation WiC no rows — — 0 Compare

Papers archive 2025-07-28

20 shown of 20 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 206. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Achieving Dimension-Free Communication in Federated Learning via Zeroth-Order Optimization 1 2 24 May 2024 ran 1 of 1 samples (0 unverified)
The CoT Collection: Improving Zero-shot and Few-shot Learning of Language Models via Chain-of-Thought Fine-Tuning 2 1 23 May 2023 not harvested
PaLM 2 Technical Report 1 3 17 May 2023 not harvested
LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions 1 5 27 Apr 2023 not harvested
Exploring the Benefits of Training Expert Language Models over Instruction Tuning 2 1 7 Feb 2023 ran 3 of 3 samples (0 unverified; 3 pointer-only for licence)
Hungry Hungry Hippos: Towards Language Modeling with State Space Models 3 3 28 Dec 2022 ran 7 of 15 samples (8 unverified)
Toward Efficient Language Model Pretraining and Downstream Adaptation via Self-Evolution: A Case Study on SuperGLUE 0 2 4 Dec 2022 not harvested
Knowledge-in-Context: Towards Knowledgeable Semi-Parametric Language Models 0 1 28 Oct 2022 not harvested
Guess the Instruction! Flipped Learning Makes Language Models Stronger Zero-Shot Learners 1 1 6 Oct 2022 not harvested
AlexaTM 20B: Few-Shot Learning Using a Large-Scale Multilingual Seq2Seq Model 1 1 2 Aug 2022 ran 1 of 1 samples (0 unverified)
N-Grammer: Augmenting Transformers with latent n-grams 2 1 13 Jul 2022 ran 0 of 6 samples (6 unverified)
UL2: Unifying Language Learning Paradigms 2 2 10 May 2022 ran 0 of 16 samples (16 unverified)
PaLM: Scaling Language Modeling with Pathways 7 1 5 Apr 2022 ran 30 of 37 samples (7 unverified)
ST-MoE: Designing Stable and Transferable Sparse Expert Models 3 2 17 Feb 2022 ran 5 of 5 samples (0 unverified; 5 pointer-only for licence)
Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self-Training Approach 1 1 15 Oct 2020 ran 0 of 5 samples (5 unverified)
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 14 2 5 Jun 2020 ran 4 of 13 samples (9 unverified; 3 pointer-only for licence)
Language Models are Few-Shot Learners 67 1 28 May 2020 ran 15 of 65 samples (50 unverified; 4 pointer-only for licence)
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer 57 1 23 Oct 2019 ran 2 of 31 samples (29 unverified)
SenseBERT: Driving Some Sense into BERT 0 2 15 Aug 2019 not harvested
WiC: the Word-in-Context Dataset for Evaluating Context-Sensitive Meaning Representations 0 6 28 Aug 2018 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

CC BY-NC 4.0

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • WiC

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections