Papers › Investigating Multi-source Active Learning for Natural Language Inference

Investigating Multi-source Active Learning for Natural Language Inference

14 Feb 2023arXiv:2302.06976archive 2025-07-28

Ard Snijders, Douwe Kiela, Katerina Margatina

In recent years, active learning has been successfully applied to an array of NLP tasks. However, prior work often assumes that training and test data are drawn from the same distribution. This is problematic, as in real-life settings data may stem from several sources of varying relevance and quality. We show that four popular active learning schemes fail to outperform random selection when applied to unlabelled pools comprised of multiple data sources on the task of natural language inference. We reveal that uncertainty-based strategies perform poorly due to the acquisition of collective outliers, i.e., hard-to-learn instances that hamper learning and generalization. When outliers are removed, strategies are found to recover and outperform random baselines. In further analysis, we find that collective outliers vary in form between sources, and show that hard-to-learn data is not always categorically harmful. Lastly, we leverage dataset cartography to introduce difficulty-stratified testing and find that different strategies are affected differently by example learnability and difficulty.

PaperPDFCode

Code

asnijders/multi_source_al officialmentioned in paperpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Active LearningNatural Language Inference

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

Testfail

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections