Papers › CURE: A dataset for Clinical Understanding & Retrieval Evaluation

CURE: A dataset for Clinical Understanding & Retrieval Evaluation

9 Dec 2024arXiv:2412.06954archive 2025-07-28

Nadia Sheikh, Anne-Laure Jousse, Daniel Buades Marcos, Akintunde Oladipo, Olivier Rousseau, Jimmy Lin

Given the dominance of dense retrievers that do not generalize well beyond their training dataset distributions, domain-specific test sets are essential in evaluating retrieval. There are few test datasets for retrieval systems intended for use by healthcare providers in a point-of-care setting. To fill this gap we have collaborated with medical professionals to create CURE, an ad-hoc retrieval test dataset for passage ranking with 2000 queries spanning 10 medical domains with a monolingual (English) and two cross-lingual (French/Spanish -> English) conditions. In this paper, we describe how CURE was constructed and provide baseline results to showcase its effectiveness as an evaluation tool. CURE is published with a Creative Commons Attribution Non Commercial 4.0 license and can be accessed on Hugging Face.

PaperPDFCode

Code

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Passage RankingRetrieval

Datasets

Introduced by this paper, per the archive.

CURE

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

NON

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections