Home › Datasets › task › Few-Shot Learning
Few-Shot Learning datasets
archive 2025-07-28
45 datasets carry the task tag "Few-Shot Learning" (the task itself: Few-Shot Learning), ordered by the archive's paper count. Page 1 of 1: 45 shown of 45. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Few-Shot Learning datasets 1–45 of 45
The ImageNet dataset contains 14,197,122 annotated images according to the WordNet hierarchy.
15,430 papers · 52 benchmarks
GLUE (General Language Understanding Evaluation benchmark)
General Language Understanding Evaluation (GLUE) benchmark is a collection of nine natural language understanding tasks, including single-sentence tasks CoLA and SST-2, similarity and paraphrasing tasks MRPC, STS-B and QQP, and natural…
3,197 papers · 13 benchmarks
SST (Stanford Sentiment Treebank)
The Stanford Sentiment Treebank is a corpus with fully labeled parse trees that allows for a complete analysis of the compositional effects of sentiment in language.
2,354 papers · 6 benchmarks
The Caltech-UCSD Birds-200-2011 (CUB-200-2011) dataset is the most widely-used dataset for fine-grained visual categorization task.
2,235 papers · 47 benchmarks
UCF101 (UCF101 Human Actions dataset)
UCF101 dataset is an extension of UCF50 and consists of 13,320 video clips, which are classified into 101 categories.
1,863 papers · 23 benchmarks
The Stanford Sentiment Treebank is a corpus with fully labeled parse trees that allows for a complete analysis of the compositional effects of sentiment in language.
1,808 papers · 2 benchmarks
mini-Imagenet is proposed by Matching Networks for One Shot Learning .
1,345 papers · 21 benchmarks
Oxford 102 Flower is an image classification dataset consisting of 102 flower categories.
1,307 papers · 16 benchmarks
DTD (Describable Textures Dataset)
The Describable Textures Dataset (DTD) contains 5640 texture images in the wild.
870 papers · 8 benchmarks
The Food-101 dataset consists of 101 food categories with 750 training and 250 test images per category, making a total of 101k images.
805 papers · 14 benchmarks
The Stanford Cars dataset consists of 196 classes of cars with a total of 16,185 images, taken from the rear.
790 papers · 13 benchmarks
MRPC (Microsoft Research Paraphrase Corpus)
Microsoft Research Paraphrase Corpus (MRPC) is a corpus consists of 5,801 sentence pairs collected from newswire articles.
786 papers · 4 benchmarks
Eurosat is a dataset and deep learning benchmark for land use and land cover classification.
687 papers · 8 benchmarks
FGVC-Aircraft contains 10,200 images of aircraft, with 100 images for each of 102 different aircraft model variants, most of which are airplanes.
520 papers · 12 benchmarks
The task of PubMedQA is to answer research questions with yes/no/maybe (e.g.: Do preoperative statins reduce atrial fibrillation after coronary artery bypass grafting?) using the corresponding abstracts.
276 papers · 3 benchmarks
PASCAL-5i is a dataset used to evaluate few-shot segmentation.
177 papers · 1 benchmark
FLEURS (Few-shot Learning Evaluation of Universal Representations of Speech)
We introduce FLEURS, the Few-shot Learning Evaluation of Universal Representations of Speech benchmark.
141 papers · 1 benchmark
The Meta-Dataset benchmark is a large few-shot learning benchmark and consists of multiple datasets of different data distributions.
128 papers · 2 benchmarks
The Scene UNderstanding (SUN) database contains 899 categories and 130,519 images.
52 papers · 8 benchmarks
MR Movie Reviews is a dataset for use in sentiment-analysis experiments.
28 papers · 3 benchmarks
CaseHOLD (Case Holdings On Legal Decisions)
CaseHOLD (Case Holdings On Legal Decisions) is a law dataset comprised of over 53,000+ multiple choice questions to identify the relevant holding of a cited case.
27 papers · 2 benchmarks
This dataset is a Wikipedia dump, split by relations to perform Few-Shot Knowledge Graph Completion.
16 papers · 0 benchmarks
The Paris-Lille-3D is a Benchmark on Point Cloud Classification.
15 papers · 1 benchmark
MedConceptsQA - Open Source Medical Concepts QA Benchmark The benchmark can be found here: https://huggingface.co/datasets/ofir408/MedConceptsQA
13 papers · 2 benchmarks
MedNLI (Medical Natural Language Inference)
The MedNLI dataset consists of the sentence pairs developed by Physicians from the Past Medical History section of MIMIC-III clinical notes annotated for Definitely True, Maybe True and Definitely False.
9 papers · 2 benchmarks
GINC (Generative IN-Context learning Dataset)
GINC (Generative In-Context learning Dataset) is a small-scale synthetic dataset for studying in-context learning.
7 papers · 0 benchmarks
ORBIT is a real-world few-shot dataset and benchmark grounded in a real-world application of teachable object recognizers for people who are blind/low vision.
7 papers · 2 benchmarks
N-Digit MNIST is a multi-digit MNIST-like dataset.
5 papers · 0 benchmarks
ExVo2022 (ICML ExVo 2022 Workshop & Competition Data)
Baseline code for the three tracks of ExVo 2022 competition.
4 papers · 0 benchmarks
F-SIOL-310 is a robotic dataset and benchmark for Few-Shot Incremental Object Learning, which is used to test incremental learning capabilities for robotic vision from a few examples.
4 papers · 0 benchmarks
The FIGR-8 database is a dataset containing 17,375 classes of 1,548,256 images representing pictograms, ideograms, icons, emoticons or object or conception depictions.
4 papers · 0 benchmarks
FewSOL (A Dataset for Few-Shot Object Learning in Robotic Environments)
The Few-Shot Object Learning (FewSOL) dataset can be used for object recognition with a few images per object.
4 papers · 0 benchmarks
N-Omniglot is a neuromorphic dataset for few-shot learning.
4 papers · 0 benchmarks
Contains 3,689,229 English news articles on politics, gathered from 11 United States (US) media outlets covering a broad ideological spectrum.
3 papers · 0 benchmarks
Bongard-OpenWorld is a new benchmark for evaluating real-world few-shot reasoning for machine vision.
3 papers · 1 benchmark
Intended to provide freely available data sets in various formats together with basic annotation to be useful for applications in computational linguistics, translation studies and cross-linguistic corpus studies.
3 papers · 0 benchmarks
carecall is a Korean dialogue dataset for role-satisfying dialogue systems.
2 papers · 0 benchmarks
"We built a large lung CT scan dataset for COVID-19 by curating data from 7 public datasets listed in the acknowledgements.
2 papers · 2 benchmarks
ToM-in-AMC is a novel NLP benchmark, Short for Theory-of-Mind meta-learning Assessment with Movie Characters.
2 papers · 0 benchmarks
Millions of people around the world have low or no vision.
1 paper · 0 benchmarks
Introduction The FewGLUE64labeled dataset is a new version of FewGLUE dataset.
1 paper · 0 benchmarks
ISEKAI dataset’s images are generated by Midjourney’s text-to-image model using well-crafted instructions.
1 paper · 0 benchmarks
A large-scale reference dataset for bioacoustics.
1 paper · 1 benchmark
Open MIC (Open Museum Identification Challenge)
Open MIC (Open Museum Identification Challenge) contains photos of exhibits captured in 10 distinct exhibition spaces of several museums which showcase paintings, timepieces, sculptures, glassware, relics, science exhibits, natural history…
1 paper · 0 benchmarks
A dataset specifically tailored to the biotech news sector, aiming to transcend the limitations of existing benchmarks.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.