Home › Datasets › task › Face Recognition

Face Recognition datasets

archive 2025-07-28

67 datasets carry the task tag "Face Recognition" (the task itself: Face Recognition), ordered by the archive's paper count. Page 1 of 2: 48 shown of 67. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Face Recognition datasets 1–48 of 67

LFW (Labeled Faces in the Wild)
The LFW dataset contains 13,233 images of faces collected from the web.
820 papers · 12 benchmarks
The CASIA-WebFace dataset is used for face verification and face identification tasks.
415 papers · 2 benchmarks
The MS-Celeb-1M dataset is a large-scale face recognition dataset consists of 100K identities, and each identity has about 100 facial images.
257 papers · 0 benchmarks
The Extended Yale B database contains 2414 frontal-face images with size 192×168 over 38 subjects and about 64 images per subject.
185 papers · 1 benchmark
MORPH is a facial age estimation dataset, which contains 55,134 facial images of 13,617 subjects ranging from 16 to 77 years old.
180 papers · 8 benchmarks
IJB-B (IARPA Janus Benchmark-B)
The IJB-B dataset is a template-based face dataset that contains 1845 subjects with 11,754 images, 55,025 frames and 7,011 videos where a template consists of a varying number of still images and video frames from different sources.
163 papers · 5 benchmarks
The VGG Face dataset is face identity recognition dataset that consists of 2,622 identities.
94 papers · 0 benchmarks
RFW (Racial Faces in-the-Wild)
To validate the racial bias of four commercial APIs and four state-of-the-art (SOTA) algorithms.
73 papers · 0 benchmarks
CASIA-FASD is a small face anti-spoofing dataset containing 50 subjects.
42 papers · 0 benchmarks
Partial REID is a specially designed partial person reidentification dataset that includes 600 images from 60 people, with 5 full-body images and 5 occluded images per person.
41 papers · 1 benchmark
The color FERET database is a dataset for face recognition.
36 papers · 3 benchmarks
description withheld: archive row vandalised before snapshot
35 papers · 6 benchmarks
UMDFaces is a face dataset divided into two parts: Still Images - 367,888 face annotations for 8,277 subjects.
34 papers · 0 benchmarks
TinyFace is a large scale face recognition benchmark to facilitate the investigation of natively LRFR (Low Resolution Face Recognition) at large scales (large gallery population sizes) in deep learning.
32 papers · 0 benchmarks
CelebA-Spoof is a large-scale face anti-spoofing dataset with the following properties: 1.
28 papers · 0 benchmarks
Dataset for face anti-spoofing in terms of both subjects and modalities.
24 papers · 0 benchmarks
IMDb-Face is large-scale noise-controlled dataset for face recognition research.
22 papers · 0 benchmarks
DigiFace-1M is a synthetic dataset for face recognition, obtained by rendering digital faces using a computer graphics pipeline.
21 papers · 0 benchmarks
WebFace260M is a million-scale face benchmark, which is constructed for the research community towards closing the data gap behind the industry.
20 papers · 0 benchmarks
The Replay-Mobile Database for face spoofing consists of 1190 video clips of photo and video attack attempts to 40 clients, under different lighting conditions.
18 papers · 0 benchmarks
CALFW (Cross-Age LFW)
A renovation of Labeled Faces in the Wild (LFW), the de facto standard testbed for unconstraint face verification.
17 papers · 4 benchmarks
CPLFW (Cross-Pose LFW)
A renovation of Labeled Faces in the Wild (LFW), the de facto standard testbed for unconstraint face verification.
17 papers · 4 benchmarks
MFR (Ongoing version of ICCV-2021 Masked Face Recognition Challenge & Workshop(MFR))
During the COVID-19 coronavirus epidemic, almost everyone wears a facial mask, which poses a huge challenge to face recognition.
16 papers · 1 benchmark
XQLFW (Cross-Quality Labeled Faces in the Wild)
An evaluation protocol for face verification focusing on a large intra-pair image quality difference.
16 papers · 1 benchmark
The IDiff-Face dataset was proposed in the paper "IDiff-Face: Synthetic-based Face Recognition through Fizzy Identity-Conditioned Diffusion Models".
13 papers · 0 benchmarks
A new face annotation dataset with balanced distribution between genders and ethnic origins.
11 papers · 2 benchmarks
RMFD (Real-World Masked Face Dataset)
Real-World Masked Face Dataset (RMFD) is a large dataset for masked face detection.
11 papers · 0 benchmarks
Although deep face recognition has achieved impressive results in recent years, there is increasing controversy regarding racial and gender bias of the models, questioning their trustworthiness and deployment into sensitive scenarios.
10 papers · 0 benchmarks
QMUL-SurvFace is a surveillance face recognition benchmark that contains 463,507 face images of 15,573 distinct identities captured in real-world uncooperative surveillance scenes over wide space and time.
10 papers · 1 benchmark
The iCartoonFace dataset is a large-scale dataset that can be used for two different tasks: cartoon face detection and cartoon face recognition.
9 papers · 1 benchmark
MeGlass is an eyeglass dataset originally designed for eyeglass face recognition evaluation.
8 papers · 0 benchmarks
DFW (Disguised Faces in the Wild)
Contains over 11000 images of 1000 identities with different types of disguise accessories.
7 papers · 3 benchmarks
DroneSURF (DroneSURF: Benchmark Dataset for Drone-based Face Recognition)
Drone Surveillance of Faces, is a large-scale drone dataset intended to facilitate research for face recognition using drones.
6 papers · 1 benchmark
MCXFACE (Multi-Channel Heterogeneous Face Recognition dataset)
MCXFace is a heterogeneous face recognition dataset consisting of multi-channel image samples for 51 subjects.
6 papers · 0 benchmarks
MLFW (Masked LFW)
The Masked LFW (MLFW), based on Cross-Age LFW (CALFW) database, is built using a simple but effective tool that generates masked faces from unmasked faces automatically.
6 papers · 1 benchmark
Aims to facilitate research in caricature recognition.
6 papers · 0 benchmarks
iQIYI-VID dataset, which comprises video clips from iQIYI variety shows, films, and television dramas.
6 papers · 0 benchmarks
A multimodal database for eye blink detection and attention level estimation.
6 papers · 0 benchmarks
BTS3.1 (Expanding Accurate Person Recognition to New Altitudes and Ranges: The BRIAR Dataset)
Large, multimodal biometric dataset: It contains still images and videos of over 1,000 people captured at various ranges (up to 1,000 meters) and elevations (up to 400 meters) using a diverse set of cameras (commercial, military-grade,…
5 papers · 2 benchmarks
ROF (Real World Occluded Faces)
ROF is a dataset for occluded face recognition that contains faces with both upper face occlusion, due to sunglasses, and lower face occlusion, due to masks.
5 papers · 0 benchmarks
Description: 1,995 People Face Images Data (Asian race).
4 papers · 0 benchmarks
The COVID-19 pandemic raises the problem of adapting face recognition systems to the new reality, where people may wear surgical masks to cover their noses and mouths.
4 papers · 1 benchmark
The COVID-19 pandemic raises the problem of adapting face recognition systems to the new reality, where people may wear surgical masks to cover their noses and mouths.
4 papers · 1 benchmark
CASIA-Face-Africa is a face image database which contains 38,546 images of 1,183 African subjects.
3 papers · 0 benchmarks
IJB-S (IARPA Janus Benchmark-S)
Paper Abstract We present IJB–S dataset, an open-source IARPA Janus Surveillance Video Benchmark and associated protocols.
3 papers · 1 benchmark
We have cleaned the noisy IMDB-WIKI dataset using a constrained clustering method, resulting this new benchmark for in-the-wild age estimation.
3 papers · 1 benchmark
Description: 5,011 Images – Human Frontal face Data (Male).
2 papers · 0 benchmarks
Celeb-HQ Face Gender Recognition Dataset This dataset is curated for the face gender classification task.
2 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.