Home › Datasets › task › Binarization

Binarization datasets

archive 2025-07-28

17 datasets carry the task tag "Binarization" (the task itself: Binarization), ordered by the archive's paper count. Page 1 of 1: 17 shown of 17. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets

Binarization datasets 1–17 of 17

description withheld: archive row vandalised before snapshot
16,145 papers · 91 benchmarks
The ImageNet dataset contains 14,197,122 annotated images according to the WordNet hierarchy.
15,430 papers · 52 benchmarks
The CIFAR-100 dataset (Canadian Institute for Advanced Research, 100 classes) is a subset of the Tiny Images dataset and consists of 60000 32x32 color images.
9,045 papers · 51 benchmarks
(JHU-CROWD) a crowd counting dataset that contains 4,250 images with 1.11 million annotations.
22 papers · 0 benchmarks
FDST (Fudan-ShanghaiTech)
The Fudan-ShanghaiTech dataset (FDST) is a dataset for video crowd counting.
18 papers · 0 benchmarks
DIBCO and H_DIBCO ((Handwritten) Document Image Binarization Competition (DIBCO))
The contest of binarization using a popular document database was organized called as Document Image Binarization Contest (DIBCO) from 2009 to 2019, except for 2015.
14 papers · 0 benchmarks
DIBCO 2011 is the International Document Image Binarization Contest organized in the context of ICDAR 2011 conference.
7 papers · 0 benchmarks
H-DIBCO 2016 is the international Handwritten Document Image Binarization Contest organized in the context of ICFHR 2016 conference
7 papers · 0 benchmarks
DIBCO 2017 is the international Competition on Document Image Binarization organized in conjunction with the ICDAR 2017 conference.
6 papers · 0 benchmarks
H-DIBCO 2014 is the International Document Image Binarization Competition which is dedicated to handwritten document images organized in conjunction with ICFHR 2014 conference.
6 papers · 0 benchmarks
H-DIBCO 2018 is the international Handwritten Document Image Binarization Contest organized in the context of ICFHR 2018 conference.
6 papers · 0 benchmarks
DIBCO 2013 is the international Document Image Binarization Contest organized in the context of ICDAR 2013 conference.
5 papers · 0 benchmarks
H-DIBCO 2012 is the International Document Image Binarization Competition which is dedicated to handwritten document images organized in conjunction with ICFHR 2012 conference.
5 papers · 0 benchmarks
DIBCO 2009 is the first International Document Image Binarization Contest organized in the context of ICDAR 2009 conference.
4 papers · 0 benchmarks
DIBCO 2019 is the international Competition on Document Image Binarization organized in conjunction with the ICDAR 2019 conference.
4 papers · 0 benchmarks
H-DIBCO 2010 is the International Document Image Binarization Contest which is dedicated to handwritten document images organized in conjunction with ICFHR 2010 conference.
4 papers · 0 benchmarks
This is a dataset is composed of full-document images, groundtruth, and tools to perform an evaluation of binarization algorithms.
1 paper · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.