Home › Datasets › task › Handwriting Recognition
Handwriting Recognition datasets
archive 2025-07-28
22 datasets carry the task tag "Handwriting Recognition" (the task itself: Handwriting Recognition), ordered by the archive's paper count. Page 1 of 1: 22 shown of 22. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 51 task tags shown of 3,717, by dataset count; the full filter by modality, task and language is on /datasets
Handwriting Recognition datasets 1–22 of 22
The IAM database contains 13,353 images of handwritten lines of text created by 657 writers.
198 papers · 1 benchmark
A new dataset of handwritten text with fine-grained annotations at the character level and report results from an initial user evaluation.
10 papers · 0 benchmarks
RIMES (Reconnaissance & Indexation de données Manuscrites et de fac similÉS / Recognition & Indexing of handwritten documents & faxes)
The RIMES database (Reconnaissance et Indexation de données Manuscrites et de fac similÉS / Recognition and Indexing of handwritten documents and faxes) was created to evaluate automatic systems of recognition and indexing of handwritten…
8 papers · 0 benchmarks
This dataset arises from the READ project (Horizon 2020).
5 papers · 1 benchmark
Bentham manuscripts refers to a large set of documents that were written by the renowned English philosopher and reformer Jeremy Bentham (1748-1832).
4 papers · 1 benchmark
HKR (Handwritten Kazakh and Russian (HKR) Database for Text Recognition)
The database is written in Cyrillic and shares the same 33 characters.
4 papers · 1 benchmark
BRUSH (Brown University Stylus Handwriting)
The BRUSH dataset (BRown University Stylus Handwriting) contains 27,649 online handwriting samples from a total of 170 writers.
3 papers · 0 benchmarks
This dataset contains Bangla handwritten numerals, basic characters and compound characters.
3 papers · 2 benchmarks
KOHTD (Kazakh Offline Handwritten Text Dataset)
Kazakh offline Handwritten Text dataset (KOHTD) has 3000 handwritten exam papers and more than 140335 segmented images and there are approximately 922010 symbols.
3 papers · 1 benchmark
Konzil dataset was created by specialists of the University of Greifswald.
3 papers · 0 benchmarks
Patzig contains handwritten texts written in modern German.
3 papers · 0 benchmarks
Ricordi contains handwritten texts written in Italian.
3 papers · 0 benchmarks
Schiller contains handwritten texts written in modern German.
3 papers · 0 benchmarks
Schwerin contains handwritten texts written in medieval German.
3 papers · 0 benchmarks
BN-HTRd (BN-HTRd: A Benchmark Dataset for Document Level Offline Bangla Handwritten Text Recognition (HTR))
We introduce a new Dataset (BN-HTRd) for offline Handwritten Text Recognition (HTR) from images of Bangla scripts comprising words, lines, and document-level annotations.
2 papers · 2 benchmarks
Saint Gall dataset contains handwritten historical manuscripts written in Latin that date back to the 9th century.
2 papers · 1 benchmark
Data collection: Finding a suitable source of data is considered a first step toward building a database.
1 paper · 1 benchmark
Calliar is a dataset for Arabic calligraphy.
1 paper · 0 benchmarks
KHATT (KFUPM Handwritten Arabic TexT Database)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
MatriVasha the largest dataset of handwritten Bangla compound characters for research on handwritten Bangla compound character recognition.
1 paper · 0 benchmarks
The Burmese Handwritten Digit Dataset (BHDD) is a dataset project specifically created for recognizing handwritten Burmese digits.
0 papers · 0 benchmarks
DigiLeTs (Digit- and Letter Trajectories)
A dataset with 23 870 digital trajectories (i.e.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.