Datasets › ICDAR 2013

ICDAR 2013

Introduced in ICDAR 2013 Robust Reading Competition1 Jan 2013 archive 2025-07-28

The ICDAR 2013 dataset consists of 229 training images and 233 testing images, with word-level annotations provided. It is the standard benchmark dataset for evaluating near-horizontal text detection.

Source: Single Shot Text Detector with Regional Attention Image Source: https://plos.figshare.com/articles/Detection_examples_of_the_proposed_method_on_the_ICDAR_2013_dataset_17_/5325856

Benchmarks archive 2025-07-28

All 3 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

Papers archive 2025-07-28

30 shown of 50 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 246. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
An Empirical Study of Scaling Law for OCR 1 1 29 Dec 2023 not harvested
DTrOCR: Decoder-only Transformer for Optical Character Recognition 1 1 30 Aug 2023 not harvested
DiffusionSTR: Diffusion Model for Scene Text Recognition 0 1 29 Jun 2023 not harvested
CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model 1 3 23 May 2023 ran 3 of 10 samples (7 unverified)
Self-supervised Character-to-Character Distillation for Text Recognition 1 3 1 Nov 2022 not harvested
Multi-Granularity Prediction for Scene Text Recognition 3 1 8 Sep 2022 not harvested
Scene Text Recognition with Permuted Autoregressive Sequence Models 2 1 14 Jul 2022 ran 5 of 8 samples (3 unverified)
SVTR: Scene Text Recognition with a Single Visual Model 4 4 30 Apr 2022 not harvested
Self-supervised Implicit Glyph Attention for Text Recognition 1 1 7 Mar 2022 not harvested
SAFL: A Self-Attention Scene Text Recognizer with Focal Loss 1 1 1 Jan 2022 not harvested
Visual Semantics Allow for Textual Reasoning Better in Scene Text Recognition 1 1 24 Dec 2021 not harvested
Multi-modal Text Recognition Networks: Interactive Enhancements between Visual and Semantic Features 3 1 30 Nov 2021 not harvested
CDistNet: Perceiving Multi-Domain Character Distance for Robust Text Recognition 3 1 22 Nov 2021 not harvested
Look Back Again: Dual Parallel Attention Network for Accurate and Robust Scene Text Recognition 2 1 1 Aug 2021 not harvested
Why You Should Try the Real Data for the Scene Text Recognition 1 1 29 Jul 2021 not harvested
Representation and Correlation Enhanced Encoder-Decoder Framework for Scene Text Recognition 1 1 13 Jun 2021 not harvested
Vision Transformer for Fast and Efficient Scene Text Recognition 3 1 18 May 2021 not harvested
Revisiting Classification Perspective on Scene Text Recognition 1 1 22 Feb 2021 not harvested
CDeC-Net: Composite Deformable Cascade Network for Table Detection in Document Images 3 1 25 Aug 2020 ran 0 of 6 samples (6 unverified)
SEED: Semantics Enhanced Encoder-Decoder Framework for Scene Text Recognition 3 1 22 May 2020 not harvested
TextFuseNet: Scene Text Detection with Richer Fused Features 6 1 17 May 2020 not harvested
CascadeTabNet: An approach for end to end table detection and structure recognition from image-based documents 3 1 27 Apr 2020 ran 0 of 2 samples (2 unverified)
Towards Accurate Scene Text Recognition with Semantic Reasoning Networks 3 1 27 Mar 2020 not harvested
TableNet: Deep Learning model for end-to-end Table detection and Tabular data extraction from Scanned Document Images 5 1 6 Jan 2020 ran 0 of 2 samples (2 unverified)
TextScanner: Reading Characters in Order for Robust Scene Text Recognition 0 1 28 Dec 2019 not harvested
Decoupled Attention Network for Text Recognition 5 1 21 Dec 2019 not harvested
On Recognizing Texts of Arbitrary Shapes with 2D Self-Attention 2 1 10 Oct 2019 not harvested
Unsharp Masking Layer: Injecting Prior Knowledge in Convolutional Networks for Image Classification 1 1 29 Sep 2019 not harvested
What Is Wrong With Scene Text Recognition Model Comparisons? Dataset and Model Analysis 13 1 3 Apr 2019 ran 0 of 19 samples (19 unverified)
Character Region Awareness for Text Detection 18 1 3 Apr 2019 ran 6 of 40 samples (34 unverified; 4 pointer-only for licence)

The full list of 50 is in the JSON twin.

Dataset loaders archive 2025-07-28

3 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

Variants archive 2025-07-28

  • ICDAR 2013
  • ICDAR2013

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections