Datasets › CUTE80

CUTE80

archive 2025-07-28

The CUTE80 dataset is a lightweight collection of images specifically designed for text detection in natural scene images. It contains a total of 13,000 annotated page images across five different popular categories: 1) Table 2) Figure 3) Natural image 4) Logo 5) ignature

Benchmarks archive 2025-07-28

All 1 leaderboard whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Scene Text Recognition CUTE80 CPPD Accuracy 99.7 Context Perception Parallel Decoder for Scene Text Recognition PaddlePaddle/PaddleOCR +1 18 Compare

Papers archive 2025-07-28

14 shown of 14 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 17. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
An Empirical Study of Scaling Law for OCR 1 1 29 Dec 2023 not harvested
DTrOCR: Decoder-only Transformer for Optical Character Recognition 1 1 30 Aug 2023 not harvested
Context Perception Parallel Decoder for Scene Text Recognition 2 1 23 Jul 2023 not harvested
DiffusionSTR: Diffusion Model for Scene Text Recognition 0 1 29 Jun 2023 not harvested
CLIP4STR: A Simple Baseline for Scene Text Recognition with Pre-trained Vision-Language Model 1 3 23 May 2023 ran 3 of 10 samples (7 unverified)
TPS++: Attention-Enhanced Thin-Plate Spline for Scene Text Recognition 1 1 9 May 2023 not harvested
Self-supervised Character-to-Character Distillation for Text Recognition 1 3 1 Nov 2022 not harvested
Multi-Granularity Prediction for Scene Text Recognition 3 1 8 Sep 2022 not harvested
Scene Text Recognition with Permuted Autoregressive Sequence Models 2 1 14 Jul 2022 ran 5 of 8 samples (3 unverified)
Self-supervised Implicit Glyph Attention for Text Recognition 1 1 7 Mar 2022 not harvested
Visual Semantics Allow for Textual Reasoning Better in Scene Text Recognition 1 1 24 Dec 2021 not harvested
Multi-modal Text Recognition Networks: Interactive Enhancements between Visual and Semantic Features 3 1 30 Nov 2021 not harvested
CDistNet: Perceiving Multi-Domain Character Distance for Robust Text Recognition 3 1 22 Nov 2021 not harvested
Look Back Again: Dual Parallel Attention Network for Accurate and Robust Scene Text Recognition 2 1 1 Aug 2021 not harvested

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

No modality tagged.

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

No variants listed.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections