Papers › LAREX - A semi-automatic open-source Tool for Layout Analysis and Region Extraction on...

LAREX - A semi-automatic open-source Tool for Layout Analysis and Region Extraction on Early Printed Books

20 Jan 2017arXiv:1701.07396archive 2025-07-28

Christian Reul, Uwe Springmann, Frank Puppe

A semi-automatic open-source tool for layout analysis on early printed books is presented. LAREX uses a rule based connected components approach which is very fast, easily comprehensible for the user and allows an intuitive manual correction if necessary. The PageXML format is used to support integration into existing OCR workflows. Evaluations showed that LAREX provides an efficient and flexible way to segment pages of early printed books.

PaperPDFCode

Code

chreul/LAREX officialmentioned in papermentioned on GitHub report
OCR4all/LAREX mentioned on GitHub report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Optical Character Recognition (OCR)

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections