Browse State-of-the-Art › Optical Character Recognition (OCR) › Papers, page 5
Optical Character Recognition (OCR)
Papers archive 2025-07-28
archive papers tagged: 1,243 · with a code link: 462 · where Syntology ran a sample: 76 (64 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (76 of 1,243 tagged: 64 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument)
Page 5 of 13: papers 401 to 500 of 1,243, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
14 May 2020 1 repository listed
-
14 May 2020 1 repository listed
-
13 May 2020 1 repository listed
-
7 May 2020 1 repository listed
-
1 May 2020 1 repository listed
-
27 Apr 2020 1 repository listed
-
23 Apr 2020 1 repository listed
-
15 Apr 2020 1 repository listed
-
19 Feb 2020 1 repository listed
-
15 Dec 2019 1 repository listed
-
13 Nov 2019 1 repository listed
-
12 Oct 2019 1 repository listed
-
9 Oct 2019 1 repository listed
-
7 Oct 2019 1 repository listed
-
1 Oct 2019 1 repository listed
-
27 Sep 2019 1 repository listed
-
20 Sep 2019 1 repository listed
-
14 Sep 2019 1 repository listed
-
14 Sep 2019 1 repository listed
-
4 Sep 2019 1 repository listed
-
29 Aug 2019 1 repository listed
-
15 Aug 2019 1 repository listed
-
5 Aug 2019 1 repository listed
-
17 Jul 2019 1 repository listed
-
2 Jul 2019 1 repository listed
-
12 Jun 2019 1 repository listed
-
1 Jun 2019 1 repository listed
-
8 May 2019 1 repository listed
-
1 May 2019 1 repository listed
-
1 Jan 2019 1 repository listed
-
31 Dec 2018 1 repository listed
-
6 Dec 2018 1 repository listed
-
15 Nov 2018 1 repository listed
-
8 Oct 2018 1 repository listed
-
1 Oct 2018 1 repository listed
-
6 Sep 2018 1 repository listed
-
1 Sep 2018 1 repository listed
-
11 Jul 2018 1 repository listed
-
5 Jul 2018 1 repository listed
-
23 May 2018 1 repository listed
-
1 May 2018 1 repository listed
-
1 May 2018 1 repository listed
-
27 Feb 2018 1 repository listed
-
27 Feb 2018 1 repository listed
-
15 Feb 2018 1 repository listed
-
15 Dec 2017 1 repository listed
-
1 Dec 2017 1 repository listed
-
27 Nov 2017 1 repository listed
-
12 Nov 2017 1 repository listed
-
22 Jun 2017 1 repository listed
-
14 Jun 2017 1 repository listed
-
3 Dec 2016 1 repository listed
-
1 Nov 2016 1 repository listed
-
20 Sep 2016 1 repository listed
-
1 Aug 2016 1 repository listed
-
13 May 2016 1 repository listed
-
24 Feb 2016 1 repository listed
-
1 Dec 2015 1 repository listed
-
1 May 2014 1 repository listed
-
25 Jul 2009 1 repository listed
-
DeQA-Doc: Adapting DeQA-Score to Document Image Quality Assessment17 Jul 2025 0 repositories listed
-
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning17 Jul 2025 0 repositories listed
-
Seeing the Signs: A Survey of Edge-Deployable OCR Models for Billboard Visibility Analysis15 Jul 2025 0 repositories listed
-
A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends14 Jul 2025 0 repositories listed
-
Design and Implementation of an OCR-Powered Pipeline for Table Extraction from Invoices9 Jul 2025 0 repositories listed
-
TextPixs: Glyph-Conditioned Diffusion with Character-Aware Attention and OCR-Guided Supervision8 Jul 2025 0 repositories listed
-
Logios : An open source Greek Polytonic Optical Character Recognition system26 Jun 2025 0 repositories listed
-
Engineering RAG Systems for Real-World Applications: Design, Development, and Evaluation25 Jun 2025 0 repositories listed
-
Seeing is Believing? Mitigating OCR Hallucinations in Multimodal Large Language Models25 Jun 2025 0 repositories listed
-
Unfolding the Past: A Comprehensive Deep Learning Approach to Analyzing Incunabula Pages22 Jun 2025 0 repositories listed
-
An accurate and revised version of optical character recognition-based speech synthesis using LabVIEW18 Jun 2025 0 repositories listed
-
FormGym: Doing Paperwork with Agents17 Jun 2025 0 repositories listed
-
AdaVideoRAG: Omni-Contextual Adaptive Retrieval-Augmented Efficient Long Video Understanding16 Jun 2025 0 repositories listed
-
Efficient Medical VIE via Reinforcement Learning16 Jun 2025 0 repositories listed
-
MultiFinBen: A Multilingual, Multimodal, and Difficulty-Aware Benchmark for Financial LLM Evaluation16 Jun 2025 0 repositories listed
-
Generalization or Hallucination? Understanding Out-of-Context Reasoning in Transformers12 Jun 2025 0 repositories listed
-
Intelligent Automation for FDI Facilitation: Optimizing Tariff Exemption Processes with OCR And Large Language Models12 Jun 2025 0 repositories listed
-
Task-driven real-world super-resolution of document scans8 Jun 2025 0 repositories listed
-
Reading in the Dark with Foveated Event Vision7 Jun 2025 0 repositories listed
-
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing7 Jun 2025 0 repositories listed
-
A Survey on Vietnamese Document Analysis and Recognition: Challenges and Future Directions5 Jun 2025 0 repositories listed
-
SARD: A Large-Scale Synthetic Arabic OCR Dataset for Book-Style Text Recognition30 May 2025 0 repositories listed
-
ChartMind: A Comprehensive Benchmark for Complex Real-world Multimodal Chart Question Answering29 May 2025 0 repositories listed
-
TextSR: Diffusion Super-Resolution with Multilingual OCR Guidance29 May 2025 0 repositories listed
-
E2E Process Automation Leveraging Generative AI and IDP-Based Automation Agent: A Case Study on Corporate Expense Processing27 May 2025 0 repositories listed
-
MT³: Scaling MLLM-based Text Image Machine Translation via Multi-Task Reinforcement Learning26 May 2025 0 repositories listed
-
TextDiffuser-RL: Efficient and Robust Text Layout Optimization for High-Fidelity Text-to-Image Synthesis25 May 2025 0 repositories listed
-
Words as Geometric Features: Estimating Homography using Optical Character Recognition as Compressed Image Representation25 May 2025 0 repositories listed
-
One RL to See Them All: Visual Triple Unified Reinforcement Learning23 May 2025 0 repositories listed
-
TextFlux: An OCR-Free DiT Model for High-Fidelity Multilingual Scene Text Synthesis23 May 2025 0 repositories listed
-
TokBench: Evaluating Your Visual Tokenizer before Visual Generation23 May 2025 0 repositories listed
-
22 May 2025 0 repositories listed Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 5 pointer-only (licence)
-
What Media Frames Reveal About Stance: A Dataset and Study about Memes in Climate Change Discourse22 May 2025 0 repositories listed
-
How Do Large Vision-Language Models See Text in Image? Unveiling the Distinctive Role of OCR Heads21 May 2025 0 repositories listed
-
Every Pixel Tells a Story: End-to-End Urdu Newspaper OCR20 May 2025 0 repositories listed
-
The Hidden Structure -- Improving Legal Document Understanding Through Explicit Text Formatting19 May 2025 0 repositories listed
-
Towards Self-Improvement of Diffusion Models via Group Preference Optimization16 May 2025 0 repositories listed
Syntology lines on 4 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.