Browse State-of-the-Art › Optical Character Recognition (OCR) › Papers, page 8
Optical Character Recognition (OCR)
Papers archive 2025-07-28
archive papers tagged: 1,243 · with a code link: 462 · where Syntology ran a sample: 76 (64 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (76 of 1,243 tagged: 64 with a run with no instrument failure, 12 where every run was a failure of Syntology's instrument)
Page 8 of 13: papers 701 to 800 of 1,243, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
What Large Language Models Bring to Text-rich VQA?13 Nov 2023 0 repositories listed
-
DONUT-hole: DONUT Sparsification by Harnessing Knowledge and Optimizing Learning Efficiency9 Nov 2023 0 repositories listed
-
MultiCoNER v2: a Large Multilingual dataset for Fine-grained and Noisy Named Entity Recognition20 Oct 2023 0 repositories listed
-
EfficientOCR: An Extensible, Open-Source Package for Efficiently Digitizing World Knowledge16 Oct 2023 0 repositories listed
-
Towards reducing hallucination in extracting information from financial reports using Large Language Models16 Oct 2023 0 repositories listed
-
Exploring Sparse Spatial Relation in Graph Inference for Text-Based VQA13 Oct 2023 0 repositories listed
-
Invisible Threats: Backdoor Attack in OCR Systems12 Oct 2023 0 repositories listed
-
Solution for SMART-101 Challenge of ICCV Multi-modal Algorithmic Reasoning Task 202310 Oct 2023 0 repositories listed
-
Constructing Image-Text Pair Dataset from Books3 Oct 2023 0 repositories listed
-
Comprehensive Overview of Named Entity Recognition: Models, Domain-Specific Applications and Challenges25 Sep 2023 0 repositories listed
-
Bengali Document Layout Analysis -- A YOLOV8 Based Ensembling Approach2 Sep 2023 0 repositories listed
-
Enhancing OCR Performance through Post-OCR Models: Adopting Glyph Embedding for Improved Correction29 Aug 2023 0 repositories listed
-
Bengali Document Layout Analysis with Detectron226 Aug 2023 0 repositories listed
-
26 Aug 2023 0 repositories listed
-
DISGO: Automatic End-to-End Evaluation for Scene Text OCR25 Aug 2023 0 repositories listed
-
American Stories: A Large-Scale Structured Text Dataset of Historical U.S. Newspapers24 Aug 2023 0 repositories listed
-
CNN based Cuneiform Sign Detection Learned from Annotated 3D Renderings and Mapped Photographs with Illumination Augmentation22 Aug 2023 0 repositories listed
-
Extraction of Text from Optic Nerve Optical Coherence Tomography Reports21 Aug 2023 0 repositories listed
-
OCR Language Models with Custom Vocabularies18 Aug 2023 0 repositories listed
-
Multimodal Analysis Of Google Bard And GPT-Vision: Experiments In Visual Reasoning17 Aug 2023 0 repositories listed
-
Training BERT Models to Carry Over a Coding System Developed on One Corpus to Another7 Aug 2023 0 repositories listed
-
CTP-Net: Character Texture Perception Network for Document Image Forgery Localization4 Aug 2023 0 repositories listed
-
Making the V in Text-VQA Matter1 Aug 2023 0 repositories listed
-
Toward Zero-shot Character Recognition: A Gold Standard Dataset with Radical-level Annotations1 Aug 2023 0 repositories listed
-
Optimizing the Neural Network Training for OCR Error Correction of Historical Hebrew Texts30 Jul 2023 0 repositories listed
-
Toward a Period-Specific Optimized Neural Network for OCR Error Correction of Historical Hebrew Texts30 Jul 2023 0 repositories listed
-
MataDoc: Margin and Text Aware Document Dewarping for Arbitrary Boundary24 Jul 2023 0 repositories listed
-
A comparative analysis of SRGAN models18 Jul 2023 0 repositories listed
-
Handwritten and Printed Text Segmentation: A Signature Case Study15 Jul 2023 0 repositories listed
-
Handwritten Text Recognition Using Convolutional Neural Network11 Jul 2023 0 repositories listed
-
A Novel Pipeline for Improving Optical Character Recognition through Post-processing Using Natural Language Processing9 Jul 2023 0 repositories listed
-
Artificial Eye for the Blind7 Jul 2023 0 repositories listed
-
Estimating Post-OCR Denoising Complexity on Numerical Texts3 Jul 2023 0 repositories listed
-
Fraunhofer SIT at CheckThat! 2023: Mixing Single-Modal Classifiers to Estimate the Check-Worthiness of Multi-Modal Tweets2 Jul 2023 0 repositories listed
-
Resume Information Extraction via Post-OCR Text Processing23 Jun 2023 0 repositories listed
-
Weakly supervised information extraction from inscrutable handwritten document images12 Jun 2023 0 repositories listed
-
Transformer-Based UNet with Multi-Headed Cross-Attention Skip Connections to Eliminate Artifacts in Scanned Documents5 Jun 2023 0 repositories listed
-
Improving Handwritten OCR with Training Samples Generated by Glyph Conditional Denoising Diffusion Probabilistic Model31 May 2023 0 repositories listed
-
People and Places of Historical Europe: Bootstrapping Annotation Pipeline and a New Corpus of Named Entities in Late Medieval Texts26 May 2023 0 repositories listed
-
23 May 2023 0 repositories listed
-
TextDiffuser: Diffusion Models as Text Painters18 May 2023 0 repositories listed
-
Sequence-to-Sequence Pre-training with Unified Modality Masking for Visual Document Understanding16 May 2023 0 repositories listed
-
Text Reading Order in Uncontrolled Conditions by Sparse Graph Segmentation4 May 2023 0 repositories listed
-
Evaluating BERT-based Scientific Relation Classifiers for Scholarly Knowledge Graph Construction on Digital Library Collections3 May 2023 0 repositories listed
-
ICDAR 2023 Competition on Reading the Seal Title24 Apr 2023 0 repositories listed
-
Multimodal Short Video Rumor Detection System Based on Contrastive Learning17 Apr 2023 0 repositories listed
-
Cleansing Jewel: A Neural Spelling Correction Model Built On Google OCR-ed Tibetan Manuscripts7 Apr 2023 0 repositories listed
-
Linking Representations with Multimodal Contrastive Learning7 Apr 2023 0 repositories listed
-
A semi-automatic method for document classification in the shipping industry29 Mar 2023 0 repositories listed
-
CLIP-ReIdent: Contrastive Training for Player Re-Identification21 Mar 2023 0 repositories listed
-
Optical Character Recognition and Transcription of Berber Signs from Images in a Low-Resource Language Amazigh21 Mar 2023 0 repositories listed
-
The System Description of dun_oscar team for The ICPR MSR Challenge13 Mar 2023 0 repositories listed
-
Meme Sentiment Analysis Enhanced with Multimodal Spatial Encoding and Facial Embedding3 Mar 2023 0 repositories listed
-
User-Centric Evaluation of OCR Systems for Kwak'wala26 Feb 2023 0 repositories listed
-
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning9 Feb 2023 0 repositories listed
-
DEVICE: DEpth and VIsual ConcEpts Aware Transformer for TextCaps3 Feb 2023 0 repositories listed
-
SPARLING: Learning Latent Representations with Extremely Sparse Activations3 Feb 2023 0 repositories listed
-
On the feasibility of attacking Thai LPR systems with adversarial examples13 Jan 2023 0 repositories listed
-
Improving Inference Performance of Machine Learning with the Divide-and-Conquer Principle12 Jan 2023 0 repositories listed
-
Semantic rule Web-based Diagnosis and Treatment of Vector-Borne Diseases using SWRL rules8 Jan 2023 0 repositories listed
-
Bengali Handwritten Digit Recognition using CNN with Explainable AI23 Dec 2022 0 repositories listed
-
Towards Robust Handwritten Text Recognition with On-the-fly User Participation17 Dec 2022 0 repositories listed
-
Geometric Rectification of Creased Document Images based on Isometric Mapping16 Dec 2022 0 repositories listed
-
SceneGATE: Scene-Graph based co-Attention networks for TExt visual question answering16 Dec 2022 0 repositories listed
-
Extending TrOCR for Text Localization-Free OCR of Full-Page Scanned Receipt Images11 Dec 2022 0 repositories listed
-
PACMAN: a framework for pulse oximeter digit detection and reading in a low-resource setting9 Dec 2022 0 repositories listed
-
OCR-RTPS: An OCR-based real-time positioning system for the valet parking8 Dec 2022 0 repositories listed
-
Information Retrieval from the Digitized Books2 Dec 2022 0 repositories listed
-
Chart-RCNN: Efficient Line Chart Data Extraction from Camera Images25 Nov 2022 0 repositories listed
-
Look, Read and Ask: Learning to Ask Questions by Reading Text in Images23 Nov 2022 0 repositories listed
-
Out-of-Candidate Rectification for Weakly Supervised Semantic Segmentation22 Nov 2022 0 repositories listed
-
Text-Aware Dual Routing Network for Visual Question Answering17 Nov 2022 0 repositories listed
-
ChartParser: Automatic Chart Parsing for Print-Impaired16 Nov 2022 0 repositories listed
-
27 Oct 2022 0 repositories listed
-
A Late Multi-Modal Fusion Model for Detecting Hybrid Spam E-mail26 Oct 2022 0 repositories listed
-
15 Oct 2022 0 repositories listed
-
MenuAI: Restaurant Food Recommendation System via a Transformer-based Deep Learning Model15 Oct 2022 0 repositories listed
-
Key Information Extraction in Purchase Documents using Deep Learning and Rule-based Corrections7 Oct 2022 0 repositories listed
-
EraseNet: A Recurrent Residual Network for Supervised Document Cleaning3 Oct 2022 0 repositories listed
-
Synthesizing Annotated Image and Video Data Using a Rendering-Based Pipeline for Improved License Plate Recognition28 Sep 2022 0 repositories listed
-
3D Rendering Framework for Data Augmentation in Optical Character Recognition27 Sep 2022 0 repositories listed
-
Toward 3D Spatial Reasoning for Human-like Text-based Visual Question Answering21 Sep 2022 0 repositories listed
-
Out-of-Vocabulary Challenge Report14 Sep 2022 0 repositories listed
-
Computer vision based vehicle tracking as a complementary and scalable approach to RFID tagging13 Sep 2022 0 repositories listed
-
Document Image Binarization in JPEG Compressed Domain using Dual Discriminator Generative Adversarial Networks13 Sep 2022 0 repositories listed
-
OCR for TIFF Compressed Document Images Directly in Compressed Domain Using Text segmentation and Hidden Markov Model13 Sep 2022 0 repositories listed
-
PreSTU: Pre-Training for Scene-Text Understanding12 Sep 2022 0 repositories listed
-
A Masked Bounding-Box Selection Based ResNet Predictor for Text Rotation Prediction6 Sep 2022 0 repositories listed
-
You’ve translated it, now what?1 Sep 2022 0 repositories listed
-
A Black-Box Attack on Optical Character Recognition Systems30 Aug 2022 0 repositories listed
-
An Energy Activity Dataset for Smart Homes29 Aug 2022 0 repositories listed
-
Effectiveness of Mining Audio and Text Pairs from Public Data for Improving ASR Systems for Low-Resource Languages26 Aug 2022 0 repositories listed
-
Visual Subtitle Feature Enhanced Video Outline Generation24 Aug 2022 0 repositories listed
-
An End-to-End OCR Framework for Robust Arabic-Handwriting Recognition using a Novel Transformers-based Model and an Innovative 270 Million-Words Multi-Font Corpus of Classical Arabic with Diacritics20 Aug 2022 0 repositories listed
-
To show or not to show: Redacting sensitive text from videos of electronic displays19 Aug 2022 0 repositories listed
-
Information Extraction from Scanned Invoice Images using Text Analysis and Layout Features8 Aug 2022 0 repositories listed
-
Optimal Boxes: Boosting End-to-End Scene Text Recognition by Adjusting Annotated Bounding Boxes via Reinforcement Learning25 Jul 2022 0 repositories listed
-
DEXTER: An end-to-end system to extract table contents from electronic medical health documents14 Jul 2022 0 repositories listed
-
GMN: Generative Multi-modal Network for Practical Document Information Extraction11 Jul 2022 0 repositories listed
-
Towards Multimodal Vision-Language Models Generating Non-Generic Text9 Jul 2022 0 repositories listed