Browse State-of-the-Art › document understanding › Papers, page 3
document understanding
Papers archive 2025-07-28
archive papers tagged: 309 · with a code link: 140 · where Syntology ran a sample: 39 (30 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (39 of 309 tagged: 30 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)
Page 3 of 4: papers 201 to 300 of 309, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications12 Jun 2024 0 repositories listed
-
Retrieval Augmented Structured Generation: Business Document Information Extraction As Tool Use30 May 2024 0 repositories listed
-
Notes on Applicability of GPT-4 to Document Understanding28 May 2024 0 repositories listed
-
CREPE: Coordinate-Aware End-to-End Document Parser1 May 2024 0 repositories listed
-
16 Apr 2024 0 repositories listed
-
Towards Efficient Resume Understanding: A Multi-Granularity Multi-Modal Pre-Training Approach13 Apr 2024 0 repositories listed
-
HRVDA: High-Resolution Visual Document Assistant10 Apr 2024 0 repositories listed
-
BuDDIE: A Business Document Dataset for Multi-task Information Extraction5 Apr 2024 0 repositories listed
-
Can AI Models Appreciate Document Aesthetics? An Exploration of Legibility and Layout Quality in Relation to Prediction Confidence27 Mar 2024 0 repositories listed
-
LayoutLLM: Large Language Model Instruction Tuning for Visually Rich Document Understanding21 Mar 2024 0 repositories listed
-
Enhancing Visual Document Understanding with Contrastive Learning in Large Visual-Language Models29 Feb 2024 0 repositories listed
-
Read and Think: An Efficient Step-wise Multimodal Language Model for Document Understanding and Reasoning26 Feb 2024 0 repositories listed
-
RJUA-MedDQA: A Multimodal Benchmark for Medical Document Question Answering and Clinical Reasoning19 Feb 2024 0 repositories listed
-
15 Feb 2024 0 repositories listed
-
LongFin: A Multimodal Document Understanding Model for Long Financial Domain Documents26 Jan 2024 0 repositories listed
-
DocGraphLM: Documental Graph Language Model for Information Extraction5 Jan 2024 0 repositories listed
-
On Scaling Up a Multilingual Vision and Language Model1 Jan 2024 0 repositories listed
-
DocLLM: A layout-aware generative language model for multimodal document understanding31 Dec 2023 0 repositories listed
-
SLJP: Semantic Extraction based Legal Judgment Prediction13 Dec 2023 0 repositories listed
-
DocPedia: Unleashing the Power of Large Multimodal Model in the Frequency Domain for Versatile Document Understanding20 Nov 2023 0 repositories listed
-
Efficient End-to-End Visual Document Understanding with Rationale Distillation16 Nov 2023 0 repositories listed
-
DONUT-hole: DONUT Sparsification by Harnessing Knowledge and Optimizing Learning Efficiency9 Nov 2023 0 repositories listed
-
A Multi-Modal Multilingual Benchmark for Document Image Classification25 Oct 2023 0 repositories listed
-
Reinforced UI Instruction Grounding: Towards a Generic UI Task Automation API7 Oct 2023 0 repositories listed
-
ProtoNER: Few shot Incremental Learning for Named Entity Recognition using Prototypical Networks3 Oct 2023 0 repositories listed
-
Finding Pragmatic Differences Between Disciplines30 Sep 2023 0 repositories listed
-
Document Understanding for Healthcare Referrals22 Sep 2023 0 repositories listed
-
21 Sep 2023 0 repositories listed Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
KOSMOS-2.5: A Multimodal Literate Model20 Sep 2023 0 repositories listed
-
11 Sep 2023 0 repositories listed
-
Attention Where It Matters: Rethinking Visual Document Understanding with Selective Region Concentration3 Sep 2023 0 repositories listed
-
Workshop on Document Intelligence Understanding31 Jul 2023 0 repositories listed
-
MataDoc: Margin and Text Aware Document Dewarping for Arbitrary Boundary24 Jul 2023 0 repositories listed
-
A Survey and Approach to Chart Classification9 Jul 2023 0 repositories listed
-
DocumentNet: Bridging the Data Gap in Document Pre-Training15 Jun 2023 0 repositories listed
-
30 May 2023 0 repositories listed
-
AWESOME: GPU Memory-constrained Long Document Summarization using Memory Mechanism and Global Salient Content24 May 2023 0 repositories listed
-
23 May 2023 0 repositories listed
-
Fast-StrucTexT: An Efficient Hourglass Transformer with Modality-guided Dynamic Token Merge for Document Understanding19 May 2023 0 repositories listed
-
DLUE: Benchmarking Document Language Understanding16 May 2023 0 repositories listed
-
Sequence-to-Sequence Pre-training with Unified Modality Masking for Visual Document Understanding16 May 2023 0 repositories listed
-
M⁶Doc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout Analysis15 May 2023 0 repositories listed
-
Two to Five Truths in Non-Negative Matrix Factorization6 May 2023 0 repositories listed
-
FormNetV2: Multimodal Graph Contrastive Learning for Form Document Information Extraction4 May 2023 0 repositories listed
-
Revisiting Table Detection Datasets for Visually Rich Documents4 May 2023 0 repositories listed
-
What Makes a Good Dataset for Symbol Description Reading?17 Apr 2023 0 repositories listed
-
13 Apr 2023 0 repositories listed
-
29 Nov 2022 0 repositories listed
-
VRDU: A Benchmark for Visually-rich Document Understanding15 Nov 2022 0 repositories listed
-
QueryForm: A Simple Zero-shot Form Entity Query Framework14 Nov 2022 0 repositories listed
-
Unimodal and Multimodal Representation Training for Relation Extraction11 Nov 2022 0 repositories listed
-
16 Oct 2022 0 repositories listed
-
ERNIE-mmLayout: Multi-grained MultiModal Transformer for Document Understanding18 Sep 2022 0 repositories listed
-
One-Shot Doc Snippet Detection: Powering Search in Document Beyond Text12 Sep 2022 0 repositories listed
-
Improving Keyphrase Extraction with Data Augmentation and Information Filtering11 Sep 2022 0 repositories listed
-
DeeperDive: The Unreasonable Effectiveness of Weak Supervision in Document Understanding A Case Study in Collaboration with UiPath Inc17 Aug 2022 0 repositories listed
-
Understanding Long Documents with Different Position-Aware Attentions17 Aug 2022 0 repositories listed
-
Towards Complex Document Understanding By Discrete Reasoning25 Jul 2022 0 repositories listed
-
Bi-VLDoc: Bidirectional Vision-Language Modeling for Visually-Rich Document Understanding27 Jun 2022 0 repositories listed
-
Test-Time Adaptation for Visual Document Understanding15 Jun 2022 0 repositories listed
-
RDU: A Region-based Approach to Form-style Document Understanding14 Jun 2022 0 repositories listed
-
Génération de question à partir d’analyse sémantique pour l’adaptation non supervisée de modèles de compréhension de documents (Question generation from semantic analysis for unsupervised adaptation of document understanding models)1 Jun 2022 0 repositories listed
-
MATrIX -- Modality-Aware Transformer for Information eXtraction17 May 2022 0 repositories listed
-
XFUND: A Benchmark Dataset for Multilingual Visually Rich Form Understanding1 May 2022 0 repositories listed
-
22 Apr 2022 0 repositories listed
-
Robust Text Line Detection in Historical Documents: Learning and Evaluation Methods23 Mar 2022 0 repositories listed
-
FormNet: Structural Encoding beyond Sequential Modeling in Form Document Information Extraction16 Mar 2022 0 repositories listed
-
Hierarchical BERT for Medical Document Understanding11 Mar 2022 0 repositories listed
-
WebFormer: The Web-page Transformer for Structure Information Extraction1 Feb 2022 0 repositories listed
-
Efficient layout-aware pretraining for multimodal form understanding16 Jan 2022 0 repositories listed
-
LoPE: Learnable Sinusoidal Positional Encoding for Improving Document Transformer Model16 Jan 2022 0 repositories listed
-
UniDoc: Unified Pretraining Framework for Document Understanding1 Dec 2021 0 repositories listed
-
PSG: Prompt-based Sequence Generation for Acronym Extraction29 Nov 2021 0 repositories listed
-
SimCLAD: A Simple Framework for Contrastive Learning of Acronym Disambiguation29 Nov 2021 0 repositories listed
-
Document Layout Analysis with Aesthetic-Guided Image Augmentation27 Nov 2021 0 repositories listed
-
Handling tree-structured text: parsing directory pages24 Nov 2021 0 repositories listed
-
Probing Position-Aware Attention Mechanism in Long Document Understanding16 Nov 2021 0 repositories listed
-
OPAD: An Optimized Policy-based Active Learning Framework for Document Content Analysis1 Oct 2021 0 repositories listed
-
Position Masking for Improved Layout-Aware Document Understanding1 Sep 2021 0 repositories listed
-
The Law of Large Documents: Understanding the Structure of Legal Contracts Using Visual Cues16 Jul 2021 0 repositories listed
-
Leveraging Domain Agnostic and Specific Knowledge for Acronym Disambiguation1 Jul 2021 0 repositories listed
-
27 Apr 2021 0 repositories listed
-
LAMPRET: Layout-Aware Multimodal PreTraining for Document Understanding16 Apr 2021 0 repositories listed
-
Automatic Knowledge Extraction with Human Interface9 Apr 2021 0 repositories listed
-
Decontextualization: Making Sentences Stand-Alone9 Feb 2021 0 repositories listed
-
AT-BERT: Adversarial Training BERT for Acronym Identification Winning Solution for SDU@AAAI-2111 Jan 2021 0 repositories listed
-
BROS: A Pre-trained Language Model for Understanding Texts in Document1 Jan 2021 0 repositories listed
-
Acronym Identification and Disambiguation Shared Tasks for Scientific Document Understanding22 Dec 2020 0 repositories listed
-
Merge and Recognize: A Geometry and 2D Context Aware Graph Model for Named Entity Recognition from Visual Documents1 Dec 2020 0 repositories listed
-
Friendly Topic Assistant for Transformer Based Abstractive Summarization1 Nov 2020 0 repositories listed
-
Attention-Based Graph Neural Network with Global Context Awareness for Document Understanding1 Oct 2020 0 repositories listed
-
Hierarchical GPT with Congruent Transformers for Multi-Sentence Language Models18 Sep 2020 0 repositories listed
-
Multi-modal Information Extraction from Text, Semi-structured, and Tabular Data on the Web1 Jul 2020 0 repositories listed
-
Scalable Cross Lingual Pivots to Model Pronoun Gender for Translation16 Jun 2020 0 repositories listed
-
Table Structure Extraction with Bi-directional Gated Recurrent Unit Networks8 Jan 2020 0 repositories listed
-
BERT-AL: BERT for Arbitrarily Long Document Understanding1 Jan 2020 0 repositories listed
-
Table-Of-Contents generation on contemporary documents20 Nov 2019 0 repositories listed
-
A Retrospective Recount of Computer Architecture Research with a Data-Driven Study of Over Four Decades of ISCA Publications22 Jun 2019 0 repositories listed
-
A User-Centered Concept Mining System for Query and Document Understanding at Tencent21 May 2019 0 repositories listed
-
Graph Convolution for Multimodal Information Extraction from Visually Rich Documents27 Mar 2019 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.