Browse State-of-the-Art › Form
Form
436 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 436 papers with code (1,618 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
10 Jun 2018 30 repositories listed Syntology ran 1 of 7 samples · 6 unverifiedWe present a generative image inpainting system to complete images with free-form mask and guidance.
-
19 Feb 2018 12 repositories listed Syntology ran 10 of 16 samples · 6 unverified · 3 pointer-only (licence)Photorealistic image stylization concerns transferring style of a reference photo to a content photo with the constraint that the stylized photo should remain photorealistic.
-
13 Jul 2020 11 repositories listed Syntology ran 7 of 14 samples · 7 unverifiedA rich set of interpretable dimensions has been shown to emerge in the latent space of the Generative Adversarial Networks (GANs) trained for synthesizing images.
-
2 Oct 2018 7 repositories listed Syntology ran 4 of 5 samples · 1 unverifiedThe result is a continuous-time invertible generative model with unbiased density estimation and one-pass sampling, while allowing unrestricted neural network architectures.
-
18 Apr 2021 6 repositories listedIn this paper, we present LayoutXLM, a multimodal pre-trained model for multilingual document understanding, which aims to bridge the language barriers for visually-rich document understanding.
-
21 May 2018 5 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedThe main idea is to teach a deep network to use standard machine learning tools, such as ridge regression, as part of its own internal model, enabling it to quickly adapt to novel data.
-
6 Jan 2016 5 repositories listed Syntology ran 1 of 7 samples · 6 unverified · 1 pointer-only (licence)Semantic parsing aims at mapping natural language to machine interpretable meaning representations.
-
23 May 2023 4 repositories listed Syntology ran 10 of 15 samples · 5 unverifiedEvaluating the factuality of long-form text generated by large language models (LMs) is non-trivial because (1) generations often contain a mixture of supported and unsupported pieces of information, making binary…
-
23 Oct 2019 4 repositories listed Syntology ran 3 of 3 samples · 0 unverifiedDespite the ability to produce human-level speech for in-domain text, attention-based end-to-end text-to-speech (TTS) systems suffer from text alignment failures that increase in frequency for out-of-domain text.
-
18 Feb 2019 4 repositories listedWe present a novel image editing system that generates images as the user provides free-form mask, sketch and color as an input.
-
27 Mar 2024 3 repositories listed Syntology ran 3 of 5 samples · 2 unverified · 4 pointer-only (licence)Empirically, we demonstrate that LLM agents can outperform crowdsourced human annotators - on a set of ~16k individual facts, SAFE agrees with crowdsourced human annotators 72% of the time, and on a random subset of 100…
-
21 Mar 2024 3 repositories listedDespite their simplicity, linear models perform well at time series forecasting, even when pitted against deeper and more expensive models.
-
3 Jul 2023 3 repositories listed Syntology ran 2 of 2 samples · 0 unverifiedLarge Language Models (LLMs) show promising results in language generation and instruction following but frequently "hallucinate", making their outputs less reliable.
-
18 May 2023 3 repositories listedThat is, the embedding space of head classes severely compresses that of tail classes, which is not conducive to subsequent classifier learning.
-
14 Dec 2022 3 repositories listedFor the retriever, we adopt a number-aware negative sampling strategy to enable the retriever to be more discriminative on key numerical facts.
-
22 Oct 2022 3 repositories listedTo this end, we propose a comprehensive logical reasoning explanation form.
-
16 Mar 2022 3 repositories listed Syntology ran 8 of 14 samples · 6 unverifiedUnsupervised semantic segmentation aims to discover and localize semantically meaningful categories within image corpora without any form of annotation.
-
17 May 2021 3 repositories listedFUDGE edits the graph structure by combining text segments (graph vertices) and pruning edges in an iterative fashion to obtain the final text entities and relationships.
-
5 Sep 2019 3 repositories listedAutomatic, template-free extraction of information from form images is challenging due to the variety of form layouts.
-
22 Jul 2019 3 repositories listed Syntology ran 4 of 11 samples · 7 unverified · 4 pointer-only (licence)We introduce the first large-scale corpus for long-form question answering, a task requiring elaborate and in-depth answers to open-ended questions.
-
3 Jun 2019 3 repositories listedAlthough plenty of methods have been proposed, a theoretical analysis of feature transform is still missing.
-
27 May 2019 3 repositories listedWe present a new dataset for form understanding in noisy scanned documents (FUNSD) that aims at extracting and structuring the textual content of forms.
-
3 Jan 2019 3 repositories listedWe investigate adversarial learning in the case when only an unnormalized form of the density can be accessed, rather than samples.
-
3 Sep 2018 3 repositories listedWe propose a novel methodology to generate domain-specific large-scale question answering (QA) datasets by re-purposing existing annotations for other NLP tasks.
-
7 Feb 2025 2 repositories listedTo address this issue, we propose Self-Loop Latent Swap, a frame-level bidirectional swap applied to the overlapping region of adjacent views.
-
7 Jan 2025 2 repositories listedThis paper presents ICAT, an evaluation framework for measuring coverage of diverse factual information in long-form text generation.
-
19 Sep 2024 2 repositories listedTo train CodePlan, we construct a large-scale dataset of 2M examples that integrate code-form plans with standard prompt-response pairs from existing corpora.
-
3 Sep 2024 2 repositories listed Syntology ran 8 of 8 samples · 0 unverified · 8 pointer-only (licence)Current benchmarks like Needle-in-a-Haystack (NIAH), Ruler, and Needlebench focus on models' ability to understand long-context input sequences but fail to capture a critical dimension: the generation of high-quality…
-
28 Jun 2024 2 repositories listed Syntology ran 5 of 7 samples · 2 unverifiedThere is a growing body of work seeking to replicate the success of machine learning (ML) on domains like computer vision (CV) and natural language processing (NLP) to applications involving biophysical data.
-
27 May 2024 2 repositories listedIn the domain of Document AI, parsing semi-structured image form is a crucial Key Information Extraction (KIE) task.
Syntology lines on 14 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections