Browse State-of-the-Art › Articles
Articles
1,123 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 1,123 papers with code (4,012 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
28 May 2020 67 repositories listed Syntology ran 15 of 65 samples · 50 unverified · 4 pointer-only (licence)By contrast, humans can generally perform a new language task from only a few examples or from simple instructions - something which current NLP systems still largely struggle to do.
-
9 Jan 2019 37 repositories listed Syntology ran 63 of 143 samples · 80 unverified · 43 pointer-only (licence)Transformers have a potential of learning longer-term dependency, but are limited by a fixed-length context in the setting of language modeling.
-
16 Jun 2016 21 repositories listed Syntology ran 4 of 6 samples · 2 unverified · 6 pointer-only (licence)We present the Stanford Question Answering Dataset (SQuAD), a new reading comprehension dataset consisting of 100, 000+ questions posed by crowdworkers on a set of Wikipedia articles, where the answer to each question…
-
28 Feb 2010 12 repositories listed Syntology ran 1 of 3 samples · 2 unverified · 1 pointer-only (licence)In this work, we model personalized recommendation of news articles as a contextual bandit problem, a principled approach in which a learning algorithm sequentially selects articles to serve users based on contextual…
-
31 Mar 2017 10 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)This paper proposes to tackle open- domain question answering using Wikipedia as the unique knowledge source: the answer to any factoid question is a text span in a Wikipedia article.
-
18 Oct 2018 9 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Sequence-to-sequence models have recently gained the state of the art performance in summarization.
-
21 Jul 2016 8 repositories listed Syntology ran 0 of 8 samples · 8 unverifiedGeometrically, gender bias is first shown to be captured by a direction in the word embedding.
-
8 Feb 2024 6 repositories listed Syntology ran 9 of 18 samples · 9 unverifiedIn this survey, we carefully review over 300 articles, focusing on KG-aware research in two principal aspects: KG-driven Multi-Modal (KG4MM) learning, where KGs support multi-modal tasks, and Multi-Modal Knowledge Graph…
-
25 Nov 2019 6 repositories listed Syntology ran 2 of 6 samples · 4 unverifiedIn addition, we propose a new Tree-Edit-Distance-based Similarity (TEDS) metric for table recognition, which more appropriately captures multi-hop cell misalignment and OCR errors than the pre-established metric.
-
29 Aug 2019 6 repositories listedWe introduce a new classification task for scientific statements and release a large-scale dataset for supervised learning.
-
16 Aug 2019 6 repositories listed Syntology ran 7 of 13 samples · 6 unverified · 13 pointer-only (licence)Deep neural networks that are developed for computer vision have been proven to be an effective method to analyze layout of document images.
-
10 Jul 2019 6 repositories listed Syntology ran 1 of 3 samples · 2 unverifiedWe present an approach based on multilingual sentence embeddings to automatically extract parallel sentences from the content of Wikipedia articles in 85 languages, including several dialects or low-resource languages.
-
30 Apr 2018 6 repositories listedWe present NEWSROOM, a summarization dataset of 1.
-
2 Mar 2023 5 repositories listedTherefore, training an effective generalist biomedical model requires high-quality multimodal data, such as parallel image-text pairs.
-
29 Aug 2018 5 repositories listedWe introduce a multi-task setup of identifying and classifying entities, relations, and coreference clusters in scientific articles.
-
5 Feb 2018 5 repositories listedThis paper is an attempt to explain all the matrix calculus you need in order to understand the training of deep neural networks.
-
25 Mar 2024 4 repositories listedNSINA is the largest news corpus for Sinhala, available up to date.
-
26 Jan 2023 4 repositories listed Syntology ran 3 of 9 samples · 6 unverified · 2 pointer-only (licence)In this paper, we identify a property of the structure of an LLM's probability function that is useful for such detection.
-
22 Mar 2020 4 repositories listedIn this paper, we model the problem of finding the relationship between two documents as a pairwise document classification task.
-
20 Dec 2019 4 repositories listedIn this paper, we propose a novel deep neural network DP-LSTM for stock price prediction, which incorporates the news articles as hidden information and integrates difference news sources through the differential…
-
16 Oct 2019 4 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)An alternative to building large monolingual training datasets is to develop cross-lingual systems which can transfer to a target language without requiring training data in that language.
-
31 Aug 2019 4 repositories listed Syntology ran 2 of 6 samples · 4 unverified · 3 pointer-only (licence)0Shot-TC aims to associate an appropriate label with a piece of text, irrespective of the text domain and the aspect (e.
-
18 Nov 2018 4 repositories listedA significant improvement is WGAN, with the help of 1-Lipschitz constraint on discriminator to prevent from gradient vanishing.
-
22 Oct 2018 4 repositories listedSentence embeddings have become an essential part of today's natural language processing (NLP) systems, especially together advanced deep learning methods.
-
4 Apr 2018 4 repositories listedWord embeddings are a popular approach to unsupervised learning of word relationships that are widely used in natural language processing.
-
30 Jan 2018 4 repositories listedWe show that generating English Wikipedia articles can be approached as a multi- document summarization of source documents.
-
5 Dec 2015 4 repositories listedWe describe an application of an encoder-decoder recurrent neural network with LSTM units and attention to generating headlines from the text of news articles.
-
10 May 2015 4 repositories listedIn recent years, There has been a variety of research on discourse parsing, particularly RST discourse parsing.
-
29 Sep 2024 3 repositories listedThis paper shows that fine-tuning a language model on synthetic data using an LM and using a character level Markov corruption process can significantly improve the ability to correct OCR errors.
-
8 Jul 2024 3 repositories listedData owners may request the removal of their data from a trained model due to privacy or copyright concerns.
Syntology lines on 14 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections