Datasets › SQuAD › Papers where code ran, page 1

SQuAD (Stanford Question Answering Dataset)

Papers archive 2025-07-28

papers with a benchmark row: 82 · with a code link: 63 · where Syntology ran a sample: 32 (29 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology

Show: all papers with a benchmark rowonly where code ran (32 of 82 with a benchmark row: 29 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)

Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.

Page 1 of 1: papers 1 to 32 of the 32 papers with a benchmark row here where Syntology ran at least one harvested sample (29 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.

The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 151. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record.

PaperCodeResultsDateSamples run Syntology
Blended RAG: Improving RAG (Retriever-Augmented Generation) Accuracy with Semantic Search and Hybrid Query-Based Retrievers 1 2 22 Mar 2024 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
TextBox 2.0: A Text Generation Library with Pre-trained Language Models 1 2 26 Dec 2022 official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result
DyREx: Dynamic Query Representation for Extractive Question Answering 1 1 26 Oct 2022 official (archive's flag): 1 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence)
LinkBERT: Pretraining Language Models with Document Links 1 1 29 Mar 2022 official: harvested, nothing ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 11 unverified
ZeroGen: Efficient Zero-shot Learning via Dataset Generation 3 1 16 Feb 2022 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (7 pointer-only for licence)
Pay Attention to MLPs 20 1 17 May 2021 34 ran (of which 15 constructed an object rather than computing a result; 31 with no instrument failure: 1 honoured, 4 violated, 26 with no contract checked; 3 where Syntology's instrument failed) · 10 unverified (11 pointer-only for licence)
LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attention 9 7 2 Oct 2020 official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified
DeBERTa: Decoding-enhanced BERT with Disentangled Attention 14 1 5 Jun 2020 official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (3 pointer-only for licence)
Generating Diverse and Consistent QA pairs from Contexts with Information-Maximizing Hierarchical Conditional VAEs 1 2 28 May 2020 official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (3 pointer-only for licence)
UniLMv2: Pseudo-Masked Language Models for Unified Language Model Pre-Training 3 1 28 Feb 2020 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
Dice Loss for Data-imbalanced NLP Tasks 4 2 7 Nov 2019 official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension 47 1 29 Oct 2019 40 ran (of which 7 constructed an object rather than computing a result; 31 with no instrument failure: 1 honoured, 1 violated, 29 with no contract checked; 9 where Syntology's instrument failed) · 13 unverified (13 pointer-only for licence)
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer 57 5 23 Oct 2019 21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 1 honoured, 0 violated, 19 with no contract checked; 1 where Syntology's instrument failed) · 10 unverified
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter 37 2 2 Oct 2019 official (archive's flag): 1 ran · 21 ran (of which 5 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 1 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 6 unverified (2 pointer-only for licence)
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations 48 6 26 Sep 2019 official (archive's flag): 6 ran · 81 ran (of which 17 constructed an object rather than computing a result; 59 with no instrument failure: 4 honoured, 0 violated, 55 with no contract checked; 22 where Syntology's instrument failed) · 45 unverified (28 pointer-only for licence)
Semantics-aware BERT for Language Understanding 1 4 5 Sep 2019 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (3 pointer-only for licence)
RoBERTa: A Robustly Optimized BERT Pretraining Approach 67 2 26 Jul 2019 community repositories only · 37 ran (of which 11 constructed an object rather than computing a result; 36 with no instrument failure: 0 honoured, 0 violated, 36 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (24 pointer-only for licence)
SpanBERT: Improving Pre-training by Representing and Predicting Spans 6 3 24 Jul 2019 official (archive's flag): 2 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence)
XLNet: Generalized Autoregressive Pretraining for Language Understanding 27 4 19 Jun 2019 official (archive's flag): 1 ran · 15 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 5 where Syntology's instrument failed) · 9 unverified (4 pointer-only for licence)
Large Batch Optimization for Deep Learning: Training BERT in 76 minutes 32 1 1 Apr 2019 community repositories only · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (9 pointer-only for licence)
End-to-End Open-Domain Question Answering with BERTserini 1 1 5 Feb 2019 official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding 534 6 11 Oct 2018 official: no sample here; runs from other or unrecorded repositories · 300 ran (of which 75 constructed an object rather than computing a result; 235 with no instrument failure: 17 honoured, 4 violated, 214 with no contract checked; 65 where Syntology's instrument failed) · 359 unverified (164 pointer-only for licence)
QANet: Combining Local Convolution with Global Self-Attention for Reading Comprehension 15 4 23 Apr 2018 14 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 3 honoured, 0 violated, 7 with no contract checked; 4 where Syntology's instrument failed) · 5 unverified (3 pointer-only for licence)
Deep contextualized word representations 46 4 15 Feb 2018 30 ran (of which 12 constructed an object rather than computing a result; 24 with no instrument failure: 2 honoured, 0 violated, 22 with no contract checked; 6 where Syntology's instrument failed) · 28 unverified (25 pointer-only for licence)
Simple and Effective Multi-Paragraph Reading Comprehension 1 1 29 Oct 2017 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified
Simple Recurrent Units for Highly Parallelizable Recurrence 11 2 8 Sep 2017 community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
Learned in Translation: Contextualized Word Vectors 5 2 1 Aug 2017 community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
Neural Question Generation from Text: A Preliminary Study 6 1 6 Apr 2017 community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Reading Wikipedia to Answer Open-Domain Questions 10 3 31 Mar 2017 community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence)
Making Neural QA as Simple as Possible but not Simpler 3 3 14 Mar 2017 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
Bidirectional Attention Flow for Machine Comprehension 27 3 5 Nov 2016 official (archive's flag): 1 ran · 9 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 2 violated, 1 with no contract checked; 5 where Syntology's instrument failed) · 2 unverified (7 pointer-only for licence)
Machine Comprehension Using Match-LSTM and Answer Pointer 5 5 29 Aug 2016 official: no sample here; runs from other or unrecorded repositories · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)