Datasets › Quora Question Pairs › Papers where code ran, page 1
Quora Question Pairs
Papers archive 2025-07-28
papers with a benchmark row: 45 · with a code link: 38 · where Syntology ran a sample: 25 (22 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument) Syntology
Show:
all papers with a benchmark row only where code ran (25 of 45 with a benchmark row: 22 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.
Page 1 of 1: papers 1 to 25 of the 25 papers with a benchmark row here where Syntology ran at least one harvested sample (22 with a run with no instrument failure, 3 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
A run counts when it is a run with no instrument failure any run, instrument failures included The first hides, in your browser, the papers where every run was a failure of Syntology's instrument; the second shows every paper on this page.
The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 55. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since 2026-09-28 , the first build that kept a record date for it; when this build read Syntology's graph is in the build record .
Paper Code Results Date Samples run Syntology
BM25S: Orders of magnitude faster lexical search via eager sparse scoring
3
5
4 Jul 2024
official (archive's flag): 10 ran · 15 ran (of which 0 constructed an object rather than computing a result; 15 with no instrument failure: 0 honoured, 0 violated, 15 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified
Hierarchical Sketch Induction for Paraphrase Generation
1
1
7 Mar 2022
official (archive's flag): 3 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language
12
1
7 Feb 2022
community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified
Charformer: Fast Character Transformers via Gradient-based Subword Tokenization
2
1
23 Jun 2021
community repositories only · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified
FNet: Mixing Tokens with Fourier Transforms
12
1
9 May 2021
community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence )
Entailment as Few-Shot Learner
3
1
29 Apr 2021
1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified
Intrinsic Dimensionality Explains the Effectiveness of Language Model Fine-Tuning
2
1
22 Dec 2020
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence )
Big Bird: Transformers for Longer Sequences
14
1
28 Jul 2020
official (archive's flag): 1 ran · 10 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 8 where Syntology's instrument failed) · 5 unverified (11 pointer-only for licence )
SqueezeBERT: What can computer vision teach NLP about efficient neural networks?
6
1
19 Jun 2020
community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
DeBERTa: Decoding-enhanced BERT with Disentangled Attention
14
1
5 Jun 2020
official: no sample here; runs from other or unrecorded repositories · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 9 unverified (3 pointer-only for licence )
ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators
19
1
23 Mar 2020
official: no sample here; runs from other or unrecorded repositories · 31 ran (of which 7 constructed an object rather than computing a result; 18 with no instrument failure: 2 honoured, 2 violated, 14 with no contract checked; 13 where Syntology's instrument failed) · 9 unverified (10 pointer-only for licence )
SMART: Robust and Efficient Fine-Tuning for Pre-trained Natural Language Models through Principled Regularized Optimization
6
3
8 Nov 2019
official: no sample here; runs from other or unrecorded repositories · 8 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 2 honoured, 1 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence )
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
57
5
23 Oct 2019
21 ran (of which 0 constructed an object rather than computing a result; 20 with no instrument failure: 1 honoured, 0 violated, 19 with no contract checked; 1 where Syntology's instrument failed) · 10 unverified
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
37
1
2 Oct 2019
official (archive's flag): 1 ran · 21 ran (of which 5 constructed an object rather than computing a result; 13 with no instrument failure: 3 honoured, 1 violated, 9 with no contract checked; 8 where Syntology's instrument failed) · 6 unverified (2 pointer-only for licence )
ALBERT: A Lite BERT for Self-supervised Learning of Language Representations
48
1
26 Sep 2019
official (archive's flag): 6 ran · 81 ran (of which 17 constructed an object rather than computing a result; 59 with no instrument failure: 4 honoured, 0 violated, 55 with no contract checked; 22 where Syntology's instrument failed) · 45 unverified (28 pointer-only for licence )
Simple and Effective Text Matching with Richer Alignment Features
3
2
1 Aug 2019
community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence )
RoBERTa: A Robustly Optimized BERT Pretraining Approach
67
1
26 Jul 2019
community repositories only · 37 ran (of which 11 constructed an object rather than computing a result; 36 with no instrument failure: 0 honoured, 0 violated, 36 with no contract checked; 1 where Syntology's instrument failed) · 11 unverified (24 pointer-only for licence )
SpanBERT: Improving Pre-training by Representing and Predicting Spans
6
1
24 Jul 2019
official (archive's flag): 2 ran · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 3 where Syntology's instrument failed) · 6 unverified (6 pointer-only for licence )
XLNet: Generalized Autoregressive Pretraining for Language Understanding
27
2
19 Jun 2019
official (archive's flag): 1 ran · 15 ran (of which 2 constructed an object rather than computing a result; 10 with no instrument failure: 2 honoured, 0 violated, 8 with no contract checked; 5 where Syntology's instrument failed) · 9 unverified (4 pointer-only for licence )
ERNIE: Enhanced Language Representation with Informative Entities
2
1
17 May 2019
official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (1 pointer-only for licence )
Multi-Task Deep Neural Networks for Natural Language Understanding
7
1
31 Jan 2019
community repositories only · 9 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 1 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (4 pointer-only for licence )
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
534
1
11 Oct 2018
official: no sample here; runs from other or unrecorded repositories · 300 ran (of which 75 constructed an object rather than computing a result; 235 with no instrument failure: 17 honoured, 4 violated, 214 with no contract checked; 65 where Syntology's instrument failed) · 359 unverified (164 pointer-only for licence )
Learning General Purpose Distributed Sentence Representations via Large Scale Multi-task Learning
4
1
30 Mar 2018
official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence )
Natural Language Inference over Interaction Space
2
1
13 Sep 2017
community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
Bilateral Multi-Perspective Matching for Natural Language Sentences
10
1
13 Feb 2017
1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (1 pointer-only for licence )