Datasets › Arcade Learning Environment › Papers where code ran, page 1

Arcade Learning Environment

Papers archive 2025-07-28

papers with a benchmark row: 66 · with a code link: 56 · where Syntology ran a sample: 35 (26 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument) Syntology

Show: all papers with a benchmark rowonly where code ran (35 of 66 with a benchmark row: 26 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument)

Syntology We ran code from the paper's repository; we did not run it on this dataset or check it against this dataset's benchmarks.

Page 1 of 1: papers 1 to 35 of the 35 papers with a benchmark row here where Syntology ran at least one harvested sample (26 with a run with no instrument failure, 9 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.

The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset, not that list; the archive's count for this dataset is 366. The Syntology column is from Syntology's graph, stated per sample; it is not part of any archive number. A line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified” (C is a part of N, never taken away from it); the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. When the archive marks a repository official for the paper, the cell starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record.

PaperCodeResultsDateSamples run Syntology
IQ-Learn: Inverse soft-Q Learning for Imitation 5 4 23 Jun 2021 community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence)
Decision Transformer: Reinforcement Learning via Sequence Modeling 20 4 2 Jun 2021 community repositories only · 17 ran (of which 10 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 4 where Syntology's instrument failed) · 9 unverified (8 pointer-only for licence)
Adaptive Rational Activations to Boost Deep Reinforcement Learning 4 26 18 Feb 2021 official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
Reinforcement Learning with Latent Flow 2 2 6 Jan 2021 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
Mastering Atari with Discrete World Models 9 50 5 Oct 2020 community repositories only · 11 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified
Munchausen Reinforcement Learning 6 1 28 Jul 2020 community repositories only · 12 ran (of which 11 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified
First return, then explore 2 10 27 Apr 2020 official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence)
CURL: Contrastive Unsupervised Representations for Reinforcement Learning 7 24 8 Apr 2020 official (archive's flag): 6 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence)
Agent57: Outperforming the Atari Human Benchmark 5 51 30 Mar 2020 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model 18 51 19 Nov 2019 43 ran (of which 36 constructed an object rather than computing a result; 43 with no instrument failure: 4 honoured, 0 violated, 39 with no contract checked; 0 where Syntology's instrument failed) · 21 unverified (62 pointer-only for licence)
Fully Parameterized Quantile Function for Distributional Reinforcement Learning 6 23 5 Nov 2019 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified
Soft Actor-Critic for Discrete Action Settings 13 18 16 Oct 2019 official: harvested, nothing ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (3 pointer-only for licence)
Recurrent Independent Mechanisms 3 3 24 Sep 2019 14 ran (of which 9 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (20 pointer-only for licence)
Go-Explore: a New Approach for Hard-Exploration Problems 3 2 30 Jan 2019 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence)
Exploration by Random Network Distillation 22 5 30 Oct 2018 community repositories only · 31 ran (of which 12 constructed an object rather than computing a result; 24 with no instrument failure: 2 honoured, 1 violated, 21 with no contract checked; 7 where Syntology's instrument failed) · 12 unverified (16 pointer-only for licence)
Large-Scale Study of Curiosity-Driven Learning 5 5 13 Aug 2018 community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence)
Count-Based Exploration with the Successor Representation 2 6 31 Jul 2018 official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified
RUDDER: Return Decomposition for Delayed Rewards 2 3 20 Jun 2018 official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (9 pointer-only for licence)
Self-Imitation Learning 4 45 14 Jun 2018 community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (3 pointer-only for licence)
Evolving simple programs for playing Atari games 2 50 14 Jun 2018 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified
Implicit Quantile Networks for Distributional Reinforcement Learning 19 51 14 Jun 2018 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified
Distributed Prioritized Experience Replay 15 51 2 Mar 2018 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (3 pointer-only for licence)
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures 24 52 5 Feb 2018 community repositories only · 16 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 18 unverified (3 pointer-only for licence)
Distributional Reinforcement Learning with Quantile Regression 17 51 27 Oct 2017 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)
Rainbow: Combining Improvements in Deep Reinforcement Learning 34 4 6 Oct 2017 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence)
Noisy Networks for Exploration 15 48 30 Jun 2017 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (3 pointer-only for licence)
Evolution Strategies as a Scalable Alternative to Reinforcement Learning 23 41 10 Mar 2017 official (archive's flag): 4 ran · 16 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 13 unverified (2 pointer-only for licence)
Count-Based Exploration with Neural Density Models 1 9 3 Mar 2017 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified
Deep Exploration via Bootstrapped DQN 6 45 15 Feb 2016 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence)
Asynchronous Methods for Deep Reinforcement Learning 70 138 4 Feb 2016 60 ran (of which 20 constructed an object rather than computing a result; 51 with no instrument failure: 2 honoured, 1 violated, 48 with no contract checked; 9 where Syntology's instrument failed) · 35 unverified (20 pointer-only for licence)
Dueling Network Architectures for Deep Reinforcement Learning 73 185 20 Nov 2015 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 3 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (7 pointer-only for licence)
Prioritized Experience Replay 77 91 18 Nov 2015 80 ran (of which 62 constructed an object rather than computing a result; 72 with no instrument failure: 4 honoured, 0 violated, 68 with no contract checked; 8 where Syntology's instrument failed) · 31 unverified (43 pointer-only for licence)
Deep Reinforcement Learning with Double Q-learning 97 183 22 Sep 2015 56 ran (of which 38 constructed an object rather than computing a result; 55 with no instrument failure: 0 honoured, 0 violated, 55 with no contract checked; 1 where Syntology's instrument failed) · 50 unverified (57 pointer-only for licence)
Playing Atari with Deep Reinforcement Learning 112 6 19 Dec 2013 64 ran (of which 24 constructed an object rather than computing a result; 46 with no instrument failure: 5 honoured, 0 violated, 41 with no contract checked; 18 where Syntology's instrument failed) · 53 unverified (56 pointer-only for licence)
The Arcade Learning Environment: An Evaluation Platform for General Agents 3 98 19 Jul 2012 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence)