| IQ-Learn: Inverse soft-Q Learning for Imitation |
5 |
4 |
23 Jun 2021 |
community repositories only · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence) |
| Decision Transformer: Reinforcement Learning via Sequence Modeling |
20 |
4 |
2 Jun 2021 |
community repositories only · 17 ran (of which 10 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 0 violated, 13 with no contract checked; 4 where Syntology's instrument failed) · 9 unverified (8 pointer-only for licence) |
| Adaptive Rational Activations to Boost Deep Reinforcement Learning |
4 |
26 |
18 Feb 2021 |
official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence) |
| Reinforcement Learning with Latent Flow |
2 |
2 |
6 Jan 2021 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified |
| Mastering Atari with Discrete World Models |
9 |
50 |
5 Oct 2020 |
community repositories only · 11 ran (of which 1 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 1 violated, 9 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified |
| Munchausen Reinforcement Learning |
6 |
1 |
28 Jul 2020 |
community repositories only · 12 ran (of which 11 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 1 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 10 unverified |
| First return, then explore |
2 |
10 |
27 Apr 2020 |
official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence) |
| CURL: Contrastive Unsupervised Representations for Reinforcement Learning |
7 |
24 |
8 Apr 2020 |
official (archive's flag): 6 ran · 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 5 where Syntology's instrument failed) · 1 unverified (3 pointer-only for licence) |
| Agent57: Outperforming the Atari Human Benchmark |
5 |
51 |
30 Mar 2020 |
8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified |
| Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model |
18 |
51 |
19 Nov 2019 |
43 ran (of which 36 constructed an object rather than computing a result; 43 with no instrument failure: 4 honoured, 0 violated, 39 with no contract checked; 0 where Syntology's instrument failed) · 21 unverified (62 pointer-only for licence) |
| Fully Parameterized Quantile Function for Distributional Reinforcement Learning |
6 |
23 |
5 Nov 2019 |
6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified |
| Soft Actor-Critic for Discrete Action Settings |
13 |
18 |
16 Oct 2019 |
official: harvested, nothing ran · 12 ran (of which 0 constructed an object rather than computing a result; 12 with no instrument failure: 2 honoured, 1 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (3 pointer-only for licence) |
| Recurrent Independent Mechanisms |
3 |
3 |
24 Sep 2019 |
14 ran (of which 9 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (20 pointer-only for licence) |
| Go-Explore: a New Approach for Hard-Exploration Problems |
3 |
2 |
30 Jan 2019 |
2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (4 pointer-only for licence) |
| Exploration by Random Network Distillation |
22 |
5 |
30 Oct 2018 |
community repositories only · 31 ran (of which 12 constructed an object rather than computing a result; 24 with no instrument failure: 2 honoured, 1 violated, 21 with no contract checked; 7 where Syntology's instrument failed) · 12 unverified (16 pointer-only for licence) |
| Large-Scale Study of Curiosity-Driven Learning |
5 |
5 |
13 Aug 2018 |
community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (3 pointer-only for licence) |
| Count-Based Exploration with the Successor Representation |
2 |
6 |
31 Jul 2018 |
official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified |
| RUDDER: Return Decomposition for Delayed Rewards |
2 |
3 |
20 Jun 2018 |
official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 1 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (9 pointer-only for licence) |
| Self-Imitation Learning |
4 |
45 |
14 Jun 2018 |
community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (3 pointer-only for licence) |
| Evolving simple programs for playing Atari games |
2 |
50 |
14 Jun 2018 |
7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified |
| Implicit Quantile Networks for Distributional Reinforcement Learning |
19 |
51 |
14 Jun 2018 |
2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified |
| Distributed Prioritized Experience Replay |
15 |
51 |
2 Mar 2018 |
9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (3 pointer-only for licence) |
| IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures |
24 |
52 |
5 Feb 2018 |
community repositories only · 16 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 18 unverified (3 pointer-only for licence) |
| Distributional Reinforcement Learning with Quantile Regression |
17 |
51 |
27 Oct 2017 |
2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence) |
| Rainbow: Combining Improvements in Deep Reinforcement Learning |
34 |
4 |
6 Oct 2017 |
5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (1 pointer-only for licence) |
| Noisy Networks for Exploration |
15 |
48 |
30 Jun 2017 |
1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (3 pointer-only for licence) |
| Evolution Strategies as a Scalable Alternative to Reinforcement Learning |
23 |
41 |
10 Mar 2017 |
official (archive's flag): 4 ran · 16 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 1 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 13 unverified (2 pointer-only for licence) |
| Count-Based Exploration with Neural Density Models |
1 |
9 |
3 Mar 2017 |
1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified |
| Deep Exploration via Bootstrapped DQN |
6 |
45 |
15 Feb 2016 |
1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (2 pointer-only for licence) |
| Asynchronous Methods for Deep Reinforcement Learning |
70 |
138 |
4 Feb 2016 |
60 ran (of which 20 constructed an object rather than computing a result; 51 with no instrument failure: 2 honoured, 1 violated, 48 with no contract checked; 9 where Syntology's instrument failed) · 35 unverified (20 pointer-only for licence) |
| Dueling Network Architectures for Deep Reinforcement Learning |
73 |
185 |
20 Nov 2015 |
7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 3 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (7 pointer-only for licence) |
| Prioritized Experience Replay |
77 |
91 |
18 Nov 2015 |
80 ran (of which 62 constructed an object rather than computing a result; 72 with no instrument failure: 4 honoured, 0 violated, 68 with no contract checked; 8 where Syntology's instrument failed) · 31 unverified (43 pointer-only for licence) |
| Deep Reinforcement Learning with Double Q-learning |
97 |
183 |
22 Sep 2015 |
56 ran (of which 38 constructed an object rather than computing a result; 55 with no instrument failure: 0 honoured, 0 violated, 55 with no contract checked; 1 where Syntology's instrument failed) · 50 unverified (57 pointer-only for licence) |
| Playing Atari with Deep Reinforcement Learning |
112 |
6 |
19 Dec 2013 |
64 ran (of which 24 constructed an object rather than computing a result; 46 with no instrument failure: 5 honoured, 0 violated, 41 with no contract checked; 18 where Syntology's instrument failed) · 53 unverified (56 pointer-only for licence) |
| The Arcade Learning Environment: An Evaluation Platform for General Agents |
3 |
98 |
19 Jul 2012 |
2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (2 pointer-only for licence) |