| Atari Games |
Atari 2600 Freeway |
TRPO-hash Score 34.0 |
#Exploration: A Study of Count-Based Exploration for... |
uoe-agents/derl +2 |
59 |
Compare |
| Atari Games |
Atari 2600 Breakout |
GDI-H3(200M frames) Score 864.00 |
Generalized Data Distribution Iteration |
— |
58 |
Compare |
| Atari Games |
Atari 2600 Q*Bert |
Agent57 Score 580328.14 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
57 |
Compare |
| Atari Games |
Atari 2600 Seaquest |
GDI-H3(200M frames) Score 1000000 |
Generalized Data Distribution Iteration |
— |
57 |
Compare |
| Atari Games |
Atari 2600 Space Invaders |
GDI-H3(200M frames) Score 154380 |
Generalized Data Distribution Iteration |
— |
55 |
Compare |
| Atari Games |
Atari 2600 Venture |
Agent57 Score 2623.71 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
55 |
Compare |
| Atari Games |
Atari 2600 Frostbite |
MuZero Score 631378.53 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
53 |
Compare |
| Atari Games |
Atari 2600 Gravitar |
Agent57 Score 19213.96 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
53 |
Compare |
| Atari Games |
Atari 2600 Pong |
Duel noop Score 21.0 |
Dueling Network Architectures for Deep Reinforcement Learning |
labmlai/annotated_deep_learning_paper_implementations +72 |
52 |
Compare |
| Atari Games |
Atari 2600 Private Eye |
Go-Explore Score 95756 |
First return, then explore |
uber-research/go-explore +1 |
52 |
Compare |
| Atari Games |
Atari 2600 Montezuma's Revenge |
Go-Explore Score 43791 |
First return, then explore |
uber-research/go-explore +1 |
50 |
Compare |
| Atari Games |
Atari 2600 Alien |
MuZero Score 741812.63 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
49 |
Compare |
| Atari Games |
Atari 2600 Beam Rider |
MuZero Score 454993.53 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
49 |
Compare |
| Atari Games |
Atari 2600 Crazy Climber |
Agent57 Score 565909.85 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
49 |
Compare |
| Atari Games |
Atari 2600 Amidar |
Agent57 Score 29660.08 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
48 |
Compare |
| Atari Games |
Atari 2600 Battle Zone |
Agent57 Score 934134.88 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
47 |
Compare |
| Atari Games |
Atari 2600 Kangaroo |
Agent57 Score 24034.16 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
47 |
Compare |
| Atari Games |
Atari 2600 Ms. Pacman |
MuZero Score 243401.10 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
47 |
Compare |
| Atari Games |
Atari 2600 Demon Attack |
GDI-H3 Score 787985 |
Generalized Data Distribution Iteration |
— |
46 |
Compare |
| Atari Games |
Atari 2600 Assault |
MuZero Score 143972.03 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
45 |
Compare |
| Atari Games |
Atari 2600 Bank Heist |
MuZero (Res2 Adam) Score 27219.8 |
Online and Offline Reinforcement Learning by Planning... |
DHDev0/Muzero-unplugged +1 |
45 |
Compare |
| Atari Games |
Atari 2600 Centipede |
Go-Explore Score 1422628 |
First return, then explore |
uber-research/go-explore +1 |
45 |
Compare |
| Atari Games |
Atari 2600 Chopper Command |
GDI-H3 Score 999999 |
GDI: Rethinking What Makes Reinforcement Learning... |
— |
45 |
Compare |
| Atari Games |
Atari 2600 James Bond |
GDI-H3 Score 620780 |
Generalized Data Distribution Iteration |
— |
45 |
Compare |
| Atari Games |
Atari 2600 Krull |
GDI-H3 Score 594540 |
Generalized Data Distribution Iteration |
— |
45 |
Compare |
| Atari Games |
Atari 2600 Bowling |
MuZero Score 260.13 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
44 |
Compare |
| Atari Games |
Atari 2600 Fishing Derby |
MuZero Score 91.16 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
44 |
Compare |
| Atari Games |
Atari 2600 HERO |
Agent57 Score 114736.26 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
44 |
Compare |
| Atari Games |
Atari 2600 Road Runner |
GDI-H3 Score 999999 |
Generalized Data Distribution Iteration |
— |
44 |
Compare |
| Atari Games |
Atari 2600 Time Pilot |
MuZero Score 476763.90 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
44 |
Compare |
| Atari Games |
Atari 2600 Tutankham |
Agent57 Score 2354.91 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
44 |
Compare |
| Atari Games |
Atari 2600 Up and Down |
GDI-I3 Score 986440 |
GDI: Rethinking What Makes Reinforcement Learning... |
— |
44 |
Compare |
| Atari Games |
Atari 2600 Asteroids |
GDI-H3 Score 760005 |
Generalized Data Distribution Iteration |
— |
43 |
Compare |
| Atari Games |
Atari 2600 Double Dunk |
UCT Score 24 |
The Arcade Learning Environment: An Evaluation Platform... |
mgbellemare/Arcade-Learning-Environment +2 |
43 |
Compare |
| Atari Games |
Atari 2600 Gopher |
GDI-I3 Score 488830 |
Generalized Data Distribution Iteration |
— |
43 |
Compare |
| Atari Games |
Atari 2600 Ice Hockey |
GDI-H3 Score 481.9 |
Generalized Data Distribution Iteration |
— |
43 |
Compare |
| Atari Games |
Atari 2600 Kung-Fu Master |
GDI-H3 Score 1666665 |
Generalized Data Distribution Iteration |
— |
43 |
Compare |
| Atari Games |
Atari 2600 Name This Game |
MuZero Score 157177.85 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
43 |
Compare |
| Atari Games |
Atari 2600 Tennis |
GDI-I3 Score 24 |
GDI: Rethinking What Makes Reinforcement Learning... |
— |
43 |
Compare |
| Atari Games |
Atari 2600 Atlantis |
GDI-H3 Score 3837300 |
Generalized Data Distribution Iteration |
— |
42 |
Compare |
| Atari Games |
Atari 2600 Robotank |
MuZero Score 131.13 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
42 |
Compare |
| Atari Games |
Atari 2600 Video Pinball |
R2D2 Score 999383.2 |
Recurrent Experience Replay in Distributed Reinforcement Learning |
opendilab/DI-engine +2 |
42 |
Compare |
| Atari Games |
Atari 2600 River Raid |
MuZero Score 323417.18 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
41 |
Compare |
| Atari Games |
Atari 2600 Wizard of Wor |
MuZero Score 197126.00 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
41 |
Compare |
| Atari Games |
Atari 2600 Zaxxon |
MuZero Score 725853.90 |
Mastering Atari, Go, Chess and Shogi by Planning with a... |
werner-duvaud/muzero-general +17 |
41 |
Compare |
| Atari Games |
Atari 2600 Berzerk |
Go-Explore Score 197376 |
First return, then explore |
uber-research/go-explore +1 |
39 |
Compare |
| Atari Games |
Atari 2600 Pitfall! |
Go-Explore Score 102571 |
Go-Explore: a New Approach for Hard-Exploration Problems |
uber-research/go-explore +2 |
23 |
Compare |
| Atari Games |
Atari 2600 Skiing |
Best Learner Score 0 |
The Arcade Learning Environment: An Evaluation Platform... |
mgbellemare/Arcade-Learning-Environment +2 |
23 |
Compare |
| Atari Games |
Atari 2600 Phoenix |
GDI-H3 Score 959580 |
Generalized Data Distribution Iteration |
— |
21 |
Compare |
| Atari Games |
Atari 2600 Yars Revenge |
Agent57 Score 998532.37 |
Agent57: Outperforming the Atari Human Benchmark |
michaelnny/deep_rl_zoo +4 |
17 |
Compare |
| Atari Games |
Atari 2600 Surround |
NoisyNet-Dueling Score 10 |
Noisy Networks for Exploration |
opendilab/DI-engine +14 |
15 |
Compare |
| Atari Games |
Atari-57 |
LBC Mean Human Normalized Score 10077.52% |
Learnable Behavior Control: Breaking Atari Human World... |
— |
11 |
Compare |
| Atari Games |
Atari 2600 Elevator Action |
Persistent AL Score 29100 |
Increasing the Action Gap: New Operators for... |
janhuenermann/neurojs +1 |
3 |
Compare |
| Atari Games |
Atari 2600 Pooyan |
UCT Score 17763.4 |
The Arcade Learning Environment: An Evaluation Platform... |
mgbellemare/Arcade-Learning-Environment +2 |
3 |
Compare |
| Montezuma's Revenge |
Atari 2600 Montezuma's Revenge |
Flare Average Return (NoOp) 1668 |
Reinforcement Learning with Latent Flow |
WendyShang/flare +1 |
3 |
Compare |
| Atari Games |
Atari 2600 Carnival |
UCT Score 5132.0 |
The Arcade Learning Environment: An Evaluation Platform... |
mgbellemare/Arcade-Learning-Environment +2 |
1 |
Compare |
| Atari Games |
Atari 2600 Journey Escape |
UCT Score 7683.3 |
The Arcade Learning Environment: An Evaluation Platform... |
mgbellemare/Arcade-Learning-Environment +2 |
1 |
Compare |