Methods › Reinforcement Learning › Replay Memory › Prioritized Experience Replay › Papers, page 2
Prioritized Experience Replay
Papers archive 2025-07-28
archive papers tagged: 138 · with a code link: 61 · where Syntology ran a sample: 19 (17 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (19 of 138 tagged: 17 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 138 of 138, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly 15 Oct 2020 · 0 repositories · arXiv:2010.08052
-
A New Approach for Tactical Decision Making in Lane Changing: Sample Efficient Deep Q Learning with a Safety Feedback Reward 24 Sep 2020 · 0 repositories · arXiv:2009.11905
-
Understanding and Mitigating the Limitations of Prioritized Experience Replay 19 Jul 2020 · 2 repositories · arXiv:2007.09569
-
Learning to Sample with Local and Global Contexts in Experience Replay Buffer 14 Jul 2020 · 0 repositories · arXiv:2007.07358
-
SUNRISE: A Simple Unified Framework for Ensemble Learning in Deep Reinforcement Learning 9 Jul 2020 · 1 repository · arXiv:2007.04938
-
Double Prioritized State Recycled Experience Replay 8 Jul 2020 · 0 repositories · arXiv:2007.03961
-
The LoCA Regret: A Consistent Metric to Evaluate Model-Based Behavior in Reinforcement Learning 7 Jul 2020 · 2 repositories · arXiv:2007.03158Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Distributed Uplink Beamforming in Cell-Free Networks Using Deep Reinforcement Learning 26 Jun 2020 · 0 repositories · arXiv:2006.15138
-
Continuous Control for Searching and Planning with a Learned Model 12 Jun 2020 · 0 repositories · arXiv:2006.07430
-
Balancing a CartPole System with Reinforcement Learning -- A Tutorial 8 Jun 2020 · 0 repositories · arXiv:2006.04938
-
Manipulating the Distributions of Experience used for Self-Play Learning in Expert Iteration 30 May 2020 · 1 repository · arXiv:2006.00283
-
STDPG: A Spatio-Temporal Deterministic Policy Gradient Agent for Dynamic Routing in SDN 21 Apr 2020 · 0 repositories · arXiv:2004.09783
-
Dynamic Experience Replay 4 Mar 2020 · 0 repositories · arXiv:2003.02372
-
Deep Reinforcement Learning Based Intelligent Reflecting Surface for Secure Wireless Communications 27 Feb 2020 · 0 repositories · arXiv:2002.12271
-
Fast Reinforcement Learning for Anti-jamming Communications 13 Feb 2020 · 0 repositories · arXiv:2002.05364
-
Stacked Auto Encoder Based Deep Reinforcement Learning for Online Resource Scheduling in Large-Scale MEC Networks 24 Jan 2020 · 0 repositories · arXiv:2001.09223
-
Sample-based Distributional Policy Gradient 8 Jan 2020 · 0 repositories · arXiv:2001.02652
-
Which Channel to Ask My Question? Personalized Customer Service RequestStream Routing using DeepReinforcement Learning 24 Nov 2019 · 0 repositories · arXiv:1911.10521
-
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model 19 Nov 2019 · 18 repositories · arXiv:1911.08265Syntology 43 ran (of which 36 constructed an object rather than computing a result; 43 with no instrument failure: 4 honoured, 0 violated, 39 with no contract checked; 0 where Syntology's instrument failed) · 21 unverified (of 64 harvested samples) · 62 pointer-only (licence)
-
Placement Optimization of Aerial Base Stations with Deep Reinforcement Learning 19 Nov 2019 · 0 repositories · arXiv:1911.08111
-
Deep Reinforcement Learning Based Dynamic Trajectory Control for UAV-assisted Mobile Edge Computing 10 Nov 2019 · 0 repositories · arXiv:1911.03887
-
Task-Oriented Language Grounding for Language Input with Multiple Sub-Goals of Non-Linear Order 27 Oct 2019 · 1 repository · arXiv:1910.12354
-
Google Research Football: A Novel Reinforcement Learning Environment 25 Jul 2019 · 1 repository · arXiv:1907.11180Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Prioritized Guidance for Efficient Multi-Agent Reinforcement Learning Exploration 18 Jul 2019 · 0 repositories · arXiv:1907.07847
-
Prioritized Sequence Experience Replay 25 May 2019 · 0 repositories · arXiv:1905.12726
-
Generative Adversarial Imagination for Sample Efficient Deep Reinforcement Learning 30 Apr 2019 · 0 repositories · arXiv:1904.13255
-
TF-Replicator: Distributed Machine Learning for Researchers 1 Feb 2019 · 1 repository · arXiv:1902.00465Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Macro action selection with deep reinforcement learning in StarCraft 2 Dec 2018 · 1 repository · arXiv:1812.00336
-
An Intriguing Failing of Convolutional Neural Networks and the CoordConv Solution 9 Jul 2018 · 24 repositories · arXiv:1807.03247Syntology community repositories only · 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Deep Curiosity Search: Intra-Life Exploration Can Improve Performance on Challenging Deep Reinforcement Learning Problems 1 Jun 2018 · 0 repositories · arXiv:1806.00553
-
Advances in Experience Replay 15 May 2018 · 1 repository · arXiv:1805.05536
-
Distributed Distributional Deterministic Policy Gradients 23 Apr 2018 · 5 repositories · arXiv:1804.08617
-
Distributed Prioritized Experience Replay 2 Mar 2018 · 15 repositories · arXiv:1803.00933Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples) · 3 pointer-only (licence)
-
ScreenerNet: Learning Self-Paced Curriculum for Deep Neural Networks 3 Jan 2018 · 0 repositories · arXiv:1801.00904
-
ViZDoom: DRQN with Prioritized Experience Replay, Double-Q Learning, & Snapshot Ensembling 3 Jan 2018 · 0 repositories · arXiv:1801.01000
-
Rainbow: Combining Improvements in Deep Reinforcement Learning 6 Oct 2017 · 34 repositories · arXiv:1710.02298Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
A novel DDPG method with prioritized experience replay 1 Oct 2017 · 1 repository
-
Prioritized Experience Replay 18 Nov 2015 · 77 repositories · arXiv:1511.05952Syntology 80 ran (of which 62 constructed an object rather than computing a result; 72 with no instrument failure: 4 honoured, 0 violated, 68 with no contract checked; 8 where Syntology's instrument failed) · 31 unverified (of 111 harvested samples) · 43 pointer-only (licence)