Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers, page 7
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 7 of 9: papers 601 to 700 of 865, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Student-Initiated Action Advising via Advice Novelty 1 Oct 2020 · 1 repository · arXiv:2010.00381
-
Lucid Dreaming for Experience Replay: Refreshing Past States with the Current Policy 29 Sep 2020 · 1 repository · arXiv:2009.13736
-
Knowledge-Assisted Deep Reinforcement Learning in 5G Scheduler Design: From Theoretical Framework to Implementation 17 Sep 2020 · 0 repositories · arXiv:2009.08346
-
Meta-Learning with Sparse Experience Replay for Lifelong Language Learning 10 Sep 2020 · 1 repository · arXiv:2009.04891
-
Sample-Efficient Automated Deep Reinforcement Learning 3 Sep 2020 · 1 repository · arXiv:2009.01555Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Online Class-Incremental Continual Learning with Adversarial Shapley Value 31 Aug 2020 · 3 repositories · arXiv:2009.00093
-
Sample Efficiency in Sparse Reinforcement Learning: Or Your Money Back 28 Aug 2020 · 0 repositories · arXiv:2008.12693
-
Curriculum Learning with Hindsight Experience Replay for Sequential Object Manipulation Tasks 21 Aug 2020 · 0 repositories · arXiv:2008.09377
-
Chrome Dino Run using Reinforcement Learning 15 Aug 2020 · 0 repositories · arXiv:2008.06799
-
Reinforcement Learning with Quantum Variational Circuits 15 Aug 2020 · 2 repositories · arXiv:2008.07524
-
Hierarchical Reinforcement Learning in StarCraft II with Human Expertise in Subgoals Selection 8 Aug 2020 · 0 repositories · arXiv:2008.03444
-
Deep Q-Network Based Multi-agent Reinforcement Learning with Binary Action Agents 6 Aug 2020 · 0 repositories · arXiv:2008.04109
-
Optimization-driven Hierarchical Learning Framework for Wireless Powered Backscatter-aided Relay Communications 4 Aug 2020 · 0 repositories · arXiv:2008.01366
-
Complex Robotic Manipulation via Graph-Based Hindsight Goal Generation 27 Jul 2020 · 1 repository · arXiv:2007.13486
-
Learning Compositional Neural Programs for Continuous Control 27 Jul 2020 · 0 repositories · arXiv:2007.13363
-
EMaQ: Expected-Max Q-Learning Operator for Simple Yet Effective Offline and Online RL 21 Jul 2020 · 0 repositories · arXiv:2007.11091
-
Understanding and Mitigating the Limitations of Prioritized Experience Replay 19 Jul 2020 · 2 repositories · arXiv:2007.09569
-
Collision Avoidance Robotics Via Meta-Learning (CARML) 16 Jul 2020 · 1 repository · arXiv:2007.08616
-
Human-like Energy Management Based on Deep Reinforcement Learning and Historical Driving Experiences 16 Jul 2020 · 0 repositories · arXiv:2007.10126
-
Learning to Sample with Local and Global Contexts in Experience Replay Buffer 14 Jul 2020 · 0 repositories · arXiv:2007.07358
-
Revisiting Fundamentals of Experience Replay 13 Jul 2020 · 2 repositories · arXiv:2007.06700Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
An Equivalence between Loss Functions and Non-Uniform Sampling in Experience Replay 12 Jul 2020 · 2 repositories · arXiv:2007.06049Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
Batch-level Experience Replay with Review for Continual Learning 11 Jul 2020 · 1 repository · arXiv:2007.05683
-
Double Prioritized State Recycled Experience Replay 8 Jul 2020 · 0 repositories · arXiv:2007.03961
-
Self-Supervised Policy Adaptation during Deployment 8 Jul 2020 · 2 repositories · arXiv:2007.04309Syntology 16 ran (of which 9 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Counterfactual Data Augmentation using Locally Factored Dynamics 6 Jul 2020 · 1 repository · arXiv:2007.02863Syntology 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Developing cooperative policies for multi-stage tasks 1 Jul 2020 · 0 repositories · arXiv:2007.00203
-
Regularly Updated Deterministic Policy Gradient Algorithm 1 Jul 2020 · 0 repositories · arXiv:2007.00169
-
UAV Path Planning for Wireless Data Harvesting: A Deep Reinforcement Learning Approach 1 Jul 2020 · 3 repositories · arXiv:2007.00544
-
Distributed Uplink Beamforming in Cell-Free Networks Using Deep Reinforcement Learning 26 Jun 2020 · 0 repositories · arXiv:2006.15138
-
Some approaches used to overcome overestimation in Deep Reinforcement Learning algorithms 25 Jun 2020 · 0 repositories · arXiv:2006.14167
-
Experience Replay with Likelihood-free Importance Weights 23 Jun 2020 · 1 repository · arXiv:2006.13169Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
The Effect of Multi-step Methods on Overestimation in Deep Reinforcement Learning 23 Jun 2020 · 0 repositories · arXiv:2006.12692
-
AutoOD: Automated Outlier Detection via Curiosity-guided Search and Self-imitation Learning 19 Jun 2020 · 0 repositories · arXiv:2006.11321
-
WD3: Taming the Estimation Bias in Deep Reinforcement Learning 18 Jun 2020 · 0 repositories · arXiv:2006.12622
-
Forgetful Experience Replay in Hierarchical Reinforcement Learning from Demonstrations 17 Jun 2020 · 1 repository · arXiv:2006.09939
-
Reinforcement Learning with Uncertainty Estimation for Tactical Decision-Making in Intersections 17 Jun 2020 · 0 repositories · arXiv:2006.09786
-
Least Squares Regression with Markovian Data: Fundamental Limits and Algorithms 16 Jun 2020 · 0 repositories · arXiv:2006.08916
-
Reinforcement Learning Control of Robotic Knee with Human in the Loop by Flexible Policy Iteration 16 Jun 2020 · 0 repositories · arXiv:2006.09008
-
An online evolving framework for advancing reinforcement-learning based automated vehicle control 15 Jun 2020 · 0 repositories · arXiv:2006.08092
-
Learning Heuristic Selection with Dynamic Algorithm Configuration 15 Jun 2020 · 1 repository · arXiv:2006.08246
-
Human and Multi-Agent collaboration in a human-MARL teaming framework 12 Jun 2020 · 0 repositories · arXiv:2006.07301
-
Balancing a CartPole System with Reinforcement Learning -- A Tutorial 8 Jun 2020 · 0 repositories · arXiv:2006.04938
-
Dynamic Algorithm Configuration: Foundation of a New Meta-Algorithmic Framework 1 Jun 2020 · 1 repository
-
PlanGAN: Model-based Planning With Sparse Rewards and Multiple Goals 1 Jun 2020 · 1 repository · arXiv:2006.00900
-
Manipulating the Distributions of Experience used for Self-Play Learning in Expert Iteration 30 May 2020 · 1 repository · arXiv:2006.00283
-
Optimization-driven Deep Reinforcement Learning for Robust Beamforming in IRS-assisted Wireless Communications 25 May 2020 · 0 repositories · arXiv:2005.11885
-
Experience Augmentation: Boosting and Accelerating Off-Policy Multi-Agent Reinforcement Learning 19 May 2020 · 0 repositories · arXiv:2005.09453
-
Automating Turbulence Modeling by Multi-Agent Reinforcement Learning 18 May 2020 · 0 repositories · arXiv:2005.09023
-
Proxy Experience Replay: Federated Distillation for Distributed Reinforcement Learning 13 May 2020 · 0 repositories · arXiv:2005.06105
-
Smooth Exploration for Robotic Reinforcement Learning 12 May 2020 · 4 repositories · arXiv:2005.05719Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
TOMA: Topological Map Abstraction for Reinforcement Learning 11 May 2020 · 0 repositories · arXiv:2005.06061
-
An FPGA-Based On-Device Reinforcement Learning Approach using Online Sequential Learning 10 May 2020 · 0 repositories · arXiv:2005.04646
-
Discrete-to-Deep Supervised Policy Learning 5 May 2020 · 1 repository · arXiv:2005.02057
-
Delay-aware Resource Allocation in Fog-assisted IoT Networks Through Reinforcement Learning 30 Apr 2020 · 0 repositories · arXiv:2005.04097
-
DSAC: Distributional Soft Actor Critic for Risk-Sensitive Reinforcement Learning 30 Apr 2020 · 0 repositories · arXiv:2004.14547
-
Reinforcement Learning with Augmented Data 30 Apr 2020 · 2 repositories · arXiv:2004.14990Syntology 21 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 12 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 21 pointer-only (licence)
-
PBCS : Efficient Exploration and Exploitation Using a Synergy between Reinforcement Learning and Motion Planning 24 Apr 2020 · 0 repositories · arXiv:2004.11667
-
Mean-Variance Policy Iteration for Risk-Averse Reinforcement Learning 22 Apr 2020 · 1 repository · arXiv:2004.10888
-
STDPG: A Spatio-Temporal Deterministic Policy Gradient Agent for Dynamic Routing in SDN 21 Apr 2020 · 0 repositories · arXiv:2004.09783
-
Dark Experience for General Continual Learning: a Strong, Simple Baseline 15 Apr 2020 · 3 repositories · arXiv:2004.07211
-
Model-based actor-critic: GAN (model generator) + DRL (actor-critic) => AGI 4 Apr 2020 · 0 repositories · arXiv:2004.04574
-
Obstacle Avoidance and Navigation Utilizing Reinforcement Learning with Reward Shaping 28 Mar 2020 · 1 repository · arXiv:2003.12863
-
Multi-Agent Reinforcement Learning for Problems with Combined Individual and Team Reward 24 Mar 2020 · 0 repositories · arXiv:2003.10598
-
Evolutionary Population Curriculum for Scaling Multi-Agent Reinforcement Learning 23 Mar 2020 · 1 repository · arXiv:2003.10423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Overcoming Catastrophic Forgetting in Graph Neural Networks with Experience Replay 22 Mar 2020 · 0 repositories · arXiv:2003.09908
-
Accelerating Deep Reinforcement Learning With the Aid of Partial Model: Energy-Efficient Predictive Video Streaming 21 Mar 2020 · 0 repositories · arXiv:2003.09708
-
Adversarial Continual Learning 21 Mar 2020 · 1 repository · arXiv:2003.09553Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 5 pointer-only (licence)
-
Online Continual Learning on Sequences 20 Mar 2020 · 0 repositories · arXiv:2003.09114
-
Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations 19 Mar 2020 · 4 repositories · arXiv:2003.08938
-
PFPN: Continuous Control of Physically Simulated Characters using Particle Filtering Policy Network 16 Mar 2020 · 1 repository · arXiv:2003.06959
-
FACMAC: Factored Multi-Agent Centralised Policy Gradients 14 Mar 2020 · 3 repositories · arXiv:2003.06709Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Sample Efficient Reinforcement Learning through Learning from Demonstrations in Minecraft 12 Mar 2020 · 3 repositories · arXiv:2003.06066Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Online Meta-Critic Learning for Off-Policy Actor-Critic Methods 11 Mar 2020 · 1 repository · arXiv:2003.05334Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Dynamic Experience Replay 4 Mar 2020 · 0 repositories · arXiv:2003.02372
-
Contention Window Optimization in IEEE 802.11ax Networks with Deep Reinforcement Learning 3 Mar 2020 · 1 repository · arXiv:2003.01492
-
Reinforcement co-Learning of Deep and Spiking Neural Networks for Energy-Efficient Mapless Navigation with Neuromorphic Hardware 2 Mar 2020 · 1 repository · arXiv:2003.01157Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
A Self-Tuning Actor-Critic Algorithm 28 Feb 2020 · 0 repositories · arXiv:2002.12928
-
Deep Reinforcement Learning Based Intelligent Reflecting Surface for Secure Wireless Communications 27 Feb 2020 · 0 repositories · arXiv:2002.12271
-
Data Freshness and Energy-Efficient UAV Navigation Optimization: A Deep Reinforcement Learning Approach 21 Feb 2020 · 0 repositories · arXiv:2003.04816
-
Disentangling Controllable Object through Video Prediction Improves Visual Reinforcement Learning 21 Feb 2020 · 0 repositories · arXiv:2002.09136
-
Using Hindsight to Anchor Past Knowledge in Continual Learning 19 Feb 2020 · 1 repository · arXiv:2002.08165
-
Adaptive Experience Selection for Policy Gradient 17 Feb 2020 · 0 repositories · arXiv:2002.06946
-
Fast Reinforcement Learning for Anti-jamming Communications 13 Feb 2020 · 0 repositories · arXiv:2002.05364
-
XCS Classifier System with Experience Replay 13 Feb 2020 · 0 repositories · arXiv:2002.05628
-
Soft Hindsight Experience Replay 6 Feb 2020 · 2 repositories · arXiv:2002.02089
-
Bootstrapping a DQN Replay Memory with Synthetic Experiences 4 Feb 2020 · 0 repositories · arXiv:2002.01370
-
Stacked Auto Encoder Based Deep Reinforcement Learning for Online Resource Scheduling in Large-Scale MEC Networks 24 Jan 2020 · 0 repositories · arXiv:2001.09223
-
Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning 23 Jan 2020 · 4 repositories · arXiv:2001.08726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Cooperative Highway Work Zone Merge Control based on Reinforcement Learning in A Connected and Automated Environment 21 Jan 2020 · 0 repositories · arXiv:2001.08581
-
Discriminator Soft Actor Critic without Extrinsic Rewards 19 Jan 2020 · 1 repository · arXiv:2001.06808
-
Effects of sparse rewards of different magnitudes in the speed of learning of model-based actor critic methods 18 Jan 2020 · 0 repositories · arXiv:2001.06725
-
Continuous-action Reinforcement Learning for Playing Racing Games: Comparing SPG to PPO 15 Jan 2020 · 1 repository · arXiv:2001.05270
-
Deep Reinforcement Learning for Complex Manipulation Tasks with Sparse Feedback 12 Jan 2020 · 0 repositories · arXiv:2001.03877
-
Reward Engineering for Object Pick and Place Training 11 Jan 2020 · 1 repository · arXiv:2001.03792
-
Population-Guided Parallel Policy Search for Reinforcement Learning 9 Jan 2020 · 1 repository · arXiv:2001.02907
-
SLM Lab: A Comprehensive Benchmark and Modular Software Framework for Reproducible Deep Reinforcement Learning 28 Dec 2019 · 1 repository · arXiv:1912.12482Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Discrete and Continuous Action Representation for Practical RL in Video Games 23 Dec 2019 · 1 repository · arXiv:1912.11077Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Exploiting the potential of deep reinforcement learning for classification tasks in high-dimensional and unstructured data 20 Dec 2019 · 0 repositories · arXiv:1912.09595
-
Recruitment-imitation Mechanism for Evolutionary Reinforcement Learning 13 Dec 2019 · 0 repositories · arXiv:1912.06310