Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers, page 8
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 8 of 9: papers 701 to 800 of 865, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning Sparse Representations Incrementally in Deep Reinforcement Learning 9 Dec 2019 · 0 repositories · arXiv:1912.04002
-
Policy Optimization Reinforcement Learning with Entropy Regularization 2 Dec 2019 · 0 repositories · arXiv:1912.01557
-
Better Exploration with Optimistic Actor Critic 1 Dec 2019 · 1 repository
-
Curriculum-guided Hindsight Experience Replay 1 Dec 2019 · 1 repository
-
Learning Reward Machines for Partially Observable Reinforcement Learning 1 Dec 2019 · 1 repository
-
Online Continual Learning with Maximal Interfered Retrieval 1 Dec 2019 · 2 repositories
-
Reconciling λ-Returns with Experience Replay 1 Dec 2019 · 1 repository
-
IMPACT: Importance Weighted Asynchronous Architectures with Clipped Target Networks 30 Nov 2019 · 0 repositories · arXiv:1912.00167
-
Distributed Soft Actor-Critic with Multivariate Reward Representation and Knowledge Distillation 29 Nov 2019 · 1 repository · arXiv:1911.13056
-
Which Channel to Ask My Question? Personalized Customer Service RequestStream Routing using DeepReinforcement Learning 24 Nov 2019 · 0 repositories · arXiv:1911.10521
-
Placement Optimization of Aerial Base Stations with Deep Reinforcement Learning 19 Nov 2019 · 0 repositories · arXiv:1911.08111
-
Data-efficient Co-Adaptation of Morphology and Behaviour with Deep Reinforcement Learning 15 Nov 2019 · 0 repositories · arXiv:1911.06832
-
Improved Exploration through Latent Trajectory Optimization in Deep Deterministic Policy Gradient 15 Nov 2019 · 0 repositories · arXiv:1911.06833
-
Deep Reinforcement Learning Based Dynamic Trajectory Control for UAV-assisted Mobile Edge Computing 10 Nov 2019 · 0 repositories · arXiv:1911.03887
-
Policy Continuation with Hindsight Inverse Dynamics 30 Oct 2019 · 1 repository · arXiv:1910.14055
-
Overcoming Catastrophic Interference in Online Reinforcement Learning with Dynamic Self-Organizing Maps 29 Oct 2019 · 0 repositories · arXiv:1910.13213
-
Better Exploration with Optimistic Actor-Critic 28 Oct 2019 · 0 repositories · arXiv:1910.12807
-
Task-Oriented Language Grounding for Language Input with Multiple Sub-Goals of Non-Linear Order 27 Oct 2019 · 1 repository · arXiv:1910.12354
-
HIGhER : Improving instruction following with Hindsight Generation for Experience Replay 21 Oct 2019 · 0 repositories · arXiv:1910.09451
-
Reverse Experience Replay 19 Oct 2019 · 0 repositories · arXiv:1910.08780
-
Towards More Sample Efficiency in Reinforcement Learning with Data Augmentation 19 Oct 2019 · 1 repository · arXiv:1910.09959
-
Ctrl-Z: Recovering from Instability in Reinforcement Learning 9 Oct 2019 · 0 repositories · arXiv:1910.03732
-
TorchBeast: A PyTorch Platform for Distributed RL 8 Oct 2019 · 3 repositories · arXiv:1910.03552Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
Striving for Simplicity and Performance in Off-Policy DRL: Output Normalization and Non-Uniform Sampling 5 Oct 2019 · 3 repositories · arXiv:1910.02208
-
QuaRL: Quantization for Fast and Environmentally Sustainable Reinforcement Learning 2 Oct 2019 · 1 repository · arXiv:1910.01055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Composite Q-learning: Multi-scale Q-function Decomposition and Separable Optimization 30 Sep 2019 · 0 repositories · arXiv:1909.13518
-
Off-Policy Actor-Critic with Shared Experience Replay 25 Sep 2019 · 0 repositories · arXiv:1909.11583
-
Invariant Transform Experience Replay: Data Augmentation for Deep Reinforcement Learning 24 Sep 2019 · 1 repository · arXiv:1909.10707
-
Constrained Attractor Selection Using Deep Reinforcement Learning 23 Sep 2019 · 0 repositories · arXiv:1909.10500
-
AC-Teach: A Bayesian Actor-Critic Method for Policy Learning with an Ensemble of Suboptimal Teachers 9 Sep 2019 · 1 repository · arXiv:1909.04121Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
Deterministic Value-Policy Gradients 9 Sep 2019 · 0 repositories · arXiv:1909.03939
-
Deep Reinforcement Learning for Control of Probabilistic Boolean Networks 7 Sep 2019 · 2 repositories · arXiv:1909.03331
-
Efficient Automatic Meta Optimization Search for Few-Shot Learning 6 Sep 2019 · 0 repositories · arXiv:1909.03817
-
Hierarchical Control for Bipedal Locomotion using Central Pattern Generators and Neural Networks 2 Sep 2019 · 1 repository · arXiv:1909.00732
-
Iterative Update and Unified Representation for Multi-Agent Reinforcement Learning 16 Aug 2019 · 0 repositories · arXiv:1908.06758
-
Online Continual Learning with Maximally Interfered Retrieval 11 Aug 2019 · 1 repository · arXiv:1908.04742Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Incremental Reinforcement Learning --- a New Continuous Reinforcement Learning Frame Based on Stochastic Differential Equation methods 8 Aug 2019 · 0 repositories · arXiv:1908.02974
-
Attention Control with Metric Learning Alignment for Image Set-based Recognition 5 Aug 2019 · 0 repositories · arXiv:1908.01872
-
Prioritized Guidance for Efficient Multi-Agent Reinforcement Learning Exploration 18 Jul 2019 · 0 repositories · arXiv:1907.07847
-
Improved Reinforcement Learning through Imitation Learning Pretraining Towards Image-based Autonomous Driving 16 Jul 2019 · 0 repositories · arXiv:1907.06838
-
Shapley Q-value: A Local Reward Approach to Solve Global Reward Games 11 Jul 2019 · 2 repositories · arXiv:1907.05707Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Dependency-aware Attention Control for Unconstrained Face Recognition with Image Sets 5 Jul 2019 · 0 repositories · arXiv:1907.03030
-
Modified Actor-Critics 2 Jul 2019 · 0 repositories · arXiv:1907.01298
-
Variational Quantum Circuits for Deep Reinforcement Learning 30 Jun 2019 · 1 repository · arXiv:1907.00397
-
Optimal Use of Experience in First Person Shooter Environments 24 Jun 2019 · 0 repositories · arXiv:1906.09734
-
Proximal Distilled Evolutionary Reinforcement Learning 24 Jun 2019 · 1 repository · arXiv:1906.09807Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Exploring Model-based Planning with Policy Networks 20 Jun 2019 · 1 repository · arXiv:1906.08649
-
Experience Replay Optimization 19 Jun 2019 · 0 repositories · arXiv:1906.08387
-
Evolutionary Reinforcement Learning for Sample-Efficient Multiagent Coordination 18 Jun 2019 · 0 repositories · arXiv:1906.07315
-
Goal-conditioned Imitation Learning 13 Jun 2019 · 1 repository · arXiv:1906.05838Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Reinforcement Learning for Unmanned Aerial Vehicle-Assisted Vehicular Networks 12 Jun 2019 · 0 repositories · arXiv:1906.05015
-
Boosting Soft Actor-Critic: Emphasizing Recent Experience without Forgetting the Past 10 Jun 2019 · 3 repositories · arXiv:1906.04009Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Exploration via Hindsight Goal Generation 10 Jun 2019 · 1 repository · arXiv:1906.04279
-
Improving Exploration in Soft-Actor-Critic with Normalizing Flows Policies 6 Jun 2019 · 1 repository · arXiv:1906.02771Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Exploration with Unreliable Intrinsic Reward in Multi-Agent Reinforcement Learning 5 Jun 2019 · 0 repositories · arXiv:1906.02138
-
Episodic Memory in Lifelong Language Learning 3 Jun 2019 · 2 repositories · arXiv:1906.01076Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Harnessing Reinforcement Learning for Neural Motion Planning 1 Jun 2019 · 1 repository · arXiv:1906.00214
-
Prioritized Sequence Experience Replay 25 May 2019 · 0 repositories · arXiv:1905.12726
-
Maximum Entropy-Regularized Multi-Goal Reinforcement Learning 21 May 2019 · 3 repositories · arXiv:1905.08786
-
Combining Experience Replay with Exploration by Random Network Distillation 18 May 2019 · 1 repository · arXiv:1905.07579
-
Bias-Reduced Hindsight Experience Replay with Virtual Goal Prioritization 14 May 2019 · 1 repository · arXiv:1905.05498
-
Deep Residual Reinforcement Learning 3 May 2019 · 1 repository · arXiv:1905.01072
-
Collaborative Evolutionary Reinforcement Learning 2 May 2019 · 1 repository · arXiv:1905.00976Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
ACTRCE: Augmenting Experience via Teacher’s Advice 1 May 2019 · 0 repositories
-
CEM-RL: Combining evolutionary and gradient-based methods for policy search 1 May 2019 · 0 repositories
-
DHER: Hindsight Experience Replay for Dynamic Goals 1 May 2019 · 1 repository
-
Learning agents with prioritization and parameter noise in continuous state and action space 1 May 2019 · 0 repositories
-
Learning Goal-Conditioned Value Functions with one-step Path rewards rather than Goal-Rewards 1 May 2019 · 0 repositories
-
Towards Combining On-Off-Policy Methods for Real-World Applications 24 Apr 2019 · 0 repositories · arXiv:1904.10642
-
Personalized Cancer Chemotherapy Schedule: a numerical comparison of performance and robustness in model-based and model-free scheduling methodologies 2 Apr 2019 · 0 repositories · arXiv:1904.01200
-
Deep Reinforcement Learning with Feedback-based Exploration 14 Mar 2019 · 2 repositories · arXiv:1903.06151
-
Complementary Learning for Overcoming Catastrophic Forgetting Using Experience Replay 11 Mar 2019 · 0 repositories · arXiv:1903.04566
-
Asynchronous Episodic Deep Deterministic Policy Gradient: Towards Continuous Control in Computationally Complex Environments 3 Mar 2019 · 1 repository · arXiv:1903.00827
-
Deep Reinforcement Learning using Genetic Algorithm for Parameter Optimization 19 Feb 2019 · 2 repositories · arXiv:1905.04100Syntology official (archive's flag): 8 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity 14 Feb 2019 · 5 repositories · arXiv:1902.05605
-
ACTRCE: Augmenting Experience via Teacher's Advice For Multi-Goal Reinforcement Learning 12 Feb 2019 · 0 repositories · arXiv:1902.04546
-
Competitive Experience Replay 1 Feb 2019 · 0 repositories · arXiv:1902.00528
-
Addressing Sample Complexity in Visual Tasks Using HER and Hallucinatory GANs 31 Jan 2019 · 2 repositories · arXiv:1901.11529
-
Reward Shaping via Meta-Learning 27 Jan 2019 · 0 repositories · arXiv:1901.09330
-
On-Policy Trust Region Policy Optimisation with Replay Buffers 18 Jan 2019 · 2 repositories · arXiv:1901.06212
-
Transfer Learning for Prosthetics Using Imitation Learning 15 Jan 2019 · 1 repository · arXiv:1901.04772
-
A Theoretical Analysis of Deep Q-Learning 1 Jan 2019 · 0 repositories · arXiv:1901.00137
-
Double Deep Q-Learning for Optimal Execution 17 Dec 2018 · 0 repositories · arXiv:1812.06600
-
Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach 16 Dec 2018 · 2 repositories · arXiv:1812.07394
-
Soft Actor-Critic Algorithms and Applications 13 Dec 2018 · 52 repositories · arXiv:1812.05905Syntology official: no sample here; runs from other or unrecorded repositories · 38 ran (of which 0 constructed an object rather than computing a result; 35 with no instrument failure: 4 honoured, 1 violated, 30 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 40 harvested samples) · 5 pointer-only (licence)
-
Off-Policy Deep Reinforcement Learning without Exploration 7 Dec 2018 · 10 repositories · arXiv:1812.02900Syntology community repositories only · 14 ran (of which 12 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Deep Reinforcement Learning and the Deadly Triad 6 Dec 2018 · 0 repositories · arXiv:1812.02648
-
Resource Constrained Deep Reinforcement Learning 3 Dec 2018 · 0 repositories · arXiv:1812.00600
-
Deep Multi-Agent Reinforcement Learning with Relevance Graphs 30 Nov 2018 · 1 repository · arXiv:1811.12557Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples)
-
Deep Reinforcement Learning for Autonomous Driving 28 Nov 2018 · 1 repository · arXiv:1811.11329
-
Experience Replay for Continual Learning 28 Nov 2018 · 0 repositories · arXiv:1811.11682
-
Modelling the Dynamic Joint Policy of Teammates with Attention Multi-agent DDPG 13 Nov 2018 · 0 repositories · arXiv:1811.07029
-
ACE: An Actor Ensemble Algorithm for Continuous Control with Tree Search 6 Nov 2018 · 1 repository · arXiv:1811.02696
-
Learning to Learn without Forgetting by Maximizing Transfer and Minimizing Interference 29 Oct 2018 · 3 repositories · arXiv:1810.11910Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Reconciling λ-Returns with Experience Replay 23 Oct 2018 · 1 repository · arXiv:1810.09967
-
Hierarchical Approaches for Reinforcement Learning in Parameterized Action Space 23 Oct 2018 · 0 repositories · arXiv:1810.09656
-
CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning 15 Oct 2018 · 1 repository · arXiv:1810.06284
-
Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space 10 Oct 2018 · 5 repositories · arXiv:1810.06394Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Energy-Based Hindsight Experience Prioritization 2 Oct 2018 · 2 repositories · arXiv:1810.01363
-
Hierarchical Deep Multiagent Reinforcement Learning with Temporal Abstraction 25 Sep 2018 · 0 repositories · arXiv:1809.09332