Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers, page 6
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 6 of 9: papers 501 to 600 of 865, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Learning Expected Emphatic Traces for Deep RL 12 Jul 2021 · 0 repositories · arXiv:2107.05405
-
MHER: Model-based Hindsight Experience Replay 1 Jul 2021 · 0 repositories · arXiv:2107.00306
-
AI-Based Secure NOMA and Cognitive Radio enabled Green Communications: Channel State Information and Battery Value Uncertainties 30 Jun 2021 · 0 repositories · arXiv:2106.15964
-
Unsupervised Continual Learning via Self-Adaptive Deep Clustering Approach 28 Jun 2021 · 0 repositories · arXiv:2106.14563
-
A Reinforcement Learning Approach for an IRS-assisted NOMA Network 17 Jun 2021 · 0 repositories · arXiv:2106.09611
-
Deep Reinforcement Learning Based Optimization for IRS Based UAV-NOMA Downlink Networks 17 Jun 2021 · 0 repositories · arXiv:2106.09616
-
Many Agent Reinforcement Learning Under Partial Observability 17 Jun 2021 · 0 repositories · arXiv:2106.09825
-
Unbiased Methods for Multi-Goal Reinforcement Learning 16 Jun 2021 · 0 repositories · arXiv:2106.08863
-
Efficient Continuous Control with Double Actors and Regularized Critics 6 Jun 2021 · 1 repository · arXiv:2106.03050
-
Deep Reinforcement Learning-based UAV Navigation and Control: A Soft Actor-Critic with Hindsight Experience Replay Approach 2 Jun 2021 · 0 repositories · arXiv:2106.01016
-
Variational Empowerment as Representation Learning for Goal-Based Reinforcement Learning 2 Jun 2021 · 0 repositories · arXiv:2106.01404
-
Adversarial Intrinsic Motivation for Reinforcement Learning 27 May 2021 · 1 repository · arXiv:2105.13345Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
Learning to Optimize Industry-Scale Dynamic Pickup and Delivery Problems 27 May 2021 · 0 repositories · arXiv:2105.12899
-
Optimistic Reinforcement Learning by Forward Kullback-Leibler Divergence Optimization 27 May 2021 · 0 repositories · arXiv:2105.12991
-
Near-optimal Offline and Streaming Algorithms for Learning Non-Linear Dynamical Systems 24 May 2021 · 0 repositories · arXiv:2105.11558
-
Improved Exploring Starts by Kernel Density Estimation-Based State-Space Coverage Acceleration in Reinforcement Learning 19 May 2021 · 1 repository · arXiv:2105.08990
-
Make Bipedal Robots Learn How to Imitate 15 May 2021 · 0 repositories · arXiv:2105.07193
-
Regret Minimization Experience Replay in Off-Policy Reinforcement Learning 15 May 2021 · 1 repository · arXiv:2105.07253
-
Hierarchical RNNs-Based Transformers MADDPG for Mixed Cooperative-Competitive Environments 11 May 2021 · 0 repositories · arXiv:2105.04888
-
Context-Based Soft Actor Critic for Environments with Non-stationary Dynamics 7 May 2021 · 1 repository · arXiv:2105.03310
-
Time-Aware Q-Networks: Resolving Temporal Irregularity for Deep Reinforcement Learning 6 May 2021 · 0 repositories · arXiv:2105.02580
-
End-to-end grasping policies for human-in-the-loop robots via deep reinforcement learning 26 Apr 2021 · 1 repository · arXiv:2104.12842
-
Class-Incremental Experience Replay for Continual Learning under Concept Drift 24 Apr 2021 · 1 repository · arXiv:2104.11861
-
Independent Reinforcement Learning for Weakly Cooperative Multiagent Traffic Control Problem 22 Apr 2021 · 1 repository · arXiv:2104.10917
-
Deep Deterministic Path Following 13 Apr 2021 · 0 repositories · arXiv:2104.06014
-
Deep Reinforcement Learning Based Controller for Active Heave Compensation 12 Apr 2021 · 0 repositories · arXiv:2104.05599
-
New Insights on Reducing Abrupt Representation Change in Online Continual Learning 11 Apr 2021 · 2 repositories · arXiv:2104.05025Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Progressive extension of reinforcement learning action dimension for asymmetric assembly tasks 6 Apr 2021 · 0 repositories · arXiv:2104.04078
-
Self-adaptive Torque Vectoring Controller Using Reinforcement Learning 27 Mar 2021 · 1 repository · arXiv:2103.14892
-
Continual Speaker Adaptation for Text-to-Speech Synthesis 26 Mar 2021 · 0 repositories · arXiv:2103.14512
-
Online Lifelong Generalized Zero-Shot Learning 19 Mar 2021 · 1 repository · arXiv:2103.10741
-
Simulation Studies on Deep Reinforcement Learning for Building Control with Human Interaction 14 Mar 2021 · 0 repositories · arXiv:2103.07919
-
Streaming Linear System Identification with Reverse Experience Replay 10 Mar 2021 · 0 repositories · arXiv:2103.05896
-
Learning to Play Soccer From Scratch: Sample-Efficient Emergent Coordination through Curriculum-Learning and Competition 9 Mar 2021 · 1 repository · arXiv:2103.05174
-
Selective Replay Enhances Learning in Online Continual Analogical Reasoning 6 Mar 2021 · 1 repository · arXiv:2103.03987
-
Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings 4 Mar 2021 · 1 repository · arXiv:2103.02886
-
Data-driven control of room temperature and bidirectional EV charging using deep reinforcement learning: simulations and experiments 2 Mar 2021 · 0 repositories · arXiv:2103.01886
-
A-DeepPixBis: Attentional Angular Margin for Face Anti-Spoofing 1 Mar 2021 · 0 repositories
-
Multi-Agent Path Planning based on MPC and DDPG 26 Feb 2021 · 0 repositories · arXiv:2102.13283
-
Twin actor twin delayed deep deterministic policy gradient (TATD3) learning for batch process control 25 Feb 2021 · 0 repositories · arXiv:2102.13012
-
Bias-reduced Multi-step Hindsight Experience Replay for Efficient Multi-goal Reinforcement Learning 25 Feb 2021 · 0 repositories · arXiv:2102.12962
-
Improved Regret Bound and Experience Replay in Regularized Policy Iteration 25 Feb 2021 · 0 repositories · arXiv:2102.12611
-
Hybrid Car-Following Strategy based on Deep Deterministic Policy Gradient and Cooperative Adaptive Cruise Control 24 Feb 2021 · 0 repositories · arXiv:2103.03796
-
Memory-based Deep Reinforcement Learning for POMDPs 24 Feb 2021 · 1 repository · arXiv:2102.12344Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Greedy-Step Off-Policy Reinforcement Learning 23 Feb 2021 · 0 repositories · arXiv:2102.11717
-
Escaping from Zero Gradient: Revisiting Action-Constrained Reinforcement Learning via Frank-Wolfe Policy Optimization 22 Feb 2021 · 0 repositories · arXiv:2102.11055
-
Stratified Experience Replay: Correcting Multiplicity Bias in Off-Policy Reinforcement Learning 22 Feb 2021 · 0 repositories · arXiv:2102.11319
-
Accelerated Sim-to-Real Deep Reinforcement Learning: Learning Collision Avoidance from Human Player 21 Feb 2021 · 1 repository · arXiv:2102.10711
-
Learning Memory-Dependent Continuous Control from Demonstrations 18 Feb 2021 · 0 repositories · arXiv:2102.09208
-
Adaptive Rational Activations to Boost Deep Reinforcement Learning 18 Feb 2021 · 4 repositories · arXiv:2102.09407Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
Deep Reinforcement Learning with Symmetric Prior for Predictive Power Allocation to Mobile Users 10 Feb 2021 · 0 repositories · arXiv:2103.13298
-
Reverb: A Framework For Experience Replay 9 Feb 2021 · 1 repository · arXiv:2102.04736
-
Explainable Reinforcement Learning for Longitudinal Control 6 Feb 2021 · 1 repository
-
Revisiting Prioritized Experience Replay: A Value Perspective 5 Feb 2021 · 2 repositories · arXiv:2102.03261
-
A review of motion planning algorithms for intelligent robotics 4 Feb 2021 · 0 repositories · arXiv:2102.02376
-
Learning Skills to Navigate without a Master: A Sequential Multi-Policy Reinforcement Learning Algorithm 30 Jan 2021 · 0 repositories · arXiv:2102.00168
-
OffCon³: What is state of the art anyway? 27 Jan 2021 · 1 repository · arXiv:2101.11331
-
GST: Group-Sparse Training for Accelerating Deep Reinforcement Learning 24 Jan 2021 · 0 repositories · arXiv:2101.09650
-
Learning Synthetic Environments for Reinforcement Learning with Evolution Strategies 24 Jan 2021 · 1 repository · arXiv:2101.09721
-
Deep Reinforcement Learning with Quantum-inspired Experience Replay 6 Jan 2021 · 0 repositories · arXiv:2101.02034
-
Compute- and Memory-Efficient Reinforcement Learning with Latent Experience Replay 1 Jan 2021 · 0 repositories
-
Factored Action Spaces in Deep Reinforcement Learning 1 Jan 2021 · 0 repositories
-
Hellinger Distance Constrained Regression 1 Jan 2021 · 0 repositories
-
Hindsight Curriculum Generation Based Multi-Goal Experience Replay 1 Jan 2021 · 0 repositories
-
Online Continual Learning Under Domain Shift 1 Jan 2021 · 0 repositories
-
PGPS : Coupling Policy Gradient with Population-based Search 1 Jan 2021 · 0 repositories
-
Playing Atari with Capsule Networks: A systematic comparison of CNN and CapsNets-based agents. 1 Jan 2021 · 0 repositories
-
Reinforcement Learning for Control of Valves 29 Dec 2020 · 2 repositories · arXiv:2012.14668
-
myGym: Modular Toolkit for Visuomotor Robotic Tasks 21 Dec 2020 · 0 repositories · arXiv:2012.11643
-
Mobile Robots Autonomous Exploration with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Policy Gradient for items Recommendation on Virtual Taobao 14 Dec 2020 · 0 repositories
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Ranking Items in Large-Scale Item Search Engines with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Virtual Autonomous Driving with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
OPAC: Opportunistic Actor-Critic 11 Dec 2020 · 0 repositories · arXiv:2012.06555
-
Efficient Reservoir Management through Deep Reinforcement Learning 7 Dec 2020 · 0 repositories · arXiv:2012.03822
-
Self-correcting Q-Learning 2 Dec 2020 · 0 repositories · arXiv:2012.01100
-
Deep Reinforcement Learning for Crowdsourced Urban Delivery: System States Characterization, Heuristics-guided Action Choice, and Rule-Interposing Integration 29 Nov 2020 · 0 repositories · arXiv:2011.14430
-
Episodic Self-Imitation Learning with Hindsight 26 Nov 2020 · 1 repository · arXiv:2011.13467
-
Learning from Simulation, Racing in Reality 26 Nov 2020 · 0 repositories · arXiv:2011.13332
-
Predictive PER: Balancing Priority and Diversity towards Stable Deep Reinforcement Learning 26 Nov 2020 · 0 repositories · arXiv:2011.13093
-
Reinforcement Learning for Robust Missile Autopilot Design 26 Nov 2020 · 0 repositories · arXiv:2011.12956
-
Consolidation via Policy Information Regularization in Deep RL for Multi-Agent Games 23 Nov 2020 · 0 repositories · arXiv:2011.11517
-
FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance 19 Nov 2020 · 6 repositories · arXiv:2011.09607
-
ACDER: Augmented Curiosity-Driven Experience Replay 16 Nov 2020 · 0 repositories · arXiv:2011.08027
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
A deep Q-Learning based Path Planning and Navigation System for Firefighting Environments 12 Nov 2020 · 0 repositories · arXiv:2011.06450
-
Reinforcement Learning Experiments and Benchmark for Solving Robotic Reaching Tasks 11 Nov 2020 · 1 repository · arXiv:2011.05782
-
RealAnt: An Open-Source Low-Cost Quadruped for Education and Research in Real-World Reinforcement Learning 5 Nov 2020 · 1 repository · arXiv:2011.03085
-
Deep Reinforcement Learning in Electricity Generation Investment for the Minimization of Long-Term Carbon Emissions and Electricity Costs 2 Nov 2020 · 0 repositories · arXiv:2011.02342
-
Self-Driving Network and Service Coordination Using Deep Reinforcement Learning 2 Nov 2020 · 1 repository
-
Optimizing Coverage and Capacity in Cellular Networks using Machine Learning 22 Oct 2020 · 0 repositories · arXiv:2010.13710
-
Deep Surrogate Q-Learning for Autonomous Driving 21 Oct 2020 · 0 repositories · arXiv:2010.11278
-
A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly 15 Oct 2020 · 0 repositories · arXiv:2010.08052
-
Rethinking Experience Replay: a Bag of Tricks for Continual Learning 12 Oct 2020 · 3 repositories · arXiv:2010.05595Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Deep RL With Information Constrained Policies: Generalization in Continuous Control 9 Oct 2020 · 0 repositories · arXiv:2010.04646
-
Hindsight Experience Replay with Kronecker Product Approximate Curvature 9 Oct 2020 · 0 repositories · arXiv:2010.06142
-
Neural Mask Generator: Learning to Generate Adaptive Word Maskings for Language Model Adaptation 6 Oct 2020 · 1 repository · arXiv:2010.02705
-
The Effectiveness of Memory Replay in Large Scale Continual Learning 6 Oct 2020 · 0 repositories · arXiv:2010.02418
-
An Empirical Investigation Towards Efficient Multi-Domain Language Model Pre-training 1 Oct 2020 · 1 repository · arXiv:2010.00784