Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers, page 4
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 4 of 9: papers 301 to 400 of 865, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
A Memory Efficient Deep Reinforcement Learning Approach For Snake Game Autonomous Agents 27 Jan 2023 · 0 repositories · arXiv:2301.11977
-
A Novel Deep Reinforcement Learning-based Approach for Enhancing Spectral Efficiency of IRS-assisted Wireless Systems 24 Jan 2023 · 0 repositories · arXiv:2302.14706
-
On Multi-Agent Deep Deterministic Policy Gradients and their Explainability for SMARTS Environment 20 Jan 2023 · 0 repositories · arXiv:2301.09420
-
Automated deep reinforcement learning for real-time scheduling strategy of multi-energy system integrated with post-carbon and direct-air carbon captured system 18 Jan 2023 · 0 repositories · arXiv:2301.07768
-
PDVN: A Patch-based Dual-view Network for Face Liveness Detection using Light Field Focal Stack 17 Jan 2023 · 0 repositories
-
Online Class-Incremental Learning For Real-World Food Image Classification 12 Jan 2023 · 1 repository · arXiv:2301.05246
-
Actor-Director-Critic: A Novel Deep Reinforcement Learning Framework 10 Jan 2023 · 0 repositories · arXiv:2301.03887
-
Hint assisted reinforcement learning: an application in radio astronomy 10 Jan 2023 · 1 repository · arXiv:2301.03933
-
Extreme Q-Learning: MaxEnt RL without Entropy 5 Jan 2023 · 4 repositories · arXiv:2301.02328Syntology 8 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 7 pointer-only (licence)
-
On the Geometry of Reinforcement Learning in Continuous State and Action Spaces 29 Dec 2022 · 0 repositories · arXiv:2301.00009
-
Lifelong Reinforcement Learning with Modulating Masks 21 Dec 2022 · 4 repositories · arXiv:2212.11110
-
Neighboring state-based RL Exploration 21 Dec 2022 · 0 repositories · arXiv:2212.10712
-
Efficient Exploration in Resource-Restricted Reinforcement Learning 14 Dec 2022 · 0 repositories · arXiv:2212.06988
-
Off-Policy Deep Reinforcement Learning Algorithms for Handling Various Robotic Manipulator Tasks 11 Dec 2022 · 0 repositories · arXiv:2212.05572
-
TD3 with Reverse KL Regularizer for Offline Reinforcement Learning from Mixed Datasets 5 Dec 2022 · 1 repository · arXiv:2212.02125
-
Automatic Discovery of Multi-perspective Process Model using Reinforcement Learning 30 Nov 2022 · 0 repositories · arXiv:2211.16687
-
The Effectiveness of World Models for Continual Reinforcement Learning 29 Nov 2022 · 2 repositories · arXiv:2211.15944
-
Learning from Good Trajectories in Offline Multi-Agent Reinforcement Learning 28 Nov 2022 · 0 repositories · arXiv:2211.15612
-
RL-Based Guidance in Outpatient Hysteroscopy Training: A Feasibility Study 26 Nov 2022 · 0 repositories · arXiv:2211.14541
-
Double Deep Q-Learning in Opponent Modeling 24 Nov 2022 · 0 repositories · arXiv:2211.15384
-
CACTO: Continuous Actor-Critic with Trajectory Optimization -- Towards global optimality 12 Nov 2022 · 0 repositories · arXiv:2211.06625
-
Deep Reinforcement Learning Microgrid Optimization Strategy Considering Priority Flexible Demand Side 11 Nov 2022 · 0 repositories · arXiv:2211.05946
-
The Benefits of Model-Based Generalization in Reinforcement Learning 4 Nov 2022 · 1 repository · arXiv:2211.02222Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Wind Power Forecasting Considering Data Privacy Protection: A Federated Deep Reinforcement Learning Approach 2 Nov 2022 · 0 repositories · arXiv:2211.02674
-
Using Contrastive Samples for Identifying and Leveraging Possible Causal Relationships in Reinforcement Learning 28 Oct 2022 · 0 repositories · arXiv:2210.17296
-
Leveraging Demonstrations with Latent Space Priors 26 Oct 2022 · 1 repository · arXiv:2210.14685Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
AACHER: Assorted Actor-Critic Deep Reinforcement Learning with Hindsight Experience Replay 24 Oct 2022 · 1 repository · arXiv:2210.12892
-
MEET: A Monte Carlo Exploration-Exploitation Trade-off for Buffer Sampling 24 Oct 2022 · 1 repository · arXiv:2210.13545
-
Navigating Memory Construction by Global Pseudo-Task Simulation for Continual Learning 16 Oct 2022 · 1 repository · arXiv:2210.08442
-
Multiagent Reinforcement Learning Based on Fusion-Multiactor-Attention-Critic for Multiple-Unmanned-Aerial-Vehicle Navigation Control 10 Oct 2022 · 1 repository
-
Algorithmic Trading Using Continuous Action Space Deep Reinforcement Learning 7 Oct 2022 · 0 repositories · arXiv:2210.03469
-
Elastic Step DQN: A novel multi-step algorithm to alleviate overestimation in Deep QNetworks 7 Oct 2022 · 0 repositories · arXiv:2210.03325
-
A Novel Entropy-Maximizing TD3-based Reinforcement Learning for Automatic PID Tuning 5 Oct 2022 · 0 repositories · arXiv:2210.02381
-
How Relevant is Selective Memory Population in Lifelong Language Learning? 3 Oct 2022 · 0 repositories · arXiv:2210.00940
-
Understanding Hindsight Goal Relabeling from a Divergence Minimization Perspective 26 Sep 2022 · 0 repositories · arXiv:2209.13046
-
Minimizing Human Assistance: Augmenting a Single Demonstration for Deep Reinforcement Learning 22 Sep 2022 · 0 repositories · arXiv:2209.11275
-
Continual VQA for Disaster Response Systems 21 Sep 2022 · 1 repository · arXiv:2209.10320
-
A Deep Reinforcement Learning-Based Charging Scheduling Approach with Augmented Lagrangian for Electric Vehicle 20 Sep 2022 · 0 repositories · arXiv:2209.09772
-
M²DQN: A Robust Method for Accelerating Deep Q-learning Network 16 Sep 2022 · 1 repository · arXiv:2209.07809
-
Data efficient reinforcement learning and adaptive optimal perimeter control of network traffic dynamics 13 Sep 2022 · 0 repositories · arXiv:2209.05726
-
Online Continual Learning via the Meta-learning Update with Multi-scale Knowledge Distillation and Data Augmentation 12 Sep 2022 · 0 repositories · arXiv:2209.06107
-
Experimental Study on The Effect of Multi-step Deep Reinforcement Learning in POMDPs 12 Sep 2022 · 1 repository · arXiv:2209.04999
-
Selecting Related Knowledge via Efficient Channel Attention for Online Continual Learning 9 Sep 2022 · 0 repositories · arXiv:2209.04212
-
A New Approach to Training Multiple Cooperative Agents for Autonomous Driving 5 Sep 2022 · 0 repositories · arXiv:2209.02157
-
Reinforced Continual Learning for Graphs 4 Sep 2022 · 0 repositories · arXiv:2209.01556
-
Actor Prioritized Experience Replay 1 Sep 2022 · 1 repository · arXiv:2209.00532
-
Cluster-based Sampling in Hindsight Experience Replay for Robotic Tasks (Student Abstract) 31 Aug 2022 · 0 repositories · arXiv:2208.14741
-
Digital Twin Assisted Risk-Aware Sleep Mode Management Using Deep Q-Networks 30 Aug 2022 · 0 repositories · arXiv:2208.14380
-
Goal-Conditioned Q-Learning as Knowledge Distillation 28 Aug 2022 · 1 repository · arXiv:2208.13298
-
Variance Reduction based Experience Replay for Policy Optimization 25 Aug 2022 · 1 repository · arXiv:2208.12341
-
Bayesian Soft Actor-Critic: A Directed Acyclic Strategy Graph Based Deep Reinforcement Learning 11 Aug 2022 · 2 repositories · arXiv:2208.06033
-
Plug-and-Play Model-Agnostic Counterfactual Policy Synthesis for Deep Reinforcement Learning based Recommendation 10 Aug 2022 · 0 repositories · arXiv:2208.05142
-
Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling 1 Aug 2022 · 1 repository · arXiv:2208.00728
-
Relay Hindsight Experience Replay: Self-Guided Continual Reinforcement Learning for Sequential Object Manipulation Tasks with Sparse Rewards 1 Aug 2022 · 1 repository · arXiv:2208.00843
-
Distributional Actor-Critic Ensemble for Uncertainty-Aware Continuous Control 27 Jul 2022 · 0 repositories · arXiv:2207.13730
-
Safe and Robust Experience Sharing for Deterministic Policy Gradient Algorithms 27 Jul 2022 · 1 repository · arXiv:2207.13453
-
Abstract Demonstrations and Adaptive Exploration for Efficient and Stable Multi-step Sparse Reward Reinforcement Learning 19 Jul 2022 · 1 repository · arXiv:2207.09243
-
Associative Memory Based Experience Replay for Deep Reinforcement Learning 16 Jul 2022 · 0 repositories · arXiv:2207.07791
-
Multi-Agent Deep Reinforcement Learning-Driven Mitigation of Adverse Effects of Cyber-Attacks on Electric Vehicle Charging Station 14 Jul 2022 · 0 repositories · arXiv:2207.07041
-
Consistency is the key to further mitigating catastrophic forgetting in continual learning 11 Jul 2022 · 1 repository · arXiv:2207.04998Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Asynchronous Curriculum Experience Replay: A Deep Reinforcement Learning Approach for UAV Autonomous Motion Control in Unknown Dynamic Environments 4 Jul 2022 · 0 repositories · arXiv:2207.01251
-
Memory Population in Continual Learning via Outlier Elimination 4 Jul 2022 · 1 repository · arXiv:2207.01145Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Distributed Online System Identification for LTI Systems Using Reverse Experience Replay 3 Jul 2022 · 0 repositories · arXiv:2207.01062
-
USHER: Unbiased Sampling for Hindsight Experience Replay 3 Jul 2022 · 1 repository · arXiv:2207.01115Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
Deep Reinforcement Learning with Swin Transformers 30 Jun 2022 · 1 repository · arXiv:2206.15269
-
EnvPool: A Highly Parallel Reinforcement Learning Environment Execution Engine 21 Jun 2022 · 3 repositories · arXiv:2206.10558Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
FedER: Federated Learning through Experience Replay and Privacy-Preserving Data Synthesis 20 Jun 2022 · 1 repository · arXiv:2206.10048
-
MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer 20 Jun 2022 · 1 repository · arXiv:2206.10607
-
Two-Hop Age of Information Scheduling for Multi-UAV Assisted Mobile Edge Computing: FRL vs MADDPG 19 Jun 2022 · 0 repositories · arXiv:2206.09488
-
Autonomous Platoon Control with Integrated Deep Reinforcement Learning and Dynamic Programming 15 Jun 2022 · 0 repositories · arXiv:2206.07536
-
Computation Offloading and Resource Allocation in F-RANs: A Federated Deep Reinforcement Learning Approach 13 Jun 2022 · 0 repositories · arXiv:2206.05881
-
SYNERgy between SYNaptic consolidation and Experience Replay for general continual learning 8 Jun 2022 · 1 repository · arXiv:2206.04016Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Introspective Experience Replay: Look Back When Surprised 7 Jun 2022 · 1 repository · arXiv:2206.03171
-
Balancing Profit, Risk, and Sustainability for Portfolio Management 6 Jun 2022 · 0 repositories · arXiv:2207.02134
-
Goal-Space Planning with Subgoal Models 6 Jun 2022 · 0 repositories · arXiv:2206.02902
-
Robust Adversarial Attacks Detection based on Explainable Deep Reinforcement Learning For UAV Guidance and Planning 6 Jun 2022 · 0 repositories · arXiv:2206.02670
-
Joint Energy Dispatch and Unit Commitment in Microgrids Based on Deep Reinforcement Learning 3 Jun 2022 · 0 repositories · arXiv:2206.01663
-
Equivariant Reinforcement Learning for Quadrotor UAV 2 Jun 2022 · 0 repositories · arXiv:2206.01233
-
Stabilizing Q-learning with Linear Architectures for Provably Efficient Learning 1 Jun 2022 · 0 repositories · arXiv:2206.00796
-
Truly Deterministic Policy Optimization 30 May 2022 · 1 repository · arXiv:2205.15379
-
Task-Agnostic Continual Reinforcement Learning: Gaining Insights and Overcoming Challenges 28 May 2022 · 2 repositories · arXiv:2205.14495Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Multiple Domain Cyberspace Attack and Defense Game Based on Reward Randomization Reinforcement Learning 23 May 2022 · 0 repositories · arXiv:2205.10990
-
Memory-efficient Reinforcement Learning with Value-based Knowledge Consolidation 22 May 2022 · 1 repository · arXiv:2205.10868Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Dexterous Robotic Manipulation using Deep Reinforcement Learning and Knowledge Transfer for Complex Sparse Reward-based Tasks 19 May 2022 · 1 repository · arXiv:2205.09683Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Neighborhood Mixup Experience Replay: Local Convex Interpolation for Improved Sample Efficiency in Continuous Control Tasks 18 May 2022 · 1 repository · arXiv:2205.09117
-
Optimal Adaptive Prediction Intervals for Electricity Load Forecasting in Distribution Systems via Reinforcement Learning 18 May 2022 · 1 repository · arXiv:2205.08698
-
Efficient Off-Policy Reinforcement Learning via Brain-Inspired Computing 14 May 2022 · 0 repositories · arXiv:2205.06978
-
Hybrid Reinforcement Learning for STAR-RISs: A Coupled Phase-Shift Model Based Beamformer 10 May 2022 · 0 repositories · arXiv:2205.05029
-
Variance Reduction based Partial Trajectory Reuse to Accelerate Policy Gradient Optimization 6 May 2022 · 1 repository · arXiv:2205.02976
-
Revisiting Gaussian mixture critics in off-policy reinforcement learning: a sample-based approach 21 Apr 2022 · 1 repository · arXiv:2204.10256
-
Continual Predictive Learning from Videos 12 Apr 2022 · 1 repository · arXiv:2204.05624
-
Confidence Estimation Transformer for Long-term Renewable Energy Forecasting in Reinforcement Learning-based Power Grid Dispatching 10 Apr 2022 · 1 repository · arXiv:2204.04612
-
MA-Dreamer: Coordination and communication through shared imagination 10 Apr 2022 · 0 repositories · arXiv:2204.04687
-
Semantic Exploration from Language Abstractions and Pretrained Representations 8 Apr 2022 · 0 repositories · arXiv:2204.05080
-
Is Word Error Rate a good evaluation metric for Speech Recognition in Indic Languages? 30 Mar 2022 · 0 repositories · arXiv:2203.16601
-
Topological Experience Replay 29 Mar 2022 · 1 repository · arXiv:2203.15845
-
Collaborative Intelligent Reflecting Surface Networks with Multi-Agent Reinforcement Learning 26 Mar 2022 · 0 repositories · arXiv:2203.14152
-
Non-Parametric Stochastic Policy Gradient with Strategic Retreat for Non-Stationary Environment 24 Mar 2022 · 0 repositories · arXiv:2203.14905
-
Remember and Forget Experience Replay for Multi-Agent Reinforcement Learning 24 Mar 2022 · 0 repositories · arXiv:2203.13319
-
Continual Sequence Generation with Adaptive Compositional Modules 20 Mar 2022 · 2 repositories · arXiv:2203.10652Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 1 pointer-only (licence)