Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers, page 3
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Page 3 of 9: papers 201 to 300 of 865, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Using Experience Classification for Training Non-Markovian Tasks 18 Oct 2023 · 0 repositories · arXiv:2310.11678
-
Sim-to-Real Transfer of Adaptive Control Parameters for AUV Stabilization under Current Disturbance 17 Oct 2023 · 0 repositories · arXiv:2310.11075
-
Distributional Soft Actor-Critic with Three Refinements 9 Oct 2023 · 2 repositories · arXiv:2310.05858Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Energy Management in a Cooperative Energy Harvesting Wireless Sensor Network 9 Oct 2023 · 0 repositories · arXiv:2310.05911
-
GEAR: A GPU-Centric Experience Replay System for Large Reinforcement Learning Models 8 Oct 2023 · 1 repository · arXiv:2310.05205
-
Class-Incremental Learning Using Generative Experience Replay Based on Time-aware Regularization 5 Oct 2023 · 0 repositories · arXiv:2310.03898
-
Continual Contrastive Spoken Language Understanding 4 Oct 2023 · 0 repositories · arXiv:2310.02699
-
A Deep Reinforcement Learning Approach for Interactive Search with Sentence-level Feedback 3 Oct 2023 · 0 repositories · arXiv:2310.03043
-
Differentially Encoded Observation Spaces for Perceptive Reinforcement Learning 3 Oct 2023 · 1 repository · arXiv:2310.01767
-
Learning and reusing primitive behaviours to improve Hindsight Experience Replay sample efficiency 3 Oct 2023 · 1 repository · arXiv:2310.01827
-
CAD: Clustering And Deep Reinforcement Learning Based Multi-Period Portfolio Management Strategy 2 Oct 2023 · 0 repositories · arXiv:2310.01319
-
Cleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform 29 Sep 2023 · 1 repository · arXiv:2310.00036Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
E2Net: Resource-Efficient Continual Learning with Elastic Expansion Network 28 Sep 2023 · 1 repository · arXiv:2309.16117
-
Deep Reinforcement Learning for the Heat Transfer Control of Pulsating Impinging Jets 25 Sep 2023 · 0 repositories · arXiv:2309.13955
-
Enhancing data efficiency in reinforcement learning: a novel imagination mechanism based on mesh information propagation 25 Sep 2023 · 2 repositories · arXiv:2309.14243
-
A Long N-step Surrogate Stage Reward for Deep Reinforcement Learning 21 Sep 2023 · 0 repositories
-
Belief Projection-Based Reinforcement Learning for Environments with Delayed Feedback 21 Sep 2023 · 1 repository
-
DIFFER:Decomposing Individual Reward for Fair Experience Replay in Multi-Agent Reinforcement Learning 21 Sep 2023 · 0 repositories
-
Double Gumbel Q-Learning 21 Sep 2023 · 1 repository
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ϵ-Greedy Exploration 21 Sep 2023 · 0 repositories
-
Prioritizing Samples in Reinforcement Learning with Reducible Loss 21 Sep 2023 · 0 repositories
-
Safe Hierarchical Reinforcement Learning for CubeSat Task Scheduling Based on Energy Consumption 21 Sep 2023 · 0 repositories · arXiv:2309.12004
-
Taylor TD-learning 21 Sep 2023 · 0 repositories
-
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions 20 Sep 2023 · 0 repositories · arXiv:2309.10980
-
Distributional Estimation of Data Uncertainty for Surveillance Face Anti-spoofing 18 Sep 2023 · 0 repositories · arXiv:2309.09485
-
Attention Loss Adjusted Prioritized Experience Replay 13 Sep 2023 · 0 repositories · arXiv:2309.06684
-
UER: A Heuristic Bias Addressing Approach for Online Continual Learning 8 Sep 2023 · 0 repositories · arXiv:2309.04081
-
Hybrid of representation learning and reinforcement learning for dynamic and complex robotic motion planning 7 Sep 2023 · 0 repositories · arXiv:2309.03758
-
A Safe Deep Reinforcement Learning Approach for Energy Efficient Federated Learning in Wireless Communication Networks 21 Aug 2023 · 0 repositories · arXiv:2308.10664
-
A Fusion of Variational Distribution Priors and Saliency Map Replay for Continual 3D Reconstruction 17 Aug 2023 · 0 repositories · arXiv:2308.08812
-
AdaER: An Adaptive Experience Replay Approach for Continual Lifelong Learning 7 Aug 2023 · 0 repositories · arXiv:2308.03810
-
Vehicles Control: Collision Avoidance using Federated Deep Reinforcement Learning 4 Aug 2023 · 0 repositories · arXiv:2308.02614
-
Caching-at-STARS: the Next Generation Edge Caching 1 Aug 2023 · 0 repositories · arXiv:2308.00562
-
ETHER: Aligning Emergent Communication for Hindsight Experience Replay 28 Jul 2023 · 0 repositories · arXiv:2307.15494
-
Exploring reinforcement learning techniques for discrete and continuous control tasks in the MuJoCo environment 20 Jul 2023 · 1 repository · arXiv:2307.11166
-
Replay to Remember: Continual Layer-Specific Fine-tuning for German Speech Recognition 14 Jul 2023 · 0 repositories · arXiv:2307.07280
-
Safe Reinforcement Learning for Strategic Bidding of Virtual Power Plants in Day-Ahead Markets 11 Jul 2023 · 0 repositories · arXiv:2307.05812
-
Interpretable and Secure Trajectory Optimization for UAV-Assisted Communication 5 Jul 2023 · 0 repositories · arXiv:2307.02002
-
Safety-Aware Task Composition for Discrete and Continuous Reinforcement Learning 29 Jun 2023 · 0 repositories · arXiv:2306.17033
-
Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning 28 Jun 2023 · 0 repositories · arXiv:2306.15968
-
Curious Replay for Model-based Adaptation 28 Jun 2023 · 1 repository · arXiv:2306.15934
-
MRHER: Model-based Relay Hindsight Experience Replay for Sequential Object Manipulation Tasks with Sparse Rewards 28 Jun 2023 · 1 repository · arXiv:2306.16061
-
Evolutionary Strategy Guided Reinforcement Learning via MultiBuffer Communication 20 Jun 2023 · 0 repositories · arXiv:2306.11535
-
Safe, Efficient, Comfort, and Energy-saving Automated Driving through Roundabout Based on Deep Reinforcement Learning 20 Jun 2023 · 0 repositories · arXiv:2306.11465
-
Vanishing Bias Heuristic-guided Reinforcement Learning Algorithm 17 Jun 2023 · 0 repositories · arXiv:2306.10216
-
Temporal Difference Learning with Experience Replay 16 Jun 2023 · 0 repositories · arXiv:2306.09746
-
Decoupled Prioritized Resampling for Offline RL 8 Jun 2023 · 2 repositories · arXiv:2306.05412Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Progression Cognition Reinforcement Learning with Prioritized Experience for Multi-Vehicle Pursuit 8 Jun 2023 · 1 repository · arXiv:2306.05016
-
RLtools: A Fast, Portable Deep Reinforcement Learning Library for Continuous Control 6 Jun 2023 · 1 repository · arXiv:2306.03530
-
For SALE: State-Action Representation Learning for Deep Reinforcement Learning 4 Jun 2023 · 2 repositories · arXiv:2306.02451Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
ReLU to the Rescue: Improve Your On-Policy Actor-Critic with Positive Advantages 2 Jun 2023 · 1 repository · arXiv:2306.01460
-
Multi-environment lifelong deep reinforcement learning for medical imaging 31 May 2023 · 0 repositories · arXiv:2306.00188
-
Exploring the Promise and Limits of Real-Time Recurrent Learning 30 May 2023 · 1 repository · arXiv:2305.19044
-
History Repeats: Overcoming Catastrophic Forgetting For Event-Centric Temporal Knowledge Graph Completion 30 May 2023 · 0 repositories · arXiv:2305.18675
-
Optimizing Attention and Cognitive Control Costs Using Temporally-Layered Architectures 30 May 2023 · 1 repository · arXiv:2305.18701
-
DoMo-AC: Doubly Multi-step Off-policy Actor-Critic Algorithm 29 May 2023 · 0 repositories · arXiv:2305.18501
-
Reinforcement Learning With Reward Machines in Stochastic Games 27 May 2023 · 0 repositories · arXiv:2305.17372
-
Continual Learning with Strong Experience Replay 23 May 2023 · 1 repository · arXiv:2305.13622
-
OER: Offline Experience Replay for Continual Offline Reinforcement Learning 23 May 2023 · 0 repositories · arXiv:2305.13804
-
Sharing Lifelong Reinforcement Learning Knowledge via Modulating Masks 18 May 2023 · 2 repositories · arXiv:2305.10997
-
Attention-based QoE-aware Digital Twin Empowered Edge Computing for Immersive Virtual Reality 15 May 2023 · 0 repositories · arXiv:2305.08569
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm 11 May 2023 · 0 repositories · arXiv:2305.06657
-
Extracting Diagnosis Pathways from Electronic Health Records Using Deep Reinforcement Learning 10 May 2023 · 1 repository · arXiv:2305.06295
-
Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity 7 May 2023 · 1 repository · arXiv:2305.04180
-
Simple Noisy Environment Augmentation for Reinforcement Learning 4 May 2023 · 1 repository · arXiv:2305.02882
-
Map-based Experience Replay: A Memory-Efficient Solution to Catastrophic Forgetting in Reinforcement Learning 3 May 2023 · 1 repository · arXiv:2305.02054
-
Replay Memory as An Empirical MDP: Combining Conservative Estimation with Experience Replay 1 May 2023 · 1 repository
-
Federated Deep Reinforcement Learning for THz-Beam Search with Limited CSI 25 Apr 2023 · 0 repositories · arXiv:2304.13109
-
Regularizing Second-Order Influences for Continual Learning 20 Apr 2023 · 1 repository · arXiv:2304.10177Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Quantum deep Q learning with distributed prioritized experience replay 19 Apr 2023 · 0 repositories · arXiv:2304.09648
-
Active RIS-aided EH-NOMA Networks: A Deep Reinforcement Learning Approach 11 Apr 2023 · 0 repositories · arXiv:2304.12184
-
Understanding Reinforcement Learning Algorithms: The Progress from Basic Q-learning to Proximal Policy Optimization 31 Mar 2023 · 0 repositories · arXiv:2304.00026
-
Exploring Continual Learning of Diffusion Models 27 Mar 2023 · 0 repositories · arXiv:2303.15342
-
Preserving Linear Separability in Continual Learning by Backward Feature Projection 26 Mar 2023 · 1 repository · arXiv:2303.14595
-
Assessor-Guided Learning for Continual Environments 21 Mar 2023 · 1 repository · arXiv:2303.11624
-
Energy Management of Multi-mode Plug-in Hybrid Electric Vehicle using Multi-agent Deep Reinforcement Learning 16 Mar 2023 · 0 repositories · arXiv:2303.09658
-
SVDE: Scalable Value-Decomposition Exploration for Cooperative Multi-Agent Reinforcement Learning 16 Mar 2023 · 0 repositories · arXiv:2303.09058
-
Synthetic Experience Replay 12 Mar 2023 · 1 repository · arXiv:2303.06614Syntology official (archive's flag): 1 ran · 4 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
A Strategy-Oriented Bayesian Soft Actor-Critic Model 7 Mar 2023 · 0 repositories · arXiv:2303.04193
-
Swim: A General-Purpose, High-Performing, and Efficient Activation Function for Locomotion Control Tasks 5 Mar 2023 · 1 repository · arXiv:2303.02640
-
Eventual Discounting Temporal Logic Counterfactual Experience Replay 3 Mar 2023 · 0 repositories · arXiv:2303.02135
-
Hindsight States: Blending Sim and Real Task Elements for Efficient Reinforcement Learning 3 Mar 2023 · 0 repositories · arXiv:2303.02234
-
Can BERT Refrain from Forgetting on Sequential Tasks? A Probing Study 2 Mar 2023 · 1 repository · arXiv:2303.01081
-
The Ladder in Chaos: A Simple and Effective Improvement to General DRL Algorithms by Policy Path Trimming and Boosting 2 Mar 2023 · 0 repositories · arXiv:2303.01391
-
Combating Uncertainties in Wind and Distributed PV Energy Sources Using Integrated Reinforcement Learning and Time-Series Forecasting 27 Feb 2023 · 0 repositories · arXiv:2302.14094
-
Detachedly Learn a Classifier for Class-Incremental Learning 23 Feb 2023 · 0 repositories · arXiv:2302.11730
-
Revisiting the Gumbel-Softmax in MADDPG 23 Feb 2023 · 1 repository · arXiv:2302.11793
-
Selective experience replay compression using coresets for lifelong deep reinforcement learning in medical imaging 22 Feb 2023 · 0 repositories · arXiv:2302.11510
-
MAC-PO: Multi-Agent Experience Replay via Collective Priority Optimization 21 Feb 2023 · 1 repository · arXiv:2302.10418
-
UAV Path Planning Employing MPC- Reinforcement Learning Method Considering Collision Avoidance 21 Feb 2023 · 0 repositories · arXiv:2302.10669
-
Understanding the effect of varying amounts of replay per step 20 Feb 2023 · 0 repositories · arXiv:2302.10311
-
Prioritized offline Goal-swapping Experience Replay 15 Feb 2023 · 0 repositories · arXiv:2302.07741
-
Smart 6G Sky for Green Mobile IOT Networks 15 Feb 2023 · 1 repository · arXiv:2302.09022
-
Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning 13 Feb 2023 · 1 repository · arXiv:2302.06548Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
A Song of Ice and Fire: Analyzing Textual Autotelic Agents in ScienceWorld 10 Feb 2023 · 0 repositories · arXiv:2302.05244
-
RIS-Assisted Jamming Rejection and Path Planning for UAV-Borne IoT Platform: A New Deep Reinforcement Learning Framework 10 Feb 2023 · 0 repositories · arXiv:2302.04994
-
On the Limitation and Experience Replay for GNNs in Continual Learning 7 Feb 2023 · 0 repositories · arXiv:2302.03534
-
Deep Reinforcement Learning for Traffic Light Control in Intelligent Transportation Systems 4 Feb 2023 · 0 repositories · arXiv:2302.03669
-
V2N Service Scaling with Deep Reinforcement Learning 30 Jan 2023 · 0 repositories · arXiv:2301.13324
-
Distilling Internet-Scale Vision-Language Models into Embodied Agents 29 Jan 2023 · 0 repositories · arXiv:2301.12507