Methods › Reinforcement Learning › Policy Gradient Methods › DDPG › Papers, page 2
Deep Deterministic Policy Gradient
DDPG
Papers archive 2025-07-28
archive papers tagged: 218 · with a code link: 71 · where Syntology ran a sample: 14 (13 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 218 tagged: 13 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument)
Page 2 of 3: papers 101 to 200 of 218, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
3DPG: Distributed Deep Deterministic Policy Gradient Algorithms for Networked Multi-Agent Systems 3 Jan 2022 · 0 repositories · arXiv:2201.00570
-
Toward Pareto Efficient Fairness-Utility Trade-off inRecommendation through Reinforcement Learning 1 Jan 2022 · 0 repositories · arXiv:2201.00140
-
Deep Reinforcement Learning for Optimal Power Flow with Renewables Using Graph Information 22 Dec 2021 · 0 repositories · arXiv:2112.11461
-
Automating Control of Overestimation Bias for Reinforcement Learning 26 Oct 2021 · 0 repositories · arXiv:2110.13523
-
Recurrent Off-policy Baselines for Memory-based Continuous Control 25 Oct 2021 · 1 repository · arXiv:2110.12628
-
Computationally Efficient Safe Reinforcement Learning for Power Systems 20 Oct 2021 · 0 repositories · arXiv:2110.10333
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations 3 Oct 2021 · 0 repositories · arXiv:2110.01101
-
Bootstrapped Hindsight Experience replay with Counterintuitive Prioritization 29 Sep 2021 · 0 repositories
-
Experience Replay More When It's a Key Transition in Deep Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Meta Attention For Off-Policy Actor-Critic 29 Sep 2021 · 0 repositories
-
SPP-RL: State Planning Policy Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Deep Reinforcement Learning Based Multidimensional Resource Management for Energy Harvesting Cognitive NOMA Communications 17 Sep 2021 · 0 repositories · arXiv:2109.09503
-
Responsive Regulation of Dynamic UAV Communication Networks Based on Deep Reinforcement Learning 25 Aug 2021 · 1 repository · arXiv:2108.11012
-
Safe Deep Reinforcement Learning for Multi-Agent Systems with Continuous Action Spaces 9 Aug 2021 · 1 repository · arXiv:2108.03952Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
AI-Based Secure NOMA and Cognitive Radio enabled Green Communications: Channel State Information and Battery Value Uncertainties 30 Jun 2021 · 0 repositories · arXiv:2106.15964
-
A Reinforcement Learning Approach for an IRS-assisted NOMA Network 17 Jun 2021 · 0 repositories · arXiv:2106.09611
-
Deep Reinforcement Learning Based Optimization for IRS Based UAV-NOMA Downlink Networks 17 Jun 2021 · 0 repositories · arXiv:2106.09616
-
Efficient Continuous Control with Double Actors and Regularized Critics 6 Jun 2021 · 1 repository · arXiv:2106.03050
-
Deep Reinforcement Learning-based UAV Navigation and Control: A Soft Actor-Critic with Hindsight Experience Replay Approach 2 Jun 2021 · 0 repositories · arXiv:2106.01016
-
Improved Exploring Starts by Kernel Density Estimation-Based State-Space Coverage Acceleration in Reinforcement Learning 19 May 2021 · 1 repository · arXiv:2105.08990
-
Deep Deterministic Path Following 13 Apr 2021 · 0 repositories · arXiv:2104.06014
-
Deep Reinforcement Learning Based Controller for Active Heave Compensation 12 Apr 2021 · 0 repositories · arXiv:2104.05599
-
Progressive extension of reinforcement learning action dimension for asymmetric assembly tasks 6 Apr 2021 · 0 repositories · arXiv:2104.04078
-
Self-adaptive Torque Vectoring Controller Using Reinforcement Learning 27 Mar 2021 · 1 repository · arXiv:2103.14892
-
Simulation Studies on Deep Reinforcement Learning for Building Control with Human Interaction 14 Mar 2021 · 0 repositories · arXiv:2103.07919
-
Data-driven control of room temperature and bidirectional EV charging using deep reinforcement learning: simulations and experiments 2 Mar 2021 · 0 repositories · arXiv:2103.01886
-
Multi-Agent Path Planning based on MPC and DDPG 26 Feb 2021 · 0 repositories · arXiv:2102.13283
-
Hybrid Car-Following Strategy based on Deep Deterministic Policy Gradient and Cooperative Adaptive Cruise Control 24 Feb 2021 · 0 repositories · arXiv:2103.03796
-
Escaping from Zero Gradient: Revisiting Action-Constrained Reinforcement Learning via Frank-Wolfe Policy Optimization 22 Feb 2021 · 0 repositories · arXiv:2102.11055
-
Accelerated Sim-to-Real Deep Reinforcement Learning: Learning Collision Avoidance from Human Player 21 Feb 2021 · 1 repository · arXiv:2102.10711
-
Deep Reinforcement Learning with Symmetric Prior for Predictive Power Allocation to Mobile Users 10 Feb 2021 · 0 repositories · arXiv:2103.13298
-
Explainable Reinforcement Learning for Longitudinal Control 6 Feb 2021 · 1 repository
-
A review of motion planning algorithms for intelligent robotics 4 Feb 2021 · 0 repositories · arXiv:2102.02376
-
Reinforcement Learning for Control of Valves 29 Dec 2020 · 2 repositories · arXiv:2012.14668
-
myGym: Modular Toolkit for Visuomotor Robotic Tasks 21 Dec 2020 · 0 repositories · arXiv:2012.11643
-
Policy Gradient for items Recommendation on Virtual Taobao 14 Dec 2020 · 0 repositories
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Virtual Autonomous Driving with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Efficient Reservoir Management through Deep Reinforcement Learning 7 Dec 2020 · 0 repositories · arXiv:2012.03822
-
FinRL: A Deep Reinforcement Learning Library for Automated Stock Trading in Quantitative Finance 19 Nov 2020 · 6 repositories · arXiv:2011.09607
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Deep Reinforcement Learning in Electricity Generation Investment for the Minimization of Long-Term Carbon Emissions and Electricity Costs 2 Nov 2020 · 0 repositories · arXiv:2011.02342
-
Self-Driving Network and Service Coordination Using Deep Reinforcement Learning 2 Nov 2020 · 1 repository
-
Optimizing Coverage and Capacity in Cellular Networks using Machine Learning 22 Oct 2020 · 0 repositories · arXiv:2010.13710
-
A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly 15 Oct 2020 · 0 repositories · arXiv:2010.08052
-
Hindsight Experience Replay with Kronecker Product Approximate Curvature 9 Oct 2020 · 0 repositories · arXiv:2010.06142
-
Knowledge-Assisted Deep Reinforcement Learning in 5G Scheduler Design: From Theoretical Framework to Implementation 17 Sep 2020 · 0 repositories · arXiv:2009.08346
-
Optimization-driven Hierarchical Learning Framework for Wireless Powered Backscatter-aided Relay Communications 4 Aug 2020 · 0 repositories · arXiv:2008.01366
-
Human-like Energy Management Based on Deep Reinforcement Learning and Historical Driving Experiences 16 Jul 2020 · 0 repositories · arXiv:2007.10126
-
Regularly Updated Deterministic Policy Gradient Algorithm 1 Jul 2020 · 0 repositories · arXiv:2007.00169
-
Distributed Uplink Beamforming in Cell-Free Networks Using Deep Reinforcement Learning 26 Jun 2020 · 0 repositories · arXiv:2006.15138
-
Some approaches used to overcome overestimation in Deep Reinforcement Learning algorithms 25 Jun 2020 · 0 repositories · arXiv:2006.14167
-
The Effect of Multi-step Methods on Overestimation in Deep Reinforcement Learning 23 Jun 2020 · 0 repositories · arXiv:2006.12692
-
WD3: Taming the Estimation Bias in Deep Reinforcement Learning 18 Jun 2020 · 0 repositories · arXiv:2006.12622
-
An online evolving framework for advancing reinforcement-learning based automated vehicle control 15 Jun 2020 · 0 repositories · arXiv:2006.08092
-
Optimization-driven Deep Reinforcement Learning for Robust Beamforming in IRS-assisted Wireless Communications 25 May 2020 · 0 repositories · arXiv:2005.11885
-
PBCS : Efficient Exploration and Exploitation Using a Synergy between Reinforcement Learning and Motion Planning 24 Apr 2020 · 0 repositories · arXiv:2004.11667
-
Model-based actor-critic: GAN (model generator) + DRL (actor-critic) => AGI 4 Apr 2020 · 0 repositories · arXiv:2004.04574
-
Obstacle Avoidance and Navigation Utilizing Reinforcement Learning with Reward Shaping 28 Mar 2020 · 1 repository · arXiv:2003.12863
-
Accelerating Deep Reinforcement Learning With the Aid of Partial Model: Energy-Efficient Predictive Video Streaming 21 Mar 2020 · 0 repositories · arXiv:2003.09708
-
Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations 19 Mar 2020 · 4 repositories · arXiv:2003.08938
-
PFPN: Continuous Control of Physically Simulated Characters using Particle Filtering Policy Network 16 Mar 2020 · 1 repository · arXiv:2003.06959
-
Online Meta-Critic Learning for Off-Policy Actor-Critic Methods 11 Mar 2020 · 1 repository · arXiv:2003.05334Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Dynamic Experience Replay 4 Mar 2020 · 0 repositories · arXiv:2003.02372
-
Contention Window Optimization in IEEE 802.11ax Networks with Deep Reinforcement Learning 3 Mar 2020 · 1 repository · arXiv:2003.01492
-
Reinforcement co-Learning of Deep and Spiking Neural Networks for Energy-Efficient Mapless Navigation with Neuromorphic Hardware 2 Mar 2020 · 1 repository · arXiv:2003.01157Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning 23 Jan 2020 · 4 repositories · arXiv:2001.08726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Improved Exploration through Latent Trajectory Optimization in Deep Deterministic Policy Gradient 15 Nov 2019 · 0 repositories · arXiv:1911.06833
-
Ctrl-Z: Recovering from Instability in Reinforcement Learning 9 Oct 2019 · 0 repositories · arXiv:1910.03732
-
QuaRL: Quantization for Fast and Environmentally Sustainable Reinforcement Learning 2 Oct 2019 · 1 repository · arXiv:1910.01055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Constrained Attractor Selection Using Deep Reinforcement Learning 23 Sep 2019 · 0 repositories · arXiv:1909.10500
-
AC-Teach: A Bayesian Actor-Critic Method for Policy Learning with an Ensemble of Suboptimal Teachers 9 Sep 2019 · 1 repository · arXiv:1909.04121Syntology 0 ran · 3 unverified (of 3 harvested samples)
-
Deterministic Value-Policy Gradients 9 Sep 2019 · 0 repositories · arXiv:1909.03939
-
Hierarchical Control for Bipedal Locomotion using Central Pattern Generators and Neural Networks 2 Sep 2019 · 1 repository · arXiv:1909.00732
-
Incremental Reinforcement Learning --- a New Continuous Reinforcement Learning Frame Based on Stochastic Differential Equation methods 8 Aug 2019 · 0 repositories · arXiv:1908.02974
-
Improved Reinforcement Learning through Imitation Learning Pretraining Towards Image-based Autonomous Driving 16 Jul 2019 · 0 repositories · arXiv:1907.06838
-
Shapley Q-value: A Local Reward Approach to Solve Global Reward Games 11 Jul 2019 · 2 repositories · arXiv:1907.05707Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Deep Reinforcement Learning for Unmanned Aerial Vehicle-Assisted Vehicular Networks 12 Jun 2019 · 0 repositories · arXiv:1906.05015
-
Harnessing Reinforcement Learning for Neural Motion Planning 1 Jun 2019 · 1 repository · arXiv:1906.00214
-
Deep Residual Reinforcement Learning 3 May 2019 · 1 repository · arXiv:1905.01072
-
CEM-RL: Combining evolutionary and gradient-based methods for policy search 1 May 2019 · 0 repositories
-
Learning agents with prioritization and parameter noise in continuous state and action space 1 May 2019 · 0 repositories
-
Personalized Cancer Chemotherapy Schedule: a numerical comparison of performance and robustness in model-based and model-free scheduling methodologies 2 Apr 2019 · 0 repositories · arXiv:1904.01200
-
Deep Reinforcement Learning with Feedback-based Exploration 14 Mar 2019 · 2 repositories · arXiv:1903.06151
-
Asynchronous Episodic Deep Deterministic Policy Gradient: Towards Continuous Control in Computationally Complex Environments 3 Mar 2019 · 1 repository · arXiv:1903.00827
-
CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity 14 Feb 2019 · 5 repositories · arXiv:1902.05605
-
Reward Shaping via Meta-Learning 27 Jan 2019 · 0 repositories · arXiv:1901.09330
-
On-Policy Trust Region Policy Optimisation with Replay Buffers 18 Jan 2019 · 2 repositories · arXiv:1901.06212
-
Transfer Learning for Prosthetics Using Imitation Learning 15 Jan 2019 · 1 repository · arXiv:1901.04772
-
Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach 16 Dec 2018 · 2 repositories · arXiv:1812.07394
-
Off-Policy Deep Reinforcement Learning without Exploration 7 Dec 2018 · 10 repositories · arXiv:1812.02900Syntology community repositories only · 14 ran (of which 12 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Resource Constrained Deep Reinforcement Learning 3 Dec 2018 · 0 repositories · arXiv:1812.00600
-
Deep Reinforcement Learning for Autonomous Driving 28 Nov 2018 · 1 repository · arXiv:1811.11329
-
Modelling the Dynamic Joint Policy of Teammates with Attention Multi-agent DDPG 13 Nov 2018 · 0 repositories · arXiv:1811.07029
-
ACE: An Actor Ensemble Algorithm for Continuous Control with Tree Search 6 Nov 2018 · 1 repository · arXiv:1811.02696
-
Hierarchical Approaches for Reinforcement Learning in Parameterized Action Space 23 Oct 2018 · 0 repositories · arXiv:1810.09656
-
CURIOUS: Intrinsically Motivated Modular Multi-Goal Reinforcement Learning 15 Oct 2018 · 1 repository · arXiv:1810.06284
-
Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space 10 Oct 2018 · 5 repositories · arXiv:1810.06394Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Curriculum goal masking for continuous deep reinforcement learning 17 Sep 2018 · 0 repositories · arXiv:1809.06146
-
Adversarial Deep Reinforcement Learning in Portfolio Management 29 Aug 2018 · 5 repositories · arXiv:1808.09940