Methods › Reinforcement Learning › Policy Gradient Methods › DDPG › Papers, page 3
Deep Deterministic Policy Gradient
DDPG
Papers archive 2025-07-28
archive papers tagged: 218 · with a code link: 71 · where Syntology ran a sample: 14 (13 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (14 of 218 tagged: 13 with a run with no instrument failure, 1 where every run was a failure of Syntology's instrument)
Page 3 of 3: papers 201 to 218 of 218, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Bipedal Walking Robot using Deep Deterministic Policy Gradient 16 Jul 2018 · 3 repositories · arXiv:1807.05924
-
Deterministic Policy Gradients With General State Transitions 10 Jul 2018 · 0 repositories · arXiv:1807.03708
-
Learning to Explore via Meta-Policy Gradient 1 Jul 2018 · 0 repositories
-
Stroke-based Character Reconstruction 23 Jun 2018 · 1 repository · arXiv:1806.08990
-
Randomized Value Functions via Multiplicative Normalizing Flows 6 Jun 2018 · 2 repositories · arXiv:1806.02315
-
Advances in Experience Replay 15 May 2018 · 1 repository · arXiv:1805.05536
-
Learning to Explore with Meta-Policy Gradient 13 Mar 2018 · 0 repositories · arXiv:1803.05044
-
Distributed Prioritized Experience Replay 2 Mar 2018 · 15 repositories · arXiv:1803.00933Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples) · 3 pointer-only (licence)
-
GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms 14 Feb 2018 · 1 repository · arXiv:1802.05054Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Pretraining Deep Actor-Critic Reinforcement Learning Algorithms With Expert Demonstrations 31 Jan 2018 · 0 repositories · arXiv:1801.10459
-
Learning to Run with Actor-Critic Ensemble 25 Dec 2017 · 2 repositories · arXiv:1712.08987
-
A novel DDPG method with prioritized experience replay 1 Oct 2017 · 1 repository
-
Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards 27 Jul 2017 · 4 repositories · arXiv:1707.08817
-
The Intentional Unintentional Agent: Learning to Solve Many Continuous Control Tasks Simultaneously 11 Jul 2017 · 0 repositories · arXiv:1707.03300
-
Parameter Space Noise for Exploration 6 Jun 2017 · 10 repositories · arXiv:1706.01905Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Discrete Sequential Prediction of Continuous Actions for Deep RL 14 May 2017 · 0 repositories · arXiv:1705.05035
-
Actor-critic versus direct policy search: a comparison based on sample complexity 29 Jun 2016 · 1 repository · arXiv:1606.09152
-
Continuous control with deep reinforcement learning 9 Sep 2015 · 161 repositories · arXiv:1509.02971Syntology 163 ran (of which 126 constructed an object rather than computing a result; 152 with no instrument failure: 3 honoured, 0 violated, 149 with no contract checked; 11 where Syntology's instrument failed) · 143 unverified (of 306 harvested samples) · 163 pointer-only (licence)