Methods › Reinforcement Learning › Policy Gradient Methods › D4PG

Distributed Distributional DDPG

D4PG

11 papers tagged archive 2025-07-28

Introduced by Gabriel Barth-Maron et al. in Distributed Distributional Deterministic Policy Gradients

archive 2025-07-28 Description, source and code snippet are the archive's method entry.

D4PG, or Distributed Distributional DDPG, is a policy gradient algorithm that extends upon the DDPG. The improvements include a distributional updates to the DDPG algorithm, combined with the use of multiple distributed workers all writing into the same replay table. The biggest performance gain of other simpler changes was the use of N-step returns. The authors found that the use of prioritized experience replay was less crucial to the overall D4PG algorithm especially on harder problems.

PaperSource

Papers archive 2025-07-28

11 shown of 11, newest first. Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code; it is per sample and not a correctness claim.

Tasks archive 2025-07-28

15 tasks the archive attaches to papers tagged with this method, by distinct papers. A task without a page in the catalog is plain text.

TaskPapers
Reinforcement Learning (RL)7
reinforcement-learning6
Continuous Control4
Reinforcement Learning4
continuous-control4
Deep Reinforcement Learning3
OpenAI Gym3
Distributional Reinforcement Learning2
BIG-bench Machine Learning1
Benchmarking1
Image Generation1
Policy Gradient Methods1
Position1
Q-Learning1
quantile regression1

Usage over time archive 2025-07-28

Papers per year tagged with D4PG: 2018 to 2024, peak 3 3 0 2018: 1 paper 2018 2019: 1 paper 2019 2020: 3 papers 2020 2021: 0 papers 2021 2022: 2 papers 2022 2023: 3 papers 2023 2024: 1 paper 2024
Papers per year the archive tags with this method, by the paper's archive date (11 dated). Bars are counts, not a trend claim.

Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).

Categories archive 2025-07-28

Policy Gradient Methods

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections