Methods › Reinforcement Learning › Offline Reinforcement Learning Methods
Offline Reinforcement Learning Methods
The archive carries no description for this collection.
Methods
All 5 methods in this collection, most-tagged first. Year is the archive's introduced_year; the archive stores 2000 when it has none, shown here as “–”. Papers counts distinct papers the archive tags with the method. Click a heading to sort.
| DPO Direct Preference Optimization | – | 409 |
| URL Umbrella Reinforcement Learning | – | 34 |
| R2D2 Recurrent Replay Distributed DQN | – | 25 |
| IQL Implicit Q-Learning | – | 16 |
| DeepCubeAI DeepCubeA + Imagination | – | 1 |