Methods › Reinforcement Learning › Offline Reinforcement Learning Methods

Offline Reinforcement Learning Methods

5 methods 485 papers tagged archive 2025-07-28

The archive carries no description for this collection.

Methods

All 5 methods in this collection, most-tagged first. Year is the archive's introduced_year; the archive stores 2000 when it has none, shown here as “–”. Papers counts distinct papers the archive tags with the method. Click a heading to sort.

DPO Direct Preference Optimization – 409
URL Umbrella Reinforcement Learning – 34
R2D2 Recurrent Replay Distributed DQN – 25
IQL Implicit Q-Learning – 16
DeepCubeAI DeepCubeA + Imagination – 1