Methods › Reinforcement Learning › Off-Policy TD Control
Off-Policy TD Control
The archive carries no description for this collection.
Methods
All 5 methods in this collection, most-tagged first. Year is the archive's introduced_year; the archive stores 2000 when it has none, shown here as “–”. Papers counts distinct papers the archive tags with the method. Click a heading to sort.
| Q-Learning | 1984 | 1,734 |
| Clipped Double Q-learning | – | 122 |
| Double Q-learning | – | 112 |
| REM Random Ensemble Mixture | – | 48 |
| Expected Sarsa | – | 9 |