Methods › Reinforcement Learning › On-Policy TD Control › True Online TD Lambda
True Online TD Lambda
archive 2025-07-28 Description, source and code snippet are the archive's method entry.
True Online TD(λ) seeks to approximate the ideal online λ-return algorithm. It seeks to invert this ideal forward-view algorithm to produce an efficient backward-view algorithm using eligibility traces. It uses dutch traces rather than accumulating traces.
Source: Sutton and Seijen
Papers archive 2025-07-28
The archive tags no paper with this method.
Tasks archive 2025-07-28
The archive attaches no task to a paper tagged with this method.
Usage over time archive 2025-07-28
No dated papers to chart.
Components: the archive holds no method-to-method composition, so PwC's Components table cannot be rebuilt; the Papers list carries no Results column for the same reason (the archive does not join its leaderboard rows to method tags).
Categories archive 2025-07-28
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections