Papers › Generalized Policy Improvement Algorithms with Theoretically Supported Sample Reuse

Generalized Policy Improvement Algorithms with Theoretically Supported Sample Reuse

28 Jun 2022arXiv:2206.13714archive 2025-07-28

James Queeney, Ioannis Ch. Paschalidis, Christos G. Cassandras

We develop a new class of model-free deep reinforcement learning algorithms for data-driven, learning-based control. Our Generalized Policy Improvement algorithms combine the policy improvement guarantees of on-policy methods with the efficiency of sample reuse, addressing a trade-off between two important deployment requirements for real-world control: (i) practical performance guarantees and (ii) data efficiency. We demonstrate the benefits of this new class of algorithms through extensive experimental analysis on a broad range of simulated control tasks.

PaperPDFCode

Code

jqueeney/gpi officialmentioned in papermentioned on GitHubtf report
jqueeney/geppo mentioned on GitHubtfGPL-3.0 report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Continuous ControlDecision MakingDeep Reinforcement LearningReinforcement Learningreinforcement-learning

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections