Papers › Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling

Performance Comparison of Deep RL Algorithms for Energy Systems Optimal Scheduling

1 Aug 2022arXiv:2208.00728archive 2025-07-28

Hou Shengren, Edgar Mauricio Salazar, Pedro P. Vergara, Peter Palensky

Taking advantage of their data-driven and model-free features, Deep Reinforcement Learning (DRL) algorithms have the potential to deal with the increasing level of uncertainty due to the introduction of renewable-based generation. To deal simultaneously with the energy systems' operational cost and technical constraints (e.g, generation-demand power balance) DRL algorithms must consider a trade-off when designing the reward function. This trade-off introduces extra hyperparameters that impact the DRL algorithms' performance and capability of providing feasible solutions. In this paper, a performance comparison of different DRL algorithms, including DDPG, TD3, SAC, and PPO, are presented. We aim to provide a fair comparison of these DRL algorithms for energy systems optimal scheduling problems. Results show DRL algorithms' capability of providing in real-time good-quality solutions, even in unseen operational scenarios, when compared with a mathematical programming model of the energy system optimal scheduling problem. Nevertheless, in the case of large peak consumption, these algorithms failed to provide feasible solutions, which can impede their practical implementation.

PaperPDFCode

Code

shengrenhou/drl-for-energy-systems-optimal-scheduling officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Deep Reinforcement LearningReinforcement Learning (RL)Schedulingenergy management

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

1x1 ConvolutionAdamAverage PoolingBatch NormalizationClipped Double Q-learningConvolutionDDPGDense ConnectionsDilated ConvolutionEntropy RegularizationExperience ReplayGlobal Average PoolingPPOReLUSACTD3Target Policy SmoothingWeight Decay

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections