AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article

Enhanced reinforcement learning-model predictive control for distributed energy systems: Overcoming local and global optimization limitations

Hua Meng1Huaijiang Bin1Fanyue Qian1Tingting Xu1Chaoliang Wang2Wei Liu2Yuting Yao1( )Yingjun Ruan1( )
College of Mechanical and Energy Engineering, Tongji University, Shanghai 200092, China
State Grid Zhejiang Marketing Service Centre, Hangzhou 310014, China
Show Author Information

Abstract

The complex structures of distributed energy systems (DES) and uncertainties arising from renewable energy sources and user load variations pose significant operational challenges. Model predictive control (MPC) and reinforcement learning (RL) are widely used to optimize DES by predicting future outcomes based on the current state. However, MPC’s real-time application is constrained by its computational demands, making it less suitable for complex systems with extended predictive horizons. Meanwhile, RL’s model-free approach leads to suboptimal data utilization, limiting its overall performance. To address these issues, this study proposes an improved reinforcement learning-model predictive control (RL-MPC) algorithm that combines the high-precision local optimization of MPC with the global optimization capability of RL. In this study, we enhance the existing RL-MPC algorithm by increasing the number of optimization steps performed by the MPC component. We evaluated RL, MPC, and the enhanced RL-MPC on a DES comprising a photovoltaic (PV) and battery energy storage system (BESS). The results indicate the following: (1) The twin delayed deep deterministic policy gradient (TD3) algorithm outperforms other RL algorithms in energy cost optimization, but is outperformed in all cases by RL-MPC. (2) For both MPC and RL-MPC, when the mean absolute percentage error (MAPE) of the first-step prediction is 5%, the total cost increases by ~1.2% compared to that when the MAPE is 0%. However, if the accuracy of the initial prediction data remains constant while only the error gradient of the data sequence increases, the total cost remains nearly unchanged, with an increase of only ~0.1%. (3) Within a 12 h predictive horizon, RL-MPC outperforms MPC, suggesting it as a suitable alternative to MPC when high-accuracy prediction data are limited.

References

【1】
【1】
 
 
Building Simulation
Pages 547-567

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Meng H, Bin H, Qian F, et al. Enhanced reinforcement learning-model predictive control for distributed energy systems: Overcoming local and global optimization limitations. Building Simulation, 2025, 18(3): 547-567. https://doi.org/10.1007/s12273-024-1227-9

2514

Views

15

Crossref

15

Web of Science

19

Scopus

3

CSCD

Received: 20 August 2024
Revised: 14 November 2024
Accepted: 06 December 2024
Published: 16 January 2025
© Tsinghua University Press 2025