AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (2.7 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Publishing Language: Chinese | Open Access

A power flow convergence adjustment based on deep reinforcement learning with hybrid action space

Tao WU1Haohao WANG2Tianran LI1
School of Electrical & Automation Engineering, Nanjing Normal University, Nanjing 210023, China
State Grid Electric Power Research Institute (NARI Group Corporation), Nanjing 211106, China
Show Author Information

Abstract

As the operation modes of power systems become increasingly complex, the difficulty of power flow convergence adjustment also increases. Traditional methods that rely on human expertise suffer from delayed response and low efficiency, making them ill-suited for complex scenarios involving diverse and high-dimensional control variables. To address this challenge, a power flow adjustment method based on deep reinforcement learning with a hybrid action space is proposed. Firstly, a power flow convergence discriminator is developed by integrating physical priors with data-driven techniques to enable real-time identification of whether the power flow converges. The output convergence probability is used as a reward guidance signal in deep reinforcement learning. Then, a reinforcement learning environment for convergence adjustment is defined. The state space integrates system-level statistical features with node-level individual features. The action space encompasses both continuous and discrete control variables. And the reward function combines convergence identification results with feedback from the adjustment process to guide policy optimization toward the feasible region. Next, an Actor-Critic network with a hybrid action space is constructed, which hierarchically decouples and models subtasks including device selection, continuous regulation, and discrete control. Finally, simulation analysis and comparative experiments are conducted on the improved IEEE 39-bus and 118-bus systems. The results demonstrate that the power flow convergence discriminator enhanced with physical priors significantly improves both the accuracy and generalization capability of convergence identification compared to traditional models. Furthermore, the proposed coordinated optimization strategy, which integrates continuous-discrete actions, achieves notable improvements in power flow adjustment efficiency and convergence success rate over existing methods.

CLC number: TM744 Document code: A

References

【1】
【1】
 
 
Electric Power Engineering Technology
Pages 50-60

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
WU T, WANG H, LI T. A power flow convergence adjustment based on deep reinforcement learning with hybrid action space. Electric Power Engineering Technology, 2026, 45(5): 50-60. https://doi.org/10.12158/j.2096-3203.2026.05.005

4

Views

0

Downloads

0

Crossref

0

Scopus

Received: 09 October 2025
Revised: 26 December 2025
Published: 30 May 2026
© After publication of the article, the authors shall own the right of signature. 2026.

The authors can use or share the published article under the Attribution-Non Commercial 4.0 International (CC BY-NC 4.0) license.