AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
PDF (3.4 MB)
Collect
Submit Manuscript AI Chat Paper
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article

Causal graph-aided reinforcement learning for HVAC energy consumption optimization

Shihang Gao1,2Xu Yang1,2( )Rang Tu3Jian Huang1,2Tao Zhang1,2Qing Li1,2
Key Laboratory of Knowledge Automation for Industrial Processes of Ministry of Education, School of Automation and Electrical Engineering, University of Science and Technology Beijing, 30 Xueyuan Road, Beijing 100083, China
Shunde Innovation School, University of Science and Technology Beijing, Foshan, Guangdong, China
School of Resources and Safety Engineering, University of Science and Technology Beijing, 30 Xueyuan Road, Beijing 100083, China
Show Author Information

Abstract

Scheduling the load of multiple chillers and distributing cooling capacity across multiple zones simultaneously remains challenging for large building HVAC (heating, ventilation and air-conditioning) systems. Meanwhile, the dynamic cooling load faced by HVAC systems places high demands on the generalization of energy optimization methods. To this end, a causal graph-aided reinforcement learning energy consumption optimization approach is proposed for large building HVAC systems. Firstly, an HVAC system with multiple chillers and multiple zones is analyzed, and the causal graph of the HVAC system is obtained. Secondly, a causal graph-aided network structure is designed to extract causal features between nodes in the causal graph. Causal structural information can help reinforcement learning achieve highly generalizable decisions and improve learning speed. Thirdly, an improved Soft Actor-Critic method is proposed with double experience replay mechanism and bias-based state augmentation to improve the resilience to external disturbances. Lastly, comparison experiments and ablation experiments based on three scenarios are conducted. Compared to other baseline reinforcement learning methods, the average reward of the proposed method increased by about 10%. Consequently, the proposed method achieves an average energy savings of 6% in different scenarios without sacrificing the indoor comfort. Theoretical analysis and experimental results confirm that the proposed method offers significant performance advantages in addressing dynamic supply-demand matching and energy consumption optimization for large-scale HVAC systems.

References

【1】
【1】
 
 
Building Simulation
Pages 1505-1522

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Gao S, Yang X, Tu R, et al. Causal graph-aided reinforcement learning for HVAC energy consumption optimization. Building Simulation, 2026, 19(6): 1505-1522. https://doi.org/10.1007/s12273-026-1451-y

3

Views

0

Downloads

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 11 February 2026
Revised: 27 March 2026
Accepted: 05 April 2026
Published: 20 June 2026
© Tsinghua University Press 2026