DOE OSTI · 3484826
Designing reinforcement learning algorithms for building HVAC control: From experimental observation to simulation comparisons
Abstract
Advanced supervisory-level control with reinforcement learning (RL) is regarded as a promising solution for HVAC systems to minimize energy consumption while maintaining thermal comfort and indoor air quality. However, most RL applications were conducted in the simulation environment rather than real-world HVAC systems. This paper developed a value-based RL controller termed Deep Q-Network (DQN) for a typical central HVAC system and evaluated its performance in a building test facility. By comparing DQN with a rule-based controller, the study not only demonstrated the cases where DQN could properly maintain indoor comfort but also discussed possible reasons why DQN failed in some other situations. Recognizing the limitations of value-based RL algorithms from the experimental tests, a simulation study was conducted to compare DQN with an alternative RL approach, an actor–critic algorithm termed Deep Deterministic Policy Gradient (DDPG). In scenarios with a relatively large action space, DDPG outperformed DQN by requiring fewer computational resources and achieving better thermal comfort, lower energy consumption, and more stable control actions. The findings suggest that the ability of DDPG to handle continuous control variables more effectively allows for faster convergence in training and more precise control in practice, which enhances the overall efficiency and reliability of the HVAC system.
Keep this discovery
Explore connections, maps & timelines
Guo, Fangzhou, Ham, Sang woo, Kim, Donghun, Kim, Sun Ho, Moon, Hyeun Jun. 2025-07-01. Designing reinforcement learning algorithms for building HVAC control: From experimental observation to simulation comparisons. https://doi.org/10.1016/j.applthermaleng.2025.126106
Cite the original work for its findings. Save a collection to share your selection of sources.