Green Data Center Cooling Control via Physics-guided Safe Reinforcement Learning

强化学习 中心(范畴论) 钢筋 数据中心 控制(管理) 物理 计算机科学 工程类 人工智能 操作系统 结构工程 化学 结晶学
作者
Ruihang Wang,Zhiwei Cao,Xin Zhou,Yonggang Wen,Rui Tan
出处
期刊:ACM Transactions on Cyber-Physical Systems [Association for Computing Machinery]
卷期号:8 (2): 1-26 被引量:13
标识
DOI:10.1145/3582577
摘要

Deep reinforcement learning (DRL) has shown good performance in tackling Markov decision process (MDP) problems. As DRL optimizes a long-term reward, it is a promising approach to improving the energy efficiency of data-center cooling. However, enforcement of thermal safety constraints during DRL’s state exploration is a main challenge. The widely adopted reward-shaping approach adds negative reward when the exploratory action results in unsafety. Thus, it needs to experience sufficient unsafe states before it learns how to prevent unsafety. In this article, we propose a safety-aware DRL framework for data-center cooling control. It applies offline imitation learning and online post-hoc rectification to holistically prevent thermal unsafety during online DRL. In particular, the post-hoc rectification searches for the minimum modification to the DRL-recommended action such that the rectified action will not result in unsafety. The rectification is designed based on a thermal state transition model that is fitted using historical safe operation traces and able to extrapolate the transitions to unsafe states explored by DRL. Extensive evaluation for chilled water and direct expansion-cooled data centers in two climate conditions show that our approach saves 18% to 26.6% of total data-center power compared with conventional control and reduces safety violations by 94.5% to 99% compared with reward shaping. We also extend the proposed framework to address data centers with non-uniform temperature distributions for detailed safety considerations. The evaluation shows that our approach saves 14% power usage compared with the PID control while addressing safety compliance during the training.
最长约 10秒,即可获得该文献文件

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
1秒前
2秒前
愤怒的微笑完成签到,获得积分10
3秒前
3秒前
3秒前
豆子完成签到 ,获得积分10
3秒前
arniu2008应助整齐画板采纳,获得20
3秒前
Vicente完成签到,获得积分10
3秒前
4秒前
xuejingling应助yunyun采纳,获得10
4秒前
SciGPT应助流年采纳,获得10
6秒前
6秒前
7秒前
7秒前
8秒前
8秒前
maguodrgon发布了新的文献求助10
8秒前
9秒前
zhuzhuyang发布了新的文献求助10
10秒前
11秒前
Nole应助乐观的箭头采纳,获得50
11秒前
小星星发布了新的文献求助10
12秒前
陈欣瑶发布了新的文献求助10
15秒前
爱放屁的美娇娘完成签到 ,获得积分10
15秒前
16秒前
牧青发布了新的文献求助10
17秒前
麻师加药发布了新的文献求助10
17秒前
印第安老斑鸠应助zm采纳,获得10
17秒前
17秒前
19秒前
20秒前
小马甲应助王聪颖采纳,获得10
20秒前
21秒前
21秒前
zzdj发布了新的文献求助10
22秒前
贪玩的寻冬完成签到,获得积分10
22秒前
落寞的善若完成签到,获得积分10
22秒前
24秒前
zx发布了新的文献求助10
26秒前
27秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
Autoparametric Resonance in Mechanical Systems 1000
Effects of Two Weeks of Red Light Therapy on Choroidal Thickness and Axial Length in Young Adults 700
Cosmos as Art Object: Studies in Plato's Timaeus and Other Dialogues 600
Management and the Arts 510
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
Rutherford's Vascular Surgery and Endovascular Therapy, 2‑Volume Set, 11th Edition 480
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7665540
求助须知:如何正确求助?哪些是违规求助? 9235468
关于积分的说明 19873813
捐赠科研通 7234686
什么是DOI,文献DOI怎么找? 3283560
关于科研通互助平台的介绍 2442341
邀请新用户注册赠送积分活动 2284608