DocumentCode :
3035811
Title :
Thermal Faults Modeling Using a RC Model with an Application to Web Farms
Author :
Ferreira, Alexandre P. ; Mossé, Daniel ; Oh, Jae C.
Author_Institution :
Univ. of Pittsburgh, Pittsburgh
fYear :
2007
fDate :
4-6 July 2007
Firstpage :
113
Lastpage :
124
Abstract :
Today´s CPUs consume a significant amount of power and generate a high amount of heat, requiring an active cooling system to support reliable operations. In case of cooling system failures, these CPUs can reduce clock speed to prevent damage due to overheating. Unfortunately, when these CPUs are used in a real-time system, a clock control based on frequency-throttling can cause missed deadlines. In this paper, we first develop and validate a system-wide thermal model that can account for various thermal fault types such as failure of a CPU fan, faults in the case fan and air-conditioning malfunctions. Then we validate the thermal model through experimentation and measurements in AMD Linux boxes. Our soft real-time power-aware load-distribution algorithm for data centers incorporates a thermal model to minimize the number of missed deadlines that can be caused by thermal faults. We implemented the algorithm in a webserver farm simulator to test the efficacy of thermal-aware load-balancing. Our results show that the new algorithm helps keep CPU temperatures within the desired thermal envelope, even in the presence of thermal faults. When thermal faults occur, our algorithm improves the QoS, at the expense of higher energy consumption.
Keywords :
Internet; Linux; quality of service; resource allocation; AMD Linux boxes; CPU; QoS; Web farms; clock speed reduction; energy consumption; frequency-throttling; real-time power-aware load-distribution algorithm; thermal faults modeling; thermal-aware load-balancing; Clocks; Control systems; Cooling; Frequency; Linux; Power generation; Power system modeling; Power system reliability; Real time systems; Thermal loading;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
Real-Time Systems, 2007. ECRTS '07. 19th Euromicro Conference on
Conference_Location :
Pisa
ISSN :
1068-3070
Print_ISBN :
0-7695-2914-3
Type :
conf
DOI :
10.1109/ECRTS.2007.36
Filename :
4271686
Link To Document :
بازگشت