DocumentCode :
2227989
Title :
Disk infant mortality in large storage systems
Author :
Xin, Qin ; Schwarz, Thomas J E ; Miller, Ethan L.
Author_Institution :
Storage Syst. Res. Center, California Univ., Santa Cruz, CA, USA
fYear :
2005
fDate :
27-29 Sept. 2005
Firstpage :
125
Lastpage :
134
Abstract :
As disk drives have dropped in price relative to tape, the desire for the convenience and speed of online access to large data repositories has, led to the deployment of petabyte-scale disk farms with thousands of disks. Unfortunately, the very large size of these repositories renders them vulnerable to previously rare failure modes such as multiple, unrelated disk failures leading to data loss. While some business models, such as free email servers, may be able to tolerate some occurrence of data loss, others, including premium online services and storage of simulation results at a national laboratory, cannot. This paper describes the effect of infant mortality on long-term failure rates of systems that must preserve their data for decades. Our failure models incorporate the well-known "bathtub curve," which reflects the higher failure rates of new disk drives, a lower, constant failure rate during the remainder of the design life span, and increased failure rates as components wear out. Large systems are vulnerable to the "cohort effect" that occurs when many disks are simultaneously replaced by new disks. Our more accurate disk models and simulations have yielded predictions of system lifetimes that are more pessimistic than existing models that assume a constant disk failure rate. Thus, larger system scale requires designers to take disk infant mortality into account.
Keywords :
RAID; disc drives; storage management; bathtub curve; cohort effect; data repository; disk drive; disk infant mortality; online access speed; petabyte-scale disk farm; storage system; Costs; Data engineering; Disk drives; Electronic mail; Laboratories; Large-scale systems; Libraries; Predictive models; Redundancy; Satellites;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
Modeling, Analysis, and Simulation of Computer and Telecommunication Systems, 2005. 13th IEEE International Symposium on
ISSN :
1526-7539
Print_ISBN :
0-7695-2458-3
Type :
conf
DOI :
10.1109/MASCOTS.2005.27
Filename :
1521125
Link To Document :
بازگشت