DocumentCode
1671683
Title
Disk scrubbing in large archival storage systems
Author
Schwarz, T.J.E. ; Xin, Qin ; Miller, Ethan L. ; Long, Darrell D E ; Hospodor, Andy ; Ng, Spencer
Author_Institution
Comput. Eng. Dept., Santa Clara Univ., CA, USA
fYear
2004
Firstpage
409
Lastpage
418
Abstract
Large archival storage systems experience long periods of idleness broken up by rare data accesses. In such systems, disks may remain powered off for long periods of time. These systems can lose data for a variety of reasons, including failures at both the device level and the block level. To deal with these failures, we must detect them early enough to be able to use the redundancy built into the storage system. We propose a process called "disk scrubbing" in a system in which drives are periodically accessed to detect drive failure. By scrubbing all of the data stored on all of the disks, we can detect block failures and compensate for them by rebuilding the affected blocks. Our research shows how the scheduling of disk scrubbing affects overall system reliability, and that "opportunistic" scrubbing, in which the system scrubs disks only when they are powered on for other reasons, performs very well without the need to power on disks solely to check them.
Keywords
disc storage; scheduling; block failure detection; disk scrubbing; large archival storage systems; opportunistic scrubbing; scheduling; system reliability; Analytical models; Computational modeling; Contracts; Data engineering; Laboratories; Large-scale systems; Mirrors; Redundancy; Reliability; Telecommunication computing;
fLanguage
English
Publisher
ieee
Conference_Titel
Modeling, Analysis, and Simulation of Computer and Telecommunications Systems, 2004. (MASCOTS 2004). Proceedings. The IEEE Computer Society's 12th Annual International Symposium on
ISSN
1526-7539
Print_ISBN
0-7695-2251-3
Type
conf
DOI
10.1109/MASCOT.2004.1348296
Filename
1348296
Link To Document