• DocumentCode
    3234607
  • Title

    VRM: a failure-aware grid resource management system

  • Author

    Burchard, Lars-Olof ; De Rose, César A F ; Heiss, Hans-Ulrich ; Linnert, Barry ; Schneider, Jörg

  • Author_Institution
    Technische Univ., Berlin, Germany
  • fYear
    2005
  • fDate
    24-27 Oct. 2005
  • Firstpage
    218
  • Lastpage
    225
  • Abstract
    For resource management in grid environments, advance reservations turned out to be very useful and hence are supported by a variety of grid toolkits. However, failure recovery for such systems has not yet received the attention it deserves. In this paper, we address the problem of remapping reservations to other resources, when the originally selected resource fails. Instead of dealing with jobs already running, which usually means checkpointing and migration, our focus is on jobs that are scheduled on the failed resource for a specific future period of time but not started yet. The most critical factor when solving this problem is the estimation of the downtime. We avoid the drawbacks of under- or overestimating the downtime by a dynamic load-based approach that is evaluated by extensive simulations in a grid environment and shows superior performance compared to estimation-based approaches.
  • Keywords
    checkpointing; grid computing; processor scheduling; resource allocation; failure recovery; failure-aware grid resource management system; grid environment; grid toolkit; remapping reservation; system checkpointing; system migration; virtual resource manager; Application software; Checkpointing; Computer networks; Context-aware services; Data processing; Data visualization; Environmental management; Grid computing; Quality of service; Resource management;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Computer Architecture and High Performance Computing, 2005. SBAC-PAD 2005. 17th International Symposium on
  • ISSN
    1550-6533
  • Print_ISBN
    0-7695-2446-X
  • Type

    conf

  • DOI
    10.1109/CAHPC.2005.41
  • Filename
    1592576