direkt zum Inhalt springen

direkt zum Hauptnavigationsmenü

Sie sind hier

TU Berlin

Inhalt des Dokuments

Publikationen

VRM: A failure-aware Grid resource management system
Zitatschlüssel Burchard2008
Autor Lars-Olof Burchard and Hans-Ulrich Heiss and Barry Linnert and Jörg Schneider and Cesar A.F. De Rose
Seiten 215-226(12)
Jahr 2008
DOI doi:10.1504/IJHPCN.2008.022298
Journal International Journal of High Performance Computing and Networking
Jahrgang 5
Zusammenfassung For resource management in Grid environments, advance reservations turned out to be very useful and hence are supported by a variety of Grid toolkits. However, failure recovery for such systems has not yet received the attention it deserves. In this paper, we address the problem of remapping reservations to other resources, when the originally selected resource fails. Instead of dealing with jobs already running, which usually means checkpointing and migration, our focus is on jobs that are scheduled on the failed resource for a specific future period of time but not started yet. The most critical factor when solving this problem is the estimation of the downtime. We avoid the drawbacks of under- or over-estimating the downtime by a dynamic load-based approach that is evaluated by extensive simulations in a Grid environment and shows superior performance compared to estimation-based approaches.
Link zur Publikation Link zur Originalpublikation Download Bibtex Eintrag

Zusatzinformationen / Extras

Direktzugang

Schnellnavigation zur Seite über Nummerneingabe

Ansprechpartner

Jörg Schneider
+49 30 314-73388
Raum EN 357

Webseite