File(s) not publicly available
Checkpointing schemes for grid workflow systems
One of the major challenges in wide use of Grid workflow systems is fault tolerance and avoidance. Checkpointing schemes provide a way of fault detection and recovery. In our research, we focus on performance optimization of checkpointing schemes and DVS (Dynamic Voltage Scaling) for Grid workflow systems. We propose offline checkpointing schemes with DVS and online adaptive checkpointing schemes that dynamically adjust the checkpointing intervals by using store-checkpoints (SCPs) and compare-checkpoints (CCPs). When combined with DVS, offline adaptive checkpointing schemes not only are fault tolerant but also lead to reduce average execution time of tasks. These schemes can efficiently utilize comparison and storage operations and significantly improve the performance. Further, these schemes can calculate the optimal numbers of checkpoints by which minimize the mean execution time. We also expand the online adaptive checkpointing schemes from single-task execution scenarios to multi-task execution scenarios. Simulation results show these online schemes outstandingly increase the likelihood of timely task completion when faults occur.
Funding
Category 1 - Australian Competitive Grants (this includes ARC, NHMRC)
History
Volume
20Issue
15Start Page
1773End Page
1790Number of Pages
18ISSN
1532-0626Location
Chichester, UKPublisher
John Wiley and SonsPublisher DOI
Full Text URL
Language
en-ausPeer Reviewed
- Yes
Open Access
- No
External Author Affiliations
Faculty of Business and Informatics; Not affiliated to a Research Institute; Xiamen da xue;Era Eligible
- Yes