2010 · 29 citations · 27 references
Cluster ComputingEngineeringComputer ArchitectureFault ToleranceHardware SecurityHardware VirtualizationSystems EngineeringParallel ComputingHypervisor-based Fault ToleranceMemory AccessesComputer EngineeringVirtualization SupportComputer ScienceVirtual MemoryStorage VirtualizationVirtualization TechnologyProgram AnalysisCloud ComputingParallel ProgrammingVirtual Machine
Hypervisor-based fault tolerance (HBFT), a checkpoint-recovery mechanism, is an emerging approach to sustaining mission-critical applications. Based on virtualization technology, HBFT provides an economic and transparent solution. However, the advantages currently come at the cost of substantial overhead during failure-free, especially for memory intensive applications. This paper presents an in-depth examination of HBFT and options to improve its performance. Based on the behavior of memory accesses among checkpointing epochs, we introduce two optimizations, read fault reduction and write fault prediction, for the memory tracking mechanism. These two optimizations improve the mechanism by 31.1% and 21.4% respectively for some application. Then, we present software-superpage which efficiently maps large memory regions between virtual machines (VM). By the above optimizations, HBFT is improved by a factor of 1.4 to 2.2 and it achieves a performance which is about 60% of that of the native VM.
27
Xen and the art of virtualization
Paul Barham, Boris Dragovic, Keir Fraser et al. · ACM SIGOPS Operating Systems Review · 2003 · 5.9K citations
Live migration of virtual machines
Christopher J. Clark, Keir Fraser, Steven Hand et al. · Networked Systems Design and Implementation · 2005 · 2.6K citations
Xen and the art of virtualization
Paul Barham, Boris Dragovic, Keir Fraser et al. · 2003 · 1.5K citations