|
| This Article | ||
| ||
| Share | ||
| Bibliographic References | ||
| Add to: | ||
| | ||
| Search | ||
| ||
27th International Conference on Distributed Computing Systems (ICDCS '07)
Fault Tolerance in Multiprocessor Systems Via Application Cloning
Toronto, Canada
June 25-June 27
ISBN: 0-7695-2837-3
| ASCII Text | x | ||
| Philippe Bergheaud, Dinesh Subhraveti, Marc Vertes, "Fault Tolerance in Multiprocessor Systems Via Application Cloning," 2012 IEEE 32nd International Conference on Distributed Computing Systems, pp. 21, 27th International Conference on Distributed Computing Systems (ICDCS '07), 2007. | |||
| BibTex | x | ||
| @article{ 10.1109/ICDCS.2007.111, author = {Philippe Bergheaud and Dinesh Subhraveti and Marc Vertes}, title = {Fault Tolerance in Multiprocessor Systems Via Application Cloning}, journal ={2012 IEEE 32nd International Conference on Distributed Computing Systems}, volume = {0}, year = {2007}, isbn = {0-7695-2837-3}, pages = {21}, doi = {http://doi.ieeecomputersociety.org/10.1109/ICDCS.2007.111}, publisher = {IEEE Computer Society}, address = {Los Alamitos, CA, USA}, } | |||
| RefWorks Procite/RefMan/Endnote | x | ||
| TY - CONF JO - 2012 IEEE 32nd International Conference on Distributed Computing Systems TI - Fault Tolerance in Multiprocessor Systems Via Application Cloning SN - 0-7695-2837-3 SP EP A1 - Philippe Bergheaud, A1 - Dinesh Subhraveti, A1 - Marc Vertes, PY - 2007 VL - 0 JA - 2012 IEEE 32nd International Conference on Distributed Computing Systems ER - | |||
Record and Replay (RR) is a software based state replication solution designed to support recording and subsequent replay of the execution of unmodified applications running on multiprocessor systems for fault-tolerance. Multiple instances of the application are simultaneously executed in separate virtualized environments called Containers. Containers facilitate state replication between the application instances by resolving the resource conflicts and providing a uniform view of the underlying operating system across all clones. The virtualization layer that creates the container abstraction actively monitors the primary instance of the application and synchronizes its state with that of the clones by transferring the necessary information to enforce identical state among them. In particular, we address the replication of relevant operating system state, such as network state to preserve network connections across failures, and the state that results from nondeterministic interleaved accesses to shared memory in SMP systems. We have implemented RR?s state replication mechanisms in the Linux operating system by making novel use of existing features on the Intel and PowerPC architectures.
Citation:
Philippe Bergheaud, Dinesh Subhraveti, Marc Vertes, "Fault Tolerance in Multiprocessor Systems Via Application Cloning," icdcs, pp.21, 27th International Conference on Distributed Computing Systems (ICDCS '07), 2007
Usage of this product signifies your acceptance of the Terms of Use.
