Title
CPPC-G: fault-tolerant applications on the grid
Abstract
The Grid community has made an important effort in developing middleware to provide different functionalities, such as resource discovery, resource management, job submission, execution monitoring. As part of this effort this paper addresses the design and implementation of an architecture (CPPC-G) based on services to manage the execution of fault tolerant applications on Grids. The CPPC (Controller/Precompiler for Portable Checkpointing) framework is used to insert checkpoint instrumentation into the application code. Designed services will be in charge of submission and monitoring of the execution of the application, management of checkpoint files and detection and automatic restart of failed executions.
Year
DOI
Venue
2007
10.1007/978-3-540-68111-3_90
PPAM
Keywords
Field
DocType
checkpoint file,checkpoint instrumentation,job submission,application code,execution monitoring,important effort,failed execution,resource discovery,fault-tolerant application,fault tolerant application,resource management,mpi,resource manager,middleware,fault tolerance,fault tolerant,grid computing,application management
Resource management,Middleware,Control theory,Architecture,Grid computing,Computer science,Fault tolerance,Grid,Operating system,Distributed computing
Conference
Volume
ISSN
ISBN
4967
0302-9743
3-540-68105-1
Citations 
PageRank 
References 
0
0.34
5
Authors
5
Name
Order
Citations
PageRank
Daniel Díaz1434.19
Xoán C. Pardo2225.92
María J. Martín317427.68
Patricia González47813.06
Gabriel Rodríguez59213.73