{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,10]],"date-time":"2026-06-10T15:55:28Z","timestamp":1781106928577,"version":"3.54.1"},"reference-count":24,"publisher":"IGI Global Scientific Publishing","issue":"2","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2017,4]]},"abstract":"<jats:p>For the distributed computing system, excessive or deficient checkpointing operations would result in severe performance degradation. To minimize the expected computation execution of the long-running application with a general failure distribution, the optimal equidistant checkpoint interval for fault tolerant performance optimization is analyzed and derived in this paper. More precisely, the optimal checkpointing period to determine the proper checkpoint sequence is proposed, and the derivation of the expected effective rate of the defined computation cycle is introduced. Corresponding to the maximal expected effective rate, the constraint of the optimal checkpoint sequence can be obtained. From the constraint of optimality, the optimal equidistant checkpoint interval can be obtained according to the minimal fault tolerant overhead ratio. By the numerical results, the proposal is practical to determine a proper equidistant checkpoint interval for fault tolerant performance optimization.<\/jats:p>","DOI":"10.4018\/ijapuc.2017040103","type":"journal-article","created":{"date-parts":[[2017,5,31]],"date-time":"2017-05-31T13:52:07Z","timestamp":1496238727000},"page":"45-54","source":"Crossref","is-referenced-by-count":0,"title":["The Optimal Checkpoint Interval for the Long-Running Application"],"prefix":"10.4018","volume":"9","author":[{"given":"Yongning","family":"Zhai","sequence":"first","affiliation":[{"name":"Jiangsu Automation Research Institute, Lianyungang, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Weiwei","family":"Li","sequence":"additional","affiliation":[{"name":"Jiangsu Automation Research Institute, Lianyungang, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"2432","reference":[{"key":"IJAPUC.2017040103-0","doi-asserted-by":"publisher","DOI":"10.1504\/IJCNDS.2014.062226"},{"key":"IJAPUC.2017040103-1","doi-asserted-by":"publisher","DOI":"10.1016\/S0166-5316(01)00037-2"},{"key":"IJAPUC.2017040103-2","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.1975.6312824"},{"key":"IJAPUC.2017040103-3","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2004.11.016"},{"key":"IJAPUC.2017040103-4","doi-asserted-by":"crossref","unstructured":"Duda A. (1983, June). The Effects of Checkpointing on Program Execution time. Information Processing Letters, 16, 221\u2013229.","DOI":"10.1016\/0020-0190(83)90093-5"},{"key":"IJAPUC.2017040103-5","doi-asserted-by":"publisher","DOI":"10.1145\/568522.568525"},{"key":"IJAPUC.2017040103-6","doi-asserted-by":"publisher","DOI":"10.1007\/s10723-014-9297-4"},{"key":"IJAPUC.2017040103-7","doi-asserted-by":"publisher","DOI":"10.1145\/2687357.2687364"},{"key":"IJAPUC.2017040103-8","doi-asserted-by":"publisher","DOI":"10.1109\/12.2197"},{"key":"IJAPUC.2017040103-9","doi-asserted-by":"publisher","DOI":"10.1109\/12.936236"},{"key":"IJAPUC.2017040103-10","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTR.2007.4629264"},{"key":"IJAPUC.2017040103-11","doi-asserted-by":"crossref","unstructured":"Mendizabal, O. M., Marandi, P. J., & Dotti, F. L. (2014). Checkpointing in parallel state-machine replication. Proceedings of the18th International Conference on Principles of Distributed Systems (pp. 123-138).","DOI":"10.1007\/978-3-319-14472-6_9"},{"issue":"3","key":"IJAPUC.2017040103-12","doi-asserted-by":"crossref","first-page":"131","DOI":"10.3233\/JHS-140492","article-title":"Lightweight coordinated checkpointing in cloud computing.","volume":"20","author":"B.Meroufel","year":"2014","journal-title":"Journal of High Speed Networks"},{"key":"IJAPUC.2017040103-13","doi-asserted-by":"publisher","DOI":"10.1016\/j.camwa.2005.11.002"},{"key":"IJAPUC.2017040103-14","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-68129-8_10"},{"key":"IJAPUC.2017040103-15","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2009.06.058"},{"key":"IJAPUC.2017040103-16","doi-asserted-by":"publisher","DOI":"10.1109\/ICPADS.2005.22"},{"key":"IJAPUC.2017040103-17","doi-asserted-by":"publisher","DOI":"10.1109\/PRDC.2004.1276566"},{"key":"IJAPUC.2017040103-18","doi-asserted-by":"publisher","DOI":"10.1016\/j.peva.2008.11.003"},{"key":"IJAPUC.2017040103-19","unstructured":"Treaster, M. \u2018A survey of fault-tolerance and fault-recovery techniques in parallel systems. Technical Report cs.DC\/0501002, ACM Computing Research Repository, January 2005"},{"key":"IJAPUC.2017040103-20","doi-asserted-by":"publisher","DOI":"10.1109\/12.609281"},{"key":"IJAPUC.2017040103-21","doi-asserted-by":"publisher","DOI":"10.1145\/361147.361115"},{"key":"IJAPUC.2017040103-22","doi-asserted-by":"publisher","DOI":"10.1109\/12.641939"},{"key":"IJAPUC.2017040103-23","doi-asserted-by":"publisher","DOI":"10.1109\/12.620479"}],"container-title":["International Journal of Advanced Pervasive and Ubiquitous Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.igi-global.com\/viewtitle.aspx?TitleId=182526","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,5,5]],"date-time":"2022-05-05T19:53:51Z","timestamp":1651780431000},"score":1,"resource":{"primary":{"URL":"http:\/\/services.igi-global.com\/resolvedoi\/resolve.aspx?doi=10.4018\/IJAPUC.2017040103"}},"subtitle":[""],"short-title":[],"issued":{"date-parts":[[2017,4]]},"references-count":24,"journal-issue":{"issue":"2"},"URL":"https:\/\/doi.org\/10.4018\/ijapuc.2017040103","relation":{},"ISSN":["1937-965X","1937-9668"],"issn-type":[{"value":"1937-965X","type":"print"},{"value":"1937-9668","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,4]]}}}