{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,21]],"date-time":"2026-03-21T19:24:12Z","timestamp":1774121052395,"version":"3.50.1"},"reference-count":38,"publisher":"Springer Science and Business Media LLC","issue":"10","license":[{"start":{"date-parts":[[2017,3,31]],"date-time":"2017-03-31T00:00:00Z","timestamp":1490918400000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"TU Wien (TUW)"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J Supercomput"],"published-print":{"date-parts":[[2017,10]]},"DOI":"10.1007\/s11227-017-2023-9","type":"journal-article","created":{"date-parts":[[2017,3,31]],"date-time":"2017-03-31T14:10:10Z","timestamp":1490969410000},"page":"4390-4406","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["Modelling parallel overhead from simple run-time records"],"prefix":"10.1007","volume":"73","author":[{"given":"Siegfried","family":"H\u00f6finger","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ernst","family":"Haunschmid","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2017,3,31]]},"reference":[{"key":"2023_CR1","doi-asserted-by":"crossref","unstructured":"Perarnau S, Thakur R, Iskra K, Raffenetti K, Cappello F, Gupta R, Beckman PH, Snir M, Hoffmann H, Schulz M, Rountree B (2015) Distributed monitoring and management of exascale systems in the argo project. In: Distributed applications and interoperable systems, vol 9038. Springer, New York, pp 173\u2013178","DOI":"10.1007\/978-3-319-19129-4_14"},{"key":"2023_CR2","unstructured":"Message Passing Interface Forum MPI: a message-passing interface standard version 3.1. High Performance Computing Center Stuttgart (HLRS), Stuttgart (2015)"},{"issue":"4","key":"2023_CR3","doi-asserted-by":"crossref","first-page":"65","DOI":"10.1145\/1498765.1498785","volume":"52","author":"S Williams","year":"2009","unstructured":"Williams S, Waterman A, Patterson D (2009) Roofline: an insightful visual performance model for multicore architectures. Commun ACM 52(4):65\u201376","journal-title":"Commun ACM"},{"key":"2023_CR4","doi-asserted-by":"crossref","unstructured":"Culler D, Karp R, Patterson D, Sahay A, Schauser KE, Santos E, Subramonian R, von Eicken T (1993) LogP: towards a realistic model of parallel computation. In: PPOPP \u201993 Proceedings of the Fourth ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming, ACM, pp 1\u201312","DOI":"10.1145\/155332.155333"},{"key":"2023_CR5","doi-asserted-by":"crossref","unstructured":"Kielmann T, Bal HE, Verstoep K (2000) Fast measurement of LogP parameters for message passing platforms. In: IPDPS \u201900 Proceedings of the 15 IPDPS 2000 Workshops on Parallel and Distributed Processing, Springer, UK, 2000, pp 1176\u20131183","DOI":"10.1007\/3-540-45591-4_162"},{"key":"2023_CR6","doi-asserted-by":"crossref","first-page":"202","DOI":"10.1006\/jpdc.2000.1677","volume":"61","author":"K Al-Tawil","year":"2001","unstructured":"Al-Tawil K, Moritz CA (2001) Performance modeling and evaluation of MPI. J Parallel Distrib Comput 61:202\u2013223","journal-title":"J Parallel Distrib Comput"},{"key":"2023_CR7","doi-asserted-by":"crossref","unstructured":"Alexandrov A, Ionescu MF, Schauser KE, Scheiman C (1995) LogGP: incorporating long messages into the LogP model\u2014one step closer towards a realistic model for parallel computation. In: Proceedings of the Seventh Annual ACM Symposium on Parallel Algorithms and Architectures, ACM, pp 95\u2013105","DOI":"10.1145\/215399.215427"},{"issue":"4","key":"2023_CR8","doi-asserted-by":"crossref","first-page":"404","DOI":"10.1109\/71.920589","volume":"12","author":"CA Moritz","year":"2001","unstructured":"Moritz CA, Frank MI (2001) LogGPC: modeling network contention in message-passing programs. IEEE Trans Parallel Distrib Syst 12(4):404\u2013415","journal-title":"IEEE Trans Parallel Distrib Syst"},{"key":"2023_CR9","doi-asserted-by":"crossref","unstructured":"Snavely A, Carrington L, Wolter N, Labarta J, Badia R, Purkayastha A (2002) A framework for performance modeling and prediction. In: SC \u201902: Proceedings of the 2002 ACM\/IEEE Conference on Supercomputing, IEEE","DOI":"10.1109\/SC.2002.10004"},{"issue":"4","key":"2023_CR10","doi-asserted-by":"crossref","first-page":"481","DOI":"10.1016\/j.simpat.2006.11.014","volume":"15","author":"L Adhianto","year":"2007","unstructured":"Adhianto L, Chapman B (2007) Performance modeling of communication and computation in hybrid MPI and OpenMP applications. Simul Model Pract Theory 15(4):481\u2013491","journal-title":"Simul Model Pract Theory"},{"issue":"11","key":"2023_CR11","doi-asserted-by":"crossref","first-page":"42","DOI":"10.1109\/MC.2009.372","volume":"42","author":"KJ Barker","year":"2009","unstructured":"Barker KJ, Davis K, Hoisie A, Kerbyson DJ, Lang M, Pakin S, Sancho JC (2009) Using performance modeling to design large-scale systems. Computer 42(11):42\u201349","journal-title":"Computer"},{"key":"2023_CR12","doi-asserted-by":"crossref","unstructured":"K\u00fchnemann M, Rauber T, R\u00fcnger G (2004) A source code analyzer for performance prediction. In: 18 International Parallel and Distributed Processing Symposium 2004, IEEE","DOI":"10.1109\/IPDPS.2004.1303333"},{"key":"2023_CR13","doi-asserted-by":"crossref","unstructured":"Pllana S, Benkner S, Xhafa F, Barolli L (2008) Hybrid performance modeling and prediction of large-scale computing systems. In: Complex, Intelligent and Software Intensive Systems, IEEE","DOI":"10.1109\/CISIS.2008.20"},{"key":"2023_CR14","doi-asserted-by":"publisher","unstructured":"Appelhans DJ, Manteuffel T, McCormick S, Ruge J (2016) A low-communication, parallel algorithm for solving PDEs based on range decomposition. Numer Linear Algebra Appl. doi:\n                        10.1002\/nla.2041","DOI":"10.1002\/nla.2041"},{"key":"2023_CR15","volume-title":"Introduction to parallel computing","author":"A Grama","year":"2003","unstructured":"Grama A, Gupta A, Karypis G, Kumar V (2003) Introduction to parallel computing, 2nd edn. Addison Wesley, Reading","edition":"2"},{"key":"2023_CR16","unstructured":"B\u00f6hme D, Hermanns MA, Wolf F (2012) In: Entwicklung und Evolution von Forschungssoftware, Rolduc, November 2011. Aachener Informatik-Berichte, Software Engineering, Shaker, pp 43\u201348"},{"key":"2023_CR17","unstructured":"Vetter J, Chambreau C (2014) mpiP: Lightweight, Scalable MPI Profiling, version 3.4.1. \n                        http:\/\/mpip.sourceforge.net"},{"key":"2023_CR18","unstructured":"MAP, release 6.0.6. Allinea Software Ltd. \n                        http:\/\/www.allinea.com\/products\/map\n                        \n                     (2016)"},{"key":"2023_CR19","unstructured":"Tuning and Analysis Utilities. \n                        https:\/\/www.cs.uoregon.edu\/research\/tau\/home.php\n                        \n                     (2013)"},{"key":"2023_CR20","unstructured":"Paradyn Tools Project. \n                        http:\/\/www.paradyn.org\n                        \n                     (2013)"},{"key":"2023_CR21","unstructured":"Petitet A, Whaley C, Dongarra J, Cleary A, Luszczek P (2012) HPL 2.1\u2014a portable implementation of the high-performance Linpack benchmark for distributed-memory computers. \n                        http:\/\/www.netlib.org\/benchmark\/hpl\/hpl-2.1.tar.gz"},{"key":"2023_CR22","doi-asserted-by":"crossref","first-page":"19","DOI":"10.1016\/j.softx.2015.06.001","volume":"12","author":"MJ Abraham","year":"2015","unstructured":"Abraham MJ, Murtola T, Schulz R, P\u00e1ll S, Smith JC, Hess B, Lindahl E (2015) GROMACS: high performance molecular simulations through multi-level parallelism from laptops to supercomputers. SoftwareX 12:19\u201325","journal-title":"SoftwareX"},{"issue":"16","key":"2023_CR23","doi-asserted-by":"crossref","first-page":"1668","DOI":"10.1002\/jcc.20290","volume":"26","author":"DA Case","year":"2005","unstructured":"Case DA, Cheatham-III TE, Darden T, Gohlke H, Luo R, Merz-Jr KM, Onufriev A, Simmerling C, Wang B, Woods RJ (2005) The Amber biomolecular simulation programs. J Comput Chem 26(16):1668\u20131688","journal-title":"J Comput Chem"},{"issue":"1","key":"2023_CR24","doi-asserted-by":"crossref","first-page":"558","DOI":"10.1103\/PhysRevB.47.558","volume":"47","author":"G Kresse","year":"1993","unstructured":"Kresse G, Hafner J (1993) Ab initio molecular dynamics for liquid metals. Phys Rev B 47(1):558\u2013561","journal-title":"Phys Rev B"},{"issue":"39","key":"2023_CR25","doi-asserted-by":"crossref","first-page":"395502","DOI":"10.1088\/0953-8984\/21\/39\/395502","volume":"21","author":"P Giannozzi","year":"2009","unstructured":"Giannozzi P, Baroni S, Bonini N, Calandra M, Car R, Cavazzoni C, Ceresoli D, Chiarotti GL, Cococcioni M, Dabo I, Corso AD, de Gironcoli S, Fabris S, Fratesi G, Gebauer R, Gerstmann U, Gougoussis C, Kokalj A, Lazzeri M, Martin-Samos L, Marzari N, Mauri F, Mazzarello R, Paolini S, Pasquarello A, Paulatto L, Sbraccia C, Scandolo S, Sclauzero G, Seitsonen AP, Smogunov A, Umari P, Wentzcovitch RM (2009) QUANTUM ESPRESSO: a modular and open-source software project for quantum simulations of materials. J Phys Condens Matter 21(39):395502","journal-title":"J Phys Condens Matter"},{"issue":"1","key":"2023_CR26","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1006\/jcph.1995.1039","volume":"117","author":"S Plimpton","year":"1995","unstructured":"Plimpton S (1995) Fast parallel algorithms for short-range molecular dynamics. J Comput Phys 117(1):1\u201319","journal-title":"J Comput Phys"},{"key":"2023_CR27","doi-asserted-by":"crossref","first-page":"163","DOI":"10.1007\/3-540-49164-3_16","volume":"1557","author":"S H\u00f6finger","year":"1999","unstructured":"H\u00f6finger S, Steinhauser O, Zinterhof P (1999) Performance analysis and derived parallelization strategy for a SCF program at the Hartree Fock Level. Lect Notes Comput Sci 1557:163\u2013172","journal-title":"Lect Notes Comput Sci"},{"issue":"1","key":"2023_CR28","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1142\/S0129183108011899","volume":"19","author":"R Mahajan","year":"2008","unstructured":"Mahajan R, Kranzlm\u00fcller D, Volkert J, Hansmann UH, H\u00f6finger S (2008) Detecting secondary bottlenecks in parallel quantum chemistry applications using MPI. Int J Mod Phys C 19(1):1\u201313","journal-title":"Int J Mod Phys C"},{"key":"2023_CR29","doi-asserted-by":"publisher","first-page":"47","DOI":"10.1016\/j.entcs.2007.10.020","volume":"198","author":"J Ezekiel","year":"2008","unstructured":"Ezekiel J, L\u00fcttgen G (2008) Measuring and evaluating parallel state-space exploration algorithms. Electron Notes Theor Comput Sci 198:47\u201361. doi:\n                        10.1016\/j.entcs.2007.10.020","journal-title":"Electron Notes Theor Comput Sci"},{"key":"2023_CR30","first-page":"105","volume":"P\u2013199","author":"E Belikov","year":"2012","unstructured":"Belikov E, Loidl HW, Michaelson G, Trinder P (2012) Architecture-aware cost modelling for parallel performance portability. Lect Notes Inf P\u2013199:105\u2013120","journal-title":"Lect Notes Inf"},{"key":"2023_CR31","doi-asserted-by":"crossref","first-page":"331","DOI":"10.1007\/11846802_46","volume":"4192","author":"D Doerfler","year":"2006","unstructured":"Doerfler D, Brightwell R (2006) Measuring MPI Send and Receive Overhead and Application Availability in High Performance Network Interfaces. Lect Notes Comput Sci 4192:331\u2013338","journal-title":"Lect Notes Comput Sci"},{"key":"2023_CR32","doi-asserted-by":"crossref","unstructured":"Amdahl GM (1967) Validity of the single processor approach to achieving large scale computing capabilities. In: Proceedings AFIPS \u201967 (Spring) Joint Computer Conference, ACM, New York, pp 483\u2013485","DOI":"10.1145\/1465482.1465560"},{"issue":"7","key":"2023_CR33","doi-asserted-by":"crossref","first-page":"33","DOI":"10.1109\/MC.2008.209","volume":"41","author":"MD Hill","year":"2008","unstructured":"Hill MD, Marty MR (2008) Amdahl\u2019s law in the multicore era. Computer 41(7):33\u201338","journal-title":"Computer"},{"issue":"1","key":"2023_CR34","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/j.parco.2013.11.001","volume":"40","author":"L Yavits","year":"2014","unstructured":"Yavits L, Morad A, Ginosar R (2014) The effect of communication and synchronization on Amdahls law in multicore systems. Parallel Comput 40(1):1\u201316","journal-title":"Parallel Comput"},{"issue":"2","key":"2023_CR35","doi-asserted-by":"crossref","first-page":"183","DOI":"10.1016\/j.jpdc.2009.05.002","volume":"70","author":"XH Sun","year":"2010","unstructured":"Sun XH, Chen Y (2010) Reevaluating Amdahl\u2019s law in the multicore era. J Parallel Distrib Comput 70(2):183\u2013188","journal-title":"J Parallel Distrib Comput"},{"key":"2023_CR36","unstructured":"Vienna Scientific Cluster, generation 3. \n                        http:\/\/vsc.ac.at\/systems\/vsc-3"},{"key":"2023_CR37","unstructured":"Green Revolution Cooling. \n                        http:\/\/www.grcooling.com"},{"key":"2023_CR38","unstructured":"Br\u00f6ker HB, Campbell J, Cunningham R, Denholm D, Elber G, Fearick R, Grammes C, Hart L, Heckingu L, Juh\u00e1sz P, Koenig T, Kotz D, Kubaitis E, Lang R, Lecomte T, Lehmann A, Mai A, M\u00e4rkisch B, Merritt EA, Mikul\u00edk P, Steger C, Takeno S, Tkacik T, der Woude JV, Zandt JRV, Woo A, Zellner J (2014) Gnuplot 4.6\u2014an interactive plotting program. \n                        http:\/\/sourceforge.net\/projects\/gnuplot"}],"container-title":["The Journal of Supercomputing"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s11227-017-2023-9\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-017-2023-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11227-017-2023-9.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2017,9,15]],"date-time":"2017-09-15T07:43:23Z","timestamp":1505461403000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11227-017-2023-9"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,3,31]]},"references-count":38,"journal-issue":{"issue":"10","published-print":{"date-parts":[[2017,10]]}},"alternative-id":["2023"],"URL":"https:\/\/doi.org\/10.1007\/s11227-017-2023-9","relation":{},"ISSN":["0920-8542","1573-0484"],"issn-type":[{"value":"0920-8542","type":"print"},{"value":"1573-0484","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,3,31]]}}}