{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,11]],"date-time":"2026-07-11T15:43:17Z","timestamp":1783784597351,"version":"3.55.0"},"reference-count":158,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2017,6,29]],"date-time":"2017-06-29T00:00:00Z","timestamp":1498694400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001602","name":"Science Foundation Ireland","doi-asserted-by":"crossref","award":["14\/IA\/2474"],"award-info":[{"award-number":["14\/IA\/2474"]}],"id":[{"id":"10.13039\/501100001602","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Programme for Research in Third Level Institutions (PRTLI) Cycle 5"},{"name":"EU under the COST Program Action IC1305: Network for Sustainable Ultrascale Computing"},{"name":"Structured PhD in Simulation Science"},{"DOI":"10.13039\/501100008530","name":"European Regional Development Fund","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100008530","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2018,5,31]]},"abstract":"<jats:p>Power and energy efficiency are now critical concerns in extreme-scale high-performance scientific computing. Many extreme-scale computing systems today (for example: Top500) have tight integration of multicore CPU processors and accelerators (mix of Graphical Processing Units, Intel Xeon Phis, or Field Programmable Gate Arrays) empowering them to provide not just unprecedented computational power but also to address these concerns. However, such integration renders these systems highly heterogeneous and hierarchical, thereby necessitating design of novel performance, power, and energy models to accurately capture these inherent characteristics.<\/jats:p>\n          <jats:p>There are now several extensive research efforts focusing exclusively on power and energy efficiency models and techniques for the processors composing these extreme-scale computing systems. This article synthesizes these research efforts with absolute concentration on predictive power and energy models and prime emphasis on node architecture. Through this survey, we also intend to highlight the shortcomings of these models to correctly and comprehensively predict the power and energy consumptions by taking into account the hierarchical and heterogeneous nature of these tightly integrated high-performance computing systems.<\/jats:p>","DOI":"10.1145\/3078811","type":"journal-article","created":{"date-parts":[[2017,6,30]],"date-time":"2017-06-30T12:36:19Z","timestamp":1498826179000},"page":"1-38","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":69,"title":["A Survey of Power and Energy Predictive Models in HPC Systems and Applications"],"prefix":"10.1145","volume":"50","author":[{"given":"Kenneth","family":"O\u2019brien","sequence":"first","affiliation":[{"name":"University College Dublin, Ireland"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ilia","family":"Pietri","sequence":"additional","affiliation":[{"name":"University of Manchester, Oxford Road, Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ravi","family":"Reddy","sequence":"additional","affiliation":[{"name":"University College Dublin, Belfield, Ireland"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Alexey","family":"Lastovetsky","sequence":"additional","affiliation":[{"name":"University College Dublin, Belfield, Ireland"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Rizos","family":"Sakellariou","sequence":"additional","affiliation":[{"name":"University of Manchester, Oxford Road, Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2017,6,29]]},"reference":[{"key":"e_1_2_2_1_1","unstructured":"2015. Compute Unified Device Architecture. Retrieved from http:\/\/www.nvidia.com\/object\/cuda_home_new.html.  2015. Compute Unified Device Architecture. Retrieved from http:\/\/www.nvidia.com\/object\/cuda_home_new.html."},{"key":"e_1_2_2_2_1","unstructured":"2015. Energy Efficiency in Data Centers. Retrieved from http:\/\/www.google.co.in\/about\/datacenters\/efficiency\/.  2015. Energy Efficiency in Data Centers. Retrieved from http:\/\/www.google.co.in\/about\/datacenters\/efficiency\/."},{"key":"e_1_2_2_3_1","unstructured":"ACPI. 2015. Advanced Configuration and Power Interface Specification Version 6.0. Retrieved from http:\/\/www.uefi.org\/sites\/default\/files\/resources\/ACPI_6.0.pdf.  ACPI. 2015. Advanced Configuration and Power Interface Specification Version 6.0. Retrieved from http:\/\/www.uefi.org\/sites\/default\/files\/resources\/ACPI_6.0.pdf."},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-16214-0_6"},{"key":"e_1_2_2_5_1","unstructured":"AMDHT. 2001. HyperTransport. Retrieved from https:\/\/en.wikipedia.org\/wiki\/HyperTransport.  AMDHT. 2001. HyperTransport. Retrieved from https:\/\/en.wikipedia.org\/wiki\/HyperTransport."},{"key":"e_1_2_2_6_1","unstructured":"AVS. 2015. Adaptive voltage scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Adaptive_voltage_scaling.  AVS. 2015. Adaptive voltage scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Adaptive_voltage_scaling."},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2007.443"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/2318716.2318718"},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2208828.2208840"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/SECON.2010.5453824"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/566726.566736"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2012.08.003"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1155\/2012\/752910"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1810085.1810108"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2012.97"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPT.2010.5681761"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1020958815308"},{"key":"e_1_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2011.47"},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/ReConFig.2012.6416735"},{"key":"e_1_2_2_20_1","unstructured":"BLAS. 2015. BLAS (Basic Linear Algebra Subprograms). Retrieved from http:\/\/www.netlib.org\/blas\/.  BLAS. 2015. BLAS (Basic Linear Algebra Subprograms). Retrieved from http:\/\/www.netlib.org\/blas\/."},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-012-0224-2"},{"key":"e_1_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.5555\/645764.666798"},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/339647.339657"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1456190.1456199"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/GRID.2005.1542730"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/IGCC.2011.6008582"},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2014.54"},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2013.77"},{"key":"e_1_2_2_29_1","unstructured":"CPUFreq. 2015. CPU frequency scaling - ondemand Governor. Retrieved from https:\/\/wiki.archlinux.org\/index.php\/CPU_frequency_scaling.  CPUFreq. 2015. CPU frequency scaling - ondemand Governor. Retrieved from https:\/\/wiki.archlinux.org\/index.php\/CPU_frequency_scaling."},{"key":"e_1_2_2_30_1","unstructured":"CUPTI. 2015. CUDA Profiling Tools Interface. Retrieved from https:\/\/developer.nvidia.com\/cuda-profiling-tools-interface.  CUPTI. 2015. CUDA Profiling Tools Interface. Retrieved from https:\/\/developer.nvidia.com\/cuda-profiling-tools-interface."},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2014.2315629"},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/1840845.1840883"},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/COMST.2015.2481183"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2013.32"},{"key":"e_1_2_2_35_1","unstructured":"Dennard. 1974. Dennard scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Dennard_scaling.  Dennard. 1974. Dennard scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Dennard_scaling."},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/2488551.2488565"},{"key":"e_1_2_2_37_1","unstructured":"DOE. 2010. The Opportunities and Challenges of Exascale Computing. Retrieved from http:\/\/science.energy.gov\/\/media\/ascr\/\/pdf\/reports\/Exascale_subcommittee_report.pdf.  DOE. 2010. The Opportunities and Challenges of Exascale Computing. Retrieved from http:\/\/science.energy.gov\/\/media\/ascr\/\/pdf\/reports\/Exascale_subcommittee_report.pdf."},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/CGC.2012.113"},{"key":"e_1_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2013.07.004"},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPT.2011.6132672"},{"key":"e_1_2_2_41_1","unstructured":"DVFS. 2015. Dynamic voltage scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Dynamic_voltage_scaling.  DVFS. 2015. Dynamic voltage scaling. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Dynamic_voltage_scaling."},{"key":"e_1_2_2_42_1","volume-title":"Proceedings of the Workshop on Modeling, Benchmarking, and Simulation. 70--77","author":"Economou Dimitris","year":"2006","unstructured":"Dimitris Economou , Suzanne Rivoire , Christos Kozyrakis , and Partha Ranganathan . 2006 . Full-system power analysis and modeling for server environments . In Proceedings of the Workshop on Modeling, Benchmarking, and Simulation. 70--77 . Dimitris Economou, Suzanne Rivoire, Christos Kozyrakis, and Partha Ranganathan. 2006. Full-system power analysis and modeling for server environments. In Proceedings of the Workshop on Modeling, Benchmarking, and Simulation. 70--77."},{"key":"e_1_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/2000064.2000108"},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/1250662.1250665"},{"key":"e_1_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/1964179.1964192"},{"key":"e_1_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1145\/2400682.2400684"},{"key":"e_1_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/2503210.2503303"},{"key":"e_1_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2009.76"},{"key":"e_1_2_2_49_1","unstructured":"GK210. 2014. An Overview of Kepler GK110 and GK210 Architecture. Retrieved from http:\/\/international.download.nvidia.com\/pdf\/kepler\/NVIDIA-Kepler-GK110-GK210-Architecture-Whitepaper.pdf.  GK210. 2014. An Overview of Kepler GK110 and GK210 Architecture. Retrieved from http:\/\/international.download.nvidia.com\/pdf\/kepler\/NVIDIA-Kepler-GK110-GK210-Architecture-Whitepaper.pdf."},{"key":"e_1_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/GREENCOMP.2010.5598313"},{"key":"e_1_2_2_51_1","volume-title":"The Green500 List--November","year":"2015","unstructured":"Green500. 2015. The Green500 List--November 2015 . Retrieved from http:\/\/www.green500.org\/news\/green500-list-november-2015. Green500. 2015. The Green500 List--November 2015. Retrieved from http:\/\/www.green500.org\/news\/green500-list-november-2015."},{"key":"e_1_2_2_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/PDP.2014.112"},{"key":"e_1_2_2_53_1","unstructured":"GSLMLR. 2015. Multi-parameter fitting. Retrieved from https:\/\/www.gnu.org\/software\/gsl\/manual\/html_node\/Multi_002dparameter-fit ting.html.  GSLMLR. 2015. Multi-parameter fitting. Retrieved from https:\/\/www.gnu.org\/software\/gsl\/manual\/html_node\/Multi_002dparameter-fit ting.html."},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.5555\/874076.876487"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1145\/42411.42415"},{"key":"e_1_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/AHS.2009.55"},{"key":"e_1_2_2_57_1","unstructured":"Haswell. 2013. Performance Monitoring Counters for Haswell Architecture. Retrieved from https:\/\/code.google.com\/p\/likwid\/wiki\/Haswell.  Haswell. 2013. Performance Monitoring Counters for Haswell Architecture. Retrieved from https:\/\/code.google.com\/p\/likwid\/wiki\/Haswell."},{"key":"e_1_2_2_58_1","doi-asserted-by":"publisher","DOI":"10.1145\/1065944.1065969"},{"key":"e_1_2_2_59_1","doi-asserted-by":"publisher","DOI":"10.1145\/1816038.1815998"},{"key":"e_1_2_2_60_1","unstructured":"HPL. 2008. HPL\u2014A Portable Implementation of the High-Performance Linpack Benchmark for Distributed-Memory Computers. Retrieved from http:\/\/www.netlib.org\/benchmark\/hpl\/.  HPL. 2008. HPL\u2014A Portable Implementation of the High-Performance Linpack Benchmark for Distributed-Memory Computers. Retrieved from http:\/\/www.netlib.org\/benchmark\/hpl\/."},{"key":"e_1_2_2_61_1","unstructured":"HPRC. 2015. High-Performance Reconfigurable Computing. Retrieved from http:\/\/www.chrec.org\/.  HPRC. 2015. High-Performance Reconfigurable Computing. Retrieved from http:\/\/www.chrec.org\/."},{"key":"e_1_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1109\/ReConFig.2011.49"},{"key":"e_1_2_2_63_1","doi-asserted-by":"publisher","DOI":"10.1504\/IJNVO.2014.067878"},{"key":"e_1_2_2_64_1","unstructured":"IntelMPSS. 2014. Intel Manycore Platform Software Stack (Intel MPSS). Retrieved from https:\/\/software.intel.com\/en-us\/articles\/intel-manycore-platform-software-stack-mpss.  IntelMPSS. 2014. Intel Manycore Platform Software Stack (Intel MPSS). Retrieved from https:\/\/software.intel.com\/en-us\/articles\/intel-manycore-platform-software-stack-mpss."},{"key":"e_1_2_2_65_1","unstructured":"IntelPCM. 2012. Intel Performance Counter Monitor\u2014A better way to measure CPU utilization. Retrieved from https:\/\/software.intel.com\/en-us\/articles\/intel-performance-counter-monitor.  IntelPCM. 2012. Intel Performance Counter Monitor\u2014A better way to measure CPU utilization. Retrieved from https:\/\/software.intel.com\/en-us\/articles\/intel-performance-counter-monitor."},{"key":"e_1_2_2_66_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2003.1253186"},{"key":"e_1_2_2_67_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2013.07.012"},{"key":"e_1_2_2_68_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2008.4771804"},{"key":"e_1_2_2_69_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2008.4536223"},{"key":"e_1_2_2_70_1","doi-asserted-by":"publisher","DOI":"10.1145\/1453175.1453180"},{"key":"e_1_2_2_71_1","doi-asserted-by":"publisher","DOI":"10.1109\/SAAHPC.2012.26"},{"key":"e_1_2_2_72_1","volume-title":"Computer Architecture Techniques for Power-Efficiency","author":"Kaxiras Stefanos","unstructured":"Stefanos Kaxiras and Margaret Martonosi . 2008. Computer Architecture Techniques for Power-Efficiency ( 1 st ed.). Morgan and Claypool Publishers . Stefanos Kaxiras and Margaret Martonosi. 2008. Computer Architecture Techniques for Power-Efficiency (1st ed.). Morgan and Claypool Publishers.","edition":"1"},{"key":"e_1_2_2_73_1","doi-asserted-by":"publisher","DOI":"10.1145\/2491661.2481429"},{"key":"e_1_2_2_74_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2013.6704670"},{"key":"e_1_2_2_75_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISVLSI.2010.84"},{"key":"e_1_2_2_76_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2009.8"},{"key":"e_1_2_2_77_1","unstructured":"LAPACK. 2013. Linear Algebra Package. Retrieved from http:\/\/http:\/\/www.netlib.org\/lapack\/.  LAPACK. 2013. Linear Algebra Package. Retrieved from http:\/\/http:\/\/www.netlib.org\/lapack\/."},{"key":"e_1_2_2_78_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2016.2608824"},{"key":"e_1_2_2_79_1","doi-asserted-by":"publisher","DOI":"10.1145\/1168917.1168881"},{"key":"e_1_2_2_80_1","doi-asserted-by":"publisher","DOI":"10.5555\/1855610.1855614"},{"key":"e_1_2_2_81_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2014.162"},{"key":"e_1_2_2_82_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPPW.2011.25"},{"key":"e_1_2_2_83_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11227-015-1463-3"},{"key":"e_1_2_2_84_1","doi-asserted-by":"publisher","DOI":"10.1145\/2445572.2445577"},{"key":"e_1_2_2_85_1","doi-asserted-by":"publisher","DOI":"10.1145\/885651.781048"},{"key":"e_1_2_2_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/2611758"},{"key":"e_1_2_2_87_1","doi-asserted-by":"publisher","DOI":"10.1145\/1346281.1346318"},{"key":"e_1_2_2_88_1","doi-asserted-by":"publisher","DOI":"10.1145\/2541228.2555291"},{"key":"e_1_2_2_89_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-013-0239-3"},{"key":"e_1_2_2_90_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-011-0190-0"},{"key":"e_1_2_2_91_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-011-0191-z"},{"key":"e_1_2_2_92_1","volume-title":"Proceedings of the ACM SOSP Workshop on Power Aware Computing and Systems (HotPower\u201909)","author":"Ma Xiaohan","year":"2009","unstructured":"Xiaohan Ma , Mian Dong , Lin Zhong , and Zhigang Deng . 2009 . Statistical power consumption analysis and modeling for GPU-based computing . In Proceedings of the ACM SOSP Workshop on Power Aware Computing and Systems (HotPower\u201909) . Xiaohan Ma, Mian Dong, Lin Zhong, and Zhigang Deng. 2009. Statistical power consumption analysis and modeling for GPU-based computing. In Proceedings of the ACM SOSP Workshop on Power Aware Computing and Systems (HotPower\u201909)."},{"key":"e_1_2_2_93_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-011-0189-6"},{"key":"e_1_2_2_94_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.compeleceng.2013.04.016"},{"key":"e_1_2_2_95_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10723-015-9345-8"},{"key":"e_1_2_2_96_1","volume-title":"Proceedings of the 2011 USENIX Conference on USENIX Annual Technical Conference (USENIXATC\u201911)","author":"McCullough John C.","unstructured":"John C. McCullough , Yuvraj Agarwal , Jaideep Chandrashekar , Sathyanarayan Kuppuswamy , Alex C. Snoeren , and Rajesh K. Gupta . 2011. Evaluating the effectiveness of model-based power characterization . In Proceedings of the 2011 USENIX Conference on USENIX Annual Technical Conference (USENIXATC\u201911) . USENIX Association. John C. McCullough, Yuvraj Agarwal, Jaideep Chandrashekar, Sathyanarayan Kuppuswamy, Alex C. Snoeren, and Rajesh K. Gupta. 2011. Evaluating the effectiveness of model-based power characterization. In Proceedings of the 2011 USENIX Conference on USENIX Annual Technical Conference (USENIXATC\u201911). USENIX Association."},{"key":"e_1_2_2_97_1","doi-asserted-by":"publisher","DOI":"10.1145\/2788396"},{"key":"e_1_2_2_98_1","doi-asserted-by":"publisher","DOI":"10.1145\/2636342"},{"key":"e_1_2_2_99_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2013.183"},{"key":"e_1_2_2_100_1","doi-asserted-by":"publisher","DOI":"10.1109\/GREENCOMP.2010.5598316"},{"key":"e_1_2_2_101_1","doi-asserted-by":"crossref","unstructured":"Moore. 1965. Moore\u2019s Law. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Moore's_law.  Moore. 1965. Moore\u2019s Law. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Moore's_law.","DOI":"10.2307\/1370998"},{"key":"e_1_2_2_102_1","volume-title":"Gamut: Generic Application EMULaTOR.","author":"Moore J.","year":"2004","unstructured":"J. Moore . 2004 . Gamut: Generic Application EMULaTOR. (2004). J. Moore. 2004. Gamut: Generic Application EMULaTOR. (2004)."},{"key":"e_1_2_2_103_1","unstructured":"MPICH. 2016. MPICH--High performance portable MPI. Retrieved from http:\/\/www.mpich.org\/.  MPICH. 2016. MPICH--High performance portable MPI. Retrieved from http:\/\/www.mpich.org\/."},{"key":"e_1_2_2_104_1","doi-asserted-by":"publisher","DOI":"10.1109\/MICRO.2007.30"},{"key":"e_1_2_2_105_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.suscom.2012.03.002"},{"key":"e_1_2_2_106_1","doi-asserted-by":"publisher","DOI":"10.1109\/GREENCOMP.2010.5598315"},{"key":"e_1_2_2_107_1","unstructured":"NAS. 2015. NAS Parallel Benchmarks. Retrieved from https:\/\/www.nas.nasa.gov\/publications\/npb.html.  NAS. 2015. NAS Parallel Benchmarks. Retrieved from https:\/\/www.nas.nasa.gov\/publications\/npb.html."},{"key":"e_1_2_2_108_1","unstructured":"NVML. 2011. NVIDIA Management Library (NVML). Retrieved from https:\/\/developer.nvidia.com\/nvidia-management-library-nvml.  NVML. 2011. NVIDIA Management Library (NVML). Retrieved from https:\/\/developer.nvidia.com\/nvidia-management-library-nvml."},{"key":"e_1_2_2_109_1","unstructured":"OpenMPI. 2016. OpenMPI - Open source high performance computing. Retrieved from https:\/\/www.open-mpi.org\/.  OpenMPI. 2016. OpenMPI - Open source high performance computing. Retrieved from https:\/\/www.open-mpi.org\/."},{"key":"e_1_2_2_110_1","doi-asserted-by":"publisher","DOI":"10.1145\/2532637"},{"key":"e_1_2_2_111_1","volume-title":"Proceedings of the IEEE International SOC Conference","author":"Ou Jingzhao","year":"2004","unstructured":"Jingzhao Ou and V. K. Prasanna . 2004. Rapid energy estimation of computations on FPGA based soft processors . In Proceedings of the IEEE International SOC Conference , 2004 . Jingzhao Ou and V. K. Prasanna. 2004. Rapid energy estimation of computations on FPGA based soft processors. In Proceedings of the IEEE International SOC Conference, 2004."},{"key":"e_1_2_2_112_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2014.6983056"},{"key":"e_1_2_2_113_1","unstructured":"PAPI. 2015. Performance Application Programming Interface 5.4.1. Retrieved from http:\/\/icl.cs.utk.edu\/papi\/.  PAPI. 2015. Performance Application Programming Interface 5.4.1. Retrieved from http:\/\/icl.cs.utk.edu\/papi\/."},{"key":"e_1_2_2_114_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2011.120"},{"key":"e_1_2_2_115_1","unstructured":"PCIE. 2003. Peripheral Component Interconnect Express. Retrieved from https:\/\/en.wikipedia.org\/wiki\/PCI_Express.  PCIE. 2003. Peripheral Component Interconnect Express. Retrieved from https:\/\/en.wikipedia.org\/wiki\/PCI_Express."},{"key":"e_1_2_2_116_1","unstructured":"PLASMA. 2015. Parallel Linear Algebra for Scalable Multi-core Architectures. Retrieved from http:\/\/http:\/\/icl.cs.utk.edu\/plasma\/.  PLASMA. 2015. Parallel Linear Algebra for Scalable Multi-core Architectures. Retrieved from http:\/\/http:\/\/icl.cs.utk.edu\/plasma\/."},{"key":"e_1_2_2_117_1","doi-asserted-by":"publisher","DOI":"10.1145\/1059876.1059881"},{"key":"e_1_2_2_118_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2009.4798264"},{"key":"e_1_2_2_119_1","unstructured":"PowerAPI. 2016. Sandia National Laboratories: High Performance Computing Power Application Programming Interface (API) Specification. Retrieved from http:\/\/powerapi.sandia.gov\/.  PowerAPI. 2016. Sandia National Laboratories: High Performance Computing Power Application Programming Interface (API) Specification. Retrieved from http:\/\/powerapi.sandia.gov\/."},{"key":"e_1_2_2_120_1","unstructured":"QPI. 2008. Intel QuickPath Interconnect. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Intel_QuickPath_Interconnect.  QPI. 2008. Intel QuickPath Interconnect. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Intel_QuickPath_Interconnect."},{"key":"e_1_2_2_122_1","doi-asserted-by":"publisher","DOI":"10.5555\/1855610.1855613"},{"key":"e_1_2_2_123_1","unstructured":"Rodinia. 2015. The Rodinia Benchmark Suite version 3.0. Retrieved from https:\/\/www.cs.virginia.edu\/skadron\/wiki\/rodinia\/index.php\/Rodinia:Accelerating_Compute-Intensive_Applications_with_Accelerators.  Rodinia. 2015. The Rodinia Benchmark Suite version 3.0. Retrieved from https:\/\/www.cs.virginia.edu\/skadron\/wiki\/rodinia\/index.php\/Rodinia:Accelerating_Compute-Intensive_Applications_with_Accelerators."},{"key":"e_1_2_2_124_1","doi-asserted-by":"publisher","DOI":"10.5555\/1855610.1855621"},{"key":"e_1_2_2_125_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2012.12"},{"key":"e_1_2_2_126_1","doi-asserted-by":"publisher","DOI":"10.1145\/2422436.2422470"},{"key":"e_1_2_2_127_1","doi-asserted-by":"publisher","DOI":"10.1109\/ReConFig.2011.11"},{"key":"e_1_2_2_128_1","doi-asserted-by":"publisher","DOI":"10.5555\/822080.822800"},{"key":"e_1_2_2_129_1","doi-asserted-by":"publisher","DOI":"10.5555\/2648668.2648758"},{"key":"e_1_2_2_130_1","doi-asserted-by":"publisher","DOI":"10.1145\/1577129.1577137"},{"key":"e_1_2_2_131_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2013.73"},{"key":"e_1_2_2_132_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2007.43"},{"key":"e_1_2_2_133_1","unstructured":"SPEC. 2015. SPEC\u2019s Benchmarks. Retrieved from https:\/\/www.spec.org\/benchmarks.html.  SPEC. 2015. SPEC\u2019s Benchmarks. Retrieved from https:\/\/www.spec.org\/benchmarks.html."},{"key":"e_1_2_2_134_1","doi-asserted-by":"publisher","DOI":"10.1109\/IGCC.2011.6008552"},{"key":"e_1_2_2_135_1","volume-title":"STREAM: Sustainable Memory Bandwidth in High Performance Computers.","year":"2015","unstructured":"Stream. 2015 . STREAM: Sustainable Memory Bandwidth in High Performance Computers. Retrieved from https:\/\/www.cs.virginia.edu\/stream\/. Stream. 2015. STREAM: Sustainable Memory Bandwidth in High Performance Computers. Retrieved from https:\/\/www.cs.virginia.edu\/stream\/."},{"key":"e_1_2_2_136_1","doi-asserted-by":"publisher","DOI":"10.1109\/GreenCom-CPSCom.2010.138"},{"key":"e_1_2_2_137_1","unstructured":"SYSTEMG. 2015. System G cluster. Retrieved from https:\/\/www.cs.vt.edu\/facilities\/systemg.  SYSTEMG. 2015. System G cluster. Retrieved from https:\/\/www.cs.vt.edu\/facilities\/systemg."},{"key":"e_1_2_2_138_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2014.09.001"},{"key":"e_1_2_2_139_1","unstructured":"TDP. 2015. Thermal design power. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Thermal_design_power.  TDP. 2015. Thermal design power. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Thermal_design_power."},{"key":"e_1_2_2_140_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-39924-7_38"},{"key":"e_1_2_2_141_1","doi-asserted-by":"publisher","DOI":"10.1145\/1508128.1508139"},{"key":"e_1_2_2_142_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2012.121"},{"key":"e_1_2_2_143_1","volume-title":"The List--November","author":"Top","year":"2015","unstructured":"Top500. 2015. Top 500. The List--November 2015 . Retrieved from http:\/\/top500.org\/lists\/2015\/11. Top500. 2015. Top 500. The List--November 2015. Retrieved from http:\/\/top500.org\/lists\/2015\/11."},{"key":"e_1_2_2_144_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2012.47"},{"key":"e_1_2_2_145_1","doi-asserted-by":"publisher","DOI":"10.1145\/339647.339659"},{"key":"e_1_2_2_146_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2015.02.002"},{"key":"e_1_2_2_147_1","volume-title":"Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing. USENIX Association.","author":"Wang Hongyi","year":"2010","unstructured":"Hongyi Wang , Qingfeng Jing , Rishan Chen , Bingsheng He , Zhengping Qian , and Lidong Zhou . 2010 . Distributed systems meet economics: Pricing in the cloud . In Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing. USENIX Association. Hongyi Wang, Qingfeng Jing, Rishan Chen, Bingsheng He, Zhengping Qian, and Lidong Zhou. 2010. Distributed systems meet economics: Pricing in the cloud. In Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing. USENIX Association."},{"key":"e_1_2_2_149_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCD.2006.4380849"},{"key":"e_1_2_2_150_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPPW.2012.39"},{"key":"e_1_2_2_151_1","doi-asserted-by":"publisher","DOI":"10.1145\/1498765.1498785"},{"key":"e_1_2_2_152_1","unstructured":"WINXPERF. 2015. Windows Performance Toolkit Technical Reference. Retrieved from https:\/\/msdn.microsoft.com\/en-us\/library\/windows\/hardware\/hh162945.aspx.  WINXPERF. 2015. Windows Performance Toolkit Technical Reference. Retrieved from https:\/\/msdn.microsoft.com\/en-us\/library\/windows\/hardware\/hh162945.aspx."},{"key":"e_1_2_2_153_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2012.06.003"},{"key":"e_1_2_2_154_1","doi-asserted-by":"publisher","DOI":"10.1145\/1146909.1147053"},{"key":"e_1_2_2_155_1","unstructured":"XEONPHI. 2015. Intel Many Integrated Core Architecture. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Xeon_Phi.  XEONPHI. 2015. Intel Many Integrated Core Architecture. Retrieved from https:\/\/en.wikipedia.org\/wiki\/Xeon_Phi."},{"key":"e_1_2_2_156_1","volume-title":"Proceedings of the 2012 7th International Conference on Computing and Convergence Technology (ICCCT\u201912)","author":"Xie Qiyao","year":"2012","unstructured":"Qiyao Xie , Tian Huang , Zhihai Zou , Liang Xia , Yongxin Zhu , and Jiang Jiang . 2012 . An accurate power model for GPU processors . In Proceedings of the 2012 7th International Conference on Computing and Convergence Technology (ICCCT\u201912) . IEEE, 1141--1146. Qiyao Xie, Tian Huang, Zhihai Zou, Liang Xia, Yongxin Zhu, and Jiang Jiang. 2012. An accurate power model for GPU processors. In Proceedings of the 2012 7th International Conference on Computing and Convergence Technology (ICCCT\u201912). IEEE, 1141--1146."},{"key":"e_1_2_2_157_1","unstructured":"XPE. 2015. Xilinx Power Estimator (XPE). Retrieved from http:\/\/www.xilinx.com\/products\/technology\/power\/xpe.html.  XPE. 2015. Xilinx Power Estimator (XPE). Retrieved from http:\/\/www.xilinx.com\/products\/technology\/power\/xpe.html."},{"key":"e_1_2_2_158_1","doi-asserted-by":"publisher","DOI":"10.1109\/NAS.2011.51"},{"key":"e_1_2_2_159_1","doi-asserted-by":"publisher","DOI":"10.1145\/2541228.2541231"},{"key":"e_1_2_2_160_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.1913"}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3078811","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3078811","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:03:08Z","timestamp":1750215788000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3078811"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,6,29]]},"references-count":158,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,5,31]]}},"alternative-id":["10.1145\/3078811"],"URL":"https:\/\/doi.org\/10.1145\/3078811","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,6,29]]},"assertion":[{"value":"2015-12-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-04-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-06-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}