{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:15:31Z","timestamp":1750306531082,"version":"3.41.0"},"reference-count":37,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2015,2,18]],"date-time":"2015-02-18T00:00:00Z","timestamp":1424217600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Parallel Comput."],"published-print":{"date-parts":[[2015,2,18]]},"abstract":"<jats:p>Networks are among major power consumers in large-scale parallel systems. During execution of common parallel applications, a sizeable fraction of the links in the high-radix interconnects are either never used or are underutilized. We propose a runtime system based adaptive approach to turn off unused links, which has various advantages over the previously proposed hardware and compiler based approaches. We discuss why the runtime system is the best system component to accomplish this task, and test the effectiveness of our approach using real applications (including NAMD, MILC), and application benchmarks (including NAS Parallel Benchmarks, Stencil). These codes are simulated on representative topologies such as 6-D Torus and multilevel directly connected network (similar to IBM PERCS in Power 775 and Dragonfly in Cray Aries). For common applications with near-neighbor communication pattern, our approach can save up to 20% of total machine's power and energy, without any performance penalty.<\/jats:p>","DOI":"10.1145\/2687001","type":"journal-article","created":{"date-parts":[[2015,2,23]],"date-time":"2015-02-23T15:32:19Z","timestamp":1424705539000},"page":"1-21","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":13,"title":["Power Management of Extreme-Scale Networks with On\/Off Links in Runtime Systems"],"prefix":"10.1145","volume":"1","author":[{"given":"Ehsan","family":"Totoni","sequence":"first","affiliation":[{"name":"University of Illinois at Urbana-Champaign"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nikhil","family":"Jain","sequence":"additional","affiliation":[{"name":"University of Illinois at Urbana-Champaign"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Laxmikant V.","family":"Kale","sequence":"additional","affiliation":[{"name":"University of Illinois at Urbana-Champaign"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,2,18]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1815961.1816004"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2011.21"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.5555\/1898699.1898826"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2010.16"},{"volume-title":"Proceedings of the Supercomputing Conference.","author":"Bailey D. H.","key":"e_1_2_1_5_1","unstructured":"D. H. Bailey , E. Barszcz , L. Dagum , and H. D. Simon . 1992. NAS parallel benchmark results . In Proceedings of the Supercomputing Conference. D. H. Bailey, E. Barszcz, L. Dagum, and H. D. Simon. 1992. NAS parallel benchmark results. In Proceedings of the Supercomputing Conference."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1103\/PhysRevD.61.111502"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.1637"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2063384.2063486"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1142\/S0129626409000419"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/2063384.2063419"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the IEEE International Parallel and Distributed Processing Symposium. 1--8. DOI:http:\/\/dx.doi.org\/10","author":"Conner S.","year":"2007","unstructured":"S. Conner , S. Akioka , M. J. Irwin , and P. Raghavan . 2007. Link shutdown opportunities during collective communications in 3-D torus nets . In Proceedings of the IEEE International Parallel and Distributed Processing Symposium. 1--8. DOI:http:\/\/dx.doi.org\/10 .1109\/IPDPS. 2007 .370534. 10.1109\/IPDPS.2007.370534 S. Conner, S. Akioka, M. J. Irwin, and P. Raghavan. 2007. Link shutdown opportunities during collective communications in 3-D torus nets. In Proceedings of the IEEE International Parallel and Distributed Processing Symposium. 1--8. DOI:http:\/\/dx.doi.org\/10.1109\/IPDPS.2007.370534."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.5555\/2388996.2389136"},{"key":"e_1_2_1_14_1","unstructured":"Megan Gilge. 2013. Blue Gene\/Q application development. http:\/\/www.redbooks.ibm.com\/abstracts\/sg247948.html.  Megan Gilge. 2013. Blue Gene\/Q application development. http:\/\/www.redbooks.ibm.com\/abstracts\/sg247948.html."},{"key":"e_1_2_1_15_1","volume-title":"Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation.","author":"Heller Brandon","year":"2010","unstructured":"Brandon Heller , Srini Seetharaman , Priya Mahadevan , Yiannis Yiakoumis , Puneet Sharma , Sujata Banerjee , and Nick McKeown . 2010 . ElasticTree: Saving energy in data center networks . In Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation. Brandon Heller, Srini Seetharaman, Priya Mahadevan, Yiannis Yiakoumis, Puneet Sharma, Sujata Banerjee, and Nick McKeown. 2010. ElasticTree: Saving energy in data center networks. In Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2013.190"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1029\/2004GB002349"},{"key":"e_1_2_1_18_1","volume-title":"Phillips","author":"Kale Laxmikant V.","year":"2011","unstructured":"Laxmikant V. Kale , Abhinav Bhatele , Eric J. Bohm , and James C . Phillips . 2011 . NAnoscale Molecular Dynamics (NAMD). In Encyclopedia of Parallel Computing, D. Padua, Ed., Springer . Laxmikant V. Kale, Abhinav Bhatele, Eric J. Bohm, and James C. Phillips. 2011. NAnoscale Molecular Dynamics (NAMD). In Encyclopedia of Parallel Computing, D. Padua, Ed., Springer."},{"key":"e_1_2_1_19_1","volume-title":"Kale and Gengbin Zheng","author":"Laxmikant","year":"2009","unstructured":"Laxmikant V. Kale and Gengbin Zheng . 2009 . Charm++ and AMPI: Adaptive runtime strategies via migratable objects. In Advanced Computational Infrastructures for Parallel and Distributed Applications, M. Parashar, Ed., Wiley-Interscience , 265--282. Laxmikant V. Kale and Gengbin Zheng. 2009. Charm++ and AMPI: Adaptive runtime strategies via migratable objects. In Advanced Computational Infrastructures for Parallel and Distributed Applications, M. Parashar, Ed., Wiley-Interscience, 265--282."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPADS.2012.81"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/1394608.1382129"},{"key":"e_1_2_1_22_1","unstructured":"P. M. Kogge. 2008. Architectural challenges at the exascale frontier (invited talk). In Simulating the Future: Using One Million Cores and Beyond.  P. M. Kogge. 2008. Architectural challenges at the exascale frontier (invited talk). In Simulating the Future: Using One Million Cores and Beyond."},{"key":"e_1_2_1_23_1","unstructured":"Peter Kogge Keren Bergman Shekhar Borkar etal 2008. ExaScale computing study: Technology challenges in achieving exascale systems.  Peter Kogge Keren Bergman Shekhar Borkar et al. 2008. ExaScale computing study: Technology challenges in achieving exascale systems."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/SBAC-PAD.2012.48"},{"volume-title":"Proceedings of the 20th High Performance Computing Symposium.","author":"Laros James","key":"e_1_2_1_25_1","unstructured":"James Laros , Kevin Pedretti , S. Kelly , Wei Shu , and C. Vaughan . 2012. Energy based performance tuning for large scale high performance computing systems . In Proceedings of the 20th High Performance Computing Symposium. James Laros, Kevin Pedretti, S. Kelly, Wei Shu, and C. Vaughan. 2012. Energy based performance tuning for large scale high performance computing systems. In Proceedings of the 20th High Performance Computing Symposium."},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.5555\/2014698.2014904"},{"key":"e_1_2_1_27_1","volume-title":"Proceedings of the IEEE INFOCOM Workshops. 1--6. DOI:http:\/\/dx.doi.org\/10","author":"Mahadevan P.","year":"2009","unstructured":"P. Mahadevan , P. Sharma , S. Banerjee , and P. Ranganathan . 2009. Energy Aware Network Operations . In Proceedings of the IEEE INFOCOM Workshops. 1--6. DOI:http:\/\/dx.doi.org\/10 .1109\/INFCOMW. 2009 .5072138. 10.1109\/INFCOMW.2009.5072138 P. Mahadevan, P. Sharma, S. Banerjee, and P. Ranganathan. 2009. Energy Aware Network Operations. In Proceedings of the IEEE INFOCOM Workshops. 1--6. DOI:http:\/\/dx.doi.org\/10.1109\/INFCOMW.2009.5072138."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2013.83"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2012.143"},{"key":"e_1_2_1_30_1","volume-title":"Proceedings of the 8th International Symposium on High-Performance Computer Architecture. 91--102","author":"Shang Li","year":"2003","unstructured":"Li Shang , Li-Shiuan Peh , and N. K. Jha . 2003. Dynamic voltage scaling with links for power optimization of interconnection networks . In Proceedings of the 8th International Symposium on High-Performance Computer Architecture. 91--102 . DOI:http:\/\/dx.doi.org\/10.1109\/HPCA. 2003 .1183527. 10.1109\/HPCA.2003.1183527 Li Shang, Li-Shiuan Peh, and N. K. Jha. 2003. Dynamic voltage scaling with links for power optimization of interconnection networks. In Proceedings of the 8th International Symposium on High-Performance Computer Architecture. 91--102. DOI:http:\/\/dx.doi.org\/10.1109\/HPCA.2003.1183527."},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/1216544.1216548"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2007.43"},{"key":"e_1_2_1_33_1","unstructured":"Top500 2013. Top500 supercomputing sites. http:\/\/top500.org.  Top500 2013. Top500 supercomputing sites. http:\/\/top500.org."},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISPASS.2012.6189208"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPADS.2011.121"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2013.191"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1177\/1094342010394383"},{"volume-title":"Proceedings of the 18th International Parallel and Distributed Processing Symposium. 78","author":"Zheng Gengbin","key":"e_1_2_1_38_1","unstructured":"Gengbin Zheng , Gunavardhan Kakulapati , and Laxmikant V. Kal\u00e9 . 2004. BigSim: A parallel simulator for performance prediction of extremely large parallel machines . In Proceedings of the 18th International Parallel and Distributed Processing Symposium. 78 . Gengbin Zheng, Gunavardhan Kakulapati, and Laxmikant V. Kal\u00e9. 2004. BigSim: A parallel simulator for performance prediction of extremely large parallel machines. In Proceedings of the 18th International Parallel and Distributed Processing Symposium. 78."}],"container-title":["ACM Transactions on Parallel Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2687001","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2687001","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T06:12:13Z","timestamp":1750227133000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2687001"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,2,18]]},"references-count":37,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2015,2,18]]}},"alternative-id":["10.1145\/2687001"],"URL":"https:\/\/doi.org\/10.1145\/2687001","relation":{},"ISSN":["2329-4949","2329-4957"],"issn-type":[{"type":"print","value":"2329-4949"},{"type":"electronic","value":"2329-4957"}],"subject":[],"published":{"date-parts":[[2015,2,18]]},"assertion":[{"value":"2013-08-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2014-08-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2015-02-18","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}