{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:34:56Z","timestamp":1750221296850,"version":"3.41.0"},"reference-count":42,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2018,8,28]],"date-time":"2018-08-28T00:00:00Z","timestamp":1535414400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Archit. Code Optim."],"published-print":{"date-parts":[[2018,9,30]]},"abstract":"<jats:p>Prefetching is a well-known technique to mitigate scalability challenges in the Partitioned Global Address Space (PGAS) model. It has been studied as either an automated compiler optimization or a manual programmer optimization. Using the PGAS locality awareness, we define a hybrid tradeoff. Specifically, we introduce locality-aware productive prefetching support for PGAS. Our novel, user-driven approach strikes a balance between the ease-of-use of compiler-based automated prefetching and the high performance of the laborious manual prefetching. Our prototype implementation in Chapel shows that significant scalability and performance improvements can be achieved with minimal effort in common applications.<\/jats:p>","DOI":"10.1145\/3233299","type":"journal-article","created":{"date-parts":[[2018,8,30]],"date-time":"2018-08-30T13:45:11Z","timestamp":1535636711000},"page":"1-26","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":9,"title":["LAPPS"],"prefix":"10.1145","volume":"15","author":[{"given":"Engin","family":"Kayraklioglu","sequence":"first","affiliation":[{"name":"The George Washington University, Street NW, Washington, D.C., USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Michael P.","family":"Ferguson","sequence":"additional","affiliation":[{"name":"Cray Inc., USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tarek","family":"El-Ghazawi","sequence":"additional","affiliation":[{"name":"The George Washington University, Street NW, Washington, D.C., USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2018,8,28]]},"reference":[{"key":"e_1_2_1_1_1","first-page":"984","volume":"0","year":"2017","unstructured":"2017 . Chapel Language Spefications - Version 0 . 984 . Retrieved February 02, 2018 from https:\/\/chapel-lang.org\/spec\/spec-0.98.pdf. 2017. Chapel Language Spefications - Version 0.984. Retrieved February 02, 2018 from https:\/\/chapel-lang.org\/spec\/spec-0.98.pdf.","journal-title":"Chapel Language Spefications - Version"},{"volume-title":"prefetch\/noprefetch &verbar","year":"2018","key":"e_1_2_1_2_1","unstructured":"2018. prefetch\/noprefetch &verbar ; Intel Software. Retrieved February 20, 2018 from https:\/\/software.intel.com\/en-us\/node\/524554. 2018. prefetch\/noprefetch &verbar; Intel Software. Retrieved February 20, 2018 from https:\/\/software.intel.com\/en-us\/node\/524554."},{"volume-title":"Retrieved","year":"2018","key":"e_1_2_1_3_1","unstructured":"2018. Using the GNU Compiler Collection (GCC): Other Builtins . Retrieved February 20, 2018 from https:\/\/gcc.gnu.org\/onlinedocs\/gcc\/Other-Builtins.html. 2018. Using the GNU Compiler Collection (GCC): Other Builtins. Retrieved February 20, 2018 from https:\/\/gcc.gnu.org\/onlinedocs\/gcc\/Other-Builtins.html."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/2464996.2465006"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/2464996.2467277"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/1465482.1465560"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/PGAS.2015.16"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/2897783"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2011.105"},{"volume-title":"The Design and Implementation of a Region-Based Parallel Programming Language","author":"Chamberlain Bradford L.","key":"e_1_2_1_11_1","unstructured":"Bradford L. Chamberlain . 2001. The Design and Implementation of a Region-Based Parallel Programming Language . University of Washington. Bradford L. Chamberlain. 2001. The Design and Implementation of a Region-Based Parallel Programming Language. University of Washington."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1177\/1094342007078442"},{"key":"e_1_2_1_13_1","volume-title":"Proceedings of Cray Users Group.","author":"Chamberlain Bradford L.","year":"2011","unstructured":"Bradford L. Chamberlain , Sung-eun Choi, Steven J. Deitz , David Iten , and Vassily Litvinov . 2011 . Authoring user-defined domain maps in chapel . In Proceedings of Cray Users Group. Bradford L. Chamberlain, Sung-eun Choi, Steven J. Deitz, David Iten, and Vassily Litvinov. 2011. Authoring user-defined domain maps in chapel. In Proceedings of Cray Users Group."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1345206.1345211"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020373.2020375"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/PACT.2005.13"},{"volume-title":"Proceedings of the 1997 International Conference on Parallel Processing (Cat. No.97TB100162)","author":"Choi Sung-Eun","key":"e_1_2_1_17_1","unstructured":"Sung-Eun Choi and L. Snyder . 1997. Quantifying the effects of communication optimizations . In Proceedings of the 1997 International Conference on Parallel Processing (Cat. No.97TB100162) . 218--222. Sung-Eun Choi and L. Snyder. 1997. Quantifying the effects of communication optimizations. In Proceedings of the 1997 International Conference on Parallel Processing (Cat. No.97TB100162). 218--222."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/1065944.1065950"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.5555\/762761.762821"},{"key":"e_1_2_1_20_1","volume-title":"UPC: Distributed Shared-Memory Programming","author":"El-Ghazawi Tarek","year":"2003","unstructured":"Tarek El-Ghazawi , William Carlson , Thomas Sterling , and Katherine Yelick . 2003 . UPC: Distributed Shared-Memory Programming . Wiley-Interscience . Tarek El-Ghazawi, William Carlson, Thomas Sterling, and Katherine Yelick. 2003. UPC: Distributed Shared-Memory Programming. Wiley-Interscience."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.5555\/645535.657163"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/PGAS.2015.10"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.5555\/3019040.3019044"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/2833157.2833164"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/1375527.1375567"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2017.126"},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2016.193"},{"key":"e_1_2_1_28_1","volume-title":"Zosel","author":"Koelbel Charles H.","year":"1993","unstructured":"Charles H. Koelbel and Mary E . Zosel . 1993 . The High Performance FORTRAN Handbook. MIT Press , Cambridge, MA. Charles H. Koelbel and Mary E. Zosel. 1993. The High Performance FORTRAN Handbook. MIT Press, Cambridge, MA."},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.5555\/977395.977673"},{"key":"e_1_2_1_30_1","volume-title":"Slingwine","author":"McKenney Paul E.","year":"1998","unstructured":"Paul E. McKenney and John D . Slingwine . 1998 . Read-copy update: Using execution history to solve concurrency problems. In Parallel and Distributed Computing and Systems . 509--518. Paul E. McKenney and John D. Slingwine. 1998. Read-copy update: Using execution history to solve concurrency problems. In Parallel and Distributed Computing and Systems. 509--518."},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/109625.109637"},{"volume-title":"High Performance Computing in Science and Engineering 99","author":"M\u00fcller Matthias M.","key":"e_1_2_1_32_1","unstructured":"Matthias M. M\u00fcller . 1999. KaHPF: Compiler generated data prefetching for HPF . In High Performance Computing in Science and Engineering 99 . Springer , Berlin , 474--482. Matthias M. M\u00fcller. 1999. KaHPF: Compiler generated data prefetching for HPF. In High Performance Computing in Science and Engineering 99. Springer, Berlin, 474--482."},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/277830.277919"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/289918.289920"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/2076021.2048088"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/1408643.1408645"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/SBAC-PAD.2012.18"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2008.4536359"},{"key":"e_1_2_1_39_1","volume-title":"Srinivas Sridharan, Timothy G. Mattson, John Abercrombie, and Jacob Nelson.","author":"Van der Wijngaart Rob F.","year":"2016","unstructured":"Rob F. Van der Wijngaart , Abdullah Kayi , Jeff R. Hammond , Gabriele Jost , Tom St John , Srinivas Sridharan, Timothy G. Mattson, John Abercrombie, and Jacob Nelson. 2016 . Comparing runtime systems with exascale ambitions using the parallel research kernels. In High Performance Computing. Springer , Cham, 321--339. http:\/\/link.springer.com\/chapter\/10.1007\/978-3-319-41321-1_17 Rob F. Van der Wijngaart, Abdullah Kayi, Jeff R. Hammond, Gabriele Jost, Tom St John, Srinivas Sridharan, Timothy G. Mattson, John Abercrombie, and Jacob Nelson. 2016. Comparing runtime systems with exascale ambitions using the parallel research kernels. In High Performance Computing. Springer, Cham, 321--339. http:\/\/link.springer.com\/chapter\/10.1007\/978-3-319-41321-1_17"},{"key":"e_1_2_1_40_1","volume-title":"Parallel Research Kernels. Retrieved","author":"Rob","year":"2017","unstructured":"Rob F. Van der Wijngaart, Tim Mattson, Jeff Hammond, Srinivas Sridharan, and Evangelos Georganas. 2017 . Parallel Research Kernels. Retrieved September 11, 2017 from https:\/\/github.com\/ParRes\/Kernels\/blob\/master\/doc\/par-res-kern-report-v1.3.pdf. Rob F. Van der Wijngaart, Tim Mattson, Jeff Hammond, Srinivas Sridharan, and Evangelos Georganas. 2017. Parallel Research Kernels. Retrieved September 11, 2017 from https:\/\/github.com\/ParRes\/Kernels\/blob\/master\/doc\/par-res-kern-report-v1.3.pdf."},{"volume-title":"Proceedings of the 2014 IEEE High Performance Extreme Computing Conference (HPEC\u201914)","author":"Rob","key":"e_1_2_1_41_1","unstructured":"Rob F. Van der Wijngaart and Tim G. Mattson. 2014. The parallel research kernels . In Proceedings of the 2014 IEEE High Performance Extreme Computing Conference (HPEC\u201914) . 1--6. Rob F. Van der Wijngaart and Tim G. Mattson. 2014. The parallel research kernels. In Proceedings of the 2014 IEEE High Performance Extreme Computing Conference (HPEC\u201914). 1--6."},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/PGAS.2015.24"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2014.115"}],"container-title":["ACM Transactions on Architecture and Code Optimization"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3233299","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3233299","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:13:12Z","timestamp":1750212792000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3233299"}},"subtitle":["Locality-Aware Productive Prefetching Support for PGAS"],"short-title":[],"issued":{"date-parts":[[2018,8,28]]},"references-count":42,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,9,30]]}},"alternative-id":["10.1145\/3233299"],"URL":"https:\/\/doi.org\/10.1145\/3233299","relation":{},"ISSN":["1544-3566","1544-3973"],"issn-type":[{"type":"print","value":"1544-3566"},{"type":"electronic","value":"1544-3973"}],"subject":[],"published":{"date-parts":[[2018,8,28]]},"assertion":[{"value":"2017-11-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-06-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2018-08-28","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}