{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,11]],"date-time":"2026-07-11T15:43:15Z","timestamp":1783784595644,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":29,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,2,25]],"date-time":"2023-02-25T00:00:00Z","timestamp":1677283200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,2,25]]},"DOI":"10.1145\/3582514.3582519","type":"proceedings-article","created":{"date-parts":[[2023,2,24]],"date-time":"2023-02-24T17:27:41Z","timestamp":1677259661000},"page":"50-59","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":15,"title":["MPI-based Remote OpenMP Offloading: A More Efficient and Easy-to-use Implementation"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-1022-3801","authenticated-orcid":false,"given":"Baodi","family":"Shan","sequence":"first","affiliation":[{"name":"Institute for Advanced Computational Science, Stony Brook University, Stony Brook, New York, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5537-1840","authenticated-orcid":false,"given":"Mauricio","family":"Araya-Polo","sequence":"additional","affiliation":[{"name":"TotalEnergies EP Research &amp; Technology US, Houston, Texas, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0706-2824","authenticated-orcid":false,"given":"Abid M.","family":"Malik","sequence":"additional","affiliation":[{"name":"Brookhaven National Laboratory, Upton, New York, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8449-8579","authenticated-orcid":false,"given":"Barbara","family":"Chapman","sequence":"additional","affiliation":[{"name":"Institute for Advanced Computational Science, Stony Brook University, Stony Brook, New York, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,2,25]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2014.58"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2019.00104"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2012.71"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1177\/1094342007078442"},{"key":"e_1_3_2_1_5_1","unstructured":"gRPC community. [n.d.]. gRPC. https:\/\/grpc.io\/about\/.  gRPC community. [n.d.]. gRPC. https:\/\/grpc.io\/about\/."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW50202.2020.00104"},{"key":"e_1_3_2_1_7_1","volume-title":"Tong Chen, Zehra Sura, Kevin O'Brien, and Michael Wong.","author":"Jacob Arpith C.","year":"2015","unstructured":"Arpith C. Jacob , Ravi Nair , Alexandre E. Eichenberger , Samuel F. Antao , Carlo Bertolli , Tong Chen, Zehra Sura, Kevin O'Brien, and Michael Wong. 2015 . Exploiting Fine- and Coarse-Grained Parallelism Using a Directive Based Approach. In OpenMP: Heterogenous Execution and Data Movements, Christian Terboven, Bronis R. de Supinski, Pablo Reble, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer International Publishing , Cham, 30--41. Arpith C. Jacob, Ravi Nair, Alexandre E. Eichenberger, Samuel F. Antao, Carlo Bertolli, Tong Chen, Zehra Sura, Kevin O'Brien, and Michael Wong. 2015. Exploiting Fine- and Coarse-Grained Parallelism Using a Directive Based Approach. In OpenMP: Heterogenous Execution and Data Movements, Christian Terboven, Bronis R. de Supinski, Pablo Reble, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer International Publishing, Cham, 30--41."},{"key":"e_1_3_2_1_8_1","unstructured":"Kokkos. [n.d.]. Kokkos Remote Spaces. https:\/\/github.com\/kokkos\/kokkos-remote-spaces  Kokkos. [n.d.]. Kokkos Remote Spaces. https:\/\/github.com\/kokkos\/kokkos-remote-spaces"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/PAW-ATM49560.2019.00010"},{"key":"e_1_3_2_1_10_1","volume-title":"OpenMP in a Modern World: From Multi-device Support to Meta Programming, Michael Klemm, Bronis R","author":"Lu Wenbin","unstructured":"Wenbin Lu , Baodi Shan , Eric Raut , Jie Meng , Mauricio Araya-Polo , Johannes Doerfert , Abid M. Malik , and Barbara Chapman . 2022. Towards Efficient Remote OpenMP Offloading . In OpenMP in a Modern World: From Multi-device Support to Meta Programming, Michael Klemm, Bronis R . de Supinski, Jannis Klinkenberg, and Brandon Neth (Eds.). Springer International Publishing , Cham , 17--31. Wenbin Lu, Baodi Shan, Eric Raut, Jie Meng, Mauricio Araya-Polo, Johannes Doerfert, Abid M. Malik, and Barbara Chapman. 2022. Towards Efficient Remote OpenMP Offloading. In OpenMP in a Modern World: From Multi-device Support to Meta Programming, Michael Klemm, Bronis R. de Supinski, Jannis Klinkenberg, and Brandon Neth (Eds.). Springer International Publishing, Cham, 17--31."},{"key":"e_1_3_2_1_11_1","volume-title":"Minimod: A Finite Difference solver for Seismic Modeling. arXiv","author":"Meng Jie","year":"2020","unstructured":"Jie Meng , Andreas Atle , Henri Calandra , and Mauricio Araya-Polo . 2020 . Minimod: A Finite Difference solver for Seismic Modeling. arXiv (2020). arXiv:2007.06048 [cs.DC] https:\/\/arxiv.org\/abs\/2007.06048 Jie Meng, Andreas Atle, Henri Calandra, and Mauricio Araya-Polo. 2020. Minimod: A Finite Difference solver for Seismic Modeling. arXiv (2020). arXiv:2007.06048 [cs.DC] https:\/\/arxiv.org\/abs\/2007.06048"},{"key":"e_1_3_2_1_12_1","unstructured":"NVIDIA. [n.d.]. NVIDIA CUDA GPUDirect RDMA. https:\/\/docs.nvidia.com\/cuda\/gpudirect-rdma\/index.html.  NVIDIA. [n.d.]. NVIDIA CUDA GPUDirect RDMA. https:\/\/docs.nvidia.com\/cuda\/gpudirect-rdma\/index.html."},{"key":"e_1_3_2_1_13_1","unstructured":"NVIDIA. [n.d.]. NVIDIA Nsight Systems. https:\/\/developer.nvidia.com\/nsight-systems.  NVIDIA. [n.d.]. NVIDIA Nsight Systems. https:\/\/developer.nvidia.com\/nsight-systems."},{"key":"e_1_3_2_1_14_1","unstructured":"OpenMP Architecture Review Board. 2018. OpenMP Application Programming Interface. https:\/\/www.openmp.org\/wp-content\/uploads\/OpenMP-API-Specification-5.0.pdf Version 5.0.  OpenMP Architecture Review Board. 2018. OpenMP Application Programming Interface. https:\/\/www.openmp.org\/wp-content\/uploads\/OpenMP-API-Specification-5.0.pdf Version 5.0."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-07312-0_16"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ESPM254806.2021.00011"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ESPM254806.2021.00011"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3448290.3448559"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58144-2_5"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/2830013.2830015"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.anucene.2012.06.040"},{"key":"e_1_3_2_1_22_1","unstructured":"Ryuichi Sai John Mellor-Crummey Xiaozhu Meng Mauricio Araya-Polo and Jie Meng. 2020. Accelerating High-Order Stencils on GPUs. arXiv:2009.04619 [cs.DC]  Ryuichi Sai John Mellor-Crummey Xiaozhu Meng Mauricio Araya-Polo and Jie Meng. 2020. Accelerating High-Order Stencils on GPUs. arXiv:2009.04619 [cs.DC]"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/HOTI.2015.13"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.5555\/1789826.1789833"},{"key":"e_1_3_2_1_25_1","volume-title":"Languages and Compilers for Parallel Computing","author":"Tian Shilei","unstructured":"Shilei Tian , Johannes Doerfert , and Barbara Chapman . 2022. Concurrent Execution of Deferred OpenMP Target Tasks with Hidden Helper Threads . In Languages and Compilers for Parallel Computing , Barbara Chapman and Jos\u00e9 Moreira (Eds.). Springer International Publishing , Cham, 41--56. Shilei Tian, Johannes Doerfert, and Barbara Chapman. 2022. Concurrent Execution of Deferred OpenMP Target Tasks with Hidden Helper Threads. In Languages and Compilers for Parallel Computing, Barbara Chapman and Jos\u00e9 Moreira (Eds.). Springer International Publishing, Cham, 41--56."},{"key":"e_1_3_2_1_26_1","volume-title":"PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114","author":"Tramm John R","year":"2014","unstructured":"John R Tramm , Andrew R Siegel , Tanzima Islam , and Martin Schulz . 2014 . XS-Bench - The Development and Verification of a Performance Abstraction for Monte Carlo Reactor Analysis . In PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114 .pdf John R Tramm, Andrew R Siegel, Tanzima Islam, and Martin Schulz. 2014. XS-Bench - The Development and Verification of a Performance Abstraction for Monte Carlo Reactor Analysis. In PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114.pdf"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2021.3097283"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3226112"},{"key":"e_1_3_2_1_29_1","volume-title":"The OpenMP Cluster Programming Model. 51st International Conference on Parallel Processing Workshop Proceedings (ICPP Workshops 22)","author":"Yviquel Herv\u00e9","year":"2022","unstructured":"Herv\u00e9 Yviquel , Marcio Pereira , Em\u00edlio Francesquini , Guilherme Valarini , Pedro Rosso Gustavo Leite , Rodrigo Ceccato , Carla Cusihualpa , Vitoria Dias , Sandro Rigo , Alan Souza , and Guido Araujo . 2022 . The OpenMP Cluster Programming Model. 51st International Conference on Parallel Processing Workshop Proceedings (ICPP Workshops 22) (2022). Herv\u00e9 Yviquel, Marcio Pereira, Em\u00edlio Francesquini, Guilherme Valarini, Pedro Rosso Gustavo Leite, Rodrigo Ceccato, Carla Cusihualpa, Vitoria Dias, Sandro Rigo, Alan Souza, and Guido Araujo. 2022. The OpenMP Cluster Programming Model. 51st International Conference on Parallel Processing Workshop Proceedings (ICPP Workshops 22) (2022)."}],"event":{"name":"PMAM'23: 14th International Workshop on Programming Models and Applications for Multicores and Manycores","location":"Montreal QC Canada","acronym":"PMAM'23","sponsor":["SIGHPC ACM Special Interest Group on High Performance Computing","SIGPLAN ACM Special Interest Group on Programming Languages"]},"container-title":["Proceedings of the 14th International Workshop on Programming Models and Applications for Multicores and Manycores"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3582514.3582519","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:47:14Z","timestamp":1750178834000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3582514.3582519"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,2,25]]},"references-count":29,"alternative-id":["10.1145\/3582514.3582519","10.1145\/3582514"],"URL":"https:\/\/doi.org\/10.1145\/3582514.3582519","relation":{},"subject":[],"published":{"date-parts":[[2023,2,25]]},"assertion":[{"value":"2023-02-25","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}