{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,24]],"date-time":"2026-02-24T18:14:54Z","timestamp":1771956894496,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":16,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,8,9]],"date-time":"2021-08-09T00:00:00Z","timestamp":1628467200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,8,9]]},"DOI":"10.1145\/3458744.3473356","type":"proceedings-article","created":{"date-parts":[[2021,9,23]],"date-time":"2021-09-23T16:38:30Z","timestamp":1632415110000},"page":"1-7","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":13,"title":["A Virtual GPU as Developer-Friendly OpenMP Offload Target"],"prefix":"10.1145","author":[{"given":"Atmn","family":"Patel","sequence":"first","affiliation":[{"name":"University of Waterloo, Canada"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shilei","family":"Tian","sequence":"additional","affiliation":[{"name":"Stony Brook University, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Johannes","family":"Doerfert","sequence":"additional","affiliation":[{"name":"Argonne National Laboratory, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Barbara","family":"Chapman","sequence":"additional","affiliation":[{"name":"Stony Brook University, United States of America"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,9,23]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Offloading Support for OpenMP in Clang and LLVM. In The Workshop on the LLVM Compiler Infrastructure in HPC (LLVM-HPC)","author":"Antao F.","year":"2016","unstructured":"Samuel\u00a0 F. Antao , Alexey Bataev , Arpith\u00a0 C. Jacob , Gheorghe-Teodor Bercea , Alexandre\u00a0 E. Eichenberger , Georgios Rokos , Matt Martineau , Tian Jin , Guray Ozen , Zehra Sura , Tong Chen , Hyojin Sung , Carlo Bertolli , and Kevin O\u2019Brien . 2016 . Offloading Support for OpenMP in Clang and LLVM. In The Workshop on the LLVM Compiler Infrastructure in HPC (LLVM-HPC) . Salt Lake City, UT, USA, 1\u201311. Samuel\u00a0F. Antao, Alexey Bataev, Arpith\u00a0C. Jacob, Gheorghe-Teodor Bercea, Alexandre\u00a0E. Eichenberger, Georgios Rokos, Matt Martineau, Tian Jin, Guray Ozen, Zehra Sura, Tong Chen, Hyojin Sung, Carlo Bertolli, and Kevin O\u2019Brien. 2016. Offloading Support for OpenMP in Clang and LLVM. In The Workshop on the LLVM Compiler Infrastructure in HPC (LLVM-HPC). Salt Lake City, UT, USA, 1\u201311."},{"key":"e_1_3_2_1_2_1","unstructured":"OpenMP Architecture\u00a0Review Board. 2020. OpenMP Application Programming Interface Version 5.1. https:\/\/www.openmp.org\/spec-html\/5.1\/openmp.html  OpenMP Architecture\u00a0Review Board. 2020. OpenMP Application Programming Interface Version 5.1. https:\/\/www.openmp.org\/spec-html\/5.1\/openmp.html"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2009.5306797"},{"key":"e_1_3_2_1_4_1","unstructured":"Fujitsu. 2021. Supercomputer Fugaki. https:\/\/www.fujitsu.com\/global\/about\/innovation\/fugaku\/  Fujitsu. 2021. Supercomputer Fugaki. https:\/\/www.fujitsu.com\/global\/about\/innovation\/fugaku\/"},{"key":"e_1_3_2_1_5_1","volume-title":"GDB: The GNU Project Debugger. https:\/\/www.gnu.org\/software\/gdb\/","author":"GNU.","year":"2021","unstructured":"GNU. 2021 . GDB: The GNU Project Debugger. https:\/\/www.gnu.org\/software\/gdb\/ GNU. 2021. GDB: The GNU Project Debugger. https:\/\/www.gnu.org\/software\/gdb\/"},{"key":"e_1_3_2_1_6_1","unstructured":"GNU. 2021. libstdc++ - GCC. https:\/\/gcc.gnu.org\/onlinedocs\/libstdc++\/manual\/status.html#status.iso.2020  GNU. 2021. libstdc++ - GCC. https:\/\/gcc.gnu.org\/onlinedocs\/libstdc++\/manual\/status.html#status.iso.2020"},{"key":"e_1_3_2_1_7_1","unstructured":"GNU. 2021. The LLDB Debugger. https:\/\/lldb.llvm.org\/  GNU. 2021. The LLDB Debugger. https:\/\/lldb.llvm.org\/"},{"key":"e_1_3_2_1_8_1","unstructured":"LLVM\u00a0Developer Group. 2021. OpenMP Support \u2014 Clang 13 documentation. https:\/\/clang.llvm.org\/docs\/OpenMPSupport.html  LLVM\u00a0Developer Group. 2021. OpenMP Support \u2014 Clang 13 documentation. https:\/\/clang.llvm.org\/docs\/OpenMPSupport.html"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10766-014-0320-y"},{"key":"e_1_3_2_1_10_1","unstructured":"LLVM. 2021. \u201clibc++\u201d C++ Standard Library. https:\/\/libcxx.llvm.org\/  LLVM. 2021. \u201clibc++\u201d C++ Standard Library. https:\/\/libcxx.llvm.org\/"},{"key":"e_1_3_2_1_11_1","unstructured":"Justin Luitjens. 2014. Faster Parallel Reductions on Kepler. https:\/\/developer.nvidia.com\/blog\/faster-parallel-reductions-kepler\/  Justin Luitjens. 2014. Faster Parallel Reductions on Kepler. https:\/\/developer.nvidia.com\/blog\/faster-parallel-reductions-kepler\/"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-85262-7_11"},{"key":"e_1_3_2_1_13_1","volume-title":"Concurrent Execution of Deferred OpenMP Target Tasks with Hidden Helper Threads. In The Workshop on Languages and Compilers for Parallel Computing (LCPC)","author":"Tian Shilei","year":"2020","unstructured":"Shilei Tian , Johannes Doerfert , and Barbara Chapman . 2020 . Concurrent Execution of Deferred OpenMP Target Tasks with Hidden Helper Threads. In The Workshop on Languages and Compilers for Parallel Computing (LCPC) . Stony Brook, NY, USA. Shilei Tian, Johannes Doerfert, and Barbara Chapman. 2020. Concurrent Execution of Deferred OpenMP Target Tasks with Hidden Helper Threads. In The Workshop on Languages and Compilers for Parallel Computing (LCPC). Stony Brook, NY, USA."},{"key":"e_1_3_2_1_14_1","volume-title":"EASC 2014 - Solving Software Challenges for Exascale","author":"Tramm R.","year":"2014","unstructured":"John\u00a0 R. Tramm , Andrew\u00a0 R. Siegel , Benoit Forget , and Colin Josey . 2014 . Performance Analysis of a Reduced Data Movement Algorithm for Neutron Cross Section Data in Monte Carlo Simulations . In EASC 2014 - Solving Software Challenges for Exascale . Stockholm. https:\/\/doi.org\/10.1007\/978-3-319-15976-8_3 10.1007\/978-3-319-15976-8_3 John\u00a0R. Tramm, Andrew\u00a0R. Siegel, Benoit Forget, and Colin Josey. 2014. Performance Analysis of a Reduced Data Movement Algorithm for Neutron Cross Section Data in Monte Carlo Simulations. In EASC 2014 - Solving Software Challenges for Exascale. Stockholm. https:\/\/doi.org\/10.1007\/978-3-319-15976-8_3"},{"key":"e_1_3_2_1_15_1","volume-title":"PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114","author":"Tramm R","year":"2014","unstructured":"John\u00a0 R Tramm , Andrew\u00a0 R Siegel , Tanzima Islam , and Martin Schulz . 2014 . XSBench - The Development and Verification of a Performance Abstraction for Monte Carlo Reactor Analysis . In PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114 .pdf John\u00a0R Tramm, Andrew\u00a0R Siegel, Tanzima Islam, and Martin Schulz. 2014. XSBench - The Development and Verification of a Performance Abstraction for Monte Carlo Reactor Analysis. In PHYSOR 2014 - The Role of Reactor Physics toward a Sustainable Future. Kyoto. https:\/\/www.mcs.anl.gov\/papers\/P5064-0114.pdf"},{"key":"e_1_3_2_1_16_1","volume-title":"Supporting Data Shuffle Between Threads in OpenMP","author":"Wang Anjia","unstructured":"Anjia Wang , Xinyao Yi , and Yonghong Yan . 2020. Supporting Data Shuffle Between Threads in OpenMP . In OpenMP: Portable Multi-Level Parallelism on Modern Systems, Kent Milfeld, Bronis\u00a0R. de\u00a0Supinski, Lars Koesterke, and Jannis Klinkenberg (Eds.). Springer International Publishing , Cham, 98\u2013112. Anjia Wang, Xinyao Yi, and Yonghong Yan. 2020. Supporting Data Shuffle Between Threads in OpenMP. In OpenMP: Portable Multi-Level Parallelism on Modern Systems, Kent Milfeld, Bronis\u00a0R. de\u00a0Supinski, Lars Koesterke, and Jannis Klinkenberg (Eds.). Springer International Publishing, Cham, 98\u2013112."}],"event":{"name":"ICPP 2021: 50th International Conference on Parallel Processing","location":"Lemont IL USA","acronym":"ICPP 2021"},"container-title":["50th International Conference on Parallel Processing Workshop"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3458744.3473356","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3458744.3473356","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T17:49:06Z","timestamp":1750268946000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3458744.3473356"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,8,9]]},"references-count":16,"alternative-id":["10.1145\/3458744.3473356","10.1145\/3458744"],"URL":"https:\/\/doi.org\/10.1145\/3458744.3473356","relation":{},"subject":[],"published":{"date-parts":[[2021,8,9]]},"assertion":[{"value":"2021-09-23","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}