{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,31]],"date-time":"2025-12-31T14:46:54Z","timestamp":1767192414702,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":27,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,7,5]],"date-time":"2021-07-05T00:00:00Z","timestamp":1625443200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100000015","name":"U.S. Department of Energy","doi-asserted-by":"publisher","award":["DE-SC0012704"],"award-info":[{"award-number":["DE-SC0012704"]}],"id":[{"id":"10.13039\/100000015","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,7,5]]},"DOI":"10.1145\/3468267.3470613","type":"proceedings-article","created":{"date-parts":[[2021,8,26]],"date-time":"2021-08-26T16:07:54Z","timestamp":1629994074000},"page":"1-11","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["Solving DWF dirac equation using multi-splitting preconditioned conjugate gradient with tensor cores on NVIDIA GPUs"],"prefix":"10.1145","author":[{"given":"Jiqun","family":"Tu","sequence":"first","affiliation":[{"name":"NVIDIA Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"M. A.","family":"Clark","sequence":"additional","affiliation":[{"name":"NVIDIA Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chulwoo","family":"Jung","sequence":"additional","affiliation":[{"name":"Brookhaven National Laboratory"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Robert D.","family":"Mawhinney","sequence":"additional","affiliation":[{"name":"Columbia University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,8,26]]},"reference":[{"doi-asserted-by":"publisher","key":"e_1_3_2_1_1_1","DOI":"10.1103\/PhysRevLett.105.201602"},{"doi-asserted-by":"crossref","unstructured":"R Babich M A Clark B Joo G Shi R C Brower and S Gottlieb. 2011. Scaling Lattice QCD beyond 100 GPUs. (2011). arXiv:1109.2935  R Babich M A Clark B Joo G Shi R C Brower and S Gottlieb. 2011. Scaling Lattice QCD beyond 100 GPUs. (2011). arXiv:1109.2935","key":"e_1_3_2_1_2_1","DOI":"10.1145\/2063384.2063478"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_3_1","DOI":"10.1103\/PhysRevLett.100.041601"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_4_1","DOI":"10.1016\/j.nuclphysbps.2006.01.047"},{"key":"e_1_3_2_1_5_1","volume-title":"Weinberg","author":"Brower Richard C.","year":"2020","unstructured":"Richard C. Brower , M.A. Clark , Dean Howarth , and Evan S . Weinberg . 2020 . Multigrid for Chiral Lattice Fermions : Domain Wall . (4 2020). arXiv:2004.07732 [heplat] Richard C. Brower, M.A. Clark, Dean Howarth, and Evan S. Weinberg. 2020. Multigrid for Chiral Lattice Fermions: Domain Wall. (4 2020). arXiv:2004.07732 [heplat]"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_6_1","DOI":"10.1016\/j.nuclphysbps.2004.11.180"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_7_1","DOI":"10.1016\/j.cpc.2010.05.002"},{"key":"e_1_3_2_1_8_1","volume-title":"A 2020 Pre-Exascale GPU-accelerated System for NERSC. Architecture and Early Application Performance Optimization Results. Paper presented at the meeting of GPU Technology Conference 2019","author":"Deslippe Jack","year":"2019","unstructured":"Jack Deslippe . [n.d.]. Perlmutter - A 2020 Pre-Exascale GPU-accelerated System for NERSC. Architecture and Early Application Performance Optimization Results. Paper presented at the meeting of GPU Technology Conference 2019 , San Jose CA. https:\/\/on-demand.gputechconf.com\/supercomputing\/ 2019 \/pdf\/sc1919-perlmutter-a-2020-pre-exascale-gpu-accelerated-system-for-nersc.-architecture-and-early-application-performance-optimization-results.pdf Jack Deslippe. [n.d.]. Perlmutter - A 2020 Pre-Exascale GPU-accelerated System for NERSC. Architecture and Early Application Performance Optimization Results. Paper presented at the meeting of GPU Technology Conference 2019, San Jose CA. https:\/\/on-demand.gputechconf.com\/supercomputing\/2019\/pdf\/sc1919-perlmutter-a-2020-pre-exascale-gpu-accelerated-system-for-nersc.-architecture-and-early-application-performance-optimization-results.pdf"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_9_1","DOI":"10.1016\/0370-2693(87)91197-X"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_10_1","DOI":"10.1142\/S012918319400115X"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_11_1","DOI":"10.1137\/s1064827597323415"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_12_1","DOI":"10.6028\/jres.049.044"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_14_1","DOI":"10.1007\/s11227-011-0563-y"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_15_1","DOI":"10.1016\/S0010-4655(03)00486-7"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_16_1","DOI":"10.1016\/j.cpc.2004.10.004"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_18_1","DOI":"10.1103\/PhysRevD.102.034507"},{"unstructured":"Nolan Miller et al. 2020. Scale setting the M{\\\"o}bius Domain Wall Fermion on gradient-flowed HISQ action using the omega baryon mass and the gradient-flow scale w0. (11 2020). arXiv:2011.12166 [hep-lat]  Nolan Miller et al. 2020. Scale setting the M{\\\"o}bius Domain Wall Fermion on gradient-flowed HISQ action using the omega baryon mass and the gradient-flow scale w 0 . (11 2020). arXiv:2011.12166 [hep-lat]","key":"e_1_3_2_1_19_1"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_20_1","DOI":"10.1137\/0606062"},{"unstructured":"Yusuke Osaki and Ken-ichi Ishikawa. 2010. Domain Decomposition method on GPU cluster. (2010) 1--7. arXiv:1011.3318  Yusuke Osaki and Ken-ichi Ishikawa. 2010. Domain Decomposition method on GPU cluster. (2010) 1--7. arXiv:1011.3318","key":"e_1_3_2_1_21_1"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_22_1","DOI":"10.1137\/0717059"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_23_1","DOI":"10.1103\/PhysRevD.91.114511"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_24_1","DOI":"10.1137\/080725532"},{"key":"e_1_3_2_1_25_1","volume-title":"Proceedings of the 30th International Conference on Machine Learning (Proceedings of Machine Learning Research","volume":"1147","author":"Sutskever Ilya","year":"2013","unstructured":"Ilya Sutskever , James Martens , George Dahl , and Geoffrey Hinton . 2013 . On the importance of initialization and momentum in deep learning . In Proceedings of the 30th International Conference on Machine Learning (Proceedings of Machine Learning Research , Vol. 28), Sanjoy Dasgupta and David McAllester (Eds.). PMLR, Atlanta, Georgia, USA, 1139-- 1147 . http:\/\/proceedings.mlr.press\/v28\/sutskever13.html Ilya Sutskever, James Martens, George Dahl, and Geoffrey Hinton. 2013. On the importance of initialization and momentum in deep learning. In Proceedings of the 30th International Conference on Machine Learning (Proceedings of Machine Learning Research, Vol. 28), Sanjoy Dasgupta and David McAllester (Eds.). PMLR, Atlanta, Georgia, USA, 1139--1147. http:\/\/proceedings.mlr.press\/v28\/sutskever13.html"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_26_1","DOI":"10.1137\/0913035"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_27_1","DOI":"10.1137\/s1064827599353865"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_28_1","DOI":"10.22323\/1.256.0374"},{"doi-asserted-by":"publisher","key":"e_1_3_2_1_29_1","DOI":"10.22323\/1.139.0051"}],"event":{"sponsor":["SIGHPC ACM Special Interest Group on High Performance Computing, Special Interest Group on High Performance Computing","CSCS Swiss National Supercomputing Centre"],"acronym":"PASC '21","name":"PASC '21: Platform for Advanced Scientific Computing Conference","location":"Geneva Switzerland"},"container-title":["Proceedings of the Platform for Advanced Scientific Computing Conference"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3468267.3470613","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3468267.3470613","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3468267.3470613","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:17:18Z","timestamp":1750191438000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3468267.3470613"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,7,5]]},"references-count":27,"alternative-id":["10.1145\/3468267.3470613","10.1145\/3468267"],"URL":"https:\/\/doi.org\/10.1145\/3468267.3470613","relation":{},"subject":[],"published":{"date-parts":[[2021,7,5]]},"assertion":[{"value":"2021-08-26","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}