{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,12]],"date-time":"2026-03-12T08:44:58Z","timestamp":1773305098949,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":51,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,1,7]],"date-time":"2022-01-07T00:00:00Z","timestamp":1641513600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100017156","name":"Swedish e-Science Research Centre","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100017156","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Vetenskapsr\u00e5det","award":["2018-05973"],"award-info":[{"award-number":["2018-05973"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,1,7]]},"DOI":"10.1145\/3492805.3492808","type":"proceedings-article","created":{"date-parts":[[2022,1,7]],"date-time":"2022-01-07T23:45:23Z","timestamp":1641599123000},"page":"125-136","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":9,"title":["A High-Fidelity Flow Solver for Unstructured Meshes on Field-Programmable Gate Arrays: Design, Evaluation, and Future Challenges"],"prefix":"10.1145","author":[{"given":"Martin","family":"Karp","sequence":"first","affiliation":[{"name":"KTH Royal Institute of Technology, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Artur","family":"Podobas","sequence":"additional","affiliation":[{"name":"KTH Royal Institute of Technology, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tobias","family":"Kenter","sequence":"additional","affiliation":[{"name":"Paderborn University, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Niclas","family":"Jansson","sequence":"additional","affiliation":[{"name":"KTH Royal Institute of Technology, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Christian","family":"Plessl","sequence":"additional","affiliation":[{"name":"Paderborn University, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Philipp","family":"Schlatter","sequence":"additional","affiliation":[{"name":"KTH Royal Institute of Technology, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Stefano","family":"Markidis","sequence":"additional","affiliation":[{"name":"KTH Royal Institute of Technology, Sweden"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,1,7]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"crossref","unstructured":"Richard Barrett Michael Berry Tony\u00a0F Chan James Demmel June Donato Jack Dongarra Victor Eijkhout Roldan Pozo Charles Romine and Henk Van\u00a0der Vorst. 1994. Templates for the solution of linear systems: building blocks for iterative methods. SIAM.  Richard Barrett Michael Berry Tony\u00a0F Chan James Demmel June Donato Jack Dongarra Victor Eijkhout Roldan Pozo Charles Romine and Henk Van\u00a0der Vorst. 1994. Templates for the solution of linear systems: building blocks for iterative methods. SIAM.","DOI":"10.1137\/1.9781611971538"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.sysarc.2014.07.001"},{"key":"e_1_3_2_1_3_1","volume-title":"FPGA Based Accelerators for Financial Applications","author":"Becker Tobias","unstructured":"Tobias Becker , Oskar Mencer , Stephen Weston , and Georgi Gaydadjiev . 2015. Maxeler data-flow in computational finance . In FPGA Based Accelerators for Financial Applications . Springer , 243\u2013266. Tobias Becker, Oskar Mencer, Stephen Weston, and Georgi Gaydadjiev. 2015. Maxeler data-flow in computational finance. In FPGA Based Accelerators for Financial Applications. Springer, 243\u2013266."},{"key":"e_1_3_2_1_4_1","volume-title":"Tal Ben-Nun, and Torsten Hoefler.","author":"Besta Maciej","year":"2019","unstructured":"Maciej Besta , Dimitri Stanojevic , Johannes De\u00a0Fine Licht , Tal Ben-Nun, and Torsten Hoefler. 2019 . Graph processing on fpgas: Taxonomy , survey, challenges. arXiv preprint arXiv:1903.06697(2019). Maciej Besta, Dimitri Stanojevic, Johannes De\u00a0Fine Licht, Tal Ben-Nun, and Torsten Hoefler. 2019. Graph processing on fpgas: Taxonomy, survey, challenges. arXiv preprint arXiv:1903.06697(2019)."},{"key":"e_1_3_2_1_5_1","volume-title":"FPGA Acceleration of Fluid-Flow Kernels. In 2020 IEEE\/ACM International Workshop on Heterogeneous High-performance Reconfigurable Computing (H2RC). IEEE, 29\u201337","author":"Blanchard Ryan","year":"2020","unstructured":"Ryan Blanchard , Greg Stitt , and Herman Lam . 2020 . FPGA Acceleration of Fluid-Flow Kernels. In 2020 IEEE\/ACM International Workshop on Heterogeneous High-performance Reconfigurable Computing (H2RC). IEEE, 29\u201337 . Ryan Blanchard, Greg Stitt, and Herman Lam. 2020. FPGA Acceleration of Fluid-Flow Kernels. In 2020 IEEE\/ACM International Workshop on Heterogeneous High-performance Reconfigurable Computing (H2RC). IEEE, 29\u201337."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2012.25"},{"key":"e_1_3_2_1_7_1","volume-title":"Rematerialization. In Proceedings of the ACM SIGPLAN 1992 conference on Programming language design and implementation. 311\u2013321","author":"Briggs Preston","year":"1992","unstructured":"Preston Briggs , Keith\u00a0 D Cooper , and Linda Torczon . 1992 . Rematerialization. In Proceedings of the ACM SIGPLAN 1992 conference on Programming language design and implementation. 311\u2013321 . Preston Briggs, Keith\u00a0D Cooper, and Linda Torczon. 1992. Rematerialization. In Proceedings of the ACM SIGPLAN 1992 conference on Programming language design and implementation. 311\u2013321."},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/H2RC51942.2020.00008"},{"key":"e_1_3_2_1_9_1","volume-title":"https:\/\/github.com\/cass-support\/clfortran Accessed","author":"Butrashvily Mordechai","year":"2021","unstructured":"Mordechai Butrashvily . [n.d.]. CLFORTRAN. https:\/\/github.com\/cass-support\/clfortran Accessed : Aug. 27, 2021 . Mordechai Butrashvily. [n.d.]. CLFORTRAN. https:\/\/github.com\/cass-support\/clfortran Accessed: Aug. 27, 2021."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1950413.1950423"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2009.5272454"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2012.6339272"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2020.3039409"},{"key":"e_1_3_2_1_14_1","volume-title":"High-order methods for incompressible fluid flow. Vol.\u00a09","author":"Deville O","unstructured":"Michel\u00a0 O Deville , Paul\u00a0 F Fischer , Paul\u00a0 F Fischer , EH Mund , 2002. High-order methods for incompressible fluid flow. Vol.\u00a09 . Cambridge university press . Michel\u00a0O Deville, Paul\u00a0F Fischer, Paul\u00a0F Fischer, EH Mund, 2002. High-order methods for incompressible fluid flow. Vol.\u00a09. Cambridge university press."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2612669.2612694"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/2693656"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"crossref","unstructured":"Paul Fischer Stefan Kerkemeier Misun Min Yu-Hsiang Lan Malachi Phillips Thilina Rathnayake Elia Merzari Ananias Tomboulides Ali Karakus Noel Chalmers and Tim Warburton. 2021. NekRS a GPU-Accelerated Spectral Element Navier-Stokes Solver. arxiv:2104.05829\u00a0[cs.PF]  Paul Fischer Stefan Kerkemeier Misun Min Yu-Hsiang Lan Malachi Phillips Thilina Rathnayake Elia Merzari Ananias Tomboulides Ali Karakus Noel Chalmers and Tim Warburton. 2021. NekRS a GPU-Accelerated Spectral Element Navier-Stokes Solver. arxiv:2104.05829\u00a0[cs.PF]","DOI":"10.1016\/j.parco.2022.102982"},{"key":"e_1_3_2_1_18_1","unstructured":"Paul\u00a0F Fischer James\u00a0W Lottes and Stefan\u00a0G Kerkemeier. 2008. nek5000 Web page.  Paul\u00a0F Fischer James\u00a0W Lottes and Stefan\u00a0G Kerkemeier. 2008. nek5000 Web page."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2016.7577352"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cosrev.2018.11.002"},{"key":"e_1_3_2_1_21_1","volume-title":"Proceedings of Machine Learning and Systems 3","author":"Ivanov Andrei","year":"2021","unstructured":"Andrei Ivanov , Nikoli Dryden , Tal Ben-Nun , Shigang Li , and Torsten Hoefler . 2021 . Data Movement Is All You Need: A Case Study on Optimizing Transformers . Proceedings of Machine Learning and Systems 3 (2021). Andrei Ivanov, Nikoli Dryden, Tal Ben-Nun, Shigang Li, and Torsten Hoefler. 2021. Data Movement Is All You Need: A Case Study on Optimizing Transformers. Proceedings of Machine Learning and Systems 3 (2021)."},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL50879.2020.00031"},{"key":"e_1_3_2_1_23_1","volume-title":"Neko: A Modern, Portable, and Scalable Framework for High-Fidelity Computational Fluid Dynamics. arxiv:2107.01243\u00a0[cs.MS]","author":"Jansson Niclas","year":"2021","unstructured":"Niclas Jansson , Martin Karp , Artur Podobas , Stefano Markidis , and Philipp Schlatter . 2021 . Neko: A Modern, Portable, and Scalable Framework for High-Fidelity Computational Fluid Dynamics. arxiv:2107.01243\u00a0[cs.MS] Niclas Jansson, Martin Karp, Artur Podobas, Stefano Markidis, and Philipp Schlatter. 2021. Neko: A Modern, Portable, and Scalable Framework for High-Fidelity Computational Fluid Dynamics. arxiv:2107.01243\u00a0[cs.MS]"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/800076.802486"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2018.032271057"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS49936.2021.00116"},{"key":"e_1_3_2_1_27_1","volume-title":"Parmetis: Parallel graph partitioning and sparse matrix ordering library.","author":"Karypis George","year":"1997","unstructured":"George Karypis , Kirk Schloegel , and Vipin Kumar . 1997 . Parmetis: Parallel graph partitioning and sparse matrix ordering library. (1997). George Karypis, Kirk Schloegel, and Vipin Kumar. 1997. Parmetis: Parallel graph partitioning and sparse matrix ordering library. (1997)."},{"key":"e_1_3_2_1_28_1","volume-title":"Proc. Int. Workshop on FPGAs for Software Programmers (FSP), collocated with Int. Conf. on Field Programmable Logic and Applications (FPL).","author":"Kenter Tobias","year":"2019","unstructured":"Tobias Kenter . 2019 . Invited Tutorial: OpenCL design flows for Intel and Xilinx FPGAs: Using common design patterns and dealing with vendor-specific differences . In Proc. Int. Workshop on FPGAs for Software Programmers (FSP), collocated with Int. Conf. on Field Programmable Logic and Applications (FPL). Tobias Kenter. 2019. Invited Tutorial: OpenCL design flows for Intel and Xilinx FPGAs: Using common design patterns and dealing with vendor-specific differences. In Proc. Int. Workshop on FPGAs for Software Programmers (FSP), collocated with Int. Conf. on Field Programmable Logic and Applications (FPL)."},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2018.00037"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3468267.3470617"},{"key":"e_1_3_2_1_31_1","volume-title":"FPGA architecture: Survey and challenges","author":"Kuon Ian","unstructured":"Ian Kuon , Russell Tessier , and Jonathan Rose . 2008. FPGA architecture: Survey and challenges . Now Publishers Inc . Ian Kuon, Russell Tessier, and Jonathan Rose. 2008. FPGA architecture: Survey and challenges. Now Publishers Inc."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3409964.3461796"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/3295500.3356181"},{"key":"e_1_3_2_1_34_1","volume-title":"International Workshop on Performance Modeling, Benchmarking and Simulation of High Performance Computer Systems. Springer, 172\u2013192","author":"Marjanovi\u0107 Vladimir","year":"2014","unstructured":"Vladimir Marjanovi\u0107 , Jos\u00e9 Gracia , and Colin\u00a0 W Glass . 2014 . Performance modeling of the HPCG benchmark . In International Workshop on Performance Modeling, Benchmarking and Simulation of High Performance Computer Systems. Springer, 172\u2013192 . Vladimir Marjanovi\u0107, Jos\u00e9 Gracia, and Colin\u00a0W Glass. 2014. Performance modeling of the HPCG benchmark. In International Workshop on Performance Modeling, Benchmarking and Simulation of High Performance Computer Systems. Springer, 172\u2013192."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/H2RC51942.2020.00007"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2012.6339276"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/SASP.2009.5226333"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2013.6645550"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/SAMOS.2016.7818354"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPT.2017.8280154"},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.3012084"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASAP49362.2020.00010"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICESS.2019.8782524"},{"key":"e_1_3_2_1_44_1","volume-title":"High-Performance Computing Using FPGAs","author":"Sano Kentaro","unstructured":"Kentaro Sano . 2013. FPGA-based systolic computational-memory array for scalable stencil computations . In High-Performance Computing Using FPGAs . Springer , 279\u2013303. Kentaro Sano. 2013. FPGA-based systolic computational-memory array for scalable stencil computations. In High-Performance Computing Using FPGAs. Springer, 279\u2013303."},{"key":"e_1_3_2_1_45_1","unstructured":"Catherine\u00a0D Schuman Thomas\u00a0E Potok Robert\u00a0M Patton J\u00a0Douglas Birdwell Mark\u00a0E Dean Garrett\u00a0S Rose and James\u00a0S Plank. 2017. A survey of neuromorphic computing and neural networks in hardware. arXiv preprint arXiv:1705.06963(2017).  Catherine\u00a0D Schuman Thomas\u00a0E Potok Robert\u00a0M Patton J\u00a0Douglas Birdwell Mark\u00a0E Dean Garrett\u00a0S Rose and James\u00a0S Plank. 2017. A survey of neuromorphic computing and neural networks in hardware. arXiv preprint arXiv:1705.06963(2017)."},{"key":"e_1_3_2_1_46_1","unstructured":"Jeffrey\u00a0P Slotnick Abdollah Khodadoust Juan Alonso David Darmofal William Gropp Elizabeth Lurie and Dimitri\u00a0J Mavriplis. 2014. CFD vision 2030 study: a path to revolutionary computational aerosciences. (2014).  Jeffrey\u00a0P Slotnick Abdollah Khodadoust Juan Alonso David Darmofal William Gropp Elizabeth Lurie and Dimitri\u00a0J Mavriplis. 2014. CFD vision 2030 study: a path to revolutionary computational aerosciences. (2014)."},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2017.3211127"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1038\/530144a"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.micpro.2014.03.007"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/H2RC49586.2019.00007"},{"key":"e_1_3_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/3174243.3174248"}],"event":{"name":"HPC Asia2022: International Conference on High Performance Computing in Asia-Pacific Region","location":"Virtual Event Japan","acronym":"HPC Asia2022"},"container-title":["International Conference on High Performance Computing in Asia-Pacific Region"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3492805.3492808","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3492805.3492808","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:48:26Z","timestamp":1750193306000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3492805.3492808"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,1,7]]},"references-count":51,"alternative-id":["10.1145\/3492805.3492808","10.1145\/3492805"],"URL":"https:\/\/doi.org\/10.1145\/3492805.3492808","relation":{},"subject":[],"published":{"date-parts":[[2022,1,7]]},"assertion":[{"value":"2022-01-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}