{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:10:02Z","timestamp":1750306202487,"version":"3.41.0"},"reference-count":25,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2017,4,6]],"date-time":"2017-04-06T00:00:00Z","timestamp":1491436800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"European Union Horizon 2020 Research and Innovation Programme","award":["671653"],"award-info":[{"award-number":["671653"]}]},{"name":"HiPEAC NoE"},{"name":"Maxeler University Programme, Altera, UK EPSRC","award":["EP\/I012036\/1, EP\/L00058X\/1 and EP\/N031768\/1"],"award-info":[{"award-number":["EP\/I012036\/1, EP\/L00058X\/1 and EP\/N031768\/1"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Reconfigurable Technol. Syst."],"published-print":{"date-parts":[[2017,6,30]]},"abstract":"<jats:p>The Finite Element Method (FEM) is a common numerical technique used for solving Partial Differential Equations on large and unstructured domain geometries. Numerical methods for FEM typically use algorithms and data structures which exhibit an unstructured memory access pattern. This makes acceleration of FEM on Field-Programmable Gate Arrays using an efficient, deeply pipelined architecture particularly challenging. In this work, we focus on implementing and optimising a vector assembly operation which, in the context of FEM, induces the unstructured memory access. We propose a dataflow architecture, graph-based theoretical model, and design flow for optimising the assembly operation for spectral\/hp finite element method on reconfigurable accelerators. We evaluate the proposed approach on two benchmark meshes and show that the graph-theoretic method of generating a static data access schedule results in a significant improvement in resource utilisation compared to prior work. This enables supporting larger FEM meshes on FPGA than previously possible.<\/jats:p>","DOI":"10.1145\/3024064","type":"journal-article","created":{"date-parts":[[2017,4,7]],"date-time":"2017-04-07T12:30:24Z","timestamp":1491568224000},"page":"1-22","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Efficient Assembly for High-Order Unstructured FEM Meshes (FPL 2015)"],"prefix":"10.1145","volume":"10","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-3055-5037","authenticated-orcid":false,"given":"Pavel","family":"Burovskiy","sequence":"first","affiliation":[{"name":"Maxeler Technologies, Albion Pl, London"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Paul","family":"Grigoras","sequence":"additional","affiliation":[{"name":"Imperial College London, Queen's Gate, London"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Spencer","family":"Sherwin","sequence":"additional","affiliation":[{"name":"Imperial College London, South Kensington Campus, London"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wayne","family":"Luk","sequence":"additional","affiliation":[{"name":"Imperial College London, Queen's Gate, London"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2017,4,6]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2015.7293749"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cpc.2015.02.008"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jcp.2013.10.019"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2014.6927464"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.2514\/2.1"},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1137\/0720013"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/800195.805928"},{"volume-title":"van der Vorst","year":"1993","author":"Demmel James W.","key":"e_1_2_1_8_1"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cpc.2007.11.014"},{"volume-title":"Rods and plates","series-title":"Series solution of some problems in elastic equilibrium of rods and plates","author":"Galerkin Boris G.","key":"e_1_2_1_10_1"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/2847263.2847338"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/FPL.2016.7577352"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2015.30"},{"volume-title":"Proceedings of the International Conference on Field Programmable Logic and Applications (FPL\u201908)","year":"2008","author":"Hu Jing","key":"e_1_2_1_14_1"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1007\/s40194-015-0235-2"},{"key":"e_1_2_1_16_1","unstructured":"George Karniadakis and Spencer Sherwin. 2013. Spectral\/hp Element Methods for Computational Fluid Dynamics. Oxford University Press.  George Karniadakis and Spencer Sherwin. 2013. Spectral\/hp Element Methods for Computational Fluid Dynamics. Oxford University Press."},{"volume-title":"Proceedings of the COMSOL Multiphysics User\u2019s Conference.","year":"2005","author":"Lienhart Gerhard","key":"e_1_2_1_17_1"},{"volume-title":"Taylor","year":"2015","author":"Lombard Jean-Eloi W.","key":"e_1_2_1_18_1"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1002\/fld.3648"},{"volume-title":"Unstructured Mesh Fluid Dynamics Using the Flux Reconstruction Method on an FPGA Dataflow Engine. Master\u2019s thesis","author":"Piechotka Maciej","key":"e_1_2_1_20_1"},{"volume-title":"Fix","year":"1973","author":"Strang Gilbert","key":"e_1_2_1_22_1"},{"volume-title":"Sparse Matrix Vector Multiplication on a Field Programmable Gate Array. Master\u2019s thesis","author":"van der Veen Marcel","key":"e_1_2_1_23_1"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10915-010-9420-z"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jcp.2010.03.031"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSII.2013.2278111"}],"container-title":["ACM Transactions on Reconfigurable Technology and Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3024064","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3024064","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:50:31Z","timestamp":1750218631000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3024064"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,4,6]]},"references-count":25,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2017,6,30]]}},"alternative-id":["10.1145\/3024064"],"URL":"https:\/\/doi.org\/10.1145\/3024064","relation":{},"ISSN":["1936-7406","1936-7414"],"issn-type":[{"type":"print","value":"1936-7406"},{"type":"electronic","value":"1936-7414"}],"subject":[],"published":{"date-parts":[[2017,4,6]]},"assertion":[{"value":"2016-04-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2016-11-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-04-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}