{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,24]],"date-time":"2026-01-24T23:02:03Z","timestamp":1769295723424,"version":"3.49.0"},"reference-count":45,"publisher":"Wiley","issue":"2","license":[{"start":{"date-parts":[[2026,1,18]],"date-time":"2026-01-18T00:00:00Z","timestamp":1768694400000},"content-version":"vor","delay-in-days":17,"URL":"http:\/\/onlinelibrary.wiley.com\/termsAndConditions#vor"},{"start":{"date-parts":[[2026,1,1]],"date-time":"2026-01-01T00:00:00Z","timestamp":1767225600000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/doi.wiley.com\/10.1002\/tdm_license_1.1"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61872422"],"award-info":[{"award-number":["61872422"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["onlinelibrary.wiley.com"],"crossmark-restriction":true},"short-container-title":["Concurrency and Computation"],"published-print":{"date-parts":[[2026,1]]},"abstract":"<jats:title>ABSTRACT<\/jats:title>\n                  <jats:p>The factorized sparse approximate inverse (FSAI) preconditioner has been proven to be effective in accelerating the convergence of iterative methods. Due to the high cost of constructing the FSAI preconditioner, accelerating it on graphics processing unit (GPU) has attracted considerable attention. However, despite the development of some existing FSAI preconditioning algorithms on GPU, their performance will significantly decrease when they encounter matrix types that are not suitable for them. This motivates us to investigate how to design an effective FSAI preconditioning algorithm on GPU. In this paper, we propose an adaptive FSAI preconditioning algorithm on GPU, called GFSAI\u2010Adaptive, to address the above problem. In GFSAI\u2010Adaptive, first, two adaptive thread allocation strategies are proposed for two special types of SPD matrices to ensure that the allocated threads can be fully utilized. Second, based on the proposed two thread allocation strategies, two FSAI kernels, called GFSAII and GFSAIII, are presented. Third, we construct a new graph convolutional network, and thus propose a search engine to select the optimal kernel from GFSAII and GFSAIII for matrices that do not belong to two special types based on it. Experimental results show that our proposed GFSAI\u2010Adaptive is effective and outperforms a popular preconditioning algorithm in the public CUSPARSE library and a recent parallel static FSAI preconditioning algorithm on GPU.<\/jats:p>","DOI":"10.1002\/cpe.70552","type":"journal-article","created":{"date-parts":[[2026,1,19]],"date-time":"2026-01-19T06:10:52Z","timestamp":1768803052000},"update-policy":"https:\/\/doi.org\/10.1002\/crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["<scp>GFSAI<\/scp>\n                    : An Adaptive Factorized Sparse Approximate Inverse Preconditioning Algorithm on\n                    <scp>GPU<\/scp>"],"prefix":"10.1002","volume":"38","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-6168-5118","authenticated-orcid":false,"given":"Yizhou","family":"Wang","sequence":"first","affiliation":[{"name":"Ministry of Education Key Lab for NSLSCS, School of Computer and Electronic Information Nanjing Normal University  Nanjing China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yige","family":"Zhang","sequence":"additional","affiliation":[{"name":"Ministry of Education Key Lab for NSLSCS, School of Computer and Electronic Information Nanjing Normal University  Nanjing China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiaquan","family":"Gao","sequence":"additional","affiliation":[{"name":"Ministry of Education Key Lab for NSLSCS, School of Computer and Electronic Information Nanjing Normal University  Nanjing China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"311","published-online":{"date-parts":[[2026,1,18]]},"reference":[{"key":"e_1_2_9_2_1","doi-asserted-by":"publisher","DOI":"10.1137\/0913035"},{"key":"e_1_2_9_3_1","doi-asserted-by":"publisher","DOI":"10.1137\/0907058"},{"key":"e_1_2_9_4_1","doi-asserted-by":"publisher","DOI":"10.1137\/S1064827599363976"},{"key":"e_1_2_9_5_1","doi-asserted-by":"publisher","DOI":"10.1137\/030601880"},{"key":"e_1_2_9_6_1","doi-asserted-by":"publisher","DOI":"10.1137\/1.9780898718003"},{"key":"e_1_2_9_7_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2013.10.002"},{"key":"e_1_2_9_8_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2017.05.006"},{"key":"e_1_2_9_9_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2016.06.004"},{"key":"e_1_2_9_10_1","doi-asserted-by":"publisher","DOI":"10.1137\/S1064827594271421"},{"key":"e_1_2_9_11_1","doi-asserted-by":"publisher","DOI":"10.1137\/S106482759833913X"},{"key":"e_1_2_9_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2012.286"},{"key":"e_1_2_9_13_1","doi-asserted-by":"publisher","DOI":"10.1137\/15M1026419"},{"key":"e_1_2_9_14_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.5598"},{"key":"e_1_2_9_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2020.102724"},{"key":"e_1_2_9_16_1","doi-asserted-by":"publisher","DOI":"10.1137\/140968896"},{"key":"e_1_2_9_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ScalA.2016.011"},{"key":"e_1_2_9_18_1","doi-asserted-by":"publisher","DOI":"10.1002\/nla.2183"},{"key":"e_1_2_9_19_1","doi-asserted-by":"publisher","DOI":"10.1137\/0614004"},{"key":"e_1_2_9_20_1","doi-asserted-by":"publisher","DOI":"10.1137\/S1064827595294691"},{"key":"e_1_2_9_21_1","doi-asserted-by":"publisher","DOI":"10.1137\/S1064827598339372"},{"key":"e_1_2_9_22_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cam.2013.07.049"},{"key":"e_1_2_9_23_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2015.05.238"},{"key":"e_1_2_9_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/2629475"},{"key":"e_1_2_9_25_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-85665-6_34"},{"key":"e_1_2_9_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3502181.3531472"},{"key":"e_1_2_9_27_1","unstructured":"NVIDIA \u201cCUDA C Programming Guide \u201dv1.0 2007 https:\/\/developer.nvidia.com\/content\/cuda\u201010."},{"key":"e_1_2_9_28_1","unstructured":"NVIDIA Corporation \u201cNvidia Hopper Architecture Whitepaper \u201d2022 https:\/\/www.nvidia.com."},{"key":"e_1_2_9_29_1","unstructured":"NVIDIA Corporation \u201cNvidia Ada Lovelace Architecture Whitepaper \u201d2023 https:\/\/www.nvidia.com."},{"key":"e_1_2_9_30_1","volume-title":"Proceedings of the 2024 IEEE International Symposium on High\u2010Performance Computer Architecture (HPCA)","author":"Venkataraman S.","year":"2024"},{"key":"e_1_2_9_31_1","doi-asserted-by":"publisher","DOI":"10.1137\/15M1027826"},{"key":"e_1_2_9_32_1","doi-asserted-by":"publisher","DOI":"10.1137\/18M1197461"},{"key":"e_1_2_9_33_1","doi-asserted-by":"publisher","DOI":"10.1177\/10943420211017188"},{"key":"e_1_2_9_34_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2019.102599"},{"key":"e_1_2_9_35_1","volume-title":"Proceedings of the 2022 IEEE International Parallel and Distributed Processing Symposium (IPDPS)","author":"Zhao Y.","year":"2022"},{"key":"e_1_2_9_36_1","volume-title":"Graph Neural Network\u2013Guided Kernel Selection for Sparse Tensor Operations","author":"Liu Y.","year":"2023"},{"key":"e_1_2_9_37_1","volume-title":"Learning Preconditioners With Graph Neural Networks","author":"Ashby W.","year":"2023"},{"key":"e_1_2_9_38_1","unstructured":"Y.Peng J.Wang andJ.Chen \u201cNeural Preconditioners for Iterative Linear Solvers \u201d. Proceedings of ICML 2024."},{"key":"e_1_2_9_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3584373"},{"key":"e_1_2_9_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3178487.3178495"},{"key":"e_1_2_9_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/2049662.2049663"},{"key":"e_1_2_9_42_1","doi-asserted-by":"publisher","DOI":"10.21105\/joss.01244"},{"key":"e_1_2_9_43_1","unstructured":"NVIDIA \u201cCUDA C Programming Guide \u201dv12.8 2025 https:\/\/docs.nvidia.com\/cuda\/index.html."},{"key":"e_1_2_9_44_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.3936"},{"key":"e_1_2_9_45_1","unstructured":"NVIDIA \u201cCUSPARSE Library \u201dv12.8 2025 https:\/\/docs.nvidia.com\/cuda\/cusparse\/index.html."},{"key":"e_1_2_9_46_1","unstructured":"NVIDIA \u201cCUBLAS Library \u201dv12.8 2025 https:\/\/docs.nvidia.com\/cuda\/cublas\/index.html."}],"container-title":["Concurrency and Computation: Practice and Experience"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1002\/cpe.70552","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/full-xml\/10.1002\/cpe.70552","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1002\/cpe.70552","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,1,23]],"date-time":"2026-01-23T12:30:35Z","timestamp":1769171435000},"score":1,"resource":{"primary":{"URL":"https:\/\/onlinelibrary.wiley.com\/doi\/10.1002\/cpe.70552"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,1]]},"references-count":45,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2026,1]]}},"alternative-id":["10.1002\/cpe.70552"],"URL":"https:\/\/doi.org\/10.1002\/cpe.70552","archive":["Portico"],"relation":{},"ISSN":["1532-0626","1532-0634"],"issn-type":[{"value":"1532-0626","type":"print"},{"value":"1532-0634","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,1]]},"assertion":[{"value":"2025-10-16","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-12-29","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-01-18","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}],"article-number":"e70552"}}