{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T23:02:34Z","timestamp":1777676554682,"version":"3.51.4"},"reference-count":60,"publisher":"SAGE Publications","issue":"3","license":[{"start":{"date-parts":[[2015,3,30]],"date-time":"2015-03-30T00:00:00Z","timestamp":1427673600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["The International Journal of High Performance Computing Applications"],"published-print":{"date-parts":[[2015,8]]},"abstract":"<jats:p>Direct and iterative methods are often used to solve linear systems in engineering. The matrices involved can be large, which leads to heavy computations on the central processing unit. A graphics processing unit can be used to accelerate these computations. In this paper, we propose a new library, named Alinea, for advanced linear algebra. This library is implemented in C++, CUDA and OpenCL. It includes several linear algebra operations and numerous algorithms for solving linear systems. For both central processing unit and graphic processing unit devices, there are different matrix storage formats, and real and complex arithmetics in single- and double-precision. The CUDA version includes a self-tuning of the grid, i.e. threading distribution, depending upon the hardware configuration and the size of the problems. Numerical experiments and comparison with existing libraries illustrates the efficiency, accuracy and robustness of the proposed library.<\/jats:p>","DOI":"10.1177\/1094342015576774","type":"journal-article","created":{"date-parts":[[2015,4,1]],"date-time":"2015-04-01T11:17:05Z","timestamp":1427887025000},"page":"284-310","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":24,"title":["Alinea: An Advanced Linear Algebra Library for Massively Parallel Computations on Graphics Processing Units"],"prefix":"10.1177","volume":"29","author":[{"given":"Fr\u00e9d\u00e9ric","family":"Magoul\u00e8s","sequence":"first","affiliation":[{"name":"Ecole Centrale Paris, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Abal-Kassim Cheik","family":"Ahamed","sequence":"additional","affiliation":[{"name":"Ecole Centrale Paris, France"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"179","published-online":{"date-parts":[[2015,3,30]]},"reference":[{"key":"bibr1-1094342015576774","series-title":"Lecture Notes in Computer Science","doi-asserted-by":"crossref","first-page":"162","DOI":"10.1007\/978-3-642-28151-8_16","volume-title":"PARA (1)","volume":"7133","author":"Aliaga JI","year":"2010"},{"key":"bibr2-1094342015576774","unstructured":"Bell N, Garland M (2012) Cusp: Generic Parallel Algorithms for Sparse Matrix and Graph Computations. Nvidia Corporation. Available on line at: http:\/\/cusplibrary.github.io\/."},{"key":"bibr3-1094342015576774","series-title":"Lectures in Applied Mathematics","first-page":"99","volume-title":"1995 AMS\u2013SIAM summer seminar in applied mathematics","volume":"32","author":"Berry MW","year":"1996"},{"key":"bibr4-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/567806.567807"},{"key":"bibr5-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-5041-2940-4_9"},{"key":"bibr6-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1006\/jsco.1996.0125"},{"key":"bibr7-1094342015576774","first-page":"233","volume-title":"SPAA","author":"Bulu\u00e7 A","year":"2009"},{"key":"bibr8-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1007\/BF01933580"},{"key":"bibr9-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/HPCC.2012.193"},{"key":"bibr10-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/HPCC.2012.118"},{"key":"bibr11-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/DCABES.2013.26"},{"key":"bibr12-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/DCABES.2014.13"},{"key":"bibr13-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/DCABES.2014.7"},{"key":"bibr14-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-05789-7_66"},{"key":"bibr15-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2002.1016646"},{"key":"bibr16-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/SC.Companion.2012.138"},{"key":"bibr17-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/2049662.2049670"},{"key":"bibr18-1094342015576774","first-page":"197","volume-title":"Second international conference on mathematical and numerical aspects of wave propagation","author":"Despr\u00e9s B","year":"1993"},{"key":"bibr19-1094342015576774","volume-title":"ICT innovations 2013 web proceedings","author":"Djinevski L","year":"2013"},{"key":"bibr20-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.728"},{"key":"bibr21-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/PDP.2010.55"},{"key":"bibr22-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/1377603.1377607"},{"key":"bibr23-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-010-0111-7"},{"key":"bibr24-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/ICCIS.2010.285"},{"key":"bibr25-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/1555754.1555775"},{"key":"bibr26-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1137\/090779760"},{"key":"bibr27-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/200979.200981"},{"key":"bibr28-1094342015576774","series-title":"Technical Report","volume-title":"Analysis of memory latency factors and their impact on KSR1 MPP performance","volume":"159","author":"Kahhaleh B","year":"1993"},{"key":"bibr29-1094342015576774","unstructured":"Khronos (2010) OpenCL. Available at: http:\/\/www.khronos.org\/opencl\/."},{"key":"bibr30-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/2228360.2228519"},{"key":"bibr31-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/2514641.2514648"},{"key":"bibr32-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1049\/el.2010.1680"},{"key":"bibr33-1094342015576774","unstructured":"Li N, Suchomel B, Osei-Kuffuor D, Li R, Saad Y (2010) ITSOL. Available at: http:\/\/www-users.cs.umn.edu\/saad\/software\/ITSOL\/index.html."},{"key":"bibr34-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1137\/S1064827597327334"},{"key":"bibr35-1094342015576774","first-page":"1","volume-title":"Proceedings of the first international symposium on domain decomposition methods for partial differential equations","author":"Lions PL","year":"1988"},{"key":"bibr36-1094342015576774","first-page":"45","volume-title":"2nd international workshop on GPUs and scientific applications (GPUScA 2011)","author":"Luo L","year":"2011"},{"key":"bibr37-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/PDCAT.2012.18"},{"key":"bibr38-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2013.06.020"},{"key":"bibr39-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.cma.2005.01.025"},{"key":"bibr40-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.apm.2005.05.020"},{"key":"bibr41-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.cma.2005.05.059"},{"key":"bibr42-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1080\/00207160.2014.930137"},{"key":"bibr43-1094342015576774","first-page":"1","author":"Magoul\u00e8s F","year":"2015","journal-title":"International Journal of Computer Mathematics"},{"key":"bibr44-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.apm.2005.06.016"},{"key":"bibr45-1094342015576774","unstructured":"NVidia Corporation (2011a) CUDA Programming Guide. 4.0 edition. Available at: http:\/\/developer.nvidia.com\/cuda-toolkit-40."},{"key":"bibr46-1094342015576774","unstructured":"NVidia Corporation (2011b) CUDA Toolkit 4.0, CUBLAS Library. Available at: http:\/\/developer.nvidia.com\/cuda-toolkit-40."},{"key":"bibr47-1094342015576774","unstructured":"Corporation (2011c) CUDA Toolkit 4.0, CUSPARSE Library. Available at: http:\/\/developer.nvidia.com\/cuda-toolkit-40."},{"key":"bibr48-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/S0167-8191(97)00026-4"},{"key":"bibr49-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/j.cma.2011.01.013"},{"key":"bibr50-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/APS.2001.960019"},{"key":"bibr51-1094342015576774","doi-asserted-by":"crossref","DOI":"10.1093\/oso\/9780198501787.001.0001","volume-title":"Domain Decomposition Methods for Partial Differential Equations","author":"Quarteroni A","year":"1999"},{"key":"bibr52-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1137\/1.9780898718003"},{"key":"bibr53-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1016\/0141-1195(80)90051-0"},{"key":"bibr54-1094342015576774","unstructured":"Volkov V, Demmel JW (2008) LU, QR and Cholesky factorizations using vector capabilities of GPUs. Technical Report No. UCB\/EECS-2008-49, EECS Department University of California, Berkeley. Available at: http:\/\/www.eecs.berkeley.edu\/Pubs\/TechRpts\/2008\/EECS-2008-49.pdf."},{"key":"bibr55-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/ICICISYS.2009.5358155"},{"key":"bibr56-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1145\/1555754.1555761"},{"key":"bibr57-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/DCABES.2010.162"},{"key":"bibr58-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.1732"},{"key":"bibr59-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/IFCSTA.2009.68"},{"key":"bibr60-1094342015576774","doi-asserted-by":"publisher","DOI":"10.1109\/ICCSNT.2012.6526361"}],"container-title":["The International Journal of High Performance Computing Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/1094342015576774","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.1177\/1094342015576774","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/1094342015576774","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T08:19:24Z","timestamp":1777450764000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.1177\/1094342015576774"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,3,30]]},"references-count":60,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2015,8]]}},"alternative-id":["10.1177\/1094342015576774"],"URL":"https:\/\/doi.org\/10.1177\/1094342015576774","relation":{},"ISSN":["1094-3420","1741-2846"],"issn-type":[{"value":"1094-3420","type":"print"},{"value":"1741-2846","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,3,30]]}}}