{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,21]],"date-time":"2026-08-21T16:39:27Z","timestamp":1787330367839,"version":"build-2736575974"},"reference-count":33,"publisher":"Society for Industrial & Applied Mathematics (SIAM)","issue":"6","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["SIAM J. Sci. Comput."],"published-print":{"date-parts":[[2015,1]]},"abstract":"<jats:p>In this paper, a task-scheduling approach to efficiently calculating sparse symmetric matrix-vector products and designed to run on graphics processing units (GPUs) is presented. The main premise is that, for many sparse symmetric matrices occurring in common applications, it is possible to obtain significant reductions in memory usage and improvements in performance when the matrix is prepared in certain ways prior to computation. The preprocessing proposed in this paper employs task scheduling to overcome the difficulties that have suppressed the development of methods taking advantage of the symmetry of sparse matrices. The performance of the proposed task-scheduling method is verified using a Kepler (Tesla K40c) graphics accelerator, and is compared to the performance of cuSPARSE library functions on a GPU and to functions from the Intel MKL on central processing units (CPUs) executed in the parallel mode. The obtained results indicate that the proposed approach for sparse symmetric matrix-vector products results in up to a 40% reduction in memory usage, as compared to nonsymmetric matrix storage formats, while retaining good throughput. Compared to cuSPARSE and Intel MKL functions for sparse symmetric matrices, the proposed TSMV approach allowed us to achieve a significant speedup (of over one order of magnitude).<\/jats:p>","DOI":"10.1137\/14097135x","type":"journal-article","created":{"date-parts":[[2015,12,2]],"date-time":"2015-12-02T14:18:21Z","timestamp":1449065901000},"page":"C643-C666","source":"Crossref","is-referenced-by-count":6,"title":["A Task-Scheduling Approach for Efficient Sparse Symmetric Matrix-Vector Multiplication on a GPU"],"prefix":"10.1137","volume":"37","author":[{"given":"P.","family":"Mironowicz","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"A.","family":"Dziekonski","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"M.","family":"Mrozowski","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"351","published-online":{"date-parts":[[2015,11,19]]},"reference":[{"key":"atypb1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2011.73"},{"key":"atypb2","doi-asserted-by":"publisher","DOI":"10.2528\/PIER11031607"},{"key":"atypb3","doi-asserted-by":"publisher","DOI":"10.1109\/LAWP.2011.2159769"},{"key":"atypb4","doi-asserted-by":"publisher","DOI":"10.1002\/nme.4452"},{"key":"atypb5","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-11515-8_10"},{"key":"atypb6","doi-asserted-by":"publisher","DOI":"10.1145\/331532.331562"},{"key":"atypb7","unstructured":"C. Damhaug,\n                      Matrix: CNVS\/shipsec1\n                      , http:\/\/www.cise.ufl.edu\/research\/sparse\/matrices\/DNVS\/shipsec1.html, (1999)."},{"key":"atypb8","doi-asserted-by":"publisher","DOI":"10.4208\/cicp.2009.v6.p342"},{"key":"atypb9","doi-asserted-by":"publisher","DOI":"10.1016\/B978-044482851-4.50030-X"},{"key":"atypb10","doi-asserted-by":"publisher","DOI":"10.1145\/800195.805928"},{"key":"atypb11","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2011.08.003"},{"key":"atypb12","volume-title":"Technical report UCB\/CSD-00-1104, EECS Department","author":"Im E.-J.","year":"2000"},{"key":"atypb13","doi-asserted-by":"publisher","DOI":"10.1137\/10079906X"},{"key":"atypb14","first-page":"415","volume":"26","author":"Nesetril J.","year":"1985","journal-title":"Comment. Math. Univ. Carolin."},{"key":"atypb15","doi-asserted-by":"publisher","DOI":"10.1137\/130930352"},{"key":"atypb16","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2014.03.008"},{"key":"atypb17","doi-asserted-by":"publisher","DOI":"10.1007\/BF02523189"},{"key":"atypb18","volume-title":"NVIDIA Technical report NVR-2008-004","author":"Bell N.","year":"2008"},{"key":"atypb19","unstructured":"NVIDIA Corporation,\n                      CUDA C Best Practices Guide\n                      , http:\/\/docs.nvidia.com\/cuda\/cuda-c-best-practices-guide\/ (2015)."},{"key":"atypb22","doi-asserted-by":"publisher","DOI":"10.2514\/3.12012"},{"key":"atypb23","volume-title":"Techniques for Optimizing Applications - High Performance Computing","author":"Garg R. P.","year":"2002"},{"key":"atypb24","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2008.12.006"},{"key":"atypb25","doi-asserted-by":"publisher","DOI":"10.1137\/1.9780898718003"},{"key":"atypb26","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2013.43"},{"key":"atypb27","doi-asserted-by":"publisher","DOI":"10.1147\/rd.416.0711"},{"key":"atypb28","volume-title":"GPU Technology Conference 2010 (GTC 2010)","author":"Volkov V."},{"key":"atypb29","unstructured":"R. W. Vuduc,\n                      Automatic Performance Tuning of Sparse Matrix Kernels\n                      , Ph.D. thesis, University of Califironia, Berkeley, CA, 2003."},{"key":"atypb30","doi-asserted-by":"publisher","DOI":"10.1007\/s006070050015"},{"key":"atypb31","doi-asserted-by":"publisher","DOI":"10.1007\/BF01933580"},{"key":"atypb32","doi-asserted-by":"publisher","DOI":"10.1145\/2503210.2503234"},{"key":"atypb34","doi-asserted-by":"publisher","DOI":"10.1145\/2464996.2465013"},{"key":"atypb35","doi-asserted-by":"publisher","DOI":"10.1137\/120900216"},{"key":"atypb36","doi-asserted-by":"publisher","DOI":"10.1145\/567112.567114"}],"container-title":["SIAM Journal on Scientific Computing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/epubs.siam.org\/doi\/pdf\/10.1137\/14097135X","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,8,21]],"date-time":"2026-08-21T16:00:42Z","timestamp":1787328042000},"score":1,"resource":{"primary":{"URL":"https:\/\/epubs.siam.org\/doi\/10.1137\/14097135X"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,1]]},"references-count":33,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2015,1]]}},"alternative-id":["10.1137\/14097135X"],"URL":"https:\/\/doi.org\/10.1137\/14097135x","relation":{},"ISSN":["1064-8275","1095-7197"],"issn-type":[{"value":"1064-8275","type":"print"},{"value":"1095-7197","type":"electronic"}],"subject":[],"published":{"date-parts":[[2015,1]]}}}