{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,13]],"date-time":"2026-07-13T22:33:42Z","timestamp":1783982022498,"version":"3.55.0"},"reference-count":32,"publisher":"SAGE Publications","issue":"3","license":[{"start":{"date-parts":[[2012,8,1]],"date-time":"2012-08-01T00:00:00Z","timestamp":1343779200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/journals.sagepub.com\/page\/policies\/text-and-data-mining-license"}],"content-domain":{"domain":["journals.sagepub.com"],"crossmark-restriction":true},"short-container-title":["The International Journal of High Performance Computing Applications"],"published-print":{"date-parts":[[2013,8]]},"abstract":"<jats:p>The aim of most microphone array applications is to localize sound sources in a noisy and reverberant environment. For that purpose, many different sound source localization (SSL) algorithms have been proposed, where the SRP-PHAT (steered response power using the phase transform) has been known as one of the state-of-the-art methods. Its original formulation allows two different practical implementations, one that is computed in the frequency domain (FDSP) and another in the time domain (TDSP), which can be enhanced by interpolation. However, the main problem of this algorithm is its high computational cost due to intensive grid scan in search for the sound source. Considering the power of graphics processing units (GPUs) for working with massively parallelizable compute-intensive algorithms, we present two highly scalable GPU-based versions of the SRP-PHAT, one for each formulation, and also an implementation of the cubic splines interpolation in the GPU. These approaches exploit the parallel aspects of the SRP-PHAT, allowing real-time execution for large search grids. Comparing our GPU approaches against traditional multithreaded CPU approaches, results show a speed up of [Formula: see text] for the FDSP, and [Formula: see text] for the TDSP with interpolation, when comparing high-end GPUs with high-end CPUs.<\/jats:p>","DOI":"10.1177\/1094342012452166","type":"journal-article","created":{"date-parts":[[2012,8,2]],"date-time":"2012-08-02T01:31:02Z","timestamp":1343871062000},"page":"291-306","update-policy":"https:\/\/doi.org\/10.1177\/sage-journals-update-policy","source":"Crossref","is-referenced-by-count":13,"title":["GPU-based approaches for real-time sound source localization using the SRP-PHAT algorithm"],"prefix":"10.1177","volume":"27","author":[{"given":"Vicente Peruffo","family":"Minotto","sequence":"first","affiliation":[{"name":"Instituto de Inform\u00e1tica, Universidade Federal do Rio Grande do Sul, RS, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Claudio Rosito","family":"Jung","sequence":"additional","affiliation":[{"name":"Instituto de Inform\u00e1tica, Universidade Federal do Rio Grande do Sul, RS, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"suffix":"Jr","given":"Luiz Gonzaga","family":"da Silveira","sequence":"additional","affiliation":[{"name":"Applied Computing Graduate Program (PIPCA), Universidade do Vale do Rio dos Sinos, RS, Brazil"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bowon","family":"Lee","sequence":"additional","affiliation":[{"name":"Mobile and Immersive Experience Lab, Hewlett-Packard Laboratories, Palo Alto, CA, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"179","published-online":{"date-parts":[[2012,8,1]]},"reference":[{"key":"bibr1-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1006\/csla.1995.0009"},{"key":"bibr2-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-662-04619-7"},{"key":"bibr3-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/89.554268"},{"key":"bibr4-1094342012452166","volume-title":"Using OpenMP: Portable Shared Memory Parallel Programming (Scientific and Engineering Computation)","author":"Chapman B","year":"2007"},{"key":"bibr5-1094342012452166","unstructured":"da Silveira LG, Minotto VP, Jung CR, Lee B (2010) A GPU Implementation of the SRP-PHAT Sound Source Localization Algorithm. In International Workshop on Acoustic Echo and Noise Control, http:\/\/www.iwaenc.org\/proceedings\/2010\/HTML\/Uploads\/1062.pdf."},{"key":"bibr6-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1145\/2070481.2070498"},{"key":"bibr7-1094342012452166","volume-title":"A High-Accuracy, Low-Latency Technique for Talker Localization in Reverberant Environments Using Microphone Arrays","author":"DiBiase JH","year":"2000"},{"key":"bibr8-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/ASPAA.2007.4392976"},{"key":"bibr9-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2007.366631"},{"key":"bibr10-1094342012452166","volume-title":"NVIDIA\u2019s Fermi: The First Complete GPU Computing Architecture","author":"Glaskowsky PN","year":"2009"},{"key":"bibr11-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2008.5213922"},{"key":"bibr12-1094342012452166","unstructured":"Harris M (2011) Optimizing Parallel Reduction in CUDA. NVIDIA Corporation. http:\/\/developer.nvidia.com\/cuda-downloads. Accessed February 2012."},{"key":"bibr13-1094342012452166","volume-title":"Programming Massively Parallel Processors: A Hands-on Approach","author":"Kirk DB","year":"2010"},{"key":"bibr14-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1976.1162830"},{"key":"bibr15-1094342012452166","unstructured":"Lee B, Kalker T (2010) A vectorized method for computationally efficient SRP-PHAT sound source localization. In International Workshop on Acoustic Echo and Noise Control, http:\/\/www.iwaenc.org\/proceedings\/2010\/HTML\/Uploads\/1086.pdf."},{"key":"bibr16-1094342012452166","unstructured":"Lee B, Kalker T, Schafer RW (2008) Maximum-likelihood sound source localization with a multivariate complex Laplacian distribution. In International Workshop on Acoustic Echo and Noise Control, http:\/\/www.iwaenc.org\/proceedings\/2008\/contents\/papers\/9053.pdf."},{"key":"bibr17-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2008.31"},{"key":"bibr18-1094342012452166","volume-title":"NVIDIA GeForce 8800 GPU Architecture Overview","author":"NVIDIA Corporation","year":"2006"},{"key":"bibr19-1094342012452166","unstructured":"NVIDIA Corporation (2011a) CUDA Best Practices Guide. NVIDIA Corporation. http:\/\/developer.nvidia.com\/cuda-downloads. Accessed February 2012."},{"key":"bibr20-1094342012452166","unstructured":"NVIDIA Corporation (2011b) CUDA CUFFT Library. NVIDIA Corporation. http:\/\/developer.nvidia.com\/cuda-downloads. Accessed February 2012."},{"key":"bibr21-1094342012452166","unstructured":"NVIDIA Corporation (2011c) CUDA Programming Guide. NVIDIA Corporation. http:\/\/developer.nvidia.com\/cuda-downloads. Accessed February 2012."},{"key":"bibr22-1094342012452166","unstructured":"NVIDIA Corporation (2011d) Fermi Compatibility Guide. NVIDIA Corporation. http:\/\/developer.download.nvidia.com\/compute\/DevZone\/docs\/html\/C\/doc\/Fermi_Compatibility_Guide.pdf. Accessed April 2012."},{"key":"bibr23-1094342012452166","volume-title":"NVIDIA\u2019s Next Generation CUDA Compute Architecture: Fermi","author":"NVIDIA Corporation","year":"2011"},{"key":"bibr24-1094342012452166","volume-title":"Spoken Dialogues with Computers","author":"Omologo M","year":"1997"},{"key":"bibr25-1094342012452166","first-page":"3","volume":"59","author":"Savioja L","year":"2011","journal-title":"Journal of the Audio Engineering Society"},{"key":"bibr26-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/TSA.2005.848875"},{"key":"bibr27-1094342012452166","unstructured":"Tervo S, Lokki T (2008) Interpolation Methods for the SRP-PHAT Algorithm. In International Workshop on Acoustic Echo and Noise Control, http:\/\/www.iwaenc.org\/proceedings\/2008\/contents\/papers\/9037.pdf."},{"issue":"59","key":"bibr28-1094342012452166","first-page":"1","volume":"127","author":"Tsingos N","year":"2009","journal-title":"Audio Engineering Society Convention"},{"key":"bibr29-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1109\/SIPS.1999.822369"},{"key":"bibr30-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1002\/9780470714089"},{"key":"bibr31-1094342012452166","first-page":"I","volume-title":"IEEE International Conference of Acoustics, Speech and Signal Processing, 2007 (ICASSP 2007)","volume":"1","author":"Zhang C","year":"2007"},{"key":"bibr32-1094342012452166","doi-asserted-by":"publisher","DOI":"10.1145\/1693453.1693472"}],"container-title":["The International Journal of High Performance Computing Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/1094342012452166","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/full-xml\/10.1177\/1094342012452166","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/journals.sagepub.com\/doi\/pdf\/10.1177\/1094342012452166","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T08:19:11Z","timestamp":1777450751000},"score":1,"resource":{"primary":{"URL":"https:\/\/journals.sagepub.com\/doi\/10.1177\/1094342012452166"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,8,1]]},"references-count":32,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2013,8]]}},"alternative-id":["10.1177\/1094342012452166"],"URL":"https:\/\/doi.org\/10.1177\/1094342012452166","relation":{},"ISSN":["1094-3420","1741-2846"],"issn-type":[{"value":"1094-3420","type":"print"},{"value":"1741-2846","type":"electronic"}],"subject":[],"published":{"date-parts":[[2012,8,1]]}}}