{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,6]],"date-time":"2026-05-06T14:55:12Z","timestamp":1778079312767,"version":"3.51.4"},"reference-count":24,"publisher":"Springer Science and Business Media LLC","issue":"7","license":[{"start":{"date-parts":[[2025,3,11]],"date-time":"2025-03-11T00:00:00Z","timestamp":1741651200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,3,11]],"date-time":"2025-03-11T00:00:00Z","timestamp":1741651200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Circuits Syst Signal Process"],"published-print":{"date-parts":[[2025,7]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Dysarthria is a neurological speech disorder that affects the speech intelligibility of the speaker. Speech assistive aids are developed to support their communication needs. Successful speech assistive aids are developed using automatic speech recognition systems trained using their own speech data. The effectiveness and usefulness of speech recognition systems depend on the amount of speech data used for training. However, collecting a large amount of dysarthric speech data is difficult. Data augmentation involves applying transformation techniques to increase the quantity of available speech data. Adding noise data is also one of the approaches to make such transformations and create a new volume of data. However, care should be taken while using noise data for the transformation of the dysarthric speech data since dysarthria on its own is disordered data, and adding even more distortion reduces its quality of it. However, by performing a proper analysis of the noisy data, noise can also be used as a source to create new samples of dysarthric speech data. This paper concentrates on identifying noise characteristics and finding the suitability of using noise as a source for data augmentation in dysarthric speech. With the noise-augmented dysarthric speech data, dysarthric speech recognition systems were trained to evaluate the quality of the augmented data. It was noted that for dysarthric speakers, especially with the severe category, the low-frequency noise selection approach has resulted in a lower WER than the without augmentation by 12.29%.<\/jats:p>","DOI":"10.1007\/s00034-025-03054-4","type":"journal-article","created":{"date-parts":[[2025,3,11]],"date-time":"2025-03-11T02:50:27Z","timestamp":1741661427000},"page":"5202-5219","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":3,"title":["Analysis for Using Noise as a Source of Data Augmentation for Dysarthric Speech Recognition"],"prefix":"10.1007","volume":"44","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-3290-9975","authenticated-orcid":false,"given":"Sarkhell Sirwan","family":"Nawroly","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Decebal","family":"Popescu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"T. A.","family":"Mariya Celin","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"M. P.","family":"Actlin Jeeva","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2025,3,11]]},"reference":[{"issue":"6","key":"3054_CR1","first-page":"5539","volume":"44","author":"A Al Fahoum","year":"2023","unstructured":"A. Al Fahoum, Enhanced cardiac arrhythmia detection utilizing deep learning architectures and multi-scale ECG analysis. J. Propul. Technol. 44(6), 5539\u20135554 (2023)","journal-title":"J. Propul. Technol."},{"issue":"6","key":"3054_CR2","doi-asserted-by":"publisher","first-page":"4660","DOI":"10.1121\/1.4986746","volume":"141","author":"SA Borrie","year":"2017","unstructured":"S.A. Borrie, M. Baese-Berk, K. Van Engen, T. Bent, A relationship between processing speech in noise and dysarthric speech. J. Acoust. Soc. Am. 141(6), 4660\u20134667 (2017)","journal-title":"J. Acoust. Soc. Am."},{"issue":"2","key":"3054_CR3","first-page":"346","volume":"14","author":"TAM Celin","year":"2020","unstructured":"T.A.M. Celin, T. Nagarajan, P. Vijayalakshmi, Data augmentation using virtual microphone array synthesis and multi-resolution feature extraction for isolated word dysarthric speech recognition. IEEE J. Sel. Top. Signal Process. 14(2), 346\u2013354 (2020)","journal-title":"IEEE J. Sel. Top. Signal Process."},{"issue":"1","key":"3054_CR4","doi-asserted-by":"publisher","first-page":"601","DOI":"10.1007\/s00034-022-02156-7","volume":"42","author":"TAM Celin","year":"2023","unstructured":"T.A.M. Celin, T. Nagarajan, P. Vijayalakshmi, Data augmentation techniques for transfer learning-based continuous dysarthric speech recognition. Circuits Syst. Signal Process. 42(1), 601\u2013622 (2023)","journal-title":"Circuits Syst. Signal Process."},{"key":"3054_CR5","volume-title":"Motor Speech Disorders","author":"FL Darley","year":"1975","unstructured":"F.L. Darley, A. Aronson, J.R. Brown, Motor Speech Disorders, 1st edn. (WB Saunders, Philadelphia, 1975)","edition":"1"},{"issue":"4","key":"3054_CR6","doi-asserted-by":"publisher","first-page":"523","DOI":"10.1007\/s10579-011-9145-0","volume":"46","author":"R Frank","year":"2012","unstructured":"R. Frank, N. Aravind Kumar, W. Talya, The TORGO database of acoustic and articulatory speech from speakers with dysarthria. Lang. Resour. Eval. 46(4), 523\u2013541 (2012)","journal-title":"Lang. Resour. Eval."},{"issue":"1","key":"3054_CR7","first-page":"696","volume":"1","author":"M Geng","year":"2020","unstructured":"M. Geng, X. Xie, S. Liu, J. Yu, S. Hu, X. Liu, H. Meng, Investigation of data augmentation techniques for disordered speech recognition. Proc. INTERSPEECH 1(1), 696\u2013700 (2020)","journal-title":"Proc. INTERSPEECH"},{"key":"3054_CR8","unstructured":"K. Heejin, M. Hasegawa-Johnson, A. Perlman, J. Gunderson, T.S. Huang, K. Watkin, S. Frame, Dysarthric speech database for universal access research, in Proceedings of the 9th Annual Conference of the International Speech Communication Association, pp. 1741\u20131744 (2008)"},{"key":"3054_CR9","doi-asserted-by":"publisher","first-page":"588","DOI":"10.1016\/j.specom.2006.12.006","volume":"49","author":"YJ Hu","year":"2007","unstructured":"Y.J. Hu, Subjective evaluation and comparison of speech enhancement algorithms. Speech Commun. 49, 588\u2013601 (2007)","journal-title":"Speech Commun."},{"issue":"5","key":"3054_CR10","doi-asserted-by":"publisher","first-page":"288","DOI":"10.1049\/iet-spr.2019.0226","volume":"14","author":"MPA Jeeva","year":"2020","unstructured":"M.P.A. Jeeva, T. Nagarajan, P. Vijayalakshmi, Adaptive multi-band filter structure-based far-end speech enhancement. IET Signal Proc. 14(5), 288\u2013299 (2020). https:\/\/doi.org\/10.1049\/iet-spr.2019.0226","journal-title":"IET Signal Proc."},{"issue":"8","key":"3054_CR11","doi-asserted-by":"publisher","first-page":"965","DOI":"10.1049\/iet-spr.2016.0125","volume":"10","author":"MPA Jeeva","year":"2016","unstructured":"M.P.A. Jeeva, T. Nagarajan, P. Vijayalakshmi, Discrete cosine transform-derived spectrum-based speech enhancement algorithm using temporal-domain multiband filtering. IET Signal Process. 10(8), 965\u2013980 (2016)","journal-title":"IET Signal Process."},{"key":"3054_CR12","doi-asserted-by":"crossref","unstructured":"Y. Jiao, M. Tu, V. Berisha, J. Liss, Simulating dysarthric speech for training data augmentation in clinical speech applications, in Proceedings of 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 6009\u20136013 (2018)","DOI":"10.1109\/ICASSP.2018.8462290"},{"key":"3054_CR13","doi-asserted-by":"crossref","unstructured":"Z. Jin, M. Geng, X. Xie, J. Yu, S. Liu, X. Liu, H. Meng, Adversarial data augmentation for disordered speech recognition, in Proceedings of Interspeech, pp. 4803\u20134807 (2021)","DOI":"10.21437\/Interspeech.2021-168"},{"issue":"5","key":"3054_CR14","doi-asserted-by":"publisher","first-page":"561","DOI":"10.1109\/PROC.1975.9792","volume":"63","author":"J Makhoul","year":"1975","unstructured":"J. Makhoul, Linear prediction: a tutorial review. Proc. IEEE 63(5), 561\u2013580 (1975)","journal-title":"Proc. IEEE"},{"key":"3054_CR15","doi-asserted-by":"crossref","unstructured":"Y. Matsuzaka, R. Takashima, C. Sasaki, T. Takiguchi, Data augmentation for dysarthric speech recognition based on text-to-speech synthesis , in 2022 IEEE 4th Global Conference on Life Sciences and Technologies (LifeTech), pp. 399\u2013400 (2022)","DOI":"10.1109\/LifeTech53646.2022.9754798"},{"key":"3054_CR16","doi-asserted-by":"crossref","unstructured":"X. Menendez Pidal, J.B. Polikoff, S.M. Peters, J.E. Leonzio, H.T. Bunnell, The Nemours database of dysarthric speech, in Proceedings of 4th International Conference on Spoken Language Processing, vol. 3(1), pp. 1962\u20131965 (1996)","DOI":"10.21437\/ICSLP.1996-503"},{"key":"3054_CR17","unstructured":"D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, The kaldi speech recognition toolkit, in Automatic Speech Recognition and Understanding Workshop, vol. 1(1), pp. 1\u20134 (2011)"},{"issue":"3","key":"3054_CR18","doi-asserted-by":"publisher","first-page":"279","DOI":"10.1109\/LSP.2017.2657381","volume":"24","author":"J Salamon","year":"2017","unstructured":"J. Salamon, J.P. Bello, Deep convolutional neural networks and data augmentation for environmental sound classification. IEEE Signal Process. Lett. 24(3), 279\u2013283 (2017)","journal-title":"IEEE Signal Process. Lett."},{"key":"3054_CR19","doi-asserted-by":"publisher","first-page":"3407","DOI":"10.1109\/TNSRE.2023.3307020","volume":"31","author":"SR Shahamiri","year":"2023","unstructured":"S.R. Shahamiri, V. Lal, D. Shah, Dysarthric speech transformer: a sequence-to-sequence dysarthric speech recognition system. IEEE Trans. Neural Syst. Rehabil. Eng. 31, 3407\u20133416 (2023)","journal-title":"IEEE Trans. Neural Syst. Rehabil. Eng."},{"issue":"6","key":"3054_CR20","doi-asserted-by":"publisher","first-page":"1147","DOI":"10.1016\/j.csl.2012.10.002","volume":"27","author":"HV Sharma","year":"2013","unstructured":"H.V. Sharma, M. Hasegawa-Johnson, Acoustic model adaptation using in-domain background models for dysarthric speech recognition. Comput. Speech Lang. 27(6), 1147\u20131162 (2013)","journal-title":"Comput. Speech Lang."},{"key":"3054_CR21","doi-asserted-by":"crossref","unstructured":"R. Takashima, T. Takiguchi, Y. Ariki, Twostep acoustic model adaptation for dysarthric speech recognition, in Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 6104\u20136108 (2020)","DOI":"10.1109\/ICASSP40776.2020.9053725"},{"key":"3054_CR22","doi-asserted-by":"crossref","unstructured":"B. Vachhani, C. Bhat, S.K. Kopparapu, Data augmentation using healthy speech for dysarthric speech recognition, in Proceedings of INTERSPEECH, pp. 471\u2013475 (2019)","DOI":"10.21437\/Interspeech.2018-1751"},{"issue":"3","key":"3054_CR23","doi-asserted-by":"publisher","first-page":"247","DOI":"10.1016\/0167-6393(93)90095-3","volume":"12","author":"A Varga","year":"1993","unstructured":"A. Varga, H.J. Steeneken, Assessment for automatic speech recognition: NOISEX-92: a database and an experiment to study the effect of additive noise on speech recognition systems. Speech Commun. 12(3), 247\u2013251 (1993)","journal-title":"Speech Commun."},{"key":"3054_CR24","doi-asserted-by":"crossref","unstructured":"F. Xiong, J. Barker, Z. Yue, H. Christensen, Source domain data selection for improved transfer learning targeting dysarthric speech recognition, in Proceedings of IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pp. 7424\u20137428 (2020)","DOI":"10.1109\/ICASSP40776.2020.9054694"}],"container-title":["Circuits, Systems, and Signal Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00034-025-03054-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00034-025-03054-4\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00034-025-03054-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,1]],"date-time":"2025-07-01T21:02:08Z","timestamp":1751403728000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00034-025-03054-4"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,3,11]]},"references-count":24,"journal-issue":{"issue":"7","published-print":{"date-parts":[[2025,7]]}},"alternative-id":["3054"],"URL":"https:\/\/doi.org\/10.1007\/s00034-025-03054-4","relation":{},"ISSN":["0278-081X","1531-5878"],"issn-type":[{"value":"0278-081X","type":"print"},{"value":"1531-5878","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,3,11]]},"assertion":[{"value":"23 April 2024","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 February 2025","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 February 2025","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"11 March 2025","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}