{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,7]],"date-time":"2026-06-07T16:56:21Z","timestamp":1780851381708,"version":"3.54.1"},"reference-count":41,"publisher":"Springer Science and Business Media LLC","issue":"3","license":[{"start":{"date-parts":[[2022,8,1]],"date-time":"2022-08-01T00:00:00Z","timestamp":1659312000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2022,8,1]],"date-time":"2022-08-01T00:00:00Z","timestamp":1659312000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Indo-Norwegian","award":["287918"],"award-info":[{"award-number":["287918"]}]},{"DOI":"10.13039\/501100012704","name":"University of Agder","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100012704","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Speech Technol"],"published-print":{"date-parts":[[2022,9]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Speech enables easy human-to-human communication as well as human-to-machine interaction. However, the quality of speech degrades due to background noise in the environment, such as drone noise embedded in speech during search and rescue operations. Similarly, helicopter noise, airplane noise, and station noise reduce the quality of speech. Speech enhancement algorithms reduce background noise, resulting in a crystal clear and noise-free conversation. For many applications, it is also necessary to process these noisy speech signals at the edge node level. Thus, we propose implicit Wiener filter-based algorithm for speech enhancement using edge computing system. In the proposed algorithm, a first order recursive equation is used to estimate the noise. The performance of the proposed algorithm is evaluated for two speech utterances, one uttered by a male speaker and the other by a female speaker. Both utterances are degraded by different types of non-stationary noises such as exhibition, station, drone, helicopter, airplane, and white Gaussian stationary noise with different signal-to-noise ratios. Further, we compare the performance of the proposed speech enhancement algorithm with the conventional spectral subtraction algorithm. Performance evaluations using objective speech quality measures demonstrate that the proposed speech enhancement algorithm outperforms the spectral subtraction algorithm in estimating the clean speech from the noisy speech. Finally, we implement the proposed speech enhancement algorithm, in addition to the spectral subtraction algorithm, on the Raspberry Pi 4 Model B, which is a low power edge computing device.<\/jats:p>","DOI":"10.1007\/s10772-022-09987-4","type":"journal-article","created":{"date-parts":[[2022,8,1]],"date-time":"2022-08-01T16:02:57Z","timestamp":1659369777000},"page":"745-758","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":21,"title":["Single-channel speech enhancement using implicit Wiener filter for high-quality speech communication"],"prefix":"10.1007","volume":"25","author":[{"given":"Rahul Kumar","family":"Jaiswal","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sreenivasa Reddy","family":"Yeduri","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1023-2118","authenticated-orcid":false,"given":"Linga Reddy","family":"Cenkeramaddi","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2022,8,1]]},"reference":[{"issue":"1","key":"9987_CR1","doi-asserted-by":"publisher","first-page":"53","DOI":"10.1007\/s10772-013-9205-5","volume":"17","author":"MA Abd El-Fattah","year":"2014","unstructured":"Abd El-Fattah, M. A., Dessouky, M. I., Abbas, A. M., Diab, S. M., El-Rabaie, E. S. M., Al-Nuaimy, W., Alshebeili, S. A., & Abd El-Samie, F. E. (2014). Speech enhancement with an adaptive Wiener filter. International Journal of Speech Technology, 17(1), 53\u201364.","journal-title":"International Journal of Speech Technology"},{"key":"9987_CR2","doi-asserted-by":"crossref","unstructured":"Al-Emadi, S., Al-Ali, A., Mohammad, A., & Al-Ali, A. (2019). Audio based drone detection and identification using deep learning. In Proceedings of the international wireless communications & mobile computing conference (pp. 459\u2013464).","DOI":"10.1109\/IWCMC.2019.8766732"},{"key":"9987_CR3","doi-asserted-by":"crossref","unstructured":"Ali, Y. S. E., Parsa, V., Doyle, P., & Berkane, S. (2020). Low-complexity disordered speech quality estimation. International Journal of Speech Technology, 23(3), 585-594.","DOI":"10.1007\/s10772-020-09688-w"},{"issue":"5","key":"9987_CR4","doi-asserted-by":"publisher","first-page":"497","DOI":"10.1109\/89.861364","volume":"8","author":"F Asano","year":"2000","unstructured":"Asano, F., Hayamizu, S., Yamada, T., & Nakamura, S. (2000). Speech enhancement based on the subspace method. IEEE Transactions on Speech and Audio Processing, 8(5), 497\u2013507.","journal-title":"IEEE Transactions on Speech and Audio Processing"},{"key":"9987_CR5","doi-asserted-by":"crossref","unstructured":"Azarpour, M., Siska, J., & Enzner, G. (2017). Real-time binaural speech enhancement demo on raspberry pi. In Proceedings of the IEEE international conference on acoustics, speech and signal processing (ICASSP) (pp. 6572\u20136573).","DOI":"10.1109\/ICASSP.2017.8005296"},{"key":"9987_CR6","doi-asserted-by":"crossref","unstructured":"Bhowmick, A., & Chandra, M. (2017). Speech enhancement using voiced speech probability based wavelet decomposition. Computers & Electrical Engineering, 62, 706\u2013718.","DOI":"10.1016\/j.compeleceng.2017.01.013"},{"key":"9987_CR7","doi-asserted-by":"crossref","unstructured":"Boll, S. (1979). Suppression of acoustic noise in speech using spectral subtraction. IEEE Transactions on Acoustics, Speech, and Signal Processing, 27(2), 113\u2013120.","DOI":"10.1109\/TASSP.1979.1163209"},{"issue":"5","key":"9987_CR8","doi-asserted-by":"publisher","first-page":"1170","DOI":"10.1109\/TASL.2010.2087750","volume":"19","author":"W Charoenruengkit","year":"2010","unstructured":"Charoenruengkit, W., & Erd\u00f6l, N. (2010). The effect of spectral estimation on speech enhancement performance. IEEE Transactions on Audio, Speech, and Language Processing, 19(5), 1170\u20131179.","journal-title":"IEEE Transactions on Audio, Speech, and Language Processing"},{"key":"9987_CR9","doi-asserted-by":"publisher","first-page":"46","DOI":"10.1016\/j.specom.2019.03.005","volume":"109","author":"RA Chiea","year":"2019","unstructured":"Chiea, R. A., Costa, M. H., & Barrault, G. (2019). New insights on the optimality of parameterized wiener filters for speech enhancement applications. Speech Communication, 109, 46\u201354.","journal-title":"Speech Communication"},{"issue":"1","key":"9987_CR10","doi-asserted-by":"publisher","first-page":"53","DOI":"10.1109\/MSP.2017.2765202","volume":"35","author":"A Creswell","year":"2018","unstructured":"Creswell, A., White, T., Dumoulin, V., Arulkumaran, K., Sengupta, B., & Bharath, A. A. (2018). Generative adversarial networks: An overview. IEEE Signal Processing Magazine, 35(1), 53\u201365.","journal-title":"IEEE Signal Processing Magazine"},{"issue":"6","key":"9987_CR11","doi-asserted-by":"publisher","first-page":"3066","DOI":"10.1109\/TSP.2010.2044260","volume":"58","author":"A Daher","year":"2010","unstructured":"Daher, A., Baghious, E. H., Burel, G., & Radoi, E. (2010). Overlap-save and overlap-add filters: Optimal design and comparison. IEEE Transactions on Signal Processing, 58(6), 3066\u20133075.","journal-title":"IEEE Transactions on Signal Processing"},{"key":"9987_CR12","doi-asserted-by":"crossref","unstructured":"Das, N., Chakraborty, S., Chaki, J., Padhy, N., & Dey, N. (2020). Fundamentals, present and future perspectives of speech enhancement. International Journal of Speech Technology, 24(4), 883\u2013901.","DOI":"10.1007\/s10772-020-09674-2"},{"issue":"5","key":"9987_CR13","doi-asserted-by":"publisher","first-page":"138","DOI":"10.1109\/MSP.2019.2924687","volume":"36","author":"A Deleforge","year":"2019","unstructured":"Deleforge, A., Di Carlo, D., Strauss, M., Serizel, R., & Marcenaro, L. (2019). Audio-based search and rescue with a drone: highlights From the IEEE Signal Processing Cup 2019 Student Competition [SP Competitions]. IEEE Signal Processing Magazine, 36(5), 138\u2013144. https:\/\/doi.org\/10.1109\/MSP.2019.2924687.","journal-title":"IEEE Signal Processing Magazine"},{"key":"9987_CR14","unstructured":"Drakopoulos, F., Baby, D., & Verhulst, S. (2019). Real-time audio processing on a Raspberry Pi using deep neural networks. In Proceedings of the international congress on acoustics."},{"key":"9987_CR15","unstructured":"Haykin, S. (1996). Adaptive filter theory (5th ed.). Prentice-Hall."},{"key":"9987_CR16","unstructured":"Hirsch, H. G., & Pearce, D. (2000). The Aurora experimental framework for the performance evaluation of speech recognition systems under noisy conditions. In Proceedings of the automatic speech recognition: Challenges for the new millenium, ISCA Tutorial and Research Workshop (ITRW)."},{"issue":"1","key":"9987_CR17","doi-asserted-by":"publisher","first-page":"59","DOI":"10.1109\/TSA.2003.819949","volume":"12","author":"Y Hu","year":"2004","unstructured":"Hu, Y., & Loizou, P. C. (2004). Speech enhancement based on wavelet thresholding the multitaper spectrum. IEEE Transactions on Speech and Audio Processing, 12(1), 59\u201367.","journal-title":"IEEE Transactions on Speech and Audio Processing"},{"key":"9987_CR18","unstructured":"Hu, Y., & Loizou, P. C. (2006). Subjective comparison of speech enhancement algorithms. In Proceedings of the EEE international conference on acoustics speech and signal processing proceedings (Vol. 1, pp. 153\u2013156)."},{"issue":"1","key":"9987_CR19","doi-asserted-by":"publisher","first-page":"229","DOI":"10.1109\/TASL.2007.911054","volume":"16","author":"Y Hu","year":"2007","unstructured":"Hu, Y., & Loizou, P. C. (2007). Evaluation of objective quality measures for speech enhancement. IEEE Transactions on Audio, Speech, and Language Processing, 16(1), 229\u2013238.","journal-title":"IEEE Transactions on Audio, Speech, and Language Processing"},{"key":"9987_CR20","unstructured":"Islam, M. T., Shahnaz, C., Zhu, W. P., Ahmad, M. O., et al. (2018). Speech enhancement in adverse environments based on non-stationary noise-driven spectral subtraction and snr-dependent phase compensation. arXiv preprint arXiv:1803.00396."},{"key":"9987_CR21","doi-asserted-by":"crossref","unstructured":"Jaiswal, R., & Romero, D. (2021). Implicit Wiener filtering for speech enhancement in non-stationary noise. In 11th international conference on information science and technology (ICIST), IEEE (pp. 39\u201347).","DOI":"10.1109\/ICIST52614.2021.9440639"},{"key":"9987_CR22","doi-asserted-by":"crossref","unstructured":"Kamath, S., Loizou, P., (2002). A multi-band spectral subtraction method for enhancing speech corrupted by colored noise. In ICASSP. IEEE.","DOI":"10.1109\/ICASSP.2002.5745591"},{"key":"9987_CR23","unstructured":"Kanehara, S., Saruwatari, H., Miyazaki, R., Shikano, K., & Kondo, K. (2012). Comparative study on various noise reduction methods with decision-directed a priori snr estimator via higher-order statistics. In Proceedings of The Asia Pacific Signal and Information Processing Association Annual Summit and Conference, IEEE (pp. 1\u20136)."},{"key":"9987_CR24","doi-asserted-by":"crossref","unstructured":"Kleijn, W. B., Lim, F. S., Luebs, A., Skoglund, J., Stimberg, F., Wang, Q., & Walters, T. C. (2018). Wavenet based low rate speech coding. In Proceedings of the IEEE international conference on acoustics, speech and signal processing (ICASSP) (pp. 676\u2013680).","DOI":"10.1109\/ICASSP.2018.8462529"},{"issue":"12","key":"9987_CR25","doi-asserted-by":"publisher","first-page":"1586","DOI":"10.1109\/PROC.1979.11540","volume":"67","author":"JS Lim","year":"1979","unstructured":"Lim, J. S., & Oppenheim, A. V. (1979). Enhancement and bandwidth compression of noisy speech. Proceedings of the IEEE, 67(12), 1586\u20131604.","journal-title":"Proceedings of the IEEE"},{"key":"9987_CR26","doi-asserted-by":"crossref","unstructured":"Loizou, P. C. (2013). Speech enhancement: Theory and practice (2nd ed.). CRC Press.","DOI":"10.1201\/b14529"},{"key":"9987_CR27","doi-asserted-by":"publisher","first-page":"574","DOI":"10.1016\/j.csl.2016.11.003","volume":"46","author":"AH Moore","year":"2017","unstructured":"Moore, A. H., Parada, P. P., & Naylor, P. A. (2017). Speech enhancement for robust automatic speech recognition: Evaluation using a baseline system and instrumental measures. Computer Speech & Language, 46, 574\u2013584.","journal-title":"Computer Speech & Language"},{"key":"9987_CR28","doi-asserted-by":"crossref","unstructured":"Ogunfunmi, T., Togneri, R., & Narasimha, M. (2015). Speech and audio processing for coding, enhancement and recognition. Springer.","DOI":"10.1007\/978-1-4939-1456-2"},{"key":"9987_CR29","doi-asserted-by":"publisher","first-page":"10","DOI":"10.1016\/j.specom.2019.09.001","volume":"114","author":"S Pascual","year":"2019","unstructured":"Pascual, S., Serr\u00e0, J., & Bonafonte, A. (2019). Time-domain speech enhancement using generative adversarial networks. Speech Communication, 114, 10\u201321. https:\/\/doi.org\/10.1016\/j.specom.2019.09.001","journal-title":"Speech Communication"},{"key":"9987_CR30","doi-asserted-by":"crossref","unstructured":"Piczak, K. J. (2015). ESC: Dataset for environmental sound classification. In Proceedings of the ACM international conference on multimedia (pp. 1015\u20131018).","DOI":"10.1145\/2733373.2806390"},{"key":"9987_CR31","doi-asserted-by":"crossref","unstructured":"Saldanha, J. C., & Shruthi, O. R. (2016). Reduction of noise for speech signal enhancement using spectral subtraction method. In Proceedings of the IEEE international conference on information science (ICIS) (pp. 44\u201347).","DOI":"10.1109\/INFOSCI.2016.7845298"},{"key":"9987_CR32","doi-asserted-by":"crossref","unstructured":"Schultz, B. G., Tarigoppula, V. S. A., Noffs, G., Rojas, S., van der Walt, A., Grayden, D. B., & Vogel, A. P. (2021). Automatic speech recognition in neurodegenerative disease. International Journal of Speech Technology 24(3) , 771\u2013779.","DOI":"10.1007\/s10772-021-09836-w"},{"issue":"1","key":"9987_CR33","doi-asserted-by":"publisher","first-page":"562","DOI":"10.1121\/1.2918540","volume":"124","author":"S Sheft","year":"2008","unstructured":"Sheft, S., Ardoint, M., & Lorenzi, C. (2008). Speech identification based on temporal fine structure cues. The Journal of the Acoustical Society of America, 124(1), 562\u2013575.","journal-title":"The Journal of the Acoustical Society of America"},{"key":"9987_CR34","doi-asserted-by":"publisher","first-page":"53040","DOI":"10.1109\/ACCESS.2019.2912200","volume":"7","author":"A Shrestha","year":"2019","unstructured":"Shrestha, A., & Mahmood, A. (2019). Review of deep learning algorithms and architectures. IEEE Access, 7, 53040\u201353065.","journal-title":"IEEE Access"},{"key":"9987_CR35","doi-asserted-by":"crossref","unstructured":"Srinivasarao, V., & Ghanekar, U. (2020). Speech intelligibility enhancement: A hybrid Wiener approach. International Journal of Speech Technology, 23(3), 517\u2013525.","DOI":"10.1007\/s10772-020-09737-4"},{"key":"9987_CR36","doi-asserted-by":"crossref","unstructured":"Vaseghi, S. V. (2008). Advanced digital signal processing and noise reduction (4th ed.). Wiley.","DOI":"10.1002\/9780470740156"},{"key":"9987_CR37","doi-asserted-by":"publisher","unstructured":"Yamazaki, Y., Tamaki, M., Premachandra, C., Perera, C. J., Sumathipala, S., & Sudantha, B. H. (2019). Victim detection using UAV with on-board voice recognition system. In Proceedings of the IEEE international conference on robotic computing (IRC) (pp. 555\u2013559). https:\/\/doi.org\/10.1109\/IRC.2019.00114","DOI":"10.1109\/IRC.2019.00114"},{"key":"9987_CR38","doi-asserted-by":"publisher","first-page":"35","DOI":"10.1016\/j.specom.2020.06.005","volume":"123","author":"X Yan","year":"2020","unstructured":"Yan, X., Yang, Z., Wang, T., & Guo, H. (2020). An iterative graph spectral subtraction method for speech enhancement. Speech Communication, 123, 35\u201342. https:\/\/doi.org\/10.1016\/j.specom.2020.06.005","journal-title":"Speech Communication"},{"key":"9987_CR39","doi-asserted-by":"publisher","first-page":"30","DOI":"10.1016\/j.specom.2017.08.007","volume":"94","author":"CH You","year":"2017","unstructured":"You, C. H., & Ma, B. (2017). Spectral-domain speech enhancement for speech recognition. Speech Communication, 94, 30\u201341. https:\/\/doi.org\/10.1016\/j.specom.2017.08.007","journal-title":"Speech Communication"},{"key":"9987_CR40","doi-asserted-by":"publisher","first-page":"142","DOI":"10.1016\/j.specom.2020.10.007","volume":"125","author":"H Yu","year":"2020","unstructured":"Yu, H., Zhu, W. P., & Champagne, B. (2020). Speech enhancement using a DNN-augmented colored-noise Kalman filter. Speech Communication, 125, 142\u2013151. https:\/\/doi.org\/10.1016\/j.specom.2020.10.007.","journal-title":"Speech Communication"},{"key":"9987_CR41","doi-asserted-by":"publisher","first-page":"75","DOI":"10.1016\/j.specom.2020.09.002","volume":"124","author":"W Yuan","year":"2020","unstructured":"Yuan, W. (2020). A time-frequency smoothing neural network for speech enhancement. Speech Communication, 124, 75\u201384. https:\/\/doi.org\/10.1016\/j.specom.2020.09.002","journal-title":"Speech Communication"}],"container-title":["International Journal of Speech Technology"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-022-09987-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10772-022-09987-4\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-022-09987-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,8,10]],"date-time":"2022-08-10T11:46:35Z","timestamp":1660131995000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10772-022-09987-4"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,8,1]]},"references-count":41,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2022,9]]}},"alternative-id":["9987"],"URL":"https:\/\/doi.org\/10.1007\/s10772-022-09987-4","relation":{},"ISSN":["1381-2416","1572-8110"],"issn-type":[{"value":"1381-2416","type":"print"},{"value":"1572-8110","type":"electronic"}],"subject":[],"published":{"date-parts":[[2022,8,1]]},"assertion":[{"value":"24 August 2021","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"17 June 2022","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 August 2022","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare that they have no conflict of interest.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of Interest"}}]}}