{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,30]],"date-time":"2026-03-30T19:30:16Z","timestamp":1774899016969,"version":"3.50.1"},"reference-count":55,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2024,10,7]],"date-time":"2024-10-07T00:00:00Z","timestamp":1728259200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"},{"start":{"date-parts":[[2024,10,7]],"date-time":"2024-10-07T00:00:00Z","timestamp":1728259200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.springernature.com\/gp\/researchers\/text-and-data-mining"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Speech Technol"],"published-print":{"date-parts":[[2024,12]]},"DOI":"10.1007\/s10772-024-10148-y","type":"journal-article","created":{"date-parts":[[2024,10,7]],"date-time":"2024-10-07T15:03:04Z","timestamp":1728313384000},"page":"997-1012","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Deep learning for speech denoising with improved Wiener approach"],"prefix":"10.1007","volume":"27","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2110-4672","authenticated-orcid":false,"given":"Ouardia","family":"Abdelli","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fatiha","family":"Merazka","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2024,10,7]]},"reference":[{"issue":"4","key":"10148_CR1","doi-asserted-by":"publisher","first-page":"1211","DOI":"10.1109\/LCOMM.2020.3045665","volume":"25","author":"I Ahmed","year":"2020","unstructured":"Ahmed, I., Alam, S., Hossain, J., & Kaddoum, G. (2020). Deep learning for MMSE estimation of a Gaussian source in the presence of bursty impulsive noise. IEEE Communications Letters, 25(4), 1211\u20131215. https:\/\/doi.org\/10.1109\/LCOMM.2020.3045665","journal-title":"IEEE Communications Letters"},{"key":"10148_CR2","doi-asserted-by":"publisher","first-page":"7","DOI":"10.1109\/TASLP.2018.2868407","volume":"27","author":"F Bao","year":"2019","unstructured":"Bao, F., & Abdulla, W. H. (2019). A new ratio mask representation for casa-based speech enhancement. IEEE\/ACM Trans Audio Speech Lang Process, 27, 7\u201319.","journal-title":"IEEE\/ACM Trans Audio Speech Lang Process"},{"key":"10148_CR3","doi-asserted-by":"publisher","first-page":"e0196924","DOI":"10.1371\/journal.pone.0196924","volume":"13","author":"T Bentsen","year":"2018","unstructured":"Bentsen, T., May, T., Kressner, A. A., & Dau, T. (2018). The benefit of combining a deep neural network architecture with ideal ratio mask estimation in computational speech segregation to improve speech intelligibility. PLoS ONE, 13, e0196924.","journal-title":"PLoS ONE"},{"key":"10148_CR4","first-page":"207","volume-title":"DNN based mask estimation for supervised speech separation","author":"J Chen","year":"2018","unstructured":"Chen, J., & Wang, D. (2018). DNN based mask estimation for supervised speech separation (pp. 207\u2013235). Springer."},{"key":"10148_CR5","unstructured":"Chung, H. (2018). Speech enhancement using training-based non-negative matrix factorization techniques. Master Thesis, Department of Electrical & Computer Engineering McGill University Montreal."},{"key":"10148_CR6","doi-asserted-by":"crossref","unstructured":"Dean, D. B., Sridharan, S., Vogt, R. J. & Mason, M. W. (2010). TheQUT-NOISETIMIT corpus for the evaluation of voice activity detection algorithms. In Proceedings of Interspeech, (pp. 3110\u20133113).","DOI":"10.21437\/Interspeech.2010-774"},{"issue":"5","key":"10148_CR7","doi-asserted-by":"publisher","first-page":"1085","DOI":"10.1109\/TASLP.2017.2687829","volume":"25","author":"M Delfarah","year":"2017","unstructured":"Delfarah, M., & Wang, D. L. (2017). Features for masking-based monaural speech separation in reverberant conditions. IEEE\/ACM Transactions on Audio, Speech, and Language Processing, 25(5), 1085\u20131094.","journal-title":"IEEE\/ACM Trans Audio Speech Lang Process"},{"key":"10148_CR8","doi-asserted-by":"crossref","unstructured":"Dimitriadis, D., Maragos, P., & Potamianos, A. (2005) Auditory Teager energy cepstrum coeffi-cients for robust speech recognition. In Proceedings Eurospeech.","DOI":"10.21437\/Interspeech.2005-142"},{"key":"10148_CR9","doi-asserted-by":"crossref","unstructured":"Duong, H. T. T., Nguyen, Q. C., Nguyen, C. P., Tran, T. H., & Duong, N. Q. (2015). Speech enhancement based on non negative matrix factorization with mixed group sparsity constraint. In 6th ACM international symposium on information and communication technology, (pp. 247\u2013251).","DOI":"10.1145\/2833258.2833276"},{"issue":"6","key":"10148_CR10","doi-asserted-by":"publisher","first-page":"1109","DOI":"10.1109\/TASSP.1984.1164453","volume":"32","author":"Y Ephraim","year":"1984","unstructured":"Ephraim, Y., & Malah, D. (1984). Speech enhancement using a minimum-mean square error short-time spectral amplitude estimator. IEEE Transaction Acoustics, Speech, and Signal Processing, 32(6), 1109\u20131121.","journal-title":"IEEE Transaction Acoustics, Speech, and Signal Process."},{"issue":"2","key":"10148_CR11","doi-asserted-by":"publisher","first-page":"443","DOI":"10.1109\/TASSP.1985.1164550","volume":"33","author":"Y Ephraim","year":"1985","unstructured":"Ephraim, Y., & Malah, D. (1985). Speech enhancement using a minimum mean-square error log spectral amplitude estimator. IEEE Transaction on Acoustics, Speech, and Signal Processing, 33(2), 443\u2013445. https:\/\/doi.org\/10.1109\/TASSP.1985.1164550","journal-title":"IEEE Transaction on Acoustics, Speech, and Signal Processing"},{"key":"10148_CR12","doi-asserted-by":"crossref","unstructured":"Garofolo, J. S., Lamel, L. F., Fisher, W. M., Fiscus, J. G., & Pallett, D. S. (1993). DARPA TIMIT acoustic-phonetic continuous speech corpus CDROM. NIST speech disc 1\u20131.1. NASA STI\/Recon Techn. Rep. N, vol. 93","DOI":"10.6028\/NIST.IR.4930"},{"key":"10148_CR13","doi-asserted-by":"publisher","first-page":"235","DOI":"10.1109\/CC.2018.8456465","volume":"15","author":"B Haichuan","year":"2018","unstructured":"Haichuan, B., Fengpei, G., & Yonghong, Y. (2018). DNN-based speech enhancement using soft audible noise masking for wind noise reduction. China Commun, 15, 235\u2013243.","journal-title":"China Commun"},{"key":"10148_CR14","doi-asserted-by":"crossref","unstructured":"Han, W., Zhang, X., Min, G., Zhou, X., & Zhang, W. (2016). Perceptual weighting deep neural networks for single-channel speech enhancement. Intelligent Control and Automation, 446\u2013450","DOI":"10.1109\/WCICA.2016.7578300"},{"key":"10148_CR15","doi-asserted-by":"publisher","first-page":"1738","DOI":"10.1121\/1.399423","volume":"87","author":"H Hermansky","year":"1990","unstructured":"Hermansky, H. (1990). Perceptual linear predictive (PLP) analysis of speech. The Journal of the Acoustical Society of America, 87, 1738\u20131752.","journal-title":"J. Acoust. Soc. Amer."},{"issue":"4","key":"10148_CR16","doi-asserted-by":"publisher","first-page":"578","DOI":"10.1109\/89.326616","volume":"2","author":"H Hermansky","year":"1994","unstructured":"Hermansky, H., & Morgan, N. (1994). RASTA processing of speech. IEEE Transaction Speech Audio Processing, 2(4), 578\u2013589.","journal-title":"IEEE Transaction Speech Audio Processing"},{"issue":"5","key":"10148_CR17","doi-asserted-by":"publisher","first-page":"9","DOI":"10.1016\/j.sysarc.2019.02.008","volume":"95","author":"TY Hsiao","year":"2019","unstructured":"Hsiao, T. Y., Chang, Y. C., Chou, H. H., & Lin, C. T. (2019). Filter-based deep compression with global average pooling for convolutional networks. Journal of Systems Architecture, 95(5), 9\u201318. https:\/\/doi.org\/10.1016\/j.sysarc.2019.02.008","journal-title":"Journal of Systems Architecture"},{"key":"10148_CR18","doi-asserted-by":"publisher","first-page":"153","DOI":"10.1016\/j.procs.2020.12.020","volume":"179","author":"N Jamal","year":"2021","unstructured":"Jamal, N., Fuad, N., Sha\u2019abani, M. N. A. H., & Shanta, S. (2021). Comparative study of IBM and IRM target mask for supervised Malay speech separation from noisy background. Procedia Computer Science, 179, 153\u2013160.","journal-title":"Procedia Computer Science."},{"key":"10148_CR19","doi-asserted-by":"publisher","first-page":"107666","DOI":"10.1016\/j.apacoust.2020.107666","volume":"171","author":"H Jia","year":"2021","unstructured":"Jia, H., Wang, W., & Mei, S. (2021). Combining adaptive sparse NMF feature extraction and soft mask to optimize DNN for speech enhancement. Applied Acoustics, 171, 107666.","journal-title":"Applied Acoustics"},{"key":"10148_CR20","doi-asserted-by":"publisher","first-page":"229","DOI":"10.1109\/LSP.2014.2354456","volume":"22","author":"TG Kang","year":"2015","unstructured":"Kang, T. G., Kwon, K., Shin, J. W., & Kim, N. S. (2015). NMF-based target source separation using deep neural network. IEEE Signal Processing Letters, 22, 229\u2013233.","journal-title":"IEEE Signal Processing Letters"},{"key":"10148_CR21","doi-asserted-by":"publisher","first-page":"102","DOI":"10.1016\/j.dsp.2017.12.002","volume":"74","author":"TG Kang","year":"2018","unstructured":"Kang, T. G., Shin, J. W., & Kim, N. S. (2018). DNN-based monaural speech enhancement with temporal and spectral variations equalization. Digital Signal Process, 74, 102\u2013110.","journal-title":"Digital Signal Process"},{"issue":"3","key":"10148_CR22","doi-asserted-by":"publisher","first-page":"1581","DOI":"10.1121\/1.3619790","volume":"130","author":"G Kim","year":"2011","unstructured":"Kim, G., & Loizou, P. C. (2011). Gain-induced speech distortions and the absence of intelligibility benefit with existing noise-reduction algorithms. The Journal of the Acoustical Society of America, 130(3), 1581\u20131596.","journal-title":"The Journal of the Acoustical Society of America."},{"key":"10148_CR23","doi-asserted-by":"publisher","first-page":"1486","DOI":"10.1121\/1.3184603","volume":"126","author":"G Kim","year":"2009","unstructured":"Kim, G., Lu, Y., Hu, Y., & Loizou, P. (2009). An algorithm that improves speech intelligibility in noise for normal-hearing listeners. The Journal of the Acoustical Society of America, 126, 1486\u20131494.","journal-title":"The Journal of the Acoustical Society of America"},{"issue":"5","key":"10148_CR24","doi-asserted-by":"publisher","first-page":"770","DOI":"10.1109\/LSP.2019.2905660","volume":"26","author":"J Kim","year":"2019","unstructured":"Kim, J., & Hahn, M. (2019). Speech enhancement using a two-stage network for an efficient boosting strategy. IEEE Signal Processing Letters, 26(5), 770\u2013774. https:\/\/doi.org\/10.1109\/LSP.2019.2905660","journal-title":"IEEE Signal Processing Letters"},{"key":"10148_CR25","doi-asserted-by":"publisher","unstructured":"Lea, C., Vidal, R., Reiter, A., & Hager, G. D. (2016). Temporal convolutional networks: A unified approach to action segmentation. In European conference on computer vision, (pp. 47\u201354). https:\/\/doi.org\/10.1007\/978-3-319-49409-8_7","DOI":"10.1007\/978-3-319-49409-8_7"},{"key":"10148_CR26","unstructured":"Mohammadiha, N. (2013). Speech enhancement using non-negative matrix factorization and hidden Markov models, PHD Thesis, Communication Theory Laboratory, School of Electrical Engineering, KTH Royal Institute of Technology."},{"issue":"8","key":"10148_CR27","doi-asserted-by":"publisher","first-page":"44","DOI":"10.1016\/j.specom.2019.06.002","volume":"111","author":"A Nicolson","year":"2019","unstructured":"Nicolson, A., & Paliwal, K. K. (2019). Deep learning for minimum mean square error approaches to speech enhancement. Speech Communication, 111(8), 44\u201355. https:\/\/doi.org\/10.1016\/j.specom.2019.06.002","journal-title":"Speech Communication"},{"key":"10148_CR28","doi-asserted-by":"publisher","first-page":"403","DOI":"10.1016\/j.csl.2019.06.004","volume":"58","author":"O Novotny","year":"2018","unstructured":"Novotny, O., Plchot, O., Glembek, O., \u010cernock\u00fd, J., & Burget, L. (2018). Analysis of DNN speech signal enhancement for robust speaker recognition. Computer Speech & Language, 58, 403\u2013421.","journal-title":"Computer Speech & Language"},{"issue":"1","key":"10148_CR29","first-page":"70","volume":"3","author":"A Ouardia","year":"2020","unstructured":"Ouardia, A., & Merazka, F. (2020). Denoising of speech signal using decision directed approach. International Journal of Informatics and Applied Mathematics, 3(1), 70\u201383.","journal-title":"International Journal of Informatics and Applied Mathematics"},{"key":"10148_CR30","doi-asserted-by":"publisher","unstructured":"Plapous, C., Marro, C., Mauuary, L., & Scalart, P. (2004). A two-step noise reduction technique. In IEEE international conference on acoustics, speech, and signal processing, (pp. 289\u2013292), Montreal. https:\/\/doi.org\/10.1109\/ICASSP.2004.1325979.","DOI":"10.1109\/ICASSP.2004.1325979"},{"key":"10148_CR31","doi-asserted-by":"crossref","unstructured":"Plapous, C., Marro, C., & Scalart, P. (2005). Speech enhancement using harmonic regeneration. In IEEE international conference on acoustics, speech, and signal processing.","DOI":"10.1109\/ICASSP.2005.1415074"},{"issue":"6","key":"10148_CR32","doi-asserted-by":"publisher","first-page":"2098","DOI":"10.1109\/TASL.2006","volume":"14","author":"C Plapous","year":"2006","unstructured":"Plapous, C., Marro, C., & Scalart, P. (2006). Improved signal-to-noise ratio estimation for speech enhancement. IEEE Transactions on Audio, Speech, and Language Processing, 14(6), 2098\u20132108. https:\/\/doi.org\/10.1109\/TASL.2006","journal-title":"IEEE Transactions on Audio, Speech, and Language Processing"},{"key":"10148_CR33","first-page":"749","volume":"2","author":"AW Rix","year":"2001","unstructured":"Rix, A. W., Beerends, J. G., Hollier, M. P., & Hekstra, A. P. (2001). Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs. In International Conference on Acoustics, Speech, and Signal Processing (ICASSP 2001) (pp. 749\u2013752).","journal-title":"IEEE Int Conf Acoust"},{"key":"10148_CR34","doi-asserted-by":"publisher","first-page":"225","DOI":"10.1109\/TAU.1969.1162058","volume":"17","author":"EH Rothauser","year":"1969","unstructured":"Rothauser, E. H., Chapman, W. D., Guttman, N., et al. (1969). IEEE recommended pratice for speech quality measurements. IEEE Transactions on Audio and Electroacoustics, 17, 225\u2013246.","journal-title":"IEEE Transactions on Audio and Electroacoustics"},{"key":"10148_CR35","doi-asserted-by":"publisher","DOI":"10.4316\/AECE.2022.02009","author":"M Salehi","year":"2022","unstructured":"Salehi, M., & Mirzakuchaki, S. (2022). Novel approach to speech enhancement based on deep neural networks. Advances in Electrical and Computer Engineering. https:\/\/doi.org\/10.4316\/AECE.2022.02009","journal-title":"Advances in Electrical and Computer Engineering"},{"key":"10148_CR36","doi-asserted-by":"crossref","unstructured":"Scalart, P. & Filho, J. V. (2016) Speech enhancement based on a priori signal to noise estimation. In Proceedings of the IEEE international conference on acoustics, speech, and signal processing (ICASSP) (pp. 629\u2013632).","DOI":"10.1109\/ICASSP.1996.543199"},{"key":"10148_CR37","doi-asserted-by":"publisher","first-page":"257","DOI":"10.1016\/j.apacoust.2016.04.024","volume":"117","author":"L Seongjae","year":"2017","unstructured":"Seongjae, L., David, K. H., & Hanseok, K. (2017). Single-channel speech enhancement method using reconstructive NMF with spectrotemporal speech presence probabilities. Applied Acoustics, 117, 257\u2013262.","journal-title":"Applied Acoustics"},{"key":"10148_CR38","doi-asserted-by":"publisher","unstructured":"Shekar, S., & Ravi, D. J. (2017). Denoising of a speech signal using wiener filter. In Proceedings of the international conference on current trends in engineering, science and technology. https:\/\/doi.org\/10.21647\/ICCTEST\/2017\/48935","DOI":"10.21647\/ICCTEST\/2017\/48935"},{"key":"10148_CR39","doi-asserted-by":"publisher","first-page":"663","DOI":"10.1016\/j.compeleceng.2017.02.021","volume":"62","author":"V Sunnydayal","year":"2017","unstructured":"Sunnydayal, V., & Kishore, K. T. (2017). Speech enhancement using posterior regularized NMF with bases update. Computers & Electrical Engineering, 62, 663\u2013675.","journal-title":"Computers & Electrical Engineering"},{"key":"10148_CR40","doi-asserted-by":"crossref","unstructured":"Taal, C. H., Hendriks, R. C., Heusdens, R., & Jensen, J. (2010) A short-time objective intelligibility measure for time-frequency weighted noisy speech. In IEEE international conference on acoustics, speech and signal processing, (pp. 4214\u20137).","DOI":"10.1109\/ICASSP.2010.5495701"},{"issue":"7","key":"10148_CR41","doi-asserted-by":"publisher","first-page":"2125","DOI":"10.1109\/TASL.2011.2114881","volume":"17","author":"CH Taal","year":"2011","unstructured":"Taal, C. H., Hendriks, R. C., Heusdens, R., & Jensen, J. (2011). An algorithm for intelligibility prediction of time\u2013frequency weighted noisy speech. IEEE Transactions on Audio, Speech, and Language Processing, 17(7), 2125\u20132136. https:\/\/doi.org\/10.1109\/TASL.2011.2114881","journal-title":"IEEE Transactions on Audio, Speech, and Language Processing"},{"issue":"1","key":"10148_CR42","doi-asserted-by":"publisher","first-page":"165","DOI":"10.1007\/s10772-020-09786-9","volume":"24","author":"YG Thimmaraja","year":"2021","unstructured":"Thimmaraja, Y. G., Nagaraja, B., & Jayanna, H. (2021). Speech enhancement and encoding by combining SS-VAD and LPC. International Journal of Speech Technology, 24(1), 165\u2013172. https:\/\/doi.org\/10.1007\/s10772-020-09786-9","journal-title":"International Journal of Speech Technology"},{"key":"10148_CR43","doi-asserted-by":"publisher","first-page":"205","DOI":"10.1016\/j.specom.2012.08.005","volume":"55","author":"H Veisi","year":"2013","unstructured":"Veisi, H., & Sameti, H. (2013). Speech enhancement using hidden Markov models in Mel-frequency domain. Speech Communication, 55, 205\u2013220. https:\/\/doi.org\/10.1016\/j.specom.2012.08.005","journal-title":"Speech Communication"},{"issue":"10","key":"10148_CR44","doi-asserted-by":"publisher","first-page":"1702","DOI":"10.1109\/TASLP.2018.2842159","volume":"26","author":"D Wang","year":"2018","unstructured":"Wang, D., & Chen, J. (2018). Supervised speech separation based on deep learning: An overview. IEEE\/ACM Transactions on Audio, Speech, and Language Processing, 26(10), 1702\u20131726.","journal-title":"IEEE\/ACM Transactions on Audio, Speech, and Language Processing."},{"key":"10148_CR45","unstructured":"Wang, D. & Chen, J. (2022). Supervised speech separation based on deep. In 2022 IEEE international conference on acoustics, speech and signal processing (ICASSP 2022) (pp. 1\u201327). https:\/\/ieeexplore.ieee.org\/xpl\/conhome\/9745891\/proceeding"},{"key":"10148_CR47","doi-asserted-by":"publisher","first-page":"2336","DOI":"10.1121\/1.3083233","volume":"125","author":"DL Wang","year":"2009","unstructured":"Wang, D. L., Kjems, U., Pedersen, M. S., Boldt, J. B., & Lunner, T. (2009). Speech intelligibility in background noise with ideal binary time-frequency masking. Journal of the Acoustical Society of America, 125, 2336\u20132347.","journal-title":"Journal of the Acoustical Society of America"},{"key":"10148_CR46","doi-asserted-by":"crossref","unstructured":"Wang, J., Yang, C., Yan, L., Huang, M., & Sang, J. (2018). Guangzhou University, Guangzhou, ChinaSpeech Enhancement Algorithm of Binary Mask Estimation Based on a Priori SNR Constraints Proceedings, APSIPA Annual Summit and Conference","DOI":"10.23919\/APSIPA.2018.8659475"},{"issue":"7","key":"10148_CR48","doi-asserted-by":"publisher","first-page":"1185","DOI":"10.1109\/TASLP.2018.2817798","volume":"26","author":"Q Wang","year":"2018","unstructured":"Wang, Q., Du, J., Dai, L. R., & Lee, C. H. (2018). A multiobjective learning and ensembling approach to high-performance speech enhancement with compact neural network architectures. IEEE\/ACM Transactions on Audio, Speech, and Language Processing, 26(7), 1185\u20131197. https:\/\/doi.org\/10.1109\/TASLP.2018.2817798","journal-title":"IEEE\/ACM Transactions on Audio, Speech, and Language Processing"},{"issue":"12","key":"10148_CR50","doi-asserted-by":"publisher","first-page":"1849","DOI":"10.1109\/TASLP.2014.2352935","volume":"22","author":"Y Wang","year":"2014","unstructured":"Wang, Y., Narayanan, A., & Wang, D. (2014). On training targets for supervised speech separation. IEEE\/ACM Transactions on Audio, Speech, and Language Processing, 22(12), 1849\u20131858.","journal-title":"IEEE\/ACM Transactions on Audio, Speech, and Language Processing."},{"issue":"7","key":"10148_CR51","doi-asserted-by":"publisher","first-page":"1381","DOI":"10.1109\/TASL.2013.2250961","volume":"21","author":"Y Wang","year":"2013","unstructured":"Wang, Y., & Wang, D. L. (2013). Towards scaling up classification-based speech separation. IEEE Transactions on Audio, Speech and Language Processing, 21(7), 1381\u20131390.","journal-title":"IEEE Transactions on Audio, Speech and Language Processing"},{"key":"10148_CR52","doi-asserted-by":"crossref","unstructured":"Yan, B., Bao, C., & Bai, Z. (2018). DNN-based speech enhancement via integrating nmf and casa. In International conference on audio, language and image processing (ICALIP) (pp. 435\u2013439).","DOI":"10.1109\/ICALIP.2018.8455780"},{"key":"10148_CR53","doi-asserted-by":"publisher","unstructured":"Yu, R. A. (2009). A low-complexity noise estimation algorithm based on smoothing of noise power estimation and estimation bias correction. In IEEE international conference on acoustics, speech and signal processing, (pp. 4421\u20134424). https:\/\/doi.org\/10.1109\/ICASSP.2009.4960610","DOI":"10.1109\/ICASSP.2009.4960610"},{"issue":"4","key":"10148_CR54","doi-asserted-by":"publisher","first-page":"1404","DOI":"10.1109\/TASLP.2020.2987441","volume":"28","author":"Q Zhang","year":"2020","unstructured":"Zhang, Q., Nicolson, A., Wang, M., Paliwal, K. K., & Wang, C. (2020). Deep MMSE: A deep learning approach to MMSE-based noise power spectral density estimation. IEEE\/ACM Transactions on Audio, Speech, and Language Processing, 28(4), 1404\u20131415. https:\/\/doi.org\/10.1109\/TASLP.2020.2987441","journal-title":"IEEE\/ACM Transactions on Audio, Speech, and Language Processing"},{"issue":"2","key":"10148_CR55","doi-asserted-by":"publisher","first-page":"252","DOI":"10.1109\/TASLP.2015.2505415","volume":"24","author":"XL Zhang","year":"2015","unstructured":"Zhang, X. L., & Wang, D. (2015). Boosting contextual information for deep neural network based voice activity detection. IEEE\/ACM Transaction on Audio, Speech, and Language Processing, 24(2), 252\u2013264. https:\/\/doi.org\/10.1109\/TASLP.2015.2505415","journal-title":"IEEE\/ACM Transaction on Audio, Speech, and Language Processing"},{"key":"10148_CR56","doi-asserted-by":"publisher","unstructured":"Zhao, Y., Wang, Z. Q., & Wang, D. (2017) A two-stage algorithm for noisy and reverberant speech enhancement. In Proceedings of the international conference on acoustics, speech and signal processing, (pp. 5580\u20135584). https:\/\/doi.org\/10.1109\/ICASSP.2017.7953224.","DOI":"10.1109\/ICASSP.2017.7953224"}],"container-title":["International Journal of Speech Technology"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-024-10148-y.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10772-024-10148-y\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-024-10148-y.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,12,16]],"date-time":"2024-12-16T10:07:02Z","timestamp":1734343622000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10772-024-10148-y"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,10,7]]},"references-count":55,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2024,12]]}},"alternative-id":["10148"],"URL":"https:\/\/doi.org\/10.1007\/s10772-024-10148-y","relation":{},"ISSN":["1381-2416","1572-8110"],"issn-type":[{"value":"1381-2416","type":"print"},{"value":"1572-8110","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,10,7]]},"assertion":[{"value":"19 May 2024","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"9 September 2024","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"7 October 2024","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}