{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,6]],"date-time":"2026-06-06T16:46:13Z","timestamp":1780764373895,"version":"3.54.1"},"reference-count":24,"publisher":"Springer Science and Business Media LLC","issue":"5","license":[{"start":{"date-parts":[[2024,1,22]],"date-time":"2024-01-22T00:00:00Z","timestamp":1705881600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,1,22]],"date-time":"2024-01-22T00:00:00Z","timestamp":1705881600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Curtin University"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Circuits Syst Signal Process"],"published-print":{"date-parts":[[2024,5]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Speech presence probability (SPP) and gain functions such as Wiener filter or MMSE estimators require an estimate of the a-priori signal-to-noise ratio (SNR). However, the estimation of the a-priori SNR is computationally involved and sensitive to noise variations. This paper proposes to approximate the SPP and the overall gain function of a speech enhancement system by using sigmoid functions to reduce the need of estimating the a-prior SNR. By applying an approximation via the sigmoid functions it is shown that only the a-posteriori estimate of SNR is needed, resulting in a low complexity system. The sigmoid function is designed with an optimization algorithm to optimize its parameters with respect to speech quality measures. The optimization algorithm is based on the idea that the solution obtained for a given problem should move towards the best solution and avoid the worst solution. The proposed algorithm requires minimal control parameters and does not require any algorithm specific parameters. Simulation results show that the proposed sigmoid functions achieve good results in terms of speech quality measures when compared with existing methods while providing significantly lower complexity for implementation.<\/jats:p>","DOI":"10.1007\/s00034-023-02549-2","type":"journal-article","created":{"date-parts":[[2024,1,22]],"date-time":"2024-01-22T20:02:03Z","timestamp":1705953723000},"page":"2891-2908","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":4,"title":["Optimized Sigmoid Functions for Speech Presence Probability and Gain Function in Speech Enhancement"],"prefix":"10.1007","volume":"43","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-9244-7533","authenticated-orcid":false,"given":"Hai Huyen","family":"Dam","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sven","family":"Nordholm","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Pei Chee","family":"Yong","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Siow Yong","family":"Low","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,1,22]]},"reference":[{"key":"2549_CR1","doi-asserted-by":"publisher","first-page":"113","DOI":"10.1109\/TASSP.1979.1163209","volume":"27","author":"S Boll","year":"1979","unstructured":"S. Boll, Suppression of acoustic noise in speech using spectral subtraction. IEEE Trans. Acoust. Speech Signal Process. 27, 113\u2013120 (1979)","journal-title":"IEEE Trans. Acoust. Speech Signal Process."},{"issue":"4","key":"2549_CR2","doi-asserted-by":"publisher","first-page":"478","DOI":"10.1109\/LSP.2014.2303498","volume":"21","author":"KY Chan","year":"2014","unstructured":"K.Y. Chan, S. Nordholm, S.Y. Low, P.C. Yong, K.F.C. Yiu, A hybrid descent method for optimal sigmoid filter design. IEEE Signal Process. Lett. 21(4), 478\u2013482 (2014)","journal-title":"IEEE Signal Process. Lett."},{"issue":"5","key":"2549_CR3","doi-asserted-by":"publisher","first-page":"466","DOI":"10.1109\/TSA.2003.811544","volume":"11","author":"I Cohen","year":"2003","unstructured":"I. Cohen, Noise spectrum estimation in adverse environments: Improved minima controlled recursive averaging. IEEE Trans. Speech Audio Process. 11(5), 466\u2013475 (2003)","journal-title":"IEEE Trans. Speech Audio Process."},{"issue":"12","key":"2549_CR4","doi-asserted-by":"publisher","first-page":"2289","DOI":"10.1109\/TASLP.2018.2862641","volume":"26","author":"G Enzner","year":"2018","unstructured":"G. Enzner, P. Thune, Bayesian MMSE filtering of noisy speech by SNR marginalization with global PSD priors. IEEE\/ACM Trans. Audio Speech Lang. Process. 26(12), 2289\u20132304 (2018)","journal-title":"IEEE\/ACM Trans. Audio Speech Lang. Process."},{"issue":"6","key":"2549_CR5","doi-asserted-by":"publisher","first-page":"1109","DOI":"10.1109\/TASSP.1984.1164453","volume":"32","author":"Y Ephraim","year":"1984","unstructured":"Y. Ephraim, D. Malah, Speech enhancement using a minimummean square error short-time spectral amplitude estimator. IEEE Trans. Acoust. Speech Signal Process. 32(6), 1109\u20131121 (1984)","journal-title":"IEEE Trans. Acoust. Speech Signal Process."},{"issue":"6","key":"2549_CR6","doi-asserted-by":"publisher","first-page":"1109","DOI":"10.1109\/TASSP.1984.1164453","volume":"32","author":"Y Ephraim","year":"1984","unstructured":"Y. Ephraim, D. Malah, Speech enhancement using a minimum-mean square error short-time spectral amplitude estimator. IEEE Trans. Acoust. Speech Signal Process. 32(6), 1109\u20131121 (1984)","journal-title":"IEEE Trans. Acoust. Speech Signal Process."},{"issue":"2","key":"2549_CR7","doi-asserted-by":"publisher","first-page":"443","DOI":"10.1109\/TASSP.1985.1164550","volume":"33","author":"Y Ephraim","year":"1985","unstructured":"Y. Ephraim, D. Malah, Speech enhancement using a minimum mean-square error log-spectral amplitude estimator. IEEE Trans. Acoust. Speech Signal Process. 33(2), 443\u2013445 (1985)","journal-title":"IEEE Trans. Acoust. Speech Signal Process."},{"issue":"4","key":"2549_CR8","doi-asserted-by":"publisher","first-page":"1383","DOI":"10.1109\/TASL.2011.2180896","volume":"20","author":"T Gerkmann","year":"2012","unstructured":"T. Gerkmann, R.C. Hendriks, Unbiased MMSE-based noise power estimation with low complexity and low tracking Delay. IEEE Trans. Audio Speech Language Process. 20(4), 1383\u20131393 (2012)","journal-title":"IEEE Trans. Audio Speech Language Process."},{"key":"2549_CR9","doi-asserted-by":"publisher","first-page":"588","DOI":"10.1016\/j.specom.2006.12.006","volume":"49","author":"Y Hu","year":"2007","unstructured":"Y. Hu, P. Loizou, Subjective evaluation and comparison of speech enhancement algorithms. Speech Commun. 49, 588\u2013601 (2007)","journal-title":"Speech Commun."},{"key":"2549_CR10","doi-asserted-by":"publisher","DOI":"10.1201\/9781420015836","volume-title":"Speech Enhancement Theory and Practice","author":"P Loizou","year":"2007","unstructured":"P. Loizou, Speech Enhancement Theory and Practice (CRC Press, Boca Raton, FL, 2007)"},{"key":"2549_CR11","doi-asserted-by":"crossref","unstructured":"S. Y. Low, An insight into the rise time of exponential smoothing for speech enhancement methods, in IEEE International Conference Signal Image Process Applications, pp. 30\u201333 (2021)","DOI":"10.1109\/ICSIPA52582.2021.9576801"},{"key":"2549_CR12","doi-asserted-by":"publisher","first-page":"504","DOI":"10.1109\/89.928915","volume":"9","author":"R Martin","year":"2001","unstructured":"R. Martin, Noise power spectral density estimation based on optimal smoothing and minimum statistics. IEEE Trans. Speech Audio Process. 9, 504\u2013512 (2001)","journal-title":"IEEE Trans. Speech Audio Process."},{"key":"2549_CR13","first-page":"1","volume":"1","author":"L Nahma","year":"2019","unstructured":"L. Nahma, P.C. Yong, H.H. Dam, S. Nordholm, An adaptive a-priori SNR estimator for perceptual speech enhancement. EURASIP J. Audio Speech Music Process. 1, 1 (2019)","journal-title":"EURASIP J. Audio Speech Music Process."},{"issue":"2","key":"2549_CR14","doi-asserted-by":"publisher","first-page":"282","DOI":"10.1016\/j.specom.2011.09.003","volume":"54","author":"K Paliwal","year":"2012","unstructured":"K. Paliwal, B. Schwerin, K. Wo, Speech enhancement using a minimum mean-square error short-time spectral modulation magnitude estimator. Speech Commun. 54(2), 282\u2013305 (2012)","journal-title":"Speech Commun."},{"key":"2549_CR15","volume-title":"Objective Measures of Speech Quality","author":"S Quackenbush","year":"1988","unstructured":"S. Quackenbush, T. Barnwell, M. Clements, Objective Measures of Speech Quality (Prientice Hall, Englewood Cliffs, 1988)"},{"key":"2549_CR16","first-page":"19","volume":"7","author":"RV Rao","year":"2016","unstructured":"R.V. Rao, Jaya: a simple and new optimizaton algorithm for solving constrained and unconstrained optimization problems. Int. J. Eng. Comput. 7, 19\u201334 (2016)","journal-title":"Int. J. Eng. Comput."},{"key":"2549_CR17","first-page":"749","volume":"2","author":"AW Rix","year":"2001","unstructured":"A.W. Rix, J.G. Beerends, M.P. Hollier, A.P. Hekstra, Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codec. IEEE Int. Conf. Acoust. Speech Signal Process. 2, 749\u2013752 (2001)","journal-title":"IEEE Int. Conf. Acoust. Speech Signal Process."},{"key":"2549_CR18","unstructured":"T. Rohdenburg, V. Hohmann, B. Kollmeier, Objective perceptual quality measures for the evaluation of noise reduction schemes, in 9th International Workshop on Acoustic Echo and Noise Control, pp. 169\u2013172 (2005)"},{"key":"2549_CR19","doi-asserted-by":"crossref","unstructured":"P. Scalart, Speech enhancement based on a-priori signal to noise estimation, in IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP\u201996), 629\u2013632 (1996)","DOI":"10.1109\/ICASSP.1996.543199"},{"key":"2549_CR20","doi-asserted-by":"publisher","first-page":"81","DOI":"10.1016\/j.specom.2017.11.008","volume":"96","author":"MK Singh","year":"2018","unstructured":"M.K. Singh, S.Y. Low, S. Nordholm, Z. Zang, Bayesian noise estimation in the modulation domain. Speech Commun. 96, 81\u201392 (2018)","journal-title":"Speech Commun."},{"issue":"7","key":"2549_CR21","doi-asserted-by":"publisher","first-page":"125","DOI":"10.1109\/TASL.2011.2114881","volume":"19","author":"CH Taal","year":"2011","unstructured":"C.H. Taal, R.C. Hendriks, R. Heusdens, J. Jensen, An algorithm for intelligibility prediction of time frequency weighted noisy speech. IEEE Trans. Audio Speech Lang. Process. 19(7), 125\u20132136 (2011)","journal-title":"IEEE Trans. Audio Speech Lang. Process."},{"issue":"2","key":"2549_CR22","doi-asserted-by":"publisher","first-page":"358","DOI":"10.1016\/j.specom.2012.09.004","volume":"55","author":"PC Yong","year":"2012","unstructured":"P.C. Yong, S. Nordholm, H.H. Dam, Optimization and evaluation of sigmoid function with a priori SNR estimate. Speech Commun. 55(2), 358\u2013376 (2012)","journal-title":"Speech Commun."},{"key":"2549_CR23","unstructured":"P. C. Yong, S. Nordholm, H. H. Dam, Noise estimation based on soft decisions and conditional smoothing for speech enhancement, in International Workshop on Acoustic Signal Enhancement (2012)"},{"issue":"2","key":"2549_CR24","doi-asserted-by":"publisher","first-page":"358","DOI":"10.1016\/j.specom.2012.09.004","volume":"55","author":"PC Yong","year":"2013","unstructured":"P.C. Yong, S. Nordholm, H.H. Dam, Optimization and evaluation of sigmoid function with a priori SNR estimate for real-time speech enhancement. Speech Commun. 55(2), 358\u2013376 (2013)","journal-title":"Speech Commun."}],"container-title":["Circuits, Systems, and Signal Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00034-023-02549-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00034-023-02549-2\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00034-023-02549-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,11,8]],"date-time":"2024-11-08T21:45:56Z","timestamp":1731102356000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00034-023-02549-2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,1,22]]},"references-count":24,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2024,5]]}},"alternative-id":["2549"],"URL":"https:\/\/doi.org\/10.1007\/s00034-023-02549-2","relation":{},"ISSN":["0278-081X","1531-5878"],"issn-type":[{"value":"0278-081X","type":"print"},{"value":"1531-5878","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,1,22]]},"assertion":[{"value":"23 November 2022","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"18 October 2023","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"21 October 2023","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"22 January 2024","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}