{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,1]],"date-time":"2025-12-01T11:23:24Z","timestamp":1764588204163,"version":"3.37.3"},"reference-count":25,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2021,7,27]],"date-time":"2021-07-27T00:00:00Z","timestamp":1627344000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2021,7,27]],"date-time":"2021-07-27T00:00:00Z","timestamp":1627344000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61901227"],"award-info":[{"award-number":["61901227"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Natural Science Foundation of the Jiangsu Higher Education Institutions","award":["19KJB510049"],"award-info":[{"award-number":["19KJB510049"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J AUDIO SPEECH MUSIC PROC."],"published-print":{"date-parts":[[2021,12]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>To improve the performance of speech enhancement in a complex noise environment, a joint constrained dictionary learning method for single-channel speech enhancement is proposed, which solves the \u201ccross projection\u201d problem of signals in the joint dictionary. In the method, the new optimization function not only constrains the sparse representation of the noisy signal in the joint dictionary, and controls the projection error of the speech signal and noise signal on the corresponding sub-dictionary, but also minimizes the cross projection error and the correlation between the sub-dictionaries. In addition, the adjustment factors are introduced to balance the weight of constraint terms to obtain the joint dictionary more discriminatively. When the method is applied to the single-channel speech enhancement, speech components of the noisy signal can be more projected onto the clean speech sub-dictionary of the joint dictionary without being affected by the noise sub-dictionary, which makes the quality and intelligibility of the enhanced speech higher. The experimental results verify that our algorithm has better performance than the speech enhancement algorithm based on discriminative dictionary learning under white noise and colored noise environments in time domain waveform, spectrogram, global signal-to-noise ratio, subjective evaluation of speech quality, and logarithmic spectrum distance.<\/jats:p>","DOI":"10.1186\/s13636-021-00218-3","type":"journal-article","created":{"date-parts":[[2021,7,27]],"date-time":"2021-07-27T07:03:11Z","timestamp":1627369391000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":6,"title":["Single-channel speech enhancement based on joint constrained dictionary learning"],"prefix":"10.1186","volume":"2021","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-9442-9964","authenticated-orcid":false,"given":"Linhui","family":"Sun","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yunyi","family":"Bu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pingan","family":"Li","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zihao","family":"Wu","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2021,7,27]]},"reference":[{"issue":"6","key":"218_CR1","doi-asserted-by":"publisher","first-page":"641","DOI":"10.1049\/iet-spr.2015.0182","volume":"10","author":"S Samui","year":"2016","unstructured":"S. Samui, I. Chakrabarti, S.K. Ghosh, Improved single channel phase-aware speech enhancement technique for low signal-to-noise ratio signal. IET Signal Process. 10(6), 641\u2013650 (2016). https:\/\/doi.org\/10.1049\/iet-spr.2015.0182","journal-title":"IET Signal Process."},{"key":"218_CR2","doi-asserted-by":"publisher","unstructured":"T. Lavanya, T. Nagarajan, P. Vijayalakshmi, Multi-level single-channel speech enhancement using a unified framework for estimating magnitude and phase spectra, IEEE\/ACM Transactions on Audio, Speech, and Language Processing. 28, 1315-1327 (2020). https:\/\/doi.org\/10.1109\/TASLP.2020.2986877","DOI":"10.1109\/TASLP.2020.2986877"},{"key":"218_CR3","doi-asserted-by":"publisher","unstructured":"P. Kajla, N.V. George, Speech quality enhancement using a two channel sparse adaptive filtering approach, Applied Acoustics. 158 (2020). https:\/\/doi.org\/10.1016\/j.apacoust.2019.107035","DOI":"10.1016\/j.apacoust.2019.107035"},{"key":"218_CR4","doi-asserted-by":"publisher","first-page":"333","DOI":"10.1016\/j.apacoust.2018.07.027","volume":"141","author":"N Saleem","year":"2018","unstructured":"N. Saleem, M.I. Khattak, M. Shafi, Unsupervised speech enhancement in low SNR environments via sparseness and temporal gradient regularization. Appl. Acoustics. 141, 333\u2013347 (2018). https:\/\/doi.org\/10.1016\/j.apacoust.2018.07.027","journal-title":"Appl. Acoustics."},{"key":"218_CR5","doi-asserted-by":"publisher","first-page":"30","DOI":"10.1016\/j.specom.2017.08.007","volume":"94","author":"C You","year":"2017","unstructured":"C. You, B. Ma, Spectral-domain speech enhancement for speech recognition. Speech Commun.. 94, 30\u201341 (2017). https:\/\/doi.org\/10.1016\/j.specom.2017.08.007","journal-title":"Speech Commun.."},{"key":"218_CR6","doi-asserted-by":"publisher","unstructured":"W. Zaw, A. T. H. Soe, Speaker identification using power spectral subtraction method, 2019 16th International Conference on Electrical Engineering\/Electronics, Computer, Telecommunications and Information Technology (ECTI-CON). 625-628 (2019). https:\/\/doi.org\/10.1109\/ECTI-CON47248.2019.8955344.","DOI":"10.1109\/ECTI-CON47248.2019.8955344"},{"issue":"3","key":"218_CR7","doi-asserted-by":"publisher","first-page":"477","DOI":"10.1016\/j.specom.2011.10.009","volume":"54","author":"J Choi","year":"2012","unstructured":"J. Choi, J. Chang, On using acoustic environment classification for statistical model-based speech enhancement. Speech Communication. 54(3), 477\u2013490 (2012). https:\/\/doi.org\/10.1016\/j.specom.2011.10.009","journal-title":"Speech Communication."},{"key":"218_CR8","doi-asserted-by":"publisher","unstructured":"A. Salman, E. Muhammad, K. Khurshid, A subspace approach for speech enhancement using frame-level AdaBoost classification, 2007 International Conference on Electrical Engineering. 1-6 (2007). https:\/\/doi.org\/10.1109\/ICEE.2007.4287303","DOI":"10.1109\/ICEE.2007.4287303"},{"issue":"12","key":"218_CR9","doi-asserted-by":"publisher","first-page":"1195","DOI":"10.1109\/LSP.2013.2285218","volume":"20","author":"M Sadeghi","year":"2013","unstructured":"M. Sadeghi, M. Babaie-Zadeh, C. Jutten, Dictionary learning for sparse representation: a novel approach. IEEE Signal Process. Lett. 20(12), 1195\u20131198 (2013). https:\/\/doi.org\/10.1109\/LSP.2013.2285218","journal-title":"IEEE Signal Process. Lett."},{"issue":"6","key":"218_CR10","doi-asserted-by":"publisher","first-page":"1045","DOI":"10.1109\/JPROC.2010.2040551","volume":"98","author":"R Rubinstein","year":"2010","unstructured":"R. Rubinstein, A.M. Bruckstein, M. Elad, Dictionaries for sparse representation modeling. Proc. IEEE. 98(6), 1045\u20131057 (2010). https:\/\/doi.org\/10.1109\/JPROC.2010.2040551","journal-title":"Proc. IEEE."},{"key":"218_CR11","doi-asserted-by":"publisher","first-page":"71","DOI":"10.1016\/j.specom.2016.09.004","volume":"85","author":"V Abrol","year":"2016","unstructured":"V. Abrol, P. Sharma, A.K. Sao, Greedy double sparse dictionary learning for sparse representation of speech signals. Speech Commun. 85, 71\u201382 (2016). https:\/\/doi.org\/10.1016\/j.specom.2016.09.004","journal-title":"Speech Commun."},{"issue":"11","key":"218_CR12","doi-asserted-by":"publisher","first-page":"4311","DOI":"10.1109\/TSP.2006.881199","volume":"54","author":"M Aharon","year":"2006","unstructured":"M. Aharon, M. Elad, A. Bruckstein, K-SVD: an algorithm for designing overcomplete dictionaries for sparse representations. IEEE Transact. Signal Process. 54(11), 4311\u20134322 (2006). https:\/\/doi.org\/10.1109\/TSP.2006.881199","journal-title":"IEEE Transact. Signal Process."},{"issue":"6","key":"218_CR13","doi-asserted-by":"publisher","first-page":"3055","DOI":"10.1109\/TSP.2010.2044251","volume":"58","author":"BV Gowreesunker","year":"2010","unstructured":"B.V. Gowreesunker, A.H. Tewfik, Learning sparse representation using iterative subspace identification. IEEE Transact. Signal Process. 58(6), 3055\u20133065 (2010). https:\/\/doi.org\/10.1109\/TSP.2010.2044251","journal-title":"IEEE Transact. Signal Process."},{"key":"218_CR14","doi-asserted-by":"publisher","first-page":"85","DOI":"10.1016\/j.specom.2018.11.008","volume":"106","author":"L Sun","year":"2019","unstructured":"L. Sun, K. Xie, T. Gu, J. Chen, Z. Yang, Joint dictionary learning using a new optimization method for single-channel blind source separation. Speech Commun. 106, 85\u201394 (2019). https:\/\/doi.org\/10.1016\/j.specom.2018.11.008","journal-title":"Speech Commun."},{"key":"218_CR15","doi-asserted-by":"publisher","unstructured":"M. Islam, Y. Zhu, M. I. Hossain, R. Ullan, Z. Ye, Supervised single channel dual domains speech enhancement using sparse non-negative matrix factorization, Digital Signal Process. 100 (2020). https:\/\/doi.org\/10.1016\/j.dsp.2020.102697","DOI":"10.1016\/j.dsp.2020.102697"},{"issue":"6","key":"218_CR16","doi-asserted-by":"publisher","first-page":"1698","DOI":"10.1109\/TASL.2012.2187194","volume":"20","author":"CD Sigg","year":"2012","unstructured":"C.D. Sigg, T. Dikk, J.M. Buhmann, Speech enhancement using generative dictionary learning, IEEE Transactions on Audio. Speech Language Process. 20(6), 1698\u20131712 (2012). https:\/\/doi.org\/10.1109\/TASL.2012.2187194","journal-title":"Speech Language Process."},{"issue":"10","key":"218_CR17","doi-asserted-by":"publisher","first-page":"2140","DOI":"10.1109\/TASL.2013.2270369","volume":"21","author":"N Mohammadiha","year":"2017","unstructured":"N. Mohammadiha, P. Smaragdis, A. Leijon, Supervised and unsupervised speech enhancement using nonnegative matrix factorization. IEEE Transact Audio Speech Language Process. 21(10), 2140\u20132151 (2017). https:\/\/doi.org\/10.1109\/TASL.2013.2270369","journal-title":"IEEE Transact Audio Speech Language Process."},{"key":"218_CR18","doi-asserted-by":"publisher","unstructured":"D. Baby, T. Virtanen, J. F. Gemmeke, H. Van hamme, Coupled dictionaries for exemplar-based speech enhancement and automatic speech recognition, IEEE Transact. Audio Speech Language Process. 23(11), 1788-1799 (2015). https:\/\/doi.org\/10.1109\/TASLP.2015.2450491","DOI":"10.1109\/TASLP.2015.2450491"},{"key":"218_CR19","doi-asserted-by":"publisher","unstructured":"P. Sprechmann, A. Bronstein, M. Bronstein, G. Sapiro, Learnable low rank sparse models for speech denoising, 2013 IEEE International Conference on Acoustics. Speech and Signal Process., 136\u2013140 (2013). https:\/\/doi.org\/10.1109\/ICASSP.2013.6637624","DOI":"10.1109\/ICASSP.2013.6637624"},{"key":"218_CR20","doi-asserted-by":"publisher","unstructured":"L. Zhang, G. Bao, Y. Luo, Z. Ye, Monaural speech enhancement using joint dictionary learning with cross-coherence penalties, international symposium on computational intelligence & design. 518-522 (2015). https:\/\/doi.org\/10.1109\/ISCID.2015.162","DOI":"10.1109\/ISCID.2015.162"},{"key":"218_CR21","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1016\/j.apacoust.2017.11.005","volume":"132","author":"J Fu","year":"2018","unstructured":"J. Fu, L. Zhang, Z. Ye, Supervised monaural speech enhancement using two-level complementary joint sparse representations. Appl. Acoustics. 132, 1\u20137 (2018). https:\/\/doi.org\/10.1016\/j.apacoust.2017.11.005","journal-title":"Appl. Acoustics."},{"issue":"03","key":"218_CR22","first-page":"74","volume":"46","author":"H Jia","year":"2019","unstructured":"H. Jia, W. Wang, Y. Wang, J. Pei, Speech enhancement based on discriminative joint sparse dictionary alternate optimization. J. Xidian Univ. 46(03), 74\u201381 (2019)","journal-title":"J. Xidian Univ."},{"key":"218_CR23","doi-asserted-by":"publisher","unstructured":"L. Sun, C. Zhao, M. Su, F. Wang, Single-channel blind source separation based on joint dictionary with common sub-dictionary, Int. J. Speech Technol. 21 19\u201327 (2018). https:\/\/doi.org\/10.1007\/s10772-017-9469-2","DOI":"10.1007\/s10772-017-9469-2"},{"key":"218_CR24","doi-asserted-by":"publisher","unstructured":"F. F. Firouzeh, S. Ghorshi, S. Salsabili, Compressed sensing based speech enhancement, 2014 8th International Conference on Signal Processing and Communication Systems (ICSPCS). 1-6 (2014). https:\/\/doi.org\/10.1109\/ICSPCS.2014.7021068","DOI":"10.1109\/ICSPCS.2014.7021068"},{"key":"218_CR25","doi-asserted-by":"publisher","unstructured":"P. Qi, W. Zhou, J. Han, A method for stochastic L-BFGS optimization, 2017 IEEE 2nd International Conference on Cloud Computing and Big Data Analysis (ICCCBDA). 156-160 (2017). https:\/\/doi.org\/10.1109\/ICCCBDA.2017.7951902","DOI":"10.1109\/ICCCBDA.2017.7951902"}],"container-title":["EURASIP Journal on Audio, Speech, and Music Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13636-021-00218-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s13636-021-00218-3\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s13636-021-00218-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,7,27]],"date-time":"2021-07-27T07:28:41Z","timestamp":1627370921000},"score":1,"resource":{"primary":{"URL":"https:\/\/asmp-eurasipjournals.springeropen.com\/articles\/10.1186\/s13636-021-00218-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,7,27]]},"references-count":25,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2021,12]]}},"alternative-id":["218"],"URL":"https:\/\/doi.org\/10.1186\/s13636-021-00218-3","relation":{},"ISSN":["1687-4722"],"issn-type":[{"type":"electronic","value":"1687-4722"}],"subject":[],"published":{"date-parts":[[2021,7,27]]},"assertion":[{"value":"30 May 2021","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 July 2021","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"27 July 2021","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"Not applicable.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"29"}}