{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2024,2,15]],"date-time":"2024-02-15T04:40:20Z","timestamp":1707972020814},"reference-count":54,"publisher":"Springer Science and Business Media LLC","issue":"1","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["EURASIP J. Adv. Signal Process."],"published-print":{"date-parts":[[2007,12]]},"DOI":"10.1155\/2007\/64102","type":"journal-article","created":{"date-parts":[[2007,6,19]],"date-time":"2007-06-19T07:21:11Z","timestamp":1182237671000},"source":"Crossref","is-referenced-by-count":5,"title":["A Comprehensive Noise Robust Speech Parameterization Algorithm Using Wavelet Packet Decomposition-Based Denoising and Speech Feature Representation Techniques"],"prefix":"10.1186","volume":"2007","author":[{"given":"Bojan","family":"Kotnik","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zdravko","family":"Ka\u010di\u010d","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2007,6,10]]},"reference":[{"key":"2042_CR1","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-4613-1297-0","volume-title":"Robustness in Automatic Speech Recognition","author":"J-C Junqua","year":"1996","unstructured":"Junqua J-C, Haton JP: Robustness in Automatic Speech Recognition. Kluwer Academic, Boston, Mass, USA; 1996."},{"issue":"4","key":"2042_CR2","doi-asserted-by":"publisher","first-page":"567","DOI":"10.1109\/89.326615","volume":"2","author":"JB Allen","year":"1994","unstructured":"Allen JB: How do humans process and recognize speech? IEEE Transactions on Speech and Audio Processing 1994,2(4):567-577. 10.1109\/89.326615","journal-title":"IEEE Transactions on Speech and Audio Processing"},{"key":"2042_CR3","volume-title":"Model-based techniques for noise robust speech recognition, Ph.D. thesis","author":"MJF Gales","year":"1996","unstructured":"Gales MJF: Model-based techniques for noise robust speech recognition, Ph.D. thesis. University of Cambridge, Cambridge, UK; 1996."},{"issue":"3","key":"2042_CR4","doi-asserted-by":"publisher","first-page":"261","DOI":"10.1016\/0167-6393(94)00059-J","volume":"16","author":"Y Gong","year":"1995","unstructured":"Gong Y: Speech recognition in noisy environments: a survey. Speech Communication 1995,16(3):261-291. 10.1016\/0167-6393(94)00059-J","journal-title":"Speech Communication"},{"key":"2042_CR5","unstructured":"ETSI standard document - ETSI ES 201 108 v1.1.1 : Speech Processing, Transmission and Quality aspects (STQ), Distributed speech recognition, Front-end feature extraction algorithm, Compression algorithm. 2000."},{"key":"2042_CR6","unstructured":"ETSI standard document - ETSI ES 202 050 v1.1.1 : Speech Processing, Transmission and Quality aspects (STQ), Distributed speech recognition, Advanced front-end feature extraction algorithm, Compression algorithm. 2002."},{"issue":"4","key":"2042_CR7","doi-asserted-by":"publisher","first-page":"357","DOI":"10.1109\/TASSP.1980.1163420","volume":"28","author":"SB Davis","year":"1980","unstructured":"Davis SB, Mermelstein P: Comparison of parametric representations for monosyllabic word recognition in continuously spoken sentences. IEEE Transactions on Acoustics, Speech, and Signal Processing 1980,28(4):357-366. 10.1109\/TASSP.1980.1163420","journal-title":"IEEE Transactions on Acoustics, Speech, and Signal Processing"},{"key":"2042_CR8","first-page":"1251","volume":"2","author":"H Bourlard","year":"1997","unstructured":"Bourlard H, Dupont S: Subband-based speech recognition. Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '97), April 1997, Munich, Germany 2: 1251-1254.","journal-title":"Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '97)"},{"key":"2042_CR9","first-page":"1351","volume":"3","author":"JN Gowdy","year":"2000","unstructured":"Gowdy JN, Tufekci Z: Mel-scaled discrete wavelet coefficients for speech recognition. Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '00), June 2000, Istanbul, Turkey 3: 1351-1354.","journal-title":"Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '00)"},{"key":"2042_CR10","first-page":"445","volume-title":"Proceedings of IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU '01)","author":"M Gupta","year":"2001","unstructured":"Gupta M, Gilbert A: Robust speech recognition using wavelet coefficient features. Proceedings of IEEE Workshop on Automatic Speech Recognition and Understanding (ASRU '01), December 2001, Madonna di Campiglio, Trento, Italy 445-448."},{"issue":"4","key":"2042_CR11","doi-asserted-by":"publisher","first-page":"1738","DOI":"10.1121\/1.399423","volume":"87","author":"H Hermansky","year":"1990","unstructured":"Hermansky H: Perceptual linear predictive (PLP) analysis of speech. The Journal of the Acoustical Society of America 1990,87(4):1738-1752. 10.1121\/1.399423","journal-title":"The Journal of the Acoustical Society of America"},{"issue":"2","key":"2042_CR12","doi-asserted-by":"publisher","first-page":"80","DOI":"10.1016\/1051-2004(92)90028-W","volume":"2","author":"KK Paliwal","year":"1992","unstructured":"Paliwal KK: On the use of line spectral frequency parameters for speech recognition. Digital Signal Processing 1992,2(2):80-87. 10.1016\/1051-2004(92)90028-W","journal-title":"Digital Signal Processing"},{"key":"2042_CR13","volume-title":"Discrete-Time Processing of Speech Signals","author":"JR Deller","year":"1993","unstructured":"Deller JR, Proakis JG, Hansen JHL: Discrete-Time Processing of Speech Signals. Macmillan, New York, NY, USA; 1993."},{"key":"2042_CR14","volume-title":"Fundamentals of Speech Recognition","author":"L Rabiner","year":"1993","unstructured":"Rabiner L, Juang B-H: Fundamentals of Speech Recognition. Prentice Hall, Upper Saddle River, NJ, USA; 1993. section 4.5"},{"issue":"2, part 2","key":"2042_CR15","doi-asserted-by":"publisher","first-page":"713","DOI":"10.1109\/18.119732","volume":"38","author":"RR Coifman","year":"1992","unstructured":"Coifman RR, Wickerhauser MV: Entropy-based algorithms for best basis selection. IEEE Transactions on Information Theory 1992,38(2, part 2):713-718. 10.1109\/18.119732","journal-title":"IEEE Transactions on Information Theory"},{"key":"2042_CR16","volume-title":"Ten Lectures on Wavelets","author":"I Daubechies","year":"1997","unstructured":"Daubechies I: Ten Lectures on Wavelets. SIAM, Philadelphia, Pa, USA; 1997."},{"issue":"7","key":"2042_CR17","doi-asserted-by":"publisher","first-page":"674","DOI":"10.1109\/34.192463","volume":"11","author":"SG Mallat","year":"1989","unstructured":"Mallat SG: A theory for multiresolution signal decomposition: the wavelet representation. IEEE Transactions on Pattern Analysis and Machine Intelligence 1989,11(7):674-693. 10.1109\/34.192463","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"2042_CR18","volume-title":"Wavelets and Filter Banks","author":"G Strang","year":"1997","unstructured":"Strang G, Nguyen T: Wavelets and Filter Banks. Wellesley-Cambridge Press, Wellesley, Mass, USA; 1997."},{"issue":"2-3","key":"2042_CR19","doi-asserted-by":"publisher","first-page":"409","DOI":"10.1016\/S0167-6393(03)00011-6","volume":"41","author":"C-T Lu","year":"2003","unstructured":"Lu C-T, Wang H-C: Enhancement of single channel speech based on masking property and wavelet transform. Speech Communication 2003,41(2-3):409-427. 10.1016\/S0167-6393(03)00011-6","journal-title":"Speech Communication"},{"key":"2042_CR20","first-page":"81","volume-title":"Proceedings of the 3rd IEEE Nordic Signal Processing Symposium (NORSIG '98)","author":"R Sarikaya","year":"1998","unstructured":"Sarikaya R, Pellom BL, Hansen JHL: Wavelet packet transform features with application to speaker identification. Proceedings of the 3rd IEEE Nordic Signal Processing Symposium (NORSIG '98), June 1998, Vigs\u00f8, Denmark 81-84."},{"key":"2042_CR21","doi-asserted-by":"crossref","first-page":"1855","DOI":"10.21437\/Eurospeech.2001-438","volume-title":"Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01)","author":"H Sheikhzadeh","year":"2001","unstructured":"Sheikhzadeh H, Abutalebi HR: An improved wavelet-based speech enhancement system. Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01), September 2001, Aalborg, Denmark 1855-1858."},{"issue":"4","key":"2042_CR22","doi-asserted-by":"publisher","first-page":"541","DOI":"10.1109\/5.488699","volume":"84","author":"K Ramchandran","year":"1996","unstructured":"Ramchandran K, Vetterli M, Herley C: Wavelets, subband coding, and best bases. Proceedings of the IEEE 1996,84(4):541-560. 10.1109\/5.488699","journal-title":"Proceedings of the IEEE"},{"issue":"5","key":"2042_CR23","doi-asserted-by":"publisher","first-page":"919","DOI":"10.1016\/S0165-1684(02)00489-9","volume":"83","author":"NR Reyes","year":"2003","unstructured":"Reyes NR, Zurera MR, Ferreras FL, Amores PJ: Adaptive wavelet-packet analysis for audio coding purposes. Signal Processing 2003,83(5):919-929. 10.1016\/S0165-1684(02)00489-9","journal-title":"Signal Processing"},{"issue":"1","key":"2042_CR24","doi-asserted-by":"publisher","first-page":"10","DOI":"10.1109\/97.889636","volume":"8","author":"M Bahoura","year":"2001","unstructured":"Bahoura M, Rouat J: Wavelet speech enhancement based on the Teager energy operator. IEEE Signal Processing Letters 2001,8(1):10-12. 10.1109\/97.889636","journal-title":"IEEE Signal Processing Letters"},{"issue":"3","key":"2042_CR25","doi-asserted-by":"publisher","first-page":"613","DOI":"10.1109\/18.382009","volume":"41","author":"DL Donoho","year":"1995","unstructured":"Donoho DL: De-noising by soft-thresholding. IEEE Transactions on Information Theory 1995,41(3):613-627. 10.1109\/18.382009","journal-title":"IEEE Transactions on Information Theory"},{"key":"2042_CR26","doi-asserted-by":"crossref","first-page":"569","DOI":"10.21437\/Eurospeech.2003-228","volume-title":"Proceedings of the 8th European Conference on Speech Communication and Technology (EUROSPEECH '03)","author":"E Jafer","year":"2003","unstructured":"Jafer E, Mahdi AE: Wavelet-based perceptual speech enhancement using adaptive threshold estimation. Proceedings of the 8th European Conference on Speech Communication and Technology (EUROSPEECH '03), September 2003, Geneva, Switzerland 569-572."},{"key":"2042_CR27","volume-title":"Wavelet thresholding and noise reduction, Ph.D. thesis","author":"M Jansen","year":"2000","unstructured":"Jansen M: Wavelet thresholding and noise reduction, Ph.D. thesis. Katholieke Universiteit Leuven, Leuven, Belgium; 2000."},{"key":"2042_CR28","doi-asserted-by":"crossref","first-page":"193","DOI":"10.21437\/Eurospeech.2001-71","volume-title":"Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01)","author":"B Andrassy","year":"2001","unstructured":"Andrassy B, Vlaj D, Beaugeant C: Recognition performance of the siemens front-end with and without frame dropping on the aurora 2 database. Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01), September 2001, Aalborg, Denmark 193-196."},{"key":"2042_CR29","first-page":"181","volume-title":"Proceedings of the Automatic Speech Recognition: Challanges for the New Millennium (ISCA ITRW ASR '00)","author":"H-G Hirsch","year":"2000","unstructured":"Hirsch H-G, Pearce D: The aurora experimental framework for the performance evaluation of speech recognition systems under noisy conditions. Proceedings of the Automatic Speech Recognition: Challanges for the New Millennium (ISCA ITRW ASR '00), September 2000, Paris, France 181-188."},{"key":"2042_CR30","volume-title":"Proceedings of Applied Voice Input\/Output Society Conference (AVIOS '00)","author":"D Pearce","year":"2000","unstructured":"Pearce D: Enabling new speech driven services for mobile devices: an overview of the ETSI standards activities for distributed speech recognition front-ends. Proceedings of Applied Voice Input\/Output Society Conference (AVIOS '00), May 2000, San Jose, Calif, USA"},{"key":"2042_CR31","unstructured":"AU\/225\/00 : Baseline Results for Subset of SpeechDat-Car Finnish Database for ETSI STQ WI008 Advanced Front-end Evaluation. Nokia, Janurary 2000"},{"key":"2042_CR32","unstructured":"AU\/271\/00 : Spanish SDC-Aurora Database for ETSI STQ Aurora WI008 Advanced DSR Front-End Evaluation: Description and Baseline Results. UPC, November 2000"},{"key":"2042_CR33","unstructured":"AU\/273\/00 : Description and Baseline Results for the Subset of the Speechdat-Car German Database used for ETSI STQ Aurora WI008 Advanced DSR Front-end Evaluation. Texas Instruments, December 2001"},{"key":"2042_CR34","unstructured":"AU\/378\/01 : Danish SpeechDat-Car Digits Database for ETSI STQ-Aurora Advanced DSR. Aalborg University, January 2001"},{"key":"2042_CR35","first-page":"17","volume-title":"Proceedings of the 7th International Conference on Spoken Language Processing (ICSLP '02)","author":"D Macho","year":"2002","unstructured":"Macho D, Mauuary L, Noe B, et al.: Evaluation of a noise-robust DSR front-end on aurora database. Proceedings of the 7th International Conference on Spoken Language Processing (ICSLP '02), September 2002, Denver, Colo, USA 17-20."},{"issue":"6","key":"2042_CR36","doi-asserted-by":"publisher","first-page":"1109","DOI":"10.1109\/TASSP.1984.1164453","volume":"32","author":"Y Ephraim","year":"1984","unstructured":"Ephraim Y, Malah D: Speech enhancement using a minimum mean-square error short-time spectral amplitude estimator. IEEE Transactions on Acoustics, Speech, and Signal Processing 1984,32(6):1109-1121. 10.1109\/TASSP.1984.1164453","journal-title":"IEEE Transactions on Acoustics, Speech, and Signal Processing"},{"issue":"3","key":"2042_CR37","doi-asserted-by":"publisher","first-page":"205","DOI":"10.1023\/A:1023410018862","volume":"6","author":"B Kotnik","year":"2003","unstructured":"Kotnik B, Vlaj D, Horvat B: Efficient noise robust feature extraction algorithms for distributed speech recognition (DSR) systems. International Journal of Speech Technology 2003,6(3):205-219. 10.1023\/A:1023410018862","journal-title":"International Journal of Speech Technology"},{"key":"2042_CR38","first-page":"1182","volume-title":"Proceedings of the European Signal Processing Conference (EUSIPCO '94)","author":"R Martin","year":"1994","unstructured":"Martin R: Spectral subtraction based on minimum statistics. Proceedings of the European Signal Processing Conference (EUSIPCO '94), September 1994, Edinburgh, UK 1182-1185."},{"issue":"2","key":"2042_CR39","doi-asserted-by":"publisher","first-page":"46","DOI":"10.1109\/35.17653","volume":"27","author":"D O'Shaughnessy","year":"1989","unstructured":"O'Shaughnessy D: Enhancing speech degraded by additive noise or interfering speakers. IEEE Communications Magazine 1989,27(2):46-52. 10.1109\/35.17653","journal-title":"IEEE Communications Magazine"},{"issue":"6","key":"2042_CR40","doi-asserted-by":"publisher","first-page":"697","DOI":"10.1109\/TCT.1973.1083764","volume":"20","author":"JH McClellan","year":"1973","unstructured":"McClellan JH, Parks TW: A unified approach to the design of optimum FIR linear-phase digital filters. IEEE Transactions on Circuits Theory 1973,20(6):697-701.","journal-title":"IEEE Transactions on Circuits Theory"},{"issue":"8","key":"2042_CR41","doi-asserted-by":"publisher","first-page":"550","DOI":"10.1109\/82.318943","volume":"41","author":"O Rioul","year":"1994","unstructured":"Rioul O, Duhamel P: A Remez exchange algorithm for orthonormal wavelets. IEEE Transactions on Circuits and Systems II: Analog and Digital Signal Processing 1994,41(8):550-560. 10.1109\/82.318943","journal-title":"IEEE Transactions on Circuits and Systems II: Analog and Digital Signal Processing"},{"issue":"2","key":"2042_CR42","doi-asserted-by":"publisher","first-page":"113","DOI":"10.1109\/TASSP.1979.1163209","volume":"27","author":"SF Boll","year":"1979","unstructured":"Boll SF: Suppression of acoustic noise in speech using spectral subtraction. IEEE Transactions on Acoustics, Speech, and Signal Processing 1979,27(2):113-120. 10.1109\/TASSP.1979.1163209","journal-title":"IEEE Transactions on Acoustics, Speech, and Signal Processing"},{"issue":"4","key":"2042_CR43","doi-asserted-by":"publisher","first-page":"694","DOI":"10.1044\/jshr.3604.694","volume":"36","author":"JM Hillenbrand","year":"1993","unstructured":"Hillenbrand JM, Gayvert RT: Vowel classification based on fundamental frequency and formant frequencies. Journal of Speech and Hearing Research 1993,36(4):694-700.","journal-title":"Journal of Speech and Hearing Research"},{"key":"2042_CR44","volume-title":"A Study of Voice Activity Detectors. Speech Communications 304-523B","author":"M Klein","year":"2000","unstructured":"Klein M: A Study of Voice Activity Detectors. Speech Communications 304-523B, McGill University, Computer and Electrical Engineering Department, April 2000"},{"key":"2042_CR45","first-page":"269","volume":"1","author":"B Mak","year":"1992","unstructured":"Mak B, Junqua J-C, Reaves B: A robust speech\/non-speech detection algorithm using time and frequency-based features. Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '92), March 1992, San Francisco, Calif, USA 1: 269-272.","journal-title":"Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '92)"},{"issue":"3","key":"2042_CR46","doi-asserted-by":"publisher","first-page":"217","DOI":"10.1109\/89.905996","volume":"9","author":"E Nemer","year":"2001","unstructured":"Nemer E, Gourbran R, Mahmoud S: Robust voice activity detection using higher-order statistics in the LPC residual domain. IEEE Transactions on Speech and Audio Processing 2001,9(3):217-231. 10.1109\/89.905996","journal-title":"IEEE Transactions on Speech and Audio Processing"},{"key":"2042_CR47","first-page":"365","volume":"1","author":"J Sohn","year":"1998","unstructured":"Sohn J, Sung W: A voice activity detector employing soft decision based noise spectrum adaptation. Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '98), May 1998, Seattle, Wash, USA 1: 365-368.","journal-title":"Proceedings of IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP '98)"},{"issue":"1","key":"2042_CR48","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1109\/97.736233","volume":"6","author":"J Sohn","year":"1999","unstructured":"Sohn J, Kim NS, Sung W: A statistical model-based voice activity detection. IEEE Signal Processing Letters 1999,6(1):1-3. 10.1109\/97.736233","journal-title":"IEEE Signal Processing Letters"},{"key":"2042_CR49","volume-title":"The HTK Book\u2014Version 3.0","author":"S Young","year":"2000","unstructured":"Young S, Kershaw D, Odell J, Ollason D, Valtchev V, Woodland P: The HTK Book\u2014Version 3.0. Microsoft, Redmond, Wash, USA; 2000."},{"issue":"6","key":"2042_CR50","doi-asserted-by":"publisher","first-page":"1304","DOI":"10.1121\/1.1914702","volume":"55","author":"BS Atal","year":"1974","unstructured":"Atal BS: Effectiveness of linear prediction characteristics of the speech wave for automatic speaker identification and verification. The Journal of the Acoustical Society of America 1974,55(6):1304-1312. 10.1121\/1.1914702","journal-title":"The Journal of the Acoustical Society of America"},{"key":"2042_CR51","doi-asserted-by":"crossref","first-page":"865","DOI":"10.21437\/Eurospeech.2001-264","volume-title":"Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01)","author":"F de Wet","year":"2001","unstructured":"de Wet F, Cranen B, de Veth J, Boves L: A comparison of LPC and FFT-based acoustic features for noise robust ASR. Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01), September 2001, Aalborg, Denmark 865-868."},{"key":"2042_CR52","doi-asserted-by":"crossref","first-page":"687","DOI":"10.21437\/Eurospeech.2001-194","volume-title":"Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01)","author":"R Sarikaya","year":"2001","unstructured":"Sarikaya R, Hansen JHL: Analysis of the root-cepstrum for acoustic modeling and fast decoding in speech recognition. Proceedings of the 7th European Conference on Speech Communication and Technology (EUROSPEECH '01), September 2001, Aalborg, Denmark 687-690."},{"key":"2042_CR53","first-page":"2083","volume-title":"Proceedings the 4th International Conference on Language Resources and Evaluation (LREC '04)","author":"B Kotnik","year":"2004","unstructured":"Kotnik B, Ka\u010di\u010d Z, Horvat B: Development and integration of the LDA-toolkit into the COST249 speechdat (II) SIG reference recognizer. Proceedings the 4th International Conference on Language Resources and Evaluation (LREC '04), May 2004, Lisbon, Portugal 2083-2086."},{"key":"2042_CR54","volume-title":"Merkmalsextraction in spracherkennungssystemen f\u00fcr grossen wortschatz, Ph.D. thesis","author":"L Welling","year":"1999","unstructured":"Welling L: Merkmalsextraction in spracherkennungssystemen f\u00fcr grossen wortschatz, Ph.D. thesis. Rheinisch-Westf\u00e4lische Technische Hochschule, Aachen, Germany; 1999."}],"container-title":["EURASIP Journal on Advances in Signal Processing"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1155\/2007\/64102.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,2,15]],"date-time":"2024-02-15T04:01:04Z","timestamp":1707969664000},"score":1,"resource":{"primary":{"URL":"https:\/\/asp-eurasipjournals.springeropen.com\/articles\/10.1155\/2007\/64102"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2007,6,10]]},"references-count":54,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2007,12]]}},"alternative-id":["2042"],"URL":"https:\/\/doi.org\/10.1155\/2007\/64102","relation":{},"ISSN":["1687-6180"],"issn-type":[{"value":"1687-6180","type":"electronic"}],"subject":[],"published":{"date-parts":[[2007,6,10]]},"article-number":"064102"}}