{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2022,3,30]],"date-time":"2022-03-30T05:19:09Z","timestamp":1648617549739},"reference-count":36,"publisher":"Springer Science and Business Media LLC","issue":"3","license":[{"start":{"date-parts":[[2011,5,11]],"date-time":"2011-05-11T00:00:00Z","timestamp":1305072000000},"content-version":"tdm","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Int J Speech Technol"],"published-print":{"date-parts":[[2011,9]]},"DOI":"10.1007\/s10772-011-9092-6","type":"journal-article","created":{"date-parts":[[2011,5,10]],"date-time":"2011-05-10T13:51:49Z","timestamp":1305035509000},"page":"147-155","source":"Crossref","is-referenced-by-count":0,"title":["Robust features for multilingual acoustic modeling"],"prefix":"10.1007","volume":"14","author":[{"given":"C.","family":"Santhosh\u00a0Kumar","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"V. P.","family":"Mohandas","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2011,5,11]]},"reference":[{"key":"9092_CR1","first-page":"1451","volume-title":"Proc. IEEE int. conf. on acoustics, speech and signal processing","author":"U. Bub","year":"1997","unstructured":"Bub, U., Kohler, J., & Imperl, B. (1997). In-service adaptation of multilingual hidden Markov models. In Proc. IEEE int. conf. on acoustics, speech and signal processing, Munich (pp.\u00a01451\u20131454)."},{"key":"9092_CR2","volume-title":"Proc. IEEE int. conf. on acoustics, speech and signal processing","author":"L. Burget","year":"2010","unstructured":"Burget, L. et al. (2010). Multilingual acoustic modeling for speech recognition based on subspace Gaussian mixture models. In Proc. IEEE int. conf. on acoustics, speech and signal processing, Denver, USA."},{"key":"9092_CR3","volume-title":"Proc. int. conf. on spoken language processing","author":"N. Chatzichrisafis","year":"2004","unstructured":"Chatzichrisafis, N., Digalakis, V., Diakoloukas, V., & Harizakis, C. (2004). Rapid acoustic model development using Gaussian mixture clustering and language adaptation. In Proc. int. conf. on spoken language processing, Jeja Island, Korea."},{"key":"9092_CR4","doi-asserted-by":"crossref","first-page":"268","DOI":"10.1109\/PROC.1973.9030","volume":"61","author":"G. D. Forney","year":"1973","unstructured":"Forney, G. D. (1973). The Viterbi algorithm. Proceedings of the IEEE, 61, 268\u2013278.","journal-title":"Proceedings of the IEEE"},{"key":"9092_CR5","volume-title":"Proc. int. conf. on acoustics, speech and signal processing","author":"F. Grezl","year":"2008","unstructured":"Grezl, F., & Fousek, F. (2008). Optimizing bottle-neck features for LVCSR. In Proc. int. conf. on acoustics, speech and signal processing, Las Vegas, USA."},{"key":"9092_CR6","volume-title":"Proc. international conference on spoken language processing","author":"H. Hermansky","year":"1998","unstructured":"Hermansky, H., & Sharma, S. (1998). TRAPS\u2014classifiers of temporal patterns. In Proc. international conference on spoken language processing, Sydney, Australia."},{"key":"9092_CR7","volume-title":"Proc. Interspeech","author":"H. Hongbing","year":"2008","unstructured":"Hongbing, H., & Zahorian, S. A. (2008). A neural network based non-linear feature transformation for speech recognition. In Proc. Interspeech, Brisbane, Australia."},{"key":"9092_CR8","volume-title":"Proc. international conference on spoken language processing","author":"S. Itahashi","year":"2004","unstructured":"Itahashi, S., Zhu, S., & Yamamoto, M. (2004). Constructing family trees of multilingual speech using Gaussian mixture models. In Proc. international conference on spoken language processing, Jeju Island, Japan."},{"key":"9092_CR9","volume-title":"Fundamentals of digital image processing","author":"A. K. Jain","year":"1989","unstructured":"Jain, A. K. (1989). Fundamentals of digital image processing. Englewood Cliffs: Prentice Hall."},{"issue":"2","key":"9092_CR10","doi-asserted-by":"crossref","first-page":"391","DOI":"10.1002\/j.1538-7305.1985.tb00439.x","volume":"64","author":"B. H. Juang","year":"1985","unstructured":"Juang, B. H., & Rabiner, L. R. (1985). A probabilistic distance measure for hidden Markov models. AT&T Technical Journal, 64(2), 391\u2013408.","journal-title":"AT&T Technical Journal"},{"key":"9092_CR11","unstructured":"Ketabdar, H. (2008). Enhancing posterior based speech recognition systems. Ph.D. thesis, IDIAP, Research Institute, Switzerland."},{"key":"9092_CR12","volume-title":"Proc. international conference on acoustics, speech and signal processing","author":"H. Ketabdar","year":"2008","unstructured":"Ketabdar, H., & Boulard, H. (2008). Hierarchical integration of phonetic and lexical knowledge in phone posterior estimation. In Proc. international conference on acoustics, speech and signal processing, Las Vegas, USA."},{"key":"9092_CR13","unstructured":"Kirchoff, K. (1999). Robust speech recognition using articulatory features. Ph.D. thesis, University of Bielefield."},{"key":"9092_CR14","volume-title":"Proc. of the workshop on phonetics and phonology in ASR, parameters and features, and their implications","author":"K. Kirchoff","year":"2000","unstructured":"Kirchoff, K. (2000). Integrating articulatory features into acoustic models for speech recognition. In Proc. of the workshop on phonetics and phonology in ASR, parameters and features, and their implications, Saarbrucken, Germany."},{"key":"9092_CR15","volume-title":"Proc. international conferences on spoken language processing (ICSLP)","author":"J. Kohler","year":"1996","unstructured":"Kohler, J. (1996). Multi-lingual phoneme recognition exploiting acoustic-phonetic similarities of sounds. In Proc. international conferences on spoken language processing (ICSLP), Philadelphia, USA."},{"key":"9092_CR16","first-page":"417","volume-title":"Proc. IEEE int. conf. on acoustics, speech and signal processing","author":"J. Kohler","year":"1998","unstructured":"Kohler, J. (1998). Language adaptation of multilingual phone models for vocabulary independent multilingual speech recognition. In Proc. IEEE int. conf. on acoustics, speech and signal processing, Seattle, USA (pp.\u00a0417\u2013420)."},{"key":"9092_CR17","doi-asserted-by":"crossref","first-page":"21","DOI":"10.1016\/S0167-6393(00)00093-5","volume":"35","author":"J. Kohler","year":"2001","unstructured":"Kohler, J. (2001). Multilingual phone models for vocabulary independent speech recognition. Speech Communication, 35, 21\u201330.","journal-title":"Speech Communication"},{"key":"9092_CR18","volume-title":"Information theory and statistics","author":"S. Kullback","year":"1958","unstructured":"Kullback, S. (1958). Information theory and statistics. New York: Wiley."},{"key":"9092_CR19","volume-title":"Proc. Interspeech","author":"C. H. Lee","year":"2007","unstructured":"Lee, C. H. et al. (2007). An overview of automatic speech attribute transcription. In Proc. Interspeech, Antwerp, Belgium."},{"key":"9092_CR20","volume-title":"Proc. Interspeech","author":"J. Li","year":"2005","unstructured":"Li, J., & Lee, C. H. (2005). On designing and evaluating speech event detectors. In Proc. Interspeech, Lisbon, Portugal."},{"key":"9092_CR21","volume-title":"Proc. IEEE int. conf. on acoustics, speech and signal processing","author":"H. Lin","year":"2009","unstructured":"Lin, H., Deng, L., Yu, D., Gong, Y., Acero, A., & Lee, C. H. (2009). A study on multilingual acoustic modeling for large vocabulary ASR. In Proc. IEEE int. conf. on acoustics, speech and signal processing, Taipei, Taiwan."},{"key":"9092_CR22","volume-title":"Proc. Interspeech","author":"D. Lyu","year":"2008","unstructured":"Lyu, D., Siniscalchi, S. M., Kim, T. Y., & Li, C. H. (2008). Continuous speech recognition without target language data. In Proc. Interspeech, Brisbane, Australia."},{"key":"9092_CR23","unstructured":"Odell, J. J. (1995). The use of context in large vocabulary speech recognition. Ph.D. Thesis, Engineering Department, Cambridge University."},{"issue":"2","key":"9092_CR24","doi-asserted-by":"crossref","first-page":"257","DOI":"10.1109\/5.18626","volume":"77","author":"L. R. Rabiner","year":"1989","unstructured":"Rabiner, L. R. (1989). A tutorial on hidden Markov models and selected applications in speech recognition. Proceedings of the IEEE, 77(2), 257\u2013286.","journal-title":"Proceedings of the IEEE"},{"issue":"4","key":"9092_CR25","doi-asserted-by":"crossref","first-page":"533","DOI":"10.1038\/323533a0","volume":"323","author":"D. E. Rumelhart","year":"1986","unstructured":"Rumelhart, D. E., Hintont, G. E., & Williams, R. J. (1986). Learning representations by back-propagating errors. Nature, 323(4), 533\u2013536.","journal-title":"Nature"},{"key":"9092_CR26","volume-title":"Multilingual speech processing","year":"2006","unstructured":"Schultz, T., & Krirchoff, K. (Eds.) (2006). Multilingual speech processing. New York: Elsevier."},{"key":"9092_CR27","first-page":"85","volume-title":"Workshop on multilingual interoperability in speech technology","author":"T. Schultz","year":"1999","unstructured":"Schultz, T., & Waibel, A. (1999). Language adaptive LVCSR through poly-phone decision tree specialization. In Workshop on multilingual interoperability in speech technology, Leusden, The Netherlands (pp. 85\u201390)."},{"key":"9092_CR28","doi-asserted-by":"crossref","unstructured":"Schultz, T., & Waibel, A. (2001). Language independent and language adaptive acoustic modeling for speech recognition. Speech Communication, 31\u201351.","DOI":"10.1016\/S0167-6393(00)00094-7"},{"key":"9092_CR29","unstructured":"Schwarz, P. (2008). Phoneme recognition using long temporal block. Ph.D. thesis, Brno University of Technology, Czech Republic."},{"key":"9092_CR30","volume-title":"Proceedings of 7th international conference text, speech and dialogue","author":"P. Schwarz","year":"2004","unstructured":"Schwarz, P., Mat\u011bjka, P., & \u010cernock\u00fd, J. (2004). Towards lower error rates in phoneme recognition. In Proceedings of 7th international conference text, speech and dialogue, Brno, Czech Republic."},{"key":"9092_CR31","volume-title":"Proc. Eurospeech","author":"S. Stuker","year":"2003","unstructured":"Stuker, S., Metze, F., Schultz, T., & Waibel, A. (2003). Integrating multilingual articulatory features into speech recognition. In Proc. Eurospeech, Geneva."},{"key":"9092_CR32","volume-title":"Proc. international conference on acoustics, speech and signal processing","author":"S. Stuker","year":"2007","unstructured":"Stuker, S., Schultz, T., Meize, F., & Waibel, A. (2007). Multilingual articulatory features. In Proc. international conference on acoustics, speech and signal processing, Honolulu, USA."},{"key":"9092_CR33","volume-title":"Proc. Interspeech","author":"L. Toth","year":"2008","unstructured":"Toth, L., Frankel, J., Gosziolya, G., & King, S. (2008). Cross-lingual portability of MLP based tandem features\u2014A case study for English and Hungarian. In Proc. Interspeech, Brisbane, Australia."},{"issue":"2","key":"9092_CR34","doi-asserted-by":"crossref","first-page":"260","DOI":"10.1109\/TIT.1967.1054010","volume":"13","author":"A. J. Viterbi","year":"1967","unstructured":"Viterbi, A. J. (1967). Error bounds for convolutional codes and an asymptotically optimal decoding algorithm. IEEE Transactions on Information Theory, 13(2), 260\u2013269.","journal-title":"IEEE Transactions on Information Theory"},{"key":"9092_CR35","first-page":"1297","volume-title":"Proc. IEEE","author":"A. Waibel","year":"2000","unstructured":"Waibel, A., Geutner, P., Mayfield, L., Schultz, T., & Woszczyna, M. (2000). Multilinguality in speech and spoken language systems. In Proc. IEEE (Vol.\u00a088, pp. 1297\u20131313). Special issue on spoken language processing."},{"key":"9092_CR36","volume-title":"The HTK book","author":"S. Young","year":"2003","unstructured":"Young, S., Jansen, J., Odell, J., Ollason, D., & Woodland, P. (2003). The HTK book. Cambridge: Cambridge University Engineering Department."}],"container-title":["International Journal of Speech Technology"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-011-9092-6.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s10772-011-9092-6\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-011-9092-6","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2019,6,10]],"date-time":"2019-06-10T17:41:52Z","timestamp":1560188512000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s10772-011-9092-6"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,5,11]]},"references-count":36,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2011,9]]}},"alternative-id":["9092"],"URL":"https:\/\/doi.org\/10.1007\/s10772-011-9092-6","relation":{},"ISSN":["1381-2416","1572-8110"],"issn-type":[{"value":"1381-2416","type":"print"},{"value":"1572-8110","type":"electronic"}],"subject":[],"published":{"date-parts":[[2011,5,11]]}}}