{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,21]],"date-time":"2026-05-21T10:25:20Z","timestamp":1779359120324,"version":"3.51.4"},"reference-count":22,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2017,8,7]],"date-time":"2017-08-07T00:00:00Z","timestamp":1502064000000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/www.springer.com\/tdm"}],"funder":[{"DOI":"10.13039\/501100004110","name":"Shanghai Normal University","doi-asserted-by":"publisher","award":["DCL201702"],"award-info":[{"award-number":["DCL201702"]}],"id":[{"id":"10.13039\/501100004110","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003399","name":"Science and Technology Commission of Shanghai Municipality","doi-asserted-by":"publisher","award":["14YF1409300"],"award-info":[{"award-number":["14YF1409300"]}],"id":[{"id":"10.13039\/501100003399","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Speech Technol"],"published-print":{"date-parts":[[2017,12]]},"DOI":"10.1007\/s10772-017-9447-8","type":"journal-article","created":{"date-parts":[[2017,8,7]],"date-time":"2017-08-07T07:38:42Z","timestamp":1502091522000},"page":"753-759","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":7,"title":["Articulatory movement features for short-duration text-dependent speaker verification"],"prefix":"10.1007","volume":"20","author":[{"given":"Yan","family":"Zhang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yanhua","family":"Long","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiangrong","family":"Shen","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haoran","family":"Wei","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Min","family":"Yang","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hong","family":"Ye","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hongwei","family":"Mao","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2017,8,7]]},"reference":[{"key":"9447_CR1","doi-asserted-by":"crossref","unstructured":"Akgul, Y. S., Kambhamettu, C., & Stone, M. (1998). Extraction and tracking of the tongue surface from ultrasound image sequences. In Proceedings of the\u00a0IEEE Computer Society Conference on Computer Vision and Pattern Recognition (pp.\u00a0298\u2013303).","DOI":"10.1109\/CVPR.1998.698623"},{"key":"9447_CR2","doi-asserted-by":"crossref","unstructured":"Alam, M. J., Kenny, P., & Stafylakis, T. (2015). Combining amplitude and phase-based features for speaker verification with short duration utterances. In Proceedings of the\u00a0Interspeech (pp.\u00a0249\u2013253).","DOI":"10.21437\/Interspeech.2015-94"},{"issue":"1","key":"9447_CR3","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1016\/0730-725X(87)90477-2","volume":"5","author":"T Baer","year":"1987","unstructured":"Baer, T., Gore, J. C., Boyce, S., et al. (1987). Application of MRI to the analysis of speech production. Magnetic Resonance Imaging, 5(1), 1\u20137.","journal-title":"Magnetic Resonance Imaging"},{"issue":"9","key":"9447_CR4","doi-asserted-by":"crossref","first-page":"1200","DOI":"10.1016\/j.specom.2005.08.008","volume":"48","author":"M BenZeghiba","year":"2006","unstructured":"BenZeghiba, M., & Bourland, H. (2006). User-customized password speaker verification using multiple reference and background models. Speech Communication, 48(9), 1200\u20131213.","journal-title":"Speech Communication"},{"key":"9447_CR5","doi-asserted-by":"crossref","unstructured":"Cai, M. Q., Ling, Z. H., & Dai, L. R. (2012). Target-filtering model based articulatory movement prediction for articulatory control of HMM-based speech synthesis. In Proceedings of the\u00a0IEEE International Conference on Signal Processing (pp.\u00a0605\u2013608).","DOI":"10.1109\/ICoSP.2012.6491560"},{"key":"9447_CR6","doi-asserted-by":"crossref","unstructured":"Fu, T., Qian, Y., Liu, Y., & Yu, K. (2014). Tandem deep features for text-dependent speaker verification. In Proceedings of the\u00a0Interspeech (pp.\u00a01327\u20131331).","DOI":"10.21437\/Interspeech.2014-329"},{"issue":"2","key":"9447_CR7","doi-asserted-by":"crossref","first-page":"254","DOI":"10.1109\/TASSP.1981.1163530","volume":"29","author":"S Furui","year":"1981","unstructured":"Furui, S. (1981). Cepstral analysis technique for automatic speaker verification. IEEE Transactions on Acoustics, Speech and Signal Processing, 29(2), 254\u2013272.\u00a0","journal-title":"IEEE Transactions on Acoustics, Speech and Signal Processing"},{"key":"9447_CR8","doi-asserted-by":"crossref","unstructured":"Ganapathy, S., Pelecanos, J., & Omar, M. K. (2011). Feature normalization for speaker verification in room reverberation. In Proceedings of the ICASSP (pp.\u00a04836\u20134839).","DOI":"10.1109\/ICASSP.2011.5947438"},{"key":"9447_CR9","doi-asserted-by":"crossref","unstructured":"Guo, J., Yeung, G., Muralidharan, D., et al. (2016). Speaker verification using short utterances with DNN-based estimation of subglottal acoustic features. In Proceedings of the Interspeech (pp.\u00a02219\u20132222).","DOI":"10.21437\/Interspeech.2016-282"},{"key":"9447_CR10","unstructured":"H\u00e9bert, M. (2008). Text-dependent speaker recognition. In Springer handbook of speech processing. Berlin: Springer."},{"issue":"1","key":"9447_CR11","doi-asserted-by":"crossref","first-page":"12","DOI":"10.1016\/j.specom.2009.08.009","volume":"52","author":"T Kinnunen","year":"2010","unstructured":"Kinnunen, T., & Li, H. (2010). An overview of text-independent speaker recognition: From features to supervectors. Speech Communication, 52(1), 12\u201340.","journal-title":"Speech Communication"},{"issue":"2","key":"9447_CR12","doi-asserted-by":"crossref","first-page":"119","DOI":"10.1016\/0167-6393(86)90003-8","volume":"5","author":"S Kiritani","year":"2010","unstructured":"Kiritani, S. (1986). X-ray microbeam method for measurement of articulatory dynamics-techniques and results. Speech Communication, 5(2), 119\u2013140.","journal-title":"Speech Communication"},{"issue":"6","key":"9447_CR13","doi-asserted-by":"crossref","first-page":"1171","DOI":"10.1109\/TASL.2009.2014796","volume":"17","author":"ZH Ling","year":"2009","unstructured":"Ling, Z. H., Richmond, K., Yamagishi, J., & Wang, R. H. (2009). Integrating articulatory features into HMM-based parametric speech synthesis. IEEE Transactions on Audio Speech and Language Processing, 17(6), 1171\u20131185.","journal-title":"IEEE Transactions on Audio Speech and Language Processing"},{"key":"9447_CR14","doi-asserted-by":"crossref","unstructured":"Long, Y., Yan, Z. J., Soong, F. K., Dai, L., & Guo, W. (2011). Speaker characterization using spectral subband energy ratio based on harmonic plus noise model. In Proceedings of the ICASSP (pp.\u00a04520\u20134523).","DOI":"10.1109\/ICASSP.2011.5947359"},{"key":"9447_CR15","unstructured":"Muda, L., Begam, M., & Elamvazuthi, I. (2010). Voice recognition algorithms using mel frequency cepstral coefficient (MFCC) and dynamic time warping (DTW) techniques.\u00a0Journal of Computing, 2(3), 138\u2013143."},{"key":"9447_CR16","doi-asserted-by":"crossref","unstructured":"Naik, J., Netsch, L., & Doddington, G. (1989). Speaker verification over long distance telephone lines. In Proceedings of the ICASSP (pp.\u00a0524\u2013527).","DOI":"10.1109\/ICASSP.1989.266479"},{"key":"9447_CR17","doi-asserted-by":"crossref","unstructured":"Qian, Y., Tao, J., Suendermann-Oeft, D., Evanini, K., & Ivanov, A. V. (2016). Noise and metadata sensitive bottleneck features for improving speaker recognition with non-native speech input. In Proceedings of the Interspeech (pp.\u00a03648\u20133652).","DOI":"10.21437\/Interspeech.2016-548"},{"issue":"1","key":"9447_CR18","doi-asserted-by":"crossref","first-page":"26","DOI":"10.1016\/0093-934X(87)90058-7","volume":"31","author":"PW Sch\u00f6nle","year":"1987","unstructured":"Sch\u00f6nle, P. W., Gr\u00e4be, K., Wenig, P., H\u00f6hne, J., Schrader, J., & Conrad, B. (1987). Electromagnetic articulography: Use of alternating magnetic fields for tracking movements of multiple points inside and outside the vocal tract. Brain and Language, 31(1), 26\u201335.","journal-title":"Brain and Language"},{"key":"9447_CR19","unstructured":"Summerfield, Q. (1987). Some preliminaries to a comprehensive account of audio\u2013visual speech perception. In B. Dodd & R. Campbell (Eds.), Hearing by eye: The psychology of lip-reading (pp. 3\u201351). Hove, UK: Lawrence Earlbaum Associates."},{"issue":"11","key":"9447_CR20","doi-asserted-by":"crossref","first-page":"2484","DOI":"10.1093\/ietisy\/e88-d.11.2484","volume":"88","author":"M Tachibana","year":"2005","unstructured":"Tachibana, M., Yamagishi, J., Masuko, T., & Kobayashi, T. (2005). Speech synthesis with various emotional expressions and speaking styles by style interpolation and morphing. IEICE Transactions on Information and Systems, 88(11), 2484\u20132491.","journal-title":"IEICE Transactions on Information and Systems"},{"key":"9447_CR21","unstructured":"Toda, T., \u00a0Black, A. W., &\u00a0Tokuda, K.\u00a0(2004). Mapping from articulatory movements to vocal tract spectrum with gaussian mixture model for articulatory speech synthesis. In 5th\u00a0ISCA Speech Synthesis Workshop (pp. 31\u201336)."},{"key":"9447_CR22","volume-title":"The HTK book","author":"S Young","year":"2002","unstructured":"Young, S., Evermann, G., & Gales, M. J. F. (2002). The HTK book. Cambridge: Cambridge University Engineering Department."}],"container-title":["International Journal of Speech Technology"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s10772-017-9447-8\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-017-9447-8.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s10772-017-9447-8.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,7,31]],"date-time":"2022-07-31T21:02:08Z","timestamp":1659301328000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s10772-017-9447-8"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,8,7]]},"references-count":22,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2017,12]]}},"alternative-id":["9447"],"URL":"https:\/\/doi.org\/10.1007\/s10772-017-9447-8","relation":{},"ISSN":["1381-2416","1572-8110"],"issn-type":[{"value":"1381-2416","type":"print"},{"value":"1572-8110","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,8,7]]}}}