{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:24:17Z","timestamp":1750220657552,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":35,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,11,30]],"date-time":"2020-11-30T00:00:00Z","timestamp":1606694400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,11,30]]},"DOI":"10.1145\/3428757.3429971","type":"proceedings-article","created":{"date-parts":[[2021,1,27]],"date-time":"2021-01-27T11:29:07Z","timestamp":1611746947000},"page":"55-61","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Support software for Automatic Speech Recognition systems targeted for non-native speech"],"prefix":"10.1145","author":[{"given":"Kacper","family":"Radzikowski","sequence":"first","affiliation":[{"name":"Waseda University, Graduate School of Information, Production and Systems Kitakyushu, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Osamu","family":"Yoshie","sequence":"additional","affiliation":[{"name":"Waseda University, Graduate School of Information, Production and Systems Kitakyushu, Japan"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Robert","family":"Nowak","sequence":"additional","affiliation":[{"name":"Warsaw University of Technology Warsaw, Poland"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,1,27]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2013.6638947"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASLP.2014.2339736"},{"key":"e_1_3_2_1_3_1","unstructured":"Dario Amodei Rishita Anubhai Eric Battenberg Carl Case Jared Casper Bryan Catanzaro Jingdong Chen Mike Chrzanowski Adam Coates Greg Diamos Erich Elsen Jesse Engel Linxi Fan Christopher Fougner Tony Han Awni Hannun Billy Jun Patrick LeGresley Libby Lin Sharan Narang Andrew Ng Sherjil Ozair Ryan Prenger Jonathan Raiman Sanjeev Satheesh David Seetapun Shubho Sengupta Yi Wang Zhiqian Wang Chong Wang Bo Xiao Dani Yogatama Jun Zhan and Zhenyao Zhu. 2015. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin. arXiv:arXiv:1512.02595  Dario Amodei Rishita Anubhai Eric Battenberg Carl Case Jared Casper Bryan Catanzaro Jingdong Chen Mike Chrzanowski Adam Coates Greg Diamos Erich Elsen Jesse Engel Linxi Fan Christopher Fougner Tony Han Awni Hannun Billy Jun Patrick LeGresley Libby Lin Sharan Narang Andrew Ng Sherjil Ozair Ryan Prenger Jonathan Raiman Sanjeev Satheesh David Seetapun Shubho Sengupta Yi Wang Zhiqian Wang Chong Wang Bo Xiao Dani Yogatama Jun Zhan and Zhenyao Zhu. 2015. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin. arXiv:arXiv:1512.02595"},{"key":"e_1_3_2_1_4_1","unstructured":"Dario Amodei Rishita Anubhai Eric Battenberg Carl Case Jared Casper Bryan Catanzaro Jingdong Chen Mike Chrzanowski Adam Coates Greg Diamos Erich Elsen Jesse Engel Linxi Fan Christopher Fougner Tony Han Awni Hannun Billy Jun Patrick LeGresley Libby Lin Sharan Narang Andrew Ng Sherjil Ozair Ryan Prenger Jonathan Raiman Sanjeev Satheesh David Seetapun Shubho Sengupta Yi Wang Zhiqian Wang Chong Wang Bo Xiao Dani Yogatama Jun Zhan and Zhenyao Zhu. 2015. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin.  Dario Amodei Rishita Anubhai Eric Battenberg Carl Case Jared Casper Bryan Catanzaro Jingdong Chen Mike Chrzanowski Adam Coates Greg Diamos Erich Elsen Jesse Engel Linxi Fan Christopher Fougner Tony Han Awni Hannun Billy Jun Patrick LeGresley Libby Lin Sharan Narang Andrew Ng Sherjil Ozair Ryan Prenger Jonathan Raiman Sanjeev Satheesh David Seetapun Shubho Sengupta Yi Wang Zhiqian Wang Chong Wang Bo Xiao Dani Yogatama Jun Zhan and Zhenyao Zhu. 2015. Deep Speech 2: End-to-End Speech Recognition in English and Mandarin."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2015.7178819"},{"key":"e_1_3_2_1_6_1","unstructured":"P.W.D. Charles. 2019. keras. https:\/\/github.com\/charlespwd\/project-title.  P.W.D. Charles. 2019. keras. https:\/\/github.com\/charlespwd\/project-title."},{"volume-title":"International Journal for Advance Research in Engineering and Technology 1 (7","year":"2013","author":"Dave N","key":"e_1_3_2_1_7_1"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASL.2010.2064307"},{"key":"e_1_3_2_1_9_1","unstructured":"Sadaoki Furui. 2005. 50 Years of Progress in Speech and Speaker Recognition Research.  Sadaoki Furui. 2005. 50 Years of Progress in Speech and Speaker Recognition Research."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSA.2005.858512"},{"key":"e_1_3_2_1_11_1","unstructured":"Leon A. Gatys Alexander S. Ecker and Matthias Bethge. 2015. A Neural Algorithm of Artistic Style. arXiv:arXiv:1508.06576  Leon A. Gatys Alexander S. Ecker and Matthias Bethge. 2015. A Neural Algorithm of Artistic Style. arXiv:arXiv:1508.06576"},{"key":"e_1_3_2_1_12_1","unstructured":"G.E.Hinton L.Deng D.Yu G.E.Dahl A.Mohamed and N.Jaitly etal [n.d.]. Deep neural networks for acoustic modeling in speech recognition: the shared views of four research groups. ([n. d.]).  G.E.Hinton L.Deng D.Yu G.E.Dahl A.Mohamed and N.Jaitly et al. [n.d.]. Deep neural networks for acoustic modeling in speech recognition: the shared views of four research groups. ([n. d.])."},{"key":"e_1_3_2_1_13_1","unstructured":"Eric Grinstein Ngoc Duong Alexey Ozerov and Patrick P\u00e9rez. 2017. Audio style transfer. (2017). https:\/\/doi.org\/10.1109\/ICASSP.2018.8461711arXiv:arXiv:1710.11385  Eric Grinstein Ngoc Duong Alexey Ozerov and Patrick P\u00e9rez. 2017. Audio style transfer. (2017). https:\/\/doi.org\/10.1109\/ICASSP.2018.8461711arXiv:arXiv:1710.11385"},{"volume-title":"A Systematic Analysis of Automatic Speech Recognition: An Overview. 4 (06","year":"2014","author":"Gulzar Taabish","key":"e_1_3_2_1_14_1"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"crossref","unstructured":"Ben Hixon Eric Schneider and Susan Epstein. 2011. Phonemic Similarity Metrics to Compare Pronunciation Methods. 825--828.  Ben Hixon Eric Schneider and Susan Epstein. 2011. Phonemic Similarity Metrics to Compare Pronunciation Methods. 825--828.","DOI":"10.21437\/Interspeech.2011-305"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2015.7178765"},{"key":"e_1_3_2_1_17_1","volume-title":"International Conference on Acoustics, Speech, and Signal Processing. 123--126","volume":"1","author":"Hon Lee","year":"1988"},{"volume-title":"2016 IEEE International Conference on Consumer Electronics (ICCE). 383--384","year":"2016","author":"Lee S.","key":"e_1_3_2_1_18_1"},{"volume-title":"IEEE International Conference on Acoustics, Speech and Signal Processing","year":"2000","author":"Livescu K.","key":"e_1_3_2_1_19_1"},{"volume-title":"Fifteenth Annual Conference of the International Speech Communication Association.","year":"2014","author":"Metallinou Angeliki","key":"e_1_3_2_1_20_1"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1121\/1.385120"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1186\/s13636-018-0146-4"},{"volume-title":"Proceedings of the conference of institute of electrical engineers of japan, electronics and information systems division","year":"2017","author":"Radzikowski Kacper Yoshie Osamu","key":"e_1_3_2_1_24_1"},{"volume-title":"Proceedings of the conference of institute of electrical engineers of japan, electronics and information systems division","year":"2017","author":"Radzikowski Kacper Yoshie Osamu","key":"e_1_3_2_1_25_1"},{"volume-title":"Proceedings of the Conference of Institute of Electrical Engineers of Japan, Electronics and Information Systems Division.","year":"2016","author":"Kacper Radzikowski","key":"e_1_3_2_1_26_1"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/3011141.3011169"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3011141.3011169"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/3011141.3011169"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2003.1198887"},{"volume-title":"Proc. Interspeech","year":"2014","author":"Geiger J. T.","key":"e_1_3_2_1_31_1"},{"volume-title":"Proc. ICASSP","year":"2007","author":"Tan T.","key":"e_1_3_2_1_32_1"},{"key":"e_1_3_2_1_33_1","unstructured":"Xu Tian Jun Zhang Zejun Ma Yi He Juan Wei Peihao Wu Wenchang Situ Shuai Li and Yang Zhang. 2017. Deep LSTM for Large Vocabulary Continuous Speech Recognition. In Arxiv. https:\/\/arxiv.org\/abs\/1703.07090  Xu Tian Jun Zhang Zejun Ma Yi He Juan Wei Peihao Wu Wenchang Situ Shuai Li and Yang Zhang. 2017. Deep LSTM for Large Vocabulary Continuous Speech Recognition. In Arxiv. https:\/\/arxiv.org\/abs\/1703.07090"},{"key":"e_1_3_2_1_35_1","unstructured":"A\u00e4ron van den Oord Sander Dieleman Heiga Zen Karen Simonyan Oriol Vinyals Alexander Graves Nal Kalchbrenner Andrew Senior and Koray Kavukcuoglu. 2016. WaveNet: A Generative Model for Raw Audio.. In Arxiv. https:\/\/arxiv.org\/abs\/1609.03499  A\u00e4ron van den Oord Sander Dieleman Heiga Zen Karen Simonyan Oriol Vinyals Alexander Graves Nal Kalchbrenner Andrew Senior and Koray Kavukcuoglu. 2016. WaveNet: A Generative Model for Raw Audio.. In Arxiv. https:\/\/arxiv.org\/abs\/1609.03499"},{"volume-title":"Smith","year":"2018","author":"Verma Prateek","key":"e_1_3_2_1_36_1"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2018.8461870"}],"event":{"name":"iiWAS '20: The 22nd International Conference on Information Integration and Web-based Applications & Services","sponsor":["Johannes Kepler University, Linz, Austria"],"location":"Chiang Mai Thailand","acronym":"iiWAS '20"},"container-title":["Proceedings of the 22nd International Conference on Information Integration and Web-based Applications &amp; Services"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3428757.3429971","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3428757.3429971","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:02:32Z","timestamp":1750197752000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3428757.3429971"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,11,30]]},"references-count":35,"alternative-id":["10.1145\/3428757.3429971","10.1145\/3428757"],"URL":"https:\/\/doi.org\/10.1145\/3428757.3429971","relation":{},"subject":[],"published":{"date-parts":[[2020,11,30]]},"assertion":[{"value":"2021-01-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}