{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:16:15Z","timestamp":1750220175923,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":34,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,10,16]],"date-time":"2021-10-16T00:00:00Z","timestamp":1634342400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100004608","name":"Natural Science Foundation of Jiangsu Province","doi-asserted-by":"publisher","award":["BK20170278, BK20190623"],"award-info":[{"award-number":["BK20170278, BK20190623"]}],"id":[{"id":"10.13039\/501100004608","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61901003"],"award-info":[{"award-number":["61901003"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,10,16]]},"DOI":"10.1145\/3503181.3503218","type":"proceedings-article","created":{"date-parts":[[2022,3,1]],"date-time":"2022-03-01T23:13:40Z","timestamp":1646176420000},"page":"160-165","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Voice spoofing detection with raw waveform based on Dual Path Res2net"],"prefix":"10.1145","author":[{"given":"Xin","family":"Fang","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Haijia","family":"Du","sequence":"additional","affiliation":[{"name":"China University of Mining and Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Tian","family":"Gao","sequence":"additional","affiliation":[{"name":"iFLYTEK Research, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Liang","family":"Zou","sequence":"additional","affiliation":[{"name":"China University of Mining and Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhenhua","family":"Ling","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,3]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"crossref","unstructured":"Zhuxin Chen Zhifeng Xie Weibin Zhang and Xiangmin Xu. 2017. ResNet and Model Fusion for Automatic Spoofing Detection.. In Interspeech. 102\u2013106.  Zhuxin Chen Zhifeng Xie Weibin Zhang and Xiangmin Xu. 2017. ResNet and Model Fusion for Automatic Spoofing Detection.. In Interspeech. 102\u2013106.","DOI":"10.21437\/Interspeech.2017-1085"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11265-016-1148-z"},{"key":"e_1_3_2_1_3_1","unstructured":"Rohan\u00a0Kumar Das Xiaohai Tian Tomi Kinnunen and Haizhou Li. 2020. The attacker\u2019s perspective on automatic speaker verification: An overview. arXiv preprint arXiv:2004.08849(2020).  Rohan\u00a0Kumar Das Xiaohai Tian Tomi Kinnunen and Haizhou Li. 2020. The attacker\u2019s perspective on automatic speaker verification: An overview. arXiv preprint arXiv:2004.08849(2020)."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.3390\/s20236784"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2019.8682327"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.3000641"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1984.1164317"},{"key":"e_1_3_2_1_8_1","unstructured":"Sarfaraz Jelil Abhishek Shrivastava Rohan\u00a0Kumar Das SR\u00a0Mahadeva Prasanna and Rohit Sinha. 2019. SpeechMarker: A Voice Based Multi-Level Attendance Application.. In Interspeech. 3665\u20133666.  Sarfaraz Jelil Abhishek Shrivastava Rohan\u00a0Kumar Das SR\u00a0Mahadeva Prasanna and Rohit Sinha. 2019. SpeechMarker: A Voice Based Multi-Level Attendance Application.. In Interspeech. 3665\u20133666."},{"key":"e_1_3_2_1_9_1","volume-title":"Rawnet: Advanced end-to-end deep neural network using raw waveforms for text-independent speaker verification. arXiv preprint arXiv:1904.08104(2019).","author":"Heo Hee-Soo","year":"2019","unstructured":"Jee-weon Jung, Hee-Soo Heo , Ju-ho Kim, Hye-jin Shim, and Ha-Jin Yu . 2019 . Rawnet: Advanced end-to-end deep neural network using raw waveforms for text-independent speaker verification. arXiv preprint arXiv:1904.08104(2019). Jee-weon Jung, Hee-Soo Heo, Ju-ho Kim, Hye-jin Shim, and Ha-Jin Yu. 2019. Rawnet: Advanced end-to-end deep neural network using raw waveforms for text-independent speaker verification. arXiv preprint arXiv:1904.08104(2019)."},{"key":"e_1_3_2_1_10_1","volume-title":"Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980(2014).","author":"Kingma P","year":"2014","unstructured":"Diederik\u00a0 P Kingma and Jimmy Ba . 2014 . Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980(2014). Diederik\u00a0P Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980(2014)."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"crossref","unstructured":"Tomi Kinnunen Jaime Lorenzo-Trueba Junichi Yamagishi Tomoki Toda Daisuke Saito Fernando Villavicencio and Zhenhua Ling. 2018. A spoofing benchmark for the 2018 voice conversion challenge: Leveraging from spoofing countermeasures for speech artifact assessment. arXiv preprint arXiv:1804.08438(2018).  Tomi Kinnunen Jaime Lorenzo-Trueba Junichi Yamagishi Tomoki Toda Daisuke Saito Fernando Villavicencio and Zhenhua Ling. 2018. A spoofing benchmark for the 2018 voice conversion challenge: Leveraging from spoofing countermeasures for speech artifact assessment. arXiv preprint arXiv:1804.08438(2018).","DOI":"10.21437\/Odyssey.2018-27"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.21437\/Interspeech.2014-539"},{"key":"e_1_3_2_1_13_1","volume-title":"Interspeech","author":"Lapidot Itshak","year":"2019","unstructured":"Itshak Lapidot and Jean-Fran\u00e7ois Bonastre . 2019. Effects of waveform pmf on anti-spoofing detection . In Interspeech 2019 . ISCA , 2853\u20132857. Itshak Lapidot and Jean-Fran\u00e7ois Bonastre. 2019. Effects of waveform pmf on anti-spoofing detection. In Interspeech 2019. ISCA, 2853\u20132857."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"crossref","unstructured":"Galina Lavrentyeva Sergey Novoselov Andzhukaev Tseren Marina Volkova Artem Gorlanov and Alexandr Kozlov. 2019. STC antispoofing systems for the ASVspoof2019 challenge. arXiv preprint arXiv:1904.05576(2019).  Galina Lavrentyeva Sergey Novoselov Andzhukaev Tseren Marina Volkova Artem Gorlanov and Alexandr Kozlov. 2019. STC antispoofing systems for the ASVspoof2019 challenge. arXiv preprint arXiv:1904.05576(2019).","DOI":"10.21437\/Interspeech.2019-1768"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP39728.2021.9413828"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"crossref","unstructured":"Jaime Lorenzo-Trueba Fuming Fang Xin Wang Isao Echizen Junichi Yamagishi and Tomi Kinnunen. 2018. Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama\u2019s voice using GAN WaveNet and low-quality found data. arXiv preprint arXiv:1803.00860(2018).  Jaime Lorenzo-Trueba Fuming Fang Xin Wang Isao Echizen Junichi Yamagishi and Tomi Kinnunen. 2018. Can we steal your vocal identity from the Internet?: Initial investigation of cloning Obama\u2019s voice using GAN WaveNet and low-quality found data. arXiv preprint arXiv:1803.00860(2018).","DOI":"10.21437\/Odyssey.2018-34"},{"key":"e_1_3_2_1_17_1","volume-title":"Odyssey 2018 - The Speaker and Language Recognition Workshop, Vol.\u00a02018","author":"Lorenzo-Trueba Jaime","year":"2016","unstructured":"Jaime Lorenzo-Trueba , Junichi Yamagishi , Tomoki Toda , Daisuke Saito , Fernando Villavicencio , Tomi Kinnunen , and Zhenhua Ling . 2016 . The voice conversion challenge 2018: Promoting development of parallel and nonparallel methods . In Odyssey 2018 - The Speaker and Language Recognition Workshop, Vol.\u00a02018 . Jaime Lorenzo-Trueba, Junichi Yamagishi, Tomoki Toda, Daisuke Saito, Fernando Villavicencio, Tomi Kinnunen, and Zhenhua Ling. 2016. The voice conversion challenge 2018: Promoting development of parallel and nonparallel methods. In Odyssey 2018 - The Speaker and Language Recognition Workshop, Vol.\u00a02018."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2006.1660175"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1587\/transinf.2015EDP7457"},{"key":"e_1_3_2_1_20_1","volume-title":"Wavenet: A generative model for raw audio. arXiv preprint arXiv:1609.03499(2016).","author":"van\u00a0den Oord Aaron","year":"2016","unstructured":"Aaron van\u00a0den Oord , Sander Dieleman , Heiga Zen , Karen Simonyan , Oriol Vinyals , Alex Graves , Nal Kalchbrenner , Andrew Senior , and Koray Kavukcuoglu . 2016 . Wavenet: A generative model for raw audio. arXiv preprint arXiv:1609.03499(2016). Aaron van\u00a0den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu. 2016. Wavenet: A generative model for raw audio. arXiv preprint arXiv:1609.03499(2016)."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"crossref","unstructured":"Tanvina\u00a0B Patel and Hemant\u00a0A Patil. 2015. Combining evidences from mel cepstral cochlear filter cepstral and instantaneous frequency features for detection of natural vs. spoofed speech. In Sixteenth annual conference of the international speech communication association.  Tanvina\u00a0B Patel and Hemant\u00a0A Patil. 2015. Combining evidences from mel cepstral cochlear filter cepstral and instantaneous frequency features for detection of natural vs. spoofed speech. In Sixteenth annual conference of the international speech communication association.","DOI":"10.21437\/Interspeech.2015-467"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00643"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/SLT.2018.8639636"},{"volume-title":"Odyssey 2016 - The Speaker and Language Recognition Workshop.","author":"Todisco M.","key":"e_1_3_2_1_24_1","unstructured":"M. Todisco , H. Delgado , and N. Evans . 2016. A New Feature for Automatic Speaker Verification Anti-Spoofing: Constant Q Cepstral Coefficients . In Odyssey 2016 - The Speaker and Language Recognition Workshop. M. Todisco, H. Delgado, and N. Evans. 2016. A New Feature for Automatic Speaker Verification Anti-Spoofing: Constant Q Cepstral Coefficients. In Odyssey 2016 - The Speaker and Language Recognition Workshop."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"crossref","unstructured":"Massimiliano Todisco Xin Wang Ville Vestman Md Sahidullah H\u00e9ctor Delgado Andreas Nautsch Junichi Yamagishi Nicholas Evans Tomi Kinnunen and Kong\u00a0Aik Lee. 2019. ASVspoof 2019: Future horizons in spoofed and fake audio detection. arXiv preprint arXiv:1904.05441(2019).  Massimiliano Todisco Xin Wang Ville Vestman Md Sahidullah H\u00e9ctor Delgado Andreas Nautsch Junichi Yamagishi Nicholas Evans Tomi Kinnunen and Kong\u00a0Aik Lee. 2019. ASVspoof 2019: Future horizons in spoofed and fake audio detection. arXiv preprint arXiv:1904.05441(2019).","DOI":"10.21437\/Interspeech.2019-2249"},{"key":"e_1_3_2_1_26_1","unstructured":"Francis Tom Mohit Jain and Prasenjit Dey. 2018. End-To-End Audio Replay Attack Detection Using Deep Convolutional Networks with Attention.. In Interspeech. 681\u2013685.  Francis Tom Mohit Jain and Prasenjit Dey. 2018. End-To-End Audio Replay Attack Detection Using Deep Convolutional Networks with Attention.. In Interspeech. 681\u2013685."},{"key":"e_1_3_2_1_27_1","unstructured":"Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan\u00a0N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998\u20136008.  Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan\u00a0N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998\u20136008."},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/LSP.2018.2822810"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2019.8682298"},{"key":"e_1_3_2_1_30_1","unstructured":"Zhenzong Wu Rohan\u00a0Kumar Das Jichen Yang and Haizhou Li. 2020. Light convolutional neural network with feature genuinization for detection of synthetic speech attacks. arXiv preprint arXiv:2009.09637(2020).  Zhenzong Wu Rohan\u00a0Kumar Das Jichen Yang and Haizhou Li. 2020. Light convolutional neural network with feature genuinization for detection of synthetic speech attacks. arXiv preprint arXiv:2009.09637(2020)."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","unstructured":"Zhizheng Wu Nicholas Evans Tomi Kinnunen Junichi Yamagishi Federico Alegre and Haizhou Li. 2015. Spoofing and countermeasures for speaker verification: A survey. Speech communication 66(2015) 130\u2013153.  Zhizheng Wu Nicholas Evans Tomi Kinnunen Junichi Yamagishi Federico Alegre and Haizhou Li. 2015. Spoofing and countermeasures for speaker verification: A survey. Speech communication 66(2015) 130\u2013153.","DOI":"10.1016\/j.specom.2014.10.005"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-015-3080-9"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/JSTSP.2017.2671435"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"crossref","unstructured":"R. Xiao. 2021. Adaptive Margin Circle Loss for Speaker Verification. In Interspeech.  R. Xiao. 2021. Adaptive Margin Circle Loss for Speaker Verification. In Interspeech.","DOI":"10.21437\/Interspeech.2021-1043"}],"event":{"name":"ICCSE '21: 5th International Conference on Crowd Science and Engineering","acronym":"ICCSE '21","location":"Jinan China"},"container-title":["5th International Conference on Crowd Science and Engineering"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503181.3503218","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503181.3503218","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:49Z","timestamp":1750186849000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503181.3503218"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,10,16]]},"references-count":34,"alternative-id":["10.1145\/3503181.3503218","10.1145\/3503181"],"URL":"https:\/\/doi.org\/10.1145\/3503181.3503218","relation":{},"subject":[],"published":{"date-parts":[[2021,10,16]]},"assertion":[{"value":"2022-03-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}