{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,4]],"date-time":"2026-07-04T16:54:39Z","timestamp":1783184079520,"version":"3.54.6"},"publisher-location":"New York, NY, USA","reference-count":69,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,11,7]],"date-time":"2022-11-07T00:00:00Z","timestamp":1667779200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/100000001","name":"NSF (National Science Foundation)","doi-asserted-by":"publisher","award":["CNS-1916926,IIS-1905558,ECCS-2020289,CNS-2038995,CNS-2154930"],"award-info":[{"award-number":["CNS-1916926,IIS-1905558,ECCS-2020289,CNS-2038995,CNS-2154930"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000183","name":"Army Research Office","doi-asserted-by":"publisher","award":["W911NF1810305,W911NF1910241,W911NF1810208"],"award-info":[{"award-number":["W911NF1810305,W911NF1910241,W911NF1810208"]}],"id":[{"id":"10.13039\/100000183","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,11,7]]},"DOI":"10.1145\/3548606.3560671","type":"proceedings-article","created":{"date-parts":[[2022,11,7]],"date-time":"2022-11-07T11:41:28Z","timestamp":1667821288000},"page":"2009-2023","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":15,"title":["When Evil Calls"],"prefix":"10.1145","author":[{"given":"Han","family":"Liu","sequence":"first","affiliation":[{"name":"Washington University in St. Louis, St. Louis, MO, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhiyuan","family":"Yu","sequence":"additional","affiliation":[{"name":"Washington University in St. Louis, St. Louis, MO, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mingming","family":"Zha","sequence":"additional","affiliation":[{"name":"Indiana University Bloomington, Bloomington, IN, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"XiaoFeng","family":"Wang","sequence":"additional","affiliation":[{"name":"Indiana University Bloomington, Bloomington, IN, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"William","family":"Yeoh","sequence":"additional","affiliation":[{"name":"Washington University in St. Louis, St. Louis, MO, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yevgeniy","family":"Vorobeychik","sequence":"additional","affiliation":[{"name":"Washington University in St. Louis, St. Louis, MO, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ning","family":"Zhang","sequence":"additional","affiliation":[{"name":"Washington University in St. Louis, St. Louis, MO, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,11,7]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"84\n    current video conferencing statistics for the 2021 market. https:\/\/www. trustradius.com\/vendor-blog\/web-conferencing-statistics-trends.  84 current video conferencing statistics for the 2021 market. https:\/\/www. trustradius.com\/vendor-blog\/web-conferencing-statistics-trends."},{"key":"e_1_3_2_1_2_1","unstructured":"90-day security plan progress report: April 22. https:\/\/blog.zoom.us\/90-daysecurity-plan-progress-report-april-22\/.  90-day security plan progress report: April 22. https:\/\/blog.zoom.us\/90-daysecurity-plan-progress-report-april-22\/."},{"key":"e_1_3_2_1_3_1","unstructured":"Alexa skills. https:\/\/developer.amazon.com\/en-US\/alexa\/alexa-skills-kit. Accessed: 2022-07--25.  Alexa skills. https:\/\/developer.amazon.com\/en-US\/alexa\/alexa-skills-kit. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_4_1","unstructured":"Amazon echo. https:\/\/www.amazon.com\/echo-3rd-generation\/s?k=echo3rd generation. Accessed: 2022-07--25.  Amazon echo. https:\/\/www.amazon.com\/echo-3rd-generation\/s?k=echo3rd generation. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_5_1","unstructured":"Amazon transcribe. https:\/\/aws.amazon.com\/transcribe\/. Accessed: 2022-07--25.  Amazon transcribe. https:\/\/aws.amazon.com\/transcribe\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_6_1","volume-title":"Apr 28","author":"Every","year":"2022","unstructured":"Every company going remote permanently : Apr 28 , 2022 update. https:\/\/ buildremote.co\/companies\/companies-going-remote-permanently\/. Accessed : 2022-03--30. Every company going remote permanently: Apr 28, 2022 update. https:\/\/ buildremote.co\/companies\/companies-going-remote-permanently\/. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_7_1","unstructured":"Google assistant. https:\/\/assistant.google.com\/. Accessed: 2022-03--30.  Google assistant. https:\/\/assistant.google.com\/. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_8_1","unstructured":"Google cloud speech-to-text. https:\/\/cloud.google.com\/speech-to-text\/. Accessed: 2022-07--25.  Google cloud speech-to-text. https:\/\/cloud.google.com\/speech-to-text\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_9_1","unstructured":"Google cloud speech-to-text price. https:\/\/cloud.google.com\/speech-to-text\/ pricing. Accessed: 2022-07--25.  Google cloud speech-to-text price. https:\/\/cloud.google.com\/speech-to-text\/ pricing. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_10_1","volume-title":"meet. https:\/\/meet.google.com\/. Accessed: 2022-07--25","author":"Google","unstructured":"Google meet. https:\/\/meet.google.com\/. Accessed: 2022-07--25 . Google meet. https:\/\/meet.google.com\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_11_1","unstructured":"How google meet supports two million new users each day. https: \/\/cloud.google.com\/blog\/products\/g-suite\/how-google-meet-supports-twomillion-new-users-each-day.  How google meet supports two million new users each day. https: \/\/cloud.google.com\/blog\/products\/g-suite\/how-google-meet-supports-twomillion-new-users-each-day."},{"key":"e_1_3_2_1_12_1","unstructured":"How many decibels does a human speak normally. https:\/\/decibelpro.app\/blog\/ how-many-decibels-does-a-human-speak-normally\/. Accessed: 2022-07--25.  How many decibels does a human speak normally. https:\/\/decibelpro.app\/blog\/ how-many-decibels-does-a-human-speak-normally\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_13_1","unstructured":"Ibm speech to text. https:\/\/www.ibm.com\/watson\/services\/speech-to-text\/. Accessed: 2022-07--25.  Ibm speech to text. https:\/\/www.ibm.com\/watson\/services\/speech-to-text\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_14_1","unstructured":"Microsoft azure speech to text. https:\/\/azure.microsoft.com\/en-us\/services\/ cognitive-services\/speech-to-text\/. Accessed: 2022-07--25.  Microsoft azure speech to text. https:\/\/azure.microsoft.com\/en-us\/services\/ cognitive-services\/speech-to-text\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_15_1","unstructured":"Microsoft cortana. https:\/\/www.microsoft.com\/en-us\/cortana\/. Accessed: 2022- 07--25.  Microsoft cortana. https:\/\/www.microsoft.com\/en-us\/cortana\/. Accessed: 2022- 07--25."},{"key":"e_1_3_2_1_16_1","unstructured":"Microsoft teams. https:\/\/www.microsoft.com\/en-us\/microsoft-teams. Accessed: 2022-07--25.  Microsoft teams. https:\/\/www.microsoft.com\/en-us\/microsoft-teams. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_17_1","volume-title":"https:\/\/www.businessofapps. com\/data\/microsoft-teams-statistics\/#: :text=Microsoft%20Teams%20saw% 20a%20huge,Zoom%20from%20February%20to%20June","author":"Microsoft","year":"2022","unstructured":"Microsoft teams revenue and usage statistics ( 2022 ). https:\/\/www.businessofapps. com\/data\/microsoft-teams-statistics\/#: :text=Microsoft%20Teams%20saw% 20a%20huge,Zoom%20from%20February%20to%20June . Microsoft teams revenue and usage statistics (2022). https:\/\/www.businessofapps. com\/data\/microsoft-teams-statistics\/#: :text=Microsoft%20Teams%20saw% 20a%20huge,Zoom%20from%20February%20to%20June."},{"key":"e_1_3_2_1_18_1","unstructured":"Network link conditioner. https:\/\/nshipster.com\/network-link-conditioner. Accessed: 2022-07--25.  Network link conditioner. https:\/\/nshipster.com\/network-link-conditioner. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_19_1","unstructured":"Number of voice assistant users in the united states from 2017 to 2022. https: \/\/www.statista.com\/statistics\/1029573\/us-voice-assistant-users\/. Accessed: 2022- 03--30.  Number of voice assistant users in the united states from 2017 to 2022. https: \/\/www.statista.com\/statistics\/1029573\/us-voice-assistant-users\/. Accessed: 2022- 03--30."},{"key":"e_1_3_2_1_20_1","unstructured":"Scientists want virtual meetings to stay after the covid pandemic. https:\/\/www. nature.com\/articles\/d41586-021-00513--1. Accessed: 2022-03--30.  Scientists want virtual meetings to stay after the covid pandemic. https:\/\/www. nature.com\/articles\/d41586-021-00513--1. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_21_1","unstructured":"Skype. https:\/\/www.skype.com\/en. Accessed: 2022-07--25.  Skype. https:\/\/www.skype.com\/en. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_22_1","unstructured":"tc: Traffic control in the linux kernel. https:\/\/linux.die.net\/man\/8\/tc. Accessed: 2022-03--30.  tc: Traffic control in the linux kernel. https:\/\/linux.die.net\/man\/8\/tc. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_23_1","unstructured":"Trickle: A lightweight userspace bandwidth shaper. https:\/\/linux.die.net\/man\/1\/ trickle. Accessed: 2022-03--30.  Trickle: A lightweight userspace bandwidth shaper. https:\/\/linux.die.net\/man\/1\/ trickle. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_24_1","unstructured":"Voip adoption statistics for 2019 & beyond. https:\/\/wisdomplexus.com\/blogs\/voipadoption-statistics-2019-beyond\/. Accessed: 2022-03--30.  Voip adoption statistics for 2019 & beyond. https:\/\/wisdomplexus.com\/blogs\/voipadoption-statistics-2019-beyond\/. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_25_1","unstructured":"Webex. https:\/\/www.webex.com\/. Accessed: 2022-07--25.  Webex. https:\/\/www.webex.com\/. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_26_1","unstructured":"Zoom. https:\/\/zoom.us. Accessed: 2022-07--25.  Zoom. https:\/\/zoom.us. Accessed: 2022-07--25."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2019.23362"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/SP40001.2021.00009"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/SP40001.2021.00014"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCNC.2013.6504214"},{"key":"e_1_3_2_1_31_1","first-page":"4312","volume-title":"Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21 (8 2021), Z.-H. Zhou, Ed., International Joint Conferences on Artificial Intelligence Organization","author":"Bai T.","unstructured":"Bai , T. , Luo , J. , Zhao , J. , Wen , B. , and Wang , Q . Recent advances in adversarial training for adversarial robustness . In Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21 (8 2021), Z.-H. Zhou, Ed., International Joint Conferences on Artificial Intelligence Organization , pp. 4312 -- 4321 . Survey Track. Bai, T., Luo, J., Zhao, J., Wen, B., and Wang, Q. Recent advances in adversarial training for adversarial robustness. In Proceedings of the Thirtieth International Joint Conference on Artificial Intelligence, IJCAI-21 (8 2021), Z.-H. Zhou, Ed., International Joint Conferences on Artificial Intelligence Organization, pp. 4312-- 4321. Survey Track."},{"key":"e_1_3_2_1_32_1","first-page":"513","volume-title":"25th USENIX security symposium (USENIX security 16)","author":"Carlini N.","year":"2016","unstructured":"Carlini , N. , Mishra , P. , Vaidya , T. , Zhang , Y. , Sherr , M. , Shields , C. , Wagner , D. , and Zhou , W . Hidden voice commands . In 25th USENIX security symposium (USENIX security 16) ( 2016 ), pp. 513 -- 530 . Carlini, N., Mishra, P., Vaidya, T., Zhang, Y., Sherr, M., Shields, C., Wagner, D., and Zhou, W. Hidden voice commands. In 25th USENIX security symposium (USENIX security 16) (2016), pp. 513--530."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/SPW.2018.00009"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"crossref","first-page":"1277","DOI":"10.1109\/SP40000.2020.00045","volume-title":"Hopskipjumpattack: A queryefficient decision-based attack. In 2020 ieee symposium on security and privacy (sp)","author":"Chen J.","year":"2020","unstructured":"Chen , J. , Jordan , M. I. , and Wainwright , M. J . Hopskipjumpattack: A queryefficient decision-based attack. In 2020 ieee symposium on security and privacy (sp) ( 2020 ), IEEE , pp. 1277 -- 1294 . Chen, J., Jordan, M. I., and Wainwright, M. J. Hopskipjumpattack: A queryefficient decision-based attack. In 2020 ieee symposium on security and privacy (sp) (2020), IEEE, pp. 1277--1294."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2020.23055"},{"key":"e_1_3_2_1_36_1","first-page":"2667","volume-title":"USENIX Security Symposium","author":"Chen Y.","year":"2020","unstructured":"Chen , Y. , Yuan , X. , Zhang , J. , Zhao , Y. , Zhang , S. , Chen , K. , and Wang , X . Devil's whisper: A general approach for physical adversarial attacks against commercial black-box speech recognition devices . In USENIX Security Symposium ( 2020 ), pp. 2667 -- 2684 . Chen, Y., Yuan, X., Zhang, J., Zhao, Y., Zhang, S., Chen, K., and Wang, X. Devil's whisper: A general approach for physical adversarial attacks against commercial black-box speech recognition devices. In USENIX Security Symposium (2020), pp. 2667--2684."},{"key":"e_1_3_2_1_37_1","volume-title":"8th International Conference on Learning Representations, ICLR 2020","author":"Cheng M.","year":"2020","unstructured":"Cheng , M. , Singh , S. , Chen , P. H. , Chen , P. , Liu , S. , and Hsieh , C . Sign-opt: A query-efficient hard-label adversarial attack . In 8th International Conference on Learning Representations, ICLR 2020 , Addis Ababa, Ethiopia, April 26--30 , 2020 (2020), OpenReview.net. Cheng, M., Singh, S., Chen, P. H., Chen, P., Liu, S., and Hsieh, C. Sign-opt: A query-efficient hard-label adversarial attack. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26--30, 2020 (2020), OpenReview.net."},{"key":"e_1_3_2_1_38_1","first-page":"15210","volume-title":"Advances in Neural Information Processing Systems 32: Annual Conference on Neural Information Processing Systems 2019","author":"Cutkosky A.","year":"2019","unstructured":"Cutkosky , A. , and Orabona , F . Momentum-based variance reduction in nonconvex SGD . In Advances in Neural Information Processing Systems 32: Annual Conference on Neural Information Processing Systems 2019 , NeurIPS 2019 , December 8--14, 2019, Vancouver, BC, Canada (2019), pp. 15210 -- 15219 . Cutkosky, A., and Orabona, F. Momentum-based variance reduction in nonconvex SGD. In Advances in Neural Information Processing Systems 32: Annual Conference on Neural Information Processing Systems 2019, NeurIPS 2019, December 8--14, 2019, Vancouver, BC, Canada (2019), pp. 15210--15219."},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3320269.3384733"},{"key":"e_1_3_2_1_40_1","volume-title":"8th International Conference on Learning Representations, ICLR 2020","author":"Engel J. H.","year":"2020","unstructured":"Engel , J. H. , Hantrakul , L. , Gu , C. , and Roberts , A . DDSP: differentiable digital signal processing . In 8th International Conference on Learning Representations, ICLR 2020 , Addis Ababa, Ethiopia, April 26--30 , 2020 (2020), OpenReview.net. Engel, J. H., Hantrakul, L., Gu, C., and Roberts, A. DDSP: differentiable digital signal processing. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26--30, 2020 (2020), OpenReview.net."},{"key":"e_1_3_2_1_41_1","volume-title":"Audio engineering society convention 108","author":"Farina A.","year":"2000","unstructured":"Farina , A. Simultaneous measurement of impulse response and distortion with a swept-sine technique . In Audio engineering society convention 108 ( 2000 ), Audio Engineering Society . Farina, A. Simultaneous measurement of impulse response and distortion with a swept-sine technique. In Audio engineering society convention 108 (2000), Audio Engineering Society."},{"key":"e_1_3_2_1_42_1","volume-title":"Fsd50k: an open dataset of human-labeled sound events","author":"Fonseca E.","year":"2022","unstructured":"Fonseca , E. , Favory , X. , Pons , J. , Font , F. , and Serra , X . Fsd50k: an open dataset of human-labeled sound events . IEEE\/ACM Transactions on Audio, Speech, and Language Processing 30 ( 2022 ), 829--852. Fonseca, E., Favory, X., Pons, J., Font, F., and Serra, X. Fsd50k: an open dataset of human-labeled sound events. IEEE\/ACM Transactions on Audio, Speech, and Language Processing 30 (2022), 829--852."},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.6028\/NIST.IR.4930"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3422622"},{"key":"e_1_3_2_1_45_1","volume-title":"3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7--9, 2015, Conference Track Proceedings","author":"Goodfellow I. J.","year":"2015","unstructured":"Goodfellow , I. J. , Shlens , J. , and Szegedy , C . Explaining and harnessing adversarial examples . In 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7--9, 2015, Conference Track Proceedings ( 2015 ). Goodfellow, I. J., Shlens, J., and Szegedy, C. Explaining and harnessing adversarial examples. In 3rd International Conference on Learning Representations, ICLR 2015, San Diego, CA, USA, May 7--9, 2015, Conference Track Proceedings (2015)."},{"key":"e_1_3_2_1_46_1","volume-title":"Visqol: an objective speech quality model. EURASIP Journal on Audio, Speech, and Music Processing","author":"Hines A.","year":"2015","unstructured":"Hines , A. , Skoglund , J. , Kokaram , A. C. , and Harte , N . Visqol: an objective speech quality model. EURASIP Journal on Audio, Speech, and Music Processing 2015 , 1 (2015), 1--18. Hines, A., Skoglund, J., Kokaram, A. C., and Harte, N. Visqol: an objective speech quality model. EURASIP Journal on Audio, Speech, and Music Processing 2015, 1 (2015), 1--18."},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-07974-5_2"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3487552.3487842"},{"key":"e_1_3_2_1_49_1","unstructured":"Mendes E. and Hogan K. Defending against imperceptible audio adversarial examples using proportional additive gaussian noise.  Mendes E. and Hogan K. Defending against imperceptible audio adversarial examples using proportional additive gaussian noise."},{"key":"e_1_3_2_1_50_1","first-page":"2127","volume-title":"22nd Annual Conference of the International Speech Communication Association","author":"Mittag G.","year":"2021","unstructured":"Mittag , G. , Naderi , B. , Chehadi , A. , and M\u00f6ller , S . NISQA: A deep cnnself-attention model for multidimensional speech quality prediction with crowdsourced datasets. In Interspeech 2021 , 22nd Annual Conference of the International Speech Communication Association , Brno, Czechia , 30 August - 3 September 2021 (2021), ISCA, pp. 2127 -- 2131 . Mittag, G., Naderi, B., Chehadi, A., and M\u00f6ller, S. NISQA: A deep cnnself-attention model for multidimensional speech quality prediction with crowdsourced datasets. In Interspeech 2021, 22nd Annual Conference of the International Speech Communication Association, Brno, Czechia, 30 August - 3 September 2021 (2021), ISCA, pp. 2127--2131."},{"key":"e_1_3_2_1_51_1","volume-title":"Distance is not a barrier: The use of videoconferencing to develop a community of practice. The Journal of Mental Health Training, Education and Practice","author":"Page R.","year":"2018","unstructured":"Page , R. , Hynes , F. , and Reed , J . Distance is not a barrier: The use of videoconferencing to develop a community of practice. The Journal of Mental Health Training, Education and Practice ( 2018 ). Page, R., Hynes, F., and Reed, J. Distance is not a barrier: The use of videoconferencing to develop a community of practice. The Journal of Mental Health Training, Education and Practice (2018)."},{"key":"e_1_3_2_1_52_1","volume-title":"On the momentum term in gradient descent learning algorithms. Neural networks 12, 1","author":"Qian N.","year":"1999","unstructured":"Qian , N. On the momentum term in gradient descent learning algorithms. Neural networks 12, 1 ( 1999 ), 145--151. Qian, N. On the momentum term in gradient descent learning algorithms. Neural networks 12, 1 (1999), 145--151."},{"key":"e_1_3_2_1_53_1","unstructured":"Research V. M. Intelligent virtual assistant market size worth $ 50.9 billion globally by 2028 at 30 cagr: Verified market research. Accessed: 2022-03--30.  Research V. M. Intelligent virtual assistant market size worth $ 50.9 billion globally by 2028 at 30 cagr: Verified market research. Accessed: 2022-03--30."},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2019.23288"},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2013.02.026"},{"key":"e_1_3_2_1_56_1","volume-title":"Tesseract ocr engine. Lecture. Google Code","author":"Smith R.","year":"2007","unstructured":"Smith , R. , Tesseract ocr engine. Lecture. Google Code . Google Inc ( 2007 ). Smith, R., et al. Tesseract ocr engine. Lecture. Google Code. Google Inc (2007)."},{"key":"e_1_3_2_1_57_1","first-page":"4506","volume-title":"21st Annual Conference of the International Speech Communication Association, Virtual Event","author":"Su J.","year":"2020","unstructured":"Su , J. , Jin , Z. , and Finkelstein , A . Hifi-gan: High-fidelity denoising and dereverberation based on speech deep features in adversarial networks. In Interspeech 2020 , 21st Annual Conference of the International Speech Communication Association, Virtual Event , Shanghai, China, 25- -29 October 2020 (2020), ISCA, pp. 4506 -- 4510 . Su, J., Jin, Z., and Finkelstein, A. Hifi-gan: High-fidelity denoising and dereverberation based on speech deep features in adversarial networks. In Interspeech 2020, 21st Annual Conference of the International Speech Communication Association, Virtual Event, Shanghai, China, 25--29 October 2020 (2020), ISCA, pp. 4506--4510."},{"key":"e_1_3_2_1_58_1","first-page":"1327","volume-title":"29th USENIX Security Symposium (USENIX Security 20)","author":"Suya F.","year":"2020","unstructured":"Suya , F. , Chi , J. , Evans , D. , and Tian , Y . Hybrid batch attacks: Finding black-box adversarial examples with limited queries . In 29th USENIX Security Symposium (USENIX Security 20) ( 2020 ), pp. 1327 -- 1344 . Suya, F., Chi, J., Evans, D., and Tian, Y. Hybrid batch attacks: Finding black-box adversarial examples with limited queries. In 29th USENIX Security Symposium (USENIX Security 20) (2020), pp. 1327--1344."},{"key":"e_1_3_2_1_59_1","volume-title":"Digital signal processing: fundamentals and applications","author":"Tan L.","year":"2018","unstructured":"Tan , L. , and Jiang , J . Digital signal processing: fundamentals and applications . Academic Press , 2018 . Tan, L., and Jiang, J. Digital signal processing: fundamentals and applications. Academic Press, 2018."},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"crossref","first-page":"15","DOI":"10.1109\/SPW.2019.00016","volume-title":"Targeted adversarial examples for black box audio systems. In 2019 IEEE security and privacy workshops (SPW)","author":"Taori R.","year":"2019","unstructured":"Taori , R. , Kamsetty , A. , Chu , B. , and Vemuri , N . Targeted adversarial examples for black box audio systems. In 2019 IEEE security and privacy workshops (SPW) ( 2019 ), IEEE , pp. 15 -- 20 . Taori, R., Kamsetty, A., Chu, B., and Vemuri, N. Targeted adversarial examples for black box audio systems. In 2019 IEEE security and privacy workshops (SPW) (2019), IEEE, pp. 15--20."},{"key":"e_1_3_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00046"},{"key":"e_1_3_2_1_62_1","first-page":"5334","volume-title":"Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 (7 2019), International Joint Conferences on Artificial Intelligence Organization","author":"Yakura H.","unstructured":"Yakura , H. , and Sakuma , J . Robust audio adversarial example for a physical attack . In Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 (7 2019), International Joint Conferences on Artificial Intelligence Organization , pp. 5334 -- 5341 . Yakura, H., and Sakuma, J. Robust audio adversarial example for a physical attack. In Proceedings of the Twenty-Eighth International Joint Conference on Artificial Intelligence, IJCAI-19 (7 2019), International Joint Conferences on Artificial Intelligence Organization, pp. 5334--5341."},{"key":"e_1_3_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2020.24068"},{"key":"e_1_3_2_1_64_1","doi-asserted-by":"publisher","DOI":"10.1109\/COMST.2021.3081450"},{"key":"e_1_3_2_1_65_1","first-page":"1","article-title":"On the influence of momentum acceleration on online learning","volume":"17","author":"Yuan K.","year":"2016","unstructured":"Yuan , K. , Ying , B. , and Sayed , A. H . On the influence of momentum acceleration on online learning . The Journal of Machine Learning Research 17 , 1 ( 2016 ), 6602-- 6667. Yuan, K., Ying, B., and Sayed, A. H. On the influence of momentum acceleration on online learning. The Journal of Machine Learning Research 17, 1 (2016), 6602-- 6667.","journal-title":"The Journal of Machine Learning Research"},{"key":"e_1_3_2_1_66_1","first-page":"49","volume-title":"27th {USENIX} Security Symposium ({USENIX} Security 18)","author":"Yuan X.","year":"2018","unstructured":"Yuan , X. , Chen , Y. , Zhao , Y. , Long , Y. , Liu , X. , Chen , K. , Zhang , S. , Huang , H. , Wang , X. , and Gunter , C. A . Commandersong: A systematic approach for practical adversarial voice recognition . In 27th {USENIX} Security Symposium ({USENIX} Security 18) ( 2018 ), pp. 49 -- 64 . Yuan, X., Chen, Y., Zhao, Y., Long, Y., Liu, X., Chen, K., Zhang, S., Huang, H., Wang, X., and Gunter, C. A. Commandersong: A systematic approach for practical adversarial voice recognition. In 27th {USENIX} Security Symposium ({USENIX} Security 18) (2018), pp. 49--64."},{"key":"e_1_3_2_1_67_1","volume-title":"Soundstream: An end-to-end neural audio codec","author":"Zeghidour N.","year":"2021","unstructured":"Zeghidour , N. , Luebs , A. , Omran , A. , Skoglund , J. , and Tagliasacchi , M . Soundstream: An end-to-end neural audio codec . IEEE\/ACM Transactions on Audio, Speech, and Language Processing ( 2021 ). Zeghidour, N., Luebs, A., Omran, A., Skoglund, J., and Tagliasacchi, M. Soundstream: An end-to-end neural audio codec. IEEE\/ACM Transactions on Audio, Speech, and Language Processing (2021)."},{"key":"e_1_3_2_1_68_1","doi-asserted-by":"publisher","DOI":"10.1145\/3133956.3134052"},{"key":"e_1_3_2_1_69_1","doi-asserted-by":"publisher","DOI":"10.1145\/3460120.3485383"}],"event":{"name":"CCS '22: 2022 ACM SIGSAC Conference on Computer and Communications Security","location":"Los Angeles CA USA","acronym":"CCS '22","sponsor":["SIGSAC ACM Special Interest Group on Security, Audit, and Control"]},"container-title":["Proceedings of the 2022 ACM SIGSAC Conference on Computer and Communications Security"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3548606.3560671","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3548606.3560671","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3548606.3560671","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:59Z","timestamp":1750182539000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3548606.3560671"}},"subtitle":["Targeted Adversarial Voice over IP Network"],"short-title":[],"issued":{"date-parts":[[2022,11,7]]},"references-count":69,"alternative-id":["10.1145\/3548606.3560671","10.1145\/3548606"],"URL":"https:\/\/doi.org\/10.1145\/3548606.3560671","relation":{},"subject":[],"published":{"date-parts":[[2022,11,7]]},"assertion":[{"value":"2022-11-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}