{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,11,15]],"date-time":"2025-11-15T10:27:16Z","timestamp":1763202436784,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":43,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,11,16]],"date-time":"2020-11-16T00:00:00Z","timestamp":1605484800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,11,16]]},"DOI":"10.1145\/3417313.3429385","type":"proceedings-article","created":{"date-parts":[[2020,11,5]],"date-time":"2020-11-05T16:05:52Z","timestamp":1604592352000},"page":"48-54","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["A Low footprint Automatic Speech Recognition System For Resource Constrained Edge Devices"],"prefix":"10.1145","author":[{"given":"Swarnava","family":"Dey","sequence":"first","affiliation":[{"name":"TCS Research and Innovation, Tata Consultancy Services Ltd. Kolkata, West Bengal, India"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jeet","family":"Dutta","sequence":"additional","affiliation":[{"name":"TCS Research and Innovation, Tata Consultancy Services Ltd. Kolkata, West Bengal, India"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,11,16]]},"reference":[{"key":"e_1_3_2_1_1_1","first-page":"2019","volume-title":"20th Annual Conference of the International Speech Communication Association","author":"Bhosale Swapnil","year":"2019","unstructured":"Swapnil Bhosale , Imran Sheikh , Sri Harsha Dumpala , and Sunil Kumar Kopparapu . 2019 . End-to-End Spoken Language Understanding: Bootstrapping in Low Resource Scenarios. In Interspeech 2019 , 20th Annual Conference of the International Speech Communication Association , Graz, Austria, 15- -19 September 2019, Gernot Kubin and Zdravko Kacic (Eds.). ISCA, 1188--1192. https:\/\/doi.org\/10.21437\/Interspeech. 2019 - 2366 10.21437\/Interspeech.2019-2366 Swapnil Bhosale, Imran Sheikh, Sri Harsha Dumpala, and Sunil Kumar Kopparapu. 2019. End-to-End Spoken Language Understanding: Bootstrapping in Low Resource Scenarios. In Interspeech 2019, 20th Annual Conference of the International Speech Communication Association, Graz, Austria, 15--19 September 2019, Gernot Kubin and Zdravko Kacic (Eds.). ISCA, 1188--1192. https:\/\/doi.org\/10.21437\/Interspeech.2019-2366"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASRU46091.2019.9003854"},{"key":"e_1_3_2_1_3_1","volume-title":"Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio.","author":"Cho Kyunghyun","year":"2014","unstructured":"Kyunghyun Cho , Bart Van Merri\u00ebnboer , Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 . Learning phrase representations using RNN encoder-decoder for statistical machine translation. arXiv preprint arXiv:1406.1078 (2014). Kyunghyun Cho, Bart Van Merri\u00ebnboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014. Learning phrase representations using RNN encoder-decoder for statistical machine translation. arXiv preprint arXiv:1406.1078 (2014)."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.195"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/PERCOMW.2019.8730817"},{"volume-title":"Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 818--823","author":"Dey S.","key":"e_1_3_2_1_6_1","unstructured":"S. Dey and A. Mukherjee . 2018 . Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 818--823 . S. Dey and A. Mukherjee. 2018. Implementing Deep Learning and Inferencing on Fog and Edge Computing Systems. In 2018 IEEE International Conference on Pervasive Computing and Communications Workshops (PerCom Workshops). 818--823."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3277893.3277899"},{"key":"e_1_3_2_1_8_1","volume-title":"Practice: Case for Model Partitioning (SenSys-ML","author":"Dey Swarnava","year":"2019","unstructured":"Swarnava Dey , Arijit Mukherjee , Arpan Pal , and Balamuralidhar P . 2019 . Embedded Deep Inference in Practice: Case for Model Partitioning (SenSys-ML 2019). Association for Computing Machinery , New York, NY, USA, 25--30. https:\/\/doi.org\/10.1145\/3362743.3362964 10.1145\/3362743.3362964 Swarnava Dey, Arijit Mukherjee, Arpan Pal, and Balamuralidhar P. 2019. Embedded Deep Inference in Practice: Case for Model Partitioning (SenSys-ML 2019). Association for Computing Machinery, New York, NY, USA, 25--30. https:\/\/doi.org\/10.1145\/3362743.3362964"},{"key":"e_1_3_2_1_9_1","volume-title":"Retrieved","author":"Dominik","year":"2020","unstructured":"Dominik domcross. 2020 . DeepSpeech 0.6.0 for Jetson Nano . Retrieved September 17, 2020 from https:\/\/github.com\/domcross\/DeepSpeech-for-Jetson-Nano\/releases\/tag\/v0.6.0 Dominik domcross. 2020. DeepSpeech 0.6.0 for Jetson Nano. Retrieved September 17, 2020 from https:\/\/github.com\/domcross\/DeepSpeech-for-Jetson-Nano\/releases\/tag\/v0.6.0"},{"key":"e_1_3_2_1_10_1","volume-title":"Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP).","author":"Pavankumar Dubagunta S.","year":"2019","unstructured":"S. Pavankumar Dubagunta and Mathew Magimai .-Doss. 2019 . Segment-level training of ANNs based on acoustic confidence measures for hybrid HMM\/ANN Speech Recognition . In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). S. Pavankumar Dubagunta and Mathew Magimai.-Doss. 2019. Segment-level training of ANNs based on acoustic confidence measures for hybrid HMM\/ANN Speech Recognition. In Proceedings of the IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP)."},{"volume-title":"2017 Ninth International Conference on Ubiquitous and Future Networks (ICUFN). 119--121","author":"Fayjie A. R.","key":"e_1_3_2_1_11_1","unstructured":"A. R. Fayjie , A. Ramezani , D. Oualid , and D. J. Lee . 2017. Voice enabled smart drone control . In 2017 Ninth International Conference on Ubiquitous and Future Networks (ICUFN). 119--121 . A. R. Fayjie, A. Ramezani, D. Oualid, and D. J. Lee. 2017. Voice enabled smart drone control. In 2017 Ninth International Conference on Ubiquitous and Future Networks (ICUFN). 119--121."},{"key":"e_1_3_2_1_12_1","volume-title":"based Embedded Speech to Text Using Deepspeech. ArXiv abs\/2002.12830","author":"Firmansyah Muhammad Hafidh","year":"2020","unstructured":"Muhammad Hafidh Firmansyah , A. Paul , D. Bhattacharya , and Gul Malik Urfa . 2020. A. I. based Embedded Speech to Text Using Deepspeech. ArXiv abs\/2002.12830 ( 2020 ). Muhammad Hafidh Firmansyah, A. Paul, D. Bhattacharya, and Gul Malik Urfa. 2020. A.I. based Embedded Speech to Text Using Deepspeech. ArXiv abs\/2002.12830 (2020)."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"crossref","unstructured":"Alex Graves. 2012. Sequence Transduction with Recurrent Neural Networks. arXiv:1211.3711 [cs.NE]  Alex Graves. 2012. Sequence Transduction with Recurrent Neural Networks. arXiv:1211.3711 [cs.NE]","DOI":"10.1007\/978-3-642-24797-2"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1143844.1143891"},{"key":"e_1_3_2_1_15_1","volume-title":"Ng","author":"Hannun Awni","year":"2014","unstructured":"Awni Hannun , Carl Case , Jared Casper , Bryan Catanzaro , Greg Diamos , Erich Elsen , Ryan Prenger , Sanjeev Satheesh , Shubho Sengupta , Adam Coates , and Andrew Y . Ng . 2014 . Deep Speech : Scaling up end-to-end speech recognition. http:\/\/arxiv.org\/abs\/1412.5567 cite arxiv:1412.5567. Awni Hannun, Carl Case, Jared Casper, Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, and Andrew Y. Ng. 2014. Deep Speech: Scaling up end-to-end speech recognition. http:\/\/arxiv.org\/abs\/1412.5567 cite arxiv:1412.5567."},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.21437\/Interspeech.2006-449"},{"key":"e_1_3_2_1_17_1","volume-title":"Long short-term memory. Neural computation 9, 8","author":"Hochreiter Sepp","year":"1997","unstructured":"Sepp Hochreiter and J\u00fcrgen Schmidhuber . 1997. Long short-term memory. Neural computation 9, 8 ( 1997 ), 1735--1780. Sepp Hochreiter and J\u00fcrgen Schmidhuber. 1997. Long short-term memory. Neural computation 9, 8 (1997), 1735--1780."},{"key":"e_1_3_2_1_18_1","volume-title":"Retrieved","author":"Hong YoonSeok","year":"2020","unstructured":"YoonSeok Hong . 2020 . NeurIPS 2020 Demo: MONICA . Retrieved September 17, 2020 from https:\/\/www.youtube.com\/watch?v=nT4loeidjFc YoonSeok Hong. 2020. NeurIPS 2020 Demo: MONICA. Retrieved September 17, 2020 from https:\/\/www.youtube.com\/watch?v=nT4loeidjFc"},{"key":"e_1_3_2_1_19_1","unstructured":"iamjanvijay \/ rnnt. 2020. RNN-Transducer Loss. Retrieved August 6 2020 from https:\/\/github.com\/iamjanvijay\/rnnt  iamjanvijay \/ rnnt. 2020. RNN-Transducer Loss. Retrieved August 6 2020 from https:\/\/github.com\/iamjanvijay\/rnnt"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"crossref","unstructured":"Hirofumi Inaguma Yashesh Gaur Liang Lu Jinyu Li and Yifan Gong. 2020. Minimum latency training strategies for streaming sequence-to-sequence ASR. In ICASSP. https:\/\/www.microsoft.com\/en-us\/research\/publication\/minimum-latency-training-strategies-for-streaming-sequence-to-sequence-asr\/  Hirofumi Inaguma Yashesh Gaur Liang Lu Jinyu Li and Yifan Gong. 2020. Minimum latency training strategies for streaming sequence-to-sequence ASR. In ICASSP. https:\/\/www.microsoft.com\/en-us\/research\/publication\/minimum-latency-training-strategies-for-streaming-sequence-to-sequence-asr\/","DOI":"10.1109\/ICASSP40776.2020.9054098"},{"key":"e_1_3_2_1_21_1","volume-title":"Retrieved","author":"Karpathy Andrej","year":"2020","unstructured":"Andrej Karpathy . 2020 . The Unreasonable Effectiveness of Recurrent Neural Networks . Retrieved August 28, 2020 fromhttp:\/\/karpathy.github.io\/2015\/05\/21\/rnn-effectiveness\/ Andrej Karpathy. 2020. The Unreasonable Effectiveness of Recurrent Neural Networks. Retrieved August 28, 2020 fromhttp:\/\/karpathy.github.io\/2015\/05\/21\/rnn-effectiveness\/"},{"key":"e_1_3_2_1_22_1","unstructured":"Andrej Karpathy Justin Johnson and Li Fei-Fei. 2015. Visualizing and Understanding Recurrent Networks. arXiv:1506.02078 [cs.LG]  Andrej Karpathy Justin Johnson and Li Fei-Fei. 2015. Visualizing and Understanding Recurrent Networks. arXiv:1506.02078 [cs.LG]"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3029798.3038329"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASRU46091.2019.9003972"},{"key":"e_1_3_2_1_25_1","volume-title":"Visualizing memorization in RNNs. Distill","author":"Madsen Andreas","year":"2019","unstructured":"Andreas Madsen . 2019. Visualizing memorization in RNNs. Distill ( 2019 ). https:\/\/doi.org\/10.23915\/distill.00016 https:\/\/distill.pub\/2019\/memorization-in-rnns. 10.23915\/distill.00016 Andreas Madsen. 2019. Visualizing memorization in RNNs. Distill (2019). https:\/\/doi.org\/10.23915\/distill.00016 https:\/\/distill.pub\/2019\/memorization-in-rnns."},{"volume-title":"mozilla \/ DeepSpeech -- deepspeech 0.6.1. Retrieved","year":"2020","key":"e_1_3_2_1_26_1","unstructured":"mozilla \/ DeepSpeech. 2020. mozilla \/ DeepSpeech -- deepspeech 0.6.1. Retrieved August 6, 2020 from https:\/\/github.com\/mozilla\/DeepSpeech mozilla \/ DeepSpeech. 2020. mozilla \/ DeepSpeech -- deepspeech 0.6.1. Retrieved August 6, 2020 from https:\/\/github.com\/mozilla\/DeepSpeech"},{"key":"e_1_3_2_1_27_1","volume-title":"kaldi-asr. Retrieved","author":"Nagendra Kumar Goel Vimal Manohar","year":"2020","unstructured":"Vimal Manohar Nagendra Kumar Goel . 2017. kaldi-asr. Retrieved August 6, 2020 from https:\/\/github.com\/kaldi-asr\/kaldi\/blob\/master\/egs\/swbd\/s5c\/local\/run_asr_segmentation.sh Vimal Manohar Nagendra Kumar Goel. 2017. kaldi-asr. Retrieved August 6, 2020 from https:\/\/github.com\/kaldi-asr\/kaldi\/blob\/master\/egs\/swbd\/s5c\/local\/run_asr_segmentation.sh"},{"key":"e_1_3_2_1_28_1","volume-title":"Retrieved","author":"NVIDIA.","year":"2020","unstructured":"NVIDIA. 2020 . Jetson Nano Developer Kit . Retrieved September 18, 2020 from https:\/\/developer.nvidia.com\/embedded\/jetson-nano-developer-kit NVIDIA. 2020. Jetson Nano Developer Kit. Retrieved September 18, 2020 from https:\/\/developer.nvidia.com\/embedded\/jetson-nano-developer-kit"},{"key":"e_1_3_2_1_29_1","volume-title":"Kite: Automatic speech recognition for unmanned aerial vehicles. CoRR abs\/1907.01195","author":"Oneata Dan","year":"2019","unstructured":"Dan Oneata and Horia Cucu . 2019 . Kite: Automatic speech recognition for unmanned aerial vehicles. CoRR abs\/1907.01195 (2019). arXiv:1907.01195 http:\/\/arxiv.org\/abs\/1907.01195 Dan Oneata and Horia Cucu. 2019. Kite: Automatic speech recognition for unmanned aerial vehicles. CoRR abs\/1907.01195 (2019). arXiv:1907.01195 http:\/\/arxiv.org\/abs\/1907.01195"},{"volume-title":"2015 IEEE International Conference on Acoustics, Speech and Signal Processing(ICASSP). 5206--5210","author":"Panayotov V.","key":"e_1_3_2_1_30_1","unstructured":"V. Panayotov , G. Chen , D. Povey , and S. Khudanpur . 2015. Librispeech: An ASR corpus based on public domain audio books . In 2015 IEEE International Conference on Acoustics, Speech and Signal Processing(ICASSP). 5206--5210 . V. Panayotov, G. Chen, D. Povey, and S. Khudanpur. 2015. Librispeech: An ASR corpus based on public domain audio books. In 2015 IEEE International Conference on Acoustics, Speech and Signal Processing(ICASSP). 5206--5210."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.5555\/3327546.3327722"},{"key":"e_1_3_2_1_32_1","volume-title":"On the Compression of Recurrent Neural Networks with an Application to LVCSR acoustic modeling for Embedded Speech Recognition. CoRR abs\/1603.08042","author":"Prabhavalkar Rohit","year":"2016","unstructured":"Rohit Prabhavalkar , Ouais Alsharif , Antoine Bruguier , and Ian McGraw . 2016. On the Compression of Recurrent Neural Networks with an Application to LVCSR acoustic modeling for Embedded Speech Recognition. CoRR abs\/1603.08042 ( 2016 ). arXiv:1603.08042 http:\/\/arxiv.org\/abs\/1603.08042 Rohit Prabhavalkar, Ouais Alsharif, Antoine Bruguier, and Ian McGraw. 2016. On the Compression of Recurrent Neural Networks with an Application to LVCSR acoustic modeling for Embedded Speech Recognition. CoRR abs\/1603.08042 (2016). arXiv:1603.08042 http:\/\/arxiv.org\/abs\/1603.08042"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"crossref","unstructured":"Vineel Pratap Qiantong Xu Jacob Kahn Gilad Avidov Tatiana Likhomanenko Awni Hannun Vitaliy Liptchinsky Gabriel Synnaeve and Ronan Collobert. 2020. Scaling Up Online Speech Recognition Using ConvNets. arXiv:2001.09727 [cs.CL]  Vineel Pratap Qiantong Xu Jacob Kahn Gilad Avidov Tatiana Likhomanenko Awni Hannun Vitaliy Liptchinsky Gabriel Synnaeve and Ronan Collobert. 2020. Scaling Up Online Speech Recognition Using ConvNets. arXiv:2001.09727 [cs.CL]","DOI":"10.21437\/Interspeech.2020-2840"},{"key":"e_1_3_2_1_34_1","unstructured":"Kanishka Rao Ha\u015fim Sak and Rohit Prabhavalkar. 2018. Exploring Architectures Data and Units For Streaming End-to-End Speech Recognition with RNN-Transducer. arXiv:1801.00841 [cs.CL]  Kanishka Rao Ha\u015fim Sak and Rohit Prabhavalkar. 2018. Exploring Architectures Data and Units For Streaming End-to-End Speech Recognition with RNN-Transducer. arXiv:1801.00841 [cs.CL]"},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2017.2647747"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"crossref","unstructured":"Tara Sainath Yanzhang (Ryan) He Bo Li Arun Narayanan Ruoming Pang Antoine Bruguier Shuo yiin Chang Wei Li Raziel Alvarez Zhifeng Chen Chung-Cheng Chiu David Garcia Alex Gruenstein Kevin Hu Minho Jin Anjuli Kannan Qiao Liang Ian McGraw Cal Peyser Rohit Prabhavalkar Golan Pundak David Rybach (June) Yuan Shangguan Yash Sheth Trevor Strohman Mirk\u00f3 Visontai Yonghui Wu Yu Zhang and Ding Zhao. 2020. A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency.  Tara Sainath Yanzhang (Ryan) He Bo Li Arun Narayanan Ruoming Pang Antoine Bruguier Shuo yiin Chang Wei Li Raziel Alvarez Zhifeng Chen Chung-Cheng Chiu David Garcia Alex Gruenstein Kevin Hu Minho Jin Anjuli Kannan Qiao Liang Ian McGraw Cal Peyser Rohit Prabhavalkar Golan Pundak David Rybach (June) Yuan Shangguan Yash Sheth Trevor Strohman Mirk\u00f3 Visontai Yonghui Wu Yu Zhang and Ding Zhao. 2020. A Streaming On-Device End-to-End Model Surpassing Server-Side Conventional Model Quality and Latency.","DOI":"10.1109\/ICASSP40776.2020.9054188"},{"key":"e_1_3_2_1_37_1","volume-title":"Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models. CoRR abs\/1708.08296","author":"Samek Wojciech","year":"2017","unstructured":"Wojciech Samek , Thomas Wiegand , and Klaus-Robert M\u00fcller . 2017. Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models. CoRR abs\/1708.08296 ( 2017 ). arXiv:1708.08296 http:\/\/arxiv.org\/abs\/1708.08296 Wojciech Samek, Thomas Wiegand, and Klaus-Robert M\u00fcller. 2017. Explainable Artificial Intelligence: Understanding, Visualizing and Interpreting Deep Learning Models. CoRR abs\/1708.08296 (2017). arXiv:1708.08296 http:\/\/arxiv.org\/abs\/1708.08296"},{"key":"e_1_3_2_1_38_1","volume-title":"Inverted Residuals and Linear Bottlenecks: Mobile Networks for Classification, Detection and Segmentation. CoRR abs\/1801.04381","author":"Sandler Mark","year":"2018","unstructured":"Mark Sandler , Andrew G. Howard , Menglong Zhu , Andrey Zhmoginov , and Liang-Chieh Chen . 2018. Inverted Residuals and Linear Bottlenecks: Mobile Networks for Classification, Detection and Segmentation. CoRR abs\/1801.04381 ( 2018 ). arXiv:1801.04381 http:\/\/arxiv.org\/abs\/1801.04381 Mark Sandler, Andrew G. Howard, Menglong Zhu, Andrey Zhmoginov, and Liang-Chieh Chen. 2018. Inverted Residuals and Linear Bottlenecks: Mobile Networks for Classification, Detection and Segmentation. CoRR abs\/1801.04381 (2018). arXiv:1801.04381 http:\/\/arxiv.org\/abs\/1801.04381"},{"volume-title":"2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). 5149--5152","author":"Schuster M.","key":"e_1_3_2_1_39_1","unstructured":"M. Schuster and K. Nakajima . 2012. Japanese and Korean voice search . In 2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). 5149--5152 . M. Schuster and K. Nakajima. 2012. Japanese and Korean voice search. In 2012 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). 5149--5152."},{"volume-title":"2014 Third ICT International Student Project Conference (ICT-ISPC). 107--110","author":"Supimros S.","key":"e_1_3_2_1_40_1","unstructured":"S. Supimros and S. Wongthanavasu . 2014. Speech recognition -- based control system for Drone . In 2014 Third ICT International Student Project Conference (ICT-ISPC). 107--110 . S. Supimros and S. Wongthanavasu. 2014. Speech recognition -- based control system for Drone. In 2014 Third ICT International Student Project Conference (ICT-ISPC). 107--110."},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2009.4960701"},{"volume-title":"Python interface to the WebRTC Voice Activity Detector. Retrieved","year":"2020","key":"e_1_3_2_1_42_1","unstructured":"wiseman \/ py webrtcvad. 2019. Python interface to the WebRTC Voice Activity Detector. Retrieved August 6, 2020 from https:\/\/github.com\/wiseman\/py-webrtcvad wiseman \/ py webrtcvad. 2019. Python interface to the WebRTC Voice Activity Detector. Retrieved August 6, 2020 from https:\/\/github.com\/wiseman\/py-webrtcvad"},{"key":"e_1_3_2_1_43_1","volume-title":"Retrieved","author":"XPRIZE.","year":"2020","unstructured":"XPRIZE. 2020 . Anywhere is possible . Retrieved August 28, 2020 from https:\/\/avatar.xprize.org\/prizes\/avatar XPRIZE. 2020. Anywhere is possible. Retrieved August 28, 2020 from https:\/\/avatar.xprize.org\/prizes\/avatar"}],"event":{"name":"SenSys '20: The 18th ACM Conference on Embedded Networked Sensor Systems","sponsor":["SIGMETRICS ACM Special Interest Group on Measurement and Evaluation","SIGCOMM ACM Special Interest Group on Data Communication","SIGMOBILE ACM Special Interest Group on Mobility of Systems, Users, Data and Computing","SIGOPS ACM Special Interest Group on Operating Systems","SIGBED ACM Special Interest Group on Embedded Systems","SIGARCH ACM Special Interest Group on Computer Architecture"],"location":"Virtual Event Japan","acronym":"SenSys '20"},"container-title":["Proceedings of the 2nd International Workshop on Challenges in Artificial Intelligence and Machine Learning for Internet of Things"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3417313.3429385","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3417313.3429385","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:48:14Z","timestamp":1750193294000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3417313.3429385"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,11,16]]},"references-count":43,"alternative-id":["10.1145\/3417313.3429385","10.1145\/3417313"],"URL":"https:\/\/doi.org\/10.1145\/3417313.3429385","relation":{},"subject":[],"published":{"date-parts":[[2020,11,16]]},"assertion":[{"value":"2020-11-16","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}