{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,9]],"date-time":"2026-05-09T05:52:46Z","timestamp":1778305966244,"version":"3.51.4"},"reference-count":58,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2017,9,11]],"date-time":"2017-09-11T00:00:00Z","timestamp":1505088000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Interact. Mob. Wearable Ubiquitous Technol."],"published-print":{"date-parts":[[2017,9,11]]},"abstract":"<jats:p>Continuous audio analysis from embedded and mobile devices is an increasingly important application domain. More and more, appliances like the Amazon Echo, along with smartphones and watches, and even research prototypes seek to perform multiple discriminative tasks simultaneously from ambient audio; for example, monitoring background sound classes (e.g., music or conversation), recognizing certain keywords (\u2018Hey Siri' or \u2018Alexa'), or identifying the user and her emotion from speech. The use of deep learning algorithms typically provides state-of-the-art model performances for such general audio tasks. However, the large computational demands of deep learning models are at odds with the limited processing, energy and memory resources of mobile, embedded and IoT devices.<\/jats:p>\n          <jats:p>In this paper, we propose and evaluate a novel deep learning modeling and optimization framework that specifically targets this category of embedded audio sensing tasks. Although the supported tasks are simpler than the task of speech recognition, this framework aims at maintaining accuracies in predictions while minimizing the overall processor resource footprint. The proposed model is grounded in multi-task learning principles to train shared deep layers and exploits, as input layer, only statistical summaries of audio filter banks to further lower computations.<\/jats:p>\n          <jats:p>We find that for embedded audio sensing tasks our framework is able to maintain similar accuracies, which are observed in comparable deep architectures that use single-task learning and typically more complex input layers. Most importantly, on an average, this approach provides almost a 2.1\u00d7 reduction in runtime, energy, and memory for four separate audio sensing tasks, assuming a variety of task combinations.<\/jats:p>","DOI":"10.1145\/3131895","type":"journal-article","created":{"date-parts":[[2017,9,11]],"date-time":"2017-09-11T12:12:26Z","timestamp":1505131946000},"page":"1-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":44,"title":["Low-resource Multi-task Audio Sensing for Mobile and Embedded Devices via Shared Deep Neural Network Representations"],"prefix":"10.1145","volume":"1","author":[{"given":"Petko","family":"Georgiev","sequence":"first","affiliation":[{"name":"University of Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sourav","family":"Bhattacharya","sequence":"additional","affiliation":[{"name":"Nokia Bell Labs"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nicholas D.","family":"Lane","sequence":"additional","affiliation":[{"name":"University College London and Nokia Bell Labs"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Cecilia","family":"Mascolo","sequence":"additional","affiliation":[{"name":"University of Cambridge"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2017,9,11]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=188864","unstructured":"2013. Recent Advances in Deep Learning for Speech Research at Microsoft . IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=188864 2013. Recent Advances in Deep Learning for Speech Research at Microsoft. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP). http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=188864"},{"key":"e_1_2_1_2_1","unstructured":"2017. https:\/\/www.qualcomm.com\/products\/snapdragon\/processors\/400. (2017).  2017. https:\/\/www.qualcomm.com\/products\/snapdragon\/processors\/400. (2017)."},{"key":"e_1_2_1_3_1","unstructured":"2017. Amazon Echo. http:\/\/www.amazon.com\/Amazon-Echo-Bluetooth-Speaker-with-WiFi-Alexa\/dp\/B00X4WHP5E. (2017).  2017. Amazon Echo. http:\/\/www.amazon.com\/Amazon-Echo-Bluetooth-Speaker-with-WiFi-Alexa\/dp\/B00X4WHP5E. (2017)."},{"key":"e_1_2_1_4_1","unstructured":"2017. Auto Shazam. https:\/\/support.shazam.com\/hc\/en-us\/articles\/204457738-Auto-Shazam-iPhone-. (2017).  2017. Auto Shazam. https:\/\/support.shazam.com\/hc\/en-us\/articles\/204457738-Auto-Shazam-iPhone-. (2017)."},{"key":"e_1_2_1_5_1","unstructured":"2017. Fitbit Surge. https:\/\/www.fitbit.com\/uk\/surge. (2017).  2017. Fitbit Surge. https:\/\/www.fitbit.com\/uk\/surge. (2017)."},{"key":"e_1_2_1_6_1","unstructured":"2017. Google Home. https:\/\/home.google.com\/. (2017).  2017. Google Home. https:\/\/home.google.com\/. (2017)."},{"key":"e_1_2_1_7_1","unstructured":"2017. Motorola Moto 360 Smartwatch. http:\/\/www.motorola.com\/us\/products\/moto-360. (2017).  2017. Motorola Moto 360 Smartwatch. http:\/\/www.motorola.com\/us\/products\/moto-360. (2017)."},{"key":"e_1_2_1_8_1","unstructured":"2017. Qualcomm Snapdragon 800 MDP. http:\/\/goo.gl\/ySfCFl. (2017).  2017. Qualcomm Snapdragon 800 MDP. http:\/\/goo.gl\/ySfCFl. (2017)."},{"key":"e_1_2_1_9_1","unstructured":"2017. TensorFlow. https:\/\/www.tensorflow.org\/. (2017).  2017. TensorFlow. https:\/\/www.tensorflow.org\/. (2017)."},{"key":"e_1_2_1_10_1","unstructured":"2017. Torch. http:\/\/torch.ch\/. (2017).  2017. Torch. http:\/\/torch.ch\/. (2017)."},{"key":"e_1_2_1_11_1","volume-title":"Workshop on Sensing Systems and Applications Using Wrist Worn Smart Devices (WristSense'16)","author":"Bhattacharya Sourav","unstructured":"Sourav Bhattacharya and Nicholas D. Lane . 2016. From Smart to Deep: Robust Activity Recognition on Smartwatches using Deep Learning . In Workshop on Sensing Systems and Applications Using Wrist Worn Smart Devices (WristSense'16) . Sourav Bhattacharya and Nicholas D. Lane. 2016. From Smart to Deep: Robust Activity Recognition on Smartwatches using Deep Learning. In Workshop on Sensing Systems and Applications Using Wrist Worn Smart Devices (WristSense'16)."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/2994551.2994564"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1007379606734"},{"key":"e_1_2_1_14_1","doi-asserted-by":"crossref","unstructured":"Guoguo Chen Carolina Parada and Georg Heigold. 2014. Small-footprint Keyword Spotting Using Deep Neural Networks (ICASSP'14).  Guoguo Chen Carolina Parada and Georg Heigold. 2014. Small-footprint Keyword Spotting Using Deep Neural Networks (ICASSP'14).","DOI":"10.1109\/ICASSP.2014.6854370"},{"key":"e_1_2_1_15_1","volume-title":"Compressing Neural Networks with the Hashing Trick. ICML-15","author":"Chen Wenlin","year":"2015","unstructured":"Wenlin Chen , James T. Wilson , Stephen Tyree , Kilian Q. Weinberger , and Yixin Chen . 2015. Compressing Neural Networks with the Hashing Trick. ICML-15 ( 2015 ). http:\/\/arxiv.org\/abs\/1504.04788 Wenlin Chen, James T. Wilson, Stephen Tyree, Kilian Q. Weinberger, and Yixin Chen. 2015. Compressing Neural Networks with the Hashing Trick. ICML-15 (2015). http:\/\/arxiv.org\/abs\/1504.04788"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/1390156.1390177"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2013.6639344"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1007\/BF02943243"},{"key":"e_1_2_1_19_1","volume-title":"Bourdev","author":"Gong Yunchao","year":"2015","unstructured":"Yunchao Gong , Liu Liu , Ming Yang , and Lubomir D . Bourdev . 2015 . Compressing Deep Convolutional Networks using Vector Quantization. ICLR- 15 (2015). http:\/\/arxiv.org\/abs\/1412.6115 Yunchao Gong, Liu Liu, Ming Yang, and Lubomir D. Bourdev. 2015. Compressing Deep Convolutional Networks using Vector Quantization. ICLR-15 (2015). http:\/\/arxiv.org\/abs\/1412.6115"},{"key":"e_1_2_1_20_1","doi-asserted-by":"crossref","unstructured":"Nils Hammerla James Fisher Peter Andras Lynn Rochester Richard Walker and Thomas Ploetz. 2015. PD Disease State Assessment in Naturalistic Environments Using Deep Learning. (2015). http:\/\/www.aaai.org\/ocs\/index.php\/AAAI\/AAAI15\/paper\/view\/9930   Nils Hammerla James Fisher Peter Andras Lynn Rochester Richard Walker and Thomas Ploetz. 2015. PD Disease State Assessment in Naturalistic Environments Using Deep Learning. (2015). http:\/\/www.aaai.org\/ocs\/index.php\/AAAI\/AAAI15\/paper\/view\/9930","DOI":"10.1609\/aaai.v29i1.9484"},{"key":"e_1_2_1_21_1","volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI'16)","author":"Hammerla Nils Y.","year":"2016","unstructured":"Nils Y. Hammerla , Shane Halloran , and Thomas Pl\u00f6tz . 2016 . Deep, Convolutional, and Recurrent Models for Human Activity Recognition Using Wearables . In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI'16) . AAAI Press. http:\/\/www.ijcai.org\/Abstract\/16\/220 Nils Y. Hammerla, Shane Halloran, and Thomas Pl\u00f6tz. 2016. Deep, Convolutional, and Recurrent Models for Human Activity Recognition Using Wearables. In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI'16). AAAI Press. http:\/\/www.ijcai.org\/Abstract\/16\/220"},{"key":"e_1_2_1_22_1","doi-asserted-by":"crossref","unstructured":"Kun Han Dong Yu and Ivan Tashev. 2014. Speech Emotion Recognition Using Deep Neural Network and Extreme Learning Machine. In Interspeech-14. http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=230136  Kun Han Dong Yu and Ivan Tashev. 2014. Speech Emotion Recognition Using Deep Neural Network and Extreme Learning Machine. In Interspeech-14. http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=230136","DOI":"10.21437\/Interspeech.2014-57"},{"key":"e_1_2_1_23_1","volume-title":"Dally","author":"Han Song","year":"2015","unstructured":"Song Han , Jeff Pool , John Tran , and William J . Dally . 2015 . Learning both Weights and Connections for Efficient Neural Networks. NIPS- 15 (2015). http:\/\/arxiv.org\/abs\/1506.02626 Song Han, Jeff Pool, John Tran, and William J. Dally. 2015. Learning both Weights and Connections for Efficient Neural Networks. NIPS-15 (2015). http:\/\/arxiv.org\/abs\/1506.02626"},{"key":"e_1_2_1_24_1","volume-title":"Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385","author":"He Kaiming","year":"2015","unstructured":"Kaiming He , Xiangyu Zhang , Shaoqing Ren , and Jian Sun . 2015. Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385 ( 2015 ). http:\/\/arxiv.org\/abs\/1512.03385 Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015. Deep Residual Learning for Image Recognition. CoRR abs\/1512.03385 (2015). http:\/\/arxiv.org\/abs\/1512.03385"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP.2014.6853595"},{"key":"e_1_2_1_26_1","first-page":"4","article-title":"Perceptual Linear Predictive (PLP) Analysis of Speech","volume":"57","author":"Hermansky H.","year":"1990","unstructured":"H. Hermansky . 1990 . Perceptual Linear Predictive (PLP) Analysis of Speech . J. Acoust. Soc. Am. 57 , 4 (April 1990), 1738--52. H. Hermansky. 1990. Perceptual Linear Predictive (PLP) Analysis of Speech. J. Acoust. Soc. Am. 57, 4 (April 1990), 1738--52.","journal-title":"J. Acoust. Soc. Am."},{"key":"e_1_2_1_27_1","doi-asserted-by":"crossref","unstructured":"Jui-Ting Huang Jinyu Li Dong Yu Li Deng and Yifan Gong. 2013. Cross-language Knowledge Transfer using Multilingual Deep Neural Network with Shared Hidden Layers. In ICASSP-13. http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=189250  Jui-Ting Huang Jinyu Li Dong Yu Li Deng and Yifan Gong. 2013. Cross-language Knowledge Transfer using Multilingual Deep Neural Network with Shared Hidden Layers. In ICASSP-13. http:\/\/research.microsoft.com\/apps\/pubs\/default.aspx?id=189250","DOI":"10.1109\/ICASSP.2013.6639081"},{"key":"e_1_2_1_28_1","volume-title":"Hinton","author":"Krizhevsky Alex","year":"2012","unstructured":"Alex Krizhevsky , Ilya Sutskever , and Geoffrey E . Hinton . 2012 . ImageNet Classification with Deep Convolutional Neural Networks. In NIPS- 12. http:\/\/papers.nips.cc\/paper\/4824-imagenet-classification-with-deep-convolutional-neural-networks.pdf Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012. ImageNet Classification with Deep Convolutional Neural Networks. In NIPS-12. http:\/\/papers.nips.cc\/paper\/4824-imagenet-classification-with-deep-convolutional-neural-networks.pdf"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.4108\/eai.30-11-2016.2267463"},{"key":"e_1_2_1_30_1","volume-title":"DeepX: A Software Accelerator for Low-Power Deep Learning Inference on Mobile Devices. In International Conference on Information Processing in Sensor Networks (IPSN '16)","author":"Lane Nicholas D.","year":"2016","unstructured":"Nicholas D. Lane , Sourav Bhattacharya , Petko Georgiev , Claudio Forlivesi , Lei Jiao , Lorena Qendro , and Fahim Kawsar . 2016 . DeepX: A Software Accelerator for Low-Power Deep Learning Inference on Mobile Devices. In International Conference on Information Processing in Sensor Networks (IPSN '16) . Nicholas D. Lane, Sourav Bhattacharya, Petko Georgiev, Claudio Forlivesi, Lei Jiao, Lorena Qendro, and Fahim Kawsar. 2016. DeepX: A Software Accelerator for Low-Power Deep Learning Inference on Mobile Devices. In International Conference on Information Processing in Sensor Networks (IPSN '16)."},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/2699343.2699349"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/2750858.2804262"},{"key":"e_1_2_1_33_1","volume-title":"Ng","author":"Lee Honglak","year":"2009","unstructured":"Honglak Lee , Peter Pham , Yan Largman , and Andrew Y . Ng . 2009 . Unsupervised feature learning for audio classification using convolutional deep belief networks. In NIPS-09. Curran Associates, Inc., 1096--1104. http:\/\/papers.nips.cc\/paper\/3674-unsupervised-feature-learning-for-audio-classification-using-convolutional-deep-belief-networks.pdf Honglak Lee, Peter Pham, Yan Largman, and Andrew Y. Ng. 2009. Unsupervised feature learning for audio classification using convolutional deep belief networks. In NIPS-09. Curran Associates, Inc., 1096--1104. http:\/\/papers.nips.cc\/paper\/3674-unsupervised-feature-learning-for-audio-classification-using-convolutional-deep-belief-networks.pdf"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2462456.2465426"},{"key":"e_1_2_1_35_1","volume-title":"DeepSaliency: Multi-Task Deep Neural Network Model for Salient Object Detection. CoRR abs\/1510.05484","author":"Li Xi","year":"2015","unstructured":"Xi Li , Liming Zhao , Lina Wei , MingHsuan Yang , Fei Wu , Yueting Zhuang , Haibin Ling , and Jingdong Wang . 2015. DeepSaliency: Multi-Task Deep Neural Network Model for Salient Object Detection. CoRR abs\/1510.05484 ( 2015 ). http:\/\/arxiv.org\/abs\/1510.05484 Xi Li, Liming Zhao, Lina Wei, MingHsuan Yang, Fei Wu, Yueting Zhuang, Haibin Ling, and Jingdong Wang. 2015. DeepSaliency: Multi-Task Deep Neural Network Model for Salient Object Detection. CoRR abs\/1510.05484 (2015). http:\/\/arxiv.org\/abs\/1510.05484"},{"key":"e_1_2_1_36_1","unstructured":"Mark Liberman Kelly Davis Murray Grossman Nii Martey and John Bell. 2002. Emotional Prosody Speech and Transcripts. (2002).  Mark Liberman Kelly Davis Murray Grossman Nii Martey and John Bell. 2002. Emotional Prosody Speech and Transcripts. (2002)."},{"key":"e_1_2_1_37_1","volume-title":"NAACL HLT","author":"Liu Xiaodong","year":"2015","unstructured":"Xiaodong Liu , Jianfeng Gao , Xiaodong He , Li Deng , Kevin Duh , and Ye-Yi Wang . 2015. Representation Learning Using Multi-Task Deep Neural Networks for Semantic Classification and Information Retrieval . In NAACL HLT 2015 , The 2015 Conference of the North American Chapter of the Association for Computational Linguistics : Human Language Technologies, Denver, Colorado, USA, May 31 - June 5, 2015. 912--921. http:\/\/aclweb.org\/anthology\/N\/N15\/N15-1092.pdf Xiaodong Liu, Jianfeng Gao, Xiaodong He, Li Deng, Kevin Duh, and Ye-Yi Wang. 2015. Representation Learning Using Multi-Task Deep Neural Networks for Semantic Classification and Information Retrieval. In NAACL HLT 2015, The 2015 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Denver, Colorado, USA, May 31 - June 5, 2015. 912--921. http:\/\/aclweb.org\/anthology\/N\/N15\/N15-1092.pdf"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.5555\/2021975.2021992"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/2370216.2370270"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/1555816.1555834"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1869983.1869992"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/1869983.1869992"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/2517351.2517353"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3081333.3081359"},{"key":"e_1_2_1_45_1","volume-title":"Proceedings of the 27th International Conference on Machine Learning (ICML-10)","author":"Nair Vinod","year":"2010","unstructured":"Vinod Nair and Geoffrey E. Hinton . 2010. Rectified Linear Units Improve Restricted Boltzmann Machines . In Proceedings of the 27th International Conference on Machine Learning (ICML-10) , Johannes F\u00fcrnkranz and Thorsten Joachims (Eds.). Omnipress, 807--814. http:\/\/www.icml 2010 .org\/papers\/432.pdf Vinod Nair and Geoffrey E. Hinton. 2010. Rectified Linear Units Improve Restricted Boltzmann Machines. In Proceedings of the 27th International Conference on Machine Learning (ICML-10), Johannes F\u00fcrnkranz and Thorsten Joachims (Eds.). Omnipress, 807--814. http:\/\/www.icml2010.org\/papers\/432.pdf"},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.5555\/2283516.2283683"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/1864349.1864393"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/1864349.1864393"},{"key":"e_1_2_1_49_1","volume-title":"Histogram of gradients of Time-Frequency Representations for Audio scene detection. CoRR abs\/1508.04909","author":"Rakotomamonjy Alain","year":"2015","unstructured":"Alain Rakotomamonjy and Gilles Gasso . 2015. Histogram of gradients of Time-Frequency Representations for Audio scene detection. CoRR abs\/1508.04909 ( 2015 ). http:\/\/arxiv.org\/abs\/1508.04909 Alain Rakotomamonjy and Gilles Gasso. 2015. Histogram of gradients of Time-Frequency Representations for Audio scene detection. CoRR abs\/1508.04909 (2015). http:\/\/arxiv.org\/abs\/1508.04909"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASSP.1987.1165139"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.220"},{"key":"e_1_2_1_52_1","volume-title":"Theano: A Python framework for fast computation of mathematical expressions. arXiv e-prints abs\/1605.02688 (May","author":"Team Theano Development","year":"2016","unstructured":"Theano Development Team . 2016 . Theano: A Python framework for fast computation of mathematical expressions. arXiv e-prints abs\/1605.02688 (May 2016). http:\/\/arxiv.org\/abs\/1605.02688 Theano Development Team. 2016. Theano: A Python framework for fast computation of mathematical expressions. arXiv e-prints abs\/1605.02688 (May 2016). http:\/\/arxiv.org\/abs\/1605.02688"},{"key":"e_1_2_1_53_1","volume-title":"Ignacio Lopez Moreno, and Javier Gonzalez-Dominguez","author":"Variani Ehsan","year":"2014","unstructured":"Ehsan Variani , Xin Lei , Erik McDermott , Ignacio Lopez Moreno, and Javier Gonzalez-Dominguez . 2014 . Deep neural networks for small footprint text-dependent speaker verification. In ICASSP-14. IEEE , 4052--4056. Ehsan Variani, Xin Lei, Erik McDermott, Ignacio Lopez Moreno, and Javier Gonzalez-Dominguez. 2014. Deep neural networks for small footprint text-dependent speaker verification. In ICASSP-14. IEEE, 4052--4056."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.5555\/2976040.2976226"},{"key":"e_1_2_1_55_1","volume-title":"INTERSPEECH 2015, Automatic Speaker Verification Spoofing and Countermeasures Challenge, colocated with INTERSPEECH 2015","author":"Wu Zhizheng","year":"2015","unstructured":"Zhizheng Wu , Tomi Kinnunen , Nicholas Evans , Junichi Yamagishi , Cemal Hanilci , Md Sahidullah , and Aleksandr Sizov . 2015 . ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge . In INTERSPEECH 2015, Automatic Speaker Verification Spoofing and Countermeasures Challenge, colocated with INTERSPEECH 2015 , September 6-10, 2015, Dresden, Germany. Dresden, ALLEMAGNE. http:\/\/www.eurecom.fr\/publication\/4573 Zhizheng Wu, Tomi Kinnunen, Nicholas Evans, Junichi Yamagishi, Cemal Hanilci, Md Sahidullah, and Aleksandr Sizov. 2015. ASVspoof 2015: the first automatic speaker verification spoofing and countermeasures challenge. In INTERSPEECH 2015, Automatic Speaker Verification Spoofing and Countermeasures Challenge, colocated with INTERSPEECH 2015, September 6-10, 2015, Dresden, Germany. Dresden, ALLEMAGNE. http:\/\/www.eurecom.fr\/publication\/4573"},{"key":"e_1_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/2493432.2493435"},{"key":"e_1_2_1_57_1","volume-title":"Restructuring of deep neural network acoustic models with singular value decomposition","author":"Xue Jian","unstructured":"Jian Xue , Jinyu Li , and Yifan Gong . 2013. Restructuring of deep neural network acoustic models with singular value decomposition .. In INTERSPEECH, Fr\u00e9d\u00e9ric Bimbot, Christophe Cerisara, C\u00e9cile Fougeron, Guillaume Gravier, Lori Lamel, Fran\u00e7ois Pellegrino, and Pascal Perrier (Eds.). ISCA , 2365--2369. Jian Xue, Jinyu Li, and Yifan Gong. 2013. Restructuring of deep neural network acoustic models with singular value decomposition.. In INTERSPEECH, Fr\u00e9d\u00e9ric Bimbot, Christophe Cerisara, C\u00e9cile Fougeron, Guillaume Gravier, Lori Lamel, Fran\u00e7ois Pellegrino, and Pascal Perrier (Eds.). ISCA, 2365--2369."},{"key":"e_1_2_1_58_1","volume-title":"How transferable are features in deep neural networks? NIPS-14","author":"Yosinski Jason","year":"2014","unstructured":"Jason Yosinski , Jeff Clune , Yoshua Bengio , and Hod Lipson . 2014. How transferable are features in deep neural networks? NIPS-14 ( 2014 ). http:\/\/arxiv.org\/abs\/1411.1792 Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson. 2014. How transferable are features in deep neural networks? NIPS-14 (2014). http:\/\/arxiv.org\/abs\/1411.1792"}],"container-title":["Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3131895","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3131895","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:30:33Z","timestamp":1750217433000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3131895"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,9,11]]},"references-count":58,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2017,9,11]]}},"alternative-id":["10.1145\/3131895"],"URL":"https:\/\/doi.org\/10.1145\/3131895","relation":{},"ISSN":["2474-9567"],"issn-type":[{"value":"2474-9567","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,9,11]]},"assertion":[{"value":"2016-11-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-06-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-09-11","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}