{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:15:03Z","timestamp":1750306503845,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":17,"publisher":"ACM","license":[{"start":{"date-parts":[[2015,10,30]],"date-time":"2015-10-30T00:00:00Z","timestamp":1446163200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2015,10,30]]},"DOI":"10.1145\/2814815.2814816","type":"proceedings-article","created":{"date-parts":[[2016,10,25]],"date-time":"2016-10-25T12:46:35Z","timestamp":1477399595000},"page":"19-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Insights into Audio-Based Multimedia Event Classification with Neural Networks"],"prefix":"10.1145","author":[{"given":"Mirco","family":"Ravanelli","sequence":"first","affiliation":[{"name":"Fondazione Bruno Kessler, Trento, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Benjamin","family":"Elizalde","sequence":"additional","affiliation":[{"name":"International Computer Science Institute, Berkeley, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Julia","family":"Bernd","sequence":"additional","affiliation":[{"name":"International Computer Science Institute, Berkeley, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Gerald","family":"Friedland","sequence":"additional","affiliation":[{"name":"International Computer Science Institute, Berkeley, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2015,10,30]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2671188.2749396"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/72.286885"},{"key":"e_1_3_2_1_4_1","volume-title":"Proceedings of TRECVID 2012","author":"Cheng H.","year":"2012","unstructured":"H. Cheng , J. Liu , S. Ali , O. Javed , Q. Yu , A. Tamrakar , A. Divakaran , H. S. Sawhney , R. Manmatha , J. Allan , A. Hauptmann , M. Shah , S. Bhattacharya , A. Dehghan , G. Friedland , B. M. Elizalde , T. Darrell , M. Witbrock , and J. Curtis . SRI-Sarnoff AURORA system at TRECVID 2012: Multimedia event detection and recounting . In Proceedings of TRECVID 2012 . NIST, USA, 2012 . H. Cheng, J. Liu, S. Ali, O. Javed, Q. Yu, A. Tamrakar, A. Divakaran, H. S. Sawhney, R. Manmatha, J. Allan, A. Hauptmann, M. Shah, S. Bhattacharya, A. Dehghan, G. Friedland, B. M. Elizalde, T. Darrell, M. Witbrock, and J. Curtis. SRI-Sarnoff AURORA system at TRECVID 2012: Multimedia event detection and recounting. In Proceedings of TRECVID 2012. NIST, USA, 2012."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASL.2011.2134090"},{"key":"e_1_3_2_1_6_1","volume-title":"Proceedings of SLAM@INTERSPEECH","author":"Elizalde B.","year":"2014","unstructured":"B. Elizalde , M. Ravanelli , and G. Friedland . Audio-concept features and hidden Markov models for multimedia event detection . In Proceedings of SLAM@INTERSPEECH , 2014 . B. Elizalde, M. Ravanelli, and G. Friedland. Audio-concept features and hidden Markov models for multimedia event detection. In Proceedings of SLAM@INTERSPEECH, 2014."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/WASPAA.2013.6701819"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1162\/neco.2006.18.7.1527"},{"key":"e_1_3_2_1_9_1","volume-title":"Proceedings of TRECVID 2013","author":"Lan Z.","year":"2013","unstructured":"Z. Lan , L. Jiang , S.-I. Yu , C. Gao , S. Rawat , Y. Cai , S. Xu , H. Shen , X. Li , Y. Wang , W. Sze , Y. Yan , Z. Ma , N. Ballas , D. Meng , W. Tong , Y. Yang , S. Burger , F. Metze , R. Singh , B. Raj , R. Stern , T. Mitamura , E. Nyberg , and A. Hauptmann . Informedia @ TRECVID 2013 . In Proceedings of TRECVID 2013 . NIST, USA, 2013 . Z. Lan, L. Jiang, S.-I. Yu, C. Gao, S. Rawat, Y. Cai, S. Xu, H. Shen, X. Li, Y. Wang, W. Sze, Y. Yan, Z. Ma, N. Ballas, D. Meng, W. Tong, Y. Yang, S. Burger, F. Metze, R. Singh, B. Raj, R. Stern, T. Mitamura, E. Nyberg, and A. Hauptmann. Informedia @ TRECVID 2013. In Proceedings of TRECVID 2013. NIST, USA, 2013."},{"key":"e_1_3_2_1_10_1","volume-title":"NIPS Workshop on Deep Learning for Speech Recognition and Related Applications","author":"Mohamed A.","year":"2009","unstructured":"A. Mohamed , G. E. Dahl , and G. E. Hinton . Deep belief networks for phone recognition . In NIPS Workshop on Deep Learning for Speech Recognition and Related Applications , Vancouver,Canada , 2009 . A. Mohamed, G. E. Dahl, and G. E. Hinton. Deep belief networks for phone recognition. In NIPS Workshop on Deep Learning for Speech Recognition and Related Applications, Vancouver,Canada, 2009."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASL.2011.2116010"},{"key":"e_1_3_2_1_12_1","volume-title":"Advances in Neural Information Processing Systems","author":"Nair V.","year":"2009","unstructured":"V. Nair and G. E. Hinton . 3-D object recognition with deep belief nets . In Advances in Neural Information Processing Systems , 2009 . V. Nair and G. E. Hinton. 3-D object recognition with deep belief nets. In Advances in Neural Information Processing Systems, 2009."},{"key":"e_1_3_2_1_13_1","volume-title":"BBN VISER TRECVID 2012 multimedia event detection and multimedia event recounting systems. In Proceedings of TRECVID 2012. NIST, USA","author":"Natarajan P.","year":"2012","unstructured":"P. Natarajan , P. Natarajan , S. Wu , X. Zhuang , A. Vazquez Reina , S. N. Vitaladevuni , K. Tsourides , C. Andersen , R. Prasad , G. Ye , D. Liu , S.-F. Chang , I. Saleemi , M. Shah , Y. Ng , B. White , L. Davis , A. Gupta , and I. Haritaoglu . BBN VISER TRECVID 2012 multimedia event detection and multimedia event recounting systems. In Proceedings of TRECVID 2012. NIST, USA , 2012 . P. Natarajan, P. Natarajan, S. Wu, X. Zhuang, A. Vazquez Reina, S. N. Vitaladevuni, K. Tsourides, C. Andersen, R. Prasad, G. Ye, D. Liu, S.-F. Chang, I. Saleemi, M. Shah, Y. Ng, B. White, L. Davis, A. Gupta, and I. Haritaoglu. BBN VISER TRECVID 2012 multimedia event detection and multimedia event recounting systems. In Proceedings of TRECVID 2012. NIST, USA, 2012."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/5.18626"},{"key":"e_1_3_2_1_15_1","volume-title":"Proceedings of EUSIPCO","author":"Ravanelli M.","year":"2014","unstructured":"M. Ravanelli , B. Elizalde , K. Ni , and G. Friedland . Audio concept classification with hierarchical deep neural networks . In Proceedings of EUSIPCO , 2014 . M. Ravanelli, B. Elizalde, K. Ni, and G. Friedland. Audio concept classification with hierarchical deep neural networks. In Proceedings of EUSIPCO, 2014."},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijar.2008.11.006"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.5555\/1887176.1887235"},{"key":"e_1_3_2_1_18_1","volume-title":"Proceedings of INTERSPEECH","author":"Vinyals O.","year":"2013","unstructured":"O. Vinyals and N. Morgan . Deep vs. wide: Depth on a budget for robust speech recognition . In Proceedings of INTERSPEECH , 2013 . O. Vinyals and N. Morgan. Deep vs. wide: Depth on a budget for robust speech recognition. In Proceedings of INTERSPEECH, 2013."}],"event":{"name":"MM '15: ACM Multimedia Conference","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Brisbane Australia","acronym":"MM '15"},"container-title":["Proceedings of the 2015 Workshop on Community-Organized Multimodal Mining: Opportunities for Novel Solutions"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2814815.2814816","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2814815.2814816","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T05:48:53Z","timestamp":1750225733000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2814815.2814816"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2015,10,30]]},"references-count":17,"alternative-id":["10.1145\/2814815.2814816","10.1145\/2814815"],"URL":"https:\/\/doi.org\/10.1145\/2814815.2814816","relation":{},"subject":[],"published":{"date-parts":[[2015,10,30]]},"assertion":[{"value":"2015-10-30","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}