{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,21]],"date-time":"2026-02-21T19:13:09Z","timestamp":1771701189696,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":27,"publisher":"ACM","license":[{"start":{"date-parts":[[2019,9,9]],"date-time":"2019-09-09T00:00:00Z","timestamp":1567987200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2019,9,9]]},"DOI":"10.1145\/3349801.3349821","type":"proceedings-article","created":{"date-parts":[[2019,9,25]],"date-time":"2019-09-25T12:58:02Z","timestamp":1569416282000},"page":"1-6","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Accurate Single-Stream Action Detection in Real-Time"],"prefix":"10.1145","author":[{"given":"Yu","family":"Liu","sequence":"first","affiliation":[{"name":"Laboratoire ImViA, Univ. Bourgogne, Franche-Comt\u00e9, Dijon, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fan","family":"Yang","sequence":"additional","affiliation":[{"name":"Laboratoire ImViA, Univ. Bourgogne, Franche-Comt\u00e9, Dijon, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dominique","family":"Ginhac","sequence":"additional","affiliation":[{"name":"Laboratoire ImViA, Univ. Bourgogne, Franche-Comt\u00e9, Dijon, France"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2019,9,9]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00815"},{"key":"e_1_3_2_1_2_1","unstructured":"Jifeng Dai Yi Li Kaiming He and Jian Sun. 2016. R-FCN: Object detection via region-based fully convolutional networks. In Advances in Neural Information Processing Systems.  Jifeng Dai Yi Li Kaiming He and Jian Sun. 2016. R-FCN: Object detection via region-based fully convolutional networks. In Advances in Neural Information Processing Systems."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.316"},{"key":"e_1_3_2_1_4_1","unstructured":"Alaaeldin El-Nouby and Graham W Taylor. 2018. Real-time end-to-end action detection with two-stream networks. arXiv preprint arXiv:1802.08362.  Alaaeldin El-Nouby and Graham W Taylor. 2018. Real-time end-to-end action detection with two-stream networks. arXiv preprint arXiv:1802.08362."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.213"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.330"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.477"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_3_2_1_9_1","volume-title":"Prajit Ramachandran, Mohammad Babaeizadeh, Honghui Shi, Jianan Li, Shuicheng Yan, and Thomas S Huang.","author":"Han Wei","year":"2016","unstructured":"Wei Han , Pooya Khorrami , Tom Le Paine , Prajit Ramachandran, Mohammad Babaeizadeh, Honghui Shi, Jianan Li, Shuicheng Yan, and Thomas S Huang. 2016 . Seq-NMS for video object detection. arXiv preprint arXiv:1602.08465. Wei Han, Pooya Khorrami, Tom Le Paine, Prajit Ramachandran, Mohammad Babaeizadeh, Honghui Shi, Jianan Li, Shuicheng Yan, and Thomas S Huang. 2016. Seq-NMS for video object detection. arXiv preprint arXiv:1602.08465."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_1_11_1","unstructured":"Max Jaderberg Karen Simonyan Andrew Zisserman etal 2015. Spatial transformer networks. In Advances in neural information processing systems.  Max Jaderberg Karen Simonyan Andrew Zisserman et al. 2015. Spatial transformer networks. In Advances in neural information processing systems."},{"key":"e_1_3_2_1_12_1","volume-title":"IEEE International Conference on Computer Vision.","author":"Kalogeiton Vicky","year":"2017","unstructured":"Vicky Kalogeiton , Philippe Weinzaepfel , Vittorio Ferrari , and Cordelia Schmid . 2017 . Action tubelet detector for spatiotemporal action localization . In IEEE International Conference on Computer Vision. Vicky Kalogeiton, Philippe Weinzaepfel, Vittorio Ferrari, and Cordelia Schmid. 2017. Action tubelet detector for spatiotemporal action localization. In IEEE International Conference on Computer Vision."},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2017.2736553"},{"key":"e_1_3_2_1_14_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition.","author":"Liu Mason","year":"2018","unstructured":"Mason Liu and Menglong Zhu . 2018 . Mobile video object detection with temporally-aware feature maps . In IEEE Conference on Computer Vision and Pattern Recognition. Mason Liu and Menglong Zhu. 2018. Mobile video object detection with temporally-aware feature maps. In IEEE Conference on Computer Vision and Pattern Recognition."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"e_1_3_2_1_16_1","volume-title":"IEEE Winter Conference on Applications of Computer Vision.","author":"Park Eunbyung","year":"2016","unstructured":"Eunbyung Park , Xufeng Han , Tamara L Berg , and Alexander C Berg . 2016 . Combining multiple sources of knowledge in deep cnns for action recognition . In IEEE Winter Conference on Applications of Computer Vision. Eunbyung Park, Xufeng Han, Tamara L Berg, and Alexander C Berg. 2016. Combining multiple sources of knowledge in deep cnns for action recognition. In IEEE Winter Conference on Applications of Computer Vision."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46493-0_45"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.91"},{"key":"e_1_3_2_1_19_1","unstructured":"Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster R-CNN: Towards real-time object detection with region proposal networks. In Advances in Neural Information Processing Systems.  Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster R-CNN: Towards real-time object detection with region proposal networks. In Advances in Neural Information Processing Systems."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"crossref","unstructured":"Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein Alexander C. Berg and Li Fei-Fei. 2015. ImageNet Large Scale Visual Recognition Challenge. In International Journal of Computer Vision.  Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein Alexander C. Berg and Li Fei-Fei. 2015. ImageNet Large Scale Visual Recognition Challenge. In International Journal of Computer Vision.","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.5244\/C.30.58"},{"key":"e_1_3_2_1_22_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Two-stream convolutional networks for action recognition in videos. In Advances in Neural Information Processing Systems.  Karen Simonyan and Andrew Zisserman. 2014. Two-stream convolutional networks for action recognition in videos. In Advances in Neural Information Processing Systems."},{"key":"e_1_3_2_1_23_1","volume-title":"Proceedings of the European Conference on Computer Vision.","author":"Singh Gurkirt","year":"2018","unstructured":"Gurkirt Singh , Suman Saha , and Fabio Cuzzolin . 2018 . Predicting Action Tubes . In Proceedings of the European Conference on Computer Vision. Gurkirt Singh, Suman Saha, and Fabio Cuzzolin. 2018. Predicting Action Tubes. In Proceedings of the European Conference on Computer Vision."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.393"},{"key":"e_1_3_2_1_25_1","volume-title":"Amir Roshan Zamir, and Mubarak Shah","author":"Soomro Khurram","year":"2012","unstructured":"Khurram Soomro , Amir Roshan Zamir, and Mubarak Shah . 2012 . UCF101: A dataset of 101 human actions classes from videos in the wild. CRCV-TR- 12-01. Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah. 2012. UCF101: A dataset of 101 human actions classes from videos in the wild. CRCV-TR-12-01."},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00151"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.441"}],"event":{"name":"ICDSC 2019: 13th International Conference on Distributed Smart Cameras","location":"Trento Italy","acronym":"ICDSC 2019","sponsor":["University of Trento"]},"container-title":["Proceedings of the 13th International Conference on Distributed Smart Cameras"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3349801.3349821","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3349801.3349821","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:23:20Z","timestamp":1750202600000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3349801.3349821"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,9,9]]},"references-count":27,"alternative-id":["10.1145\/3349801.3349821","10.1145\/3349801"],"URL":"https:\/\/doi.org\/10.1145\/3349801.3349821","relation":{},"subject":[],"published":{"date-parts":[[2019,9,9]]},"assertion":[{"value":"2019-09-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}