{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,30]],"date-time":"2026-07-30T14:23:29Z","timestamp":1785421409591,"version":"3.56.0"},"publisher-location":"New York, NY, USA","reference-count":42,"publisher":"ACM","license":[{"start":{"date-parts":[[2019,10,15]],"date-time":"2019-10-15T00:00:00Z","timestamp":1571097600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2019,10,15]]},"DOI":"10.1145\/3343031.3350849","type":"proceedings-article","created":{"date-parts":[[2019,10,21]],"date-time":"2019-10-21T16:32:26Z","timestamp":1571675546000},"page":"2414-2422","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":73,"title":["Lossy Intermediate Deep Learning Feature Compression and Evaluation"],"prefix":"10.1145","author":[{"given":"Zhuo","family":"Chen","sequence":"first","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kui","family":"Fan","sequence":"additional","affiliation":[{"name":"Peking University, Shenzhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shiqi","family":"Wang","sequence":"additional","affiliation":[{"name":"City University of Hong Kong, Hong Kong, Hong Kong"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ling-Yu","family":"Duan","sequence":"additional","affiliation":[{"name":"Peking University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Weisi","family":"Lin","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Alex","family":"Kot","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2019,10,15]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"Davide Del Testa","author":"Bojarski Mariusz","year":"2016","unstructured":"Mariusz Bojarski , Davide Del Testa , Daniel Dworakowski, Bernhard Firner , Beat Flepp, Prasoon Goyal, Lawrence D Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, et almbox. 2016 . End to end learning for self-driving cars. arXiv preprint arXiv:1604.07316 (2016). Mariusz Bojarski, Davide Del Testa, Daniel Dworakowski, Bernhard Firner, Beat Flepp, Prasoon Goyal, Lawrence D Jackel, Mathew Monfort, Urs Muller, Jiakai Zhang, et almbox. 2016. End to end learning for self-driving cars. arXiv preprint arXiv:1604.07316 (2016)."},{"key":"e_1_3_2_1_2_1","volume-title":"JVET-L1001-v9","author":"Bross Benjamin","year":"2018","unstructured":"Benjamin Bross , Jianle Chen , and Shan Liu . 2018. Working Draft 3 of Versatile Video Coding. Joint Video Exploration Team of ITU?T SG16 WP3 and ISO\/IEC JTC1\/SC29\/WG11 , JVET-L1001-v9 ( 2018 ). Benjamin Bross, Jianle Chen, and Shan Liu. 2018. Working Draft 3 of Versatile Video Coding. Joint Video Exploration Team of ITU?T SG16 WP3 and ISO\/IEC JTC1\/SC29\/WG11, JVET-L1001-v9 (2018)."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.sigpro.2016.05.021"},{"key":"e_1_3_2_1_4_1","volume-title":"An Implementation of Faster RCNN with Study for Region Sampling. arXiv preprint arXiv:1702.02138","author":"Chen Xinlei","year":"2017","unstructured":"Xinlei Chen and Abhinav Gupta . 2017. An Implementation of Faster RCNN with Study for Region Sampling. arXiv preprint arXiv:1702.02138 ( 2017 ). Xinlei Chen and Abhinav Gupta. 2017. An Implementation of Faster RCNN with Study for Region Sampling. arXiv preprint arXiv:1702.02138 (2017)."},{"key":"e_1_3_2_1_5_1","volume-title":"Intermediate deep feature compression: the next battlefield of intelligent sensing. arXiv preprint arXiv:1809.06196","author":"Chen Zhuo","year":"2018","unstructured":"Zhuo Chen , Weisi Lin , Shiqi Wang , Lingyu Duan , and Alex C Kot . 2018. Intermediate deep feature compression: the next battlefield of intelligent sensing. arXiv preprint arXiv:1809.06196 ( 2018 ). Zhuo Chen, Weisi Lin, Shiqi Wang, Lingyu Duan, and Alex C Kot. 2018. Intermediate deep feature compression: the next battlefield of intelligent sensing. arXiv preprint arXiv:1809.06196 (2018)."},{"key":"e_1_3_2_1_6_1","volume-title":"Deep feature compression for collaborative object detection. arXiv preprint arXiv:1802.03931","author":"Choi Hyomin","year":"2018","unstructured":"Hyomin Choi and Ivan V Bajic . 2018a. Deep feature compression for collaborative object detection. arXiv preprint arXiv:1802.03931 ( 2018 ). Hyomin Choi and Ivan V Bajic. 2018a. Deep feature compression for collaborative object detection. arXiv preprint arXiv:1802.03931 (2018)."},{"key":"e_1_3_2_1_7_1","volume-title":"Near-Lossless Deep Feature Compression for Collaborative Intelligence. arXiv preprint arXiv:1804.09963","author":"Choi Hyomin","year":"2018","unstructured":"Hyomin Choi and Ivan V Bajic . 2018b. Near-Lossless Deep Feature Compression for Collaborative Intelligence. arXiv preprint arXiv:1804.09963 ( 2018 ). Hyomin Choi and Ivan V Bajic. 2018b. Near-Lossless Deep Feature Compression for Collaborative Intelligence. arXiv preprint arXiv:1804.09963 (2018)."},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/QoMEX.2016.7498955"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2015.2500034"},{"key":"e_1_3_2_1_10_1","volume-title":"Alex Chichung Kot, and Wen Gao","author":"Duan Ling-Yu","year":"2017","unstructured":"Ling-Yu Duan , Vijay Chandrasekhar , Shiqi Wang , Yihang Lou , Jie Lin , Yan Bai , Tiejun Huang , Alex Chichung Kot, and Wen Gao . 2017 . Compact Descriptors for Video Analysis: the Emerging MPEG Standard . arXiv preprint arXiv:1704.08141 (2017). Ling-Yu Duan, Vijay Chandrasekhar, Shiqi Wang, Yihang Lou, Jie Lin, Yan Bai, Tiejun Huang, Alex Chichung Kot, and Wen Gao. 2017. Compact Descriptors for Video Analysis: the Emerging MPEG Standard. arXiv preprint arXiv:1704.08141 (2017)."},{"key":"e_1_3_2_1_11_1","unstructured":"M. Everingham L. Van Gool C. K. I. Williams J. Winn and A. Zisserman. [n. d.]. The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results. http:\/\/www.pascal-network.org\/challenges\/VOC\/voc2007\/workshop\/index.html.  M. Everingham L. Van Gool C. K. I. Williams J. Winn and A. Zisserman. [n. d.]. The PASCAL Visual Object Classes Challenge 2007 (VOC2007) Results. http:\/\/www.pascal-network.org\/challenges\/VOC\/voc2007\/workshop\/index.html."},{"key":"e_1_3_2_1_12_1","volume-title":"Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach.","author":"Fukui Akira","year":"2016","unstructured":"Akira Fukui , Dong Huk Park , Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach. 2016 . Multimodal compact bilinear pooling for visual question answering and visual grounding. arXiv preprint arXiv:1606.01847(2016). Akira Fukui, Dong Huk Park, Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach. 2016. Multimodal compact bilinear pooling for visual question answering and visual grounding. arXiv preprint arXiv:1606.01847(2016)."},{"key":"e_1_3_2_1_13_1","volume-title":"Fast r-cnn. arXiv preprint arXiv:1504.08083","author":"Girshick Ross","year":"2015","unstructured":"Ross Girshick . 2015. Fast r-cnn. arXiv preprint arXiv:1504.08083 ( 2015 ). Ross Girshick. 2015. Fast r-cnn. arXiv preprint arXiv:1504.08083 (2015)."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2015.2437384"},{"key":"e_1_3_2_1_15_1","volume-title":"Stack-captioning: Coarse-to-fine learning for image captioning. arXiv preprint arXiv:1709.03376","author":"Gu Jiuxiang","year":"2017","unstructured":"Jiuxiang Gu , Jianfei Cai , Gang Wang , and Tsuhan Chen . 2017 . Stack-captioning: Coarse-to-fine learning for image captioning. arXiv preprint arXiv:1709.03376 (2017). Jiuxiang Gu, Jianfei Cai, Gang Wang, and Tsuhan Chen. 2017. Stack-captioning: Coarse-to-fine learning for image captioning. arXiv preprint arXiv:1709.03376 (2017)."},{"key":"e_1_3_2_1_16_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. [n. d.]. Deep Residual Learning for Image Recognition. https:\/\/github.com\/KaimingHe\/deep-residual-networks.  Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. [n. d.]. Deep Residual Learning for Image Recognition. https:\/\/github.com\/KaimingHe\/deep-residual-networks."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_1_18_1","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition","volume":"1","author":"Huang Gao","unstructured":"Gao Huang , Zhuang Liu , Kilian Q Weinberger , and Laurens van der Maaten. 2017. Densely connected convolutional networks . In Proceedings of the IEEE conference on computer vision and pattern recognition , Vol. 1 . 3. Gao Huang, Zhuang Liu, Kilian Q Weinberger, and Laurens van der Maaten. 2017. Densely connected convolutional networks. In Proceedings of the IEEE conference on computer vision and pattern recognition, Vol. 1. 3."},{"key":"e_1_3_2_1_19_1","unstructured":"Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems. 1097--1105.  Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. Imagenet classification with deep convolutional neural networks. In Advances in neural information processing systems. 1097--1105."},{"key":"e_1_3_2_1_20_1","volume-title":"Face Recognition in Low Quality Images: A Survey. arXiv preprint arXiv:1805.11519","author":"Li Pei","year":"2018","unstructured":"Pei Li , Loreto Prieto , Domingo Mery , and Patrick Flynn . 2018. Face Recognition in Low Quality Images: A Survey. arXiv preprint arXiv:1805.11519 ( 2018 ). Pei Li, Loreto Prieto, Domingo Mery, and Patrick Flynn. 2018. Face Recognition in Low Quality Images: A Survey. arXiv preprint arXiv:1805.11519 (2018)."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2017.2713410"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.238"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2019.2902112"},{"key":"e_1_3_2_1_24_1","unstructured":"Jiasen Lu Jianwei Yang Dhruv Batra and Devi Parikh. 2016. Hierarchical question-image co-attention for visual question answering. In Advances In Neural Information Processing Systems. 289--297.  Jiasen Lu Jianwei Yang Dhruv Batra and Devi Parikh. 2016. Hierarchical question-image co-attention for visual question answering. In Advances In Neural Information Processing Systems. 289--297."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.257"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICMLA.2016.0054"},{"key":"e_1_3_2_1_27_1","volume-title":"YOLOv3: An Incremental Improvement. arXiv preprint arXiv:1804.02767","author":"Redmon Joseph","year":"2018","unstructured":"Joseph Redmon and Ali Farhadi . 2018. YOLOv3: An Incremental Improvement. arXiv preprint arXiv:1804.02767 ( 2018 ). Joseph Redmon and Ali Farhadi. 2018. YOLOv3: An Incremental Improvement. arXiv preprint arXiv:1804.02767 (2018)."},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMC.2016.2519340"},{"key":"e_1_3_2_1_29_1","unstructured":"Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems. 91--99.  Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems. 91--99."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_2_1_31_1","unstructured":"K. Simonyan and A. Zisserman. [n. d.]. ILSVRC-2014 model (VGG team) with 16 weight layers. https:\/\/gist.github.com\/ksimonyan\/211839e770f7b538e2d8.  K. Simonyan and A. Zisserman. [n. d.]. ILSVRC-2014 model (VGG team) with 16 weight layers. https:\/\/gist.github.com\/ksimonyan\/211839e770f7b538e2d8."},{"key":"e_1_3_2_1_32_1","volume-title":"Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 ( 2014 ). Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014)."},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2012.2221191"},{"key":"e_1_3_2_1_34_1","unstructured":"Yi Sun Yuheng Chen Xiaogang Wang and Xiaoou Tang. 2014. Deep learning face representation by joint identification-verification. In Advances in neural information processing systems. 1988--1996.  Yi Sun Yuheng Chen Xiaogang Wang and Xiaoou Tang. 2014. Deep learning face representation by joint identification-verification. In Advances in neural information processing systems. 1988--1996."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.220"},{"key":"e_1_3_2_1_36_1","volume-title":"Cider: Consensus-based image description evaluation. In CVPR.","author":"Vedantam Ramakrishna","year":"2015","unstructured":"Ramakrishna Vedantam , C Lawrence Zitnick , and Devi Parikh . 2015 . Cider: Consensus-based image description evaluation. In CVPR. Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015. Cider: Consensus-based image description evaluation. In CVPR."},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.357"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2017.2655449"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2003.815165"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.140"},{"key":"e_1_3_2_1_41_1","volume-title":"International Conference on Machine Learning. 2048--2057","author":"Xu Kelvin","year":"2015","unstructured":"Kelvin Xu , Jimmy Ba , Ryan Kiros , Kyunghyun Cho , Aaron Courville , Ruslan Salakhudinov , Rich Zemel , and Yoshua Bengio . 2015 . Show, attend and tell: Neural image caption generation with visual attention . In International Conference on Machine Learning. 2048--2057 . Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio. 2015. Show, attend and tell: Neural image caption generation with visual attention. In International Conference on Machine Learning. 2048--2057."},{"key":"e_1_3_2_1_42_1","volume-title":"The Framework and Test Condition for Lossy Compression of Deep Feature Maps. Audio Video Coding Standard (AVS) document AI M1061","author":"Zhuo Chen","year":"2018","unstructured":"Chen Zhuo , Fan Kui , Lin Weisi , Duan Lingyu , Kot Alex , C., and Huang Tiejun . 2018. The Framework and Test Condition for Lossy Compression of Deep Feature Maps. Audio Video Coding Standard (AVS) document AI M1061 ( 2018 ). Chen Zhuo, Fan Kui, Lin Weisi, Duan Lingyu, Kot Alex, C., and Huang Tiejun. 2018. The Framework and Test Condition for Lossy Compression of Deep Feature Maps. Audio Video Coding Standard (AVS) document AI M1061 (2018)."}],"event":{"name":"MM '19: The 27th ACM International Conference on Multimedia","location":"Nice France","acronym":"MM '19","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 27th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3343031.3350849","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3343031.3350849","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T23:13:25Z","timestamp":1750202005000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3343031.3350849"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,10,15]]},"references-count":42,"alternative-id":["10.1145\/3343031.3350849","10.1145\/3343031"],"URL":"https:\/\/doi.org\/10.1145\/3343031.3350849","relation":{},"subject":[],"published":{"date-parts":[[2019,10,15]]},"assertion":[{"value":"2019-10-15","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}