{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:15:57Z","timestamp":1750220157819,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":40,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3548242","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:42:35Z","timestamp":1665416555000},"page":"648-658","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Mixed Supervision for Instance Learning in Object Detection with Few-shot Annotation"],"prefix":"10.1145","author":[{"given":"Yi","family":"Zhong","sequence":"first","affiliation":[{"name":"Sun Yat-Sen University, Guangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chengyao","family":"Wang","sequence":"additional","affiliation":[{"name":"Sun Yat-Sen University, Guangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shiyong","family":"Li","sequence":"additional","affiliation":[{"name":"AI Application Research Center, Huawei, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhu","family":"Zhou","sequence":"additional","affiliation":[{"name":"AI Application Research Center, Huawei, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yaowei","family":"Wang","sequence":"additional","affiliation":[{"name":"Pengcheng Laboratory, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei-Shi","family":"Zheng","sequence":"additional","affiliation":[{"name":"Sun Yat-Sen University, Guangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.49"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58598-3_3"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.311"},{"key":"e_1_3_2_2_4_1","volume-title":"Spatial likelihood voting with self-knowledge distillation for weakly supervised object detection. Image and Vision Computing","author":"Chen Ze","year":"2021","unstructured":"Ze Chen , Zhihang Fu , Jianqiang Huang , Mingyuan Tao , Rongxin Jiang , Xiang Tian , Yaowu Chen , and Xian-Sheng Hua . 2021. Spatial likelihood voting with self-knowledge distillation for weakly supervised object detection. Image and Vision Computing ( 2021 ), 104314. Ze Chen, Zhihang Fu, Jianqiang Huang, Mingyuan Tao, Rongxin Jiang, Xiang Tian, Yaowu Chen, and Xian-Sheng Hua. 2021. Spatial likelihood voting with self-knowledge distillation for weakly supervised object detection. Image and Vision Computing (2021), 104314."},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2020.2987161"},{"key":"e_1_3_2_2_6_1","volume-title":"Weakly supervised object localization with multi-fold multiple instance learning","author":"Cinbis Ramazan Gokberk","year":"2016","unstructured":"Ramazan Gokberk Cinbis , Jakob Verbeek , and Cordelia Schmid . 2016. Weakly supervised object localization with multi-fold multiple instance learning . IEEE transactions on pattern analysis and machine intelligence, Vol. 39 , 1 ( 2016 ), 189--203. Ramazan Gokberk Cinbis, Jakob Verbeek, and Cordelia Schmid. 2016. Weakly supervised object localization with multi-fold multiple instance learning. IEEE transactions on pattern analysis and machine intelligence, Vol. 39, 1 (2016), 189--203."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_2_8_1","volume-title":"Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman.","author":"Everingham Mark","year":"2015","unstructured":"Mark Everingham , SM Ali Eslami , Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman. 2015 . The pascal visual object classes challenge: A retrospective. International journal of computer vision, Vol. 111 , 1 (2015), 98--136. Mark Everingham, SM Ali Eslami, Luc Van Gool, Christopher KI Williams, John Winn, and Andrew Zisserman. 2015. The pascal visual object classes challenge: A retrospective. International journal of computer vision, Vol. 111, 1 (2015), 98--136."},{"key":"e_1_3_2_2_9_1","volume-title":"EHSOD: CAM-Guided End-to-end Hybrid-Supervised Object Detection with Cascade Refinement. arXiv preprint arXiv:2002.07421","author":"Fang Linpu","year":"2020","unstructured":"Linpu Fang , Hang Xu , Zhili Liu , Sarah Parisot , and Zhenguo Li . 2020 . EHSOD: CAM-Guided End-to-end Hybrid-Supervised Object Detection with Cascade Refinement. arXiv preprint arXiv:2002.07421 (2020). Linpu Fang, Hang Xu, Zhili Liu, Sarah Parisot, and Zhenguo Li. 2020. EHSOD: CAM-Guided End-to-end Hybrid-Supervised Object Detection with Cascade Refinement. arXiv preprint arXiv:2002.07421 (2020)."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00960"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.169"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_14_1","volume-title":"Comprehensive Attention Self-Distillation for Weakly-Supervised Object Detection. arXiv preprint arXiv:2010.12023","author":"Huang Zeyi","year":"2020","unstructured":"Zeyi Huang , Yang Zou , Vijayakumar Bhagavatula , and Dong Huang . 2020. Comprehensive Attention Self-Distillation for Weakly-Supervised Object Detection. arXiv preprint arXiv:2010.12023 ( 2020 ). Zeyi Huang, Yang Zou, Vijayakumar Bhagavatula, and Dong Huang. 2020. Comprehensive Attention Self-Distillation for Weakly-Supervised Object Detection. arXiv preprint arXiv:2010.12023 (2020)."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2021.02.018"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46454-1_22"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3065386"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.324"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46448-0_2"},{"key":"e_1_3_2_2_21_1","volume-title":"Contrastive Proposal Extension with Sequential Network for Weakly Supervised Object Detection. arXiv preprint arXiv:2110.07511","author":"Lv Pei","year":"2021","unstructured":"Pei Lv , Suqi Hu , Tianran Hao , Haohan Ji , Lisha Cui , Haoyi Fan , Mingliang Xu , and Changsheng Xu. 2021. Contrastive Proposal Extension with Sequential Network for Weakly Supervised Object Detection. arXiv preprint arXiv:2110.07511 ( 2021 ). Pei Lv, Suqi Hu, Tianran Hao, Haohan Ji, Lisha Cui, Haoyi Fan, Mingliang Xu, and Changsheng Xu. 2021. Contrastive Proposal Extension with Sequential Network for Weakly Supervised Object Detection. arXiv preprint arXiv:2110.07511 (2021)."},{"key":"e_1_3_2_2_22_1","unstructured":"Oded Maron and Tom\u00e1s Lozano-P\u00e9rez. 1998. A framework for multiple-instance learning. In Advances in neural information processing systems. 570--576.  Oded Maron and Tom\u00e1s Lozano-P\u00e9rez. 1998. A framework for multiple-instance learning. In Advances in neural information processing systems. 570--576."},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.24963\/ijcai.2019\/125"},{"key":"e_1_3_2_2_24_1","volume-title":"Baod: Budget-aware object detection. arXiv preprint arXiv:1904.05443","author":"Pardo Alejandro","year":"2019","unstructured":"Alejandro Pardo , Mengmeng Xu , Ali Thabet , Pablo Arbelaez , and Bernard Ghanem . 2019 . Baod: Budget-aware object detection. arXiv preprint arXiv:1904.05443 (2019). Alejandro Pardo, Mengmeng Xu, Ali Thabet, Pablo Arbelaez, and Bernard Ghanem. 2019. Baod: Budget-aware object detection. arXiv preprint arXiv:1904.05443 (2019)."},{"key":"e_1_3_2_2_25_1","volume-title":"Yolov3: An incremental improvement. arXiv preprint arXiv:1804.02767","author":"Redmon Joseph","year":"2018","unstructured":"Joseph Redmon and Ali Farhadi . 2018. Yolov3: An incremental improvement. arXiv preprint arXiv:1804.02767 ( 2018 ). Joseph Redmon and Ali Farhadi. 2018. Yolov3: An incremental improvement. arXiv preprint arXiv:1804.02767 (2018)."},{"key":"e_1_3_2_2_26_1","unstructured":"Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015a. Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems. 91--99.  Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015a. Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems. 91--99."},{"key":"e_1_3_2_2_27_1","volume-title":"Weakly supervised large scale object localization with multiple instance learning and bag splitting","author":"Ren Weiqiang","year":"2015","unstructured":"Weiqiang Ren , Kaiqi Huang , Dacheng Tao , and Tieniu Tan . 2015b. Weakly supervised large scale object localization with multiple instance learning and bag splitting . IEEE transactions on pattern analysis and machine intelligence, Vol. 38 , 2 ( 2015 ), 405--416. Weiqiang Ren, Kaiqi Huang, Dacheng Tao, and Tieniu Tan. 2015b. Weakly supervised large scale object localization with multiple instance learning and bag splitting. IEEE transactions on pattern analysis and machine intelligence, Vol. 38, 2 (2015), 405--416."},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01061"},{"key":"e_1_3_2_2_29_1","volume-title":"A Unified Framework towards Omni-supervised Object Detection. arXiv preprint arXiv:2010.10804","author":"Ren Zhongzheng","year":"2020","unstructured":"Zhongzheng Ren , Zhiding Yu , Xiaodong Yang , Ming-Yu Liu , Alexander G Schwing , and Jan Kautz . 2020b. UFO$^2$ : A Unified Framework towards Omni-supervised Object Detection. arXiv preprint arXiv:2010.10804 ( 2020 ). Zhongzheng Ren, Zhiding Yu, Xiaodong Yang, Ming-Yu Liu, Alexander G Schwing, and Jan Kautz. 2020b. UFO$^2$: A Unified Framework towards Omni-supervised Object Detection. arXiv preprint arXiv:2010.10804 (2020)."},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.89"},{"key":"e_1_3_2_2_31_1","volume-title":"Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 ( 2014 ). Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014)."},{"key":"e_1_3_2_2_32_1","volume-title":"Pcl: Proposal cluster learning for weakly supervised object detection","author":"Tang Peng","year":"2018","unstructured":"Peng Tang , Xinggang Wang , Song Bai , Wei Shen , Xiang Bai , Wenyu Liu , and Alan Yuille . 2018 . Pcl: Proposal cluster learning for weakly supervised object detection . IEEE transactions on pattern analysis and machine intelligence, Vol. 42 , 1 (2018), 176--191. Peng Tang, Xinggang Wang, Song Bai, Wei Shen, Xiang Bai, Wenyu Liu, and Alan Yuille. 2018. Pcl: Proposal cluster learning for weakly supervised object detection. IEEE transactions on pattern analysis and machine intelligence, Vol. 42, 1 (2018), 176--191."},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.326"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2017.05.001"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00121"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00230"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2018.8546088"},{"key":"e_1_3_2_2_38_1","volume-title":"C-MIDN: Coupled Multiple Instance Detection Network With Segmentation Guidance for Weakly Supervised Object Detection. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV)","author":"Yan G.","year":"2019","unstructured":"G. Yan , B. Liu , Nan Guo , Xiaochun Ye , Fang Wan , Haihang You , and Dongrui Fan . 2019 . C-MIDN: Coupled Multiple Instance Detection Network With Segmentation Guidance for Weakly Supervised Object Detection. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV) (2019), 9833--9842. G. Yan, B. Liu, Nan Guo, Xiaochun Ye, Fang Wan, Haihang You, and Dongrui Fan. 2019. C-MIDN: Coupled Multiple Instance Detection Network With Segmentation Guidance for Weakly Supervised Object Detection. 2019 IEEE\/CVF International Conference on Computer Vision (ICCV) (2019), 9833--9842."},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00838"},{"key":"e_1_3_2_2_40_1","volume-title":"European conference on computer vision. 391--405","author":"Lawrence Zitnick C","year":"2014","unstructured":"C Lawrence Zitnick and Piotr Doll\u00e1r . 2014 . Edge boxes: Locating object proposals from edges . In European conference on computer vision. 391--405 . C Lawrence Zitnick and Piotr Doll\u00e1r. 2014. Edge boxes: Locating object proposals from edges. In European conference on computer vision. 391--405."}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Lisboa Portugal","acronym":"MM '22"},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548242","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3548242","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:21Z","timestamp":1750186821000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548242"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":40,"alternative-id":["10.1145\/3503161.3548242","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3548242","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}