{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:35:27Z","timestamp":1750221327293,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":52,"publisher":"ACM","license":[{"start":{"date-parts":[[2017,10,19]],"date-time":"2017-10-19T00:00:00Z","timestamp":1508371200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2017,10,19]]},"DOI":"10.1145\/3123266.3123455","type":"proceedings-article","created":{"date-parts":[[2017,10,20]],"date-time":"2017-10-20T13:04:26Z","timestamp":1508504666000},"page":"279-287","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":27,"title":["A Dual-Network Progressive Approach to Weakly Supervised Object Detection"],"prefix":"10.1145","author":[{"given":"Xuanyi","family":"Dong","sequence":"first","affiliation":[{"name":"University of Technology Sydney, Sydney, Australia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Deyu","family":"Meng","sequence":"additional","affiliation":[{"name":"Xi'an Jiaotong University, Xi'an, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fan","family":"Ma","sequence":"additional","affiliation":[{"name":"Xi'an Jiaotong University, Xi'an, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yi","family":"Yang","sequence":"additional","affiliation":[{"name":"University of Technology Sydney, Sydney, Australia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2017,10,19]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"Maria-Florina Balcan Avrim Blum and Ke Yang. 2004. Co-training and expansion: Towards bridging theory and practice NIPS.   Maria-Florina Balcan Avrim Blum and Ke Yang. 2004. Co-training and expansion: Towards bridging theory and practice NIPS."},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"crossref","unstructured":"Loris Bazzani Alessandra Bergamo Dragomir Anguelov and Lorenzo Torresani. 2016. Self-taught object localization with deep networks WACV.  Loris Bazzani Alessandra Bergamo Dragomir Anguelov and Lorenzo Torresani. 2016. Self-taught object localization with deep networks WACV.","DOI":"10.1109\/WACV.2016.7477688"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1553374.1553380"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"crossref","unstructured":"Hakan Bilen Marco Pedersoli and Tinne Tuytelaars. 2014. Weakly supervised object detection with posterior regularization BMVC.  Hakan Bilen Marco Pedersoli and Tinne Tuytelaars. 2014. Weakly supervised object detection with posterior regularization BMVC.","DOI":"10.5244\/C.28.52"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"crossref","unstructured":"Hakan Bilen Marco Pedersoli and Tinne Tuytelaars. 2015. Weakly supervised object detection with convex clustering CVPR.  Hakan Bilen Marco Pedersoli and Tinne Tuytelaars. 2015. Weakly supervised object detection with convex clustering CVPR.","DOI":"10.1109\/CVPR.2015.7298711"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"crossref","unstructured":"H. Bilen and A. Vedaldi. 2016. Weakly supervised deep detection networks. In CVPR.  H. Bilen and A. Vedaldi. 2016. Weakly supervised deep detection networks. In CVPR.","DOI":"10.1109\/CVPR.2016.311"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/279943.279962"},{"key":"e_1_3_2_1_8_1","unstructured":"Ramazan Gokberk Cinbis Jakob Verbeek and Cordelia Schmid. 2014. Multi-fold MIL training for weakly supervised object localization CVPR.  Ramazan Gokberk Cinbis Jakob Verbeek and Cordelia Schmid. 2014. Multi-fold MIL training for weakly supervised object localization CVPR."},{"key":"e_1_3_2_1_9_1","unstructured":"Jifeng Dai Yi Li Kaiming He and Jian Sun. 2016. R-FCN: object detection via region-based fully convolutional networks NIPS.  Jifeng Dai Yi Li Kaiming He and Jian Sun. 2016. R-FCN: object detection via region-based fully convolutional networks NIPS."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"crossref","unstructured":"Thomas Deselaers Bogdan Alexe and Vittorio Ferrari. 2010. Localizing objects while learning their appearance ECCV.   Thomas Deselaers Bogdan Alexe and Vittorio Ferrari. 2010. Localizing objects while learning their appearance ECCV.","DOI":"10.1007\/978-3-642-15561-1_33"},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-012-0538-3"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"crossref","unstructured":"Ali Diba Vivek Sharma Ali Pazandeh Hamed Pirsiavash and Luc Van Gool. 2017. Weakly supervised cascaded convolutional networks. CVPR.  Ali Diba Vivek Sharma Ali Pazandeh Hamed Pirsiavash and Luc Van Gool. 2017. Weakly supervised cascaded convolutional networks. CVPR.","DOI":"10.1109\/CVPR.2017.545"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"crossref","unstructured":"Xuanyi Dong Junshi Huang Yi Yang and Shuicheng Yan. 2017. More is Less: A More Complicated Network with Less Inference Complexity CVPR.  Xuanyi Dong Junshi Huang Yi Yang and Shuicheng Yan. 2017. More is Less: A More Complicated Network with Less Inference Complexity CVPR.","DOI":"10.1109\/CVPR.2017.205"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-014-0733-5"},{"key":"e_1_3_2_1_15_1","unstructured":"Ma Fan Meng Deyu Xie Qi Zina Li and Xuanyi Dong. 2017. Self-paced Cotraining ICML.  Ma Fan Meng Deyu Xie Qi Zina Li and Xuanyi Dong. 2017. Self-paced Cotraining ICML."},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.169"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.81"},{"key":"e_1_3_2_1_18_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR.  Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2005.259"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/2647868.2654918"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"crossref","unstructured":"Zequn Jie Yunchao Wei Xiaojie Jin Jiashi Feng and Wei Liu. 2017. Deep self-taught learning for weakly supervised object localization CVPR.  Zequn Jie Yunchao Wei Xiaojie Jin Jiashi Feng and Wei Liu. 2017. Deep self-taught learning for weakly supervised object localization CVPR.","DOI":"10.1109\/CVPR.2017.457"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/2502081.2502159"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"crossref","unstructured":"Vadim Kantorov Maxime Oquab Minsu Cho and Ivan Laptev. 2016. ContextLocNet: Context-aware deep network models for weakly supervised localization ECCV.  Vadim Kantorov Maxime Oquab Minsu Cho and Ivan Laptev. 2016. ContextLocNet: Context-aware deep network models for weakly supervised localization ECCV.","DOI":"10.1007\/978-3-319-46454-1_22"},{"key":"e_1_3_2_1_24_1","unstructured":"Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. ImageNet classification with deep convolutional neural networks NIPS.   Alex Krizhevsky Ilya Sutskever and Geoffrey E Hinton. 2012. ImageNet classification with deep convolutional neural networks NIPS."},{"key":"e_1_3_2_1_25_1","unstructured":"M Pawan Kumar Benjamin Packer and Daphne Koller. 2010. Self-paced learning for latent variable models. In NIPS.   M Pawan Kumar Benjamin Packer and Daphne Koller. 2010. Self-paced learning for latent variable models. In NIPS."},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"crossref","unstructured":"Anat Levin Paul Viola and Yoav Freund. 2003. Unsupervised improvement of visual detectors using cotraining ICCV.   Anat Levin Paul Viola and Yoav Freund. 2003. Unsupervised improvement of visual detectors using cotraining ICCV.","DOI":"10.1109\/ICCV.2003.1238406"},{"key":"e_1_3_2_1_27_1","unstructured":"Dong Li Jia-Bin Huang Yali Li Shengjin Wang and Ming-Hsuan Yang. 2016. Weakly Supervised Object Localization with Progressive Domain Adaptation CVPR.  Dong Li Jia-Bin Huang Yali Li Shengjin Wang and Ming-Hsuan Yang. 2016. Weakly Supervised Object Localization with Progressive Domain Adaptation CVPR."},{"volume-title":"SSD: Single Shot MultiBox Detector. In ECCV.","year":"2016","author":"Liu Wei","key":"e_1_3_2_1_28_1"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2017.05.043"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2964284.2971479"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","unstructured":"Maxime Oquab L\u00e9on Bottou Ivan Laptev and Josef Sivic. 2015. Is object localization for free? Weakly-supervised learning with convolutional neural networks CVPR.  Maxime Oquab L\u00e9on Bottou Ivan Laptev and Josef Sivic. 2015. Is object localization for free? Weakly-supervised learning with convolutional neural networks CVPR.","DOI":"10.1109\/CVPR.2015.7298668"},{"volume-title":"et almbox","year":"2014","author":"Oquab Maxime","key":"e_1_3_2_1_32_1"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"crossref","unstructured":"Joseph Redmon Santosh Divvala Ross Girshick and Ali Farhadi. 2016. You only look once: Unified real-time object detection CVPR.  Joseph Redmon Santosh Divvala Ross Girshick and Ali Farhadi. 2016. You only look once: Unified real-time object detection CVPR.","DOI":"10.1109\/CVPR.2016.91"},{"key":"e_1_3_2_1_34_1","unstructured":"Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster R-CNN: Towards real-time object detection with region proposal networks NIPS.   Shaoqing Ren Kaiming He Ross Girshick and Jian Sun. 2015. Faster R-CNN: Towards real-time object detection with region proposal networks NIPS."},{"key":"e_1_3_2_1_35_1","unstructured":"Miaojing Shi and Vittorio Ferrari. 2016. Weakly supervised object localization using size estimates ECCV.  Miaojing Shi and Vittorio Ferrari. 2016. Weakly supervised object localization using size estimates ECCV."},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"crossref","unstructured":"Abhinav Shrivastava Abhinav Gupta and Ross Girshick. 2016. Training region-based object detectors with online hard example mining CVPR.  Abhinav Shrivastava Abhinav Gupta and Ross Girshick. 2016. Training region-based object detectors with online hard example mining CVPR.","DOI":"10.1109\/CVPR.2016.89"},{"key":"e_1_3_2_1_37_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2015. Very deep convolutional networks for large-scale image recognition ICLR.  Karen Simonyan and Andrew Zisserman. 2015. Very deep convolutional networks for large-scale image recognition ICLR."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33712-3_43"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.416"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126261"},{"key":"e_1_3_2_1_41_1","unstructured":"Hyun Oh Song Ross B Girshick Stefanie Jegelka Julien Mairal Zaid Harchaoui Trevor Darrell et almbox.. 2014 a. On learning to localize objects with minimal supervision ICML.   Hyun Oh Song Ross B Girshick Stefanie Jegelka Julien Mairal Zaid Harchaoui Trevor Darrell et almbox.. 2014 a. On learning to localize objects with minimal supervision ICML."},{"key":"e_1_3_2_1_42_1","unstructured":"Hyun Oh Song Yong Jae Lee Stefanie Jegelka and Trevor Darrell. 2014 b. Weakly-supervised discovery of visual pattern configurations NIPS.   Hyun Oh Song Yong Jae Lee Stefanie Jegelka and Trevor Darrell. 2014 b. Weakly-supervised discovery of visual pattern configurations NIPS."},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.308"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"crossref","unstructured":"Christian Szegedy Vincent Vanhoucke Sergey Ioffe Jon Shlens and Zbigniew Wojna. 2016. Rethinking the inception architecture for computer vision CVPR.  Christian Szegedy Vincent Vanhoucke Sergey Ioffe Jon Shlens and Zbigniew Wojna. 2016. Rethinking the inception architecture for computer vision CVPR.","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0620-5"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"crossref","unstructured":"Chong Wang Weiqiang Ren Kaiqi Huang and Tieniu Tan. 2014. Weakly supervised object localization with latent category learning ECCV.  Chong Wang Weiqiang Ren Kaiqi Huang and Tieniu Tan. 2014. Weakly supervised object localization with latent category learning ECCV.","DOI":"10.1007\/978-3-319-10599-4_28"},{"key":"e_1_3_2_1_47_1","unstructured":"Wei Wang and Zhi-Hua Zhou. 2010. A new analysis of co-training. In ICML.   Wei Wang and Zhi-Hua Zhou. 2010. A new analysis of co-training. In ICML."},{"key":"e_1_3_2_1_48_1","unstructured":"Saining Xie Ross Girshick Piotr Doll\u00e1r Zhuowen Tu and Kaiming He. 2017. Aggregated Residual Transformations for Deep Neural Networks. CVPR.  Saining Xie Ross Girshick Piotr Doll\u00e1r Zhuowen Tu and Kaiming He. 2017. Aggregated Residual Transformations for Deep Neural Networks. CVPR."},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2012.2237023"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/2964284.2967274"},{"key":"e_1_3_2_1_51_1","unstructured":"Dingwen Zhang Deyu Meng Long Zhao and Junwei Han. 2016. Bridging saliency detection to weakly supervised object detection based on self-paced curriculum learning. In IJCAI.   Dingwen Zhang Deyu Meng Long Zhao and Junwei Han. 2016. Bridging saliency detection to weakly supervised object detection based on self-paced curriculum learning. In IJCAI."},{"key":"e_1_3_2_1_52_1","doi-asserted-by":"crossref","unstructured":"C Lawrence Zitnick and Piotr Doll\u00e1r. 2014. Edge boxes: locating object proposals from edges. ECCV.  C Lawrence Zitnick and Piotr Doll\u00e1r. 2014. Edge boxes: locating object proposals from edges. ECCV.","DOI":"10.1007\/978-3-319-10602-1_26"}],"event":{"name":"MM '17: ACM Multimedia Conference","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Mountain View California USA","acronym":"MM '17"},"container-title":["Proceedings of the 25th ACM international conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3123266.3123455","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3123266.3123455","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:14:04Z","timestamp":1750212844000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3123266.3123455"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,10,19]]},"references-count":52,"alternative-id":["10.1145\/3123266.3123455","10.1145\/3123266"],"URL":"https:\/\/doi.org\/10.1145\/3123266.3123455","relation":{},"subject":[],"published":{"date-parts":[[2017,10,19]]},"assertion":[{"value":"2017-10-19","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}