{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,12]],"date-time":"2025-12-12T13:37:09Z","timestamp":1765546629655,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":62,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413812","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T12:26:18Z","timestamp":1602505578000},"page":"3302-3310","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Panoptic Image Annotation with a Collaborative Assistant"],"prefix":"10.1145","author":[{"given":"Jasper R.R.","family":"Uijlings","sequence":"first","affiliation":[{"name":"Google Research, Zurich, Switzerland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mykhaylo","family":"Andriluka","sequence":"additional","affiliation":[{"name":"Google Research, Zurich, Switzerland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Vittorio","family":"Ferrari","sequence":"additional","affiliation":[{"name":"Google Research, Zurich, Switzerland"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","unstructured":"[n.d.]. COCO panoptic dataset. http:\/\/cocodataset.org\/index.htm#panoptic-2018.  [n.d.]. COCO panoptic dataset. http:\/\/cocodataset.org\/index.htm#panoptic-2018."},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"crossref","unstructured":"D. Acuna H. Ling A. Kar and S. Fidler. 2018. Efficient Interactive Annotation of Segmentation Datasets with Polygon-RNN+. In CVPR.  D. Acuna H. Ling A. Kar and S. Fidler. 2018. Efficient Interactive Annotation of Segmentation Datasets with Polygon-RNN+. In CVPR.","DOI":"10.1109\/CVPR.2018.00096"},{"key":"e_1_3_2_2_3_1","volume-title":"Jasper RR Uijlings, and Vittorio Ferrari","author":"Agustsson Eirikur","year":"2019","unstructured":"Eirikur Agustsson , Jasper RR Uijlings, and Vittorio Ferrari . 2019 . Interactive Full Image Segmentation by Considering All Regions Jointly. In CVPR. Eirikur Agustsson, Jasper RR Uijlings, and Vittorio Ferrari. 2019. Interactive Full Image Segmentation by Considering All Regions Jointly. In CVPR."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3240508.3241916"},{"key":"e_1_3_2_2_5_1","volume-title":"Geodesic matting: A framework for fast interactive image and video segmentation and matting. IJCV","author":"Bai Xue","year":"2009","unstructured":"Xue Bai and Guillermo Sapiro . 2009. Geodesic matting: A framework for fast interactive image and video segmentation and matting. IJCV ( 2009 ). Xue Bai and Guillermo Sapiro. 2009. Geodesic matting: A framework for fast interactive image and video segmentation and matting. IJCV (2009)."},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"crossref","unstructured":"Dan Banica and Cristian Sminchisescu. 2015. Second-Order Constrained Parametric Proposals and Sequential Search-Based Structured Prediction for Semantic Segmentation in RGB-D Images. (2015).  Dan Banica and Cristian Sminchisescu. 2015. Second-Order Constrained Parametric Proposals and Sequential Search-Based Structured Prediction for Semantic Segmentation in RGB-D Images. (2015).","DOI":"10.1109\/CVPR.2015.7298974"},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"crossref","unstructured":"D. Batra A. Kowdle D. Parikh J. Luo and T. Chen. 2011. Interactively Co-segmentating Topically Related Images with Intelligent Scribble Guidance. IJCV (2011).  D. Batra A. Kowdle D. Parikh J. Luo and T. Chen. 2011. Interactively Co-segmentating Topically Related Images with Intelligent Scribble Guidance. IJCV (2011).","DOI":"10.1007\/s11263-010-0415-x"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"crossref","unstructured":"Rodrigo Benenson Stefan Popov and Vittorio Ferrari. 2019. Large-scale interactive object segmentation with human annotators. In CVPR.  Rodrigo Benenson Stefan Popov and Vittorio Ferrari. 2019. Large-scale interactive object segmentation with human annotators. In CVPR.","DOI":"10.1109\/CVPR.2019.01197"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"crossref","unstructured":"Arijit Biswas and Devi Parikh. 2013. Simultaneous active learning of classifiers & attributes via relative feedback. In CVPR.  Arijit Biswas and Devi Parikh. 2013. Simultaneous active learning of classifiers & attributes via relative feedback. In CVPR.","DOI":"10.1109\/CVPR.2013.89"},{"key":"e_1_3_2_2_10_1","unstructured":"Y. Boykov and M. P. Jolly. 2001. Interactive Graph Cuts for Optimal Boundary and Region Segmentation of Objects in N-D Images. In ICCV.  Y. Boykov and M. P. Jolly. 2001. Interactive Graph Cuts for Optimal Boundary and Region Segmentation of Objects in N-D Images. In ICCV."},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"crossref","unstructured":"S. Branson K.E. Hj\u00f6rleifsson and P. Perona. 2014. Active annotation translation. In CVPR.  S. Branson K.E. Hj\u00f6rleifsson and P. Perona. 2014. Active annotation translation. In CVPR.","DOI":"10.1109\/CVPR.2014.473"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"crossref","unstructured":"Steve Branson Catherine Wah Florian Schroff Boris Babenko Peter Welinder Pietro Perona and Serge Belongie. 2010. Visual recognition with humans in the loop. In ECCV.  Steve Branson Catherine Wah Florian Schroff Boris Babenko Peter Welinder Pietro Perona and Serge Belongie. 2010. Visual recognition with humans in the loop. In ECCV.","DOI":"10.1007\/978-3-642-15561-1_32"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"crossref","unstructured":"Holger Caesar Jasper Uijlings and Vittorio Ferrari. 2018. COCO-Stuff: Thing and Stuff Classes in Context. In CVPR.  Holger Caesar Jasper Uijlings and Vittorio Ferrari. 2018. COCO-Stuff: Thing and Stuff Classes in Context. In CVPR.","DOI":"10.1109\/CVPR.2018.00132"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"crossref","unstructured":"L. Castrejon K. Kundu R. Urtasun and S. Fidler. 2017. Annotating Object Instances with a Polygon-RNN. In CVPR.  L. Castrejon K. Kundu R. Urtasun and S. Fidler. 2017. Annotating Object Instances with a Polygon-RNN. In CVPR.","DOI":"10.1109\/CVPR.2017.477"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"crossref","unstructured":"L-C. Chen G. Papandreou I. Kokkinos K. Murphy and A.L. Yuille. 2018a. DeepLab: Semantic Image Segmentation with Deep Convolutional Nets Atrous Convolution and Fully Connected CRFs. IEEE Trans. on PAMI (2018).  L-C. Chen G. Papandreou I. Kokkinos K. Murphy and A.L. Yuille. 2018a. DeepLab: Semantic Image Segmentation with Deep Convolutional Nets Atrous Convolution and Fully Connected CRFs. IEEE Trans. on PAMI (2018).","DOI":"10.1109\/TPAMI.2017.2699184"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"crossref","unstructured":"Yuhua Chen Jordi Pont-Tuset Alberto Montes and Luc Van Gool. 2018b. Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning. In CVPR.  Yuhua Chen Jordi Pont-Tuset Alberto Montes and Luc Van Gool. 2018b. Blazingly Fast Video Object Segmentation with Pixel-Wise Metric Learning. In CVPR.","DOI":"10.1109\/CVPR.2018.00130"},{"key":"e_1_3_2_2_17_1","volume-title":"DenseCut: Densely Connected CRFs for Realtime GrabCut. Computer Graphics Forum","author":"Cheng Ming-Ming","year":"2015","unstructured":"Ming-Ming Cheng , V A Prisacariu , Shuai Zheng , Philip H. S. Torr , and Carsten Rother . 2015. DenseCut: Densely Connected CRFs for Realtime GrabCut. Computer Graphics Forum ( 2015 ). Ming-Ming Cheng, V A Prisacariu, Shuai Zheng, Philip H. S. Torr, and Carsten Rother. 2015. DenseCut: Densely Connected CRFs for Realtime GrabCut. Computer Graphics Forum (2015)."},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"crossref","unstructured":"L. Cohen and R. Kimmel. 1996. Global Minimum for Active Contour Models: A Minimal Path Approach. In CVPR.  L. Cohen and R. Kimmel. 1996. Global Minimum for Active Contour Models: A Minimal Path Approach. In CVPR.","DOI":"10.1109\/CVPR.1996.517144"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"crossref","unstructured":"M. Cordts M. Omran S. Ramos T. Rehfeld M. Enzweiler R. Benenson U. Franke S. Roth and B. Schiele. 2016. The Cityscapes Dataset for Semantic Urban Scene Understanding. In CVPR.  M. Cordts M. Omran S. Ramos T. Rehfeld M. Enzweiler R. Benenson U. Franke S. Roth and B. Schiele. 2016. The Cityscapes Dataset for Semantic Urban Scene Understanding. In CVPR.","DOI":"10.1109\/CVPR.2016.350"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"crossref","unstructured":"A. Criminisi T. Sharp C. Rother and P Perez. 2010. Geodesic Image and Video Editing.  A. Criminisi T. Sharp C. Rother and P Perez. 2010. Geodesic Image and Video Editing.","DOI":"10.1145\/1857907.1857910"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10994-009-5106-x"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"crossref","unstructured":"Georgia Gkioxari Alexander Toshev and Navdeep Jaitly. 2016. Chained predictions using convolutional neural networks. (2016).  Georgia Gkioxari Alexander Toshev and Navdeep Jaitly. 2016. Chained predictions using convolutional neural networks. (2016).","DOI":"10.1007\/978-3-319-46493-0_44"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"crossref","unstructured":"V. Gulshan C. Rother A. Criminisi A. Blake and A. Zisserman. 2010. Geodesic star convexity for interactive image segmentation. In CVPR.  V. Gulshan C. Rother A. Criminisi A. Blake and A. Zisserman. 2010. Geodesic star convexity for interactive image segmentation. In CVPR.","DOI":"10.1109\/CVPR.2010.5540073"},{"key":"e_1_3_2_2_24_1","unstructured":"Kaiming He Georgia Gkioxari Piotr Doll\u00e1r and Ross Girshick. 2017. Mask R-CNN. In ICCV.  Kaiming He Georgia Gkioxari Piotr Doll\u00e1r and Ross Girshick. 2017. Mask R-CNN. In ICCV."},{"key":"e_1_3_2_2_25_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR.  Kaiming He Xiangyu Zhang Shaoqing Ren and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR."},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"crossref","unstructured":"G. Heitz and D. Koller. 2008. Learning Spatial Context: Using Stuff to Find Things. In ECCV.  G. Heitz and D. Koller. 2008. Learning Spatial Context: Using Stuff to Find Things. In ECCV.","DOI":"10.1007\/978-3-540-88682-2_4"},{"key":"e_1_3_2_2_27_1","unstructured":"Ronghang Hu Piotr Doll\u00e1r Kaiming He Trevor Darrell and Ross Girshick. 2018. Learning to Segment Every Thing. In CVPR.  Ronghang Hu Piotr Doll\u00e1r Kaiming He Trevor Darrell and Ross Girshick. 2018. Learning to Segment Every Thing. In CVPR."},{"key":"e_1_3_2_2_28_1","volume-title":"A fully convolutional two-stream fusion network for interactive image segmentation. Neural Networks","author":"Hu Yang","year":"2019","unstructured":"Yang Hu , Andrea Soltoggio , Russell Lock , and Steve Carter . 2019. A fully convolutional two-stream fusion network for interactive image segmentation. Neural Networks ( 2019 ). Yang Hu, Andrea Soltoggio, Russell Lock, and Steve Carter. 2019. A fully convolutional two-stream fusion network for interactive image segmentation. Neural Networks (2019)."},{"key":"e_1_3_2_2_29_1","volume-title":"Kingma and Jimmy Lei Ba","author":"Diederik","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Lei Ba . 2015 . Adam : A Method for Stochastic Optimization. In ICLR. Diederik P. Kingma and Jimmy Lei Ba. 2015. Adam: A Method for Stochastic Optimization. In ICLR."},{"key":"e_1_3_2_2_30_1","unstructured":"Thomas N Kipf and Max Welling. 2017. Semi-Supervised Classification with Graph Convolutional Networks. (2017).  Thomas N Kipf and Max Welling. 2017. Semi-Supervised Classification with Graph Convolutional Networks. (2017)."},{"key":"e_1_3_2_2_31_1","unstructured":"A. Kirillov. [n.d.]. Panoptic Challenge Intro. COCO+Mapillary Joing Recognition Challenge Workshop. http:\/\/presentations.cocodataset.org\/ECCV18\/COCO18-Panoptic-Overview.pdf.  A. Kirillov. [n.d.]. Panoptic Challenge Intro. COCO+Mapillary Joing Recognition Challenge Workshop. http:\/\/presentations.cocodataset.org\/ECCV18\/COCO18-Panoptic-Overview.pdf."},{"key":"e_1_3_2_2_32_1","first-page":"00868","article-title":"Panoptic Segmentation","volume":"1801","author":"Kirillov A.","year":"2018","unstructured":"A. Kirillov , K. He , R. Girshick , C. Rother , and P. Dollar . 2018 . Panoptic Segmentation . CVPR , Vol. 1801 . 00868 . A. Kirillov, K. He, R. Girshick, C. Rother, and P. Dollar. 2018. Panoptic Segmentation. CVPR, Vol. 1801.00868.","journal-title":"CVPR"},{"key":"e_1_3_2_2_33_1","unstructured":"K. Konyushkova J.R.R. Uijlings C. Lampert and V. Ferrari. 2018. Learning Intelligent Dialogs for Bounding Box Annotation. In CVPR.  K. Konyushkova J.R.R. Uijlings C. Lampert and V. Ferrari. 2018. Learning Intelligent Dialogs for Bounding Box Annotation. In CVPR."},{"key":"e_1_3_2_2_34_1","unstructured":"Hoang Le Long Mai Brian Price Scott Cohen Hailin Jin and Feng Liu. 2018. Interactive Boundary Prediction for Object Selection. In ECCV.  Hoang Le Long Mai Brian Price Scott Cohen Hailin Jin and Feng Liu. 2018. Interactive Boundary Prediction for Object Selection. In ECCV."},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"crossref","unstructured":"Z. Li Q. Chen and V. Koltun. 2018. Interactive Image Segmentation with Latent Diversity. In CVPR.  Z. Li Q. Chen and V. Koltun. 2018. Interactive Image Segmentation with Latent Diversity. In CVPR.","DOI":"10.1109\/CVPR.2018.00067"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"J.H. Liew Y. Wei W. Xiong S-H. Ong and J. Feng. 2017. Regional interactive image segmentation networks. In ICCV.  J.H. Liew Y. Wei W. Xiong S-H. Ong and J. Feng. 2017. Regional interactive image segmentation networks. In ICCV.","DOI":"10.1109\/ICCV.2017.297"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"crossref","unstructured":"D. Lin Y. Ji D. Lischinski D. Cohen and H. Huang. 2019. Multi-Scale Context Intertwining for Semantic Segmentation. In ECCV.  D. Lin Y. Ji D. Lischinski D. Cohen and H. Huang. 2019. Multi-Scale Context Intertwining for Semantic Segmentation. In ECCV.","DOI":"10.1007\/978-3-030-01219-9_37"},{"key":"e_1_3_2_2_38_1","unstructured":"Tsung-Yi Lin Michael Maire Serge Belongie Lubomir Bourdev Ross Girshick James Hays Pietro Perona Deva Ramanan C. Lawrence Zitnick and Piotr Doll\u00e1r. 2014. Microsoft COCO: Common Objects in Context. In ECCV.  Tsung-Yi Lin Michael Maire Serge Belongie Lubomir Bourdev Ross Girshick James Hays Pietro Perona Deva Ramanan C. Lawrence Zitnick and Piotr Doll\u00e1r. 2014. Microsoft COCO: Common Objects in Context. In ECCV."},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"crossref","unstructured":"J. Long E. Shelhamer and T. Darrell. 2015. Fully Convolutional Networks for Semantic Segmentation. In CVPR.  J. Long E. Shelhamer and T. Darrell. 2015. Fully Convolutional Networks for Semantic Segmentation. In CVPR.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"e_1_3_2_2_40_1","unstructured":"S. Mahadevan P. Voigtlaender and B. Leibe. 2018. Iteratively Trained Interactive Segmentation. In BMVC.  S. Mahadevan P. Voigtlaender and B. Leibe. 2018. Iteratively Trained Interactive Segmentation. In BMVC."},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"crossref","unstructured":"K.-K. Maninis S. Caelles J. Pont-Tuset and L. Van Gool. 2018. Deep Extreme Cut: From Extreme Points to Object Segmentation. In CVPR.  K.-K. Maninis S. Caelles J. Pont-Tuset and L. Van Gool. 2018. Deep Extreme Cut: From Extreme Points to Object Segmentation. In CVPR.","DOI":"10.1109\/CVPR.2018.00071"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"crossref","unstructured":"Davide Modolo Alexander Vezhnevets and Vittorio Ferrari. 2015. Context Forest for Object Class Detection. In BMVC.  Davide Modolo Alexander Vezhnevets and Vittorio Ferrari. 2015. Context Forest for Object Class Detection. In BMVC.","DOI":"10.5244\/C.29.188"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"crossref","unstructured":"R. Mottaghi X. Chen X. Liu N.-G. Cho S.-W. Lee S. Fidler R. Urtasun and A. Yuille. 2014. The role of context for object detection and semantic segmentation in the wild. In CVPR.  R. Mottaghi X. Chen X. Liu N.-G. Cho S.-W. Lee S. Fidler R. Urtasun and A. Yuille. 2014. The role of context for object detection and semantic segmentation in the wild. In CVPR.","DOI":"10.1109\/CVPR.2014.119"},{"key":"e_1_3_2_2_44_1","volume-title":"Trees: A Graphical Model Relating Features, Objects, and Scenes. In NIPS.","author":"Murphy K.","year":"2003","unstructured":"K. Murphy , A. Torralba , and W. T. Freeman . 2003 . Using the Forest to See the Trees: A Graphical Model Relating Features, Objects, and Scenes. In NIPS. K. Murphy, A. Torralba, and W. T. Freeman. 2003. Using the Forest to See the Trees: A Graphical Model Relating Features, Objects, and Scenes. In NIPS."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"crossref","unstructured":"N. S. Nagaraja F. R. Schmidt and T. Brox. 2015. Video Segmentation with Just a Few Strokes. In ICCV.  N. S. Nagaraja F. R. Schmidt and T. Brox. 2015. Video Segmentation with Just a Few Strokes. In ICCV.","DOI":"10.1109\/ICCV.2015.370"},{"key":"e_1_3_2_2_46_1","volume-title":"Frank Keller, and Vittorio Ferrari.","author":"Papadopoulos Dim P","year":"2017","unstructured":"Dim P Papadopoulos , Jasper RR Uijlings , Frank Keller, and Vittorio Ferrari. 2017 . Extreme clicking for efficient object annotation. In ICCV. Dim P Papadopoulos, Jasper RR Uijlings, Frank Keller, and Vittorio Ferrari. 2017. Extreme clicking for efficient object annotation. In ICCV."},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"crossref","unstructured":"D. P. Papadopoulos Jasper R. R. Uijlings F. Keller and V. Ferrari. 2016. We don't need no bounding-boxes: Training object class detectors using only human verification. In CVPR.  D. P. Papadopoulos Jasper R. R. Uijlings F. Keller and V. Ferrari. 2016. We don't need no bounding-boxes: Training object class detectors using only human verification. In CVPR.","DOI":"10.1109\/CVPR.2016.99"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"crossref","unstructured":"Amar Parkash and Devi Parikh. 2012. Attributes for classifier feedback. In ECCV.  Amar Parkash and Devi Parikh. 2012. Attributes for classifier feedback. In ECCV.","DOI":"10.1007\/978-3-642-33712-3_26"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2016.2537320"},{"key":"e_1_3_2_2_50_1","doi-asserted-by":"crossref","unstructured":"A. Rabinovich A. Vedaldi C. Galleguillos E. Wiewiora and S. Belongie. 2007. Objects in Context. In ICCV.  A. Rabinovich A. Vedaldi C. Galleguillos E. Wiewiora and S. Belongie. 2007. Objects in Context. In ICCV.","DOI":"10.1109\/ICCV.2007.4408986"},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"crossref","unstructured":"C. Rother V. Kolmogorov and A. Blake. 2004. GrabCut: Interactive Foreground Extraction using Iterated Graph Cuts. In SIGGRAPH.  C. Rother V. Kolmogorov and A. Blake. 2004. GrabCut: Interactive Foreground Extraction using Iterated Graph Cuts. In SIGGRAPH.","DOI":"10.1145\/1186562.1015720"},{"key":"e_1_3_2_2_52_1","volume-title":"Guide Me: Interacting with Deep Networks. In CVPR.","author":"Rupprecht Christian","year":"2018","unstructured":"Christian Rupprecht , Iro Laina , Nassir Navab , Gregory D. Hager , and Federico Tombari . 2018 . Guide Me: Interacting with Deep Networks. In CVPR. Christian Rupprecht, Iro Laina, Nassir Navab, Gregory D. Hager, and Federico Tombari. 2018. Guide Me: Interacting with Deep Networks. In CVPR."},{"key":"e_1_3_2_2_53_1","doi-asserted-by":"crossref","unstructured":"Olga Russakovsky Li-Jia Li and Li Fei-Fei. 2015. Best of both worlds: human-machine collaboration for object annotation. In CVPR.  Olga Russakovsky Li-Jia Li and Li Fei-Fei. 2015. Best of both worlds: human-machine collaboration for object annotation. In CVPR.","DOI":"10.1109\/CVPR.2015.7298824"},{"key":"e_1_3_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-007-0090-8"},{"key":"e_1_3_2_2_55_1","unstructured":"Adam Santoro David Raposo David G Barrett Mateusz Malinowski Razvan Pascanu Peter Battaglia and Timothy Lillicrap. 2017. A simple neural network module for relational reasoning. In NIPS.  Adam Santoro David Raposo David G Barrett Mateusz Malinowski Razvan Pascanu Peter Battaglia and Timothy Lillicrap. 2017. A simple neural network module for relational reasoning. In NIPS."},{"volume-title":"AAAI Human Computation Workshop.","author":"Su H.","key":"e_1_3_2_2_56_1","unstructured":"H. Su , J. Deng , and L. Fei-Fei . 2012. Crowdsourcing annotations for visual object detection . In AAAI Human Computation Workshop. H. Su, J. Deng, and L. Fei-Fei. 2012. Crowdsourcing annotations for visual object detection. In AAAI Human Computation Workshop."},{"key":"e_1_3_2_2_57_1","doi-asserted-by":"crossref","unstructured":"J. Tighe and S. Lazebnik. 2011. Understanding Scenes on Many Levels. In ICCV.  J. Tighe and S. Lazebnik. 2011. Understanding Scenes on Many Levels. In ICCV.","DOI":"10.1109\/ICCV.2011.6126260"},{"key":"e_1_3_2_2_58_1","volume-title":"Selective search for object recognition. IJCV","author":"Uijlings J. R. R.","year":"2013","unstructured":"J. R. R. Uijlings , K. E. A. van de Sande , T. Gevers , and A. W. M. Smeulders . 2013. Selective search for object recognition. IJCV ( 2013 ). J. R. R. Uijlings, K. E. A. van de Sande, T. Gevers, and A. W. M. Smeulders. 2013. Selective search for object recognition. IJCV (2013)."},{"key":"e_1_3_2_2_59_1","doi-asserted-by":"crossref","unstructured":"Sudheendra Vijayanarasimhan and Kristen Grauman. 2009. What's it going to cost you?: Predicting effort vs. informativeness for multi-label image annotations. In CVPR.  Sudheendra Vijayanarasimhan and Kristen Grauman. 2009. What's it going to cost you?: Predicting effort vs. informativeness for multi-label image annotations. In CVPR.","DOI":"10.1109\/CVPR.2009.5206705"},{"key":"e_1_3_2_2_60_1","volume-title":"Steve Branson, Subhrajyoti Maji, Pietro Perona, and Serge Belongie.","author":"Wah Catherine","year":"2014","unstructured":"Catherine Wah , Grant Van Horn , Steve Branson, Subhrajyoti Maji, Pietro Perona, and Serge Belongie. 2014 . Similarity comparisons for interactive fine-grained categorization. In CVPR. Catherine Wah, Grant Van Horn, Steve Branson, Subhrajyoti Maji, Pietro Perona, and Serge Belongie. 2014. Similarity comparisons for interactive fine-grained categorization. In CVPR."},{"key":"e_1_3_2_2_61_1","doi-asserted-by":"crossref","unstructured":"N. Xu B. Price S. Cohen J. Yang and T.S. Huang. 2016. Deep interactive object selection. In CVPR.  N. Xu B. Price S. Cohen J. Yang and T.S. Huang. 2016. Deep interactive object selection. In CVPR.","DOI":"10.1109\/CVPR.2016.47"},{"key":"e_1_3_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-018-1140-0"}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Seattle WA USA","acronym":"MM '20"},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413812","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413812","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:01:17Z","timestamp":1750197677000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413812"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":62,"alternative-id":["10.1145\/3394171.3413812","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413812","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}