{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:08:27Z","timestamp":1750219707010,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":29,"publisher":"ACM","license":[{"start":{"date-parts":[[2023,12,6]],"date-time":"2023-12-06T00:00:00Z","timestamp":1701820800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2023,12,6]]},"DOI":"10.1145\/3595916.3626447","type":"proceedings-article","created":{"date-parts":[[2024,1,1]],"date-time":"2024-01-01T16:34:41Z","timestamp":1704126881000},"page":"1-7","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["MontageNet: Annotated Dataset of Furniture Components in Real-World Images"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0002-7693-2579","authenticated-orcid":false,"given":"Iuan Kai","family":"Fang","sequence":"first","affiliation":[{"name":"National Tsing Hua University, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1426-9743","authenticated-orcid":false,"given":"Bo Hao","family":"Zhang","sequence":"additional","affiliation":[{"name":"National Tsing Hua University, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-9689-0041","authenticated-orcid":false,"given":"Te Lun","family":"Liu","sequence":"additional","affiliation":[{"name":"National Tsing Hua University, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0005-9637-1504","authenticated-orcid":false,"given":"Hao","family":"Tan","sequence":"additional","affiliation":[{"name":"National Tsing Hua University, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-4101-2099","authenticated-orcid":false,"given":"Wei Syun","family":"Chen","sequence":"additional","affiliation":[{"name":"National Tsing Hua University, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3940-4478","authenticated-orcid":false,"given":"Che-Rung","family":"Lee","sequence":"additional","affiliation":[{"name":"National Tsing Hua University\\t, TW"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,1]]},"reference":[{"volume-title":"Proceedings of 28th IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Song S.","key":"e_1_3_2_1_1_1","unstructured":"S. Song , S. Lichtenberg , and J. Xiao . 2015. SUN RGB-D: A RGB-D Scene Understanding Benchmark Suite . In Proceedings of 28th IEEE Conference on Computer Vision and Pattern Recognition (CVPR) S. Song, S. Lichtenberg, and J. Xiao. 2015. SUN RGB-D: A RGB-D Scene Understanding Benchmark Suite. In Proceedings of 28th IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Sun Xingyuan","key":"e_1_3_2_1_2_1","unstructured":"Xingyuan Sun , Jiajun Wu , Xiuming Zhang , Zhoutong Zhang , Chengkai Zhang , Tianfan Xue , Joshua B. Tenenbaum , and William T. Freeman . 2018. Pix3D: Dataset and Methods for Single-Image 3D Shape Modeling . In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Xingyuan Sun, Jiajun Wu, Xiuming Zhang, Zhoutong Zhang, Chengkai Zhang, Tianfan Xue, Joshua B. Tenenbaum, and William T. Freeman. 2018. Pix3D: Dataset and Methods for Single-Image 3D Shape Modeling. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2014.6836101"},{"key":"e_1_3_2_1_4_1","unstructured":"Yongzhi Su Mingxin Liu Jason Rambach Antonia Pehrson Anton Berg Didier Stricker. 2021. IKEA Object State Dataset: A 6DoF object pose estimation dataset and benchmark for multi-state assembly objects. arXiv:2111.08614.  Yongzhi Su Mingxin Liu Jason Rambach Antonia Pehrson Anton Berg Didier Stricker. 2021. IKEA Object State Dataset: A 6DoF object pose estimation dataset and benchmark for multi-state assembly objects. arXiv:2111.08614."},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_2_1_6_1","unstructured":"Tsung-Yi Lin Michael Maire Serge Belongie Lubomir Bourdev Ross Girshick James Hays Pietro Perona Deva Ramanan C. Lawrence Zitnick Piotr Doll\u00e1r. 2014. Microsoft COCO: Common Objects in Context. arXiv:1405.0312v3.  Tsung-Yi Lin Michael Maire Serge Belongie Lubomir Bourdev Ross Girshick James Hays Pietro Perona Deva Ramanan C. Lawrence Zitnick Piotr Doll\u00e1r. 2014. Microsoft COCO: Common Objects in Context. arXiv:1405.0312v3."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"crossref","unstructured":"Andreas Geiger Philip Lenz Christoph Stiller Raquel Urtasun. 2013. Vision meets Robotics: The KITTI Dataset. In International Journal of Robotics Research (IJRR)  Andreas Geiger Philip Lenz Christoph Stiller Raquel Urtasun. 2013. Vision meets Robotics: The KITTI Dataset. In International Journal of Robotics Research (IJRR)","DOI":"10.1177\/0278364913491297"},{"key":"e_1_3_2_1_8_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zhou Bolei","year":"2017","unstructured":"Bolei Zhou , Hang Zhao , Xavier Puig , Sanja Fidler , Adela Barriuso , Antonio Torralba . 2017 . Scene Parsing Through ADE20K Dataset . In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Bolei Zhou, Hang Zhao, Xavier Puig, Sanja Fidler, Adela Barriuso, Antonio Torralba. 2017. Scene Parsing Through ADE20K Dataset. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33715-4_54"},{"volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Xiao J.","key":"e_1_3_2_1_10_1","unstructured":"J. Xiao , J. Hays , K. Ehinger , A. Oliva , and A. Torralba . 2010. SUN Database: Large-scale Scene Recognition from Abbey to Zoo . In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) J. Xiao, J. Hays, K. Ehinger, A. Oliva, and A. Torralba. 2010. SUN Database: Large-scale Scene Recognition from Abbey to Zoo. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_11_1","volume-title":"Places: A 10 million Image Database for Scene Recognition","author":"Zhou B.","year":"2017","unstructured":"B. Zhou , A. Lapedriza , A. Khosla , A. Oliva , and A. Torralba . 2017 . Places: A 10 million Image Database for Scene Recognition . In IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) B. Zhou, A. Lapedriza, A. Khosla, A. Oliva, and A. Torralba. 2017. Places: A 10 million Image Database for Scene Recognition. In IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/219717.219748"},{"key":"e_1_3_2_1_13_1","volume-title":"European Conference on Computer Vision (ECCV)","author":"Xiang Yu","year":"2016","unstructured":"Yu Xiang , Wonhui Kim , Wei Chen , Jingwei Ji , Christopher Choy , Hao Su , Roozbeh Mottaghi , Leonidas Guibas and Silvio Savarese . 2016 . ObjectNet3D: A Large Scale Database for 3D Object Recognition . In European Conference on Computer Vision (ECCV) Yu Xiang, Wonhui Kim, Wei Chen, Jingwei Ji, Christopher Choy, Hao Su, Roozbeh Mottaghi, Leonidas Guibas and Silvio Savarese. 2016. ObjectNet3D: A Large Scale Database for 3D Object Recognition. In European Conference on Computer Vision (ECCV)"},{"key":"e_1_3_2_1_14_1","unstructured":"Angel X. Chang Thomas Funkhouser Leonidas Guibas Pat Hanrahan Qi-Xing Huang Zimo Li Silvio Savarese Manolis Savva Shuran Song Hao Su Jianxiong Xiao Li Yi Fisher Yu. 2015. ShapeNet: An Information-Rich 3D Model Repository. arXiv:1512.03012.  Angel X. Chang Thomas Funkhouser Leonidas Guibas Pat Hanrahan Qi-Xing Huang Zimo Li Silvio Savarese Manolis Savva Shuran Song Hao Su Jianxiong Xiao Li Yi Fisher Yu. 2015. ShapeNet: An Information-Rich 3D Model Repository. arXiv:1512.03012."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"crossref","first-page":"443","DOI":"10.1007\/s10851-022-01083-1","article-title":"Elastic 3D\u20132D Image Registration","volume":"64","author":"Paul","year":"2022","unstructured":"Paul Striewski1, Benedikt Wirth . 2022 . Elastic 3D\u20132D Image Registration . J Math Imaging Vis 64 , 443 \u2013 462 . Paul Striewski1, Benedikt Wirth. 2022. Elastic 3D\u20132D Image Registration. J Math Imaging Vis 64, 443\u2013462.","journal-title":"J Math Imaging Vis"},{"key":"e_1_3_2_1_16_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Mo Kaichun","year":"2019","unstructured":"Kaichun Mo , Shilin Zhu , Angel X. Chang , Li Yi , Subarna Tripathi , Leonidas J. Guibas , Hao Su . 2019 . PartNet: A Large-scale Benchmark for Fine-grained and Hierarchical Part-level 3D Object Understanding . In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Kaichun Mo, Shilin Zhu, Angel X. Chang, Li Yi, Subarna Tripathi, Leonidas J. Guibas, Hao Su. 2019. PartNet: A Large-scale Benchmark for Fine-grained and Hierarchical Part-level 3D Object Understanding. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_17_1","volume-title":"Proceeding of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Cordts Marius","year":"2016","unstructured":"Marius Cordts , Mohamed Omran , Sebastian Ramos , Timo Rehfeld , Markus Enzweiler , Rodrigo Benenson , Uwe Franke , Stefan Roth , Bernt Schiele . 2016 . The Cityscapes Dataset for Semantic Urban Scene Understanding . In Proceeding of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, Bernt Schiele. 2016. The Cityscapes Dataset for Semantic Urban Scene Understanding. In Proceeding of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"crossref","unstructured":"Shervin Minaee Yuri Boykov Fatih Porikli Antonio Plaza Nasser Kehtarnavaz Demetri Terzopoulos. 2020. Image Segmentation Using Deep Learning: A Survey. arXiv:2001.05566.  Shervin Minaee Yuri Boykov Fatih Porikli Antonio Plaza Nasser Kehtarnavaz Demetri Terzopoulos. 2020. Image Segmentation Using Deep Learning: A Survey. arXiv:2001.05566.","DOI":"10.1109\/TPAMI.2021.3059968"},{"key":"e_1_3_2_1_19_1","volume-title":"COCO-Stuff: Thing and Stuff Classes in Context. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Caesar Holger","year":"2018","unstructured":"Holger Caesar , Jasper Uijlings , Vittorio Ferrari . 2018 . COCO-Stuff: Thing and Stuff Classes in Context. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Holger Caesar, Jasper Uijlings, Vittorio Ferrari. 2018. COCO-Stuff: Thing and Stuff Classes in Context. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_20_1","volume-title":"The Role of Context for Object Detection and Semantic Segmentation in the Wild. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Mottaghi Roozbeh","year":"2014","unstructured":"Roozbeh Mottaghi , Xianjie Chen , Xiaobai Liu , Nam-Gyu Cho , Seong-Whan Lee , Sanja Fidler , Raquel Urtasun , Alan Yuille . 2014 . The Role of Context for Object Detection and Semantic Segmentation in the Wild. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Roozbeh Mottaghi, Xianjie Chen, Xiaobai Liu, Nam-Gyu Cho, Seong-Whan Lee, Sanja Fidler, Raquel Urtasun, Alan Yuille. 2014. The Role of Context for Object Detection and Semantic Segmentation in the Wild. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_21_1","volume-title":"Christopher K. I. Williams, John Winn, and Andrew Zisserman.","author":"Everingham Mark","year":"2010","unstructured":"Mark Everingham , Luc Van Gool , Christopher K. I. Williams, John Winn, and Andrew Zisserman. 2010 . The PASCAL Visual Object Classes (VOC) Challenge. In International Journal of Computer Vision (IJCV) Mark Everingham, Luc Van Gool, Christopher K. I. Williams, John Winn, and Andrew Zisserman. 2010. The PASCAL Visual Object Classes (VOC) Challenge. In International Journal of Computer Vision (IJCV)"},{"key":"e_1_3_2_1_22_1","volume-title":"Pyramid Scene Parsing Network. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Zhao Hengshuang","year":"2017","unstructured":"Hengshuang Zhao , Jianping Shi , Xiaojuan Qi , Xiaogang Wang , Jiaya Jia . 2017 . Pyramid Scene Parsing Network. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, Jiaya Jia. 2017. Pyramid Scene Parsing Network. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_23_1","unstructured":"Kaiming He Xiangyu Zhang Shaoqing Ren Jian Sun. 2015. Deep Residual Learning for Image Recognition. In Computing Research Repository (CoRR)  Kaiming He Xiangyu Zhang Shaoqing Ren Jian Sun. 2015. Deep Residual Learning for Image Recognition. In Computing Research Repository (CoRR)"},{"key":"e_1_3_2_1_24_1","volume-title":"Unified Perceptual Parsing for Scene Understanding. In European Conference on Computer Vision (ECCV)","author":"Xiao Tete","year":"2018","unstructured":"Tete Xiao , Yingcheng Liu , Bolei Zhou , Yuning Jiang , Jian Sun . 2018 . Unified Perceptual Parsing for Scene Understanding. In European Conference on Computer Vision (ECCV) Tete Xiao, Yingcheng Liu, Bolei Zhou, Yuning Jiang, Jian Sun. 2018. Unified Perceptual Parsing for Scene Understanding. In European Conference on Computer Vision (ECCV)"},{"key":"e_1_3_2_1_25_1","volume-title":"BEiT: BERT Pre-Training of Image Transformers. In International Conference on Learning Representations (ICLR)","author":"Bao Hangbo","year":"2022","unstructured":"Hangbo Bao , Li Dong , Songhao Piao , Furu Wei . 2022 . BEiT: BERT Pre-Training of Image Transformers. In International Conference on Learning Representations (ICLR) Hangbo Bao, Li Dong, Songhao Piao, Furu Wei. 2022. BEiT: BERT Pre-Training of Image Transformers. In International Conference on Learning Representations (ICLR)"},{"key":"e_1_3_2_1_26_1","volume-title":"Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation. In European Conference on Computer Vision (ECCV)","author":"Chen Liang-Chieh","year":"2018","unstructured":"Liang-Chieh Chen , Yukun Zhu , George Papandreou , Florian Schroff , Hartwig Adam . 2018 . Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation. In European Conference on Computer Vision (ECCV) Liang-Chieh Chen, Yukun Zhu, George Papandreou, Florian Schroff, Hartwig Adam. 2018. Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation. In European Conference on Computer Vision (ECCV)"},{"key":"e_1_3_2_1_27_1","volume-title":"Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation. In European Conference on Computer Vision (ECCV)","author":"Yuan Yuhui","year":"2020","unstructured":"Yuhui Yuan , Xiaokang Chen , Xilin Chen , Jingdong Wang . 2020 . Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation. In European Conference on Computer Vision (ECCV) Yuhui Yuan, Xiaokang Chen, Xilin Chen, Jingdong Wang. 2020. Segmentation Transformer: Object-Contextual Representations for Semantic Segmentation. In European Conference on Computer Vision (ECCV)"},{"key":"e_1_3_2_1_28_1","volume-title":"Deep High-Resolution Representation Learning for Human Pose Estimation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)","author":"Sun Ke","year":"2019","unstructured":"Ke Sun , Bin Xiao , Dong Liu , Jingdong Wang . 2019 . Deep High-Resolution Representation Learning for Human Pose Estimation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Ke Sun, Bin Xiao, Dong Liu, Jingdong Wang. 2019. Deep High-Resolution Representation Learning for Human Pose Estimation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)"},{"key":"e_1_3_2_1_29_1","unstructured":"MMSegmentation Contributors. 2020. OpenMMLab Semantic Segmentation Toolbox and Benchmark. https:\/\/github.com\/open-mmlab\/mmsegmentation  MMSegmentation Contributors. 2020. OpenMMLab Semantic Segmentation Toolbox and Benchmark. https:\/\/github.com\/open-mmlab\/mmsegmentation"}],"event":{"name":"MMAsia '23: ACM Multimedia Asia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Tainan Taiwan","acronym":"MMAsia '23"},"container-title":["ACM Multimedia Asia 2023"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3595916.3626447","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3595916.3626447","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T16:35:56Z","timestamp":1750178156000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3595916.3626447"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,6]]},"references-count":29,"alternative-id":["10.1145\/3595916.3626447","10.1145\/3595916"],"URL":"https:\/\/doi.org\/10.1145\/3595916.3626447","relation":{},"subject":[],"published":{"date-parts":[[2023,12,6]]},"assertion":[{"value":"2024-01-01","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}