{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,8]],"date-time":"2026-07-08T15:53:43Z","timestamp":1783526023415,"version":"3.55.0"},"reference-count":60,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2023,2,6]],"date-time":"2023-02-06T00:00:00Z","timestamp":1675641600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["61972067\/U21A20491"],"award-info":[{"award-number":["61972067\/U21A20491"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Innovation Technology Funding of Dalian","award":["2020JJ26GX036"],"award-info":[{"award-number":["2020JJ26GX036"]}]},{"name":"Open Research Fund of Beijing Key Laboratory of Big Data Technology for Food Safety","award":["BTBD-2018KF"],"award-info":[{"award-number":["BTBD-2018KF"]}]},{"name":"National Key Research and Development Program of China","award":["2022ZD0210500"],"award-info":[{"award-number":["2022ZD0210500"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2023,5,31]]},"abstract":"<jats:p>Most matting research resorts to advanced semantics to achieve high-quality alpha mattes, and a direct low-level features combination is usually explored to complement alpha details. However, we argue that appearance-agnostic integration can only provide biased foreground (FG) details and that alpha mattes require different-level feature aggregation for better pixel-wise opacity perception. In this article, we propose an end-to-end hierarchical and progressive attention matting network (HAttMatting++), which can better predict the opacity of the FG from single RGB images without additional input. Specifically, we utilize channel-wise attention (CA) to distill pyramidal features and employ spatial attention (SA) at different levels to filter appearance cues. This progressive attention mechanism can estimate alpha mattes from adaptive semantics and semantics-indicated boundaries. We also introduce a hybrid loss function fusing structural similarity, mean square error, adversarial loss, and sentry supervision to guide the network to further improve the overall FG structure. In addition, we construct a large-scale and challenging image matting dataset comprised of 59,000 training images and 1,000 test images (a total of 646 distinct FG alpha mattes), which can further improve the robustness of our hierarchical and progressive aggregation model. Extensive experiments demonstrate that the proposed HAttMatting++ can capture sophisticated FG structures and achieve state-of-the-art performance with single RGB images as input.<\/jats:p>","DOI":"10.1145\/3540201","type":"journal-article","created":{"date-parts":[[2022,6,11]],"date-time":"2022-06-11T22:41:04Z","timestamp":1654987264000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":14,"title":["Hierarchical and Progressive Image Matting"],"prefix":"10.1145","volume":"19","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7205-2924","authenticated-orcid":false,"given":"Yu","family":"Qiao","sequence":"first","affiliation":[{"name":"Dalian University of Technology, Dalian, Liaoning, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0550-4788","authenticated-orcid":false,"given":"Yuhao","family":"Liu","sequence":"additional","affiliation":[{"name":"Dalian University of Technology, Dalian, Liaoning, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9402-8386","authenticated-orcid":false,"given":"Ziqi","family":"Wei","sequence":"additional","affiliation":[{"name":"Institute of Automation, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5133-3978","authenticated-orcid":false,"given":"Yuxin","family":"Wang","sequence":"additional","affiliation":[{"name":"Dalian University of Technology, Dalian, Liaoning, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3865-9252","authenticated-orcid":false,"given":"Qiang","family":"Cai","sequence":"additional","affiliation":[{"name":"Beijing Technology and Business University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9525-9573","authenticated-orcid":false,"given":"Guofeng","family":"Zhang","sequence":"additional","affiliation":[{"name":"Wonxing Technology, Shenzhen, Guangdong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8046-722X","authenticated-orcid":false,"given":"Xin","family":"Yang","sequence":"additional","affiliation":[{"name":"Dalian University of Technology, Dalian, Liaoning, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,2,6]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"228","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Aksoy Yagiz","year":"2017","unstructured":"Yagiz Aksoy, Tunc Ozan Aydin, and Marc Pollefeys. 2017. Designing effective inter-pixel information flow for natural image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 228\u2013236."},{"key":"e_1_3_1_3_2","article-title":"Semantic soft segmentation","volume":"37","author":"Aksoy Ya\u011f\u0131z","year":"2018","unstructured":"Ya\u011f\u0131z Aksoy, Tae-Hyun Oh, Sylvain Paris, Marc Pollefeys, and Wojciech Matusik. 2018. Semantic soft segmentation. ACM Transactions on Graphics 37, 4 (2018), Article 72.","journal-title":"ACM Transactions on Graphics"},{"key":"e_1_3_1_4_2","first-page":"8818","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919)","author":"Cai Shaofan","year":"2019","unstructured":"Shaofan Cai, Xiaoshuai Zhang, Haoqiang Fan, Haibin Huang, Jiangyu Liu, Jiaming Liu, Jiaying Liu, Jue Wang, and Jian Sun. 2019. Disentangled image matting. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919). 8818\u20138827."},{"key":"e_1_3_1_5_2","first-page":"6298","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Chen Long","year":"2017","unstructured":"Long Chen, Hanwang Zhang, Jun Xiao, Liqiang Nie, Jian Shao, Wei Liu, and Tat-Seng Chua. 2017. SCA-CNN: Spatial and channel-wise attention in convolutional networks for image captioning. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 6298\u20136306."},{"issue":"4","key":"e_1_3_1_6_2","doi-asserted-by":"crossref","first-page":"834","DOI":"10.1109\/TPAMI.2017.2699184","article-title":"DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs","volume":"40","author":"Chen L. C.","year":"2018","unstructured":"L. C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille. 2018. DeepLab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected CRFs. IEEE Transactions on Pattern Analysis and Machine Intelligence 40, 4 (2018), 834\u2013848.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"618","DOI":"10.1145\/3240508.3240610","volume-title":"Proceedings of the ACM International Conference on Multimedia (MM\u201918)","author":"Chen Quan","year":"2018","unstructured":"Quan Chen, Tiezheng Ge, Yanyu Xu, Zhiqiang Zhang, Xinxin Yang, and Kun Gai. 2018. Semantic human matting. In Proceedings of the ACM International Conference on Multimedia (MM\u201918). 618\u2013626."},{"issue":"9","key":"e_1_3_1_8_2","doi-asserted-by":"crossref","first-page":"2175","DOI":"10.1109\/TPAMI.2013.18","article-title":"KNN matting","volume":"35","author":"Chen Qifeng","year":"2013","unstructured":"Qifeng Chen, Dingzeyu Li, and Chi Keung Tang. 2013. KNN matting. IEEE Transactions on Pattern Analysis and Machine Intelligence 35, 9 (2013), 2175\u20132188.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"8","key":"e_1_3_1_9_2","doi-asserted-by":"crossref","first-page":"1504","DOI":"10.1109\/TPAMI.2016.2606397","article-title":"Automatic trimap generation and consistent matting for light-field images","volume":"39","author":"Cho D.","year":"2016","unstructured":"D. Cho, S. Kim, Y. W. Tai, and I. S. Kweon. 2016. Automatic trimap generation and consistent matting for light-field images. IEEE Transactions on Pattern Analysis and Machine Intelligence 39, 8 (2016), 1504\u20131517.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"3","key":"e_1_3_1_10_2","doi-asserted-by":"crossref","first-page":"1054","DOI":"10.1109\/TIP.2018.2872925","article-title":"Deep convolutional neural network for natural image matting using initial alpha mattes","volume":"28","author":"Cho Donghyeon","year":"2019","unstructured":"Donghyeon Cho, Yu-Wing Tai, and In So Kweon. 2019. Deep convolutional neural network for natural image matting using initial alpha mattes. IEEE Transactions on Image Processing 28, 3 (2019), 1054\u20131067.","journal-title":"IEEE Transactions on Image Processing"},{"key":"e_1_3_1_11_2","first-page":"6841","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921)","author":"Dai Yutong","year":"2021","unstructured":"Yutong Dai, Hao Lu, and Chunhua Shen. 2021. Learning affinity-aware upsampling for deep image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921). 6841\u20136850."},{"issue":"2","key":"e_1_3_1_12_2","doi-asserted-by":"crossref","first-page":"303","DOI":"10.1007\/s11263-009-0275-4","article-title":"The PASCAL Visual Object Classes (VOC) challenge","volume":"88","author":"Everingham Mark","year":"2010","unstructured":"Mark Everingham, Luc Van Gool, Christopher K. I. Williams, John Winn, and Andrew Zisserman. 2010. The PASCAL Visual Object Classes (VOC) challenge. International Journal of Computer Vision 88, 2 (2010), 303\u2013338.","journal-title":"International Journal of Computer Vision"},{"issue":"2","key":"e_1_3_1_13_2","doi-asserted-by":"crossref","first-page":"575","DOI":"10.1111\/j.1467-8659.2009.01627.x","article-title":"Shared sampling for real-time alpha matting","volume":"29","author":"Gastal Eduardo S. L.","year":"2010","unstructured":"Eduardo S. L. Gastal and Manuel M. Oliveira. 2010. Shared sampling for real-time alpha matting. Computer Graphics Forum 29, 2 (2010), 575\u2013584.","journal-title":"Computer Graphics Forum"},{"key":"e_1_3_1_14_2","first-page":"2672","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems (NeurIPS\u201914)","author":"Goodfellow Ian J.","year":"2014","unstructured":"Ian J. Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Xu Bing, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative adversarial nets. In Proceedings of the International Conference on Neural Information Processing Systems (NeurIPS\u201914). 2672\u20132680."},{"key":"e_1_3_1_15_2","first-page":"3203","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Hou Qibin","year":"2017","unstructured":"Qibin Hou, Ming-Ming Cheng, Xiaowei Hu, Ali Borji, Zhuowen Tu, and Philip H. S. Torr. 2017. Deeply supervised salient object detection with short connections. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 3203\u20133212."},{"key":"e_1_3_1_16_2","first-page":"4129","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919)","author":"Hou Qiqi","year":"2019","unstructured":"Qiqi Hou and Feng Liu. 2019. Context-aware image matting for simultaneous foreground and alpha estimation. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919). 4129\u20134138."},{"key":"e_1_3_1_17_2","first-page":"5967","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Isola Phillip","year":"2017","unstructured":"Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros. 2017. Image-to-image translation with conditional adversarial networks. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 5967\u20135976."},{"key":"e_1_3_1_18_2","first-page":"424","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201915)","author":"Karacan L.","year":"2015","unstructured":"L. Karacan, A. Erdem, and E. Erdem. 2015. Image matting with KL-divergence based sparse sampling. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201915). 424\u2013432."},{"key":"e_1_3_1_19_2","first-page":"2193","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201911)","author":"Lee P.","year":"2011","unstructured":"P. Lee and Ying Wu. 2011. Nonlocal matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201911). 2193\u20132200."},{"issue":"2","key":"e_1_3_1_20_2","doi-asserted-by":"crossref","first-page":"228","DOI":"10.1109\/TPAMI.2007.1177","article-title":"A closed-form solution to natural image matting","volume":"30","author":"Levin Anat","year":"2007","unstructured":"Anat Levin, Dani Lischinski, and Yair Weiss. 2007. A closed-form solution to natural image matting. IEEE Transactions on Pattern Analysis and Machine Intelligence 30, 2 (2007), 228\u2013242.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"10","key":"e_1_3_1_21_2","doi-asserted-by":"crossref","first-page":"1699","DOI":"10.1109\/TPAMI.2008.168","article-title":"Spectral matting","volume":"30","author":"Levin Anat","year":"2008","unstructured":"Anat Levin, Alex Rav-Acha, and Dani Lischinski. 2008. Spectral matting. IEEE Transactions on Pattern Analysis and Machine Intelligence 30, 10 (2008), 1699\u20131712.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_1_22_2","first-page":"11450","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence (AAAI\u201920)","author":"Li Yaoyi","year":"2020","unstructured":"Yaoyi Li and Hongtao Lu. 2020. Natural image matting via guided contextual attention. In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI\u201920). 11450\u201311457."},{"key":"e_1_3_1_23_2","first-page":"8762","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921)","author":"Lin Shanchuan","year":"2021","unstructured":"Shanchuan Lin, Andrey Ryabtsev, Soumyadip Sengupta, Brian L. Curless, Steven M. Seitz, and Ira Kemelmacher-Shlizerman. 2021. Real-time high-resolution background matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921). 8762\u20138771."},{"key":"e_1_3_1_24_2","first-page":"740","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201914)","author":"Lin Tsung-Yi","year":"2014","unstructured":"Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll\u00e1r, and C. Lawrence Zitnick. 2014. Microsoft COCO: Common objects in context. In Proceedings of the European Conference on Computer Vision (ECCV\u201914). 740\u2013755."},{"key":"e_1_3_1_25_2","article-title":"ParseNet: Looking wider to see better","author":"Liu Wei","year":"2015","unstructured":"Wei Liu, Andrew Rabinovich, and Alexander C. Berg. 2015. ParseNet: Looking wider to see better. arXiv preprint arXiv:1506.04579 (2015).","journal-title":"arXiv preprint arXiv:1506.04579"},{"key":"e_1_3_1_26_2","first-page":"3265","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919)","author":"Lu Hao","year":"2019","unstructured":"Hao Lu, Yutong Dai, Chunhua Shen, and Songcen Xu. 2019. Indices matter: Learning to index for deep image matting. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201919). 3265\u20133274."},{"key":"e_1_3_1_27_2","first-page":"259","volume-title":"Proceedings of the British Machine Vision Conference (BMVC\u201918)","author":"Lutz Sebastian","year":"2018","unstructured":"Sebastian Lutz, Konstantinos Amplianitis, and Aljoscha Smolic. 2018. AlphaGAN: Generative adversarial networks for natural image matting. In Proceedings of the British Machine Vision Conference (BMVC\u201918). 259."},{"key":"e_1_3_1_28_2","article-title":"Exploring dense context for salient object detection","author":"Mei Haiyang","year":"2021","unstructured":"Haiyang Mei, Yuanyuan Liu, Ziqi Wei, Dongsheng Zhou, Xiaopeng Xiaopeng, Qiang Zhang, and Xin Yang. 2021. Exploring dense context for salient object detection. IEEE Transactions on Circuits and Systems for Video Technology 32, 3 (2021), 1378\u20131389.","journal-title":"IEEE Transactions on Circuits and Systems for Video Technology"},{"key":"e_1_3_1_29_2","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201920)","author":"Mei Haiyang","year":"2020","unstructured":"Haiyang Mei, Xin Yang, Yang Wang, Yuanyuan Liu, Shengfeng He, Qiang Zhang, Xiaopeng Wei, and Rynson W. H. Lau. 2020. Don\u2019t hit me! Glass detection in real-world scenes. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201920)."},{"key":"e_1_3_1_30_2","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201920)","author":"Qiao Yu","year":"2020","unstructured":"Yu Qiao, Yuhao Liu, Xin Yang, Dongsheng Zhou, Mingliang Xu, Qiang Zhang, and Xiaopeng Wei. 2020. Attention-guided hierarchical structure aggregation for image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201920)."},{"key":"e_1_3_1_31_2","first-page":"565","volume-title":"Computer Graphics Forum","author":"Qiao Yu","year":"2020","unstructured":"Yu Qiao, Yuhao Liu, Qiang Zhu, Xin Yang, Yuxin Wang, Qiang Zhang, and Xiaopeng Wei. 2020. Multi-scale information assembly for image matting. Computer Graphics Forum 39 (2020), 565\u2013574."},{"key":"e_1_3_1_32_2","first-page":"7471","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919)","author":"Qin Xuebin","year":"2019","unstructured":"Xuebin Qin, Zichen Zhang, Chenyang Huang, Chao Gao, Masood Dehghan, and Martin Jagersand. 2019. BASNet: Boundary-aware salient object detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919). 7471\u20137481."},{"key":"e_1_3_1_33_2","first-page":"2049","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201911)","author":"Rhemann C.","year":"2011","unstructured":"C. Rhemann and C. Rother. 2011. A global sampling method for alpha matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201911). 2049\u20132056."},{"key":"e_1_3_1_34_2","doi-asserted-by":"crossref","first-page":"1826","DOI":"10.1109\/CVPR.2009.5206503","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201909)","author":"Rhemann Christoph","year":"2009","unstructured":"Christoph Rhemann, Carsten Rother, Jue Wang, Margrit Gelautz, Pushmeet Kohli, and Pamela Rott. 2009. A perceptually motivated online benchmark for image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201909). 1826\u20131833."},{"key":"e_1_3_1_35_2","doi-asserted-by":"crossref","first-page":"3752","DOI":"10.1109\/CVPR.2018.00395","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918)","author":"Sankaranarayanan Swami","year":"2018","unstructured":"Swami Sankaranarayanan, Yogesh Balaji, Arpit Jain, Ser Nam Lim, and Rama Chellappa. 2018. Learning from synthetic data: Addressing domain shift for semantic segmentation. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918). 3752\u20133761."},{"key":"e_1_3_1_36_2","first-page":"2288","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201920)","author":"Sengupta Soumyadip","year":"2020","unstructured":"Soumyadip Sengupta, Vivek Jayaram, Brian Curless, Steven M. Seitz, and Ira Kemelmacher-Shlizerman. 2020. Background matting: The world is your green screen. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201920). 2288\u20132297."},{"key":"e_1_3_1_37_2","doi-asserted-by":"crossref","first-page":"636","DOI":"10.1109\/CVPR.2013.88","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201913)","author":"Shahrian Ehsan","year":"2013","unstructured":"Ehsan Shahrian, Deepu Rajan, Brian Price, and Scott Cohen. 2013. Improving image matting using comprehensive sampling sets. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201913). 636\u2013643."},{"key":"e_1_3_1_38_2","first-page":"92","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201916)","author":"Shen Xiaoyong","year":"2016","unstructured":"Xiaoyong Shen, Xin Tao, Hongyun Gao, Chao Zhou, and Jiaya Jia. 2016. Deep automatic portrait matting. In Proceedings of the European Conference on Computer Vision (ECCV\u201916). 92\u2013107."},{"key":"e_1_3_1_39_2","first-page":"11120","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921)","author":"Sun Yanan","year":"2021","unstructured":"Yanan Sun, Chi-Keung Tang, and Yu-Wing Tai. 2021. Semantic image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921). 11120\u201311129."},{"key":"e_1_3_1_40_2","first-page":"3050","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919)","author":"Tang Jingwei","year":"2019","unstructured":"Jingwei Tang, Yagiz Aksoy, Cengiz Oztireli, Markus Gross, and Tunc Ozan Aydin. 2019. Learning-based sampling for natural image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919). 3050\u20133058."},{"key":"e_1_3_1_41_2","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201922)","author":"Tian Xin","year":"2022","unstructured":"Xin Tian, Ke Xu, Xin Yang, Lin Du, Baocai Yin, and Rynson W. H. Lau. 2022. Bi-directional object-context prioritization learning for saliency ranking. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR\u201922)."},{"key":"e_1_3_1_42_2","volume-title":"Proceedings of the British Machine Vision Conference (BMVC\u201920)","author":"Tian Xin","year":"2020","unstructured":"Xin Tian, Ke Xu, Xin Yang, Baocai Yin, and Rynson W. H. Lau. 2020. Weakly-supervised salient instance detection. In Proceedings of the British Machine Vision Conference (BMVC\u201920)."},{"key":"e_1_3_1_43_2","article-title":"Learning to detect instance-level salient objects using complementary image labels","author":"Tian Xin","year":"2021","unstructured":"Xin Tian, Ke Xu, Xin Yang, Baocai Yin, and Rynson W. H. Lau. 2021. Learning to detect instance-level salient objects using complementary image labels. International Journal of Computer Vision 130 (2021), 729\u2013746.","journal-title":"International Journal of Computer Vision"},{"key":"e_1_3_1_44_2","first-page":"4777","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918)","author":"Wan Renjie","year":"2018","unstructured":"Renjie Wan, Boxin Shi, Ling-Yu Duan, Ah-Hwee Tan, and Alex C. Kot. 2018. CRRN: Multi-scale guided concurrent reflection removal network. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918). 4777\u20134785."},{"key":"e_1_3_1_45_2","first-page":"1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201907)","author":"Wang Jue","year":"2007","unstructured":"Jue Wang and Michael F. Cohen. 2007. Optimized color sampling for robust matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201907). 1\u20138."},{"key":"e_1_3_1_46_2","first-page":"999","volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI\u201918)","author":"Wang Yu","year":"2018","unstructured":"Yu Wang, Yi Niu, Peiyong Duan, Jianwei Lin, and Yuanjie Zheng. 2018. Deep propagation based image matting. In Proceedings of the International Joint Conference on Artificial Intelligence (IJCAI\u201918). 999\u20131006."},{"issue":"4","key":"e_1_3_1_47_2","doi-asserted-by":"crossref","first-page":"600","DOI":"10.1109\/TIP.2003.819861","article-title":"Image quality assessment: From error visibility to structural similarity","volume":"13","author":"Wang Zhou","year":"2004","unstructured":"Zhou Wang, Alan C. Bovik, Hamid R. Sheikh, and Eero P. Simoncelli. 2004. Image quality assessment: From error visibility to structural similarity. IEEE Transactions on Image Processing 13, 4 (2004), 600\u2013612.","journal-title":"IEEE Transactions on Image Processing"},{"key":"e_1_3_1_48_2","first-page":"15374","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921)","author":"Wei Tianyi","year":"2021","unstructured":"Tianyi Wei, Dongdong Chen, Wenbo Zhou, Jing Liao, Hanqing Zhao, Weiming Zhang, and Nenghai Yu. 2021. Improved image matting via real-time user clicks and uncertainty estimation. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921). 15374\u201315383."},{"key":"e_1_3_1_49_2","first-page":"3","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV\u201918)","author":"Woo Sanghyun","year":"2018","unstructured":"Sanghyun Woo, Jongchan Park, Joon-Young Lee, and In So Kweon. 2018. CBAM: Convolutional block attention module. In Proceedings of the European Conference on Computer Vision (ECCV\u201918). 3\u201319."},{"key":"e_1_3_1_50_2","volume-title":"Proceedings of the International Conference on Learning Representations (ICLR\u201920)","author":"Xiangli Yuanbo","year":"2020","unstructured":"Yuanbo Xiangli, Yubin Deng, Bo Dai, Chen Change Loy, and Dahua Lin. 2020. Real or not real, that is the question. In Proceedings of the International Conference on Learning Representations (ICLR\u201920)."},{"key":"e_1_3_1_51_2","first-page":"5987","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Xie Saining","year":"2017","unstructured":"Saining Xie, Ross Girshick, Piotr Dollar, Zhuowen Tu, and Kaiming He. 2017. Aggregated residual transformations for deep neural networks. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 5987\u20135995."},{"key":"e_1_3_1_52_2","first-page":"311","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917)","author":"Xu Ning","year":"2017","unstructured":"Ning Xu, Brian Price, Scott Cohen, and Thomas Huang. 2017. Deep image matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201917). 311\u2013320."},{"issue":"4","key":"e_1_3_1_53_2","first-page":"Article 121, 21","article-title":"Smart scribbles for image matting","volume":"16","author":"Yang Xin","year":"2020","unstructured":"Xin Yang, Yu Qiao, Shaozhe Chen, Shengfeng He, Baocai Yin, Qiang Zhang, Xiaopeng Wei, and Rynson W. H. Lau. 2020. Smart scribbles for image matting. ACM Transactions on Multimedia Computing Communications and Applications 16, 4 (2020), Article 121, 21 pages.","journal-title":"ACM Transactions on Multimedia Computing Communications and Applications"},{"key":"e_1_3_1_54_2","first-page":"4590","volume-title":"Proceedings of the International Conference on Neural Information Processing Systems (NeurIPS\u201918)","author":"Yang Xin","year":"2018","unstructured":"Xin Yang, Ke Xu, Shaozhe Chen, Shengfeng He, Baocai Yin Yin, and Rynson Lau. 2018. Active matting. In Proceedings of the International Conference on Neural Information Processing Systems (NeurIPS\u201918). 4590\u20134600."},{"issue":"3","key":"e_1_3_1_55_2","doi-asserted-by":"crossref","first-page":"567","DOI":"10.1111\/j.1467-8659.2006.00976.x","article-title":"Easy matting\u2014A stroke based approach for continuous image matting","volume":"25","author":"Yu Guan","year":"2006","unstructured":"Guan Yu, Wei Chen, Xiao Liang, Zi\u2019ang Ding, and Qunsheng Peng. 2006. Easy matting\u2014A stroke based approach for continuous image matting. Computer Graphics Forum 25, 3 (2006), 567\u2013576.","journal-title":"Computer Graphics Forum"},{"key":"e_1_3_1_56_2","first-page":"3217","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence (AAAI\u201921)","author":"Yu Haichao","year":"2021","unstructured":"Haichao Yu, Ning Xu, Zilong Huang, Yuqian Zhou, and Humphrey Shi. 2021. High-resolution deep image matting. In Proceedings of the AAAI Conference on Artificial Intelligence (AAAI\u201921). 3217\u20133224."},{"key":"e_1_3_1_57_2","first-page":"5505","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918)","author":"Yu Jiahui","year":"2018","unstructured":"Jiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen, Xin Lu, and Thomas S. Huang. 2018. Generative image inpainting with contextual attention. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201918). 5505\u20135514."},{"key":"e_1_3_1_58_2","first-page":"1154","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921)","author":"Yu Qihang","year":"2021","unstructured":"Qihang Yu, Jianming Zhang, He Zhang, Yilin Wang, Zhe Lin, Ning Xu, Yutong Bai, and Alan Yuille. 2021. Mask guided matting via progressive refinement network. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201921). 1154\u20131163."},{"key":"e_1_3_1_59_2","first-page":"7461","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919)","author":"Zhang Yunke","year":"2019","unstructured":"Yunke Zhang, Lixue Gong, Lubin Fan, Peiran Ren, Qixing Huang, Hujun Bao, and Weiwei Xu. 2019. A late fusion CNN for digital matting. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR\u201919). 7461\u20137470."},{"key":"e_1_3_1_60_2","first-page":"889","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201909)","author":"Zheng Yuanjie","year":"2009","unstructured":"Yuanjie Zheng and Chandra Kambhamettu. 2009. Learning based digital matting. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201909). 889\u2013896."},{"key":"e_1_3_1_61_2","first-page":"2242","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201917)","author":"Zhu Jun-Yan","year":"2017","unstructured":"Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros. 2017. Unpaired image-to-image translation using cycle-consistent adversarial networks. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV\u201917). 2242\u20132251."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3540201","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3540201","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:51:01Z","timestamp":1750182661000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3540201"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,2,6]]},"references-count":60,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,5,31]]}},"alternative-id":["10.1145\/3540201"],"URL":"https:\/\/doi.org\/10.1145\/3540201","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,2,6]]},"assertion":[{"value":"2021-08-29","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-05-23","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-02-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}