{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,29]],"date-time":"2026-05-29T14:31:35Z","timestamp":1780065095854,"version":"3.54.0"},"reference-count":76,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2023,2,6]],"date-time":"2023-02-06T00:00:00Z","timestamp":1675641600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62022075 and 62021001"],"award-info":[{"award-number":["62022075 and 62021001"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"crossref","award":["WK3490000006"],"award-info":[{"award-number":["WK3490000006"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2023,5,31]]},"abstract":"<jats:p>The capability of image semantic segmentation may be deteriorated due to the noisy input image, where image denoising prior to segmentation may help. Both image denoising and semantic segmentation have been developed significantly with the advance of deep learning. In this work, we are interested in the synergy between these two tasks by using a holistic deep model. We observe that not only denoising helps combat the drop of segmentation accuracy due to the noisy input, but also pixel-wise semantic information boosts the capability of denoising. We then propose a boosting network to perform denoising and segmentation alternately. The proposed network is composed of multiple segmentation and denoising blocks (SDBs), each of which estimates a semantic map and then uses the map to regularize denoising. Experimental results show that the denoised image quality is improved substantially and the segmentation accuracy is improved to close to that on clean images, and segmentation and denoising are both boosted as the number of SDBs increases. On the Cityscapes dataset, using three SDBs improves the denoising quality to 34.42 dB in PSNR, and the segmentation accuracy to 66.5 in mIoU, when the additive white Gaussian noise level is 50.<\/jats:p>","DOI":"10.1145\/3548459","type":"journal-article","created":{"date-parts":[[2022,7,14]],"date-time":"2022-07-14T11:17:31Z","timestamp":1657797451000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["Synergy between Semantic Segmentation and Image Denoising via Alternate Boosting"],"prefix":"10.1145","volume":"19","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7342-0124","authenticated-orcid":false,"given":"Shunxin","family":"Xu","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, Hefei, Anhui, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6742-3484","authenticated-orcid":false,"given":"Ke","family":"Sun","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, Anhui, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9100-2906","authenticated-orcid":false,"given":"Dong","family":"Liu","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, Anhui, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9787-7460","authenticated-orcid":false,"given":"Zhiwei","family":"Xiong","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, Anhui, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2510-8993","authenticated-orcid":false,"given":"Zheng-Jun","family":"Zha","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, Anhui, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,2,6]]},"reference":[{"key":"e_1_3_1_2_2","first-page":"1692","volume-title":"CVPR","author":"Abdelhamed Abdelrahman","year":"2018","unstructured":"Abdelrahman Abdelhamed, Stephen Lin, and Michael S. Brown. 2018. A high-quality denoising dataset for smartphone cameras. In CVPR. IEEE, 1692\u20131700."},{"key":"e_1_3_1_3_2","article-title":"Unsupervised hyperspectral stimulated Raman microscopy image enhancement: Denoising and segmentation via one-shot deep learning","author":"Abdolghader Pedram","year":"2021","unstructured":"Pedram Abdolghader, Andrew Ridsdale, Tassos Grammatikopoulos, Gavin Resch, Francois Legare, Albert Stolow, Adrian F. Pegoraro, and Isaac Tamblyn. 2021. Unsupervised hyperspectral stimulated Raman microscopy image enhancement: Denoising and segmentation via one-shot deep learning. arXiv preprint arXiv:2104.08338 (2021).","journal-title":"arXiv preprint arXiv:2104.08338"},{"key":"e_1_3_1_4_2","first-page":"3155","volume-title":"ICCV","author":"Anwar Saeed","year":"2019","unstructured":"Saeed Anwar and Nick Barnes. 2019. Real image denoising with feature attention. In ICCV. IEEE, 3155\u20133164."},{"key":"e_1_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2017.2733739"},{"key":"e_1_3_1_6_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2010.161"},{"key":"e_1_3_1_7_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2016.2644615"},{"key":"e_1_3_1_8_2","first-page":"324","volume-title":"ECCV","author":"Buchholz Tim-Oliver","year":"2020","unstructured":"Tim-Oliver Buchholz, Mangal Prakash, Deborah Schmidt, Alexander Krull, and Florian Jug. 2020. DenoiSeg: Joint denoising and segmentation. In ECCV. Springer, 324\u2013337."},{"key":"e_1_3_1_9_2","first-page":"452","volume-title":"CISS","author":"Charest Michael R.","year":"2006","unstructured":"Michael R. Charest, Michael Elad, and Peyman Milanfar. 2006. A general iterative regularization framework for image denoising. In CISS. IEEE, 452\u2013457."},{"issue":"12","key":"e_1_3_1_10_2","doi-asserted-by":"crossref","first-page":"3071","DOI":"10.1109\/TPAMI.2019.2921548","article-title":"Real-world image denoising with deep boosting","volume":"42","author":"Chen Chang","year":"2019","unstructured":"Chang Chen, Zhiwei Xiong, Xinmei Tian, Zheng-Jun Zha, and Feng Wu. 2019. Real-world image denoising with deep boosting. IEEE Trans. Patt. Anal. Mach. Intell. 42, 12 (2019), 3071\u20133087.","journal-title":"IEEE Trans. Patt. Anal. Mach. Intell."},{"key":"e_1_3_1_11_2","first-page":"182","volume-title":"CVPR","author":"Chen Liangyu","year":"2021","unstructured":"Liangyu Chen, Xin Lu, Jie Zhang, Xiaojie Chu, and Chengpeng Chen. 2021. HINet: Half instance normalization network for image restoration. In CVPR. 182\u2013192."},{"key":"e_1_3_1_12_2","article-title":"Semantic image segmentation with deep convolutional nets and fully connected CRFs","author":"Chen Liang-Chieh","year":"2014","unstructured":"Liang-Chieh Chen, George Papandreou, Iasonas Kokkinos, Kevin Murphy, and Alan L. Yuille. 2014. Semantic image segmentation with deep convolutional nets and fully connected CRFs. arXiv preprint arXiv:1412.7062 (2014).","journal-title":"arXiv preprint arXiv:1412.7062"},{"key":"e_1_3_1_13_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2017.2699184"},{"key":"e_1_3_1_14_2","first-page":"3213","volume-title":"CVPR","author":"Cordts Marius","year":"2016","unstructured":"Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele. 2016. The cityscapes dataset for semantic urban scene understanding. In CVPR. IEEE, 3213\u20133223."},{"key":"e_1_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2007.901238"},{"key":"e_1_3_1_16_2","first-page":"457","volume-title":"CVPR","author":"Dong Weisheng","year":"2011","unstructured":"Weisheng Dong, Xin Li, Lei Zhang, and Guangming Shi. 2011. Sparsity-based image denoising via dictionary learning and structural clustering. In CVPR. IEEE, 457\u2013464."},{"key":"e_1_3_1_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2006.881969"},{"key":"e_1_3_1_18_2","first-page":"249","volume-title":"Proceedings of the 13th International Conference on Artificial Intelligence and Statistics","author":"Glorot Xavier","year":"2010","unstructured":"Xavier Glorot and Yoshua Bengio. 2010. Understanding the difficulty of training deep feedforward neural networks. In Proceedings of the 13th International Conference on Artificial Intelligence and Statistics. JMLR Workshop and Conference Proceedings, 249\u2013256."},{"key":"e_1_3_1_19_2","first-page":"1712","volume-title":"CVPR","author":"Guo Shi","year":"2019","unstructured":"Shi Guo, Zifei Yan, Kai Zhang, Wangmeng Zuo, and Lei Zhang. 2019. Toward convolutional blind denoising of real photographs. In CVPR. IEEE, 1712\u20131722."},{"key":"e_1_3_1_20_2","first-page":"109","volume-title":"NIPS","author":"Han Shizhong","year":"2016","unstructured":"Shizhong Han, Zibo Meng, Ahmed-Shehab Khan, and Yan Tong. 2016. Incremental boosting convolutional neural network for facial action unit recognition. In NIPS. 109\u2013117."},{"key":"e_1_3_1_21_2","first-page":"1026","volume-title":"ICCV","author":"He Kaiming","year":"2015","unstructured":"Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015. Delving deep into rectifiers: Surpassing human-level performance on ImageNet classification. In ICCV. 1026\u20131034."},{"key":"e_1_3_1_22_2","first-page":"770","volume-title":"CVPR","author":"He Kaiming","year":"2016","unstructured":"Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016. Deep residual learning for image recognition. In CVPR. IEEE, 770\u2013778."},{"key":"e_1_3_1_23_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2015.2494461"},{"key":"e_1_3_1_24_2","first-page":"448","volume-title":"ICML","author":"Ioffe Sergey","year":"2015","unstructured":"Sergey Ioffe and Christian Szegedy. 2015. Batch normalization: Accelerating deep network training by reducing internal covariate shift. In ICML. 448\u2013456."},{"key":"e_1_3_1_25_2","first-page":"694","volume-title":"ECCV","author":"Johnson Justin","year":"2016","unstructured":"Justin Johnson, Alexandre Alahi, and Li Fei-Fei. 2016. Perceptual losses for real-time style transfer and super-resolution. In ECCV. Springer, 694\u2013711."},{"key":"e_1_3_1_26_2","first-page":"3482","volume-title":"CVPR","author":"Kim Yoonsik","year":"2020","unstructured":"Yoonsik Kim, Jae Woong Soh, Gu Yong Park, and Nam Ik Cho. 2020. Transfer learning from synthetic to real-noise denoising with adaptive instance normalization. In CVPR. IEEE, 3482\u20133492."},{"key":"e_1_3_1_27_2","article-title":"Adam: A method for stochastic optimization","author":"Kingma Diederik P.","year":"2014","unstructured":"Diederik P. Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014).","journal-title":"arXiv preprint arXiv:1412.6980"},{"key":"e_1_3_1_28_2","first-page":"2129","volume-title":"CVPR","author":"Krull Alexander","year":"2019","unstructured":"Alexander Krull, Tim-Oliver Buchholz, and Florian Jug. 2019. Noise2Void-learning denoising from single noisy images. In CVPR. IEEE, 2129\u20132137."},{"key":"e_1_3_1_29_2","first-page":"585","volume-title":"MICCAI","author":"Larrazabal Agostina J.","year":"2019","unstructured":"Agostina J. Larrazabal, Cesar Martinez, and Enzo Ferrante. 2019. Anatomical priors for image segmentation via post-processing with denoising autoencoders. In MICCAI. Springer, 585\u2013593."},{"key":"e_1_3_1_30_2","doi-asserted-by":"publisher","DOI":"10.14419\/ijet.v7i2.3.9964"},{"key":"e_1_3_1_31_2","first-page":"2965","volume-title":"ICML","author":"Lehtinen Jaakko","year":"2018","unstructured":"Jaakko Lehtinen, Jacob Munkberg, Jon Hasselgren, Samuli Laine, Tero Karras, Miika Aittala, and Timo Aila. 2018. Noise2Noise: Learning image restoration without clean data. In ICML. 2965\u20132974."},{"key":"e_1_3_1_32_2","first-page":"1833","volume-title":"ICCV","author":"Liang Jingyun","year":"2021","unstructured":"Jingyun Liang, Jiezhang Cao, Guolei Sun, Kai Zhang, Luc Van Gool, and Radu Timofte. 2021. SwinIR: Image restoration using Swin transformer. In ICCV. 1833\u20131844."},{"key":"e_1_3_1_33_2","first-page":"1925","volume-title":"CVPR","author":"Lin Guosheng","year":"2017","unstructured":"Guosheng Lin, Anton Milan, Chunhua Shen, and Ian Reid. 2017. RefineNet: Multi-path refinement networks for high-resolution semantic segmentation. In CVPR. IEEE, 1925\u20131934."},{"key":"e_1_3_1_34_2","first-page":"740","volume-title":"ECCV","author":"Lin Tsung-Yi","year":"2014","unstructured":"Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll\u00e1r, and C. Lawrence Zitnick. 2014. Microsoft COCO: Common objects in context. In ECCV. Springer, 740\u2013755."},{"key":"e_1_3_1_35_2","article-title":"Learning degraded image classification with restoration data fidelity","author":"Lin Xiaoyu","year":"2021","unstructured":"Xiaoyu Lin. 2021. Learning degraded image classification with restoration data fidelity. arXiv preprint arXiv:2101.09606 (2021).","journal-title":"arXiv preprint arXiv:2101.09606"},{"key":"e_1_3_1_36_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2020.2964518"},{"key":"e_1_3_1_37_2","first-page":"842","volume-title":"IJCAI","author":"Liu Ding","year":"2018","unstructured":"Ding Liu, Bihan Wen, Xianming Liu, Zhangyang Wang, and Thomas S. Huang. 2018. When image denoising meets high-level vision tasks: A deep learning approach. In IJCAI. 842\u2013848."},{"key":"e_1_3_1_38_2","first-page":"10012","volume-title":"ICCV","author":"Liu Ze","year":"2021","unstructured":"Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo. 2021. Swin transformer: Hierarchical vision transformer using shifted windows. In ICCV. 10012\u201310022."},{"key":"e_1_3_1_39_2","first-page":"3431","volume-title":"CVPR","author":"Long Jonathan","year":"2015","unstructured":"Jonathan Long, Evan Shelhamer, and Trevor Darrell. 2015. Fully convolutional networks for semantic segmentation. In CVPR. IEEE, 3431\u20133440."},{"key":"e_1_3_1_40_2","first-page":"2272","volume-title":"ICCV","author":"Mairal Julien","year":"2009","unstructured":"Julien Mairal, Francis Bach, Jean Ponce, Guillermo Sapiro, and Andrew Zisserman. 2009. Non-local sparse models for image restoration. In ICCV. IEEE, 2272\u20132279."},{"key":"e_1_3_1_41_2","first-page":"2802","volume-title":"NIPS","author":"Mao Xiaojiao","year":"2016","unstructured":"Xiaojiao Mao, Chunhua Shen, and Yu-Bin Yang. 2016. Image restoration using very deep convolutional encoder-decoder networks with symmetric skip connections. In NIPS. 2802\u20132810."},{"key":"e_1_3_1_42_2","first-page":"1","volume-title":"BMVC","author":"Moghimi Mohammad","year":"2016","unstructured":"Mohammad Moghimi, Serge J. Belongie, Mohammad J. Saberian, Jian Yang, Nuno Vasconcelos, and Li-Jia Li. 2016. Boosted convolutional neural networks. In BMVC. 1\u20136."},{"key":"e_1_3_1_43_2","first-page":"8024","volume-title":"NeurIPS","author":"Paszke Adam","year":"2019","unstructured":"Adam Paszke, Sam Gross, Francisco Massa, et\u00a0al. 2019. PyTorch: An imperative style, high-performance deep learning library. In NeurIPS. 8024\u20138035."},{"key":"e_1_3_1_44_2","first-page":"1586","volume-title":"CVPR","author":"Plotz Tobias","year":"2017","unstructured":"Tobias Plotz and Stefan Roth. 2017. Benchmarking denoising algorithms with real photographs. In CVPR. IEEE, 1586\u20131595."},{"key":"e_1_3_1_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2859044"},{"key":"e_1_3_1_46_2","first-page":"1077","volume-title":"ICCV","author":"Ren Wenqi","year":"2017","unstructured":"Wenqi Ren, Jinshan Pan, Xiaochun Cao, and Ming-Hsuan Yang. 2017. Video deblurring via semantic segmentation and pixel-wise non-linear kernel. In ICCV. IEEE, 1077\u20131085."},{"key":"e_1_3_1_47_2","first-page":"234","volume-title":"MICCAI","author":"Ronneberger Olaf","year":"2015","unstructured":"Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015. U-Net: Convolutional networks for biomedical image segmentation. In MICCAI. Springer, 234\u2013241."},{"key":"e_1_3_1_48_2","unstructured":"Jie Shao Kai Hu Changhu Wang Xiangyang Xue and Bhiksha Raj. 2020. Is normalization indispensable for training deep neural network? In NeurIPS . 13434\u201313444."},{"key":"e_1_3_1_49_2","first-page":"4033","volume-title":"CVPR","author":"Sharma Vivek","year":"2018","unstructured":"Vivek Sharma, Ali Diba, Davy Neven, Michael S. Brown, Luc Van Gool, and Rainer Stiefelhagen. 2018. Classification-driven dynamic image enhancement. In CVPR. IEEE, 4033\u20134041."},{"key":"e_1_3_1_50_2","article-title":"Very deep convolutional networks for large-scale image recognition","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very deep convolutional networks for large-scale image recognition. arXiv preprint arXiv:1409.1556 (2014).","journal-title":"arXiv preprint arXiv:1409.1556"},{"key":"e_1_3_1_51_2","first-page":"372","volume-title":"ICIP","author":"Singh Maneesh","year":"1999","unstructured":"Maneesh Singh, Prakash Ishwar, Krishna Ratakonda, and Narendra Ahuja. 1999. Segmentation based denoising using multiple compaction domains. In ICIP. IEEE, 372\u2013375."},{"key":"e_1_3_1_52_2","first-page":"7262","volume-title":"ICCV","author":"Strudel Robin","year":"2021","unstructured":"Robin Strudel, Ricardo Garcia, Ivan Laptev, and Cordelia Schmid. 2021. Segmenter: Transformer for semantic segmentation. In ICCV. 7262\u20137272."},{"issue":"4","key":"e_1_3_1_53_2","doi-asserted-by":"crossref","first-page":"1470","DOI":"10.1109\/TIP.2012.2231691","article-title":"How to SAIF-ly boost denoising performance","volume":"22","author":"Talebi Hossein","year":"2012","unstructured":"Hossein Talebi, Xiang Zhu, and Peyman Milanfar. 2012. How to SAIF-ly boost denoising performance. IEEE Trans. Image Process. 22, 4 (2012), 1470\u20131485.","journal-title":"IEEE Trans. Image Process."},{"key":"e_1_3_1_54_2","first-page":"9446","volume-title":"CVPR","author":"Ulyanov Dmitry","year":"2018","unstructured":"Dmitry Ulyanov, Andrea Vedaldi, and Victor Lempitsky. 2018. Deep image prior. In CVPR. IEEE, 9446\u20139454."},{"key":"e_1_3_1_55_2","article-title":"Attention is all you need","volume":"30","author":"Vaswani Ashish","year":"2017","unstructured":"Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, \u0141ukasz Kaiser, and Illia Polosukhin. 2017. Attention is all you need. Adv. Neural Inf. Process. 30 (2017).","journal-title":"Adv. Neural Inf. Process."},{"key":"e_1_3_1_56_2","first-page":"561","article-title":"Denoising and segmentation of 3D brain images.","volume":"9","author":"Vatsa Mayank","year":"2009","unstructured":"Mayank Vatsa, Richa Singh, and Afzel Noore. 2009. Denoising and segmentation of 3D brain images.Image Process. Comput. Vis. Patt. Recog. 9 (2009), 561\u2013567.","journal-title":"Image Process. Comput. Vis. Patt. Recog."},{"key":"e_1_3_1_57_2","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2020.2983686"},{"key":"e_1_3_1_58_2","first-page":"3774","volume-title":"CVPR","author":"Wang Li","year":"2020","unstructured":"Li Wang, Dong Li, Yousong Zhu, Lu Tian, and Yi Shan. 2020. Dual super-resolution learning for semantic segmentation. In CVPR. IEEE, 3774\u20133783."},{"key":"e_1_3_1_59_2","article-title":"Segmentation-aware image denoising without knowing true segmentation","author":"Wang Sicheng","year":"2019","unstructured":"Sicheng Wang, Bihan Wen, Junru Wu, Dacheng Tao, and Zhangyang Wang. 2019. Segmentation-aware image denoising without knowing true segmentation. arXiv preprint arXiv:1905.08965 (2019).","journal-title":"arXiv preprint arXiv:1905.08965"},{"key":"e_1_3_1_60_2","first-page":"568","volume-title":"ICCV","author":"Wang Wenhai","year":"2021","unstructured":"Wenhai Wang, Enze Xie, Xiang Li, Deng-Ping Fan, Kaitao Song, Ding Liang, Tong Lu, Ping Luo, and Ling Shao. 2021. Pyramid vision transformer: A versatile backbone for dense prediction without convolutions. In ICCV. 568\u2013578."},{"key":"e_1_3_1_61_2","first-page":"606","volume-title":"CVPR","author":"Wang Xintao","year":"2018","unstructured":"Xintao Wang, Ke Yu, Chao Dong, and Chen Change Loy. 2018. Recovering realistic texture in image super-resolution by deep spatial feature transform. In CVPR. IEEE, 606\u2013615."},{"key":"e_1_3_1_62_2","article-title":"Uformer: A general u-shaped transformer for image restoration","author":"Wang Zhendong","year":"2021","unstructured":"Zhendong Wang, Xiaodong Cun, Jianmin Bao, and Jianzhuang Liu. 2021. Uformer: A general u-shaped transformer for image restoration. arXiv preprint arXiv:2106.03106 (2021).","journal-title":"arXiv preprint arXiv:2106.03106"},{"key":"e_1_3_1_63_2","first-page":"698","volume-title":"MICCAI","author":"Xu Ziyue","year":"2014","unstructured":"Ziyue Xu, Ulas Bagci, Jurgen Seidel, David Thomasson, Jeff Solomon, and Daniel J. Mollura. 2014. Segmentation based denoising of PET images: An iterative approach via regional means and affinity propagation. In MICCAI. Springer, 698\u2013705."},{"key":"e_1_3_1_64_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2018.03.007"},{"key":"e_1_3_1_65_2","first-page":"3684","volume-title":"CVPR","author":"Yang Maoke","year":"2018","unstructured":"Maoke Yang, Kun Yu, Chi Zhang, Zhiwei Li, and Kuiyuan Yang. 2018. DenseASPP for semantic segmentation in street scenes. In CVPR. IEEE, 3684\u20133692."},{"key":"e_1_3_1_66_2","article-title":"Multi-scale context aggregation by dilated convolutions","author":"Yu Fisher","year":"2015","unstructured":"Fisher Yu and Vladlen Koltun. 2015. Multi-scale context aggregation by dilated convolutions. arXiv preprint arXiv:1511.07122 (2015).","journal-title":"arXiv preprint arXiv:1511.07122"},{"key":"e_1_3_1_67_2","first-page":"173","volume-title":"ECCV","author":"Yuan Yuhui","year":"2020","unstructured":"Yuhui Yuan, Xilin Chen, and Jingdong Wang. 2020. Object-contextual representations for semantic segmentation. In ECCV. Springer, 173\u2013190."},{"key":"e_1_3_1_68_2","article-title":"Restormer: Efficient transformer for high-resolution image restoration","author":"Zamir Syed Waqas","year":"2021","unstructured":"Syed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat, Fahad Shahbaz Khan, and Ming-Hsuan Yang. 2021. Restormer: Efficient transformer for high-resolution image restoration. arXiv preprint arXiv:2111.09881 (2021).","journal-title":"arXiv preprint arXiv:2111.09881"},{"key":"e_1_3_1_69_2","first-page":"14821","volume-title":"CVPR","author":"Zamir Syed Waqas","year":"2021","unstructured":"Syed Waqas Zamir, Aditya Arora, Salman Khan, Munawar Hayat, Fahad Shahbaz Khan, Ming-Hsuan Yang, and Ling Shao. 2021. Multi-stage progressive image restoration. In CVPR. 14821\u201314831."},{"key":"e_1_3_1_70_2","first-page":"8799","volume-title":"ICCV","author":"Zhang Haochen","year":"2019","unstructured":"Haochen Zhang, Dong Liu, and Zhiwei Xiong. 2019. Two-stream action recognition-oriented video super-resolution. In ICCV. IEEE, 8799\u20138808."},{"key":"e_1_3_1_71_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2017.2662206"},{"key":"e_1_3_1_72_2","first-page":"3929","volume-title":"CVPR","author":"Zhang Kai","year":"2017","unstructured":"Kai Zhang, Wangmeng Zuo, Shuhang Gu, and Lei Zhang. 2017. Learning deep CNN denoiser prior for image restoration. In CVPR. IEEE, 3929\u20133938."},{"key":"e_1_3_1_73_2","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2839891"},{"key":"e_1_3_1_74_2","first-page":"235","volume-title":"ECCV","author":"Zhang Zhenyu","year":"2018","unstructured":"Zhenyu Zhang, Zhen Cui, Chunyan Xu, Zequn Jie, Xiang Li, and Jian Yang. 2018. Joint task-recursive learning for semantic segmentation and depth estimation. In ECCV. Springer, 235\u2013251."},{"key":"e_1_3_1_75_2","first-page":"2881","volume-title":"CVPR","author":"Zhao Hengshuang","year":"2017","unstructured":"Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia. 2017. Pyramid scene parsing network. In CVPR. IEEE, 2881\u20132890."},{"key":"e_1_3_1_76_2","first-page":"6881","volume-title":"CVPR","author":"Zheng Sixiao","year":"2021","unstructured":"Sixiao Zheng, Jiachen Lu, Hengshuang Zhao, Xiatian Zhu, Zekun Luo, Yabiao Wang, Yanwei Fu, Jianfeng Feng, Tao Xiang, Philip H. S. Torr, et\u00a0al. 2021. Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers. In CVPR. 6881\u20136890."},{"key":"e_1_3_1_77_2","first-page":"633","volume-title":"CVPR","author":"Zhou Bolei","year":"2017","unstructured":"Bolei Zhou, Hang Zhao, Xavier Puig, Sanja Fidler, Adela Barriuso, and Antonio Torralba. 2017. Scene parsing through ADE20K dataset. In CVPR. IEEE, 633\u2013641."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3548459","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3548459","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T18:43:29Z","timestamp":1750272209000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3548459"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,2,6]]},"references-count":76,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,5,31]]}},"alternative-id":["10.1145\/3548459"],"URL":"https:\/\/doi.org\/10.1145\/3548459","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,2,6]]},"assertion":[{"value":"2021-12-13","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-07-07","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-02-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}