{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:13:07Z","timestamp":1750219987252,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":32,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,11,17]],"date-time":"2022-11-17T00:00:00Z","timestamp":1668643200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,11,17]]},"DOI":"10.1145\/3581807.3581827","type":"proceedings-article","created":{"date-parts":[[2023,5,23]],"date-time":"2023-05-23T00:02:28Z","timestamp":1684800148000},"page":"136-141","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Swin Transformer with Multi-Scale Residual Attention for Semantic Segmentation of Remote Sensing Images"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-2959-6018","authenticated-orcid":false,"given":"Yuanyang","family":"Lin","sequence":"first","affiliation":[{"name":"Fujian Key Laboratory of Pattern Recognition and Image Understanding, Xiamen University of Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5901-0778","authenticated-orcid":false,"given":"Da-han","family":"Wang","sequence":"additional","affiliation":[{"name":"Fujian Key Laboratory of Pattern Recognition and Image Understanding, Xiamen University of Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2476-7746","authenticated-orcid":false,"given":"Yun","family":"Wu","sequence":"additional","affiliation":[{"name":"Fujian Key Laboratory of Pattern Recognition and Image Understanding, Xiamen University of Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9715-4281","authenticated-orcid":false,"given":"Shunzhi","family":"Zhu","sequence":"additional","affiliation":[{"name":"Fujian Key Laboratory of Pattern Recognition and Image Understanding, Xiamen University of Technology, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2023,5,22]]},"reference":[{"key":"e_1_3_2_1_1_1","first-page":"1","article-title":"The ISPRS benchmark on urban object classification and 3D building reconstruction","volume":"3","author":"Rottensteiner Franz","year":"2012","unstructured":"Rottensteiner , Franz , \" The ISPRS benchmark on urban object classification and 3D building reconstruction .\" \u00a0ISPRS Annals of the Photogrammetry, Remote Sensing and Spatial Information Sciences I-3 ( 2012 ), Nr. 1\u00a0 1 .1 (2012): 293-298. Rottensteiner, Franz, \"The ISPRS benchmark on urban object classification and 3D building reconstruction.\"\u00a0ISPRS Annals of the Photogrammetry, Remote Sensing and Spatial Information Sciences I-3 (2012), Nr. 1\u00a01.1 (2012): 293-298.","journal-title":"\u00a0ISPRS Annals of the Photogrammetry, Remote Sensing and Spatial Information Sciences"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"crossref","unstructured":"Volpi Michele and Vittorio Ferrari. \"Semantic segmentation of urban scenes by learning local class interactions.\"\u00a0Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 2015.  Volpi Michele and Vittorio Ferrari. \"Semantic segmentation of urban scenes by learning local class interactions.\"\u00a0Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops. 2015.","DOI":"10.1109\/CVPRW.2015.7301377"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"crossref","unstructured":"Kampffmeyer Michael Arnt-Borre Salberg and Robert Jenssen. \"Semantic segmentation of small objects and modeling of uncertainty in urban remote sensing images using deep convolutional neural networks.\"\u00a0Proceedings of the IEEE conference on computer vision and pattern recognition workshops. 2016.  Kampffmeyer Michael Arnt-Borre Salberg and Robert Jenssen. \"Semantic segmentation of small objects and modeling of uncertainty in urban remote sensing images using deep convolutional neural networks.\"\u00a0Proceedings of the IEEE conference on computer vision and pattern recognition workshops. 2016.","DOI":"10.1109\/CVPRW.2016.90"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"crossref","first-page":"3511","DOI":"10.1109\/TGRS.2010.2047260","article-title":"\"Using aerial imagery and GIS in automated building footprint extraction and shape recognition for earthquake risk assessment of urban inventories","volume":"9","author":"Sahar Liora","year":"2010","unstructured":"Sahar , Liora , Subrahmanyam Muthukumar , and Steven P. French . \"Using aerial imagery and GIS in automated building footprint extraction and shape recognition for earthquake risk assessment of urban inventories .\" \u00a0IEEE Transactions on Geoscience and Remote Sensing\u00a048 . 9 ( 2010 ): 3511 - 3520 . Sahar, Liora, Subrahmanyam Muthukumar, and Steven P. French. \"Using aerial imagery and GIS in automated building footprint extraction and shape recognition for earthquake risk assessment of urban inventories.\"\u00a0IEEE Transactions on Geoscience and Remote Sensing\u00a048.9 (2010): 3511-3520.","journal-title":"\u00a0IEEE Transactions on Geoscience and Remote Sensing\u00a048"},{"key":"e_1_3_2_1_5_1","volume-title":"106971","author":"Liu Ganchao","year":"2019","unstructured":"Liu , Ganchao , \" Stacked Fisher autoencoder for SAR change detection.\"\u00a0Pattern Recognition\u00a096 ( 2019 ): 106971 . Liu, Ganchao, \"Stacked Fisher autoencoder for SAR change detection.\"\u00a0Pattern Recognition\u00a096 (2019): 106971."},{"issue":"5","key":"e_1_3_2_1_6_1","doi-asserted-by":"crossref","first-page":"901","DOI":"10.3390\/rs13050901","article-title":"Crop row segmentation and detection in paddy fields based on treble-classification otsu and double-dimensional clustering method","volume":"13","author":"Yu Yue","year":"2021","unstructured":"Yu , Yue , \" Crop row segmentation and detection in paddy fields based on treble-classification otsu and double-dimensional clustering method .\" Remote Sensing 13 . 5 ( 2021 ): 901 . Sheikh, Rasha, \"Gradient and log-based active learning for semantic segmentation of crop and weed for agricultural robots.\"\u00a02020 IEEE International Conference on Robotics and Automation (ICRA). IEEE, 2020. Yu, Yue, \"Crop row segmentation and detection in paddy fields based on treble-classification otsu and double-dimensional clustering method.\" Remote Sensing 13.5 (2021): 901. Sheikh, Rasha, \"Gradient and log-based active learning for semantic segmentation of crop and weed for agricultural robots.\"\u00a02020 IEEE International Conference on Robotics and Automation (ICRA). IEEE, 2020.","journal-title":"Remote Sensing"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"crossref","unstructured":"Long Jonathan Evan Shelhamer and Trevor Darrell. \"Fully convolutional networks for semantic segmentation.\"\u00a0Proceedings of the IEEE conference on computer vision and pattern recognition. 2015.  Long Jonathan Evan Shelhamer and Trevor Darrell. \"Fully convolutional networks for semantic segmentation.\"\u00a0Proceedings of the IEEE conference on computer vision and pattern recognition. 2015.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"e_1_3_2_1_8_1","volume-title":"Convolutional networks for biomedical image segmentation.\" International Conference on Medical image computing and computer-assisted intervention","author":"Ronneberger Olaf","year":"2015","unstructured":"Ronneberger , Olaf , Philipp Fischer , and Thomas Brox . \"U-net : Convolutional networks for biomedical image segmentation.\" International Conference on Medical image computing and computer-assisted intervention . Springer , Cham , 2015 . Ronneberger, Olaf, Philipp Fischer, and Thomas Brox. \"U-net: Convolutional networks for biomedical image segmentation.\" International Conference on Medical image computing and computer-assisted intervention. Springer, Cham, 2015."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"crossref","unstructured":"Chen Liang-Chieh \"Deeplab: Semantic image segmentation with deep convolutional nets atrous convolution and fully connected crfs.\" IEEE transactions on pattern analysis and machine intelligence 40.4 (2017):  834-848.  Chen Liang-Chieh \"Deeplab: Semantic image segmentation with deep convolutional nets atrous convolution and fully connected crfs.\" IEEE transactions on pattern analysis and machine intelligence 40.4 (2017): 834-848.","DOI":"10.1109\/TPAMI.2017.2699184"},{"key":"e_1_3_2_1_10_1","unstructured":"Chen Liang-Chieh \"Rethinking atrous convolution for semantic image segmentation.\" arXiv preprint arXiv:1706.05587 (2017).  Chen Liang-Chieh \"Rethinking atrous convolution for semantic image segmentation.\" arXiv preprint arXiv:1706.05587 (2017)."},{"key":"e_1_3_2_1_11_1","volume-title":"Transformers for image recognition at scale.\" arXiv preprint arXiv:2010.11929","author":"Dosovitskiy Alexey","year":"2020","unstructured":"Dosovitskiy , Alexey , \"An image is worth 16x16 words : Transformers for image recognition at scale.\" arXiv preprint arXiv:2010.11929 ( 2020 ). Dosovitskiy, Alexey, \"An image is worth 16x16 words: Transformers for image recognition at scale.\" arXiv preprint arXiv:2010.11929 (2020)."},{"key":"e_1_3_2_1_12_1","unstructured":"Vaswani Ashish \"Attention is all you need.\" Advances in neural information processing systems 30 (2017).  Vaswani Ashish \"Attention is all you need.\" Advances in neural information processing systems 30 (2017)."},{"key":"e_1_3_2_1_13_1","volume-title":"Hierarchical vision transformer using shifted windows.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision","author":"Liu Ze","year":"2021","unstructured":"Liu , Ze , \"Swin transformer : Hierarchical vision transformer using shifted windows.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision . 2021 . Liu, Ze, \"Swin transformer: Hierarchical vision transformer using shifted windows.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision. 2021."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"crossref","unstructured":"Zheng Sixiao \"Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers.\" Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2021.  Zheng Sixiao \"Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers.\" Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2021.","DOI":"10.1109\/CVPR46437.2021.00681"},{"issue":"4","key":"e_1_3_2_1_15_1","first-page":"637","article-title":"PEGNet: Progressive edge guidance network for semantic segmentation of remote sensing images","volume":"18","author":"Pan Shaoming","year":"2020","unstructured":"Pan , Shaoming , \" PEGNet: Progressive edge guidance network for semantic segmentation of remote sensing images .\" IEEE Geoscience and Remote Sensing Letters 18 . 4 ( 2020 ): 637 - 641 . Pan, Shaoming, \"PEGNet: Progressive edge guidance network for semantic segmentation of remote sensing images.\" IEEE Geoscience and Remote Sensing Letters 18.4 (2020): 637-641.","journal-title":"IEEE Geoscience and Remote Sensing Letters"},{"key":"e_1_3_2_1_16_1","first-page":"1","article-title":"Progressive Guidance Edge Perception Network for Semantic Segmentation of Remote-Sensing Images","volume":"19","author":"Pan Shaoming","year":"2021","unstructured":"Pan , Shaoming , \" Progressive Guidance Edge Perception Network for Semantic Segmentation of Remote-Sensing Images .\" IEEE Geoscience and Remote Sensing Letters 19 ( 2021 ): 1 - 5 . Pan, Shaoming, \"Progressive Guidance Edge Perception Network for Semantic Segmentation of Remote-Sensing Images.\" IEEE Geoscience and Remote Sensing Letters 19 (2021): 1-5.","journal-title":"IEEE Geoscience and Remote Sensing Letters"},{"key":"e_1_3_2_1_17_1","first-page":"1","article-title":"Factseg: Foreground activation-driven small object semantic segmentation in large-scale remote sensing imagery","volume":"60","author":"Ma Ailong","year":"2021","unstructured":"Ma , Ailong , \" Factseg: Foreground activation-driven small object semantic segmentation in large-scale remote sensing imagery .\" IEEE Transactions on Geoscience and Remote Sensing 60 ( 2021 ): 1 - 16 . Ma, Ailong, \"Factseg: Foreground activation-driven small object semantic segmentation in large-scale remote sensing imagery.\" IEEE Transactions on Geoscience and Remote Sensing 60 (2021): 1-16.","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"crossref","unstructured":"Zhao Hengshuang \"Pyramid scene parsing network.\" Proceedings of the IEEE conference on computer vision and pattern recognition. 2017.  Zhao Hengshuang \"Pyramid scene parsing network.\" Proceedings of the IEEE conference on computer vision and pattern recognition. 2017.","DOI":"10.1109\/CVPR.2017.660"},{"key":"e_1_3_2_1_19_1","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition.","author":"Wang Xiaolong","year":"2018","unstructured":"Wang , Xiaolong , \"Non-local neural networks.\" Proceedings of the IEEE conference on computer vision and pattern recognition. 2018 . Wang, Xiaolong, \"Non-local neural networks.\" Proceedings of the IEEE conference on computer vision and pattern recognition. 2018."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"crossref","unstructured":"Fu Jun \"Dual attention network for scene segmentation.\" Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2019.  Fu Jun \"Dual attention network for scene segmentation.\" Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2019.","DOI":"10.1109\/CVPR.2019.00326"},{"key":"e_1_3_2_1_21_1","volume-title":"Criss-cross attention for semantic segmentation.\" Proceedings of the IEEE\/CVF international conference on computer vision","author":"Huang Zilong","year":"2019","unstructured":"Huang , Zilong , \" Ccnet : Criss-cross attention for semantic segmentation.\" Proceedings of the IEEE\/CVF international conference on computer vision . 2019 . Huang, Zilong, \"Ccnet: Criss-cross attention for semantic segmentation.\" Proceedings of the IEEE\/CVF international conference on computer vision. 2019."},{"key":"e_1_3_2_1_22_1","volume-title":"Transformers make strong encoders for medical image segmentation.\" arXiv preprint arXiv:2102.04306","author":"Chen Jieneng","year":"2021","unstructured":"Chen , Jieneng , \" Transunet : Transformers make strong encoders for medical image segmentation.\" arXiv preprint arXiv:2102.04306 ( 2021 ). Chen, Jieneng, \"Transunet: Transformers make strong encoders for medical image segmentation.\" arXiv preprint arXiv:2102.04306 (2021)."},{"key":"e_1_3_2_1_23_1","first-page":"12077","article-title":"SegFormer: Simple and efficient design for semantic segmentation with transformers","volume":"34","author":"Xie Enze","year":"2021","unstructured":"Xie , Enze , \" SegFormer: Simple and efficient design for semantic segmentation with transformers .\" Advances in Neural Information Processing Systems 34 ( 2021 ): 12077 - 12090 . Xie, Enze, \"SegFormer: Simple and efficient design for semantic segmentation with transformers.\" Advances in Neural Information Processing Systems 34 (2021): 12077-12090.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_24_1","volume-title":"A versatile backbone for dense prediction without convolutions.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision","author":"Wang Wenhai","year":"2021","unstructured":"Wang , Wenhai , \"Pyramid vision transformer : A versatile backbone for dense prediction without convolutions.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision . 2021 . Wang, Wenhai, \"Pyramid vision transformer: A versatile backbone for dense prediction without convolutions.\" Proceedings of the IEEE\/CVF International Conference on Computer Vision. 2021."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"crossref","unstructured":"Mou Lichao Yuansheng Hua and Xiao Xiang Zhu. \"A relation-augmented fully convolutional network for semantic segmentation in aerial scenes.\" Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 2019.  Mou Lichao Yuansheng Hua and Xiao Xiang Zhu. \"A relation-augmented fully convolutional network for semantic segmentation in aerial scenes.\" Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 2019.","DOI":"10.1109\/CVPR.2019.01270"},{"key":"e_1_3_2_1_26_1","volume-title":"Flowing semantics through points for aerial image segmentation.\" Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Li Xiangtai","year":"2021","unstructured":"Li , Xiangtai , \" PointFlow : Flowing semantics through points for aerial image segmentation.\" Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition . 2021 . Li, Xiangtai, \"PointFlow: Flowing semantics through points for aerial image segmentation.\" Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 2021."},{"issue":"9","key":"e_1_3_2_1_27_1","doi-asserted-by":"crossref","first-page":"1339","DOI":"10.3390\/rs10091339","article-title":"ERN: Edge loss reinforced semantic segmentation network for remote sensing images","volume":"10","author":"Liu Shuo","year":"2018","unstructured":"Liu , Shuo , \" ERN: Edge loss reinforced semantic segmentation network for remote sensing images .\" Remote Sensing 10 . 9 ( 2018 ): 1339 . Liu, Shuo, \"ERN: Edge loss reinforced semantic segmentation network for remote sensing images.\" Remote Sensing 10.9 (2018): 1339.","journal-title":"Remote Sensing"},{"key":"e_1_3_2_1_28_1","first-page":"1","article-title":"Boundary enhancement semantic segmentation for building extraction from remote sensed image","volume":"60","author":"Jung Hoin","year":"2021","unstructured":"Jung , Hoin , Han-Soo Choi , and Myungjoo Kang . \" Boundary enhancement semantic segmentation for building extraction from remote sensed image .\" IEEE Transactions on Geoscience and Remote Sensing 60 ( 2021 ): 1 - 12 . Jung, Hoin, Han-Soo Choi, and Myungjoo Kang. \"Boundary enhancement semantic segmentation for building extraction from remote sensed image.\" IEEE Transactions on Geoscience and Remote Sensing 60 (2021): 1-12.","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"},{"key":"e_1_3_2_1_29_1","first-page":"1","article-title":"Multiattention network for semantic segmentation of fine-resolution remote sensing images","volume":"60","author":"Li Rui","year":"2021","unstructured":"Li , Rui , \" Multiattention network for semantic segmentation of fine-resolution remote sensing images .\" IEEE Transactions on Geoscience and Remote Sensing 60 ( 2021 ): 1 - 13 . Li, Rui, \"Multiattention network for semantic segmentation of fine-resolution remote sensing images.\" IEEE Transactions on Geoscience and Remote Sensing 60 (2021): 1-13.","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Li Rui \"ABCNet: Attentive bilateral contextual network for efficient semantic segmentation of Fine-Resolution remotely sensed imagery.\" ISPRS journal of photogrammetry and remote sensing 181 (2021):  84-98.  Li Rui \"ABCNet: Attentive bilateral contextual network for efficient semantic segmentation of Fine-Resolution remotely sensed imagery.\" ISPRS journal of photogrammetry and remote sensing 181 (2021): 84-98.","DOI":"10.1016\/j.isprsjprs.2021.09.005"},{"issue":"16","key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","first-page":"3065","DOI":"10.3390\/rs13163065","article-title":"Transformer meets convolution: A bilateral awareness network for semantic segmentation of very fine resolution urban scene images","volume":"13","author":"Wang Libo","year":"2021","unstructured":"Wang , Libo , \" Transformer meets convolution: A bilateral awareness network for semantic segmentation of very fine resolution urban scene images .\" Remote Sensing 13 . 16 ( 2021 ): 3065 . Wang, Libo, \"Transformer meets convolution: A bilateral awareness network for semantic segmentation of very fine resolution urban scene images.\" Remote Sensing 13.16 (2021): 3065.","journal-title":"Remote Sensing"},{"key":"e_1_3_2_1_32_1","first-page":"1","article-title":"Swin Transformer Embedding UNet for Remote Sensing Image Semantic Segmentation","volume":"60","author":"He Xin","year":"2022","unstructured":"He , Xin , \" Swin Transformer Embedding UNet for Remote Sensing Image Semantic Segmentation .\" IEEE Transactions on Geoscience and Remote Sensing 60 ( 2022 ): 1 - 15 . He, Xin, \"Swin Transformer Embedding UNet for Remote Sensing Image Semantic Segmentation.\" IEEE Transactions on Geoscience and Remote Sensing 60 (2022): 1-15.","journal-title":"IEEE Transactions on Geoscience and Remote Sensing"}],"event":{"name":"ICCPR 2022: 2022 11th International Conference on Computing and Pattern Recognition","acronym":"ICCPR 2022","location":"Beijing China"},"container-title":["Proceedings of the 2022 11th International Conference on Computing and Pattern Recognition"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3581807.3581827","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3581807.3581827","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:49:29Z","timestamp":1750182569000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3581807.3581827"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,17]]},"references-count":32,"alternative-id":["10.1145\/3581807.3581827","10.1145\/3581807"],"URL":"https:\/\/doi.org\/10.1145\/3581807.3581827","relation":{},"subject":[],"published":{"date-parts":[[2022,11,17]]},"assertion":[{"value":"2023-05-22","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}