{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,4]],"date-time":"2026-06-04T23:27:35Z","timestamp":1780615655674,"version":"3.54.1"},"reference-count":42,"publisher":"MDPI AG","issue":"22","license":[{"start":{"date-parts":[[2021,11,11]],"date-time":"2021-11-11T00:00:00Z","timestamp":1636588800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["4207513"],"award-info":[{"award-number":["4207513"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Remote Sensing"],"abstract":"<jats:p>Cloud detection is a key step in the preprocessing of optical satellite remote sensing images. In the existing literature, cloud detection methods are roughly divided into threshold methods and deep-learning methods. Most of the traditional threshold methods are based on the spectral characteristics of clouds, so it is easy to lose the spatial location information in the high-reflection area, resulting in misclassification. Besides, due to the lack of generalization, the traditional deep-learning network also easily loses the details and spatial information if it is directly applied to cloud detection. In order to solve these problems, we propose a deep-learning model, Cloud Detection UNet (CDUNet), for cloud detection. The characteristics of the network are that it can refine the division boundary of the cloud layer and capture its spatial position information. In the proposed model, we introduced a High-frequency Feature Extractor (HFE) and a Multiscale Convolution (MSC) to refine the cloud boundary and predict fragmented clouds. Moreover, in order to improve the accuracy of thin cloud detection, the Spatial Prior Self-Attention (SPSA) mechanism was introduced to establish the cloud spatial position information. Additionally, a dual-attention mechanism is proposed to reduce the proportion of redundant information in the model and improve the overall performance of the model. The experimental results showed that our model can cope with complex cloud cover scenes and has excellent performance on cloud datasets and SPARCS datasets. Its segmentation accuracy is better than the existing methods, which is of great significance for cloud-detection-related work.<\/jats:p>","DOI":"10.3390\/rs13224533","type":"journal-article","created":{"date-parts":[[2021,11,11]],"date-time":"2021-11-11T23:04:46Z","timestamp":1636671886000},"page":"4533","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":58,"title":["CDUNet: Cloud Detection UNet for Remote Sensing Imagery"],"prefix":"10.3390","volume":"13","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7181-9935","authenticated-orcid":false,"given":"Kai","family":"Hu","sequence":"first","affiliation":[{"name":"Collaborative Innovation Center on Atmospheric Environment and Equipment Technology, Nanjing University of Information Science and Technology, Nanjing 210044, China"},{"name":"Jiangsu Key Laboratory of Big Data Analysis Technology, Nanjing University of Information Science and Technology, Nanjing 210044, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Dongsheng","family":"Zhang","sequence":"additional","affiliation":[{"name":"Collaborative Innovation Center on Atmospheric Environment and Equipment Technology, Nanjing University of Information Science and Technology, Nanjing 210044, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4681-9129","authenticated-orcid":false,"given":"Min","family":"Xia","sequence":"additional","affiliation":[{"name":"Collaborative Innovation Center on Atmospheric Environment and Equipment Technology, Nanjing University of Information Science and Technology, Nanjing 210044, China"},{"name":"Jiangsu Key Laboratory of Big Data Analysis Technology, Nanjing University of Information Science and Technology, Nanjing 210044, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2021,11,11]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","unstructured":"Kegelmeyer, W.P. (1994). Extraction of Cloud Statistics from Whole Sky Imaging Cameras (No. SAND-94-8222).","DOI":"10.2172\/10141846"},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1016\/j.rse.2011.10.028","article-title":"Object-based cloud and cloud shadow detection in Landsat imagery","volume":"118","author":"Zhu","year":"2012","journal-title":"Remote Sens. Environ."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Candra, D.S., Phinn, S., and Scarth, P. (2019). Automated cloud and cloud shadow masking for landsat 8 using multitemporal images in a variety of environments. Remote Sens., 11.","DOI":"10.3390\/rs11172060"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"3322","DOI":"10.1109\/TGRS.2017.2669341","article-title":"Automatic Road Detection and Centerline Extraction via Cascaded End-to-End Convolutional Neural Network","volume":"55","author":"Cheng","year":"2017","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_5","first-page":"102597","article-title":"SUACDNet: Attentional change detection network based on siamese U-shaped structure","volume":"105","author":"Song","year":"2021","journal-title":"Int. J. Appl. Earth Obs. Geoinf."},{"key":"ref_6","first-page":"9548325","article-title":"An All-Scale Feature Fusion Network With Boundary Point Prediction for Cloud Detection","volume":"2021","author":"Wang","year":"2021","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"2022","DOI":"10.1080\/01431161.2020.1849852","article-title":"Cloud\/shadow segmentation based on global attention feature fusion residual network for remote sensing imagery","volume":"42","author":"Xia","year":"2021","journal-title":"Int. J. Remote Sens."},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Mohajerani, S., and Saeedi, P. (August, January 28). Cloud-Net: An End-To-End Cloud Detection Algorithm for Landsat 8 Imagery. Proceedings of the IGARSS 2019\u20142019 IEEE International Geoscience and Remote Sensing Symposium, Yokohama, Japan.","DOI":"10.1109\/IGARSS.2019.8898776"},{"key":"ref_9","doi-asserted-by":"crossref","unstructured":"Long, J., Shelhamer, E., and Darrell, T. (2015, January 7\u201312). Fully convolutional networks for semantic segmentation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","DOI":"10.1109\/CVPR.2015.7298965"},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"1785","DOI":"10.1109\/LGRS.2017.2735801","article-title":"Distinguishing Cloud and Snow in Satellite Images via Deep Convolutional Network","volume":"14","author":"Zhan","year":"2017","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_11","doi-asserted-by":"crossref","unstructured":"Dr\u00f6nner, J., Korfhage, N., Egli, S., M\u00fchling, M., Thies, B., Bendix, J., Freisleben, B., and Seeger, B. (2018). Fast Cloud Segmentation Using Convolutional Neural Networks. Remote Sens., 10.","DOI":"10.3390\/rs10111782"},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"307","DOI":"10.1016\/j.rse.2019.03.007","article-title":"Cloud and cloud shadow detection in Landsat imagery based on deep convolutional neural networks","volume":"225","author":"Chai","year":"2019","journal-title":"Remote Sens. Environ."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"2481","DOI":"10.1109\/TPAMI.2016.2644615","article-title":"SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation","volume":"39","author":"Badrinarayanan","year":"2017","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"key":"ref_14","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-Net: Convolutional Networks for Biomedical Image Segmentation. Proceedings of the International Conference on Medical Image Computing and Computer-Assisted Intervention, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_15","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"104940","DOI":"10.1016\/j.cageo.2021.104940","article-title":"Strip pooling channel spatial attention network for the segmentation of cloud and cloud shadow","volume":"157","author":"Qu","year":"2021","journal-title":"Comput. Geosci."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"113669","DOI":"10.1016\/j.eswa.2020.113669","article-title":"Non-intrusive load disaggregation based on composite deep long short-term memory network","volume":"160","author":"Xia","year":"2020","journal-title":"Expert Syst. Appl."},{"key":"ref_18","unstructured":"Wang, Z., Xia, M., Lu, M., Pan, L., and Liu, J. (2021). Parameter Identification in Power Transmission Systems Based on Graph Convolution Network. IEEE Trans. Power Deliv., 1."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"156","DOI":"10.1080\/01431161.2018.1508917","article-title":"Cloud\/snow recognition for multispectral satellite imagery based on a multidimensional deep residual network","volume":"40","author":"Xia","year":"2018","journal-title":"Int. J. Remote Sens."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"2417","DOI":"10.1109\/TIFS.2020.2969552","article-title":"Multi-Stage Feature Constraints Learning for Age Estimation","volume":"15","author":"Xia","year":"2020","journal-title":"IEEE Trans. Inf. Forensics Secur."},{"key":"ref_21","unstructured":"Lin, M., and Chen, Q. (2013). Network in network. arXiv."},{"key":"ref_22","doi-asserted-by":"crossref","first-page":"2278","DOI":"10.1109\/5.726791","article-title":"Gradient-based learning applied to document recognition","volume":"86","author":"LeCun","year":"1998","journal-title":"Proc. IEEE"},{"key":"ref_23","unstructured":"Chen, C.F., Fan, Q., Mallinar, N., Sercu, T., and Feris, R. (2018). Big-little net: An efficient multiscale feature representation for visual and speech recognition. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Chen, Y., Fan, H., Xu, B., Yan, Z., Kalantidis, Y., Rohrbach, M., Shuicheng, Y., and Feng, J. (2019, January 27\u201328). Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks With Octave Convolution. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Korea.","DOI":"10.1109\/ICCV.2019.00353"},{"key":"ref_25","unstructured":"Cheng, B., Xiao, R., Wang, J., Huang, T., and Zhang, L. (2019). High frequency residual learning for multiscale image classification. arXiv."},{"key":"ref_26","unstructured":"Luo, W., Li, Y., Urtasun, R., and Zemel, R. (2016, January 5\u201310). Understanding the effective receptive field in deep convolutional neural networks. Proceedings of the 30th International Conference on Neural Information Processing Systems, Barcelona, Spain."},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Wang, X., Girshick, R., Gupta, A., and He, K. (2018, January 18\u201322). Non-local Neural Networks. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00813"},{"key":"ref_28","doi-asserted-by":"crossref","unstructured":"Fu, J., Liu, J., Tian, H., Li, Y., Bao, Y., Fang, Z., and Lu, H. (2019, January 16\u201320). Dual Attention Network for Scene Segmentation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00326"},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Huang, Z., Wang, X., Huang, L., Huang, C., Wei, Y., and Liu, W. (November, January 27). CCNet: Criss-Cross Attention for Semantic Segmentation. Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Korea.","DOI":"10.1109\/ICCV.2019.00069"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Li, X., Wang, W., Hu, X., and Yang, J. (2019, January 15\u201320). Selective kernel networks. Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00060"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., and Lee, J.-Y. (2018). CBAM: Convolutional Block Attention Module. arXiv.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_32","unstructured":"Hughes, M. (2016). L8 SPARCS Cloud Validation Masks."},{"key":"ref_33","doi-asserted-by":"crossref","first-page":"4907","DOI":"10.3390\/rs6064907","article-title":"Automated detection of cloud and cloud shadow in single-date Landsat imagery using neural networks and spatial postprocessing","volume":"6","author":"Hughes","year":"2014","journal-title":"Remote Sens."},{"key":"ref_34","unstructured":"Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A. (2017, January 9). Automatic Differentiation in Pytorch. Proceedings of the NIPS 2017 Workshop Autodiff Submission, Long Beach, CA, USA."},{"key":"ref_35","unstructured":"Kingma, D.P., and Ba, J. (2015, January 5\u20138). Adam: A method for stochastic optimization. Proceedings of the International Conference Learn (ICLR), San Diego, CA, USA."},{"key":"ref_36","doi-asserted-by":"crossref","unstructured":"Zhao, H., Shi, J., Qi, X., Wang, X., and Jia, J. (2017;, January 21\u201326). Pyramid Scene Parsing Network. Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","DOI":"10.1109\/CVPR.2017.660"},{"key":"ref_37","unstructured":"Howard, A.G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Wey, T., Andreetto, M., and Adam, H. (2017). Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv."},{"key":"ref_38","doi-asserted-by":"crossref","first-page":"211","DOI":"10.1007\/s11263-015-0816-y","article-title":"Imagenet large scale visual recognition challenge","volume":"115","author":"Russakovsky","year":"2015","journal-title":"Int. J. Comput. Vis."},{"key":"ref_39","doi-asserted-by":"crossref","first-page":"3051","DOI":"10.1007\/s11263-021-01515-2","article-title":"BiSeNet V2: Bilateral Network with Guided Aggregation for Real-Time Semantic Segmentation","volume":"129","author":"Yu","year":"2021","journal-title":"Int. J. Comput. Vis."},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Yang, M., Yu, K., Zhang, C., Li, Z., and Yang, K. (2018, January 18\u201323). DenseASPP for Semantic Segmentation in Street Scenes. Proceedings of the 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00388"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Chen, L.C., Zhu, Y., Papandreou, G., Schroff, F., and Adam, H. (2018, January 8\u201314). Encoder-decoder with atrous separable convolution for semantic image segmentation. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_49"},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Sun, K., Xiao, B., Liu, D., and Wang, J. (2019, January 16\u201320). Deep High-Resolution Representation Learning for Human Pose Estimation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00584"}],"container-title":["Remote Sensing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2072-4292\/13\/22\/4533\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T07:28:43Z","timestamp":1760167723000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2072-4292\/13\/22\/4533"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,11,11]]},"references-count":42,"journal-issue":{"issue":"22","published-online":{"date-parts":[[2021,11]]}},"alternative-id":["rs13224533"],"URL":"https:\/\/doi.org\/10.3390\/rs13224533","relation":{},"ISSN":["2072-4292"],"issn-type":[{"value":"2072-4292","type":"electronic"}],"subject":[],"published":{"date-parts":[[2021,11,11]]}}}