{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,31]],"date-time":"2026-01-31T04:40:08Z","timestamp":1769834408220,"version":"3.49.0"},"reference-count":88,"publisher":"MDPI AG","issue":"5","license":[{"start":{"date-parts":[[2023,2,27]],"date-time":"2023-02-27T00:00:00Z","timestamp":1677456000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/501100012166","name":"National Key Research and Development Program of China","doi-asserted-by":"publisher","award":["2017YFA0700800"],"award-info":[{"award-number":["2017YFA0700800"]}],"id":[{"id":"10.13039\/501100012166","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Remote Sensing"],"abstract":"<jats:p>The reconstruction of missing pixels is essential for remote sensing images, as they often suffer from problems such as covering, dead pixels, and scan line corrector (SLC)-off. Image inpainting techniques can solve these problems, as they can generate realistic content for the unknown regions of an image based on the known regions. Recently, convolutional neural network (CNN)-based inpainting methods have integrated the attention mechanism to improve inpainting performance, as they can capture long-range dependencies and adapt to inputs in a flexible manner. However, to obtain the attention map for each feature, they compute the similarities between the feature and the entire feature map, which may introduce noise from irrelevant features. To address this problem, we propose a novel adaptive attention (Ada-attention) that uses an offset position subnet to adaptively select the most relevant keys and values based on self-attention. This enables the attention to be focused on essential features and model more informative dependencies on the global range. Ada-attention first employs an offset subnet to predict offset position maps on the query feature map; then, it samples the most relevant features from the input feature map based on the offset position; next, it computes key and value maps for self-attention using the sampled features; finally, using the query, key and value maps, the self-attention outputs the reconstructed feature map. Based on Ada-attention, we customized a u-shaped adaptive-attention completing network (AACNet) to reconstruct missing regions. Experimental results on several digital remote sensing and natural image datasets, using two image inpainting models and two remote sensing image reconstruction approaches, demonstrate that the proposed AACNet achieves a good quantitative performance and good visual restoration results with regard to object integrity, texture\/edge detail, and structural consistency. Ablation studies indicate that Ada-attention outperforms self-attention in terms of PSNR by 0.66%, SSIM by 0.74%, and MAE by 3.9%, and can focus on valuable global features using the adaptive offset subnet. Additionally, our approach has also been successfully applied to remove real clouds in remote sensing images, generating credible content for cloudy regions.<\/jats:p>","DOI":"10.3390\/rs15051321","type":"journal-article","created":{"date-parts":[[2023,3,6]],"date-time":"2023-03-06T03:02:32Z","timestamp":1678071752000},"page":"1321","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":11,"title":["Adaptive-Attention Completing Network for Remote Sensing Image"],"prefix":"10.3390","volume":"15","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8636-044X","authenticated-orcid":false,"given":"Wenli","family":"Huang","sequence":"first","affiliation":[{"name":"The Institute of Artificial Intelligence and Robotic, Xi\u2019an Jiaotong University, Xian Ning West Road No. 28, Xi\u2019an 710049, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ye","family":"Deng","sequence":"additional","affiliation":[{"name":"The Institute of Artificial Intelligence and Robotic, Xi\u2019an Jiaotong University, Xian Ning West Road No. 28, Xi\u2019an 710049, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Siqi","family":"Hui","sequence":"additional","affiliation":[{"name":"The Institute of Artificial Intelligence and Robotic, Xi\u2019an Jiaotong University, Xian Ning West Road No. 28, Xi\u2019an 710049, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jinjun","family":"Wang","sequence":"additional","affiliation":[{"name":"The Institute of Artificial Intelligence and Robotic, Xi\u2019an Jiaotong University, Xian Ning West Road No. 28, Xi\u2019an 710049, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2023,2,27]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"61","DOI":"10.1109\/MGRS.2015.2441912","article-title":"Missing information reconstruction of remote sensing data: A technical review","volume":"3","author":"Shen","year":"2015","journal-title":"IEEE Geosci. Remote Sens. Mag."},{"key":"ref_2","first-page":"1","article-title":"Context-based multiscale unified network for missing data reconstruction in remote sensing images","volume":"19","author":"Shao","year":"2020","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_3","doi-asserted-by":"crossref","unstructured":"Pathak, D., Krahenbuhl, P., Donahue, J., Darrell, T., and Efros, A.A. (2016, January 27\u201330). Context Encoders: Feature Learning by Inpainting. Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.278"},{"key":"ref_4","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3072959.3073659","article-title":"Globally and locally consistent image completion","volume":"36","author":"Iizuka","year":"2017","journal-title":"ACM Trans. Graph. (ToG)"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"1784","DOI":"10.1109\/TIP.2020.3048629","article-title":"Dynamic selection network for image inpainting","volume":"30","author":"Wang","year":"2021","journal-title":"IEEE Trans. Image Process."},{"key":"ref_6","doi-asserted-by":"crossref","unstructured":"Yu, J., Lin, Z., Yang, J., Shen, X., Lu, X., and Huang, T.S. (2018, January 18\u201323). Generative Image Inpainting with Contextual Attention. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00577"},{"key":"ref_7","doi-asserted-by":"crossref","unstructured":"Zeng, Y., Fu, J., Chao, H., and Guo, B. (2019, January 15\u201320). Learning Pyramid-Context Encoder Network for High-Quality Image Inpainting. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00158"},{"key":"ref_8","doi-asserted-by":"crossref","unstructured":"Ronneberger, O., Fischer, P., and Brox, T. (2015, January 5\u20139). U-net: Convolutional networks for biomedical image segmentation. Proceedings of the International Conference on Medical Image Computing and Computer-Assisted Intervention, Munich, Germany.","DOI":"10.1007\/978-3-319-24574-4_28"},{"key":"ref_9","first-page":"768","article-title":"Dead pixel completion of aqua MODIS band 6 using a robust M-estimator multiregression","volume":"11","author":"Li","year":"2013","journal-title":"IEEE Geosci. Remote Sens. Lett."},{"key":"ref_10","doi-asserted-by":"crossref","first-page":"7901","DOI":"10.1109\/TGRS.2020.3038878","article-title":"Spatial\u2013spectral radial basis function-based interpolation for Landsat ETM+ SLC-off image gap filling","volume":"59","author":"Wang","year":"2020","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"182","DOI":"10.1016\/j.rse.2012.12.012","article-title":"Recovering missing pixels for Landsat ETM+ SLC-off imagery using multi-temporal regression analysis and a regularization method","volume":"131","author":"Zeng","year":"2013","journal-title":"Remote Sens. Environ."},{"key":"ref_12","doi-asserted-by":"crossref","first-page":"4274","DOI":"10.1109\/TGRS.2018.2810208","article-title":"Missing data reconstruction in remote sensing image with a unified spatial\u2013temporal\u2013spectral deep convolutional neural network","volume":"56","author":"Zhang","year":"2018","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"894","DOI":"10.1109\/TGRS.2013.2245509","article-title":"Compressed sensing-based inpainting of aqua moderate resolution imaging spectroradiometer band 6 using adaptive spectrum-weighted sparse Bayesian dictionary learning","volume":"52","author":"Shen","year":"2013","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_14","unstructured":"Scaramuzza, P., and Barsi, J. (2005, January 23\u201327). Landsat 7 scan line corrector-off gap-filled product development. Proceedings of the Pecora, Sioux Falls, SD, USA."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"1053","DOI":"10.1016\/j.rse.2010.12.010","article-title":"A simple and effective method for filling gaps in Landsat ETM+ SLC-off images","volume":"115","author":"Chen","year":"2011","journal-title":"Remote Sens. Environ."},{"key":"ref_16","doi-asserted-by":"crossref","first-page":"3629","DOI":"10.1109\/JSTARS.2016.2533547","article-title":"Patch matching-based multitemporal group sparse representation for the missing information reconstruction of remote-sensing images","volume":"9","author":"Li","year":"2016","journal-title":"IEEE J. Sel. Top. Appl. Earth Obs. Remote Sens."},{"key":"ref_17","doi-asserted-by":"crossref","first-page":"54","DOI":"10.1016\/j.isprsjprs.2014.02.015","article-title":"Cloud removal for remotely sensed images by similar pixel replacement guided with a spatio-temporal MRF model","volume":"92","author":"Cheng","year":"2014","journal-title":"ISPRS J. Photogramm. Remote Sens."},{"key":"ref_18","doi-asserted-by":"crossref","first-page":"215","DOI":"10.1364\/OPTCON.439671","article-title":"Remote sensing image cloud removal by deep image prior with a multitemporal constraint","volume":"1","author":"Zhang","year":"2022","journal-title":"Opt. Contin."},{"key":"ref_19","doi-asserted-by":"crossref","first-page":"3047","DOI":"10.1109\/TGRS.2018.2790262","article-title":"Nonlocal tensor completion for multitemporal remotely sensed images\u2019 inpainting","volume":"56","author":"Ji","year":"2018","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_20","doi-asserted-by":"crossref","first-page":"3367","DOI":"10.1109\/TGRS.2017.2670021","article-title":"An adaptive weighted tensor completion method for the recovery of remote sensing images with missing data","volume":"55","author":"Ng","year":"2017","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Yu, C., Chen, L., Su, L., Fan, M., and Li, S. (2011, January 24\u201326). Kriging interpolation method and its application in retrieval of MODIS aerosol optical depth. Proceedings of the 19th International Conference on Geoinformatics, Shanghai, China.","DOI":"10.1109\/GeoInformatics.2011.5981052"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Bertalmio, M., Sapiro, G., Caselles, V., and Ballester, C. (2000, January 23\u201328). Image inpainting. Proceedings of the 27th Annual Conference on Computer Graphics and Interactive Techniques, New Orleans, LA, USA.","DOI":"10.1145\/344779.344972"},{"key":"ref_23","doi-asserted-by":"crossref","unstructured":"Barcelos, C.A.Z., and Batista, M.A. (2003, January 12\u201315). Image inpainting and denoising by nonlinear partial differential equations. Proceedings of the 16th Brazilian Symposium on Computer Graphics and Image Processing, Sao Carlos, Brazil.","DOI":"10.1109\/SIBGRA.2003.1241021"},{"key":"ref_24","doi-asserted-by":"crossref","first-page":"1200","DOI":"10.1109\/TIP.2004.833105","article-title":"Region Filling and Object Removal by Exemplar-Based Image Inpainting","volume":"13","author":"Criminisi","year":"2004","journal-title":"IEEE Trans. Image Process."},{"key":"ref_25","doi-asserted-by":"crossref","first-page":"24","DOI":"10.1145\/1531326.1531330","article-title":"PatchMatch: A randomized correspondence algorithm for structural image editing","volume":"28","author":"Barnes","year":"2009","journal-title":"ACM Trans. Graph."},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Singh, P., and Komodakis, N. (2018, January 22\u201327). Cloud-gan: Cloud removal for sentinel-2 imagery using a cyclic consistent generative adversarial networks. Proceedings of the 2018 IEEE International Geoscience and Remote Sensing Symposium, Valencia, Spain.","DOI":"10.1109\/IGARSS.2018.8519033"},{"key":"ref_27","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1109\/TGRS.2022.3208339","article-title":"Efficient Pyramidal GAN for Versatile Missing Data Reconstruction in Remote Sensing Images","volume":"60","author":"Shao","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_28","unstructured":"Pan, H. (2020). Cloud removal for remote sensing imagery via spatial attention generative adversarial network. arXiv."},{"key":"ref_29","doi-asserted-by":"crossref","first-page":"333","DOI":"10.1016\/j.isprsjprs.2020.05.013","article-title":"Cloud removal in Sentinel-2 imagery using a deep residual neural network and SAR-optical data fusion","volume":"166","author":"Meraner","year":"2020","journal-title":"ISPRS J. Photogramm. Remote Sens."},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Nazeri, K., Ng, E., Joseph, T., Qureshi, F., and Ebrahimi, M. (2019, January 27\u201328). Edgeconnect: Structure guided image inpainting using edge prediction. Proceedings of the IEEE\/CVF International Conference on Computer Vision Workshops (ICCVW), Seoul, Republic of Korea.","DOI":"10.1109\/ICCVW.2019.00408"},{"key":"ref_31","doi-asserted-by":"crossref","unstructured":"Ren, Y., Yu, X., Zhang, R., Li, T.H., and Li, G. (2019, January 27\u201328). StructureFlow: Image Inpainting via Structure-aware Appearance Flow. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00027"},{"key":"ref_32","doi-asserted-by":"crossref","unstructured":"Xiong, W., Yu, J., Lin, Z., Yang, J., Lu, X., Barnes, C., and Luo, J. (2019, January 15\u201320). Foreground-Aware Image Inpainting. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Long Beach, CA, USA.","DOI":"10.1109\/CVPR.2019.00599"},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Peng, J., Liu, D., Xu, S., and Li, H. (2021, January 19\u201325). Generating Diverse Structure for Image Inpainting With Hierarchical VQ-VAE. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Virtual.","DOI":"10.1109\/CVPR46437.2021.01063"},{"key":"ref_34","doi-asserted-by":"crossref","unstructured":"Liu, H., Wan, Z., Huang, W., Song, Y., Han, X., and Liao, J. (2021, January 19\u201325). PD-GAN: Probabilistic Diverse GAN for Image Inpainting. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Virtual.","DOI":"10.1109\/CVPR46437.2021.00925"},{"key":"ref_35","doi-asserted-by":"crossref","unstructured":"Liu, Q., Tan, Z., Chen, D., Chu, Q., Dai, X., Chen, Y., Liu, M., Yuan, L., and Yu, N. (2022, January 19\u201320). Reduce Information Loss in Transformers for Pluralistic Image Inpainting. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), New Orleans, LA, USA.","DOI":"10.1109\/CVPR52688.2022.01106"},{"key":"ref_36","first-page":"1","article-title":"A Coarse-to-Fine Deep Generative Model with Spatial Semantic Attention for High-Resolution Remote Sensing Image Inpainting","volume":"60","author":"Du","year":"2022","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_37","doi-asserted-by":"crossref","unstructured":"Li, J., Wang, N., Zhang, L., Du, B., and Tao, D. (2020, January 13\u201319). Recurrent Feature Reasoning for Image Inpainting. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00778"},{"key":"ref_38","doi-asserted-by":"crossref","unstructured":"Zhang, H., Hu, Z., Luo, C., Zuo, W., and Wang, M. (2018, January 22\u201326). Semantic image inpainting with progressive generative networks. Proceedings of the 26th ACM international conference on Multimedia, Seoul, Republic of Korea.","DOI":"10.1145\/3240508.3240625"},{"key":"ref_39","doi-asserted-by":"crossref","unstructured":"Wang, W., Zhang, J., Niu, L., Ling, H., Yang, X., and Zhang, L. (2021, January 10\u201317). Parallel multi-resolution fusion network for image inpainting. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.01429"},{"key":"ref_40","doi-asserted-by":"crossref","unstructured":"Guo, X., Yang, H., and Huang, D. (2021, January 10\u201317). Image Inpainting via Conditional Texture and Structure Dual Generation. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.01387"},{"key":"ref_41","doi-asserted-by":"crossref","unstructured":"Liu, G., Reda, F.A., Shih, K.J., Wang, T.C., Tao, A., and Catanzaro, B. (2018, January 8\u201314). Image Inpainting for Irregular Holes Using Partial Convolutions. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01252-6_6"},{"key":"ref_42","doi-asserted-by":"crossref","unstructured":"Yu, J., Lin, Z., Yang, J., Shen, X., Lu, X., and Huang, T.S. (2019, January 27\u201328). Free-form image inpainting with gated convolution. Proceedings of the IEEE\/CVF International Conference on Computer Vision, Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00457"},{"key":"ref_43","doi-asserted-by":"crossref","unstructured":"Huang, W., Deng, Y., Hui, S., and Wang, J. (2022). Image Inpainting with Bilateral Convolution. Remote Sens., 14.","DOI":"10.3390\/rs14236140"},{"key":"ref_44","doi-asserted-by":"crossref","unstructured":"Suvorov, R., Logacheva, E., Mashikhin, A., Remizova, A., Ashukha, A., Silvestrov, A., Kong, N., Goka, H., Park, K., and Lempitsky, V. (2022, January 4\u20138). Resolution-robust large mask inpainting with fourier convolutions. Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision, Waikoloa, HI, USA.","DOI":"10.1109\/WACV51458.2022.00323"},{"key":"ref_45","unstructured":"Yu, T., Guo, Z., Jin, X., Wu, S., Chen, Z., Li, W., Zhang, Z., and Liu, S. (2020, January 7\u201312). Region normalization for image inpainting. Proceedings of the AAAI Conference on Artificial Intelligence, New York, NY, USA."},{"key":"ref_46","doi-asserted-by":"crossref","unstructured":"Ma, X., Zhou, X., Huang, H., Chai, Z., Wei, X., and He, R. (2021, January 10\u201315). Free-form image inpainting via contrastive attention network. Proceedings of the 2020 25th International Conference on Pattern Recognition (ICPR), Milan, Italy.","DOI":"10.1109\/ICPR48806.2021.9412028"},{"key":"ref_47","doi-asserted-by":"crossref","first-page":"103155","DOI":"10.1016\/j.cviu.2020.103155","article-title":"Multi-scale attention network for image inpainting","volume":"204","author":"Qin","year":"2021","journal-title":"Comput. Vis. Image Underst."},{"key":"ref_48","doi-asserted-by":"crossref","unstructured":"Liu, H., Jiang, B., Xiao, Y., and Yang, C. (2019, January 27\u201328). Coherent semantic attention for image inpainting. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Seoul, Republic of Korea.","DOI":"10.1109\/ICCV.2019.00427"},{"key":"ref_49","first-page":"5998","article-title":"Attention is all you need","volume":"30","author":"Vaswani","year":"2017","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_50","doi-asserted-by":"crossref","unstructured":"Hu, J., Shen, L., and Sun, G. (2018, January 18\u201323). Squeeze-and-excitation networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00745"},{"key":"ref_51","doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.Y., and Kweon, I.S. (2018, January 8\u201314). Cbam: Convolutional block attention module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","DOI":"10.1007\/978-3-030-01234-2_1"},{"key":"ref_52","unstructured":"Zhu, X., Su, W., Lu, L., Li, B., Wang, X., and Dai, J. (2021, January 3\u20137). Deformable DETR: Deformable transformers for end-to-end object detection. Proceedings of the 9th International Conference on Learning Representations (ICLR), Virtual Event, Austria."},{"key":"ref_53","doi-asserted-by":"crossref","unstructured":"Liu, Z., Lin, Y., Cao, Y., Hu, H., Wei, Y., Zhang, Z., Lin, S., and Guo, B. (2021, January 10\u201317). Swin transformer: Hierarchical vision transformer using shifted windows. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Montreal, QC, Canada.","DOI":"10.1109\/ICCV48922.2021.00986"},{"key":"ref_54","unstructured":"Child, R., Gray, S., Radford, A., and Sutskever, I. (2019). Generating long sequences with sparse transformers. arXiv."},{"key":"ref_55","unstructured":"Wang, S., Li, B.Z., Khabsa, M., Fang, H., and Ma, H. (2020). Linformer: Self-attention with linear complexity. arXiv."},{"key":"ref_56","unstructured":"Kitaev, N., Kaiser, \u0141., and Levskaya, A. (2020). Reformer: The efficient transformer. arXiv."},{"key":"ref_57","doi-asserted-by":"crossref","first-page":"53","DOI":"10.1162\/tacl_a_00353","article-title":"Efficient content-based sparse attention with routing transformers","volume":"9","author":"Roy","year":"2021","journal-title":"Trans. Assoc. Comput. Linguist."},{"key":"ref_58","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 27\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2016.90"},{"key":"ref_59","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 11\u201314). Identity mappings in deep residual networks. Proceedings of the European Conference on Computer Vision (ECCV), Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46493-0_38"},{"key":"ref_60","unstructured":"Ulyanov, D., Vedaldi, A., and Lempitsky, V. (2016). Instance normalization: The missing ingredient for fast stylization. arXiv."},{"key":"ref_61","unstructured":"Agarap, A.F. (2018). Deep learning using rectified linear units (relu). arXiv."},{"key":"ref_62","doi-asserted-by":"crossref","unstructured":"Dai, J., Qi, H., Xiong, Y., Li, Y., Zhang, G., Hu, H., and Wei, Y. (2017, January 22\u201329). Deformable convolutional networks. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","DOI":"10.1109\/ICCV.2017.89"},{"key":"ref_63","unstructured":"Howard, A.G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., and Adam, H. (2017). Mobilenets: Efficient convolutional neural networks for mobile vision applications. arXiv."},{"key":"ref_64","unstructured":"Ba, J.L., Kiros, J.R., and Hinton, G.E. (2016). Layer normalization. arXiv."},{"key":"ref_65","unstructured":"Hendrycks, D., and Gimpel, K. (2016). Bridging nonlinearities and stochastic regularizers with gaussian error linear units. CoRR."},{"key":"ref_66","doi-asserted-by":"crossref","unstructured":"Kirkland, E.J. (2010). Advanced Computing in Electron Microscopy, Springer.","DOI":"10.1007\/978-1-4419-6533-2"},{"key":"ref_67","doi-asserted-by":"crossref","unstructured":"Johnson, J., Alahi, A., and Fei-Fei, L. (2016, January 11-\u201314). Perceptual Losses for Real-Time Style Transfer and Super-Resolution. Proceedings of the European Conference on Computer Vision (ECCV), Amsterdam, The Netherlands.","DOI":"10.1007\/978-3-319-46475-6_43"},{"key":"ref_68","doi-asserted-by":"crossref","unstructured":"Isola, P., Zhu, J.Y., Zhou, T., and Efros, A.A. (2016, January 27\u201330). Image-to-Image Translation with Conditional Adversarial Networks. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA.","DOI":"10.1109\/CVPR.2017.632"},{"key":"ref_69","unstructured":"Simonyan, K., and Zisserman, A. (2014). Very deep convolutional networks for large-scale image recognition. arXiv."},{"key":"ref_70","first-page":"2153","article-title":"On the Nystr\u00f6m Method for Approximating a Gram Matrix for Improved Kernel-Based Learning","volume":"6","author":"Drineas","year":"2005","journal-title":"J. Mach. Learn. Res."},{"key":"ref_71","doi-asserted-by":"crossref","first-page":"131579","DOI":"10.1109\/ACCESS.2022.3227387","article-title":"Multi-receptions and multi-gradients discriminator for Image Inpainting","volume":"10","author":"Huang","year":"2022","journal-title":"IEEE Access"},{"key":"ref_72","doi-asserted-by":"crossref","first-page":"3965","DOI":"10.1109\/TGRS.2017.2685945","article-title":"AID: A benchmark data set for performance evaluation of aerial scene classification","volume":"55","author":"Xia","year":"2017","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_73","doi-asserted-by":"crossref","first-page":"197","DOI":"10.1016\/j.isprsjprs.2018.01.004","article-title":"PatternNet: A benchmark dataset for performance evaluation of remote sensing image retrieval","volume":"145","author":"Zhou","year":"2018","journal-title":"ISPRS J. Photogramm. Remote Sens."},{"key":"ref_74","doi-asserted-by":"crossref","first-page":"1865","DOI":"10.1109\/JPROC.2017.2675998","article-title":"Remote sensing image scene classification: Benchmark and state of the art","volume":"105","author":"Cheng","year":"2017","journal-title":"Proc. IEEE"},{"key":"ref_75","doi-asserted-by":"crossref","first-page":"103","DOI":"10.1145\/2830541","article-title":"What makes paris look like paris?","volume":"58","author":"Doersch","year":"2015","journal-title":"Commun. ACM"},{"key":"ref_76","doi-asserted-by":"crossref","unstructured":"Lee, C.H., Liu, Z., Wu, L., and Luo, P. (2020, January 13\u201319). MaskGAN: Towards Diverse and Interactive Facial Image Manipulation. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Seattle, WA, USA.","DOI":"10.1109\/CVPR42600.2020.00559"},{"key":"ref_77","unstructured":"Lin, D., Xu, G., Wang, X., Wang, Y., Sun, X., and Fu, K. (2019). A remote sensing image dataset for cloud removal. arXiv."},{"key":"ref_78","doi-asserted-by":"crossref","first-page":"5866","DOI":"10.1109\/TGRS.2020.3024744","article-title":"Multisensor data fusion for cloud removal in global and all-season sentinel-2 imagery","volume":"59","author":"Ebel","year":"2020","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_79","doi-asserted-by":"crossref","unstructured":"Liu, Z., Luo, P., Wang, X., and Tang, X. (2015, January 7\u201313). Deep learning face attributes in the wild. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Santiago, Chile.","DOI":"10.1109\/ICCV.2015.425"},{"key":"ref_80","doi-asserted-by":"crossref","unstructured":"Korhonen, J., and You, J. (2012, January 5\u20137). Peak signal-to-noise ratio revisited: Is simple beautiful?. Proceedings of the 2012 Fourth International Workshop on Quality of Multimedia Experience, Melbourne, Australia.","DOI":"10.1109\/QoMEX.2012.6263880"},{"key":"ref_81","first-page":"7","article-title":"Structural similarity measure for color images","volume":"43","author":"Hassan","year":"2012","journal-title":"Int. J. Comput. Appl."},{"key":"ref_82","unstructured":"Wang, Z., Simoncelli, E.P., and Bovik, A.C. (2003, January 9\u201312). Multiscale structural similarity for image quality assessment. Proceedings of the Thirty-Seventh Asilomar Conference on Signals, Systems & Computers, Pacific Grove, CA, USA."},{"key":"ref_83","doi-asserted-by":"crossref","first-page":"2378","DOI":"10.1109\/TIP.2011.2109730","article-title":"FSIM: A feature similarity index for image quality assessment","volume":"20","author":"Zhang","year":"2011","journal-title":"IEEE Trans. Image Process."},{"key":"ref_84","doi-asserted-by":"crossref","first-page":"1247","DOI":"10.5194\/gmd-7-1247-2014","article-title":"Root mean square error (RMSE) or mean absolute error (MAE)?\u2014Arguments against avoiding RMSE in the literature","volume":"7","author":"Chai","year":"2014","journal-title":"Geosci. Model Dev."},{"key":"ref_85","doi-asserted-by":"crossref","unstructured":"Zhang, R., Isola, P., Efros, A.A., Shechtman, E., and Wang, O. (2018, January 18\u201323). The unreasonable effectiveness of deep features as a perceptual metric. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Salt Lake City, UT, USA.","DOI":"10.1109\/CVPR.2018.00068"},{"key":"ref_86","first-page":"8024","article-title":"Pytorch: An imperative style, high-performance deep learning library","volume":"32","author":"Paszke","year":"2019","journal-title":"Adv. Neural Inf. Process. Syst."},{"key":"ref_87","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1145\/3065386","article-title":"Imagenet classification with deep convolutional neural networks","volume":"60","author":"Krizhevsky","year":"2017","journal-title":"Commun. ACM"},{"key":"ref_88","unstructured":"Loshchilov, I., and Hutter, F. (2017). Decoupled weight decay regularization. arXiv."}],"container-title":["Remote Sensing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2072-4292\/15\/5\/1321\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,10]],"date-time":"2025-10-10T18:43:50Z","timestamp":1760121830000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2072-4292\/15\/5\/1321"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,2,27]]},"references-count":88,"journal-issue":{"issue":"5","published-online":{"date-parts":[[2023,3]]}},"alternative-id":["rs15051321"],"URL":"https:\/\/doi.org\/10.3390\/rs15051321","relation":{},"ISSN":["2072-4292"],"issn-type":[{"value":"2072-4292","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,2,27]]}}}