{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,11]],"date-time":"2026-06-11T17:08:33Z","timestamp":1781197713758,"version":"3.54.1"},"reference-count":34,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2026,2,13]],"date-time":"2026-02-13T00:00:00Z","timestamp":1770940800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,4,17]],"date-time":"2026-04-17T00:00:00Z","timestamp":1776384000000},"content-version":"vor","delay-in-days":63,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["J. King Saud Univ. Comput. Inf. Sci."],"published-print":{"date-parts":[[2026,5]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Image inpainting is crucial for restoring missing or corrupted regions in images, with diverse applications in computer vision and digital content creation. Existing deep learning-based methods, however, often struggle with large or irregular missing regions due to limited contextual cues, leading to structural inconsistencies and semantic misalignment. To address these challenges, we propose Structure-Aware Image Inpainting (SAIN), a novel two-stage framework that integrates multi-scale structural priors and a mask-aware gated attention mechanism to guide the inpainting process. A structured encoder-decoder architecture is designed to extract multi-scale structural features of the complete edge maps generated by the edge generation network. These features from decoder-side reconstructions, rich in semantic and geometric information, are adaptively injected as structural priors into the content inpainting network, helping the model to balance fine detail restoration with global shape consistency. Both stages of SAIN are enhanced with a mask-aware gated attention module that explicitly encodes the geometry of missing regions and applies gated attention. This mechanism enables the network to focus on reliable contextual features while suppressing misleading signals near hole boundaries caused by corruption. This combination allows for the effective handling of large missing regions, ensuring robust inpainting while maintaining spatial and semantic coherence, even under complex mask conditions. We evaluate SAIN on multiple datasets and conduct extensive experiments, demonstrating its superior performance in inpainting complex structures. The proposed method significantly enhances geometric consistency, semantic fidelity, and robustness across a variety of degradation scenarios.<\/jats:p>","DOI":"10.1007\/s44443-026-00545-5","type":"journal-article","created":{"date-parts":[[2026,2,13]],"date-time":"2026-02-13T16:05:09Z","timestamp":1770998709000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["SAIN:structure-aware image inpainting for large missing areas"],"prefix":"10.1007","volume":"38","author":[{"given":"Dong","family":"Wang","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yaowen","family":"Kang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yibin","family":"Chen","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuefang","family":"Gao","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Songhua","family":"Xu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,2,13]]},"reference":[{"key":"545_CR1","doi-asserted-by":"publisher","DOI":"10.1109\/tmm.2024.3369897\/mm2","author":"S Chen","year":"2024","unstructured":"Chen S, Atapour-Abarghouei A, Shum HP (2024) Hint: high-quality inpainting transformer with mask-aware encoding and enhanced attention. IEEE Trans Multimedia. https:\/\/doi.org\/10.1109\/tmm.2024.3369897\/mm2","journal-title":"IEEE Trans Multimedia"},{"key":"545_CR2","doi-asserted-by":"crossref","unstructured":"Cao F, Xu Q, Ye H (2025) Adaptive prior and long-range dependency-based learners for image inpainting. IEEE Trans Circuits Syst Video Technol","DOI":"10.1109\/TCSVT.2025.3574529"},{"key":"545_CR3","doi-asserted-by":"publisher","unstructured":"Deng Y, Hui S, Zhou S, Meng D, Wang J (2022) T-former: an efficient transformer for image inpainting. In: Proceedings of the 30th ACM international conference on multimedia, pp 6559\u20136568. https:\/\/doi.org\/10.1145\/3503161.3548446","DOI":"10.1145\/3503161.3548446"},{"key":"545_CR4","doi-asserted-by":"publisher","first-page":"126717","DOI":"10.1016\/j.eswa.2025.126717","volume":"272","author":"Y Gui","year":"2025","unstructured":"Gui Y, Liu Y, Yan C, Kuang L, Chen Z (2025) Lightweight structure-guided network with hydra interaction attention and global\u2013local gating mechanism for high-resolution image inpainting. Expert Syst Appl 272:126717","journal-title":"Expert Syst Appl"},{"key":"545_CR5","doi-asserted-by":"publisher","unstructured":"Guo X, Yang H, Huang D (2021) Image inpainting via conditional texture and structure dual generation. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 14134\u201314143. https:\/\/doi.org\/10.1109\/iccv48922.2021.01387","DOI":"10.1109\/iccv48922.2021.01387"},{"issue":"4","key":"545_CR6","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3072959.3073659","volume":"36","author":"S Iizuka","year":"2017","unstructured":"Iizuka S, Simo-Serra E, Ishikawa H (2017) Globally and locally consistent image completion. ACM Trans Graphics (ToG) 36(4):1\u201314. https:\/\/doi.org\/10.1145\/3072959.3073659","journal-title":"ACM Trans Graphics (ToG)"},{"key":"545_CR7","unstructured":"Jocher G, Chaurasia A, Qiu J (2023) Ultralytics yolo (version 8.0. 0)[computer software]"},{"key":"545_CR8","doi-asserted-by":"publisher","unstructured":"Jo Y, Park J (2019) Sc-fegan: face editing generative adversarial network with user\u2019s sketch and color. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 1745\u20131753. https:\/\/doi.org\/10.1109\/iccv.2019.00183","DOI":"10.1109\/iccv.2019.00183"},{"key":"545_CR9","unstructured":"Kingma DP, Ba J (2017) Adam: a method for stochastic optimization. https:\/\/arxiv.org\/abs\/1412.6980"},{"key":"545_CR10","doi-asserted-by":"crossref","unstructured":"Kirillov A, Mintun E, Ravi N, Mao H, Rolland C, Gustafson L, Xiao T, Whitehead S, Berg AC, Lo W-Y et al (2023) Segment anything. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 4015\u20134026","DOI":"10.1109\/ICCV51070.2023.00371"},{"key":"545_CR11","doi-asserted-by":"publisher","unstructured":"Li Ha, Hu L, Hua Q, Yang M, Li X (2022) Image inpainting based on contextual coherent attention gan. J Circuits Syst Comput 31(12):2250209. https:\/\/doi.org\/10.1142\/S0218126622502097","DOI":"10.1142\/S0218126622502097"},{"key":"545_CR12","doi-asserted-by":"publisher","unstructured":"Liao L, Hu R, Xiao J, Wang Z (2018) Edge-aware context encoder for image inpainting. In: 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, pp 3156\u20133160. https:\/\/doi.org\/10.1109\/icassp.2018.8462549","DOI":"10.1109\/icassp.2018.8462549"},{"key":"545_CR13","doi-asserted-by":"publisher","unstructured":"Liu G, Reda FA, Shih KJ, Wang T-C, Tao A, Catanzaro B (2018) Image inpainting for irregular holes using partial convolutions. In: Proceedings of the European Conference on Computer Vision (ECCV), pp 85\u2013100. https:\/\/doi.org\/10.1007\/978-3-030-01252-6_6","DOI":"10.1007\/978-3-030-01252-6_6"},{"key":"545_CR14","doi-asserted-by":"publisher","unstructured":"Nazeri K, Ng E, Joseph T, Qureshi F, Ebrahimi M (2019) Edgeconnect: structure guided image inpainting using edge prediction. In: Proceedings of the IEEE\/CVF international conference on computer vision workshops, pp 0\u20130. https:\/\/doi.org\/10.1109\/iccvw.2019.00408","DOI":"10.1109\/iccvw.2019.00408"},{"key":"545_CR15","doi-asserted-by":"crossref","unstructured":"Portenier T, Hu Q, Szab\u00f3 A, Bigdeli SA, Favaro P, Zwicker M (2018) FaceShop: deep sketch-based face image editing. https:\/\/arxiv.org\/abs\/1804.08972","DOI":"10.1145\/3197517.3201393"},{"key":"545_CR16","doi-asserted-by":"publisher","unstructured":"Pathak D, Kr\u00e4henb\u00fchl P, Donahue J, Darrell T, Efros AA (2016) Context encoders: feature learning by inpainting. In: 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp 2536\u20132544. https:\/\/doi.org\/10.1109\/CVPR.2016.278","DOI":"10.1109\/CVPR.2016.278"},{"key":"545_CR17","doi-asserted-by":"crossref","unstructured":"Suvorov R, Logacheva E, Mashikhin A, Remizova A, Ashukha A, Silvestrov A, Kong N, Goka H, Park K, Lempitsky V (2021) Resolution-robust large mask inpainting with fourier convolutions. https:\/\/arxiv.org\/abs\/2109.07161","DOI":"10.1109\/WACV51458.2022.00323"},{"key":"545_CR18","doi-asserted-by":"crossref","unstructured":"Suvorov R, Logacheva E, Mashikhin A, Remizova A, Ashukha A, Silvestrov A, Kong N, Goka H, Park K, Lempitsky V (2022) Resolution-robust large mask inpainting with fourier convolutions. In: Proceedings of the IEEE\/CVF winter conference on applications of computer vision, pp 2149\u20132159","DOI":"10.1109\/WACV51458.2022.00323"},{"key":"545_CR19","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1109\/lgrs.2020.3021116","volume":"19","author":"M Shao","year":"2020","unstructured":"Shao M, Wang C, Wu T, Meng D (2020) Luo J (2020) Context-based multiscale unified network for missing data reconstruction in remote sensing images. IEEE Geosci Remote Sens Lett 19:1\u20135. https:\/\/doi.org\/10.1109\/lgrs.2020.3021116","journal-title":"IEEE Geosci Remote Sens Lett"},{"key":"545_CR20","unstructured":"Song Y, Yang C, Shen Y, Wang P, Huang Q, Kuo C-CJ (2018) SPG-Net: segmentation prediction and guidance network for image inpainting. https:\/\/arxiv.org\/abs\/1805.03356"},{"key":"545_CR21","doi-asserted-by":"publisher","unstructured":"Wei J, Long C, Zou H, Xiao C (2019) Shadow inpainting and removal using generative adversarial networks with slice convolutions 38(7):381\u2013392. https:\/\/doi.org\/10.1111\/cgf.13845. Wiley Online Library","DOI":"10.1111\/cgf.13845"},{"key":"545_CR22","doi-asserted-by":"crossref","unstructured":"Wan Z, Zhang J, Chen D, Liao J (2021) High-fidelity pluralistic image completion with transformers. In: 2021 IEEE\/CVF International Conference on Computer Vision (ICCV)","DOI":"10.1109\/ICCV48922.2021.00465"},{"key":"545_CR23","doi-asserted-by":"crossref","unstructured":"Xing M, Feng Z, Su Y, Oh C (2024) Learning by erasing: conditional entropy based transferable out-of-distribution detection. In: Proceedings of the AAAI conference on artificial intelligence, vol 38, pp 6261\u20136269","DOI":"10.1609\/aaai.v38i6.28444"},{"key":"545_CR24","doi-asserted-by":"publisher","first-page":"266","DOI":"10.1016\/j.isprsjprs.2020.11.020","volume":"171","author":"H Xu","year":"2021","unstructured":"Xu H, Tang X, Ai B, Gao X, Yang F, Wen Z (2021) Missing data reconstruction in vhr images based on progressive structure prediction and texture generation. ISPRS J Photogrammetry Remote Sens 171:266\u2013277. https:\/\/doi.org\/10.1016\/j.isprsjprs.2020.11.020","journal-title":"ISPRS J Photogrammetry Remote Sens"},{"key":"545_CR25","doi-asserted-by":"publisher","unstructured":"Xiong W, Yu J, Lin Z, Yang J, Lu X, Barnes C, Luo J (2019) Foreground-aware image inpainting. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 5840\u20135848. https:\/\/doi.org\/10.1109\/cvpr.2019.00599","DOI":"10.1109\/cvpr.2019.00599"},{"key":"545_CR26","doi-asserted-by":"publisher","unstructured":"Yu Y, Du D, Zhang L, Luo T (2022) Unbiased multi-modality guidance for image inpainting. In: European conference on computer vision. Springer, pp 668\u2013684. https:\/\/doi.org\/10.1007\/978-3-031-19787-1_38","DOI":"10.1007\/978-3-031-19787-1_38"},{"key":"545_CR27","doi-asserted-by":"publisher","unstructured":"Yu F, Koltun V (2015) Multi-scale context aggregation by dilated convolutions. arXiv:1511.07122. https:\/\/doi.org\/10.48550\/arXiv.1511.07122","DOI":"10.48550\/arXiv.1511.07122"},{"key":"545_CR28","doi-asserted-by":"publisher","unstructured":"Yu J, Lin Z, Yang J, Shen X, Lu X, Huang TS (2019) Free-form image inpainting with gated convolution. In: Proceedings of the IEEE\/CVF international conference on computer vision, pp 4471\u20134480. https:\/\/doi.org\/10.1109\/iccv.2019.00457","DOI":"10.1109\/iccv.2019.00457"},{"key":"545_CR29","doi-asserted-by":"crossref","unstructured":"Yang Y, Su Y, An S (2022) Vdssa: ventral & dorsal sequential self-attention autoencoder for cognitive-consistency disentanglement. In: Chinese Conference on Pattern Recognition and Computer Vision (PRCV). Springer, pp 693\u2013705","DOI":"10.1007\/978-3-031-18910-4_55"},{"key":"545_CR30","first-page":"13294","volume":"36","author":"Z Yue","year":"2023","unstructured":"Yue Z, Wang J, Loy CC (2023) Resshift: efficient diffusion model for image super-resolution by residual shifting. Adv Neural Inf Process Syst 36:13294\u201313307","journal-title":"Adv Neural Inf Process Syst"},{"key":"545_CR31","doi-asserted-by":"publisher","unstructured":"Zamir SW, Arora A, Khan, S. Hayat M, Khan FS, Yang M-H (2022) Restormer: efficient transformer for high-resolution image restoration. In: Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition, pp 5728\u20135739. https:\/\/doi.org\/10.1109\/cvpr52688.2022.00564","DOI":"10.1109\/cvpr52688.2022.00564"},{"issue":"6","key":"545_CR32","doi-asserted-by":"publisher","first-page":"1452","DOI":"10.1109\/TPAMI.2017.2723009","volume":"40","author":"B Zhou","year":"2018","unstructured":"Zhou B, Lapedriza A, Khosla A, Oliva A, Torralba A (2018) Places: a 10 million image database for scene recognition. IEEE Trans Pattern Anal Mach Intell 40(6):1452\u20131464. https:\/\/doi.org\/10.1109\/TPAMI.2017.2723009","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"545_CR33","doi-asserted-by":"publisher","unstructured":"Zheng H, Lin Z, Lu J, Cohen S, Shechtman E, Barnes C, Zhang J, Xu N, Amirghodsi S, Luo J (2022) Image inpainting with cascaded modulation gan and object-aware training. In: European conference on computer vision. Springer, pp 277\u2013296. https:\/\/doi.org\/10.1007\/978-3-031-19787-1_16","DOI":"10.1007\/978-3-031-19787-1_16"},{"key":"545_CR34","doi-asserted-by":"crossref","unstructured":"Zhang J, Liu J, Zhang H, Zhang J, Lian J (2025) Structural-prior guided bi-generative network for image inpainting. Pattern Recogn, 112432","DOI":"10.1016\/j.patcog.2025.112432"}],"container-title":["Journal of King Saud University Computer and Information Sciences"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s44443-026-00545-5","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44443-026-00545-5.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44443-026-00545-5.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,11]],"date-time":"2026-06-11T16:44:01Z","timestamp":1781196241000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s44443-026-00545-5"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,2,13]]},"references-count":34,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2026,5]]}},"alternative-id":["545"],"URL":"https:\/\/doi.org\/10.1007\/s44443-026-00545-5","relation":{},"ISSN":["1319-1578","2213-1248"],"issn-type":[{"value":"1319-1578","type":"print"},{"value":"2213-1248","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,2,13]]},"assertion":[{"value":"6 November 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 February 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"13 February 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"The authors declare no competing interests.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"130"}}