{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,21]],"date-time":"2026-07-21T22:56:56Z","timestamp":1784674616085,"version":"3.55.0"},"reference-count":40,"publisher":"MDPI AG","issue":"8","license":[{"start":{"date-parts":[[2025,8,4]],"date-time":"2025-08-04T00:00:00Z","timestamp":1754265600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Natural Science Foundation of China","award":["42171438"],"award-info":[{"award-number":["42171438"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["IJGI"],"abstract":"<jats:p>Traditional map style transfer methods are mostly based on GAN, which are either overly artistic at the expense of conveying information, or insufficiently aesthetic by simply changing the color scheme of the map image. These methods often struggle to balance style transfer with semantic preservation and lack consistency in their transfer effects. In recent years, diffusion models have made significant progress in the field of image processing and have shown great potential in image-style transfer tasks. Inspired by these advances, this paper presents a method for transferring real-world 3D scenes to a cartoon style without the need for additional input condition guidance. The method combines pre-trained LDM with LoRA models to achieve stable and high-quality style infusion. By integrating DDIM Inversion, ControlNet, and MultiDiffusion strategies, it achieves the cartoon style transfer of real-world 3D scenes through initial noise control, detail redrawing, and global coordination. Qualitative and quantitative analyses, as well as user studies, indicate that our method effectively injects a cartoon style while preserving the semantic content of the real-world 3D scene, maintaining a high degree of consistency in style transfer. This paper offers a new perspective for map style transfer.<\/jats:p>","DOI":"10.3390\/ijgi14080303","type":"journal-article","created":{"date-parts":[[2025,8,4]],"date-time":"2025-08-04T15:30:06Z","timestamp":1754321406000},"page":"303","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":6,"title":["Diffusion Model-Based Cartoon Style Transfer for Real-World 3D Scenes"],"prefix":"10.3390","volume":"14","author":[{"given":"Yuhang","family":"Chen","sequence":"first","affiliation":[{"name":"School of Geography and Information Engineering, China University of Geosciences, Wuhan 430078, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haoran","family":"Zhou","sequence":"additional","affiliation":[{"name":"Aerospace Information Research Institute, Chinese Academy of Sciences, Beijing 100190, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jing","family":"Chen","sequence":"additional","affiliation":[{"name":"School of Geography and Information Engineering, China University of Geosciences, Wuhan 430078, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5306-1163","authenticated-orcid":false,"given":"Nai","family":"Yang","sequence":"additional","affiliation":[{"name":"School of Geography and Information Engineering, China University of Geosciences, Wuhan 430078, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jing","family":"Zhao","sequence":"additional","affiliation":[{"name":"Provincial Surveying and Mapping Production Archives of Hubei, Wuhan 430074, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7286-5439","authenticated-orcid":false,"given":"Yi","family":"Chao","sequence":"additional","affiliation":[{"name":"School of Geography and Information Engineering, China University of Geosciences, Wuhan 430078, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"1968","published-online":{"date-parts":[[2025,8,4]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"179","DOI":"10.1179\/000870409X12488753453453","article-title":"Stylistic Diversity in European State 1: 50 000 Topographic Maps","volume":"46","author":"Kent","year":"2009","journal-title":"Cartogr. J."},{"key":"ref_2","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1080\/00087041.2019.1633103","article-title":"Cartographic Design as Visual Storytelling: Synthesis and Review of Map-Based Narratives, Genres, and Tropes","volume":"58","author":"Roth","year":"2021","journal-title":"Cartogr. J."},{"key":"ref_3","first-page":"61","article-title":"Expressive Map Design Based on Pop Art: Revisit of Semiology of Graphics?","volume":"73","author":"Christophe","year":"2012","journal-title":"Cartogr. Perspect."},{"key":"ref_4","first-page":"18","article-title":"Neural map style transfer exploration with GANs","volume":"8","author":"Christophe","year":"2022","journal-title":"Int. J. Cartogr."},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"9","DOI":"10.5194\/ica-proc-2-9-2019","article-title":"Projecting emotions from artworks to maps using neural style transfer","volume":"2","author":"Bogucka","year":"2019","journal-title":"Proc. ICA"},{"key":"ref_6","first-page":"115","article-title":"Transferring multiscale map styles using generative adversarial networks","volume":"5","author":"Kang","year":"2019","journal-title":"Int. J. Cartogr."},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"4388","DOI":"10.1109\/TGRS.2020.3021819","article-title":"SMAPGAN: Generative Adversarial Network-Based Semisupervised Styled Map Tile Generation Method","volume":"59","author":"Chen","year":"2021","journal-title":"IEEE Trans. Geosci. Remote Sens."},{"key":"ref_8","unstructured":"Ganguli, S., Garzon, P., and Glaser, N. (2019). GeoGAN: A Conditional GAN with Reconstruction and Style Loss to Generate Standard Layer of Maps from Satellite Images. arXiv."},{"key":"ref_9","doi-asserted-by":"crossref","first-page":"012041","DOI":"10.1088\/1742-6596\/1903\/1\/012041","article-title":"Map style transfer using pixel-to-pixel model","volume":"1903","author":"Jin","year":"2021","journal-title":"J. Phys. Conf. Ser."},{"key":"ref_10","doi-asserted-by":"crossref","unstructured":"Li, Z., Guan, R., Yu, Q., Chiang, Y.-Y., and Knoblock, C.A. (2021, January 2). Synthetic Map Generation to Provide Unlimited Training Data for Historical Map Text Detection. Proceedings of the 4th ACM SIGSPATIAL International Workshop on AI for Geographic Knowledge Discovery, Beijing, China.","DOI":"10.1145\/3486635.3491070"},{"key":"ref_11","unstructured":"Ho, J., Jain, A., and Abbeel, P. (2020). Denoising Diffusion Probabilistic Models. Advances in Neural Information Processing Systems 33, Proceedings of the Annual Conference on Neural Information Processing Systems, Virtual, 6\u201312 December 2020, Curran Associates, Inc.. Available online: https:\/\/proceedings.neurips.cc\/paper\/2020\/hash\/4c5bcfec8584af0d967f1ab10179ca4b-Abstract.html."},{"key":"ref_12","doi-asserted-by":"crossref","unstructured":"Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B. (2022, January 18\u201324). High-Resolution Image Synthesis with Latent Diffusion Models. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), New Orleans, LA, USA. Available online: https:\/\/openaccess.thecvf.com\/content\/CVPR2022\/html\/Rombach_High-Resolution_Image_Synthesis_With_Latent_Diffusion_Models_CVPR_2022_paper.html.","DOI":"10.1109\/CVPR52688.2022.01042"},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"266","DOI":"10.1109\/TVCG.2004.1272726","article-title":"Efficient example-based painting and synthesis of 2D directional texture","volume":"10","author":"Wang","year":"2004","journal-title":"IEEE Trans. Vis. Comput. Graph."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"1594","DOI":"10.1109\/TMM.2013.2265675","article-title":"Style Transfer Via Image Component Analysis","volume":"15","author":"Zhang","year":"2013","journal-title":"IEEE Trans. Multimed."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"3365","DOI":"10.1109\/TVCG.2019.2921336","article-title":"Neural Style Transfer: A Review","volume":"26","author":"Jing","year":"2020","journal-title":"IEEE Trans. Vis. Comput. Graph."},{"key":"ref_16","unstructured":"Zhang, C., Zhu, Y., and Zhu, S.-C. (February, January 27). MetaStyle: Three-Way Trade-off among Speed, Flexibility, and Quality in Neural Style Transfer. Proceedings of the AAAI Conference on Artificial Intelligence, Honolulu, HI, USA."},{"key":"ref_17","doi-asserted-by":"crossref","unstructured":"Zhang, Y., Tang, F., Dong, W., Huang, H., Ma, C., Lee, T.-Y., and Xu, C. (2022, January 7\u201311). Domain Enhanced Arbitrary Image Style Transfer via Contrastive Learning. Proceedings of the SIGGRAPH \u201822: Special Interest Group on Computer Graphics and Interactive Techniques Conference, Vancouver, BC, Canada.","DOI":"10.1145\/3528233.3530736"},{"key":"ref_18","doi-asserted-by":"crossref","unstructured":"Zhu, J.-Y., Park, T., Isola, P., and Efros, A.A. (2017, January 22\u201329). Unpaired Image-To-Image Translation Using Cycle-Consistent Adversarial Networks. Proceedings of the IEEE International Conference on Computer Vision (ICCV), Venice, Italy. Available online: https:\/\/openaccess.thecvf.com\/content_iccv_2017\/html\/Zhu_Unpaired_Image-To-Image_Translation_ICCV_2017_paper.html.","DOI":"10.1109\/ICCV.2017.244"},{"key":"ref_19","unstructured":"Brock, A., Donahue, J., and Simonyan, K. (2019). Large Scale GAN Training for High Fidelity Natural Image Synthesis. arXiv."},{"key":"ref_20","unstructured":"Miyato, T., Kataoka, T., Koyama, M., and Yoshida, Y. (2018). Spectral Normalization for Generative Adversarial Networks. arXiv."},{"key":"ref_21","doi-asserted-by":"crossref","unstructured":"Huang, N., Tang, F., Dong, W., and Xu, C. (2022, January 10\u201314). Draw Your Art Dream: Diverse Digital Art Synthesis with Multimodal Guided Diffusion. Proceedings of the 30th ACM International Conference on Multimedia, Lisbon, Portugal.","DOI":"10.1145\/3503161.3548282"},{"key":"ref_22","doi-asserted-by":"crossref","unstructured":"Wang, Z., Zhao, L., and Xing, W. (2023, January 1\u20136). StyleDiffusion: Controllable Disentangled Style Transfer via Diffusion Models. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Paris, France. Available online: https:\/\/openaccess.thecvf.com\/content\/ICCV2023\/html\/Wang_StyleDiffusion_Controllable_Disentangled_Style_Transfer_via_Diffusion_Models_ICCV_2023_paper.html.","DOI":"10.1109\/ICCV51070.2023.00706"},{"key":"ref_23","unstructured":"Li, S. (2024). DiffStyler: Diffusion-based Localized Image Style Transfer. arXiv."},{"key":"ref_24","doi-asserted-by":"crossref","unstructured":"Mokady, R., Hertz, A., Aberman, K., Pritch, Y., and Cohen-Or, D. (2023, January 17\u201324). NULL-Text Inversion for Editing Real Images Using Guided Diffusion Models. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Vancouver, BC, Canada. Available online: https:\/\/openaccess.thecvf.com\/content\/CVPR2023\/html\/Mokady_NULL-Text_Inversion_for_Editing_Real_Images_Using_Guided_Diffusion_Models_CVPR_2023_paper.html.","DOI":"10.1109\/CVPR52729.2023.00585"},{"key":"ref_25","doi-asserted-by":"crossref","unstructured":"Zhang, Y., Huang, N., Tang, F., Huang, H., Ma, C., Dong, W., and Xu, C. (2023, January 17\u201324). Inversion-Based Style Transfer with Diffusion Models. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Vancouver, BC, Canada. Available online: https:\/\/openaccess.thecvf.com\/content\/CVPR2023\/html\/Zhang_Inversion-Based_Style_Transfer_With_Diffusion_Models_CVPR_2023_paper.html.","DOI":"10.1109\/CVPR52729.2023.00978"},{"key":"ref_26","doi-asserted-by":"crossref","unstructured":"Alaluf, Y., Garibi, D., Patashnik, O., Averbuch-Elor, H., and Cohen-Or, D. (2023). Cross-Image Attention for Zero-Shot Appearance Transfer. arXiv.","DOI":"10.1145\/3641519.3657423"},{"key":"ref_27","doi-asserted-by":"crossref","unstructured":"Hertz, A., Voynov, A., Fruchter, S., and Cohen-Or, D. (2024). Style Aligned Image Generation via Shared Attention. arXiv.","DOI":"10.1109\/CVPR52733.2024.00457"},{"key":"ref_28","unstructured":"He, F., Li, G., Zhang, M., Yan, L., Si, L., and Li, F. (2024). FreeStyle: Free Lunch for Text-Guided Style Transfer Using Diffusion Models. arXiv."},{"key":"ref_29","doi-asserted-by":"crossref","unstructured":"Friedmannov\u00e1, L. (2009). What Can We Learn from the Masters? Color Schemas on Paintings as the Source for Color Ranges Applicable in Cartography. Cartography and Art, Springer.","DOI":"10.1007\/978-3-540-68569-2_9"},{"key":"ref_30","doi-asserted-by":"crossref","unstructured":"Isola, P., Zhu, J.-Y., Zhou, T., and Efros, A.A. (2017, January 21\u201326). Image-To-Image Translation with Conditional Adversarial Networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA. Available online: https:\/\/openaccess.thecvf.com\/content_cvpr_2017\/html\/Isola_Image-To-Image_Translation_With_CVPR_2017_paper.html.","DOI":"10.1109\/CVPR.2017.632"},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"5538","DOI":"10.1109\/TVCG.2023.3295122","article-title":"Adaptive color transfer from images to terrain visualizations","volume":"30","author":"Wu","year":"2023","journal-title":"IEEE Trans. Vis. Comput. Graph."},{"key":"ref_32","doi-asserted-by":"crossref","first-page":"289","DOI":"10.1080\/15230406.2021.1982009","article-title":"Adaptive transfer of color from images to maps and visualizations","volume":"49","author":"Wu","year":"2022","journal-title":"Cartogr. Geogr. Inf. Sci."},{"key":"ref_33","doi-asserted-by":"crossref","unstructured":"Zhang, L., Rao, A., and Agrawala, M. (2023, January 1\u20136). Adding Conditional Control to Text-to-Image Diffusion Models. Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV), Paris, France. Available online: https:\/\/openaccess.thecvf.com\/content\/ICCV2023\/html\/Zhang_Adding_Conditional_Control_to_Text-to-Image_Diffusion_Models_ICCV_2023_paper.html.","DOI":"10.1109\/ICCV51070.2023.00355"},{"key":"ref_34","unstructured":"Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., and Ganguli, S. (2015, January 6\u201311). Deep Unsupervised Learning using Nonequilibrium Thermodynamics. Proceedings of the 32nd International Conference on Machine Learning, Lille, France. Available online: https:\/\/proceedings.mlr.press\/v37\/sohl-dickstein15.html."},{"key":"ref_35","unstructured":"Kingma, D.P., and Welling, M. (2022). Auto-Encoding Variational Bayes. arXiv."},{"key":"ref_36","unstructured":"Hu, E.J., Shen, Y., Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W. (2021). LoRA: Low-Rank Adaptation of Large Language Models. arXiv."},{"key":"ref_37","unstructured":"Ho, J., and Salimans, T. (2022). Classifier-Free Diffusion Guidance. arXiv."},{"key":"ref_38","unstructured":"Bar-Tal, O., Yariv, L., Lipman, Y., and Dekel, T. (2023, January 23\u201329). MultiDiffusion: Fusing diffusion paths for controlled image generation. Proceedings of the 40th International Conference on Machine Learning, Honolulu, HI, USA."},{"key":"ref_39","unstructured":"Hertz, A., Mokady, R., Tenenbaum, J., Aberman, K., Pritch, Y., and Cohen-Or, D. (2022). Prompt-to-Prompt Image Editing with Cross Attention Control. arXiv."},{"key":"ref_40","unstructured":"Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., and Clark, J. (2021, January 18\u201324). Learning Transferable Visual Models from Natural Language Supervision. Proceedings of the 38th International Conference on Machine Learning, Virtual. Available online: https:\/\/proceedings.mlr.press\/v139\/radford21a.html."}],"container-title":["ISPRS International Journal of Geo-Information"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/2220-9964\/14\/8\/303\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T18:23:05Z","timestamp":1760034185000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/2220-9964\/14\/8\/303"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,8,4]]},"references-count":40,"journal-issue":{"issue":"8","published-online":{"date-parts":[[2025,8]]}},"alternative-id":["ijgi14080303"],"URL":"https:\/\/doi.org\/10.3390\/ijgi14080303","relation":{},"ISSN":["2220-9964"],"issn-type":[{"value":"2220-9964","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,8,4]]}}}