{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,24]],"date-time":"2026-06-24T16:13:48Z","timestamp":1782317628748,"version":"3.54.5"},"reference-count":60,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2024,9,3]],"date-time":"2024-09-03T00:00:00Z","timestamp":1725321600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,9,3]],"date-time":"2024-09-03T00:00:00Z","timestamp":1725321600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Vis. Intell."],"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Face reshaping aims to adjust the shape of a face in a portrait image to make the face aesthetically beautiful, which has many potential applications. Existing methods 1) operate on the pre-defined facial landmarks, leading to artifacts and distortions due to the limited number of landmarks, 2) synthesize new faces based on segmentation masks or sketches, causing generated faces to look dissatisfied due to the losses of skin details and difficulties in dealing with hair and background blurring, and 3) project the positions of the deformed feature points from the 3D face model to the 2D image, making the results unrealistic because of the misalignment between feature points. In this paper, we propose a novel method named face shape transfer (FST) via semantic warping, which can transfer both the overall face and individual components (e.g., eyes, nose, and mouth) of a reference image to the source image. To achieve controllability at the component level, we introduce five encoding networks, which are designed to learn feature embedding specific to different face components. To effectively exploit the features obtained from semantic parsing maps at different scales, we employ a straightforward method of directly connecting all layers within the global dense network. This direct connection facilitates maximum information flow between layers, efficiently utilizing diverse scale semantic parsing information. To avoid deformation artifacts, we introduce a spatial transformer network, allowing the network to handle different types of semantic warping effectively. To facilitate extensive evaluation, we construct a large-scale high-resolution face dataset, which contains 14,000 images with a resolution of 1024 \u00d7 1024. Superior performance of our method is demonstrated by qualitative and quantitative experiments on the benchmark dataset.<\/jats:p>","DOI":"10.1007\/s44267-024-00058-7","type":"journal-article","created":{"date-parts":[[2024,9,3]],"date-time":"2024-09-03T08:02:21Z","timestamp":1725350541000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":11,"title":["Face shape transfer via semantic warping"],"prefix":"10.1007","volume":"2","author":[{"given":"Zonglin","family":"Li","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiaoqian","family":"Lv","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wei","family":"Yu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qinglin","family":"Liu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jingbo","family":"Lin","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shengping","family":"Zhang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,9,3]]},"reference":[{"issue":"4","key":"58_CR1","doi-asserted-by":"publisher","first-page":"711","DOI":"10.1007\/s12650-017-0420-z","volume":"20","author":"G. K. Suryanarayana","year":"2017","unstructured":"Suryanarayana, G. K., & Dubey, R. (2017). Image analyses of supersonic air-intake buzz and control by natural ventilation. Journal of Visualization, 20(4), 711\u2013727.","journal-title":"Journal of Visualization"},{"key":"58_CR2","doi-asserted-by":"publisher","DOI":"10.1016\/j.image.2020.116127","volume":"93","author":"L. Liu","year":"2021","unstructured":"Liu, L., Yu, H., Wang, S., Wan, L., & Han, S. (2021). Learning shape and texture progression for young child face aging. Signal Processing. Image Communication, 93, 116127.","journal-title":"Signal Processing. Image Communication"},{"key":"58_CR3","doi-asserted-by":"publisher","first-page":"134","DOI":"10.1016\/j.neucom.2014.09.093","volume":"172","author":"X. Fan","year":"2016","unstructured":"Fan, X., Chai, Z., Feng, Y., Wang, Y., Wang, S., & Luo, Z. (2016). An efficient mesh-based face beautifier on mobile devices. Neurocomputing, 172, 134\u2013142.","journal-title":"Neurocomputing"},{"key":"58_CR4","first-page":"1","volume-title":"Proceedings of the 13th European conference on computer vision","author":"J. Zhang","year":"2014","unstructured":"Zhang, J., Shan, S., Kan, M., & Chen, X. (2014). Coarse-to-fine auto-encoder networks (CFAN) for real-time face alignment. In D. J. Fleet, T. Pajdla, B. Schiele, et al. (Eds.), Proceedings of the 13th European conference on computer vision (pp. 1\u201316). Cham: Springer."},{"issue":"4","key":"58_CR5","doi-asserted-by":"publisher","first-page":"889","DOI":"10.1007\/s12650-017-0424-8","volume":"20","author":"F. J. A. Alvarez","year":"2017","unstructured":"Alvarez, F. J. A., Parra, E. B. B., & Tubio, F. M. (2017). Improving graphic expression training with 3D models. Journal of Visualization, 20(4), 889\u2013904.","journal-title":"Journal of Visualization"},{"key":"58_CR6","first-page":"1","volume-title":"Proceedings of the international conference on computer graphics and interactive techniques","author":"D. Vlasic","year":"2006","unstructured":"Vlasic, D., Brand, M., Pfister, H., & Popovic, J. (2006). Face transfer with multilinear models. In J. W. Finnegan & D. Shreiner (Eds.), Proceedings of the international conference on computer graphics and interactive techniques (pp. 1\u20138). New York: ACM."},{"key":"58_CR7","first-page":"2387","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition","author":"J. Thies","year":"2016","unstructured":"Thies, J., Zollhofer, M., Stamminger, M., Theobalt, C., & Nie\u00dfner, M. (2016). Face2face: real-time face capture and reenactment of RGB videos. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 2387\u20132395). Piscataway: IEEE."},{"issue":"6","key":"58_CR8","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2508363.2508417","volume":"32","author":"T. Neumann","year":"2013","unstructured":"Neumann, T., Varanasi, K., Wenger, S., Wacker, M., Magnor, M., & Theobalt, C. (2013). Sparse localized deformation components. ACM Transactions on Graphics, 32(6), 1\u201310.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR9","first-page":"40","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"H. Chang","year":"2018","unstructured":"Chang, H., Lu, J., Yu, F., & Finkelstein, A. (2018). Pairedcyclegan: asymmetric style transfer for applying and removing makeup. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 40\u201348). Piscataway: IEEE."},{"key":"58_CR10","first-page":"31","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"H. Yang","year":"2018","unstructured":"Yang, H., Huang, D., Wang, Y., & Jain, A. K. (2018). Learning face age progression: a pyramid architecture of GANs. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 31\u201339). Piscataway: IEEE."},{"key":"58_CR11","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/1360612.1360637","volume":"27","author":"T. Leyvand","year":"2008","unstructured":"Leyvand, T., Cohen-Or, D., Dror, G., & Lischinski, D. (2008). Data-driven enhancement of facial attractiveness. ACM Transactions on Graphics, 27, 1\u20139.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR12","first-page":"4217","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition","author":"P. Garrido","year":"2014","unstructured":"Garrido, P., Valgaerts, L., Rehmsen, O., Thormahlen, T., Perez, P., & Theobalt, C. (2014). Automatic face reenactment. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 4217\u20134224). Piscataway: IEEE."},{"key":"58_CR13","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/1360612.1360638","volume":"27","author":"D. Bitouk","year":"2008","unstructured":"Bitouk, D., Kumar, N., Dhillon, S., Belhumeur, P., & Nayar, S. K. (2008). Face swapping: automatically replacing faces in photographs. ACM Transactions on Graphics, 27, 1\u20138.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR14","first-page":"3436","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"S. Gu","year":"2019","unstructured":"Gu, S., Bao, J., Yang, H., Chen, D., Wen, F., & Yuan, L. (2019). Mask-guided portrait editing with conditional GANs. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 3436\u20133445). Piscataway: IEEE."},{"key":"58_CR15","first-page":"8188","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"Y. Choi","year":"2020","unstructured":"Choi, Y., Uh, Y., Yoo, J., & Ha, J.-W. (2020). Stargan v2: diverse image synthesis for multiple domains. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 8188\u20138197). Piscataway: IEEE."},{"key":"58_CR16","first-page":"13518","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"Z. Chen","year":"2020","unstructured":"Chen, Z., Wang, C., Yuan, B., & Tao, D. (2020). PuppeteerGAN: arbitrary portrait animation with semantic-aware appearance transformation. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 13518\u201313527). Piscataway: IEEE."},{"key":"58_CR17","unstructured":"Ling, H., Kreis, K., Li, D., Kim, S. W., Torralba, A., & Fidler, S. (2021). EditGAN: high-precision semantic image editing. arXiv preprint. arXiv:2111.03186."},{"issue":"3","key":"58_CR18","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3447648","volume":"40","author":"R. Abdal","year":"2021","unstructured":"Abdal, R., Zhu, P., Mitra, N. J., & Wonka, P. (2021). Styleflow: attribute-conditioned exploration of stylegan-generated images using conditional continuous normalizing flows. ACM Transactions on Graphics, 40(3), 1\u201321.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR19","doi-asserted-by":"crossref","unstructured":"Chen, S.-Y., Liu, F.-L., Lai, Y.-K., Rosin, P. L., Li, C., Fu, H., et\u00a0al. (2021). Deepfaceediting: deep face generation and editing with disentangled geometry and appearance control. ArXiv preprint. arXiv:2105.08935.","DOI":"10.1145\/3450626.3459760"},{"issue":"11","key":"58_CR20","doi-asserted-by":"publisher","first-page":"139","DOI":"10.1145\/3422622","volume":"63","author":"I. Goodfellow","year":"2020","unstructured":"Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., et al. (2020). Generative adversarial networks. Communications of the ACM, 63(11), 139\u2013144.","journal-title":"Communications of the ACM"},{"issue":"5","key":"58_CR21","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2629494","volume":"33","author":"J. Liao","year":"2014","unstructured":"Liao, J., Lima, R. S., Nehab, D., Hoppe, H., Sander, P. V., & Yu, J. (2014). Automating image morphing using structural similarity on a halfway domain. ACM Transactions on Graphics, 33(5), 1\u201312.","journal-title":"ACM Transactions on Graphics"},{"issue":"1","key":"58_CR22","doi-asserted-by":"publisher","first-page":"77","DOI":"10.1109\/MCG.2018.011461529","volume":"38","author":"H. Zhao","year":"2018","unstructured":"Zhao, H., Jin, X., Huang, X., Chai, M., & Zhou, K. (2018). Parametric reshaping of portrait images for weight-change. IEEE Computer Graphics and Applications, 38(1), 77\u201390.","journal-title":"IEEE Computer Graphics and Applications"},{"key":"58_CR23","first-page":"187","volume-title":"Proceedings of the 26th annual conference on computer graphics and interactive techniques","author":"V. Blanz","year":"1999","unstructured":"Blanz, V., & Vetter, T. (1999). A morphable model for the synthesis of 3d faces. In Proceedings of the 26th annual conference on computer graphics and interactive techniques (pp. 187\u2013194). Piscataway: IEEE."},{"key":"58_CR24","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2010324.1964971","volume":"30","author":"J. R. Tena","year":"2011","unstructured":"Tena, J. R., De la Torre, F., & Matthews, I. (2011). Interactive region-based linear 3D face models. ACM Transactions on Graphics, 30, 1\u201310.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR25","doi-asserted-by":"publisher","first-page":"246","DOI":"10.1016\/j.eswa.2019.02.008","volume":"126","author":"L. Male\u0161","year":"2019","unstructured":"Male\u0161, L., Mar\u010deti\u0107, D., & Ribari\u0107, S. (2019). A multi-agent dynamic system for robust multi-face tracking. Expert Systems with Applications, 126, 246\u2013264.","journal-title":"Expert Systems with Applications"},{"issue":"4","key":"58_CR26","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/2010324.1964956","volume":"30","author":"I. Kemelmacher-Shlizerman","year":"2011","unstructured":"Kemelmacher-Shlizerman, I., Shechtman, E., Garg, R., & Seitz, S. M. (2011). Exploring photobios. ACM Transactions on Graphics, 30(4), 1\u201310.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR27","first-page":"3607","volume-title":"Proceedings of the IEEE international conference on computer vision","author":"T. Hassner","year":"2013","unstructured":"Hassner, T. (2013). Viewing real-world faces in 3d. In Proceedings of the IEEE international conference on computer vision (pp. 3607\u20133614). Piscataway: IEEE."},{"key":"58_CR28","first-page":"5771","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"E. Collins","year":"2020","unstructured":"Collins, E., Bala, R., Price, B., & Susstrunk, S. (2020). Editing in style: uncovering the local semantics of gans. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 5771\u20135780). Piscataway: IEEE."},{"key":"58_CR29","first-page":"5134","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"Y. Alharbi","year":"2020","unstructured":"Alharbi, Y., & Wonka, P. (2020). Disentangled image generation through structured noise injection. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 5134\u20135142). Piscataway: IEEE."},{"key":"58_CR30","first-page":"5104","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"P. Zhu","year":"2020","unstructured":"Zhu, P., Abdal, R., Qin, Y., & Wonka, P. (2020). Sean: image synthesis with semantic region-adaptive normalization. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 5104\u20135113). Piscataway: IEEE."},{"key":"58_CR31","first-page":"3653","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"F. Zhan","year":"2019","unstructured":"Zhan, F., Zhu, H., & Lu, S. (2019). Spatial fusion GAN for image synthesis. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 3653\u20133662). Piscataway: IEEE."},{"key":"58_CR32","first-page":"7457","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"A. Shocher","year":"2020","unstructured":"Shocher, A., Gandelsman, Y., Mosseri, I., Yarom, M., Irani, M., Freeman, W. T., et al. (2020). Semantic pyramid for image generation. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 7457\u20137466). Piscataway: IEEE."},{"key":"58_CR33","first-page":"9243","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"Y. Shen","year":"2020","unstructured":"Shen, Y., Gu, J., Tang, X., & Zhou, B. (2020). Interpreting the latent space of GANs for semantic face editing. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 9243\u20139252). Piscataway: IEEE."},{"key":"58_CR34","first-page":"9786","volume-title":"International conference on machine learning","author":"A. Voynov","year":"2020","unstructured":"Voynov, A., & Babenko, A. (2020). Unsupervised discovery of interpretable directions in the GAN latent space. In International conference on machine learning (pp. 9786\u20139796). Stroudsburg: International Machine Learning Society."},{"key":"58_CR35","first-page":"7198","volume-title":"Proceedings of the 34th international conference on neural information processing systems","author":"T. Park","year":"2020","unstructured":"Park, T., Zhu, J.-Y., Wang, O., Lu, J., Shechtman, E., Efros, A., et al. (2020). Swapping autoencoder for deep image manipulation. In Proceedings of the 34th international conference on neural information processing systems (Vol.\u00a033, pp. 7198\u20137211). Red Hook: Curran Associates."},{"key":"58_CR36","first-page":"2337","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"T. Park","year":"2019","unstructured":"Park, T., Liu, M.-Y., Wang, T.-C., & Zhu, J.-Y. (2019). Semantic image synthesis with spatially-adaptive normalization. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 2337\u20132346). Piscataway: IEEE."},{"issue":"4","key":"58_CR37","doi-asserted-by":"publisher","first-page":"72","DOI":"10.1145\/3386569.3392386","volume":"39","author":"S.-Y. Chen","year":"2020","unstructured":"Chen, S.-Y., Su, W., Gao, L., Xia, S., & Fu, H. (2020). DeepFaceDrawing: deep generation of face images from sketches. ACM Transactions on Graphics, 39(4), 72.","journal-title":"ACM Transactions on Graphics"},{"issue":"1","key":"58_CR38","first-page":"1","volume":"41","author":"A. Chen","year":"2022","unstructured":"Chen, A., Liu, R., Xie, L., Chen, Z., Su, H., & Yu, J. (2022). SofGAN: a portrait image generator with dynamic styling. ACM Transactions on Graphics, 41(1), 1\u201326.","journal-title":"ACM Transactions on Graphics"},{"key":"58_CR39","first-page":"3282","volume-title":"Proceedings of the 2019 IEEE international conference on image processing","author":"W. Chu","year":"2019","unstructured":"Chu, W., Hung, W.-C., Tsai, Y.-H., Cai, D., & Yang, M.-H. (2019). Weakly-supervised caricature face parsing through domain adaptation. In Proceedings of the 2019 IEEE international conference on image processing (pp. 3282\u20133286). Piscataway: IEEE."},{"key":"58_CR40","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2023.103740","volume":"234","author":"Z. Li","year":"2023","unstructured":"Li, Z., Zhang, S., Zhang, Z., Meng, Q., Liu, Q., & Zhou, H. (2023). Attention guided domain alignment for conditional face image generation. Computer Vision and Image Understanding, 234, 103740.","journal-title":"Computer Vision and Image Understanding"},{"key":"58_CR41","first-page":"8798","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"T.-C. Wang","year":"2018","unstructured":"Wang, T.-C., Liu, M.-Y., Zhu, J.-Y., Tao, A., Kautz, J., & Catanzaro, B. (2018). High-resolution image synthesis and semantic manipulation with conditional GANs. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 8798\u20138807). Piscataway: IEEE."},{"issue":"5","key":"58_CR42","doi-asserted-by":"publisher","first-page":"603","DOI":"10.1109\/34.1000236","volume":"24","author":"D. Comaniciu","year":"2002","unstructured":"Comaniciu, D., & Meer, P. (2002). Mean shift: a robust approach toward feature space analysis. IEEE Transactions on Pattern Analysis and Machine Intelligence, 24(5), 603\u2013619.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"issue":"86","key":"58_CR43","first-page":"2579","volume":"9","author":"L. van der Maaten","year":"2008","unstructured":"van der Maaten, L., & Hinton, G. (2008). Visualizing data using t-SNE. Journal of Machine Learning Research, 9(86), 2579\u20132605.","journal-title":"Journal of Machine Learning Research"},{"key":"58_CR44","first-page":"2017","volume-title":"Proceedings of the 29th international conference on neural information processing systems","author":"M. Jaderberg","year":"2015","unstructured":"Jaderberg, M., Simonyan, K., Zisserman, A., & Kavukcuoglu, K. (2015). Spatial transformer networks. In Proceedings of the 29th international conference on neural information processing systems (pp. 2017\u20132025). Red Hook: Curran Associates."},{"key":"58_CR45","first-page":"9455","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"C.-H. Lin","year":"2018","unstructured":"Lin, C.-H., Yumer, E., Wang, O., Shechtman, E., & Lucey, S. (2018). ST-GAN: spatial transformer generative adversarial networks for image compositing. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 9455\u20139464). Piscataway: IEEE."},{"key":"58_CR46","first-page":"4681","volume-title":"Proceedings of the 14thProceedings of the IEEE conference on computer vision and pattern recognition","author":"C. Ledig","year":"2017","unstructured":"Ledig, C., Theis, L., Husz\u00e1r, F., Caballero, J., Cunningham, A., Acosta, A., et al. (2017). Photo-realistic single image super-resolution using a generative adversarial network. In B. Leibe, J. Matas, N. Sebe, et al. (Eds.), Proceedings of the 14thProceedings of the IEEE conference on computer vision and pattern recognition (pp. 4681\u20134690). Piscataway: IEEE."},{"key":"58_CR47","first-page":"694","volume-title":"Proceedings of the 14th European conference on computer vision","author":"J. Johnson","year":"2016","unstructured":"Johnson, J., Alahi, A., & Li, F.-F. (2016). Perceptual losses for real-time style transfer and super-resolution. In Proceedings of the 14th European conference on computer vision (pp. 694\u2013711). Cham: Springer."},{"key":"58_CR48","first-page":"1","volume-title":"Proceedings of the 3rd international conference on learning representations","author":"K. Simonyan","year":"2015","unstructured":"Simonyan, K., & Zisserman, A. (2015). Very deep convolutional networks for large-scale image recognition. In Y. Bengio & Y. LeCun (Eds.), Proceedings of the 3rd international conference on learning representations (pp. 1\u201314). San Diego, USA."},{"issue":"9","key":"58_CR49","doi-asserted-by":"publisher","first-page":"2663","DOI":"10.1007\/s11263-021-01489-1","volume":"129","author":"W. Chu","year":"2021","unstructured":"Chu, W., Hung, W.-C., Tsai, Y.-H., Chang, Y.-T., Li, Y., Cai, D., et al. (2021). Learning to caricature via semantic shape transform. International Journal of Computer Vision, 129(9), 2663\u20132679.","journal-title":"International Journal of Computer Vision"},{"key":"58_CR50","first-page":"4401","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition","author":"T. Karras","year":"2019","unstructured":"Karras, T., Laine, S., & Aila, T. (2019). A style-based generator architecture for generative adversarial networks. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition (pp. 4401\u20134410). Piscataway: IEEE."},{"key":"58_CR51","first-page":"214","volume-title":"International conference on machine learning","author":"M. Arjovsky","year":"2017","unstructured":"Arjovsky, M., Chintala, S., & Bottou, L. (2017). Wasserstein generative adversarial networks. In International conference on machine learning (pp. 214\u2013223). Stroudsburg: International Machine Learning Society."},{"key":"58_CR52","first-page":"2961","volume-title":"Proceedings of the IEEE international conference on computer vision","author":"K. He","year":"2017","unstructured":"He, K., Gkioxari, G., Doll\u00e1r, P., & Girshick, R. (2017). Mask R-CNN. In Proceedings of the IEEE international conference on computer vision (pp. 2961\u20132969). Piscataway: IEEE."},{"key":"58_CR53","doi-asserted-by":"crossref","unstructured":"Lee, C.-H., Liu, Z., Wu, L., & Luo, P. (2019). MaskGAN: towards diverse and interactive facial image manipulation. arXiv preprint. arXiv:1907.11922.","DOI":"10.1109\/CVPR42600.2020.00559"},{"key":"58_CR54","first-page":"2439","volume-title":"Proceedings of the IEEE international conference on computer vision","author":"R. Huang","year":"2017","unstructured":"Huang, R., Zhang, S., Li, T., & He, R. (2017). Beyond face rotation: global and local perception GAN for photorealistic and identity preserving frontal view synthesis. In Proceedings of the IEEE international conference on computer vision (pp. 2439\u20132448). Piscataway: IEEE."},{"key":"58_CR55","first-page":"1","volume-title":"Proceedings of the 3rd international conference on learning representations","author":"D. P. Kingma","year":"2015","unstructured":"Kingma, D. P., & Ba, J. (2015). Adam: a method for stochastic optimization. In Y. Bengio & Y. LeCun (Eds.), Proceedings of the 3rd international conference on learning representations (pp. 1\u201315). San Diego, USA."},{"key":"58_CR56","doi-asserted-by":"publisher","DOI":"10.1016\/j.image.2019.115699","volume":"81","author":"P. Nousi","year":"2020","unstructured":"Nousi, P., Papadopoulos, S., Tefas, A., & Pitas, I. (2020). Deep autoencoders for attribute preserving face de-identification. Signal Processing. Image Communication, 81, 115699.","journal-title":"Signal Processing. Image Communication"},{"key":"58_CR57","doi-asserted-by":"crossref","unstructured":"Sun, J., Wang, X., Zhang, Y., Li, X., Zhang, Q., Liu, Y., et\u00a0al. (2021). FENeRF: face editing in neural radiance fields. arXiv preprint. arXiv:2111.15490.","DOI":"10.1109\/CVPR52688.2022.00752"},{"key":"58_CR58","first-page":"6626","volume-title":"Proceedings of the 31st international conference on neural information processing systems","author":"M. Heusel","year":"2017","unstructured":"Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., & Hochreiter, S. (2017). GANs trained by a two time-scale update rule converge to a local Nash equilibrium. In I. Guyon, U. von Luxburg, S. Bengio, et al. (Eds.), Proceedings of the 31st international conference on neural information processing systems (pp. 6626\u20136637). Red Hook: Curran Associates."},{"key":"58_CR59","first-page":"1","volume-title":"Proceedings of the 6th international conference on learning representations","author":"T. Karras","year":"2017","unstructured":"Karras, T., Aila, T., Laine, S., & Lehtinen, J. (2017). Progressive growing of GANs for improved quality, stability, and variation. In Proceedings of the 6th international conference on learning representations (pp. 1\u201326). Retrieved June 30, 2024, from https:\/\/openreview.net\/forum?id=Hk99zCeAb."},{"key":"58_CR60","doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., & Sun, J. (2015). Deep residual learning for image recognition. arXiv preprint. arXiv:1512.03385.","DOI":"10.1109\/CVPR.2016.90"}],"container-title":["Visual Intelligence"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44267-024-00058-7.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s44267-024-00058-7\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44267-024-00058-7.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,9,3]],"date-time":"2024-09-03T09:06:12Z","timestamp":1725354372000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s44267-024-00058-7"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,9,3]]},"references-count":60,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2024,12]]}},"alternative-id":["58"],"URL":"https:\/\/doi.org\/10.1007\/s44267-024-00058-7","relation":{},"ISSN":["2731-9008"],"issn-type":[{"value":"2731-9008","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,9,3]]},"assertion":[{"value":"4 September 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"22 July 2024","order":2,"name":"revised","label":"Revised","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"23 July 2024","order":3,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"3 September 2024","order":4,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare no competing interests.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"26"}}