{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,12]],"date-time":"2025-10-12T01:08:42Z","timestamp":1760231322124,"version":"build-2065373602"},"reference-count":31,"publisher":"MDPI AG","issue":"18","license":[{"start":{"date-parts":[[2022,9,6]],"date-time":"2022-09-06T00:00:00Z","timestamp":1662422400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"National Natural Science Foundation of China","award":["62071071"],"award-info":[{"award-number":["62071071"]}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Sensors"],"abstract":"<jats:p>The synthesis between face sketches and face photos has important application values in law enforcement and digital entertainment. In cases of a lack of paired sketch-photo data, this paper proposes an unsupervised model to solve the problems of missing key facial details and a lack of realism in the synthesized images of existing methods. The model is built on the CycleGAN architecture. To retain more semantic information in the target domain, a multi-scale feature extraction module is inserted before the generator. In addition, the convolutional block attention module is introduced into the generator to enhance the ability of the model to extract important feature information. Via CBAM, the model improves the quality of the converted image and reduces the artifacts caused by image background interference. Next, in order to preserve more identity information in the generated photo, this paper constructs the multi-level cycle consistency loss function. Qualitative experiments on CUFS and CUFSF public datasets show that the facial details and edge structures synthesized by our model are clearer and more realistic. Meanwhile the performance indexes of structural similarity and peak signal-to-noise ratio in quantitative experiments are also significantly improved compared with other methods.<\/jats:p>","DOI":"10.3390\/s22186725","type":"journal-article","created":{"date-parts":[[2022,9,8]],"date-time":"2022-09-08T04:18:32Z","timestamp":1662610712000},"page":"6725","update-policy":"https:\/\/doi.org\/10.3390\/mdpi_crossmark_policy","source":"Crossref","is-referenced-by-count":2,"title":["Multi-Level Cycle-Consistent Adversarial Networks with Attention Mechanism for Face Sketch-Photo Synthesis"],"prefix":"10.3390","volume":"22","author":[{"given":"Danping","family":"Ren","sequence":"first","affiliation":[{"name":"Hebei Key Laboratory of Security Protection Information Sensing and Processing, Handan 056038, China"},{"name":"School of Information and Electrical Engineering, Hebei University of Engineering, Handan 056038, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiajun","family":"Yang","sequence":"additional","affiliation":[{"name":"Hebei Key Laboratory of Security Protection Information Sensing and Processing, Handan 056038, China"},{"name":"School of Information and Electrical Engineering, Hebei University of Engineering, Handan 056038, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhongcheng","family":"Wei","sequence":"additional","affiliation":[{"name":"Hebei Key Laboratory of Security Protection Information Sensing and Processing, Handan 056038, China"},{"name":"School of Information and Electrical Engineering, Hebei University of Engineering, Handan 056038, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"1968","published-online":{"date-parts":[[2022,9,6]]},"reference":[{"key":"ref_1","doi-asserted-by":"crossref","first-page":"1264","DOI":"10.1109\/TIP.2017.2651375","article-title":"Bayesian face sketch synthesis","volume":"26","author":"Wang","year":"2017","journal-title":"IEEE Trans. Image Process."},{"doi-asserted-by":"crossref","unstructured":"Zhu, J.Y., Park, T., Isola, P., and Efros, A.A. (2017, January 22\u201329). Unpaired image-to-image translation using cycle-consistent adversarial networks. Proceedings of the IEEE International Conference on Computer Vision, Venice, Italy.","key":"ref_2","DOI":"10.1109\/ICCV.2017.244"},{"doi-asserted-by":"crossref","unstructured":"Woo, S., Park, J., Lee, J.Y., and Kweon, I.S. (2018, January 8\u201314). Cbam: Convolutional block attention module. Proceedings of the European Conference on Computer Vision (ECCV), Munich, Germany.","key":"ref_3","DOI":"10.1007\/978-3-030-01234-2_1"},{"unstructured":"Liu, Q., Tang, X., Jin, H., Lu, H., and Ma, S. (2005, January 20\u201325). A nonlinear approach for face sketch synthesis and recognition. Proceedings of the 2005 IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR\u201905), San Diego, CA, USA.","key":"ref_4"},{"key":"ref_5","doi-asserted-by":"crossref","first-page":"1955","DOI":"10.1109\/TPAMI.2008.222","article-title":"Face photo-sketch synthesis and recognition","volume":"31","author":"Wang","year":"2008","journal-title":"IEEE Trans. Pattern Anal. Mach. Intell."},{"unstructured":"Zhou, H., Kuang, Z., and Wong, K.Y.K. (2012, January 16\u201321). Markov weight fields for face sketch synthesis. Proceedings of the 2012 IEEE Conference on Computer Vision and Pattern Recognition, Providence, RI, USA.","key":"ref_6"},{"key":"ref_7","doi-asserted-by":"crossref","first-page":"220","DOI":"10.1109\/TIP.2015.2501755","article-title":"Robust face sketch style synthesis","volume":"25","author":"Zhang","year":"2015","journal-title":"IEEE Trans. Image Process."},{"key":"ref_8","doi-asserted-by":"crossref","first-page":"328","DOI":"10.1109\/TIP.2016.2623485","article-title":"Content-adaptive sketch portrait generation by decompositional representation learning","volume":"26","author":"Zhang","year":"2016","journal-title":"IEEE Trans. Image Process."},{"unstructured":"Goodfellow, I., Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2014, January 8\u201313). Generative Adversarial Nets. Proceedings of the Advances in Neural Information Processing Systems 27 (NIPS 2014), Montreal, QC, Canada.","key":"ref_9"},{"doi-asserted-by":"crossref","unstructured":"Wang, L., Sindagi, V., and Patel, V. (2018, January 15\u201319). High-quality facial photo-sketch synthesis using multi-adversarial networks. Proceedings of the 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition, Xi\u2019an, China.","key":"ref_10","DOI":"10.1109\/FG.2018.00022"},{"key":"ref_11","doi-asserted-by":"crossref","first-page":"107249","DOI":"10.1016\/j.patcog.2020.107249","article-title":"Identity-aware CycleGAN for face photo-sketch synthesis and recognition","volume":"102","author":"Fang","year":"2020","journal-title":"Pattern Recognit."},{"doi-asserted-by":"crossref","unstructured":"Chao, W., Chang, L., Wang, X., Cheng, J., Deng, X., and Duan, F. (2019, January 22\u201325). High-fidelity face sketch-to-photo synthesis using generative adversarial network. Proceedings of the 2019 IEEE International Conference on Image Processing (ICIP), Taipei, Taiwan.","key":"ref_12","DOI":"10.1109\/ICIP.2019.8803549"},{"key":"ref_13","doi-asserted-by":"crossref","first-page":"3096","DOI":"10.1109\/TNNLS.2018.2890018","article-title":"A deep collaborative framework for face photo\u2013sketch synthesis","volume":"30","author":"Zhu","year":"2019","journal-title":"IEEE Trans. Neural Netw. Learn. Syst."},{"key":"ref_14","doi-asserted-by":"crossref","first-page":"4350","DOI":"10.1109\/TCYB.2020.2972944","article-title":"Toward Realistic Face Photo\u2013Sketch Synthesis via Composition-Aided GANs","volume":"51","author":"Yu","year":"2020","journal-title":"IEEE Trans. Cybern."},{"key":"ref_15","doi-asserted-by":"crossref","first-page":"108077","DOI":"10.1016\/j.patcog.2021.108077","article-title":"IsGAN: Identity-sensitive generative adversarial network for face photo-sketch synthesis","volume":"119","author":"Yan","year":"2021","journal-title":"Pattern Recognit."},{"unstructured":"Bahdanau, D., Cho, K., and Bengio, Y. (2014). Neural machine translation by jointly learning to align and translate. arXiv.","key":"ref_16"},{"doi-asserted-by":"crossref","unstructured":"Hu, J., Shen, L., and Sun, G. (2018, January 18\u201322). Squeeze-and-excitation networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Salt Lake City, UT, USA.","key":"ref_17","DOI":"10.1109\/CVPR.2018.00745"},{"doi-asserted-by":"crossref","unstructured":"Fu, J., Liu, J., Tian, H., Li, Y., Bao, Y., Fang, Z., and Lu, H. (2019, January 16\u201317). Dual attention network for scene segmentation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Long Beach, CA, USA.","key":"ref_18","DOI":"10.1109\/CVPR.2019.00326"},{"unstructured":"Zhang, H., Goodfellow, I., Metaxas, D., and Odena, A. (2019, January 9\u201315). Self-attention generative adversarial networks. Proceedings of the International Conference on Machine Learning, Long Beach, CA, USA.","key":"ref_19"},{"doi-asserted-by":"crossref","unstructured":"Isola, P., Zhu, J.Y., Zhou, T., and Efros, A.A. (2017, January 21\u201326). Image-to-Image Translation with Conditional Adversarial Networks. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Honolulu, HI, USA.","key":"ref_20","DOI":"10.1109\/CVPR.2017.632"},{"doi-asserted-by":"crossref","unstructured":"He, K., Zhang, X., Ren, S., and Sun, J. (2016, January 26\u201330). Deep residual learning for image recognition. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Las Vegas, NV, USA.","key":"ref_21","DOI":"10.1109\/CVPR.2016.90"},{"doi-asserted-by":"crossref","unstructured":"Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., and Rabinovich, A. (2015, January 7\u201312). Going deeper with convolutions. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, Boston, MA, USA.","key":"ref_22","DOI":"10.1109\/CVPR.2015.7298594"},{"doi-asserted-by":"crossref","unstructured":"Mao, X.D., Li, Q., Xie, H.R., Lau, R.Y., Wang, Z., and Smolley, S.P. (2017, January 22\u201329). Least Squares Generative Adversarial Networks. Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV), Venice, Italy.","key":"ref_23","DOI":"10.1109\/ICCV.2017.304"},{"doi-asserted-by":"crossref","unstructured":"Johnson, J., Alahi, A., and Fei-Fei, L. (2016, January 11\u201314). Perceptual losses for real-time style transfer and super-resolution. Proceedings of the European Conference on Computer Vision, Amsterdam, The Netherlands.","key":"ref_24","DOI":"10.1007\/978-3-319-46475-6_43"},{"doi-asserted-by":"crossref","unstructured":"Zhang, W., Wang, X., and Tang, X. (2011, January 20\u201325). Coupled information-theoretic encoding for face photo-sketch recognition. Proceedings of the 2011 IEEE Conference on Computer Vision and Pattern Recognition, Washington, DC, USA.","key":"ref_25","DOI":"10.1109\/CVPR.2011.5995324"},{"doi-asserted-by":"crossref","unstructured":"Liu, R., Ge, Y., Choi, C.L., Wang, X., and Li, H. (2021, January 10\u201317). Divco: Diverse conditional image synthesis via contrastive generative adversarial network. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, Montreal, QC, Canada.","key":"ref_26","DOI":"10.1109\/CVPR46437.2021.01611"},{"doi-asserted-by":"crossref","unstructured":"Xu, Y., Xie, S., Wu, W., Zhang, K., Gong, M., and Batmanghelich, K. (2022, January 19\u201323). Maximum Spatial Perturbation Consistency for Unpaired Image-to-Image Translation. Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition, New Orleans, LO, USA.","key":"ref_27","DOI":"10.1109\/CVPR52688.2022.01777"},{"key":"ref_28","first-page":"5","article-title":"The AR face database","volume":"3","author":"Benavente","year":"2007","journal-title":"Comput. Vis. Cent."},{"unstructured":"Messer, K., Matas, J., Kittler, J., Luettin, J., and Maitre, G. (1999, January 22\u201324). XM2VTSDB: The extended M2VTS database. Proceedings of the Second International Conference on Audio and Video-Based Biometric Person Authentication, Washington, DC, USA.","key":"ref_29"},{"key":"ref_30","doi-asserted-by":"crossref","first-page":"295","DOI":"10.1016\/S0262-8856(97)00070-X","article-title":"The FERET database and evaluation procedure for face-recognition algorithms","volume":"16","author":"Phillips","year":"1998","journal-title":"Image Vis. Comput."},{"key":"ref_31","doi-asserted-by":"crossref","first-page":"600","DOI":"10.1109\/TIP.2003.819861","article-title":"Image quality assessment: From error visibility to structural similarity","volume":"13","author":"Wang","year":"2004","journal-title":"IEEE Trans. Image Process."}],"container-title":["Sensors"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/18\/6725\/pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,11]],"date-time":"2025-10-11T00:24:07Z","timestamp":1760142247000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.mdpi.com\/1424-8220\/22\/18\/6725"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,9,6]]},"references-count":31,"journal-issue":{"issue":"18","published-online":{"date-parts":[[2022,9]]}},"alternative-id":["s22186725"],"URL":"https:\/\/doi.org\/10.3390\/s22186725","relation":{},"ISSN":["1424-8220"],"issn-type":[{"type":"electronic","value":"1424-8220"}],"subject":[],"published":{"date-parts":[[2022,9,6]]}}}