{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,13]],"date-time":"2026-06-13T09:24:50Z","timestamp":1781342690048,"version":"3.54.1"},"reference-count":58,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2023,1,6]],"date-time":"2023-01-06T00:00:00Z","timestamp":1672963200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2023,1,31]]},"abstract":"<jats:p>Face-mask occluded restoration aims at restoring the masked region of a human face, which has attracted increasing attention in the context of the COVID-19 pandemic. One major challenge of this task is the large visual variance of masks in the real world. To solve it we first construct a large-scale Face-mask Occluded Restoration (FMOR) dataset, which contains 5,500 unmasked images and 5,500 face-mask occluded images with various illuminations, and involves 1,100 subjects of different races, face orientations, and mask types. Moreover, we propose a Face-Mask Occluded Detection and Restoration (FMODR) framework, which can detect face-mask regions with large visual variations and restore them to realistic human faces. In particular, our FMODR contains a self-adaptive contextual attention module specifically designed for this task, which is able to exploit the contextual information and correlations of adjacent pixels for achieving high realism of the restored faces, which are however often neglected in existing contextual attention models. Our framework achieves state-of-the-art results of face restoration on three datasets, including CelebA, AR, and our FMOR datasets. Moreover, experimental results on AR and FMOR datasets demonstrate that our framework can significantly improve masked face recognition and verification performance.<\/jats:p>","DOI":"10.1145\/3524137","type":"journal-article","created":{"date-parts":[[2022,3,25]],"date-time":"2022-03-25T13:08:35Z","timestamp":1648213715000},"page":"1-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":13,"title":["Toward High-quality Face-Mask Occluded Restoration"],"prefix":"10.1145","volume":"19","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-7013-6913","authenticated-orcid":false,"given":"Lu","family":"Feihong","sequence":"first","affiliation":[{"name":"Xidian University, Xi\u2019an, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2330-0594","authenticated-orcid":false,"given":"Chen","family":"Hang","sequence":"additional","affiliation":[{"name":"Xidian University, Xi\u2019an, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0748-1583","authenticated-orcid":false,"given":"Li","family":"Kang","sequence":"additional","affiliation":[{"name":"Xidian University, Xi\u2019an, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0373-1274","authenticated-orcid":false,"given":"Deng","family":"Qiliang","sequence":"additional","affiliation":[{"name":"Xidian University, Xi\u2019an, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3508-756X","authenticated-orcid":false,"given":"Zhao","family":"Jian","sequence":"additional","affiliation":[{"name":"Institute of North Electronic Equipment, Beijing, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6105-6532","authenticated-orcid":false,"given":"Zhang","family":"Kaipeng","sequence":"additional","affiliation":[{"name":"The University of Tokyo, Tokyo, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8019-3740","authenticated-orcid":false,"given":"Han","family":"Hong","sequence":"additional","affiliation":[{"name":"Xidian University, Xi\u2019an, PRC"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,1,6]]},"reference":[{"key":"e_1_3_1_2_2","volume-title":"Proceedings of the IEEE CGC Workshop on Computational Geometry","author":"Arya Sunil","year":"1998","unstructured":"Sunil Arya and D. M. Mount. 1998. ANN: Library for approximate nearest neighbor searching. In Proceedings of the IEEE CGC Workshop on Computational Geometry, Providence, RI."},{"key":"e_1_3_1_3_2","doi-asserted-by":"crossref","first-page":"217","DOI":"10.1145\/364338.364405","volume-title":"Proceedings of the 2001 Symposium on Interactive 3D Graphics","author":"Ashikhmin Michael","year":"2001","unstructured":"Michael Ashikhmin. 2001. Synthesizing natural textures. In Proceedings of the 2001 Symposium on Interactive 3D Graphics. 217\u2013226."},{"key":"e_1_3_1_4_2","unstructured":"Samik Banerjee and Sukhendu Das. 2020. SD-GAN: Structural and denoising GAN reveals facial parts under occlusion[J]. arXiv preprint arXiv:2002.08448(2020)."},{"issue":"3","key":"e_1_3_1_5_2","first-page":"24","article-title":"PatchMatch: A randomized correspondence algorithm for structural image editing","volume":"28","author":"Barnes Connelly","year":"2009","unstructured":"Connelly Barnes, Eli Shechtman, Adam Finkelstein, and Dan B. Goldman. 2009. PatchMatch: A randomized correspondence algorithm for structural image editing. ACM Transactions on Graphics 28, 3 (2009), 24.","journal-title":"ACM Transactions on Graphics"},{"key":"e_1_3_1_6_2","first-page":"355","volume-title":"Proceedings of the 10th ACM International Conference on Multimedia","author":"Bornard Rapha\u00ebl","year":"2002","unstructured":"Rapha\u00ebl Bornard, Emmanuelle Lecan, Louis Laborelli, and Jean-Hugues Chenot. 2002. Missing data correction in still images and image sequences. In Proceedings of the 10th ACM International Conference on Multimedia. 355\u2013361."},{"issue":"3","key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/1970392.1970395","article-title":"Robust principal component analysis?","volume":"58","author":"Cand\u00e8s Emmanuel J.","year":"2011","unstructured":"Emmanuel J. Cand\u00e8s, Xiaodong Li, Yi Ma, and John Wright. 2011. Robust principal component analysis? The Journal of the ACM 58, 3 (2011), 1\u201337.","journal-title":"The Journal of the ACM"},{"key":"e_1_3_1_8_2","first-page":"11896","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Chen Chaofeng","year":"2021","unstructured":"Chaofeng Chen, Xiaoming Li, Lingbo Yang, Xianhui Lin, Lei Zhang, and Kwan-Yee K. Wong. 2021. Progressive semantic-aware style transformation for blind face restoration. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 11896\u201311905."},{"key":"e_1_3_1_9_2","doi-asserted-by":"crossref","first-page":"14567","DOI":"10.1109\/ACCESS.2018.2803787","article-title":"From eyes to face synthesis: A new approach for human-centered smart surveillance","volume":"6","author":"Chen Xiang","year":"2018","unstructured":"Xiang Chen, Linbo Qing, Xiaohai He, Jie Su, and Yonghong Peng. 2018. From eyes to face synthesis: A new approach for human-centered smart surveillance. IEEE Access 6 (2018), 14567\u201314575.","journal-title":"IEEE Access"},{"issue":"9","key":"e_1_3_1_10_2","doi-asserted-by":"crossref","first-page":"1200","DOI":"10.1109\/TIP.2004.833105","article-title":"Region filling and object removal by exemplar-based image inpainting","volume":"13","author":"Criminisi Antonio","year":"2004","unstructured":"Antonio Criminisi, Patrick P\u00e9rez, and Kentaro Toyama. 2004. Region filling and object removal by exemplar-based image inpainting. IEEE Transactions on Image Processing 13, 9 (2004), 1200\u20131212.","journal-title":"IEEE Transactions on Image Processing"},{"key":"e_1_3_1_11_2","first-page":"741","volume-title":"Proceedings of the European Conference on Computer Vision","author":"Deng Jiankang","year":"2020","unstructured":"Jiankang Deng, Jia Guo, Tongliang Liu, Mingming Gong, and Stefanos Zafeiriou. 2020. Sub-center arcface: Boosting face recognition by large-scale noisy web faces. In Proceedings of the European Conference on Computer Vision. Springer, 741\u2013757."},{"key":"e_1_3_1_12_2","first-page":"341","volume-title":"Proceedings of the CVPR","author":"Efros Alexei A.","year":"2001","unstructured":"Alexei A. Efros and William T. Freeman. 2001. Image quilting for texture synthesis and transfer. In Proceedings of the CVPR. ACM, 341\u2013346."},{"key":"e_1_3_1_13_2","first-page":"1033","volume-title":"Proceedings of the ICCV","author":"Efros Alexei A.","year":"1999","unstructured":"Alexei A. Efros and Thomas K. Leung. 1999. Texture synthesis by non-parametric sampling. In Proceedings of the ICCV. IEEE, 1033\u20131038."},{"issue":"1","key":"e_1_3_1_14_2","doi-asserted-by":"crossref","first-page":"149","DOI":"10.1109\/TSMCA.2007.909557","article-title":"The CAS-PEAL large-scale chinese face database and baseline evaluations","volume":"38","author":"Gao Wen","year":"2007","unstructured":"Wen Gao, Bo Cao, Shiguang Shan, Xilin Chen, Delong Zhou, Xiaohua Zhang, and Debin Zhao. 2007. The CAS-PEAL large-scale chinese face database and baseline evaluations. IEEE Transactions on Systems, Man, and Cybernetics-Part A: Systems and Humans 38, 1 (2007), 149\u2013161.","journal-title":"IEEE Transactions on Systems, Man, and Cybernetics-Part A: Systems and Humans"},{"key":"e_1_3_1_15_2","first-page":"2682","volume-title":"Proceedings of the CVPR","author":"Ge Shiming","year":"2017","unstructured":"Shiming Ge, Jia Li, Qiting Ye, and Zhao Luo. 2017. Detecting masked faces in the wild with lle-cnns. In Proceedings of the CVPR. 2682\u20132690."},{"key":"e_1_3_1_16_2","first-page":"2672","volume-title":"Proceedings of the NeurIPS","author":"Goodfellow Ian","year":"2014","unstructured":"Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative adversarial nets. In Proceedings of the NeurIPS. 2672\u20132680."},{"key":"e_1_3_1_17_2","doi-asserted-by":"crossref","first-page":"301","DOI":"10.1007\/0-387-27257-7_14","volume-title":"Proceedings of the Handbook of Face Recognition","author":"Gross Ralph","year":"2005","unstructured":"Ralph Gross. 2005. Face databases. In Proceedings of the Handbook of Face Recognition. Springer, 301\u2013327."},{"issue":"5","key":"e_1_3_1_18_2","doi-asserted-by":"crossref","first-page":"807","DOI":"10.1016\/j.imavis.2009.08.002","article-title":"Multi-pie","volume":"28","author":"Gross Ralph","year":"2010","unstructured":"Ralph Gross, Iain Matthews, Jeffrey Cohn, Takeo Kanade, and Simon Baker. 2010. Multi-pie. Image and Vision Computing 28, 5 (2010), 807\u2013813.","journal-title":"Image and Vision Computing"},{"key":"e_1_3_1_19_2","first-page":"14134","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision","author":"Guo Xiefan","year":"2021","unstructured":"Xiefan Guo, Hongyu Yang, and Di Huang. 2021. Image inpainting via conditional texture and structure dual generation. In Proceedings of the IEEE\/CVF International Conference on Computer Vision. 14134\u201314143."},{"key":"e_1_3_1_20_2","first-page":"16","volume-title":"Proceedings of the ECCV","author":"He Kaiming","year":"2012","unstructured":"Kaiming He and Jian Sun. 2012. Statistics of patch offsets for image completion. In Proceedings of the ECCV. Springer, 16\u201329."},{"key":"e_1_3_1_21_2","first-page":"1","volume-title":"Proceedings of the BTAS","author":"He Lingxiao","year":"2016","unstructured":"Lingxiao He, Haiqing Li, Qi Zhang, Zhenan Sun, and Zhaofeng He. 2016. Multiscale representation for partial face recognition under near infrared illumination. In Proceedings of the BTAS. IEEE, 1\u20137."},{"key":"e_1_3_1_22_2","first-page":"153","volume-title":"Proceedings of the Advances in Neural Information Processing Systems.","author":"He Xiaofei","year":"2004","unstructured":"Xiaofei He and Partha Niyogi. 2004. Locality preserving projections. In Proceedings of the Advances in Neural Information Processing Systems.153\u2013160."},{"key":"e_1_3_1_23_2","unstructured":"Alexander Hermans Lucas Beyer and Bastian Leibe. 2017. In defense of the triplet loss for person re-identification[J]. arXiv preprint arXiv:1703.07737(2017)."},{"issue":"4","key":"e_1_3_1_24_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3072959.3073659","article-title":"Globally and locally consistent image completion","volume":"36","author":"Iizuka Satoshi","year":"2017","unstructured":"Satoshi Iizuka, Edgar Simo-Serra, and Hiroshi Ishikawa. 2017. Globally and locally consistent image completion. ACM Transactions on Graphics 36, 4 (2017), 1\u201314.","journal-title":"ACM Transactions on Graphics"},{"key":"e_1_3_1_25_2","first-page":"1125","volume-title":"Proceedings of the CVPR","author":"Isola Phillip","year":"2017","unstructured":"Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros. 2017. Image-to-image translation with conditional adversarial networks. In Proceedings of the CVPR. 1125\u20131134."},{"key":"e_1_3_1_26_2","article-title":"Ways to retouch photos","author":"Jian-Ke Wen","year":"2013","unstructured":"Wen Jian-Ke. 2013. Ways to retouch photos. Laboratory Ence 16, 4 (2013), 14\u201317.","journal-title":"Laboratory Ence"},{"issue":"2","key":"e_1_3_1_27_2","doi-asserted-by":"crossref","first-page":"553","DOI":"10.1109\/TCE.2012.6227460","article-title":"Robust exemplar-based inpainting algorithm using region segmentation","volume":"58","author":"Lee Jino","year":"2012","unstructured":"Jino Lee, Dong-Kyu Lee, and Rae-Hong Park. 2012. Robust exemplar-based inpainting algorithm using region segmentation. IEEE Transactions on Consumer Electronics 58, 2 (2012), 553\u2013561.","journal-title":"IEEE Transactions on Consumer Electronics"},{"key":"e_1_3_1_28_2","first-page":"1","volume-title":"Proceedings of the International Conference on Automatic Face Gesture Recognition.","author":"Lei Zhen","year":"2008","unstructured":"Zhen Lei, Shengcai Liao, Ran He, Matti Pietikainen, and Stan Z. Li. 2008. Gabor volume based local binary pattern for face representation and recognition. In Proceedings of the International Conference on Automatic Face Gesture Recognition. IEEE, 1\u20136."},{"key":"e_1_3_1_29_2","first-page":"1","volume-title":"Proceedings of the IJCNN","author":"Li Ang","year":"2019","unstructured":"Ang Li, Jianzhong Qi, Rui Zhang, and Ramamohanarao Kotagiri. 2019. Boosted gan with semantically interpretable information for image inpainting. In Proceedings of the IJCNN. IEEE, 1\u20138."},{"key":"e_1_3_1_30_2","doi-asserted-by":"crossref","first-page":"3016","DOI":"10.1145\/3394171","volume-title":"Proceedings of the 28th ACM International Conference on Multimedia","author":"Li Chenyu","year":"2020","unstructured":"Chenyu Li, Shiming Ge, Daichi Zhang, and Jia Li. 2020. Look through masks: Towards masked face recognition with de-occlusion distillation. In Proceedings of the 28th ACM International Conference on Multimedia. 3016\u20133024."},{"key":"e_1_3_1_31_2","first-page":"3911","volume-title":"Proceedings of the CVPR","author":"Li Yijun","year":"2017","unstructured":"Yijun Li, Sifei Liu, Jimei Yang, and Ming-Hsuan Yang. 2017. Generative face completion. In Proceedings of the CVPR. 3911\u20133919."},{"key":"e_1_3_1_32_2","first-page":"I\u2013I","volume-title":"Proceedings of the CVPR","author":"Liu Ce","year":"2001","unstructured":"Ce Liu, Heung-Yeung Shum, and Chang-Shui Zhang. 2001. A two-step approach to hallucinating faces: Global parametric model and local nonparametric model. In Proceedings of the CVPR. IEEE, I\u2013I."},{"key":"e_1_3_1_33_2","first-page":"3730","volume-title":"Proceedings of the ICCV","author":"Liu Ziwei","year":"2015","unstructured":"Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang. 2015. Deep learning face attributes in the wild. In Proceedings of the ICCV. 3730\u20133738."},{"key":"e_1_3_1_34_2","doi-asserted-by":"crossref","unstructured":"Omkar M. Parkhi Andrea Vedaldi and Andrew Zisserman. 2015. Deep Face Recognition . British Machine Vision Association.","DOI":"10.5244\/C.29.41"},{"key":"e_1_3_1_35_2","first-page":"2536","volume-title":"Proceedings of the CVPR","author":"Pathak Deepak","year":"2016","unstructured":"Deepak Pathak, Philipp Krahenbuhl, Jeff Donahue, Trevor Darrell, and Alexei A. Efros. 2016. Context encoders: Feature learning by inpainting. In Proceedings of the CVPR. 2536\u20132544."},{"key":"e_1_3_1_36_2","article-title":"LabelMe","author":"Russell Bryan Christopher","year":"2008","unstructured":"Bryan Christopher Russell, Antonio J. Torralba, Kevin Patrick Murphy, and William T. Freeman. 2008. LabelMe. International Journal of Computer Vision 77, 1 (2008), 157\u2013173.","journal-title":"International Journal of Computer Vision"},{"issue":"11","key":"e_1_3_1_37_2","doi-asserted-by":"crossref","first-page":"1958","DOI":"10.1109\/TPAMI.2008.128","article-title":"80 million tiny images: A large data set for nonparametric object and scene recognition","volume":"30","author":"Torralba Antonio","year":"2008","unstructured":"Antonio Torralba, Rob Fergus, and William T. Freeman. 2008. 80 million tiny images: A large data set for nonparametric object and scene recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence 30, 11 (2008), 1958\u20131970.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_1_38_2","first-page":"1415","volume-title":"Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition","author":"Tran Luan","year":"2017","unstructured":"Luan Tran, Xi Yin, and Xiaoming Liu. 2017. Disentangled representation learning gan for pose-invariant face recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition. 1415\u20131424."},{"key":"e_1_3_1_39_2","first-page":"586","volume-title":"Proceedings of the CVPR","author":"Turk Matthew A.","year":"1991","unstructured":"Matthew A. Turk and Alex P. Pentland. 1991. Face recognition using eigenfaces. In Proceedings of the CVPR. IEEE Computer Society, 586\u2013587."},{"issue":"1","key":"e_1_3_1_40_2","doi-asserted-by":"crossref","first-page":"55","DOI":"10.13176\/11.15","article-title":"Eye detection in facial images with unconstrained background","volume":"1","author":"Wang Qiong","year":"2006","unstructured":"Qiong Wang and Jingyu Yang. 2006. Eye detection in facial images with unconstrained background. JPRR 1, 1 (2006), 55\u201362.","journal-title":"JPRR"},{"key":"e_1_3_1_41_2","first-page":"9168","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Wang Xintao","year":"2021","unstructured":"Xintao Wang, Yu Li, Honglun Zhang, and Ying Shan. 2021. Towards real-world blind face restoration with generative facial prior. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 9168\u20139178."},{"key":"e_1_3_1_42_2","first-page":"331","volume-title":"Proceedings of the NeurIPS","author":"Wang Yi","year":"2018","unstructured":"Yi Wang, Xin Tao, Xiaojuan Qi, Xiaoyong Shen, and Jiaya Jia. 2018. Image inpainting via generative multi-column convolutional neural networks. In Proceedings of the NeurIPS. 331\u2013340."},{"key":"e_1_3_1_43_2","unstructured":"Zhongyuan Wang Guangcheng Wang Baojin Huang Zhangyang Xiong Qi Hong Hao Wu Peng Yi Kui Jiang Nanxi Wang Yingjiao Pei Heling Chen Yu Miao Zhibing Huang and Jinbi Liang. 2020. Masked face recognition dataset and application. arXiv preprint arXiv:2003.09093 (2020)."},{"issue":"2","key":"e_1_3_1_44_2","doi-asserted-by":"crossref","first-page":"210","DOI":"10.1109\/TPAMI.2008.79","article-title":"Robust face recognition via sparse representation","volume":"31","author":"Wright John","year":"2008","unstructured":"John Wright, Allen Y. Yang, Arvind Ganesh, S. Shankar Sastry, and Yi Ma. 2008. Robust face recognition via sparse representation. IEEE Transactions on Pattern Analysis and Machine Intelligence 31, 2 (2008), 210\u2013227.","journal-title":"IEEE Transactions on Pattern Analysis and Machine Intelligence"},{"key":"e_1_3_1_45_2","unstructured":"Weihao Xia Yujiu Yang Jing-Hao Xue and Baoyuan Wu. 2021. Towards open-world text-guided face image generation and manipulation[J]. arXiv preprint arXiv:2104.08910 (2021)."},{"key":"e_1_3_1_46_2","first-page":"6721","volume-title":"Proceedings of the CVPR","author":"Yang Chao","year":"2017","unstructured":"Chao Yang, Xin Lu, Zhe Lin, Eli Shechtman, Oliver Wang, and Hao Li. 2017. High-resolution image inpainting using multi-scale neural patch synthesis. In Proceedings of the CVPR. 6721\u20136729."},{"key":"e_1_3_1_47_2","first-page":"9348","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision","author":"Yin Bangjie","year":"2019","unstructured":"Bangjie Yin, Luan Tran, Haoxiang Li, Xiaohui Shen, and Xiaoming Liu. 2019. Towards interpretable face recognition. In Proceedings of the IEEE\/CVF International Conference on Computer Vision. 9348\u20139357."},{"key":"e_1_3_1_48_2","first-page":"5505","volume-title":"Proceedings of the CVPR","author":"Yu Jiahui","year":"2018","unstructured":"Jiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen, Xin Lu, and Thomas S. Huang. 2018. Generative image inpainting with contextual attention. In Proceedings of the CVPR. 5505\u20135514."},{"key":"e_1_3_1_49_2","first-page":"4471","volume-title":"Proceedings of the ICCV","author":"Yu Jiahui","year":"2019","unstructured":"Jiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen, Xin Lu, and Thomas S. Huang. 2019. Free-form image inpainting with gated convolution. In Proceedings of the ICCV. 4471\u20134480."},{"key":"e_1_3_1_50_2","article-title":"Multimodal learning for temporally coherent talking face generation with articulator synergy","author":"Yu Lingyun","year":"2021","unstructured":"Lingyun Yu, Hongtao Xie, and Yongdong Zhang. 2021. Multimodal learning for temporally coherent talking face generation with articulator synergy. IEEE Transactions on Multimedia 24 (2021), 2950\u20132962.","journal-title":"IEEE Transactions on Multimedia"},{"key":"e_1_3_1_51_2","first-page":"12733","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Yu Tao","year":"2020","unstructured":"Tao Yu, Zongyu Guo, Xin Jin, Shilin Wu, Zhibo Chen, Weiping Li, Zhizheng Zhang, and Sen Liu. 2020. Region normalization for image inpainting. In Proceedings of the AAAI Conference on Artificial Intelligence. 12733\u201312740."},{"key":"e_1_3_1_52_2","doi-asserted-by":"crossref","first-page":"107626","DOI":"10.1016\/j.asoc.2021.107626","article-title":"Face inpainting based on GAN by facial prediction and fusion as guidance information","volume":"111","author":"Zhang Xian","year":"2021","unstructured":"Xian Zhang, Canghong Shi, Xin Wang, Xi Wu, Xiaojie Li, Jiancheng Lv, and Imran Mumtaz. 2021. Face inpainting based on GAN by facial prediction and fusion as guidance information. Applied Soft Computing 111 (2021), 107626.","journal-title":"Applied Soft Computing"},{"issue":"2","key":"e_1_3_1_53_2","doi-asserted-by":"crossref","first-page":"778","DOI":"10.1109\/TIP.2017.2771408","article-title":"Robust lstm-autoencoders for face de-occlusion in the wild","volume":"27","author":"Zhao Fang","year":"2017","unstructured":"Fang Zhao, Jiashi Feng, Jian Zhao, Wenhan Yang, and Shuicheng Yan. 2017. Robust lstm-autoencoders for face de-occlusion in the wild. IEEE Transactions on Image Processing 27, 2 (2017), 778\u2013790.","journal-title":"IEEE Transactions on Image Processing"},{"key":"e_1_3_1_54_2","first-page":"2207","volume-title":"Proceedings of the CVPR","author":"Zhao Jian","year":"2018","unstructured":"Jian Zhao, Yu Cheng, Yan Xu, Lin Xiong, Jianshu Li, Fang Zhao, Karlekar Jayashree, Sugiri Pranata, Shengmei Shen, Junliang Xing, et\u00a0al. 2018. Towards pose invariant face recognition in the wild. In Proceedings of the CVPR. 2207\u20132216."},{"key":"e_1_3_1_55_2","doi-asserted-by":"crossref","unstructured":"Jian Zhao Jianshu Li Xiaoguang Tu Fang Zhao Yuan Xin Junliang Xing Hengzhu Liu Shuicheng Yan and Jiashi Feng. 2019. Multi-prototype networks for unconstrained set-based face recognition. In Proceedings of the 28th International Joint Conference on Artificial Intelligence . 4397\u20134403.","DOI":"10.24963\/ijcai.2019\/611"},{"key":"e_1_3_1_56_2","first-page":"66","volume-title":"Proceedings of the NeurIPS","author":"Zhao Jian","year":"2017","unstructured":"Jian Zhao, Lin Xiong, Panasonic Karlekar Jayashree, Jianshu Li, Fang Zhao, Zhecan Wang, Panasonic Sugiri Pranata, Panasonic Shengmei Shen, Shuicheng Yan, and Jiashi Feng. 2017. Dual-agent gans for photorealistic and identity preserving profile face synthesis. In Proceedings of the NeurIPS. 66\u201376."},{"key":"e_1_3_1_57_2","first-page":"7680","volume-title":"Proceedings of the CVPR","author":"Zhou Tong","year":"2020","unstructured":"Tong Zhou, Changxing Ding, Shaowen Lin, Xinchao Wang, and Dacheng Tao. 2020. Learning oracle attention for high-fidelity face completion. In Proceedings of the CVPR. 7680\u20137689."},{"issue":"3","key":"e_1_3_1_58_2","doi-asserted-by":"crossref","first-page":"542","DOI":"10.1109\/TAFFC.2018.2828819","article-title":"Visually interpretable representation learning for depression recognition from facial images","volume":"11","author":"Zhou Xiuzhuang","year":"2018","unstructured":"Xiuzhuang Zhou, Kai Jin, Yuanyuan Shang, and Guodong Guo. 2018. Visually interpretable representation learning for depression recognition from facial images. IEEE Transactions on Affective Computing 11, 3 (2018), 542\u2013552.","journal-title":"IEEE Transactions on Affective Computing"},{"key":"e_1_3_1_59_2","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Zhu Zheng","year":"2021","unstructured":"Zheng Zhu, Guan Huang, Jiankang Deng, Yun Ye, Junjie Huang, Xinze Chen, Jiagang Zhu, Tian Yang, Jiwen Lu, Dalong Du, and Jie Zhou. 2021. WebFace260M: A benchmark unveiling the power of million-scale deep face recognition. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 10492\u201310502."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3524137","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3524137","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:31:05Z","timestamp":1750188665000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3524137"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,1,6]]},"references-count":58,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2023,1,31]]}},"alternative-id":["10.1145\/3524137"],"URL":"https:\/\/doi.org\/10.1145\/3524137","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,1,6]]},"assertion":[{"value":"2021-07-05","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2022-03-06","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-01-06","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}