{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,18]],"date-time":"2026-07-18T22:28:35Z","timestamp":1784413715013,"version":"3.55.0"},"reference-count":76,"publisher":"Springer Science and Business Media LLC","issue":"5","license":[{"start":{"date-parts":[[2020,2,22]],"date-time":"2020-02-22T00:00:00Z","timestamp":1582329600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2020,2,22]],"date-time":"2020-02-22T00:00:00Z","timestamp":1582329600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100000761","name":"Imperial College London","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100000761","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Comput Vis"],"published-print":{"date-parts":[[2020,5]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>This paper presents a novel approach for synthesizing facial affect; either in terms of the six basic expressions (i.e., anger, disgust, fear, joy, sadness and surprise), or in terms of valence (i.e., how positive or negative is an emotion) and arousal (i.e., power of the emotion activation). The proposed approach accepts the following inputs:(i) a neutral 2D image of a person; (ii) a basic facial expression or a pair of valence-arousal (VA) emotional state descriptors to be generated, or a path of affect in the 2D VA space to be generated as an image sequence. In order to synthesize affect in terms of VA, for this person, 600,000 frames from the 4DFAB database were annotated. The affect synthesis is implemented by fitting a 3D Morphable Model on the neutral image, then deforming the reconstructed face and adding the inputted affect, and blending the new face with the given affect into the original image. Qualitative experiments illustrate the generation of realistic images, when the neutral image is sampled from fifteen well known lab-controlled or in-the-wild databases, including Aff-Wild, AffectNet, RAF-DB; comparisons with generative adversarial networks (GANs) show the higher quality achieved by the proposed approach. Then, quantitative experiments are conducted, in which the synthesized images are used for data augmentation in training deep neural networks to perform affect recognition over all databases; greatly improved performances are achieved when compared with state-of-the-art methods, as well as with GAN-based data augmentation, in all cases.<\/jats:p>","DOI":"10.1007\/s11263-020-01304-3","type":"journal-article","created":{"date-parts":[[2020,2,22]],"date-time":"2020-02-22T09:03:01Z","timestamp":1582362181000},"page":"1455-1484","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":84,"title":["Deep Neural Network Augmentation: Generating Faces for Affect Analysis"],"prefix":"10.1007","volume":"128","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8188-3751","authenticated-orcid":false,"given":"Dimitrios","family":"Kollias","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shiyang","family":"Cheng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Evangelos","family":"Ververas","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Irene","family":"Kotsia","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Stefanos","family":"Zafeiriou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2020,2,22]]},"reference":[{"key":"1304_CR1","doi-asserted-by":"crossref","unstructured":"Abbasnejad, I., Sridharan, S., Nguyen, D., Denman, S., Fookes, C., & Lucey, S. (2017). Using synthetic data to improve facial expression analysis with 3d convolutional networks. In Proceedings of the IEEE international conference on computer vision (pp. 1609\u20131618).","DOI":"10.1109\/ICCVW.2017.189"},{"key":"1304_CR2","doi-asserted-by":"publisher","unstructured":"Alabort-i-Medina, J., Antonakos, E., Booth, J., Snape, P., & Zafeiriou, S. (2014). Menpo: A comprehensive platform for parametric image alignment and visual deformable models. In Proceedings of the ACM international conference on multimedia, MM\u201914 (pp. 679\u2013682). New York, NY, USA: ACM. https:\/\/doi.org\/10.1145\/2647868.2654890. http:\/\/doi.acm.org\/10.1145\/2647868.2654890.","DOI":"10.1145\/2647868.2654890."},{"issue":"1","key":"1304_CR3","doi-asserted-by":"publisher","first-page":"26","DOI":"10.1007\/s11263-016-0916-3","volume":"121","author":"J Alabort-i Medina","year":"2017","unstructured":"Alabort-i Medina, J., & Zafeiriou, S. (2017). A unified framework for compositional fitting of active appearance models. International Journal of Computer Vision, 121(1), 26\u201364.","journal-title":"International Journal of Computer Vision"},{"key":"1304_CR4","doi-asserted-by":"crossref","unstructured":"Amberg, B., Romdhani, S., & Vetter, T.: Optimal step nonrigid icp algorithms for surface registration. In 2007 IEEE conference on computer vision and pattern recognition (pp. 1\u20138). IEEE (2007).","DOI":"10.1109\/CVPR.2007.383165"},{"key":"1304_CR5","unstructured":"Antoniou, A., Storkey, A., & Edwards, H. (2017). Data augmentation generative adversarial networks. arXiv preprint arXiv:1711.04340."},{"issue":"6","key":"1304_CR6","doi-asserted-by":"publisher","first-page":"196","DOI":"10.1145\/3130800.3130818","volume":"36","author":"H Averbuch-Elor","year":"2017","unstructured":"Averbuch-Elor, H., Cohen-Or, D., Kopf, J., & Cohen, M. F. (2017). Bringing portraits to life. ACM Transactions on Graphics (TOG), 36(6), 196.","journal-title":"ACM Transactions on Graphics (TOG)"},{"issue":"1","key":"1304_CR7","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1561\/2200000015","volume":"4","author":"F Bach","year":"2012","unstructured":"Bach, F., Jenatton, R., Mairal, J., & Obozinski, G. (2012). Optimization with sparsity-inducing penalties. Foundations and Trends in Machine Learning, 4(1), 1\u2013106. https:\/\/doi.org\/10.1561\/2200000015.","journal-title":"Foundations and Trends in Machine Learning"},{"key":"1304_CR8","doi-asserted-by":"crossref","unstructured":"Blanz, V., Basso, C., Poggio, T., & Vetter, T. (2003). Reanimating faces in images and video. In Computer graphics forum (Vol.\u00a022, pp. 641\u2013650). Wiley Online Library.","DOI":"10.1111\/1467-8659.t01-1-00712"},{"key":"1304_CR9","unstructured":"Booth, J., Antonakos, E., Ploumpis, S., Trigeorgis, G., Panagakis, Y., & Zafeiriou, S. (2017). 3d face morphable models \u201cin-the-wild\u201d. In IEEE Conference on computer vision and pattern recognition (CVPR). https:\/\/arxiv.org\/abs\/1701.05360."},{"issue":"2\u20134","key":"1304_CR10","doi-asserted-by":"publisher","first-page":"233","DOI":"10.1007\/s11263-017-1009-7","volume":"126","author":"J Booth","year":"2018","unstructured":"Booth, J., Roussos, A., Ponniah, A., Dunaway, D., & Zafeiriou, S. (2018). Large scale 3d morphable models. International Journal of Computer Vision, 126(2\u20134), 233\u2013254.","journal-title":"International Journal of Computer Vision"},{"key":"1304_CR11","doi-asserted-by":"crossref","unstructured":"Booth, J., & Zafeiriou, S. (2014). Optimal uv spaces for facial morphable model construction. In 2014 IEEE international conference on image processing (ICIP) (pp. 4672\u20134676). IEEE.","DOI":"10.1109\/ICIP.2014.7025947"},{"issue":"4","key":"1304_CR12","first-page":"43","volume":"33","author":"C Cao","year":"2014","unstructured":"Cao, C., Hou, Q., & Zhou, K. (2014). Displaced dynamic expression regression for real-time facial tracking and animation. ACM Transactions on Graphics (TOG), 33(4), 43.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"1304_CR13","doi-asserted-by":"crossref","unstructured":"Cao, Q., Shen, L., Xie, W., Parkhi, O.M., & Zisserman, A. (2018). Vggface2: A dataset for recognising faces across pose and age. In 2018 13th IEEE international conference on automatic face & gesture recognition (FG 2018) (pp. 67\u201374). IEEE.","DOI":"10.1109\/FG.2018.00020"},{"key":"1304_CR14","doi-asserted-by":"crossref","unstructured":"Chang, W. Y., Hsu, S. H., & Chien, J. H. (2017). Fatauva-net : An integrated deep learning framework for facial attribute recognition, action unit (au) detection, and valence-arousal estimation. In Proceedings of the IEEE conference on computer vision and pattern recognition workshop.","DOI":"10.1109\/CVPRW.2017.246"},{"key":"1304_CR15","unstructured":"Cheng, S., Kotsia, I., Pantic, M., Zafeiriou, S., & (2018). 4dfab: A large scale 4d database for facial expression analysis and biometric applications. In IEEE conference on computer vision and pattern recognition (CVPR 2018). Utah, US: Salt Lake City."},{"issue":"4","key":"1304_CR16","doi-asserted-by":"publisher","first-page":"1006","DOI":"10.1109\/TSMCB.2012.2194485","volume":"42","author":"SW Chew","year":"2012","unstructured":"Chew, S. W., Lucey, P., Lucey, S., Saragih, J., Cohn, J. F., Matthews, I., et al. (2012). In the pursuit of effective affective computing: The relationship between features and registration. IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics), 42(4), 1006\u20131016.","journal-title":"IEEE Transactions on Systems, Man, and Cybernetics, Part B (Cybernetics)"},{"key":"1304_CR17","doi-asserted-by":"crossref","unstructured":"Choi, Y., Choi, M., Kim, M., Ha, J.W., Kim, S., & Choo, J. (2018). Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 8789\u20138797).","DOI":"10.1109\/CVPR.2018.00916"},{"key":"1304_CR18","doi-asserted-by":"crossref","unstructured":"Cosker, D., Krumhuber, E., & Hilton, A. (2011). A facs valid 3d dynamic action unit database with applications to 3d dynamic morphable facial modeling. In 2011 international conference on computer vision (pp. 2296\u20132303). IEEE.","DOI":"10.1109\/ICCV.2011.6126510"},{"key":"1304_CR19","doi-asserted-by":"publisher","unstructured":"Deng, J., Zhou, Y., Cheng, S., & Zaferiou, S.: Cascade multi-view hourglass model for robust 3d face alignment, pp. 399\u2013403 (2018). https:\/\/doi.org\/10.1109\/FG.2018.00064.","DOI":"10.1109\/FG.2018.00064"},{"key":"1304_CR20","doi-asserted-by":"crossref","unstructured":"Dhall, A., Goecke, R., Ghosh, S., Joshi, J., Hoey, J., & Gedeon, T. (2017). From individual to group-level emotion recognition: Emotiw 5.0. In Proceedings of the 19th ACM international conference on multimodal interaction (pp. 524\u2013528). ACM.","DOI":"10.1145\/3136755.3143004"},{"key":"1304_CR21","doi-asserted-by":"crossref","unstructured":"Ding, H., Sricharan, K., & Chellappa, R. (2018). Exprgan: Facial expression editing with controllable expression intensity. In Thirty-second AAAI conference on artificial intelligence.","DOI":"10.1609\/aaai.v32i1.12277"},{"issue":"12","key":"1304_CR22","doi-asserted-by":"publisher","first-page":"2170","DOI":"10.1109\/TIFS.2014.2359646","volume":"9","author":"E Eidinger","year":"2014","unstructured":"Eidinger, E., Enbar, R., & Hassner, T. (2014). Age and gender estimation of unfiltered faces. IEEE Transactions on Information Forensics and Security, 9(12), 2170\u20132179.","journal-title":"IEEE Transactions on Information Forensics and Security"},{"issue":"4","key":"1304_CR23","doi-asserted-by":"publisher","first-page":"128","DOI":"10.1145\/2897824.2925933","volume":"35","author":"O Fried","year":"2016","unstructured":"Fried, O., Shechtman, E., Goldman, D. B., & Finkelstein, A. (2016). Perspective-aware manipulation of portrait photos. ACM Transactions on Graphics (TOG), 35(4), 128.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"1304_CR24","doi-asserted-by":"crossref","unstructured":"Garrido, P., Valgaerts, L., Rehmsen, O., Thormahlen, T., Perez, P., & Theobalt, C. (2014). Automatic face reenactment. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 4217\u20134224).","DOI":"10.1109\/CVPR.2014.537"},{"key":"1304_CR25","doi-asserted-by":"crossref","unstructured":"Genova, K., Cole, F., Maschinot, A., Sarna, A., Vlasic, D., & Freeman, W. T. (2018). Unsupervised training for 3d morphable model regression. In The IEEE conference on computer vision and pattern recognition (CVPR).","DOI":"10.1109\/CVPR.2018.00874"},{"key":"1304_CR26","unstructured":"Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., et al. (2014). Generative adversarial nets. In Advances in neural information processing systems (pp. 2672\u20132680)."},{"key":"1304_CR27","doi-asserted-by":"crossref","unstructured":"Goodfellow, I. J., Erhan, D., Carrier, P. L., Courville, A., Mirza, M., Hamner, B., et\u00a0al. (2013). Challenges in representation learning: A report on three machine learning contests. In International conference on neural information processing (pp. 117\u2013124). Springer.","DOI":"10.1007\/978-3-642-42051-1_16"},{"issue":"5","key":"1304_CR28","doi-asserted-by":"publisher","first-page":"807","DOI":"10.1016\/j.imavis.2009.08.002","volume":"28","author":"R Gross","year":"2010","unstructured":"Gross, R., Matthews, I., Cohn, J., Kanade, T., & Baker, S. (2010). Multi-pie. Image and Vision Computing, 28(5), 807\u2013813.","journal-title":"Image and Vision Computing"},{"key":"1304_CR29","unstructured":"Knyazev, B., Shvetsov, R., Efremova, N., & Kuharenko, A. (2017). Convolutional neural networks pretrained on large face recognition datasets for emotion classification from video. arXiv preprint arXiv:1711.04598."},{"key":"1304_CR30","doi-asserted-by":"crossref","unstructured":"Kollias, D., Nicolaou, M. A., Kotsia, I., Zhao, G., & Zafeiriou, S. (2017). Recognition of affect in the wild using deep neural networks. In 2017 IEEE conference on computer vision and pattern recognition workshops (CVPRW) (pp. 1972\u20131979). IEEE.","DOI":"10.1109\/CVPRW.2017.247"},{"issue":"6\u20137","key":"1304_CR31","doi-asserted-by":"publisher","first-page":"907","DOI":"10.1007\/s11263-019-01158-4","volume":"127","author":"D Kollias","year":"2019","unstructured":"Kollias, D., Tzirakis, P., Nicolaou, M. A., Papaioannou, A., Zhao, G., Schuller, B., et al. (2019). Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond. International Journal of Computer Vision, 127(6\u20137), 907\u2013929.","journal-title":"International Journal of Computer Vision"},{"key":"1304_CR32","doi-asserted-by":"publisher","first-page":"23","DOI":"10.1016\/j.imavis.2017.02.001","volume":"65","author":"J Kossaifi","year":"2017","unstructured":"Kossaifi, J., Tzimiropoulos, G., Todorovic, S., & Pantic, M. (2017). AFEW-VA database for valence and arousal estimation in-the-wild. Image and Vision Computing, 65, 23\u201336.","journal-title":"Image and Vision Computing"},{"key":"1304_CR33","doi-asserted-by":"crossref","DOI":"10.1515\/9780691211701","volume-title":"Quaternions and rotation sequences","author":"JB Kuipers","year":"1999","unstructured":"Kuipers, J. B., et al. (1999). Quaternions and rotation sequences (Vol. 66). Princeton: Princeton University Press."},{"issue":"8","key":"1304_CR34","doi-asserted-by":"publisher","first-page":"1377","DOI":"10.1080\/02699930903485076","volume":"24","author":"O Langner","year":"2010","unstructured":"Langner, O., Dotsch, R., Bijlstra, G., Wigboldus, D. H., Hawk, S. T., & Van Knippenberg, A. (2010). Presentation and validation of the radboud faces database. Cognition and Emotion, 24(8), 1377\u20131388.","journal-title":"Cognition and Emotion"},{"issue":"1","key":"1304_CR35","doi-asserted-by":"publisher","first-page":"255","DOI":"10.2307\/2532051","volume":"45","author":"I Lawrence","year":"1989","unstructured":"Lawrence, I., & Lin, K. (1989). A concordance correlation coefficient to evaluate reproducibility. Biometrics, 45(1), 255\u2013268.","journal-title":"Biometrics"},{"key":"1304_CR36","doi-asserted-by":"crossref","unstructured":"Li, S., Deng, W., & Du, J. (2017). Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild. In 2017 IEEE conference on computer vision and pattern recognition (CVPR) (pp. 2584\u20132593). IEEE.","DOI":"10.1109\/CVPR.2017.277"},{"issue":"3\u20134","key":"1304_CR37","doi-asserted-by":"publisher","first-page":"235","DOI":"10.1002\/cav.248","volume":"19","author":"X Liu","year":"2008","unstructured":"Liu, X., Mao, T., Xia, S., Yu, Y., & Wang, Z. (2008). Facial animation by optimized blendshapes from motion capture data. Computer Animation and Virtual Worlds, 19(3\u20134), 235\u2013245.","journal-title":"Computer Animation and Virtual Worlds"},{"key":"1304_CR38","doi-asserted-by":"crossref","unstructured":"Ma, L., & Deng, Z. (2019). Real-time facial expression transformation for monocular rgb video. In Computer Graphics Forum (Vol.\u00a038, pp. 470\u2013481). Wiley Online Library.","DOI":"10.1111\/cgf.13586"},{"key":"1304_CR39","doi-asserted-by":"crossref","unstructured":"Maimon, O., & Rokach, L. (2005). Data mining and knowledge discovery handbook. Springer.","DOI":"10.1007\/b107408"},{"key":"1304_CR40","unstructured":"Mirza, M., & Osindero, S. (2014). Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784."},{"issue":"3","key":"1304_CR41","doi-asserted-by":"publisher","first-page":"57","DOI":"10.1145\/1531326.1531363","volume":"28","author":"U Mohammed","year":"2009","unstructured":"Mohammed, U., Prince, S. J., & Kautz, J. (2009). Visio-lization: Generating novel facial images. ACM Transactions on Graphics (TOG), 28(3), 57.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"1304_CR42","unstructured":"Mollahosseini, A., Hasani, B., & Mahoor, M. H. (2017). Affectnet: A database for facial expression, valence, and arousal computing in the wild. arXiv preprint arXiv:1708.03985."},{"issue":"6","key":"1304_CR43","doi-asserted-by":"publisher","first-page":"179","DOI":"10.1145\/2508363.2508417","volume":"32","author":"T Neumann","year":"2013","unstructured":"Neumann, T., Varanasi, K., Wenger, S., Wacker, M., Magnor, M., & Theobalt, C. (2013). Sparse localized deformation components. ACM Transactions on Graphics (TOG), 32(6), 179.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"1304_CR44","doi-asserted-by":"crossref","unstructured":"Parkhi, O. M., Vedaldi, A., & Zisserman, A. (2015). Deep face recognition. In BMVC (Vol. 1, p. 6).","DOI":"10.5244\/C.29.41"},{"key":"1304_CR45","doi-asserted-by":"crossref","unstructured":"Paysan, P., Knothe, R., Amberg, B., Romdhani, S., & Vetter, T. (2009). A 3d face model for pose and illumination invariant face recognition. In 2009 sixth IEEE international conference on advanced video and signal based surveillance (pp. 296\u2013301). IEEE.","DOI":"10.1109\/AVSS.2009.58"},{"key":"1304_CR46","doi-asserted-by":"publisher","unstructured":"P\u00e9rez, P., Gangnet, M., & Blake, A. (2003). Poisson image editing. In ACM SIGGRAPH 2003 Papers, SIGGRAPH\u201903 (pp. 313\u2013318). ACM, New York, NY, USA. https:\/\/doi.org\/10.1145\/1201775.882269. http:\/\/doi.acm.org\/10.1145\/1201775.882269.","DOI":"10.1145\/1201775.882269."},{"key":"1304_CR47","unstructured":"Pham, H. X., Wang, Y., & Pavlovic, V. (2018). Generative adversarial talking head: Bringing portraits to life with a weakly supervised neural network. arXiv preprint arXiv:1803.07716."},{"key":"1304_CR48","doi-asserted-by":"crossref","unstructured":"Pumarola, A., Agudo, A., Martinez, A. M., Sanfeliu, A., & Moreno-Noguer, F. (2018). Ganimation: Anatomically-aware facial animation from a single image. In Proceedings of the European conference on computer vision (ECCV) (pp. 818\u2013833).","DOI":"10.1007\/978-3-030-01249-6_50"},{"key":"1304_CR49","unstructured":"Qiao, F., Yao, N., Jiao, Z., Li, Z., Chen, H., & Wang, H. (2018). Geometry-contrastive gan for facial expression transfer. arXiv preprint arXiv:1802.01822."},{"key":"1304_CR50","unstructured":"Reed, S., Sohn, K., Zhang, Y., & Lee, H. (2014). Learning to disentangle factors of variation with manifold interaction. In International conference on machine learning (pp. 1431\u20131439)."},{"key":"1304_CR51","doi-asserted-by":"crossref","unstructured":"Ringeval, F., Sonderegger, A., Sauer, J., & Lalanne, D. (2013). Introducing the recola multimodal corpus of remote collaborative and affective interactions. In 2013 10th IEEE international conference and workshops on automatic face and gesture recognition (FG) (pp. 1\u20138). IEEE.","DOI":"10.1109\/FG.2013.6553805"},{"key":"1304_CR52","doi-asserted-by":"crossref","unstructured":"Rothe, R., Timofte, R., & Van\u00a0Gool, L. (2015). Dex: Deep expectation of apparent age from a single image. In Proceedings of the IEEE international conference on computer vision workshops (pp. 10\u201315).","DOI":"10.1109\/ICCVW.2015.41"},{"issue":"10","key":"1304_CR53","doi-asserted-by":"publisher","first-page":"1152","DOI":"10.1037\/0022-3514.36.10.1152","volume":"36","author":"JA Russell","year":"1978","unstructured":"Russell, J. A. (1978). Evidence of convergent validity on the dimensions of affect. Journal of Personality and Social Psychology, 36(10), 1152.","journal-title":"Journal of Personality and Social Psychology"},{"key":"1304_CR54","doi-asserted-by":"crossref","unstructured":"Savran, A., Aly\u00fcz, N., Dibeklio\u011flu, H., \u00c7eliktutan, O., G\u00f6kberk, B., Sankur, B., et al. (2008). Bosphorus database for 3d face analysis. In European workshop on biometrics and identity management (pp. 47\u201356). Springer.","DOI":"10.1007\/978-3-540-89991-4_6"},{"key":"1304_CR55","doi-asserted-by":"publisher","unstructured":"Shang, F., Liu, Y., Cheng, J., & Cheng, H. (2014). Robust principal component analysis with missing data. In Proceedings of the 23rd ACM international conference on conference on information and knowledge management, CIKM\u201914 (pp. 1149\u20131158). New York, NY, USA: ACM. https:\/\/doi.org\/10.1145\/2661829.2662083. http:\/\/doi.acm.org\/10.1145\/2661829.2662083.","DOI":"10.1145\/2661829.2662083"},{"key":"1304_CR56","unstructured":"Sohn, K., Lee, H., & Yan, X. (2015). Learning structured output representation using deep conditional generative models. In Advances in Neural Information Processing Systems (pp. 3483\u20133491)."},{"key":"1304_CR57","doi-asserted-by":"crossref","unstructured":"Song, L., Lu, Z., He, R., Sun, Z., & Tan, T. (2018). Geometry guided adversarial facial expression synthesis. In 2018 ACM multimedia conference on multimedia conference (pp. 627\u2013635). ACM.","DOI":"10.1145\/3240508.3240612"},{"key":"1304_CR58","unstructured":"Sun, Y., Chen, Y., Wang, X., & Tang, X. (2014). Deep learning face representation by joint identification\u2013verification. In Advances in neural information processing systems (pp. 1988\u20131996)."},{"key":"1304_CR59","unstructured":"Susskind, J. M., Hinton, G. E., Movellan, J. R., & Anderson, A. K. (2008). Generating facial expressions with deep belief nets. In Affective computing. IntechOpen."},{"key":"1304_CR60","doi-asserted-by":"crossref","unstructured":"Taigman, Y., Yang, M., Ranzato, M., & Wolf, L. (2014). Deepface: Closing the gap to human-level performance in face verification. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 1701\u20131708).","DOI":"10.1109\/CVPR.2014.220"},{"key":"1304_CR61","doi-asserted-by":"crossref","unstructured":"Thies, J., Zollhofer, M., Stamminger, M., Theobalt, C., & Nie\u00dfner, M. (2016). Face2face: Real-time face capture and reenactment of rgb videos. In Proceedings of the IEEE conference on computer vision and pattern recognition (pp. 2387\u20132395).","DOI":"10.1145\/2929464.2929475"},{"issue":"4","key":"1304_CR62","doi-asserted-by":"publisher","first-page":"164","DOI":"10.1145\/3197517.3201350","volume":"37","author":"J Thies","year":"2018","unstructured":"Thies, J., Zollh\u00f6fer, M., Theobalt, C., Stamminger, M., & Nie\u00dfner, M. (2018). Headon: real-time reenactment of human portrait videos. ACM Transactions on Graphics (TOG), 37(4), 164.","journal-title":"ACM Transactions on Graphics (TOG)"},{"issue":"6","key":"1304_CR63","doi-asserted-by":"publisher","first-page":"902","DOI":"10.1016\/j.imavis.2009.11.005","volume":"28","author":"CE Thomaz","year":"2010","unstructured":"Thomaz, C. E., & Giraldi, G. A. (2010). A new ranking method for principal components analysis and its application to face image analysis. Image and Vision Computing, 28(6), 902\u2013913.","journal-title":"Image and Vision Computing"},{"key":"1304_CR64","doi-asserted-by":"crossref","unstructured":"Valstar, M., Gratch, J., Schuller, B., Ringeval, F., Lalanne, D., Torres\u00a0Torres, M., et al. (2016). Avec 2016: Depression, mood, and emotion recognition workshop and challenge. In Proceedings of the 6th international workshop on audio\/visual emotion challenge (pp. 3\u201310). ACM.","DOI":"10.1145\/2988257.2988258"},{"key":"1304_CR65","unstructured":"Vielzeuf, V., Pateux, S., & Jurie, F. (2017). Temporal multimodal fusion for video emotion classification in the wild. arXiv preprint arXiv:1709.07200."},{"key":"1304_CR66","unstructured":"Wheeler, M. D., & Ikeuchi, K. (1995). Iterative estimation of rotation and translation using the quaternion. Department of Computer Science, Carnegie-Mellon University."},{"key":"1304_CR67","doi-asserted-by":"crossref","unstructured":"Whissell, C. M. (1989). The dictionary of affect in language. In The measurement of emotions (pp. 113\u2013131). Elsevier.","DOI":"10.1016\/B978-0-12-558704-4.50011-6"},{"issue":"7","key":"1304_CR68","doi-asserted-by":"publisher","first-page":"2479","DOI":"10.1109\/TSP.2009.2016892","volume":"57","author":"SJ Wright","year":"2009","unstructured":"Wright, S. J., Nowak, R. D., & Figueiredo, M. A. T. (2009). Sparse reconstruction by separable approximation. IEEE Transactions on Signal Processing, 57(7), 2479\u20132493. https:\/\/doi.org\/10.1109\/TSP.2009.2016892.","journal-title":"IEEE Transactions on Signal Processing"},{"key":"1304_CR69","doi-asserted-by":"crossref","unstructured":"Wu, W., Zhang, Y., Li, C., Qian, C., & Change\u00a0Loy, C. (2018). Reenactgan: Learning to reenact faces via boundary transfer. In Proceedings of the European conference on computer vision (ECCV) (pp. 603\u2013619).","DOI":"10.1007\/978-3-030-01246-5_37"},{"key":"1304_CR70","unstructured":"Yin, L., Wei, X., Sun, Y., Wang, J., & Rosato, M. J. (2006). A 3d facial expression database for facial behavior research. In 7th international conference on automatic face and gesture recognition, 2006. FGR 2006 (pp. 211\u2013216). IEEE."},{"key":"1304_CR71","doi-asserted-by":"crossref","unstructured":"Zafeiriou, S., Kollias, D., Nicolaou, M. A., Papaioannou, A., Zhao, G., & Kotsia, I. (2017). Aff-wild: Valence and arousal \u2018in-the-wild\u2019challenge. In 2017 IEEE conference on computer vision and pattern recognition workshops (CVPRW) (pp. 1980\u20131987). IEEE.","DOI":"10.1109\/CVPRW.2017.248"},{"key":"1304_CR72","unstructured":"Zagoruyko, S., & Komodakis, N. (2016). Wide residual networks. arXiv preprint arXiv:1605.07146."},{"issue":"10","key":"1304_CR73","doi-asserted-by":"publisher","first-page":"1499","DOI":"10.1109\/LSP.2016.2603342","volume":"23","author":"K Zhang","year":"2016","unstructured":"Zhang, K., Zhang, Z., Li, Z., & Qiao, Y. (2016). Joint face detection and alignment using multitask cascaded convolutional networks. IEEE Signal Processing Letters, 23(10), 1499\u20131503. https:\/\/doi.org\/10.1109\/LSP.2016.2603342.","journal-title":"IEEE Signal Processing Letters"},{"key":"1304_CR74","doi-asserted-by":"crossref","unstructured":"Zhou, Y., & Shi, B. E. (2017). Photorealistic facial expression synthesis by the conditional difference adversarial autoencoder. In 2017 seventh international conference on affective computing and intelligent interaction (ACII) (pp. 370\u2013376). IEEE.","DOI":"10.1109\/ACII.2017.8273626"},{"key":"1304_CR75","doi-asserted-by":"crossref","unstructured":"Zhu, J. Y., Park, T., Isola, P., & Efros, A. A. (2017). Unpaired image-to-image translation using cycle-consistent adversarial networks. In Proceedings of the IEEE international conference on computer vision (pp. 2223\u20132232).","DOI":"10.1109\/ICCV.2017.244"},{"key":"1304_CR76","doi-asserted-by":"crossref","unstructured":"Zhu, X., Liu, Y., Li, J., Wan, T., & Qin, Z. (2018). Emotion classification with data augmentation using generative adversarial networks. In Pacific-Asia conference on knowledge discovery and data mining (pp. 349\u2013360). Springer.","DOI":"10.1007\/978-3-319-93040-4_28"}],"container-title":["International Journal of Computer Vision"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-020-01304-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s11263-020-01304-3\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s11263-020-01304-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,10,16]],"date-time":"2022-10-16T08:03:07Z","timestamp":1665907387000},"score":1,"resource":{"primary":{"URL":"http:\/\/link.springer.com\/10.1007\/s11263-020-01304-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,2,22]]},"references-count":76,"journal-issue":{"issue":"5","published-print":{"date-parts":[[2020,5]]}},"alternative-id":["1304"],"URL":"https:\/\/doi.org\/10.1007\/s11263-020-01304-3","relation":{},"ISSN":["0920-5691","1573-1405"],"issn-type":[{"value":"0920-5691","type":"print"},{"value":"1573-1405","type":"electronic"}],"subject":[],"published":{"date-parts":[[2020,2,22]]},"assertion":[{"value":"31 October 2018","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"5 February 2020","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"22 February 2020","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}