{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,2]],"date-time":"2026-06-02T09:28:15Z","timestamp":1780392495784,"version":"3.54.1"},"reference-count":102,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2023,7,26]],"date-time":"2023-07-26T00:00:00Z","timestamp":1690329600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"ONR","award":["N000142012529"],"award-info":[{"award-number":["N000142012529"]}]},{"name":"ONR","award":["N000142312526"],"award-info":[{"award-number":["N000142312526"]}]},{"name":"NSF","award":["IIS 2110409"],"award-info":[{"award-number":["IIS 2110409"]}]},{"DOI":"10.13039\/100000185","name":"Defense Advanced Research Projects Agency","doi-asserted-by":"publisher","award":["HR0011-20-3-0005"],"award-info":[{"award-number":["HR0011-20-3-0005"]}],"id":[{"id":"10.13039\/100000185","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2023,8]]},"abstract":"<jats:p>We present a one-shot method to infer and render a photorealistic 3D representation from a single unposed image (e.g., face portrait) in real-time. Given a single RGB input, our image encoder directly predicts a canonical triplane representation of a neural radiance field for 3D-aware novel view synthesis via volume rendering. Our method is fast (24 fps) on consumer hardware, and produces higher quality results than strong GAN-inversion baselines that require test-time optimization. To train our triplane encoder pipeline, we use only synthetic data, showing how to distill the knowledge from a pretrained 3D GAN into a feedforward encoder. Technical contributions include a Vision Transformer-based triplane encoder, a camera data augmentation strategy, and a well-designed loss function for synthetic data training. We benchmark against the state-of-the-art methods, demonstrating significant improvements in robustness and image quality in challenging real-world settings. We showcase our results on portraits of faces (FFHQ) and cats (AFHQ), but our algorithm can also be applied in the future to other categories with a 3D-aware image generator.<\/jats:p>","DOI":"10.1145\/3592460","type":"journal-article","created":{"date-parts":[[2023,7,26]],"date-time":"2023-07-26T15:47:45Z","timestamp":1690386465000},"page":"1-15","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":69,"title":["Real-Time Radiance Fields for Single-Image Portrait View Synthesis"],"prefix":"10.1145","volume":"42","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-9734-7232","authenticated-orcid":false,"given":"Alex","family":"Trevithick","sequence":"first","affiliation":[{"name":"UC San Diego, La Jolla, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-1704-1013","authenticated-orcid":false,"given":"Matthew","family":"Chan","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-6234-936X","authenticated-orcid":false,"given":"Michael","family":"Stengel","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-9691-8213","authenticated-orcid":false,"given":"Eric","family":"Chan","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-5751-8723","authenticated-orcid":false,"given":"Chao","family":"Liu","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1776-996X","authenticated-orcid":false,"given":"Zhiding","family":"Yu","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8068-7807","authenticated-orcid":false,"given":"Sameh","family":"Khamis","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4683-2454","authenticated-orcid":false,"given":"Manmohan","family":"Chandraker","sequence":"additional","affiliation":[{"name":"UC San Diego, La Jolla, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3993-5789","authenticated-orcid":false,"given":"Ravi","family":"Ramamoorthi","sequence":"additional","affiliation":[{"name":"UC San Diego, La Jolla, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6815-4864","authenticated-orcid":false,"given":"Koki","family":"Nagano","sequence":"additional","affiliation":[{"name":"NVIDIA, Santa Clara, United States of America"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,7,26]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00453"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00664"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01972"},{"key":"e_1_2_2_4_1","volume-title":"FWD: Real-time Novel View Synthesis with Forward Warping and Depth. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Cao Ang","year":"2022","unstructured":"Ang Cao, Chris Rockwell, and Justin Johnson. 2022. FWD: Real-time Novel View Synthesis with Forward Warping and Depth. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01565"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00574"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01386"},{"key":"e_1_2_2_8_1","volume-title":"Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:1706.05587","author":"Chen Liang-Chieh","year":"2017","unstructured":"Liang-Chieh Chen, George Papandreou, Florian Schroff, and Hartwig Adam. 2017. Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:1706.05587 (2017)."},{"key":"e_1_2_2_9_1","doi-asserted-by":"crossref","unstructured":"S Chen and L Williams. 1993. View Interpolation for Image Synthesis. In SIGGRAPH 93. 279--288.","DOI":"10.1145\/166117.166153"},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00609"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00821"},{"key":"e_1_2_2_12_1","volume-title":"ArcFace: Additive Angular Margin Loss for Deep Face Recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Deng Jiankang","year":"2019","unstructured":"Jiankang Deng, Jia Guo, Xue Niannan, and Stefanos Zafeiriou. 2019a. ArcFace: Additive Angular Margin Loss for Deep Face Recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01041"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2019.00038"},{"key":"e_1_2_2_15_1","volume-title":"Simoncelli","author":"Ding Keyan","year":"2022","unstructured":"Keyan Ding, Kede Ma, Shiqi Wang, and Eero P. Simoncelli. 2022. Image Quality Assessment: Unifying Structure and Texture Similarity. IEEE Transactions on Pattern Analysis and Machine Intelligence (2022)."},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01110"},{"key":"e_1_2_2_17_1","volume-title":"HeadGAN: One-shot Neural Head Synthesis and Editing. In IEEE International Conference on Computer Vision (ICCV).","author":"Doukas Michail Christos","year":"2021","unstructured":"Michail Christos Doukas, Stefanos Zafeiriou, and Viktoriia Sharmanska. 2021. HeadGAN: One-shot Neural Head Synthesis and Editing. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_2_2_18_1","volume-title":"Megaportraits: One-shot megapixel neural head avatars. arXiv preprint arXiv:2207.07621","author":"Drobyshev Nikita","year":"2022","unstructured":"Nikita Drobyshev, Jenya Chelishev, Taras Khakhulin, Aleksei Ivakhnenko, Victor Lempitsky, and Egor Zakharov. 2022. Megaportraits: One-shot megapixel neural head avatars. arXiv preprint arXiv:2207.07621 (2022)."},{"key":"e_1_2_2_19_1","volume-title":"Near perfect gan inversion. arXiv preprint arXiv:2202.11833","author":"Feng Qianli","year":"2022","unstructured":"Qianli Feng, Viraj Shah, Raghudeep Gadde, Pietro Perona, and Aleix Martinez. 2022. Near perfect gan inversion. arXiv preprint arXiv:2202.11833 (2022)."},{"key":"e_1_2_2_20_1","volume-title":"Portrait Neural Radiance Fields from a Single Image. arXiv preprint arXiv:2012.05903","author":"Gao Chen","year":"2020","unstructured":"Chen Gao, Yichang Shih, Wei-Sheng Lai, Chia-Kai Liang, and Jia-Bin Huang. 2020. Portrait Neural Radiance Fields from a Single Image. arXiv preprint arXiv:2012.05903 (2020)."},{"key":"e_1_2_2_21_1","unstructured":"Ian Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron Courville and Yoshua Bengio. 2014. Generative Adversarial Nets. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_22_1","doi-asserted-by":"crossref","unstructured":"S Gortler R Grzeszczuk R Szeliski and M Cohen. 1996. The Lumigraph. In SIGGRAPH 96. 43--54.","DOI":"10.1145\/237170.237200"},{"key":"e_1_2_2_23_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Groueix Thibault","year":"2018","unstructured":"Thibault Groueix, Matthew Fisher, Vladimir G. Kim, Bryan C. Russell, and Mathieu Aubry. 2018. AtlasNet: A Papier-M\u00e2ch\u00e9 Approach to Learning 3D Surface Generation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_24_1","volume-title":"StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image Synthesis. arXiv preprint arXiv:2110.08985","author":"Gu Jiatao","year":"2021","unstructured":"Jiatao Gu, Lingjie Liu, Peng Wang, and Christian Theobalt. 2021. StyleNeRF: A Style-based 3D-Aware Generator for High-resolution Image Synthesis. arXiv preprint arXiv:2110.08985 (2021)."},{"key":"e_1_2_2_25_1","unstructured":"Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler G\u00fcnter Klambauer and Sepp Hochreiter. 2017. GANs Trained by a Two Time-Scale Update Rule Converge to a Nash Equilibrium. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_26_1","volume-title":"Depth-Aware Generative Adversarial Network for Talking Head Video Generation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Hong Fa-Ting","year":"2022","unstructured":"Fa-Ting Hong, Longhao Zhang, Li Shen, and Dan Xu. 2022b. Depth-Aware Generative Adversarial Network for Talking Head Video Generation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_27_1","volume-title":"HeadNeRF: A Real-time NeRF-based Parametric Head Model. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Hong Yang","year":"2022","unstructured":"Yang Hong, Bo Peng, Haiyao Xiao, Ligang Liu, and Juyong Zhang. 2022a. HeadNeRF: A Real-time NeRF-based Parametric Head Model. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01271"},{"key":"e_1_2_2_29_1","doi-asserted-by":"crossref","unstructured":"N. Khademi Kalantari T. Wang and R. Ramamoorthi. 2016. Learning-based view synthesis for light field cameras. ACM Transactions on Graphics (SIGGRAPH Asia 16) 35 6 (2016) 193:1--193:10.","DOI":"10.1145\/2980179.2980251"},{"key":"e_1_2_2_30_1","volume-title":"International Conference on Machine Learning. PMLR, 1771--1779","author":"Kalchbrenner Nal","year":"2017","unstructured":"Nal Kalchbrenner, A\u00e4ron Oord, Karen Simonyan, Ivo Danihelka, Oriol Vinyals, Alex Graves, and Koray Kavukcuoglu. 2017. Video pixel networks. In International Conference on Machine Learning. PMLR, 1771--1779."},{"key":"e_1_2_2_31_1","unstructured":"Tero Karras Miika Aittala Samuli Laine Erik H\u00e4rk\u00f6nen Janne Hellsten Jaakko Lehtinen and Timo Aila. 2021. Alias-Free Generative Adversarial Networks. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00453"},{"key":"e_1_2_2_33_1","volume-title":"Analyzing and Improving the Image Quality of StyleGAN. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Karras Tero","year":"2020","unstructured":"Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila. 2020. Analyzing and Improving the Image Quality of StyleGAN. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_34_1","volume-title":"Realistic One-shot Mesh-based Head Avatars. In European Conference on Computer Vision (ECCV).","author":"Khakhulin Taras","year":"2022","unstructured":"Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, and Egor Zakharov. 2022. Realistic One-shot Mesh-based Head Avatars. In European Conference on Computer Vision (ECCV)."},{"key":"e_1_2_2_35_1","volume-title":"Deep Video Portraits. ACM Transactions on Graphics (SIGGRAPH)","author":"Kim Hyeongwoo","year":"2018","unstructured":"Hyeongwoo Kim, Pablo Garrido, Ayush Tewari, Weipeng Xu, Justus Thies, Matthias Nie\u00dfner, Patrick P\u00e9rez, Christian Richardt, Michael Zoll\u00f6fer, and Christian Theobalt. 2018. Deep Video Portraits. ACM Transactions on Graphics (SIGGRAPH) (2018)."},{"key":"e_1_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV56688.2023.00298"},{"key":"e_1_2_2_37_1","volume-title":"Freestylegan: Free-view editable portrait rendering with the camera manifold. arXiv preprint arXiv:2109.09378","author":"Leimk\u00fchler Thomas","year":"2021","unstructured":"Thomas Leimk\u00fchler and George Drettakis. 2021. Freestylegan: Free-view editable portrait rendering with the camera manifold. arXiv preprint arXiv:2109.09378 (2021)."},{"key":"e_1_2_2_38_1","doi-asserted-by":"crossref","unstructured":"M Levoy and P Hanrahan. 1996. Light Field Rendering. In SIGGRAPH 96. 31--42.","DOI":"10.1145\/237170.237199"},{"key":"e_1_2_2_39_1","volume-title":"Asian Conference on Computer Vision (ACCV).","author":"Li Xingyi","year":"2022","unstructured":"Xingyi Li, Chaoyi Hong, Yiran Wang, Zhiguo Cao, Ke Xian, and Guosheng Lin. 2022. SymmNeRF: Learning to Explore Symmetry Prior for Single-View View Synthesis. In Asian Conference on Computer Vision (ACCV)."},{"key":"e_1_2_2_40_1","volume-title":"ECCV Workshop on Learning to Generate 3D Shapes and Scenes.","author":"Lin C.Z.","unstructured":"C.Z. Lin, D.B. Lindell, E.R. Chan, and G. Wetzstein. 2022. 3D GAN Inversion for Controllable Portrait Image Animation. In ECCV Workshop on Learning to Generate 3D Shapes and Scenes."},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV56688.2023.00087"},{"key":"e_1_2_2_42_1","doi-asserted-by":"crossref","unstructured":"N. Max. 1995. Optical models for direct volume rendering. IEEE Transactions on Visualization and Computer Graphics (TVCG) (1995).","DOI":"10.1109\/2945.468400"},{"key":"e_1_2_2_43_1","first-page":"39","article-title":"Plenoptic Modeling","volume":"95","author":"McMillan L","year":"1995","unstructured":"L McMillan and G Bishop. 1995. Plenoptic Modeling: An Image-Based Rendering System. In SIGGRAPH 95. 39--46.","journal-title":"An Image-Based Rendering System. In SIGGRAPH"},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01400"},{"key":"e_1_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00459"},{"key":"e_1_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-19784-0_11"},{"key":"e_1_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58452-8_24"},{"key":"e_1_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3355089.3356568"},{"key":"e_1_2_2_49_1","volume-title":"PaGAN: Real-Time Avatars Using Dynamic Textures. ACM Transactions on Graphics (SIGGRAPH ASIA)","author":"Nagano Koki","year":"2018","unstructured":"Koki Nagano, Jaewoo Seo, Jun Xing, Lingyu Wei, Zimo Li, Shunsuke Saito, Aviral Agarwal, Jens Fursund, and Hao Li. 2018. PaGAN: Real-Time Avatars Using Dynamic Textures. ACM Transactions on Graphics (SIGGRAPH ASIA) (2018)."},{"key":"e_1_2_2_50_1","volume-title":"IEEE International Conference on Computer Vision (ICCV).","author":"Nguyen-Phuoc Thu","year":"2019","unstructured":"Thu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt, and Yong-Liang Yang. 2019. HoloGAN: Unsupervised learning of 3D representations from natural images. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01129"},{"key":"e_1_2_2_52_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Or-El Roy","year":"2022","unstructured":"Roy Or-El, Xuan Luo, Mengyi Shan, Eli Shechtman, Jeong Joon Park, and Ira Kemelmacher-Shlizerman. 2022. StyleSDF: High-Resolution 3D-Consistent Image and Geometry Generation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_53_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Pan Xingang","year":"2021","unstructured":"Xingang Pan, Bo Dai, Ziwei Liu, Chen Change Loy, and Ping Luo. 2021. Do 2D GANs Know 3D Shape? Unsupervised 3D Shape Reconstruction from 2D Image GANs. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_2_2_54_1","volume-title":"DeepSDF: Learning Continuous Signed Distance Functions for Shape Representation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Park Jeong Joon","year":"2019","unstructured":"Jeong Joon Park, Peter Florence, Julian Straub, Richard Newcombe, and Steven Lovegrove. 2019. DeepSDF: Learning Continuous Signed Distance Functions for Shape Representation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_55_1","volume-title":"GAN-Supervised Dense Visual Alignment. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Peebles William","year":"2022","unstructured":"William Peebles, Jun-Yan Zhu, Richard Zhang, Antonio Torralba, Alexei Efros, and Eli Shechtman. 2022. GAN-Supervised Dense Visual Alignment. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00557"},{"key":"e_1_2_2_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00161"},{"key":"e_1_2_2_58_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00355"},{"key":"e_1_2_2_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00232"},{"key":"e_1_2_2_60_1","volume-title":"Pivotal Tuning for Latent-based Editing of Real Images. arXiv preprint arXiv:2106.05744","author":"Roich Daniel","year":"2021","unstructured":"Daniel Roich, Ron Mokady, Amit H Bermano, and Daniel Cohen-Or. 2021. Pivotal Tuning for Latent-based Editing of Real Images. arXiv preprint arXiv:2106.05744 (2021)."},{"key":"e_1_2_2_61_1","volume-title":"IEEE International Conference on Computer Vision (ICCV).","author":"Rombach R.","unstructured":"R. Rombach, P. Esser, and B. Ommer. 2021. Geometry-Free View Synthesis: Transformers and no 3D Priors. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00613"},{"key":"e_1_2_2_63_1","volume-title":"GRAF: Generative Radiance Fields for 3D-Aware Image Synthesis. In Advances in Neural Information Processing Systems (NeurIPS).","author":"Schwarz Katja","year":"2020","unstructured":"Katja Schwarz, Yiyi Liao, Michael Niemeyer, and Andreas Geiger. 2020. GRAF: Generative Radiance Fields for 3D-Aware Image Synthesis. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_64_1","volume-title":"J. P. Lewis, and Junyong Noh.","author":"Seol Yeongho","year":"2011","unstructured":"Yeongho Seol, Jaewoo Seo, Paul Hyunjin Kim, J. P. Lewis, and Junyong Noh. 2011. Artist Friendly Facial Animation Retargeting. In ACM Transactions on Graphics (SIGGRAPH ASIA)."},{"key":"e_1_2_2_65_1","unstructured":"Xingjian Shi Zhourong Chen Hao Wang Dit-Yan Yeung Wai-Kin Wong and Wangchun Woo. 2015. Convolutional LSTM network: A machine learning approach for precipitation nowcasting. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_66_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00619"},{"key":"e_1_2_2_67_1","unstructured":"Vincent Sitzmann Michael Zollh\u00f6fer and Gordon Wetzstein. 2019. Scene Representation Networks: Continuous 3D-Structure-Aware Neural Scene Representations. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_68_1","volume-title":"3D generation on ImageNet. arXiv preprint arXiv:2303.01416","author":"Skorokhodov Ivan","year":"2023","unstructured":"Ivan Skorokhodov, Aliaksandr Siarohin, Yinghao Xu, Jian Ren, Hsin-Ying Lee, Peter Wonka, and Sergey Tulyakov. 2023. 3D generation on ImageNet. arXiv preprint arXiv:2303.01416 (2023)."},{"key":"e_1_2_2_69_1","unstructured":"Ivan Skorokhodov Sergey Tulyakov Yiqun Wang and Peter Wonka. 2022. EpiGRAF: Rethinking training of 3D GANs. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_70_1","volume-title":"International Conference on Computer Vision (ICCV). 2262--2270","author":"Srinivasan P.","unstructured":"P. Srinivasan, T. Wang, A. Sreelal, R. Ramamoorthi, and R. Ng. 2017. Learning to Synthesize a 4D RGBD Light Field from a Single Image. In International Conference on Computer Vision (ICCV). 2262--2270."},{"key":"e_1_2_2_71_1","volume-title":"International conference on machine learning. PMLR, 843--852","author":"Srivastava Nitish","year":"2015","unstructured":"Nitish Srivastava, Elman Mansimov, and Ruslan Salakhudinov. 2015. Unsupervised learning of video representations using lstms. In International conference on machine learning. PMLR, 843--852."},{"key":"e_1_2_2_72_1","doi-asserted-by":"publisher","DOI":"10.1145\/3550454.3555506"},{"key":"e_1_2_2_73_1","volume-title":"NeLF: Neural Light-transport Field for Portrait View Synthesis and Relighting. In Eurographics Symposium on Rendering.","author":"Sun Tiancheng","year":"2021","unstructured":"Tiancheng Sun, Kai-En Lin, Sai Bi, Zexiang Xu, and Ravi Ramamoorthi. 2021. NeLF: Neural Light-transport Field for Portrait View Synthesis and Relighting. In Eurographics Symposium on Rendering."},{"key":"e_1_2_2_74_1","volume-title":"Designing an Encoder for StyleGAN Image Manipulation. ACM Transactions on Graphics (SIGGRAPH)","author":"Tov Omer","year":"2021","unstructured":"Omer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik, and Daniel Cohen-Or. 2021. Designing an Encoder for StyleGAN Image Manipulation. ACM Transactions on Graphics (SIGGRAPH) (2021)."},{"key":"e_1_2_2_75_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01490"},{"key":"e_1_2_2_76_1","volume-title":"Repurposing GANs for One-shot Semantic Part Segmentation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Tritrong Nontawat","year":"2021","unstructured":"Nontawat Tritrong, Pitchaporn Rewatbowornwong, and Supasorn Suwajanakorn. 2021. Repurposing GANs for One-shot Semantic Part Segmentation. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_77_1","volume-title":"MoRF: Morphable Radiance Fields for Multiview Neural Head Modeling. In ACM SIGGRAPH 2022 Conference Proceedings.","author":"Wang Daoye","year":"2022","unstructured":"Daoye Wang, Prashanth Chandran, Gaspard Zoss, Derek Bradley, and Paulo Gotardo. 2022a. MoRF: Morphable Radiance Fields for Multiview Neural Head Modeling. In ACM SIGGRAPH 2022 Conference Proceedings."},{"key":"e_1_2_2_78_1","volume-title":"IBRNet: Learning Multi-View Image-Based Rendering. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Wang Qianqian","year":"2021","unstructured":"Qianqian Wang, Zhicheng Wang, Kyle Genova, Pratul Srinivasan, Howard Zhou, Jonathan T. Barron, Ricardo Martin-Brualla, Noah Snavely, and Thomas Funkhouser. 2021b. IBRNet: Learning Multi-View Image-Based Rendering. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_79_1","volume-title":"High-Fidelity GAN Inversion for Image Attribute Editing. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Wang Tengfei","year":"2022","unstructured":"Tengfei Wang, Yong Zhang, Yanbo Fan, Jue Wang, and Qifeng Chen. 2022c. High-Fidelity GAN Inversion for Image Attribute Editing. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_80_1","volume-title":"One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Wang Ting-Chun","year":"2021","unstructured":"Ting-Chun Wang, Arun Mallya, and Ming-Yu Liu. 2021a. One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_81_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Wang Yaohui","year":"2022","unstructured":"Yaohui Wang, Di Yang, Francois Bremond, and Antitza Dantcheva. 2022b. Latent Image Animator: Learning to Animate Images via Latent Space Navigation. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_2_2_82_1","doi-asserted-by":"crossref","unstructured":"Z. Wang A. C. Bovik H. R. Sheikh and E. P. Simoncelli. 2004. Image Quality Assessment: From Error Visibility to Structural Similarity. TIP (2004).","DOI":"10.1109\/TIP.2003.819861"},{"key":"e_1_2_2_83_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00749"},{"key":"e_1_2_2_84_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-19778-9_10"},{"key":"e_1_2_2_85_1","volume-title":"Fake It Till You Make It: Face Analysis in the Wild Using Synthetic Data Alone. In IEEE International Conference on Computer Vision (ICCV).","author":"Wood Erroll","year":"2021","unstructured":"Erroll Wood, Tadas Baltru\u0161aitis, Charlie Hewitt, Sebastian Dziadzio, Thomas J. Cashman, and Jamie Shotton. 2021. Fake It Till You Make It: Face Analysis in the Wild Using Synthetic Data Alone. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_2_2_86_1","volume-title":"Gram-hd: 3d-consistent image generation at high resolution with generative radiance manifolds. arXiv preprint arXiv:2206.07255","author":"Xiang Jianfeng","year":"2022","unstructured":"Jianfeng Xiang, Jiaolong Yang, Yu Deng, and Xin Tong. 2022. Gram-hd: 3d-consistent image generation at high resolution with generative radiance manifolds. arXiv preprint arXiv:2206.07255 (2022)."},{"key":"e_1_2_2_87_1","unstructured":"Enze Xie Wenhai Wang Zhiding Yu Anima Anandkumar Jose M Alvarez and Ping Luo. 2021. SegFormer: Simple and efficient design for semantic segmentation with transformers. In Advances in Neural Information Processing Systems (NeurIPS)."},{"key":"e_1_2_2_88_1","volume-title":"High-fidelity 3D GAN Inversion by Pseudo-multi-view Optimization. arXiv preprint arXiv:2211.15662","author":"Xie Jiaxin","year":"2022","unstructured":"Jiaxin Xie, Hao Ouyang, Jingtan Piao, Chenyang Lei, and Qifeng Chen. 2022a. High-fidelity 3D GAN Inversion by Pseudo-multi-view Optimization. arXiv preprint arXiv:2211.15662 (2022)."},{"key":"e_1_2_2_89_1","volume-title":"Computer Graphics Forum","author":"Xie Yiheng","unstructured":"Yiheng Xie, Towaki Takikawa, Shunsuke Saito, Or Litany, Shiqin Yan, Numair Khan, Federico Tombari, James Tompkin, Vincent Sitzmann, and Srinath Sridhar. 2022b. Neural fields in visual computing and beyond. In Computer Graphics Forum, Vol. 41. Wiley Online Library."},{"key":"e_1_2_2_90_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-20047-2_42"},{"key":"e_1_2_2_91_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01788"},{"key":"e_1_2_2_92_1","volume-title":"Learning to Relight Portrait Images via a Virtual Light Stage and Synthetic-to-Real Adaptation. ACM Transactions on Graphics (SIGGRAPH ASIA)","author":"Yeh Yu-Ying","year":"2022","unstructured":"Yu-Ying Yeh, Koki Nagano, Sameh Khamis, Jan Kautz, Ming-Yu Liu, and Ting-Chun Wang. 2022. Learning to Relight Portrait Images via a Virtual Light Stage and Synthetic-to-Real Adaptation. ACM Transactions on Graphics (SIGGRAPH ASIA) (2022)."},{"key":"e_1_2_2_93_1","volume-title":"3D GAN Inversion with Facial Symmetry Prior. arxiv:2211.16927","author":"Yin Fei","year":"2022","unstructured":"Fei Yin, Yong Zhang, Xuan Wang, Tengfei Wang, Xiaoyu Li, Yuan Gong, Yanbo Fan, Xiaodong Cun, \u00d6ztireli Cengiz, and Yujiu Yang. 2022. 3D GAN Inversion with Facial Symmetry Prior. arxiv:2211.16927 (2022)."},{"key":"e_1_2_2_94_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00455"},{"key":"e_1_2_2_95_1","volume-title":"Voxel- and Surface-Aligned Radiance Field for Single-Image Novel View Synthesis. In ACM International Conference on Multimedia.","author":"Yu Xianggang","year":"2022","unstructured":"Xianggang Yu, Jiapeng Tang, Yipeng Qin, Chenghong Li, Xiaoguang Han, Linchao Bao, and Shuguang Cui. 2022. PVSeRF: Joint Pixel-, Voxel- and Surface-Aligned Radiance Field for Single-Image Novel View Synthesis. In ACM International Conference on Multimedia."},{"key":"e_1_2_2_96_1","volume-title":"Fast Bi-layer Neural Synthesis of One-Shot Realistic Head Avatars. In European Conference on Computer Vision (ECCV).","author":"Zakharov Egor","year":"2020","unstructured":"Egor Zakharov, Aleksei Ivakhnenko, Aliaksandra Shysheya, and Victor Lempitsky. 2020. Fast Bi-layer Neural Synthesis of One-Shot Realistic Head Avatars. In European Conference on Computer Vision (ECCV)."},{"key":"e_1_2_2_97_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00068"},{"key":"e_1_2_2_98_1","volume-title":"Jacobs","author":"Zhang Xuaner","year":"2020","unstructured":"Xuaner Zhang, Jonathan T. Barron, Yun-Ta Tsai, Rohit Pandey, Xiuming Zhang, Ren Ng, and David E. Jacobs. 2020. Portrait Shadow Manipulation. ACM Transactions on Graphics (SIGGRAPH)."},{"key":"e_1_2_2_99_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01790"},{"key":"e_1_2_2_100_1","volume-title":"DatasetGAN: Efficient Labeled Data Factory with Minimal Human Effort. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Zhang Yuxuan","year":"2021","unstructured":"Yuxuan Zhang, Huan Ling, Jun Gao, Kangxue Yin, Jean-Francois Lafleche, Adela Barriuso, Antonio Torralba, and Sanja Fidler. 2021. DatasetGAN: Efficient Labeled Data Factory with Minimal Human Effort. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_101_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00364"},{"key":"e_1_2_2_102_1","volume-title":"CIPS-3D: A 3D-Aware Generator of GANs Based on Conditionally-Independent Pixel Synthesis. arXiv preprint arXiv:2110.09788","author":"Zhou Peng","year":"2021","unstructured":"Peng Zhou, Lingxi Xie, Bingbing Ni, and Qi Tian. 2021. CIPS-3D: A 3D-Aware Generator of GANs Based on Conditionally-Independent Pixel Synthesis. arXiv preprint arXiv:2110.09788 (2021)."}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3592460","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3592460","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3592460","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:59Z","timestamp":1750182539000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3592460"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,7,26]]},"references-count":102,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2023,8]]}},"alternative-id":["10.1145\/3592460"],"URL":"https:\/\/doi.org\/10.1145\/3592460","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,7,26]]},"assertion":[{"value":"2023-07-26","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}