{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,8]],"date-time":"2026-07-08T15:56:31Z","timestamp":1783526191143,"version":"3.55.0"},"reference-count":49,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2019,7,12]],"date-time":"2019-07-12T00:00:00Z","timestamp":1562889600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"DOI":"10.13039\/100000185","name":"DARPA","doi-asserted-by":"crossref","award":["FA8750-18-C000"],"award-info":[{"award-number":["FA8750-18-C000"]}],"id":[{"id":"10.13039\/100000185","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/100000001","name":"NSF","doi-asserted-by":"publisher","award":["1524817"],"award-info":[{"award-number":["1524817"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Graph."],"published-print":{"date-parts":[[2019,8,31]]},"abstract":"<jats:p>Despite the recent success of GANs in synthesizing images conditioned on inputs such as a user sketch, text, or semantic labels, manipulating the high-level attributes of an existing natural photograph with GANs is challenging for two reasons. First, it is hard for GANs to precisely reproduce an input image. Second, after manipulation, the newly synthesized pixels often do not fit the original image. In this paper, we address these issues by adapting the image prior learned by GANs to image statistics of an individual image. Our method can accurately reconstruct the input image and synthesize new content, consistent with the appearance of the input image. We demonstrate our interactive system on several semantic image editing tasks, including synthesizing new objects consistent with background, removing unwanted objects, and changing the appearance of an object. Quantitative and qualitative comparisons against several existing methods demonstrate the effectiveness of our method.<\/jats:p>","DOI":"10.1145\/3306346.3323023","type":"journal-article","created":{"date-parts":[[2019,7,12]],"date-time":"2019-07-12T19:04:08Z","timestamp":1562958248000},"page":"1-11","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":199,"title":["Semantic photo manipulation with a generative image prior"],"prefix":"10.1145","volume":"38","author":[{"given":"David","family":"Bau","sequence":"first","affiliation":[{"name":"MIT CSAIL"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Hendrik","family":"Strobelt","sequence":"additional","affiliation":[{"name":"IBM Research"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"William","family":"Peebles","sequence":"additional","affiliation":[{"name":"MIT CSAIL"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jonas","family":"Wulff","sequence":"additional","affiliation":[{"name":"MIT CSAIL"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bolei","family":"Zhou","sequence":"additional","affiliation":[{"name":"The Chinese University of Hong Kong"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jun-Yan","family":"Zhu","sequence":"additional","affiliation":[{"name":"MIT CSAIL"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Antonio","family":"Torralba","sequence":"additional","affiliation":[{"name":"MIT CSAIL"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2019,7,12]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1360612.1360639"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276390"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1531326.1531330"},{"key":"e_1_2_2_4_1","unstructured":"David Bau Jun-Yan Zhu Hendrik Strobelt Zhou Bolei Joshua B. Tenenbaum William T. Freeman and Antonio Torralba. 2019. GAN Dissection: Visualizing and Understanding Generative Adversarial Networks. In ICLR.  David Bau Jun-Yan Zhu Hendrik Strobelt Zhou Bolei Joshua B. Tenenbaum William T. Freeman and Antonio Torralba. 2019. GAN Dissection: Visualizing and Understanding Generative Adversarial Networks. In ICLR."},{"key":"e_1_2_2_5_1","unstructured":"Andrew Brock Jeff Donahue and Karen Simonyan. 2019. Large scale gan training for high fidelity natural image synthesis. (2019).  Andrew Brock Jeff Donahue and Karen Simonyan. 2019. Large scale gan training for high fidelity natural image synthesis. (2019)."},{"key":"e_1_2_2_6_1","unstructured":"Andrew Brock Theodore Lim James M Ritchie and Nick Weston. 2017. Neural photo editing with introspective adversarial networks. In ICLR.  Andrew Brock Theodore Lim James M Ritchie and Nick Weston. 2017. Neural photo editing with introspective adversarial networks. In ICLR."},{"key":"e_1_2_2_7_1","volume-title":"Infogan: Interpretable representation learning by information maximizing generative adversarial nets. In NIPS.","author":"Chen Xi","year":"2016","unstructured":"Xi Chen , Yan Duan , Rein Houthooft , John Schulman , Ilya Sutskever , and Pieter Abbeel . 2016 . Infogan: Interpretable representation learning by information maximizing generative adversarial nets. In NIPS. Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel. 2016. Infogan: Interpretable representation learning by information maximizing generative adversarial nets. In NIPS."},{"key":"e_1_2_2_8_1","unstructured":"Alexey Dosovitskiy and Thomas Brox. 2016. Generating images with perceptual similarity metrics based on deep networks. In NIPS.   Alexey Dosovitskiy and Thomas Brox. 2016. Generating images with perceptual similarity metrics based on deep networks. In NIPS."},{"key":"e_1_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/566654.566574"},{"key":"e_1_2_2_10_1","volume-title":"Image Style Transfer Using Convolutional Neural Networks. CVPR","author":"Gatys Leon A","year":"2016","unstructured":"Leon A Gatys , Alexander S Ecker , and Matthias Bethge . 2016. Image Style Transfer Using Convolutional Neural Networks. CVPR ( 2016 ). Leon A Gatys, Alexander S Ecker, and Matthias Bethge. 2016. Image Style Transfer Using Convolutional Neural Networks. CVPR (2016)."},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/3272127.3275043"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073592"},{"key":"e_1_2_2_13_1","unstructured":"Ian Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron Courville and Yoshua Bengio. 2014. Generative adversarial nets. In NIPS.   Ian Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron Courville and Yoshua Bengio. 2014. Generative adversarial nets. In NIPS."},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/2897824.2925974"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073659"},{"key":"e_1_2_2_16_1","doi-asserted-by":"crossref","unstructured":"Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR.  Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR.","DOI":"10.1109\/CVPR.2017.632"},{"key":"e_1_2_2_17_1","unstructured":"Tero Karras Timo Aila Samuli Laine and Jaakko Lehtinen. 2018. Progressive growing of gans for improved quality stability and variation. In ICLR.  Tero Karras Timo Aila Samuli Laine and Jaakko Lehtinen. 2018. Progressive growing of gans for improved quality stability and variation. In ICLR."},{"key":"e_1_2_2_18_1","doi-asserted-by":"crossref","unstructured":"Tero Karras Samuli Laine and Timo Aila. 2019. A style-based generator architecture for generative adversarial networks. In CVPR.  Tero Karras Samuli Laine and Timo Aila. 2019. A style-based generator architecture for generative adversarial networks. In CVPR.","DOI":"10.1109\/CVPR.2019.00453"},{"key":"e_1_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/2070781.2024191"},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/2601097.2601209"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201377"},{"key":"e_1_2_2_22_1","volume-title":"Adam: A method for stochastic optimization. In ICLR.","author":"Kingma Diederik","year":"2015","unstructured":"Diederik Kingma and Jimmy Ba . 2015 . Adam: A method for stochastic optimization. In ICLR. Diederik Kingma and Jimmy Ba. 2015. Adam: A method for stochastic optimization. In ICLR."},{"key":"e_1_2_2_23_1","volume-title":"Auto-encoding variational bayes. ICLR","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Max Welling . 2014. Auto-encoding variational bayes. ICLR ( 2014 ). Diederik P Kingma and Max Welling. 2014. Auto-encoding variational bayes. ICLR (2014)."},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276381"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/1015706.1015780"},{"key":"e_1_2_2_26_1","unstructured":"Yijun Li Ming-Yu Liu Xueting Li Ming-Hsuan Yang and Jan Kautz. 2018. A closed-form solution to photorealistic image stylization. In ECCV.  Yijun Li Ming-Yu Liu Xueting Li Ming-Hsuan Yang and Jan Kautz. 2018. A closed-form solution to photorealistic image stylization. In ECCV."},{"key":"e_1_2_2_27_1","unstructured":"Takeru Miyato Toshiki Kataoka Masanori Koyama and Yuichi Yoshida. 2018. Spectral normalization for generative adversarial networks. In ICLR.  Takeru Miyato Toshiki Kataoka Masanori Koyama and Yuichi Yoshida. 2018. Spectral normalization for generative adversarial networks. In ICLR."},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3272127.3275075"},{"key":"e_1_2_2_29_1","doi-asserted-by":"crossref","unstructured":"Taesung Park Ming-Yu Liu Ting-Chun Wang and Jun-Yan Zhu. 2019. Semantic Image Synthesis with Spatially-Adaptive Normalization. In CVPR.  Taesung Park Ming-Yu Liu Ting-Chun Wang and Jun-Yan Zhu. 2019. Semantic Image Synthesis with Spatially-Adaptive Normalization. In CVPR.","DOI":"10.1109\/CVPR.2019.00244"},{"key":"e_1_2_2_30_1","doi-asserted-by":"crossref","unstructured":"Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context Encoders:Feature Learning by Inpainting. CVPR (2016).  Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context Encoders:Feature Learning by Inpainting. CVPR (2016).","DOI":"10.1109\/CVPR.2016.278"},{"key":"e_1_2_2_31_1","volume-title":"NIPS Workshop on Adversarial Training.","author":"Perarnau Guim","unstructured":"Guim Perarnau , Joost van de Weijer, Bogdan Raducanu, and Jose M \u00c1lvarez. 2016. Invertible conditional gans for image editing . In NIPS Workshop on Adversarial Training. Guim Perarnau, Joost van de Weijer, Bogdan Raducanu, and Jose M \u00c1lvarez. 2016. Invertible conditional gans for image editing. In NIPS Workshop on Adversarial Training."},{"key":"e_1_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/882262.882269"},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/3197517.3201393"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/38.946629"},{"key":"e_1_2_2_35_1","volume-title":"Scribbler: Controlling Deep Image Synthesis with Sketch and Color. In CVPR.","author":"Sangkloy Patsorn","year":"2017","unstructured":"Patsorn Sangkloy , Jingwan Lu , Chen Fang , Fisher Yu , and James Hays . 2017 . Scribbler: Controlling Deep Image Synthesis with Sketch and Color. In CVPR. Patsorn Sangkloy, Jingwan Lu, Chen Fang, Fisher Yu, and James Hays. 2017. Scribbler: Controlling Deep Image Synthesis with Sketch and Color. In CVPR."},{"key":"e_1_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Assaf Shocher Nadav Cohen and Michal Irani. 2018. \"Zero-Shot\" Super-Resolution using Deep Internal Learning. In CVPR.  Assaf Shocher Nadav Cohen and Michal Irani. 2018. \"Zero-Shot\" Super-Resolution using Deep Internal Learning. In CVPR.","DOI":"10.1109\/CVPR.2018.00329"},{"key":"e_1_2_2_37_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2015. Very deep convolutional networks for large-scale image recognition. In ICLR.  Karen Simonyan and Andrew Zisserman. 2015. Very deep convolutional networks for large-scale image recognition. In ICLR."},{"key":"e_1_2_2_38_1","unstructured":"Michael W Tao Micah K Johnson and Sylvain Paris. 2010. Error-tolerant image compositing. In ECCV.   Michael W Tao Micah K Johnson and Sylvain Paris. 2010. Error-tolerant image compositing. In ECCV."},{"key":"e_1_2_2_39_1","unstructured":"Dmitry Ulyanov Andrea Vedaldi and Victor Lempitsky. 2018. Deep image prior. In CVPR.  Dmitry Ulyanov Andrea Vedaldi and Victor Lempitsky. 2018. Deep image prior. In CVPR."},{"key":"e_1_2_2_40_1","unstructured":"Ting-Chun Wang Ming-Yu Liu Jun-Yan Zhu Andrew Tao Jan Kautz and Bryan Catanzaro. 2018. High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs. In CVPR.  Ting-Chun Wang Ming-Yu Liu Jun-Yan Zhu Andrew Tao Jan Kautz and Bryan Catanzaro. 2018. High-Resolution Image Synthesis and Semantic Manipulation with Conditional GANs. In CVPR."},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/2185520.2185580"},{"key":"e_1_2_2_42_1","volume-title":"Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop. arXiv preprint arXiv:1506.03365","author":"Yu Fisher","year":"2015","unstructured":"Fisher Yu , Ari Seff , Yinda Zhang , Shuran Song , Thomas Funkhouser , and Jianxiong Xiao . 2015 . Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop. arXiv preprint arXiv:1506.03365 (2015). Fisher Yu, Ari Seff, Yinda Zhang, Shuran Song, Thomas Funkhouser, and Jianxiong Xiao. 2015. Lsun: Construction of a large-scale image dataset using deep learning with humans in the loop. arXiv preprint arXiv:1506.03365 (2015)."},{"key":"e_1_2_2_43_1","volume-title":"Generative Image Inpainting With Contextual Attention. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Yu Jiahui","unstructured":"Jiahui Yu , Zhe Lin , Jimei Yang , Xiaohui Shen , Xin Lu , and Thomas S. Huang . 2018 . Generative Image Inpainting With Contextual Attention. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Jiahui Yu, Zhe Lin, Jimei Yang, Xiaohui Shen, Xin Lu, and Thomas S. Huang. 2018. Generative Image Inpainting With Contextual Attention. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/2980179.2982432"},{"key":"e_1_2_2_45_1","doi-asserted-by":"crossref","unstructured":"Han Zhang Tao Xu Hongsheng Li Shaoting Zhang Xiaogang Wang Xiaolei Huang and Dimitris Metaxas. 2017a. StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks. In ICCV.  Han Zhang Tao Xu Hongsheng Li Shaoting Zhang Xiaogang Wang Xiaolei Huang and Dimitris Metaxas. 2017a. StackGAN: Text to Photo-realistic Image Synthesis with Stacked Generative Adversarial Networks. In ICCV.","DOI":"10.1109\/ICCV.2017.629"},{"key":"e_1_2_2_46_1","doi-asserted-by":"crossref","unstructured":"Richard Zhang Phillip Isola and Alexei A Efros. 2016b. Colorful Image Colorization. In ECCV.  Richard Zhang Phillip Isola and Alexei A Efros. 2016b. Colorful Image Colorization. In ECCV.","DOI":"10.1007\/978-3-319-46487-9_40"},{"key":"e_1_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073703"},{"key":"e_1_2_2_48_1","volume-title":"Efros","author":"Zhu Jun-Yan","year":"2016","unstructured":"Jun-Yan Zhu , Philipp Kr\u00e4henb\u00fchl , Eli Shechtman , and Alexei A . Efros . 2016 . Generative Visual Manipulation on the Natural Image Manifold. In ECCV. Jun-Yan Zhu, Philipp Kr\u00e4henb\u00fchl, Eli Shechtman, and Alexei A. Efros. 2016. Generative Visual Manipulation on the Natural Image Manifold. In ECCV."},{"key":"e_1_2_2_49_1","unstructured":"Jun-Yan Zhu Taesung Park Phillip Isola and Alexei A Efros. 2017. Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. In ICCV.  Jun-Yan Zhu Taesung Park Phillip Isola and Alexei A Efros. 2017. Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. In ICCV."}],"container-title":["ACM Transactions on Graphics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3306346.3323023","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3306346.3323023","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3306346.3323023","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T00:25:52Z","timestamp":1750206352000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3306346.3323023"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2019,7,12]]},"references-count":49,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2019,8,31]]}},"alternative-id":["10.1145\/3306346.3323023"],"URL":"https:\/\/doi.org\/10.1145\/3306346.3323023","relation":{},"ISSN":["0730-0301","1557-7368"],"issn-type":[{"value":"0730-0301","type":"print"},{"value":"1557-7368","type":"electronic"}],"subject":[],"published":{"date-parts":[[2019,7,12]]},"assertion":[{"value":"2019-07-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}