{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,13]],"date-time":"2026-01-13T03:53:41Z","timestamp":1768276421614,"version":"3.49.0"},"publisher-location":"New York, NY, USA","reference-count":48,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413505","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T13:10:44Z","timestamp":1602508244000},"page":"1357-1365","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":52,"title":["Describe What to Change"],"prefix":"10.1145","author":[{"given":"Yahui","family":"Liu","sequence":"first","affiliation":[{"name":"University of Trento &amp; FBK, Trento, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Marco","family":"De Nadai","sequence":"additional","affiliation":[{"name":"FBK, Trento, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Deng","family":"Cai","sequence":"additional","affiliation":[{"name":"The Chinese University of Hong Kong, Hong Kong, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Huayang","family":"Li","sequence":"additional","affiliation":[{"name":"Tencent AI Lab, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xavier","family":"Alameda-Pineda","sequence":"additional","affiliation":[{"name":"Inria Grenoble, Grenoble, France"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Nicu","family":"Sebe","sequence":"additional","affiliation":[{"name":"University of Trento &amp; Huawei Ireland, Trento, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Bruno","family":"Lepri","sequence":"additional","affiliation":[{"name":"FBK, Trento, Italy"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","unstructured":"Martin Arjovsky Soumith Chintala and L\u00e9on Bottou. 2017. Wasserstein generative adversarial networks. In ICML. 214--223.  Martin Arjovsky Soumith Chintala and L\u00e9on Bottou. 2017. Wasserstein generative adversarial networks. In ICML. 214--223."},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3306346.3323023"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00051"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"crossref","unstructured":"Jianbo Chen Yelong Shen Jianfeng Gao Jingjing Liu and Xiaodong Liu. 2018. Language-based image editing with recurrent attentive models. In CVPR.  Jianbo Chen Yelong Shen Jianfeng Gao Jingjing Liu and Xiaodong Liu. 2018. Language-based image editing with recurrent attentive models. In CVPR.","DOI":"10.1109\/CVPR.2018.00909"},{"key":"e_1_3_2_2_5_1","unstructured":"Yu Cheng Zhe Gan Yitong Li Jingjing Liu and Jianfeng Gao. 2018. Sequential attention gan for interactive image editing via dialogue. In AAAI.  Yu Cheng Zhe Gan Yitong Li Jingjing Liu and Jianfeng Gao. 2018. Sequential attention gan for interactive image editing via dialogue. In AAAI."},{"key":"e_1_3_2_2_6_1","volume-title":"Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR. 8789--8797.","author":"Choi Yunjey","year":"2018","unstructured":"Yunjey Choi , Minje Choi , Munyoung Kim , Jung-Woo Ha , Sunghun Kim , and Jaegul Choo . 2018 . Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR. 8789--8797. Yunjey Choi, Minje Choi, Munyoung Kim, Jung-Woo Ha, Sunghun Kim, and Jaegul Choo. 2018. Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR. 8789--8797."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"crossref","unstructured":"Yunjey Choi Youngjung Uh Jaejun Yoo and Jung-Woo Ha. 2020. StarGAN v2: Diverse Image Synthesis for Multiple Domains. In CVPR.  Yunjey Choi Youngjung Uh Jaejun Yoo and Jung-Woo Ha. 2020. StarGAN v2: Diverse Image Synthesis for Multiple Domains. In CVPR.","DOI":"10.1109\/CVPR42600.2020.00821"},{"key":"e_1_3_2_2_8_1","volume-title":"Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor.","author":"El-Nouby Alaaeldin","year":"2018","unstructured":"Alaaeldin El-Nouby , Shikhar Sharma , Hannes Schulz , Devon Hjelm , Layla El Asri , Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor. 2018 . Keep Drawing It: Iterative language-based image generation and editing. In NIPS. Alaaeldin El-Nouby, Shikhar Sharma, Hannes Schulz, Devon Hjelm, Layla El Asri, Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor. 2018. Keep Drawing It: Iterative language-based image generation and editing. In NIPS."},{"key":"e_1_3_2_2_9_1","volume-title":"Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor.","author":"El-Nouby Alaaeldin","year":"2019","unstructured":"Alaaeldin El-Nouby , Shikhar Sharma , Hannes Schulz , Devon Hjelm , Layla El Asri , Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor. 2019 . Tell, draw, and repeat: Generating and modifying images based on continual linguistic instruction. In ICCV. 10304--10312. Alaaeldin El-Nouby, Shikhar Sharma, Hannes Schulz, Devon Hjelm, Layla El Asri, Samira Ebrahimi Kahou, Yoshua Bengio, and Graham W Taylor. 2019. Tell, draw, and repeat: Generating and modifying images based on continual linguistic instruction. In ICCV. 10304--10312."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/3306346.3323028"},{"key":"e_1_3_2_2_11_1","volume-title":"Joost Van De Weijer, and Yoshua Bengio","author":"Gonzalez-Garcia Abel","year":"2018","unstructured":"Abel Gonzalez-Garcia , Joost Van De Weijer, and Yoshua Bengio . 2018 . Image-to-image translation for cross-domain disentanglement. In Advances in neural information processing systems. 1287--1298. Abel Gonzalez-Garcia, Joost Van De Weijer, and Yoshua Bengio. 2018. Image-to-image translation for cross-domain disentanglement. In Advances in neural information processing systems. 1287--1298."},{"key":"e_1_3_2_2_12_1","unstructured":"Ian Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron Courville and Yoshua Bengio. 2014. Generative adversarial nets. In NIPS. 2672--2680.  Ian Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron Courville and Yoshua Bengio. 2014. Generative adversarial nets. In NIPS. 2672--2680."},{"key":"e_1_3_2_2_13_1","unstructured":"Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In NIPS.  Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In NIPS."},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"crossref","unstructured":"Xun Huang and Serge Belongie. 2017. Arbitrary style transfer in real-time with adaptive instance normalization. In ICCV. 1501--1510.  Xun Huang and Serge Belongie. 2017. Arbitrary style transfer in real-time with adaptive instance normalization. In ICCV. 1501--1510.","DOI":"10.1109\/ICCV.2017.167"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"crossref","unstructured":"Xun Huang Ming-Yu Liu Serge Belongie and Jan Kautz. 2018. Multimodal unsupervised image-to-image translation. In ECCV. 172--189.  Xun Huang Ming-Yu Liu Serge Belongie and Jan Kautz. 2018. Multimodal unsupervised image-to-image translation. In ECCV. 172--189.","DOI":"10.1007\/978-3-030-01219-9_11"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"crossref","unstructured":"Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR. 1125--1134.  Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR. 1125--1134.","DOI":"10.1109\/CVPR.2017.632"},{"key":"e_1_3_2_2_17_1","unstructured":"Youngjoo Jo and Jongyoul Park. 2019. SC-FEGAN: Face Editing Generative Adversarial Network with User's Sketch and Color. In ICCV.  Youngjoo Jo and Jongyoul Park. 2019. SC-FEGAN: Face Editing Generative Adversarial Network with User's Sketch and Color. In ICCV."},{"key":"e_1_3_2_2_18_1","volume-title":"One Model To Learn Them All. ArXiV","author":"Kaiser Lukasz","year":"2017","unstructured":"Lukasz Kaiser , Aidan N Gomez , Noam Shazeer , Ashish Vaswani , Niki Parmar , Llion Jones , and Jakob Uszkoreit . 2017. One Model To Learn Them All. ArXiV ( 2017 ). Lukasz Kaiser, Aidan N Gomez, Noam Shazeer, Ashish Vaswani, Niki Parmar, Llion Jones, and Jakob Uszkoreit. 2017. One Model To Learn Them All. ArXiV (2017)."},{"key":"e_1_3_2_2_19_1","volume-title":"2019 b. Controllable Text-to-Image Generation. arXiv preprint arXiv:1909.07083","author":"Li Bowen","year":"2019","unstructured":"Bowen Li , Xiaojuan Qi , Thomas Lukasiewicz , and Philip HS Torr . 2019 b. Controllable Text-to-Image Generation. arXiv preprint arXiv:1909.07083 ( 2019 ). Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, and Philip HS Torr. 2019 b. Controllable Text-to-Image Generation. arXiv preprint arXiv:1909.07083 (2019)."},{"key":"e_1_3_2_2_20_1","volume-title":"2019 c. ManiGAN: Text-Guided Image Manipulation. arXiv preprint arXiv:1912.06203","author":"Li Bowen","year":"2019","unstructured":"Bowen Li , Xiaojuan Qi , Thomas Lukasiewicz , and Philip HS Torr . 2019 c. ManiGAN: Text-Guided Image Manipulation. arXiv preprint arXiv:1912.06203 ( 2019 ). Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, and Philip HS Torr. 2019 c. ManiGAN: Text-Guided Image Manipulation. arXiv preprint arXiv:1912.06203 (2019)."},{"key":"e_1_3_2_2_21_1","unstructured":"Yitong Li Zhe Gan Yelong Shen Jingjing Liu Yu Cheng Yuexin Wu Lawrence Carin David Carlson and Jianfeng Gao. 2019 a. StoryGAN: A Sequential Conditional GAN for Story Visualization. In CVPR. 6329--6338.  Yitong Li Zhe Gan Yelong Shen Jingjing Liu Yu Cheng Yuexin Wu Lawrence Carin David Carlson and Jianfeng Gao. 2019 a. StoryGAN: A Sequential Conditional GAN for Story Visualization. In CVPR. 6329--6338."},{"key":"e_1_3_2_2_22_1","volume-title":"Jian Yao, Nicu Sebe, Bruno Lepri, and Xavier Alameda-Pineda.","author":"Liu Yahui","year":"2020","unstructured":"Yahui Liu , Marco De Nadai , Jian Yao, Nicu Sebe, Bruno Lepri, and Xavier Alameda-Pineda. 2020 . GMM-UNIT: Unsupervised Multi-Domain and Multi-Modal Image-to-Image Translation via Attribute Gaussian Mixture Modeling . arXiv preprint arXiv:2003.06788 (2020). Yahui Liu, Marco De Nadai, Jian Yao, Nicu Sebe, Bruno Lepri, and Xavier Alameda-Pineda. 2020. GMM-UNIT: Unsupervised Multi-Domain and Multi-Modal Image-to-Image Translation via Attribute Gaussian Mixture Modeling. arXiv preprint arXiv:2003.06788 (2020)."},{"key":"e_1_3_2_2_23_1","volume-title":"Gloria Zen, Nicu Sebe, and Bruno Lepri.","author":"Liu Yahui","year":"2019","unstructured":"Yahui Liu , Marco De Nadai , Gloria Zen, Nicu Sebe, and Bruno Lepri. 2019 . Gesture-to-gesture translation in the wild via category-independent conditional maps. ACM MM ( 2019). Yahui Liu, Marco De Nadai, Gloria Zen, Nicu Sebe, and Bruno Lepri. 2019. Gesture-to-gesture translation in the wild via category-independent conditional maps. ACM MM (2019)."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"crossref","unstructured":"Ziwei Liu Ping Luo Xiaogang Wang and Xiaoou Tang. 2015. Deep learning face attributes in the wild. In ICCV. 3730--3738.  Ziwei Liu Ping Luo Xiaogang Wang and Xiaoou Tang. 2015. Deep learning face attributes in the wild. In ICCV. 3730--3738.","DOI":"10.1109\/ICCV.2015.425"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"crossref","unstructured":"Qi Mao Hsin-Ying Lee Hung-Yu Tseng Siwei Ma and Ming-Hsuan Yang. 2019. Mode seeking generative adversarial networks for diverse image synthesis. In CVPR. 1429--1437.  Qi Mao Hsin-Ying Lee Hung-Yu Tseng Siwei Ma and Ming-Hsuan Yang. 2019. Mode seeking generative adversarial networks for diverse image synthesis. In CVPR. 1429--1437.","DOI":"10.1109\/CVPR.2019.00152"},{"key":"e_1_3_2_2_26_1","volume-title":"Zhen Wang, and Stephen Paul Smolley.","author":"Mao Xudong","year":"2017","unstructured":"Xudong Mao , Qing Li , Haoran Xie , Raymond YK Lau , Zhen Wang, and Stephen Paul Smolley. 2017 . Least squares generative adversarial networks. In ICCV. Xudong Mao, Qing Li, Haoran Xie, Raymond YK Lau, Zhen Wang, and Stephen Paul Smolley. 2017. Least squares generative adversarial networks. In ICCV."},{"key":"e_1_3_2_2_27_1","unstructured":"Youssef Alami Mejjati Christian Richardt James Tompkin Darren Cosker and Kwang In Kim. 2018. Unsupervised attention-guided image-to-image translation. In Advances in Neural Information Processing Systems. 3693--3703.  Youssef Alami Mejjati Christian Richardt James Tompkin Darren Cosker and Kwang In Kim. 2018. Unsupervised attention-guided image-to-image translation. In Advances in Neural Information Processing Systems. 3693--3703."},{"key":"e_1_3_2_2_28_1","volume-title":"Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784","author":"Mirza Mehdi","year":"2014","unstructured":"Mehdi Mirza and Simon Osindero . 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 ( 2014 ). Mehdi Mirza and Simon Osindero. 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 (2014)."},{"key":"e_1_3_2_2_29_1","unstructured":"Sangwoo Mo Minsu Cho and Jinwoo Shin. 2019. Instance-aware Image-to-Image Translation. In ICLR.  Sangwoo Mo Minsu Cho and Jinwoo Shin. 2019. Instance-aware Image-to-Image Translation. In ICLR."},{"key":"e_1_3_2_2_30_1","volume-title":"Trends in Integration of Vision and Language Research: A Survey of Tasks, Datasets, and Methods. arXiv preprint arXiv:1907.09358","author":"Mogadala Aditya","year":"2019","unstructured":"Aditya Mogadala , Marimuthu Kalimuthu , and Dietrich Klakow . 2019. Trends in Integration of Vision and Language Research: A Survey of Tasks, Datasets, and Methods. arXiv preprint arXiv:1907.09358 ( 2019 ). Aditya Mogadala, Marimuthu Kalimuthu, and Dietrich Klakow. 2019. Trends in Integration of Vision and Language Research: A Survey of Tasks, Datasets, and Methods. arXiv preprint arXiv:1907.09358 (2019)."},{"key":"e_1_3_2_2_31_1","unstructured":"Seonghyeon Nam Yunji Kim and Seon Joo Kim. 2018. Text-adaptive generative adversarial networks: manipulating images with natural language. In NIPS.  Seonghyeon Nam Yunji Kim and Seon Joo Kim. 2018. Text-adaptive generative adversarial networks: manipulating images with natural language. In NIPS."},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"crossref","unstructured":"Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR.  Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR.","DOI":"10.1109\/CVPR.2016.278"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01249-6_50"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"crossref","unstructured":"Tingting Qiao Jing Zhang Duanqing Xu and Dacheng Tao. 2019. MirrorGAN: Learning Text-to-image Generation by Redescription. In CVPR. 1505--1514.  Tingting Qiao Jing Zhang Duanqing Xu and Dacheng Tao. 2019. MirrorGAN: Learning Text-to-image Generation by Redescription. In CVPR. 1505--1514.","DOI":"10.1109\/CVPR.2019.00160"},{"key":"e_1_3_2_2_35_1","unstructured":"Scott Reed Zeynep Akata Xinchen Yan Lajanugen Logeswaran Bernt Schiele and Honglak Lee. 2016b. Generative Adversarial Text to Image Synthesis. In ICML.  Scott Reed Zeynep Akata Xinchen Yan Lajanugen Logeswaran Bernt Schiele and Honglak Lee. 2016b. Generative Adversarial Text to Image Synthesis. In ICML."},{"key":"e_1_3_2_2_36_1","unstructured":"Scott E Reed Zeynep Akata Santosh Mohan Samuel Tenka Bernt Schiele and Honglak Lee. 2016a. Learning what and where to draw. In NIPS. 217--225.  Scott E Reed Zeynep Akata Santosh Mohan Samuel Tenka Bernt Schiele and Honglak Lee. 2016a. Learning what and where to draw. In NIPS. 217--225."},{"key":"e_1_3_2_2_37_1","unstructured":"Tim Salimans Ian Goodfellow Wojciech Zaremba Vicki Cheung Alec Radford and Xi Chen. 2016. Improved techniques for training gans. In NIPS. 2234--2242.  Tim Salimans Ian Goodfellow Wojciech Zaremba Vicki Cheung Alec Radford and Xi Chen. 2016. Improved techniques for training gans. In NIPS. 2234--2242."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"crossref","unstructured":"Christian Szegedy Vincent Vanhoucke Sergey Ioffe Jon Shlens and Zbigniew Wojna. 2016. Rethinking the inception architecture for computer vision. In CVPR.  Christian Szegedy Vincent Vanhoucke Sergey Ioffe Jon Shlens and Zbigniew Wojna. 2016. Rethinking the inception architecture for computer vision. In CVPR.","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_3_2_2_40_1","volume-title":"Learning to Globally Edit Images with Textual Description. arXiv preprint arXiv:1810.05786","author":"Wang Hai","year":"2018","unstructured":"Hai Wang , Jason D Williams , and SingBing Kang . 2018. Learning to Globally Edit Images with Textual Description. arXiv preprint arXiv:1810.05786 ( 2018 ). Hai Wang, Jason D Williams, and SingBing Kang. 2018. Learning to Globally Edit Images with Textual Description. arXiv preprint arXiv:1810.05786 (2018)."},{"key":"e_1_3_2_2_41_1","volume-title":"Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR. 1316--1324.","author":"Xu Tao","year":"2018","unstructured":"Tao Xu , Pengchuan Zhang , Qiuyuan Huang , Han Zhang , Zhe Gan , Xiaolei Huang , and Xiaodong He . 2018 . Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR. 1316--1324. Tao Xu, Pengchuan Zhang, Qiuyuan Huang, Han Zhang, Zhe Gan, Xiaolei Huang, and Xiaodong He. 2018. Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR. 1316--1324."},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"crossref","unstructured":"Ting Yao Yingwei Pan Yehao Li Zhaofan Qiu and Tao Mei. 2017. Boosting image captioning with attributes. In ICCV. 4894--4902.  Ting Yao Yingwei Pan Yehao Li Zhaofan Qiu and Tao Mei. 2017. Boosting image captioning with attributes. In ICCV. 4894--4902.","DOI":"10.1109\/ICCV.2017.524"},{"key":"e_1_3_2_2_43_1","volume-title":"Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In CVPR. 5907--5915.","author":"Zhang Han","year":"2017","unstructured":"Han Zhang , Tao Xu , Hongsheng Li , Shaoting Zhang , Xiaogang Wang , Xiaolei Huang , and Dimitris N Metaxas . 2017 . Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In CVPR. 5907--5915. Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris N Metaxas. 2017. Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In CVPR. 5907--5915."},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2018.2856256"},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"crossref","unstructured":"Richard Zhang Phillip Isola Alexei A Efros Eli Shechtman and Oliver Wang. 2018a. The unreasonable effectiveness of deep features as a perceptual metric. In CVPR. 586--595.  Richard Zhang Phillip Isola Alexei A Efros Eli Shechtman and Oliver Wang. 2018a. The unreasonable effectiveness of deep features as a perceptual metric. In CVPR. 586--595.","DOI":"10.1109\/CVPR.2018.00068"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"crossref","unstructured":"Chuanxia Zheng Tat-Jen Cham and Jianfei Cai. 2019. Pluralistic Image Completion. In CVPR. 1438--1447.  Chuanxia Zheng Tat-Jen Cham and Jianfei Cai. 2019. Pluralistic Image Completion. In CVPR. 1438--1447.","DOI":"10.1109\/CVPR.2019.00153"},{"key":"e_1_3_2_2_47_1","unstructured":"Jun-Yan Zhu Taesung Park Phillip Isola and Alexei A Efros. 2017a. Unpaired image-to-image translation using cycle-consistent adversarial networks. In ICCV.  Jun-Yan Zhu Taesung Park Phillip Isola and Alexei A Efros. 2017a. Unpaired image-to-image translation using cycle-consistent adversarial networks. In ICCV."},{"key":"e_1_3_2_2_48_1","unstructured":"Jun-Yan Zhu Richard Zhang Deepak Pathak Trevor Darrell Alexei A Efros Oliver Wang and Eli Shechtman. 2017b. Toward multimodal image-to-image translation. In NIPS. 465--476.  Jun-Yan Zhu Richard Zhang Deepak Pathak Trevor Darrell Alexei A Efros Oliver Wang and Eli Shechtman. 2017b. Toward multimodal image-to-image translation. In NIPS. 465--476."},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/3355089.3356561"}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","location":"Seattle WA USA","acronym":"MM '20","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413505","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413505","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:47:13Z","timestamp":1750193233000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413505"}},"subtitle":["A Text-guided Unsupervised Image-to-image Translation Approach"],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":48,"alternative-id":["10.1145\/3394171.3413505","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413505","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}