{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T16:48:40Z","timestamp":1777654120015,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":42,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Alibaba-NTU Singapore Joint Research Institute"},{"name":"Alibaba Group"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547935","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:42:46Z","timestamp":1665416566000},"page":"3637-3645","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":31,"title":["Towards Counterfactual Image Manipulation via CLIP"],"prefix":"10.1145","author":[{"given":"Yingchen","family":"Yu","sequence":"first","affiliation":[{"name":"Nanyang Technological University &amp; Alibaba Group, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Fangneng","family":"Zhan","sequence":"additional","affiliation":[{"name":"Max Planck Institute for Informatics, Saarbr\u00fccken, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Rongliang","family":"Wu","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiahui","family":"Zhang","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shijian","family":"Lu","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Miaomiao","family":"Cui","sequence":"additional","affiliation":[{"name":"Alibaba Group, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xuansong","family":"Xie","sequence":"additional","affiliation":[{"name":"Alibaba Group, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xian-Sheng","family":"Hua","sequence":"additional","affiliation":[{"name":"Zhejiang University, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chunyan","family":"Miao","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00453"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447648"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00664"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW54120.2021.00220"},{"key":"e_1_3_2_2_5_1","volume-title":"Explaining image classifiers by counterfactual generation. arXiv preprint arXiv:1807.08024","author":"Chang Chun-Hao","year":"2018","unstructured":"Chun-Hao Chang , Elliot Creager , Anna Goldenberg , and David Duvenaud . 2018. Explaining image classifiers by counterfactual generation. arXiv preprint arXiv:1807.08024 ( 2018 ). Chun-Hao Chang, Elliot Creager, Anna Goldenberg, and David Duvenaud. 2018. Explaining image classifiers by counterfactual generation. arXiv preprint arXiv:1807.08024 (2018)."},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00821"},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00581"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00482"},{"key":"e_1_3_2_2_9_1","volume-title":"Stylegan-nada: Clip-guided domain adaptation of image generators. arXiv preprint arXiv:2108.00946","author":"Gal Rinon","year":"2021","unstructured":"Rinon Gal , Or Patashnik , Haggai Maron , Gal Chechik , and Daniel Cohen-Or . 2021 . Stylegan-nada: Clip-guided domain adaptation of image generators. arXiv preprint arXiv:2108.00946 (2021). Rinon Gal, Or Patashnik, Haggai Maron, Gal Chechik, and Daniel Cohen-Or. 2021. Stylegan-nada: Clip-guided domain adaptation of image generators. arXiv preprint arXiv:2108.00946 (2021)."},{"key":"e_1_3_2_2_10_1","volume-title":"Generative adversarial nets. Advances in neural information processing systems 27","author":"Goodfellow Ian","year":"2014","unstructured":"Ian Goodfellow , Jean Pouget-Abadie , Mehdi Mirza , Bing Xu , David Warde-Farley , Sherjil Ozair , Aaron Courville , and Yoshua Bengio . 2014. Generative adversarial nets. Advances in neural information processing systems 27 ( 2014 ). Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative adversarial nets. Advances in neural information processing systems 27 (2014)."},{"key":"e_1_3_2_2_11_1","first-page":"9841","article-title":"Ganspace: Discovering interpretable gan controls","volume":"33","author":"H\u00e4rk\u00f6nen Erik","year":"2020","unstructured":"Erik H\u00e4rk\u00f6nen , Aaron Hertzmann , Jaakko Lehtinen , and Sylvain Paris . 2020 . Ganspace: Discovering interpretable gan controls . Advances in Neural Information Processing Systems 33 (2020), 9841 -- 9850 . Erik H\u00e4rk\u00f6nen, Aaron Hertzmann, Jaakko Lehtinen, and Sylvain Paris. 2020. Ganspace: Discovering interpretable gan controls. Advances in Neural Information Processing Systems 33 (2020), 9841--9850.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46475-6_43"},{"key":"e_1_3_2_2_13_1","volume-title":"International Conference on Learning Representations","author":"Karras Tero","year":"2018","unstructured":"Tero Karras , Timo Aila , Samuli Laine , and Jaakko Lehtinen . 2018 . Progressive growing of gans for improved quality, stability, and variation . International Conference on Learning Representations (2018). Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen. 2018. Progressive growing of gans for improved quality, stability, and variation. International Conference on Learning Representations (2018)."},{"key":"e_1_3_2_2_14_1","volume-title":"Alias-free generative adversarial networks. Advances in Neural Information Processing Systems 34","author":"Karras Tero","year":"2021","unstructured":"Tero Karras , Miika Aittala , Samuli Laine , Erik H\u00e4rk\u00f6nen , Janne Hellsten , Jaakko Lehtinen , and Timo Aila . 2021. Alias-free generative adversarial networks. Advances in Neural Information Processing Systems 34 ( 2021 ). Tero Karras, Miika Aittala, Samuli Laine, Erik H\u00e4rk\u00f6nen, Janne Hellsten, Jaakko Lehtinen, and Timo Aila. 2021. Alias-free generative adversarial networks. Advances in Neural Information Processing Systems 34 (2021)."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00453"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00813"},{"key":"e_1_3_2_2_17_1","volume-title":"Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Jimmy Ba . 2014 . Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014). Diederik P Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_3_2_2_18_1","volume-title":"Clipstyler: Image style transfer with a single text condition. arXiv preprint arXiv:2112.00374","author":"Kwon Gihyun","year":"2021","unstructured":"Gihyun Kwon and Jong Chul Ye . 2021 . Clipstyler: Image style transfer with a single text condition. arXiv preprint arXiv:2112.00374 (2021). Gihyun Kwon and Jong Chul Ye. 2021. Clipstyler: Image style transfer with a single text condition. arXiv preprint arXiv:2112.00374 (2021)."},{"key":"e_1_3_2_2_19_1","volume-title":"FuseDream: Training-Free Text-to-Image Generation with Improved CLIP GAN Space Optimization. arXiv preprint arXiv:2112.01573","author":"Liu Xingchao","year":"2021","unstructured":"Xingchao Liu , Chengyue Gong , Lemeng Wu , Shujian Zhang , Hao Su , and Qiang Liu . 2021. FuseDream: Training-Free Text-to-Image Generation with Improved CLIP GAN Space Optimization. arXiv preprint arXiv:2112.01573 ( 2021 ). Xingchao Liu, Chengyue Gong, Lemeng Wu, Shujian Zhang, Hao Su, and Qiang Liu. 2021. FuseDream: Training-Free Text-to-Image Generation with Improved CLIP GAN Space Optimization. arXiv preprint arXiv:2112.01573 (2021)."},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.425"},{"key":"e_1_3_2_2_21_1","volume-title":"Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784","author":"Mirza Mehdi","year":"2014","unstructured":"Mehdi Mirza and Simon Osindero . 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 ( 2014 ). Mehdi Mirza and Simon Osindero. 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 (2014)."},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58545-7_19"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00209"},{"key":"e_1_3_2_2_24_1","volume-title":"International Conference on Machine Learning. PMLR, 8748--8763","author":"Radford Alec","year":"2021","unstructured":"Alec Radford , Jong Wook Kim , Chris Hallacy , Aditya Ramesh , Gabriel Goh , Sandhini Agarwal , Girish Sastry , Amanda Askell , Pamela Mishkin , Jack Clark , 2021 . Learning transferable visual models from natural language supervision . In International Conference on Machine Learning. PMLR, 8748--8763 . Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021. Learning transferable visual models from natural language supervision. In International Conference on Machine Learning. PMLR, 8748--8763."},{"key":"e_1_3_2_2_25_1","volume-title":"International Conference on Machine Learning. PMLR, 8821--8831","author":"Ramesh Aditya","year":"2021","unstructured":"Aditya Ramesh , Mikhail Pavlov , Gabriel Goh , Scott Gray , Chelsea Voss , Alec Radford , Mark Chen , and Ilya Sutskever . 2021 . Zero-shot text-to-image generation . In International Conference on Machine Learning. PMLR, 8821--8831 . Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021. Zero-shot text-to-image generation. In International Conference on Machine Learning. PMLR, 8821--8831."},{"key":"e_1_3_2_2_26_1","unstructured":"Scott Reed Zeynep Akata Xinchen Yan Lajanugen Logeswaran Bernt Schiele and Honglak Lee. 2016. Generative adversarial text to image synthesis. In Inter- national conference on machine learning. PMLR 1060--1069.  Scott Reed Zeynep Akata Xinchen Yan Lajanugen Logeswaran Bernt Schiele and Honglak Lee. 2016. Generative adversarial text to image synthesis. In Inter- national conference on machine learning. PMLR 1060--1069."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00232"},{"key":"e_1_3_2_2_28_1","volume-title":"Counterfactual generative networks. arXiv preprint arXiv:2101.06046","author":"Sauer Axel","year":"2021","unstructured":"Axel Sauer and Andreas Geiger . 2021. Counterfactual generative networks. arXiv preprint arXiv:2101.06046 ( 2021 ). Axel Sauer and Andreas Geiger. 2021. Counterfactual generative networks. arXiv preprint arXiv:2101.06046 (2021)."},{"key":"e_1_3_2_2_29_1","volume-title":"Interfacegan: Interpreting the disentangled face representation learned by gans","author":"Shen Yujun","year":"2020","unstructured":"Yujun Shen , Ceyuan Yang , Xiaoou Tang , and Bolei Zhou . 2020 . Interfacegan: Interpreting the disentangled face representation learned by gans . IEEE transactions on pattern analysis and machine intelligence (2020). Yujun Shen, Ceyuan Yang, Xiaoou Tang, and Bolei Zhou. 2020. Interfacegan: Interpreting the disentangled face representation learned by gans. IEEE transactions on pattern analysis and machine intelligence (2020)."},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3450626.3459838"},{"key":"e_1_3_2_2_31_1","volume-title":"Representation learning with contrastive predictive coding. arXiv e-prints","author":"den Oord Aaron Van","year":"2018","unstructured":"Aaron Van den Oord , Yazhe Li , and Oriol Vinyals . 2018. Representation learning with contrastive predictive coding. arXiv e-prints ( 2018 ), arXiv--1807. Aaron Van den Oord, Yazhe Li, and Oriol Vinyals. 2018. Representation learning with contrastive predictive coding. arXiv e-prints (2018), arXiv--1807."},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58610-2_46"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00507"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01267"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00229"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00143"},{"key":"e_1_3_2_2_37_1","volume-title":"Bi-level feature alignment for versatile image translation and manipulation. arXiv preprint arXiv:2107.03021","author":"Zhan Fangneng","year":"2021","unstructured":"Fangneng Zhan , Yingchen Yu , Rongliang Wu , Kaiwen Cui , Aoran Xiao , Shijian Lu , and Ling Shao . 2021. Bi-level feature alignment for versatile image translation and manipulation. arXiv preprint arXiv:2107.03021 ( 2021 ). Fangneng Zhan, Yingchen Yu, Rongliang Wu, Kaiwen Cui, Aoran Xiao, Shijian Lu, and Ling Shao. 2021. Bi-level feature alignment for versatile image translation and manipulation. arXiv preprint arXiv:2107.03021 (2021)."},{"key":"e_1_3_2_2_38_1","volume-title":"Multimodal image synthesis and editing: A survey. arXiv preprint arXiv:2112.13592","author":"Zhan Fangneng","year":"2021","unstructured":"Fangneng Zhan , Yingchen Yu , Rongliang Wu , Jiahui Zhang , and Shijian Lu. 2021. Multimodal image synthesis and editing: A survey. arXiv preprint arXiv:2112.13592 ( 2021 ). Fangneng Zhan, Yingchen Yu, Rongliang Wu, Jiahui Zhang, and Shijian Lu. 2021. Multimodal image synthesis and editing: A survey. arXiv preprint arXiv:2112.13592 (2021)."},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01040"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01774"},{"key":"e_1_3_2_2_41_1","volume-title":"Blind image super-resolution via contrastive representation learning. arXiv preprint arXiv:2107.00708","author":"Zhang Jiahui","year":"2021","unstructured":"Jiahui Zhang , Shijian Lu , Fangneng Zhan , and Yingchen Yu. 2021. Blind image super-resolution via contrastive representation learning. arXiv preprint arXiv:2107.00708 ( 2021 ). Jiahui Zhang, Shijian Lu, Fangneng Zhan, and Yingchen Yu. 2021. Blind image super-resolution via contrastive representation learning. arXiv preprint arXiv:2107.00708 (2021)."},{"key":"e_1_3_2_2_42_1","volume-title":"Generative visual manipulation on the natural image manifold","author":"Zhu Jun-Yan","unstructured":"Jun-Yan Zhu , Philipp Kr\u00e4henb\u00fchl , Eli Shechtman , and Alexei A Efros . 2016. Generative visual manipulation on the natural image manifold . In ECCV. Springer , 597--613. Jun-Yan Zhu, Philipp Kr\u00e4henb\u00fchl, Eli Shechtman, and Alexei A Efros. 2016. Generative visual manipulation on the natural image manifold. In ECCV. Springer, 597--613."}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547935","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547935","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:31Z","timestamp":1750186831000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547935"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":42,"alternative-id":["10.1145\/3503161.3547935","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547935","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}