{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,3]],"date-time":"2026-07-03T16:18:26Z","timestamp":1783095506809,"version":"3.54.6"},"publisher-location":"New York, NY, USA","reference-count":61,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,10,17]],"date-time":"2021-10-17T00:00:00Z","timestamp":1634428800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Alibaba-NTU Singapore Joint Research Institute","award":["Nil"],"award-info":[{"award-number":["Nil"]}]},{"name":"Alibaba Group","award":["Nil"],"award-info":[{"award-number":["Nil"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,10,17]]},"DOI":"10.1145\/3474085.3475436","type":"proceedings-article","created":{"date-parts":[[2021,10,18]],"date-time":"2021-10-18T04:59:18Z","timestamp":1634533158000},"page":"69-78","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":133,"title":["Diverse Image Inpainting with Bidirectional and Autoregressive Transformers"],"prefix":"10.1145","author":[{"given":"Yingchen","family":"Yu","sequence":"first","affiliation":[{"name":"Nanyang Technological University &amp; Alibaba Group, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fangneng","family":"Zhan","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Rongliang","family":"WU","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jianxiong","family":"Pan","sequence":"additional","affiliation":[{"name":"DAMO Academy, Alibaba Group, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kaiwen","family":"Cui","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shijian","family":"Lu","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Feiying","family":"Ma","sequence":"additional","affiliation":[{"name":"DAMO Academy, Alibaba Group, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xuansong","family":"Xie","sequence":"additional","affiliation":[{"name":"DAMO Academy, Alibaba Group, Hangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chunyan","family":"Miao","sequence":"additional","affiliation":[{"name":"Nanyang Technological University, Singapore, Singapore"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2021,10,17]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/83.935036"},{"key":"e_1_3_2_1_2_1","first-page":"10","article-title":"b. A variational model for filling-in gray level and color images","volume":"1","author":"Ballester Coloma","year":"2001","unstructured":"Coloma Ballester , Vicent Caselles , Joan Verdera , Marcelo Bertalmio , and Guillermo Sapiro . 2001 b. A variational model for filling-in gray level and color images . In ICCV , Vol. 1. 10 -- 16 . Coloma Ballester, Vicent Caselles, Joan Verdera, Marcelo Bertalmio, and Guillermo Sapiro. 2001 b. A variational model for filling-in gray level and color images. In ICCV, Vol. 1. 10--16.","journal-title":"ICCV"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1576246.1531330"},{"key":"e_1_3_2_1_4_1","first-page":"355","article-title":"Navier-stokes, fluid dynamics, and image and video inpainting","volume":"1","author":"Bertalmio Marcelo","year":"2001","unstructured":"Marcelo Bertalmio , Andrea L Bertozzi , and Guillermo Sapiro . 2001 . Navier-stokes, fluid dynamics, and image and video inpainting . In CVPR , Vol. 1. 355 -- 355 . Marcelo Bertalmio, Andrea L Bertozzi, and Guillermo Sapiro. 2001. Navier-stokes, fluid dynamics, and image and video inpainting. In CVPR, Vol. 1. 355--355.","journal-title":"CVPR"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/344779.344972"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2003.815261"},{"key":"e_1_3_2_1_7_1","volume-title":"End-to-End Object Detection with Transformers. arXiv preprint arXiv:2005.12872","author":"Carion Nicolas","year":"2020","unstructured":"Nicolas Carion , Francisco Massa , Gabriel Synnaeve , Nicolas Usunier , Alexander Kirillov , and Sergey Zagoruyko . 2020. End-to-End Object Detection with Transformers. arXiv preprint arXiv:2005.12872 ( 2020 ). Nicolas Carion, Francisco Massa, Gabriel Synnaeve, Nicolas Usunier, Alexander Kirillov, and Sergey Zagoruyko. 2020. End-to-End Object Detection with Transformers. arXiv preprint arXiv:2005.12872 (2020)."},{"key":"e_1_3_2_1_8_1","volume-title":"International Conference on Machine Learning. PMLR, 1691--1703","author":"Chen Mark","year":"2020","unstructured":"Mark Chen , Alec Radford , Rewon Child , Jeffrey Wu , Heewoo Jun , David Luan , and Ilya Sutskever . 2020 . Generative pretraining from pixels . In International Conference on Machine Learning. PMLR, 1691--1703 . Mark Chen, Alec Radford, Rewon Child, Jeffrey Wu, Heewoo Jun, David Luan, and Ilya Sutskever. 2020. Generative pretraining from pixels. In International Conference on Machine Learning. PMLR, 1691--1703."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/2185520.2185578"},{"key":"e_1_3_2_1_10_1","volume-title":"Imagenet: A large-scale hierarchical image database. In CVPR. Ieee, 248--255.","author":"Deng Jia","year":"2009","unstructured":"Jia Deng , Wei Dong , Richard Socher , Li-Jia Li , Kai Li , and Li Fei-Fei . 2009 . Imagenet: A large-scale hierarchical image database. In CVPR. Ieee, 248--255. Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009. Imagenet: A large-scale hierarchical image database. In CVPR. Ieee, 248--255."},{"key":"e_1_3_2_1_11_1","volume-title":"Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805","author":"Devlin Jacob","year":"2018","unstructured":"Jacob Devlin , Ming-Wei Chang , Kenton Lee , and Kristina Toutanova . 2018 . Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018). Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv preprint arXiv:1810.04805 (2018)."},{"key":"e_1_3_2_1_12_1","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly etal 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 (2020).  Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly et al. 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 (2020)."},{"key":"e_1_3_2_1_13_1","volume-title":"Taming Transformers for High-Resolution Image Synthesis. arXiv preprint arXiv:2012.09841","author":"Esser Patrick","year":"2020","unstructured":"Patrick Esser , Robin Rombach , and Bj\u00f6rn Ommer . 2020. Taming Transformers for High-Resolution Image Synthesis. arXiv preprint arXiv:2012.09841 ( 2020 ). Patrick Esser, Robin Rombach, and Bj\u00f6rn Ommer. 2020. Taming Transformers for High-Resolution Image Synthesis. arXiv preprint arXiv:2012.09841 (2020)."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.5555\/2969033.2969125"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/1276377.1276382"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.5555\/3295222.3295408"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073659"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"crossref","unstructured":"Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR. 1125--1134.  Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017. Image-to-image translation with conditional adversarial networks. In CVPR. 1125--1134.","DOI":"10.1109\/CVPR.2017.632"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46475-6_43"},{"key":"e_1_3_2_1_20_1","volume-title":"International Conference on Learning Representations","author":"Karras Tero","year":"2018","unstructured":"Tero Karras , Timo Aila , Samuli Laine , and Jaakko Lehtinen . 2018 . Progressive growing of gans for improved quality, stability, and variation . International Conference on Learning Representations (2018). Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen. 2018. Progressive growing of gans for improved quality, stability, and variation. International Conference on Learning Representations (2018)."},{"key":"e_1_3_2_1_21_1","volume-title":"Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Jimmy Ba . 2014 . Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014). Diederik P Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_3_2_1_22_1","volume-title":"Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114","author":"Kingma Diederik P","year":"2013","unstructured":"Diederik P Kingma and Max Welling . 2013. Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 ( 2013 ). Diederik P Kingma and Max Welling. 2013. Auto-encoding variational bayes. arXiv preprint arXiv:1312.6114 (2013)."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00985"},{"key":"e_1_3_2_1_24_1","unstructured":"Guilin Liu Fitsum A Reda Kevin J Shih Ting-Chun Wang Andrew Tao and Bryan Catanzaro. 2018. Image inpainting for irregular holes using partial convolutions. In ECCV. 85--100.  Guilin Liu Fitsum A Reda Kevin J Shih Ting-Chun Wang Andrew Tao and Bryan Catanzaro. 2018. Image inpainting for irregular holes using partial convolutions. In ECCV. 85--100."},{"key":"e_1_3_2_1_25_1","unstructured":"Hongyu Liu Bin Jiang Yibing Song Wei Huang and Chao Yang. 2020 a. Rethinking Image Inpainting via a Mutual Encoder-Decoder with Feature Equalizations. In ECCV .  Hongyu Liu Bin Jiang Yibing Song Wei Huang and Chao Yang. 2020 a. Rethinking Image Inpainting via a Mutual Encoder-Decoder with Feature Equalizations. In ECCV ."},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"crossref","unstructured":"Hongyu Liu Bin Jiang Yibing Song Wei Huang and Chao Yang. 2020 b. Rethinking image inpainting via a mutual encoder-decoder with feature equalizations. In ECCV. 725--741.  Hongyu Liu Bin Jiang Yibing Song Wei Huang and Chao Yang. 2020 b. Rethinking image inpainting via a mutual encoder-decoder with feature equalizations. In ECCV. 725--741.","DOI":"10.1007\/978-3-030-58536-5_43"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.425"},{"key":"e_1_3_2_1_28_1","volume-title":"Decoupled weight decay regularization. arXiv preprint arXiv:1711.05101","author":"Loshchilov Ilya","year":"2017","unstructured":"Ilya Loshchilov and Frank Hutter . 2017. Decoupled weight decay regularization. arXiv preprint arXiv:1711.05101 ( 2017 ). Ilya Loshchilov and Frank Hutter. 2017. Decoupled weight decay regularization. arXiv preprint arXiv:1711.05101 (2017)."},{"key":"e_1_3_2_1_29_1","volume-title":"Edgeconnect: Generative image inpainting with adversarial edge learning. arXiv preprint arXiv:1901.00212","author":"Nazeri Kamyar","year":"2019","unstructured":"Kamyar Nazeri , Eric Ng , Tony Joseph , Faisal Z Qureshi , and Mehran Ebrahimi . 2019 . Edgeconnect: Generative image inpainting with adversarial edge learning. arXiv preprint arXiv:1901.00212 (2019). Kamyar Nazeri, Eric Ng, Tony Joseph, Faisal Z Qureshi, and Mehran Ebrahimi. 2019. Edgeconnect: Generative image inpainting with adversarial edge learning. arXiv preprint arXiv:1901.00212 (2019)."},{"key":"e_1_3_2_1_30_1","doi-asserted-by":"crossref","unstructured":"Taesung Park Ming-Yu Liu Ting-Chun Wang and Jun-Yan Zhu. 2019. Semantic image synthesis with spatially-adaptive normalization. In CVPR. 2337--2346.  Taesung Park Ming-Yu Liu Ting-Chun Wang and Jun-Yan Zhu. 2019. Semantic image synthesis with spatially-adaptive normalization. In CVPR. 2337--2346.","DOI":"10.1109\/CVPR.2019.00244"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"crossref","unstructured":"Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR. 2536--2544.  Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR. 2536--2544.","DOI":"10.1109\/CVPR.2016.278"},{"key":"e_1_3_2_1_32_1","volume-title":"Generating Diverse Structure for Image Inpainting With Hierarchical VQ-VAE. arXiv preprint arXiv:2103.10022","author":"Peng Jialun","year":"2021","unstructured":"Jialun Peng , Dong Liu , Songcen Xu , and Houqiang Li. 2021. Generating Diverse Structure for Image Inpainting With Hierarchical VQ-VAE. arXiv preprint arXiv:2103.10022 ( 2021 ). Jialun Peng, Dong Liu, Songcen Xu, and Houqiang Li. 2021. Generating Diverse Structure for Image Inpainting With Hierarchical VQ-VAE. arXiv preprint arXiv:2103.10022 (2021)."},{"key":"e_1_3_2_1_33_1","unstructured":"Alec Radford Jeff Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language Models are Unsupervised Multitask Learners. (2019).  Alec Radford Jeff Wu Rewon Child David Luan Dario Amodei and Ilya Sutskever. 2019. Language Models are Unsupervised Multitask Learners. (2019)."},{"key":"e_1_3_2_1_34_1","volume-title":"Structureflow: Image inpainting via structure-aware appearance flow. In ICCV. 181--190.","author":"Ren Yurui","year":"2019","unstructured":"Yurui Ren , Xiaoming Yu , Ruonan Zhang , Thomas H Li , Shan Liu , and Ge Li . 2019 . Structureflow: Image inpainting via structure-aware appearance flow. In ICCV. 181--190. Yurui Ren, Xiaoming Yu, Ruonan Zhang, Thomas H Li, Shan Liu, and Ge Li. 2019. Structureflow: Image inpainting via structure-aware appearance flow. In ICCV. 181--190."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.241"},{"key":"e_1_3_2_1_36_1","volume-title":"Mpnet: Masked and permuted pre-training for language understanding. arXiv preprint arXiv:2004.09297","author":"Song Kaitao","year":"2020","unstructured":"Kaitao Song , Xu Tan , Tao Qin , Jianfeng Lu , and Tie-Yan Liu . 2020 . Mpnet: Masked and permuted pre-training for language understanding. arXiv preprint arXiv:2004.09297 (2020). Kaitao Song, Xu Tan, Tao Qin, Jianfeng Lu, and Tie-Yan Liu. 2020. Mpnet: Masked and permuted pre-training for language understanding. arXiv preprint arXiv:2004.09297 (2020)."},{"key":"e_1_3_2_1_37_1","volume-title":"Training data-efficient image transformers & distillation through attention. arXiv preprint arXiv:2012.12877","author":"Touvron Hugo","year":"2020","unstructured":"Hugo Touvron , Matthieu Cord , Matthijs Douze , Francisco Massa , Alexandre Sablayrolles , and Herv\u00e9 J\u00e9gou . 2020. Training data-efficient image transformers & distillation through attention. arXiv preprint arXiv:2012.12877 ( 2020 ). Hugo Touvron, Matthieu Cord, Matthijs Douze, Francisco Massa, Alexandre Sablayrolles, and Herv\u00e9 J\u00e9gou. 2020. Training data-efficient image transformers & distillation through attention. arXiv preprint arXiv:2012.12877 (2020)."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.5555\/3295222.3295349"},{"key":"e_1_3_2_1_39_1","volume-title":"High-Fidelity Pluralistic Image Completion with Transformers. arXiv preprint arXiv:2103.14031","author":"Wan Ziyu","year":"2021","unstructured":"Ziyu Wan , Jingbo Zhang , Dongdong Chen , and Jing Liao . 2021. High-Fidelity Pluralistic Image Completion with Transformers. arXiv preprint arXiv:2103.14031 ( 2021 ). Ziyu Wan, Jingbo Zhang, Dongdong Chen, and Jing Liao. 2021. High-Fidelity Pluralistic Image Completion with Transformers. arXiv preprint arXiv:2103.14031 (2021)."},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2003.819861"},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58610-2_46"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00507"},{"key":"e_1_3_2_1_43_1","volume-title":"E2I: Generative Inpainting from Edge to Image","author":"Xu Shunxin","year":"2020","unstructured":"Shunxin Xu , Dong Liu , and Zhiwei Xiong . 2020. E2I: Generative Inpainting from Edge to Image . IEEE Trans. Circuit Syst. Video Technol . ( 2020 ). Shunxin Xu, Dong Liu, and Zhiwei Xiong. 2020. E2I: Generative Inpainting from Edge to Image. IEEE Trans. Circuit Syst. Video Technol. (2020)."},{"key":"e_1_3_2_1_44_1","volume-title":"Xlnet: Generalized autoregressive pretraining for language understanding. arXiv preprint arXiv:1906.08237","author":"Yang Zhilin","year":"2019","unstructured":"Zhilin Yang , Zihang Dai , Yiming Yang , Jaime Carbonell , Ruslan Salakhutdinov , and Quoc V Le . 2019 . Xlnet: Generalized autoregressive pretraining for language understanding. arXiv preprint arXiv:1906.08237 (2019). Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Ruslan Salakhutdinov, and Quoc V Le. 2019. Xlnet: Generalized autoregressive pretraining for language understanding. arXiv preprint arXiv:1906.08237 (2019)."},{"key":"e_1_3_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00577"},{"key":"e_1_3_2_1_46_1","unstructured":"Jiahui Yu Zhe Lin Jimei Yang Xiaohui Shen Xin Lu and Thomas S Huang. 2019. Free-form image inpainting with gated convolution. In ICCV. 4471--4480.  Jiahui Yu Zhe Lin Jimei Yang Xiaohui Shen Xin Lu and Thomas S Huang. 2019. Free-form image inpainting with gated convolution. In ICCV. 4471--4480."},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00216"},{"key":"e_1_3_2_1_48_1","volume-title":"Proceedings of the Asian Conference on Computer Vision .","author":"Zhan Fangneng","year":"2020","unstructured":"Fangneng Zhan , Shijian Lu , Changgong Zhang , Feiying Ma , and Xuansong Xie . 2020 a. Adversarial Image Composition with Auxiliary Illumination . In Proceedings of the Asian Conference on Computer Vision . Fangneng Zhan, Shijian Lu, Changgong Zhang, Feiying Ma, and Xuansong Xie. 2020 a. Adversarial Image Composition with Auxiliary Illumination. In Proceedings of the Asian Conference on Computer Vision ."},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00920"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01478"},{"key":"e_1_3_2_1_51_1","volume-title":"2021 b. GMLight: Lighting Estimation via Geometric Distribution Approximation. arXiv preprint arXiv:2102.10244","author":"Zhan Fangneng","year":"2021","unstructured":"Fangneng Zhan , Yingchen Yu , Rongliang Wu , Changgong Zhang , Shijian Lu , Ling Shao , Feiying Ma , and Xuansong Xie . 2021 b. GMLight: Lighting Estimation via Geometric Distribution Approximation. arXiv preprint arXiv:2102.10244 ( 2021 ). Fangneng Zhan, Yingchen Yu, Rongliang Wu, Changgong Zhang, Shijian Lu, Ling Shao, Feiying Ma, and Xuansong Xie. 2021 b. GMLight: Lighting Estimation via Geometric Distribution Approximation. arXiv preprint arXiv:2102.10244 (2021)."},{"key":"e_1_3_2_1_52_1","volume-title":"Proceedings of the International Conference on Pattern Recognition","author":"Zhan Fangneng","year":"2020","unstructured":"Fangneng Zhan and Changgong Zhang . 2020 . Spatial-Aware GAN for Unsupervised Person Re-identification . Proceedings of the International Conference on Pattern Recognition (2020). Fangneng Zhan and Changgong Zhang. 2020. Spatial-Aware GAN for Unsupervised Person Re-identification. Proceedings of the International Conference on Pattern Recognition (2020)."},{"key":"e_1_3_2_1_53_1","volume-title":"2020 b. EMLight: Lighting Estimation via Spherical Distribution Approximation. AAAI","author":"Zhan Fangneng","year":"2020","unstructured":"Fangneng Zhan , Changgong Zhang , Yingchen Yu , Yuan Chang , Shijian Lu , Feiying Ma , and Xuansong Xie . 2020 b. EMLight: Lighting Estimation via Spherical Distribution Approximation. AAAI ( 2020 ). Fangneng Zhan, Changgong Zhang, Yingchen Yu, Yuan Chang, Shijian Lu, Feiying Ma, and Xuansong Xie. 2020 b. EMLight: Lighting Estimation via Spherical Distribution Approximation. AAAI (2020)."},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00377"},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV48630.2021.00257"},{"key":"e_1_3_2_1_56_1","doi-asserted-by":"crossref","unstructured":"Richard Zhang Phillip Isola Alexei A Efros Eli Shechtman and Oliver Wang. 2018. The unreasonable effectiveness of deep features as a perceptual metric. In CVPR. 586--595.  Richard Zhang Phillip Isola Alexei A Efros Eli Shechtman and Oliver Wang. 2018. The unreasonable effectiveness of deep features as a perceptual metric. In CVPR. 586--595.","DOI":"10.1109\/CVPR.2018.00068"},{"key":"e_1_3_2_1_57_1","volume-title":"UCTGAN: Diverse Image Inpainting Based on Unsupervised Cross-Space Translation. In CVPR. 5741--5750.","author":"Zhao Lei","year":"2020","unstructured":"Lei Zhao , Qihang Mo , Sihuan Lin , Zhizhong Wang , Zhiwen Zuo , Haibo Chen , Wei Xing , and Dongming Lu . 2020 . UCTGAN: Diverse Image Inpainting Based on Unsupervised Cross-Space Translation. In CVPR. 5741--5750. Lei Zhao, Qihang Mo, Sihuan Lin, Zhizhong Wang, Zhiwen Zuo, Haibo Chen, Wei Xing, and Dongming Lu. 2020. UCTGAN: Diverse Image Inpainting Based on Unsupervised Cross-Space Translation. In CVPR. 5741--5750."},{"key":"e_1_3_2_1_58_1","volume-title":"International Conference on Learning Representations (ICLR) .","author":"Zhao Shengyu","year":"2021","unstructured":"Shengyu Zhao , Jonathan Cui , Yilun Sheng , Yue Dong , Xiao Liang , Eric I Chang , and Yan Xu . 2021 . Large Scale Image Completion via Co-Modulated Generative Adversarial Networks . In International Conference on Learning Representations (ICLR) . Shengyu Zhao, Jonathan Cui, Yilun Sheng, Yue Dong, Xiao Liang, Eric I Chang, and Yan Xu. 2021. Large Scale Image Completion via Co-Modulated Generative Adversarial Networks. In International Conference on Learning Representations (ICLR) ."},{"key":"e_1_3_2_1_59_1","volume-title":"Towards deeper understanding of variational autoencoding models. arXiv preprint arXiv:1702.08658","author":"Zhao Shengjia","year":"2017","unstructured":"Shengjia Zhao , Jiaming Song , and Stefano Ermon . 2017. Towards deeper understanding of variational autoencoding models. arXiv preprint arXiv:1702.08658 ( 2017 ). Shengjia Zhao, Jiaming Song, and Stefano Ermon. 2017. Towards deeper understanding of variational autoencoding models. arXiv preprint arXiv:1702.08658 (2017)."},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"crossref","unstructured":"Chuanxia Zheng Tat-Jen Cham and Jianfei Cai. 2019. Pluralistic image completion. In CVPR. 1438--1447.  Chuanxia Zheng Tat-Jen Cham and Jianfei Cai. 2019. Pluralistic image completion. In CVPR. 1438--1447.","DOI":"10.1109\/CVPR.2019.00153"},{"key":"e_1_3_2_1_61_1","volume-title":"Places: A 10 million Image Database for Scene Recognition","author":"Zhou Bolei","year":"2017","unstructured":"Bolei Zhou , Agata Lapedriza , Aditya Khosla , Aude Oliva , and Antonio Torralba . 2017 . Places: A 10 million Image Database for Scene Recognition . IEEE Transactions on Pattern Analysis and Machine Intelligence ( 2017). Bolei Zhou, Agata Lapedriza, Aditya Khosla, Aude Oliva, and Antonio Torralba. 2017. Places: A 10 million Image Database for Scene Recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence (2017)."}],"event":{"name":"MM '21: ACM Multimedia Conference","location":"Virtual Event China","acronym":"MM '21","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 29th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3474085.3475436","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3474085.3475436","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:48:33Z","timestamp":1750193313000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3474085.3475436"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,10,17]]},"references-count":61,"alternative-id":["10.1145\/3474085.3475436","10.1145\/3474085"],"URL":"https:\/\/doi.org\/10.1145\/3474085.3475436","relation":{},"subject":[],"published":{"date-parts":[[2021,10,17]]},"assertion":[{"value":"2021-10-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}