{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,7]],"date-time":"2026-07-07T15:57:46Z","timestamp":1783439866349,"version":"3.54.6"},"publisher-location":"New York, NY, USA","reference-count":54,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100012659","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61876177"],"award-info":[{"award-number":["61876177"]}],"id":[{"id":"10.13039\/501100012659","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Beijing Natural Science Foundation","award":["4202034"],"award-info":[{"award-number":["4202034"]}]},{"name":"Fundamental Research Funds for the Central Universities and Zhejiang Lab","award":["2019KD0AB04"],"award-info":[{"award-number":["2019KD0AB04"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413854","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T13:12:00Z","timestamp":1602508320000},"page":"165-173","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["InteractGAN: Learning to Generate Human-Object Interaction"],"prefix":"10.1145","author":[{"given":"Chen","family":"Gao","sequence":"first","affiliation":[{"name":"Institute of Information Engineering, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Si","family":"Liu","sequence":"additional","affiliation":[{"name":"Beihang University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Defa","family":"Zhu","sequence":"additional","affiliation":[{"name":"Institute of Information Engineering, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Quan","family":"Liu","sequence":"additional","affiliation":[{"name":"Beihang University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jie","family":"Cao","sequence":"additional","affiliation":[{"name":"Institute of Automation, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haoqian","family":"He","sequence":"additional","affiliation":[{"name":"Beihang University, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ran","family":"He","sequence":"additional","affiliation":[{"name":"Institute of Automation, Chinese Academy of Sciences, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shuicheng","family":"Yan","sequence":"additional","affiliation":[{"name":"Yitu Technology, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"crossref","unstructured":"Oron Ashual and Lior Wolf. 2019. Specifying object attributes and relations in interactive scene generation. In ICCV.  Oron Ashual and Lior Wolf. 2019. Specifying object attributes and relations in interactive scene generation. In ICCV.","DOI":"10.1109\/ICCV.2019.00466"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"crossref","unstructured":"Navaneeth Bodla Gang Hua and Rama Chellappa. 2018. Semi-supervised fusedgan for conditional image generation. In ECCV.  Navaneeth Bodla Gang Hua and Rama Chellappa. 2018. Semi-supervised fusedgan for conditional image generation. In ECCV.","DOI":"10.1007\/978-3-030-01228-1_41"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"crossref","unstructured":"Zhe Cao Tomas Simon Shih-En Wei and Yaser Sheikh. 2017. Realtime Multi-person 2D Pose Estimation Using Part Affinity Fields. In CVPR.  Zhe Cao Tomas Simon Shih-En Wei and Yaser Sheikh. 2017. Realtime Multi-person 2D Pose Estimation Using Part Affinity Fields. In CVPR.","DOI":"10.1109\/CVPR.2017.143"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"crossref","unstructured":"Yu-Wei Chao Yunfan Liu Xieyang Liu Huayi Zeng and Jia Deng. 2018. Learning to Detect Human-Object Interactions. In WACV.  Yu-Wei Chao Yunfan Liu Xieyang Liu Huayi Zeng and Jia Deng. 2018. Learning to Detect Human-Object Interactions. In WACV.","DOI":"10.1109\/WACV.2018.00048"},{"key":"e_1_3_2_2_5_1","volume-title":"Hico: A benchmark for recognizing human-object interactions in images. In ICCV.","author":"Chao Yu-Wei","year":"2015","unstructured":"Yu-Wei Chao , Zhan Wang , Yugeng He , Jiaxuan Wang , and Jia Deng . 2015 . Hico: A benchmark for recognizing human-object interactions in images. In ICCV. Yu-Wei Chao, Zhan Wang, Yugeng He, Jiaxuan Wang, and Jia Deng. 2015. Hico: A benchmark for recognizing human-object interactions in images. In ICCV."},{"key":"e_1_3_2_2_6_1","volume-title":"Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR.","author":"Choi Yunjey","year":"2018","unstructured":"Yunjey Choi , Minje Choi , Munyoung Kim , Jung-Woo Ha , Sunghun Kim , and Jaegul Choo . 2018 . Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR. Yunjey Choi, Minje Choi, Munyoung Kim, Jung-Woo Ha, Sunghun Kim, and Jaegul Choo. 2018. Stargan: Unified generative adversarial networks for multi-domain image-to-image translation. In CVPR."},{"key":"e_1_3_2_2_7_1","volume-title":"CVPR-Workshops","author":"Desai Chaitanya","unstructured":"Chaitanya Desai , Deva Ramanan , and Charless Fowlkes . 2010. Discriminative models for static human-object interactions . In CVPR-Workshops . IEEE. Chaitanya Desai, Deva Ramanan, and Charless Fowlkes. 2010. Discriminative models for static human-object interactions. In CVPR-Workshops. IEEE."},{"key":"e_1_3_2_2_8_1","volume-title":"Nice: Non-linear independent components estimation. arXiv preprint arXiv:1410.8516","author":"Dinh Laurent","year":"2014","unstructured":"Laurent Dinh , David Krueger , and Yoshua Bengio . 2014 . Nice: Non-linear independent components estimation. arXiv preprint arXiv:1410.8516 (2014). Laurent Dinh, David Krueger, and Yoshua Bengio. 2014. Nice: Non-linear independent components estimation. arXiv preprint arXiv:1410.8516 (2014)."},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"crossref","unstructured":"Chen Gao Yunpeng Chen Si Liu Zhenxiong Tan and Shuicheng Yan. 2020 a. AdversarialNAS: Adversarial Neural Architecture Search for GANs. In CVPR.  Chen Gao Yunpeng Chen Si Liu Zhenxiong Tan and Shuicheng Yan. 2020 a. AdversarialNAS: Adversarial Neural Architecture Search for GANs. In CVPR.","DOI":"10.1109\/CVPR42600.2020.00572"},{"key":"e_1_3_2_2_10_1","volume-title":"2020 b. Recapture as You Want. arXiv preprint arXiv:2006.01435","author":"Gao Chen","year":"2020","unstructured":"Chen Gao , Si Liu , Ran He , Shuicheng Yan , and Bo Li . 2020 b. Recapture as You Want. arXiv preprint arXiv:2006.01435 ( 2020 ). Chen Gao, Si Liu, Ran He, Shuicheng Yan, and Bo Li. 2020 b. Recapture as You Want. arXiv preprint arXiv:2006.01435 (2020)."},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"crossref","unstructured":"Georgia Gkioxari Ross Girshick and Jitendra Malik. 2015. Actions and attributes from wholes and parts. In ICCV.  Georgia Gkioxari Ross Girshick and Jitendra Malik. 2015. Actions and attributes from wholes and parts. In ICCV.","DOI":"10.1109\/ICCV.2015.284"},{"key":"e_1_3_2_2_12_1","unstructured":"Ian J. Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron C. Courville and Yoshua Bengio. 2014. Generative Adversarial Nets. In NIPS.  Ian J. Goodfellow Jean Pouget-Abadie Mehdi Mirza Bing Xu David Warde-Farley Sherjil Ozair Aaron C. Courville and Yoshua Bengio. 2014. Generative Adversarial Nets. In NIPS."},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2009.83"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"crossref","unstructured":"Ran He Bao-Gang Hu and Xiao-Tong Yuan. 2009. Robust discriminant analysis based on nonparametric maximum entropy. In ACML.  Ran He Bao-Gang Hu and Xiao-Tong Yuan. 2009. Robust discriminant analysis based on nonparametric maximum entropy. In ACML.","DOI":"10.1007\/978-3-642-05224-8_11"},{"key":"e_1_3_2_2_15_1","unstructured":"Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium. In NIPS.  Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium. In NIPS."},{"key":"e_1_3_2_2_16_1","volume-title":"Reducing the dimensionality of data with neural networks. science","author":"Hinton Geoffrey E","year":"2006","unstructured":"Geoffrey E Hinton and Ruslan R Salakhutdinov . 2006. Reducing the dimensionality of data with neural networks. science , Vol. 313 , 5786 ( 2006 ), 504--507. Geoffrey E Hinton and Ruslan R Salakhutdinov. 2006. Reducing the dimensionality of data with neural networks. science, Vol. 313, 5786 (2006), 504--507."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"crossref","unstructured":"Seunghoon Hong Dingdong Yang Jongwook Choi and Honglak Lee. 2018. Inferring semantic layout for hierarchical text-to-image synthesis. In CVPR.  Seunghoon Hong Dingdong Yang Jongwook Choi and Honglak Lee. 2018. Inferring semantic layout for hierarchical text-to-image synthesis. In CVPR.","DOI":"10.1109\/CVPR.2018.00833"},{"key":"e_1_3_2_2_18_1","unstructured":"Jian-Fang Hu Wei-Shi Zheng Jianhuang Lai Shaogang Gong and Tao Xiang. 2013. Recognising human-object interaction via exemplar based modelling. In ICCV.  Jian-Fang Hu Wei-Shi Zheng Jianhuang Lai Shaogang Gong and Tao Xiang. 2013. Recognising human-object interaction via exemplar based modelling. In ICCV."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"crossref","unstructured":"Xun Huang Yixuan Li Omid Poursaeed John Hopcroft and Serge Belongie. 2017. Stacked generative adversarial networks. In CVPR.  Xun Huang Yixuan Li Omid Poursaeed John Hopcroft and Serge Belongie. 2017. Stacked generative adversarial networks. In CVPR.","DOI":"10.1109\/CVPR.2017.202"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"crossref","unstructured":"Nazli Ikizler R Gokberk Cinbis Selen Pehlivan and Pinar Duygulu. 2008. Recognizing actions from still images. In ICPR.  Nazli Ikizler R Gokberk Cinbis Selen Pehlivan and Pinar Duygulu. 2008. Recognizing actions from still images. In ICPR.","DOI":"10.1109\/ICPR.2008.4761663"},{"key":"e_1_3_2_2_21_1","volume-title":"Efros","author":"Isola Phillip","year":"2017","unstructured":"Phillip Isola , Jun-Yan Zhu , Tinghui Zhou , and Alexei A . Efros . 2017 a. Image-to-Image Translation with Conditional Adversarial Networks. In CVPR. Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros. 2017a. Image-to-Image Translation with Conditional Adversarial Networks. In CVPR."},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"crossref","unstructured":"Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017b. Image-to-image translation with conditional adversarial networks. In CVPR.  Phillip Isola Jun-Yan Zhu Tinghui Zhou and Alexei A Efros. 2017b. Image-to-image translation with conditional adversarial networks. In CVPR.","DOI":"10.1109\/CVPR.2017.632"},{"key":"e_1_3_2_2_23_1","unstructured":"Max Jaderberg Karen Simonyan Andrew Zisserman and Koray Kavukcuoglu. 2015. Spatial Transformer Networks. In NIPS.  Max Jaderberg Karen Simonyan Andrew Zisserman and Koray Kavukcuoglu. 2015. Spatial Transformer Networks. In NIPS."},{"key":"e_1_3_2_2_24_1","volume-title":"PSGAN: Pose and Expression Robust Spatial-Aware GAN for Customizable Makeup Transfer. In CVPR.","author":"Jiang Wentao","year":"2020","unstructured":"Wentao Jiang , Si Liu , Chen Gao , Jie Cao , Ran He , Jiashi Feng , and Shuicheng Yan . 2020 . PSGAN: Pose and Expression Robust Spatial-Aware GAN for Customizable Makeup Transfer. In CVPR. Wentao Jiang, Si Liu, Chen Gao, Jie Cao, Ran He, Jiashi Feng, and Shuicheng Yan. 2020. PSGAN: Pose and Expression Robust Spatial-Aware GAN for Customizable Makeup Transfer. In CVPR."},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"crossref","unstructured":"Justin Johnson Agrim Gupta and Li Fei-Fei. 2018. Image generation from scene graphs. In CVPR.  Justin Johnson Agrim Gupta and Li Fei-Fei. 2018. Image generation from scene graphs. In CVPR.","DOI":"10.1109\/CVPR.2018.00133"},{"key":"e_1_3_2_2_26_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2015","unstructured":"Diederik P. Kingma and Jimmy Ba . 2015 . Adam : A Method for Stochastic Optimization. CoRR , Vol. abs\/ 1412 .6980 (2015). Diederik P. Kingma and Jimmy Ba. 2015. Adam: A Method for Stochastic Optimization. CoRR, Vol. abs\/1412.6980 (2015)."},{"key":"e_1_3_2_2_27_1","volume-title":"Glow: Generative flow with invertible 1x1 convolutions. In NIPS.","author":"Kingma Durk P","year":"2018","unstructured":"Durk P Kingma and Prafulla Dhariwal . 2018 . Glow: Generative flow with invertible 1x1 convolutions. In NIPS. Durk P Kingma and Prafulla Dhariwal. 2018. Glow: Generative flow with invertible 1x1 convolutions. In NIPS."},{"key":"e_1_3_2_2_28_1","volume-title":"Kingma and Max Welling","author":"Diederik","year":"2013","unstructured":"Diederik P. Kingma and Max Welling . 2013 . Auto-Encoding Variational Bayes. CoRR , Vol. abs\/ 1312 .6114 (2013). Diederik P. Kingma and Max Welling. 2013. Auto-Encoding Variational Bayes. CoRR, Vol. abs\/1312.6114 (2013)."},{"key":"e_1_3_2_2_29_1","volume-title":"Xing","author":"Liang Xiaodan","year":"2017","unstructured":"Xiaodan Liang , Lisa Lee , and Eric P . Xing . 2017 . Deep Variation-Structured Reinforcement Learning for Visual Relationship and Attribute Detection. In CVPR. Xiaodan Liang, Lisa Lee, and Eric P. Xing. 2017. Deep Variation-Structured Reinforcement Learning for Visual Relationship and Attribute Detection. In CVPR."},{"key":"e_1_3_2_2_30_1","volume-title":"GPS: Group People Segmentation with Detailed Part Inference. In ICME.","author":"Liao Yue","year":"2019","unstructured":"Yue Liao , Si Liu , Tianrui Hui , Chen Gao , Yao Sun , Hefei Ling , and Bo Li . 2019 . GPS: Group People Segmentation with Detailed Part Inference. In ICME. Yue Liao, Si Liu, Tianrui Hui, Chen Gao, Yao Sun, Hefei Ling, and Bo Li. 2019. GPS: Group People Segmentation with Detailed Part Inference. In ICME."},{"key":"e_1_3_2_2_31_1","volume-title":"Ppdm: Parallel point detection and matching for real-time human-object interaction detection. In CVPR.","author":"Liao Yue","year":"2020","unstructured":"Yue Liao , Si Liu , Fei Wang , Yanjie Chen , Chen Qian , and Jiashi Feng . 2020 . Ppdm: Parallel point detection and matching for real-time human-object interaction detection. In CVPR. Yue Liao, Si Liu, Fei Wang, Yanjie Chen, Chen Qian, and Jiashi Feng. 2020. Ppdm: Parallel point detection and matching for real-time human-object interaction detection. In CVPR."},{"key":"e_1_3_2_2_32_1","unstructured":"Ming-Yu Liu Thomas Breuel and Jan Kautz. 2017. Unsupervised Image-to-Image Translation Networks. In NIPS.  Ming-Yu Liu Thomas Breuel and Jan Kautz. 2017. Unsupervised Image-to-Image Translation Networks. In NIPS."},{"key":"e_1_3_2_2_33_1","unstructured":"Cewu Lu Ranjay Krishna Michael S. Bernstein and Li Fei-Fei. 2016. Visual Relationship Detection with Language Priors. In ECCV.  Cewu Lu Ranjay Krishna Michael S. Bernstein and Li Fei-Fei. 2016. Visual Relationship Detection with Language Priors. In ECCV."},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"crossref","unstructured":"Arun Mallya and Svetlana Lazebnik. 2016. Learning models for actions and person-object interactions with transfer to question answering. In ECCV.  Arun Mallya and Svetlana Lazebnik. 2016. Learning models for actions and person-object interactions with transfer to question answering. In ECCV.","DOI":"10.1007\/978-3-319-46448-0_25"},{"key":"e_1_3_2_2_35_1","volume-title":"Spectral Normalization for Generative Adversarial Networks. CoRR","author":"Miyato Takeru","year":"2018","unstructured":"Takeru Miyato , Toshiki Kataoka , Masanori Koyama , and Yuichi Yoshida . 2018. Spectral Normalization for Generative Adversarial Networks. CoRR , Vol. abs\/ 1802 .05957 ( 2018 ). Takeru Miyato, Toshiki Kataoka, Masanori Koyama, and Yuichi Yoshida. 2018. Spectral Normalization for Generative Adversarial Networks. CoRR, Vol. abs\/1802.05957 (2018)."},{"key":"e_1_3_2_2_36_1","volume-title":"cGANs with projection discriminator. arXiv preprint arXiv:1802.05637","author":"Miyato Takeru","year":"2018","unstructured":"Takeru Miyato and Masanori Koyama . 2018. cGANs with projection discriminator. arXiv preprint arXiv:1802.05637 ( 2018 ). Takeru Miyato and Masanori Koyama. 2018. cGANs with projection discriminator. arXiv preprint arXiv:1802.05637 (2018)."},{"key":"e_1_3_2_2_37_1","volume-title":"Zero-Shot Generation of Human-Object Interaction Videos. arXiv preprint arXiv:1912.02401","author":"Nawhal Megha","year":"2019","unstructured":"Megha Nawhal , Mengyao Zhai , Andreas Lehrmann , and Leonid Sigal . 2019. Zero-Shot Generation of Human-Object Interaction Videos. arXiv preprint arXiv:1912.02401 ( 2019 ). Megha Nawhal, Mengyao Zhai, Andreas Lehrmann, and Leonid Sigal. 2019. Zero-Shot Generation of Human-Object Interaction Videos. arXiv preprint arXiv:1912.02401 (2019)."},{"key":"e_1_3_2_2_38_1","unstructured":"Augustus Odena Christopher Olah and Jonathon Shlens. 2017. Conditional Image Synthesis With Auxiliary Classifier GANs. In ICML.  Augustus Odena Christopher Olah and Jonathon Shlens. 2017. Conditional Image Synthesis With Auxiliary Classifier GANs. In ICML."},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"crossref","unstructured":"Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR.  Deepak Pathak Philipp Krahenbuhl Jeff Donahue Trevor Darrell and Alexei A Efros. 2016. Context encoders: Feature learning by inpainting. In CVPR.","DOI":"10.1109\/CVPR.2016.278"},{"key":"e_1_3_2_2_40_1","volume-title":"TPAMI","volume":"34","author":"Prest Alessandro","year":"2011","unstructured":"Alessandro Prest , Cordelia Schmid , and Vittorio Ferrari . 2011 . Weakly supervised learning of interactions between humans and objects . TPAMI , Vol. 34 , 3 (2011). Alessandro Prest, Cordelia Schmid, and Vittorio Ferrari. 2011. Weakly supervised learning of interactions between humans and objects. TPAMI, Vol. 34, 3 (2011)."},{"key":"e_1_3_2_2_41_1","volume-title":"Scene Graph Generation With Hierarchical Context. TNNLS","author":"Ren Guanghui","year":"2020","unstructured":"Guanghui Ren , Lejian Ren , Yue Liao , Si Liu , Bo Li , Jizhong Han , and Shuicheng Yan . 2020. Scene Graph Generation With Hierarchical Context. TNNLS ( 2020 ). Guanghui Ren, Lejian Ren, Yue Liao, Si Liu, Bo Li, Jizhong Han, and Shuicheng Yan. 2020. Scene Graph Generation With Hierarchical Context. TNNLS (2020)."},{"key":"e_1_3_2_2_42_1","unstructured":"Tim Salimans Ian J. Goodfellow Wojciech Zaremba Vicki Cheung Alec Radford and Xi Chen. 2016. Improved Techniques for Training GANs. In NIPS.  Tim Salimans Ian J. Goodfellow Wojciech Zaremba Vicki Cheung Alec Radford and Xi Chen. 2016. Improved Techniques for Training GANs. In NIPS."},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"crossref","unstructured":"Aliaksandr Siarohin Enver Sangineto St\u00e9phane Lathuili\u00e8re and Nicu Sebe. 2018. Deformable GANs for Pose-Based Human Image Generation. In CVPR.  Aliaksandr Siarohin Enver Sangineto St\u00e9phane Lathuili\u00e8re and Nicu Sebe. 2018. Deformable GANs for Pose-Based Human Image Generation. In CVPR.","DOI":"10.1109\/CVPR.2018.00359"},{"key":"e_1_3_2_2_44_1","unstructured":"Kihyuk Sohn Honglak Lee and Xinchen Yan. 2015. Learning structured output representation using deep conditional generative models. In NIPS.  Kihyuk Sohn Honglak Lee and Xinchen Yan. 2015. Learning structured output representation using deep conditional generative models. In NIPS."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"crossref","unstructured":"Christian Thurau and V\u00e1clav Hlav\u00e1c. 2008. Pose primitive based human action recognition in videos or still images. In CVPR.  Christian Thurau and V\u00e1clav Hlav\u00e1c. 2008. Pose primitive based human action recognition in videos or still images. In CVPR.","DOI":"10.1109\/CVPR.2008.4587721"},{"key":"e_1_3_2_2_46_1","volume-title":"Computer Graphics Forum","author":"Wang He","unstructured":"He Wang , S\u00f6ren Pirk , Ersin Yumer , Vladimir G Kim , Ozan Sener , Srinath Sridhar , and Leonidas J Guibas . 2019. Learning a Generative Model for Multi-Step Human-Object Interactions from Videos . In Computer Graphics Forum , Vol. 38 . Wiley Online Library , 367--378. He Wang, S\u00f6ren Pirk, Ersin Yumer, Vladimir G Kim, Ozan Sener, Srinath Sridhar, and Leonidas J Guibas. 2019. Learning a Generative Model for Multi-Step Human-Object Interactions from Videos. In Computer Graphics Forum, Vol. 38. Wiley Online Library, 367--378."},{"key":"e_1_3_2_2_47_1","unstructured":"Ting-Chun Wang Ming-Yu Liu Jun-Yan Zhu Andrew Tao Jan Kautz and Bryan Catanzaro. 2018. High-resolution image synthesis and semantic manipulation with conditional gans. In CVPR.  Ting-Chun Wang Ming-Yu Liu Jun-Yan Zhu Andrew Tao Jan Kautz and Bryan Catanzaro. 2018. High-resolution image synthesis and semantic manipulation with conditional gans. In CVPR."},{"key":"e_1_3_2_2_48_1","volume-title":"Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR.","author":"Xu Tao","year":"2018","unstructured":"Tao Xu , Pengchuan Zhang , Qiuyuan Huang , Han Zhang , Zhe Gan , Xiaolei Huang , and Xiaodong He . 2018 . Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR. Tao Xu, Pengchuan Zhang, Qiuyuan Huang, Han Zhang, Zhe Gan, Xiaolei Huang, and Xiaodong He. 2018. Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In CVPR."},{"key":"e_1_3_2_2_49_1","volume-title":"Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In ICCV.","author":"Zhang Han","year":"2017","unstructured":"Han Zhang , Tao Xu , Hongsheng Li , Shaoting Zhang , Xiaogang Wang , Xiaolei Huang , and Dimitris N Metaxas . 2017 . Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In ICCV. Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris N Metaxas. 2017. Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In ICCV."},{"key":"e_1_3_2_2_50_1","volume-title":"Image Generation from Layout. CoRR","author":"Zhao Bo","year":"2018","unstructured":"Bo Zhao , Lili Meng , Weiping Yin , and Leonid Sigal . 2018. Image Generation from Layout. CoRR , Vol. abs\/ 1811 .11389 ( 2018 ). Bo Zhao, Lili Meng, Weiping Yin, and Leonid Sigal. 2018. Image Generation from Layout. CoRR, Vol. abs\/1811.11389 (2018)."},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"crossref","unstructured":"Liang Zheng Liyue Shen Lu Tian Shengjin Wang Jingdong Wang and Qi Tian. 2015. Scalable Person Re-identification: A Benchmark. In ICCV.  Liang Zheng Liyue Shen Lu Tian Shengjin Wang Jingdong Wang and Qi Tian. 2015. Scalable Person Re-identification: A Benchmark. In ICCV.","DOI":"10.1109\/ICCV.2015.133"},{"key":"e_1_3_2_2_52_1","volume-title":"Yi Yang, and Qi Tian.","author":"Zheng Liang","year":"2017","unstructured":"Liang Zheng , Hengheng Zhang , Shaoyan Sun , Manmohan Krishna Chandraker , Yi Yang, and Qi Tian. 2017 . Person Re-identification in the Wild. In CVPR. Liang Zheng, Hengheng Zhang, Shaoyan Sun, Manmohan Krishna Chandraker, Yi Yang, and Qi Tian. 2017. Person Re-identification in the Wild. In CVPR."},{"key":"e_1_3_2_2_53_1","volume-title":"UGAN: Untraceable GAN for Multi-Domain Face Translation. arXiv preprint arXiv:1907.11418","author":"Zhu Defa","year":"2019","unstructured":"Defa Zhu , Si Liu , Wentao Jiang , Chen Gao , Tianyi Wu , and Guodong Guo . 2019 . UGAN: Untraceable GAN for Multi-Domain Face Translation. arXiv preprint arXiv:1907.11418 (2019). Defa Zhu, Si Liu, Wentao Jiang, Chen Gao, Tianyi Wu, and Guodong Guo. 2019. UGAN: Untraceable GAN for Multi-Domain Face Translation. arXiv preprint arXiv:1907.11418 (2019)."},{"key":"e_1_3_2_2_54_1","volume-title":"Efros","author":"Zhu Jun-Yan","year":"2017","unstructured":"Jun-Yan Zhu , Taesung Park , Phillip Isola , and Alexei A . Efros . 2017 . Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In ICCV. Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros. 2017. Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In ICCV."}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","location":"Seattle WA USA","acronym":"MM '20","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413854","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413854","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T22:01:18Z","timestamp":1750197678000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413854"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":54,"alternative-id":["10.1145\/3394171.3413854","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413854","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}