{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,20]],"date-time":"2026-06-20T16:21:09Z","timestamp":1781972469650,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":53,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["U19A2057, 61876223"],"award-info":[{"award-number":["U19A2057, 61876223"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003999","name":"Science Fund for Creative Research Groups","doi-asserted-by":"publisher","award":["62121002"],"award-info":[{"award-number":["62121002"]}],"id":[{"id":"10.13039\/501100003999","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"publisher","award":["WK3480000008, WK3480000010"],"award-info":[{"award-number":["WK3480000008, WK3480000010"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547881","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:42:35Z","timestamp":1665416555000},"page":"4345-4354","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":25,"title":["DSE-GAN: Dynamic Semantic Evolution Generative Adversarial Network for Text-to-Image Generation"],"prefix":"10.1145","author":[{"given":"Mengqi","family":"Huang","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhendong","family":"Mao","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Penghui","family":"Wang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Quan","family":"Wang","sequence":"additional","affiliation":[{"name":"Beijing University of Posts and Telecommunications, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yongdong","family":"Zhang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Institute of Artificial Intelligence, Hefei Comprehensive National Science Center, Hefei, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297","author":"Bengio Emmanuel","year":"2015","unstructured":"Emmanuel Bengio , Pierre-Luc Bacon , Joelle Pineau , and Doina Precup . 2015. Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297 ( 2015 ). Emmanuel Bengio, Pierre-Luc Bacon, Joelle Pineau, and Doina Precup. 2015. Conditional computation in neural networks for faster models. arXiv preprint arXiv:1511.06297 (2015)."},{"key":"e_1_3_2_2_2_1","volume-title":"International Conference on Machine Learning. PMLR, 527--536","author":"Bolukbasi Tolga","year":"2017","unstructured":"Tolga Bolukbasi , Joseph Wang , Ofer Dekel , and Venkatesh Saligrama . 2017 . Adaptive neural networks for efficient inference . In International Conference on Machine Learning. PMLR, 527--536 . Tolga Bolukbasi, Joseph Wang, Ofer Dekel, and Venkatesh Saligrama. 2017. Adaptive neural networks for efficient inference. In International Conference on Machine Learning. PMLR, 527--536."},{"key":"e_1_3_2_2_3_1","volume-title":"Asian Conference on Computer Vision. Springer, 100--116","author":"Chen Kevin","year":"2018","unstructured":"Kevin Chen , Christopher B Choy , Manolis Savva , Angel X Chang , Thomas Funkhouser , and Silvio Savarese . 2018 . Text2shape: Generating shapes from natural language by learning joint embeddings . In Asian Conference on Computer Vision. Springer, 100--116 . Kevin Chen, Christopher B Choy, Manolis Savva, Angel X Chang, Thomas Funkhouser, and Silvio Savarese. 2018. Text2shape: Generating shapes from natural language by learning joint embeddings. In Asian Conference on Computer Vision. Springer, 100--116."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01092"},{"key":"e_1_3_2_2_5_1","unstructured":"Ming Ding Zhuoyi Yang Wenyi Hong Wendi Zheng Chang Zhou Da Yin Junyang Lin Xu Zou Zhou Shao Hongxia Yang etal 2021. CogView: Mastering Text-to-Image Generation via Transformers. arXiv preprint arXiv:2105.13290 (2021).  Ming Ding Zhuoyi Yang Wenyi Hong Wendi Zheng Chang Zhou Da Yin Junyang Lin Xu Zou Zhou Shao Hongxia Yang et al. 2021. CogView: Mastering Text-to-Image Generation via Transformers. arXiv preprint arXiv:2105.13290 (2021)."},{"key":"e_1_3_2_2_6_1","volume-title":"Generative adversarial nets. Advances in neural information processing systems","author":"Goodfellow Ian","year":"2014","unstructured":"Ian Goodfellow , Jean Pouget-Abadie , Mehdi Mirza , Bing Xu , David Warde-Farley , Sherjil Ozair , Aaron Courville , and Yoshua Bengio . 2014. Generative adversarial nets. Advances in neural information processing systems , Vol. 27 ( 2014 ). Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative adversarial nets. Advances in neural information processing systems , Vol. 27 (2014)."},{"key":"e_1_3_2_2_7_1","volume-title":"Dynamic neural networks: A survey","author":"Han Yizeng","year":"2021","unstructured":"Yizeng Han , Gao Huang , Shiji Song , Le Yang , Honghui Wang , and Yulin Wang . 2021. Dynamic neural networks: A survey . IEEE Transactions on Pattern Analysis and Machine Intelligence ( 2021 ). Yizeng Han, Gao Huang, Shiji Song, Le Yang, Honghui Wang, and Yulin Wang. 2021. Dynamic neural networks: A survey. IEEE Transactions on Pattern Analysis and Machine Intelligence (2021)."},{"key":"e_1_3_2_2_8_1","volume-title":"Gans trained by a two time-scale update rule converge to a local nash equilibrium. Advances in neural information processing systems","author":"Heusel Martin","year":"2017","unstructured":"Martin Heusel , Hubert Ramsauer , Thomas Unterthiner , Bernhard Nessler , and Sepp Hochreiter . 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. Advances in neural information processing systems , Vol. 30 ( 2017 ). Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. Advances in neural information processing systems , Vol. 30 (2017)."},{"key":"e_1_3_2_2_9_1","volume-title":"Generating multiple objects at spatially distinct locations. arXiv preprint arXiv:1901.00686","author":"Hinz Tobias","year":"2019","unstructured":"Tobias Hinz , Stefan Heinrich , and Stefan Wermter . 2019a. Generating multiple objects at spatially distinct locations. arXiv preprint arXiv:1901.00686 ( 2019 ). Tobias Hinz, Stefan Heinrich, and Stefan Wermter. 2019a. Generating multiple objects at spatially distinct locations. arXiv preprint arXiv:1901.00686 (2019)."},{"key":"e_1_3_2_2_10_1","volume-title":"Semantic Object Accuracy for Generative Text-to-Image Synthesis","author":"Hinz Tobias","year":"2019","unstructured":"Tobias Hinz , Stefan Heinrich , and Stefan Wermter . 2019b. Semantic Object Accuracy for Generative Text-to-Image Synthesis . IEEE transactions on pattern analysis and machine intelligence ( 2019 ). Tobias Hinz, Stefan Heinrich, and Stefan Wermter. 2019b. Semantic Object Accuracy for Generative Text-to-Image Synthesis. IEEE transactions on pattern analysis and machine intelligence (2019)."},{"key":"e_1_3_2_2_11_1","volume-title":"Semantic object accuracy for generative text-to-image synthesis. arXiv preprint arXiv:1910.13321","author":"Hinz Tobias","year":"2019","unstructured":"Tobias Hinz , Stefan Heinrich , and Stefan Wermter . 2019c. Semantic object accuracy for generative text-to-image synthesis. arXiv preprint arXiv:1910.13321 ( 2019 ). Tobias Hinz, Stefan Heinrich, and Stefan Wermter. 2019c. Semantic object accuracy for generative text-to-image synthesis. arXiv preprint arXiv:1910.13321 (2019)."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2020.3021209"},{"key":"e_1_3_2_2_13_1","volume-title":"Laurens Van Der Maaten, and Kilian Q Weinberger","author":"Huang Gao","year":"2017","unstructured":"Gao Huang , Danlu Chen , Tianhong Li , Felix Wu , Laurens Van Der Maaten, and Kilian Q Weinberger . 2017 . Multi-scale de nse networks for resource efficient image classification. arXiv preprint arXiv:1703.09844 (2017). Gao Huang, Danlu Chen, Tianhong Li, Felix Wu, Laurens Van Der Maaten, and Kilian Q Weinberger. 2017. Multi-scale dense networks for resource efficient image classification. arXiv preprint arXiv:1703.09844 (2017)."},{"key":"e_1_3_2_2_14_1","volume-title":"Transgan: Two transformers can make one strong gan. arXiv preprint arXiv:2102.07074","author":"Jiang Yifan","year":"2021","unstructured":"Yifan Jiang , Shiyu Chang , and Zhangyang Wang . 2021 . Transgan: Two transformers can make one strong gan. arXiv preprint arXiv:2102.07074 , Vol. 1 , 3 (2021). Yifan Jiang, Shiyu Chang, and Zhangyang Wang. 2021. Transgan: Two transformers can make one strong gan. arXiv preprint arXiv:2102.07074 , Vol. 1, 3 (2021)."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00133"},{"key":"e_1_3_2_2_16_1","volume-title":"Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Jimmy Ba . 2014 . Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014). Diederik P Kingma and Jimmy Ba. 2014. Adam: A method for stochastic optimization. arXiv preprint arXiv:1412.6980 (2014)."},{"key":"e_1_3_2_2_17_1","volume-title":"Controllable text-to-image generation. arXiv preprint arXiv:1909.07083","author":"Li Bowen","year":"2019","unstructured":"Bowen Li , Xiaojuan Qi , Thomas Lukasiewicz , and Philip HS Torr . 2019a. Controllable text-to-image generation. arXiv preprint arXiv:1909.07083 ( 2019 ). Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, and Philip HS Torr. 2019a. Controllable text-to-image generation. arXiv preprint arXiv:1909.07083 (2019)."},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01245"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00858"},{"key":"e_1_3_2_2_20_1","volume-title":"Geometric gan. arXiv preprint arXiv:1705.02894","author":"Lim Jae Hyun","year":"2017","unstructured":"Jae Hyun Lim and Jong Chul Ye. 2017. Geometric gan. arXiv preprint arXiv:1705.02894 ( 2017 ). Jae Hyun Lim and Jong Chul Ye. 2017. Geometric gan. arXiv preprint arXiv:1705.02894 (2017)."},{"key":"e_1_3_2_2_21_1","volume-title":"Runtime neural pruning. Advances in neural information processing systems","author":"Lin Ji","year":"2017","unstructured":"Ji Lin , Yongming Rao , Jiwen Lu , and Jie Zhou . 2017. Runtime neural pruning. Advances in neural information processing systems , Vol. 30 ( 2017 ). Ji Lin, Yongming Rao, Jiwen Lu, and Jie Zhou. 2017. Runtime neural pruning. Advances in neural information processing systems , Vol. 30 (2017)."},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i3.16305"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.298"},{"key":"e_1_3_2_2_25_1","volume-title":"Neural discrete representation learning. arXiv preprint arXiv:1711.00937","author":"van den Oord Aaron","year":"2017","unstructured":"Aaron van den Oord , Oriol Vinyals , and Koray Kavukcuoglu . 2017. Neural discrete representation learning. arXiv preprint arXiv:1711.00937 ( 2017 ). Aaron van den Oord, Oriol Vinyals, and Koray Kavukcuoglu. 2017. Neural discrete representation learning. arXiv preprint arXiv:1711.00937 (2017)."},{"key":"e_1_3_2_2_26_1","volume-title":"Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1).","author":"Park Dong Huk","year":"2021","unstructured":"Dong Huk Park , Samaneh Azadi , Xihui Liu , Trevor Darrell , and Anna Rohrbach . 2021 . Benchmark for compositional text-to-image synthesis . In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1). Dong Huk Park, Samaneh Azadi, Xihui Liu, Trevor Darrell, and Anna Rohrbach. 2021. Benchmark for compositional text-to-image synthesis. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1)."},{"key":"e_1_3_2_2_27_1","volume-title":"PyTorch: An Imperative Style","author":"Paszke Adam","year":"1912","unstructured":"Adam Paszke , Sam Gross , Francisco Massa , Adam Lerer , James Bradbury , Gregory Chanan , Trevor Killeen , Zeming Lin , Natalia Gimelshein , Luca Antiga , Alban Desmaison , Andreas K\u00f6pf , Edward Yang , Zach DeVito , Martin Raison , Alykhan Tejani , Sasank Chilamkurthy , Benoit Steiner , Lu Fang , Junjie Bai , and Soumith Chintala . 2019. PyTorch: An Imperative Style , High-Performance Deep Learning Library . arxiv: 1912 .01703 [cs.LG] Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas K\u00f6pf, Edward Yang, Zach DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019. PyTorch: An Imperative Style, High-Performance Deep Learning Library. arxiv: 1912.01703 [cs.LG]"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01258-8_16"},{"key":"e_1_3_2_2_29_1","first-page":"887","article-title":"Learn, imagine and create: Text-to-image generation from prior knowledge","volume":"32","author":"Qiao Tingting","year":"2019","unstructured":"Tingting Qiao , Jing Zhang , Duanqing Xu , and Dacheng Tao . 2019 a. Learn, imagine and create: Text-to-image generation from prior knowledge . Advances in Neural Information Processing Systems , Vol. 32 (2019), 887 -- 897 . Tingting Qiao, Jing Zhang, Duanqing Xu, and Dacheng Tao. 2019a. Learn, imagine and create: Text-to-image generation from prior knowledge. Advances in Neural Information Processing Systems , Vol. 32 (2019), 887--897.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00160"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3474085.3475363"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394171.3413961"},{"key":"e_1_3_2_2_33_1","volume-title":"Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al.","author":"Radford Alec","year":"2021","unstructured":"Alec Radford , Jong Wook Kim , Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021 . Learning transferable visual models from natural language supervision. arXiv preprint arXiv:2103.00020 (2021). Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021. Learning transferable visual models from natural language supervision. arXiv preprint arXiv:2103.00020 (2021)."},{"key":"e_1_3_2_2_34_1","volume-title":"Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092","author":"Ramesh Aditya","year":"2021","unstructured":"Aditya Ramesh , Mikhail Pavlov , Gabriel Goh , Scott Gray , Chelsea Voss , Alec Radford , Mark Chen , and Ilya Sutskever . 2021. Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092 ( 2021 ). Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021. Zero-shot text-to-image generation. arXiv preprint arXiv:2102.12092 (2021)."},{"key":"e_1_3_2_2_35_1","volume-title":"International Conference on Machine Learning. PMLR, 1060--1069","author":"Reed Scott","year":"2016","unstructured":"Scott Reed , Zeynep Akata , Xinchen Yan , Lajanugen Logeswaran , Bernt Schiele , and Honglak Lee . 2016 . Generative adversarial text to image synthesis . In International Conference on Machine Learning. PMLR, 1060--1069 . Scott Reed, Zeynep Akata, Xinchen Yan, Lajanugen Logeswaran, Bernt Schiele, and Honglak Lee. 2016. Generative adversarial text to image synthesis. In International Conference on Machine Learning. PMLR, 1060--1069."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01370"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_3_2_2_38_1","volume-title":"Semantics-Enhanced Adversarial Nets for Text-to-Image Synthesis. In 2019 IEEE\/CVF International Conference on Computer Vision (ICCV). IEEE Computer Society, 10500--10509","author":"Tan Hongchen","year":"2019","unstructured":"Hongchen Tan , Xiuping Liu , Xin Li , Yi Zhang , and Baocai Yin . 2019 . Semantics-Enhanced Adversarial Nets for Text-to-Image Synthesis. In 2019 IEEE\/CVF International Conference on Computer Vision (ICCV). IEEE Computer Society, 10500--10509 . Hongchen Tan, Xiuping Liu, Xin Li, Yi Zhang, and Baocai Yin. 2019. Semantics-Enhanced Adversarial Nets for Text-to-Image Synthesis. In 2019 IEEE\/CVF International Conference on Computer Vision (ICCV). IEEE Computer Society, 10500--10509."},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2020.3026728"},{"key":"e_1_3_2_2_40_1","volume-title":"DF-GAN: Deep Fusion Generative Adversarial Networks for Text-to-Image Synthesis. arXiv e-prints","author":"Tao Ming","year":"2020","unstructured":"Ming Tao , Hao Tang , Songsong Wu , Nicu Sebe , Fei Wu , and Xiao-Yuan Jing . 2020. DF-GAN: Deep Fusion Generative Adversarial Networks for Text-to-Image Synthesis. arXiv e-prints ( 2020 ), arXiv--2008. Ming Tao, Hao Tang, Songsong Wu, Nicu Sebe, Fei Wu, and Xiao-Yuan Jing. 2020. DF-GAN: Deep Fusion Generative Adversarial Networks for Text-to-Image Synthesis. arXiv e-prints (2020), arXiv--2008."},{"key":"e_1_3_2_2_41_1","unstructured":"Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998--6008.  Ashish Vaswani Noam Shazeer Niki Parmar Jakob Uszkoreit Llion Jones Aidan N Gomez \u0141ukasz Kaiser and Illia Polosukhin. 2017. Attention is all you need. In Advances in neural information processing systems. 5998--6008."},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01246-5_1"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.193"},{"key":"e_1_3_2_2_44_1","unstructured":"Catherine Wah Steve Branson Peter Welinder Pietro Perona and Serge Belongie. 2011. The caltech-ucsd birds-200--2011 dataset. (2011).  Catherine Wah Steve Branson Peter Welinder Pietro Perona and Serge Belongie. 2011. The caltech-ucsd birds-200--2011 dataset. (2011)."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00143"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00244"},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2021.3055062"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00243"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00089"},{"key":"e_1_3_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.629"},{"key":"e_1_3_2_2_51_1","volume-title":"Stackgan: Realistic image synthesis with stacked generative adversarial networks","author":"Zhang Han","year":"2018","unstructured":"Han Zhang , Tao Xu , Hongsheng Li , Shaoting Zhang , Xiaogang Wang , Xiaolei Huang , and Dimitris N Metaxas . 2018 . Stackgan: Realistic image synthesis with stacked generative adversarial networks . IEEE transactions on pattern analysis and machine intelligence, Vol. 41 , 8 (2018), 1947--1962. Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris N Metaxas. 2018. Stackgan: Realistic image synthesis with stacked generative adversarial networks. IEEE transactions on pattern analysis and machine intelligence, Vol. 41, 8 (2018), 1947--1962."},{"key":"e_1_3_2_2_52_1","first-page":"256","article-title":"PixelBrush: Art Generation from text with GANs. In Cl. Proj. Stanford CS231N Convolutional Neural Networks Vis. Recognition","volume":"2017","author":"Zhi Jiale","year":"2017","unstructured":"Jiale Zhi . 2017 . PixelBrush: Art Generation from text with GANs. In Cl. Proj. Stanford CS231N Convolutional Neural Networks Vis. Recognition , Sprint 2017. 256 . Jiale Zhi. 2017. PixelBrush: Art Generation from text with GANs. In Cl. Proj. Stanford CS231N Convolutional Neural Networks Vis. Recognition, Sprint 2017. 256.","journal-title":"Sprint"},{"key":"e_1_3_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00595"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547881","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547881","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:30Z","timestamp":1750186830000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547881"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":53,"alternative-id":["10.1145\/3503161.3547881","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547881","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}