{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,23]],"date-time":"2025-08-23T05:25:03Z","timestamp":1755926703572,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":49,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Innovation Program for Quantum Science and Technology","award":["No. 2021ZD0302902"],"award-info":[{"award-number":["No. 2021ZD0302902"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["No. 62172385"],"award-info":[{"award-number":["No. 62172385"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3548159","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:43:12Z","timestamp":1665416592000},"page":"2211-2219","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["AtHom: Two Divergent Attentions Stimulated By Homomorphic Training in Text-to-Image Synthesis"],"prefix":"10.1145","author":[{"given":"Zhenbo","family":"Shi","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhi","family":"Chen","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhenbo","family":"Xu","sequence":"additional","affiliation":[{"name":"Hangzhou Innovation Institute, Beihang University, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Wei","family":"Yang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Liusheng","family":"Huang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2017.8296650"},{"key":"e_1_3_2_2_2_1","unstructured":"Ben Bolker. 2020. Maximum likelihood estimation and analysis with the bbmle package.  Ben Bolker. 2020. Maximum likelihood estimation and analysis with the bbmle package."},{"key":"e_1_3_2_2_3_1","unstructured":"Konstantinos Bousmalis George Trigeorgis Nathan Silberman Dilip Krishnan and Dumitru Erhan. 2016. Domain separation networks. In Advances in neural information processing systems. 343--351.  Konstantinos Bousmalis George Trigeorgis Nathan Silberman Dilip Krishnan and Dumitru Erhan. 2016. Domain separation networks. In Advances in neural information processing systems. 343--351."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01092"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.5555\/2946645.2946704"},{"key":"e_1_3_2_2_6_1","unstructured":"Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In Advances in Neural Information Processing Systems. 6626--6637.  Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In Advances in Neural Information Processing Systems. 6626--6637."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2020.3021209"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00766"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1214\/18-AAP1421"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00790"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01245"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00500-020-04951-3"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2020.2977573"},{"key":"e_1_3_2_2_15_1","volume-title":"Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784","author":"Mirza Mehdi","year":"2014","unstructured":"Mehdi Mirza and Simon Osindero . 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 ( 2014 ). Mehdi Mirza and Simon Osindero. 2014. Conditional generative adversarial nets. arXiv preprint arXiv:1411.1784 (2014)."},{"key":"e_1_3_2_2_16_1","unstructured":"Volodymyr Mnih Nicolas Heess Alex Graves etal 2014. Recurrent models of visual attention. In Advances in neural information processing systems. 2204--2212.  Volodymyr Mnih Nicolas Heess Alex Graves et al. 2014. Recurrent models of visual attention. In Advances in neural information processing systems. 2204--2212."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0022-2496(02)00028-7"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.5555\/3305890.3305954"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00160"},{"key":"e_1_3_2_2_20_1","volume-title":"Generative adversarial text to image synthesis. arXiv preprint arXiv:1605.05396","author":"Reed Scott","year":"2016","unstructured":"Scott Reed , Zeynep Akata , Xinchen Yan , Lajanugen Logeswaran , Bernt Schiele , and Honglak Lee . 2016b. Generative adversarial text to image synthesis. arXiv preprint arXiv:1605.05396 ( 2016 ). Scott Reed, Zeynep Akata, Xinchen Yan, Lajanugen Logeswaran, Bernt Schiele, and Honglak Lee. 2016b. Generative adversarial text to image synthesis. arXiv preprint arXiv:1605.05396 (2016)."},{"key":"e_1_3_2_2_21_1","volume-title":"Ziyu Wang, Dan Belov, and Nando De Freitas.","author":"Reed Scott","year":"2017","unstructured":"Scott Reed , A\"aron van den Oord, Nal Kalchbrenner , Sergio G\u00f3mez Colmenarejo , Ziyu Wang, Dan Belov, and Nando De Freitas. 2017 . Parallel multiscale autoregressive density estimation. arXiv preprint arXiv:1703.03664 (2017). Scott Reed, A\"aron van den Oord, Nal Kalchbrenner, Sergio G\u00f3mez Colmenarejo, Ziyu Wang, Dan Belov, and Nando De Freitas. 2017. Parallel multiscale autoregressive density estimation. arXiv preprint arXiv:1703.03664 (2017)."},{"key":"e_1_3_2_2_22_1","unstructured":"Scott E Reed Zeynep Akata Santosh Mohan Samuel Tenka Bernt Schiele and Honglak Lee. 2016a. Learning what and where to draw. In Advances in Neural Information Processing Systems. 217--225.  Scott E Reed Zeynep Akata Santosh Mohan Samuel Tenka Bernt Schiele and Honglak Lee. 2016a. Learning what and where to draw. In Advances in Neural Information Processing Systems. 217--225."},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"crossref","unstructured":"Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein etal 2015. Imagenet large scale visual recognition challenge. International journal of computer vision Vol. 115 3 (2015) 211--252.  Olga Russakovsky Jia Deng Hao Su Jonathan Krause Sanjeev Satheesh Sean Ma Zhiheng Huang Andrej Karpathy Aditya Khosla Michael Bernstein et al. 2015. Imagenet large scale visual recognition challenge. International journal of computer vision Vol. 115 3 (2015) 211--252.","DOI":"10.1007\/s11263-015-0816-y"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISBI.2019.8759152"},{"key":"e_1_3_2_2_25_1","volume-title":"International Conference on Machine Learning. PMLR, 8624--8633","author":"Shankar Tanmay","year":"2020","unstructured":"Tanmay Shankar and Abhinav Gupta . 2020 . Learning robot skills with temporal variational inference . In International Conference on Machine Learning. PMLR, 8624--8633 . Tanmay Shankar and Abhinav Gupta. 2020. Learning robot skills with temporal variational inference. In International Conference on Machine Learning. PMLR, 8624--8633."},{"key":"e_1_3_2_2_26_1","volume-title":"International Conference on Artificial Intelligence and Statistics. PMLR","author":"Shi Jiaxin","year":"2020","unstructured":"Jiaxin Shi , Michalis Titsias , and Andriy Mnih . 2020 . Sparse orthogonal variational inference for gaussian processes . In International Conference on Artificial Intelligence and Statistics. PMLR , 1932--1942. Jiaxin Shi, Michalis Titsias, and Andriy Mnih. 2020. Sparse orthogonal variational inference for gaussian processes. In International Conference on Artificial Intelligence and Statistics. PMLR, 1932--1942."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v36i8.20802"},{"key":"e_1_3_2_2_28_1","volume-title":"Adversarial Attacks on Object Detectors with Limited Perturbations. In ICASSP 2021--2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 1375--1379","author":"Shi Zhenbo","year":"2021","unstructured":"Zhenbo Shi , Wei Yang , Zhenbo Xu , Zhi Chen , Yingjie Li , Haoran Zhu , and Liusheng Huang . 2021 . Adversarial Attacks on Object Detectors with Limited Perturbations. In ICASSP 2021--2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 1375--1379 . Zhenbo Shi, Wei Yang, Zhenbo Xu, Zhi Chen, Yingjie Li, Haoran Zhu, and Liusheng Huang. 2021. Adversarial Attacks on Object Detectors with Limited Perturbations. In ICASSP 2021--2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). IEEE, 1375--1379."},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.308"},{"key":"e_1_3_2_2_30_1","volume-title":"Intriguing properties of neural networks. arXiv preprint arXiv:1312.6199","author":"Szegedy Christian","year":"2013","unstructured":"Christian Szegedy , Wojciech Zaremba , Ilya Sutskever , Joan Bruna , Dumitru Erhan , Ian Goodfellow , and Rob Fergus . 2013. Intriguing properties of neural networks. arXiv preprint arXiv:1312.6199 ( 2013 ). Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus. 2013. Intriguing properties of neural networks. arXiv preprint arXiv:1312.6199 (2013)."},{"key":"e_1_3_2_2_31_1","volume-title":"Hysteresis loop reversing by applying Langevin approximation. COMPEL-The international journal for computation and mathematics in electrical and electronic engineering","author":"Takacs Jeno","year":"2017","unstructured":"Jeno Takacs . 2017. Hysteresis loop reversing by applying Langevin approximation. COMPEL-The international journal for computation and mathematics in electrical and electronic engineering ( 2017 ). Jeno Takacs. 2017. Hysteresis loop reversing by applying Langevin approximation. COMPEL-The international journal for computation and mathematics in electrical and electronic engineering (2017)."},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.01060"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2018.2866698"},{"key":"e_1_3_2_2_34_1","unstructured":"Aaron Van den Oord Nal Kalchbrenner Lasse Espeholt Oriol Vinyals Alex Graves etal 2016. Conditional image generation with pixelcnn decoders. In Advances in neural information processing systems. 4790--4798.  Aaron Van den Oord Nal Kalchbrenner Lasse Espeholt Oriol Vinyals Alex Graves et al. 2016. Conditional image generation with pixelcnn decoders. In Advances in neural information processing systems. 4790--4798."},{"key":"e_1_3_2_2_35_1","unstructured":"Catherine Wah Steve Branson Peter Welinder Pietro Perona and Serge Belongie. 2011. The caltech-ucsd birds-200--2011 dataset. (2011).  Catherine Wah Steve Branson Peter Welinder Pietro Perona and Serge Belongie. 2011. The caltech-ucsd birds-200--2011 dataset. (2011)."},{"key":"e_1_3_2_2_36_1","volume-title":"S2IGAN: Speech-to-image generation via adversarial learning. arXiv preprint arXiv:2005.06968","author":"Wang Xinsheng","year":"2020","unstructured":"Xinsheng Wang , Tingting Qiao , Jihua Zhu , Alan Hanjalic , and Odette Scharenborg . 2020. S2IGAN: Speech-to-image generation via adversarial learning. arXiv preprint arXiv:2005.06968 ( 2020 ). Xinsheng Wang, Tingting Qiao, Jihua Zhu, Alan Hanjalic, and Odette Scharenborg. 2020. S2IGAN: Speech-to-image generation via adversarial learning. arXiv preprint arXiv:2005.06968 (2020)."},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2021.3052075"},{"key":"e_1_3_2_2_38_1","volume-title":"Attngan: Fine-grained text to image generation with attentional generative adversarial networks. arXiv preprint","author":"Xu Tao","year":"2017","unstructured":"Tao Xu , Pengchuan Zhang , Qiuyuan Huang , Han Zhang , Zhe Gan , Xiaolei Huang , and Xiaodong He . 2017 . Attngan: Fine-grained text to image generation with attentional generative adversarial networks. arXiv preprint (2017). Tao Xu, Pengchuan Zhang, Qiuyuan Huang, Han Zhang, Zhe Gan, Xiaolei Huang, and Xiaodong He. 2017. Attngan: Fine-grained text to image generation with attentional generative adversarial networks. arXiv preprint (2017)."},{"key":"e_1_3_2_2_39_1","volume-title":"Diversity-sensitive conditional generative adversarial networks. arXiv preprint arXiv:1901.09024","author":"Yang Dingdong","year":"2019","unstructured":"Dingdong Yang , Seunghoon Hong , Yunseok Jang , Tianchen Zhao , and Honglak Lee . 2019. Diversity-sensitive conditional generative adversarial networks. arXiv preprint arXiv:1901.09024 ( 2019 ). Dingdong Yang, Seunghoon Hong, Yunseok Jang, Tianchen Zhao, and Honglak Lee. 2019. Diversity-sensitive conditional generative adversarial networks. arXiv preprint arXiv:1901.09024 (2019)."},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.10"},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00243"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00457"},{"key":"e_1_3_2_2_43_1","volume-title":"International Conference on Machine Learning. PMLR, 7354--7363","author":"Zhang Han","year":"2019","unstructured":"Han Zhang , Ian Goodfellow , Dimitris Metaxas , and Augustus Odena . 2019 a. Self-attention generative adversarial networks . In International Conference on Machine Learning. PMLR, 7354--7363 . Han Zhang, Ian Goodfellow, Dimitris Metaxas, and Augustus Odena. 2019a. Self-attention generative adversarial networks. In International Conference on Machine Learning. PMLR, 7354--7363."},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.00089"},{"key":"e_1_3_2_2_45_1","volume-title":"Image de-raining using a conditional generative adversarial network","author":"Zhang He","year":"2019","unstructured":"He Zhang , Vishwanath Sindagi , and Vishal M Patel . 2019b. Image de-raining using a conditional generative adversarial network . IEEE transactions on circuits and systems for video technology, Vol. 30 , 11 ( 2019 ), 3943--3956. He Zhang, Vishwanath Sindagi, and Vishal M Patel. 2019b. Image de-raining using a conditional generative adversarial network. IEEE transactions on circuits and systems for video technology, Vol. 30, 11 (2019), 3943--3956."},{"key":"e_1_3_2_2_46_1","volume-title":"Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. arXiv preprint","author":"Zhang Han","year":"2017","unstructured":"Han Zhang , Tao Xu , Hongsheng Li , Shaoting Zhang , Xiaolei Huang , Xiaogang Wang , and Dimitris Metaxas . 2017 a. Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. arXiv preprint (2017). Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaolei Huang, Xiaogang Wang, and Dimitris Metaxas. 2017a. Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. arXiv preprint (2017)."},{"key":"e_1_3_2_2_47_1","volume-title":"Stackgan: Realistic image synthesis with stacked generative adversarial networks. arXiv preprint arXiv:1710.10916","author":"Zhang Han","year":"2017","unstructured":"Han Zhang , Tao Xu , Hongsheng Li , Shaoting Zhang , Xiaogang Wang , Xiaolei Huang , and Dimitris Metaxas . 2017 b. Stackgan: Realistic image synthesis with stacked generative adversarial networks. arXiv preprint arXiv:1710.10916 (2017). Han Zhang, Tao Xu, Hongsheng Li, Shaoting Zhang, Xiaogang Wang, Xiaolei Huang, and Dimitris Metaxas. 2017b. Stackgan: Realistic image synthesis with stacked generative adversarial networks. arXiv preprint arXiv:1710.10916 (2017)."},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00649"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00595"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Lisboa Portugal","acronym":"MM '22"},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548159","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3548159","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:00:19Z","timestamp":1750186819000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3548159"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":49,"alternative-id":["10.1145\/3503161.3548159","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3548159","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}