{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:16:33Z","timestamp":1750220193153,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":46,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547810","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:42:35Z","timestamp":1665416555000},"page":"2441-2451","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Cloud2Sketch: Augmenting Clouds with Imaginary Sketches"],"prefix":"10.1145","author":[{"given":"Zhaoyi","family":"Wan","sequence":"first","affiliation":[{"name":"University of Rochester, Rochester, NY, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Dejia","family":"Xu","sequence":"additional","affiliation":[{"name":"University of Texas at Austin, Austin, TX, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhangyang","family":"Wang","sequence":"additional","affiliation":[{"name":"University of Texas at Austin, Austin, NY, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jian","family":"Wang","sequence":"additional","affiliation":[{"name":"Snap Inc., New York City, NY, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiebo","family":"Luo","sequence":"additional","affiliation":[{"name":"University of Rochester, Rochester, NY, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.24792"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1196\/annals.1440.011"},{"key":"e_1_3_2_2_4_1","volume-title":"Learning to generate line drawings that convey geometry and semantics. arXiv preprint arXiv:2203.12691","author":"Chan Caroline","year":"2022","unstructured":"Caroline Chan , Fredo Durand , and Phillip Isola . 2022. Learning to generate line drawings that convey geometry and semantics. arXiv preprint arXiv:2203.12691 ( 2022 ). Caroline Chan, Fredo Durand, and Phillip Isola. 2022. Learning to generate line drawings that convey geometry and semantics. arXiv preprint arXiv:2203.12691 (2022)."},{"key":"e_1_3_2_2_5_1","volume-title":"Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:1706.05587","author":"Chen Liang-Chieh","year":"2017","unstructured":"Liang-Chieh Chen , George Papandreou , Florian Schroff , and Hartwig Adam . 2017. Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:1706.05587 ( 2017 ). Liang-Chieh Chen, George Papandreou, Florian Schroff, and Hartwig Adam. 2017. Rethinking atrous convolution for semantic image segmentation. arXiv preprint arXiv:1706.05587 (2017)."},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1037\/h0022011"},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP.2017.8296300"},{"key":"e_1_3_2_2_8_1","volume-title":"How do humans sketch objects? ACM Transactions on graphics (TOG) 31, 4","author":"Eitz Mathias","year":"2012","unstructured":"Mathias Eitz , James Hays , and Marc Alexa . 2012. How do humans sketch objects? ACM Transactions on graphics (TOG) 31, 4 ( 2012 ), 1--10. Mathias Eitz, James Hays, and Marc Alexa. 2012. How do humans sketch objects? ACM Transactions on graphics (TOG) 31, 4 (2012), 1--10."},{"key":"e_1_3_2_2_9_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Geirhos Robert","year":"2019","unstructured":"Robert Geirhos , Patricia Rubisch , Claudio Michaelis , Matthias Bethge , Felix A Wichmann , and Wieland Brendel . 2019 . ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness . In International Conference on Learning Representations (ICLR). Robert Geirhos, Patricia Rubisch, Claudio Michaelis, Matthias Bethge, Felix A Wichmann, and Wieland Brendel. 2019. ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness. In International Conference on Learning Representations (ICLR)."},{"volume-title":"Animals and the human imagination: a companion to animal studies","author":"Gross Aaron","key":"e_1_3_2_2_10_1","unstructured":"Aaron Gross and Anne Vallely . 2012. Animals and the human imagination: a companion to animal studies . Columbia University Press . Aaron Gross and Anne Vallely. 2012. Animals and the human imagination: a companion to animal studies. Columbia University Press."},{"key":"e_1_3_2_2_11_1","volume-title":"A neural representation of sketch drawings. arXiv preprint arXiv:1704.03477","author":"Ha David","year":"2017","unstructured":"David Ha and Douglas Eck . 2017. A neural representation of sketch drawings. arXiv preprint arXiv:1704.03477 ( 2017 ). David Ha and Douglas Eck. 2017. A neural representation of sketch drawings. arXiv preprint arXiv:1704.03477 (2017)."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cobeha.2018.12.011"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3267347","article-title":"Alignet: Partial-shape agnostic alignment via unsupervised learning","volume":"38","author":"Hanocka Rana","year":"2018","unstructured":"Rana Hanocka , Noa Fish , Zhenhua Wang , Raja Giryes , Shachar Fleishman , and Daniel Cohen-Or . 2018 . Alignet: Partial-shape agnostic alignment via unsupervised learning . ACM Transactions on Graphics (TOG) 38 , 1 (2018), 1 -- 14 . Rana Hanocka, Noa Fish, Zhenhua Wang, Raja Giryes, Shachar Fleishman, and Daniel Cohen-Or. 2018. Alignet: Partial-shape agnostic alignment via unsupervised learning. ACM Transactions on Graphics (TOG) 38, 1 (2018), 1--14.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2003.817125"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neuron.2017.06.011"},{"key":"e_1_3_2_2_16_1","volume-title":"Masked autoencoders are scalable vision learners. arXiv preprint arXiv:2111.06377","author":"He Kaiming","year":"2021","unstructured":"Kaiming He , Xinlei Chen , Saining Xie , Yanghao Li , Piotr Doll\u00e1r , and Ross Girshick . 2021. Masked autoencoders are scalable vision learners. arXiv preprint arXiv:2111.06377 ( 2021 ). Kaiming He, Xinlei Chen, Saining Xie, Yanghao Li, Piotr Doll\u00e1r, and Ross Girshick. 2021. Masked autoencoders are scalable vision learners. arXiv preprint arXiv:2111.06377 (2021)."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.276"},{"key":"e_1_3_2_2_19_1","volume-title":"Visual pattern recognition by moment invariants. IRE transactions on information theory 8, 2","author":"Ming-Kuei Hu.","year":"1962","unstructured":"Ming-Kuei Hu. 1962. Visual pattern recognition by moment invariants. IRE transactions on information theory 8, 2 ( 1962 ), 179--187. Ming-Kuei Hu. 1962. Visual pattern recognition by moment invariants. IRE transactions on information theory 8, 2 (1962), 179--187."},{"volume-title":"International Cloud Atlas Vol 2","author":"World","key":"e_1_3_2_2_20_1","unstructured":"World international organization. 1987. International Cloud Atlas Vol 2 . World Meteorological Organization . World international organization. 1987. International Cloud Atlas Vol 2. World Meteorological Organization."},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.632"},{"key":"e_1_3_2_2_22_1","unstructured":"Max Jaderberg Karen Simonyan Andrew Zisserman etal 2015. Spatial transformer networks. Advances in neural information processing systems 28 (2015).  Max Jaderberg Karen Simonyan Andrew Zisserman et al. 2015. Spatial transformer networks. Advances in neural information processing systems 28 (2015)."},{"key":"e_1_3_2_2_23_1","volume-title":"Augmented reality in education: current technologies and the potential for education. Procedia-social and behavioral sciences 47","author":"Kesim Mehmet","year":"2012","unstructured":"Mehmet Kesim and Yasin Ozarslan . 2012. Augmented reality in education: current technologies and the potential for education. Procedia-social and behavioral sciences 47 ( 2012 ), 297--302. Mehmet Kesim and Yasin Ozarslan. 2012. Augmented reality in education: current technologies and the potential for education. Procedia-social and behavioral sciences 47 (2012), 297--302."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.5555\/13437.13441"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV.2019.00154"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1175\/JTECH-D-11-00009.1"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.324"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00598"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1175\/JTECH1875.1"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v32i1.12214"},{"key":"e_1_3_2_2_32_1","first-page":"6069","article-title":"EAR: Enhanced augmented reality system for sports entertainment applications","volume":"11","author":"Mahmood Zahid","year":"2017","unstructured":"Zahid Mahmood , Tauseef Ali , Nazeer Muhammad , Nargis Bibi , Imran Shahzad , and Shoaib Azmat . 2017 . EAR: Enhanced augmented reality system for sports entertainment applications . KSII Transactions on Internet and Information Systems (TIIS) 11 , 12 (2017), 6069 -- 6091 . Zahid Mahmood, Tauseef Ali, Nazeer Muhammad, Nargis Bibi, Imran Shahzad, and Shoaib Azmat. 2017. EAR: Enhanced augmented reality system for sports entertainment applications. KSII Transactions on Internet and Information Systems (TIIS) 11, 12 (2017), 6069--6091.","journal-title":"KSII Transactions on Internet and Information Systems (TIIS)"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2001.937655"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1089\/cmb.2006.13.614"},{"key":"e_1_3_2_2_35_1","volume-title":"International Conference on Machine Learning (ICML). PMLR, 8748--8763","author":"Radford Alec","year":"2021","unstructured":"Alec Radford , Jong Wook Kim , Chris Hallacy , Aditya Ramesh , Gabriel Goh , Sandhini Agarwal , Girish Sastry , Amanda Askell , Pamela Mishkin , Jack Clark , 2021 . Learning transferable visual models from natural language supervision . In International Conference on Machine Learning (ICML). PMLR, 8748--8763 . Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. 2021. Learning transferable visual models from natural language supervision. In International Conference on Machine Learning (ICML). PMLR, 8748--8763."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/2897824.2925954"},{"key":"e_1_3_2_2_37_1","volume-title":"Chen Change Loy, and Ran He","author":"Song Linsen","year":"2021","unstructured":"Linsen Song , Wayne Wu , Chaoyou Fu , Chen Qian , Chen Change Loy, and Ran He . 2021 . Everything's Talkin': Pareidolia Face Reenactment . arXiv preprint arXiv:2104.03061 (2021). Linsen Song, Wayne Wu, Chaoyou Fu, Chen Qian, Chen Change Loy, and Ran He. 2021. Everything's Talkin': Pareidolia Face Reenactment. arXiv preprint arXiv:2104.03061 (2021)."},{"key":"e_1_3_2_2_38_1","volume-title":"An Efficient Solution for Semantic Segmentation of Three Ground-based Cloud Datasets. Earth and Space Science 7, 4","author":"Song Qianqian","year":"2020","unstructured":"Qianqian Song , Zhihui Cui , and Pu Liu . 2020. An Efficient Solution for Semantic Segmentation of Three Ground-based Cloud Datasets. Earth and Space Science 7, 4 ( 2020 ), e2019EA001040. Qianqian Song, Zhihui Cui, and Pu Liu. 2020. An Efficient Solution for Semantic Segmentation of Three Ground-based Cloud Datasets. Earth and Space Science 7, 4 (2020), e2019EA001040."},{"key":"e_1_3_2_2_39_1","volume-title":"2020 IEEE Winter Conference on Applications of Computer Vision (WACV). IEEE Computer Society","author":"Soria X.","year":"1912","unstructured":"X. Soria , E. Riba , and A. Sappa . 2020. Dense Extreme Inception Network: Towards a Robust CNN Model for Edge Detection . In 2020 IEEE Winter Conference on Applications of Computer Vision (WACV). IEEE Computer Society , Los Alamitos, CA, USA , 1912 --1921. X. Soria, E. Riba, and A. Sappa. 2020. Dense Extreme Inception Network: Towards a Robust CNN Model for Edge Detection. In 2020 IEEE Winter Conference on Applications of Computer Vision (WACV). IEEE Computer Society, Los Alamitos, CA, USA, 1912--1921."},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/FG.2018.00022"},{"key":"e_1_3_2_2_41_1","volume-title":"Face photo-sketch synthesis and recognition","author":"Wang Xiaogang","year":"2008","unstructured":"Xiaogang Wang and Xiaoou Tang . 2008. Face photo-sketch synthesis and recognition . IEEE transactions on pattern analysis and machine intelligence (TPAMI) 31, 11 ( 2008 ), 1955--1967. Xiaogang Wang and Xiaoou Tang. 2008. Face photo-sketch synthesis and recognition. IEEE transactions on pattern analysis and machine intelligence (TPAMI) 31, 11 (2008), 1955--1967."},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00760"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00844"},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/IJCNN.2019.8851809"},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.125"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIP.2019.2910398"},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.244"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Lisboa Portugal","acronym":"MM '22"},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547810","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547810","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:34Z","timestamp":1750186954000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547810"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":46,"alternative-id":["10.1145\/3503161.3547810","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547810","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}