{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,23]],"date-time":"2025-08-23T05:25:09Z","timestamp":1755926709321,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":22,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,11,7]],"date-time":"2022-11-07T00:00:00Z","timestamp":1667779200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Shenzhen Science and Technology Innovation Committee","award":["WDZC20200818121348001"],"award-info":[{"award-number":["WDZC20200818121348001"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["62076144"],"award-info":[{"award-number":["62076144"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Shenzhen Key Laboratory of next generation interactive media innovative technology","award":["ZDSYS20210623092001004"],"award-info":[{"award-number":["ZDSYS20210623092001004"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,11,7]]},"DOI":"10.1145\/3536221.3558066","type":"proceedings-article","created":{"date-parts":[[2022,11,4]],"date-time":"2022-11-04T15:54:14Z","timestamp":1667577254000},"page":"758-763","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":8,"title":["The ReprGesture entry to the GENEA Challenge 2022"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-8533-0524","authenticated-orcid":false,"given":"Sicheng","family":"Yang","sequence":"first","affiliation":[{"name":"Shenzhen International Graduate School, Tsinghua University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Zhiyong","family":"Wu","sequence":"additional","affiliation":[{"name":"Shenzhen International Graduate School, Tsinghua University, China and The Chinese University of Hong Kong, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Minglei","family":"Li","sequence":"additional","affiliation":[{"name":"Huawei Cloud Computing Technologies Co., Ltd, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mengchen","family":"Zhao","sequence":"additional","affiliation":[{"name":"Huawei Noah's Ark Lab, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiuxin","family":"Lin","sequence":"additional","affiliation":[{"name":"Shenzhen International Graduate School, Tsinghua University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Liyang","family":"Chen","sequence":"additional","affiliation":[{"name":"Shenzhen International Graduate School, Tsinghua University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Weihong","family":"Bao","sequence":"additional","affiliation":[{"name":"Shenzhen International Graduate School, Tsinghua University, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,11,7]]},"reference":[{"key":"e_1_3_2_2_1_1","volume-title":"Enriching Word Vectors with Subword Information. Transactions of the Association for Computational Linguistics 5 (06","author":"Bojanowski Piotr","year":"2017","unstructured":"Piotr Bojanowski , Edouard Grave , Armand Joulin , and Tomas Mikolov . 2017. Enriching Word Vectors with Subword Information. Transactions of the Association for Computational Linguistics 5 (06 2017 ), 135\u2013146. https:\/\/doi.org\/10.1162\/tacl_a_00051 arXiv:https:\/\/direct.mit.edu\/tacl\/article-pdf\/doi\/10.1162\/tacl_a_00051\/1567442\/tacl_a_00051.pdf 10.1162\/tacl_a_00051 Piotr Bojanowski, Edouard Grave, Armand Joulin, and Tomas Mikolov. 2017. Enriching Word Vectors with Subword Information. Transactions of the Association for Computational Linguistics 5 (06 2017), 135\u2013146. https:\/\/doi.org\/10.1162\/tacl_a_00051 arXiv:https:\/\/direct.mit.edu\/tacl\/article-pdf\/doi\/10.1162\/tacl_a_00051\/1567442\/tacl_a_00051.pdf"},{"key":"e_1_3_2_2_2_1","volume-title":"Advances in Neural Information Processing Systems, D.\u00a0Lee, M.\u00a0Sugiyama, U.\u00a0Luxburg, I.\u00a0Guyon, and R.\u00a0Garnett (Eds.). Vol.\u00a029. Curran Associates","author":"Bousmalis Konstantinos","year":"2016","unstructured":"Konstantinos Bousmalis , George Trigeorgis , Nathan Silberman , Dilip Krishnan , and Dumitru Erhan . 2016. Domain Separation Networks . In Advances in Neural Information Processing Systems, D.\u00a0Lee, M.\u00a0Sugiyama, U.\u00a0Luxburg, I.\u00a0Guyon, and R.\u00a0Garnett (Eds.). Vol.\u00a029. Curran Associates , Inc .https:\/\/proceedings.neurips.cc\/paper\/ 2016 \/file\/45fbc6d3e05ebd93369ce542e8f2322d-Paper.pdf Konstantinos Bousmalis, George Trigeorgis, Nathan Silberman, Dilip Krishnan, and Dumitru Erhan. 2016. Domain Separation Networks. In Advances in Neural Information Processing Systems, D.\u00a0Lee, M.\u00a0Sugiyama, U.\u00a0Luxburg, I.\u00a0Guyon, and R.\u00a0Garnett (Eds.). Vol.\u00a029. Curran Associates, Inc.https:\/\/proceedings.neurips.cc\/paper\/2016\/file\/45fbc6d3e05ebd93369ce542e8f2322d-Paper.pdf"},{"key":"e_1_3_2_2_3_1","unstructured":"Sanyuan Chen Chengyi Wang Zhengyang Chen Yu Wu Shujie Liu Zhuo Chen Jinyu Li Naoyuki Kanda Takuya Yoshioka Xiong Xiao Jian Wu Long Zhou Shuo Ren Yanmin Qian Yao Qian Jian Wu Michael Zeng and Furu Wei. 2021. WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing. CoRR abs\/2110.13900(2021). arXiv:2110.13900https:\/\/arxiv.org\/abs\/2110.13900  Sanyuan Chen Chengyi Wang Zhengyang Chen Yu Wu Shujie Liu Zhuo Chen Jinyu Li Naoyuki Kanda Takuya Yoshioka Xiong Xiao Jian Wu Long Zhou Shuo Ren Yanmin Qian Yao Qian Jian Wu Michael Zeng and Furu Wei. 2021. WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing. CoRR abs\/2110.13900(2021). arXiv:2110.13900https:\/\/arxiv.org\/abs\/2110.13900"},{"key":"e_1_3_2_2_4_1","volume-title":"Learning Phrase Representations using RNN Encoder\u2013Decoder for Statistical Machine Translation. empirical methods in natural language processing","author":"Cho Kyunghyun","year":"2014","unstructured":"Kyunghyun Cho , Bart van Merri\u00ebnboer , Caglar Gulcehre , Dzmitry Bahdanau , Fethi Bougares , Holger Schwenk , and Yoshua Bengio . 2014. Learning Phrase Representations using RNN Encoder\u2013Decoder for Statistical Machine Translation. empirical methods in natural language processing ( 2014 ). Kyunghyun Cho, Bart van Merri\u00ebnboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014. Learning Phrase Representations using RNN Encoder\u2013Decoder for Statistical Machine Translation. empirical methods in natural language processing (2014)."},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2019.2916887"},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394171.3413678"},{"key":"e_1_3_2_2_7_1","unstructured":"P. Itu-T. 1996. Methods for subjective determination of transmission quality. ITU-T Recommendation P.800(1996). https:\/\/www.itu.int\/rec\/T-REC-P.800-199608-I  P. Itu-T. 1996. Methods for subjective determination of transmission quality. ITU-T Recommendation P.800(1996). https:\/\/www.itu.int\/rec\/T-REC-P.800-199608-I"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/3462244.3479957"},{"key":"e_1_3_2_2_9_1","volume-title":"Adam: A Method for Stochastic Optimization. Computer Science","author":"Kingma D.","year":"2014","unstructured":"D. Kingma and J. Ba . 2014 . Adam: A Method for Stochastic Optimization. Computer Science (2014). D. Kingma and J. Ba. 2014. Adam: A Method for Stochastic Optimization. Computer Science (2014)."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/3308532.3329472"},{"key":"#cr-split#-e_1_3_2_2_11_1.1","doi-asserted-by":"crossref","unstructured":"Taras Kucherenko Dai Hasegawa Naoshi Kaneko Gustav\u00a0Eje Henter and Hedvig Kjellstr\u00f6m. 2021. Moving Fast and Slow: Analysis of Representations and Post-Processing in Speech-Driven Automatic Gesture Generation. International Journal of Human-Computer Interaction 37 14(2021) 1300-1316. https:\/\/doi.org\/10.1080\/10447318.2021.1883883 arXiv:https:\/\/doi.org\/10.1080\/10447318.2021.1883883 10.1080\/10447318.2021.1883883","DOI":"10.1080\/10447318.2021.1883883"},{"key":"#cr-split#-e_1_3_2_2_11_1.2","doi-asserted-by":"crossref","unstructured":"Taras Kucherenko Dai Hasegawa Naoshi Kaneko Gustav\u00a0Eje Henter and Hedvig Kjellstr\u00f6m. 2021. Moving Fast and Slow: Analysis of Representations and Post-Processing in Speech-Driven Automatic Gesture Generation. International Journal of Human-Computer Interaction 37 14(2021) 1300-1316. https:\/\/doi.org\/10.1080\/10447318.2021.1883883 arXiv:https:\/\/doi.org\/10.1080\/10447318.2021.1883883","DOI":"10.1080\/10447318.2021.1883883"},{"key":"e_1_3_2_2_12_1","volume-title":"Gesticulator: A Framework for Semantically-Aware Speech-Driven Gesture Generation","author":"Kucherenko Taras","year":"2020","unstructured":"Taras Kucherenko , Patrik Jonell , Sanne van Waveren , Gustav\u00a0Eje Henter , Simon Alexandersson , Iolanda Leite , and Hedvig Kjellstr\u00f6m . 2020 . Gesticulator: A Framework for Semantically-Aware Speech-Driven Gesture Generation . Association for Computing Machinery , New York, NY, USA , 242\u2013250. https:\/\/doi.org\/10.1145\/3382507.3418815 10.1145\/3382507.3418815 Taras Kucherenko, Patrik Jonell, Sanne van Waveren, Gustav\u00a0Eje Henter, Simon Alexandersson, Iolanda Leite, and Hedvig Kjellstr\u00f6m. 2020. Gesticulator: A Framework for Semantically-Aware Speech-Driven Gesture Generation. Association for Computing Machinery, New York, NY, USA, 242\u2013250. https:\/\/doi.org\/10.1145\/3382507.3418815"},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3397481.3450692"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00085"},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01110"},{"key":"e_1_3_2_2_16_1","volume-title":"BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis. ArXiv abs\/2203.05297(2022).","author":"Liu Haiyang","year":"2022","unstructured":"Haiyang Liu , Zihao Zhu , Naoya Iwamoto , Yichen Peng , Zhengqing Li , You Zhou , Elif Bozkurt , and Bo Zheng . 2022 . BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis. ArXiv abs\/2203.05297(2022). Haiyang Liu, Zihao Zhu, Naoya Iwamoto, Yichen Peng, Zhengqing Li, You Zhou, Elif Bozkurt, and Bo Zheng. 2022. BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis. ArXiv abs\/2203.05297(2022)."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.01021"},{"key":"e_1_3_2_2_18_1","unstructured":"Jing Xu Wei Zhang Yalong Bai Qi-Biao Sun and Tao Mei. 2022. Freeform Body Motion Generation from Speech. ArXiv abs\/2203.02291(2022).  Jing Xu Wei Zhang Yalong Bai Qi-Biao Sun and Tao Mei. 2022. Freeform Body Motion Generation from Speech. ArXiv abs\/2203.02291(2022)."},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/3414685.3417838"},{"key":"e_1_3_2_2_20_1","volume-title":"Robots Learn Social Skills: End-to-End Learning of Co-Speech Gesture Generation for Humanoid Robots. In 2019 International Conference on Robotics and Automation (ICRA). 4303\u20134309","author":"Yoon Youngwoo","year":"2019","unstructured":"Youngwoo Yoon , Woo-Ri Ko , Minsu Jang , Jaeyeon Lee , Jaehong Kim , and Geehyuk Lee . 2019 . Robots Learn Social Skills: End-to-End Learning of Co-Speech Gesture Generation for Humanoid Robots. In 2019 International Conference on Robotics and Automation (ICRA). 4303\u20134309 . https:\/\/doi.org\/10.1109\/ICRA.2019.8793720 10.1109\/ICRA.2019.8793720 Youngwoo Yoon, Woo-Ri Ko, Minsu Jang, Jaeyeon Lee, Jaehong Kim, and Geehyuk Lee. 2019. Robots Learn Social Skills: End-to-End Learning of Co-Speech Gesture Generation for Humanoid Robots. In 2019 International Conference on Robotics and Automation (ICRA). 4303\u20134309. https:\/\/doi.org\/10.1109\/ICRA.2019.8793720"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/3536221.3558058"}],"event":{"name":"ICMI '22: INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"],"location":"Bengaluru India","acronym":"ICMI '22"},"container-title":["Proceedings of the 2022 International Conference on Multimodal Interaction"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3536221.3558066","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3536221.3558066","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:50Z","timestamp":1750186970000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3536221.3558066"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,7]]},"references-count":22,"alternative-id":["10.1145\/3536221.3558066","10.1145\/3536221"],"URL":"https:\/\/doi.org\/10.1145\/3536221.3558066","relation":{},"subject":[],"published":{"date-parts":[[2022,11,7]]},"assertion":[{"value":"2022-11-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}