{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,28]],"date-time":"2026-07-28T14:46:17Z","timestamp":1785249977378,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":39,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"NSERC Discovery Grant","award":["RGPIN-2019-04575"],"award-info":[{"award-number":["RGPIN-2019-04575"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413635","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T13:10:44Z","timestamp":1602508244000},"page":"2021-2029","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":387,"title":["Action2Motion"],"prefix":"10.1145","author":[{"given":"Chuan","family":"Guo","sequence":"first","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xinxin","family":"Zuo","sequence":"additional","affiliation":[{"name":"University of Alberta &amp; University of Guelph, Edmonton, AB, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sen","family":"Wang","sequence":"additional","affiliation":[{"name":"University of Alberta &amp; University of Guelph, Edmonton, AB, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shihao","family":"Zou","sequence":"additional","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Qingyao","family":"Sun","sequence":"additional","affiliation":[{"name":"University of Chicago, Chicago, IL, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Annan","family":"Deng","sequence":"additional","affiliation":[{"name":"Yale University, New Haven, CT, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Minglun","family":"Gong","sequence":"additional","affiliation":[{"name":"University of Guelph, Guelph, ON, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Li","family":"Cheng","sequence":"additional","affiliation":[{"name":"University of Alberta, Edmonton, AB, Canada"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2018.8460608"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/3DV.2019.00084"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01216-8_23"},{"key":"e_1_3_2_2_4_1","volume-title":"Pau Climent-P\u00e9rez, and Francisco Fl\u00f3rez-Revuelta.","author":"Chaaraoui Alexandros Andre","year":"2014","unstructured":"Alexandros Andre Chaaraoui , Jos\u00e9 Ram\u00f3n Padilla-L\u00f3pez , Pau Climent-P\u00e9rez, and Francisco Fl\u00f3rez-Revuelta. 2014 . Evolutionary joint selection to improve human action recognition with RGB-D devices. Expert systems with applications, Vol. 41 , 3 (2014), 786--794. Alexandros Andre Chaaraoui, Jos\u00e9 Ram\u00f3n Padilla-L\u00f3pez, Pau Climent-P\u00e9rez, and Francisco Fl\u00f3rez-Revuelta. 2014. Evolutionary joint selection to improve human action recognition with RGB-D devices. Expert systems with applications, Vol. 41, 3 (2014), 786--794."},{"key":"e_1_3_2_2_5_1","unstructured":"Carnegie Mellon University. 2003. Carnegie Mellon University graphics lab motion capture database. (2003).  Carnegie Mellon University. 2003. Carnegie Mellon University graphics lab motion capture database. (2003)."},{"key":"e_1_3_2_2_6_1","volume-title":"International Conference on Machine Learning (ICML). 1174--1183","author":"Denton Emily","year":"2018","unstructured":"Emily Denton and Rob Fergus . 2018 . Stochastic Video Generation with a Learned Prior . In International Conference on Machine Learning (ICML). 1174--1183 . Emily Denton and Rob Fergus. 2018. Stochastic Video Generation with a Learned Prior. In International Conference on Machine Learning (ICML). 1174--1183."},{"key":"e_1_3_2_2_7_1","volume-title":"International workshop on automatic face-and gesture-recognition. Citeseer, 272--277","author":"Gavrila Dariu M","year":"1995","unstructured":"Dariu M Gavrila , Larry S Davis , 1995 . Towards 3-d model-based tracking and recognition of human movement: a multi-view approach . In International workshop on automatic face-and gesture-recognition. Citeseer, 272--277 . Dariu M Gavrila, Larry S Davis, et almbox. 1995. Towards 3-d model-based tracking and recognition of human movement: a multi-view approach. In International workshop on automatic face-and gesture-recognition. Citeseer, 272--277."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01225-0_48"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2017.01.011"},{"key":"e_1_3_2_2_10_1","unstructured":"Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In Advances in neural information processing systems. 6626--6637.  Martin Heusel Hubert Ramsauer Thomas Unterthiner Bernhard Nessler and Sepp Hochreiter. 2017. Gans trained by a two time-scale update rule converge to a local nash equilibrium. In Advances in neural information processing systems. 6626--6637."},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.137"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.5555\/2540128.2540483"},{"key":"e_1_3_2_2_13_1","volume-title":"Cho, and Seon Joo Kim.","author":"Kim Yunji","year":"2019","unstructured":"Yunji Kim , Seonghyeon Nam , In Cho, and Seon Joo Kim. 2019 . Unsupervised Keypoint Learning for Guiding Class-Conditional Video Prediction. In Advances in Neural Information Processing Systems. 3809--3819. Yunji Kim, Seonghyeon Nam, In Cho, and Seon Joo Kim. 2019. Unsupervised Keypoint Learning for Guiding Class-Conditional Video Prediction. In Advances in Neural Information Processing Systems. 3809--3819."},{"key":"e_1_3_2_2_14_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Kingma Diederik P","year":"2014","unstructured":"Diederik P Kingma and Max Welling . 2014 . Auto-encoding variational bayes . In International Conference on Learning Representations (ICLR). Diederik P Kingma and Max Welling. 2014. Auto-encoding variational bayes. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_2_15_1","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR).","author":"Kocabas Muhammed","unstructured":"Muhammed Kocabas , Nikos Athanasiou , and Michael J. Black . 2020. VIBE: Video Inference for Human Body Pose and Shape Estimation . In Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR). Muhammed Kocabas, Nikos Athanasiou, and Michael J. Black. 2020. VIBE: Video Inference for Human Body Pose and Shape Estimation. In Proceedings of the IEEE conference on computer vision and pattern recognition (CVPR)."},{"key":"e_1_3_2_2_16_1","unstructured":"Hsin-Ying Lee Xiaodong Yang Ming-Yu Liu Ting-Chun Wang Yu-Ding Lu Ming-Hsuan Yang and Jan Kautz. 2019. Dancing to Music. In Advances in Neural Information Processing Systems. 3581--3591.  Hsin-Ying Lee Xiaodong Yang Ming-Yu Liu Ting-Chun Wang Yu-Ding Lu Ming-Hsuan Yang and Jan Kautz. 2019. Dancing to Music. In Advances in Neural Information Processing Systems. 3581--3591."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2010.5543273"},{"key":"e_1_3_2_2_18_1","volume-title":"Proceedings of the Visually Grounded Interaction and Language Workshop at NeurIPS","author":"Lin Angela S","year":"2018","unstructured":"Angela S Lin , Lemeng Wu , Rodolfo Corona , Kevin Tai , Qixing Huang , and Raymond J Mooney . 2018 . generating animated videos of human activities from natural language descriptions . In Proceedings of the Visually Grounded Interaction and Language Workshop at NeurIPS 2018. Angela S Lin, Lemeng Wu, Rodolfo Corona, Kevin Tai, Qixing Huang, and Raymond J Mooney. 2018. generating animated videos of human activities from natural language descriptions. In Proceedings of the Visually Grounded Interaction and Language Workshop at NeurIPS 2018."},{"key":"e_1_3_2_2_19_1","volume-title":"Gang Wang, Ling-Yu Duan, and Alex Kot Chichung. 2019 a. NTU RGBD 120: A Large-Scale Benchmark for 3D Human Activity Understanding","author":"Liu Jun","year":"2019","unstructured":"Jun Liu , Amir Shahroudy , Mauricio Lisboa Perez , Gang Wang, Ling-Yu Duan, and Alex Kot Chichung. 2019 a. NTU RGBD 120: A Large-Scale Benchmark for 3D Human Activity Understanding . IEEE transactions on pattern analysis and machine intelligence ( 2019 ). Jun Liu, Amir Shahroudy, Mauricio Lisboa Perez, Gang Wang, Ling-Yu Duan, and Alex Kot Chichung. 2019 a. NTU RGBD 120: A Large-Scale Benchmark for 3D Human Activity Understanding. IEEE transactions on pattern analysis and machine intelligence (2019)."},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01024"},{"key":"e_1_3_2_2_21_1","volume-title":"Information retrieval for music and motion","author":"M\u00fcller Meinard","unstructured":"Meinard M\u00fcller . 2007. Information retrieval for music and motion . Vol. 2 . Springer . Meinard M\u00fcller. 2007. Information retrieval for music and motion. Vol. 2. Springer."},{"key":"e_1_3_2_2_22_1","unstructured":"Meinard M\u00fcller Tido R\u00f6der Michael Clausen Bernhard Eberhardt Bj\u00f6rn Kr\u00fcger and Andreas Weber. [n.d.]. Mocap database hdm05. ([n. d.]).  Meinard M\u00fcller Tido R\u00f6der Michael Clausen Bernhard Eberhardt Bj\u00f6rn Kr\u00fcger and Andreas Weber. [n.d.]. Mocap database hdm05. ([n. d.])."},{"key":"e_1_3_2_2_23_1","volume-title":"A mathematical introduction to robotic manipulation","author":"Murray Richard M","unstructured":"Richard M Murray , Zexiang Li , and S Shankar Sastry . 1994. A mathematical introduction to robotic manipulation . CRC press . Richard M Murray, Zexiang Li, and S Shankar Sastry. 1994. A mathematical introduction to robotic manipulation. CRC press."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2018.07.006"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00790"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-019-01281-2"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.450"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3125739.3132594"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/3240508.3240526"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00165"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.82"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2012.6247813"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2012.6239233"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-017-0998-6"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1006\/cviu.1998.0726"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01228-1_17"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01249-6_13"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58568-6_21"},{"key":"e_1_3_2_2_39_1","unstructured":"Shihao Zou Xinxin Zuo Yiming Qian Sen Wang Chi Xu Minglun Gong and Li Cheng. 2020 b. Polarization Human Shape and Pose Dataset.  Shihao Zou Xinxin Zuo Yiming Qian Sen Wang Chi Xu Minglun Gong and Li Cheng. 2020 b. Polarization Human Shape and Pose Dataset."}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","location":"Seattle WA USA","acronym":"MM '20","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413635","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413635","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:47:15Z","timestamp":1750193235000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413635"}},"subtitle":["Conditioned Generation of 3D Human Motions"],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":39,"alternative-id":["10.1145\/3394171.3413635","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413635","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}