{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,9]],"date-time":"2025-12-09T19:41:04Z","timestamp":1765309264582,"version":"3.46.0"},"publisher-location":"New York, NY, USA","reference-count":37,"publisher":"ACM","funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["No. 62222203 and No. 62476201"],"award-info":[{"award-number":["No. 62222203 and No. 62476201"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100012226","name":"Fundamental Research Funds for the Central Universities","doi-asserted-by":"publisher","award":["No. 22120250572"],"award-info":[{"award-number":["No. 22120250572"]}],"id":[{"id":"10.13039\/501100012226","id-type":"DOI","asserted-by":"publisher"}]},{"name":"the New Cornerstone Science Foundation through the XPLORER PRIZE"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,10,27]]},"DOI":"10.1145\/3746027.3762075","type":"proceedings-article","created":{"date-parts":[[2025,10,25]],"date-time":"2025-10-25T06:54:17Z","timestamp":1761375257000},"page":"14143-14149","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Multimodal Time Series Alignment for Error Detection in Human Robot Interactions"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2209-651X","authenticated-orcid":false,"given":"Xun","family":"Jiang","sequence":"first","affiliation":[{"name":"University of Electronic Science and Technology of China, Chengdu, China and Tongji University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0006-5176-7713","authenticated-orcid":false,"given":"Shuangle","family":"Li","sequence":"additional","affiliation":[{"name":"University of Electronic Science and Technology of China, Chengdu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-8197-178X","authenticated-orcid":false,"given":"Chong","family":"Liu","sequence":"additional","affiliation":[{"name":"University of Electronic Science and Technology of China, Chengdu, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5685-3123","authenticated-orcid":false,"given":"Xing","family":"Xu","sequence":"additional","affiliation":[{"name":"Tongji University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,10,27]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.02538"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v38i17.29884"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/TFUZZ.2024.3405541"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3731715.3733353"},{"volume-title":"Proceedings of the ACM International Conference on Multimedia, 2025. Challenge description page or workshop proceedings.","author":"Organizing Committee","key":"e_1_3_2_1_5_1","unstructured":"Organizing Committee. The err@hri 2.0 challenge: Detecting errors in human-robot interaction. In Proceedings of the ACM International Conference on Multimedia, 2025. Challenge description page or workshop proceedings."},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.3389\/frobt.2023.1202306"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1089\/dia.2022.0162"},{"key":"e_1_3_2_1_8_1","volume-title":"IEEE International Conference on Automatic Face and Gesture Recognition","author":"Zadeh Amir","year":"2018","unstructured":"Amir Zadeh, Yao Chong Lim, and Louis-Philippe Morency. Openface 2.0: facial behavior analysis toolkit tadas baltru\u0161aitis. In IEEE International Conference on Automatic Face and Gesture Recognition, 2018."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/1873951.1874246"},{"key":"e_1_3_2_1_10_1","first-page":"8748","volume-title":"International conference on machine learning","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al. Learning transferable visual models from natural language supervision. In International conference on machine learning, pages 8748-8763. PmLR, 2021."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/TMM.2024.3396272"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00806"},{"key":"e_1_3_2_1_13_1","article-title":"Resisting noise in pseudo labels: Audible video event parsing with evidential learning","author":"Jiang Xun","unstructured":"Xun Jiang, Xing Xu, Liqing Zhu, Zhe Sun, Andrzej Cichocki, and Heng Tao Shen. Resisting noise in pseudo labels: Audible video event parsing with evidential learning. IEEE Transactions on Neural Networks and Learning Systems.","journal-title":"IEEE Transactions on Neural Networks and Learning Systems."},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/3678957.3688386"},{"key":"e_1_3_2_1_15_1","volume-title":"Random forest. Journal of insurance medicine, 47(1):31-39","author":"Rigatti Steven J","year":"2017","unstructured":"Steven J Rigatti. Random forest. Journal of insurance medicine, 47(1):31-39, 2017."},{"key":"e_1_3_2_1_16_1","volume-title":"Lightgbm: A highly efficient gradient boosting decision tree. Advances in neural information processing systems, 30","author":"Ke Guolin","year":"2017","unstructured":"Guolin Ke, Qi Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and Tie-Yan Liu. Lightgbm: A highly efficient gradient boosting decision tree. Advances in neural information processing systems, 30, 2017."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447548.3467231"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3136755.3136785"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS47612.2022.9981726"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313831.3376372"},{"key":"e_1_3_2_1_21_1","first-page":"652","volume-title":"Proceedings of the 26th International Conference on Multimodal Interaction","author":"Spitale Micol","year":"2024","unstructured":"Micol Spitale, Maria Teresa Parreira, Maia Stiber, Minja Axelsson, Neval Kara, Garima Kankariya, Chien-Ming Huang, Malte Jung, Wendy Ju, and Hatice Gunes. Err@ hri 2024 challenge: Multimodal detection of errors and failures in human-robot interactions. In Proceedings of the 26th International Conference on Multimodal Interaction, pages 652-656, 2024."},{"key":"e_1_3_2_1_22_1","volume-title":"Wendy Ju, Micol Spitale, Hatice Gunes, and Chien-Ming Huang. Err@ hri 2.0 challenge: Multimodal detection of errors and failures in human-robot conversations. arXiv preprint arXiv:2507.13468","author":"Cao Shiye","year":"2025","unstructured":"Shiye Cao, Maia Stiber, Amama Mahmood, Maria Teresa Parreira, Wendy Ju, Micol Spitale, Hatice Gunes, and Chien-Ming Huang. Err@ hri 2.0 challenge: Multimodal detection of errors and failures in human-robot conversations. arXiv preprint arXiv:2507.13468, 2025."},{"key":"e_1_3_2_1_23_1","volume-title":"Ziang Xiao, Anqi Liu, and Chien-Ming Huang. Interruption handling for conversational robots. arXiv preprint arXiv:2501.01568","author":"Cao Shiye","year":"2025","unstructured":"Shiye Cao, Jiwon Moon, Amama Mahmood, Victor Nikhil Antony, Ziang Xiao, Anqi Liu, and Chien-Ming Huang. Interruption handling for conversational robots. arXiv preprint arXiv:2501.01568, 2025."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijhcs.2024.103406"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3678957.3688388"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS51168.2021.9636455"},{"key":"e_1_3_2_1_27_1","volume-title":"Reflect: Summarizing robot experiences for failure explanation and correction. arXiv preprint arXiv:2306.15724","author":"Liu Zeyi","year":"2023","unstructured":"Zeyi Liu, Arpit Bahety, and Shuran Song. Reflect: Summarizing robot experiences for failure explanation and correction. arXiv preprint arXiv:2306.15724, 2023."},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.robot.2023.104568"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/TASE.2023.3276856"},{"key":"e_1_3_2_1_30_1","volume-title":"Fam-hri: Foundation-model assisted multi-modal human-robot interaction combining gaze and speech. arXiv preprint arXiv:2503.16492","author":"Lai Yuzhi","year":"2025","unstructured":"Yuzhi Lai, Shenghai Yuan, Boya Zhang, Benjamin Kiefer, Peizheng Li, and Andreas Zell. Fam-hri: Foundation-model assisted multi-modal human-robot interaction combining gaze and speech. arXiv preprint arXiv:2503.16492, 2025."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.3390\/w9100796"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1002\/widm.1143"},{"key":"e_1_3_2_1_33_1","volume-title":"Github","author":"Oguiza Ignacio","year":"2022","unstructured":"Ignacio Oguiza. tsai-a state-of-the-art deep learning library for time series and sequential data. Github, 2022."},{"key":"e_1_3_2_1_34_1","volume-title":"et al. Scikit-learn: Machine learning in python. the Journal of machine Learning research, 12:2825-2830","author":"Pedregosa Fabian","year":"2011","unstructured":"Fabian Pedregosa, Ga\u00ebl Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, et al. Scikit-learn: Machine learning in python. the Journal of machine Learning research, 12:2825-2830, 2011."},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.3390\/s21051571"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.3389\/frobt.2019.00081"},{"key":"e_1_3_2_1_37_1","first-page":"1","volume-title":"Automation & Test in Europe Conference (DATE)","author":"Trivedi Amit Ranjan","year":"2025","unstructured":"Amit Ranjan Trivedi, Sina Tayebati, Hemant Kumawat, Nastaran Darabi, Divake Kumar, Adarsh Kumar Kosta, Yeshwanth Venkatesha, Dinithi Jayasuriya, Nethmi Jayasinghe, Priyadarshini Panda, et al. Intelligent sensing-to-action for robust autonomy at the edge: Opportunities and challenges. In 2025 Design, Automation & Test in Europe Conference (DATE), pages 1-10. IEEE, 2025."}],"event":{"name":"MM '25: The 33rd ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Dublin Ireland","acronym":"MM '25"},"container-title":["Proceedings of the 33rd ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3746027.3762075","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,9]],"date-time":"2025-12-09T19:37:31Z","timestamp":1765309051000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3746027.3762075"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,27]]},"references-count":37,"alternative-id":["10.1145\/3746027.3762075","10.1145\/3746027"],"URL":"https:\/\/doi.org\/10.1145\/3746027.3762075","relation":{},"subject":[],"published":{"date-parts":[[2025,10,27]]},"assertion":[{"value":"2025-10-27","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}