{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:08:08Z","timestamp":1750306088646,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":26,"publisher":"ACM","license":[{"start":{"date-parts":[[2017,10,27]],"date-time":"2017-10-27T00:00:00Z","timestamp":1509062400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Singapore Agency for Science Technology and Research (A*STAR)","award":["ARAP Program"],"award-info":[{"award-number":["ARAP Program"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2017,10,27]]},"DOI":"10.1145\/3132515.3132517","type":"proceedings-article","created":{"date-parts":[[2017,10,23]],"date-time":"2017-10-23T12:28:38Z","timestamp":1508761718000},"page":"47-53","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Robust Multi-Modal Cues for Dyadic Human Interaction Recognition"],"prefix":"10.1145","author":[{"given":"Rim","family":"Trabelsi","sequence":"first","affiliation":[{"name":"Advanced Digital Sciences Center &amp; National Engineering School of Gabes &amp; University of Tunis El Manar, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jagannadan","family":"Varadarajan","sequence":"additional","affiliation":[{"name":"Advanced Digital Sciences Center, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yong","family":"Pei","sequence":"additional","affiliation":[{"name":"SAP Asia Pte Ltd, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Le","family":"Zhang","sequence":"additional","affiliation":[{"name":"Advanced Digital Sciences Center, Singapore, Singapore"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Issam","family":"Jabri","sequence":"additional","affiliation":[{"name":"National Engineering School of Gabes &amp; Al Yamamah University, Gabes, Tunisia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ammar","family":"Bouallegue","sequence":"additional","affiliation":[{"name":"University of Tunis El Manar, Tunis, Tunisia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pierre","family":"Moulin","sequence":"additional","affiliation":[{"name":"University of Illinois at Urbana-Champaign, Champaign, IL, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2017,10,27]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW).","author":"Calderara Simone","year":"2012","unstructured":"Simone Calderara and Rita Cucchiara . 2012 . Understanding dyadic interactions applying proxemic theory on video surveillance trajectories . In IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW). Simone Calderara and Rita Cucchiara. 2012. Understanding dyadic interactions applying proxemic theory on video surveillance trajectories. In IEEE Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)."},{"key":"e_1_3_2_1_2_1","volume-title":"IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","author":"Du Yong","year":"2015","unstructured":"Yong Du , Wei Wang , and Liang Wang . 2015 . Hierarchical recurrent neural network for skeleton based action recognition . In IEEE Conference on Computer Vision and Pattern Recognition (CVPR). Yong Du, Wei Wang, and Liang Wang. 2015. Hierarchical recurrent neural network for skeleton based action recognition. In IEEE Conference on Computer Vision and Pattern Recognition (CVPR)."},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2015.7353446"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPR.2014.772"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.5555\/1390681.1442794"},{"key":"e_1_3_2_1_6_1","volume-title":"Interactive phrases: Semantic descriptions for human interaction recognition","author":"Fu Yun","year":"2014","unstructured":"Yun Fu , Yunde Jia , and Yu Kong . 2014. Interactive phrases: Semantic descriptions for human interaction recognition . IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) ( 2014 ). Yun Fu, Yunde Jia, and Yu Kong. 2014. Interactive phrases: Semantic descriptions for human interaction recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) (2014)."},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2007.70711"},{"key":"e_1_3_2_1_8_1","volume-title":"A system for the notation of proxemic behavior. American anthropologist 65, 5","author":"Hall Edward T","year":"1963","unstructured":"Edward T Hall . 1963. A system for the notation of proxemic behavior. American anthropologist 65, 5 ( 1963 ), 1003--1026. Edward T Hall. 1963. A system for the notation of proxemic behavior. American anthropologist 65, 5 (1963), 1003--1026."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7299172"},{"key":"e_1_3_2_1_10_1","volume-title":"Rgbd-hudaact: A color-depth video database for human daily activity recognition. In Consumer Depth Cameras for Computer Vision.","author":"Ni Bingbing","year":"2013","unstructured":"Bingbing Ni , Gang Wang , and Pierre Moulin . 2013 . Rgbd-hudaact: A color-depth video database for human daily activity recognition. In Consumer Depth Cameras for Computer Vision. Bingbing Ni, Gang Wang, and Pierre Moulin. 2013. Rgbd-hudaact: A color-depth video database for human daily activity recognition. In Consumer Depth Cameras for Computer Vision."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2013.76"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2013.98"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPAMI.2012.24"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.529"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-33885-4_12"},{"volume-title":"Spatio-Temporal Relationship Match: Video Structure Comparison for Recognition of Complex Human Activities. In IEEE International Conference on Computer Vision (ICCV).","author":"Ryoo M. S.","key":"e_1_3_2_1_16_1","unstructured":"M. S. Ryoo and J. K. Aggarwal . 2009 . Spatio-Temporal Relationship Match: Video Structure Comparison for Recognition of Complex Human Activities. In IEEE International Conference on Computer Vision (ICCV). M. S. Ryoo and J. K. Aggarwal. 2009. Spatio-Temporal Relationship Match: Video Structure Comparison for Recognition of Complex Human Activities. In IEEE International Conference on Computer Vision (ICCV)."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.5555\/1018429.1020906"},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.115"},{"key":"e_1_3_2_1_19_1","volume-title":"Neural Information Processing Systems Conference (NIPS).","author":"Simonyan Karen","year":"2014","unstructured":"Karen Simonyan and Andrew Zisserman . 2014 . Two-stream convolutional networks for action recognition in videos . In Neural Information Processing Systems Conference (NIPS). Karen Simonyan and Andrew Zisserman. 2014. Two-stream convolutional networks for action recognition in videos. In Neural Information Processing Systems Conference (NIPS)."},{"key":"e_1_3_2_1_20_1","unstructured":"K. Soomro A. Roshan Zamir and M. Shah. 2012. UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild. In CRCV-TR-12-01.  K. Soomro A. Roshan Zamir and M. Shah. 2012. UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild. In CRCV-TR-12-01."},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.82"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2013.441"},{"key":"e_1_3_2_1_23_1","volume-title":"Towards good practices for very deep two-stream convnets. arXiv preprint arXiv:1507.02159","author":"Wang Limin","year":"2015","unstructured":"Limin Wang , Yuanjun Xiong , Zhe Wang , and Yu Qiao . 2015. Towards good practices for very deep two-stream convnets. arXiv preprint arXiv:1507.02159 ( 2015 ). Limin Wang, Yuanjun Xiong, Zhe Wang, and Yu Qiao. 2015. Towards good practices for very deep two-stream convnets. arXiv preprint arXiv:1507.02159 (2015)."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.108"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2012.6239234"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.5555\/1771530.1771554"}],"event":{"name":"MM '17: ACM Multimedia Conference","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Mountain View California USA","acronym":"MM '17"},"container-title":["Proceedings of the Workshop on Multimodal Understanding of Social, Affective and Subjective Attributes"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3132515.3132517","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3132515.3132517","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:30:19Z","timestamp":1750217419000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3132515.3132517"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2017,10,27]]},"references-count":26,"alternative-id":["10.1145\/3132515.3132517","10.1145\/3132515"],"URL":"https:\/\/doi.org\/10.1145\/3132515.3132517","relation":{},"subject":[],"published":{"date-parts":[[2017,10,27]]},"assertion":[{"value":"2017-10-27","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}