{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T04:44:38Z","timestamp":1781585078110,"version":"3.54.5"},"publisher-location":"New York, NY, USA","reference-count":61,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,11,7]],"date-time":"2022-11-07T00:00:00Z","timestamp":1667779200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100010661","name":"Horizon 2020 Framework Programme","doi-asserted-by":"publisher","award":["871245"],"award-info":[{"award-number":["871245"]}],"id":[{"id":"10.13039\/100010661","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,11,7]]},"DOI":"10.1145\/3536221.3556624","type":"proceedings-article","created":{"date-parts":[[2022,11,4]],"date-time":"2022-11-04T15:54:14Z","timestamp":1667577254000},"page":"420-431","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":24,"title":["Multimodal Across Domains Gaze Target Detection"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-1938-3449","authenticated-orcid":false,"given":"Francesco","family":"Tonini","sequence":"first","affiliation":[{"name":"Department of Information Engineering and Computer Science, University of Trento, Italy"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9583-0087","authenticated-orcid":false,"given":"Cigdem","family":"Beyan","sequence":"additional","affiliation":[{"name":"Department of Information Engineering and Computer Science, University of Trento, Italy"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0228-1147","authenticated-orcid":false,"given":"Elisa","family":"Ricci","sequence":"additional","affiliation":[{"name":"Department of Information Engineering and Computer Science, University of Trento, Italy and Fondazione Bruno Kessler, Italy"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2022,11,7]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/2993148.2993175"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.18"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01225-0_38"},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01228-1_24"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00544"},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.01247"},{"key":"e_1_3_2_2_7_1","volume-title":"Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision (WACV). 1181\u20131190","author":"Victor","unstructured":"Victor G. \u00a0Turrisi da Costa, Giacomo Zara, Paolo Rota, Thiago Oliveira-Santos, Nicu Sebe, Vittorio Murino, and Elisa Ricci. 2022. Dual-Head Contrastive Domain Adaptation for Video Action Recognition . In Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision (WACV). 1181\u20131190 . Victor G.\u00a0Turrisi da Costa, Giacomo Zara, Paolo Rota, Thiago Oliveira-Santos, Nicu Sebe, Vittorio Murino, and Elisa Ricci. 2022. Dual-Head Contrastive Domain Adaptation for Video Action Recognition. In Proceedings of the IEEE\/CVF Winter Conference on Applications of Computer Vision (WACV). 1181\u20131190."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/3317697.3325118"},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-009-0275-4"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01123"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW54120.2021.00249"},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/2578153.2578190"},{"key":"e_1_3_2_2_13_1","volume-title":"Domain-adversarial training of neural networks. The journal of machine learning research 17, 1","author":"Ganin Yaroslav","year":"2016","unstructured":"Yaroslav Ganin , Evgeniya Ustinova , Hana Ajakan , Pascal Germain , Hugo Larochelle , Fran\u00e7ois Laviolette , Mario Marchand , and Victor Lempitsky . 2016. Domain-adversarial training of neural networks. The journal of machine learning research 17, 1 ( 2016 ), 2096\u20132030. Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, Fran\u00e7ois Laviolette, Mario Marchand, and Victor Lempitsky. 2016. Domain-adversarial training of neural networks. The journal of machine learning research 17, 1 (2016), 2096\u20132030."},{"key":"e_1_3_2_2_14_1","volume-title":"Proceedings of the Asian Conference on Computer Vision.","author":"Guo Zidong","year":"2020","unstructured":"Zidong Guo , Zejian Yuan , Chong Zhang , Wanchao Chi , Yonggen Ling , and Shenghao Zhang . 2020 . Domain adaptation gaze estimation by embedding with prediction consistency . In Proceedings of the Asian Conference on Computer Vision. Zidong Guo, Zejian Yuan, Chong Zhang, Wanchao Chi, Yonggen Ling, and Shenghao Zhang. 2020. Domain adaptation gaze estimation by embedding with prediction consistency. In Proceedings of the Asian Conference on Computer Vision."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.96"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2016.7487708"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIM.2022.3205664"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.engappai.2022.104924"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2009.5459462"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00701"},{"key":"e_1_3_2_2_21_1","volume-title":"C (feb","author":"Li Xiao","year":"2017","unstructured":"Xiao Li , Min Fang , Ju-Jie Zhang , and Jinqiao Wu. 2017. Domain Adaptation from RGB-D to RGB Images. Signal Process. 131 , C (feb 2017 ), 27\u201335. https:\/\/doi.org\/10.1016\/j.sigpro.2016.07.018 10.1016\/j.sigpro.2016.07.018 Xiao Li, Min Fang, Ju-Jie Zhang, and Jinqiao Wu. 2017. Domain Adaptation from RGB-D to RGB Images. Signal Process. 131, C (feb 2017), 27\u201335. https:\/\/doi.org\/10.1016\/j.sigpro.2016.07.018"},{"key":"e_1_3_2_2_22_1","unstructured":"Yin Li Miao Liu and Jame Rehg. 2021. In the eye of the beholder: Gaze and actions in first person video. IEEE Transactions on Pattern Analysis and Machine Intelligence (2021).  Yin Li Miao Liu and Jame Rehg. 2021. In the eye of the beholder: Gaze and actions in first person video. IEEE Transactions on Pattern Analysis and Machine Intelligence (2021)."},{"key":"e_1_3_2_2_23_1","volume-title":"Asian Conference on Computer Vision. Springer, 35\u201350","author":"Lian Dongze","year":"2018","unstructured":"Dongze Lian , Zehao Yu , and Shenghua Gao . 2018 . Believe it or not, we know what you are looking at! . In Asian Conference on Computer Vision. Springer, 35\u201350 . Dongze Lian, Zehao Yu, and Shenghua Gao. 2018. Believe it or not, we know what you are looking at!. In Asian Conference on Computer Vision. Springer, 35\u201350."},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-10602-1_48"},{"key":"e_1_3_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.2999633"},{"key":"e_1_3_2_2_26_1","volume-title":"International conference on machine learning. PMLR, 97\u2013105","author":"Long Mingsheng","year":"2015","unstructured":"Mingsheng Long , Yue Cao , Jianmin Wang , and Michael Jordan . 2015 . Learning transferable features with deep adaptation networks . In International conference on machine learning. PMLR, 97\u2013105 . Mingsheng Long, Yue Cao, Jianmin Wang, and Michael Jordan. 2015. Learning transferable features with deep adaptation networks. In International conference on machine learning. PMLR, 97\u2013105."},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00359"},{"key":"e_1_3_2_2_28_1","volume-title":"Tracking gaze and visual focus of attention of people involved in social interaction","author":"Mass\u00e9 Beno\u00eet","year":"2017","unstructured":"Beno\u00eet Mass\u00e9 , Sil\u00e8ye Ba , and Radu Horaud . 2017. Tracking gaze and visual focus of attention of people involved in social interaction . IEEE transactions on pattern analysis and machine intelligence 40, 11( 2017 ), 2711\u20132724. Beno\u00eet Mass\u00e9, Sil\u00e8ye Ba, and Radu Horaud. 2017. Tracking gaze and visual focus of attention of people involved in social interaction. IEEE transactions on pattern analysis and machine intelligence 40, 11(2017), 2711\u20132724."},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/FG.2019.8756555"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV48630.2021.00111"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/3379155.3391332"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3212721.3212816"},{"key":"e_1_3_2_2_33_1","volume-title":"Unsupervised Human Action Recognition with Skeletal Graph Laplacian and Self-Supervised Viewpoints Invariance. In The 32nd British Machine Vision Conference (BMVC).","author":"Paoletti Giancarlo","year":"2021","unstructured":"Giancarlo Paoletti , Jacopo Cavazza , Cigdem Beyan , and Alessio Del\u00a0Bue . 2021 . Unsupervised Human Action Recognition with Skeletal Graph Laplacian and Self-Supervised Viewpoints Invariance. In The 32nd British Machine Vision Conference (BMVC). Giancarlo Paoletti, Jacopo Cavazza, Cigdem Beyan, and Alessio Del\u00a0Bue. 2021. Unsupervised Human Action Recognition with Skeletal Graph Laplacian and Self-Supervised Viewpoints Invariance. In The 32nd British Machine Vision Conference (BMVC)."},{"key":"e_1_3_2_2_34_1","volume-title":"Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer","author":"Ranftl Ren\u00e9","year":"2020","unstructured":"Ren\u00e9 Ranftl , Katrin Lasinger , David Hafner , Konrad Schindler , and Vladlen Koltun . 2020. Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer . IEEE transactions on pattern analysis and machine intelligence ( 2020 ). Ren\u00e9 Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun. 2020. Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer. IEEE transactions on pattern analysis and machine intelligence (2020)."},{"key":"e_1_3_2_2_35_1","volume-title":"Advances in Neural Information Processing Systems, Vol.\u00a028. Curran Associates","author":"Recasens Adria","unstructured":"Adria Recasens , Aditya Khosla , Carl Vondrick , and Antonio Torralba . 2015. Where are they looking? . In Advances in Neural Information Processing Systems, Vol.\u00a028. Curran Associates , Inc . Adria Recasens, Aditya Khosla, Carl Vondrick, and Antonio Torralba. 2015. Where are they looking?. In Advances in Neural Information Processing Systems, Vol.\u00a028. Curran Associates, Inc."},{"key":"e_1_3_2_2_36_1","volume-title":"Following Gaze in Video. In 2017 IEEE International Conference on Computer Vision (ICCV). 1444\u20131452","author":"Recasens Adri\u00e0","year":"2017","unstructured":"Adri\u00e0 Recasens , Carl Vondrick , Aditya Khosla , and Antonio Torralba . 2017 . Following Gaze in Video. In 2017 IEEE International Conference on Computer Vision (ICCV). 1444\u20131452 . https:\/\/doi.org\/10.1109\/ICCV.2017.160 10.1109\/ICCV.2017.160 Adri\u00e0 Recasens, Carl Vondrick, Aditya Khosla, and Antonio Torralba. 2017. Following Gaze in Video. In 2017 IEEE International Conference on Computer Vision (ICCV). 1444\u20131452. https:\/\/doi.org\/10.1109\/ICCV.2017.160"},{"key":"e_1_3_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/IROS.2014.6942680"},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2012.6225137"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i05.6431"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3462244.3479954"},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW53098.2021.00349"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2015.463"},{"key":"e_1_3_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.316"},{"key":"e_1_3_2_2_44_1","unstructured":"Jing Wang and Kuangen Zhang. 2019. Unsupervised domain adaptation learning algorithm for rgb-d staircase recognition. arXiv preprint arXiv:1903.01212(2019).  Jing Wang and Kuangen Zhang. 2019. Unsupervised domain adaptation learning algorithm for rgb-d staircase recognition. arXiv preprint arXiv:1903.01212(2019)."},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00840"},{"key":"e_1_3_2_2_46_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01252-6_9"},{"key":"e_1_3_2_2_47_1","volume-title":"Jointly Inferring Human Attention and Intentions in Complex Tasks. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6801\u20136809","author":"Wei Ping","year":"2018","unstructured":"Ping Wei , Yang Liu , Tianmin Shu , Nanning Zheng , and Song-Chun Zhu . 2018 . Where and Why are They Looking? Jointly Inferring Human Attention and Intentions in Complex Tasks. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6801\u20136809 . https:\/\/doi.org\/10.1109\/CVPR.2018.00711 10.1109\/CVPR.2018.00711 Ping Wei, Yang Liu, Tianmin Shu, Nanning Zheng, and Song-Chun Zhu. 2018. Where and Why are They Looking? Jointly Inferring Human Attention and Intentions in Complex Tasks. In 2018 IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 6801\u20136809. https:\/\/doi.org\/10.1109\/CVPR.2018.00711"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3400066"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-012-1220-z"},{"key":"e_1_3_2_2_50_1","volume-title":"Sun database: Large-scale scene recognition from abbey to zoo. In 2010 IEEE computer society conference on computer vision and pattern recognition","author":"Xiao Jianxiong","unstructured":"Jianxiong Xiao , James Hays , Krista\u00a0 A Ehinger , Aude Oliva , and Antonio Torralba . 2010. Sun database: Large-scale scene recognition from abbey to zoo. In 2010 IEEE computer society conference on computer vision and pattern recognition . IEEE , 3485\u20133492. Jianxiong Xiao, James Hays, Krista\u00a0A Ehinger, Aude Oliva, and Antonio Torralba. 2010. Sun database: Large-scale scene recognition from abbey to zoo. In 2010 IEEE computer society conference on computer vision and pattern recognition. IEEE, 3485\u20133492."},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2019.2949697"},{"key":"e_1_3_2_2_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00151"},{"key":"e_1_3_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3096553"},{"key":"e_1_3_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2011.6126386"},{"key":"e_1_3_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01221"},{"key":"e_1_3_2_2_56_1","doi-asserted-by":"publisher","DOI":"10.1145\/2663204.2663247"},{"key":"e_1_3_2_2_57_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3174198"},{"key":"e_1_3_2_2_58_1","volume-title":"Learning deep features for scene recognition using places database. Advances in neural information processing systems 27","author":"Zhou Bolei","year":"2014","unstructured":"Bolei Zhou , Agata Lapedriza , Jianxiong Xiao , Antonio Torralba , and Aude Oliva . 2014. Learning deep features for scene recognition using places database. Advances in neural information processing systems 27 ( 2014 ). Bolei Zhou, Agata Lapedriza, Jianxiong Xiao, Antonio Torralba, and Aude Oliva. 2014. Learning deep features for scene recognition using places database. Advances in neural information processing systems 27 (2014)."},{"key":"e_1_3_2_2_59_1","volume-title":"Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In 2017 IEEE International Conference on Computer Vision (ICCV). 2242\u20132251","author":"Zhu Jun-Yan","year":"2017","unstructured":"Jun-Yan Zhu , Taesung Park , Phillip Isola , and Alexei\u00a0 A. Efros . 2017 . Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In 2017 IEEE International Conference on Computer Vision (ICCV). 2242\u20132251 . https:\/\/doi.org\/10.1109\/ICCV.2017.244 10.1109\/ICCV.2017.244 Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei\u00a0A. Efros. 2017. Unpaired Image-to-Image Translation Using Cycle-Consistent Adversarial Networks. In 2017 IEEE International Conference on Computer Vision (ICCV). 2242\u20132251. https:\/\/doi.org\/10.1109\/ICCV.2017.244"},{"key":"e_1_3_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/TCSVT.2019.2940479"},{"key":"e_1_3_2_2_61_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-019-01234-9"}],"event":{"name":"ICMI '22: INTERNATIONAL CONFERENCE ON MULTIMODAL INTERACTION","location":"Bengaluru India","acronym":"ICMI '22","sponsor":["SIGCHI ACM Special Interest Group on Computer-Human Interaction"]},"container-title":["Proceedings of the 2022 International Conference on Multimodal Interaction"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3536221.3556624","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3536221.3556624","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T17:48:53Z","timestamp":1750182533000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3536221.3556624"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,11,7]]},"references-count":61,"alternative-id":["10.1145\/3536221.3556624","10.1145\/3536221"],"URL":"https:\/\/doi.org\/10.1145\/3536221.3556624","relation":{},"subject":[],"published":{"date-parts":[[2022,11,7]]},"assertion":[{"value":"2022-11-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}