{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T17:41:59Z","timestamp":1777657319079,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":62,"publisher":"ACM","license":[{"start":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T00:00:00Z","timestamp":1665360000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2022,10,10]]},"DOI":"10.1145\/3503161.3547828","type":"proceedings-article","created":{"date-parts":[[2022,10,10]],"date-time":"2022-10-10T15:42:35Z","timestamp":1665416555000},"page":"2452-2463","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["CycleHand: Increasing 3D Pose Estimation Ability on In-the-wild Monocular Image through Cyclic Flow"],"prefix":"10.1145","author":[{"given":"Daiheng","family":"Gao","sequence":"first","affiliation":[{"name":"XR Lab, Alibaba Group, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xindi","family":"Zhang","sequence":"additional","affiliation":[{"name":"Queen Mary University of London, London, United Kingdom"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xingyu","family":"Chen","sequence":"additional","affiliation":[{"name":"Xiaobing AI, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Andong","family":"Tan","sequence":"additional","affiliation":[{"name":"Technical University of Munich, Munich, Germany"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Bang","family":"Zhang","sequence":"additional","affiliation":[{"name":"XR Lab, Alibaba Group, Hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Pan","family":"Pan","sequence":"additional","affiliation":[{"name":"Alibaba Group, beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ping","family":"Tan","sequence":"additional","affiliation":[{"name":"XR Lab, Alibaba Group, hangzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,10,10]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01110"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01231-1_41"},{"key":"e_1_3_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.00951"},{"key":"e_1_3_2_2_4_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 5714--5724","author":"Chen Ching-Hang","year":"2019","unstructured":"Ching-Hang Chen , Ambrish Tyagi , Amit Agrawal , Dylan Drover , Stefan Stojanov , and James M Rehg . 2019 . Unsupervised 3d pose estimation with geometric selfsupervision . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 5714--5724 . Ching-Hang Chen, Ambrish Tyagi, Amit Agrawal, Dylan Drover, Stefan Stojanov, and James M Rehg. 2019. Unsupervised 3d pose estimation with geometric selfsupervision. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 5714--5724."},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01269"},{"key":"e_1_3_2_2_6_1","volume-title":"International conference on machine learning. PMLR, 1597--1607","author":"Chen Ting","year":"2020","unstructured":"Ting Chen , Simon Kornblith , Mohammad Norouzi , and Geoffrey Hinton . 2020 . A simple framework for contrastive learning of visual representations . In International conference on machine learning. PMLR, 1597--1607 . Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton. 2020. A simple framework for contrastive learning of visual representations. In International conference on machine learning. PMLR, 1597--1607."},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/3DV.2016.58"},{"key":"e_1_3_2_2_8_1","volume-title":"MobRecon: Mobile-Friendly Hand Mesh Reconstruction from Monocular Image. arXiv preprint arXiv:2112.02753","author":"Chen Xingyu","year":"2021","unstructured":"Xingyu Chen , Yufeng Liu , Yajiao Dong , Xiong Zhang , Chongyang Ma , Yanmin Xiong , Yuan Zhang , and Xiaoyan Guo . 2021. MobRecon: Mobile-Friendly Hand Mesh Reconstruction from Monocular Image. arXiv preprint arXiv:2112.02753 ( 2021 ). Xingyu Chen, Yufeng Liu, Yajiao Dong, Xiong Zhang, Chongyang Ma, Yanmin Xiong, Yuan Zhang, and Xiaoyan Guo. 2021. MobRecon: Mobile-Friendly Hand Mesh Reconstruction from Monocular Image. arXiv preprint arXiv:2112.02753 (2021)."},{"key":"e_1_3_2_2_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR46437.2021.01031"},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01240-3_41"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2019.00038"},{"key":"e_1_3_2_2_12_1","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0.","author":"Drover Dylan","year":"2018","unstructured":"Dylan Drover , Ching-Hang Chen , Amit Agrawal , Ambrish Tyagi , and Cong Phuoc Huynh . 2018 . Can 3d pose be learned from 2d projections alone? . In Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0. Dylan Drover, Ching-Hang Chen, Amit Agrawal, Ambrish Tyagi, and Cong Phuoc Huynh. 2018. Can 3d pose be learned from 2d projections alone?. In Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0."},{"key":"e_1_3_2_2_13_1","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0.","author":"Drover Dylan","year":"2018","unstructured":"Dylan Drover , Ching-Hang Chen , Amit Agrawal , Ambrish Tyagi , and Cong Phuoc Huynh . 2018 . Can 3d pose be learned from 2d projections alone? . In Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0. Dylan Drover, Ching-Hang Chen, Amit Agrawal, Ambrish Tyagi, and Cong Phuoc Huynh. 2018. Can 3d pose be learned from 2d projections alone?. In Proceedings of the European Conference on Computer Vision (ECCV) Workshops. 0--0."},{"key":"e_1_3_2_2_14_1","volume-title":"Int. Conf. Knowledge Discovery and Data Mining","volume":"240","author":"Ester Martin","year":"1996","unstructured":"Martin Ester , Hans-Peter Kriegel , J\u00f6rg Sander , and Xiaowei Xu . 1996 . Density-based spatial clustering of applications with noise . In Int. Conf. Knowledge Discovery and Data Mining , Vol. 240 . 6. Martin Ester, Hans-Peter Kriegel, J\u00f6rg Sander, and Xiaowei Xu. 1996. Density-based spatial clustering of applications with noise. In Int. Conf. Knowledge Discovery and Data Mining, Vol. 240. 6."},{"key":"e_1_3_2_2_15_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3450626.3459936","article-title":"Learning an animatable detailed 3D face model from in-the-wild images","volume":"40","author":"Feng Yao","year":"2021","unstructured":"Yao Feng , Haiwen Feng , Michael J Black , and Timo Bolkart . 2021 . Learning an animatable detailed 3D face model from in-the-wild images . ACM Transactions on Graphics (TOG) 40 , 4 (2021), 1 -- 13 . Yao Feng, Haiwen Feng, Michael J Black, and Timo Bolkart. 2021. Learning an animatable detailed 3D face model from in-the-wild images. ACM Transactions on Graphics (TOG) 40, 4 (2021), 1--13.","journal-title":"ACM Transactions on Graphics (TOG)"},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW54120.2021.00256"},{"key":"e_1_3_2_2_17_1","unstructured":"Liuhao Ge Zhou Ren Yuncheng Li Zehao Xue Yingying Wang Jianfei Cai and Junsong Yuan. 2019. 3D Hand Shape and Pose Estimation from a Single RGB Image. arXiv:1903.00812 [cs.CV]  Liuhao Ge Zhou Ren Yuncheng Li Zehao Xue Yingying Wang Jianfei Cai and Junsong Yuan. 2019. 3D Hand Shape and Pose Estimation from a Single RGB Image. arXiv:1903.00812 [cs.CV]"},{"key":"e_1_3_2_2_18_1","volume-title":"International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=Bygh9j09KX","author":"Geirhos Robert","year":"2019","unstructured":"Robert Geirhos , Patricia Rubisch , Claudio Michaelis , Matthias Bethge , Felix A Wichmann , and Wieland Brendel . 2019 . ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness .. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=Bygh9j09KX Robert Geirhos, Patricia Rubisch, Claudio Michaelis, Matthias Bethge, Felix A Wichmann, and Wieland Brendel. 2019. ImageNet-trained CNNs are biased towards texture; increasing shape bias improves accuracy and robustness.. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=Bygh9j09KX"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"crossref","unstructured":"Shreyas Hampali Mahdi Rad Markus Oberweger and Vincent Lepetit. 2020. HOnnotate: A method for 3D Annotation of Hand and Object Poses. arXiv:1907.01481 [cs.CV]  Shreyas Hampali Mahdi Rad Markus Oberweger and Vincent Lepetit. 2020. HOnnotate: A method for 3D Annotation of Hand and Object Poses. arXiv:1907.01481 [cs.CV]","DOI":"10.1109\/CVPR42600.2020.00326"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00065"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01208"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00975"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"crossref","unstructured":"Justin Johnson Nikhila Ravi Jeremy Reizenstein David Novotny Shubham Tulsiani Christoph Lassner and Steve Branson. 2020. Accelerating 3d deep learning with pytorch3d. In SIGGRAPH Asia 2020 Courses. 1--1.  Justin Johnson Nikhila Ravi Jeremy Reizenstein David Novotny Shubham Tulsiani Christoph Lassner and Steve Branson. 2020. Accelerating 3d deep learning with pytorch3d. In SIGGRAPH Asia 2020 Courses. 1--1.","DOI":"10.1145\/3415263.3419160"},{"key":"e_1_3_2_2_25_1","volume-title":"Kingma and Jimmy Ba","author":"Diederik","year":"2017","unstructured":"Diederik P. Kingma and Jimmy Ba . 2017 . Adam : A Method for Stochastic Optimization . arXiv:1412.6980 [cs.LG] Diederik P. Kingma and Jimmy Ba. 2017. Adam: A Method for Stochastic Optimization. arXiv:1412.6980 [cs.LG]"},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073599"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.00117"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i07.6808"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01270"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00780"},{"key":"e_1_3_2_2_31_1","volume-title":"HandTailor: Towards High-Precision Monocular 3D Hand Recovery. arXiv preprint arXiv:2102.09244","author":"Lv Jun","year":"2021","unstructured":"Jun Lv , Wenqiang Xu , Lixin Yang , Sucheng Qian , Chongzhao Mao , and Cewu Lu. 2021. HandTailor: Towards High-Precision Monocular 3D Hand Recovery. arXiv preprint arXiv:2102.09244 ( 2021 ). Jun Lv, Wenqiang Xu, Lixin Yang, Sucheng Qian, Chongzhao Mao, and Cewu Lu. 2021. HandTailor: Towards High-Precision Monocular 3D Hand Recovery. arXiv preprint arXiv:2102.09244 (2021)."},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3072959.3073596"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58571-6_44"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00778"},{"key":"e_1_3_2_2_35_1","unstructured":"Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in pytorch. (2017).  Adam Paszke Sam Gross Soumith Chintala Gregory Chanan Edward Yang Zachary DeVito Zeming Lin Alban Desmaison Luca Antiga and Adam Lerer. 2017. Automatic differentiation in pytorch. (2017)."},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Dario Pavllo Christoph Feichtenhofer David Grangier and MichaelAuli. 2019. 3D human pose estimation in video with temporal convolutions and semi-supervised training. arXiv:1811.11742 [cs.CV]  Dario Pavllo Christoph Feichtenhofer David Grangier and MichaelAuli. 2019. 3D human pose estimation in video with temporal convolutions and semi-supervised training. arXiv:1811.11742 [cs.CV]","DOI":"10.1109\/CVPR.2019.00794"},{"key":"e_1_3_2_2_37_1","volume-title":"Drones: Self-Supervised Active Triangulation for 3D Human Pose Reconstruction. In NeurIPS.","author":"Pirinen Aleksis","year":"2019","unstructured":"Aleksis Pirinen , Erik G\u00e4rtner , and C. Sminchisescu . 2019 . Domes to Drones: Self-Supervised Active Triangulation for 3D Human Pose Reconstruction. In NeurIPS. Aleksis Pirinen, Erik G\u00e4rtner, and C. Sminchisescu. 2019. Domes to Drones: Self-Supervised Active Triangulation for 3D Human Pose Reconstruction. In NeurIPS."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"crossref","unstructured":"Neng Qian Jiayi Wang Franziska Mueller Florian Bernard Vladislav Golyanik and Christian Theobalt. 2020. Parametric Hand Texture Model for 3D Hand Reconstruction and Personalization. (2020).  Neng Qian Jiayi Wang Franziska Mueller Florian Bernard Vladislav Golyanik and Christian Theobalt. 2020. Parametric Hand Texture Model for 3D Hand Reconstruction and Personalization. (2020).","DOI":"10.1007\/978-3-030-58621-8_4"},{"key":"e_1_3_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-01249-6_46"},{"key":"e_1_3_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.1145\/3130800.3130883"},{"key":"e_1_3_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW54120.2021.00201"},{"key":"e_1_3_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/1186562.1015720"},{"key":"e_1_3_2_2_43_1","volume-title":"The anatomy and mechanics of the human hand. Artificial limbs 2, 2","author":"Schwarz Robert J","year":"1955","unstructured":"Robert J Schwarz and C Taylor . 1955. The anatomy and mechanics of the human hand. Artificial limbs 2, 2 ( 1955 ), 22--35. Robert J Schwarz and C Taylor. 1955. The anatomy and mechanics of the human hand. Artificial limbs 2, 2 (1955), 22--35."},{"key":"e_1_3_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.1109\/3DV53792.2021.00013"},{"key":"e_1_3_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR42600.2020.00989"},{"key":"e_1_3_2_2_46_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Very Deep Convolutional Networks for Large-Scale Image Recognition. http:\/\/arxiv.org\/abs\/1409.1556 cite arxiv:1409.1556.  Karen Simonyan and Andrew Zisserman. 2014. Very Deep Convolutional Networks for Large-Scale Image Recognition. http:\/\/arxiv.org\/abs\/1409.1556 cite arxiv:1409.1556."},{"key":"e_1_3_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01104"},{"key":"e_1_3_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-58520-4_13"},{"key":"e_1_3_2_2_49_1","doi-asserted-by":"publisher","DOI":"10.1080\/10867651.2004.10487596"},{"key":"e_1_3_2_2_50_1","unstructured":"Ilya Tolstikhin Neil Houlsby Alexander Kolesnikov Lucas Beyer Xiaohua Zhai Thomas Unterthiner Jessica Yung Daniel Keysers Jakob Uszkoreit Mario Lucic etal 2021. MLP-Mixer: An all-MLP architecture for vision. arXiv preprint arXiv:2105.01601 (2021).  Ilya Tolstikhin Neil Houlsby Alexander Kolesnikov Lucas Beyer Xiaohua Zhai Thomas Unterthiner Jessica Yung Daniel Keysers Jakob Uszkoreit Mario Lucic et al. 2021. MLP-Mixer: An all-MLP architecture for vision. arXiv preprint arXiv:2105.01601 (2021)."},{"key":"e_1_3_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2017.492"},{"key":"e_1_3_2_2_52_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 7782--7791","author":"Bodo Rosenhahn BastianWandt","year":"2019","unstructured":"BastianWandt and Bodo Rosenhahn . 2019 . Repnet:Weakly supervised training of an adversarial reprojection network for 3d human pose estimation . In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 7782--7791 . BastianWandt and Bodo Rosenhahn. 2019. Repnet:Weakly supervised training of an adversarial reprojection network for 3d human pose estimation. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 7782--7791."},{"key":"e_1_3_2_2_53_1","unstructured":"Yi Wang Xin Tao Xiaojuan Qi Xiaoyong Shen and Jiaya Jia. 2018. Image Inpainting via Generative Multi-column Convolutional Neural Networks. In Advances in Neural Information Processing Systems. 331--340.  Yi Wang Xin Tao Xiaojuan Qi Xiaoyong Shen and Jiaya Jia. 2018. Image Inpainting via Generative Multi-column Convolutional Neural Networks. In Advances in Neural Information Processing Systems. 331--340."},{"key":"e_1_3_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00551"},{"key":"e_1_3_2_2_55_1","volume-title":"Mediapipe hands: On-device real-time hand tracking. arXiv preprint arXiv:2006.10214","author":"Zhang Fan","year":"2020","unstructured":"Fan Zhang , Valentin Bazarevsky , Andrey Vakunov , Andrei Tkachenka , George Sung , Chuo-Ling Chang , and Matthias Grundmann . 2020. Mediapipe hands: On-device real-time hand tracking. arXiv preprint arXiv:2006.10214 ( 2020 ). Fan Zhang, Valentin Bazarevsky, Andrey Vakunov, Andrei Tkachenka, George Sung, Chuo-Ling Chang, and Matthias Grundmann. 2020. Mediapipe hands: On-device real-time hand tracking. arXiv preprint arXiv:2006.10214 (2020)."},{"key":"e_1_3_2_2_56_1","unstructured":"Jiawei Zhang Jianbo Jiao Mingliang Chen Liangqiong Qu Xiaobin Xu and Qingxiong Yang. 2016. 3D Hand Pose Tracking and Estimation Using Stereo Matching. arXiv:1610.07214 [cs.CV]  Jiawei Zhang Jianbo Jiao Mingliang Chen Liangqiong Qu Xiaobin Xu and Qingxiong Yang. 2016. 3D Hand Pose Tracking and Estimation Using Stereo Matching. arXiv:1610.07214 [cs.CV]"},{"key":"e_1_3_2_2_57_1","first-page":"2408","article-title":"Inference stage optimization for cross-scenario 3d human pose estimation","volume":"33","author":"Zhang Jianfeng","year":"2020","unstructured":"Jianfeng Zhang , Xuecheng Nie , and Jiashi Feng . 2020 . Inference stage optimization for cross-scenario 3d human pose estimation . Advances in Neural Information Processing Systems 33 (2020), 2408 -- 2419 . Jianfeng Zhang, Xuecheng Nie, and Jiashi Feng. 2020. Inference stage optimization for cross-scenario 3d human pose estimation. Advances in Neural Information Processing Systems 33 (2020), 2408--2419.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_2_58_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV48922.2021.01109"},{"key":"e_1_3_2_2_59_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.51"},{"key":"e_1_3_2_2_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2017.51"},{"key":"e_1_3_2_2_61_1","volume-title":"Contrastive Representation Learning for Hand Shape Estimation. In DAGM German Conference on Pattern Recognition. Springer, 250--264","author":"Zimmermann Christian","year":"2021","unstructured":"Christian Zimmermann , Max Argus , and Thomas Brox . 2021 . Contrastive Representation Learning for Hand Shape Estimation. In DAGM German Conference on Pattern Recognition. Springer, 250--264 . Christian Zimmermann, Max Argus, and Thomas Brox. 2021. Contrastive Representation Learning for Hand Shape Estimation. In DAGM German Conference on Pattern Recognition. Springer, 250--264."},{"key":"e_1_3_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00090"}],"event":{"name":"MM '22: The 30th ACM International Conference on Multimedia","location":"Lisboa Portugal","acronym":"MM '22","sponsor":["SIGMM ACM Special Interest Group on Multimedia"]},"container-title":["Proceedings of the 30th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547828","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3503161.3547828","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T19:02:35Z","timestamp":1750186955000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3503161.3547828"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,10,10]]},"references-count":62,"alternative-id":["10.1145\/3503161.3547828","10.1145\/3503161"],"URL":"https:\/\/doi.org\/10.1145\/3503161.3547828","relation":{},"subject":[],"published":{"date-parts":[[2022,10,10]]},"assertion":[{"value":"2022-10-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}