{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,2]],"date-time":"2025-12-02T19:52:13Z","timestamp":1764705133455,"version":"3.46.0"},"reference-count":52,"publisher":"Association for Computing Machinery (ACM)","issue":"4","funder":[{"name":"National Science Fund for Distinguished Young Scholars of China","award":["62125203"],"award-info":[{"award-number":["62125203"]}]},{"name":"Natural Science Foundation for Excellent Young Scholars of Jiangsu Province","award":["BK20220105"],"award-info":[{"award-number":["BK20220105"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation","doi-asserted-by":"crossref","award":["62172236, 62572253"],"award-info":[{"award-number":["62172236, 62572253"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"China Postdoctoral Special Grant Foundation","award":["2024T170432"],"award-info":[{"award-number":["2024T170432"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Interact. Mob. Wearable Ubiquitous Technol."],"published-print":{"date-parts":[[2025,12,2]]},"abstract":"<jats:p>Facial landmarks provide essential representations of facial states and movements, serving as the foundation for numerous face-related tasks. However, traditional optical device-based facial landmark detection (FLD) solutions suffer from limitations in low-light conditions, occlusion sensitivity, and privacy concerns. In this paper, we propose an efficient two-stage facial landmark detection system, CF-FLD, which utilizes millimeter-wave (mmWave) radar signals to reconstruct human faces. Specifically, CF-FLD is composed of coarse-grained affine transformation (CAT) and fine-grained offset transformation (FOT). To characterize large-scale and rigid facial movements caused by head poses and joint motions, CAT defines sparse and representative triangle constraints within and across different facial parts for affine transformation. Based on CAT results, FOT is presented to progressively obtain offset shifts from subtle and non-rigid facial deformations. Instead of resource-intensive area detection or search, FOT designs a multi-level region partition strategy, in which region-wise hybrid network and region-aware attention are constructed to hierarchically refine facial landmarks. Comprehensive evaluations on data collected from 20 participants in real-world environments demonstrate that CF-FLD can accurately localize facial landmarks during facial motion, achieving a mean absolute error (MAE) of 2.02 mm and a normalized mean error (NME) of 3.37% at a low cost.<\/jats:p>","DOI":"10.1145\/3770656","type":"journal-article","created":{"date-parts":[[2025,12,2]],"date-time":"2025-12-02T19:42:32Z","timestamp":1764704552000},"page":"1-25","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["From Coarse to Fine: Fast and Effortless Facial Landmark Detection via mmWave Signals"],"prefix":"10.1145","volume":"9","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-5006-3822","authenticated-orcid":false,"given":"Biyun","family":"Sheng","sequence":"first","affiliation":[{"name":"Nanjing University of Posts and Telecommunications, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-2821-3801","authenticated-orcid":false,"given":"Shuqi","family":"Sun","sequence":"additional","affiliation":[{"name":"Nanjing University of Posts and Telecommunications, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8438-3843","authenticated-orcid":false,"given":"Hui","family":"Cai","sequence":"additional","affiliation":[{"name":"Nanjing University of Posts and Telecommunications, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9035-4539","authenticated-orcid":false,"given":"Chen","family":"Dai","sequence":"additional","affiliation":[{"name":"Nanjing University of Posts and Telecommunications, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1815-2793","authenticated-orcid":false,"given":"Fu","family":"Xiao","sequence":"additional","affiliation":[{"name":"Nanjing University of Posts and Telecommunications, Nanjing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,12,2]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1134\/S105466181104002X"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3489517.3530522"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACSSC.2003.1291865"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/3503161.3548262"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1155\/2021\/5383573"},{"key":"e_1_2_1_6_1","first-page":"378","article-title":"WiFace: Facial expression recognition using Wi-Fi signals","volume":"21","author":"Chen Yanjiao","year":"2020","unstructured":"Yanjiao Chen, Runmin Ou, Zhiyang Li, and Kaishun Wu. 2020. WiFace: Facial expression recognition using Wi-Fi signals. IEEE Transactions on Mobile Computing 21, 1 (2020), 378\u2013391.","journal-title":"IEEE Transactions on Mobile Computing"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.927467"},{"key":"e_1_2_1_8_1","volume-title":"Active shape models-their training and application. Computer vision and image understanding 61, 1","author":"Cootes Timothy F","year":"1995","unstructured":"Timothy F Cootes, Christopher J Taylor, David H Cooper, and Jim Graham. 1995. Active shape models-their training and application. Computer vision and image understanding 61, 1 (1995), 38\u201359."},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.5244\/C.20.95"},{"key":"e_1_2_1_10_1","unstructured":"Vivek Dham. 2017. Programming chirp parameters in TI radar devices. Application Report SWRA553 Texas Instruments 1457 (2017)."},{"key":"e_1_2_1_11_1","unstructured":"Alexey Dosovitskiy Lucas Beyer Alexander Kolesnikov Dirk Weissenborn Xiaohua Zhai Thomas Unterthiner Mostafa Dehghani Matthias Minderer Georg Heigold Sylvain Gelly et al. 2020. An image is worth 16x16 words: Transformers for image recognition at scale. arXiv preprint arXiv:2010.11929 (2020)."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3659619"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2016.90"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICAIIC48513.2020.9065010"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijforecast.2006.03.001"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ipm.2024.103919"},{"key":"e_1_2_1_17_1","volume-title":"Rodar: Robust gesture recognition based on mmWave radar under human activity interference","author":"Jin Can","year":"2024","unstructured":"Can Jin, Xiangzhu Meng, Xuanheng Li, Jie Wang, Miao Pan, and Yuguang Fang. 2024. Rodar: Robust gesture recognition based on mmWave radar under human activity interference. IEEE Transactions on Mobile Computing (2024)."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2014.241"},{"key":"e_1_2_1_19_1","unstructured":"Davis E King. 2022. A toolkit for making real world machine learning and data analysis applications in C++."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCS60910.2024.00102"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/WACV56688.2023.00567"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52688.2022.00414"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3613904.3642613"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2021.3137387"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/3699739"},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3494985"},{"key":"e_1_2_1_27_1","volume-title":"Mobilevit: light-weight, general-purpose, and mobile-friendly vision transformer. arXiv preprint arXiv:2110.02178","author":"Mehta Sachin","year":"2021","unstructured":"Sachin Mehta and Mohammad Rastegari. 2021. Mobilevit: light-weight, general-purpose, and mobile-friendly vision transformer. arXiv preprint arXiv:2110.02178 (2021)."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3699772"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","unstructured":"Luoyu Mei Shuai Wang Yun Cheng Ruofeng Liu Zhimeng Yin Wenchao Jiang Shuai Wang and Wei Gong. 2024. ESP-PCT: enhanced VR semantic performance through efficient compression of temporal and spatial redundancies in point cloud transformers (IJCAI '24). Article 131 9 pages. doi:10.24963\/ijcai.2024\/131","DOI":"10.24963\/ijcai.2024\/131"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2024.3352890"},{"key":"e_1_2_1_31_1","unstructured":"S Rao. 2017. Mimo radar-application report swra554a. Tech. Rep (2017)."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3534605"},{"key":"e_1_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2015.7298876"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00474"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2022.3151101"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.00162"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW.2018.00246"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICIP40778.2020.9190989"},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/3664647.3681546"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1109\/COMST.2024.3398004"},{"key":"e_1_2_1_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/3699751"},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2018.00813"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-19778-9_10"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447993.3483252"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-46448-0_4"},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1145\/3581791.3596839"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1145\/3458864.3467679"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSAI48974.2019.9010478"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/3614438"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3570361.3592515"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/3643832.3661855"},{"key":"e_1_2_1_52_1","volume-title":"Chen Change Loy, and Xiaoou Tang","author":"Zhang Zhanpeng","year":"2014","unstructured":"Zhanpeng Zhang, Ping Luo, Chen Change Loy, and Xiaoou Tang. 2014. Facial landmark detection by deep multi-task learning. In Computer Vision-ECCV 2014: 13th European Conference, Zurich, Switzerland, September 6-12, 2014, Proceedings, Part VI 13. Springer, 94\u2013108."}],"container-title":["Proceedings of the ACM on Interactive, Mobile, Wearable and Ubiquitous Technologies"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3770656","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,12,2]],"date-time":"2025-12-02T19:48:00Z","timestamp":1764704880000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3770656"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,2]]},"references-count":52,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2025,12,2]]}},"alternative-id":["10.1145\/3770656"],"URL":"https:\/\/doi.org\/10.1145\/3770656","relation":{},"ISSN":["2474-9567"],"issn-type":[{"value":"2474-9567","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,12,2]]},"assertion":[{"value":"2025-12-02","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}