{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T04:20:35Z","timestamp":1750220435021,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":38,"publisher":"ACM","license":[{"start":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T00:00:00Z","timestamp":1602460800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"The National Key R & D program of China","award":["2018YFB1307102"],"award-info":[{"award-number":["2018YFB1307102"]}]},{"name":"National Natural Science Foundation of China","award":["61727809"],"award-info":[{"award-number":["61727809"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2020,10,12]]},"DOI":"10.1145\/3394171.3413547","type":"proceedings-article","created":{"date-parts":[[2020,10,12]],"date-time":"2020-10-12T12:27:38Z","timestamp":1602505658000},"page":"2991-2998","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["Exploiting Self-Supervised and Semi-Supervised Learning for Facial Landmark Tracking with Unlabeled Data"],"prefix":"10.1145","author":[{"given":"Shi","family":"Yin","sequence":"first","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Shangfei","family":"Wang","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaoping","family":"Chen","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Enhong","family":"Chen","sequence":"additional","affiliation":[{"name":"University of Science and Technology of China, Hefei, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2020,10,12]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"crossref","unstructured":"Adrian Bulat and Georgios Tzimiropoulos. 2017. How far are we from solving the 2d & 3d face alignment problem?(and a dataset of 230 000 3d facial landmarks). In ICCV. 1021--1030.  Adrian Bulat and Georgios Tzimiropoulos. 2017. How far are we from solving the 2d & 3d face alignment problem?(and a dataset of 230 000 3d facial landmarks). In ICCV. 1021--1030.","DOI":"10.1109\/ICCV.2017.116"},{"key":"e_1_3_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-013-0667-3"},{"key":"e_1_3_2_2_3_1","unstructured":"Lisha Chen Hui Su and Qiang Ji. 2019 a. Deep Structured Prediction for Facial Landmark Detection. In NeurIPS.  Lisha Chen Hui Su and Qiang Ji. 2019 a. Deep Structured Prediction for Facial Landmark Detection. In NeurIPS."},{"key":"e_1_3_2_2_4_1","doi-asserted-by":"crossref","unstructured":"Ting Chen Xiaohua Zhai Marvin Ritter Mario Lucic and Neil Houlsby. 2019 b. Self-Supervised GANs via Auxiliary Rotation Loss. In CVPR. 12154--12163.  Ting Chen Xiaohua Zhai Marvin Ritter Mario Lucic and Neil Houlsby. 2019 b. Self-Supervised GANs via Auxiliary Rotation Loss. In CVPR. 12154--12163.","DOI":"10.1109\/CVPR.2019.01243"},{"key":"e_1_3_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-017-0999-5"},{"key":"e_1_3_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.imavis.2010.12.001"},{"key":"e_1_3_2_2_7_1","doi-asserted-by":"crossref","unstructured":"Xuanyi Dong and Yi Yang. 2019. Teacher Supervises Students How to Learn From Partially Labeled Images for Facial Landmark Detection. In ICCV.  Xuanyi Dong and Yi Yang. 2019. Teacher Supervises Students How to Learn From Partially Labeled Images for Facial Landmark Detection. In ICCV.","DOI":"10.1109\/ICCV.2019.00087"},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"crossref","unstructured":"Xuanyi Dong Shoou-I Yu Xinshuo Weng Shih-En Wei Yi Yang and Yaser Sheikh. 2018. Supervision-by-Registration: An Unsupervised Approach to Improve the Precision of Facial Landmark Detectors. In CVPR. 360--368.  Xuanyi Dong Shoou-I Yu Xinshuo Weng Shih-En Wei Yi Yang and Yaser Sheikh. 2018. Supervision-by-Registration: An Unsupervised Approach to Improve the Precision of Facial Landmark Detectors. In CVPR. 360--368.","DOI":"10.1109\/CVPR.2018.00045"},{"key":"e_1_3_2_2_9_1","unstructured":"FGNET. 2014. Talking Face Video. http:\/\/www-prima.inrialpes.fr\/FGnet\/data\/01-TalkingFace\/talking_face.html.  FGNET. 2014. Talking Face Video. http:\/\/www-prima.inrialpes.fr\/FGnet\/data\/01-TalkingFace\/talking_face.html."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"crossref","unstructured":"Quan Gan Siqi Nie Shangfei Wang and Qiang Ji. 2017. Differentiating Between Posed and Spontaneous Expressions with Latent Regression Bayesian Network. In AAAI. 4039--4045.  Quan Gan Siqi Nie Shangfei Wang and Qiang Ji. 2017. Differentiating Between Posed and Spontaneous Expressions with Latent Regression Bayesian Network. In AAAI. 4039--4045.","DOI":"10.1609\/aaai.v31i1.11225"},{"key":"e_1_3_2_2_11_1","doi-asserted-by":"crossref","unstructured":"Sina Honari Pavlo Molchanov Stephen Tyree Pascal Vincent Christopher J. Pal and Jan Kautz. 2018. Improving Landmark Localization With Semi-Supervised Learning. In CVPR. 1546--1555.  Sina Honari Pavlo Molchanov Stephen Tyree Pascal Vincent Christopher J. Pal and Jan Kautz. 2018. Improving Landmark Localization With Semi-Supervised Learning. In CVPR. 1546--1555.","DOI":"10.1109\/CVPR.2018.00167"},{"key":"e_1_3_2_2_12_1","unstructured":"Junlin Hu Jiwen Lu and Yap-Peng Tan. 2014. Discriminative Deep Metric Learning for Face Verification in the Wild. In CVPR. 1875--1882.  Junlin Hu Jiwen Lu and Yap-Peng Tan. 2014. Discriminative Deep Metric Learning for Face Verification in the Wild. In CVPR. 1875--1882."},{"key":"e_1_3_2_2_13_1","doi-asserted-by":"crossref","unstructured":"Zhiwu Huang Xiaowei Zhao Shiguang Shan Ruiping Wang and Xilin Chen. 2013. Coupling alignments with recognition for still-to-video face recognition. In ICCV. 3296--3303.  Zhiwu Huang Xiaowei Zhao Shiguang Shan Ruiping Wang and Xilin Chen. 2013. Coupling alignments with recognition for still-to-video face recognition. In ICCV. 3296--3303.","DOI":"10.1109\/ICCV.2013.409"},{"key":"e_1_3_2_2_14_1","doi-asserted-by":"crossref","unstructured":"Xuhui Jia Heng Yang Kwok-Ping Chan and ioannis Patras. 2014. Structured Semi-supervised Forest for Facial Landmarks Localization with Face Mask Reasoning. In BMVC.  Xuhui Jia Heng Yang Kwok-Ping Chan and ioannis Patras. 2014. Structured Semi-supervised Forest for Facial Landmarks Localization with Face Mask Reasoning. In BMVC.","DOI":"10.5244\/C.28.85"},{"key":"e_1_3_2_2_15_1","unstructured":"L. Jing and Y. Tian. 2020. Self-supervised Visual Feature Learning with Deep Neural Networks: A Survey. TPAMI (2020).  L. Jing and Y. Tian. 2020. Self-supervised Visual Feature Learning with Deep Neural Networks: A Survey. TPAMI (2020)."},{"key":"e_1_3_2_2_16_1","doi-asserted-by":"crossref","unstructured":"Dahun Kim Donghyeon Cho and In So Kweon. 2019. Self-Supervised Video Representation Learning with Space-Time Cubic Puzzles. In AAAI. 8545--8552.  Dahun Kim Donghyeon Cho and In So Kweon. 2019. Self-Supervised Video Representation Learning with Space-Time Cubic Puzzles. In AAAI. 8545--8552.","DOI":"10.1609\/aaai.v33i01.33018545"},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"crossref","unstructured":"Hao Liu Jiwen Lu Jianjiang Feng and Jie Zhou. 2018. Two-stream transformer networks for video-based face alignment. IEEE transactions on pattern analysis and machine intelligence Vol. 40 11 (2018) 2546--2554.  Hao Liu Jiwen Lu Jianjiang Feng and Jie Zhou. 2018. Two-stream transformer networks for video-based face alignment. IEEE transactions on pattern analysis and machine intelligence Vol. 40 11 (2018) 2546--2554.","DOI":"10.1109\/TPAMI.2017.2734779"},{"volume-title":"Know More: Unsupervised Video Object Segmentation With Co-Attention Siamese Networks. In The IEEE Conference on Computer Vision and Pattern Recognition (CVPR).","year":"2019","author":"Lu Xiankai","key":"e_1_3_2_2_18_1"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"crossref","unstructured":"Xin Miao Xiantong Zhen Xianglong Liu Cheng Deng Vassilis Athitsos and Heng Huang. 2018. Direct Shape Regression Networks for End-to-End Face Alignment. In CVPR. 5040--5049.  Xin Miao Xiantong Zhen Xianglong Liu Cheng Deng Vassilis Athitsos and Heng Huang. 2018. Direct Shape Regression Networks for End-to-End Face Alignment. In CVPR. 5040--5049.","DOI":"10.1109\/CVPR.2018.00529"},{"key":"e_1_3_2_2_20_1","doi-asserted-by":"crossref","unstructured":"Marco Pedersoli Radu Timofte Tinne Tuytelaars and Luc Van Gool. 2014. Using a Deformation Field Model for Localizing Faces and Facial Points under Weak Supervision. In CVPR. 3694--3701.  Marco Pedersoli Radu Timofte Tinne Tuytelaars and Luc Van Gool. 2014. Using a Deformation Field Model for Localizing Faces and Facial Points under Weak Supervision. In CVPR. 3694--3701.","DOI":"10.1109\/CVPR.2014.472"},{"key":"e_1_3_2_2_21_1","doi-asserted-by":"crossref","unstructured":"Xi Peng Rogerio S Feris Xiaoyu Wang and Dimitris N Metaxas. 2016. A recurrent encoder-decoder network for sequential face alignment. In ECCV. 38--56.  Xi Peng Rogerio S Feris Xiaoyu Wang and Dimitris N Metaxas. 2016. A recurrent encoder-decoder network for sequential face alignment. In ECCV. 38--56.","DOI":"10.1007\/978-3-319-46448-0_3"},{"key":"e_1_3_2_2_22_1","doi-asserted-by":"crossref","unstructured":"Ilija Radosavovic Piotr Doll\u00e1 r Ross B. Girshick Georgia Gkioxari and Kaiming He. 2018. Data Distillation: Towards Omni-Supervised Learning. In CVPR. 4119--4128.  Ilija Radosavovic Piotr Doll\u00e1 r Ross B. Girshick Georgia Gkioxari and Kaiming He. 2018. Data Distillation: Towards Omni-Supervised Learning. In CVPR. 4119--4128.","DOI":"10.1109\/CVPR.2018.00433"},{"key":"e_1_3_2_2_23_1","doi-asserted-by":"crossref","unstructured":"Christos Sagonas Epameinondas Antonakos Georgios Tzimiropoulos Stefanos Zafeiriou and Maja Pantic. 2016. 300 faces in-the-wild challenge: Database and results. Image and vision computing Vol. 47 (2016) 3--18.  Christos Sagonas Epameinondas Antonakos Georgios Tzimiropoulos Stefanos Zafeiriou and Maja Pantic. 2016. 300 faces in-the-wild challenge: Database and results. Image and vision computing Vol. 47 (2016) 3--18.","DOI":"10.1016\/j.imavis.2016.01.002"},{"key":"e_1_3_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVW.2015.132"},{"key":"e_1_3_2_2_25_1","unstructured":"Karen Simonyan and Andrew Zisserman. 2014. Two-stream convolutional networks for action recognition in videos. In NIPS. 568--576.  Karen Simonyan and Andrew Zisserman. 2014. Two-stream convolutional networks for action recognition in videos. In NIPS. 568--576."},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"crossref","unstructured":"Ying Tai Yicong Liang Xiaoming Liu Lei Duan Jilin Li Chengjie Wang Feiyue Huang and Yu Chen. 2019. Towards highly accurate and stable face alignment for high-resolution videos. In AAAI. 8893--8900.  Ying Tai Yicong Liang Xiaoming Liu Lei Duan Jilin Li Chengjie Wang Feiyue Huang and Yu Chen. 2019. Towards highly accurate and stable face alignment for high-resolution videos. In AAAI. 8893--8900.","DOI":"10.1609\/aaai.v33i01.33018893"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.neucom.2018.01.080"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"crossref","unstructured":"James Thewlis Hakan Bilen and Andrea Vedaldi. 2017. Unsupervised Learning of Object Landmarks by Factorized Spatial Embeddings. In ICCV.  James Thewlis Hakan Bilen and Andrea Vedaldi. 2017. Unsupervised Learning of Object Landmarks by Factorized Spatial Embeddings. In ICCV.","DOI":"10.1109\/ICCV.2017.348"},{"key":"e_1_3_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cviu.2012.03.008"},{"key":"e_1_3_2_2_30_1","doi-asserted-by":"crossref","unstructured":"Yue Wu and Qiang Ji. 2016. Constrained joint cascade regression framework for simultaneous facial action unit recognition and facial landmark detection. In CVPR. 3400--3408.  Yue Wu and Qiang Ji. 2016. Constrained joint cascade regression framework for simultaneous facial action unit recognition and facial landmark detection. In CVPR. 3400--3408.","DOI":"10.1109\/CVPR.2016.370"},{"key":"e_1_3_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11263-018-1097-z"},{"key":"e_1_3_2_2_32_1","doi-asserted-by":"crossref","unstructured":"Yue Wu Zuoguan Wang and Qiang Ji. 2013. Facial Feature Tracking Under Varying Facial Expressions and Face Poses Based on Restricted Boltzmann Machines. In CVPR. 3452--3459.  Yue Wu Zuoguan Wang and Qiang Ji. 2013. Facial Feature Tracking Under Varying Facial Expressions and Face Poses Based on Restricted Boltzmann Machines. In CVPR. 3452--3459.","DOI":"10.1109\/CVPR.2013.443"},{"key":"e_1_3_2_2_33_1","doi-asserted-by":"crossref","unstructured":"Xuehan Xiong and Fernando De la Torre. 2013. Supervised descent method and its applications to face alignment. In CVPR. 532--539.  Xuehan Xiong and Fernando De la Torre. 2013. Supervised descent method and its applications to face alignment. In CVPR. 532--539.","DOI":"10.1109\/CVPR.2013.75"},{"key":"e_1_3_2_2_34_1","doi-asserted-by":"crossref","unstructured":"Shi Yin Shangfei Wang Guozhu Peng Xiaoping Chen and Bowen Pan. 2019. Capturing Spatial and Temporal Patterns for Facial Landmark Tracking through Adversarial Learning. In IJCAI. 1010--1017.  Shi Yin Shangfei Wang Guozhu Peng Xiaoping Chen and Bowen Pan. 2019. Capturing Spatial and Temporal Patterns for Facial Landmark Tracking through Adversarial Learning. In IJCAI. 1010--1017.","DOI":"10.24963\/ijcai.2019\/142"},{"key":"e_1_3_2_2_35_1","doi-asserted-by":"crossref","unstructured":"Zhanpeng Zhang Ping Luo Chen Change Loy and Xiaoou Tang. 2016a. Learning deep representation for face alignment with auxiliary attributes. IEEE transactions on pattern analysis and machine intelligence Vol. 38 5 (2016) 918--930.  Zhanpeng Zhang Ping Luo Chen Change Loy and Xiaoou Tang. 2016a. Learning deep representation for face alignment with auxiliary attributes. IEEE transactions on pattern analysis and machine intelligence Vol. 38 5 (2016) 918--930.","DOI":"10.1109\/TPAMI.2015.2469286"},{"key":"e_1_3_2_2_36_1","doi-asserted-by":"crossref","unstructured":"Zhanpeng Zhang Ping Luo Chen Change Loy and Xiaoou Tang. 2016b. Learning deep representation for face alignment with auxiliary attributes. IEEE transactions on pattern analysis and machine intelligence Vol. 38 5 (2016) 918--930.  Zhanpeng Zhang Ping Luo Chen Change Loy and Xiaoou Tang. 2016b. Learning deep representation for face alignment with auxiliary attributes. IEEE transactions on pattern analysis and machine intelligence Vol. 38 5 (2016) 918--930.","DOI":"10.1109\/TPAMI.2015.2469286"},{"key":"e_1_3_2_2_37_1","unstructured":"Congcong Zhu Hao Liu Zhenhua Yu and Xuehong Sun. 2020. Towards Omni-Supervised Face Alignment for Large Scale Unlabeled Videos. In AAAI.  Congcong Zhu Hao Liu Zhenhua Yu and Xuehong Sun. 2020. Towards Omni-Supervised Face Alignment for Large Scale Unlabeled Videos. In AAAI."},{"key":"e_1_3_2_2_38_1","doi-asserted-by":"crossref","unstructured":"Shizhan Zhu Cheng Li Chen Change Loy and Xiaoou Tang. 2015. Face alignment by coarse-to-fine shape searching. In CVPR. 4998--5006.  Shizhan Zhu Cheng Li Chen Change Loy and Xiaoou Tang. 2015. Face alignment by coarse-to-fine shape searching. In CVPR. 4998--5006.","DOI":"10.1109\/CVPR.2015.7299134"}],"event":{"name":"MM '20: The 28th ACM International Conference on Multimedia","sponsor":["SIGMM ACM Special Interest Group on Multimedia"],"location":"Seattle WA USA","acronym":"MM '20"},"container-title":["Proceedings of the 28th ACM International Conference on Multimedia"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413547","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3394171.3413547","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:47:13Z","timestamp":1750193233000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3394171.3413547"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2020,10,12]]},"references-count":38,"alternative-id":["10.1145\/3394171.3413547","10.1145\/3394171"],"URL":"https:\/\/doi.org\/10.1145\/3394171.3413547","relation":{},"subject":[],"published":{"date-parts":[[2020,10,12]]},"assertion":[{"value":"2020-10-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}