{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,9]],"date-time":"2026-07-09T04:28:11Z","timestamp":1783571291029,"version":"3.55.0"},"reference-count":53,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2023,12,9]],"date-time":"2023-12-09T00:00:00Z","timestamp":1702080000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Multimedia Comput. Commun. Appl."],"published-print":{"date-parts":[[2024,3,31]]},"abstract":"<jats:p>\n            The advancement of generative models has made it easier to create highly realistic Deepfake videos. This accessibility has led to a surge in research on Deepfake detection to mitigate potential misuse. Typically, Deepfake detection models utilize binary backbones, even though the training dataset contains additional exploitable information, such as the Deepfake generation method employed for each video. However, recent findings suggest that inferring a binary class from a multi-class backbone yields superior performance compared to directly employing a binary backbone. Building upon this research, our article introduces two novel methods to infer a binary class from a multi-class backbone. The first method, named\n            <jats:italic>root dummies<\/jats:italic>\n            , leverages the dummy triplet loss, which employs fixed vectors (i.e., dummies) instead of mined positives and negatives in the triplet loss. By training the multi-class backbone with these dummies, we can easily infer a binary class during testing by adjusting the number of dummies (from six during training to two during inference). Through this approach, we achieve an accuracy improvement of 0.23% compared to the existing inference method, without requiring additional training. The second proposed method is transfer learning. It involves training a classifier, such as a support vector machine, to predict binary classes based on the image embeddings generated by the multi-class backbone. Although this method necessitates additional training, it further enhances the model\u2019s performance, resulting in an accuracy increase of 1.79%. In summary, our proposed methods improve the accuracy of Deepfake detection by simply modifying the number of classes during training, making them suitable for integration into a variety of existing Deepfake training pipelines. Additionally, to foster reproducible research, we have made the source code of our solution publicly available at\n            <jats:ext-link xmlns:xlink=\"http:\/\/www.w3.org\/1999\/xlink\" xlink:href=\"https:\/\/github.com\/beuve\/DmyT\">https:\/\/github.com\/beuve\/DmyT<\/jats:ext-link>\n            .\n          <\/jats:p>","DOI":"10.1145\/3626101","type":"journal-article","created":{"date-parts":[[2023,10,5]],"date-time":"2023-10-05T10:32:16Z","timestamp":1696501936000},"page":"1-18","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Hierarchical Learning and Dummy Triplet Loss for Efficient Deepfake Detection"],"prefix":"10.1145","volume":"20","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-1371-4016","authenticated-orcid":false,"given":"Nicolas","family":"Beuve","sequence":"first","affiliation":[{"name":"University of Rennes, INSA Rennes, CNRS, IETR\u2013UMR 6164, France"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0143-1756","authenticated-orcid":false,"given":"Wassim","family":"Hamidouche","sequence":"additional","affiliation":[{"name":"University of Rennes, INSA Rennes, CNRS, IETR\u2013UMR 6164, France, and Technology Innovation Institute, Masdar City, UAE"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0750-0959","authenticated-orcid":false,"given":"Olivier","family":"D\u00e9forges","sequence":"additional","affiliation":[{"name":"University of Rennes, INSA Rennes, CNRS, IETR\u2013UMR 6164, France"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,12,9]]},"reference":[{"key":"e_1_3_1_2_2","doi-asserted-by":"crossref","DOI":"10.1109\/WIFS.2018.8630761","article-title":"MesoNet: A compact facial video forgery detection network","author":"Afchar Darius","year":"2018","unstructured":"Darius Afchar, Vincent Nozick, Junichi Yamagishi, and Isao Echizen. 2018. MesoNet: A compact facial video forgery detection network. In Proceedings of the 2018 IEEE International Workshop on Information Forensics and Security (WIFS \u201918). 1\u20137.","journal-title":"Proceedings of the 2018 IEEE International Workshop on Information Forensics and Security (WIFS \u201918)."},{"key":"e_1_3_1_3_2","doi-asserted-by":"crossref","unstructured":"Shruti Agarwal Tarek El-Gaaly Hany Farid and Ser-Nam Lim. 2020. Detecting deep-fake videos from appearance and behavior. arxiv:2004.14491 [cs.CV] (2020).","DOI":"10.1109\/WIFS49906.2020.9360904"},{"key":"e_1_3_1_4_2","doi-asserted-by":"crossref","first-page":"981","DOI":"10.1109\/CVPRW53098.2021.00109","article-title":"Detecting deep-fake videos from aural and oral dynamics","author":"Agarwal Shruti","year":"2021","unstructured":"Shruti Agarwal and Hany Farid. 2021. Detecting deep-fake videos from aural and oral dynamics. In Proceedings of the 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201921).981\u2013989.","journal-title":"Proceedings of the 2021 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201921)."},{"key":"e_1_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW53098.2021.00109"},{"key":"e_1_3_1_6_2","doi-asserted-by":"crossref","first-page":"2814","DOI":"10.1109\/CVPRW50498.2020.00338","volume-title":"Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201920)","author":"Agarwal Shruti","year":"2020","unstructured":"Shruti Agarwal, Hany Farid, Ohad Fried, and Maneesh Agrawala. 2020. Detecting deep-fake videos from phoneme-viseme mismatches. In Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201920). 2814\u20132822."},{"key":"e_1_3_1_7_2","doi-asserted-by":"crossref","first-page":"17","DOI":"10.1145\/3476099.3484316","volume-title":"Proceedings of the 1st Workshop on Synthetic Multimedia: Audiovisual Deepfake Generation and Detection","author":"Beuve Nicolas","year":"2021","unstructured":"Nicolas Beuve, Wassim Hamidouche, and Olivier Deforges. 2021. DmyT: Dummy triplet loss for deepfake detection. In Proceedings of the 1st Workshop on Synthetic Multimedia: Audiovisual Deepfake Generation and Detection. ACM, New York, NY, 17\u201324."},{"key":"e_1_3_1_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/311535.311556"},{"key":"e_1_3_1_9_2","doi-asserted-by":"crossref","first-page":"1800","DOI":"10.1109\/CVPR.2017.195","article-title":"Xception: Deep learning with depthwise separable convolutions","author":"Chollet Fran\u00e7ois","year":"2017","unstructured":"Fran\u00e7ois Chollet. 2017. Xception: Deep learning with depthwise separable convolutions. In Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201917).1800\u20131807.","journal-title":"Proceedings of the 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201917)."},{"key":"e_1_3_1_10_2","first-page":"130","volume-title":"Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201919)","author":"Cozzolino Davide","year":"2019","unstructured":"Davide Cozzolino, Giovanni Poggi, and Luisa Verdoliva. 2019. Extracting camera-based fingerprints for video forensics. In Proceedings of the 2019 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201919). 130\u2013137."},{"key":"e_1_3_1_11_2","article-title":"ID-Reveal: Identity-aware DeepFake video detection","author":"Cozzolino Davide","year":"2021","unstructured":"Davide Cozzolino, Andreas R\u00f6ssler, Justus Thies, Matthias Nie\u00dfner, and Luisa Verdoliva. 2021. ID-Reveal: Identity-aware DeepFake video detection. In Proceedings of the 2021 International Conference on Computer Vision.","journal-title":"Proceedings of the 2021 International Conference on Computer Vision."},{"key":"e_1_3_1_12_2","doi-asserted-by":"crossref","DOI":"10.1145\/3448017.3457387","article-title":"Where do deep fakes look? Synthetic face detection via gaze tracking","author":"Demir Ilke","year":"2021","unstructured":"Ilke Demir and Umur Aybars Ciftci. 2021. Where do deep fakes look? Synthetic face detection via gaze tracking. In Proceedings of the ACM Symposium on Eye Tracking Research and Applications.","journal-title":"Proceedings of the ACM Symposium on Eye Tracking Research and Applications."},{"key":"e_1_3_1_13_2","first-page":"5203","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920)","author":"Deng Jiankang","year":"2020","unstructured":"Jiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia, and Stefanos Zafeiriou. 2020. RetinaFace: Single-shot multi-level face localisation in the wild. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920). 5203\u20135212."},{"key":"e_1_3_1_14_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2019.01065"},{"key":"e_1_3_1_15_2","unstructured":"Brian Dolhansky Joanna Bitton Ben Pflaum Jikuo Lu Russ Howes Menglin Wang and Cristian Canton Ferrer. 2020. The DeepFake Detection Challenge Dataset. arxiv:2006.07397 [cs.CV] (2020)."},{"key":"e_1_3_1_16_2","article-title":"Identity-driven DeepFake detection","volume":"2012","author":"Dong Xiaoyi","year":"2020","unstructured":"Xiaoyi Dong, Jianmin Bao, Dongdong Chen, Weiming Zhang, Nenghai Yu, Dong Chen, Fang Wen, and Baining Guo. 2020. Identity-driven DeepFake detection. arXiv abs\/2012.03930 (2020).","journal-title":"arXiv"},{"key":"e_1_3_1_17_2","article-title":"Unmasking DeepFakes with simple features","volume":"1911","author":"Durall Ricard","year":"2019","unstructured":"Ricard Durall, Margret Keuper, Franz-Josef Pfreundt, and Janis Keuper. 2019. Unmasking DeepFakes with simple features. arXiv abs\/1911.00686 (2019).","journal-title":"arXiv"},{"key":"e_1_3_1_18_2","doi-asserted-by":"crossref","first-page":"1721","DOI":"10.1109\/ICCVW.2019.00213","volume-title":"Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision Workshop (ICCVW \u201919)","author":"Fernandes S.","year":"2019","unstructured":"S. Fernandes, S. Raj, E. Ortiz, I. Vintila, M. Saflter, G. Urosevic, and S. Jha. 2019. Predicting heart rate variations of deepfake videos using neural ODE. In Proceedings of the 2019 IEEE\/CVF International Conference on Computer Vision Workshop (ICCVW \u201919). 1721\u20131729."},{"key":"e_1_3_1_19_2","article-title":"Fighting deepfakes by detecting GAN DCT anomalies","volume":"7","author":"Giudice Oliver","year":"2021","unstructured":"Oliver Giudice, Luca Guarnera, and Sebastiano Battiato. 2021. Fighting deepfakes by detecting GAN DCT anomalies. Journal of Imaging 7 (2021), 128.","journal-title":"Journal of Imaging"},{"key":"e_1_3_1_20_2","volume-title":"Advances in Neural Information Processing Systems","author":"Goodfellow Ian","year":"2014","unstructured":"Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014. Generative adversarial nets. In Advances in Neural Information Processing Systems, Z. Ghahramani, M. Welling, C. Cortes, N. Lawrence, and K. Q. Weinberger (Eds.), Vol. 27. Curran Associates, 1\u20139."},{"key":"e_1_3_1_21_2","article-title":"Contributing data to deepfake detection research","author":"Blog Google AI","year":"2019","unstructured":"Google AI Blog. 2019. Contributing data to deepfake detection research. Google Research. Retrieved October 12, 2023 from https:\/\/blog.research.google\/2019\/09\/contributing-data-to-deepfake-detection.html?m=1","journal-title":"R"},{"key":"e_1_3_1_22_2","doi-asserted-by":"crossref","first-page":"2841","DOI":"10.1109\/CVPRW50498.2020.00341","article-title":"DeepFake detection by analyzing convolutional traces","author":"Guarnera Luca","year":"2020","unstructured":"Luca Guarnera, Oliver Giudice, and Sebastiano Battiato. 2020. DeepFake detection by analyzing convolutional traces. In Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201920).2841\u20132850.","journal-title":"Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW \u201920)."},{"key":"e_1_3_1_23_2","article-title":"Deepfake detection scheme based on vision transformer and distillation","volume":"2104","author":"Heo Young-Jin","year":"2021","unstructured":"Young-Jin Heo, Young-Ju Choi, Young-Woon Lee, and Byung-Gyu Kim. 2021. Deepfake detection scheme based on vision transformer and distillation. arXiv abs\/2104.01353 (2021).","journal-title":"arXiv"},{"key":"e_1_3_1_24_2","article-title":"DeepFakesON-phys: DeepFakes detection based on heart rate estimation","volume":"2010","author":"Hernandez-Ortega J.","year":"2021","unstructured":"J. Hernandez-Ortega, Rub\u00e9n Tolosana, Julian Fierrez, and Aythami Morales. 2021. DeepFakesON-phys: DeepFakes detection based on heart rate estimation. arXiv abs\/2010.00400 (2021).","journal-title":"arXiv"},{"key":"e_1_3_1_25_2","article-title":"MobileNets: Efficient convolutional neural networks for mobile vision applications","volume":"1704","author":"Howard Andrew G.","year":"2017","unstructured":"Andrew G. Howard, Menglong Zhu, Bo Chen, Dmitry Kalenichenko, Weijun Wang, Tobias Weyand, Marco Andreetto, and Hartwig Adam. 2017. MobileNets: Efficient convolutional neural networks for mobile vision applications. arXiv abs\/1704.04861 (2017).","journal-title":"arXiv"},{"key":"e_1_3_1_26_2","doi-asserted-by":"publisher","DOI":"10.1109\/MMSP53017.2021.9733468"},{"key":"e_1_3_1_27_2","volume-title":"Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920)","author":"Jiang Liming","year":"2020","unstructured":"Liming Jiang, Ren Li, Wayne Wu, Chen Qian, and Chen Change Loy. 2020. DeeperForensics-1.0: A large-scale dataset for real-world face forgery detection. In Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920)."},{"key":"e_1_3_1_28_2","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.2988660"},{"key":"e_1_3_1_29_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICASSP43922.2022.9747628"},{"key":"e_1_3_1_30_2","first-page":"5074","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition","author":"Li Lingzhi","year":"2020","unstructured":"Lingzhi Li, Jianmin Bao, Hao Yang, Dong Chen, and Fang Wen. 2020. Advancing high fidelity identity swapping for forgery detection. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition. 5074\u20135083."},{"key":"e_1_3_1_31_2","doi-asserted-by":"publisher","DOI":"10.1109\/WIFS.2018.8630787"},{"key":"e_1_3_1_32_2","first-page":"46","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPR2 \u201919)","author":"Li Yuezun","year":"2019","unstructured":"Yuezun Li and Siwei Lyu. 2019. Exposing DeepFake videos by detecting face warping artifacts. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPR2 \u201919). 46\u201352."},{"key":"e_1_3_1_33_2","first-page":"3204","volume-title":"Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920)","author":"Li Yuezun","year":"2020","unstructured":"Yuezun Li, Xin Yang, Pu Sun, Honggang Qi, and Siwei Lyu. 2020. Celeb-DF: A large-scale challenging dataset for DeepFake forensics. In Proceedings of the 2020 IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR \u201920). 3204\u20133213."},{"key":"e_1_3_1_34_2","doi-asserted-by":"publisher","DOI":"10.1111\/cgf.14062"},{"key":"e_1_3_1_35_2","first-page":"2307","volume-title":"Proceedings of the 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP \u201919)","author":"Nguyen Huy H.","year":"2019","unstructured":"Huy H. Nguyen, Junichi Yamagishi, and Isao Echizen. 2019. Capsule-forensics: Using capsule networks to detect forged images and videos. In Proceedings of the 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP \u201919). 2307\u20132311."},{"key":"e_1_3_1_36_2","article-title":"Use of a capsule network to detect fake images and videos","volume":"1910","author":"Nguyen Huy Hoang","year":"2019","unstructured":"Huy Hoang Nguyen, Junichi Yamagishi, and Isao Echizen. 2019. Use of a capsule network to detect fake images and videos. arXiv abs\/1910.12467 (2019).","journal-title":"arXiv"},{"key":"e_1_3_1_37_2","unstructured":"NVIDIA. 2020. NVIDIA Maxine. Retrieved October 12 2023 from https:\/\/developer.nvidia.com\/maxine"},{"key":"e_1_3_1_38_2","article-title":"DeepFaceLab: A simple, flexible and extensible face swapping framework","volume":"2005","author":"Petrov Ivan","year":"2020","unstructured":"Ivan Petrov, Daiheng Gao, Nikolay Chervoniy, Kunlin Liu, Sugasa Marangonda, Chris Um\u00e9, Mr. Dpfks, Luis RP, Jian Jiang, Sheng Zhang, Pingyu Wu, Bo Zhou, and Weiming Zhang. 2020. DeepFaceLab: A simple, flexible and extensible face swapping framework. arXiv abs\/2005.05535 (2020).","journal-title":"arXiv"},{"key":"e_1_3_1_39_2","first-page":"86","volume-title":"Proceedings of the European Conference on Computer Vision (ECCV \u201920)","author":"Qian Yuyang","year":"2020","unstructured":"Yuyang Qian, Guojun Yin, Lu Sheng, Zixuan Chen, and Jing Shao. 2020. Thinking in frequency: Face forgery detection by mining frequency-aware clues. In Proceedings of the European Conference on Computer Vision (ECCV \u201920). 86\u2013102."},{"key":"e_1_3_1_40_2","first-page":"1","volume-title":"Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV \u201919)","author":"Rossler Andreas","year":"2019","unstructured":"Andreas Rossler, Davide Cozzolino, Luisa Verdoliva, Christian Riess, Justus Thies, and Matthias Niessner. 2019. FaceForensics++: Learning to detect manipulated facial images. In Proceedings of the IEEE\/CVF International Conference on Computer Vision (ICCV \u201919). 1\u201311."},{"key":"e_1_3_1_41_2","doi-asserted-by":"crossref","first-page":"815","DOI":"10.1109\/CVPR.2015.7298682","volume-title":"Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201915)","author":"Schroff Florian","year":"2015","unstructured":"Florian Schroff, Dmitry Kalenichenko, and James Philbin. 2015. FaceNet: A unified embedding for face recognition and clustering. In Proceedings of the 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201915). 815\u2013823."},{"issue":"6","key":"e_1_3_1_42_2","article-title":"Seeing is believing: Is video modality more powerful in spreading fake news via online messaging apps?","volume":"26","author":"Sundar S. Shyam","year":"2021","unstructured":"S. Shyam Sundar, Maria D. Molina, and Eugene Cho. 2021. Seeing is believing: Is video modality more powerful in spreading fake news via online messaging apps? Journal of Computer-Mediated Communication 26, 6 (2021), 301\u2013319.","journal-title":"Journal of Computer-Mediated Communication"},{"key":"e_1_3_1_43_2","series-title":"Proceedings of the 36th International Conference on Machine Learning","first-page":"6105","volume":"97","author":"Tan Mingxing","year":"2019","unstructured":"Mingxing Tan and Quoc Le. 2019. EfficientNet: Rethinking model scaling for convolutional neural networks. In Proceedings of the 36th International Conference on Machine Learning, Kamalika Chaudhuri and Ruslan Salakhutdinov (Eds.). Proceedings of Machine Learning Research, Vol. 97. PMLR, 6105\u20136114."},{"key":"e_1_3_1_44_2","unstructured":"Shahroz Tariq Sangyup Lee and Simon S. Woo. 2020. A convolutional LSTM based residual network for deepfake video detection. arxiv:2009.07480 [cs.CV] (2020)."},{"key":"e_1_3_1_45_2","first-page":"12","article-title":"Deferred neural rendering: Image synthesis using neural textures","volume":"38","author":"Thies Justus","year":"2019","unstructured":"Justus Thies, Michael Zollh\u00f6fer, and Matthias Nie\u00dfner. 2019. Deferred neural rendering: Image synthesis using neural textures. ACM Transactions on Graphics 38, 4 (July 2019), Article 66, 12 pages.","journal-title":"ACM Transactions on Graphics"},{"key":"e_1_3_1_46_2","doi-asserted-by":"crossref","first-page":"2387","DOI":"10.1109\/CVPR.2016.262","article-title":"Face2Face: Real-time face capture and reenactment of RGB videos","author":"Thies Justus","year":"2016","unstructured":"Justus Thies, Michael Zollh\u00f6fer, Marc Stamminger, Christian Theobalt, and Matthias Nie\u00dfner. 2016. Face2Face: Real-time face capture and reenactment of RGB videos. In Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201916).2387\u20132395.","journal-title":"Proceedings of the 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR \u201916)."},{"key":"e_1_3_1_47_2","doi-asserted-by":"crossref","unstructured":"Junke Wang Zuxuan Wu Jingjing Chen and Yu-Gang Jiang. 2021. M2TR: Multi-modal multi-scale transformers for deepfake detection. arxiv:2104.09770 [cs.CV] (2021).","DOI":"10.1145\/3512527.3531415"},{"key":"e_1_3_1_48_2","first-page":"1136","volume-title":"Proceedings of the 30th International Joint Conference on Artificial Intelligence (IJCAI \u201921)","author":"Wang Yuhan","year":"2021","unstructured":"Yuhan Wang, Xu Chen, Junwei Zhu, Wenqing Chu, Ying Tai, Chengjie Wang, Jilin Li, Yongjian Wu, Feiyue Huang, and Rongrong Ji. 2021. HifiFace: 3D shape and semantic prior guided high fidelity face swapping. In Proceedings of the 30th International Joint Conference on Artificial Intelligence (IJCAI \u201921). 1136\u20131142."},{"key":"e_1_3_1_49_2","doi-asserted-by":"crossref","first-page":"515","DOI":"10.1109\/FG47880.2020.00089","article-title":"A video is worth more than 1000 lies. Comparing 3DCNN approaches for detecting deepfakes","author":"Wang Yaohui","year":"2020","unstructured":"Yaohui Wang and Antitza Dantcheva. 2020. A video is worth more than 1000 lies. Comparing 3DCNN approaches for detecting deepfakes. In Proceedings of the 2020 15th IEEE International Conference on Automatic Face and Gesture Recognition (FG \u201920).515\u2013519.","journal-title":"Proceedings of the 2020 15th IEEE International Conference on Automatic Face and Gesture Recognition (FG \u201920)."},{"key":"e_1_3_1_50_2","doi-asserted-by":"crossref","first-page":"2859","DOI":"10.1109\/ICCV.2017.309","volume-title":"Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV \u201917)","author":"Wu C.","year":"2017","unstructured":"C. Wu, R. Manmatha, A. J. Smola, and P. Krahenbuhl. 2017. Sampling matters in deep embedding learning. In Proceedings of the 2017 IEEE International Conference on Computer Vision (ICCV \u201917). IEEE, Los Alamitos, CA, 2859\u20132867."},{"key":"e_1_3_1_51_2","first-page":"622","volume-title":"Proceedings of the 16th European Conference on Computer Vision (ECCV \u201920)","author":"Wu Wayne","year":"2018","unstructured":"Wayne Wu, Yunxuan Zhang, Cheng Li, Chen Qian, and Chen Change Loy. 2018. ReenactGAN: Learning to reenact faces via boundary transfer. In Proceedings of the 16th European Conference on Computer Vision (ECCV \u201920). 622\u2013638."},{"key":"e_1_3_1_52_2","doi-asserted-by":"crossref","first-page":"8261","DOI":"10.1109\/ICASSP.2019.8683164","article-title":"Exposing deep fakes using inconsistent head poses","author":"Yang Xin","year":"2019","unstructured":"Xin Yang, Yuezun Li, and Siwei Lyu. 2019. Exposing deep fakes using inconsistent head poses. In Proceedings of the 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP \u201919).8261\u20138265.","journal-title":"Proceedings of the 2019 IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP \u201919)."},{"key":"e_1_3_1_53_2","first-page":"1","article-title":"Detecting and simulating artifacts in GAN fake images","author":"Zhang Xu","year":"2019","unstructured":"Xu Zhang, Svebor Karaman, and Shih-Fu Chang. 2019. Detecting and simulating artifacts in GAN fake images. In Proceedings of the 2019 IEEE International Workshop on Information Forensics and Security (WIFS \u201919).1\u20136.","journal-title":"Proceedings of the 2019 IEEE International Workshop on Information Forensics and Security (WIFS \u201919)."},{"key":"e_1_3_1_54_2","article-title":"WildDeepfake: A challenging real-world dataset for deepfake detection","author":"Zi Bojia","year":"2020","unstructured":"Bojia Zi, Minghao Chang, Jingjing Chen, Xingjun Ma, and Yugang Jiang. 2020. WildDeepfake: A challenging real-world dataset for deepfake detection. In Proceedings of the 28th ACM International Conference on Multimedia.","journal-title":"Proceedings of the 28th ACM International Conference on Multimedia."}],"container-title":["ACM Transactions on Multimedia Computing, Communications, and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3626101","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3626101","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T22:53:59Z","timestamp":1750287239000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3626101"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,9]]},"references-count":53,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2024,3,31]]}},"alternative-id":["10.1145\/3626101"],"URL":"https:\/\/doi.org\/10.1145\/3626101","relation":{},"ISSN":["1551-6857","1551-6865"],"issn-type":[{"value":"1551-6857","type":"print"},{"value":"1551-6865","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,9]]},"assertion":[{"value":"2022-06-10","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-09-18","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2023-12-09","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}