{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,31]],"date-time":"2026-03-31T21:06:54Z","timestamp":1774991214723,"version":"3.50.1"},"reference-count":99,"publisher":"Association for Computing Machinery (ACM)","issue":"FSE","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Proc. ACM Softw. Eng."],"published-print":{"date-parts":[[2025,6,19]]},"abstract":"<jats:p>Learning-based Android malware detection has earned significant recognition across industry and academia, yet its effectiveness hinges on the accuracy of labeled training data. Manual labeling, being prohibitively expensive, has prompted the use of automated methods, such as leveraging anti-virus engines like VirusTotal, which unfortunately introduces mislabeling, aka \"label noise\". The state-of-the-art label noise reduction approach, MalWhiteout, can effectively reduce random label noise but underperforms in mitigating real-world emergent malware (EM) label noise stemming from newly emerging Android malware variants overlooked by VirusTotal. To tackle this, we conceptualize EM label noise detection as an anomaly detection problem and introduce a novel tool, MalCleanse, that surpasses MalWhiteout's ability to address EM label noise. MalCleanse combines uncertainty estimation with unsupervised anomaly detection, identifying samples with high uncertainty as mislabeled, thereby enhancing its capability to remove EM label noise. Our experimental results demonstrate a significant reduction in EM label noise by approximately 25.25%, achieving an F1 Score of 80.32% for label noise detection at a noise ratio of 40.61%. Notably, MalCleanse outperforms MalWhiteout with an increase of 40.9% in overall F1 score for mitigating EM label noise. This paper pioneers the integration of deep neural network model uncertainty to refine label accuracy, thereby enhancing the reliability of malware detection systems. Our approach represents a significant step forward in addressing the challenges posed by emergent malware in automated labeling systems.<\/jats:p>","DOI":"10.1145\/3715769","type":"journal-article","created":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T15:15:34Z","timestamp":1750346134000},"page":"1136-1159","source":"Crossref","is-referenced-by-count":2,"title":["Mitigating Emergent Malware Label Noise in DNN-Based Android Malware Detection"],"prefix":"10.1145","volume":"2","author":[{"ORCID":"https:\/\/orcid.org\/0009-0004-1649-7549","authenticated-orcid":false,"given":"Haodong","family":"Li","sequence":"first","affiliation":[{"name":"Beijing University of Posts and Telecommunications, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5456-3827","authenticated-orcid":false,"given":"Xiao","family":"Cheng","sequence":"additional","affiliation":[{"name":"University of New South Wales, Sydney, Sydney, Australia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0006-1355-2324","authenticated-orcid":false,"given":"Guohan","family":"Zhang","sequence":"additional","affiliation":[{"name":"Beijing University of Posts and Telecommunications, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3310-926X","authenticated-orcid":false,"given":"Guosheng","family":"Xu","sequence":"additional","affiliation":[{"name":"Beijing University of Posts and Telecommunications, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9582-0698","authenticated-orcid":false,"given":"Guoai","family":"Xu","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1100-8633","authenticated-orcid":false,"given":"Haoyu","family":"Wang","sequence":"additional","affiliation":[{"name":"Huazhong University of Science and Technology, Wuhan, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,6,19]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","unstructured":"Orieb AbuAlghanam Hadeel Alazzam Mohammad Qatawneh Omar Aladwan Mohammad A Alsharaiah and Mohammed Amin Almaiah. 2023. Android Malware Detection System Based on Ensemble Learning. https:\/\/doi.org\/10.21203\/rs.3.rs-2521341\/v1 10.21203\/rs.3.rs-2521341\/v1","DOI":"10.21203\/rs.3.rs-2521341"},{"key":"e_1_2_1_2_1","unstructured":"G\u00f6rkem Algan and Ilkay Ulusoy. 2020. Label noise types and their effects on deep learning. arXiv preprint arXiv:2003.10471."},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-014-9352-6"},{"key":"e_1_2_1_4_1","volume-title":"Proceedings of the 13th international conference on mining software repositories. https:\/\/doi.org\/10","author":"Allix Kevin","year":"2016","unstructured":"Kevin Allix, Tegawend\u00e9 F Bissyand\u00e9, Jacques Klein, and Yves Le Traon. 2016. Androzoo: Collecting millions of Android apps for the research community. In Proceedings of the 13th international conference on mining software repositories. https:\/\/doi.org\/10.1145\/2901739.2903508 10.1145\/2901739.2903508"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/COMPSAC.2014.61"},{"key":"e_1_2_1_6_1","unstructured":"Ndz Anthony. 2023. \u201cReliable AI Models: 10 Data Reliability Checks Every Analyst Should Perform in 2023\u201d. https:\/\/www.datameer.com\/blog\/reliable-ai-models-10-data-reliability-checks-every-analyst-should-perform-in-2023\/"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2014.23247"},{"key":"e_1_2_1_8_1","first-page":"215","article-title":"Ensemble learning in Bayesian neural networks","volume":"168","author":"Barber David","year":"1998","unstructured":"David Barber and Christopher M Bishop. 1998. Ensemble learning in Bayesian neural networks. Nato ASI Series F Computer and Systems Sciences, 168 (1998), 215\u2013238.","journal-title":"Nato ASI Series F Computer and Systems Sciences"},{"key":"e_1_2_1_9_1","volume-title":"2022 IEEE Symposium on Security and Privacy (SP). https:\/\/doi.org\/10","author":"Barbero Federico","year":"2022","unstructured":"Federico Barbero, Feargus Pendlebury, Fabio Pierazzi, and Lorenzo Cavallaro. 2022. Transcending transcend: Revisiting malware classification in the presence of concept drift. In 2022 IEEE Symposium on Security and Privacy (SP). https:\/\/doi.org\/10.1109\/SP46214.2022.9833659 10.1109\/SP46214.2022.9833659"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.bpa.2005.07.009"},{"key":"e_1_2_1_11_1","volume-title":"Mixmatch: A holistic approach to semi-supervised learning. Advances in neural information processing systems, 32","author":"Berthelot David","year":"2019","unstructured":"David Berthelot, Nicholas Carlini, Ian Goodfellow, Nicolas Papernot, Avital Oliver, and Colin A Raffel. 2019. Mixmatch: A holistic approach to semi-supervised learning. Advances in neural information processing systems, 32 (2019), https:\/\/doi.org\/doi\/10.5555\/3454287.3454741"},{"key":"e_1_2_1_12_1","volume-title":"Proceedings of the eleventh annual conference on Computational learning theory. 92\u2013100","author":"Blum Avrim","year":"1998","unstructured":"Avrim Blum and Tom Mitchell. 1998. Combining labeled and unlabeled data with co-training. In Proceedings of the eleventh annual conference on Computational learning theory. 92\u2013100. https:\/\/doi.org\/10.1145\/279943.279962 10.1145\/279943.279962"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1613\/jair.606"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/1541880.1541882"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","unstructured":"Shuaike Dong Menghao Li Wenrui Diao Xiangyu Liu Jian Liu Zhou Li Fenghao Xu Kai Chen XiaoFeng Wang and Kehuan Zhang. 2018. Understanding Android Obfuscation Techniques: A Large-Scale Investigation in the Wild. In Security and Privacy in Communication Networks Raheem Beyah Bing Chang Yingjiu Li and Sencun Zhu (Eds.). https:\/\/doi.org\/10.1007\/978-3-030-01701-9_10 10.1007\/978-3-030-01701-9_10","DOI":"10.1007\/978-3-030-01701-9_10"},{"key":"e_1_2_1_16_1","unstructured":"Universit\u00e9 du Luxembourg. 2024. \u201cAndrozoo List\u201d. https:\/\/androzoo.uni.lu\/lists"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.14722\/ndss.2018.23296"},{"key":"e_1_2_1_18_1","volume-title":"Jos\u00e9 Miguel Hern\u00e1ndez-Lobato, and Richard E Turner","author":"Foong Andrew YK","year":"2019","unstructured":"Andrew YK Foong, Yingzhen Li, Jos\u00e9 Miguel Hern\u00e1ndez-Lobato, and Richard E Turner. 2019. \u2019In-Between\u2019Uncertainty in Bayesian Neural Networks. arXiv preprint arXiv:1906.11537."},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNNLS.2013.2292894"},{"key":"e_1_2_1_20_1","volume-title":"international conference on machine learning. https:\/\/dl.acm.org\/doi\/10","author":"Gal Yarin","year":"2016","unstructured":"Yarin Gal and Zoubin Ghahramani. 2016. Dropout as a bayesian approximation: Representing model uncertainty in deep learning. In international conference on machine learning. https:\/\/dl.acm.org\/doi\/10.5555\/3045390.3045502"},{"key":"e_1_2_1_21_1","volume-title":"Machine Learning: ECML-97: 9th European Conference on Machine Learning Prague, Czech Republic.","author":"Gamberger Dragan","year":"1997","unstructured":"Dragan Gamberger and Nada Lavra\u010d. 1997. Conditions for Occam\u2019s razor applicability and noise elimination. In Machine Learning: ECML-97: 9th European Conference on Machine Learning Prague, Czech Republic."},{"key":"e_1_2_1_22_1","volume-title":"Proceedings of the 2013 ACM workshop on Artificial intelligence and security. 45\u201354","author":"Gascon Hugo","year":"2013","unstructured":"Hugo Gascon, Fabian Yamaguchi, Daniel Arp, and Konrad Rieck. 2013. Structural detection of Android malware using embedded call graphs. In Proceedings of the 2013 ACM workshop on Artificial intelligence and security. 45\u201354. https:\/\/doi.org\/10.1145\/2517312.2517315 10.1145\/2517312.2517315"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10462-023-10562-9"},{"key":"e_1_2_1_24_1","volume-title":"International conference on learning representations. https:\/\/openreview.net\/forum?id=H12GRgcxg","author":"Goldberger Jacob","year":"2022","unstructured":"Jacob Goldberger and Ehud Ben-Reuven. 2022. Training deep neural-networks using a noise adaptation layer. In International conference on learning representations. https:\/\/openreview.net\/forum?id=H12GRgcxg"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-66399-9_4"},{"key":"e_1_2_1_26_1","volume-title":"International conference on machine learning. 1321\u20131330","author":"Guo Chuan","year":"2017","unstructured":"Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q Weinberger. 2017. On calibration of modern neural networks. In International conference on machine learning. 1321\u20131330. https:\/\/proceedings.mlr.press\/v70\/guo17a.html"},{"key":"e_1_2_1_27_1","volume-title":"Co-teaching: Robust training of deep neural networks with extremely noisy labels. Advances in neural information processing systems, 31","author":"Han Bo","year":"2018","unstructured":"Bo Han, Quanming Yao, Xingrui Yu, Gang Niu, Miao Xu, Weihua Hu, Ivor Tsang, and Masashi Sugiyama. 2018. Co-teaching: Robust training of deep neural networks with extremely noisy labels. Advances in neural information processing systems, 31 (2018), https:\/\/dl.acm.org\/doi\/10.5555\/3327757.3327944"},{"key":"e_1_2_1_28_1","unstructured":"Sara Hooker Yann Dauphin Aaron Courville and Andrea Frome. 2019. Selective brain damage: Measuring the disparate impact of model pruning. https:\/\/openreview.net\/forum?id=S1eIw0NFvr"},{"key":"e_1_2_1_29_1","first-page":"2017","volume-title":"Security and Communication Networks","author":"Hu Donghui","year":"2017","unstructured":"Donghui Hu, Zhongjin Ma, Xiaotian Zhang, Peipei Li, Dengpan Ye, and Baohong Ling. 2017. The concept drift problem in Android malware detection and its solution. Security and Communication Networks, 2017 (2017), https:\/\/doi.org\/10.1155\/2017\/4956386 10.1155\/2017\/4956386"},{"key":"e_1_2_1_30_1","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence. 36","author":"Huang Yingsong","year":"2022","unstructured":"Yingsong Huang, Bing Bai, Shengwei Zhao, Kun Bai, and Fei Wang. 2022. Uncertainty-aware learning against label noise on imbalanced datasets. In Proceedings of the AAAI Conference on Artificial Intelligence. 36, 6960\u20136969. https:\/\/doi.org\/10.1609\/aaai.v36i6.20654 10.1609\/aaai.v36i6.20654"},{"key":"e_1_2_1_31_1","unstructured":"Pavel Izmailov Wesley J Maddox Polina Kirichenko Timur Garipov Dmitry Vetrov and Andrew Gordon Wilson. 2020. Subspace inference for Bayesian deep learning. In Uncertainty in Artificial Intelligence. 1169\u20131179. https:\/\/proceedings.mlr.press\/v115\/izmailov20a.html"},{"key":"e_1_2_1_32_1","volume-title":"Black Hat USA","author":"Jeon Chanil","year":"2017","unstructured":"Chanil Jeon, Insu Yun, Jinho Jung, Max Wolotsky, and Taesoo Kim. 2017. AVPASS: Leaking and bypassing antivirus detection model automatically. In Black Hat USA 2017."},{"key":"e_1_2_1_33_1","volume-title":"International conference on machine learning. 2304\u20132313","author":"Jiang Lu","year":"2018","unstructured":"Lu Jiang, Zhengyuan Zhou, Thomas Leung, Li-Jia Li, and Li Fei-Fei. 2018. Mentornet: Learning data-driven curriculum for very deep neural networks on corrupted labels. In International conference on machine learning. 2304\u20132313. https:\/\/proceedings.mlr.press\/v80\/jiang18c.html"},{"key":"e_1_2_1_34_1","volume-title":"Transcend: Detecting concept drift in malware classification models. In 26th USENIX security symposium (USENIX security 17). 625\u2013642. https:\/\/dl.acm.org\/doi\/abs\/10.5555\/3241189.3241239","author":"Jordaney Roberto","year":"2017","unstructured":"Roberto Jordaney, Kumar Sharad, Santanu K Dash, Zhi Wang, Davide Papini, Ilia Nouretdinov, and Lorenzo Cavallaro. 2017. Transcend: Detecting concept drift in malware classification models. In 26th USENIX security symposium (USENIX security 17). 625\u2013642. https:\/\/dl.acm.org\/doi\/abs\/10.5555\/3241189.3241239"},{"key":"e_1_2_1_35_1","unstructured":"ElMouatez Billah Karbab Mourad Debbabi Abdelouahid Derhab and Djedjiga Mouheb. 2017. Android malware detection using deep learning on API method sequences. arXiv preprint arXiv:1712.08996."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.diin.2018.01.007"},{"key":"e_1_2_1_37_1","volume-title":"What uncertainties do we need in bayesian deep learning for computer vision? Advances in neural information processing systems, 30","author":"Kendall Alex","year":"2017","unstructured":"Alex Kendall and Yarin Gal. 2017. What uncertainties do we need in bayesian deep learning for computer vision? Advances in neural information processing systems, 30 (2017), https:\/\/dl.acm.org\/doi\/10.5555\/3295222.3295309"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIFS.2018.2866319"},{"key":"e_1_2_1_39_1","volume-title":"CVPR workshops. 33\u201337","author":"K\u00f6hler Jan M","year":"2019","unstructured":"Jan M K\u00f6hler, Maximilian Autenrieth, and William H Beluch. 2019. Uncertainty Based Detection and Relabeling of Noisy Image Labels.. In CVPR workshops. 33\u201337."},{"key":"e_1_2_1_40_1","volume-title":"Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security. https:\/\/doi.org\/10","author":"Korczynski David","year":"2017","unstructured":"David Korczynski and Heng Yin. 2017. Capturing malware propagations with code injections and code-reuse attacks. In Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security. https:\/\/doi.org\/10.1145\/3133956.3134099 10.1145\/3133956.3134099"},{"key":"e_1_2_1_41_1","doi-asserted-by":"crossref","unstructured":"Solomon Kullback and Richard A Leibler. 1951. On information and sufficiency. The annals of mathematical statistics.","DOI":"10.1214\/aoms\/1177729694"},{"key":"e_1_2_1_42_1","volume-title":"Simple and scalable predictive uncertainty estimation using deep ensembles. Advances in neural information processing systems, 30","author":"Lakshminarayanan Balaji","year":"2017","unstructured":"Balaji Lakshminarayanan, Alexander Pritzel, and Charles Blundell. 2017. Simple and scalable predictive uncertainty estimation using deep ensembles. Advances in neural information processing systems, 30 (2017), https:\/\/dl.acm.org\/doi\/10.5555\/3295222.3295387"},{"key":"e_1_2_1_43_1","volume-title":"Proceedings of the 37th Annual Computer Security Applications Conference. 596\u2013608","author":"Li Deqiang","year":"2021","unstructured":"Deqiang Li, Tian Qiu, Shuo Chen, Qianmu Li, and Shouhuai Xu. 2021. Can we leverage predictive uncertainty to detect dataset shift and adversarial examples in Android malware detection? In Proceedings of the 37th Annual Computer Security Applications Conference. 596\u2013608. https:\/\/doi.org\/10.1145\/3485832.3485916 10.1145\/3485832.3485916"},{"key":"e_1_2_1_44_1","volume-title":"Proceedings of the IEEE\/ACM 46th International Conference on Software Engineering. https:\/\/doi.org\/10","author":"Li Haodong","year":"2024","unstructured":"Haodong Li, Guosheng Xu, Liu Wang, Xusheng Xiao, Xiapu Luo, Guoai Xu, and Haoyu Wang. 2024. MalCertain: Enhancing Deep Neural Network Based Android Malware Detection by Tackling Prediction Uncertainty. In Proceedings of the IEEE\/ACM 46th International Conference on Software Engineering. https:\/\/doi.org\/10.1145\/3597503.3639122 10.1145\/3597503.3639122"},{"key":"e_1_2_1_45_1","volume-title":"Dividemix: Learning with noisy labels as semi-supervised learning. arXiv preprint arXiv:2002.07394.","author":"Li Junnan","year":"2020","unstructured":"Junnan Li, Richard Socher, and Steven CH Hoi. 2020. Dividemix: Learning with noisy labels as semi-supervised learning. arXiv preprint arXiv:2002.07394."},{"key":"e_1_2_1_46_1","volume-title":"Proceedings of the 31st IEEE\/ACM International Conference on Automated Software Engineering. 756\u2013761","author":"Li Li","year":"2016","unstructured":"Li Li, Tegawend\u00e9 F Bissyand\u00e9, Damien Octeau, and Jacques Klein. 2016. Reflection-aware static analysis of Android apps. In Proceedings of the 31st IEEE\/ACM International Conference on Automated Software Engineering. 756\u2013761. https:\/\/doi.org\/10.1145\/2970276.2970277 10.1145\/2970276.2970277"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/TIFS.2017.2656460"},{"key":"e_1_2_1_48_1","volume-title":"Pacific-Asia Conference on Knowledge Discovery and Data Mining. 611\u2013621","author":"Li Ming","year":"2005","unstructured":"Ming Li and Zhi-Hua Zhou. 2005. SETRED: Self-training with editing. In Pacific-Asia Conference on Knowledge Discovery and Data Mining. 611\u2013621. https:\/\/doi.org\/10.1007\/11430919_71 10.1007\/11430919_71"},{"key":"e_1_2_1_49_1","unstructured":"Mengting Li and Chuang Zhu. 2024. Noisy Label Processing for Classification: A Survey. arXiv preprint arXiv:2404.04159."},{"key":"e_1_2_1_50_1","volume-title":"Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2686\u20132695","author":"Li Zhizhong","year":"2020","unstructured":"Zhizhong Li and Derek Hoiem. 2020. Improving confidence estimates for unfamiliar examples. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 2686\u20132695."},{"key":"e_1_2_1_51_1","unstructured":"Shiyu Liang Yixuan Li and Rayadurgam Srikant. 2017. Enhancing the reliability of out-of-distribution image detection in neural networks. arXiv preprint arXiv:1706.02690."},{"key":"e_1_2_1_52_1","volume-title":"Proceedings of the Fifteenth ACM International Conference on Web Search and Data Mining. 608\u2013617","author":"Lu Yangdi","year":"2022","unstructured":"Yangdi Lu, Yang Bo, and Wenbo He. 2022. An Ensemble model for combating label noise. In Proceedings of the Fifteenth ACM International Conference on Web Search and Data Mining. 608\u2013617. https:\/\/doi.org\/10.1145\/3488560.3498376 10.1145\/3488560.3498376"},{"key":"e_1_2_1_53_1","unstructured":"Yueming Lyu and Ivor W Tsang. 2019. Curriculum loss: Robust learning and generalization against label corruption. arXiv preprint arXiv:1905.10045."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10922-021-09634-4"},{"key":"e_1_2_1_55_1","unstructured":"MalCleanse. 2024. \u201cMalCleanse\u201d. https:\/\/github.com\/Dirtyboy1029\/MALCLEANSE"},{"key":"e_1_2_1_56_1","volume-title":"ICML Workshop on Uncertainty and Robustness in Deep Learning.","author":"Manunza Gilberto","year":"2021","unstructured":"Gilberto Manunza, Matteo Pagliardini, Martin Jaggi, and Tatjana Chavdarova. 2021. Improved Adversarial Robustness via Uncertainty Targeted Attacks. In ICML Workshop on Uncertainty and Robustness in Deep Learning."},{"key":"e_1_2_1_57_1","volume-title":"Proceedings of the seventh ACM on conference on data and application security and privacy. 301\u2013308","author":"McLaughlin Niall","year":"2017","unstructured":"Niall McLaughlin, Jesus Martinez del Rincon, BooJoong Kang, Suleiman Yerima, Paul Miller, Sakir Sezer, Yeganeh Safaei, Erik Trickel, Ziming Zhao, and Adam Doup\u00e9. 2017. Deep Android malware detection. In Proceedings of the seventh ACM on conference on data and application security and privacy. 301\u2013308. https:\/\/doi.org\/10.1145\/3029806.3029823 10.1145\/3029806.3029823"},{"key":"e_1_2_1_58_1","unstructured":"Marcin Mo\u017cejko Mateusz Susik and Rafa\u0142 Karczewski. 2018. Inhibited softmax for uncertainty estimation in neural networks. arXiv preprint arXiv:1810.01861."},{"key":"e_1_2_1_59_1","volume-title":"When does label smoothing help? Advances in neural information processing systems, 32","author":"M\u00fcller Rafael","year":"2019","unstructured":"Rafael M\u00fcller, Simon Kornblith, and Geoffrey E Hinton. 2019. When does label smoothing help? Advances in neural information processing systems, 32 (2019), https:\/\/dl.acm.org\/doi\/10.5555\/3454287.3454709"},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2022.111359"},{"key":"e_1_2_1_61_1","doi-asserted-by":"crossref","unstructured":"Masataka Nakahara Norihiro Okui Yasuaki Kobayashi and Yutaka Miyake. 2021. Malware Detection for IoT Devices using Automatically Generated White List and Isolation Forest.. In IoTBDS. 38\u201347.","DOI":"10.5220\/0010394900380047"},{"key":"e_1_2_1_62_1","unstructured":"Abdelmonim Naway and Yuancheng Li. 2019. Using deep neural network for Android malware detection. arXiv preprint arXiv:1904.00736."},{"key":"e_1_2_1_63_1","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence. https:\/\/doi.org\/10","author":"Nguyen Andre T","year":"2022","unstructured":"Andre T Nguyen, Fred Lu, Gary Lopez Munoz, Edward Raff, Charles Nicholas, and James Holt. 2022. Out of distribution data detection using dropout bayesian neural networks. In Proceedings of the AAAI Conference on Artificial Intelligence. https:\/\/doi.org\/10.1609\/aaai.v36i7.20757 10.1609\/aaai.v36i7.20757"},{"key":"e_1_2_1_64_1","volume-title":"Proceedings of the 32nd ACM SIGSOFT International Symposium on Software Testing and Analysis. 52\u201363","author":"Nie Xu","year":"2023","unstructured":"Xu Nie, Ningke Li, Kailong Wang, Shangguang Wang, Xiapu Luo, and Haoyu Wang. 2023. Understanding and tackling label errors in deep learning-based vulnerability detection (experience paper). In Proceedings of the 32nd ACM SIGSOFT International Symposium on Software Testing and Analysis. 52\u201363. https:\/\/doi.org\/10.1145\/3597926.3598037 10.1145\/3597926.3598037"},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.1613\/jair.1.12125"},{"key":"e_1_2_1_66_1","unstructured":"Curtis G Northcutt Anish Athalye and Jonas Mueller. 2021. Pervasive label errors in test sets destabilize machine learning benchmarks. arXiv preprint arXiv:2103.14749."},{"key":"e_1_2_1_67_1","volume-title":"Proceedings of the 12th ACM Workshop on artificial intelligence and security. 37\u201348","author":"Oak Rajvardhan","year":"2019","unstructured":"Rajvardhan Oak, Min Du, David Yan, Harshvardhan Takawale, and Idan Amit. 2019. Malware detection on highly imbalanced data through sequence modeling. In Proceedings of the 12th ACM Workshop on artificial intelligence and security. 37\u201348. https:\/\/doi.org\/10.1145\/3338501.3357374 10.1145\/3338501.3357374"},{"key":"e_1_2_1_68_1","unstructured":"Yaniv Ovadia Emily Fertig Jie Ren Zachary Nado David Sculley Sebastian Nowozin Joshua Dillon Balaji Lakshminarayanan and Jasper Snoek. 2019. Can you trust your model\u2019s uncertainty? evaluating predictive uncertainty under dataset shift. Advances in neural information processing systems https:\/\/dl.acm.org\/doi\/10.5555\/3454287.3455541"},{"key":"e_1_2_1_69_1","volume-title":"2020 25th International Conference on Pattern Recognition (ICPR). https:\/\/doi.org\/10","author":"Patel Kanil","year":"2021","unstructured":"Kanil Patel, William Beluch, Dan Zhang, Michael Pfeiffer, and Bin Yang. 2021. On-manifold adversarial data augmentation improves uncertainty calibration. In 2020 25th International Conference on Pattern Recognition (ICPR). https:\/\/doi.org\/10.1109\/ICPR48806.2021.9413010 10.1109\/ICPR48806.2021.9413010"},{"key":"e_1_2_1_70_1","first-page":"262","article-title":"Label noise filtering based on the data distribution","volume":"59","author":"Qingqiang CHEN","year":"2019","unstructured":"CHEN Qingqiang, WANG Wenjian, and JIANG Gaoxia. 2019. Label noise filtering based on the data distribution. Journal of Tsinghua University (Science and Technology), 59, 4 (2019), 262\u2013269.","journal-title":"Journal of Tsinghua University (Science and Technology)"},{"key":"e_1_2_1_71_1","doi-asserted-by":"publisher","DOI":"10.1145\/3068335"},{"key":"e_1_2_1_72_1","volume-title":"Evidential deep learning to quantify classification uncertainty. Advances in neural information processing systems, 31","author":"Sensoy Murat","year":"2018","unstructured":"Murat Sensoy, Lance Kaplan, and Melih Kandemir. 2018. Evidential deep learning to quantify classification uncertainty. Advances in neural information processing systems, 31 (2018), https:\/\/dl.acm.org\/doi\/abs\/10.5555\/3327144.3327239"},{"key":"e_1_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.4169\/amer.math.monthly.122.5.490"},{"key":"e_1_2_1_74_1","doi-asserted-by":"publisher","unstructured":"Xin Su Dafang Zhang Wenjia Li and Kai Zhao. 2016. A deep learning approach to Android malware feature learning and detection. In 2016 IEEE Trustcom\/BigDataSE\/ISPA. 244\u2013251. https:\/\/doi.org\/10.1109\/TrustCom.2016.0070 10.1109\/TrustCom.2016.0070","DOI":"10.1109\/TrustCom.2016.0070"},{"key":"e_1_2_1_75_1","volume-title":"Proceedings of the IEEE conference on computer vision and pattern recognition.","author":"Szegedy Christian","year":"2016","unstructured":"Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016. Rethinking the inception architecture for computer vision. In Proceedings of the IEEE conference on computer vision and pattern recognition."},{"key":"e_1_2_1_76_1","volume-title":"Investigating Commercial Pay-Per-Install and the Distribution of Unwanted Software. In 25th USENIX Security Symposium. https:\/\/dl.acm.org\/doi\/abs\/10","author":"Thomas Kurt","year":"2016","unstructured":"Kurt Thomas, Juan A Elices Crespo, Ryan Rasti, Jean-Michel Picod, Cait Phillips, Marc-Andr\u00e9 Decoste, Chris Sharp, Fabio Tirelo, Ali Tofigh, and Marc-Antoine Courteau. 2016. Investigating Commercial Pay-Per-Install and the Distribution of Unwanted Software. In 25th USENIX Security Symposium. https:\/\/dl.acm.org\/doi\/abs\/10.5555\/3241094.3241151"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.5555\/3454287.3455533"},{"key":"e_1_2_1_78_1","doi-asserted-by":"publisher","DOI":"10.1007\/s42452-020-3132-2"},{"key":"e_1_2_1_79_1","unstructured":"Bindya Venkatesh and Jayaraman J Thiagarajan. 2019. Heteroscedastic calibration of uncertainty estimators in deep learning. arXiv preprint arXiv:1910.14179."},{"key":"e_1_2_1_80_1","volume-title":"2017 International conference on advances in computing, communications and informatics (ICACCI). 1677\u20131683","author":"Vinayakumar R","year":"2017","unstructured":"R Vinayakumar, KP Soman, and Prabaharan Poornachandran. 2017. Deep Android malware detection and classification. In 2017 International conference on advances in computing, communications and informatics (ICACCI). 1677\u20131683. https:\/\/doi.org\/10.1109\/ICACCI.2017.8126084 10.1109\/ICACCI.2017.8126084"},{"key":"e_1_2_1_81_1","doi-asserted-by":"publisher","DOI":"10.3233\/JIFS-169424"},{"key":"e_1_2_1_82_1","unstructured":"VirusTotal. 2025. \u201cVirusTotal\u201d. https:\/\/www.virustotal.com\/"},{"key":"e_1_2_1_83_1","volume-title":"Proceedings of the Internet Measurement Conference","author":"Wang Haoyu","year":"2018","unstructured":"Haoyu Wang, Zhe Liu, Jingyue Liang, Narseo Vallina-Rodriguez, Yao Guo, Li Li, Juan Tapiador, Jingcun Cao, and Guoai Xu. 2018. Beyond google play: A large-scale comparative study of chinese Android app markets. In Proceedings of the Internet Measurement Conference 2018. 293\u2013307. https:\/\/doi.org\/10.1145\/3278532.3278558 10.1145\/3278532.3278558"},{"key":"e_1_2_1_84_1","volume-title":"Proceedings of the ACM on Measurement and Analysis of Computing Systems, https:\/\/doi.org\/10","author":"Wang Liu","year":"2022","unstructured":"Liu Wang, Haoyu Wang, Ren He, Ran Tao, Guozhu Meng, Xiapu Luo, and Xuanzhe Liu. 2022. MalRadar: Demystifying Android malware in the new era. Proceedings of the ACM on Measurement and Analysis of Computing Systems, https:\/\/doi.org\/10.1145\/3530906 10.1145\/3530906"},{"key":"e_1_2_1_85_1","volume-title":"Proceedings of the 37th IEEE\/ACM International Conference on Automated Software Engineering. 1\u201313","author":"Wang Liu","year":"2022","unstructured":"Liu Wang, Haoyu Wang, Xiapu Luo, and Yulei Sui. 2022. MalWhiteout: Reducing label errors in Android malware detection. In Proceedings of the 37th IEEE\/ACM International Conference on Automated Software Engineering. 1\u201313. https:\/\/doi.org\/10.1145\/3551349.3560418 10.1145\/3551349.3560418"},{"key":"e_1_2_1_86_1","volume-title":"17th International Conference on Information Technology\u2013New Generations (ITNG","author":"Wang Yang","year":"2020","unstructured":"Yang Wang and Jun Zheng. 2020. An evaluation of one-class feature selection and classification for zero-day Android malware detection. In 17th International Conference on Information Technology\u2013New Generations (ITNG 2020). https:\/\/doi.org\/10.1007\/978-3-030-43020-7_15 10.1007\/978-3-030-43020-7_15"},{"key":"e_1_2_1_87_1","volume-title":"DIMVA 2017, Bonn, Germany, July 6-7, 2017, Proceedings 14","author":"Wei Fengguo","year":"2017","unstructured":"Fengguo Wei, Yuping Li, Sankardas Roy, Xinming Ou, and Wu Zhou. 2017. Deep ground truth analysis of current Android malware. In Detection of Intrusions and Malware, and Vulnerability Assessment: 14th International Conference, DIMVA 2017, Bonn, Germany, July 6-7, 2017, Proceedings 14. https:\/\/doi.org\/10.1007\/978-3-319-60876-1_12 10.1007\/978-3-319-60876-1_12"},{"key":"e_1_2_1_88_1","unstructured":"Andrew G Wilson and Pavel Izmailov. 2020. Bayesian deep learning and a probabilistic perspective of generalization. Advances in neural information processing systems https:\/\/dl.acm.org\/doi\/abs\/10.5555\/3495724.3496118"},{"key":"e_1_2_1_89_1","volume-title":"2023 IEEE Symposium on Security and Privacy (SP). 2602\u20132619","author":"Wu Xian","year":"2023","unstructured":"Xian Wu, Wenbo Guo, Jia Yan, Baris Coskun, and Xinyu Xing. 2023. From grim reality to practical solution: Malware classification in real-world noise. In 2023 IEEE Symposium on Security and Privacy (SP). 2602\u20132619. https:\/\/doi.org\/10.1109\/SP46215.2023.10179453 10.1109\/SP46215.2023.10179453"},{"key":"e_1_2_1_90_1","volume-title":"2019 34th IEEE\/ACM International Conference on Automated Software Engineering (ASE). 139\u2013150","author":"Wu Yueming","year":"2019","unstructured":"Yueming Wu, Xiaodi Li, Deqing Zou, Wei Yang, Xin Zhang, and Hai Jin. 2019. Malscan: Fast market-wide mobile malware scanning by social-network centrality analysis. In 2019 34th IEEE\/ACM International Conference on Automated Software Engineering (ASE). 139\u2013150. https:\/\/doi.org\/10.1109\/ASE.2019.00023 10.1109\/ASE.2019.00023"},{"key":"e_1_2_1_91_1","doi-asserted-by":"publisher","unstructured":"Jiayun Xu Yingjiu Li and Robert H Deng. 2021. Differential training: A generic framework to reduce label noises for Android malware detection. https:\/\/doi.org\/10.14722\/ndss.2021.24126 10.14722\/ndss.2021.24126","DOI":"10.14722\/ndss.2021.24126"},{"key":"e_1_2_1_92_1","volume-title":"Formal Methods and Software Engineering: 20th International Conference on Formal Engineering Methods, ICFEM 2018, Gold Coast, QLD, Australia, November 12-16, 2018, Proceedings 20","author":"Xu Zhiwu","year":"2018","unstructured":"Zhiwu Xu, Kerong Ren, Shengchao Qin, and Florin Craciun. 2018. CDGDroid: Android malware detection based on deep learning using CFG and DFG. In Formal Methods and Software Engineering: 20th International Conference on Formal Engineering Methods, ICFEM 2018, Gold Coast, QLD, Australia, November 12-16, 2018, Proceedings 20. https:\/\/doi.org\/10.1007\/978-3-030-02450-5_11 10.1007\/978-3-030-02450-5_11"},{"key":"e_1_2_1_93_1","unstructured":"Jingkang Yang Kaiyang Zhou Yixuan Li and Ziwei Liu. 2021. Generalized out-of-distribution detection: A survey. arXiv preprint arXiv:2110.11334."},{"key":"e_1_2_1_94_1","volume-title":"30th USENIX Security Symposium (USENIX Security 21)","author":"Yang Limin","year":"2021","unstructured":"Limin Yang, Wenbo Guo, Qingying Hao, Arridhana Ciptadi, Ali Ahmadzadeh, Xinyu Xing, and Gang Wang. 2021. CADE: Detecting and explaining concept drift samples for security applications. In 30th USENIX Security Symposium (USENIX Security 21). 2327\u20132344. https:\/\/www.usenix.org\/conference\/usenixsecurity21\/presentation\/yang-limin"},{"key":"e_1_2_1_95_1","volume-title":"International conference on machine learning. 7164\u20137173","author":"Yu Xingrui","year":"2019","unstructured":"Xingrui Yu, Bo Han, Jiangchao Yao, Gang Niu, Ivor Tsang, and Masashi Sugiyama. 2019. How does disagreement help generalization against label corruption? In International conference on machine learning. 7164\u20137173. https:\/\/proceedings.mlr.press\/v97\/yu19b.html"},{"key":"e_1_2_1_96_1","volume-title":"Proceedings of the ACM\/IEEE 42nd International Conference on Software Engineering. https:\/\/doi.org\/10","author":"Zhang Xiyue","year":"2020","unstructured":"Xiyue Zhang, Xiaofei Xie, Lei Ma, Xiaoning Du, Qiang Hu, Yang Liu, Jianjun Zhao, and Meng Sun. 2020. Towards characterizing adversarial defects of deep learning software from the lens of uncertainty. In Proceedings of the ACM\/IEEE 42nd International Conference on Software Engineering. https:\/\/doi.org\/10.1145\/3377811.3380368 10.1145\/3377811.3380368"},{"key":"e_1_2_1_97_1","unstructured":"Zhilu Zhang Adrian V Dalca and Mert R Sabuncu. 2019. Confidence calibration for convolutional neural networks using structured dropout. arXiv preprint arXiv:1906.09551."},{"key":"e_1_2_1_98_1","volume-title":"2012 IEEE symposium on security and privacy. 95\u2013109","author":"Zhou Yajin","year":"2012","unstructured":"Yajin Zhou and Xuxian Jiang. 2012. Dissecting Android malware: Characterization and evolution. In 2012 IEEE symposium on security and privacy. 95\u2013109. https:\/\/doi.org\/10.1109\/SP.2012.16 10.1109\/SP.2012.16"},{"key":"e_1_2_1_99_1","unstructured":"Shuofei Zhu Jianjun Shi Limin Yang Boqin Qin Ziyi Zhang Linhai Song and Gang Wang. 2020. Measuring and modeling the label dynamics of online Anti-Malware engines. In USENIX Security 20. https:\/\/dl.acm.org\/doi\/10.5555\/3489212.3489345"}],"container-title":["Proceedings of the ACM on Software Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3715769","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T15:28:16Z","timestamp":1750346896000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3715769"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,6,19]]},"references-count":99,"journal-issue":{"issue":"FSE","published-print":{"date-parts":[[2025,6,19]]}},"alternative-id":["10.1145\/3715769"],"URL":"https:\/\/doi.org\/10.1145\/3715769","relation":{},"ISSN":["2994-970X"],"issn-type":[{"value":"2994-970X","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,6,19]]}}}