{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,3,27]],"date-time":"2026-03-27T23:57:03Z","timestamp":1774655823954,"version":"3.50.1"},"publisher-location":"New York, NY, USA","reference-count":87,"publisher":"ACM","license":[{"start":{"date-parts":[[2024,10,24]],"date-time":"2024-10-24T00:00:00Z","timestamp":1729728000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Cyber Security Cooperative Research Centre"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2024,10,24]]},"DOI":"10.1145\/3674805.3686674","type":"proceedings-article","created":{"date-parts":[[2024,10,15]],"date-time":"2024-10-15T18:39:24Z","timestamp":1729017564000},"page":"119-130","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":8,"title":["Mitigating Data Imbalance for Software Vulnerability Assessment: Does Data Augmentation Help?"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-1935-037X","authenticated-orcid":false,"given":"Triet Huynh Minh","family":"Le","sequence":"first","affiliation":[{"name":"CREST - the Centre for Research on Engineering Software Technologies, The University of Adelaide, Australia"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9696-3626","authenticated-orcid":false,"given":"Muhammad","family":"Ali Babar","sequence":"additional","affiliation":[{"name":"CREST - the Centre for Research on Engineering Software Technologies, The University of Adelaide, Australia"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,10,24]]},"reference":[{"key":"e_1_3_2_1_1_1","first-page":"27865","article-title":"Self-supervised bug detection and repair","volume":"34","author":"Allamanis Miltiadis","year":"2021","unstructured":"Miltiadis Allamanis, Henry Jackson-Flux, and Marc Brockschmidt. 2021. Self-supervised bug detection and repair. Advances in Neural Information Processing Systems 34 (2021), 27865\u201327876.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_2_1","volume-title":"AI in Cybersecurity","author":"Almukaynizi Mohammed","unstructured":"Mohammed Almukaynizi, Eric Nunes, Krishna Dharaiya, Manoj Senguttuvan, Jana Shakarian, and Paulo Shakarian. 2019. Patch before exploited: An approach to identify targeted software vulnerabilities. In AI in Cybersecurity. Springer, 81\u2013113."},{"key":"e_1_3_2_1_3_1","volume-title":"Gpt4all: Training an assistant-style chatbot with large scale data distillation from gpt-3.5-turbo. GitHub","author":"Anand Yuvanesh","year":"2023","unstructured":"Yuvanesh Anand, Zach Nussbaum, Brandon Duderstadt, Benjamin Schmidt, and Andriy Mulyar. 2023. Gpt4all: Training an assistant-style chatbot with large scale data distillation from gpt-3.5-turbo. GitHub (2023)."},{"key":"e_1_3_2_1_4_1","volume-title":"Cleaning the NVD: Comprehensive quality assessment, improvements, and analyses","author":"Anwar Afsah","year":"2021","unstructured":"Afsah Anwar, Ahmed Abusnaina, Songqing Chen, Frank Li, and David Mohaisen. 2021. Cleaning the NVD: Comprehensive quality assessment, improvements, and analyses. IEEE Transactions on Dependable and Secure Computing (2021)."},{"key":"e_1_3_2_1_5_1","volume-title":"Mansooreh Zahedi, and M\u00a0Ali Babar.","author":"Arani Ali\u00a0Kazemi","year":"2024","unstructured":"Ali\u00a0Kazemi Arani, Triet Huynh\u00a0Minh Le, Mansooreh Zahedi, and M\u00a0Ali Babar. 2024. Systematic literature review on application of learning-based approaches in continuous integration. IEEE Access (2024)."},{"key":"e_1_3_2_1_6_1","unstructured":"Authors. [n.\u00a0d.]. Reproduction package. https:\/\/github.com\/lhmtriet\/DA4SVA"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3544558"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-23318-5_6"},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/1835804.1835821"},{"key":"e_1_3_2_1_10_1","volume-title":"Language models are few-shot learners. Advances in neural information processing systems 33","author":"Brown Tom","year":"2020","unstructured":"Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared\u00a0D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, 2020. Language models are few-shot learners. Advances in neural information processing systems 33 (2020), 1877\u20131901."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.5555\/1622407.1622416"},{"key":"e_1_3_2_1_12_1","unstructured":"The Conversation. [n.\u00a0d.]. What is Log4j?https:\/\/bit.ly\/log4j_the_conversation"},{"key":"e_1_3_2_1_13_1","volume-title":"Proceedings of the RSAConference","author":"Cornell Dan","year":"2012","unstructured":"Dan Cornell. 2012. Remediation statistics: what does fixing application vulnerabilities cost. Proceedings of the RSAConference, San Fransisco, CA, USA (2012)."},{"key":"e_1_3_2_1_14_1","volume-title":"Chataug: Leveraging chatgpt for text data augmentation. arXiv preprint arXiv:2302.13007","author":"Dai Haixing","year":"2023","unstructured":"Haixing Dai, Zhengliang Liu, Wenxiong Liao, Xiaoke Huang, Zihao Wu, Lin Zhao, Wei Liu, Ninghao Liu, Sheng Li, Dajiang Zhu, 2023. Chataug: Leveraging chatgpt for text data augmentation. arXiv preprint arXiv:2302.13007 (2023)."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/PRDC53464.2021.00016"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1109\/CSCloud.2015.56"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3407023.3407038"},{"key":"e_1_3_2_1_18_1","volume-title":"A survey of data augmentation approaches for NLP. arXiv preprint arXiv:2105.03075","author":"Feng Y","year":"2021","unstructured":"Steven\u00a0Y Feng, Varun Gangal, Jason Wei, Sarath Chandar, Soroush Vosoughi, Teruko Mitamura, and Eduard Hovy. 2021. A survey of data augmentation approaches for NLP. arXiv preprint arXiv:2105.03075 (2021)."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/CANDAR.2018.00009"},{"key":"e_1_3_2_1_20_1","unstructured":"Andy Field. 2013. Discovering statistics using IBM SPSS statistics. sage."},{"key":"e_1_3_2_1_21_1","unstructured":"FIRST. [n.\u00a0d.]. Common Vulnerability Scoring System. https:\/\/www.first.org\/cvss"},{"key":"e_1_3_2_1_22_1","unstructured":"FIRST. [n.\u00a0d.]. CVSS version 2. https:\/\/www.first.org\/cvss\/v2\/guide"},{"key":"e_1_3_2_1_23_1","volume-title":"Vulnerability management","author":"Foreman Park","unstructured":"Park Foreman. 2019. Vulnerability management. CRC Press."},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/APSEC60848.2023.00085"},{"key":"e_1_3_2_1_25_1","unstructured":"Recorded Future. [n.\u00a0d.]. Exploiting old vulnerabilities. https:\/\/www.recordedfuture.com\/exploiting-old-vulnerabilities\/"},{"key":"e_1_3_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1145\/3092566"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICECCS.2019.00011"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.compbiolchem.2004.09.006"},{"key":"e_1_3_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSME.2017.52"},{"key":"e_1_3_2_1_30_1","volume-title":"3rd international conference on document analysis and recognition, Vol.\u00a01. IEEE, 278\u2013282","author":"Ho Tin\u00a0Kam","year":"1995","unstructured":"Tin\u00a0Kam Ho. 1995. Random decision forests. In 3rd international conference on document analysis and recognition, Vol.\u00a01. IEEE, 278\u2013282."},{"key":"e_1_3_2_1_31_1","volume-title":"Long short-term memory. Neural computation 9, 8","author":"Hochreiter Sepp","year":"1997","unstructured":"Sepp Hochreiter and J\u00fcrgen Schmidhuber. 1997. Long short-term memory. Neural computation 9, 8 (1997), 1735\u20131780."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2013.06.040"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.acl-long.442"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/3338906.3338941"},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TDSC.2016.2644614"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/SERE.2013.31"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.24251\/HICSS.2021.841"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3343440","article-title":"A systematic review on imbalanced data challenges in machine learning: Applications and solutions","volume":"52","author":"Kaur Harsurinder","year":"2019","unstructured":"Harsurinder Kaur, Husanbir\u00a0Singh Pannu, and Avleen\u00a0Kaur Malhi. 2019. A systematic review on imbalanced data challenges in machine learning: Applications and solutions. ACM Computing Surveys (CSUR) 52, 4 (2019), 1\u201336.","journal-title":"ACM Computing Surveys (CSUR)"},{"key":"e_1_3_2_1_39_1","volume-title":"Guide to Vulnerability Analysis for Computer Networks and Systems","author":"Khan Saad","unstructured":"Saad Khan and Simon Parkinson. 2018. Review into state of the art of vulnerability assessment using artificial intelligence. In Guide to Vulnerability Analysis for Computer Networks and Systems. Springer, 3\u201332."},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.3115\/v1\/D14-1181"},{"key":"e_1_3_2_1_41_1","volume-title":"International conference on machine learning. PMLR, 1188\u20131196","author":"Le Quoc","year":"2014","unstructured":"Quoc Le and Tomas Mikolov. 2014. Distributed representations of sentences and documents. In International conference on machine learning. PMLR, 1188\u20131196."},{"key":"e_1_3_2_1_42_1","volume-title":"Towards an improved understanding of software vulnerability assessment using data-driven approaches. arXiv preprint arXiv:2207.11708","author":"HM Le.","year":"2022","unstructured":"Triet\u00a0HM Le. 2022. Towards an improved understanding of software vulnerability assessment using data-driven approaches. arXiv preprint arXiv:2207.11708 (2022)."},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3529757"},{"key":"e_1_3_2_1_44_1","volume-title":"Proceedings of the 19th International Conference on Mining Software Repositories. 621\u2013633","author":"Huynh\u00a0Minh Le Triet","year":"2022","unstructured":"Triet Huynh\u00a0Minh Le and M\u00a0Ali Babar. 2022. On the use of fine-grained vulnerable code statements for software vulnerability assessment models. In Proceedings of the 19th International Conference on Mining Software Repositories. 621\u2013633."},{"key":"e_1_3_2_1_45_1","volume-title":"Proceedings of the 28th International Conference on Evaluation and Assessment in Software Engineering. 679\u2013685","author":"Huynh\u00a0Minh Le Triet","year":"2024","unstructured":"Triet Huynh\u00a0Minh Le, M\u00a0Ali Babar, and Tung\u00a0Hoang Thai. 2024. Software vulnerability prediction in low-resource languages: An empirical study of codebert and chatgpt. In Proceedings of the 28th International Conference on Evaluation and Assessment in Software Engineering. 679\u2013685."},{"key":"e_1_3_2_1_46_1","first-page":"1","article-title":"Deep learning for source code modeling and generation: Models, applications, and challenges","volume":"53","author":"Huynh\u00a0Minh Le Triet","year":"2020","unstructured":"Triet Huynh\u00a0Minh Le, Hao Chen, and M\u00a0Ali Babar. 2020. Deep learning for source code modeling and generation: Models, applications, and challenges. ACM Computing Surveys (CSUR) 53, 3 (2020), 1\u201338.","journal-title":"ACM Computing Surveys (CSUR)"},{"key":"e_1_3_2_1_47_1","doi-asserted-by":"crossref","unstructured":"Triet Huynh\u00a0Minh Le Roland Croft David Hin and M\u00a0Ali Babar. 2021. A large-scale study of security vulnerability support on developer Q&A websites. In Evaluation and Assessment in Software Engineering. 109\u2013118.","DOI":"10.1145\/3463274.3463331"},{"key":"e_1_3_2_1_48_1","volume-title":"2024 IEEE\/ACM 21st International Conference on Mining Software Repositories (MSR). IEEE, 716\u2013727","author":"Huynh\u00a0Minh Le Triet","year":"2024","unstructured":"Triet Huynh\u00a0Minh Le, Xiaoning Du, and M\u00a0Ali Babar. 2024. Are latent vulnerabilities hidden gems for software vulnerability prediction? An empirical study. In 2024 IEEE\/ACM 21st International Conference on Mining Software Repositories (MSR). IEEE, 716\u2013727."},{"key":"e_1_3_2_1_49_1","volume-title":"the 17th International Conference on Mining Software Repositories. 350\u2013361","author":"Huynh\u00a0Minh Le Triet","year":"2020","unstructured":"Triet Huynh\u00a0Minh Le, David Hin, Roland Croft, and M\u00a0Ali Babar. 2020. PUMiner: Mining security posts from developer question and answer websites with PU learning. In the 17th International Conference on Mining Software Repositories. 350\u2013361."},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/ASE51524.2021.9678622"},{"key":"e_1_3_2_1_51_1","volume-title":"the 16th International Conference on Mining Software Repositories (MSR). IEEE, 371\u2013382","author":"Huynh\u00a0Minh Le Triet","year":"2019","unstructured":"Triet Huynh\u00a0Minh Le, Bushra Sabir, and Muhammad\u00a0Ali Babar. 2019. Automated software vulnerability assessment with concept drift. In the 16th International Conference on Mining Software Repositories (MSR). IEEE, 371\u2013382."},{"key":"e_1_3_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00207"},{"key":"e_1_3_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2022.11.349"},{"key":"e_1_3_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.3115\/1118108.1118117"},{"key":"e_1_3_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2019.02.023"},{"key":"e_1_3_2_1_56_1","unstructured":"Edward Ma. 2019. NLP Augmentation. https:\/\/github.com\/makcedward\/nlpaug."},{"key":"e_1_3_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-016-9488-7"},{"key":"e_1_3_2_1_58_1","doi-asserted-by":"publisher","DOI":"10.1145\/219717.219748"},{"key":"e_1_3_2_1_59_1","volume-title":"Data augmentation: A comprehensive survey of modern approaches. Array","author":"Mumuni Alhassan","year":"2022","unstructured":"Alhassan Mumuni and Fuseini Mumuni. 2022. Data augmentation: A comprehensive survey of modern approaches. Array (2022), 100258."},{"key":"e_1_3_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2016.02.048"},{"key":"e_1_3_2_1_61_1","unstructured":"Vinod Nair and Geoffrey\u00a0E Hinton. 2010. Rectified linear units improve restricted boltzmann machines. In Icml."},{"key":"e_1_3_2_1_62_1","unstructured":"NIST. [n.\u00a0d.]. National Vulnerability Database. https:\/\/nvd.nist.gov\/"},{"key":"e_1_3_2_1_63_1","unstructured":"NIST. [n.\u00a0d.]. The number of new software vulnerabilities. https:\/\/nvd.nist.gov\/general\/nvd-dashboard"},{"key":"e_1_3_2_1_64_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISI.2016.7745435"},{"key":"e_1_3_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.eacl-main.262"},{"key":"e_1_3_2_1_66_1","volume-title":"Scikit-learn: Machine learning in Python. the Journal of machine Learning research 12","author":"Pedregosa Fabian","year":"2011","unstructured":"Fabian Pedregosa, Ga\u00ebl Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, 2011. Scikit-learn: Machine learning in Python. the Journal of machine Learning research 12 (2011), 2825\u20132830."},{"key":"e_1_3_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPC58990.2023.00031"},{"key":"e_1_3_2_1_68_1","volume-title":"An algorithm for suffix stripping.Program 14, 3","author":"Porter F","year":"1980","unstructured":"Martin\u00a0F Porter. 1980. An algorithm for suffix stripping.Program 14, 3 (1980), 130\u2013137."},{"key":"e_1_3_2_1_69_1","volume-title":"Model evaluation, model selection, and algorithm selection in machine learning. arXiv preprint arXiv:1811.12808","author":"Raschka Sebastian","year":"2018","unstructured":"Sebastian Raschka. 2018. Model evaluation, model selection, and algorithm selection in machine learning. arXiv preprint arXiv:1811.12808 (2018)."},{"key":"e_1_3_2_1_70_1","doi-asserted-by":"publisher","DOI":"10.1145\/2601248.2601294"},{"key":"e_1_3_2_1_71_1","volume-title":"24th { USENIX} Security Symposium. 1041\u20131056.","author":"Sabottke Carl","unstructured":"Carl Sabottke, Octavian Suciu, and Tudor Dumitra\u0219. 2015. Vulnerability disclosure in the age of social media: Exploiting twitter for predicting real-world exploits. In 24th { USENIX} Security Symposium. 1041\u20131056."},{"key":"e_1_3_2_1_72_1","doi-asserted-by":"crossref","unstructured":"Sefa\u00a0Eren Sahin and Ayse Tosun. 2019. A conceptual replication on predicting the severity of software vulnerabilities. In the Evaluation and Assessment on Software Engineering. 244\u2013250.","DOI":"10.1145\/3319008.3319033"},{"key":"e_1_3_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P16-1009"},{"key":"e_1_3_2_1_74_1","doi-asserted-by":"publisher","DOI":"10.1007\/s13198-020-01021-7"},{"key":"e_1_3_2_1_75_1","doi-asserted-by":"publisher","DOI":"10.1016\/S1353-4858(17)30027-2"},{"key":"e_1_3_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2018.2836442"},{"key":"e_1_3_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2018.09.039"},{"key":"e_1_3_2_1_78_1","doi-asserted-by":"publisher","DOI":"10.1145\/3139367.3139390"},{"key":"e_1_3_2_1_79_1","unstructured":"Inc. Synopsys. [n.\u00a0d.]. Heartbleed bug. https:\/\/heartbleed.com\/"},{"key":"e_1_3_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2012.6227176"},{"key":"e_1_3_2_1_81_1","volume-title":"Predicting defective lines using a model-agnostic technique","author":"Wattanakriengkrai Supatsara","year":"2020","unstructured":"Supatsara Wattanakriengkrai, Patanamon Thongtanunam, Chakkrit Tantithamthavorn, Hideaki Hata, and Kenichi Matsumoto. 2020. Predicting defective lines using a model-agnostic technique. IEEE Transactions on Software Engineering (2020)."},{"key":"e_1_3_2_1_82_1","volume-title":"Eda: Easy data augmentation techniques for boosting performance on text classification tasks. arXiv preprint arXiv:1901.11196","author":"Wei Jason","year":"2019","unstructured":"Jason Wei and Kai Zou. 2019. Eda: Easy data augmentation techniques for boosting performance on text classification tasks. arXiv preprint arXiv:1901.11196 (2019)."},{"key":"e_1_3_2_1_83_1","volume-title":"Breakthroughs in Statistics","author":"Wilcoxon Frank","unstructured":"Frank Wilcoxon. 1992. Individual comparisons by ranking methods. In Breakthroughs in Statistics. Springer, 196\u2013202."},{"key":"e_1_3_2_1_84_1","doi-asserted-by":"publisher","DOI":"10.1109\/BADGERS.2015.018"},{"key":"e_1_3_2_1_85_1","volume-title":"mixup: Beyond empirical risk minimization. arXiv preprint arXiv:1710.09412","author":"Zhang Hongyi","year":"2017","unstructured":"Hongyi Zhang, Moustapha Cisse, Yann\u00a0N Dauphin, and David Lopez-Paz. 2017. mixup: Beyond empirical risk minimization. arXiv preprint arXiv:1710.09412 (2017)."},{"key":"e_1_3_2_1_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/3377811.3380383"},{"key":"e_1_3_2_1_87_1","volume-title":"Data augmentation approaches for source code models: A survey. arXiv preprint arXiv:2305.19915","author":"Zhuo Terry\u00a0Yue","year":"2023","unstructured":"Terry\u00a0Yue Zhuo, Zhou Yang, Zhensu Sun, Yufei Wang, Li Li, Xiaoning Du, Zhenchang Xing, and David Lo. 2023. Data augmentation approaches for source code models: A survey. arXiv preprint arXiv:2305.19915 (2023)."}],"event":{"name":"ESEM '24: ACM \/ IEEE International Symposium on Empirical Software Engineering and Measurement","location":"Barcelona Spain","acronym":"ESEM '24","sponsor":["SIGSOFT ACM Special Interest Group on Software Engineering"]},"container-title":["Proceedings of the 18th ACM\/IEEE International Symposium on Empirical Software Engineering and Measurement"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3674805.3686674","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3674805.3686674","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,8,22]],"date-time":"2025-08-22T12:55:48Z","timestamp":1755867348000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3674805.3686674"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,10,24]]},"references-count":87,"alternative-id":["10.1145\/3674805.3686674","10.1145\/3674805"],"URL":"https:\/\/doi.org\/10.1145\/3674805.3686674","relation":{},"subject":[],"published":{"date-parts":[[2024,10,24]]},"assertion":[{"value":"2024-10-24","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}