{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T05:05:52Z","timestamp":1750309552262,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":31,"publisher":"ACM","license":[{"start":{"date-parts":[[2025,2,21]],"date-time":"2025-02-21T00:00:00Z","timestamp":1740096000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/"}],"funder":[{"name":"Research and Development on the Digital Employees and Intelligent Integration Capability of Business Scenarios in CMDC","award":["R2411DRH"],"award-info":[{"award-number":["R2411DRH"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,2,21]]},"DOI":"10.1145\/3728725.3728799","type":"proceedings-article","created":{"date-parts":[[2025,6,4]],"date-time":"2025-06-04T21:00:02Z","timestamp":1749070802000},"page":"473-479","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["The Application of Membership Inference in Privacy Auditing of Large Language Models Based on Fine-Tuning Method"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-9232-2749","authenticated-orcid":false,"given":"Chengjiang","family":"Wen","sequence":"first","affiliation":[{"name":"China Mobile Group Device Co. Ltd., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-4881-8860","authenticated-orcid":false,"given":"Yang","family":"Yue","sequence":"additional","affiliation":[{"name":"China Mobile Group Device Co. Ltd., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-4858-2211","authenticated-orcid":false,"given":"Zhixiang","family":"Wang","sequence":"additional","affiliation":[{"name":"China Mobile Group Device Co. Ltd., Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,6,3]]},"reference":[{"key":"e_1_3_3_1_1_2","volume-title":"Large language models for mathematical reasoning: Progresses and challenges. arXiv preprint arXiv:2402.00157","author":"Ahn Janice","year":"2024","unstructured":"Janice Ahn, Rishu Verma, Renze Lou, Di Liu, Rui Zhang, and Wenpeng Yin. 2024. Large language models for mathematical reasoning: Progresses and challenges. arXiv preprint arXiv:2402.00157 (2024)."},{"key":"e_1_3_3_1_2_2","unstructured":"anthropic. 2024. Fine=Tuning. https:\/\/www.anthropic.com\/api."},{"key":"e_1_3_3_1_3_2","volume-title":"December 27, 20247:18 PM GMT+8.","author":"Brittain Blake","year":"2024","unstructured":"Blake Brittain. 2024. Tech companies face tough AI copyright questions in 2025. https:\/\/www.reuters.com\/legal\/litigation\/tech-companies-face-tough-aicopyright-questions-2025-2024-12-27\/. December 27, 20247:18 PM GMT+8."},{"key":"e_1_3_3_1_4_2","volume-title":"30th USENIX Security Symposium (USENIX Security 21)","author":"Carlini Nicholas","year":"2021","unstructured":"Nicholas Carlini, Florian Tram\u00e8r, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, \u00dalfar Erlingsson, Alina Oprea, and Colin Raffel. 2021. Extracting Training Data from Large Language Models. In 30th USENIX Security Symposium (USENIX Security 21). USENIX Association, 2633\u20132650. https:\/\/www.usenix.org\/conference\/usenixsecurity21\/presentation\/carlini-extracting"},{"key":"e_1_3_3_1_5_2","doi-asserted-by":"publisher","DOI":"10.1145\/3658644.3690325"},{"key":"e_1_3_3_1_6_2","volume-title":"A method to facilitate membership inference attacks in deep learning models. arXiv preprint arXiv:2407.01919","author":"Chen Zitao","year":"2024","unstructured":"Zitao Chen and Karthik Pattabiraman. 2024. A method to facilitate membership inference attacks in deep learning models. arXiv preprint arXiv:2407.01919 (2024)."},{"key":"e_1_3_3_1_7_2","unstructured":"Debeshee Das Jie Zhang and Florian Tram\u00e8r. 2024. Blind Baselines Beat Membership Inference Attacks for Foundation Models. arXiv:2406.16201 [cs.CR] https:\/\/arxiv.org\/abs\/2406.16201"},{"key":"e_1_3_3_1_8_2","volume-title":"Do membership inference attacks work on large language models? arXiv preprint arXiv:2402.07841","author":"Duan Michael","year":"2024","unstructured":"Michael Duan, Anshuman Suri, Niloofar Mireshghallah, Sewon Min, Weijia Shi, Luke Zettlemoyer, Yulia Tsvetkov, Yejin Choi, David Evans, and Hannaneh Hajishirzi. 2024. Do membership inference attacks work on large language models? arXiv preprint arXiv:2402.07841 (2024)."},{"key":"e_1_3_3_1_9_2","volume-title":"The Pile: An 800GB Dataset of Diverse Text for Language Modeling. arXiv:2101.00027 [cs.CL] https:\/\/arxiv.org\/abs\/2101.00027","author":"Gao Leo","year":"2020","unstructured":"Leo Gao, Stella Biderman, Sid Black, Laurence Golding, Travis Hoppe, Charles Foster, Jason Phang, Horace He, Anish Thite, Noa Nabeshima, Shawn Presser, and Connor Leahy. 2020. The Pile: An 800GB Dataset of Diverse Text for Language Modeling. arXiv:2101.00027 [cs.CL] https:\/\/arxiv.org\/abs\/2101.00027"},{"key":"e_1_3_3_1_10_2","unstructured":"Daya Guo Qihao Zhu Dejian Yang Zhenda Xie Kai Dong Wentao Zhang Guanting Chen Xiao Bi Yu Wu YK Li et al. 2024. DeepSeek-Coder: When the Large Language Model Meets Programming\u2013The Rise of Code Intelligence. arXiv preprint arXiv:2401.14196 (2024)."},{"key":"e_1_3_3_1_11_2","unstructured":"Mathav Raj J Kushala VM Harikrishna Warrier and Yogesh Gupta. 2024. Fine Tuning LLM for Enterprise: Practical Guidelines and Recommendations. arXiv:2404.10779 [cs.SE] https:\/\/arxiv.org\/abs\/2404.10779"},{"key":"e_1_3_3_1_12_2","volume-title":"Scaling laws for neural language models. arXiv preprint arXiv:2001.08361","author":"Kaplan Jared","year":"2020","unstructured":"Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020. Scaling laws for neural language models. arXiv preprint arXiv:2001.08361 (2020)."},{"key":"e_1_3_3_1_13_2","volume-title":"Panoramia: privacy auditing of machine learning models without retraining. arXiv preprint arXiv:2402.09477","author":"Kazmi Mishaal","year":"2024","unstructured":"Mishaal Kazmi, Hadrien Lautraite, Alireza Akbari, Qiaoyue Tang, Mauricio Soroco, Tao Wang, S\u00e9bastien Gambs, and Mathias L\u00e9cuyer. 2024. Panoramia: privacy auditing of machine learning models without retraining. arXiv preprint arXiv:2402.09477 (2024)."},{"key":"e_1_3_3_1_14_2","volume-title":"https:\/\/www.wired.com\/story\/ai-copyright-case-tracker\/.","author":"Knibbs Kate","year":"2024","unstructured":"Kate Knibbs. 2020. Every AI Copyright Lawsuit in the US, Visualized. https:\/\/www.wired.com\/story\/ai-copyright-case-tracker\/. Dec 19, 2024 1:41 PM."},{"key":"e_1_3_3_1_15_2","doi-asserted-by":"publisher","DOI":"10.1145\/3658644.3690335"},{"key":"e_1_3_3_1_16_2","doi-asserted-by":"publisher","DOI":"10.1145\/3626772.3657807"},{"key":"e_1_3_3_1_17_2","volume-title":"The Thirty-eighth Annual Conference on Neural Information Processing Systems. https:\/\/openreview.net\/forum?id=Fr9d1UMc37","author":"Maini Pratyush","year":"2024","unstructured":"Pratyush Maini, Hengrui Jia, Nicolas Papernot, and Adam Dziedzic. 2024. LLM Dataset Inference: Did you train on my dataset?. In The Thirty-eighth Annual Conference on Neural Information Processing Systems. https:\/\/openreview.net\/forum?id=Fr9d1UMc37"},{"key":"e_1_3_3_1_18_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-acl.719"},{"key":"e_1_3_3_1_19_2","volume-title":"33rd USENIX Security Symposium (USENIX Security 24)","author":"Meeus Matthieu","year":"2024","unstructured":"Matthieu Meeus, Shubham Jain, Marek Rei, and Yves-Alexandre de Montjoye. 2024. Did the neurons read your book? document-level membership inference for large language models. In 33rd USENIX Security Symposium (USENIX Security 24). 2369\u20132385."},{"key":"e_1_3_3_1_20_2","unstructured":"OpenAI. 2024. Fine=Tuning. https:\/\/platform.openai.com\/docs\/guides\/finetuning."},{"key":"e_1_3_3_1_21_2","volume-title":"Reassessing EMNLP 2024\u2019s Best Paper: Does Divergence-Based Calibration for Membership Inference Attacks Hold Up? https:\/\/www.anshumansuri.com\/blog\/2024\/calibrated-mia\/.","author":"Pratyush Maini Anshuman Suri","year":"2024","unstructured":"Anshuman Suri Pratyush Maini. 2024. Reassessing EMNLP 2024\u2019s Best Paper: Does Divergence-Based Calibration for Membership Inference Attacks Hold Up? https:\/\/www.anshumansuri.com\/blog\/2024\/calibrated-mia\/."},{"key":"e_1_3_3_1_22_2","unstructured":"Haritz Puerto Martin Gubri Sangdoo Yun and Seong Joon Oh. 2024. Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models. arXiv:2411.00154 [cs.CL] https:\/\/arxiv.org\/abs\/2411.00154"},{"key":"e_1_3_3_1_23_2","volume-title":"Detecting pretraining data from large language models. arXiv preprint arXiv:2310.16789","author":"Shi Weijia","year":"2023","unstructured":"Weijia Shi, Anirudh Ajith, Mengzhou Xia, Yangsibo Huang, Daogao Liu, Terra Blevins, Danqi Chen, and Luke Zettlemoyer. 2023. Detecting pretraining data from large language models. arXiv preprint arXiv:2310.16789 (2023)."},{"key":"e_1_3_3_1_24_2","volume-title":"Mosaic Memory: Fuzzy Duplication in Copyright Traps for Large Language Models. arXiv preprint arXiv:2405.15523","author":"Shilov Igor","year":"2024","unstructured":"Igor Shilov, Matthieu Meeus, and Yves-Alexandre de Montjoye. 2024. Mosaic Memory: Fuzzy Duplication in Copyright Traps for Large Language Models. arXiv preprint arXiv:2405.15523 (2024)."},{"key":"e_1_3_3_1_25_2","doi-asserted-by":"crossref","unstructured":"Reza Shokri Marco Stronati Congzheng Song and Vitaly Shmatikov. 2017. Membership Inference Attacks against Machine Learning Models. arXiv:1610.05820 [cs.CR] https:\/\/arxiv.org\/abs\/1610.05820","DOI":"10.1109\/SP.2017.41"},{"key":"e_1_3_3_1_26_2","volume-title":"Will we run out of data? an analysis of the limits of scaling datasets in machine learning. arXiv preprint arXiv:2211.04325 1","author":"Villalobos Pablo","year":"2022","unstructured":"Pablo Villalobos, Jaime Sevilla, Lennart Heim, Tamay Besiroglu, Marius Hobbhahn, and Anson Ho. 2022. Will we run out of data? an analysis of the limits of scaling datasets in machine learning. arXiv preprint arXiv:2211.04325 1 (2022)."},{"key":"e_1_3_3_1_27_2","doi-asserted-by":"publisher","DOI":"10.1109\/CSF.2018.00027"},{"key":"e_1_3_3_1_28_2","volume-title":"Low-Cost High-Power Membership Inference Attacks. In Forty-first International Conference on Machine Learning.","author":"Zarifzadeh Sajjad","year":"2024","unstructured":"Sajjad Zarifzadeh, Philippe Liu, and Reza Shokri. 2024. Low-Cost High-Power Membership Inference Attacks. In Forty-first International Conference on Machine Learning."},{"key":"e_1_3_3_1_29_2","volume-title":"Pretraining data detection for large language models: A divergence-based calibration method. arXiv preprint arXiv:2409.14781","author":"Zhang Weichao","year":"2024","unstructured":"Weichao Zhang, Ruqing Zhang, Jiafeng Guo, Maarten de Rijke, Yixing Fan, and Xueqi Cheng. 2024. Pretraining data detection for large language models: A divergence-based calibration method. arXiv preprint arXiv:2409.14781 (2024)."},{"volume-title":"International Conference on Machine Learning (pp. 2397-2430)","author":"Bradley Q.G.","key":"e_1_3_3_1_30_2","unstructured":"Biderman, S., Schoelkopf, H., Anthony, Q.G., Bradley, H., O'Brien, K., Hallahan, E., Khan, M.A., Purohit, S., Prashanth, U.S., Raff, E. and Skowron, A., 2023, July. Pythia: A suite for analyzing large language models across training and scaling. In International Conference on Machine Learning (pp. 2397-2430). PMLR."},{"key":"e_1_3_3_1_31_2","volume-title":"Large scale autoregressive language modeling with mesh-tensorflow.\" If you use this software, please cite it using these metadata 58, no. 2","author":"Gao Leo","year":"2021","unstructured":"Black, Sid, Leo Gao, Phil Wang, Connor Leahy, and Stella Biderman. \"Gpt-neo: Large scale autoregressive language modeling with mesh-tensorflow.\" If you use this software, please cite it using these metadata 58, no. 2 (2021)."}],"event":{"name":"GAIIS 2025: 2025 2nd International Conference on Generative Artificial Intelligence and Information Security","acronym":"GAIIS 2025","location":"Hangzhou China"},"container-title":["Proceedings of the 2025 2nd International Conference on Generative Artificial Intelligence and Information Security"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3728725.3728799","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3728725.3728799","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T01:18:48Z","timestamp":1750295928000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3728725.3728799"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,2,21]]},"references-count":31,"alternative-id":["10.1145\/3728725.3728799","10.1145\/3728725"],"URL":"https:\/\/doi.org\/10.1145\/3728725.3728799","relation":{},"subject":[],"published":{"date-parts":[[2025,2,21]]},"assertion":[{"value":"2025-06-03","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}