{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,2]],"date-time":"2026-06-02T11:59:44Z","timestamp":1780401584146,"version":"3.54.1"},"publisher-location":"New York, NY, USA","reference-count":35,"publisher":"ACM","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,12,12]]},"DOI":"10.1145\/3788149.3788245","type":"proceedings-article","created":{"date-parts":[[2026,3,20]],"date-time":"2026-03-20T06:35:19Z","timestamp":1773988519000},"page":"18-23","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":1,"title":["Dynamic Soft Data Pruning: Adaptive Selection and Scheduling for Data-Efficient Learning"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-1205-5275","authenticated-orcid":false,"given":"Junda","family":"Yu","sequence":"first","affiliation":[{"name":"Fudan University, Shanghai, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0000-4652-9637","authenticated-orcid":false,"given":"Bangyin","family":"Xiang","sequence":"additional","affiliation":[{"name":"Communication University of China, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,3,19]]},"reference":[{"key":"e_1_3_3_1_2_2","unstructured":"Patryk Chrabaszcz Ilya Loshchilov and Frank Hutter. 2017. A downsampled variant of imagenet as an alternative to the cifar datasets. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1707.08819 (Aug 2017)."},{"key":"e_1_3_3_1_3_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR.2009.5206848"},{"key":"e_1_3_3_1_4_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00365"},{"key":"e_1_3_3_1_5_2","unstructured":"Vitaly Feldman and Chiyuan Zhang. 2020. What neural networks memorize and why: Discovering the long tail via influence estimation. Advances in Neural Information Processing Systems 33 (2020) 2881\u20132891."},{"key":"e_1_3_3_1_6_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPRW63382.2024.00767"},{"key":"e_1_3_3_1_7_2","first-page":"722","volume-title":"Algorithmic Learning Theory","author":"Iyer Rishabh","year":"2021","unstructured":"Rishabh Iyer, Ninad Khargoankar, Jeff Bilmes, and Himanshu Asanani. 2021. Submodular combinatorial information measures with applications in machine learning. In Algorithmic Learning Theory. PMLR, 722\u2013754."},{"key":"e_1_3_3_1_8_2","first-page":"15356","volume-title":"International conference on machine learning","author":"Joshi Siddharth","year":"2023","unstructured":"Siddharth Joshi and Baharan Mirzasoleiman. 2023. Data-efficient contrastive self-supervised learning: Most beneficial examples for supervised learning contribute the least. In International conference on machine learning. PMLR, 15356\u201315370."},{"key":"e_1_3_3_1_9_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v35i9.16988"},{"key":"e_1_3_3_1_10_2","unstructured":"Alex Krizhevsky Geoffrey Hinton et\u00a0al. 2009. Learning multiple layers of features from tiny images. (2009)."},{"key":"e_1_3_3_1_11_2","unstructured":"Shiye Lei and Dacheng Tao. 2023. A comprehensive survey to dataset distillation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2301.05603 (2023)."},{"key":"e_1_3_3_1_12_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52729.2023.00366"},{"key":"e_1_3_3_1_13_2","doi-asserted-by":"crossref","unstructured":"Tian\u00a0Yu Liu and Baharan Mirzasoleiman. 2022. Data-efficient augmentation for training neural networks. Advances in Neural Information Processing Systems 35 (2022) 5124\u20135136.","DOI":"10.52202\/068431-0370"},{"key":"e_1_3_3_1_14_2","unstructured":"Adyasha Maharana Prateek Yadav and Mohit Bansal. 2023. D2 pruning: Message passing for balancing diversity and difficulty in data pruning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.07931 (2023)."},{"key":"e_1_3_3_1_15_2","first-page":"6950","volume-title":"International Conference on Machine Learning","author":"Mirzasoleiman Baharan","year":"2020","unstructured":"Baharan Mirzasoleiman, Jeff Bilmes, and Jure Leskovec. 2020. Coresets for data-efficient training of machine learning models. In International Conference on Machine Learning. PMLR, 6950\u20136960."},{"key":"e_1_3_3_1_16_2","volume-title":"The Eleventh International Conference on Learning Representations","author":"Nohyun Ki","year":"2023","unstructured":"Ki Nohyun, Hoyong Choi, and Hye\u00a0Won Chung. 2023. Data Valuation Without Training of a Model. In The Eleventh International Conference on Learning Representations."},{"key":"e_1_3_3_1_17_2","unstructured":"Mansheej Paul Surya Ganguli and Gintare\u00a0Karolina Dziugaite. 2021. Deep learning on a data diet: Finding important examples early in training. Advances in Neural Information Processing Systems 34 (2021) 20596\u201320607."},{"key":"e_1_3_3_1_18_2","unstructured":"Ziheng Qin Kai Wang Zangwei Zheng Jianyang Gu Xiangyu Peng Zhaopan Xu Daquan Zhou Lei Shang Baigui Sun Xuansong Xie et\u00a0al. 2023. Infobatch: Lossless training speed up by unbiased dynamic data pruning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.04947 (2023)."},{"key":"e_1_3_3_1_19_2","unstructured":"Ravi\u00a0S Raju Kyle Daruwalla and Mikko Lipasti. 2021. Accelerating deep learning with dynamic data pruning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2111.12621 (2021)."},{"key":"e_1_3_3_1_20_2","unstructured":"Srikumar Ramalingam Pranjal Awasthi and Sanjiv Kumar. 2023. A Weighted K-Center Algorithm for Data Subset Selection. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.10602 (2023)."},{"key":"e_1_3_3_1_21_2","volume-title":"Advances in Neural Information Processing Systems","author":"Sorscher Ben","year":"2022","unstructured":"Ben Sorscher, Robert Geirhos, Shashank Shekhar, Surya Ganguli, and Ari\u00a0S. Morcos. 2022. Beyond neural scaling laws: beating power law scaling via data pruning. In Advances in Neural Information Processing Systems, Alice\u00a0H. Oh, Alekh Agarwal, Danielle Belgrave, and Kyunghyun Cho (Eds.)."},{"key":"e_1_3_3_1_22_2","unstructured":"Haoru Tan Sitong Wu Fei Du Yukang Chen Zhibin Wang Fan Wang and Xiaojuan Qi. 2024. Data pruning via moving-one-sample-out. Advances in Neural Information Processing Systems 36 (2024)."},{"key":"e_1_3_3_1_23_2","unstructured":"Mariya Toneva Alessandro Sordoni Remi Tachet\u00a0des Combes Adam Trischler Yoshua Bengio and Geoffrey\u00a0J Gordon. 2018. An empirical study of example forgetting during deep neural network learning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/1812.05159 (2018)."},{"key":"e_1_3_3_1_24_2","doi-asserted-by":"publisher","DOI":"10.1145\/1553374.1553517"},{"key":"e_1_3_3_1_25_2","volume-title":"The Eleventh International Conference on Learning Representations","author":"Xia Xiaobo","year":"2023","unstructured":"Xiaobo Xia, Jiale Liu, Jun Yu, Xu Shen, Bo Han, and Tongliang Liu. 2023. Moderate Coreset: A Universal Method of Data Selection for Real-world Data-efficient Deep Learning. In The Eleventh International Conference on Learning Representations."},{"key":"e_1_3_3_1_26_2","unstructured":"Enneng Yang Li Shen Zhenyi Wang Tongliang Liu and Guibing Guo. 2023. An efficient dataset condensation plugin and its application to continual learning. Advances in Neural Information Processing Systems 36 (2023)."},{"key":"e_1_3_3_1_27_2","unstructured":"Suorong Yang Peijia Li Yujie Liu Zhiming Xu Peng Ye Wanli Ouyang Furao Shen and Dongzhan Zhou. 2025. Multimodal-Guided Dynamic Dataset Pruning for Robust and Efficient Data-Centric Learning. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2507.12750 (2025)."},{"key":"e_1_3_3_1_28_2","volume-title":"International Conference on Learning Representations","author":"Yang Shuo","year":"2023","unstructured":"Shuo Yang, Zeke Xie, Hanyu Peng, Min Xu, Mingming Sun, and Ping Li. 2023. Dataset Pruning: Reducing Training Data by Examining Generalization Influence. In International Conference on Learning Representations."},{"key":"e_1_3_3_1_29_2","unstructured":"Suorong Yang Hongchao Yang Suhan Guo Furao Shen and Jian Zhao. 2023. Not All Data Matters: An End-to-End Adaptive Dataset Pruning Framework for Enhancing Model Performance and Efficiency. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.05599 (2023)."},{"key":"e_1_3_3_1_30_2","unstructured":"Suorong Yang Peng Ye Wanli Ouyang Dongzhan Zhou and Furao Shen. 2024. A CLIP-Powered Framework for Robust and Generalizable Data Selection. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.11215 (2024)."},{"key":"e_1_3_3_1_31_2","unstructured":"Suorong Yang Peng Ye Furao Shen and Dongzhan Zhou. 2025. When Dynamic Data Selection Meets Data Augmentation. arxiv:https:\/\/arXiv.org\/abs\/2505.03809\u00a0[cs.LG] https:\/\/arxiv.org\/abs\/2505.03809"},{"key":"e_1_3_3_1_32_2","first-page":"39314","volume-title":"International Conference on Machine Learning","author":"Yang Yu","year":"2023","unstructured":"Yu Yang, Hao Kang, and Baharan Mirzasoleiman. 2023. Towards sustainable learning: Coresets for data-efficient deep learning. In International Conference on Machine Learning. PMLR, 39314\u201339330."},{"key":"e_1_3_3_1_33_2","doi-asserted-by":"publisher","DOI":"10.1109\/CVPR52733.2024.02477"},{"key":"e_1_3_3_1_34_2","first-page":"12674","volume-title":"International Conference on Machine Learning","author":"Zhao Bo","year":"2021","unstructured":"Bo Zhao and Hakan Bilen. 2021. Dataset condensation with differentiable siamese augmentation. In International Conference on Machine Learning. PMLR, 12674\u201312685."},{"key":"e_1_3_3_1_35_2","volume-title":"The Eleventh International Conference on Learning Representations","author":"Zheng Haizhong","year":"2023","unstructured":"Haizhong Zheng, Rui Liu, Fan Lai, and Atul Prakash. 2023. Coverage-centric Coreset Selection for High Pruning Rates. In The Eleventh International Conference on Learning Representations."},{"key":"e_1_3_3_1_36_2","unstructured":"Yongchao Zhou Ehsan Nezhadarya and Jimmy Ba. 2022. Dataset Distillation using Neural Feature Regression. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2206.00719 (2022)."}],"event":{"name":"CSAI 2025: 2025 The 9th International Conference on Computer Science and Artificial Intelligence","location":"Beijing China","acronym":"CSAI 2025"},"container-title":["Proceedings of the 2025 9th International Conference on Computer Science and Artificial Intelligence"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3788149.3788245","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,3,20]],"date-time":"2026-03-20T06:37:38Z","timestamp":1773988658000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3788149.3788245"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,12]]},"references-count":35,"alternative-id":["10.1145\/3788149.3788245","10.1145\/3788149"],"URL":"https:\/\/doi.org\/10.1145\/3788149.3788245","relation":{},"subject":[],"published":{"date-parts":[[2025,12,12]]},"assertion":[{"value":"2026-03-19","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}