{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T10:10:17Z","timestamp":1784196617857,"version":"3.55.0"},"reference-count":79,"publisher":"Association for Computing Machinery (ACM)","issue":"OOPSLA2","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Program. Lang."],"published-print":{"date-parts":[[2025,10,9]]},"abstract":"<jats:p>Code style transformation models built on code Language Models (code LMs) have achieved remarkable success. However, they typically focus on basis style transformations, where the target style follows a single criterion, and often struggle with combination styles, where the target style involves multiple criteria. In practice, style guides encompass multiple criteria, making the lack of effective combination style transformation a major limitation to their real-world applicability.<\/jats:p>\n                  <jats:p>\n                    In this paper, we propose\n                    <jats:sc>Absent-Basis-Combination<\/jats:sc>\n                    (abbreviated as ABC), a novel framework for code style transformation that significantly improves combination style transformation and overcomes the limitations of existing approaches. We implement four variants of ABC with parameter sizes of 0.5B, 1.3B, 1.5B, and 3B, demonstrating consistent superiority over existing approaches across all model sizes in both basis and combination style transformations. Specifically, ABC achieves performance gains of up to 86.70%, and remains superior even when baseline approaches use three times the parameters. Furthermore, to address the lack of high-quality datasets and evaluation metrics, we construct and release a new style transformation dataset,\n                    <jats:sc>Basis &amp; Combination Code Style<\/jats:sc>\n                    (abbreviated as BCC\n                    <jats:sc>Style<\/jats:sc>\n                    ), and introduce\n                    <jats:sc>Code Sequence, Syntactic, Semantic and Stylistic<\/jats:sc>\n                    BLEU (abbreviated as CS4BLEU), a novel code similarity metric that surpasses existing metrics in accuracy and consistency.\n                  <\/jats:p>","DOI":"10.1145\/3763104","type":"journal-article","created":{"date-parts":[[2025,10,9]],"date-time":"2025-10-09T08:49:50Z","timestamp":1759999790000},"page":"1512-1540","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["ABC: Towards a Universal Code Styler through Model Merging"],"prefix":"10.1145","volume":"9","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-4180-5811","authenticated-orcid":false,"given":"Yitong","family":"Chen","sequence":"first","affiliation":[{"name":"Southeast University, Nanjing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-7381-9349","authenticated-orcid":false,"given":"Zhiqiang","family":"Gao","sequence":"additional","affiliation":[{"name":"Southeast University, Nanjing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-8935-4862","authenticated-orcid":false,"given":"Chuanqi","family":"Shi","sequence":"additional","affiliation":[{"name":"Southeast University, Nanjing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-8560-5875","authenticated-orcid":false,"given":"Baixuan","family":"Li","sequence":"additional","affiliation":[{"name":"Southeast University, Nanjing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0003-5864-6445","authenticated-orcid":false,"given":"Miao","family":"Gao","sequence":"additional","affiliation":[{"name":"Southeast University, Nanjing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,10,9]]},"reference":[{"key":"e_1_3_2_2_1","unstructured":"Armen Aghajanyan Luke Zettlemoyer and Sonal Gupta. 2020. Intrinsic dimensionality explains the effectiveness of language model fine-tuning. arXiv preprint arXiv:2012.13255 (2020)."},{"key":"e_1_3_2_3_1","unstructured":"Takuya Akiba Makoto Shing Yujin Tang Qi Sun and David Ha. 2025. Evolutionary optimization of model merging recipes. Nature Machine Intelligence (2025) 1\u201310."},{"key":"e_1_3_2_4_1","doi-asserted-by":"crossref","unstructured":"Devansh Arpit Huan Wang Yingbo Zhou and Caiming Xiong. 2022. Ensemble of averages: Improving model selection and boosting performance in domain generalization. Advances in Neural Information Processing Systems 35 (2022) 8265\u20138277.","DOI":"10.52202\/068431-0601"},{"key":"e_1_3_2_5_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-20050-2_13"},{"key":"e_1_3_2_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/3586030"},{"key":"e_1_3_2_7_1","first-page":"257","volume-title":"In European Conference on Computer Vision","author":"Biggs Benjamin","year":"2024","unstructured":"Benjamin Biggs, Arjun Seshadri, Yang Zou, Achin Jain, Aditya Golatkar, Yusheng Xie, Alessandro Achille, Ashwin Swaminathan, and Stefano Soatto. 2024. Diffusion soup: Model merging for text-to-image diffusion models. In European Conference on Computer Vision. Springer, 257\u2013274."},{"key":"e_1_3_2_8_1","first-page":"1877","article-title":"Language models are few-shot learners","volume":"33","author":"Brown Tom","year":"2020","unstructured":"Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020. Language models are few-shot learners. Advances in neural information processing systems 33 (2020), 1877\u20131901.","journal-title":"Advances in neural information processing systems"},{"key":"e_1_3_2_9_1","first-page":"22405","article-title":"Swad: Domain generalization by seeking flat minima","volume":"34","author":"Cha Junbum","year":"2021","unstructured":"Junbum Cha, Sanghyuk Chun, Kyungjae Lee, Han-Cheol Cho, Seunghyun Park, Yunsung Lee, and Sungrae Park. 2021. Swad: Domain generalization by seeking flat minima. Advances in Neural Information Processing Systems 34 (2021), 22405\u201322418.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_10_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00198"},{"key":"e_1_3_2_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/ACCESS.2020.3024102"},{"key":"e_1_3_2_12_1","unstructured":"Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde De Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman et al. 2021. Evaluating large language models trained on code. arXiv preprint arXiv:2107.03374 (2021)."},{"key":"e_1_3_2_13_1","doi-asserted-by":"publisher","unstructured":"Yitong Chen. 2025. Artifact for \u201cABC: Towards a Universal Code Styler through Model Merging\u201d. https:\/\/doi.org\/10.1145\/3747412. doi:10.1145\/3747412","DOI":"10.1145\/3747412"},{"key":"e_1_3_2_14_1","unstructured":"Leshem Choshen Elad Venezian Noam Slonim and Yoav Katz. 2022. Fusing finetuned models for better pretraining. arXiv preprint arXiv:2204.03044 (2022)."},{"key":"e_1_3_2_15_1","volume-title":"Technical Report. Carnegie-Mellon University","author":"Coutaz Joelle","year":"1984","unstructured":"Joelle Coutaz et al. 1984. The box, a layout abstraction for user interface toolkits. Technical Report. Carnegie-Mellon University. Department of Computer Science."},{"key":"e_1_3_2_16_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.acllong.207"},{"key":"e_1_3_2_17_1","first-page":"270","volume-title":"In European Conference on Computer Vision","author":"Davari MohammadReza","year":"2024","unstructured":"MohammadReza Davari and Eugene Belilovsky. 2024. Model breadcrumbs: Scaling multi-task model merging with sparse masks. In European Conference on Computer Vision. Springer, 270\u2013287."},{"key":"e_1_3_2_18_1","unstructured":"Jasper Dekoninck Marc Fischer Luca Beurer-Kellner and Martin Vechev. 2023. Controlled text generation via language model arithmetic. arXiv preprint arXiv:2311.14479 (2023)."},{"key":"e_1_3_2_19_1","doi-asserted-by":"crossref","unstructured":"Shachar Don-Yehiya Elad Venezian Colin Raffel Noam Slonim Yoav Katz and Leshem Choshen. 2022. Cold fusion: Collaborative descent for distributed multitask finetuning. arXiv preprint arXiv:2212.01378 (2022).","DOI":"10.18653\/v1\/2023.acl-long.46"},{"key":"e_1_3_2_20_1","doi-asserted-by":"publisher","DOI":"10.1145\/3551349.3556903"},{"key":"e_1_3_2_21_1","unstructured":"Daya Guo Qihao Zhu Dejian Yang Zhenda Xie Kai Dong Wentao Zhang Guanting Chen Xiao Bi Yu Wu YK Li et al. 2024. DeepSeek-Coder: When the Large Language Model Meets Programming\u2014The Rise of Code Intelligence. arXiv preprint arXiv:2401.14196 (2024)."},{"key":"e_1_3_2_22_1","unstructured":"Vipul Gupta Santiago Akle Serrano and Dennis DeCoste. 2020. Stochastic Weight Averaging in Parallel: Large-Batch Training that Generalizes Well. International Conference on Learning Representations (2020)."},{"key":"e_1_3_2_23_1","unstructured":"Ari Holtzman Jan Buys Li Du Maxwell Forbes and Yejin Choi. 2020. The curious case of neural text degeneration. ICLR (2020)."},{"key":"e_1_3_2_24_1","unstructured":"Edward J Hu Yelong Shen Phillip Wallis Zeyuan Allen-Zhu Yuanzhi Li Shean Wang Lu Wang and Weizhu Chen. 2022. LoRA: Low-Rank Adaptation of Large Language Models. In International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=nZeVKeeFYf9"},{"key":"e_1_3_2_25_1","unstructured":"HuggingFace. 2024. Hugging Face Dataset Hub. https:\/\/huggingface.co\/datasets."},{"key":"e_1_3_2_26_1","unstructured":"Binyuan Hui Jian Yang Zeyu Cui Jiaxi Yang Dayiheng Liu Lei Zhang Tianyu Liu Jiajun Zhang Bowen Yu Keming Lu et al. 2024. Qwen2. 5-coder technical report. arXiv preprint arXiv:2409.12186 (2024)."},{"key":"e_1_3_2_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE-SEET58685.2023.00024"},{"key":"e_1_3_2_28_1","unstructured":"Gabriel Ilharco Marco Tulio Ribeiro Mitchell Wortsman Ludwig Schmidt Hannaneh Hajishirzi and Ali Farhadi. 2023. Editing models with task arithmetic. In The Eleventh International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=6t0Kwf8-jrj"},{"key":"e_1_3_2_29_1","unstructured":"Pavel Izmailov Dmitrii Podoprikhin Timur Garipov Dmitry Vetrov and Andrew Gordon Wilson. 2018. Averaging weights leads to wider optima and better generalization. In Conference on Uncertainty in Artificial Intelligence (UAI). https:\/\/arxiv.org\/abs\/1803.05407."},{"key":"e_1_3_2_30_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-1-4842-8844-3_4"},{"key":"e_1_3_2_31_1","unstructured":"Joel Jang Seungone Kim Bill Yuchen Lin Yizhong Wang Jack Hessel Luke Zettlemoyer Hannaneh Hajishirzi Yejin Choi and Prithviraj Ammanabrolu. 2023. Personalized soups: Personalized large language model alignment via post-hoc parameter merging. arXiv preprint arXiv:2310.11564 (2023)."},{"key":"e_1_3_2_32_1","unstructured":"Xisen Jin Xiang Ren Daniel Preotiuc-Pietro and Pengxiang Cheng. 2023. Dataless Knowledge Fusion by Merging Weights of Language Models. In The Eleventh International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=FCnohuR6AnM"},{"key":"e_1_3_2_33_1","unstructured":"Jared Kaplan Sam McCandlish Tom Henighan Tom B Brown Benjamin Chess Rewon Child Scott Gray Alec Radford Jeffrey Wu and Dario Amodei. 2020. Scaling laws for neural language models. arXiv preprint arXiv:2001.08361 (2020)."},{"key":"e_1_3_2_34_1","doi-asserted-by":"crossref","unstructured":"Vladimir Kovalenko Egor Bogomolov Timofey Bryksin and Alberto Bacchelli. 2020. Building implicit vector representations of individual coding style. In Proceedings of the IEEE\/ACM 42nd International Conference on Software Engineering Workshops. 117\u2013124.","DOI":"10.1145\/3387940.3391494"},{"key":"e_1_3_2_35_1","doi-asserted-by":"crossref","unstructured":"T Kudo. 2018. Sentencepiece: A simple and language independent subword tokenizer and detokenizer for neural text processing. arXiv preprint arXiv:1808.06226 (2018).","DOI":"10.18653\/v1\/D18-2012"},{"key":"e_1_3_2_36_1","unstructured":"Chris Lattner. 2006. Introduction to the LLVM compiler infrastructure. In Itanium conference and expo."},{"key":"e_1_3_2_37_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.acl-long.75"},{"key":"e_1_3_2_38_1","unstructured":"Margaret Li Suchin Gururangan Tim Dettmers Mike Lewis Tim Althoff Noah A Smith and Luke Zettlemoyer. 2022. Branch-train-merge: Embarrassingly parallel training of expert language models. arXiv preprint arXiv:2208.03306 (2022)."},{"key":"e_1_3_2_39_1","unstructured":"Pingzhi Li Zhenyu Zhang Prateek Yadav Yi-Lin Sung Yu Cheng Mohit Bansal and Tianlong Chen. 2023b. Merge then compress: Demystify efficient SMoe with hints from its routing policy. arXiv preprint arXiv:2310.01334 (2023)."},{"key":"e_1_3_2_40_1","unstructured":"Raymond Li Loubna Ben Allal Yangtian Zi Niklas Muennighoff Denis Kocetkov Chenghao Mou Marc Marone Christopher Akiki Jia Li Jenny Chim et al. 2023a. Starcoder: may the source be with you! arXiv preprint arXiv:2305.06161 (2023)."},{"key":"e_1_3_2_41_1","unstructured":"Xiang Li Kaixuan Huang Wenhao Yang Shusen Wang and Zhihua Zhang. 2019. On the Convergence of FedAvg on Non-IID Data. In International Conference on Learning Representations."},{"key":"e_1_3_2_42_1","unstructured":"Chin-Yew Lin. 2004. Rouge: A package for automatic evaluation of summaries. In Text summarization branches out. 74\u201381."},{"key":"e_1_3_2_43_1","unstructured":"Deyuan Liu Zecheng Wang Bingning Wang Weipeng Chen Chunshan Li Zhiying Tu Dianhui Chu Bo Li and Dianbo Sui. 2024. Checkpoint Merging via Bayesian Optimization in LLM Pretraining. arXiv preprint arXiv:2403.19390 (2024)."},{"key":"e_1_3_2_44_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-021-10107-0"},{"key":"e_1_3_2_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSR.2019.00073"},{"key":"e_1_3_2_46_1","doi-asserted-by":"publisher","DOI":"10.52202\/068431-1287"},{"key":"e_1_3_2_47_1","first-page":"1273","volume-title":"In Artificial intelligence and statistics","author":"McMahan Brendan","year":"2017","unstructured":"Brendan McMahan, Eider Moore, Daniel Ramage, Seth Hampson, and Blaise Aguera y Arcas. 2017. Communication-efficient learning of deep networks from decentralized data. In Artificial intelligence and statistics. PMLR, 1273\u20131282."},{"key":"e_1_3_2_48_1","unstructured":"Erik Nijkamp Bo Pang Hiroaki Hayashi Lifu Tu Huan Wang Yingbo Zhou Silvio Savarese and Caiming Xiong. 2023. CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis. ICLR (2023)."},{"key":"e_1_3_2_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/SANER.2018.8330253"},{"key":"e_1_3_2_50_1","unstructured":"Benjamin Paa\u00dfen. 2018. Revisiting the tree edit distance and its backtracing: A tutorial. arXiv preprint arXiv:1805.06869 (2018)."},{"key":"e_1_3_2_51_1","doi-asserted-by":"crossref","unstructured":"Kishore Papineni Salim Roukos Todd Ward and Wei-Jing Zhu. 2002. Bleu: a method for automatic evaluation of machine translation. In Proceedings of the 40th annual meeting of the Association for Computational Linguistics. 311\u2013318.","DOI":"10.3115\/1073083.1073135"},{"key":"e_1_3_2_52_1","doi-asserted-by":"crossref","unstructured":"Terence Parr and Jurgin Vinju. 2016. Technical report: Towards a universal code formatter through machine learning. arXiv preprint arXiv:1606.08866 (2016).","DOI":"10.1145\/2997364.2997383"},{"key":"e_1_3_2_53_1","article-title":"Pytorch: An imperative style, high-performance deep learning library","volume":"32","author":"Paszke Adam","year":"2019","unstructured":"Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, et al. 2019. Pytorch: An imperative style, high-performance deep learning library. Advances in neural information processing systems 32 (2019).","journal-title":"Advances in neural information processing systems"},{"key":"e_1_3_2_54_1","unstructured":"Alexandre Ram\u00e9 Kartik Ahuja Jianyu Zhang Matthieu Cord L\u00e9on Bottou and David Lopez-Paz. 2022. Model Ratatouille: Recycling Diverse Models for Out-of-Distribution Generalization. arXiv preprint arXiv:2212.10445 (2022)."},{"key":"e_1_3_2_55_1","unstructured":"Alexandre Ram\u00e9 Matthieu Kirchmeyer Thibaud Rahier Alain Rakotomamonjy Patrick Gallinari and Matthieu Cord. 2023. Diverse Weight Averaging for Out-of-Distribution Generalization. ICML (2023)."},{"key":"e_1_3_2_56_1","unstructured":"Shuo Ren Daya Guo Shuai Lu Long Zhou Shujie Liu Duyu Tang Neel Sundaresan Ming Zhou Ambrosio Blanco and Shuai Ma. 2020. Codebleu: a method for automatic evaluation of code synthesis. arXiv preprint arXiv:2009.10297 (2020)."},{"key":"e_1_3_2_57_1","doi-asserted-by":"publisher","DOI":"10.1109\/34.682181"},{"key":"e_1_3_2_58_1","unstructured":"Baptiste Roziere Jonas Gehring Fabian Gloeckle Sten Sootla Itai Gat Xiaoqing Ellen Tan Yossi Adi Jingyu Liu Romain Sauvestre Tal Remez et al. 2023. Code llama: Open foundation models for code. arXiv preprint arXiv:2308.12950 (2023)."},{"key":"e_1_3_2_59_1","doi-asserted-by":"publisher","DOI":"10.1213\/ANE.0000000000002864"},{"key":"e_1_3_2_60_1","first-page":"422","volume-title":"In European Conference on Computer Vision","author":"Shah Viraj","year":"2024","unstructured":"Viraj Shah, Nataniel Ruiz, Forrester Cole, Erika Lu, Svetlana Lazebnik, Yuanzhen Li, and Varun Jampani. 2024. Ziplora: Any subject in any style by effectively merging loras. In European Conference on Computer Vision. Springer, 422\u2013438."},{"key":"e_1_3_2_61_1","unstructured":"Bjarne Stroustrup Herb Sutter et al. 2018. C++ core guidelines. Web. Last accessed February (2018)."},{"key":"e_1_3_2_62_1","doi-asserted-by":"crossref","unstructured":"Yi-Lin Sung Linjie Li Kevin Lin Zhe Gan Mohit Bansal and Lijuan Wang. 2023. An Empirical Study of Multimodal Model Merging. Empirical Methods in Natural Language Processing (Findings) (2023).","DOI":"10.18653\/v1\/2023.findings-emnlp.105"},{"key":"e_1_3_2_63_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v37i13.27087"},{"key":"e_1_3_2_64_1","doi-asserted-by":"crossref","unstructured":"Adam Tornhill and Markus Borg. 2022. Code red: the business impact of code quality-a quantitative study of 39 proprietary production codebases. In Proceedings of the International Conference on Technical Debt. 11\u201320.","DOI":"10.1145\/3524843.3528091"},{"key":"e_1_3_2_65_1","doi-asserted-by":"publisher","DOI":"10.1109\/CSMR.2006.4"},{"key":"e_1_3_2_66_1","doi-asserted-by":"publisher","DOI":"10.1145\/226155.226156"},{"key":"e_1_3_2_67_1","doi-asserted-by":"publisher","DOI":"10.1016\/S1571-0661(04)80917-4"},{"key":"e_1_3_2_68_1","unstructured":"Yanlin Wang Tianyue Jiang Mingwei Liu Jiachi Chen and Zibin Zheng. 2024. Beyond functional correctness: Investigating coding style inconsistencies in large language models. arXiv preprint arXiv:2407.00456 (2024)."},{"key":"e_1_3_2_69_1","unstructured":"WebKit. 2025. WebKit Code Style Guidelines. Retrieved February 10 2025 from https:\/\/webkit.org\/code-style-guidelines\/"},{"key":"e_1_3_2_70_1","unstructured":"Jason Wei Yi Tay Rishi Bommasani Colin Raffel Barret Zoph Sebastian Borgeaud Dani Yogatama Maarten Bosma Denny Zhou Donald Metzler et al. 2022. Emergent abilities of large language models. arXiv preprint arXiv:2206.07682 (2022)."},{"key":"e_1_3_2_71_1","unstructured":"Benjy Weinberger Craig Silverstein Gregory Eitzmann Mark Mentovai and Tashana Landray. 2013. Google C++ style guide. Section: Line Length. url: http:\/\/google-styleguide. googlecode.com\/svn\/trunk\/cppguide. xml# Line_Length (2013)."},{"key":"e_1_3_2_72_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.emnlp-demos.6"},{"key":"e_1_3_2_73_1","first-page":"23965","volume-title":"In International conference on machine learning","author":"Wortsman Mitchell","year":"2022","unstructured":"Mitchell Wortsman, Gabriel Ilharco, Samir Ya Gadre, Rebecca Roelofs, Raphael Gontijo-Lopes, Ari S Morcos, Hongseok Namkoong, Ali Farhadi, Yair Carmon, Simon Kornblith, et al. 2022a. Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time. In International conference on machine learning. PMLR, 23965\u201323998."},{"key":"e_1_3_2_74_1","doi-asserted-by":"crossref","unstructured":"Mitchell Wortsman Gabriel Ilharco Jong Wook Kim Mike Li Simon Kornblith Rebecca Roelofs Raphael Gontijo Lopes Hannaneh Hajishirzi Ali Farhadi Hongseok Namkoong et al. 2022b. Robust fine-tuning of zero-shot models. In Proceedings of the IEEE\/CVF conference on computer vision and pattern recognition. 7959\u20137971.","DOI":"10.1109\/CVPR52688.2022.00780"},{"key":"e_1_3_2_75_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-acl.36"},{"key":"e_1_3_2_76_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-short.68"},{"key":"e_1_3_2_77_1","doi-asserted-by":"crossref","unstructured":"Prateek Yadav Derek Tam Leshem Choshen Colin A Raffel and Mohit Bansal. 2024. Ties-merging: Resolving interference when merging models. Advances in Neural Information Processing Systems 36 (2024).","DOI":"10.52202\/075280-0310"},{"key":"e_1_3_2_78_1","unstructured":"Enneng Yang Li Shen Guibing Guo Xingwei Wang Xiaochun Cao Jie Zhang and Dacheng Tao. 2024. Model merging in llms mllms and beyond: Methods theories applications and opportunities. arXiv preprint arXiv:2408.07666 (2024)."},{"key":"e_1_3_2_79_1","unstructured":"Enneng Yang Zhenyi Wang Li Shen Shiwei Liu Guibing Guo Xingwei Wang and Dacheng Tao. 2023. Adamerging: Adaptive model merging for multi-task learning. arXiv preprint arXiv:2310.02575 (2023)."},{"key":"e_1_3_2_80_1","unstructured":"Le Yu Bowen Yu Haiyang Yu Fei Huang and Yongbin Li. 2024. Language models are super mario: Absorbing abilities from homologous models as a free lunch. In Forty-first International Conference on Machine Learning."}],"container-title":["Proceedings of the ACM on Programming Languages"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3763104","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T10:02:41Z","timestamp":1784196161000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3763104"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,9]]},"references-count":79,"journal-issue":{"issue":"OOPSLA2","published-print":{"date-parts":[[2025,10,9]]}},"alternative-id":["10.1145\/3763104"],"URL":"https:\/\/doi.org\/10.1145\/3763104","relation":{},"ISSN":["2475-1421"],"issn-type":[{"value":"2475-1421","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,10,9]]},"assertion":[{"value":"2025-03-25","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-08-12","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-10-09","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}