{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T07:07:24Z","timestamp":1784185644995,"version":"3.55.0"},"reference-count":117,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:00:00Z","timestamp":1784160000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:00:00Z","timestamp":1784160000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Karlsruher Institut f\u00fcr Technologie (KIT)"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Empir Software Eng"],"published-print":{"date-parts":[[2027,2]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>The application of Large Language Models (LLMs) in Model-Driven Engineering (MDE) has emerged as a rapidly evolving research area. While existing systematic literature reviews have examined specific technical approaches, a comprehensive mapping of the broader research landscape (e.g., development trends) remains lacking. This study presents a systematic mapping study of LLM applications in MDE, analyzing 86 primary studies collected from five databases, covering publications from 2022 to early 2026. Guided by five research questions, we characterize the field across five dimensions: MDE task distribution and research contribution types, LLM technologies and interaction strategies, artifact representation and processing, validation practices, and publication landscape. Our findings reveal that current LLM4MDE research is heavily concentrated on Model Generation, while tasks such as Model Migration, DSL Engineering, and Metamodeling remain marginal. Most approaches rely on black-box OpenAI models accessed via remote APIs and adapted through prompt engineering, with fine-tuning and retrieval-augmented generation rarely employed. Inputs are predominantly natural-language artifacts, while outputs are model-oriented but usually expressed in lightweight textual formats rather than native MDE exchange formats. Validation is centered on quantitative experimentation, with 42% of studies reporting no baseline and cost efficiency reported in fewer than one quarter of studies. The field has grown rapidly, from one paper in 2022 to 42 in 2025, with research concentrated in Europe and Canada and limited industry involvement. Based on these findings, we identify gaps and opportunities across task coverage, technical configuration, and evaluation practice, offering a knowledge map to guide future work in this cross-disciplinary field.<\/jats:p>","DOI":"10.1007\/s10664-026-10921-4","type":"journal-article","created":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T06:53:47Z","timestamp":1784184827000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Large language models in model-driven engineering: a systematic mapping study"],"prefix":"10.1007","volume":"32","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-2890-6034","authenticated-orcid":false,"given":"Weixing","family":"Zhang","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bowen","family":"Jiang","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yuhong","family":"Fu","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haowei","family":"Cheng","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Maximilian","family":"Hummel","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Vincenzo","family":"Scotti","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Nathan","family":"Hagel","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jialong","family":"Li","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Georg","family":"Grossmann","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Markus","family":"Stumptner","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Regina","family":"Hebig","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Daniel","family":"Str\u00fcber","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Anne","family":"Koziolek","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,7,16]]},"reference":[{"key":"10921_CR1","unstructured":"Achiam J, Adler S, Agarwal S, Ahmad L, Akkaya I, Aleman FL, Almeida D, Altenschmidt J, Altman S, Anadkat S (2023) GPT-4 technical report. arXiv:2303.08774"},{"key":"10921_CR2","doi-asserted-by":"crossref","unstructured":"ACM SIGSOFT (2021) Empirical standards for software engineering. https:\/\/www2.sigsoft.org\/EmpiricalStandards\/, accessed: 19 Sept 2025","DOI":"10.1145\/1010763.1010767"},{"key":"10921_CR3","doi-asserted-by":"crossref","unstructured":"Ahmed K, Song J, Chen B, Wei O, Zheng B (2025) MCeT: behavioral model correctness evaluation using large language models. In: 2025 ACM\/IEEE 28th international conference on model driven engineering languages and systems (MODELS). IEEE, pp 84\u201395","DOI":"10.1109\/MODELS67397.2025.00014"},{"key":"10921_CR4","doi-asserted-by":"publisher","unstructured":"Alhanahnah M, Rashedul Hasan M, Xu L, Bagheri H (2025) An empirical evaluation of pre-trained large language models for repairing declarative formal specifications. Empir Softw Eng 30(5):149. https:\/\/doi.org\/10.1007\/s10664-025-10687-1","DOI":"10.1007\/s10664-025-10687-1"},{"issue":"4","key":"10921_CR5","doi-asserted-by":"publisher","first-page":"469","DOI":"10.1093\/reseval\/rvaa038","volume":"29","author":"B \u2019Alvarez Bornstein","year":"2020","unstructured":"\u2019Alvarez Bornstein B, Montesi M (2020) Funding acknowledgements in scientific publications: a literature review. Res Eval 29(4):469\u2013488. https:\/\/doi.org\/10.1093\/reseval\/rvaa038","journal-title":"Res Eval"},{"key":"10921_CR6","doi-asserted-by":"crossref","unstructured":"Amalfitano D, Metzger A, Autili M, Fulcini T, Hey T, Keim J, Pelliccione P, Scotti V, Koziolek A, Mirandola R (2026) A research roadmap for augmenting software engineering processes and software products with generative AI. ACM Transactions on Software Engineering and Methodology","DOI":"10.1145\/3788879"},{"key":"10921_CR7","unstructured":"Anthropic (2023) Introducing Claude. https:\/\/www.anthropic.com\/news\/introducing-claude"},{"key":"10921_CR8","unstructured":"Baltes S, Angermeir F, Arora C, Mu\u00f1oz\u00a0Bar\u00f3n M, Chen C, B\u00f6hme L, Calefato F, Ernst N, Falessi D, Fitzgerald B et\u00a0al (2025) Guidelines for empirical studies in software engineering involving large language models. arXiv e-prints pp arXiv\u20132508"},{"key":"10921_CR9","doi-asserted-by":"crossref","unstructured":"Bamouh L, B\u00e9ziers La Fosse T, Tisi M (2025) Towards diagram-based data model generation with LLMs. International Conference on Conceptual Modeling. Springer, pp 157\u2013175","DOI":"10.1007\/978-3-032-08620-4_10"},{"key":"10921_CR10","unstructured":"Brown TB, Mann B, Ryder N, Subbiah M, Kaplan J, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, Agarwal S, Herbert-Voss A, Krueger G, Henighan T, Child R, Ramesh A, Ziegler DM, Wu J, Winter C, Hesse C, Chen M, Sigler E, Litwin M, Gray S, Chess B, Clark J, Berner C, McCandlish S, Radford A, Sutskever I, Amodei D et\u00a0al (2020) Language models are few-shot learners. In: Larochelle H, Ranzato M, Hadsell R, Balcan M, Lin H (eds) Advances in neural information processing systems 33: annual conference on neural information processing systems 2020, NeurIPS 2020, December 6-12, 2020, virtual. https:\/\/proceedings.neurips.cc\/paper\/2020\/hash\/1457c0d6bfcb4967418bfb8ac142f64a-Abstract.html"},{"key":"10921_CR11","doi-asserted-by":"publisher","unstructured":"Brunello N (2026) 6 - trustworthiness of large language models: hallucinations. In: Pillai AS, Tedesco R, Scotti V (eds) Challenges and Applications of Generative Large Language Models, Morgan Kaufmann, pp 107\u2013126. https:\/\/doi.org\/10.1016\/B978-0-443-33592-1.00007-3","DOI":"10.1016\/B978-0-443-33592-1.00007-3"},{"issue":"5","key":"10921_CR12","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3712008","volume":"34","author":"L Burgue\u00f1o","year":"2025","unstructured":"Burgue\u00f1o L, Di Ruscio D, Sahraoui H, Wimmer M (2025) Automation in model-driven engineering: a look back, and ahead. ACM Trans Softw Eng Methodol 34(5):1\u201325","journal-title":"ACM Trans Softw Eng Methodol"},{"key":"10921_CR13","doi-asserted-by":"crossref","unstructured":"Cabot J, Gogolla M (2012) Object constraint language (OCL): a definitive guide. International school on formal methods for the design of computer, communication and software systems. Springer, pp 58\u201390","DOI":"10.1007\/978-3-642-30982-3_3"},{"key":"10921_CR14","doi-asserted-by":"crossref","unstructured":"C\u00e1mara J, Troya J, Burgue\u00f1o L, Vallecillo A (2023) On the assessment of generative AI in modeling tasks: an experience report with ChatGPT and UML: Jj. c\u00e1mara et al. Softw Syst Model 22(3):781\u2013793","DOI":"10.1007\/s10270-023-01105-5"},{"key":"10921_CR15","unstructured":"Chen M, Tworek J, Jun H, Yuan Q, Pinto HPDO, Kaplan J, Edwards H, Burda Y, Joseph N, Brockman G et\u00a0al (2021a) Evaluating large language models trained on code. arXiv:2107.03374"},{"key":"10921_CR16","unstructured":"Chen M, Tworek J, Jun H, Yuan Q et\u00a0al (2021b) Evaluating large language models trained on code. arXiv:2107.03374"},{"key":"10921_CR17","doi-asserted-by":"crossref","unstructured":"Chen B, Chen K, Hassani S, Yang Y, Amyot D, Lessard L, Mussbacher G, Sabetzadeh M, Varr\u00f3 D (2023) On the use of GPT-4 for creating goal models: an exploratory study. In: 2023 IEEE 31st international requirements engineering conference workshops (REW), IEEE, pp 262\u2013271","DOI":"10.1109\/REW57809.2023.00052"},{"key":"10921_CR18","unstructured":"Chen L, Guo Q, Jia H, Zeng Z, Wang X, Xu Y, Wu J, Wang Y, Gao Q, Wang J et\u00a0al (2024) A survey on evaluating large language models in code generation tasks. arXiv:2408.16498"},{"issue":"2","key":"10921_CR19","doi-asserted-by":"publisher","first-page":"141","DOI":"10.1002\/spe.70029","volume":"56","author":"H Cheng","year":"2026","unstructured":"Cheng H, Husen JH, Lu Y, Racharak T, Yoshioka N, Ubayashi N, Washizaki H (2026) Generative AI for requirements engineering: a systematic literature review. Softw Pract Exper 56(2):141\u2013170","journal-title":"Softw Pract Exper"},{"key":"10921_CR20","doi-asserted-by":"publisher","unstructured":"Chung HW, Hou L, Longpre S, Zoph B, Tay Y, Fedus W, Li E, Wang X, Dehghani M, Brahma S, Webson A, Gu SS, Dai Z, Suzgun M, Chen X, Chowdhery A, Narang S, Mishra G, Yu A, andf Yanping\u00a0Huang VYZ, Dai AM, Yu H, Petrov S, Chi EH, Dean J, Devlin J, Roberts A, Zhou D, Le QV, Wei J (2022) Scaling instruction-finetuned language models. https:\/\/doi.org\/10.48550\/arXiv.2210.11416","DOI":"10.48550\/arXiv.2210.11416"},{"key":"10921_CR21","unstructured":"Clarivate (2025) Journal citation reports. https:\/\/clarivate.com\/academia-government\/scientific-and-academic-research\/research-funding-analytics\/journal-citation-reports\/"},{"issue":"1","key":"10921_CR22","doi-asserted-by":"publisher","first-page":"37","DOI":"10.1177\/001316446002000104","volume":"20","author":"J Cohen","year":"1960","unstructured":"Cohen J (1960) A coefficient of agreement for nominal scales. Educ Psychol Measur 20(1):37\u201346","journal-title":"Educ Psychol Measur"},{"key":"10921_CR23","unstructured":"Comanici G, Bieber E, Schaekermann M, Pasupat I, Sachdeva N, Dhillon I, Blistein M, Ram O, Zhang D, Rosen E, Marris L, Petulla S, Gaffney C, Aharoni A, Lintz N, Pais TC, Jacobsson H, Szpektor I, Jiang NJ, Haridasan K, Omran A, Saunshi N, Bahri D, Mishra G, Chu E, Boyd T, Hekman B, Parisi A, Zhang C, Kawintiranon K, Bedrax-Weiss T et\u00a0al (2025) Gemini 2.5: pushing the frontier with advanced reasoning, multimodality, long context, and next generation agentic capabilities. arXiv:2507.06261"},{"key":"10921_CR24","unstructured":"Computing Research and Education Association of Australasia (2026) ICORE2026 conference rankings via the ICORE conference portal. https:\/\/portal.core.edu.au\/conf-ranks\/"},{"issue":"11","key":"10921_CR25","doi-asserted-by":"publisher","first-page":"3056","DOI":"10.1109\/TSE.2025.3607625","volume":"51","author":"S Corbo","year":"2025","unstructured":"Corbo S, Bancale L, Gennaro VD, Lestingi L, Scotti V, Camilli M (2025) How toxic can you get? search-based toxicity testing for large language models. IEEE Trans Softw Eng 51(11):3056\u20133071. https:\/\/doi.org\/10.1109\/TSE.2025.3607625","journal-title":"IEEE Trans Softw Eng"},{"issue":"3","key":"10921_CR26","doi-asserted-by":"publisher","first-page":"621","DOI":"10.1147\/sj.453.0621","volume":"45","author":"K Czarnecki","year":"2006","unstructured":"Czarnecki K, Helsen S (2006) Feature-based survey of model transformation approaches. IBM Syst J 45(3):621\u2013645","journal-title":"IBM Syst J"},{"key":"10921_CR27","doi-asserted-by":"publisher","unstructured":"Devlin J, Chang M, Lee K, Toutanova K (2019) BERT: pre-training of deep bidirectional transformers for language understanding. In: Burstein J, Doran C, Solorio T (eds) Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics: human language technologies, NAACL-HLT 2019, Minneapolis, MN, USA, June 2-7, 2019, Volume 1 (Long and Short Papers), Association for Computational Linguistics, pp 4171\u20134186. https:\/\/doi.org\/10.18653\/v1\/n19-1423","DOI":"10.18653\/v1\/n19-1423"},{"key":"10921_CR28","doi-asserted-by":"crossref","unstructured":"Di Rocco J, Iovino L, Pierantonio A (2012) Bridging state-based differencing and co-evolution. In: Proceedings of the 6th international workshop on models and evolution, pp 15\u201320","DOI":"10.1145\/2523599.2523603"},{"key":"10921_CR29","doi-asserted-by":"crossref","unstructured":"Di Rocco J, Di Ruscio D, Di Sipio C, Nguyen PT, Rubei R (2025) On the use of large language models in model-driven engineering. J Di Rocco et al. Softw Syst Model 24(3):923\u2013948","DOI":"10.1007\/s10270-025-01263-8"},{"key":"10921_CR30","doi-asserted-by":"publisher","first-page":"285","DOI":"10.1016\/j.jbusres.2021.04.070","volume":"133","author":"N Donthu","year":"2021","unstructured":"Donthu N, Kumar S, Mukherjee D, Pandey N, Lim WM (2021) How to conduct a bibliometric analysis: an overview and guidelines. J Bus Res 133:285\u2013296","journal-title":"J Bus Res"},{"key":"10921_CR31","doi-asserted-by":"crossref","unstructured":"Dvivedi SS, Vijay V, Pujari SLR, Lodh S, Kumar D 2024) A comparative analysis of large language models for code documentation generation. In: Proceedings of the 1st ACM international conference on AI-powered software, pp 65\u201373","DOI":"10.1145\/3664646.3664765"},{"key":"10921_CR32","unstructured":"EAST-ADL Association (2021) East-ADL. https:\/\/www.east-adl.info\/, Accessed February, 2023"},{"issue":"4","key":"10921_CR33","doi-asserted-by":"publisher","first-page":"479","DOI":"10.1007\/s10270-008-0095-y","volume":"8","author":"K Ehrig","year":"2009","unstructured":"Ehrig K, K\u00fcster JM, Taentzer G (2009) Generating instance models from meta models. Softw Syst Model 8(4):479\u2013500","journal-title":"Softw Syst Model"},{"key":"10921_CR34","doi-asserted-by":"publisher","unstructured":"Eysholdt M, Behrens H (2010) Xtext: implement your language faster than the quick and dirty way. In: Proceedings of the ACM international conference companion on object oriented programming systems languages and applications, ACM, OOPSLA \u201910, pp 307\u2013309. https:\/\/doi.org\/10.1145\/1869542.1869625","DOI":"10.1145\/1869542.1869625"},{"key":"10921_CR35","doi-asserted-by":"crossref","unstructured":"Fan A, Gokkaya B, Harman M, Lyubarskiy M, Sengupta S, Yoo S, Zhang JM (2023) Large language models for software engineering: Survey and open problems. In: 2023 IEEE\/ACM international conference on software engineering: future of software engineering. IEEE, pp 31\u201353","DOI":"10.1109\/ICSE-FoSE59343.2023.00008"},{"key":"10921_CR36","doi-asserted-by":"publisher","first-page":"107697","DOI":"10.1016\/j.infsof.2025.107697","volume":"181","author":"A Ferrari","year":"2025","unstructured":"Ferrari A, Spoletini P (2025) Formal requirements engineering and large language models: a two-way roadmap. Inf Softw Technol 181:107697. https:\/\/doi.org\/10.1016\/j.infsof.2025.107697","journal-title":"Inf Softw Technol"},{"key":"10921_CR37","doi-asserted-by":"crossref","unstructured":"Fleurey F, Steel J, Baudry B (2004) Validation in model-driven engineering: testing model transformations. In: Proceedings. 2004 First international workshop on model, design and validation. IEEE, pp 29\u201340","DOI":"10.1109\/MODEVA.2004.1425846"},{"key":"10921_CR38","volume-title":"Domain-Specific Languages","author":"M Fowler","year":"2010","unstructured":"Fowler M (2010) Domain-Specific Languages. Portable Documents, Pearson Education"},{"key":"10921_CR39","doi-asserted-by":"publisher","unstructured":"Gao Y, Xiong Y, Gao X, Jia K, Pan J, Bi Y, Dai Y, Sun J, Guo Q, Wang M, Wang H (2023) Retrieval-augmented generation for large language models: a survey. https:\/\/doi.org\/10.48550\/arXiv.2312.10997","DOI":"10.48550\/arXiv.2312.10997"},{"issue":"2","key":"10921_CR40","doi-asserted-by":"publisher","first-page":"685","DOI":"10.1007\/s11219-016-9350-6","volume":"26","author":"FD Giraldo","year":"2018","unstructured":"Giraldo FD, Espana S, Pastor O, Giraldo WJ (2018) Considerations about quality in model-driven engineering: current state and challenges. Software Qual J 26(2):685\u2013750","journal-title":"Software Qual J"},{"key":"10921_CR41","unstructured":"GitHub (2021) Introducing Github copilot: your AI pair programmer. https:\/\/github.blog\/news-insights\/product-news\/introducing-github-copilot-ai-pair-programmer\/"},{"key":"10921_CR42","first-page":"64","volume-title":"Systems","author":"MK G\u00f6rmez","year":"2024","unstructured":"G\u00f6rmez MK, Y\u0131lmaz M, Clarke PM (2024) Large language models for software engineering: A systematic mapping study. In: Yilmaz M, Clarke P, Riel A, Messnarz R, Greiner C, Peisl T (eds) Systems. Software and Services Process Improvement, Springer Nature Switzerland, Cham, pp 64\u201379"},{"key":"10921_CR43","doi-asserted-by":"publisher","unstructured":"Grattafiori A, Dubey A, Jauhri A, Pandey A, Kadian A, Al-Dahle A, Letman A, Mathur A, Schelten A, Yang A, Fan A, Goyal A, Hartshorn A, Yang A, Mitra A, Sravankumar A, Korenev A, Hinsvark A, Rao A, Zhang A, Rodriguez A, Gregerson A, Spataru A, Rozi\u00e8re B, Biron B, Tang B, Chern B, Caucheteux C, Nayak C, Bi C, Marra C et\u00a0al. (2024) The Llama 3 herd of models. https:\/\/doi.org\/10.48550\/arXiv.2407.21783","DOI":"10.48550\/arXiv.2407.21783"},{"key":"10921_CR44","unstructured":"Guo D, Zhu Q, Yang D, Xie Z, Dong K, Zhang W, Chen G, Bi X, Wu Y, Li YK (2024) DeepSeek-coder: when the large language model meets programming - the rise of code intelligence. arXiv:2401.14196"},{"key":"10921_CR45","doi-asserted-by":"publisher","unstructured":"Guo D, Yang D, Zhang H, Song J, Wang P, Zhu Q, Xu R, Zhang R, Ma S, Bi X, Zhang X, Yu X, Wu Y, Wu ZF, Gou Z, Shao Z, Li Z, Gao Z, Liu A, Xue B, Wang B, Wu B, Feng B, Lu C, Zhao C, Deng C, Ruan C, Dai D, Chen D, Ji D, Li E et al (2025) DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning. Nature 645(8081):633\u2013638. https:\/\/doi.org\/10.1038\/s41586-025-09422-z","DOI":"10.1038\/s41586-025-09422-z"},{"key":"10921_CR46","doi-asserted-by":"crossref","unstructured":"Hachm Z, Le Calvar T, Bruneliere H, Tisi M (2025) Towards LLM agents for model-based engineering: a case in transformation selection. In: 2025 ACM\/IEEE 28th international conference on model driven engineering languages and systems companion (MODELS-C). IEEE, pp 421\u2013431","DOI":"10.1109\/MODELS-C68889.2025.00061"},{"issue":"8","key":"10921_CR47","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3695988","volume":"33","author":"X Hou","year":"2024","unstructured":"Hou X, Zhao Y, Liu Y, Yang Z, Wang K, Li L, Luo X, Lo D, Grundy J, Wang H (2024) Large language models for software engineering: a systematic literature review. ACM Trans Softw Eng Methodol 33(8):1\u201379","journal-title":"ACM Trans Softw Eng Methodol"},{"issue":"2","key":"10921_CR48","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3703155","volume":"43","author":"L Huang","year":"2025","unstructured":"Huang L, Yu W, Ma W, Zhong W, Feng Z, Wang H, Chen Q, Peng W, Feng X, Qin B et al (2025) A survey on hallucination in large language models: principles, taxonomy, challenges, and open questions. ACM Trans Inf Syst 43(2):1\u201355","journal-title":"ACM Trans Inf Syst"},{"key":"10921_CR49","doi-asserted-by":"crossref","unstructured":"Hummel P, Grimstad J, Xia Y, Morozov A (2024) Generation of MATLAB Simscape models using large language models for training reliability evaluation methods. In: ASME international mechanical engineering congress and exposition, American society of mechanical engineers, vol 88698, p V011T14A010","DOI":"10.1115\/IMECE2024-144344"},{"key":"10921_CR50","doi-asserted-by":"crossref","unstructured":"Hutchinson J, Whittle J, Rouncefield M, Kristoffersen S (2011) Empirical assessment of MDE in industry. In: Proceedings of the 33rd international conference on software engineering, pp 471\u2013480","DOI":"10.1145\/1985793.1985858"},{"key":"10921_CR51","unstructured":"ISO (2024) ISO\/TR 8344:2024(en) Information and documentation \u2014 Issues and considerations for managing records in structured data environments. https:\/\/www.iso.org\/obp\/ui\/es\/#iso:std:iso:tr:8344:ed-1:v1:en, accessed: 25 Mar 2026"},{"key":"10921_CR52","doi-asserted-by":"publisher","unstructured":"Jiang AQ, Sablayrolles A, Mensch A, Bamford C, Chaplot DS, de\u00a0Las\u00a0Casas D, Bressand F, Lengyel G, Lample G, Saulnier L, Lavaud LR, Lachaux M, Stock P, Scao TL, Lavril T, Wang T, Lacroix T, Sayed WE (2023) Mistral 7B. https:\/\/doi.org\/10.48550\/arXiv.2310.06825","DOI":"10.48550\/arXiv.2310.06825"},{"issue":"1\u20132","key":"10921_CR53","doi-asserted-by":"publisher","first-page":"31","DOI":"10.1016\/j.scico.2007.08.002","volume":"72","author":"F Jouault","year":"2008","unstructured":"Jouault F, Allilaire F, B\u00e9zivin J, Kurtev I (2008) ATL: a model transformation tool. Sci Comput Program 72(1\u20132):31\u201339. https:\/\/doi.org\/10.1016\/j.scico.2007.08.002","journal-title":"Sci Comput Program"},{"issue":"1","key":"10921_CR54","doi-asserted-by":"publisher","first-page":"11","DOI":"10.1145\/381790.381795","volume":"21","author":"BA Kitchenham","year":"1996","unstructured":"Kitchenham BA (1996) Evaluating software engineering methods and tool part 1: the evaluation context and evaluation methods. ACM SIGSOFT Softw Eng Notes 21(1):11\u201314","journal-title":"ACM SIGSOFT Softw Eng Notes"},{"issue":"12","key":"10921_CR55","doi-asserted-by":"publisher","first-page":"2049","DOI":"10.1016\/j.infsof.2013.07.010","volume":"55","author":"B Kitchenham","year":"2013","unstructured":"Kitchenham B, Brereton P (2013) A systematic review of systematic review process research in software engineering. Inf Softw Technol 55(12):2049\u20132075","journal-title":"Inf Softw Technol"},{"key":"10921_CR56","unstructured":"Kojima T, Gu SS, Reid M, Matsuo Y, Iwasawa Y (2022) Large language models are zero-shot reasoners. In: Koyejo S, Mohamed S, Agarwal A, Belgrave D, Cho K, Oh A (eds) Advances in neural information processing systems 35: annual conference on neural information processing systems 2022, NeurIPS 2022, New Orleans, LA, USA, November 28 - December 9, 2022, http:\/\/papers.nips.cc\/paper_files\/paper\/2022\/hash\/8bb0d291acd4acf06ef112099c16f326-Abstract-Conference.html"},{"key":"10921_CR57","doi-asserted-by":"publisher","first-page":"104013","DOI":"10.1016\/j.csi.2025.104013","volume":"94","author":"Y Kong","year":"2025","unstructured":"Kong Y, Zhang N, Duan Z, Yu B (2025) Collaboration with generative ai to improve requirements change. Comput Standards Interf 94:104013","journal-title":"Comput Standards Interf"},{"key":"10921_CR58","doi-asserted-by":"crossref","unstructured":"Lamas V, Garcia-Gonzalez D, Sala L, Luaces MR (2025) DSL-Xpert 2.0: Enhancing LLM-driven code generation for domain-specific languages. Information and Software Technology 107954","DOI":"10.1016\/j.infsof.2025.107954"},{"key":"10921_CR59","unstructured":"Li R, Ben Allal L, Zi Y, Muennighoff N (2023) StarCoder: may the source be with you! Transactions on Machine Learning Research"},{"key":"10921_CR60","doi-asserted-by":"publisher","unstructured":"Li J, Zhang M, Li N, Weyns D, Jin Z, Tei K (2024) Generative AI for self-adaptive systems: State of the art and research roadmap. ACM Trans Auton Adapt Syst 19(3). https:\/\/doi.org\/10.1145\/3686803","DOI":"10.1145\/3686803"},{"key":"10921_CR61","unstructured":"Liu Y, Zhou H, Guo Z, Shareghi E, Vuli\u0107 I, Korhonen A, Collier N (2024) Aligning with human judgement: The role of pairwise preference in large language model evaluators. arXiv:2403.16950"},{"issue":"7","key":"10921_CR62","doi-asserted-by":"publisher","first-page":"615","DOI":"10.1109\/TSE.2016.2620145","volume":"43","author":"N Macedo","year":"2016","unstructured":"Macedo N, Jorge T, Cunha A (2016) A feature-based classification of model repair approaches. IEEE Trans Softw Eng 43(7):615\u2013640","journal-title":"IEEE Trans Softw Eng"},{"issue":"3","key":"10921_CR63","doi-asserted-by":"publisher","first-page":"276","DOI":"10.11613\/BM.2012.031","volume":"22","author":"ML McHugh","year":"2012","unstructured":"McHugh ML (2012) Interrater reliability: the kappa statistic. Biochemia Medica 22(3):276\u2013282","journal-title":"Biochemia Medica"},{"key":"10921_CR64","doi-asserted-by":"publisher","unstructured":"Mesnard T, Hardin C, Dadashi R, Bhupatiraju S, Pathak S, Sifre L, Rivi\u00e8re M, Kale MS, Love J, Tafti P, Hussenot L, Chowdhery A, Roberts A, Barua A, Botev A, Castro-Ros A, Slone A, H\u00e9liou A, Tacchetti A, Bulanova A, Paterson A, Tsai B, Shahriari B, Lan CL, Choquette-Choo CA, Crepy C, Cer D, Ippolito D, Reid D, Buchatskaya E, Ni E et\u00a0al. (2024) Gemma: open models based on Gemini research and technology. https:\/\/doi.org\/10.48550\/arXiv.2403.08295","DOI":"10.48550\/arXiv.2403.08295"},{"key":"10921_CR65","doi-asserted-by":"crossref","unstructured":"Moezkarimi Z, Eriksson K, Johansson AA, Bucaioni A, Sirjani M (2025) Harnessing ChatGPT for model transformation in software architecture: from UML state diagrams to Rebeca models for formal verification. In: 2025 IEEE 22nd international conference on software architecture companion (ICSA-C), IEEE, pp 387\u2013396","DOI":"10.1109\/ICSA-C65153.2025.00061"},{"key":"10921_CR66","doi-asserted-by":"crossref","unstructured":"Mohagheghi P, Fernandez MA, Martell JA, Fritzsche M, Gilani W (2008) MDE adoption in industry: challenges and success criteria. In: International conference on model driven engineering languages and systems. Springer, pp 54\u201359","DOI":"10.1007\/978-3-642-01648-6_6"},{"key":"10921_CR67","doi-asserted-by":"publisher","first-page":"e16","DOI":"10.1017\/dsj.2024.8","volume":"10","author":"JJ Norheim","year":"2024","unstructured":"Norheim JJ, Rebentisch E, Xiao D, Draeger L, Kerbrat A, de Weck OL (2024) Challenges in applying large language models to requirements engineering tasks. Design Sci 10:e16. https:\/\/doi.org\/10.1017\/dsj.2024.8","journal-title":"Design Sci"},{"key":"10921_CR68","unstructured":"Object Management Group (2019) OMG Systems Modeling Language (SysML), Version 1.6. Tech. Rep. formal\/2019-11-01, Object Management Group. https:\/\/www.omg.org\/spec\/SysML\/1.6\/"},{"key":"10921_CR69","doi-asserted-by":"publisher","unstructured":"OECD, Eurostat (2018) Oslo manual 2018: guidelines for collecting, reporting and using data on innovation, 4th Edition. The Measurement of Scientific, Technological and Innovation Activities 2018. https:\/\/doi.org\/10.1787\/9789264304604-en","DOI":"10.1787\/9789264304604-en"},{"key":"10921_CR70","unstructured":"OpenAI (2022) Introducing ChatGPT. https:\/\/openai.com\/index\/chatgpt\/"},{"key":"10921_CR71","doi-asserted-by":"publisher","unstructured":"OpenAI (2026) OpenAI GPT-5 system card. https:\/\/doi.org\/10.48550\/arXiv.2601.03267","DOI":"10.48550\/arXiv.2601.03267"},{"key":"10921_CR72","doi-asserted-by":"crossref","unstructured":"Paige RF, Brooke PJ, Ostroff JS (2007) Metamodel-based model conformance and multiview consistency checking. ACM Trans Softw Eng Methodol (TOSEM) 16(3):11\u2013es","DOI":"10.1145\/1243987.1243989"},{"issue":"1","key":"10921_CR73","doi-asserted-by":"publisher","first-page":"130","DOI":"10.1109\/TSE.2018.2884706","volume":"47","author":"JI Panach","year":"2018","unstructured":"Panach JI, Dieste O, Mar\u00edn B, Espa\u00f1a S, Vegas S, Pastor O, Juristo N (2018) Evaluating model-driven development claims with respect to quality: a family of experiments. IEEE Trans Softw Eng 47(1):130\u2013145","journal-title":"IEEE Trans Softw Eng"},{"key":"10921_CR74","doi-asserted-by":"crossref","unstructured":"Petersen K, Feldt R, Mujtaba S, Mattsson M (2008) Systematic mapping studies in software engineering. In: 12th international conference on evaluation and assessment in software engineering (EASE), BCS Learning & Development","DOI":"10.14236\/ewic\/EASE2008.8"},{"key":"10921_CR75","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1016\/j.infsof.2015.03.007","volume":"64","author":"K Petersen","year":"2015","unstructured":"Petersen K, Vakkalanka S, Kuzniarz L (2015) Guidelines for conducting systematic mapping studies in software engineering: an update. Inf Softw Technol 64:1\u201318","journal-title":"Inf Softw Technol"},{"key":"10921_CR76","unstructured":"Radford A, Narasimhan K, Salimans T, Sutskever I (2018) Improving language understanding by generative pre-training. https:\/\/cdn.openai.com\/research-covers\/language-unsupervised\/language_understanding_paper.pdf"},{"key":"10921_CR77","unstructured":"Raffel C, Shazeer N, Roberts A, Lee K, Narang S, Matena M, Zhou Y, Li W, Liu PJ (2020) Exploring the limits of transfer learning with a unified text-to-text transformer. J Mach Learn Res 21:140:1\u2013140:67. https:\/\/jmlr.org\/papers\/v21\/20-074.html"},{"key":"10921_CR78","doi-asserted-by":"crossref","unstructured":"Ratiu D, Voelter M (2016) Automated testing of DSL implementations: experiences from building mbeddr. In: Proceedings of the 11th international workshop on automation of software test, pp 15\u201321","DOI":"10.1145\/2896921.2896922"},{"key":"10921_CR79","unstructured":"Ren S, Guo D, Lu S, Zhou L, Liu S, Tang D, Sundaresan N, Zhou M, Blanco A, Ma S (2020) CodeBLEU: a method for automatic evaluation of code synthesis. arXiv:2009.10297"},{"key":"10921_CR80","doi-asserted-by":"publisher","first-page":"102452","DOI":"10.1016\/j.datak.2025.102452","volume":"159","author":"S Rizzi","year":"2025","unstructured":"Rizzi S, Francia M, Gallinucci E, Golfarelli M (2025) Conceptual design of multidimensional cubes with LLMs: an investigation. Data Knowl Eng 159:102452","journal-title":"Data Knowl Eng"},{"key":"10921_CR81","unstructured":"Rose LM, Paige RF, Kolovos DS, Polack FA (2009) An analysis of approaches to model migration. Proc. Joint MoDSE-MCCM Workshop. pp 6\u201315"},{"key":"10921_CR82","doi-asserted-by":"publisher","unstructured":"Rozi\u00e8re B, Gehring J, Gloeckle F, Sootla S, Gat I, Tan XE, Adi Y, Liu J, Remez T, Rapin J, Kozhevnikov A, Evtimov I, Bitton J, Bhatt M, Canton-Ferrer C, Grattafiori A, Xiong W, D\u00e9fossez A, Copet J, Azhar F, Touvron H, Martin L, Usunier N, Scialom T, Synnaeve G (2023) Code Llama: open foundation models for code. https:\/\/doi.org\/10.48550\/arXiv.2308.12950","DOI":"10.48550\/arXiv.2308.12950"},{"key":"10921_CR83","doi-asserted-by":"crossref","unstructured":"Sallou J, Durieux T, Panichella A (2024) Breaking the silence: the threats of using LLMs in software engineering. Proceedings of the 2024 ACM\/IEEE 44th International Conference on Software Engineering: New Ideas and Emerging Results. pp 102\u2013106","DOI":"10.1145\/3639476.3639764"},{"key":"10921_CR84","unstructured":"Sami AM, Rasheed Z, Waseem M, Zhang Z, Tomas H, Abrahamsson P (2024) A tool for test case scenarios generation using large language models. arXiv:2406.07021"},{"key":"10921_CR85","unstructured":"Sanh V, Webson A, Raffel C, Bach SH, Sutawika L, Alyafeai Z, Chaffin A, Stiegler A, Raja A, Dey M, Bari MS, Xu C, Thakker U, Sharma SS, Szczechla E, Kim T, Chhablani G, Nayak NV, Datta D, Chang J, Jiang MT, Wang H, Manica M, Shen S, Yong ZX, Pandey H, Bawden R, Wang T, Neeraj T, Rozen J, Sharma A et\u00a0al (2022) Multitask prompted training enables zero-shot task generalization. In: The Tenth international conference on learning representations, ICLR 2022, Virtual Event, April 25-29, 2022, OpenReview.net, https:\/\/openreview.net\/forum?id=9Vrb9D0WI4"},{"issue":"12","key":"10921_CR86","doi-asserted-by":"publisher","first-page":"1340","DOI":"10.1016\/j.infsof.2012.07.008","volume":"54","author":"I Santiago","year":"2012","unstructured":"Santiago I, Jim\u00e9nez A, Vara JM, De Castro V, Bollati VA, Marcos E (2012) Model-driven engineering as a new landscape for traceability management: a systematic literature review. Inf Softw Technol 54(12):1340\u20131356","journal-title":"Inf Softw Technol"},{"key":"10921_CR87","doi-asserted-by":"crossref","unstructured":"Sasaki Y, Washizaki H, Li J, Sander D, Yoshioka N, Fukazawa Y (2024) Systematic literature review of prompt engineering patterns in software engineering. In: 2024 IEEE 48th annual computers, software, and applications conference (COMPSAC). IEEE, pp 670\u2013675","DOI":"10.1109\/COMPSAC61105.2024.00096"},{"key":"10921_CR88","unstructured":"Schmid L, Hey T, Armbruster M, Corallo S, Fuch\u00df D, Keim J, Liu H, Koziolek A (2025) Software architecture meets LLMs: a systematic literature review. arxiv:2505.16697"},{"issue":"2","key":"10921_CR89","doi-asserted-by":"publisher","first-page":"25","DOI":"10.1109\/MC.2006.58","volume":"39","author":"DC Schmidt","year":"2006","unstructured":"Schmidt DC et al (2006) Model-driven engineering. Comput-IEEE Comput Soc 39(2):25","journal-title":"Comput-IEEE Comput Soc"},{"key":"10921_CR90","unstructured":"Schulhoff S, Ilie M, Balepur N, Kahadze K, Liu A, Si C, Li Y, Gupta A, Han H, Schulhoff S et\u00a0al (2024) The prompt report: a systematic survey of prompt engineering techniques. arXiv:2406.06608"},{"issue":"3","key":"10921_CR91","doi-asserted-by":"publisher","first-page":"75:1","DOI":"10.1145\/3604281","volume":"56","author":"V Scotti","year":"2024","unstructured":"Scotti V, Sbattella L, Tedesco R (2024) A primer on Seq2Seq models for generative Chatbots. ACM Comput Surv 56(3):75:1-75:58. https:\/\/doi.org\/10.1145\/3604281","journal-title":"ACM Comput Surv"},{"key":"10921_CR92","unstructured":"Scotti V, Keim J, Hey T, Metzger A, Koziolek A, Mirandola R (2025) A roadmap for tamed interactions with large language models. arXiv:2510.24819"},{"key":"10921_CR93","doi-asserted-by":"crossref","unstructured":"Shaw M (2003) Writing good software engineering research papers. In: 25th International conference on software engineering, 2003. Proceedings., IEEE, pp 726\u2013736","DOI":"10.1109\/ICSE.2003.1201262"},{"key":"10921_CR94","unstructured":"Siddiq ML, Islam-Gomes A, Sekerak N, Santos JCS (2025) Large language models for software engineering: A reproducibility crisis. arxiv:2512.00651"},{"key":"10921_CR95","doi-asserted-by":"crossref","unstructured":"Silva J, Ma Q, Cabot J, Kelsen P, Proper HA (2025) Towards human-in-the-loop LLM-enabled domain modeling. In: International conference on conceptual modeling. Springer, pp 127\u2013145","DOI":"10.1007\/978-3-032-08623-5_7"},{"key":"10921_CR96","unstructured":"Steinberg D, Budinsky F, Paternostro M, Merks E (2008) EMF: eclipse modeling framework. Addison-Wesley Professional,(2nd edn)"},{"key":"10921_CR97","unstructured":"Touvron H, Martin L, Stone K, Albert P, Almahairi A, Babaei Y, Bashlykov N, Batra S, Bhargava P, Bhosale S et\u00a0al (2023) Llama 2: open foundation and fine-tuned chat models. arXiv:2307.09288"},{"key":"10921_CR98","doi-asserted-by":"publisher","first-page":"130991","DOI":"10.1016\/j.eswa.2025.130991","volume":"308","author":"J Umre","year":"2026","unstructured":"Umre J, Parihar AS, Gupta A (2026) CSLLM: code-specific large language models\u2013a survey. Expert Syst Appl 308:130991. https:\/\/doi.org\/10.1016\/j.eswa.2025.130991","journal-title":"Expert Syst Appl"},{"key":"10921_CR99","doi-asserted-by":"publisher","first-page":"101151","DOI":"10.1016\/j.csl.2020.101151","volume":"67","author":"C Van der Lee","year":"2021","unstructured":"Van der Lee C, Gatt A, Van Miltenburg E, Krahmer E (2021) Human evaluation of automatically generated text: current trends and best practice guidelines. Comput Speech Lang 67:101151","journal-title":"Comput Speech Lang"},{"key":"10921_CR100","doi-asserted-by":"crossref","unstructured":"Van Der Straeten R, Mens T, Van Baelen S (2008) Challenges in model-driven software engineering. In: International conference on model driven engineering languages and systems. Springer, pp 35\u201347","DOI":"10.1007\/978-3-642-01648-6_4"},{"key":"10921_CR101","unstructured":"Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser L, Polosukhin I (2017) Attention is all you need. In: Guyon I, von Luxburg U, Bengio S, Wallach HM, Fergus R, Vishwanathan SVN, Garnett R (eds) Advances in neural information processing systems 30: annual conference on neural information processing systems 2017, December 4-9, 2017, Long Beach, CA, USA, pp 5998\u20136008. https:\/\/proceedings.neurips.cc\/paper\/2017\/hash\/3f5ee243547dee91fbd053c1c4a845aa-Abstract.html"},{"key":"10921_CR102","doi-asserted-by":"publisher","unstructured":"Viyovi\u0107 V, Maksimovi\u0107 M, Perisi\u0107 B (2014) Sirius: A rapid development of DSM graphical editor. In: IEEE 18th international conference on intelligent engineering systems (INES), pp 233\u2013238. https:\/\/doi.org\/10.1109\/INES.2014.6909375","DOI":"10.1109\/INES.2014.6909375"},{"key":"10921_CR103","doi-asserted-by":"crossref","unstructured":"Wagner S, Bar\u00f3n MM, Falessi D, Baltes S (2025) Towards evaluation guidelines for empirical studies involving LLMs. IEEE\/ACM International Workshop on Methodological Issues with Empirical Studies in Software Engineering (WSESE) 24\u201327IEEE","DOI":"10.1109\/WSESE66602.2025.00011"},{"key":"10921_CR104","doi-asserted-by":"publisher","unstructured":"Wang Y, Wang W, Joty S, Hoi SCH (2021) CodeT5: identifier-aware unified pre-trained encoder-decoder models for code understanding and generation. In: Proceedings of the 2021 conference on empirical methods in natural language processing (EMNLP), pp 8696\u20138708. https:\/\/doi.org\/10.18653\/v1\/2021.emnlp-main.685","DOI":"10.18653\/v1\/2021.emnlp-main.685"},{"key":"10921_CR105","unstructured":"Washizaki H (2024) Guide to the software engineering body of knowledge v4.0a. IEEE Computer Society"},{"key":"10921_CR106","doi-asserted-by":"crossref","unstructured":"Weber T, Brandmaier M, Schmidt A, Mayer S (2024) Significant productivity gains through programming with large language models. Proceedings of the ACM on Human-Computer Interaction 8:1\u201329EICS","DOI":"10.1145\/3661145"},{"key":"10921_CR107","doi-asserted-by":"crossref","unstructured":"Wohlin C (2014) Guidelines for snowballing in systematic literature studies and a replication in software engineering. In: Proceedings of the 18th international conference on evaluation and assessment in software engineering, pp 1\u201310","DOI":"10.1145\/2601248.2601268"},{"key":"10921_CR108","unstructured":"Yang A, Li A, Yang B, Zhang B, Hui B, Zheng B, Yu B, Gao C, Huang C, Lv C, Zheng C, Liu D, Zhou F, Huang F, Hu F, Ge H, Wei H, Lin H, Tang J, Yang J, Tu J, Zhang J, Yang J, Yang J, Zhou J, Zhou J, Lin J, Dang K, Bao K, Yang K, Yu Let\u00a0al (2025) Qwen3 technical report. arxiv:2505.09388"},{"key":"10921_CR109","unstructured":"Zhang W, Holtmann J (2023) Unresolved challenges and potential features in EATXT. arXiv:2312.10250 Technical report"},{"key":"10921_CR110","doi-asserted-by":"crossref","unstructured":"Zhang W, Hebig R, Stegh\u00f6fer JP, Holtmann J (2023a) Creating Python-style domain specific languages: a semi-automated approach and intermediate results. In: MODELSWARD, pp 210\u2013217","DOI":"10.5220\/0011744900003402"},{"key":"10921_CR111","doi-asserted-by":"crossref","unstructured":"Zhang W, Hebig R, Str\u00fcber D, Stegh\u00f6fer JP (2023b) Automated extraction of grammar optimization rule configurations for metamodel-grammar co-evolution. In: Proceedings of the 16th ACM SIGPLAN international conference on software language engineering, pp 84\u201396","DOI":"10.1145\/3623476.3623525"},{"key":"10921_CR112","doi-asserted-by":"publisher","DOI":"10.1016\/j.jss.2024.112069","volume":"214","author":"W Zhang","year":"2024","unstructured":"Zhang W, Holtmann J, Str\u00fcber D, Hebig R, Stegh\u00f6fer JP (2024) Supporting meta-model-based language evolution and rapid prototyping with automated grammar transformation. J Syst Softw 214:112069","journal-title":"J Syst Softw"},{"key":"10921_CR113","doi-asserted-by":"crossref","unstructured":"Zhang W, Hebig R, Str\u00fcber D (2025) Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs. In: Proceedings of the first large language models for software engineering workshop (LLM4SE). Koblenz, Germany, (co-located with STAF 2025)","DOI":"10.1007\/s10270-026-01402-9"},{"key":"10921_CR114","unstructured":"Zhang W, Jiang B, Fu Y, Cheng H, Hummel M, Scotti V, Hagel NJ, Li J, Grossmann G, Stumptner M, Hebig R, Str\u00fcber D, Koziolek A (2026) Replication package. https:\/\/osf.io\/g5by9\/overview?view_only=5c10c1e56be3480d8d25e017b4276f7a, accessed: 10 Apr 2026"},{"key":"10921_CR115","doi-asserted-by":"publisher","unstructured":"Zhang Q, Fang C, Xie Y, Zhang Y, Yu S, Sun W, Yang Y, Chen Z (2026) A survey on large language models for software engineering. SCIENCE CHINA Inf Sci 69(4):141102. https:\/\/doi.org\/10.1007\/s11432-025-4670-0","DOI":"10.1007\/s11432-025-4670-0"},{"key":"10921_CR116","doi-asserted-by":"crossref","unstructured":"Zhang W, Jiang B, Fu Y, Koziolek A, Hebig R, Str\u00fcber D (2026b) Leveraging LLMs to support co-evolution between definitions and instances of textual DSLs: a systematic evaluation. arXiv:2602.11904","DOI":"10.1007\/s10270-026-01402-9"},{"key":"10921_CR117","unstructured":"Zhao WX, Zhou K, Li J, Tang T, Wang X, Hou Y, Min Y, Zhang B, Zhang J, Dong Z et\u00a0al (2023) A survey of large language models. arXiv:2303.18223 1(2):1\u2013124"}],"container-title":["Empirical Software Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10664-026-10921-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10664-026-10921-4","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10664-026-10921-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T06:54:15Z","timestamp":1784184855000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10664-026-10921-4"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,7,16]]},"references-count":117,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2027,2]]}},"alternative-id":["10921"],"URL":"https:\/\/doi.org\/10.1007\/s10664-026-10921-4","relation":{},"ISSN":["1382-3256","1573-7616"],"issn-type":[{"value":"1382-3256","type":"print"},{"value":"1573-7616","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,7,16]]},"assertion":[{"value":"21 April 2026","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"29 June 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 July 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"Not applicable.","order":1,"name":"Ethics","label":"Ethical Approval","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","label":"Informed Consent","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":3,"name":"Ethics","label":"Clinical Trial Number","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors have no competing interests to declare that are relevant to the content of this article.","order":4,"name":"Ethics","label":"Conflict of Interest","group":{"name":"EthicsHeading","label":"Declarations"}}],"article-number":"3"}}