{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,4]],"date-time":"2026-08-04T15:39:51Z","timestamp":1785857991297,"version":"3.56.0"},"reference-count":58,"publisher":"Springer Science and Business Media LLC","issue":"2","license":[{"start":{"date-parts":[[2024,12,18]],"date-time":"2024-12-18T00:00:00Z","timestamp":1734480000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2024,12,18]],"date-time":"2024-12-18T00:00:00Z","timestamp":1734480000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Institute of Information & communications Technology Planning & Evaluation","award":["IITP2022-0-00995"],"award-info":[{"award-number":["IITP2022-0-00995"]}]},{"name":"Institute of Information & communications Technology Planning & Evaluation","award":["IITP2022-0-00995"],"award-info":[{"award-number":["IITP2022-0-00995"]}]},{"DOI":"10.13039\/501100003725","name":"National Research Foundation of Korea","doi-asserted-by":"crossref","award":["RS-2023-00208998"],"award-info":[{"award-number":["RS-2023-00208998"]}],"id":[{"id":"10.13039\/501100003725","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100003725","name":"National Research Foundation of Korea","doi-asserted-by":"crossref","award":["RS-2023-00208998"],"award-info":[{"award-number":["RS-2023-00208998"]}],"id":[{"id":"10.13039\/501100003725","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Institute of Information & communications Technology Planning & Evaluation","award":["RS-2022-00155958"],"award-info":[{"award-number":["RS-2022-00155958"]}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Empir Software Eng"],"published-print":{"date-parts":[[2025,3]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Automated debugging techniques have the potential to reduce developer effort in debugging. However, while developers want rationales for the provided automatic debugging results, existing techniques are ill-suited to provide them, as their deduction process differs significantly froof human developers. Inspired by the way developers interact with code when debugging, we propose Automated Scientific Debugging (<jats:sc>AutoSD<\/jats:sc>), a technique that prompts large language models to automatically generate hypotheses, uses debuggers to interact with buggy code, and thus automatically reach conclusions prior to patch generation. In doing so, we aim to produce explanations of how a specific patch has been generated, with the hope that these explanations will lead to enhanced developer decision-making. Our empirical analysis on three program repair benchmarks shows that <jats:sc>AutoSD<\/jats:sc>performs competitively with other program repair baselines, and that it can indicate when it is confident in its results. Furthermore, we perform a human study with 20 participants to evaluate <jats:sc>AutoSD<\/jats:sc>-generated explanations. Participants with access to explanations judged patch correctness more accurately in five out of six real-world bugs studied. Furthermore, 70% of participants answered that they wanted explanations when using repair tools, and 55% answered that they were satisfied with the Scientific Debugging presentation.<\/jats:p>","DOI":"10.1007\/s10664-024-10594-x","type":"journal-article","created":{"date-parts":[[2024,12,18]],"date-time":"2024-12-18T07:51:24Z","timestamp":1734508284000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":52,"title":["Explainable automated debugging via large language model-driven scientific debugging"],"prefix":"10.1007","volume":"30","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-0298-5320","authenticated-orcid":false,"given":"Sungmin","family":"Kang","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bei","family":"Chen","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Shin","family":"Yoo","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jian-Guang","family":"Lou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2024,12,18]]},"reference":[{"key":"10594_CR1","doi-asserted-by":"crossref","unstructured":"Alaboudi A, LaToza TD (2020) Using hypotheses as a debugging aid. In: 2020 IEEE Symposium on Visual Languages and Human-Centric Computing (VL\/HCC), IEEE, pp 1\u20139","DOI":"10.1109\/VL\/HCC50065.2020.9127278"},{"key":"10594_CR2","doi-asserted-by":"publisher","unstructured":"Allahyari M, Pouriyeh S, Assefi M, Safaei S, Trippe ED, Gutierrez JB, Kochut K (2017) Text summarization techniques: A brief survey. International Journal of Advanced Computer Science and Applications 8(10). https:\/\/doi.org\/10.14569\/IJACSA.2017.081052","DOI":"10.14569\/IJACSA.2017.081052"},{"key":"10594_CR3","doi-asserted-by":"publisher","unstructured":"Arcuri A, Briand L (2011) A practical guide for using statistical tests to assess randomized algorithms in software engineering. In: Proceedings of the 33rd International Conference on Software Engineering, Association for Computing Machinery, New York, NY, USA, ICSE \u201911, pp 1\u201310. https:\/\/doi.org\/10.1145\/1985793.1985795","DOI":"10.1145\/1985793.1985795"},{"key":"10594_CR4","doi-asserted-by":"publisher","first-page":"82","DOI":"10.1016\/j.inffus.2019.12.012","volume":"58","author":"AB Arrieta","year":"2020","unstructured":"Arrieta AB, D\u00edaz-Rodr\u00edguez N, Del Ser J, Bennetot A, Tabik S, Barbado A, Garc\u00eda S, Gil-L\u00f3pez S, Molina D, Benjamins R et al (2020) Explainable artificial intelligence (xai): Concepts, taxonomies, opportunities and challenges toward responsible ai. Information Fusion 58:82\u2013115","journal-title":"Information Fusion"},{"key":"10594_CR5","doi-asserted-by":"publisher","unstructured":"B\u00f6hme M, Soremekun EO, Chattopadhyay S, Ugherughe EJ, Zeller A (2017) How developers debug software - the dbgbench dataset. In: 2017 IEEE\/ACM 39th International Conference on Software Engineering Companion (ICSE-C), pp 244\u2013246, https:\/\/doi.org\/10.1109\/ICSE-C.2017.94","DOI":"10.1109\/ICSE-C.2017.94"},{"key":"10594_CR6","first-page":"1877","volume":"33","author":"T Brown","year":"2020","unstructured":"Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A et al (2020) Language models are few-shot learners. Adv Neural Inf Process Syst 33:1877\u20131901","journal-title":"Adv Neural Inf Process Syst"},{"key":"10594_CR7","unstructured":"Chen M, Tworek J, Jun H, Yuan Q, Pinto HPdO, Kaplan J, Edwards H, Burda Y, Joseph N, Brockman G et\u00a0al (2021) Evaluating large language models trained on code. arXiv:2107.03374"},{"key":"10594_CR8","doi-asserted-by":"crossref","unstructured":"Coles H, Laurent T, Henard C, Papadakis M, Ventresque A (2016) Pit: A practical mutation testing tool for java (demo). In: Proceedings of the 25th International Symposium on Software Testing and Analysis, Association for Computing Machinery, New York, NY, USA, ISSTA 2016, pp 449\u2013452","DOI":"10.1145\/2931037.2948707"},{"key":"10594_CR9","doi-asserted-by":"crossref","unstructured":"Dam HK, Tran T, Ghose A (2018) Explainable software analytics. In: Proceedings of the 40th International Conference on Software Engineering: New Ideas and Emerging Results, Association for Computing Machinery, New York, NY, USA, ICSE-NIER \u201918, pp 53\u201356","DOI":"10.1145\/3183399.3183424"},{"key":"10594_CR10","unstructured":"Fried D, Aghajanyan A, Lin J, Wang S, Wallace E, Shi F, Zhong R, Yih Wt, Zettlemoyer L, Lewis M (2022) Incoder: A generative model for code infilling and synthesis. arXiv preprint arXiv:2204.05999"},{"key":"10594_CR11","unstructured":"Gao L, Madaan A, Zhou S, Alon U, Liu P, Yang Y, Callan J, Neubig G (2022) Pal: Program-aided language models. arXiv:2211.10435"},{"issue":"1","key":"10594_CR12","doi-asserted-by":"publisher","first-page":"34","DOI":"10.1109\/TSE.2017.2755013","volume":"45","author":"L Gazzola","year":"2019","unstructured":"Gazzola L, Micucci D, Mariani L (2019) Automatic software repair: A survey. IEEE Trans Software Eng 45(1):34\u201367. https:\/\/doi.org\/10.1109\/TSE.2017.2755013","journal-title":"IEEE Trans Software Eng"},{"issue":"12","key":"10594_CR13","doi-asserted-by":"publisher","first-page":"56","DOI":"10.1145\/3318162","volume":"62","author":"CL Goues","year":"2019","unstructured":"Goues CL, Pradel M, Roychoudhury A (2019) Automated program repair. Commun ACM 62(12):56\u201365","journal-title":"Commun ACM"},{"key":"10594_CR14","doi-asserted-by":"publisher","unstructured":"Gould JD (1975) Some psychological evidence on how people debug computer programs. Int J Man Mach Stud 7(2):151\u2013182. https:\/\/doi.org\/10.1016\/S0020-7373(75)80005-8, URL https:\/\/www.sciencedirect.com\/science\/article\/pii\/S0020737375800058","DOI":"10.1016\/S0020-7373(75)80005-8"},{"key":"10594_CR15","doi-asserted-by":"crossref","unstructured":"Haas R, Elsner D, Juergens E, Pretschner A, Apel S (2021) How can manual testing processes be optimized? developer survey, optimization guidelines, and case studies. In: Proceedings of the 29th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2021, pp 1281\u20131291","DOI":"10.1145\/3468264.3473922"},{"key":"10594_CR16","doi-asserted-by":"crossref","unstructured":"Jiang J, Xiong Y, Zhang H, Gao Q, Chen X (2018) Shaping program repair space with existing patches and similar code. Proceedings of the 27th ACM SIGSOFT International Symposium on Software Testing and Analysis","DOI":"10.1145\/3213846.3213871"},{"key":"10594_CR17","unstructured":"Jiang N, Liu K, Lutellier T, Tan L (2023a) Impact of code language models on automated program repair. 2302.05020"},{"key":"10594_CR18","doi-asserted-by":"crossref","unstructured":"Jiang N, Lutellier T, Lou Y, Tan L, Goldwasser D, Zhang X (2023b) Knod: Domain knowledge distilled tree decoder for automated program repair. 2302.01857","DOI":"10.1109\/ICSE48619.2023.00111"},{"key":"10594_CR19","doi-asserted-by":"publisher","unstructured":"Jiang S, Armaly A, McMillan C (2017) Automatically generating commit messages from diffs using neural machine translation. In: 2017 32nd IEEE\/ACM International Conference on Automated Software Engineering (ASE), pp 135\u2013146, https:\/\/doi.org\/10.1109\/ASE.2017.8115626","DOI":"10.1109\/ASE.2017.8115626"},{"key":"10594_CR20","doi-asserted-by":"crossref","unstructured":"Jones JA, Harrold MJ, Stasko J (2002) Visualization of test information to assist fault localization. In: Proceedings of the 24th International Conference on Software Engineering, Association for Computing Machinery, New York, NY, USA, ICSE \u201902, pp 467\u2013477","DOI":"10.1145\/581396.581397"},{"key":"10594_CR21","doi-asserted-by":"publisher","unstructured":"Just R, Jalali D, Ernst MD (2014) Defects4j: A database of existing faults to enable controlled testing studies for java programs. In: Proceedings of the 2014 International Symposium on Software Testing and Analysis, Association for Computing Machinery, New York, NY, USA, ISSTA 2014, pp 437\u2013440.https:\/\/doi.org\/10.1145\/2610384.2628055","DOI":"10.1145\/2610384.2628055"},{"issue":"4","key":"10594_CR22","doi-asserted-by":"publisher","first-page":"43","DOI":"10.1109\/MS.2021.3071086","volume":"38","author":"S Kirbas","year":"2021","unstructured":"Kirbas S, Windels E, McBello O, Kells K, Pagano M, Szalanski R, Nowack V, Winter ER, Counsell S, Bowes D, Hall T, Haraldsson S, Woodward J (2021) On the introduction of automatic program repair in bloomberg. IEEE Softw 38(4):43\u201351. https:\/\/doi.org\/10.1109\/MS.2021.3071086","journal-title":"IEEE Softw"},{"issue":"1","key":"10594_CR23","doi-asserted-by":"publisher","first-page":"110","DOI":"10.1007\/s10664-013-9279-3","volume":"20","author":"AJ Ko","year":"2015","unstructured":"Ko AJ, LaToza TD, Burnett MM (2015) A practical guide to controlled experiments of software engineering tools with human participants. Empirical Softw Engg 20(1):110\u2013141. https:\/\/doi.org\/10.1007\/s10664-013-9279-3","journal-title":"Empirical Softw Engg"},{"key":"10594_CR24","doi-asserted-by":"publisher","unstructured":"Kochhar PS, Xia X, Lo D, Li S (2016) Practitioners\u2019 expectations on automated fault localization. In: Proceedings of the 25th International Symposium on Software Testing and Analysis, Association for Computing Machinery, New York, NY, USA, ISSTA 2016, pp 165\u2013176, https:\/\/doi.org\/10.1145\/2931037.2931051","DOI":"10.1145\/2931037.2931051"},{"key":"10594_CR25","doi-asserted-by":"publisher","unstructured":"Layman L, Diep M, Nagappan M, Singer J, Deline R, Venolia G (2013) Debugging revisited: Toward understanding the debugging needs of contemporary software developers. In: 2013 ACM \/ IEEE International Symposium on Empirical Software Engineering and Measurement, pp 383\u2013392, https:\/\/doi.org\/10.1109\/ESEM.2013.43","DOI":"10.1109\/ESEM.2013.43"},{"key":"10594_CR26","doi-asserted-by":"publisher","unstructured":"Lee J, Kang S, Yoon J, Yoo S (2024) The github recent bugs dataset for evaluating llm-based debugging applications. In: 2024 IEEE Conference on Software Testing, Verification and Validation (ICST), IEEE Computer Society, Los Alamitos, CA, USA, pp 442\u201344https:\/\/doi.org\/10.1109\/ICST60714.2024.00049, URL https:\/\/doi.ieeecomputersociety.org\/10.1109\/ICST60714.2024.00049","DOI":"10.1109\/ICST60714.2024.00049"},{"key":"10594_CR27","doi-asserted-by":"crossref","unstructured":"Li X, Li W, Zhang Y, Zhang L (2019) Deepfl: Integrating multiple fault diagnosis dimensions for deep fault localization. In: Proceedings of the 28th ACM SIGSOFT International Symposium on Software Testing and Analysis, Association for Computing Machinery, New York, NY, USA, ISSTA 2019, pp 169\u2013180","DOI":"10.1145\/3293882.3330574"},{"key":"10594_CR28","doi-asserted-by":"crossref","unstructured":"Lim BY, Dey AK, Avrahami D (2009) Why and why not explanations improve the intelligibility of context-aware intelligent systems. In: Proceedings of the SIGCHI Conference on Human Factors in Computing Systems, Association for Computing Machinery, New York, NY, USA, CHI \u201909, pp 2119\u20132128","DOI":"10.1145\/1518701.1519023"},{"key":"10594_CR29","doi-asserted-by":"publisher","unstructured":"Liu K, Wang S, Koyuncu A, Kim K, Bissyand\u00e9 TF, Kim D, Wu P, Klein J, Mao X, Traon YL (2020) On the efficiency of test suite based program repair: A systematic assessment of 16 automated repair systems for java programs. In: Proceedings of the ACM\/IEEE 42nd International Conference on Software Engineering, Association for Computing Machinery, New York, NY, USA, ICSE \u201920, pp 615\u2013627, https:\/\/doi.org\/10.1145\/3377811.3380338","DOI":"10.1145\/3377811.3380338"},{"key":"10594_CR30","doi-asserted-by":"publisher","unstructured":"Marginean A, Bader J, Chandra S, Harman M, Jia Y, Mao K, Mols A, Scott A (2019) Sapfix: Automated end-to-end repair at scale. In: 2019 IEEE\/ACM 41st International Conference on Software Engineering: Software Engineering in Practice (ICSE-SEIP), pp 269\u2013278, https:\/\/doi.org\/10.1109\/ICSE-SEIP.2019.00039","DOI":"10.1109\/ICSE-SEIP.2019.00039"},{"issue":"4","key":"10594_CR31","doi-asserted-by":"publisher","first-page":"486","DOI":"10.1109\/TSE.2010.93","volume":"37","author":"L Mariani","year":"2011","unstructured":"Mariani L, Pastore F, Pezze M (2011) Dynamic analysis for diagnosing integration faults. IEEE Trans Software Eng 37(4):486\u2013508. https:\/\/doi.org\/10.1109\/TSE.2010.93","journal-title":"IEEE Trans Software Eng"},{"key":"10594_CR32","doi-asserted-by":"publisher","first-page":"65","DOI":"10.1016\/j.jss.2019.01.069","volume":"151","author":"M Martinez","year":"2019","unstructured":"Martinez M, Monperrus M (2019) Astor: Exploring the design space of generate-and-validate program repair beyond genprog. J Syst Softw 151:65\u201380","journal-title":"J Syst Softw"},{"key":"10594_CR33","doi-asserted-by":"publisher","unstructured":"Mechtaev S, Yi J, Roychoudhury A (2016) Angelix: Scalable multiline program patch synthesis via symbolic analysis. In: 2016 IEEE\/ACM 38th International Conference on Software Engineering (ICSE), pp 691\u2013701,https:\/\/doi.org\/10.1145\/2884781.2884807","DOI":"10.1145\/2884781.2884807"},{"key":"10594_CR34","doi-asserted-by":"publisher","unstructured":"Monperrus M (2019) Explainable software bot contributions: Case study of automated bug fixes. In: 2019 IEEE\/ACM 1st International Workshop on Bots in Software Engineering (BotSE), IEEE Computer Society, Los Alamitos, CA, USA, pp 12\u20131https:\/\/doi.org\/10.1109\/BotSE.2019.00010, URL https:\/\/doi.ieeecomputersociety.org\/10.1109\/BotSE.2019.00010","DOI":"10.1109\/BotSE.2019.00010"},{"key":"10594_CR35","unstructured":"Monperrus M (2020) The Living Review on Automated Program Repair, URL https:\/\/hal.archives-ouvertes.fr\/hal-01956501, working paper or preprint"},{"key":"10594_CR36","doi-asserted-by":"publisher","unstructured":"Moon S, Kim Y, Kim M, Yoo S (2014) Ask the mutants: Mutating faulty programs for fault localization. In: 2014 IEEE Seventh International Conference on Software Testing, Verification and Validation, pp 153\u2013162,https:\/\/doi.org\/10.1109\/ICST.2014.28","DOI":"10.1109\/ICST.2014.28"},{"key":"10594_CR37","unstructured":"Nijkamp E, Pang B, Hayashi H, Tu L, Wang H, Zhou Y, Savarese S, Xiong C (2022) Codegen: An open large language model for code with multi-turn program synthesis. arXiv preprint"},{"key":"10594_CR38","doi-asserted-by":"publisher","unstructured":"Noller Y, Shariffdeen R, Gao X, Roychoudhury A (2022) Trust enhancement issues in program repair. In: Proceedings of the 44th International Conference on Software Engineering, Association for Computing Machinery, New York, NY, USA, ICSE \u201922, pp 2228\u20132240,https:\/\/doi.org\/10.1145\/3510003.3510040","DOI":"10.1145\/3510003.3510040"},{"key":"10594_CR39","unstructured":"OpenAI (2023) Gpt-4 technical report. 2303.08774"},{"key":"10594_CR40","unstructured":"Ouyang L, Wu J, Jiang X, Almeida D, Wainwright CL, Mishkin P, Zhang C, Agarwal S, Slama K, Ray A, et\u00a0al. (2022) Training language models to follow instructions with human feedback. arXiv preprint arXiv:2203.02155"},{"key":"10594_CR41","doi-asserted-by":"publisher","unstructured":"Rigby PC, Bird C (2013) Convergent contemporary software peer review practices. In: Proceedings of the 2013 9th Joint Meeting on Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2013, pp 202\u2013212, https:\/\/doi.org\/10.1145\/2491411.2491444","DOI":"10.1145\/2491411.2491444"},{"key":"10594_CR42","doi-asserted-by":"publisher","unstructured":"Sadowski C, S\u00f6derberg E, Church L, Sipko M, Bacchelli A (2018) Modern code review: A case study at google. In: Proceedings of the 40th International Conference on Software Engineering: Software Engineering in Practice, Association for Computing Machinery, New York, NY, USA, ICSE-SEIP \u201918, pp 181\u2013190, https:\/\/doi.org\/10.1145\/3183519.3183525","DOI":"10.1145\/3183519.3183525"},{"key":"10594_CR43","unstructured":"Shinn N, Cassano F, Labash B, Gopinath A, Narasimhan K, Yao S (2023) Reflexion: Language agents with verbal reinforcement learning. 2303.11366"},{"key":"10594_CR44","doi-asserted-by":"publisher","unstructured":"Siegmund B, Perscheid M, Taeumel M, Hirschfeld R (2014) Studying the advancement in debugging practice of professional software developers. In: 2014 IEEE International Symposium on Software Reliability Engineering Workshops, pp 269\u2013274, https:\/\/doi.org\/10.1109\/ISSREW.2014.36","DOI":"10.1109\/ISSREW.2014.36"},{"key":"10594_CR45","doi-asserted-by":"crossref","unstructured":"Sun C, Khoo SC (2013) Mining succinct predicated bug signatures. In: Proceedings of the 2013 9th Joint Meeting on Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2013, pp 576\u2013586","DOI":"10.1145\/2491411.2491449"},{"key":"10594_CR46","unstructured":"Wang X, Wei J, Schuurmans D, Le QV, Chi EH, Narang S, Chowdhery A, Zhou D (2023) Self-consistency improves chain of thought reasoning in language models. In: The Eleventh International Conference on Learning Representations, URL https:\/\/openreview.net\/forum?id=1PL1NIMMrw"},{"key":"10594_CR47","unstructured":"Wei J, Wang X, Schuurmans D, Bosma M, hsin Chi EH, Le Q, Zhou D (2022) Chain of thought prompting elicits reasoning in large language models. ArXiv abs\/2201.11903"},{"key":"10594_CR48","doi-asserted-by":"crossref","unstructured":"Widyasari R, Sim SQ, Lok C, Qi H, Phan J, Tay Q, Tan C, Wee F, Tan JE, Yieh Y, Goh B, Thung F, Kang HJ, Hoang T, Lo D, Ouh EL (2020) Bugsinpy: A database of existing bugs in python programs to enable controlled testing and debugging studies. In: Proceedings of the 28th ACM Joint Meeting on European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2020, pp 1556\u20131560","DOI":"10.1145\/3368089.3417943"},{"key":"10594_CR49","doi-asserted-by":"publisher","unstructured":"Winter ER, Nowack V, Bowes D, Counsell S, Hall T, Haraldsson S, Woodward J, Kirbas S, Windels E, McBello O, Atakishiyev A, Kells K, Pagano M (2022) Towards developer-centered automatic program repair: Findings from bloomberg. In: Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2022, pp 1578\u20131588, https:\/\/doi.org\/10.1145\/3540250.3558953","DOI":"10.1145\/3540250.3558953"},{"key":"10594_CR50","doi-asserted-by":"crossref","unstructured":"Xia CS, Zhang L (2022) Less training, more repairing please: Revisiting automated program repair via zero-shot learning. In: Proceedings of the 30th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering, Association for Computing Machinery, New York, NY, USA, ESEC\/FSE 2022, pp 959\u2013971","DOI":"10.1145\/3540250.3549101"},{"key":"10594_CR51","doi-asserted-by":"crossref","unstructured":"Xia CS, Zhang L (2023) Keep the conversation going: Fixing 162 out of 337 bugs for $0.42 each using chatgpt. arXiv preprint arXiv:2304.00385","DOI":"10.1145\/3650212.3680323"},{"key":"10594_CR52","unstructured":"Xia CS, Wei Y, Zhang L (2022) Practical program repair in the era of large pre-trained language models. arXiv preprint arXiv:2210.14179"},{"key":"10594_CR53","doi-asserted-by":"crossref","unstructured":"Xiong Y, Wang J, Yan R, Zhang J, Han S, Huang G, Zhang L (2017) Precise condition synthesis for program repair. In: 2017 IEEE\/ACM 39th International Conference on Software Engineering (ICSE), pp 416\u2013426","DOI":"10.1109\/ICSE.2017.45"},{"key":"10594_CR54","doi-asserted-by":"publisher","unstructured":"Xiong Y, Liu X, Zeng M, Zhang L, Huang G (2018) Identifying patch correctness in test-based program repair. In: Proceedings of the 40th International Conference on Software Engineering, Association for Computing Machinery, New York, NY, USA, ICSE \u201918, pp 789\u2013799, https:\/\/doi.org\/10.1145\/3180155.3180182","DOI":"10.1145\/3180155.3180182"},{"key":"10594_CR55","unstructured":"Yao S, Zhao J, Yu D, Du N, Shafran I, Narasimhan K, Cao Y (2022) React: Synergizing reasoning and acting in language models. arXiv preprint arXiv:2210.03629"},{"key":"10594_CR56","unstructured":"Yao S, Yu D, Zhao J, Shafran I, Griffiths TL, Cao Y, Narasimhan K (2023) Tree of thoughts: Deliberate problem solving with large language models. 2305.10601"},{"key":"10594_CR57","volume-title":"Why programs fail: a guide to systematic debugging","author":"A Zeller","year":"2009","unstructured":"Zeller A (2009) Why programs fail: a guide to systematic debugging. Elsevier"},{"key":"10594_CR58","first-page":"341","volume-title":"A Syntax-Guided Edit Decoder for Neural Program Repair","author":"Q Zhu","year":"2021","unstructured":"Zhu Q, Sun Z, Ya Xiao, Zhang W, Yuan K, Xiong Y, Zhang L (2021) A Syntax-Guided Edit Decoder for Neural Program Repair. Association for Computing Machinery, New York, NY, USA, pp 341\u2013353"}],"container-title":["Empirical Software Engineering"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10664-024-10594-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s10664-024-10594-x\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s10664-024-10594-x.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,5,2]],"date-time":"2025-05-02T13:45:12Z","timestamp":1746193512000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s10664-024-10594-x"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,12,18]]},"references-count":58,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2025,3]]}},"alternative-id":["10594"],"URL":"https:\/\/doi.org\/10.1007\/s10664-024-10594-x","relation":{},"ISSN":["1382-3256","1573-7616"],"issn-type":[{"value":"1382-3256","type":"print"},{"value":"1573-7616","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,12,18]]},"assertion":[{"value":"12 November 2024","order":1,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"18 December 2024","order":2,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare that Shin Yoo is a member of the EMSE Editorial board. All co-authors have seen and agree with the contents of the manuscript and there is no financial interest to report.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflicts of Interest"}}],"article-number":"45"}}