{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,2,13]],"date-time":"2026-02-13T15:21:30Z","timestamp":1770996090397,"version":"3.50.1"},"reference-count":67,"publisher":"Association for Computing Machinery (ACM)","issue":"3","funder":[{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62202306 and 62372299"],"award-info":[{"award-number":["62202306 and 62372299"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Hong Kong RGC Projects","award":["PolyU15224121 and PolyU15231223"],"award-info":[{"award-number":["PolyU15224121 and PolyU15231223"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Softw. Eng. Methodol."],"published-print":{"date-parts":[[2026,3,31]]},"abstract":"<jats:p>\n                    XML configurations are integral to the Android development framework, particularly in the realm of UI display. However, these configurations can introduce compatibility issues (bugs), resulting in divergent visual outcomes and system crashes across various Android API versions (levels). In this study, we systematically investigate LLM-based approaches for detecting and repairing configuration compatibility bugs. Our findings highlight certain limitations of LLMs in effectively identifying and resolving these bugs, while also revealing their potential in addressing complex, hard-to-repair issues that traditional tools struggle with. Leveraging these insights, we introduce the LLM-CompDroid framework, which combines the strengths of LLMs and traditional tools for bug resolution. Our experimental results demonstrate a significant enhancement in bug resolution performance by LLM-CompDroid, with LLM-CompDroid-GPT-3.5 and LLM-CompDroid-GPT-4 surpassing the state-of-the-art tool, ConfFix, by at least 9.8% and 10.4% in both\n                    <jats:italic toggle=\"yes\">Correct<\/jats:italic>\n                    and\n                    <jats:italic toggle=\"yes\">Correct@k<\/jats:italic>\n                    metrics, respectively. In addition, our real-world evaluation shows that LLM-CompDroid successfully repairs 21 configuration compatibility bugs with a 100% success rate, demonstrating its practical utility. This innovative approach holds promise for advancing the reliability and robustness of Android applications, making a valuable contribution to the field of software development.\n                  <\/jats:p>","DOI":"10.1145\/3736406","type":"journal-article","created":{"date-parts":[[2025,5,20]],"date-time":"2025-05-20T13:03:21Z","timestamp":1747746201000},"page":"1-32","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":5,"title":["LLM-CompDroid: Repairing Configuration Compatibility Bugs in Android Apps with Pre-trained Large Language Models"],"prefix":"10.1145","volume":"35","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-0972-8631","authenticated-orcid":false,"given":"Zhijie","family":"Liu","sequence":"first","affiliation":[{"name":"School of Information Science and Technology, ShanghaiTech University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-5677-4564","authenticated-orcid":false,"given":"Yutian","family":"Tang","sequence":"additional","affiliation":[{"name":"School of Computing Science, University of Glasgow, Glasgow, United Kingdom of Great Britain and Northern Ireland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-7400-2681","authenticated-orcid":false,"given":"Meiyun","family":"Li","sequence":"additional","affiliation":[{"name":"Chang\u2019an University, Xi\u2019an, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0006-1435-1604","authenticated-orcid":false,"given":"Xin","family":"Jin","sequence":"additional","affiliation":[{"name":"Chang\u2019an University, Xi\u2019an, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4407-578X","authenticated-orcid":false,"given":"Yunfei","family":"Long","sequence":"additional","affiliation":[{"name":"University of Essex, Colchester, United Kingdom of Great Britain and Northern Ireland"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3543-1524","authenticated-orcid":false,"given":"Liang Feng","family":"Zhang","sequence":"additional","affiliation":[{"name":"School of Information Science and Technology, ShanghaiTech University, Shanghai, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-9082-3208","authenticated-orcid":false,"given":"Xiapu","family":"Luo","sequence":"additional","affiliation":[{"name":"The Hong Kong Polytechnic University, Kowloon, Hong Kong"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2026,2,13]]},"reference":[{"key":"e_1_3_3_2_2","unstructured":"Music-Player-GO. 2019. Retrieved from https:\/\/github.com\/enricocid\/Music-Player-GO\/tree\/aef85dc"},{"key":"e_1_3_3_3_2","unstructured":"Tachiyomi. 2020. Retrieved from https:\/\/github.com\/CarlosEsco\/Neko\/tree\/885c7bbb103dde6ca6b1b47cfefc6b9ea5ea231c"},{"key":"e_1_3_3_4_2","unstructured":"All Android Releases. 2023. Retrieved from https:\/\/developer.android.com\/about\/versions"},{"key":"e_1_3_3_5_2","unstructured":"Developers. 2023. Retrieved from https:\/\/developer.android.com\/"},{"key":"e_1_3_3_6_2","unstructured":"GitHub. 2023. Retrieved from https:\/\/github.com\/"},{"key":"e_1_3_3_7_2","unstructured":"GitHub Issues. 2023. Retrieved from https:\/\/docs.github.com\/en\/issues\/tracking-your-work-with-issues\/about-issues"},{"key":"e_1_3_3_8_2","unstructured":"Google Android Lint. 2023. Retrieved August 30 2023 from https:\/\/developer.android.com\/studio\/write\/lint"},{"key":"e_1_3_3_9_2","unstructured":"Google Bard. 2023. Retrieved August 30 2023 from https:\/\/bard.google.com\/"},{"key":"e_1_3_3_10_2","unstructured":"OpenAI. 2023. Retrieved August 30 2023 from https:\/\/chat.openai.com\/"},{"key":"e_1_3_3_11_2","unstructured":"OpenAI Documentation. 2023. Retrieved from https:\/\/platform.openai.com\/docs\/models\/overview"},{"key":"e_1_3_3_12_2","unstructured":"Stack Overflow\u2014Where Developers Learn Share & Build Careers. 2023. Retrieved from https:\/\/stackoverflow.com\/"},{"key":"e_1_3_3_13_2","unstructured":"Online Artifact. 2024. Retrieved from https:\/\/zenodo.org\/records\/10618818"},{"key":"e_1_3_3_14_2","unstructured":"Sobriety Pull Request #113. 2024. Retrieved from https:\/\/github.com\/KiARC\/Sobriety\/pull\/113"},{"key":"e_1_3_3_15_2","unstructured":"F-Droid. 2025. Free and Open Source Android App Repository. Retrieved from https:\/\/f-droid.org\/en\/"},{"key":"e_1_3_3_16_2","unstructured":"Thorsten Brants Ashok C. Popat Peng Xu Franz J. Och and Jeffrey Dean. 2007. Large language models in machine translation. In Proceedings of the 2007 Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning (EMNLP-CoNLL) 858\u2013867."},{"key":"e_1_3_3_17_2","first-page":"2633","volume-title":"30th USENIX Security Symposium (USENIX Security \u2019","volume":"21","author":"Carlini Nicholas","year":"2021","unstructured":"Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al. 2021. Extracting training data from large language models. In 30th USENIX Security Symposium (USENIX Security \u201921), 2633\u20132650."},{"key":"e_1_3_3_18_2","volume-title":"USENIX Security Symposium","volume":"6","author":"Carlini Nicholas","unstructured":"Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom B. Brown, Dawn Song, Ulfar Erlingsson, et al. 2021. Extracting training data from large language models. In USENIX Security Symposium, Vol. 6."},{"key":"e_1_3_3_19_2","unstructured":"Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde de Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman et al. 2021. Evaluating large language models trained on code. arXiv:2107.03374. Retrieved from https:\/\/arxiv.org\/abs\/2107.03374"},{"key":"e_1_3_3_20_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597926.3598067"},{"key":"e_1_3_3_21_2","doi-asserted-by":"publisher","DOI":"10.1109\/ASE56229.2023.00217"},{"key":"e_1_3_3_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00128"},{"key":"e_1_3_3_23_2","doi-asserted-by":"publisher","DOI":"10.1109\/ASE.2017.8115644"},{"key":"e_1_3_3_24_2","doi-asserted-by":"publisher","DOI":"10.1145\/3293882.3330571"},{"key":"e_1_3_3_25_2","doi-asserted-by":"publisher","DOI":"10.1145\/3387904.3389285"},{"key":"e_1_3_3_26_2","doi-asserted-by":"publisher","DOI":"10.1007\/s10664-021-10096-0"},{"key":"e_1_3_3_27_2","doi-asserted-by":"publisher","DOI":"10.1145\/3238147.3238185"},{"key":"e_1_3_3_28_2","doi-asserted-by":"publisher","DOI":"10.1145\/3695988"},{"key":"e_1_3_3_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/3238147.3238181"},{"key":"e_1_3_3_30_2","doi-asserted-by":"publisher","DOI":"10.1109\/ASE51524.2021.9678556"},{"key":"e_1_3_3_31_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597926.3598074"},{"key":"e_1_3_3_32_2","unstructured":"Xin Jin Jonathan Larson Weiwei Yang and Zhiqiang Lin. 2023. Binary code summarization: Benchmarking ChatGPT\/gpt-4 and other large language models. arXiv:2312.09601. Retrieved from https:\/\/arxiv.org\/abs\/2312.09601"},{"key":"e_1_3_3_33_2","doi-asserted-by":"publisher","DOI":"10.1145\/3524842.3527997"},{"key":"e_1_3_3_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2019.00040"},{"key":"e_1_3_3_35_2","first-page":"22199","article-title":"Large language models are zero-shot reasoners","volume":"35","author":"Kojima Takeshi","year":"2022","unstructured":"Takeshi Kojima, Shixiang, Shane Gu, and Machel Reid, Yutaka Matsuo, and Yusuke Iwasawa. 2022. Large language models are zero-shot reasoners. In Advances in Neural Information Processing Systems, Vol. 35, 22199\u201322213.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_3_36_2","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2020.2988396"},{"key":"e_1_3_3_37_2","doi-asserted-by":"publisher","DOI":"10.1145\/3551349.3556898"},{"key":"e_1_3_3_38_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00085"},{"key":"e_1_3_3_39_2","doi-asserted-by":"publisher","DOI":"10.1109\/APSEC.2018.00042"},{"key":"e_1_3_3_40_2","doi-asserted-by":"publisher","DOI":"10.1145\/3213846.3213857"},{"key":"e_1_3_3_41_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597503.3639091"},{"key":"e_1_3_3_42_2","doi-asserted-by":"publisher","DOI":"10.1145\/3533767.3534407"},{"key":"e_1_3_3_43_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.findings-acl.229"},{"issue":"5","key":"e_1_3_3_44_2","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3643674","article-title":"Refining ChatGPT-generated code: Characterizing and mitigating code quality issues","volume":"33","author":"Liu Yue","year":"2023","unstructured":"Yue Liu, Thanh Le-Cong, Ratnadira Widyasari, Chakkrit Tantithamthavorn, Li Li, Xuan-Bach D. Le, and David Lo. 2023. Refining ChatGPT-generated code: Characterizing and mitigating code quality issues. ACM Transactions on Software Engineering and Methodology 33, 5 (2023), 1\u201326.","journal-title":"ACM Transactions on Software Engineering and Methodology"},{"key":"e_1_3_3_45_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00119"},{"key":"e_1_3_3_46_2","unstructured":"Zhe Liu Chunyang Chen Junjie Wang Mengzhuo Chen Boyu Wu Xing Che Dandan Wang and Qing Wang. 2023. Chatting with GPT-3 for zero-shot human-like mobile automated GUI testing. arXiv:2305.09434. Retrieved from https:\/\/arxiv.org\/abs\/2305.09434"},{"key":"e_1_3_3_47_2","unstructured":"Zhijie Liu Yutian Tang Xiapu Luo Yuming Zhou and Liang Feng Zhang. 2023. No need to lift a finger anymore? Assessing the quality of code generation by ChatGPT. arXiv:2308.04838. Retrieved from https:\/\/arxiv.org\/abs\/2308.04838"},{"key":"e_1_3_3_48_2","unstructured":"Yao Lu Max Bartolo Alastair Moore Sebastian Riedel and Pontus Stenetorp. 2021. Fantastically ordered prompts and where to find them: Overcoming few-shot prompt order sensitivity. arXiv:2104.08786. Retrieved from https:\/\/arxiv.org\/abs\/2104.08786"},{"key":"e_1_3_3_49_2","doi-asserted-by":"publisher","DOI":"10.1145\/3092703.3092726"},{"key":"e_1_3_3_50_2","unstructured":"Ryo Nagata Manabu Kimura and Kazuaki Hanawa. 2021. Exploring the capacity of a large-scale masked language model to recognize grammatical errors. arXiv:2108.12216. Retrieved from https:\/\/arxiv.org\/abs\/2108.12216"},{"key":"e_1_3_3_51_2","unstructured":"OpenAI. 2020. OpenAI API. Retrieved from https:\/\/openai.com\/index\/openai-api\/"},{"key":"e_1_3_3_52_2","unstructured":"OpenAI. 2023. GPT-4 Technical Report. Retrieved from https:\/\/cdn.openai.com\/papers\/gpt-4.pdf"},{"key":"e_1_3_3_53_2","doi-asserted-by":"publisher","DOI":"10.1145\/2932631"},{"key":"e_1_3_3_54_2","doi-asserted-by":"publisher","DOI":"10.1109\/SP46214.2022.9833571"},{"key":"e_1_3_3_55_2","doi-asserted-by":"publisher","DOI":"10.1109\/SP46215.2023.10179324"},{"issue":"1","key":"e_1_3_3_56_2","first-page":"1","article-title":"Exploring the limits of transfer learning with a unified text-to-text transformer","volume":"21","author":"Raffel Colin","year":"2022","unstructured":"Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2022. Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of Machine Learning Research 21, 1 (2022), 1\u201367.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_3_57_2","doi-asserted-by":"publisher","DOI":"10.1109\/MSR.2019.00055"},{"key":"e_1_3_3_58_2","unstructured":"Yuqiang Sun Daoyuan Wu Yue Xue Han Liu Haijun Wang Zhengzi Xu Xiaofei Xie and Yang Liu. 2023. When GPT meets program analysis: Towards intelligent detection of smart contract logic vulnerabilities in GPTScan. arXiv:2308.03314. Retrieved from https:\/\/arxiv.org\/abs\/2308.03314"},{"key":"e_1_3_3_59_2","unstructured":"Yutian Tang Zhijie Liu Zhichao Zhou and Xiapu Luo. 2023. ChatGPT vs SBST: A comparative assessment of unit test suite generation. arXiv:2307.00588. Retrieved from https:\/\/arxiv.org\/abs\/2307.00588"},{"key":"e_1_3_3_60_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-37057-1_15"},{"key":"e_1_3_3_61_2","first-page":"24824","article-title":"Chain-of-thought prompting elicits reasoning in large language models","volume":"35","author":"Wei Jason","year":"2022","unstructured":"Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc V. Le, and Denny Zhou. 2022. Chain-of-thought prompting elicits reasoning in large language models. In Advances in Neural Information Processing Systems, Vol. 35, 24824\u201324837.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_3_62_2","doi-asserted-by":"publisher","DOI":"10.1145\/2970276.2970312"},{"key":"e_1_3_3_63_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE.2019.00094"},{"key":"e_1_3_3_64_2","doi-asserted-by":"publisher","DOI":"10.1109\/TSE.2018.2876439"},{"key":"e_1_3_3_65_2","unstructured":"Chunqiu Steven Xia Matteo Paltenghi Jia Le Tian Michael Pradel and Lingming Zhang. 2023. Universal fuzzing via large language models. arXiv:2308.04748. Retrieved from https:\/\/arxiv.org\/abs\/2308.04748"},{"key":"e_1_3_3_66_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICSE48619.2023.00129"},{"key":"e_1_3_3_67_2","doi-asserted-by":"publisher","DOI":"10.1145\/3691620.3695529"},{"key":"e_1_3_3_68_2","doi-asserted-by":"publisher","DOI":"10.1145\/3510003.3510128"}],"container-title":["ACM Transactions on Software Engineering and Methodology"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3736406","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,2,13]],"date-time":"2026-02-13T14:36:21Z","timestamp":1770993381000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3736406"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,2,13]]},"references-count":67,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2026,3,31]]}},"alternative-id":["10.1145\/3736406"],"URL":"https:\/\/doi.org\/10.1145\/3736406","relation":{},"ISSN":["1049-331X","1557-7392"],"issn-type":[{"value":"1049-331X","type":"print"},{"value":"1557-7392","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,2,13]]},"assertion":[{"value":"2024-02-05","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-04-19","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-02-13","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}