{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,29]],"date-time":"2026-07-29T16:00:26Z","timestamp":1785340826922,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":51,"publisher":"ACM","license":[{"start":{"date-parts":[[2026,4,12]],"date-time":"2026-04-12T00:00:00Z","timestamp":1775952000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/legalcode"}],"funder":[{"name":"National\u00a0Natural\u00a0Science\u00a0Foundation\u00a0of\u00a0China","award":["62402185"],"award-info":[{"award-number":["62402185"]}]},{"name":"Fundamental\u00a0Research\u00a0Funds\u00a0for\u00a0the\u00a0Central\u00a0Universities","award":["2025ZYGXZR064"],"award-info":[{"award-number":["2025ZYGXZR064"]}]},{"name":"Science and Technology Planning Project of Guangdong Province","award":["2025B0101120003"],"award-info":[{"award-number":["2025B0101120003"]}]},{"name":"Guangdong Provincial Fund for Basic and Applied Basic Research\u2014Regional Joint Fund Project (Key Project)","award":["2023B1515120078"],"award-info":[{"award-number":["2023B1515120078"]}]},{"name":"National Natural Science Foundation of China","award":["62476097"],"award-info":[{"award-number":["62476097"]}]},{"name":"Guangdong Provincial Natural Science Foundation for Outstanding Youth Team Project","award":["2024B1515040010"],"award-info":[{"award-number":["2024B1515040010"]}]},{"name":"Fundamental Research Funds for the Central Universities, South China University of Technology","award":["x2rjD2250190"],"award-info":[{"award-number":["x2rjD2250190"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,4,12]]},"DOI":"10.1145\/3794763.3794823","type":"proceedings-article","created":{"date-parts":[[2026,7,29]],"date-time":"2026-07-29T15:18:58Z","timestamp":1785338338000},"page":"355-366","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["RepoMind: Enhancing Repository-Level Code Generation via LLM Reasoning over Structured Repository Documentation"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-1020-1230","authenticated-orcid":false,"given":"Songwen","family":"Gong","sequence":"first","affiliation":[{"name":"South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4086-2743","authenticated-orcid":false,"given":"Mengzhen","family":"Wang","sequence":"additional","affiliation":[{"name":"South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7064-6507","authenticated-orcid":false,"given":"Jiexin","family":"Wang","sequence":"additional","affiliation":[{"name":"South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1767-789X","authenticated-orcid":false,"given":"Yi","family":"Cai","sequence":"additional","affiliation":[{"name":"South China University of Technology, Guangzhou, China and Key Laboratory of Big Data and Intelligent Robot (SCUT), Ministry of Education, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,7,29]]},"reference":[{"key":"e_1_3_3_2_2_2","unstructured":"Ingeol Baek Hwan Chang Byeongjeong Kim Jimin Lee and Hwanhee Lee. 2024. Probing-rag: Self-probing to guide language models in selective document retrieval. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.13339 (2024)."},{"key":"e_1_3_3_2_3_2","doi-asserted-by":"crossref","unstructured":"Ramakrishna Bairi Atharv Sonwane Aditya Kanade Vageesh\u00a0D C Arun Iyer Suresh Parthasarathy Sriram Rajamani Balasubramanyan Ashok and Shashank Shet. 2024. Codeplan: Repository-level coding using llms and planning. Proceedings of the ACM on Software Engineering 1 FSE (2024) 675\u2013698.","DOI":"10.1145\/3643757"},{"key":"e_1_3_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1145\/1553374.1553380"},{"key":"e_1_3_3_2_5_2","unstructured":"Zhangqian Bi Yao Wan Zheng Wang Hongyu Zhang Batu Guan Fangxin Lu Zili Zhang Yulei Sui Hai Jin and Xuanhua Shi. 2024. Iterative refinement of project-level code context for precise code generation with compiler feedback. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2403.16792 (2024)."},{"key":"e_1_3_3_2_6_2","unstructured":"Egor Bogomolov Aleksandra Eliseeva Timur Galimzyanov Evgeniy Glukhov Anton Shapkin Maria Tigina Yaroslav Golubev Alexander Kovrigin Arie Van\u00a0Deursen Maliheh Izadi et\u00a0al. 2024. Long code arena: a set of benchmarks for long-context code models. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2406.11612 (2024)."},{"key":"e_1_3_3_2_7_2","unstructured":"Tuan-Dung Bui Duc-Thieu Luu-Van Thanh-Phat Nguyen Thu-Trang Nguyen Son Nguyen and Hieu\u00a0Dinh Vo. 2024. Rambo: Enhancing rag-based repository-level method body completion. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2409.15204 (2024)."},{"key":"e_1_3_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.63317\/3agxr3ouzo5m"},{"key":"e_1_3_3_2_9_2","first-page":"3043","volume-title":"Proceedings of the 31st International Conference on Computational Linguistics","author":"Cao Liuwen","year":"2025","unstructured":"Liuwen Cao, Hongkui He, Hailin Huang, Jiexin Wang, and Yi Cai. 2025. Rethinking-based Code Summarization with Chain of Comments. In Proceedings of the 31st International Conference on Computational Linguistics. 3043\u20133056."},{"key":"e_1_3_3_2_10_2","unstructured":"Mark Chen Jerry Tworek Heewoo Jun Qiming Yuan Henrique Ponde De\u00a0Oliveira Pinto Jared Kaplan Harri Edwards Yuri Burda Nicholas Joseph Greg Brockman et\u00a0al. 2021. Evaluating large language models trained on code. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2107.03374 (2021)."},{"key":"e_1_3_3_2_11_2","unstructured":"Zhaoling Chen Xiangru Tang Gangda Deng Fang Wu Jialong Wu Zhiwei Jiang Viktor Prasanna Arman Cohan and Xingyao Wang. 2025. Locagent: Graph-guided llm agents for code localization. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2503.09089 (2025)."},{"key":"e_1_3_3_2_12_2","unstructured":"Wei Cheng Yuhan Wu and Wei Hu. 2024. Dataflow-guided retrieval augmentation for repository-level code completion. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2405.19782 (2024)."},{"key":"e_1_3_3_2_13_2","unstructured":"Aryaz Eghbali and Michael Pradel. 2024. De-Hallucinator: Mitigating LLM Hallucinations in Code Generation Tasks via Iterative Grounding. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2401.01701 (2024)."},{"key":"e_1_3_3_2_14_2","unstructured":"Yunfan Gao Yun Xiong Xinyu Gao Kangxiang Jia Jinliu Pan Yuxi Bi Yixin Dai Jiawei Sun Haofen Wang and Haofen Wang. 2023. Retrieval-augmented generation for large language models: A survey. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.10997 2 1 (2023)."},{"key":"e_1_3_3_2_15_2","unstructured":"Wenchao Gu Juntao Chen Yanlin Wang Tianyue Jiang Xingzhe Li Mingwei Liu Xilin Liu Yuchi Ma and Zibin Zheng. 2025. What to Retrieve for Effective Retrieval-Augmented Code Generation? An Empirical Study and Beyond. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2503.20589 (2025)."},{"key":"e_1_3_3_2_16_2","unstructured":"Daya Guo Shuai Lu Nan Duan Yanlin Wang Ming Zhou and Jian Yin. 2022. Unixcoder: Unified cross-modal pre-training for code representation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2203.03850 (2022)."},{"key":"e_1_3_3_2_17_2","unstructured":"Daya Guo Shuo Ren Shuai Lu Zhangyin Feng Duyu Tang Shujie Liu Long Zhou Nan Duan Alexey Svyatkovskiy Shengyu Fu et\u00a0al. 2020. Graphcodebert: Pre-training code representations with data flow. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2009.08366 (2020)."},{"key":"e_1_3_3_2_18_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICPC66645.2025.00035"},{"key":"e_1_3_3_2_19_2","unstructured":"Dan Hendrycks Steven Basart Saurav Kadavath Mantas Mazeika Akul Arora Ethan Guo Collin Burns Samir Puranik Horace He Dawn Song et\u00a0al. 2021. Measuring coding challenge competence with apps. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2105.09938 (2021)."},{"key":"e_1_3_3_2_20_2","unstructured":"Aaron Hurst Adam Lerer Adam\u00a0P Goucher Adam Perelman Aditya Ramesh Aidan Clark AJ Ostrow Akila Welihinda Alan Hayes Alec Radford et\u00a0al. 2024. Gpt-4o system card. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.21276 (2024)."},{"key":"e_1_3_3_2_21_2","unstructured":"Carlos\u00a0E Jimenez John Yang Alexander Wettig Shunyu Yao Kexin Pei Ofir Press and Karthik Narasimhan. 2023. Swe-bench: Can language models resolve real-world github issues? arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2310.06770 (2023)."},{"key":"e_1_3_3_2_22_2","doi-asserted-by":"crossref","unstructured":"Jia Li Ge Li Yunfei Zhao Yongmin Li Huanyu Liu Hao Zhu Lecheng Wang Kaibo Liu Zheng Fang Lanshen Wang et\u00a0al. 2024. Deveval: A manually-annotated code generation benchmark aligned with real-world code repositories. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2405.19856 (2024).","DOI":"10.18653\/v1\/2024.findings-acl.214"},{"key":"e_1_3_3_2_23_2","unstructured":"Raymond Li Loubna\u00a0Ben Allal Yangtian Zi Niklas Muennighoff Denis Kocetkov Chenghao Mou Marc Marone Christopher Akiki Jia Li Jenny Chim et\u00a0al. 2023. Starcoder: may the source be with you! arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2305.06161 (2023)."},{"key":"e_1_3_3_2_24_2","unstructured":"Ming Liang Xiaoheng Xie Gehao Zhang Xunjin Zheng Peng Di Hongwei Chen Chengpeng Wang Gang Fan et\u00a0al. 2024. Repofuse: Repository-level code completion with fused dual context. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2402.14323 (2024)."},{"key":"e_1_3_3_2_25_2","doi-asserted-by":"crossref","unstructured":"Dianshu Liao Shidong Pan Xiaoyu Sun Xiaoxue Ren Qing Huang Zhenchang Xing Huan Jin and Qinying Li. 2024. A 3-codgen: A repository-level code generation framework for code reuse with local-aware global-aware and third-party-library-aware. IEEE Transactions on Software Engineering (2024).","DOI":"10.1109\/TSE.2024.3486195"},{"key":"e_1_3_3_2_26_2","unstructured":"Aixin Liu Bei Feng Bing Xue Bingxuan Wang Bochao Wu Chengda Lu Chenggang Zhao Chengqi Deng Chenyu Zhang Chong Ruan et\u00a0al. 2024. Deepseek-v3 technical report. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2412.19437 (2024)."},{"key":"e_1_3_3_2_27_2","unstructured":"Nelson\u00a0F Liu Kevin Lin John Hewitt Ashwin Paranjape Michele Bevilacqua Fabio Petroni and Percy Liang. 2023. Lost in the middle: How language models use long contexts. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2307.03172 (2023)."},{"key":"e_1_3_3_2_28_2","doi-asserted-by":"crossref","unstructured":"Wei Liu Ailun Yu Daoguang Zan Bo Shen Wei Zhang Haiyan Zhao Zhi Jin and Qianxiang Wang. 2024. Graphcoder: Enhancing repository-level code completion via code context graph-based retrieval and language model. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2406.07003 (2024).","DOI":"10.1145\/3691620.3695054"},{"key":"e_1_3_3_2_29_2","unstructured":"Yang Liu Li Zhang Fang Liu Zhuohang Wang Donglin Wei Zhishuo Yang Kechi Zhang Jia Li and Lin Shi. 2025. Enhancing Repository-Level Code Generation with Call Chain-Aware Multi-View Context. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2507.14791 (2025)."},{"key":"e_1_3_3_2_30_2","doi-asserted-by":"crossref","unstructured":"Qinyu Luo Yining Ye Shihao Liang Zhong Zhang Yujia Qin Yaxi Lu Yesai Wu Xin Cong Yankai Lin Yingli Zhang et\u00a0al. 2024. Repoagent: An llm-powered open-source framework for repository-level code documentation generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2402.16667 (2024).","DOI":"10.18653\/v1\/2024.emnlp-demo.46"},{"key":"e_1_3_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.1145\/3696630.3728549"},{"key":"e_1_3_3_2_32_2","unstructured":"Multi-Linguality Multi-Functionality Multi-Granularity. 2024. M3-Embedding: Multi-Linguality Multi-Functionality Multi-Granularity Text Embeddings Through Self-Knowledge Distillation."},{"key":"e_1_3_3_2_33_2","unstructured":"Siru Ouyang Wenhao Yu Kaixin Ma Zilin Xiao Zhihan Zhang Mengzhao Jia Jiawei Han Hongming Zhang and Dong Yu. 2024. Repograph: Enhancing ai software engineering with repository-level code graph. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2410.14684 (2024)."},{"key":"e_1_3_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/Forge66646.2025.00009"},{"key":"e_1_3_3_2_35_2","unstructured":"Baptiste Roziere Jonas Gehring Fabian Gloeckle Sten Sootla Itai Gat Xiaoqing\u00a0Ellen Tan Yossi Adi Jingyu Liu Romain Sauvestre Tal Remez et\u00a0al. 2023. Code llama: Open foundation models for code. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2308.12950 (2023)."},{"key":"e_1_3_3_2_36_2","unstructured":"Anton Shapkin Denis Litvinov Yaroslav Zharov Egor Bogomolov Timur Galimzyanov and Timofey Bryksin. 2023. Dynamic Retrieval-Augmented Generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2312.08976 (2023)."},{"key":"e_1_3_3_2_37_2","unstructured":"Jicheng Wang Yifeng He and Hao Chen. 2024. Repogenreflex: Enhancing repository-level code completion with verbal reinforcement and retrieval-augmented generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2409.13122 (2024)."},{"key":"e_1_3_3_2_38_2","doi-asserted-by":"publisher","unstructured":"Mengzhen Wang Yi Cai Jiayuan Xie Jiexin Wang Xin Wu Jie Liu and Ronghui Yang. 2026. Code-Enhanced Cross-Perspective Bug Question Retrieval. ACM Trans. Softw. Eng. Methodol. (Jan. 2026). 10.1145\/3789667Just Accepted.","DOI":"10.1145\/3789667"},{"key":"e_1_3_3_2_39_2","doi-asserted-by":"publisher","DOI":"10.1145\/3746027.3755425"},{"key":"e_1_3_3_2_40_2","doi-asserted-by":"publisher","DOI":"10.1109\/SANER60148.2024.00068"},{"key":"e_1_3_3_2_41_2","doi-asserted-by":"crossref","unstructured":"Yue Wang Weishi Wang Shafiq Joty and Steven\u00a0CH Hoi. 2021. Codet5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2109.00859 (2021).","DOI":"10.18653\/v1\/2021.emnlp-main.685"},{"key":"e_1_3_3_2_42_2","unstructured":"Yanlin Wang Yanli Wang Daya Guo Jiachi Chen Ruikai Zhang Yuchi Ma and Zibin Zheng. 2024. Rlcoder: Reinforcement learning for repository-level code completion. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2407.19487 (2024)."},{"key":"e_1_3_3_2_43_2","unstructured":"Di Wu Wasi\u00a0Uddin Ahmad Dejiao Zhang Murali\u00a0Krishna Ramanathan and Xiaofei Ma. 2024. Repoformer: Selective retrieval for repository-level code completion. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2403.10059 (2024)."},{"key":"e_1_3_3_2_44_2","unstructured":"An Yang Anfeng Li Baosong Yang Beichen Zhang Binyuan Hui Bo Zheng Bowen Yu Chang Gao Chengen Huang Chenxu Lv et\u00a0al. 2025. Qwen3 technical report. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2505.09388 (2025)."},{"key":"e_1_3_3_2_45_2","unstructured":"Dayu Yang Antoine Simoulin Xin Qian Xiaoyi Liu Yuwei Cao Zhaopu Teng and Grey Yang. 2025. DocAgent: A Multi-Agent System for Automated Code Documentation Generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2504.08725 (2025)."},{"key":"e_1_3_3_2_46_2","doi-asserted-by":"crossref","unstructured":"Zezhou Yang Sirong Chen Cuiyun Gao Zhenhao Li Xing Hu Kui Liu and Xin Xia. 2025. An empirical study of retrieval-augmented code generation: Challenges and opportunities. ACM Transactions on Software Engineering and Methodology (2025).","DOI":"10.1145\/3717061"},{"key":"e_1_3_3_2_47_2","doi-asserted-by":"publisher","DOI":"10.1145\/3597503.3623316"},{"key":"e_1_3_3_2_48_2","doi-asserted-by":"crossref","unstructured":"Fengji Zhang Bei Chen Yue Zhang Jacky Keung Jin Liu Daoguang Zan Yi Mao Jian-Guang Lou and Weizhu Chen. 2023. Repocoder: Repository-level code completion through iterative retrieval and generation. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2303.12570 (2023).","DOI":"10.18653\/v1\/2023.emnlp-main.151"},{"key":"e_1_3_3_2_49_2","unstructured":"Kechi Zhang Jia Li Ge Li Xianjie Shi and Zhi Jin. 2024. Codeagent: Enhancing code generation with tool-integrated agent systems for real-world repo-level coding challenges. arXiv preprint arXiv:https:\/\/arXiv.org\/abs\/2401.07339 (2024)."},{"key":"e_1_3_3_2_50_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v39i24.34782"},{"key":"e_1_3_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.1145\/3650212.3680384"},{"key":"e_1_3_3_2_52_2","volume-title":"The Eleventh International Conference on Learning Representations","author":"Zhou Shuyan","year":"2022","unstructured":"Shuyan Zhou, Uri Alon, Frank\u00a0F Xu, Zhengbao Jiang, and Graham Neubig. 2022. Docprompting: Generating code by retrieving the docs. In The Eleventh International Conference on Learning Representations."}],"event":{"name":"ICPC '26: 34th IEEE\/ACM International Conference on Program Comprehension","location":"Rio de Janeiro , Brazil","acronym":"ICPC '26","sponsor":["SIGSOFT ACM Special Interest Group on Software Engineering"]},"container-title":["Proceedings of the 2026 34th IEEE\/ACM International Conference on Program Comprehension"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3794763.3794823","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,29]],"date-time":"2026-07-29T15:19:12Z","timestamp":1785338352000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3794763.3794823"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,12]]},"references-count":51,"alternative-id":["10.1145\/3794763.3794823","10.1145\/3794763"],"URL":"https:\/\/doi.org\/10.1145\/3794763.3794823","relation":{},"subject":[],"published":{"date-parts":[[2026,4,12]]},"assertion":[{"value":"2026-07-29","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}