{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T05:17:05Z","timestamp":1784179025920,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":51,"publisher":"ACM","funder":[{"name":"National Natural Science Foundation of China","award":["62372139"],"award-info":[{"award-number":["62372139"]}]},{"name":"Natural Science Foundation of Guangdong Province","award":["2024A1515030024"],"award-info":[{"award-number":["2024A1515030024"]}]},{"name":"Research Projects of Shenzhen","award":["ZDCY20250901100302003"],"award-info":[{"award-number":["ZDCY20250901100302003"]}]},{"name":"CCF YOCSEF Shenzhen &#x5c;&quot;Large Model Inference&#x5c;&quot; series forum","award":["CCF-Yo-25-111"],"award-info":[{"award-number":["CCF-Yo-25-111"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2026,4,13]]},"DOI":"10.1145\/3774904.3792103","type":"proceedings-article","created":{"date-parts":[[2026,4,9]],"date-time":"2026-04-09T21:54:34Z","timestamp":1775771674000},"page":"1923-1934","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["PruneRAG: Confidence-Guided Query Decomposition Trees for Efficient Retrieval-Augmented Generation"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0009-5711-5910","authenticated-orcid":false,"given":"Shuguang","family":"Jiao","sequence":"first","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, Shenzhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8877-8129","authenticated-orcid":false,"given":"Xinyu","family":"Xiao","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, Shenzhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0002-3388-0481","authenticated-orcid":false,"given":"Yunfan","family":"Wei","sequence":"additional","affiliation":[{"name":"South China University of Technology, Guangzhou, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6903-145X","authenticated-orcid":false,"given":"Shuhan","family":"Qi","sequence":"additional","affiliation":[{"name":"Harbin Institute of Technology, Shenzhen, Shenzhen, China and Leanplans, Shenzhen, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-1630-424X","authenticated-orcid":false,"given":"Chengkai","family":"Huang","sequence":"additional","affiliation":[{"name":"Macquarie University, Sydney, Australia and The University of New South Wales, Sydney, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3326-4147","authenticated-orcid":false,"given":"Quan Z.","family":"Sheng","sequence":"additional","affiliation":[{"name":"Macquarie University, Sydney, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4149-839X","authenticated-orcid":false,"given":"Lina","family":"Yao","sequence":"additional","affiliation":[{"name":"The University of New South Wales, Sydney, Australia and CSIRO\u2019s Data61, Sydney, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,4,12]]},"reference":[{"key":"e_1_3_2_1_1_1","unstructured":"Akari Asai Zeqiu Wu Yizhong Wang Avirup Sil and Hannaneh Hajishirzi. 2024. Self-RAG: Learning to Retrieve Generate and Critique through Self-Reflection. In ICLR. OpenReview.net."},{"key":"e_1_3_2_1_2_1","first-page":"12541","article-title":"Probabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex Questions. In EMNLP (Findings)","author":"Cao Shulin","year":"2023","unstructured":"Shulin Cao, Jiajie Zhang, Jiaxin Shi, Xin Lv, Zijun Yao, Qi Tian, Lei Hou, and Juanzi Li. 2023. Probabilistic Tree-of-thought Reasoning for Answering Knowledge-intensive Complex Questions. In EMNLP (Findings). Association for Computational Linguistics, 12541-12560.","journal-title":"Association for Computational Linguistics"},{"key":"e_1_3_2_1_3_1","volume-title":"Towards reasoning era: A survey of long chain-of-thought for reasoning large language models. arXiv preprint arXiv:2503.09567","author":"Chen Qiguang","year":"2025","unstructured":"Qiguang Chen, Libo Qin, Jinhao Liu, Dengyun Peng, Jiannan Guan, Peng Wang, Mengkang Hu, Yuhang Zhou, Te Gao, and Wanxiang Che. 2025. Towards reasoning era: A survey of long chain-of-thought for reasoning large language models. arXiv preprint arXiv:2503.09567 (2025)."},{"key":"e_1_3_2_1_4_1","unstructured":"Wenfeng Feng Chuzhan Hao Yuewei Zhang Jingyi Song and Hao Wang. 2025. AirRAG: Activating Intrinsic Reasoning for Retrieval Augmented Generation using Tree-based Search. arXiv:2501.10053 [cs.AI]"},{"key":"e_1_3_2_1_5_1","volume-title":"Decomposing Complex Questions Makes Multi-Hop QA Easier and More Interpretable. In Findings of the Association for Computational Linguistics: EMNLP 2021","author":"Fu Ruiliu","year":"2021","unstructured":"Ruiliu Fu, Han Wang, Xuejun Zhang, Jun Zhou, and Yonghong Yan. 2021. Decomposing Complex Questions Makes Multi-Hop QA Easier and More Interpretable. In Findings of the Association for Computational Linguistics: EMNLP 2021, Virtual Event \/ Punta Cana, Dominican Republic, 16-20 November, 2021. Association for Computational Linguistics, 169-180."},{"key":"e_1_3_2_1_6_1","unstructured":"Yunfan Gao Yun Xiong Xinyu Gao Kangxiang Jia Jinliu Pan Yuxi Bi Yi Dai Jiawei Sun Meng Wang and Haofen Wang. 2024. Retrieval-Augmented Generation for Large Language Models: A Survey. arXiv:2312.10997 [cs.CL]"},{"key":"e_1_3_2_1_7_1","unstructured":"Aaron Grattafiori Abhimanyu Dubey Abhinav Jauhri Abhinav Pandey Abhishek Kadian Ahmad Al-Dahle Aiesha Letman Akhil Mathur Alan Schelten Alex Vaughan Amy Yang Angela Fan Anirudh Goyal Anthony Hartshorn Aobo Yang Archi Mitra Archie Sravankumar Artem Korenev and et al. 2024. The Llama 3 Herd of Models. arXiv:2407.21783 [cs.AI]"},{"key":"e_1_3_2_1_8_1","volume-title":"International conference on machine learning. PMLR, 3929-3938","author":"Guu Kelvin","year":"2020","unstructured":"Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Mingwei Chang. 2020. Retrieval augmented language model pre-training. In International conference on machine learning. PMLR, 3929-3938."},{"key":"e_1_3_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.coling-main.580"},{"key":"e_1_3_2_1_10_1","volume-title":"Proceedings of the 31st International Conference on Computational Linguistics. 1403-1412","author":"Huang Chengkai","year":"2025","unstructured":"Chengkai Huang, Yu Xia, Rui Wang, Kaige Xie, Tong Yu, Julian McAuley, and Lina Yao. 2025. Embedding-informed adaptive retrieval-augmented generation of large language models. In Proceedings of the 31st International Conference on Computational Linguistics. 1403-1412."},{"key":"e_1_3_2_1_11_1","volume-title":"Foundation models for recommender systems: A survey and new perspectives. arXiv preprint arXiv:2402.11143","author":"Huang Chengkai","year":"2024","unstructured":"Chengkai Huang, Tong Yu, Kaige Xie, Shuai Zhang, Lina Yao, and Julian McAuley. 2024. Foundation models for recommender systems: A survey and new perspectives. arXiv preprint arXiv:2402.11143 (2024)."},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.389"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.naacl-long.361"},{"key":"e_1_3_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/TBDATA.2019.2921572"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/P17-1147"},{"key":"e_1_3_2_1_16_1","volume-title":"Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih.","author":"Karpukhin Vladimir","year":"2020","unstructured":"Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick SH Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020. Dense Passage Retrieval for Open-Domain Question Answering.. In EMNLP (1). 6769-6781."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00276"},{"key":"e_1_3_2_1_18_1","first-page":"611","article-title":"Efficient Memory Management for Large Language Model Serving with PagedAttention","author":"Kwon Woosuk","year":"2023","unstructured":"Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph Gonzalez, Hao Zhang, and Ion Stoica. 2023. Efficient Memory Management for Large Language Model Serving with PagedAttention. In SOSP. ACM, 611-626.","journal-title":"SOSP. ACM"},{"key":"e_1_3_2_1_19_1","unstructured":"Patrick Lewis Ethan Perez Aleksandra Piktus Fabio Petroni Vladimir Karpukhin Naman Goyal Heinrich K\u00fcttler Mike Lewis Wen-tau Yih Tim Rockt\u00e4schel et al. 2020. Retrieval-augmented generation for knowledge-intensive nlp tasks. Advances in neural information processing systems Vol. 33 (2020) 9459-9474."},{"key":"e_1_3_2_1_20_1","volume-title":"Search-o1: Agentic search-enhanced large reasoning models. arXiv preprint arXiv:2501.05366","author":"Li Xiaoxi","year":"2025","unstructured":"Xiaoxi Li, Guanting Dong, Jiajie Jin, Yuyao Zhang, Yujia Zhou, Yutao Zhu, Peitian Zhang, and Zhicheng Dou. 2025. Search-o1: Agentic search-enhanced large reasoning models. arXiv preprint arXiv:2501.05366 (2025)."},{"key":"e_1_3_2_1_21_1","unstructured":"Nelson F. Liu Kevin Lin John Hewitt Ashwin Paranjape Michele Bevilacqua Fabio Petroni and Percy Liang. 2023. Lost in the Middle: How Language Models Use Long Contexts. arXiv:2307.03172 [cs.CL]"},{"key":"e_1_3_2_1_22_1","volume-title":"UltraLED: Learning to See Everything in Ultra-High Dynamic Range Scenes. CoRR","author":"Meng Yuang","year":"2025","unstructured":"Yuang Meng, Xin Jin, Lina Lei, Chun-Le Guo, and Chongyi Li. 2025. UltraLED: Learning to See Everything in Ultra-High Dynamic Range Scenes. CoRR, Vol. abs\/2510.07741 (2025)."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.acl-long.1352"},{"key":"e_1_3_2_1_24_1","volume-title":"Measuring and Narrowing the Compositionality Gap in Language Models. In Findings of the Association for Computational Linguistics: EMNLP 2023","author":"Press Ofir","year":"2023","unstructured":"Ofir Press, Muru Zhang, Sewon Min, Ludwig Schmidt, Noah A. Smith, and Mike Lewis. 2023. Measuring and Narrowing the Compositionality Gap in Language Models. In Findings of the Association for Computational Linguistics: EMNLP 2023, Singapore, December 6-10, 2023. Association for Computational Linguistics, 5687-5711."},{"key":"e_1_3_2_1_25_1","first-page":"2366","article-title":"MemoRAG","author":"Qian Hongjin","year":"2025","unstructured":"Hongjin Qian, Zheng Liu, Peitian Zhang, Kelong Mao, Defu Lian, Zhicheng Dou, and Tiejun Huang. 2025. MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation. In WWW. ACM, 2366-2377.","journal-title":"In WWW. ACM"},{"key":"e_1_3_2_1_26_1","volume-title":"Asa Cooper Stickland, Jackson Petty, Richard Yuanzhe Pang, Julien Dirani, Julian Michael, and Samuel R. Bowman.","author":"Rein David","year":"2023","unstructured":"David Rein, Betty Li Hou, Asa Cooper Stickland, Jackson Petty, Richard Yuanzhe Pang, Julien Dirani, Julian Michael, and Samuel R. Bowman. 2023. GPQA: A Graduate-Level Google-Proof Q&A Benchmark. CoRR, Vol. abs\/2311.12022 (2023)."},{"key":"e_1_3_2_1_27_1","first-page":"13773","volume-title":"ConTReGen: Context-driven Tree-structured Retrieval for Open-domain Long-form Text Generation. In Findings of the Association for Computational Linguistics: EMNLP","author":"Roy Kashob Kumar","year":"2024","unstructured":"Kashob Kumar Roy, Pritom Saha Akash, Kevin Chen-Chuan Chang, and Lucian Popa. 2024. ConTReGen: Context-driven Tree-structured Retrieval for Open-domain Long-form Text Generation. In Findings of the Association for Computational Linguistics: EMNLP 2024. Association for Computational Linguistics, Miami, Florida, USA, 13773-13784."},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3626772.3657957"},{"key":"e_1_3_2_1_29_1","unstructured":"Michael Shen Muhammad Umar Kiwan Maeng G. Edward Suh and Udit Gupta. 2024. Towards Understanding Systems Trade-offs in Retrieval-Augmented Generation Model Inference. arXiv:2412.11854 [cs.AR]"},{"key":"e_1_3_2_1_30_1","first-page":"3138","volume-title":"Proceedings of the 31st International Conference on Computational Linguistics, COLING 2025, Abu Dhabi, UAE","author":"Shen Tiesunlong","year":"2025","unstructured":"Tiesunlong Shen, Jin Wang, Xuejie Zhang, and Erik Cambria. 2025. Reasoning with Trees: Faithful Question Answering over Knowledge Graph. In Proceedings of the 31st International Conference on Computational Linguistics, COLING 2025, Abu Dhabi, UAE, January 19-24, 2025. Association for Computational Linguistics, 3138-3157."},{"key":"e_1_3_2_1_31_1","volume-title":"Retrieval Augmentation Reduces Hallucination in Conversation. In Findings of the Association for Computational Linguistics: EMNLP 2021","author":"Shuster Kurt","year":"2021","unstructured":"Kurt Shuster, Spencer Poff, Moya Chen, Douwe Kiela, and Jason Weston. 2021. Retrieval Augmentation Reduces Hallucination in Conversation. In Findings of the Association for Computational Linguistics: EMNLP 2021, Virtual Event \/ Punta Cana, Dominican Republic, 16-20 November, 2021. Association for Computational Linguistics, 3784-3803."},{"key":"e_1_3_2_1_32_1","unstructured":"Ishneet Sukhvinder Singh Ritvik Aggarwal Ibrahim Allahverdiyev Muhammad Taha Aslihan Akalin Kevin Zhu and Sean O'Brien. 2025. ChunkRAG: Novel LLM-Chunk Filtering Method for RAG Systems. arXiv:2410.19572 [cs.CL]"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.acl-long.702"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.acl-long.1175"},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/3696410.3714892"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.emnlp-main.730"},{"key":"e_1_3_2_1_37_1","volume-title":"MemoTime: Memory-Augmented Temporal Knowledge Graph Enhanced Large Language Model Reasoning. arXiv preprint arXiv:2510.13614","author":"Tan Xingyu","year":"2025","unstructured":"Xingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu, Xin Yuan, Liming Zhu, and Wenjie Zhang. 2025c. MemoTime: Memory-Augmented Temporal Knowledge Graph Enhanced Large Language Model Reasoning. arXiv preprint arXiv:2510.13614 (2025)."},{"key":"e_1_3_2_1_38_1","volume-title":"PrivGemo: Privacy-Preserving Dual-Tower Graph Retrieval for Empowering LLM Reasoning with Memory Augmentation. arXiv preprint arXiv:2601.08739","author":"Tan Xingyu","year":"2026","unstructured":"Xingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu, Xin Yuan, Liming Zhu, and Wenjie Zhang. 2026. PrivGemo: Privacy-Preserving Dual-Tower Graph Retrieval for Empowering LLM Reasoning with Memory Augmentation. arXiv preprint arXiv:2601.08739 (2026)."},{"key":"e_1_3_2_1_39_1","unstructured":"Hugo Touvron Louis Martin Kevin Stone and et al. 2023. Llama 2: Open Foundation and Fine-Tuned Chat Models. CoRR Vol. abs\/2307.09288 (2023)."},{"key":"e_1_3_2_1_40_1","volume-title":"Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions. arXiv preprint arXiv:2212.10509","author":"Trivedi Harsh","year":"2022","unstructured":"Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022a. Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions. arXiv preprint arXiv:2212.10509 (2022)."},{"key":"e_1_3_2_1_41_1","first-page":"539","article-title":"MuSiQue: Multihop Questions via Single-hop Question","volume":"10","author":"Trivedi Harsh","year":"2022","unstructured":"Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022b. MuSiQue: Multihop Questions via Single-hop Question Composition. Trans. Assoc. Comput. Linguistics, Vol. 10 (2022), 539-554.","journal-title":"Composition. Trans. Assoc. Comput. Linguistics"},{"key":"e_1_3_2_1_42_1","unstructured":"Liang Wang Nan Yang Xiaolong Huang and et al. 2022. Text Embeddings by Weakly-Supervised Contrastive Pre-training. CoRR Vol. abs\/2212.03533 (2022). arXiv:2212.03533"},{"key":"e_1_3_2_1_43_1","unstructured":"Shang Wang Tianqing Zhu Dayong Ye and Wanlei Zhou. 2024. When Machine Unlearning Meets Retrieval-Augmented Generation (RAG): Keep Secret or Forget Knowledge? arXiv:2410.15267 [cs.CR]"},{"key":"e_1_3_2_1_44_1","volume-title":"Denny Zhou, et al.","author":"Wei Jason","year":"2022","unstructured":"Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al., 2022. Chain-of-thought prompting elicits reasoning in large language models. Advances in neural information processing systems, Vol. 35 (2022), 24824-24837."},{"key":"e_1_3_2_1_45_1","volume-title":"C-Pack: Packaged Resources To Advance General Chinese Embedding. CoRR","author":"Xiao Shitao","year":"2023","unstructured":"Shitao Xiao, Zheng Liu, Peitian Zhang, and Niklas Muennighoff. 2023. C-Pack: Packaged Resources To Advance General Chinese Embedding. CoRR, Vol. abs\/2309.07597 (2023). arXiv:2309.07597"},{"key":"e_1_3_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1145\/3589334.3645363"},{"key":"e_1_3_2_1_47_1","unstructured":"An Yang Anfeng Li Baosong Yang and et al. 2025. Qwen3 Technical Report. arXiv:2505.09388 [cs.CL]"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D18-1259"},{"key":"e_1_3_2_1_49_1","volume-title":"Tree of thoughts: Deliberate problem solving with large language models. Advances in neural information processing systems","author":"Yao Shunyu","year":"2023","unstructured":"Shunyu Yao, Dian Yu, Jeffrey Zhao, Izhak Shafran, Tom Griffiths, Yuan Cao, and Karthik Narasimhan. 2023a. Tree of thoughts: Deliberate problem solving with large language models. Advances in neural information processing systems, Vol. 36 (2023), 11809-11822."},{"key":"e_1_3_2_1_50_1","volume-title":"International Conference on Learning Representations (ICLR).","author":"Yao Shunyu","year":"2023","unstructured":"Shunyu Yao, Jeffrey Zhao, Dian Yu, Nan Du, Izhak Shafran, Karthik Narasimhan, and Yuan Cao. 2023b. React: Synergizing reasoning and acting in language models. In International Conference on Learning Representations (ICLR)."},{"key":"e_1_3_2_1_51_1","volume-title":"PRoH: Dynamic Planning and Reasoning over Knowledge Hypergraphs for Retrieval-Augmented Generation. arXiv preprint arXiv:2510.12434","author":"Zai Xiangjun","year":"2025","unstructured":"Xiangjun Zai, Xingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu, and Wenjie Zhang. 2025. PRoH: Dynamic Planning and Reasoning over Knowledge Hypergraphs for Retrieval-Augmented Generation. arXiv preprint arXiv:2510.12434 (2025)."}],"event":{"name":"WWW '26: The ACM Web Conference 2026","location":"Dubai United Arab Emirates","sponsor":["SIGWEB ACM Special Interest Group on Hypertext, Hypermedia, and Web"]},"container-title":["Proceedings of the ACM Web Conference 2026"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3774904.3792103","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,4]],"date-time":"2026-07-04T07:22:55Z","timestamp":1783149775000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3774904.3792103"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,12]]},"references-count":51,"alternative-id":["10.1145\/3774904.3792103","10.1145\/3774904"],"URL":"https:\/\/doi.org\/10.1145\/3774904.3792103","relation":{},"subject":[],"published":{"date-parts":[[2026,4,12]]},"assertion":[{"value":"2026-04-12","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}