{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,11]],"date-time":"2026-07-11T02:49:51Z","timestamp":1783738191714,"version":"3.55.0"},"reference-count":85,"publisher":"Association for Computing Machinery (ACM)","issue":"6","funder":[{"name":"Hong Kong Research Grants Council\u2019s Research Impact Fund","award":["R1015-23"],"award-info":[{"award-number":["R1015-23"]}]},{"name":"Collaborative Research Fund","award":["C1043-24GF"],"award-info":[{"award-number":["C1043-24GF"]}]},{"name":"General Research Fund","award":["11218325"],"award-info":[{"award-number":["11218325"]}]},{"name":"Institute of Digital Medicine of City University of Hong Kong","award":["9229503"],"award-info":[{"award-number":["9229503"]}]},{"name":"Huawei (Huawei Innovation Research Program), Tencent"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Inf. Syst."],"published-print":{"date-parts":[[2025,11,30]]},"abstract":"<jats:p>\n            Search engines are crucial as they provide an efficient and easy way to access vast amounts of information on the Internet for diverse information needs. User queries, even with a specific need, can differ significantly. Prior research has explored the resilience of ranking models against typical query variations like paraphrasing, misspellings, and order changes. Yet, these works overlook how diverse demographics uniquely formulate identical queries. For instance, older individuals tend to construct queries more naturally and in varied order compared to other groups. This demographic diversity necessitates enhancing the adaptability of ranking models to diverse query formulations. To this end, in this article, we propose a framework that integrates a novel rewriting pipeline that rewrites queries from various demographic perspectives and a novel framework to enhance ranking robustness. To be specific, we use Chain of Thought (CoT) technology to utilize Large Language Models (LLMs) as agents to emulate various demographic profiles, then use them for efficient query rewriting, and we innovate a Robust Multi-gate Mixture-of-Experts (R-MMoE) architecture coupled with a hybrid loss function, collectively strengthening the ranking models\u2019 robustness. Our extensive experiments on both public and industrial datasets assesses the efficacy of our query rewriting approach and the enhanced accuracy and robustness of the ranking model. The findings highlight the sophistication and effectiveness of our proposed model. We release our code implementation publicly (\n            <jats:ext-link xmlns:xlink=\"http:\/\/www.w3.org\/1999\/xlink\" ext-link-type=\"uri\" xlink:href=\"https:\/\/github.com\/Applied-Machine-Learning-Lab\/ROBR\">https:\/\/github.com\/Applied-Machine-Learning-Lab\/ROBR<\/jats:ext-link>\n            ).\n          <\/jats:p>","DOI":"10.1145\/3749099","type":"journal-article","created":{"date-parts":[[2025,7,18]],"date-time":"2025-07-18T04:51:01Z","timestamp":1752814261000},"page":"1-33","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Agent4Ranking: Semantic Robust Ranking via Personalized Query Rewriting Using Multi-Agent LLMs"],"prefix":"10.1145","volume":"43","author":[{"ORCID":"https:\/\/orcid.org\/0009-0008-6162-8500","authenticated-orcid":false,"given":"Xiaopeng","family":"Li","sequence":"first","affiliation":[{"name":"City University of Hong Kong, Hong Kong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-1246-1740","authenticated-orcid":false,"given":"Lixin","family":"Su","sequence":"additional","affiliation":[{"name":"Baidu Inc, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4712-3676","authenticated-orcid":false,"given":"Pengyue","family":"Jia","sequence":"additional","affiliation":[{"name":"City University of Hong Kong, Hong Kong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-3622-3399","authenticated-orcid":false,"given":"Suqi","family":"Cheng","sequence":"additional","affiliation":[{"name":"Baidu Inc, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-7347-143X","authenticated-orcid":false,"given":"Junfeng","family":"Wang","sequence":"additional","affiliation":[{"name":"Baidu Inc, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-0684-6205","authenticated-orcid":false,"given":"Dawei","family":"Yin","sequence":"additional","affiliation":[{"name":"Baidu Inc, Beijing, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2926-4416","authenticated-orcid":false,"given":"Xiangyu","family":"Zhao","sequence":"additional","affiliation":[{"name":"City University of Hong Kong, Hong Kong, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2025,9,10]]},"reference":[{"key":"e_1_3_2_2_2","article-title":"Robust returns ranking prediction and portfolio optimization for M6","author":"Ai Hongfeng","year":"2024","unstructured":"Hongfeng Ai, Chenning Liu, and Peng Lin. 2024. Robust returns ranking prediction and portfolio optimization for M6. International Journal of Forecasting (2024).","journal-title":"International Journal of Forecasting"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.1145\/3539618.3591960"},{"key":"e_1_3_2_4_2","unstructured":"Avishek Anand Venktesh V Abhijit Anand Vinay Setty. 2023. Query understanding in the age of large language models. arXiv:2306.16004. Retrieved from https:\/\/arxiv.org\/abs\/2306.16004"},{"key":"e_1_3_2_5_2","first-page":"1609","volume-title":"Companion Proceedings of the ACM on Web Conference 2024","author":"Ayoub Michael Antonios Kruse","year":"2024","unstructured":"Michael Antonios Kruse Ayoub, Zhan Su, and Qiuchi Li. 2024. A case study of enhancing sparse retrieval using LLMs. In Companion Proceedings of the ACM on Web Conference 2024, 1609\u20131615."},{"key":"e_1_3_2_6_2","unstructured":"Daniel Campos ChengXiang Zhai and Alessandro Magnani. 2023. Noise-robust dense retrieval via contrastive alignment post training. arXiv:2304.03401. Retrieved from https:\/\/arxiv.org\/abs\/2304.03401"},{"key":"e_1_3_2_7_2","doi-asserted-by":"crossref","first-page":"14868","DOI":"10.18653\/v1\/2024.findings-acl.882","volume-title":"Findings of the Association for Computational Linguistics (ACL \u201924)","author":"Cheng Xuxin","year":"2024","unstructured":"Xuxin Cheng, Zhihong Zhu, Xianwei Zhuang, Zhanpeng Chen, Zhiqi Huang, and Yuexian Zou. 2024. MoE-SLU: Towards ASR-robust spoken language understanding via mixture-of-experts. In Findings of the Association for Computational Linguistics (ACL \u201924), 14868\u201314879."},{"key":"e_1_3_2_8_2","unstructured":"CNNIC. n.\u2009d. Statistical Report on Internet Development in China. Retrieved from https:\/\/www.cnnic.net.cn\/n4\/2022\/0401\/c88-1125.html"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3331184.3331303"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1145\/3159652.3159659"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1016\/S0004-3702(03)00079-1"},{"key":"e_1_3_2_12_2","unstructured":"Jacob Devlin Ming-Wei Chang Kenton Lee and Kristina Toutanova. 2018. Bert: Pre-training of deep bidirectional transformers for language understanding. arXiv:1810.04805. Retrieved from https:\/\/arxiv.org\/abs\/1810.04805"},{"key":"e_1_3_2_13_2","unstructured":"Jiazhan Feng Chongyang Tao Xiubo Geng Tao Shen Can Xu Guodong Long Dongyan Zhao and Daxin Jiang. 2023. Knowledge refinement via interaction between search engines and large language models. arXiv:2305.07402. Retrieved from https:\/\/arxiv.org\/abs\/2305.07402"},{"key":"e_1_3_2_14_2","unstructured":"Ivar Frisch and Mario Giulianelli. 2024. LLM agents in interaction: Measuring personality consistency and linguistic alignment in interacting populations of large language models. arXiv:2402.02896. Retrieved from https:\/\/arxiv.org\/abs\/2402.02896"},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1145\/3698878"},{"key":"e_1_3_2_16_2","doi-asserted-by":"publisher","DOI":"10.1145\/3701716.3715253"},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.1145\/3637871"},{"key":"e_1_3_2_18_2","doi-asserted-by":"crossref","unstructured":"Luyu Gao Xueguang Ma Jimmy Lin and Jamie Callan. 2022. Precise zero-shot dense retrieval without relevance labels. arXiv:2212.10496. Retrieved from https:\/\/arxiv.org\/abs\/2212.10496","DOI":"10.18653\/v1\/2023.acl-long.99"},{"key":"e_1_3_2_19_2","unstructured":"Tianyu Gao Xingcheng Yao and Danqi Chen. 2021. Simcse: Simple contrastive learning of sentence embeddings. arXiv:2104.08821. Retrieved from https:\/\/arxiv.org\/abs\/2104.08821"},{"key":"e_1_3_2_20_2","doi-asserted-by":"publisher","DOI":"10.1145\/3340531.3412330"},{"key":"e_1_3_2_21_2","doi-asserted-by":"publisher","DOI":"10.1145\/2983323.2983769"},{"key":"e_1_3_2_22_2","first-page":"475","volume-title":"Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing: Industry Track","author":"Hao Jie","year":"2022","unstructured":"Jie Hao, Yang Liu, Xing Fan, Saurabh Gupta, Saleh Soltan, Rakesh Chada, Pradeep Natarajan, Chenlei Guo, and G\u00f6khan T\u00fcr. 2022. CGF: Constrained generation framework for query rewriting in conversational AI. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing: Industry Track, 475\u2013483."},{"key":"e_1_3_2_23_2","first-page":"2790","volume-title":"Proceedings of the International Conference on Machine Learning","author":"Houlsby Neil","year":"2019","unstructured":"Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019. Parameter-efficient transfer learning for NLP. In Proceedings of the International Conference on Machine Learning. PMLR, 2790\u20132799."},{"key":"e_1_3_2_24_2","unstructured":"Hui Huang Yingqi Qu Jing Liu Muyun Yang and Tiejun Zhao. 2024. An empirical study of llm-as-a-judge for llm evaluation: Fine-tuned judge models are task-specific classifiers. arXiv:2403.02839. Retrieved from https:\/\/arxiv.org\/abs\/2403.02839"},{"key":"e_1_3_2_25_2","doi-asserted-by":"crossref","unstructured":"Gautier Izacard and Edouard Grave. 2020. Leveraging passage retrieval with generative models for open domain question answering. arXiv:2007.01282. Retrieved from https:\/\/arxiv.org\/abs\/2007.01282","DOI":"10.18653\/v1\/2021.eacl-main.74"},{"key":"e_1_3_2_26_2","unstructured":"Rolf Jagerman Honglei Zhuang Zhen Qin Xuanhui Wang and Michael Bendersky. 2023. Query expansion by prompting large language models. arXiv:2305.03653. Retrieved from https:\/\/arxiv.org\/abs\/2305.03653"},{"key":"e_1_3_2_27_2","unstructured":"Pengyue Jia Yiding Liu Xiangyu Zhao Xiaopeng Li Changying Hao Shuaiqiang Wang and Dawei Yin. 2023. MILL: Mutual verification with large language models for zero-shot query expansion. arXiv:2310.19056. Retrieved from https:\/\/arxiv.org\/abs\/2310.19056"},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v38i8.28699"},{"key":"e_1_3_2_29_2","unstructured":"Pengyue Jia Derong Xu Xiaopeng Li Zhaocheng Du Xiangyang Li Xiangyu Zhao Yichao Wang Yuhao Wang Huifeng Guo and Ruiming Tang. 2024. Bridging relevance and reasoning: Rationale distillation in retrieval-augmented generation. arXiv:2412.08519. Retrieved from https:\/\/arxiv.org\/abs\/2412.08519"},{"key":"e_1_3_2_30_2","unstructured":"Hang Jiang Xiajie Zhang Xubo Cao and Jad Kabbara. 2023. Personallm: Investigating the ability of large language models to express big five personality traits. arXiv:2305.02547. Retrieved from https:\/\/arxiv.org\/abs\/2305.02547"},{"key":"e_1_3_2_31_2","unstructured":"Yibin Lei Yu Cao Tianyi Zhou Tao Shen and Andrew Yates. 2024. Corpus-steered query expansion with large language models. arXiv:2402.18031. Retrieved from https:\/\/arxiv.org\/abs\/2402.18031"},{"key":"e_1_3_2_32_2","doi-asserted-by":"publisher","DOI":"10.1145\/3604915.3608779"},{"key":"e_1_3_2_33_2","unstructured":"Xiaopeng Li Pengyue Jia Derong Xu Yi Wen Yingyi Zhang Wenlin Zhang Wanyu Wang Yichao Wang Zhaocheng Du Xiangyang Li et al. 2025. A survey of personalization: From rag to agent. arXiv:2504.10147. Retrieved from https:\/\/arxiv.org\/abs\/2504.10147"},{"key":"e_1_3_2_34_2","unstructured":"Xiaopeng Li Xiangyang Li Hao Zhang Zhaocheng Du Pengyue Jia Yichao Wang Xiangyu Zhao Huifeng Guo and Ruiming Tang. 2024. SyNeg: LLM-driven synthetic hard-negatives for dense retrieval. arXiv:2412.17250. Retrieved from https:\/\/arxiv.org\/abs\/2412.17250"},{"key":"e_1_3_2_35_2","doi-asserted-by":"publisher","DOI":"10.1145\/3583780.3615137"},{"key":"e_1_3_2_36_2","doi-asserted-by":"publisher","DOI":"10.1145\/3543507.3583339"},{"key":"e_1_3_2_37_2","doi-asserted-by":"publisher","DOI":"10.1145\/3485447.3512078"},{"key":"e_1_3_2_38_2","unstructured":"Jie Liu and Barzan Mozafari. 2024. Query rewriting via large language models. arXiv:2403.09060. Retrieved from https:\/\/arxiv.org\/abs\/2403.09060"},{"key":"e_1_3_2_39_2","doi-asserted-by":"publisher","DOI":"10.1145\/3626772.3657722"},{"key":"e_1_3_2_40_2","unstructured":"Qidong Liu Xian Wu Xiangyu Zhao Yuanshao Zhu Zijian Zhang Feng Tian and Yefeng Zheng. 2024. Large language model distilling medication recommendation model. arXiv:2402.02803. Retrieved from https:\/\/arxiv.org\/abs\/2402.02803"},{"key":"e_1_3_2_41_2","doi-asserted-by":"publisher","DOI":"10.1145\/3543507.3583244"},{"key":"e_1_3_2_42_2","unstructured":"Yinhan Liu Myle Ott Naman Goyal Jingfei Du Mandar Joshi Danqi Chen Omer Levy Mike Lewis Luke Zettlemoyer and Veselin Stoyanov. 2019. Roberta: A robustly optimized BERT pretraining approach. arXiv:1907.11692. Retrieved from https:\/\/arxiv.org\/abs\/1907.11692"},{"key":"e_1_3_2_43_2","unstructured":"Wenxin Luo Weirui Wang Xiaopeng Li Weibo Zhou Pengyue Jia and Xiangyu Zhao. 2025. TAPO: Task-referenced adaptation for prompt optimization. arXiv:2501.06689. Retrieved from https:\/\/arxiv.org\/abs\/2501.06689"},{"key":"e_1_3_2_44_2","first-page":"484","volume-title":"European Conference on Information Retrieval","author":"Lupart Simon","year":"2023","unstructured":"Simon Lupart and St\u00e9phane Clinchant. 2023. A study on FGSM adversarial training for neural retrieval. In European Conference on Information Retrieval. Springer, 484\u2013492."},{"key":"e_1_3_2_45_2","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3220007"},{"key":"e_1_3_2_46_2","unstructured":"Xinbei Ma Yeyun Gong Pengcheng He Hai Zhao and Nan Duan. 2023. Query rewriting for retrieval-augmented large language models. arXiv:2305.14283. Retrieved from https:\/\/arxiv.org\/abs\/2305.14283"},{"key":"e_1_3_2_47_2","doi-asserted-by":"publisher","DOI":"10.1145\/3404835.3462869"},{"key":"e_1_3_2_48_2","unstructured":"Iain Mackie Shubham Chatterjee and Jeffrey Dalton. 2023. Generative and pseudo-relevant feedback for sparse dense and learned sparse retrieval. arXiv:2305.07477. Retrieved from https:\/\/arxiv.org\/abs\/2305.07477"},{"key":"e_1_3_2_49_2","unstructured":"Shengyu Mao Yong Jiang Boli Chen Xiao Li Peng Wang Xinyu Wang Pengjun Xie Fei Huang Huajun Chen and Ningyu Zhang. 2024. RaFe: Ranking feedback improves query rewriting for RAG. arXiv:2405.14431. Retrieved from https:\/\/arxiv.org\/abs\/2405.14431"},{"key":"e_1_3_2_50_2","doi-asserted-by":"publisher","unstructured":"Niklas Muennighoff Nouamane Tazi Lo\u00efc Magne and Nils Reimers. 2022. MTEB: Massive text embedding benchmark. arXiv:2210.07316. DOI: 10.48550\/ARXIV.2210.07316","DOI":"10.48550\/ARXIV.2210.07316"},{"key":"e_1_3_2_51_2","unstructured":"Jianmo Ni Gustavo Hernandez Abrego Noah Constant Ji Ma Keith B. Hall Daniel Cer and Yinfei Yang. 2021. Sentence-t5: Scalable sentence encoders from pre-trained text-to-text models. arXiv:2108.08877. Retrieved from https:\/\/arxiv.org\/abs\/2108.08877"},{"key":"e_1_3_2_52_2","first-page":"1","volume-title":"Proceedings of the 13th International Conference on World Wide Web","author":"Ntoulas Alexandros","year":"2004","unstructured":"Alexandros Ntoulas, Junghoo Cho, and Christopher Olston. 2004. What\u2019s new on the web? The evolution of the web from a search engine perspective. In Proceedings of the 13th International Conference on World Wide Web, 1\u201312."},{"key":"e_1_3_2_53_2","first-page":"397","volume-title":"European Conference on Information Retrieval","author":"Penha Gustavo","year":"2022","unstructured":"Gustavo Penha, Arthur C\u00e2mara, and Claudia Hauff. 2022. Evaluating the robustness of retrieval pipelines with query variation generators. In European Conference on Information Retrieval. Springer, 397\u2013412."},{"key":"e_1_3_2_54_2","doi-asserted-by":"crossref","unstructured":"Nils Reimers and Iryna Gurevych. 2019. Sentence-BERT: Sentence embeddings using Siamese BERT-networks. arXiv:1908.10084. Retrieved from https:\/\/arxiv.org\/abs\/1908.10084","DOI":"10.18653\/v1\/D19-1410"},{"key":"e_1_3_2_55_2","unstructured":"Ruiyang Ren Peng Qiu Yingqi Qu Jing Liu Wayne Xin Zhao Hua Wu Ji-Rong Wen and Haifeng Wang. 2024. Bases: Large-scale web search user simulation with large language model based agents. arXiv:2402.17505. Retrieved from https:\/\/arxiv.org\/abs\/2402.17505"},{"key":"e_1_3_2_56_2","doi-asserted-by":"publisher","DOI":"10.1561\/1500000019"},{"key":"e_1_3_2_57_2","doi-asserted-by":"publisher","DOI":"10.1002\/asi.4630270302"},{"key":"e_1_3_2_58_2","doi-asserted-by":"publisher","DOI":"10.1145\/361219.361220"},{"key":"e_1_3_2_59_2","doi-asserted-by":"crossref","unstructured":"Yunfan Shao Linyang Li Junqi Dai and Xipeng Qiu. 2023. Character-llm: A trainable agent for role-playing. arXiv:2310.10158. Retrieved from https:\/\/arxiv.org\/abs\/2310.10158","DOI":"10.18653\/v1\/2023.emnlp-main.814"},{"key":"e_1_3_2_60_2","unstructured":"Tao Shen Guodong Long Xiubo Geng Chongyang Tao Tianyi Zhou and Daxin Jiang. 2023. Large language models are strong zero-shot retriever. arXiv:2304.14233. Retrieved from https:\/\/arxiv.org\/abs\/2304.14233"},{"key":"e_1_3_2_61_2","first-page":"297","volume-title":"European Conference on Information Retrieval","author":"Sidiropoulos Georgios","year":"2024","unstructured":"Georgios Sidiropoulos and Evangelos Kanoulas. 2024. Improving the robustness of dense retrievers against typos via multi-positive contrastive learning. In European Conference on Information Retrieval. Springer, 297\u2013305."},{"key":"e_1_3_2_62_2","doi-asserted-by":"crossref","unstructured":"Krishna Srinivasan Karthik Raman Anupam Samanta Lingrui Liao Luca Bertelli and Mike Bendersky. 2022. QUILL: Query intent with large language models using retrieval augmentation and multi-stage distillation. arXiv:2210.15718. Retrieved from https:\/\/arxiv.org\/abs\/2210.15718","DOI":"10.18653\/v1\/2022.emnlp-industry.50"},{"key":"e_1_3_2_63_2","doi-asserted-by":"crossref","unstructured":"Weiwei Sun Lingyong Yan Xinyu Ma Shuaiqiang Wang Pengjie Ren Zhumin Chen Dawei Yin and Zhaochun Ren. 2023. Is ChatGPT good at search? Investigating large language models as re-ranking agents. arXiv:2304.09542. Retrieved from https:\/\/arxiv.org\/abs\/2304.09542","DOI":"10.18653\/v1\/2023.emnlp-main.923"},{"key":"e_1_3_2_64_2","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v34i05.6428"},{"key":"e_1_3_2_65_2","doi-asserted-by":"crossref","unstructured":"Panuthep Tasawong Wuttikorn Ponwitayarat Peerat Limkonchotiwat Can Udomcharoenchaikit Ekapol Chuangsuwanich and Sarana Nutanong. 2023. Typo-robust representation learning for dense retrieval. arXiv:2306.10348. Retrieved from https:\/\/arxiv.org\/abs\/2306.10348","DOI":"10.18653\/v1\/2023.acl-short.95"},{"key":"e_1_3_2_66_2","unstructured":"Nandan Thakur Nils Reimers Andreas R\u00fcckl\u00e9 Abhishek Srivastava and Iryna Gurevych. 2021. Beir: A heterogenous benchmark for zero-shot evaluation of information retrieval models. arXiv:2104.08663. Retrieved from https:\/\/arxiv.org\/abs\/2104.08663"},{"key":"e_1_3_2_67_2","doi-asserted-by":"crossref","unstructured":"Liang Wang Nan Yang and Furu Wei. 2023. Query2doc: Query expansion with large language models. arXiv:2303.07678. Retrieved from https:\/\/arxiv.org\/abs\/2303.07678","DOI":"10.18653\/v1\/2023.emnlp-main.585"},{"key":"e_1_3_2_68_2","doi-asserted-by":"publisher","DOI":"10.1145\/3539618.3591767"},{"key":"e_1_3_2_69_2","doi-asserted-by":"publisher","DOI":"10.1145\/3539618.3591750"},{"key":"e_1_3_2_70_2","unstructured":"Zekun Moore Wang Zhongyuan Peng Haoran Que Jiaheng Liu Wangchunshu Zhou Yuhan Wu Hongcheng Guo Ruitong Gan Zehao Ni Jian Yang et al. 2023. Rolellm: Benchmarking eliciting and enhancing role-playing abilities of large language models. arXiv:2310.00746. Retrieved from https:\/\/arxiv.org\/abs\/2310.00746"},{"issue":"2","key":"e_1_3_2_71_2","first-page":"1","article-title":"Are neural ranking models robust","volume":"41","author":"Wu Chen","year":"2022","unstructured":"Chen Wu, Ruqing Zhang, Jiafeng Guo, Yixing Fan, and Xueqi Cheng. 2022. Are neural ranking models robust? ACM Transactions on Information Systems 41, 2 (2022), 1\u201336.","journal-title":"ACM Transactions on Information Systems"},{"key":"e_1_3_2_72_2","unstructured":"Lee Xiong Chenyan Xiong Ye Li Kwok-Fung Tang Jialin Liu Paul Bennett Junaid Ahmed and Arnold Overwijk. 2020. Approximate nearest neighbor negative contrastive learning for dense text retrieval. arXiv:2007.00808. Retrieved from https:\/\/arxiv.org\/abs\/2007.00808"},{"key":"e_1_3_2_73_2","unstructured":"Derong Xu Pengyue Jia Xiaopeng Li Yingyi Zhang Maolin Wang Qidong Liu Xiangyu Zhao Yichao Wang Huifeng Guo Ruiming Tang et al. 2025. Align-GRAG: Reasoning-guided dual alignment for graph retrieval-augmented generation. arXiv:2505.16237. Retrieved from https:\/\/arxiv.org\/abs\/2505.16237"},{"key":"e_1_3_2_74_2","unstructured":"Fanghua Ye Meng Fang Shenghui Li and Emine Yilmaz. 2023. Enhancing conversational search: Large language model-aided informative query rewriting. arXiv:2310.09716. Retrieved from https:\/\/arxiv.org\/abs\/2310.09716"},{"key":"e_1_3_2_75_2","unstructured":"Wenhao Yu Dan Iter Shuohang Wang Yichong Xu Mingxuan Ju Soumya Sanyal Chenguang Zhu Michael Zeng and Meng Jiang. 2022. Generate rather than retrieve: Large language models are strong context generators. arXiv:2209.10063. Retrieved from https:\/\/arxiv.org\/abs\/2209.10063"},{"key":"e_1_3_2_76_2","unstructured":"Wenlin Zhang Xiangyang Li Kuicai Dong Yichao Wang Pengyue Jia Xiaopeng Li Yingyi Zhang Derong Xu Zhaocheng Du Huifeng Guo et al. 2025. Process vs. outcome reward: Which is better for agentic RAG reinforcement learning. arXiv:2505.14069. Retrieved from https:\/\/arxiv.org\/abs\/2505.14069"},{"key":"e_1_3_2_77_2","unstructured":"Xu Zhang Kaidi Xu Ziqing Hu and Ren Wang. 2025. Optimizing robustness and accuracy in mixture of experts: A dual-model approach. arXiv:2502.06832. Retrieved from https:\/\/arxiv.org\/abs\/2502.06832"},{"key":"e_1_3_2_78_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV51070.2023.00015"},{"key":"e_1_3_2_79_2","doi-asserted-by":"publisher","DOI":"10.1145\/3626772.3657686"},{"key":"e_1_3_2_80_2","first-page":"44880","article-title":"KuaiSim: A comprehensive simulator for recommender systems","volume":"36","author":"Zhao Kesen","year":"2023","unstructured":"Kesen Zhao, Shuchang Liu, Qingpeng Cai, Xiangyu Zhao, Ziru Liu, Dong Zheng, Peng Jiang, and Kun Gai. 2023. KuaiSim: A comprehensive simulator for recommender systems. In Advances in Neural Information Processing Systems, Vol. 36, 44880\u201344897.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_81_2","doi-asserted-by":"publisher","DOI":"10.1145\/3543507.3583418"},{"key":"e_1_3_2_82_2","doi-asserted-by":"publisher","DOI":"10.1145\/3240323.3240374"},{"key":"e_1_3_2_83_2","doi-asserted-by":"publisher","DOI":"10.1145\/3219819.3219886"},{"key":"e_1_3_2_84_2","first-page":"46595","article-title":"Judging llm-as-a-judge with mt-bench and chatbot arena","volume":"36","author":"Zheng Lianmin","year":"2023","unstructured":"Lianmin Zheng, Wei-Lin Chiang, Ying Sheng, Siyuan Zhuang, Zhanghao Wu, Yonghao Zhuang, Zi Lin, Zhuohan Li, Dacheng Li, Eric Xing. 2023. Judging llm-as-a-judge with mt-bench and chatbot arena. In Advances in Neural Information Processing Systems, Vol. 36, 46595\u201346623.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_85_2","doi-asserted-by":"crossref","unstructured":"Shengyao Zhuang Xueguang Ma Bevan Koopman Jimmy Lin and Guido Zuccon. 2024. PromptReps: Prompting large language models to generate dense and sparse representations for zero-shot document retrieval. arXiv:2404.18424. Retrieved from https:\/\/arxiv.org\/abs\/2404.18424","DOI":"10.18653\/v1\/2024.emnlp-main.250"},{"key":"e_1_3_2_86_2","doi-asserted-by":"publisher","DOI":"10.1145\/3477495.3531951"}],"container-title":["ACM Transactions on Information Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3749099","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,10]],"date-time":"2025-09-10T16:13:46Z","timestamp":1757520826000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3749099"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,9,10]]},"references-count":85,"journal-issue":{"issue":"6","published-print":{"date-parts":[[2025,11,30]]}},"alternative-id":["10.1145\/3749099"],"URL":"https:\/\/doi.org\/10.1145\/3749099","relation":{},"ISSN":["1046-8188","1558-2868"],"issn-type":[{"value":"1046-8188","type":"print"},{"value":"1558-2868","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,9,10]]},"assertion":[{"value":"2024-07-04","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-07-02","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-09-10","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}