{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,5]],"date-time":"2026-06-05T04:54:37Z","timestamp":1780635277779,"version":"3.54.1"},"reference-count":53,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2023,12,1]],"date-time":"2023-12-01T00:00:00Z","timestamp":1701388800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGIR Forum"],"published-print":{"date-parts":[[2023,12]]},"abstract":"<jats:p>\n            Modern search engines are built on a stack of different components, including query understanding, retrieval, multi-stage ranking, and question answering, among others. These components are often optimized and deployed independently. In this paper, we introduce a novel conceptual framework called\n            <jats:italic>large search model<\/jats:italic>\n            , which redefines the conventional search stack by unifying search tasks with one large language model (LLM). All tasks are formulated as autoregressive text generation problems, allowing for the customization of tasks through the use of natural language prompts. This proposed framework capitalizes on the strong language understanding and reasoning capabilities of LLMs, offering the potential to enhance search result quality while simultaneously simplifying the existing cumbersome search stack. To substantiate the feasibility of this framework, we present a series of proof-of-concept experiments and discuss the potential challenges associated with implementing this approach within real-world search systems.\n          <\/jats:p>","DOI":"10.1145\/3642979.3643006","type":"journal-article","created":{"date-parts":[[2024,1,22]],"date-time":"2024-01-22T17:05:12Z","timestamp":1705943112000},"page":"1-16","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":12,"title":["Large Search Model: Redefining Search Stack in the Era of LLMs"],"prefix":"10.1145","volume":"57","author":[{"given":"Liang","family":"Wang","sequence":"first","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Nan","family":"Yang","sequence":"additional","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiaolong","family":"Huang","sequence":"additional","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Linjun","family":"Yang","sequence":"additional","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Rangan","family":"Majumder","sequence":"additional","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Furu","family":"Wei","sequence":"additional","affiliation":[{"name":"Microsoft, USA"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2024,1,22]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Gqa: Training generalized multi-query transformer models from multi-head checkpoints. ArXiv preprint, abs\/2305.13245","author":"Ainslie Joshua","year":"2023","unstructured":"Joshua Ainslie, James Lee-Thorp, Michiel de Jong, Yury Zemlyanskiy, Federico Lebr\u00f3n, and Sumit Sanghai. Gqa: Training generalized multi-query transformer models from multi-head checkpoints. ArXiv preprint, abs\/2305.13245, 2023. URL https:\/\/arxiv.org\/abs\/2305.13245."},{"key":"e_1_2_1_2_1","first-page":"23716","article-title":"Flamingo: a visual language model for few-shot learning","volume":"35","author":"Alayrac Jean-Baptiste","year":"2022","unstructured":"Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katherine Millican, Malcolm Reynolds, et al. Flamingo: a visual language model for few-shot learning. Advances in Neural Information Processing Systems, 35:23716--23736, 2022.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_3_1","unstructured":"Sebastian Borgeaud Arthur Mensch Jordan Hoffmann Trevor Cai Eliza Rutherford Katie Millican George van den Driessche Jean-Baptiste Lespiau Bogdan Damoc Aidan Clark Diego de Las Casas Aurelia Guy Jacob Menick Roman Ring Tom Hennigan Saffron Huang Loren Maggiore Chris Jones Albin Cassirer Andy Brock Michela Paganini Geoffrey Irving Oriol Vinyals Simon Osindero Karen Simonyan Jack W. Rae Erich Elsen and Laurent Sifre. Improving language models by retrieving from trillions of tokens. In Kamalika Chaudhuri Stefanie Jegelka Le Song Csaba Szepesv\u00e1ri Gang Niu and Sivan Sabato editors International Conference on Machine Learning ICML 2022 17-23 July 2022 Baltimore Maryland USA volume 162 of Proceedings of Machine Learning Research pages 2206--2240. PMLR 2022. URL https:\/\/proceedings.mlr.press\/v162\/borgeaud22a.html."},{"key":"e_1_2_1_4_1","volume-title":"Extending context window of large language models via positional interpolation. ArXiv preprint, abs\/2306.15595","author":"Chen Shouyuan","year":"2023","unstructured":"Shouyuan Chen, Sherman Wong, Liangjian Chen, and Yuandong Tian. Extending context window of large language models via positional interpolation. ArXiv preprint, abs\/2306.15595, 2023. URL https:\/\/arxiv.org\/abs\/2306.15595."},{"key":"e_1_2_1_5_1","volume-title":"The Eleventh International Conference on Learning Representations","author":"Dai Zhuyun","year":"2022","unstructured":"Zhuyun Dai, Vincent Y Zhao, Ji Ma, Yi Luan, Jianmo Ni, Jing Lu, Anton Bakalov, Kelvin Guu, Keith Hall, and Ming-Wei Chang. Promptagator: Few-shot dense retrieval from 8 examples. In The Eleventh International Conference on Learning Representations, 2022."},{"key":"e_1_2_1_6_1","first-page":"16344","article-title":"Flashattention: Fast and memory-efficient exact attention with io-awareness","volume":"35","author":"Dao Tri","year":"2022","unstructured":"Tri Dao, Dan Fu, Stefano Ermon, Atri Rudra, and Christopher R\u00e9. Flashattention: Fast and memory-efficient exact attention with io-awareness. Advances in Neural Information Processing Systems, 35:16344--16359, 2022.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/N19-1423"},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.5555\/3586589.3586709"},{"key":"e_1_2_1_9_1","volume-title":"Red teaming language models to reduce harms: Methods, scaling behaviors, and lessons learned. ArXiv preprint, abs\/2209.07858","author":"Ganguli Deep","year":"2022","unstructured":"Deep Ganguli, Liane Lovitt, Jackson Kernion, Amanda Askell, Yuntao Bai, Saurav Kadavath, Ben Mann, Ethan Perez, Nicholas Schiefer, Kamal Ndousse, et al. Red teaming language models to reduce harms: Methods, scaling behaviors, and lessons learned. ArXiv preprint, abs\/2209.07858, 2022. URL https:\/\/arxiv.org\/abs\/2209.07858."},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.naacl-main.241"},{"key":"e_1_2_1_11_1","volume-title":"Precise zero-shot dense retrieval without relevance labels. ArXiv preprint, abs\/2212.10496","author":"Gao Luyu","year":"2022","unstructured":"Luyu Gao, Xueguang Ma, Jimmy Lin, and Jamie Callan. Precise zero-shot dense retrieval without relevance labels. ArXiv preprint, abs\/2212.10496, 2022. URL https:\/\/arxiv.org\/abs\/2212.10496."},{"key":"e_1_2_1_12_1","volume-title":"Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding. ArXiv preprint, abs\/1510.00149","author":"Han Song","year":"2015","unstructured":"Song Han, Huizi Mao, and William J Dally. Deep compression: Compressing deep neural networks with pruning, trained quantization and huffman coding. ArXiv preprint, abs\/1510.00149, 2015. URL https:\/\/arxiv.org\/abs\/1510.00149."},{"key":"e_1_2_1_13_1","volume-title":"Qiang Liu, et al. Language is not all you need: Aligning perception with language models. ArXiv preprint, abs\/2302.14045","author":"Huang Shaohan","year":"2023","unstructured":"Shaohan Huang, Li Dong, Wenhui Wang, Yaru Hao, Saksham Singhal, Shuming Ma, Tengchao Lv, Lei Cui, Owais Khan Mohammed, Qiang Liu, et al. Language is not all you need: Aligning perception with language models. ArXiv preprint, abs\/2302.14045, 2023. URL https:\/\/arxiv.org\/abs\/2302.14045."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.eacl-main.74"},{"key":"e_1_2_1_15_1","volume-title":"Scaling laws for neural language models. ArXiv preprint, abs\/2001.08361","author":"Kaplan Jared","year":"2020","unstructured":"Jared Kaplan, Sam McCandlish, T. J. Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeff Wu, and Dario Amodei. Scaling laws for neural language models. ArXiv preprint, abs\/2001.08361, 2020. URL https:\/\/arxiv.org\/abs\/2001.08361."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.emnlp-main.550"},{"key":"e_1_2_1_17_1","volume-title":"Alignment of language agents. ArXiv preprint, abs\/2103.14659","author":"Kenton Zachary","year":"2021","unstructured":"Zachary Kenton, Tom Everitt, Laura Weidinger, Iason Gabriel, Vladimir Mikulik, and Geoffrey Irving. Alignment of language agents. ArXiv preprint, abs\/2103.14659, 2021. URL https:\/\/arxiv.org\/abs\/2103.14659."},{"key":"e_1_2_1_18_1","volume-title":"8th International Conference on Learning Representations, ICLR 2020","author":"Khandelwal Urvashi","year":"2020","unstructured":"Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis. Generalization through memorization: Nearest neighbor language models. In 8th International Conference on Learning Representations, ICLR 2020, Addis Ababa, Ethiopia, April 26-30, 2020. OpenReview.net, 2020. URL https:\/\/openreview.net\/forum?id=HklBjCEKvH."},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1162\/tacl"},{"key":"e_1_2_1_20_1","first-page":"19274","volume-title":"International Conference on Machine Learning","author":"Leviathan Yaniv","year":"2023","unstructured":"Yaniv Leviathan, Matan Kalman, and Yossi Matias. Fast inference from transformers via speculative decoding. In International Conference on Machine Learning, pages 19274--19286. PMLR, 2023."},{"key":"e_1_2_1_21_1","volume-title":"Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020","author":"Lewis Patrick S. H.","year":"2020","unstructured":"Patrick S. H. Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich K\u00fcttler, Mike Lewis, Wen-tau Yih, Tim Rockt\u00e4schel, Sebastian Riedel, and Douwe Kiela. Retrieval-augmented generation for knowledge-intensive NLP tasks. In Hugo Larochelle, Marc'Aurelio Ranzato, Raia Hadsell, Maria-Florina Balcan, and Hsuan-Tien Lin, editors, Advances in Neural Information Processing Systems 33: Annual Conference on Neural Information Processing Systems 2020, NeurIPS 2020, December 6-12, 2020, virtual, 2020. URL https:\/\/proceedings.neurips.cc\/paper\/2020\/hash\/6b493230205f780e1bc26945df7481e5-Abstract.html."},{"issue":"1","key":"e_1_2_1_22_1","first-page":"29","article-title":"A proposed conceptual framework for a representational approach to information retrieval","volume":"55","author":"Lin Jimmy J.","year":"2021","unstructured":"Jimmy J. Lin. A proposed conceptual framework for a representational approach to information retrieval. ACM SIGIR Forum, 55:1 -- 29, 2021.","journal-title":"ACM SIGIR Forum"},{"key":"e_1_2_1_23_1","volume-title":"Lost in the middle: How language models use long contexts. ArXiv preprint, abs\/2307.03172","author":"Liu Nelson F.","year":"2023","unstructured":"Nelson F. Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. Lost in the middle: How language models use long contexts. ArXiv preprint, abs\/2307.03172, 2023a. URL https:\/\/arxiv.org\/abs\/2307.03172."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3580305.3599931"},{"key":"e_1_2_1_25_1","first-page":"1","volume-title":"Acm sigir forum","author":"Metzler Donald","year":"2021","unstructured":"Donald Metzler, Yi Tay, Dara Bahri, and Marc Najork. Rethinking search: making domain experts out of dilettantes. In Acm sigir forum, volume 55, pages 1--27. ACM New York, NY, USA, 2021."},{"key":"e_1_2_1_26_1","doi-asserted-by":"publisher","DOI":"10.1561\/1500000061"},{"key":"e_1_2_1_27_1","volume-title":"et al. Webgpt: Browser-assisted question-answering with human feedback. ArXiv preprint, abs\/2112.09332","author":"Nakano Reiichiro","year":"2021","unstructured":"Reiichiro Nakano, Jacob Hilton, Suchir Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, et al. Webgpt: Browser-assisted question-answering with human feedback. ArXiv preprint, abs\/2112.09332, 2021. URL https:\/\/arxiv.org\/abs\/2112.09332."},{"key":"e_1_2_1_28_1","volume-title":"Ms marco: A human-generated machine reading comprehension dataset","author":"Nguyen Tri","year":"2016","unstructured":"Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. Ms marco: A human-generated machine reading comprehension dataset. 2016."},{"key":"e_1_2_1_29_1","volume-title":"Multi-stage document ranking with bert. ArXiv preprint, abs\/1910.14424","author":"Nogueira Rodrigo","year":"2019","unstructured":"Rodrigo Nogueira, Wei Yang, Kyunghyun Cho, and Jimmy Lin. Multi-stage document ranking with bert. ArXiv preprint, abs\/1910.14424, 2019a. URL https:\/\/arxiv.org\/abs\/1910.14424."},{"key":"e_1_2_1_30_1","volume-title":"Document expansion by query prediction. ArXiv preprint, abs\/1904.08375","author":"Nogueira Rodrigo","year":"2019","unstructured":"Rodrigo Nogueira, Wei Yang, Jimmy Lin, and Kyunghyun Cho. Document expansion by query prediction. ArXiv preprint, abs\/1904.08375, 2019b. URL https:\/\/arxiv.org\/abs\/1904.08375."},{"key":"e_1_2_1_31_1","volume-title":"Gpt-4 technical report. ArXiv preprint, abs\/2303.08774","author":"AI.","year":"2023","unstructured":"OpenAI. Gpt-4 technical report. ArXiv preprint, abs\/2303.08774, 2023. URL https:\/\/arxiv.org\/abs\/2303.08774."},{"key":"e_1_2_1_32_1","first-page":"27730","article-title":"Training language models to follow instructions with human feedback","volume":"35","author":"Ouyang Long","year":"2022","unstructured":"Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. Training language models to follow instructions with human feedback. Advances in Neural Information Processing Systems, 35: 27730--27744, 2022.","journal-title":"Advances in Neural Information Processing Systems"},{"issue":"1","key":"e_1_2_1_33_1","first-page":"12","article-title":"The dilemma of the direct answer","volume":"54","author":"Potthast Martin","year":"2021","unstructured":"Martin Potthast, Matthias Hagen, and Benno Stein. The dilemma of the direct answer. ACM SIGIR Forum, 54:1 -- 12, 2021.","journal-title":"ACM SIGIR Forum"},{"key":"e_1_2_1_34_1","volume-title":"How does generative retrieval scale to millions of passages? ArXiv preprint, abs\/2305.11841","author":"Pradeep Ronak","year":"2023","unstructured":"Ronak Pradeep, Kai Hui, Jai Gupta, Adam D Lelkes, Honglei Zhuang, Jimmy Lin, Donald Metzler, and Vinh Q Tran. How does generative retrieval scale to millions of passages? ArXiv preprint, abs\/2305.11841, 2023. URL https:\/\/arxiv.org\/abs\/2305.11841."},{"key":"e_1_2_1_35_1","first-page":"8748","volume-title":"Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event, volume 139 of Proceedings of Machine Learning Research","author":"Radford Alec","year":"2021","unstructured":"Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever. Learning transferable visual models from natural language supervision. In Marina Meila and Tong Zhang, editors, Proceedings of the 38th International Conference on Machine Learning, ICML 2021, 18-24 July 2021, Virtual Event, volume 139 of Proceedings of Machine Learning Research, pages 8748--8763. PMLR, 2021. URL http:\/\/proceedings.mlr.press\/v139\/radford21a.html."},{"key":"e_1_2_1_36_1","first-page":"140","article-title":"Exploring the limits of transfer learning with a unified text-to-text transformer","volume":"21","author":"Raffel Colin","year":"2020","unstructured":"Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. Exploring the limits of transfer learning with a unified text-to-text transformer. J. Mach. Learn. Res., 21:140:1--140:67, 2020. URL http:\/\/jmlr.org\/papers\/v21\/20-074.html.","journal-title":"J. Mach. Learn. Res."},{"key":"e_1_2_1_37_1","volume-title":"In-context retrieval-augmented language models. ArXiv preprint, abs\/2302.00083","author":"Ram Ori","year":"2023","unstructured":"Ori Ram, Yoav Levine, Itay Dalmedigos, Dor Muhlgay, Amnon Shashua, Kevin Leyton-Brown, and Yoav Shoham. In-context retrieval-augmented language models. ArXiv preprint, abs\/2302.00083, 2023. URL https:\/\/arxiv.org\/abs\/2302.00083."},{"key":"e_1_2_1_38_1","volume-title":"Hierarchical text-conditional image generation with clip latents. ArXiv preprint, abs\/2204.06125","author":"Ramesh Aditya","year":"2022","unstructured":"Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. Hierarchical text-conditional image generation with clip latents. ArXiv preprint, abs\/2204.06125, 2022. URL https:\/\/arxiv.org\/abs\/2204.06125."},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/D19-1410"},{"key":"e_1_2_1_40_1","volume-title":"Fast transformer decoding: One write-head is all you need. ArXiv preprint, abs\/1911.02150","author":"Shazeer Noam","year":"2019","unstructured":"Noam Shazeer. Fast transformer decoding: One write-head is all you need. ArXiv preprint, abs\/1911.02150, 2019. URL https:\/\/arxiv.org\/abs\/1911.02150."},{"key":"e_1_2_1_41_1","volume-title":"Replug: Retrieval-augmented black-box language models. ArXiv preprint, abs\/2301.12652","author":"Shi Weijia","year":"2023","unstructured":"Weijia Shi, Sewon Min, Michihiro Yasunaga, Minjoon Seo, Rich James, Mike Lewis, Luke Zettlemoyer, and Wen-tau Yih. Replug: Retrieval-augmented black-box language models. ArXiv preprint, abs\/2301.12652, 2023. URL https:\/\/arxiv.org\/abs\/2301.12652."},{"key":"e_1_2_1_42_1","volume-title":"Is chatgpt good at search? investigating large language models as re-ranking agent. ArXiv preprint, abs\/2304.09542","author":"Sun Weiwei","year":"2023","unstructured":"Weiwei Sun, Lingyong Yan, Xinyu Ma, Pengjie Ren, Dawei Yin, and Zhaochun Ren. Is chatgpt good at search? investigating large language models as re-ranking agent. ArXiv preprint, abs\/2304.09542, 2023. URL https:\/\/arxiv.org\/abs\/2304.09542."},{"key":"e_1_2_1_43_1","first-page":"21831","article-title":"Transformer memory as a differentiable search index","volume":"35","author":"Tay Yi","year":"2022","unstructured":"Yi Tay, Vinh Tran, Mostafa Dehghani, Jianmo Ni, Dara Bahri, Harsh Mehta, Zhen Qin, Kai Hui, Zhe Zhao, Jai Gupta, et al. Transformer memory as a differentiable search index. Advances in Neural Information Processing Systems, 35:21831--21843, 2022.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_44_1","volume-title":"Beir: A heterogenous benchmark for zero-shot evaluation of information retrieval models. ArXiv preprint, abs\/2104.08663","author":"Thakur Nandan","year":"2021","unstructured":"Nandan Thakur, Nils Reimers, Andreas Ruckl'e, Abhishek Srivastava, and Iryna Gurevych. Beir: A heterogenous benchmark for zero-shot evaluation of information retrieval models. ArXiv preprint, abs\/2104.08663, 2021. URL https:\/\/arxiv.org\/abs\/2104.08663."},{"key":"e_1_2_1_45_1","volume-title":"Llama: Open and efficient foundation language models. ArXiv preprint, abs\/2302.13971","author":"Touvron Hugo","year":"2023","unstructured":"Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timoth\u00e9e Lacroix, Baptiste Rozi\u00e8re, Naman Goyal, Eric Hambro, Faisal Azhar, Aur'elien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. Llama: Open and efficient foundation language models. ArXiv preprint, abs\/2302.13971, 2023. URL https:\/\/arxiv.org\/abs\/2302.13971."},{"key":"e_1_2_1_46_1","volume-title":"Text embeddings by weakly-supervised contrastive pre-training. ArXiv preprint, abs\/2212.03533","author":"Wang Liang","year":"2022","unstructured":"Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, and Furu Wei. Text embeddings by weakly-supervised contrastive pre-training. ArXiv preprint, abs\/2212.03533, 2022. URL https:\/\/arxiv.org\/abs\/2212.03533."},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-long.125"},{"key":"e_1_2_1_48_1","volume-title":"Query2doc: Query expansion with large language models. ArXiv preprint, abs\/2303.07678","author":"Wang Liang","year":"2023","unstructured":"Liang Wang, Nan Yang, and Furu Wei. Query2doc: Query expansion with large language models. ArXiv preprint, abs\/2303.07678, 2023b. URL https:\/\/arxiv.org\/abs\/2303.07678."},{"key":"e_1_2_1_49_1","first-page":"24824","article-title":"Chain-of-thought prompting elicits reasoning in large language models","volume":"35","author":"Wei Jason","year":"2022","unstructured":"Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Fei Xia, Ed Chi, Quoc V Le, Denny Zhou, et al. Chain-of-thought prompting elicits reasoning in large language models. Advances in Neural Information Processing Systems, 35:24824--24837, 2022.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_2_1_50_1","volume-title":"9th International Conference on Learning Representations, ICLR 2021","author":"Xiong Lee","year":"2021","unstructured":"Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk. Approximate nearest neighbor negative contrastive learning for dense text retrieval. In 9th International Conference on Learning Representations, ICLR 2021, Virtual Event, Austria, May 3-7, 2021. OpenReview.net, 2021. URL https:\/\/openreview.net\/forum?id=zeFrfgyZln."},{"key":"e_1_2_1_51_1","volume-title":"Inference with reference: Lossless acceleration of large language models. ArXiv preprint, abs\/2304.04487","author":"Yang Nan","year":"2023","unstructured":"Nan Yang, Tao Ge, Liang Wang, Binxing Jiao, Daxin Jiang, Linjun Yang, Rangan Majumder, and Furu Wei. Inference with reference: Lossless acceleration of large language models. ArXiv preprint, abs\/2304.04487, 2023. URL https:\/\/arxiv.org\/abs\/2304.04487."},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2022.emnlp-main.382"},{"key":"e_1_2_1_53_1","volume-title":"Pose: Efficient context window extension of llms via positional skip-wise training. ArXiv preprint, abs\/2309.10400","author":"Zhu Dawei","year":"2023","unstructured":"Dawei Zhu, Nan Yang, Liang Wang, Yifan Song, Wenhao Wu, Furu Wei, and Sujian Li. Pose: Efficient context window extension of llms via positional skip-wise training. ArXiv preprint, abs\/2309.10400, 2023. URL https:\/\/arxiv.org\/abs\/2309.10400."}],"container-title":["ACM SIGIR Forum"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3642979.3643006","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3642979.3643006","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T16:31:21Z","timestamp":1750264281000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3642979.3643006"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12]]},"references-count":53,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2023,12]]}},"alternative-id":["10.1145\/3642979.3643006"],"URL":"https:\/\/doi.org\/10.1145\/3642979.3643006","relation":{},"ISSN":["0163-5840"],"issn-type":[{"value":"0163-5840","type":"print"}],"subject":[],"published":{"date-parts":[[2023,12]]},"assertion":[{"value":"2024-01-22","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}