{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,22]],"date-time":"2025-08-22T04:58:51Z","timestamp":1755838731656,"version":"3.41.0"},"publisher-location":"New York, NY, USA","reference-count":17,"publisher":"ACM","license":[{"start":{"date-parts":[[2018,9,10]],"date-time":"2018-09-10T00:00:00Z","timestamp":1536537600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Key R&D Program of China","award":["2016QY02D0405"],"award-info":[{"award-number":["2016QY02D0405"]}]},{"name":"973 Program of China","award":["2014CB340401"],"award-info":[{"award-number":["2014CB340401"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"publisher","award":["61773362, 61425016, 61472401, 61722211, 20180290"],"award-info":[{"award-number":["61773362, 61425016, 61472401, 61722211, 20180290"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"publisher"}]},{"name":"Youth Innovation Promotion Association CAS","award":["20144310, 2016102"],"award-info":[{"award-number":["20144310, 2016102"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2018,9,10]]},"DOI":"10.1145\/3234944.3234977","type":"proceedings-article","created":{"date-parts":[[2018,9,13]],"date-time":"2018-09-13T12:54:52Z","timestamp":1536843292000},"page":"175-178","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":20,"title":["Multi Page Search with Reinforcement Learning to Rank"],"prefix":"10.1145","author":[{"given":"Wei","family":"Zeng","sequence":"first","affiliation":[{"name":"University of Chinese Academy of Sciences &amp; Chinese Academy of Sciences, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jun","family":"Xu","sequence":"additional","affiliation":[{"name":"University of Chinese Academy of Sciences &amp; Chinese Academy of Sciences, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Yanyan","family":"Lan","sequence":"additional","affiliation":[{"name":"University of Chinese Academy of Sciences &amp; Chinese Academy of Sciences, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Jiafeng","family":"Guo","sequence":"additional","affiliation":[{"name":"University of Chinese Academy of Sciences &amp; Chinese Academy of Sciences, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xueqi","family":"Cheng","sequence":"additional","affiliation":[{"name":"University of Chinese Academy of Sciences &amp; Chinese Academy of Sciences, Beijing, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2018,9,10]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1102351.1102363"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1273496.1273513"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1390334.1390446"},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.5555\/945365.964285"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"crossref","unstructured":"Katja Hofmann Shimon Whiteson and Maarten de Rijke . 2011. Balancing exploration and exploitation in learning to rank online European Conference on Information Retrieval. 251--263.   Katja Hofmann Shimon Whiteson and Maarten de Rijke . 2011. Balancing exploration and exploitation in learning to rank online European Conference on Information Retrieval. 251--263.","DOI":"10.1007\/978-3-642-20161-5_25"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1145\/2488388.2488446"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/1229179.1229181"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1561\/1500000016"},{"key":"e_1_3_2_1_9_1","volume-title":"Letor: Benchmark dataset for research on learning to rank for information retrieval Proceedings of SIGIR 2007 workshop","author":"Liu Tie-Yan","year":"2007","unstructured":"Tie-Yan Liu , Jun Xu , Tao Qin , Wenying Xiong , and Hang Li . 2007 . Letor: Benchmark dataset for research on learning to rank for information retrieval Proceedings of SIGIR 2007 workshop , Vol. Vol. 310 . Tie-Yan Liu, Jun Xu, Tao Qin, Wenying Xiong, and Hang Li . 2007. Letor: Benchmark dataset for research on learning to rank for information retrieval Proceedings of SIGIR 2007 workshop, Vol. Vol. 310."},{"key":"e_1_3_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1645953.1645988"},{"key":"e_1_3_2_1_11_1","volume-title":"Relevance feedback in information retrieval. THE SMART RETRIEVAL SYSTEM: Experiments in Automatic Document Processing","author":"Rocchio Joseph John","year":"1971","unstructured":"Joseph John Rocchio . 1971. Relevance feedback in information retrieval. THE SMART RETRIEVAL SYSTEM: Experiments in Automatic Document Processing ( 1971 ), 313--323. Joseph John Rocchio . 1971. Relevance feedback in information retrieval. THE SMART RETRIEVAL SYSTEM: Experiments in Automatic Document Processing (1971), 313--323."},{"volume-title":"Improving retrieval performance by relevance feedback","author":"Salton Gerard","key":"e_1_3_2_1_12_1","unstructured":"Gerard Salton and Chris Buckley . 1990. Improving retrieval performance by relevance feedback . Vol. Vol. 41 . Wiley Online Library , 288--297. Gerard Salton and Chris Buckley . 1990. Improving retrieval performance by relevance feedback. Vol. Vol. 41. Wiley Online Library, 288--297."},{"volume-title":"Reinforcement learning: An introduction","author":"Sutton Richard S","key":"e_1_3_2_1_13_1","unstructured":"Richard S Sutton and Andrew G Barto . 1998. Reinforcement learning: An introduction . Vol. Vol. 1 . MIT press Cambridge . Richard S Sutton and Andrew G Barto . 1998. Reinforcement learning: An introduction. Vol. Vol. 1. MIT press Cambridge."},{"key":"e_1_3_2_1_14_1","unstructured":"Richard S Sutton David A McAllester Satinder P Singh and Yishay Mansour . 2000. Policy gradient methods for reinforcement learning with function approximation Advances in neural information processing systems. 1057--1063.   Richard S Sutton David A McAllester Satinder P Singh and Yishay Mansour . 2000. Policy gradient methods for reinforcement learning with function approximation Advances in neural information processing systems. 1057--1063."},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/3077136.3080685"},{"key":"e_1_3_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/3077136.3080775"},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2747874"}],"event":{"name":"ICTIR '18: The 2018 ACM SIGIR International Conference on the Theory of Information Retrieval","sponsor":["SIGIR ACM Special Interest Group on Information Retrieval"],"location":"Tianjin China","acronym":"ICTIR '18"},"container-title":["Proceedings of the 2018 ACM SIGIR International Conference on Theory of Information Retrieval"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3234944.3234977","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3234944.3234977","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T02:08:16Z","timestamp":1750212496000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3234944.3234977"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,9,10]]},"references-count":17,"alternative-id":["10.1145\/3234944.3234977","10.1145\/3234944"],"URL":"https:\/\/doi.org\/10.1145\/3234944.3234977","relation":{},"subject":[],"published":{"date-parts":[[2018,9,10]]},"assertion":[{"value":"2018-09-10","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}