{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,29]],"date-time":"2026-04-29T02:13:41Z","timestamp":1777428821813,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":31,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,8,14]],"date-time":"2021-08-14T00:00:00Z","timestamp":1628899200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Total Graduate Fellowship, ICME, Stanford University"},{"name":"Office of Naval Research Grant","award":["N00014-19-1-2468"],"award-info":[{"award-number":["N00014-19-1-2468"]}]},{"name":"Stanford Institute for Human-Centered Artificial Intelligence"},{"name":"PayPal Graduate Fellowship, ICME, Stanford University"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,8,14]]},"DOI":"10.1145\/3447548.3467456","type":"proceedings-article","created":{"date-parts":[[2021,8,12]],"date-time":"2021-08-12T06:12:08Z","timestamp":1628748728000},"page":"2125-2135","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":16,"title":["Off-Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits"],"prefix":"10.1145","author":[{"given":"Ruohan","family":"Zhan","sequence":"first","affiliation":[{"name":"Stanford University, Stanford, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Vitor","family":"Hadad","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"David A.","family":"Hirshberg","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Susan","family":"Athey","sequence":"additional","affiliation":[{"name":"Stanford University, Stanford, CA, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,8,14]]},"reference":[{"key":"e_1_3_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/3097983.3098155"},{"key":"e_1_3_2_2_2_1","volume-title":"International Conference on Machine Learning. PMLR, 127--135","author":"Agrawal Shipra","year":"2013","unstructured":"Shipra Agrawal and Navin Goyal . 2013 . Thompson sampling for contextual bandits with linear payoffs . In International Conference on Machine Learning. PMLR, 127--135 . Shipra Agrawal and Navin Goyal. 2013. Thompson sampling for contextual bandits with linear payoffs. In International Conference on Machine Learning. PMLR, 127--135."},{"key":"e_1_3_2_2_3_1","volume-title":"Peter J Bickel, Ya'acov Ritov, J Klaassen, Jon A Wellner, and YA'Acov Ritov.","author":"Bickel Peter J","year":"1993","unstructured":"Peter J Bickel , Chris AJ Klaassen , Peter J Bickel, Ya'acov Ritov, J Klaassen, Jon A Wellner, and YA'Acov Ritov. 1993 . Efficient and adaptive estimation for semiparametric models. Vol. 4 . Johns Hopkins University Press Baltimore . Peter J Bickel, Chris AJ Klaassen, Peter J Bickel, Ya'acov Ritov, J Klaassen, Jon A Wellner, and YA'Acov Ritov. 1993. Efficient and adaptive estimation for semiparametric models. Vol. 4. Johns Hopkins University Press Baltimore."},{"key":"e_1_3_2_2_4_1","article-title":"Counterfactual reasoning and learning systems: The example of computational advertising","volume":"14","author":"Charles Denis","year":"2013","unstructured":"Denis Charles , Max Chickering , and Patrice Simard . 2013 . Counterfactual reasoning and learning systems: The example of computational advertising . Journal of Machine Learning Research , Vol. 14 (2013). Denis Charles, Max Chickering, and Patrice Simard. 2013. Counterfactual reasoning and learning systems: The example of computational advertising. Journal of Machine Learning Research, Vol. 14 (2013).","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_2_2_5_1","volume-title":"arXiv preprint arXiv:1608.00060","author":"Chernozhukov Victor","year":"2016","unstructured":"Victor Chernozhukov , Denis Chetverikov , Mert Demirer , Esther Duflo , Christian Hansen , Whitney Newey , and James Robins . 2016. Double\/debiased machine learning for treatment and causal parameters. arXiv preprint arXiv:1608.00060 ( 2016 ). Victor Chernozhukov, Denis Chetverikov, Mert Demirer, Esther Duflo, Christian Hansen, Whitney Newey, and James Robins. 2016. Double\/debiased machine learning for treatment and causal parameters. arXiv preprint arXiv:1608.00060 (2016)."},{"key":"e_1_3_2_2_6_1","volume-title":"Estimation considerations in contextual bandits. arXiv preprint arXiv:1711.07077","author":"Dimakopoulou Maria","year":"2017","unstructured":"Maria Dimakopoulou , Zhengyuan Zhou , Susan Athey , and Guido Imbens . 2017. Estimation considerations in contextual bandits. arXiv preprint arXiv:1711.07077 ( 2017 ). Maria Dimakopoulou, Zhengyuan Zhou, Susan Athey, and Guido Imbens. 2017. Estimation considerations in contextual bandits. arXiv preprint arXiv:1711.07077 (2017)."},{"key":"e_1_3_2_2_7_1","volume-title":"Doubly robust policy evaluation and learning. arXiv preprint arXiv:1103.4601","author":"Dud\u00edk Miroslav","year":"2011","unstructured":"Miroslav Dud\u00edk , John Langford , and Lihong Li. 2011. Doubly robust policy evaluation and learning. arXiv preprint arXiv:1103.4601 ( 2011 ). Miroslav Dud\u00edk, John Langford, and Lihong Li. 2011. Doubly robust policy evaluation and learning. arXiv preprint arXiv:1103.4601 (2011)."},{"key":"e_1_3_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1073\/pnas.2014602118"},{"key":"e_1_3_2_2_9_1","volume-title":"Martingale limit theory and its application","author":"Hall Peter","unstructured":"Peter Hall and Christopher C Heyde . 2014. Martingale limit theory and its application . Academic press . Peter Hall and Christopher C Heyde. 2014. Martingale limit theory and its application. Academic press."},{"key":"e_1_3_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1080\/01621459.1952.10483446"},{"key":"e_1_3_2_2_11_1","volume-title":"nonparametric, non-asymptotic confidence sequences. arXiv preprint arXiv:1810.08240","author":"Howard Steven R","year":"2018","unstructured":"Steven R Howard , Aaditya Ramdas , Jon McAuliffe , and Jasjeet Sekhon . 2018. Uniform , nonparametric, non-asymptotic confidence sequences. arXiv preprint arXiv:1810.08240 ( 2018 ). Steven R Howard, Aaditya Ramdas, Jon McAuliffe, and Jasjeet Sekhon. 2018. Uniform, nonparametric, non-asymptotic confidence sequences. arXiv preprint arXiv:1810.08240 (2018)."},{"key":"e_1_3_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1162\/003465304323023651"},{"key":"e_1_3_2_2_13_1","volume-title":"Causal inference in statistics, social, and biomedical sciences","author":"Imbens Guido W","unstructured":"Guido W Imbens and Donald B Rubin . 2015. Causal inference in statistics, social, and biomedical sciences . Cambridge University Press . Guido W Imbens and Donald B Rubin. 2015. Causal inference in statistics, social, and biomedical sciences. Cambridge University Press."},{"key":"e_1_3_2_2_14_1","volume-title":"Tractable contextual bandits beyond realizability. arXiv preprint arXiv:2010.13013","author":"Krishnamurthy Sanath Kumar","year":"2020","unstructured":"Sanath Kumar Krishnamurthy , Vitor Hadad , and Susan Athey . 2020. Tractable contextual bandits beyond realizability. arXiv preprint arXiv:2010.13013 ( 2020 ). Sanath Kumar Krishnamurthy, Vitor Hadad, and Susan Athey. 2020. Tractable contextual bandits beyond realizability. arXiv preprint arXiv:2010.13013 (2020)."},{"key":"e_1_3_2_2_15_1","unstructured":"V. Laan and J. Mark. 2008. The Construction and Analysis of Adaptive Group Sequential Designs. In U.C. Berkeley Division of Biostatistics Working Paper Series. Working Paper 232.https:\/\/biostats.bepress.com\/ucbbiostat\/paper232.  V. Laan and J. Mark. 2008. The Construction and Analysis of Adaptive Group Sequential Designs. In U.C. Berkeley Division of Biostatistics Working Paper Series. Working Paper 232.https:\/\/biostats.bepress.com\/ucbbiostat\/paper232."},{"key":"e_1_3_2_2_16_1","volume-title":"Bandit algorithms. preprint","author":"Lattimore Tor","year":"2018","unstructured":"Tor Lattimore and Csaba Szepesv\u00e1ri . 2018. Bandit algorithms. preprint ( 2018 ), 28. Tor Lattimore and Csaba Szepesv\u00e1ri. 2018. Bandit algorithms. preprint (2018), 28."},{"key":"e_1_3_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/1935826.1935878"},{"key":"e_1_3_2_2_18_1","doi-asserted-by":"publisher","DOI":"10.1214\/15-AOS1384"},{"key":"e_1_3_2_2_19_1","doi-asserted-by":"publisher","DOI":"10.1111\/1467-9868.00389"},{"key":"e_1_3_2_2_20_1","volume-title":"Why adaptively collected data have negative bias and how to correct for it. arXiv preprint arXiv:1708.01977","author":"Nie Xinkun","year":"2017","unstructured":"Xinkun Nie , Xiaoying Tian , Jonathan Taylor , and James Zou . 2017. Why adaptively collected data have negative bias and how to correct for it. arXiv preprint arXiv:1708.01977 ( 2017 ). Xinkun Nie, Xiaoying Tian, Jonathan Taylor, and James Zou. 2017. Why adaptively collected data have negative bias and how to correct for it. arXiv preprint arXiv:1708.01977 (2017)."},{"key":"e_1_3_2_2_21_1","volume-title":"Abbas Kazerouni, Ian Osband, and Zheng Wen.","author":"Russo Daniel","year":"2017","unstructured":"Daniel Russo , Benjamin Van Roy , Abbas Kazerouni, Ian Osband, and Zheng Wen. 2017 . A tutorial on thompson sampling. arXiv preprint arXiv:1707.02038 (2017). Daniel Russo, Benjamin Van Roy, Abbas Kazerouni, Ian Osband, and Zheng Wen. 2017. A tutorial on thompson sampling. arXiv preprint arXiv:1707.02038 (2017)."},{"key":"e_1_3_2_2_22_1","unstructured":"Jaehyeok Shin Aaditya Ramdas and Alessandro Rinaldo. 2019. Are sample means in multi-armed bandits positively or negatively biased?. In Advances in Neural Information Processing Systems. 7100--7109.  Jaehyeok Shin Aaditya Ramdas and Alessandro Rinaldo. 2019. Are sample means in multi-armed bandits positively or negatively biased?. In Advances in Neural Information Processing Systems. 7100--7109."},{"key":"e_1_3_2_2_23_1","volume-title":"On conditional versus marginal bias in multi-armed bandits. arXiv preprint arXiv:2002.08422","author":"Shin Jaehyeok","year":"2020","unstructured":"Jaehyeok Shin , Aaditya Ramdas , and Alessandro Rinaldo . 2020. On conditional versus marginal bias in multi-armed bandits. arXiv preprint arXiv:2002.08422 ( 2020 ). Jaehyeok Shin, Aaditya Ramdas, and Alessandro Rinaldo. 2020. On conditional versus marginal bias in multi-armed bandits. arXiv preprint arXiv:2002.08422 (2020)."},{"key":"e_1_3_2_2_24_1","volume-title":"International Conference on Machine Learning. PMLR, 9167--9176","author":"Su Yi","year":"2020","unstructured":"Yi Su , Maria Dimakopoulou , Akshay Krishnamurthy , and Miroslav Dud\u00edk . 2020 . Doubly robust off-policy evaluation with shrinkage . In International Conference on Machine Learning. PMLR, 9167--9176 . Yi Su, Maria Dimakopoulou, Akshay Krishnamurthy, and Miroslav Dud\u00edk. 2020. Doubly robust off-policy evaluation with shrinkage. In International Conference on Machine Learning. PMLR, 9167--9176."},{"key":"e_1_3_2_2_25_1","volume-title":"International Conference on Machine Learning. PMLR, 6005--6014","author":"Su Yi","year":"2019","unstructured":"Yi Su , Lequn Wang , Michele Santacatterina , and Thorsten Joachims . 2019 . Cab: Continuous adaptive blending for policy evaluation and learning . In International Conference on Machine Learning. PMLR, 6005--6014 . Yi Su, Lequn Wang, Michele Santacatterina, and Thorsten Joachims. 2019. Cab: Continuous adaptive blending for policy evaluation and learning. In International Conference on Machine Learning. PMLR, 6005--6014."},{"key":"e_1_3_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.21105\/joss.02232"},{"key":"e_1_3_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1093\/biomet\/25.3-4.285"},{"key":"e_1_3_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/2641190.2641198"},{"key":"e_1_3_2_2_29_1","volume-title":"Multi-armed bandit models for the optimal design of clinical trials: benefits and challenges. Statistical science: a review journal of the Institute of Mathematical Statistics","author":"Villar Sof\u00eda S","year":"2015","unstructured":"Sof\u00eda S Villar , Jack Bowden , and James Wason . 2015. Multi-armed bandit models for the optimal design of clinical trials: benefits and challenges. Statistical science: a review journal of the Institute of Mathematical Statistics , Vol. 30 , 2 ( 2015 ), 199. Sof\u00eda S Villar, Jack Bowden, and James Wason. 2015. Multi-armed bandit models for the optimal design of clinical trials: benefits and challenges. Statistical science: a review journal of the Institute of Mathematical Statistics, Vol. 30, 2 (2015), 199."},{"key":"e_1_3_2_2_30_1","volume-title":"International Conference on Machine Learning. PMLR, 3589--3597","author":"Wang Yu-Xiang","year":"2017","unstructured":"Yu-Xiang Wang , Alekh Agarwal , and Miroslav Dudik . 2017 . Optimal and adaptive off-policy evaluation in contextual bandits . In International Conference on Machine Learning. PMLR, 3589--3597 . Yu-Xiang Wang, Alekh Agarwal, and Miroslav Dudik. 2017. Optimal and adaptive off-policy evaluation in contextual bandits. In International Conference on Machine Learning. PMLR, 3589--3597."},{"key":"e_1_3_2_2_31_1","volume-title":"Inference for Batched Bandits. arXiv preprint arXiv:2002.03217","author":"Zhang Kelly W","year":"2020","unstructured":"Kelly W Zhang , Lucas Janson , and Susan A Murphy . 2020. Inference for Batched Bandits. arXiv preprint arXiv:2002.03217 ( 2020 ). Kelly W Zhang, Lucas Janson, and Susan A Murphy. 2020. Inference for Batched Bandits. arXiv preprint arXiv:2002.03217 (2020)."}],"event":{"name":"KDD '21: The 27th ACM SIGKDD Conference on Knowledge Discovery and Data Mining","location":"Virtual Event Singapore","acronym":"KDD '21","sponsor":["SIGMOD ACM Special Interest Group on Management of Data","SIGKDD ACM Special Interest Group on Knowledge Discovery in Data"]},"container-title":["Proceedings of the 27th ACM SIGKDD Conference on Knowledge Discovery &amp; Data Mining"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3447548.3467456","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3447548.3467456","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:18:37Z","timestamp":1750191517000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3447548.3467456"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,8,14]]},"references-count":31,"alternative-id":["10.1145\/3447548.3467456","10.1145\/3447548"],"URL":"https:\/\/doi.org\/10.1145\/3447548.3467456","relation":{},"subject":[],"published":{"date-parts":[[2021,8,14]]},"assertion":[{"value":"2021-08-14","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}