{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,10]],"date-time":"2025-12-10T08:57:43Z","timestamp":1765357063876,"version":"3.41.2"},"reference-count":32,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2022,1,17]],"date-time":"2022-01-17T00:00:00Z","timestamp":1642377600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGMETRICS Perform. Eval. Rev."],"published-print":{"date-parts":[[2022,1,17]]},"abstract":"<jats:p>Reinforcement learning (RL) has received widespread attention across multiple communities, but the experiments have focused primarily on large-scale game playing and robotics tasks. In this paper we introduce ORSuite, an open-source library containing environments, algorithms, and instrumentation for operational problems. Our package is designed to motivate researchers in the reinforcement learning community to develop and evaluate algorithms on operational tasks, and to consider the true multi-objective nature of these problems by considering metrics beyond cumulative reward.<\/jats:p>","DOI":"10.1145\/3512798.3512819","type":"journal-article","created":{"date-parts":[[2022,1,20]],"date-time":"2022-01-20T13:13:23Z","timestamp":1642684403000},"page":"57-61","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["ORSuite"],"prefix":"10.1145","volume":"49","author":[{"given":"Christopher","family":"Archer","sequence":"first","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Siddhartha","family":"Banerjee","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mayleen","family":"Cortez","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Carrie","family":"Rucker","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sean R.","family":"Sinclair","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Max","family":"Solberg","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Qiaomin","family":"Xie","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Christina","family":"Lee Yu","sequence":"additional","affiliation":[{"name":"Cornell University"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2022,1,20]]},"reference":[{"doi-asserted-by":"publisher","key":"e_1_2_1_1_1","DOI":"10.3386\/w27102"},{"key":"e_1_2_1_2_1","volume-title":"Reinforcement learning: Theory and algorithms","author":"Agarwal Alekh","year":"2020","unstructured":"Alekh Agarwal, Nan Jiang, Sham M Kakade, and Wen Sun. Reinforcement learning: Theory and algorithms. 2020."},{"key":"e_1_2_1_3_1","volume-title":"Operations Research","author":"Banerjee Siddhartha","year":"2021","unstructured":"Siddhartha Banerjee, Daniel Freund, and Thodoris Lykouris. Pricing and optimization in shared vehicle systems: An approximation framework. Operations Research, 2021."},{"key":"e_1_2_1_4_1","volume-title":"Dynamic assignment control of a closed queueing network under complete resource pooling. arXiv e-prints","author":"Banerjee Siddhartha","year":"1803","unstructured":"Siddhartha Banerjee, Yash Kanoria, and Pengyu Qian. Dynamic assignment control of a closed queueing network under complete resource pooling. arXiv e-prints, pages arXiv--1803, 2018."},{"key":"e_1_2_1_5_1","volume-title":"An optimal on-line algorithm for metrical task system. Journal of the ACM (JACM), 39(4):745--763","author":"Borodin Allan","year":"1992","unstructured":"Allan Borodin, Nathan Linial, and Michael E Saks. An optimal on-line algorithm for metrical task system. Journal of the ACM (JACM), 39(4):745--763, 1992."},{"key":"e_1_2_1_6_1","volume-title":"Operations Research","author":"Braverman Anton","year":"2019","unstructured":"Anton Braverman, Jim G Dai, Xin Liu, and Lei Ying. Empty-car routing in ridesharing systems. Operations Research, 2019."},{"key":"e_1_2_1_7_1","volume-title":"Openai gym","author":"Brockman Greg","year":"2016","unstructured":"Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. Openai gym, 2016."},{"key":"e_1_2_1_8_1","volume-title":"Ambulance location and relocation models. European journal of operational research, 147(3):451--463","author":"Brotcorne Luce","year":"2003","unstructured":"Luce Brotcorne, Gilbert Laporte, and Frederic Semet. Ambulance location and relocation models. European journal of operational research, 147(3):451--463, 2003."},{"key":"e_1_2_1_9_1","volume-title":"https:\/\/www.foodbankst.org\/","author":"Southern Food Bank","year":"2020","unstructured":"Food Bank of the Southern Tier of New York. https:\/\/www.foodbankst.org\/, 2020."},{"key":"e_1_2_1_10_1","volume-title":"Stable baselines. https:\/\/github.com\/hill-a\/stable-baselines","author":"Hill Ashley","year":"2018","unstructured":"Ashley Hill, Antonin Raffin, Maximilian Ernestus, Adam Gleave, Anssi Kanervisto, Rene Traore, Prafulla Dhariwal, Christopher Hesse, Oleg Klimov, Alex Nichol, Matthias Plappert, Alec Radford, John Schulman, Szymon Sidor, and Yuhuai Wu. Stable baselines. https:\/\/github.com\/hill-a\/stable-baselines, 2018."},{"key":"e_1_2_1_11_1","volume-title":"Or-gym: A reinforcement learning library for operations research problem. arXiv preprint arXiv:2008.06319","author":"Hubbs Christian D","year":"2020","unstructured":"Christian D Hubbs, Hector D Perez, Owais Sarwar, Nikolaos V Sahinidis, Ignacio E Grossmann, and John M Wassick. Or-gym: A reinforcement learning library for operations research problem. arXiv preprint arXiv:2008.06319, 2020."},{"key":"e_1_2_1_12_1","first-page":"772","article-title":"A contribution to the mathematical theory of epidemics","volume":"115","author":"Kermack William Ogilvy","year":"1927","unstructured":"William Ogilvy Kermack, A. G. McKendrick, and Gilbert Thomas Walker. A contribution to the mathematical theory of epidemics. Proceedings of the Royal Society of London, 115:772, 1927. Series A, Containing Papers of a Mathematical and Physical Character.","journal-title":"Proceedings of the Royal Society of London"},{"doi-asserted-by":"publisher","key":"e_1_2_1_13_1","DOI":"10.3386\/w27532"},{"key":"e_1_2_1_14_1","volume-title":"Advances in Neural Information Processing Systems","volume":"32","author":"Mao Hongzi","year":"2019","unstructured":"Hongzi Mao, Parimarjan Negi, Akshay Narayan, Hanrui Wang, Jiacheng Yang, Haonan Wang, Ryan Marcus, ravichandra addanki, Mehrdad Khani Shirkoohi, Songtao He, Vikram Nathan, Frank Cangialosi, Shaileshh Venkatakrishnan, Wei-Hung Weng, Song Han, Tim Kraska, and Dr.Mohammad Alizadeh. Park: An open platform for learning-augmented computer systems. In H. Wallach, H. Larochelle, A. Beygelzimer, F. d'Alch\u00b4e-Buc, E. Fox, and R. Garnett, editors, Advances in Neural Information Processing Systems, volume 32. Curran Associates, Inc., 2019."},{"doi-asserted-by":"publisher","key":"e_1_2_1_15_1","DOI":"10.1287\/ijoc.1090.0345"},{"key":"e_1_2_1_16_1","volume-title":"New york state's covid-19 vaccination program","author":"New York State Department of Health.","year":"2020","unstructured":"New York State Department of Health. New york state's covid-19 vaccination program. October 2020."},{"key":"e_1_2_1_17_1","volume-title":"Reinforcement learning and stochastic optimization","author":"Powell W","year":"2019","unstructured":"W Powell. Reinforcement learning and stochastic optimization, 2019."},{"key":"e_1_2_1_18_1","volume-title":"Ecole: A library for learning inside milp solvers. arXiv preprint arXiv:2104.02828","author":"Prouvost Antoine","year":"2021","unstructured":"Antoine Prouvost, Justin Dumouchelle, Maxime Gasse, Didier Ch\u00b4etelat, and Andrea Lodi. Ecole: A library for learning inside milp solvers. arXiv preprint arXiv:2104.02828, 2021."},{"doi-asserted-by":"publisher","key":"e_1_2_1_19_1","DOI":"10.1002\/9780470316887"},{"key":"e_1_2_1_20_1","volume-title":"https:\/\/github.com\/uoa-ems-research\/JEMSS.jl","author":"Ridler Samuel","year":"2021","unstructured":"Samuel Ridler. Jemss. https:\/\/github.com\/uoa-ems-research\/JEMSS.jl, 2021."},{"key":"e_1_2_1_21_1","volume-title":"Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al. Mastering the game of go with deep neural networks and tree search. nature, 529(7587):484","author":"Silver David","year":"2016","unstructured":"David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al. Mastering the game of go with deep neural networks and tree search. nature, 529(7587):484, 2016."},{"key":"e_1_2_1_22_1","volume-title":"https:\/\/github.com\/cornell-orie\/ORSuite","author":"Sinclair Sean","year":"2021","unstructured":"Sean Sinclair, Christopher Archer, Carrie Rucker, Max Solberg, Mayleen Cortez, Shashank Pathak, Siddhartha Banerjee, and Christina Yu. Orsuite. https:\/\/github.com\/cornell-orie\/ORSuite, 2021."},{"key":"e_1_2_1_23_1","first-page":"33","article-title":"Adaptive discretization for model-based reinforcement learning","author":"Sinclair Sean","year":"2020","unstructured":"Sean Sinclair, Tianyu Wang, Gauri Jain, Siddhartha Banerjee, and Christina Yu. Adaptive discretization for model-based reinforcement learning. Advances in Neural Information Processing Systems, 33, 2020.","journal-title":"Advances in Neural Information Processing Systems"},{"doi-asserted-by":"publisher","key":"e_1_2_1_24_1","DOI":"10.1145\/3366703"},{"key":"e_1_2_1_25_1","volume-title":"Sequential fair allocation: Achieving the optimal envy-efficiency tradeoff curve","author":"Sinclair Sean R.","year":"2021","unstructured":"Sean R. Sinclair, Siddhartha Banerjee, and Christina Lee Yu. Sequential fair allocation: Achieving the optimal envy-efficiency tradeoff curve, 2021."},{"volume-title":"Sequential fair allocation of limited resources under stochastic demands. arXiv preprint 60 Performance Evaluation Review","author":"Sinclair Sean R","unstructured":"Sean R Sinclair, Gauri Jain, Siddhartha Banerjee, and Christina Lee Yu. Sequential fair allocation of limited resources under stochastic demands. arXiv preprint 60 Performance Evaluation Review, Vol. 49, No. 2, September 2021 arXiv:2011.14382, 2020.","key":"e_1_2_1_26_1"},{"doi-asserted-by":"publisher","key":"e_1_2_1_27_1","DOI":"10.1561\/9781680836219"},{"key":"e_1_2_1_28_1","volume-title":"Efficient model-free reinforcement learning in metric spaces. arXiv preprint arXiv:1905.00475","author":"Song Zhao","year":"2019","unstructured":"Zhao Song and Wen Sun. Efficient model-free reinforcement learning in metric spaces. arXiv preprint arXiv:1905.00475, 2019."},{"doi-asserted-by":"publisher","key":"e_1_2_1_29_1","DOI":"10.2307\/2215223"},{"key":"e_1_2_1_30_1","volume-title":"Reinforcement learning: An introduction","author":"Sutton Richard S","year":"2018","unstructured":"Richard S Sutton and Andrew G Barto. Reinforcement learning: An introduction. MIT press, 2018."},{"key":"e_1_2_1_31_1","volume-title":"September","author":"Varian Hal R.","year":"1974","unstructured":"Hal R. Varian. Equity, envy, and efficiency. Journal of Economic Theory, 9(1):63--91, September 1974."},{"doi-asserted-by":"publisher","key":"e_1_2_1_32_1","DOI":"10.1016\/0047-2727(76)90018-9"}],"container-title":["ACM SIGMETRICS Performance Evaluation Review"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3512798.3512819","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3512798.3512819","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,14]],"date-time":"2025-07-14T22:26:48Z","timestamp":1752532008000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3512798.3512819"}},"subtitle":["Benchmarking Suite for Sequential Operations Models"],"short-title":[],"issued":{"date-parts":[[2022,1,17]]},"references-count":32,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2022,1,17]]}},"alternative-id":["10.1145\/3512798.3512819"],"URL":"https:\/\/doi.org\/10.1145\/3512798.3512819","relation":{},"ISSN":["0163-5999"],"issn-type":[{"type":"print","value":"0163-5999"}],"subject":[],"published":{"date-parts":[[2022,1,17]]},"assertion":[{"value":"2022-01-20","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}