{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,2]],"date-time":"2025-10-02T00:47:34Z","timestamp":1759366054092,"version":"build-2065373602"},"publisher-location":"New York, NY, USA","reference-count":51,"publisher":"ACM","funder":[{"DOI":"10.13039\/100000001","name":"National Science Foundation","doi-asserted-by":"publisher","award":["2211882"],"award-info":[{"award-number":["2211882"]}],"id":[{"id":"10.13039\/100000001","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100006751","name":"U.S. Army","doi-asserted-by":"publisher","award":["W911NF-20- D-0002"],"award-info":[{"award-number":["W911NF-20- D-0002"]}],"id":[{"id":"10.13039\/100006751","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2025,10,13]]},"DOI":"10.1145\/3731569.3764846","type":"proceedings-article","created":{"date-parts":[[2025,10,1]],"date-time":"2025-10-01T12:43:24Z","timestamp":1759322604000},"page":"322-340","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["COpter: Efficient Large-Scale Resource-Allocation via Continual Optimization"],"prefix":"10.1145","author":[{"ORCID":"https:\/\/orcid.org\/0009-0001-8013-461X","authenticated-orcid":false,"given":"Suhas Jayaram","family":"Subramanya","sequence":"first","affiliation":[{"name":"Microsoft, Seattle, Washington, USA"},{"name":"Computer Science Department, Carnegie Mellon University, Pittsburgh, Pennsylvania, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0007-2941-592X","authenticated-orcid":false,"given":"Don Kurian","family":"Dennis","sequence":"additional","affiliation":[{"name":"Meta Platforms, Inc., Menlo Park, California, USA"},{"name":"Machine Learning Department, Carnegie Mellon University, Pittsburgh, Pennsylvania, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6404-2395","authenticated-orcid":false,"given":"Virginia","family":"Smith","sequence":"additional","affiliation":[{"name":"Machine Learning Department, Carnegie Mellon University, Pittsburgh, Pennsylvania, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3065-7316","authenticated-orcid":false,"given":"Gregory R.","family":"Ganger","sequence":"additional","affiliation":[{"name":"Electrical and Computer Engineering Department, Carnegie Mellon University, Pittsburgh, Pennsylvania, USA"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,10,12]]},"reference":[{"key":"e_1_3_2_1_1_1","volume-title":"18th USENIX Symposium on Networked Systems Design and Implementation (NSDI 21)","author":"Abuzaid Firas","year":"2021","unstructured":"Firas Abuzaid, Srikanth Kandula, Behnaz Arzani, Ishai Menache, Matei Zaharia, and Peter Bailis. 2021. Contracting wide-area network topologies to solve flow problems quickly. In 18th USENIX Symposium on Networked Systems Design and Implementation (NSDI 21). 175\u2013200."},{"key":"e_1_3_2_1_2_1","first-page":"20243","article-title":"Practical large-scale linear programming using primal-dual hybrid gradient","volume":"34","author":"Applegate David","year":"2021","unstructured":"David Applegate, Mateo D\u00edaz, Oliver Hinder, Haihao Lu, Miles Lubin, Brendan O'Donoghue, and Warren Schudy. 2021. Practical large-scale linear programming using primal-dual hybrid gradient. Advances in Neural Information Processing Systems 34 (2021), 20243\u201320257.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_3_1","volume-title":"International Conference on Artificial Intelligence and Statistics. PMLR","author":"Baby Dheeraj","year":"2022","unstructured":"Dheeraj Baby and Yu-Xiang Wang. 2022. Optimal dynamic regret in proper online learning with strongly convex losses and beyond. In International Conference on Artificial Intelligence and Statistics. PMLR, 1805\u20131845."},{"key":"e_1_3_2_1_4_1","doi-asserted-by":"crossref","unstructured":"Stephen Boyd Neal Parikh Eric Chu Borja Peleato Jonathan Eckstein et al. 2011. Distributed optimization and statistical learning via the alternating direction method of multipliers. Foundations and Trends\u00ae in Machine learning 3 1 (2011) 1\u2013122.","DOI":"10.1561\/2200000016"},{"key":"e_1_3_2_1_5_1","volume-title":"Projected gradient methods for linearly constrained problems. Mathematical programming 39, 1","author":"Calamai Paul H","year":"1987","unstructured":"Paul H Calamai and Jorge J Mor\u00e9. 1987. Projected gradient methods for linearly constrained problems. Mathematical programming 39, 1 (1987), 93\u2013116."},{"key":"e_1_3_2_1_6_1","volume-title":"18th USENIX Symposium on Operating Systems Design and Implementation (OSDI 24)","author":"Choudhury Arnab","year":"2024","unstructured":"Arnab Choudhury, Yang Wang, Tuomas Pelkonen, Kutta Srinivasan, Abha Jain, Shenghao Lin, Delia David, Siavash Soleimanifard, Michael Chen, Abhishek Yadav, et al. 2024. MAST: Global scheduling of ML training across Geo-Distributed datacenters at hyperscale. In 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI 24). 563\u2013580."},{"key":"e_1_3_2_1_7_1","volume-title":"Proceedings of the 2011 ACM SIGMOD International Conference on Management of data. 313\u2013324","author":"Curino Carlo","year":"2011","unstructured":"Carlo Curino, Evan PC Jones, Samuel Madden, and Hari Balakrishnan. 2011. Workload-aware database monitoring and consolidation. In Proceedings of the 2011 ACM SIGMOD International Conference on Management of data. 313\u2013324."},{"key":"e_1_3_2_1_8_1","first-page":"25","article-title":"Global linear convergence of an augmented Lagrangian algorithm for solving convex quadratic optimization problems","volume":"12","author":"Delbos Fr\u00e9d\u00e9ric","year":"2005","unstructured":"Fr\u00e9d\u00e9ric Delbos and Jean Charles Gilbert. 2005. Global linear convergence of an augmented Lagrangian algorithm for solving convex quadratic optimization problems. Journal of Convex Analysis 12, 1 (2005), 25.","journal-title":"Journal of Convex Analysis"},{"key":"e_1_3_2_1_9_1","first-page":"1","article-title":"CVXPY: A Python-embedded modeling language for convex optimization","volume":"17","author":"Diamond Steven","year":"2016","unstructured":"Steven Diamond and Stephen Boyd. 2016. CVXPY: A Python-embedded modeling language for convex optimization. Journal of Machine Learning Research 17, 83 (2016), 1\u20135.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_3_2_1_10_1","first-page":"619","article-title":"Understanding the convergence of the alternating direction method of multipliers: Theoretical and computational perspectives","volume":"11","author":"Eckstein Jonathan","year":"2015","unstructured":"Jonathan Eckstein and Wang Yao. 2015. Understanding the convergence of the alternating direction method of multipliers: Theoretical and computational perspectives. Pac. J. Optim. 11, 4 (2015), 619\u2013644.","journal-title":"Pac. J. Optim."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"crossref","first-page":"505","DOI":"10.1137\/S0895480199355754","article-title":"Approximating fractional multicommodity flow independent of the number of commodities","volume":"13","author":"Fleischer Lisa K","year":"2000","unstructured":"Lisa K Fleischer. 2000. Approximating fractional multicommodity flow independent of the number of commodities. SIAM Journal on Discrete Mathematics 13, 4 (2000), 505\u2013520.","journal-title":"SIAM Journal on Discrete Mathematics"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.5281\/zenodo.13347196"},{"key":"e_1_3_2_1_13_1","volume-title":"Proceedings of the ACM SIGCOMM 2024 Conference. 71\u201385","author":"Gui Fei","year":"2024","unstructured":"Fei Gui, Songtao Wang, Dan Li, Li Chen, Kaihui Gao, Congcong Min, and Yi Wang. 2024. RedTE: Mitigating subsecond traffic bursts with real-time and distributed traffic engineering. In Proceedings of the ACM SIGCOMM 2024 Conference. 71\u201385."},{"key":"e_1_3_2_1_14_1","unstructured":"Gurobi Optimization LLC. 2025. Gurobi Optimizer Reference Manual. https:\/\/www.gurobi.com"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"crossref","unstructured":"Elad Hazan et al. 2016. Introduction to online convex optimization. Foundations and Trends\u00ae in Optimization 2 3-4 (2016) 157\u2013325.","DOI":"10.1561\/2400000013"},{"key":"e_1_3_2_1_16_1","volume-title":"Proceedings of the ACM SIGCOMM 2013 Conference on SIGCOMM. 15\u201326","author":"Hong Chi-Yao","year":"2013","unstructured":"Chi-Yao Hong, Srikanth Kandula, Ratul Mahajan, Ming Zhang, Vijay Gill, Mohan Nanduri, and Roger Wattenhofer. 2013. Achieving high utilization with software-driven WAN. In Proceedings of the ACM SIGCOMM 2013 Conference on SIGCOMM. 15\u201326."},{"key":"e_1_3_2_1_17_1","volume-title":"Proceedings of the 2018 Conference of the ACM Special Interest Group on Data Communication. 74\u201387","author":"Hong Chi-Yao","year":"2018","unstructured":"Chi-Yao Hong, Subhasree Mandal, Mohammad Al-Fares, Min Zhu, Richard Alimi, Chandan Bhagat, Sourabh Jain, Jay Kaimal, Shiyu Liang, Kirill Mendelev, et al. 2018. B4 and after: managing hierarchy, partitioning, and asymmetry for availability and scale in google's software-defined WAN. In Proceedings of the 2018 Conference of the ACM Special Interest Group on Data Communication. 74\u201387."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"crossref","first-page":"119","DOI":"10.1007\/s12532-017-0130-5","article-title":"Parallelizing the dual revised simplex method","volume":"10","author":"Huangfu Qi","year":"2018","unstructured":"Qi Huangfu and JA Julian Hall. 2018. Parallelizing the dual revised simplex method. Mathematical Programming Computation 10, 1 (2018), 119\u2013142.","journal-title":"Mathematical Programming Computation"},{"key":"e_1_3_2_1_19_1","volume-title":"IBM ILOG CPLEX Optimization Studio V22.1.1: User's Manual","author":"IBM Corporation","unstructured":"IBM Corporation. 2022. IBM ILOG CPLEX Optimization Studio V22.1.1: User's Manual. IBM Corporation."},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1145\/2534169.2486019","article-title":"B4: Experience with a globally-deployed software defined WAN","volume":"43","author":"Jain Sushant","year":"2013","unstructured":"Sushant Jain, Alok Kumar, Subhasree Mandal, Joon Ong, Leon Poutievski, Arjun Singh, Subbaiah Venkata, Jim Wanderer, Junlan Zhou, Min Zhu, et al. 2013. B4: Experience with a globally-deployed software defined WAN. ACM SIGCOMM Computer Communication Review 43, 4 (2013), 3\u201314.","journal-title":"ACM SIGCOMM Computer Communication Review"},{"key":"e_1_3_2_1_21_1","volume-title":"Proceedings of the 29th Symposium on Operating Systems Principles. 642\u2013657","author":"Subramanya Suhas Jayaram","year":"2023","unstructured":"Suhas Jayaram Subramanya, Daiyaan Arfeen, Shouxu Lin, Aurick Qiao, Zhihao Jia, and Gregory R Ganger. 2023. Sia: Heterogeneity-aware, goodput-optimized ML-cluster scheduling. In Proceedings of the 29th Symposium on Operating Systems Principles. 642\u2013657."},{"key":"e_1_3_2_1_22_1","volume-title":"2019 USENIX Annual Technical Conference (USENIX ATC 19)","author":"Jeon Myeongjae","year":"2019","unstructured":"Myeongjae Jeon, Shivaram Venkataraman, Amar Phanishayee, Junjie Qian, Wencong Xiao, and Fan Yang. 2019. Analysis of Large-Scale Multi-Tenant GPU clusters for DNN training workloads. In 2019 USENIX Annual Technical Conference (USENIX ATC 19). 947\u2013960."},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"crossref","first-page":"151","DOI":"10.1007\/s10589-007-9096-y","article-title":"Implementation of warmstart strategies in interior-point methods for linear programming in fixed dimension","volume":"41","author":"John Elizabeth","year":"2008","unstructured":"Elizabeth John and E Alper Y\u0131ld\u0131r\u0131m. 2008. Implementation of warmstart strategies in interior-point methods for linear programming in fixed dimension. Computational Optimization and Applications 41, 2 (2008), 151\u2013183.","journal-title":"Computational Optimization and Applications"},{"key":"e_1_3_2_1_24_1","volume-title":"18th USENIX Symposium on Operating Systems Design and Implementation (OSDI 24)","author":"Kumar Neeraj","year":"2024","unstructured":"Neeraj Kumar, Pol Mauri Ruiz, Vijay Menon, Igor Kabiljo, Mayank Pundir, Andrew Newell, Daniel Lee, Liyuan Wang, and Chunqiang Tang. 2024. Optimizing resource allocation in hyperscale datacenters: Scalability, usability, and experiences. In 18th USENIX Symposium on Operating Systems Design and Implementation (OSDI 24). 507\u2013528."},{"key":"e_1_3_2_1_25_1","volume-title":"http:\/\/www.gnu.org\/s\/glpk\/glpk.html","author":"Makhorin Andrew","year":"2008","unstructured":"Andrew Makhorin. 2008. GLPK (GNU linear programming kit). http:\/\/www.gnu.org\/s\/glpk\/glpk.html (2008)."},{"key":"e_1_3_2_1_26_1","volume-title":"Proceedings of the ACM SIGCOMM 2024 Conference. 103\u2013116","author":"Miao Congcong","year":"2024","unstructured":"Congcong Miao, Zhizhen Zhong, Yunming Xiao, Feng Yang, Senkuo Zhang, Yinan Jiang, Zizhuo Bai, Chaodong Lu, Jingyi Geng, Zekun He, et al. 2024. MegaTE: Extending WAN Traffic Engineering to Millions of Endpoints in Virtualized Cloud. In Proceedings of the ACM SIGCOMM 2024 Conference. 103\u2013116."},{"key":"e_1_3_2_1_27_1","volume-title":"13th USENIX symposium on operating systems design and implementation (OSDI 18)","author":"Moritz Philipp","year":"2018","unstructured":"Philipp Moritz, Robert Nishihara, Stephanie Wang, Alexey Tumanov, Richard Liaw, Eric Liang, Melih Elibol, Zongheng Yang, William Paul, Michael I Jordan, et al. 2018. Ray: A distributed framework for emerging AI applications. In 13th USENIX symposium on operating systems design and implementation (OSDI 18). 561\u2013577."},{"key":"e_1_3_2_1_28_1","volume-title":"Proceedings of the 21st ACM Workshop on Hot Topics in Networks. 138\u2013144","author":"Namyar Pooria","year":"2022","unstructured":"Pooria Namyar, Behnaz Arzani, Ryan Beckett, Santiago Segarra, Himanshu Raj, and Srikanth Kandula. 2022. Minding the gap between fast heuristics and their optimal counterparts. In Proceedings of the 21st ACM Workshop on Hot Topics in Networks. 138\u2013144."},{"key":"e_1_3_2_1_29_1","volume-title":"Proceedings of the ACM SIGOPS 28th Symposium on Operating Systems Principles. 521\u2013537","author":"Narayanan Deepak","year":"2021","unstructured":"Deepak Narayanan, Fiodar Kazhamiaka, Firas Abuzaid, Peter Kraft, Akshay Agrawal, Srikanth Kandula, Stephen Boyd, and Matei Zaharia. 2021. Solving large-scale granular resource allocation problems efficiently with pop. In Proceedings of the ACM SIGOPS 28th Symposium on Operating Systems Principles. 521\u2013537."},{"key":"e_1_3_2_1_30_1","volume-title":"Heterogeneity-Aware Cluster Scheduling Policies for Deep Learning Workloads. In 14th USENIX Symposium on Operating Systems Design and Implementation (OSDI 20)","author":"Narayanan Deepak","year":"2020","unstructured":"Deepak Narayanan, Keshav Santhanam, Fiodar Kazhamiaka, Amar Phanishayee, and Matei Zaharia. 2020. Heterogeneity-Aware Cluster Scheduling Policies for Deep Learning Workloads. In 14th USENIX Symposium on Operating Systems Design and Implementation (OSDI 20)."},{"key":"e_1_3_2_1_31_1","volume-title":"Proceedings of the ACM SIGOPS 28th Symposium on Operating Systems Principles. 505\u2013520","author":"Newell Andrew","year":"2021","unstructured":"Andrew Newell, Dimitrios Skarlatos, Jingyuan Fan, Pavan Kumar, Maxim Khutornenko, Mayank Pundir, Yirui Zhang, Mingjun Zhang, Yuanlai Liu, Linh Le, et al. 2021. RAS: continuously optimized region-wide datacenter resource allocation. In Proceedings of the ACM SIGOPS 28th Symposium on Operating Systems Principles. 505\u2013520."},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"crossref","unstructured":"Jorge Nocedal and Stephen J. Wright (Eds.). 1999. Numerical Optimization.","DOI":"10.1007\/b98874"},{"key":"e_1_3_2_1_33_1","volume-title":"Operator splitting for conic optimization via homogeneous self-dual embedding. arXiv preprint arXiv:1312.3039 1, 2.1","author":"O'Donoghue Brendan","year":"2013","unstructured":"Brendan O'Donoghue, Eric Chu, Neal Parikh, and Stephen Boyd. 2013. Operator splitting for conic optimization via homogeneous self-dual embedding. arXiv preprint arXiv:1312.3039 1, 2.1 (2013), 3."},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"crossref","unstructured":"Neal Parikh Stephen Boyd et al. 2014. Proximal algorithms. Foundations and trends\u00ae in Optimization 1 3 (2014) 127\u2013239.","DOI":"10.1561\/2400000003"},{"key":"e_1_3_2_1_35_1","unstructured":"Laurent Perron and Vincent Furnon. 2024. OR-Tools. Google. https:\/\/developers.google.com\/optimization\/"},{"key":"e_1_3_2_1_36_1","volume-title":"DOTE: Rethinking (Predictive)WAN Traffic Engineering. In 20th USENIX Symposium on Networked Systems Design and Implementation (NSDI 23)","author":"Perry Yarin","year":"2023","unstructured":"Yarin Perry, Felipe Vieira Frujeri, Chaim Hoch, Srikanth Kandula, Ishai Menache, Michael Schapira, and Aviv Tamar. 2023. DOTE: Rethinking (Predictive)WAN Traffic Engineering. In 20th USENIX Symposium on Networked Systems Design and Implementation (NSDI 23). 1557\u20131581."},{"key":"e_1_3_2_1_37_1","volume-title":"15th USENIX Symposium on Operating Systems Design and Implementation (OSDI 21)","author":"Qiao Aurick","year":"2021","unstructured":"Aurick Qiao, Sang Keun Choe, Suhas Jayaram Subramanya, Willie Neiswanger, Qirong Ho, Hao Zhang, Gregory R Ganger, and Eric P Xing. 2021. Pollux: Co-adaptive cluster scheduling for goodput-optimized deep learning. In 15th USENIX Symposium on Operating Systems Design and Implementation (OSDI 21)."},{"key":"e_1_3_2_1_38_1","volume-title":"Optimization, learning, and games with predictable sequences. Advances in Neural Information Processing Systems 26","author":"Rakhlin Sasha","year":"2013","unstructured":"Sasha Rakhlin and Karthik Sridharan. 2013. Optimization, learning, and games with predictable sequences. Advances in Neural Information Processing Systems 26 (2013)."},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"crossref","first-page":"55","DOI":"10.1016\/j.jpdc.2020.05.021","article-title":"GPU acceleration of ADMM for large-scale quadratic programming","volume":"144","author":"Schubiger Michel","year":"2020","unstructured":"Michel Schubiger, Goran Banjac, and John Lygeros. 2020. GPU acceleration of ADMM for large-scale quadratic programming. J. Parallel and Distrib. Comput. 144 (2020), 55\u201367.","journal-title":"J. Parallel and Distrib. Comput."},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"crossref","first-page":"1035","DOI":"10.14778\/2732977.2732979","article-title":"Accordion: Elastic scalability for database systems supporting distributed transactions","volume":"7","author":"Serafini Marco","year":"2014","unstructured":"Marco Serafini, Essam Mansour, Ashraf Aboulnaga, Kenneth Salem, Taha Rafiq, and Umar Farooq Minhas. 2014. Accordion: Elastic scalability for database systems supporting distributed transactions. Proceedings of the VLDB Endowment 7, 12 (2014), 1035\u20131046.","journal-title":"Proceedings of the VLDB Endowment"},{"key":"e_1_3_2_1_41_1","doi-asserted-by":"crossref","first-page":"637","DOI":"10.1007\/s12532-020-00179-2","article-title":"OSQP: An operator splitting solver for quadratic programs","volume":"12","author":"Stellato Bartolomeo","year":"2020","unstructured":"Bartolomeo Stellato, Goran Banjac, Paul Goulart, Alberto Bemporad, and Stephen Boyd. 2020. OSQP: An operator splitting solver for quadratic programs. Mathematical Programming Computation 12, 4 (2020), 637\u2013672.","journal-title":"Mathematical Programming Computation"},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"crossref","first-page":"245","DOI":"10.14778\/2735508.2735514","article-title":"E-store: Fine-grained elastic partitioning for distributed transaction processing systems","volume":"8","author":"Taft Rebecca","year":"2014","unstructured":"Rebecca Taft, Essam Mansour, Marco Serafini, Jennie Duggan, Aaron J Elmore, Ashraf Aboulnaga, Andrew Pavlo, and Michael Stonebraker. 2014. E-store: Fine-grained elastic partitioning for distributed transaction processing systems. Proceedings of the VLDB Endowment 8, 3 (2014), 245\u2013256.","journal-title":"Proceedings of the VLDB Endowment"},{"key":"e_1_3_2_1_43_1","volume-title":"Proceedings of the Eleventh European Conference on Computer Systems. 1\u201316","author":"Tumanov Alexey","year":"2016","unstructured":"Alexey Tumanov, Timothy Zhu, Jun Woo Park, Michael A Kozuch, Mor Harchol-Balter, and Gregory R Ganger. 2016. TetriSched: global rescheduling with adaptive plan-ahead in dynamic heterogeneous clusters. In Proceedings of the Eleventh European Conference on Computer Systems. 1\u201316."},{"key":"e_1_3_2_1_44_1","volume-title":"A new alternating direction method for linear programming. Advances in neural information processing systems 30","author":"Wang Sinong","year":"2017","unstructured":"Sinong Wang and Ness Shroff. 2017. A new alternating direction method for linear programming. Advances in neural information processing systems 30 (2017)."},{"key":"e_1_3_2_1_45_1","volume-title":"19th USENIX Symposium on Networked Systems Design and Implementation (NSDI 22)","author":"Weng Qizhen","year":"2022","unstructured":"Qizhen Weng, Wencong Xiao, Yinghao Yu, Wei Wang, Cheng Wang, Jian He, Yong Li, Liping Zhang, Wei Lin, and Yu Ding. 2022. MLaaS in the wild: Workload analysis and scheduling in Large-Scale heterogeneous GPU clusters. In 19th USENIX Symposium on Networked Systems Design and Implementation (NSDI 22). 945\u2013960."},{"key":"e_1_3_2_1_46_1","volume-title":"Proceedings of the ACM SIGCOMM 2023 Conference. 378\u2013393","author":"Xu Zhiying","year":"2023","unstructured":"Zhiying Xu, Francis Y Yan, Rachee Singh, Justin T Chiu, Alexander M Rush, and Minlan Yu. 2023. Teal: Learning-accelerated optimization of wan traffic engineering. In Proceedings of the ACM SIGCOMM 2023 Conference. 378\u2013393."},{"key":"e_1_3_2_1_47_1","volume-title":"Sparse linear programming via primal and dual augmented coordinate descent. Advances in neural information processing systems 28","author":"En-Hsu Yen Ian","year":"2015","unstructured":"Ian En-Hsu Yen, Kai Zhong, Cho-Jui Hsieh, Pradeep K Ravikumar, and Inderjit S Dhillon. 2015. Sparse linear programming via primal and dual augmented coordinate descent. Advances in neural information processing systems 28 (2015)."},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"crossref","first-page":"782","DOI":"10.1137\/S1052623400369235","article-title":"Warm-start strategies in interior-point methods for linear programming","volume":"12","author":"Alper Yildirim E","year":"2002","unstructured":"E Alper Yildirim and Stephen J Wright. 2002. Warm-start strategies in interior-point methods for linear programming. SIAM Journal on Optimization 12, 3 (2002), 782\u2013810.","journal-title":"SIAM Journal on Optimization"},{"key":"e_1_3_2_1_49_1","first-page":"11573","article-title":"Efficient methods for non-stationary online learning","volume":"35","author":"Zhao Peng","year":"2022","unstructured":"Peng Zhao, Yan-Feng Xie, Lijun Zhang, and Zhi-Hua Zhou. 2022. Efficient methods for non-stationary online learning. Advances in Neural Information Processing Systems 35 (2022), 11573\u201311585.","journal-title":"Advances in Neural Information Processing Systems"},{"key":"e_1_3_2_1_50_1","unstructured":"Peng Zhao and Lijun Zhang. 2021. Improved analysis for dynamic regret of strongly convex and smooth functions. In Learning for Dynamics and Control. PMLR 48\u201359."},{"key":"e_1_3_2_1_51_1","volume-title":"20th USENIX Symposium on Networked Systems Design and Implementation (NSDI 23)","author":"Zheng Pengfei","year":"2023","unstructured":"Pengfei Zheng, Rui Pan, Tarannum Khan, Shivaram Venkataraman, and Aditya Akella. 2023. Shockwave: Fair and efficient cluster scheduling for dynamic adaptation in machine learning. In 20th USENIX Symposium on Networked Systems Design and Implementation (NSDI 23). 703\u2013723."}],"event":{"name":"SOSP '25: ACM SIGOPS 31st Symposium on Operating Systems Principles","location":"Lotte Hotel World Seoul Republic of Korea","acronym":"SOSP '25","sponsor":["SIGOPS ACM Special Interest Group on Operating Systems","USENIX"]},"container-title":["Proceedings of the ACM SIGOPS 31st Symposium on Operating Systems Principles"],"original-title":[],"deposited":{"date-parts":[[2025,10,1]],"date-time":"2025-10-01T12:47:31Z","timestamp":1759322851000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3731569.3764846"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,12]]},"references-count":51,"alternative-id":["10.1145\/3731569.3764846","10.1145\/3731569"],"URL":"https:\/\/doi.org\/10.1145\/3731569.3764846","relation":{},"subject":[],"published":{"date-parts":[[2025,10,12]]},"assertion":[{"value":"2025-10-12","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}