{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,15]],"date-time":"2026-07-15T07:12:20Z","timestamp":1784099540694,"version":"3.55.0"},"reference-count":90,"publisher":"Association for Computing Machinery (ACM)","issue":"12","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. VLDB Endow."],"published-print":{"date-parts":[[2023,8]]},"abstract":"<jats:p>The distributed data analytic system - Spark is a common choice for processing massive volumes of heterogeneous data, while it is challenging to tune its parameters to achieve high performance. Recent studies try to employ auto-tuning techniques to solve this problem but suffer from three issues: limited functionality, high overhead, and inefficient search.<\/jats:p>\n          <jats:p>In this paper, we present a general and efficient Spark tuning framework that can deal with the three issues simultaneously. First, we introduce a generalized tuning formulation, which can support multiple tuning goals and constraints conveniently, and a Bayesian optimization (BO) based solution to solve this generalized optimization problem. Second, to avoid high overhead from additional offline evaluations in existing methods, we propose to tune parameters along with the actual periodic executions of each job (i.e., online evaluations). To ensure safety during online job executions, we design a safe configuration acquisition method that models the safe region. Finally, three innovative techniques are leveraged to further accelerate the search process: adaptive sub-space generation, approximate gradient descent, and meta-learning method.<\/jats:p>\n          <jats:p>We have implemented this framework as an independent cloud service, and applied it to the data platform in Tencent. The empirical results on both public benchmarks and large-scale production tasks demonstrate its superiority in terms of practicality, generality, and efficiency. Notably, this service saves an average of 57.00% memory cost and 34.93% CPU cost on 25K in-production tasks within 20 iterations, respectively.<\/jats:p>","DOI":"10.14778\/3611540.3611548","type":"journal-article","created":{"date-parts":[[2023,9,15]],"date-time":"2023-09-15T11:32:37Z","timestamp":1694777557000},"page":"3570-3583","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":24,"title":["Towards General and Efficient Online Tuning for Spark"],"prefix":"10.14778","volume":"16","author":[{"given":"Yang","family":"Li","sequence":"first","affiliation":[{"name":"Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Huaijun","family":"Jiang","sequence":"additional","affiliation":[{"name":"Peking University &amp; Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yu","family":"Shen","sequence":"additional","affiliation":[{"name":"Peking University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yide","family":"Fang","sequence":"additional","affiliation":[{"name":"Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xiaofeng","family":"Yang","sequence":"additional","affiliation":[{"name":"Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Danqing","family":"Huang","sequence":"additional","affiliation":[{"name":"Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Xinyi","family":"Zhang","sequence":"additional","affiliation":[{"name":"Peking University"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Wentao","family":"Zhang","sequence":"additional","affiliation":[{"name":"Mila - Qu\u00e9bec AI Institute"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ce","family":"Zhang","sequence":"additional","affiliation":[{"name":"ETH Z\u00fcrich"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Peng","family":"Chen","sequence":"additional","affiliation":[{"name":"Tencent Inc."}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bin","family":"Cui","sequence":"additional","affiliation":[{"name":"Peking University"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,8]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Automatic Database Management System Tuning Through Large-scale Machine Learning. In SIGMOD Conference. ACM, 1009--1024","author":"Aken Dana Van","year":"2017","unstructured":"Dana Van Aken, Andrew Pavlo, Geoffrey J. Gordon, and Bohan Zhang. 2017. Automatic Database Management System Tuning Through Large-scale Machine Learning. In SIGMOD Conference. ACM, 1009--1024."},{"key":"e_1_2_1_2_1","volume-title":"CherryPick: Adaptively Unearthing the Best Cloud Configurations for Big Data Analytics. In 14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17)","author":"Alipourfard Omid","year":"2017","unstructured":"Omid Alipourfard, Hongqiang Harry Liu, Jianshu Chen, Shivaram Venkataraman, Minlan Yu, and Ming Zhang. 2017. CherryPick: Adaptively Unearthing the Best Cloud Configurations for Big Data Analytics. In 14th USENIX Symposium on Networked Systems Design and Implementation (NSDI 17). 469--482."},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11227-020-03307-w"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/2723372.2742797"},{"key":"e_1_2_1_5_1","volume-title":"Transfer Learning for Bayesian Optimization: A Survey. arXiv preprint arXiv:2302.05927","author":"Bai Tianyi","year":"2023","unstructured":"Tianyi Bai, Yang Li, Yu Shen, Xinyi Zhang, Wentao Zhang, and Bin Cui. 2023. Transfer Learning for Bayesian Optimization: A Survey. arXiv preprint arXiv:2302.05927 (2023)."},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData.2018.8622018"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2015.2449299"},{"key":"e_1_2_1_8_1","first-page":"281","article-title":"Random search for hyper-parameter optimization","author":"Bergstra James","year":"2012","unstructured":"James Bergstra and Yoshua Bengio. 2012. Random search for hyper-parameter optimization. Journal of Machine Learning Research 13, Feb (2012), 281--305.","journal-title":"Journal of Machine Learning Research 13"},{"key":"e_1_2_1_9_1","unstructured":"James S Bergstra R\u00e9mi Bardenet Yoshua Bengio and Bal\u00e1zs K\u00e9gl. 2011. Algorithms for hyper-parameter optimization. In Advances in neural information processing systems. 2546--2554."},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/3514221.3517882"},{"key":"e_1_2_1_11_1","volume-title":"Apache flink: Stream and batch processing in a single engine. Bulletin of the IEEE Computer Society Technical Committee on Data Engineering 36, 4","author":"Carbone Paris","year":"2015","unstructured":"Paris Carbone, Asterios Katsifodimos, Stephan Ewen, Volker Markl, Seif Haridi, and Kostas Tzoumas. 2015. Apache flink: Stream and batch processing in a single engine. Bulletin of the IEEE Computer Society Technical Committee on Data Engineering 36, 4 (2015)."},{"key":"e_1_2_1_12_1","volume-title":"d-simplexed: Adaptive delaunay triangulation for performance modeling and prediction on big data analytics","author":"Chen Yuxing","year":"2019","unstructured":"Yuxing Chen, Peter Goetsch, Mohammad A Hoque, Jiaheng Lu, and Sasu Tarkoma. 2019. d-simplexed: Adaptive delaunay triangulation for performance modeling and prediction on big data analytics. IEEE Transactions on Big Data (2019)."},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/3357384.3358090"},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/DISTRA.2018.8601018"},{"key":"e_1_2_1_15_1","volume-title":"Multi-Layer-Mesh: A Novel Topology and SDN-based Path Switching for Big Data Cluster Networks. In ICC 2019-2019 IEEE International Conference on Communications (ICC). IEEE, 1--7.","author":"de Almeida Leandro Batista","year":"2019","unstructured":"Leandro Batista de Almeida, Damien Magoni, Philip Perry, Eduardo Cunha de Almeida, John Murphy, and Anthony Ventresque. 2019. Multi-Layer-Mesh: A Novel Topology and SDN-based Path Switching for Big Data Cluster Networks. In ICC 2019-2019 IEEE International Conference on Communications (ICC). IEEE, 1--7."},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/1327452.1327492"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/2499368.2451125"},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/2644865.2541941"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPADS.2015.112"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.14778\/2367502.2367562"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00778-013-0319-9"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687627.1687767"},{"key":"e_1_2_1_23_1","volume-title":"Scalable global optimization via local bayesian optimization. Advances in neural information processing systems 32","author":"Eriksson David","year":"2019","unstructured":"David Eriksson, Michael Pearce, Jacob Gardner, Ryan D Turner, and Matthias Poloczek. 2019. Scalable global optimization via local bayesian optimization. Advances in neural information processing systems 32 (2019)."},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/3394486.3403299"},{"key":"e_1_2_1_25_1","volume-title":"Scalable meta-learning for Bayesian optimization. stat 1050","author":"Feurer Matthias","year":"2018","unstructured":"Matthias Feurer, Benjamin Letham, and Eytan Bakshy. 2018. Scalable meta-learning for Bayesian optimization. stat 1050 (2018), 6."},{"key":"e_1_2_1_26_1","volume-title":"ICML","volume":"2014","author":"Gardner Jacob R","year":"2014","unstructured":"Jacob R Gardner, Matt J Kusner, Zhixiang Eddie Xu, Kilian Q Weinberger, and John P Cunningham. 2014. Bayesian optimization with inequality constraints.. In ICML, Vol. 2014. 937--945."},{"key":"e_1_2_1_27_1","volume-title":"on-line tuning of YARN container memory and CPU parameters. In 2016 IEEE 18th International Conference on High Performance Computing and Communications","author":"Genkin Mikhail","unstructured":"Mikhail Genkin, Frank Dehne, Maria Pospelova, Yabing Chen, and Pablo Navarro. 2016. Automatic, on-line tuning of YARN container memory and CPU parameters. In 2016 IEEE 18th International Conference on High Performance Computing and Communications; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC\/SmartCity\/DSS). IEEE, 317--324."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/3097983.3098043"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2017.2647939"},{"key":"e_1_2_1_30_1","unstructured":"Spark Tuning guide. 2022. Tuning - Spark 3.2.1 Documentation. https:\/\/spark.apache.org\/docs\/latest\/tuning.html"},{"key":"e_1_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2017.07.003"},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3381027"},{"key":"e_1_2_1_33_1","first-page":"261","article-title":"Starfish: A Self-tuning System for Big Data Analytics","volume":"11","author":"Herodotou Herodotos","year":"2011","unstructured":"Herodotos Herodotou, Harold Lim, Gang Luo, Nedyalko Borisov, Liang Dong, Fatma Bilgen Cetin, and Shivnath Babu. 2011. Starfish: A Self-tuning System for Big Data Analytics.. In Cidr, Vol. 11. 261--272.","journal-title":"Cidr"},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDEW.2010.5452747"},{"key":"e_1_2_1_35_1","volume-title":"International conference on machine learning. PMLR, 754--762","author":"Hutter Frank","year":"2014","unstructured":"Frank Hutter, Holger Hoos, and Kevin Leyton-Brown. 2014. An efficient approach for assessing hyperparameter importance. In International conference on machine learning. PMLR, 754--762."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-25566-3_40"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/2967938.2967957"},{"key":"e_1_2_1_38_1","volume-title":"OpenBox: A Python Toolkit for Generalized Black-box Optimization. arXiv preprint arXiv:2304.13339","author":"Jiang Huaijun","year":"2023","unstructured":"Huaijun Jiang, Yu Shen, Yang Li, Wentao Zhang, Ce Zhang, and Bin Cui. 2023. OpenBox: A Python Toolkit for Generalized Black-box Optimization. arXiv preprint arXiv:2304.13339 (2023)."},{"key":"e_1_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1023\/A:1008306431147"},{"key":"e_1_2_1_40_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-030-33495-6_34"},{"key":"e_1_2_1_41_1","volume-title":"Lightgbm: A highly efficient gradient boosting decision tree. Advances in neural information processing systems 30","author":"Ke Guolin","year":"2017","unstructured":"Guolin Ke, Qi Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and Tie-Yan Liu. 2017. Lightgbm: A highly efficient gradient boosting decision tree. Advances in neural information processing systems 30 (2017)."},{"key":"e_1_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/2723372.2742788"},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/3318464.3380591"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.1145\/3318464.3380591"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/2371536.2371547"},{"key":"e_1_2_1_46_1","volume-title":"Yon Dohn Chung, and Bongki Moon","author":"Lee Kyong-Ha","year":"2012","unstructured":"Kyong-Ha Lee, Yoon-Joon Lee, Hyunsik Choi, Yon Dohn Chung, and Bongki Moon. 2012. Parallel data processing with MapReduce: a survey. AcM sIGMoD record 40, 4 (2012), 11--20."},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.14778\/3352063.3352129"},{"key":"e_1_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.14778\/3352063.3352129"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1145\/3534678.3539369"},{"key":"e_1_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1145\/3534678.3539255"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/3447548.3467061"},{"key":"e_1_2_1_52_1","volume-title":"VolcanoML: speeding up end-to-end AutoML via scalable search space decomposition. The VLDB Journal","author":"Li Yang","year":"2022","unstructured":"Yang Li, Yu Shen, Wentao Zhang, Ce Zhang, and Bin Cui. 2022. VolcanoML: speeding up end-to-end AutoML via scalable search space decomposition. The VLDB Journal (2022), 1--25."},{"key":"e_1_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2016.2624733"},{"key":"e_1_2_1_54_1","volume-title":"Query-based Workload Forecasting for Self-Driving Database Management Systems. In SIGMOD Conference. ACM, 631--645","author":"Ma Lin","unstructured":"Lin Ma, Dana Van Aken, Ahmed Hefny, Gustavo Mezerhane, Andrew Pavlo, and Geoffrey J. Gordon. 2018. Query-based Workload Forecasting for Self-Driving Database Management Systems. In SIGMOD Conference. ACM, 631--645."},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.5555\/2946645.2946679"},{"key":"e_1_2_1_56_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLOUD.2018.00059"},{"key":"e_1_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.14778\/3137765.3137770"},{"key":"e_1_2_1_58_1","volume-title":"TPOT: A tree-based pipeline optimization tool for automating machine learning. In Automated Machine Learning","author":"Olson Randal S","year":"2019","unstructured":"Randal S Olson and Jason H Moore. 2019. TPOT: A tree-based pipeline optimization tool for automating machine learning. In Automated Machine Learning. Springer, 151--160."},{"key":"e_1_2_1_59_1","volume-title":"INNS Conference on Big Data. Springer, 226--237","author":"Petridis Panagiotis","year":"2016","unstructured":"Panagiotis Petridis, Anastasios Gounaris, and Jordi Torres. 2016. Spark parameter tuning via trial-and-error. In INNS Conference on Big Data. Springer, 226--237."},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNSM.2020.3034824"},{"key":"e_1_2_1_61_1","volume-title":"Summer school on machine learning","author":"Rasmussen Carl Edward","unstructured":"Carl Edward Rasmussen. 2003. Gaussian processes in machine learning. In Summer school on machine learning. Springer, 63--71."},{"key":"e_1_2_1_62_1","volume-title":"An overview of gradient descent optimization algorithms. arXiv preprint arXiv:1609.04747","author":"Ruder Sebastian","year":"2016","unstructured":"Sebastian Ruder. 2016. An overview of gradient descent optimization algorithms. arXiv preprint arXiv:1609.04747 (2016)."},{"key":"e_1_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2015.2494218"},{"key":"e_1_2_1_64_1","volume-title":"Rover: An online Spark SQL tuning service via generalized transfer learning. arXiv preprint arXiv:2302.04046","author":"Shen Yu","year":"2023","unstructured":"Yu Shen, Xinyuyang Ren, Yupeng Lu, Huaijun Jiang, Huanyong Xu, Di Peng, Yang Li, Wentao Zhang, and Bin Cui. 2023. Rover: An online Spark SQL tuning service via generalized transfer learning. arXiv preprint arXiv:2302.04046 (2023)."},{"key":"e_1_2_1_65_1","volume-title":"Technology Conference on Performance Evaluation and Benchmarking. Springer, 131--146","author":"Singhal Rekha","year":"2017","unstructured":"Rekha Singhal and Praveen Singh. 2017. Performance assurance model for applications on SPARK platform. In Technology Conference on Performance Evaluation and Benchmarking. Springer, 131--146."},{"key":"e_1_2_1_66_1","unstructured":"Jasper Snoek Hugo Larochelle and Ryan P Adams. 2012. Practical bayesian optimization of machine learning algorithms. In Advances in neural information processing systems."},{"key":"e_1_2_1_67_1","volume-title":"On quasi-monte carlo integrations. Mathematics and computers in simulation 47, 2--5","author":"Sobol Ilya M","year":"1998","unstructured":"Ilya M Sobol. 1998. On quasi-monte carlo integrations. Mathematics and computers in simulation 47, 2--5 (1998), 103--112."},{"key":"e_1_2_1_68_1","unstructured":"SparkConf. 2022. Configuration - Spark 3.2.1 Documentation. https:\/\/spark.apache.org\/docs\/latest\/configuration.html"},{"key":"e_1_2_1_69_1","volume-title":"International conference on machine learning. PMLR, 997--1005","author":"Sui Yanan","year":"2015","unstructured":"Yanan Sui, Alkis Gotovos, Joel Burdick, and Andreas Krause. 2015. Safe exploration for optimization with Gaussian processes. In International conference on machine learning. PMLR, 997--1005."},{"key":"e_1_2_1_70_1","volume-title":"Apache Spark: Unified engine for large-scale data analytics. https:\/\/spark.apache.org\/","author":"Team Apache Spark","year":"2022","unstructured":"Apache Spark Team. 2022. Apache Spark: Unified engine for large-scale data analytics. https:\/\/spark.apache.org\/"},{"key":"e_1_2_1_71_1","unstructured":"Spark Streaming Team. 2022. Spark Streaming. http:\/\/spark.apache.org\/streaming\/"},{"key":"e_1_2_1_72_1","doi-asserted-by":"publisher","DOI":"10.1145\/2588555.2595641"},{"key":"e_1_2_1_73_1","volume-title":"13th USENIX Symposium on Networked Systems Design and Implementation (NSDI 16)","author":"Venkataraman Shivaram","year":"2016","unstructured":"Shivaram Venkataraman, Zongheng Yang, Michael Franklin, Benjamin Recht, and Ion Stoica. 2016. Ernest: Efficient Performance Prediction for {Large-Scale} Advanced Analytics. In 13th USENIX Symposium on Networked Systems Design and Implementation (NSDI 16). 363--378."},{"key":"e_1_2_1_74_1","volume-title":"A novel method for tuning configuration parameters of spark based on machine learning. In 2016 IEEE 18th International Conference on High Performance Computing and Communications","author":"Wang Guolu","unstructured":"Guolu Wang, Jungang Xu, and Ben He. 2016. A novel method for tuning configuration parameters of spark based on machine learning. In 2016 IEEE 18th International Conference on High Performance Computing and Communications; IEEE 14th International Conference on Smart City; IEEE 2nd International Conference on Data Science and Systems (HPCC\/SmartCity\/DSS). IEEE, 586--593."},{"key":"e_1_2_1_75_1","volume-title":"Performance prediction for apache spark platform. In 2015 IEEE 17th International Conference on High Performance Computing and Communications","author":"Wang Kewen","year":"2015","unstructured":"Kewen Wang and Mohammad Maifi Hasan Khan. 2015. Performance prediction for apache spark platform. In 2015 IEEE 17th International Conference on High Performance Computing and Communications, 2015 IEEE 7th International Symposium on Cyberspace Safety and Security, and 2015 IEEE 12th International Conference on Embedded Software and Systems. IEEE, 166--173."},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1145\/3514221.3526157"},{"key":"e_1_2_1_77_1","doi-asserted-by":"publisher","DOI":"10.1145\/2484425.2484427"},{"key":"e_1_2_1_78_1","doi-asserted-by":"publisher","DOI":"10.1145\/2463676.2465288"},{"key":"e_1_2_1_79_1","doi-asserted-by":"publisher","DOI":"10.1145\/3173162.3173187"},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData.2017.8257950"},{"key":"e_1_2_1_81_1","volume-title":"9th USENIX Symposium on Networked Systems Design and Implementation (NSDI 12)","author":"Zaharia Matei","year":"2012","unstructured":"Matei Zaharia, Mosharaf Chowdhury, Tathagata Das, Ankur Dave, Justin Ma, Murphy McCauly, Michael J Franklin, Scott Shenker, and Ion Stoica. 2012. Resilient distributed datasets: A {Fault-Tolerant} abstraction for {In-Memory} cluster computing. In 9th USENIX Symposium on Networked Systems Design and Implementation (NSDI 12). 15--28."},{"key":"e_1_2_1_82_1","doi-asserted-by":"publisher","DOI":"10.1145\/2934664"},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.14778\/3229863.3236222"},{"key":"e_1_2_1_84_1","volume-title":"An End-to-End Automatic Cloud Database Tuning System Using Deep Reinforcement Learning. In SIGMOD Conference. ACM, 415--432","author":"Zhang Ji","year":"2019","unstructured":"Ji Zhang, Yu Liu, Ke Zhou, Guoliang Li, Zhili Xiao, Bin Cheng, Jiashu Xing, Yangtao Wang, Tianheng Cheng, Li Liu, Minwei Ran, and Zekang Li. 2019. An End-to-End Automatic Cloud Database Tuning System Using Deep Reinforcement Learning. In SIGMOD Conference. ACM, 415--432."},{"key":"e_1_2_1_85_1","doi-asserted-by":"publisher","DOI":"10.14778\/3538598.3538604"},{"key":"e_1_2_1_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/3589331"},{"key":"e_1_2_1_87_1","doi-asserted-by":"publisher","DOI":"10.1145\/3448016.3457291"},{"key":"e_1_2_1_88_1","doi-asserted-by":"publisher","DOI":"10.1145\/3514221.3526176"},{"key":"e_1_2_1_89_1","doi-asserted-by":"publisher","DOI":"10.1145\/3448016.3457569"},{"key":"e_1_2_1_90_1","doi-asserted-by":"publisher","DOI":"10.1145\/3127479.3128605"}],"container-title":["Proceedings of the VLDB Endowment"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.14778\/3611540.3611548","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,10]],"date-time":"2025-09-10T22:37:26Z","timestamp":1757543846000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.14778\/3611540.3611548"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,8]]},"references-count":90,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2023,8]]}},"alternative-id":["10.14778\/3611540.3611548"],"URL":"https:\/\/doi.org\/10.14778\/3611540.3611548","relation":{},"ISSN":["2150-8097"],"issn-type":[{"value":"2150-8097","type":"print"}],"subject":[],"published":{"date-parts":[[2023,8]]},"assertion":[{"value":"2023-08-01","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}