{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,1]],"date-time":"2026-06-01T20:37:31Z","timestamp":1780346251421,"version":"3.54.1"},"reference-count":42,"publisher":"Association for Computing Machinery (ACM)","issue":"7","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. VLDB Endow."],"published-print":{"date-parts":[[2024,3]]},"abstract":"<jats:p>Current index advisors often struggle to balance efficiency and effectiveness when dealing with workload shifts. This arises from ignorance of the continual similarity and distant variety in workloads. This paper proposes a novel learning-based index advisor called BALANCE, which boosts indexing performance by leveraging knowledge obtained from dynamic and heterogeneous workloads. Our approach consists of three components. First, we build separate Lightweight Index Advisors (LIAs) on sequential chunks of similar workloads, where each LIA is trained with a small batch of workloads drawn from the chunk, and it provides direct index recommendations for all workloads in the same chunk. Second, we perform a policy transfer mechanism by adapting the LIA's index selection strategy from historical knowledge, substantially reducing the training overhead. Third, we employ a self-supervised contrastive learning method to provide an off-the-shelf workload representation, enabling the LIA to generate more accurate index recommendations. Extensive experiments across various benchmarks demonstrate that BALANCE improves the state-of-the-art learning-based index advisor, SWIRL, by 10.03% while reducing training overhead by 35.70% on average.<\/jats:p>","DOI":"10.14778\/3654621.3654631","type":"journal-article","created":{"date-parts":[[2024,5,30]],"date-time":"2024-05-30T22:21:08Z","timestamp":1717107668000},"page":"1642-1654","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":10,"title":["Leveraging Dynamic and Heterogeneous Workload Knowledge to Boost the Performance of Index Advisors"],"prefix":"10.14778","volume":"17","author":[{"given":"Zijia","family":"Wang","sequence":"first","affiliation":[{"name":"School of Informatics, Xiamen University, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haoran","family":"Liu","sequence":"additional","affiliation":[{"name":"School of Informatics, Xiamen University, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Chen","family":"Lin","sequence":"additional","affiliation":[{"name":"School of Informatics, Xiamen University, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Zhifeng","family":"Bao","sequence":"additional","affiliation":[{"name":"RMIT University, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Guoliang","family":"Li","sequence":"additional","affiliation":[{"name":"Department of Computer Science, Tsinghua University, China"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Tianqing","family":"Wang","sequence":"additional","affiliation":[{"name":"Huawei Company, China"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2024,5,30]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1109\/MSP.2017.2743240"},{"key":"e_1_2_1_2_1","unstructured":"Andr\u00e9 Barreto Will Dabney R\u00e9mi Munos Jonathan J. Hunt Tom Schaul David Silver and Hado van Hasselt. 2017. Successor Features for Transfer in Reinforcement Learning. In Advances in Neural Information Processing Systems. 4055--4065."},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1066157.1066184"},{"key":"e_1_2_1_4_1","volume-title":"2007 IEEE 23rd International Conference on Data Engineering. 826--835","author":"Bruno Nicolas","year":"2007","unstructured":"Nicolas Bruno and Surajit Chaudhuri. 2007. An Online Approach to Physical Design Tuning. In 2007 IEEE 23rd International Conference on Data Engineering. 826--835."},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.5555\/2772879.2772905"},{"key":"e_1_2_1_6_1","volume-title":"Proceedings of the 23rd International Conference on Very Large Data Bases. 146--155","author":"Chaudhuri Surajit","unstructured":"Surajit Chaudhuri and Vivek R. Narasayya. 1997. An Efficient Cost-Driven Index Selection Tool for Microsoft SQL Server. In Proceedings of the 23rd International Conference on Very Large Data Bases. 146--155."},{"key":"e_1_2_1_7_1","volume-title":"Proceedings of the 1998 ACM SIGMOD International Conference on Management of Data. 367--378","author":"Chaudhuri Surajit","unstructured":"Surajit Chaudhuri and Vivek R. Narasayya. 1998. AutoAdmin 'What-if' Index Analysis Utility. In Proceedings of the 1998 ACM SIGMOD International Conference on Management of Data. 367--378."},{"key":"e_1_2_1_8_1","volume-title":"Proceedings of the 37th International Conference on Machine Learning","volume":"119","author":"Chen Ting","unstructured":"Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey E. Hinton. 2020. A Simple Framework for Contrastive Learning of Visual Representations. In Proceedings of the 37th International Conference on Machine Learning, Vol. 119. 1597--1607."},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1002\/(SICI)1097-4571(199009)41:6<391::AID-ASI1>3.0.CO;2-9"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.findings-emnlp.139"},{"key":"e_1_2_1_11_1","volume-title":"Proceedings of the Fifth International Joint Conference on Autonomous Agents and Multiagent Systems (AAMAS '06)","author":"Fern\u00e1ndez Fernando","unstructured":"Fernando Fern\u00e1ndez and Manuela M. Veloso. 2006. Probabilistic policy reuse in a reinforcement learning agent. In Proceedings of the Fifth International Joint Conference on Autonomous Agents and Multiagent Systems (AAMAS '06). 720--727."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/509383.509385"},{"key":"e_1_2_1_13_1","volume-title":"Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR). 9726--9735","author":"He Kaiming","unstructured":"Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross B. Girshick. 2020. Momentum Contrast for Unsupervised Visual Representation Learning. In Proceedings of the IEEE\/CVF Conference on Computer Vision and Pattern Recognition (CVPR). 9726--9735."},{"key":"e_1_2_1_14_1","unstructured":"HypoPG. 2015. https:\/\/github.com\/HypoPG\/hypopg."},{"key":"e_1_2_1_15_1","volume-title":"Mohammad Zaki Zadeh, Debapriya Banerjee, and Fillia Makedon.","author":"Jaiswal Ashish","year":"2020","unstructured":"Ashish Jaiswal, Ashwin Ramesh Babu, Mohammad Zaki Zadeh, Debapriya Banerjee, and Fillia Makedon. 2020. A Survey on Contrastive Self-supervised Learning. arXiv Preprint (2020). https:\/\/arxiv.org\/abs\/2011.00362"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.14778\/3407790.3407832"},{"key":"e_1_2_1_17_1","first-page":"155","article-title":"SWIRL: Selection of Workload-aware Indexes using Reinforcement Learning","volume":"2","author":"Kossmann Jan","year":"2022","unstructured":"Jan Kossmann, Alexander Kastius, and Rainer Schlosser. 2022. SWIRL: Selection of Workload-aware Indexes using Reinforcement Learning. In EDBT. 2:155--2:168.","journal-title":"EDBT."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1145\/3340531.3412106"},{"key":"e_1_2_1_19_1","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence. 2147--2153","author":"Laroche Romain","year":"2017","unstructured":"Romain Laroche and Merwan Barlier. 2017. Transfer Reinforcement Learning with Shared Dynamics. In Proceedings of the AAAI Conference on Artificial Intelligence. 2147--2153."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.14778\/2850583.2850594"},{"key":"e_1_2_1_21_1","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence. 3562--3570","author":"Li Siyuan","year":"2018","unstructured":"Siyuan Li and Chongjie Zhang. 2018. An Optimal Online Method of Selecting Source Policies for Reinforcement Learning. In Proceedings of the AAAI Conference on Artificial Intelligence. 3562--3570."},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10489-020-01674-8"},{"key":"e_1_2_1_23_1","volume-title":"Proceedings of the 2018 International Conference on Management of Data. 631--645","author":"Ma Lin","unstructured":"Lin Ma, Dana Van Aken, Ahmed Hefny, Gustavo Mezerhane, Andrew Pavlo, and Geoffrey J. Gordon. 2018. Query-based Workload Forecasting for Self-Driving Database Management Systems. In Proceedings of the 2018 International Conference on Management of Data. 631--645."},{"key":"e_1_2_1_24_1","unstructured":"Tom\u00e1s Mikolov Ilya Sutskever Kai Chen Gregory S. Corrado and Jeffrey Dean. 2013. Distributed Representations of Words and Phrases and their Compositionality. In Advances in Neural Information Processing Systems. 3111--3119."},{"key":"e_1_2_1_25_1","volume-title":"Riedmiller","author":"Mnih Volodymyr","year":"2013","unstructured":"Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin A. Riedmiller. 2013. Playing Atari with Deep Reinforcement Learning. arXiv Preprint (2013). https:\/\/arxiv.org\/abs\/1312.5602"},{"key":"e_1_2_1_26_1","volume-title":"Proceedings of the 32nd International Conference on Very Large Data Bases. 1049--1058","author":"Nambiar Raghunath Othayoth","year":"2006","unstructured":"Raghunath Othayoth Nambiar and Meikel Poess. 2006. The Making of TPC-DS. In Proceedings of the 32nd International Conference on Very Large Data Bases. 1049--1058."},{"key":"e_1_2_1_27_1","volume-title":"2021 IEEE 37th International Conference on Data Engineering (ICDE). 600--611","author":"Perera R. Malinga","year":"2021","unstructured":"R. Malinga Perera, Bastian Oetomo, Benjamin I. P. Rubinstein, and Renata Borovica-Gajic. 2021. DBA bandits: Self-driving index tuning under ad-hoc, analytical workloads with safety guarantees. In 2021 IEEE 37th International Conference on Data Engineering (ICDE). 600--611."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/984523.984530"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1145\/369275.369291"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/3410566.3410603"},{"key":"e_1_2_1_31_1","volume-title":"2020 IEEE 36th International Conference on Data Engineering Workshops (ICDEW). 158--161","author":"Sadri Zahra","year":"2020","unstructured":"Zahra Sadri, Le Gruenwald, and Eleazar Leal. 2020. Online Index Selection Using Deep Reinforcement Learning for a Cluster Database. In 2020 IEEE 36th International Conference on Data Engineering Workshops (ICDEW). 158--161."},{"key":"e_1_2_1_32_1","volume-title":"Efficient Scalable Multi-attribute Index Selection Using Recursive Strategies. In 2019 IEEE 35th International Conference on Data Engineering (ICDE). 1238--1249","author":"Schlosser Rainer","year":"2019","unstructured":"Rainer Schlosser, Jan Kossmann, and Martin Boissier. 2019. Efficient Scalable Multi-attribute Index Selection Using Recursive Strategies. In 2019 IEEE 35th International Conference on Data Engineering (ICDE). 1238--1249."},{"key":"e_1_2_1_33_1","volume-title":"Proximal Policy Optimization Algorithms. arXiv Preprint","author":"Schulman John","year":"2017","unstructured":"John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017. Proximal Policy Optimization Algorithms. arXiv Preprint (2017). https:\/\/arxiv.org\/abs\/1707.06347"},{"key":"e_1_2_1_34_1","volume-title":"Proceedings of the 37th ACM\/SIGAPP Symposium on Applied Computing. 372--380","author":"Sharma Vishal","unstructured":"Vishal Sharma and Curtis E. Dyreson. 2022. Indexer++: workload-aware online index tuning with transformers and reinforcement learning. In Proceedings of the 37th ACM\/SIGAPP Symposium on Applied Computing. 372--380."},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/TNN.1998.712192"},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2000.839397"},{"key":"e_1_2_1_37_1","volume-title":"Index Selection in Relational Databases","author":"Whang Kyu-Young","unstructured":"Kyu-Young Whang. 1987. Index Selection in Relational Databases. In Foundations of Data Organization. 487--500."},{"key":"e_1_2_1_38_1","volume-title":"Proceedings of the 2022 International Conference on Management of Data. 1528--1541","author":"Wu Wentao","unstructured":"Wentao Wu, Chi Wang, Tarique Siddiqui, Junxiong Wang, Vivek R. Narasayya, Surajit Chaudhuri, and Philip A. Bernstein. 2022. Budget-aware Index Tuning with Reinforcement Learning. In Proceedings of the 2022 International Conference on Management of Data. 1528--1541."},{"key":"e_1_2_1_39_1","volume-title":"Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, IJCAI-20","author":"Yang Tianpei","year":"2020","unstructured":"Tianpei Yang, Jianye Hao, Zhaopeng Meng, Zongzhang Zhang, Yujing Hu, Yingfeng Chen, Changjie Fan, Weixun Wang, Wulong Liu, Zhaodong Wang, and Jiajie Peng. 2020. Efficient Deep Reinforcement Learning via Adaptive Policy Transfer. In Proceedings of the Twenty-Ninth International Joint Conference on Artificial Intelligence, IJCAI-20. 3094--3100."},{"key":"e_1_2_1_40_1","volume-title":"Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics","volume":"1","author":"Yang Zonghan","year":"2019","unstructured":"Zonghan Yang, Yong Cheng, Yang Liu, and Maosong Sun. 2019. Reducing Word Omission Errors in Neural Machine Translation: A Contrastive Learning Approach. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Vol. 1. 6191--6196."},{"key":"e_1_2_1_41_1","volume-title":"AutoIndex: An Incremental Index Management System for Dynamic Workloads. In 2022 IEEE 38th International Conference on Data Engineering (ICDE). 2196--2208","author":"Zhou Xuanhe","year":"2022","unstructured":"Xuanhe Zhou, Luyang Liu, Wenbo Li, Lianyuan Jin, Shifu Li, Tianqing Wang, and Jianhua Feng. 2022. AutoIndex: An Incremental Index Management System for Dynamic Workloads. In 2022 IEEE 38th International Conference on Data Engineering (ICDE). 2196--2208."},{"key":"e_1_2_1_42_1","volume-title":"Transfer Learning in Deep Reinforcement Learning: A Survey. arXiv Preprint","author":"Zhu Zhuangdi","year":"2020","unstructured":"Zhuangdi Zhu, Kaixiang Lin, and Jiayu Zhou. 2020. Transfer Learning in Deep Reinforcement Learning: A Survey. arXiv Preprint (2020). https:\/\/arxiv.org\/abs\/2009.07888"}],"container-title":["Proceedings of the VLDB Endowment"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.14778\/3654621.3654631","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,5,30]],"date-time":"2024-05-30T22:21:24Z","timestamp":1717107684000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.14778\/3654621.3654631"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,3]]},"references-count":42,"journal-issue":{"issue":"7","published-print":{"date-parts":[[2024,3]]}},"alternative-id":["10.14778\/3654621.3654631"],"URL":"https:\/\/doi.org\/10.14778\/3654621.3654631","relation":{},"ISSN":["2150-8097"],"issn-type":[{"value":"2150-8097","type":"print"}],"subject":[],"published":{"date-parts":[[2024,3]]},"assertion":[{"value":"2024-05-30","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}