{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,25]],"date-time":"2026-06-25T16:46:37Z","timestamp":1782405997205,"version":"3.54.5"},"reference-count":67,"publisher":"Association for Computing Machinery (ACM)","issue":"2","funder":[{"name":"Ministry of Industry and Information Technology 2024 Cloud Operating System Project, the National NSF of China","award":["62572307"],"award-info":[{"award-number":["62572307"]}]},{"name":"Shanghai Key Laboratory of Scalable Computing and Systems"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Archit. Code Optim."],"published-print":{"date-parts":[[2026,6,30]]},"abstract":"<jats:p>Multi-tenant clouds enhance resource sharing among Virtual Machines (VMs) to boost overall utilization and reduce power consumption. However, this also introduces interference among workloads from different tenants and impedes VM performance isolation. In this article, we first demonstrate that the last-level cache (LLC) in CPUs, which is inherently shared by all VMs on the same physical machine, becomes a significant contending resource for LLC-critical workloads, leading to notable performance imbalances under the default hardware caching strategy. Although recent studies on LLC scheduling have progressed, they often require detailed profiling of user workloads or rely on hyperparameter tuning, limiting their applicability to private clusters or specific scenarios. We propose SpiderSense, a software-initiated LLC partitioner for managing Virtual Machine Monitors (VMM), to address these limitations. SpiderSense leverages modern yet off-the-shelf server CPU features to adaptively orchestrate LLC allocation among running black-boxed user VMs. SpiderSense dynamically samples VMs and calculates their fair share of LLC to allocate them while fully improving performance isolation among VMs. We experiment with SpiderSense using typical LLC-critical workloads, representative of the types of applications that stress LLC performance, such as Memcached and Llama. Our results show that SpiderSense improves performance by up to 40% in numerous colocation scenarios compared to current solutions.<\/jats:p>","DOI":"10.1145\/3778358","type":"journal-article","created":{"date-parts":[[2025,12,25]],"date-time":"2025-12-25T12:07:45Z","timestamp":1766664465000},"page":"1-24","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["SpiderSense: Lightweight Last-Level Cache Management via Time Period Tagging for LLC-Critical Workloads"],"prefix":"10.1145","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0009-0009-8021-8803","authenticated-orcid":false,"given":"Zhixiang","family":"Wei","sequence":"first","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8883-5269","authenticated-orcid":false,"given":"Zhibai","family":"Huang","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-1415-0832","authenticated-orcid":false,"given":"James","family":"Yen","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0004-1035-5435","authenticated-orcid":false,"given":"Tianlei","family":"Xiong","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0001-2267-7202","authenticated-orcid":false,"given":"Kailiang","family":"Xu","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0003-5546-3168","authenticated-orcid":false,"given":"Yucheng","family":"Zheng","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-5729-3207","authenticated-orcid":false,"given":"Xingzi","family":"Yu","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5952-3426","authenticated-orcid":false,"given":"Chenggang","family":"Wu","sequence":"additional","affiliation":[{"name":"SEIEE, Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-7969-6891","authenticated-orcid":false,"given":"Yun","family":"Wang","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2730-2319","authenticated-orcid":false,"given":"Zhengwei","family":"Qi","sequence":"additional","affiliation":[{"name":"Shanghai Jiao Tong University","place":["Shanghai, China"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,6,25]]},"reference":[{"key":"e_1_3_2_2_2","doi-asserted-by":"publisher","DOI":"10.1145\/3292500.3330701"},{"key":"e_1_3_2_3_2","doi-asserted-by":"publisher","DOI":"10.3390\/pr11020349"},{"key":"e_1_3_2_4_2","doi-asserted-by":"publisher","DOI":"10.1177\/109434209100500306"},{"key":"e_1_3_2_5_2","doi-asserted-by":"publisher","DOI":"10.1109\/WCNC45663.2020.9120635"},{"key":"e_1_3_2_6_2","first-page":"1","article-title":"Enabling fair pricing on HPC systems with node sharing","author":"Breslow Alex D.","year":"2013","unstructured":"Alex D. Breslow, Ananta Tiwari, Martin Schulz, Laura Carrington, Lingjia Tang, and Jason Mars. 2013. Enabling fair pricing on HPC systems with node sharing. In Proceedings of the 2013 SC - International Conference for High Performance Computing, Networking, Storage and Analysis (SC). 1\u201312. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:629140","journal-title":"Proceedings of the 2013 SC - International Conference for High Performance Computing, Networking, Storage and Analysis (SC)"},{"key":"e_1_3_2_7_2","doi-asserted-by":"publisher","DOI":"10.1145\/1274971.1275005"},{"key":"e_1_3_2_8_2","doi-asserted-by":"publisher","DOI":"10.1145\/3719656"},{"key":"e_1_3_2_9_2","doi-asserted-by":"publisher","DOI":"10.1109\/SC41405.2020.00036"},{"key":"e_1_3_2_10_2","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2023.3303959"},{"key":"e_1_3_2_11_2","doi-asserted-by":"publisher","DOI":"10.1145\/3431379.3460648"},{"key":"e_1_3_2_12_2","doi-asserted-by":"publisher","DOI":"10.1145\/3297858.3304005"},{"key":"e_1_3_2_13_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA53966.2022.00083"},{"key":"e_1_3_2_14_2","doi-asserted-by":"publisher","DOI":"10.1145\/2556583"},{"key":"e_1_3_2_15_2","doi-asserted-by":"publisher","DOI":"10.1145\/2541940.2541941"},{"key":"e_1_3_2_16_2","volume-title":"Proceedings of the 19th USENIX Conference on Operating Systems Design and Implementation (OSDI\u201925)","author":"Domingo David","year":"2025","unstructured":"David Domingo, Hugo Barbalho, Marco Molinaro, Kuan Liu, Abhisek Pan, David Dion, Thomas Moscibroda, Sudarsun Kannan, and Ishai Menache. 2025. Kamino: Efficient VM allocation at scale with latency-driven cache-aware scheduling. In Proceedings of the 19th USENIX Conference on Operating Systems Design and Implementation (OSDI\u201925). USENIX Association, USA, Article 29, 17 pages."},{"key":"e_1_3_2_17_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2018.00019"},{"key":"e_1_3_2_18_2","doi-asserted-by":"publisher","DOI":"10.1145\/3302424.3303977"},{"key":"e_1_3_2_19_2","doi-asserted-by":"publisher","DOI":"10.1287\/educ.2018.0188"},{"key":"e_1_3_2_20_2","volume-title":"Proceedings of the USENIX Symposium on Operating Systems Design and Implementation","author":"Fried Joshua","year":"2020","unstructured":"Joshua Fried, Zhenyuan Ruan, Amy Ousterhout, and Adam Belay. 2020. Caladan: Mitigating interference at microsecond timescales. In Proceedings of the USENIX Symposium on Operating Systems Design and Implementation. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:227541336"},{"key":"e_1_3_2_21_2","first-page":"19","volume-title":"Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24)","author":"Fu Yuqi","year":"2024","unstructured":"Yuqi Fu, Ruizhe Shi, Haoliang Wang, Songqing Chen, and Yue Cheng. 2024. ALPS: An adaptive learning, priority os scheduler for serverless functions. In Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24). USENIX Association, Santa Clara, CA, 19\u201336. Retrieved from https:\/\/www.usenix.org\/conference\/atc24\/presentation\/fu"},{"key":"e_1_3_2_22_2","doi-asserted-by":"publisher","DOI":"10.1109\/JAS.2022.105845"},{"key":"e_1_3_2_23_2","unstructured":"GNU MSR Tools Package. 2013. msr-tools 1.3. Retrieved August 21 2021 from https:\/\/github.com\/intel\/msr-tools"},{"key":"e_1_3_2_24_2","doi-asserted-by":"publisher","DOI":"10.1145\/3538645"},{"key":"e_1_3_2_25_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2016.7446102"},{"key":"e_1_3_2_26_2","volume-title":"Proceedings of the International Conference on Supercomputing","author":"Iyer Ravi R.","year":"2004","unstructured":"Ravi R. Iyer. 2004. CQoS: A framework for enabling QoS in shared caches of CMP platforms. In Proceedings of the International Conference on Supercomputing. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:251503"},{"key":"e_1_3_2_27_2","doi-asserted-by":"publisher","DOI":"10.1145\/3357223.3362734"},{"key":"e_1_3_2_28_2","doi-asserted-by":"publisher","DOI":"10.1145\/3721286"},{"key":"e_1_3_2_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/3403943"},{"key":"e_1_3_2_30_2","doi-asserted-by":"publisher","DOI":"10.1145\/2541940.2541944"},{"key":"e_1_3_2_31_2","doi-asserted-by":"publisher","DOI":"10.18034\/apjee.v7i2.698"},{"key":"e_1_3_2_32_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2010.10.013"},{"key":"e_1_3_2_33_2","volume-title":"Proceedings of the USENIX Workshop on Hot Topics in Cloud Computing","author":"Lin Xing","year":"2012","unstructured":"Xing Lin, Yun Mao, Feifei Li, and Robert Ricci. 2012. Towards fair sharing of block storage in a multi-tenant cloud. In Proceedings of the USENIX Workshop on Hot Topics in Cloud Computing. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:5531028"},{"key":"e_1_3_2_34_2","doi-asserted-by":"publisher","DOI":"10.1109\/SP.2015.43"},{"key":"e_1_3_2_35_2","doi-asserted-by":"publisher","DOI":"10.1145\/3659207"},{"key":"e_1_3_2_36_2","first-page":"1","volume-title":"Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24)","author":"Liu Qingyuan","year":"2024","unstructured":"Qingyuan Liu, Yanning Yang, Dong Du, Yubin Xia, Ping Zhang, Jia Feng, James R. Larus, and Haibo Chen. 2024. Harmonizing efficiency and practicability: Optimizing resource utilization in serverless computing with JIAGU. In Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24). USENIX Association, Santa Clara, CA, 1\u201317. Retrieved from https:\/\/www.usenix.org\/conference\/atc24\/presentation\/liu-qingyuan"},{"key":"e_1_3_2_37_2","doi-asserted-by":"publisher","DOI":"10.1145\/3267809.3267830"},{"key":"e_1_3_2_38_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA56546.2023.10071128"},{"key":"e_1_3_2_39_2","doi-asserted-by":"publisher","DOI":"10.1145\/2749469.2749475"},{"key":"e_1_3_2_40_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA57654.2024.00090"},{"key":"e_1_3_2_41_2","first-page":"37","volume-title":"Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24)","author":"Luo Michael","year":"2024","unstructured":"Michael Luo, Siyuan Zhuang, Suryaprakash Vengadesan, Romil Bhardwaj, Justin Chang, Eric Friedman, Scott Shenker, and Ion Stoica. 2024. Starburst: A cost-aware scheduler for hybrid cloud. In Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24). USENIX Association, Santa Clara, CA, 37\u201357. Retrieved from https:\/\/www.usenix.org\/conference\/atc24\/presentation\/luo"},{"key":"e_1_3_2_42_2","first-page":"1","article-title":"On the usage, co-usage and migration of CI\/CD tools: A qualitative analysis","volume":"28","author":"Mazrae Pooya Rostami","year":"2023","unstructured":"Pooya Rostami Mazrae, Tom Mens, Mehdi Golzadeh, and Alexandre Decan. 2023. On the usage, co-usage and migration of CI\/CD tools: A qualitative analysis. Empirical Software Engineering 28, Article no. 52 (2023), 1\u201345. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:257367999","journal-title":"Empirical Software Engineering"},{"key":"e_1_3_2_43_2","volume-title":"Slides from Linux Kongress","author":"Melo AC De","year":"2010","unstructured":"AC De Melo. 2010. The new linux\u2019perf\u2019 tools. In Slides from Linux Kongress. Retrieved from vger.kernel.org"},{"key":"e_1_3_2_44_2","doi-asserted-by":"publisher","DOI":"10.1145\/1755913.1755938"},{"key":"e_1_3_2_45_2","doi-asserted-by":"publisher","DOI":"10.1145\/3337821.3337891"},{"key":"e_1_3_2_46_2","volume-title":"Proceedings of the Symposium on Networked Systems Design and Implementation","author":"Ousterhout Amy","year":"2019","unstructured":"Amy Ousterhout, Joshua Fried, Jonathan Behrens, Adam Belay, and H. Balakrishnan. 2019. Shenango: Achieving high CPU efficiency for latency-sensitive datacenter workloads. In Proceedings of the Symposium on Networked Systems Design and Implementation. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:73725100"},{"key":"e_1_3_2_47_2","doi-asserted-by":"publisher","DOI":"10.1145\/3302424.3303963"},{"key":"e_1_3_2_48_2","article-title":"Hypart: A hybrid technique for practical memory bandwidth partitioning on commodity servers","author":"Park Jinsu","year":"2018","unstructured":"Jinsu Park, Seongbeom Park, Myeonggyun Han, Jihoon Hyun, and Woongki Baek. 2018. Hypart: A hybrid technique for practical memory bandwidth partitioning on commodity servers. In Proceedings of the 27th International Conference on Parallel Architectures and Compilation Techniques. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:52964500","journal-title":"Proceedings of the 27th International Conference on Parallel Architectures and Compilation Techniques"},{"key":"e_1_3_2_49_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA47549.2020.00025"},{"key":"e_1_3_2_50_2","doi-asserted-by":"publisher","DOI":"10.1109\/ISCA52012.2021.00031"},{"key":"e_1_3_2_51_2","doi-asserted-by":"publisher","DOI":"10.1145\/3472883.3487006"},{"key":"e_1_3_2_52_2","first-page":"403","volume-title":"Proceedings of the 2023 USENIX Annual Technical Conference (USENIX ATC 23)","author":"Shi Jiuchen","year":"2023","unstructured":"Jiuchen Shi, Hang Zhang, Zhixin Tong, Quan Chen, Kaihua Fu, and Minyi Guo. 2023. Nodens: Enabling resource efficient and fast QoS recovery of dynamic microservice applications in datacenters. In Proceedings of the 2023 USENIX Annual Technical Conference (USENIX ATC 23). USENIX Association, Boston, MA, 403\u2013417. Retrieved from https:\/\/www.usenix.org\/conference\/atc23\/presentation\/shi"},{"key":"e_1_3_2_53_2","doi-asserted-by":"publisher","DOI":"10.1145\/3534879.3534882"},{"key":"e_1_3_2_54_2","doi-asserted-by":"publisher","DOI":"10.1145\/3545008.3545064"},{"key":"e_1_3_2_55_2","unstructured":"Hugo Touvron Louis Martin Kevin R. Stone Peter Albert Amjad Almahairi Yasmine Babaei Nikolay Bashlykov Soumya Batra Prajjwal Bhargava Shruti Bhosale et\u00a0al. 2023. Llama 2: Open foundation and fine-tuned chat models. arXiv:2307.09288. Retrieved from https:\/\/arxiv.org\/abs\/2307.09288. https:\/\/api.semanticscholar.org\/CorpusID:259950998"},{"key":"e_1_3_2_56_2","doi-asserted-by":"publisher","DOI":"10.1109\/ICCV.2019.00167"},{"key":"e_1_3_2_57_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA.2017.65"},{"key":"e_1_3_2_58_2","first-page":"59","volume-title":"Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24)","author":"Wu Hao","year":"2024","unstructured":"Hao Wu, Yue Yu, Junxiao Deng, Shadi Ibrahim, Song Wu, Hao Fan, Ziyue Cheng, and Hai Jin. 2024. StreamBox: A lightweight GPU SandBox for serverless inference workflow. In Proceedings of the 2024 USENIX Annual Technical Conference (USENIX ATC 24). USENIX Association, Santa Clara, CA, 59\u201373. Retrieved from https:\/\/www.usenix.org\/conference\/atc24\/presentation\/wu-hao"},{"key":"e_1_3_2_59_2","doi-asserted-by":"publisher","DOI":"10.1145\/3190508.3190511"},{"key":"e_1_3_2_60_2","doi-asserted-by":"publisher","DOI":"10.1145\/3190508.3190555"},{"key":"e_1_3_2_61_2","doi-asserted-by":"publisher","DOI":"10.1145\/3711875.3729163"},{"key":"e_1_3_2_62_2","doi-asserted-by":"publisher","DOI":"10.1145\/3617232.3624871"},{"key":"e_1_3_2_63_2","doi-asserted-by":"publisher","DOI":"10.1145\/3419111.3421280"},{"key":"e_1_3_2_64_2","doi-asserted-by":"publisher","DOI":"10.1145\/3445814.3446693"},{"key":"e_1_3_2_65_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA57654.2024.00077"},{"key":"e_1_3_2_66_2","article-title":"Dirigent: Enforcing QoS for latency-critical tasks on shared multicore systems","author":"Zhu Haishan","year":"2016","unstructured":"Haishan Zhu and Mattan Erez. 2016. Dirigent: Enforcing QoS for latency-critical tasks on shared multicore systems. In Proceedings of the 21st International Conference on Architectural Support for Programming Languages and Operating Systems. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:12021213","journal-title":"Proceedings of the 21st International Conference on Architectural Support for Programming Languages and Operating Systems"},{"key":"e_1_3_2_67_2","unstructured":"Hang Zhu Kostis Kaffes Zixu Chen Zhenming Liu Christos Kozyrakis Ion Stoica and Xin Jin. 2020. RackSched: A microsecond-scale scheduler for rack-scale computers (technical report). arXiv:2010.05969. Retrieved from https:\/\/arxiv.org\/abs\/2010.05969. Retrieved from https:\/\/api.semanticscholar.org\/CorpusID:222310459"},{"key":"e_1_3_2_68_2","doi-asserted-by":"publisher","DOI":"10.1109\/HPCA61900.2025.00042"}],"container-title":["ACM Transactions on Architecture and Code Optimization"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3778358","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,25]],"date-time":"2026-06-25T15:56:23Z","timestamp":1782402983000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3778358"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,6,25]]},"references-count":67,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2026,6,30]]}},"alternative-id":["10.1145\/3778358"],"URL":"https:\/\/doi.org\/10.1145\/3778358","relation":{},"ISSN":["1544-3566","1544-3973"],"issn-type":[{"value":"1544-3566","type":"print"},{"value":"1544-3973","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,6,25]]},"assertion":[{"value":"2025-05-16","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2025-11-15","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2026-06-25","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}