{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,12,20]],"date-time":"2025-12-20T22:24:19Z","timestamp":1766269459732,"version":"3.41.0"},"reference-count":108,"publisher":"Association for Computing Machinery (ACM)","issue":"1-2","license":[{"start":{"date-parts":[[2024,2,13]],"date-time":"2024-02-13T00:00:00Z","timestamp":1707782400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"National Key Research and Development Program of China","award":["2022YFB4500702"],"award-info":[{"award-number":["2022YFB4500702"]}]},{"DOI":"10.13039\/501100007129","name":"Shandong Provincial Natural Science Foundation","doi-asserted-by":"crossref","award":["ZR2022LZH018"],"award-info":[{"award-number":["ZR2022LZH018"]}],"id":[{"id":"10.13039\/501100007129","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62141218, 62372322"],"award-info":[{"award-number":["62141218, 62372322"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"name":"Zhejiang Lab","award":["2021DA0AM01\/003"],"award-info":[{"award-number":["2021DA0AM01\/003"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Comput. Syst."],"published-print":{"date-parts":[[2024,5,31]]},"abstract":"<jats:p>Cloud service providers improve resource utilization by co-locating latency-critical (LC) workloads with best-effort batch (BE) jobs in datacenters. However, they usually treat multi-component LCs as monolithic applications and treat BEs as \u201csecond-class citizens\u201d when allocating resources to them. Neglecting the inconsistent interference tolerance abilities of LC components and the inconsistent preemption loss of BE workloads can result in missed co-location opportunities for higher throughput.<\/jats:p>\n          <jats:p>\n            We present\n            <jats:monospace>Rhythm<\/jats:monospace>\n            , a co-location controller that deploys workloads and reclaims resources rhythmically for maximizing the system throughput while guaranteeing LC service\u2019s tail latency requirement. The key idea is to differentiate the BE throughput launched with each LC component, that is, components with higher interference tolerance can be deployed together with more BE jobs. It also assigns different reclamation priority values to BEs by evaluating their preemption losses into a multi-level reclamation queue. We implement and evaluate\n            <jats:monospace>Rhythm<\/jats:monospace>\n            using workloads in the form of containerized processes and microservices. Experimental results show that it can improve the system throughput by 47.3%, CPU utilization by 38.6%, and memory bandwidth utilization by 45.4% while guaranteeing the tail latency requirement.\n          <\/jats:p>","DOI":"10.1145\/3630006","type":"journal-article","created":{"date-parts":[[2023,11,18]],"date-time":"2023-11-18T11:28:07Z","timestamp":1700306887000},"page":"1-37","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Component-distinguishable Co-location and Resource Reclamation for High-throughput Computing"],"prefix":"10.1145","volume":"42","author":[{"ORCID":"https:\/\/orcid.org\/0000-0003-1967-2192","authenticated-orcid":false,"given":"Laiping","family":"Zhao","sequence":"first","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0009-9413-5187","authenticated-orcid":false,"given":"Yushuai","family":"Cui","sequence":"additional","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-2222-393X","authenticated-orcid":false,"given":"Yanan","family":"Yang","sequence":"additional","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-7772-458X","authenticated-orcid":false,"given":"Xiaobo","family":"Zhou","sequence":"additional","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-2324-2523","authenticated-orcid":false,"given":"Tie","family":"Qiu","sequence":"additional","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1758-3030","authenticated-orcid":false,"given":"Keqiu","family":"Li","sequence":"additional","affiliation":[{"name":"College of Intelligence and Computing, Tianjin University, Tianjin Key Lab. of Advanced Networking, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6565-5276","authenticated-orcid":false,"given":"Yungang","family":"Bao","sequence":"additional","affiliation":[{"name":"Inst. of Computing Technology, CAS, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,2,13]]},"reference":[{"unstructured":"2020. https:\/\/parsec.cs.princeton.edu\/","key":"e_1_3_1_2_2"},{"unstructured":"2020. Scimark: A benchmark for scientific and numerical computing. https:\/\/openbenchmarking.org\/test\/pts\/scimark2-1.3.2","key":"e_1_3_1_3_2"},{"unstructured":"2020. The SPEC Cloud IaaS 2018 benchmark is SPEC\u2019s second benchmark suite to measure cloud performance.https:\/\/www.spec.org\/","key":"e_1_3_1_4_2"},{"unstructured":"2020. Tensorflow-Bench: A benchmark framework for TensorFlow.https:\/\/github.com\/tensorflow\/benchmarks","key":"e_1_3_1_5_2"},{"key":"e_1_3_1_6_2","doi-asserted-by":"crossref","first-page":"749","DOI":"10.1109\/DSN.2007.38","volume-title":"The 37th Annual IEEE\/IFIP International Conference on Dependable Systems and Networks (DSN\u201907)","author":"Agarwala S.","year":"2007","unstructured":"S. Agarwala, F. Alegre, K. Schwan, and J. Mehalingham. 2007. E2EProf: Automated end-to-end performance management for enterprise systems. In The 37th Annual IEEE\/IFIP International Conference on Dependable Systems and Networks (DSN\u201907). 749\u2013758."},{"key":"e_1_3_1_7_2","series-title":"Proceedings of the Nineteenth ACM Symposium on Operating Systems Principles","first-page":"74","author":"Aguilera Marcos K.","year":"2003","unstructured":"Marcos K. Aguilera, Jeffrey C. Mogul, Janet L. Wiener, Patrick Reynolds, and Athicha Muthitacharoen. 2003. Performance debugging for distributed systems of black boxes. In Proceedings of the Nineteenth ACM Symposium on Operating Systems Principles (Bolton Landing, NY, USA) (SOSP \u201903). Association for Computing Machinery, New York, NY, USA, 74\u201389."},{"key":"e_1_3_1_8_2","volume-title":"Providing SLOs for Resource-Harvesting VMs in Cloud Platforms","author":"Ambati Pradeep","year":"2020","unstructured":"Pradeep Ambati, \u00cd\u00f1igo Goiri, Felipe Frujeri, Alper Gun, Ke Wang, Brian Dolan, Brian Corell, Sekhar Pasupuleti, Thomas Moscibroda, Sameh Elnikety, Marcus Fontoura, and Ricardo Bianchini. 2020. Providing SLOs for Resource-Harvesting VMs in Cloud Platforms. USENIX Association, USA."},{"key":"e_1_3_1_9_2","series-title":"Proceedings of the 24th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval","first-page":"276","author":"Aslam Javed A.","year":"2001","unstructured":"Javed A. Aslam and Mark Montague. 2001. Models for metasearch. In Proceedings of the 24th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval (New Orleans, Louisiana, USA) (SIGIR \u201901). Association for Computing Machinery, New York, NY, USA, 276\u2013284."},{"key":"e_1_3_1_10_2","first-page":"275","volume-title":"SIGIR 2001: Proceedings of the 24th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, September 9-13, 2001, New Orleans, Louisiana, USA","author":"Aslam Javed A.","year":"2001","unstructured":"Javed A. Aslam and Mark H. Montague. 2001. Models for metasearch. In SIGIR 2001: Proceedings of the 24th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, September 9-13, 2001, New Orleans, Louisiana, USA, W. Bruce Croft, David J. Harper, Donald H. Kraft, and Justin Zobel (Eds.). ACM, 275\u2013284."},{"key":"e_1_3_1_11_2","first-page":"1","volume-title":"Constellation: Automated Discovery of Service and Host Dependencies in Networked Systems","author":"Barham Paul","year":"2008","unstructured":"Paul Barham, Richard Black, Moises Goldszmidt, Rebecca Isaacs, John MacCormick, Richard Mortier, and Aleksandr Simma. 2008. Constellation: Automated Discovery of Service and Host Dependencies in Networked Systems. Technical Report MSR-TR-2008-67. 1\u201314 pages."},{"key":"e_1_3_1_12_2","first-page":"259","volume-title":"Proceedings of the Sixth USENIX Symposium on Operating Systems Design and Implementation (OSDI) 2004","author":"Barham Paul","year":"2004","unstructured":"Paul Barham, Austin Donnelly, Rebecca Isaacs, and Richard Mortier. 2004. Using Magpie for request extraction and workload modelling. In Proceedings of the Sixth USENIX Symposium on Operating Systems Design and Implementation (OSDI) 2004. 259\u2013272."},{"key":"e_1_3_1_13_2","series-title":"Proceedings of the First Annual ACM SIGMM Conference on Multimedia Systems","first-page":"35","author":"Barker Sean Kenneth","year":"2010","unstructured":"Sean Kenneth Barker and Prashant Shenoy. 2010. Empirical evaluation of latency-sensitive application performance in the cloud. In Proceedings of the First Annual ACM SIGMM Conference on Multimedia Systems (Phoenix, Arizona, USA) (MMSys \u201910). ACM, New York, NY, USA, 35\u201346."},{"key":"e_1_3_1_14_2","first-page":"1","volume-title":"2008 IEEE International Symposium on Parallel and Distributed Processing","author":"Benoit Anne","year":"2008","unstructured":"Anne Benoit, Mourad Hakem, and Yves Robert. 2008. Fault tolerant scheduling of precedence task graphs on heterogeneous platforms. In 2008 IEEE International Symposium on Parallel and Distributed Processing. 1\u20138."},{"issue":"2","key":"e_1_3_1_15_2","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1016\/j.parco.2008.11.001","article-title":"Contention awareness and fault-tolerant scheduling for precedence constrained tasks in heterogeneous systems","volume":"35","author":"Benoit Anne","year":"2009","unstructured":"Anne Benoit, Mourad Hakem, and Yves Robert. 2009. Contention awareness and fault-tolerant scheduling for precedence constrained tasks in heterogeneous systems. Parallel Comput. 35, 2 (2009), 83\u2013108.","journal-title":"Parallel Comput."},{"issue":"2","key":"e_1_3_1_16_2","doi-asserted-by":"crossref","first-page":"214","DOI":"10.1109\/TSC.2016.2607739","article-title":"CauseInfer: Automated end-to-end performance diagnosis with hierarchical causality graph in cloud environment","volume":"12","author":"Chen P.","year":"2019","unstructured":"P. Chen, Y. Qi, and D. Hou. 2019. CauseInfer: Automated end-to-end performance diagnosis with hierarchical causality graph in cloud environment. IEEE Transactions on Services Computing 12, 2 (March2019), 214\u2013230.","journal-title":"IEEE Transactions on Services Computing"},{"key":"e_1_3_1_17_2","series-title":"Proceedings of the Twenty-Fourth International Conference on Architectural Support for Programming Languages and Operating Systems","first-page":"107","author":"Chen Shuang","year":"2019","unstructured":"Shuang Chen, Christina Delimitrou, and Jos\u00e9 F. Mart\u00ednez. 2019. PARTIES: QoS-aware resource partitioning for multiple interactive services. In Proceedings of the Twenty-Fourth International Conference on Architectural Support for Programming Languages and Operating Systems (Providence, RI, USA) (ASPLOS \u201919). ACM, New York, NY, USA, 107\u2013120."},{"key":"e_1_3_1_18_2","series-title":"Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation","first-page":"117","author":"Chen Xu","year":"2008","unstructured":"Xu Chen, Ming Zhang, Z. Morley Mao, and Paramvir Bahl. 2008. Automating network application dependency discovery: Experiences, limitations, and new solutions. In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (San Diego, California) (OSDI\u201908). USENIX Association, USA, 117\u2013130."},{"key":"e_1_3_1_19_2","first-page":"217","volume-title":"11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14)","author":"Chow Michael","year":"2014","unstructured":"Michael Chow, David Meisner, Jason Flinn, Daniel Peek, and Thomas F. Wenisch. 2014. The mystery machine: End-to-end performance analysis of large-scale internet services. In 11th USENIX Symposium on Operating Systems Design and Implementation (OSDI 14). USENIX Association, Broomfield, CO, 217\u2013231."},{"unstructured":"The Internet Traffic Archive ClarkNet. 2017. http:\/\/ita.ee.lbl.gov\/html\/traces.html","key":"e_1_3_1_20_2"},{"key":"e_1_3_1_21_2","first-page":"308","volume-title":"ACM SIGARCH Computer Architecture News","author":"Cook Henry","year":"2013","unstructured":"Henry Cook, Miquel Moreto, Sarah Bird, Khanh Dao, David A. Patterson, and Krste Asanovic. 2013. A hardware evaluation of cache partitioning to improve utilization and energy-efficiency while preserving responsiveness. In ACM SIGARCH Computer Architecture News, Vol. 41. ACM, 308\u2013319."},{"issue":"2","key":"e_1_3_1_22_2","doi-asserted-by":"crossref","first-page":"74","DOI":"10.1145\/2408776.2408794","article-title":"The tail at scale","volume":"56","author":"Dean Jeffrey","year":"2013","unstructured":"Jeffrey Dean and Luiz Andr\u00e9 Barroso. 2013. The tail at scale. Commun. ACM 56, 2 (Feb.2013), 74\u201380.","journal-title":"Commun. ACM"},{"unstructured":"DeathStarBench. 2019. https:\/\/github.com\/delimitrou\/DeathStarBench","key":"e_1_3_1_23_2"},{"key":"e_1_3_1_24_2","doi-asserted-by":"crossref","first-page":"23","DOI":"10.1109\/IISWC.2013.6704667","volume-title":"2013 IEEE International Symposium on Workload Characterization (IISWC)","author":"Delimitrou Christina","year":"2013","unstructured":"Christina Delimitrou and Christos Kozyrakis. 2013. ibench: Quantifying interference for datacenter applications. In 2013 IEEE International Symposium on Workload Characterization (IISWC). IEEE, 23\u201333."},{"key":"e_1_3_1_25_2","first-page":"77","volume-title":"ACM SIGPLAN Notices","author":"Delimitrou Christina","year":"2013","unstructured":"Christina Delimitrou and Christos Kozyrakis. 2013. Paragon: QoS-aware scheduling for heterogeneous datacenters. In ACM SIGPLAN Notices, Vol. 48. ACM, 77\u201388."},{"issue":"4","key":"e_1_3_1_26_2","doi-asserted-by":"crossref","first-page":"127","DOI":"10.1145\/2644865.2541941","article-title":"Quasar: Resource-efficient and QoS-aware cluster management","volume":"49","author":"Delimitrou Christina","year":"2014","unstructured":"Christina Delimitrou and Christos Kozyrakis. 2014. Quasar: Resource-efficient and QoS-aware cluster management. ACM SIGPLAN Notices 49, 4 (2014), 127\u2013144.","journal-title":"ACM SIGPLAN Notices"},{"issue":"4","key":"e_1_3_1_27_2","doi-asserted-by":"crossref","first-page":"473","DOI":"10.1145\/2954679.2872365","article-title":"HCloud: Resource-efficient provisioning in shared cloud systems","volume":"51","author":"Delimitrou Christina","year":"2016","unstructured":"Christina Delimitrou and Christos Kozyrakis. 2016. HCloud: Resource-efficient provisioning in shared cloud systems. SIGPLAN Not. 51, 4 (March2016), 473\u2013488.","journal-title":"SIGPLAN Not."},{"unstructured":"Elasticsearch. 2021. Elasticsearch: A search engine based on the Lucene library. https:\/\/lucene.apache.org\/solr\/","key":"e_1_3_1_28_2"},{"issue":"3","key":"e_1_3_1_29_2","doi-asserted-by":"crossref","first-page":"375","DOI":"10.1145\/568522.568525","article-title":"A survey of rollback-recovery protocols in message-passing systems","volume":"34","author":"Elnozahy E. N. (Mootaz)","year":"2002","unstructured":"E. N. (Mootaz) Elnozahy, Lorenzo Alvisi, Yi-Min Wang, and David B. Johnson. 2002. A survey of rollback-recovery protocols in message-passing systems. ACM Comput. Surv. 34, 3 (Sept.2002), 375\u2013408.","journal-title":"ACM Comput. Surv."},{"issue":"4","key":"e_1_3_1_30_2","doi-asserted-by":"crossref","first-page":"37","DOI":"10.1145\/2248487.2150982","article-title":"Clearing the clouds: A study of emerging scale-out workloads on modern hardware","volume":"47","author":"Ferdman Michael","year":"2012","unstructured":"Michael Ferdman, Almutaz Adileh, Onur Kocberber, Stavros Volos, Mohammad Alisafaee, Djordje Jevdjic, Cansu Kaynak, Adrian Daniel Popescu, Anastasia Ailamaki, and Babak Falsafi. 2012. Clearing the clouds: A study of emerging scale-out workloads on modern hardware. SIGPLAN Not. 47, 4 (March2012), 37\u201348.","journal-title":"SIGPLAN Not."},{"key":"e_1_3_1_31_2","volume-title":"4th USENIX Symposium on Networked Systems Design & Implementation (NSDI 07)","author":"Fonseca Rodrigo","year":"2007","unstructured":"Rodrigo Fonseca, George Porter, Randy H. Katz, and Scott Shenker. 2007. X-trace: A pervasive network tracing framework. In 4th USENIX Symposium on Networked Systems Design & Implementation (NSDI 07). USENIX Association, Cambridge, MA."},{"issue":"2","key":"e_1_3_1_32_2","doi-asserted-by":"crossref","first-page":"155","DOI":"10.1109\/LCA.2018.2839189","article-title":"The architectural implications of cloud microservices","volume":"17","author":"Gan Yu","year":"2018","unstructured":"Yu Gan and Christina Delimitrou. 2018. The architectural implications of cloud microservices. IEEE Computer Architecture Letters 17, 2 (July2018), 155\u2013158.","journal-title":"IEEE Computer Architecture Letters"},{"key":"e_1_3_1_33_2","series-title":"Proceedings of the Twenty-Fourth International Conference on Architectural Support for Programming Languages and Operating Systems","first-page":"3","author":"Gan Yu","year":"2019","unstructured":"Yu Gan, Yanqi Zhang, Dailun Cheng, Ankitha Shetty, Priyal Rathi, Nayan Katarki, Ariana Bruno, Justin Hu, Brian Ritchken, Brendon Jackson, Kelvin Hu, Meghna Pancholi, Yuan He, Brett Clancy, Chris Colen, Fukang Wen, Catherine Leung, Siyuan Wang, Leon Zaruvinsky, Mateo Espinosa, Rick Lin, Zhongling Liu, Jake Padilla, and Christina Delimitrou. 2019. An open-source benchmark suite for microservices and their hardware-software implications for cloud & edge systems. In Proceedings of the Twenty-Fourth International Conference on Architectural Support for Programming Languages and Operating Systems (Providence, RI, USA) (ASPLOS \u201919). ACM, New York, NY, USA, 3\u201318."},{"issue":"73","key":"e_1_3_1_34_2","doi-asserted-by":"crossref","first-page":"2013","DOI":"10.1007\/s11538-010-9597-1","article-title":"Law of the minimum paradoxes","author":"Gorban Alexander N.","year":"2011","unstructured":"Alexander N. Gorban, Lyudmila I. Pokidysheva, Elena V. Smirnova, and Tatiana A. Tyukina. 2011. Law of the minimum paradoxes. Bulletin of Mathematical Biology73 (2011), 2013\u20132044.","journal-title":"Bulletin of Mathematical Biology"},{"doi-asserted-by":"crossref","unstructured":"Sriram Govindan Jie Liu Aman Kansal and Anand Sivasubramaniam. 2011. Cuanta: Quantifying effects of shared on-chip resource interference for consolidated virtual machines. In Proceedings of the 2nd ACM Symposium on Cloud Computing (Cascais Portugal) (SOCC \u201911). ACM New York NY USA 22:1\u201322:14.","key":"e_1_3_1_35_2","DOI":"10.1145\/2038916.2038938"},{"doi-asserted-by":"crossref","unstructured":"Jing Guo Zihao Chang Sa Wang Haiyang Ding Yihui Feng Liang Mao and Yungang Bao. 2019. Who limits the resource efficiency of my datacenter: An analysis of Alibaba datacenter traces. In Proceedings of the International Symposium on Quality of Service (Phoenix Arizona) (IWQoS \u201919). ACM New York NY USA Article 39 10 pages.","key":"e_1_3_1_36_2","DOI":"10.1145\/3326285.3329074"},{"unstructured":"HiBench. 2020. HiBench: HiBench is a big data benchmark suite that helps evaluate different big data frameworks in terms of speed throughput and system resource utilizations.https:\/\/github.com\/Intel-bigdata\/HiBench","key":"e_1_3_1_37_2"},{"key":"e_1_3_1_38_2","first-page":"519","volume-title":"2018 USENIX Annual Technical Conference (USENIX ATC 18)","author":"Iorgulescu Calin","year":"2018","unstructured":"Calin Iorgulescu, Reza Azimi, Youngjin Kwon, Sameh Elnikety, Manoj Syamala, Vivek Narasayya, Herodotos Herodotou, Paulo Tomita, Alex Chen, Jack Zhang, and Junhua Wang. 2018. PerfIso: Performance isolation for commercial latency-sensitive services. In 2018 USENIX Annual Technical Conference (USENIX ATC 18). USENIX Association, Boston, MA, 519\u2013532."},{"key":"e_1_3_1_39_2","doi-asserted-by":"crossref","first-page":"104","DOI":"10.1109\/CCGrid.2011.22","volume-title":"2011 11th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing","author":"Iosup Alexandru","year":"2011","unstructured":"Alexandru Iosup, Nezih Yigitbasi, and Dick Epema. 2011. On the performance variability of production cloud services. In 2011 11th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing. 104\u2013113."},{"key":"e_1_3_1_40_2","first-page":"257","volume-title":"International Conference on Supercomputing","author":"Iyer Ravi R.","year":"2004","unstructured":"Ravi R. Iyer. 2004. CQoS: A framework for enabling QoS in shared caches of CMP platforms. In International Conference on Supercomputing. 257\u2013266."},{"key":"e_1_3_1_41_2","article-title":"SystemTap: Instrumenting the Linux kernel for analyzing performance and functional problems","author":"Jacob Bart","year":"2008","unstructured":"Bart Jacob, Paul Larson, B. Leitao, and S. A. M. M. da Silva. 2008. SystemTap: Instrumenting the Linux kernel for analyzing performance and functional problems. IBM Redbook (2008).","journal-title":"IBM Redbook"},{"unstructured":"jaeger. 2019. https:\/\/www.jaegertracing.io\/","key":"e_1_3_1_42_2"},{"key":"e_1_3_1_43_2","first-page":"850","volume-title":"DAC Design Automation Conference 2012","author":"Jeong M. K.","year":"2012","unstructured":"M. K. Jeong, M. Erez, C. Sudanthi, and N. Paver. 2012. A QoS-aware memory controller for dynamically balancing GPU and CPU bandwidth use in an MPSoC. In DAC Design Automation Conference 2012. 850\u2013855."},{"key":"e_1_3_1_44_2","first-page":"34","volume-title":"Proceedings of the 26th Symposium on Operating Systems Principles, Shanghai, China, October 28-31, 2017","author":"Kaldor Jonathan","year":"2017","unstructured":"Jonathan Kaldor, Jonathan Mace, Michal Bejda, Edison Gao, Wiktor Kuropatwa, Joe O\u2019Neill, Kian Win Ong, Bill Schaller, Pingjia Shan, Brendan Viscomi, Vinod Venkataraman, Kaushik Veeraraghavan, and Yee Jiun Song. 2017. Canopy: An end-to-end performance tracing and analysis system. In Proceedings of the 26th Symposium on Operating Systems Principles, Shanghai, China, October 28-31, 2017. ACM, 34\u201350."},{"key":"e_1_3_1_45_2","first-page":"1","volume-title":"High Performance Computing, Networking, Storage and Analysis","author":"Kambadur M.","year":"2012","unstructured":"M. Kambadur, T. Moseley, R. Hank, and M. A. Kim. 2012. Measuring interference between live datacenter applications. In High Performance Computing, Networking, Storage and Analysis. 1\u201312."},{"doi-asserted-by":"crossref","unstructured":"Ram Srivatsa Kannan Lavanya Subramanian Ashwin Raju Jeongseob Ahn Jason Mars and Lingjia Tang. 2019. GrandSLAm: Guaranteeing SLAs for jobs in microservices execution frameworks. In Proceedings of the Fourteenth EuroSys Conference 2019 (Dresden Germany) (EuroSys \u201919). Association for Computing Machinery New York NY USA Article 34 16 pages.","key":"e_1_3_1_46_2","DOI":"10.1145\/3302424.3303958"},{"key":"e_1_3_1_47_2","first-page":"729","volume-title":"ACM SIGPLAN Notices","author":"Kasture Harshad","year":"2014","unstructured":"Harshad Kasture and Daniel Sanchez. 2014. Ubik: Efficient cache sharing with strict QoS for latency-critical workloads. In ACM SIGPLAN Notices, Vol. 49. ACM, 729\u2013742."},{"key":"e_1_3_1_48_2","doi-asserted-by":"crossref","first-page":"703","DOI":"10.1145\/2488388.2488450","volume-title":"Proceedings of the 22nd International Conference on World Wide Web","author":"Krushevskaja Darja","year":"2013","unstructured":"Darja Krushevskaja and Mark Sandler. 2013. Understanding latency variations of black box services. In Proceedings of the 22nd International Conference on World Wide Web. ACM, 703\u2013714."},{"unstructured":"Kubernetes. 2019. https:\/\/kubernetes.io\/","key":"e_1_3_1_49_2"},{"key":"e_1_3_1_50_2","volume-title":"Proceedings of the 30th International Symposium on High-Performance Parallel and Distributed Computing (HPDC \u201921)","author":"Li Wenyu Qu Kunlin Zhan Laiping Zhao, Fangshu","year":"2021","unstructured":"Wenyu Qu Kunlin Zhan Laiping Zhao, Fangshu Li and Qingman Zhang. 2021. AITurbo: Unified compute allocation for partial predictable training in commodity clusters. In Proceedings of the 30th International Symposium on High-Performance Parallel and Distributed Computing (HPDC \u201921). ACM, USA."},{"key":"e_1_3_1_51_2","series-title":"Proceedings of the ACM Symposium on Cloud Computing","first-page":"347","author":"Liu Qixiao","year":"2018","unstructured":"Qixiao Liu and Zhibin Yu. 2018. The elasticity and plasticity in semi-containerized co-locating cloud workload: A view from Alibaba trace. In Proceedings of the ACM Symposium on Cloud Computing (Carlsbad, CA, USA) (SoCC \u201918). ACM, New York, NY, USA, 347\u2013360."},{"key":"e_1_3_1_52_2","first-page":"450","volume-title":"ACM SIGARCH Computer Architecture News","author":"Lo David","year":"2015","unstructured":"David Lo, Liqun Cheng, Rama Govindaraju, Parthasarathy Ranganathan, and Christos Kozyrakis. 2015. Heracles: Improving resource efficiency at scale. In ACM SIGARCH Computer Architecture News, Vol. 43. ACM, 450\u2013462."},{"unstructured":"LTTng. 2019. https:\/\/lttng.org\/","key":"e_1_3_1_53_2"},{"unstructured":"Piotr Luszczek Jack J. Dongarra David Koester Rolf Rabenseifner Bob Lucas Jeremy Kepner John McCalpin David Bailey and Daisuke Takahashi. [n.d.]. Introduction to the HPC challenge benchmark suite. ([n. d.]).","key":"e_1_3_1_54_2"},{"issue":"1","key":"e_1_3_1_55_2","doi-asserted-by":"crossref","first-page":"131","DOI":"10.1145\/2786763.2694382","article-title":"Supporting differentiated services in computers via programmable architecture for resourcing-on-demand (PARD)","volume":"43","author":"Ma Jiuyue","year":"2015","unstructured":"Jiuyue Ma, Xiufeng Sui, Ninghui Sun, Yupeng Li, Zihao Yu, Bowen Huang, Tianni Xu, Zhicheng Yao, Yun Chen, Haibin Wang, Lixin Zhang, and Yungang Bao. 2015. Supporting differentiated services in computers via programmable architecture for resourcing-on-demand (PARD). SIGARCH Comput. Archit. News 43, 1 (March2015), 131\u2013143.","journal-title":"SIGARCH Comput. Archit. News"},{"key":"e_1_3_1_56_2","first-page":"589","volume-title":"12th USENIX Symposium on Networked Systems Design and Implementation (NSDI 15)","author":"Mace Jonathan","year":"2015","unstructured":"Jonathan Mace, Peter Bodik, Rodrigo Fonseca, and Madanlal Musuvathi. 2015. Retro: Targeted resource management in multi-tenant distributed systems. In 12th USENIX Symposium on Networked Systems Design and Implementation (NSDI 15). USENIX Association, Oakland, CA, 589\u2013603."},{"key":"e_1_3_1_57_2","doi-asserted-by":"crossref","first-page":"91","DOI":"10.1109\/ICAC.2015.48","volume-title":"2015 IEEE International Conference on Autonomic Computing","author":"Maji A. K.","year":"2015","unstructured":"A. K. Maji, S. Mitra, and S. Bagchi. 2015. ICE: An integrated configuration engine for interference mitigation in cloud services. In 2015 IEEE International Conference on Autonomic Computing. 91\u2013100."},{"key":"e_1_3_1_58_2","first-page":"1012","volume-title":"Proceedings of the 2013 International Conference on Software Engineering","author":"Malik Haroon","year":"2013","unstructured":"Haroon Malik, Hadi Hemmati, and Ahmed E. Hassan. 2013. Automatic detection of performance deviations in the load testing of large scale systems. In Proceedings of the 2013 International Conference on Software Engineering. IEEE Press, 1012\u20131021."},{"key":"e_1_3_1_59_2","doi-asserted-by":"crossref","first-page":"428","DOI":"10.1109\/ISCA.2012.6237037","volume-title":"Computer Architecture (ISCA), 2012 39th Annual International Symposium on","author":"Manikantan Raman","year":"2012","unstructured":"Raman Manikantan, Kaushik Rajan, and Ramaswamy Govindarajan. 2012. Probabilistic shared cache management (PriSM). In Computer Architecture (ISCA), 2012 39th Annual International Symposium on. IEEE, 428\u2013439."},{"key":"e_1_3_1_60_2","doi-asserted-by":"crossref","first-page":"248","DOI":"10.1145\/2155620.2155650","volume-title":"Proceedings of the 44th Annual IEEE\/ACM International Symposium on Microarchitecture","author":"Mars Jason","year":"2011","unstructured":"Jason Mars, Lingjia Tang, Robert Hundt, Kevin Skadron, and Mary Lou Soffa. 2011. Bubble-up: Increasing utilization in modern warehouse scale computers via sensible co-locations. In Proceedings of the 44th Annual IEEE\/ACM International Symposium on Microarchitecture. ACM, 248\u2013259."},{"key":"e_1_3_1_61_2","series-title":"Proceedings of the 44th Annual IEEE\/ACM International Symposium on Microarchitecture","first-page":"248","author":"Mars Jason","year":"2011","unstructured":"Jason Mars, Lingjia Tang, Robert Hundt, Kevin Skadron, and Mary Lou Soffa. 2011. Bubble-up: Increasing utilization in modern warehouse scale computers via sensible co-locations. In Proceedings of the 44th Annual IEEE\/ACM International Symposium on Microarchitecture (Porto Alegre, Brazil) (MICRO-44). ACM, New York, NY, USA, 248\u2013259."},{"key":"e_1_3_1_62_2","volume-title":"Proceedings of Machine Learning and Systems 2020, MLSys 2020, Austin, TX, USA, March 2-4, 2020","author":"Mattson Peter","year":"2020","unstructured":"Peter Mattson, Christine Cheng, Gregory F. Diamos, Cody Coleman, Paulius Micikevicius, David A. Patterson, Hanlin Tang, Gu-Yeon Wei, Peter Bailis, Victor Bittorf, David Brooks, Dehao Chen, Debo Dutta, Udit Gupta, Kim M. Hazelwood, Andy Hock, Xinyuan Huang, Daniel Kang, David Kanter, Naveen Kumar, Jeffery Liao, Deepak Narayanan, Tayo Oguntebi, Gennady Pekhimenko, Lillian Pentecost, Vijay Janapa Reddi, Taylor Robie, Tom St. John, Carole-Jean Wu, Lingjie Xu, Cliff Young, and Matei Zaharia. 2020. MLPerf training benchmark. In Proceedings of Machine Learning and Systems 2020, MLSys 2020, Austin, TX, USA, March 2-4, 2020, Inderjit S. Dhillon, Dimitris S. Papailiopoulos, and Vivienne Sze (Eds.). mlsys.org"},{"key":"e_1_3_1_63_2","doi-asserted-by":"crossref","first-page":"83","DOI":"10.1109\/MIC.2002.1003136","article-title":"TPC-W: A benchmark for E-commerce","volume":"6","author":"Menasce D. A.","year":"2002","unstructured":"D. A. Menasce. 2002. TPC-W: A benchmark for E-commerce. IEEE Internet Computing 6 (052002), 83\u201387.","journal-title":"IEEE Internet Computing"},{"key":"e_1_3_1_64_2","doi-asserted-by":"crossref","first-page":"237","DOI":"10.1145\/1755913.1755938","volume-title":"Proceedings of the 5th European Conference on Computer Systems","author":"Nathuji Ripal","year":"2010","unstructured":"Ripal Nathuji, Aman Kansal, and Alireza Ghaffarkhah. 2010. Q-clouds: Managing performance interference effects for QoS-aware clouds. In Proceedings of the 5th European Conference on Computer Systems. ACM, 237\u2013250."},{"key":"e_1_3_1_65_2","doi-asserted-by":"crossref","first-page":"167","DOI":"10.1109\/HPCA47549.2020.00023","volume-title":"2020 IEEE International Symposium on High Performance Computer Architecture (HPCA)","author":"Nishtala Rajiv","year":"2020","unstructured":"Rajiv Nishtala, Vinicius Petrucci, Paul Carpenter, and Magnus Sjalander. 2020. Twig : Multi-agent task management for colocated latency-critical cloud services. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, 167\u2013179."},{"key":"e_1_3_1_66_2","series-title":"Proceedings of the 2013 USENIX Conference on Annual Technical Conference","first-page":"219","author":"Novakovic Dejan","year":"2013","unstructured":"Dejan Novakovic, Nedeljko Vasic, Stanko Novakovic, Dejan Kostic, and Ricardo Bianchini. 2013. DeepDive: Transparently identifying and managing performance interference in virtualized environments. In Proceedings of the 2013 USENIX Conference on Annual Technical Conference (San Jose, CA) (USENIX ATC\u201913). USENIX Association, Berkeley, CA, USA, 219\u2013230."},{"unstructured":"Numactl. 2019. https:\/\/github.com\/numactl\/numactl","key":"e_1_3_1_67_2"},{"key":"e_1_3_1_68_2","series-title":"Proceedings of the 4th USENIX Conference on Hot Topics in Cloud Computing","first-page":"4","author":"Ou Zhonghong","year":"2012","unstructured":"Zhonghong Ou, Hao Zhuang, Jukka K. Nurminen, Antti Yl\u00e4-J\u00e4\u00e4ski, and Pan Hui. 2012. Exploiting hardware heterogeneity within the same instance type of Amazon EC2. In Proceedings of the 4th USENIX Conference on Hot Topics in Cloud Computing (Boston, MA) (HotCloud\u201912). USENIX Association, Berkeley, CA, USA, 4\u20134."},{"key":"e_1_3_1_69_2","first-page":"21","volume-title":"Proceedings of the Joined Workshops COSH 2017 and VisorHPC 2017","author":"Papadakis Ioannis","year":"2017","unstructured":"Ioannis Papadakis, Konstantinos Nikas, Vasileios Karakostas, Georgios Goumas, and Nectarios Koziris. 2017. Improving QoS and utilisation in modern multi-core servers with dynamic cache partitioning. In Proceedings of the Joined Workshops COSH 2017 and VisorHPC 2017, Carsten Clauss, Stefan Lankes, Carsten Trinitis, and Josef Weidendorfer (Eds.). Stockholm, Sweden, 21\u201326."},{"doi-asserted-by":"crossref","unstructured":"Jinsu Park Seongbeom Park and Woongki Baek. 2019. CoPart: Coordinated partitioning of last-level cache and memory bandwidth for fairness-aware workload consolidation on commodity servers. In Proceedings of the Fourteenth EuroSys Conference 2019 (Dresden Germany) (EuroSys \u201919). ACM New York NY USA Article 10 16 pages.","key":"e_1_3_1_70_2","DOI":"10.1145\/3302424.3303963"},{"key":"e_1_3_1_71_2","doi-asserted-by":"crossref","first-page":"193","DOI":"10.1109\/HPCA47549.2020.00025","volume-title":"2020 IEEE International Symposium on High Performance Computer Architecture (HPCA)","author":"Patel Tirthak","year":"2020","unstructured":"Tirthak Patel and Devesh Tiwari. 2020. CLITE : Efficient and QoS-aware co-location of multiple latency-critical jobs for warehouse scale computers. In 2020 IEEE International Symposium on High Performance Computer Architecture (HPCA). IEEE, 193\u2013206."},{"doi-asserted-by":"crossref","unstructured":"Yanghua Peng Yixin Bao Yangrui Chen Chuan Wu and Chuanxiong Guo. 2018. Optimus: An efficient dynamic resource scheduler for deep learning clusters(EuroSys \u201918). Association for Computing Machinery New York NY USA Article 3 14 pages.","key":"e_1_3_1_72_2","DOI":"10.1145\/3190508.3190517"},{"key":"e_1_3_1_73_2","first-page":"51","volume-title":"Proceedings of the 2010 IEEE 3rd International Conference on Cloud Computing (CLOUD \u201910)","author":"Pu Xing","year":"2010","unstructured":"Xing Pu, Ling Liu, Yiduo Mei, Sankaran Sivathanu, Younggyun Koh, and Calton Pu. 2010. Understanding performance interference of I\/O workload in virtualized cloud environments. In Proceedings of the 2010 IEEE 3rd International Conference on Cloud Computing (CLOUD \u201910). IEEE Computer Society, Washington, DC, USA, 51\u201358."},{"key":"e_1_3_1_74_2","series-title":"Proceedings of the 15th International Middleware Conference","first-page":"301","author":"Rameshan Navaneeth","year":"2014","unstructured":"Navaneeth Rameshan, Leandro Navarro, Enric Monte, and Vladimir Vlassov. 2014. Stay-away, protecting sensitive applications from performance interference. In Proceedings of the 15th International Middleware Conference (Bordeaux, France) (Middleware \u201914). ACM, New York, NY, USA, 301\u2013312."},{"unstructured":"Redis. 2019. Redis: An open source in-memory data structure store. https:\/\/redis.io","key":"e_1_3_1_75_2"},{"issue":"6","key":"e_1_3_1_76_2","doi-asserted-by":"crossref","first-page":"1159","DOI":"10.1109\/TPDS.2011.257","article-title":"Precise, scalable, and online request tracing for multitier services of black boxes","volume":"23","author":"Sang B.","year":"2012","unstructured":"B. Sang, J. Zhan, G. Lu, H. Wang, D. Xu, L. Wang, Z. Zhang, and Z. Jia. 2012. Precise, scalable, and online request tracing for multitier services of black boxes. IEEE Transactions on Parallel and Distributed Systems 23, 6 (June2012), 1159\u20131167.","journal-title":"IEEE Transactions on Parallel and Distributed Systems"},{"issue":"1","key":"e_1_3_1_77_2","doi-asserted-by":"crossref","first-page":"460","DOI":"10.14778\/1920841.1920902","article-title":"Runtime measurements in the cloud: Observing, analyzing, and reducing variance","volume":"3","author":"Schad J\u00f6rg","year":"2010","unstructured":"J\u00f6rg Schad, Jens Dittrich, and Jorge-Arnulfo Quian\u00e9-Ruiz. 2010. Runtime measurements in the cloud: Observing, analyzing, and reducing variance. Proc. VLDB Endow. 3, 1-2 (Sept.2010), 460\u2013471.","journal-title":"Proc. VLDB Endow."},{"doi-asserted-by":"crossref","unstructured":"Prateek Sharma Ahmed Ali-Eldin and Prashant Shenoy. 2019. Resource deflation: A new approach for transient resource reclamation. In Proceedings of the Fourteenth EuroSys Conference 2019 (Dresden Germany) (EuroSys \u201919). Association for Computing Machinery New York NY USA Article 33 17 pages.","key":"e_1_3_1_78_2","DOI":"10.1145\/3302424.3303945"},{"issue":"6","key":"e_1_3_1_79_2","doi-asserted-by":"crossref","first-page":"84","DOI":"10.1109\/MM.2003.1261391","article-title":"Discovering and exploiting program phases","volume":"23","author":"Sherwood Timothy","year":"2003","unstructured":"Timothy Sherwood, Erez Perelman, Greg Hamerly, Suleyman Sair, and Brad Calder. 2003. Discovering and exploiting program phases. IEEE Micro 23, 6 (Nov.2003), 84\u201393.","journal-title":"IEEE Micro"},{"key":"e_1_3_1_80_2","volume-title":"Dapper, a Large-Scale Distributed Systems Tracing Infrastructure","author":"Sigelman Benjamin H.","year":"2010","unstructured":"Benjamin H. Sigelman, Luiz Andr\u00e9 Barroso, Mike Burrows, Pat Stephenson, Manoj Plakal, Donald Beaver, Saul Jaspan, and Chandan Shanbhag. 2010. Dapper, a Large-Scale Distributed Systems Tracing Infrastructure. Technical Report. Google, Inc."},{"key":"e_1_3_1_81_2","doi-asserted-by":"crossref","first-page":"48","DOI":"10.1109\/TSC.2011.36","article-title":"Performance analysis of network I\/O workloads in virtualized data centers","volume":"6","author":"Sivathanu S.","year":"2013","unstructured":"S. Sivathanu, X. Pu, L. Liu, X. Dong, and Y. Mei. 2013. Performance analysis of network I\/O workloads in virtualized data centers. IEEE Transactions on Services Computing 6 (012013), 48\u201363.","journal-title":"IEEE Transactions on Services Computing"},{"unstructured":"Solr. 2021. Solr is the popular blazing-fast open source enterprise search platform built on Apache Lucene. https:\/\/solr.apache.org\/","key":"e_1_3_1_82_2"},{"unstructured":"Spark-Bench. 2020. Spark-Bench: A Benchmark Suite for Apache Spark. https:\/\/github.com\/codait\/spark-bench","key":"e_1_3_1_83_2"},{"key":"e_1_3_1_84_2","doi-asserted-by":"crossref","first-page":"517","DOI":"10.1145\/1669112.1669177","volume-title":"Proceedings of the 42nd Annual IEEE\/ACM International Symposium on Microarchitecture","author":"Srikantaiah Shekhar","year":"2009","unstructured":"Shekhar Srikantaiah, Mahmut Kandemir, and Qian Wang. 2009. SHARP control: Controlled shared cache management in chip multiprocessors. In Proceedings of the 42nd Annual IEEE\/ACM International Symposium on Microarchitecture. ACM, 517\u2013528."},{"key":"e_1_3_1_85_2","first-page":"71","volume-title":"Proceedings of the 2nd Conference on Symposium on Networked Systems Design & Implementation-Volume 2","author":"Stewart Christopher","year":"2005","unstructured":"Christopher Stewart and Kai Shen. 2005. Performance modeling and system management for multi-component online services. In Proceedings of the 2nd Conference on Symposium on Networked Systems Design & Implementation-Volume 2. USENIX Association, 71\u201384."},{"key":"e_1_3_1_86_2","first-page":"24","volume-title":"IEEE International Symposium on Performance Analysis of Systems and Software, ISPASS 2021, Stony Brook, NY, USA, March 28-30, 2021","author":"Tang Fei","year":"2021","unstructured":"Fei Tang, Wanling Gao, Jianfeng Zhan, Chuanxin Lan, Xu Wen, Lei Wang, Chunjie Luo, Zheng Cao, Xingwang Xiong, Zihan Jiang, Tianshu Hao, Fanda Fan, Fan Zhang, Yunyou Huang, Jianan Chen, Mengjia Du, Rui Ren, Chen Zheng, Daoyi Zheng, Haoning Tang, Kunlin Zhan, Biao Wang, Defei Kong, Minghe Yu, Chongkang Tan, Huan Li, Xinhui Tian, Yatao Li, Junchao Shao, Zhenyu Wang, Xiaoyu Wang, Jiahui Dai, and Hainan Ye. 2021. AIBench training: Balanced industry-standard AI training benchmarking. In IEEE International Symposium on Performance Analysis of Systems and Software, ISPASS 2021, Stony Brook, NY, USA, March 28-30, 2021. 24\u201335."},{"issue":"1","key":"e_1_3_1_87_2","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1145\/1140103.1140280","article-title":"Stardust: Tracking activity in a distributed storage system","volume":"34","author":"Thereska Eno","year":"2006","unstructured":"Eno Thereska, Brandon Salmon, John Strunk, Matthew Wachs, Michael Abd-El-Malek, Julio Lopez, and Gregory R. Ganger. 2006. Stardust: Tracking activity in a distributed storage system. SIGMETRICS Perform. Eval. Rev. 34, 1 (June2006), 3\u201314.","journal-title":"SIGMETRICS Perform. Eval. Rev."},{"doi-asserted-by":"crossref","unstructured":"Muhammad Tirmazi Adam Barker Nan Deng Md. E. Haque Zhijing Gene Qin Steven Hand Mor Harchol-Balter and John Wilkes. 2020. Borg: The next generation. In Proceedings of the Fifteenth European Conference on Computer Systems (Heraklion Greece) (EuroSys \u201920). Association for Computing Machinery New York NY USA Article 30 14 pages.","key":"e_1_3_1_88_2","DOI":"10.1145\/3342195.3387517"},{"key":"e_1_3_1_89_2","article-title":"iPerf: The TCP\/UDP bandwidth measurement tool","volume":"38","author":"Tirumala A.","year":"2005","unstructured":"A. Tirumala, F. Qin, J. Dugan, J. Ferguson, and K. Gibbs. 2005. iPerf: The TCP\/UDP bandwidth measurement tool. http.dast.nlanr.net\/Projects 38 (2005).","journal-title":"http.dast.nlanr.net\/Projects"},{"key":"e_1_3_1_90_2","series-title":"Proceedings of the 29th Conference on Information Communications","first-page":"1163","author":"Wang Guohui","year":"2010","unstructured":"Guohui Wang and T. S. Eugene Ng. 2010. The impact of virtualization on network performance of Amazon EC2 data center. In Proceedings of the 29th Conference on Information Communications (San Diego, California, USA) (INFOCOM\u201910). IEEE Press, Piscataway, NJ, USA, 1163\u20131171."},{"key":"e_1_3_1_91_2","doi-asserted-by":"crossref","first-page":"488","DOI":"10.1109\/HPCA.2014.6835958","volume-title":"2014 IEEE 20th International Symposium on High Performance Computer Architecture (HPCA)","author":"Wang Lei","year":"2014","unstructured":"Lei Wang, Jianfeng Zhan, Chunjie Luo, Yuqing Zhu, Qiang Yang, Yongqiang He, Wanling Gao, Zhen Jia, Yingjie Shi, Shujie Zhang, Chen Zheng, Gang Lu, Kent Zhan, Xiaona Li, and Bizhu Qiu. 2014. BigDataBench: A big data benchmark suite from internet services. In 2014 IEEE 20th International Symposium on High Performance Computer Architecture (HPCA). 488\u2013499."},{"issue":"8","key":"e_1_3_1_92_2","doi-asserted-by":"crossref","first-page":"2470","DOI":"10.1109\/TC.2015.2481403","article-title":"Heterogeneity and interference-aware virtual machine provisioning for predictable performance in the cloud","volume":"65","author":"Xu Fei","year":"2016","unstructured":"Fei Xu, Fangming Liu, and Hai Jin. 2016. Heterogeneity and interference-aware virtual machine provisioning for predictable performance in the cloud. IEEE Trans. Comput. 65, 8 (2016), 2470\u20132483.","journal-title":"IEEE Trans. Comput."},{"issue":"12","key":"e_1_3_1_93_2","doi-asserted-by":"crossref","first-page":"3012","DOI":"10.1109\/TC.2013.185","article-title":"iAware: Making live migration of virtual machines interference-aware in the cloud","volume":"63","author":"Xu Fei","year":"2014","unstructured":"Fei Xu, Fangming Liu, Linghui Liu, Hai Jin, Bo Li, and Baochun Li. 2014. iAware: Making live migration of virtual machines interference-aware in the cloud. IEEE Trans. Comput. 63, 12 (Dec.2014), 3012\u20133025.","journal-title":"IEEE Trans. Comput."},{"key":"e_1_3_1_94_2","doi-asserted-by":"crossref","first-page":"199","DOI":"10.1109\/ICAC.2016.38","volume-title":"2016 IEEE International Conference on Autonomic Computing (ICAC)","author":"Xu H.","year":"2016","unstructured":"H. Xu, X. Ning, H. Zhang, J. Rhee, and G. Jiang. 2016. PInfer: Learning to infer concurrent request paths from system kernel events. In 2016 IEEE International Conference on Autonomic Computing (ICAC). 199\u2013208."},{"key":"e_1_3_1_95_2","series-title":"Proceedings of the 19th International Middleware Conference","first-page":"146","author":"Xu Ran","year":"2018","unstructured":"Ran Xu, Subrata Mitra, Jason Rahman, Peter Bai, Bowen Zhou, Greg Bronevetsky, and Saurabh Bagchi. 2018. Pythia: Improving datacenter utilization via precise contention prediction for multiple co-located workloads. In Proceedings of the 19th International Middleware Conference (Rennes, France) (Middleware \u201918). ACM, New York, NY, USA, 146\u2013160."},{"key":"e_1_3_1_96_2","first-page":"1","volume-title":"Proceedings of the ACM Symposium on Cloud Computing","author":"Yadwadkar Neeraja J.","year":"2014","unstructured":"Neeraja J. Yadwadkar, Ganesh Ananthanarayanan, and Randy Katz. 2014. Wrangler: Predictable and faster jobs using fewer resources. In Proceedings of the ACM Symposium on Cloud Computing. ACM, 1\u201314."},{"issue":"3","key":"e_1_3_1_97_2","doi-asserted-by":"crossref","first-page":"607","DOI":"10.1145\/2508148.2485974","article-title":"Bubble-flux: Precise online QoS management for increased utilization in warehouse scale computers","volume":"41","author":"Yang Hailong","year":"2013","unstructured":"Hailong Yang, Alex Breslow, Jason Mars, and Lingjia Tang. 2013. Bubble-flux: Precise online QoS management for increased utilization in warehouse scale computers. ACM SIGARCH Computer Architecture News 41, 3 (2013), 607\u2013618.","journal-title":"ACM SIGARCH Computer Architecture News"},{"key":"e_1_3_1_98_2","first-page":"309","volume-title":"2016 USENIX Annual Technical Conference (USENIX ATC 16)","author":"Yang Xi","year":"2016","unstructured":"Xi Yang, Stephen M. Blackburn, and Kathryn S. McKinley. 2016. Elfen scheduling: Fine-grain principled borrowing from latency-critical workloads using simultaneous multithreading. In 2016 USENIX Annual Technical Conference (USENIX ATC 16). USENIX Association, Denver, CO, 309\u2013322."},{"unstructured":"Matei Zaharia Mosharaf Chowdhury Tathagata Das Ankur Dave Justin Ma Murphy McCauley Michael J. Franklin Scott Shenker and Ion Stoica. 2012. Resilient distributed datasets: A fault-tolerant abstraction for in-memory cluster computing. In Proceedings of the 9th USENIX Conference on Networked Systems Design and Implementation (San Jose CA) (NSDI\u201912). USENIX Association USA 2.","key":"e_1_3_1_99_2"},{"key":"e_1_3_1_100_2","doi-asserted-by":"crossref","first-page":"379","DOI":"10.1145\/2465351.2465388","volume-title":"Proceedings of the 8th ACM European Conference on Computer Systems","author":"Zhang Xiao","year":"2013","unstructured":"Xiao Zhang, Eric Tune, Robert Hagmann, Rohit Jnagal, Vrigo Gokhale, and John Wilkes. 2013. CPI 2: CPU performance isolation for shared compute clusters. In Proceedings of the 8th ACM European Conference on Computer Systems. 379\u2013391."},{"key":"e_1_3_1_101_2","doi-asserted-by":"crossref","first-page":"406","DOI":"10.1109\/MICRO.2014.53","volume-title":"2014 47th Annual IEEE\/ACM International Symposium on Microarchitecture","author":"Zhang Y.","year":"2014","unstructured":"Y. Zhang, M. A. Laurenzano, J. Mars, and L. Tang. 2014. SMiTe: Precise QoS prediction on real-system SMT processors to improve utilization in warehouse scale computers. In 2014 47th Annual IEEE\/ACM International Symposium on Microarchitecture. 406\u2013418."},{"key":"e_1_3_1_102_2","doi-asserted-by":"crossref","first-page":"406","DOI":"10.1109\/MICRO.2014.53","volume-title":"Microarchitecture (MICRO), 2014 47th Annual IEEE\/ACM International Symposium on","author":"Zhang Yunqi","year":"2014","unstructured":"Yunqi Zhang, Michael A. Laurenzano, Jason Mars, and Lingjia Tang. 2014. SMiTe: Precise QoS prediction on real-system SMT processors to improve utilization in warehouse scale computers. In Microarchitecture (MICRO), 2014 47th Annual IEEE\/ACM International Symposium on. IEEE, 406\u2013418."},{"issue":"5","key":"e_1_3_1_103_2","doi-asserted-by":"crossref","first-page":"1443","DOI":"10.1109\/TPDS.2015.2442983","article-title":"Predicting cross-core performance interference on multicore processors with regression analysis","volume":"27","author":"Zhao Jiacheng","year":"2016","unstructured":"Jiacheng Zhao, Huimin Cui, Jingling Xue, and Xiaobing Feng. 2016. Predicting cross-core performance interference on multicore processors with regression analysis. IEEE Trans. Parallel Distrib. Syst. 27, 5 (May2016), 1443\u20131456.","journal-title":"IEEE Trans. Parallel Distrib. Syst."},{"key":"e_1_3_1_104_2","series-title":"Proceedings of the 22nd International Conference on Parallel Architectures and Compilation Techniques","first-page":"201","author":"Zhao Jiacheng","year":"2013","unstructured":"Jiacheng Zhao, Huimin Cui, Jingling Xue, Xiaobing Feng, Youliang Yan, and Wensen Yang. 2013. An empirical model for predicting cross-core performance interference on multicore processors. In Proceedings of the 22nd International Conference on Parallel Architectures and Compilation Techniques (Edinburgh, Scotland, UK) (PACT \u201913). IEEE Press, Piscataway, NJ, USA, 201\u2013212."},{"key":"e_1_3_1_105_2","first-page":"653","volume-title":"2019 IEEE International Parallel and Distributed Processing Symposium (IPDPS)","author":"Zhao Wenyi","year":"2019","unstructured":"Wenyi Zhao, Quan Chen, Hao Lin, Jianfeng Zhang, Jingwen Leng, Chao Li, Wenli Zheng, Li Li, and Minyi Guo. 2019. Themis: Predicting and reining in application-level slowdown on spatial multitasking GPUs. In 2019 IEEE International Parallel and Distributed Processing Symposium (IPDPS). 653\u2013663."},{"issue":"3","key":"e_1_3_1_106_2","doi-asserted-by":"crossref","first-page":"380","DOI":"10.1109\/TC.2008.172","article-title":"On the design of fault-tolerant scheduling strategies using primary-backup approach for computational grids with low replication costs","volume":"58","author":"Zheng Qin","year":"2009","unstructured":"Qin Zheng, Bharadwaj Veeravalli, and Chen-Khong Tham. 2009. On the design of fault-tolerant scheduling strategies using primary-backup approach for computational grids with low replication costs. IEEE Trans. Comput. 58, 3 (2009), 380\u2013393.","journal-title":"IEEE Trans. Comput."},{"issue":"2","key":"e_1_3_1_107_2","doi-asserted-by":"crossref","first-page":"33","DOI":"10.1145\/2980024.2872394","article-title":"Dirigent: Enforcing QoS for latency-critical tasks on shared multicore systems","volume":"44","author":"Zhu Haishan","year":"2016","unstructured":"Haishan Zhu and Mattan Erez. 2016. Dirigent: Enforcing QoS for latency-critical tasks on shared multicore systems. ACM SIGARCH Computer Architecture News 44, 2 (2016), 33\u201347.","journal-title":"ACM SIGARCH Computer Architecture News"},{"key":"e_1_3_1_108_2","first-page":"2223","volume-title":"Proceedings of the IEEE International Conference on Computer Vision","author":"Zhu Jun-Yan","year":"2017","unstructured":"Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros. 2017. Unpaired image-to-image translation using cycle-consistent adversarial networks. In Proceedings of the IEEE International Conference on Computer Vision. 2223\u20132232."},{"unstructured":"Zipkin. 2019. https:\/\/zipkin.io\/","key":"e_1_3_1_109_2"}],"container-title":["ACM Transactions on Computer Systems"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3630006","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3630006","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T23:57:00Z","timestamp":1750291020000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3630006"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,2,13]]},"references-count":108,"journal-issue":{"issue":"1-2","published-print":{"date-parts":[[2024,5,31]]}},"alternative-id":["10.1145\/3630006"],"URL":"https:\/\/doi.org\/10.1145\/3630006","relation":{},"ISSN":["0734-2071","1557-7333"],"issn-type":[{"type":"print","value":"0734-2071"},{"type":"electronic","value":"1557-7333"}],"subject":[],"published":{"date-parts":[[2024,2,13]]},"assertion":[{"value":"2023-10-20","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-02-13","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}