{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T22:38:17Z","timestamp":1784241497256,"version":"3.55.0"},"reference-count":171,"publisher":"Association for Computing Machinery (ACM)","issue":"3","license":[{"start":{"date-parts":[[2017,6,29]],"date-time":"2017-06-29T00:00:00Z","timestamp":1498694400000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"European Commission under the Horizon 2020 program RAPID","award":["H2020-ICT-644312"],"award-info":[{"award-number":["H2020-ICT-644312"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2018,5,31]]},"abstract":"<jats:p>The integration of graphics processing units (GPUs) on high-end compute nodes has established a new accelerator-based heterogeneous computing model, which now permeates high-performance computing. The same paradigm nevertheless has limited adoption in cloud computing or other large-scale distributed computing paradigms. Heterogeneous computing with GPUs can benefit the Cloud by reducing operational costs and improving resource and energy efficiency. However, such a paradigm shift would require effective methods for virtualizing GPUs, as well as other accelerators. In this survey article, we present an extensive and in-depth survey of GPU virtualization techniques and their scheduling methods. We review a wide range of virtualization techniques implemented at the GPU library, driver, and hardware levels. Furthermore, we review GPU scheduling methods that address performance and fairness issues between multiple virtual machines sharing GPUs. We believe that our survey delivers a perspective on the challenges and opportunities for virtualization of heterogeneous computing environments.<\/jats:p>","DOI":"10.1145\/3068281","type":"journal-article","created":{"date-parts":[[2017,6,30]],"date-time":"2017-06-30T12:36:19Z","timestamp":1498826179000},"page":"1-37","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":97,"title":["GPU Virtualization and Scheduling Methods"],"prefix":"10.1145","volume":"50","author":[{"given":"Cheol-Ho","family":"Hong","sequence":"first","affiliation":[{"name":"Queen's University Belfast, Northern Ireland, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ivor","family":"Spence","sequence":"additional","affiliation":[{"name":"Queen's University Belfast, Northern Ireland, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Dimitrios S.","family":"Nikolopoulos","sequence":"additional","affiliation":[{"name":"Queen's University Belfast, Northern Ireland, United Kingdom"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2017,6,29]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1535\/itj.1003.02"},{"key":"e_1_2_1_2_1","unstructured":"EC Amazon. 2010. Amazon elastic compute cloud (Amazon EC2). https:\/\/aws.amazon.com\/ec2\/.  EC Amazon. 2010. Amazon elastic compute cloud (Amazon EC2). https:\/\/aws.amazon.com\/ec2\/."},{"key":"e_1_2_1_3_1","unstructured":"AMD. 2009. R6xx_3D_Registers.pdf. Retrieved from http:\/\/amd-dev.wpengine.netdna-cdn.com\/wordpress\/media\/2013\/10\/R6xx_3D_Registers.pdf. (2009).  AMD. 2009. R6xx_3D_Registers.pdf. Retrieved from http:\/\/amd-dev.wpengine.netdna-cdn.com\/wordpress\/media\/2013\/10\/R6xx_3D_Registers.pdf. (2009)."},{"key":"e_1_2_1_4_1","volume-title":"APS Meeting Abstracts","volume":"1","author":"Anderson Joshua","year":"2010","unstructured":"Joshua Anderson , Aaron Keys , Carolyn Phillips , Trung Dac Nguyen , and Sharon Glotzer . 2010 . HOOMD-blue, general-purpose many-body dynamics on the GPU . In APS Meeting Abstracts , Vol. 1 . 18008. Joshua Anderson, Aaron Keys, Carolyn Phillips, Trung Dac Nguyen, and Sharon Glotzer. 2010. HOOMD-blue, general-purpose many-body dynamics on the GPU. In APS Meeting Abstracts, Vol. 1. 18008."},{"key":"e_1_2_1_5_1","volume-title":"Joseph James Gebis, Parry Husbands, Kurt Keutzer, David A. Patterson, William Lester Plishker, John Shalf, Samuel Webb Williams, and others.","author":"Asanovic Krste","year":"2006","unstructured":"Krste Asanovic , Ras Bodik , Bryan Christopher Catanzaro , Joseph James Gebis, Parry Husbands, Kurt Keutzer, David A. Patterson, William Lester Plishker, John Shalf, Samuel Webb Williams, and others. 2006 . The Landscape of Parallel Computing Research: A View from Berkeley. EECS Department Technical Report UCB\/EECS-2006-183. University of California , Berkeley. Krste Asanovic, Ras Bodik, Bryan Christopher Catanzaro, Joseph James Gebis, Parry Husbands, Kurt Keutzer, David A. Patterson, William Lester Plishker, John Shalf, Samuel Webb Williams, and others. 2006. The Landscape of Parallel Computing Research: A View from Berkeley. EECS Department Technical Report UCB\/EECS-2006-183. University of California, Berkeley."},{"key":"e_1_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1118\/1.3231824"},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/1165389.945462"},{"key":"e_1_2_1_8_1","volume-title":"12th International Workshop on Image Analysis for Multimedia Interactive Services (WIAMIS\u201911)","author":"Athanasopoulos Andreas","year":"2011","unstructured":"Andreas Athanasopoulos , Anastasios Dimou , Vasileios Mezaris , and Ioannis Kompatsiaris . 2011 . GPU acceleration for support vector machines . In 12th International Workshop on Image Analysis for Multimedia Interactive Services (WIAMIS\u201911) . TU Delft; EWI; MM; PRB, Delft, The Netherlands. Andreas Athanasopoulos, Anastasios Dimou, Vasileios Mezaris, and Ioannis Kompatsiaris. 2011. GPU acceleration for support vector machines. In 12th International Workshop on Image Analysis for Multimedia Interactive Services (WIAMIS\u201911). TU Delft; EWI; MM; PRB, Delft, The Netherlands."},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/ECRTS.2012.15"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/2287076.2287090"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/90.958328"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/1141911.1141947"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/2962131"},{"key":"e_1_2_1_14_1","volume-title":"Proceedings of the USENIX Annual Technical Conference.","author":"Burtsev Anton","unstructured":"Anton Burtsev , Kiran Srinivasan , Prashanth Radhakrishnan , Kaladhar Voruganti , and Garth R. Goodson . 2009. Fido: Fast inter-virtual-machine communication for enterprise appliances . In Proceedings of the USENIX Annual Technical Conference. Anton Burtsev, Kiran Srinivasan, Prashanth Radhakrishnan, Kaladhar Voruganti, and Garth R. Goodson. 2009. Fido: Fast inter-virtual-machine communication for enterprise appliances. In Proceedings of the USENIX Annual Technical Conference."},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER.2015.23"},{"key":"e_1_2_1_16_1","volume-title":"SOAP, UDDI 8 WSDL. O\u2019Reilly Media","author":"Cerami Ethan","unstructured":"Ethan Cerami . 2002. Web Services Essentials: Distributed Applications with XML-RPC , SOAP, UDDI 8 WSDL. O\u2019Reilly Media , Inc . Ethan Cerami. 2002. Web Services Essentials: Distributed Applications with XML-RPC, SOAP, UDDI 8 WSDL. O\u2019Reilly Media, Inc."},{"key":"e_1_2_1_17_1","volume-title":"The architecture of vmware esxi. VMware White Pap. 1, 7","author":"Chaubal Charu","year":"2008","unstructured":"Charu Chaubal . 2008. The architecture of vmware esxi. VMware White Pap. 1, 7 ( 2008 ). Charu Chaubal. 2008. The architecture of vmware esxi. VMware White Pap. 1, 7 (2008)."},{"key":"e_1_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/IISWC.2009.5306797"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1109\/IWQoS.2010.5542746"},{"key":"e_1_2_1_20_1","volume-title":"Proceedings of the International Conference on Control, Automation and Systems","author":"Cho Yun Chan","year":"2007","unstructured":"Yun Chan Cho and Jae Wook Jeon . 2007 . Sharing data between processes running on different domains in para-virtualized xen . In Proceedings of the International Conference on Control, Automation and Systems , 2007 (ICCAS\u201907). IEEE, 1255--1260. Yun Chan Cho and Jae Wook Jeon. 2007. Sharing data between processes running on different domains in para-virtualized xen. In Proceedings of the International Conference on Control, Automation and Systems, 2007 (ICCAS\u201907). IEEE, 1255--1260."},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER.2011.49"},{"key":"e_1_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1145\/1496909.1496918"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1735688.1735702"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/75246.75248"},{"key":"e_1_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISPA.2012.136"},{"key":"e_1_2_1_26_1","volume-title":"Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 24","author":"Dixon Matthew","year":"2014","unstructured":"Matthew Dixon , Sabbir Ahmed Khan , and Mohammad Zubair . 2014 . Accelerating option risk analytics in R using GPUs . In Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 24 . Matthew Dixon, Sabbir Ahmed Khan, and Mohammad Zubair. 2014. Accelerating option risk analytics in R using GPUs. In Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 24."},{"key":"e_1_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.5555\/2813767.2813806"},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2012.01.020"},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.728"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/1618525.1618534"},{"key":"e_1_2_1_31_1","volume-title":"European Conference on Parallel Processing. Springer, 385--394","author":"Duato Jos\u00e9","year":"2009","unstructured":"Jos\u00e9 Duato , Francisco D. Igual , Rafael Mayo , Antonio J. Pe\u00f1a , Enrique S. Quintana-Ort\u00ed , and Federico Silla . 2009 . An efficient implementation of GPU virtualization in high performance clusters . In European Conference on Parallel Processing. Springer, 385--394 . Jos\u00e9 Duato, Francisco D. Igual, Rafael Mayo, Antonio J. Pe\u00f1a, Enrique S. Quintana-Ort\u00ed, and Federico Silla. 2009. An efficient implementation of GPU virtualization in high performance clusters. In European Conference on Parallel Processing. Springer, 385--394."},{"key":"e_1_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/HiPC.2011.6152718"},{"key":"e_1_2_1_33_1","volume-title":"Proceedings of the 1st Workshop on Language, Compiler, and Architecture Support for GPGPU.","author":"Duato Jos\u00e9","unstructured":"Jos\u00e9 Duato , Antonio J. Pe\u00f1a , Federico Silla , Rafael Mayo , and Enrique S . Quintana-Ort\u0131. 2010a. Modeling the CUDA remoting virtualization behaviour in high performance networks . In Proceedings of the 1st Workshop on Language, Compiler, and Architecture Support for GPGPU. Jos\u00e9 Duato, Antonio J. Pe\u00f1a, Federico Silla, Rafael Mayo, and Enrique S. Quintana-Ort\u0131. 2010a. Modeling the CUDA remoting virtualization behaviour in high performance networks. In Proceedings of the 1st Workshop on Language, Compiler, and Architecture Support for GPGPU."},{"key":"e_1_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCS.2010.5547126"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2011.58"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.2845"},{"key":"e_1_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1145\/2851141.2851194"},{"key":"e_1_2_1_39_1","volume-title":"pascal and stacked memory: Feeding the appetite for big data. Retrieved from Nvidia.com","author":"Foley Denis","year":"2014","unstructured":"Denis Foley . 2014. NVLink , pascal and stacked memory: Feeding the appetite for big data. Retrieved from Nvidia.com ( 2014 ). Denis Foley. 2014. NVLink, pascal and stacked memory: Feeding the appetite for big data. Retrieved from Nvidia.com (2014)."},{"key":"e_1_2_1_40_1","unstructured":"Futuremark. 1998. 3DMark Benchmarks\u2014See the Current Range of this Popular PC Graphics Card Test. Retrieved from http:\/\/www.futuremark.com\/benchmarks\/3dmark\/all?_ga&equals;1.168926249.987441096.1470653002. (1998).  Futuremark. 1998. 3DMark Benchmarks\u2014See the Current Range of this Popular PC Graphics Card Test. Retrieved from http:\/\/www.futuremark.com\/benchmarks\/3dmark\/all?_ga&equals;1.168926249.987441096.1470653002. (1998)."},{"key":"e_1_2_1_41_1","volume-title":"Proceedings of the Workshop on Hot Topics in Operating Systems (HotOS\u201905)","author":"Garfinkel Tal","year":"2005","unstructured":"Tal Garfinkel and Mendel Rosenblum . 2005 . When virtual is harder than real: Security challenges in virtual machine based computing environments . In Proceedings of the Workshop on Hot Topics in Operating Systems (HotOS\u201905) . Tal Garfinkel and Mendel Rosenblum. 2005. When virtual is harder than real: Security challenges in virtual machine based computing environments. In Proceedings of the Workshop on Hot Topics in Operating Systems (HotOS\u201905)."},{"key":"e_1_2_1_43_1","doi-asserted-by":"crossref","unstructured":"Francisco Giunta Raffaele Montella Giuliano Laccetti Florin Isaila and F. Blas. 2011. A GPU accelerated high performance cloud computing infrastructure for grid computing based virtual environmental laboratory. Adv. Grid Comput. Lecture Notes in Computer Science. Vol.&nbsp;6271. Springer Berlin Heidelberg 35--43.  Francisco Giunta Raffaele Montella Giuliano Laccetti Florin Isaila and F. Blas. 2011. A GPU accelerated high performance cloud computing infrastructure for grid computing based virtual environmental laboratory. Adv. Grid Comput. Lecture Notes in Computer Science. Vol.&nbsp;6271. Springer Berlin Heidelberg 35--43.","DOI":"10.5772\/14594"},{"key":"e_1_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.5555\/1887695.1887738"},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.cpc.2015.02.028"},{"key":"e_1_2_1_46_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.1974.6323581"},{"key":"e_1_2_1_47_1","doi-asserted-by":"publisher","DOI":"10.1109\/HPCC.and.EUC.2013.245"},{"key":"e_1_2_1_48_1","first-page":"121","article-title":"Particle simulation using cuda","volume":"6","author":"Green Simon","year":"2010","unstructured":"Simon Green . 2010 . Particle simulation using cuda . NVIDIA Whitepaper 6 (2010), 121 -- 128 . Simon Green. 2010. Particle simulation using cuda. NVIDIA Whitepaper 6 (2010), 121--128.","journal-title":"NVIDIA Whitepaper"},{"key":"e_1_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1016\/0167-8191(96)00024-5"},{"key":"e_1_2_1_50_1","first-page":"8","article-title":"The opencl specification","volume":"1","author":"Khronos OpenCL Working Group et&nbsp;al.","year":"2008","unstructured":"Khronos OpenCL Working Group et&nbsp;al. 2008 . The opencl specification . Version 1 , 29 (2008), 8 . Khronos OpenCL Working Group et&nbsp;al. 2008. The opencl specification. Version 1, 29 (2008), 8.","journal-title":"Version"},{"key":"e_1_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/1519138.1519141"},{"key":"e_1_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2014.2350499"},{"key":"e_1_2_1_53_1","volume-title":"Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911)","author":"Gupta Vishakha","year":"2011","unstructured":"Vishakha Gupta , Karsten Schwan , Niraj Tolia , Vanish Talwar , and Parthasarathy Ranganathan . 2011 . Pegasus: Coordinated scheduling for virtualized accelerator-based systems . In Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911) . 31. Vishakha Gupta, Karsten Schwan, Niraj Tolia, Vanish Talwar, and Parthasarathy Ranganathan. 2011. Pegasus: Coordinated scheduling for virtualized accelerator-based systems. In Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911). 31."},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2014.10"},{"key":"e_1_2_1_55_1","volume-title":"Proceedings of the SIGMM Workshop on Network and Operating Systems Support for Digital Audio and Video (NOSSDAV\u201907)","author":"Hansen Jacob Gorm","year":"2007","unstructured":"Jacob Gorm Hansen . 2007 . Blink: Advanced display multiplexing for virtualized applications . In Proceedings of the SIGMM Workshop on Network and Operating Systems Support for Digital Audio and Video (NOSSDAV\u201907) . Jacob Gorm Hansen. 2007. Blink: Advanced display multiplexing for virtualized applications. In Proceedings of the SIGMM Workshop on Network and Operating Systems Support for Digital Audio and Video (NOSSDAV\u201907)."},{"key":"e_1_2_1_56_1","volume-title":"Proceedings of the USENIX Annual Technical Conference. 231--242","author":"Har\u2019El Nadav","year":"2013","unstructured":"Nadav Har\u2019El , Abel Gordon , Alex Landau , Muli Ben-Yehuda , Avishay Traeger , and Razya Ladelsky . 2013 . Efficient and scalable paravirtual I\/O system . In Proceedings of the USENIX Annual Technical Conference. 231--242 . Nadav Har\u2019El, Abel Gordon, Alex Landau, Muli Ben-Yehuda, Avishay Traeger, and Razya Ladelsky. 2013. Efficient and scalable paravirtual I\/O system. In Proceedings of the USENIX Annual Technical Conference. 231--242."},{"key":"e_1_2_1_57_1","volume-title":"Proceedings of the 2015 IEEE\/ACM 8th International Conference on Utility and Cloud Computing (UCC\u201915)","author":"Haydel Nicholas","unstructured":"Nicholas Haydel , Sandra Gesing , Ian Taylor , Gregory Madey , Abdul Dakkak , Simon Garcia De Gonzalo , and Wen-Mei W. Hwu . 2015. Enhancing the usability and utilization of accelerated architectures via docker . In Proceedings of the 2015 IEEE\/ACM 8th International Conference on Utility and Cloud Computing (UCC\u201915) . IEEE, 361--367. Nicholas Haydel, Sandra Gesing, Ian Taylor, Gregory Madey, Abdul Dakkak, Simon Garcia De Gonzalo, and Wen-Mei W. Hwu. 2015. Enhancing the usability and utilization of accelerated architectures via docker. In Proceedings of the 2015 IEEE\/ACM 8th International Conference on Utility and Cloud Computing (UCC\u201915). IEEE, 361--367."},{"key":"e_1_2_1_58_1","volume-title":"NVIDIA GRID: Graphics accelerated VDI with the visual performance of a workstation. Nvidia Corp","author":"Herrera Alex","year":"2014","unstructured":"Alex Herrera . 2014 . NVIDIA GRID: Graphics accelerated VDI with the visual performance of a workstation. Nvidia Corp (2014). http:\/\/www.nvidia.com\/content\/grid\/vdi-whitepaper.pdf. Alex Herrera. 2014. NVIDIA GRID: Graphics accelerated VDI with the visual performance of a workstation. Nvidia Corp (2014). http:\/\/www.nvidia.com\/content\/grid\/vdi-whitepaper.pdf."},{"key":"e_1_2_1_59_1","doi-asserted-by":"publisher","DOI":"10.5555\/2755535.2755539"},{"key":"e_1_2_1_60_1","doi-asserted-by":"publisher","DOI":"10.1145\/2892242.2892246"},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/566654.566639"},{"key":"e_1_2_1_62_1","doi-asserted-by":"publisher","DOI":"10.4218\/etrij.13.0212.0213"},{"key":"e_1_2_1_63_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-540-92990-1_4"},{"key":"e_1_2_1_64_1","volume-title":"Int. 2013","author":"Jo Heeseung","year":"2013","unstructured":"Heeseung Jo , Jinkyu Jeong , Myoungho Lee , and Dong Hoon Choi . 2013a. Exploiting GPUs in virtual machine for biocloud. BioMed Res . Int. 2013 ( 2013 ). Heeseung Jo, Jinkyu Jeong, Myoungho Lee, and Dong Hoon Choi. 2013a. Exploiting GPUs in virtual machine for biocloud. BioMed Res. Int. 2013 (2013)."},{"key":"e_1_2_1_65_1","doi-asserted-by":"publisher","DOI":"10.4028\/www.scientific.net\/AMM.311.15"},{"key":"e_1_2_1_66_1","unstructured":"David Kanter. 2010. Intels sandy bridge microarchitecture. http:\/\/www.realworldtech.com\/sandy-bridge\/.  David Kanter. 2010. Intels sandy bridge microarchitecture. http:\/\/www.realworldtech.com\/sandy-bridge\/."},{"key":"e_1_2_1_67_1","volume-title":"https:\/\/codesign.llnl.gov\/lulesh.php","author":"Karlin Ian","year":"2013","unstructured":"Ian Karlin , Jeff Keasler , and Rob Neely . 2013. Lulesh 2.0 updates and changes. Livermore , CA ( 2013 ). https:\/\/codesign.llnl.gov\/lulesh.php . Ian Karlin, Jeff Keasler, and Rob Neely. 2013. Lulesh 2.0 updates and changes. Livermore, CA (2013). https:\/\/codesign.llnl.gov\/lulesh.php."},{"key":"e_1_2_1_68_1","volume-title":"Proceedings of the International Workshop on Operating Systems Platforms for Embedded Real-Time Applications. 23--32","author":"Kato Shinpei","year":"2011","unstructured":"Shinpei Kato , Scott Brandt , Yutaka Ishikawa , and R Rajkumar . 2011 a. Operating systems challenges for GPU resource management . In Proceedings of the International Workshop on Operating Systems Platforms for Embedded Real-Time Applications. 23--32 . Shinpei Kato, Scott Brandt, Yutaka Ishikawa, and R Rajkumar. 2011a. Operating systems challenges for GPU resource management. In Proceedings of the International Workshop on Operating Systems Platforms for Embedded Real-Time Applications. 23--32."},{"key":"e_1_2_1_69_1","doi-asserted-by":"publisher","DOI":"10.1109\/RTAS.2011.26"},{"key":"e_1_2_1_70_1","doi-asserted-by":"publisher","DOI":"10.5555\/2120959.2121120"},{"key":"e_1_2_1_71_1","volume-title":"Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911)","author":"Kato Shinpei","year":"2011","unstructured":"Shinpei Kato , Karthik Lakshmanan , Raj Rajkumar , and Yutaka Ishikawa . 2011 d. TimeGraph: GPU scheduling for real-time multi-tasking environments . In Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911) . 17. Shinpei Kato, Karthik Lakshmanan, Raj Rajkumar, and Yutaka Ishikawa. 2011d. TimeGraph: GPU scheduling for real-time multi-tasking environments. In Proceedings of the 2011 USENIX Annual Technical Conference (USENIX ATC\u201911). 17."},{"key":"e_1_2_1_72_1","volume-title":"Proceedings of the USENIX Annual Technical Conference (USENIX ATC\u201911)","author":"Kato Shinpei","unstructured":"Shinpei Kato , Michael McThrow , Carlos Maltzahn , and Scott A. Brandt . 2012. Gdev: First-class GPU resource management in the operating system .. In Proceedings of the USENIX Annual Technical Conference (USENIX ATC\u201911) . 401--412. Shinpei Kato, Michael McThrow, Carlos Maltzahn, and Scott A. Brandt. 2012. Gdev: First-class GPU resource management in the operating system.. In Proceedings of the USENIX Annual Technical Conference (USENIX ATC\u201911). 401--412."},{"key":"e_1_2_1_73_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICCVE.2013.6799789"},{"key":"e_1_2_1_74_1","unstructured":"David B. Kirk and W. Hwu Wen-mei. 2012. Programming Massively Parallel Processors: A Hands-on Approach. Newnes.   David B. Kirk and W. Hwu Wen-mei. 2012. Programming Massively Parallel Processors: A Hands-on Approach. Newnes."},{"key":"e_1_2_1_75_1","volume-title":"Proceedings of the Linux Symposium","volume":"1","author":"Kivity Avi","year":"2007","unstructured":"Avi Kivity , Yaniv Kamay , Dor Laor , Uri Lublin , and Anthony Liguori . 2007 . kvm: The Linux virtual machine monitor . In Proceedings of the Linux Symposium , Vol. 1 . 225--230. Avi Kivity, Yaniv Kamay, Dor Laor, Uri Lublin, and Anthony Liguori. 2007. kvm: The Linux virtual machine monitor. In Proceedings of the Linux Symposium, Vol. 1. 225--230."},{"key":"e_1_2_1_76_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISSCC.2010.5434033"},{"issue":"7","key":"e_1_2_1_77_1","first-page":"975","article-title":"Method and system for remote device access in virtual environment. (issued date: July 5 2011)","author":"Kuzkin Maxim A.","year":"2011","unstructured":"Maxim A. Kuzkin and Alexander G. Tormasov . 2011 . Method and system for remote device access in virtual environment. (issued date: July 5 2011) . Patent No. 7 , 975 ,017. Filed date: Feb 25, 2009. Maxim A. Kuzkin and Alexander G. Tormasov. 2011. Method and system for remote device access in virtual environment. (issued date: July 5 2011). Patent No. 7,975,017. Filed date: Feb 25, 2009.","journal-title":"Patent"},{"key":"e_1_2_1_78_1","volume-title":"Proceedings of the AMD Fusion Developer Summit","author":"Kyriazis George","year":"2012","unstructured":"George Kyriazis . 2012 . Heterogeneous system architecture: A technical review . In Proceedings of the AMD Fusion Developer Summit (2012). George Kyriazis. 2012. Heterogeneous system architecture: A technical review. In Proceedings of the AMD Fusion Developer Summit (2012)."},{"key":"e_1_2_1_79_1","volume-title":"International Conference on Parallel Processing and Applied Mathematics. Springer, 734--744","author":"Laccetti Giuliano","year":"2013","unstructured":"Giuliano Laccetti , Raffaele Montella , Carlo Palmieri , and Valentina Pelliccia . 2013 . The high performance internet of things: Using GVirtuS to share high-end GPUs with ARM based cluster computing nodes . In International Conference on Parallel Processing and Applied Mathematics. Springer, 734--744 . Giuliano Laccetti, Raffaele Montella, Carlo Palmieri, and Valentina Pelliccia. 2013. The high performance internet of things: Using GVirtuS to share high-end GPUs with ARM based cluster computing nodes. In International Conference on Parallel Processing and Applied Mathematics. Springer, 734--744."},{"key":"e_1_2_1_80_1","doi-asserted-by":"publisher","DOI":"10.1145\/1254810.1254816"},{"key":"e_1_2_1_81_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDCS.2013.51"},{"key":"e_1_2_1_82_1","unstructured":"Michael Larabel and M. Tippett. 2011. Phoronix test suite. https:\/\/www.phoronix-test-suite.com.  Michael Larabel and M. Tippett. 2011. Phoronix test suite. https:\/\/www.phoronix-test-suite.com."},{"key":"e_1_2_1_83_1","doi-asserted-by":"publisher","DOI":"10.1109\/TII.2015.2469644"},{"key":"e_1_2_1_84_1","volume-title":"Proceedings of the 3rd USENIX Workshop on Hot Topics in Cloud Computing (HotCloud\u201911)","author":"Lee Gunho","unstructured":"Gunho Lee and Randy H. Katz . 2011. Heterogeneity-aware resource allocation and scheduling in the cloud .. In Proceedings of the 3rd USENIX Workshop on Hot Topics in Cloud Computing (HotCloud\u201911) . Gunho Lee and Randy H. Katz. 2011. Heterogeneity-aware resource allocation and scheduling in the cloud.. In Proceedings of the 3rd USENIX Workshop on Hot Topics in Cloud Computing (HotCloud\u201911)."},{"key":"e_1_2_1_85_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICPP.2011.88"},{"key":"e_1_2_1_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/2212908.2212950"},{"key":"e_1_2_1_87_1","doi-asserted-by":"publisher","DOI":"10.1109\/CCGrid.2015.105"},{"key":"e_1_2_1_88_1","doi-asserted-by":"publisher","DOI":"10.1109\/WAINA.2011.82"},{"key":"e_1_2_1_89_1","doi-asserted-by":"publisher","DOI":"10.1145\/2854038.2854040"},{"key":"e_1_2_1_90_1","volume-title":"Proceedings of the 2013 USENIX Annual Technical Conference (USENIX ATC\u201913)","author":"Menychtas Konstantinos","unstructured":"Konstantinos Menychtas , Kai Shen , and Michael L. Scott . 2013. Enabling OS research by inferring interactions in the black-box GPU stack . In Proceedings of the 2013 USENIX Annual Technical Conference (USENIX ATC\u201913) . 291--296. Konstantinos Menychtas, Kai Shen, and Michael L. Scott. 2013. Enabling OS research by inferring interactions in the black-box GPU stack. In Proceedings of the 2013 USENIX Annual Technical Conference (USENIX ATC\u201913). 291--296."},{"key":"e_1_2_1_91_1","doi-asserted-by":"publisher","DOI":"10.1145\/2644865.2541963"},{"key":"e_1_2_1_92_1","doi-asserted-by":"publisher","DOI":"10.1145\/1996121.1996124"},{"key":"e_1_2_1_93_1","doi-asserted-by":"publisher","DOI":"10.1145\/2636342"},{"key":"e_1_2_1_94_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-31464-3_75"},{"key":"e_1_2_1_95_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10586-013-0341-0"},{"key":"e_1_2_1_96_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-32152-3_1"},{"key":"e_1_2_1_97_1","volume-title":"Carmine Ferraro, Valentina Pelliccia, Cheol-Ho Hong, Ivor Spence, and Dimitrios S. Nikolopoulos.","author":"Montella Raffaele","year":"2016","unstructured":"Raffaele Montella , Giulio Giunta , Giuliano Laccetti , Marco Lapegna , Carlo Palmieri , Carmine Ferraro, Valentina Pelliccia, Cheol-Ho Hong, Ivor Spence, and Dimitrios S. Nikolopoulos. 2016 b. On the virtualization of CUDA based GPU remoting on ARM and X86 machines in the GVirtuS framework. Int. J. Parallel Program . (2016), 1--22. Raffaele Montella, Giulio Giunta, Giuliano Laccetti, Marco Lapegna, Carlo Palmieri, Carmine Ferraro, Valentina Pelliccia, Cheol-Ho Hong, Ivor Spence, and Dimitrios S. Nikolopoulos. 2016b. On the virtualization of CUDA based GPU remoting on ARM and X86 machines in the GVirtuS framework. Int. J. Parallel Program. (2016), 1--22."},{"key":"e_1_2_1_98_1","doi-asserted-by":"publisher","DOI":"10.1145\/641480.641493"},{"key":"e_1_2_1_99_1","unstructured":"Nvidia. 2007a. CUDA Code Samples\u2014NVIDIA Developer. Retrieved from https:\/\/developer.nvidia.com\/cuda-code-samples.  Nvidia. 2007a. CUDA Code Samples\u2014NVIDIA Developer. Retrieved from https:\/\/developer.nvidia.com\/cuda-code-samples."},{"key":"e_1_2_1_100_1","unstructured":"NVIDIA. 2012. HyperQ Example. Retrieved from http:\/\/docs.nvidia.com\/cuda\/samples\/6_Advanced\/simpleHyperQ\/doc\/HyperQ.pdf.  NVIDIA. 2012. HyperQ Example. Retrieved from http:\/\/docs.nvidia.com\/cuda\/samples\/6_Advanced\/simpleHyperQ\/doc\/HyperQ.pdf."},{"key":"e_1_2_1_101_1","unstructured":"NVIDIA. 2016a. GP100 Pascal Whitepaper. Retrieved from https:\/\/images.nvidia.com\/content\/pdf\/tesla\/whitepaper\/pascal-architecture-whitepaper.pdf.  NVIDIA. 2016a. GP100 Pascal Whitepaper. Retrieved from https:\/\/images.nvidia.com\/content\/pdf\/tesla\/whitepaper\/pascal-architecture-whitepaper.pdf."},{"key":"e_1_2_1_102_1","unstructured":"NVIDIA. 2016b. GPU Cloud Computing Service Providers\u2014NVIDIA. Retrieved from http:\/\/www.nvidia.com\/object\/gpu-cloud-computing-services.html.  NVIDIA. 2016b. GPU Cloud Computing Service Providers\u2014NVIDIA. Retrieved from http:\/\/www.nvidia.com\/object\/gpu-cloud-computing-services.html."},{"key":"e_1_2_1_103_1","unstructured":"CUDA Nvidia. 2007b. Compute Unified Device Architecture Programming Guide. http:\/\/docs.nvidia.com\/cuda\/cuda-c-programming-guide\/index.html.  CUDA Nvidia. 2007b. Compute Unified Device Architecture Programming Guide. http:\/\/docs.nvidia.com\/cuda\/cuda-c-programming-guide\/index.html."},{"key":"e_1_2_1_104_1","volume-title":"Discrete-Time Control Systems","author":"Ogata Katsuhiko","unstructured":"Katsuhiko Ogata . 1995. Discrete-Time Control Systems . Vol. 2 . Prentice Hall, Englewood Cliffs , NJ. Katsuhiko Ogata. 1995. Discrete-Time Control Systems. Vol. 2. Prentice Hall, Englewood Cliffs, NJ."},{"key":"e_1_2_1_105_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.Companion.2012.146"},{"key":"e_1_2_1_106_1","volume-title":"4th USENIX Workshop on Hot Topics in Cloud Computing (HotCloud).","author":"Ou Zhonghong","year":"2012","unstructured":"Zhonghong Ou , Hao Zhuang , Jukka K. Nurminen , Antti Yl\u00e4-J\u00e4\u00e4ski , and Pan Hui . 2012 . Exploiting hardware heterogeneity within the same instance type of Amazon EC2 . Presented in the 4th USENIX Workshop on Hot Topics in Cloud Computing (HotCloud). Zhonghong Ou, Hao Zhuang, Jukka K. Nurminen, Antti Yl\u00e4-J\u00e4\u00e4ski, and Pan Hui. 2012. Exploiting hardware heterogeneity within the same instance type of Amazon EC2. Presented in the 4th USENIX Workshop on Hot Topics in Cloud Computing (HotCloud)."},{"key":"e_1_2_1_107_1","doi-asserted-by":"publisher","DOI":"10.5555\/2342788.2342792"},{"key":"e_1_2_1_108_1","volume-title":"Proceedings of the 10th USENEX Conference on File and Storage Technologies (FAST\u201912)","author":"Park Stan","year":"2012","unstructured":"Stan Park and Kai Shen . 2012 . FIOS: A fair, efficient flash I\/O scheduler . In Proceedings of the 10th USENEX Conference on File and Storage Technologies (FAST\u201912) . 13. Stan Park and Kai Shen. 2012. FIOS: A fair, efficient flash I\/O scheduler. In Proceedings of the 10th USENEX Conference on File and Storage Technologies (FAST\u201912). 13."},{"key":"e_1_2_1_109_1","unstructured":"PathScale. 2012. pathscale\/pscnv. Retrieved from https:\/\/github.com\/pathscale\/pscnv.  PathScale. 2012. pathscale\/pscnv. Retrieved from https:\/\/github.com\/pathscale\/pscnv."},{"key":"e_1_2_1_110_1","doi-asserted-by":"publisher","DOI":"10.1109\/NGCT.2015.7375072"},{"key":"e_1_2_1_111_1","volume-title":"The top 10 innovations in the new NVIDIA fermi architecture, and the top 3 next challenges. NVIDIA Whitepaper 47","author":"Patterson David","year":"2009","unstructured":"David Patterson . 2009. The top 10 innovations in the new NVIDIA fermi architecture, and the top 3 next challenges. NVIDIA Whitepaper 47 ( 2009 ). David Patterson. 2009. The top 10 innovations in the new NVIDIA fermi architecture, and the top 3 next challenges. NVIDIA Whitepaper 47 (2009)."},{"key":"e_1_2_1_112_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2014.09.011"},{"key":"e_1_2_1_113_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-319-39577-7_7"},{"key":"e_1_2_1_114_1","unstructured":"Antoine Petitet. 2004. HPL-A portable implementation of the high-performance Linpack benchmark for distributed-memory computers. Retrieved from http:\/\/www.netlib-.org\/-benchmark\/hpl\/.  Antoine Petitet. 2004. HPL-A portable implementation of the high-performance Linpack benchmark for distributed-memory computers. Retrieved from http:\/\/www.netlib-.org\/-benchmark\/hpl\/."},{"key":"e_1_2_1_115_1","doi-asserted-by":"publisher","DOI":"10.1002\/jcc.20289"},{"key":"e_1_2_1_116_1","volume-title":"LAMMPS-large-scale atomic\/molecular massively parallel simulator. Sandia National Laboratories 18","author":"Plimpton Steve","year":"2007","unstructured":"Steve Plimpton , Paul Crozier , and Aidan Thompson . 2007. LAMMPS-large-scale atomic\/molecular massively parallel simulator. Sandia National Laboratories 18 ( 2007 ). http:\/\/lammps.sandia.gov. Steve Plimpton, Paul Crozier, and Aidan Thompson. 2007. LAMMPS-large-scale atomic\/molecular massively parallel simulator. Sandia National Laboratories 18 (2007). http:\/\/lammps.sandia.gov."},{"key":"e_1_2_1_117_1","doi-asserted-by":"publisher","DOI":"10.1145\/2851141.2851181"},{"key":"e_1_2_1_118_1","doi-asserted-by":"publisher","DOI":"10.1145\/2632216"},{"key":"e_1_2_1_119_1","volume-title":"Toward a paravirtual vRDMA device for VMware ESXi guests. VMware Techn. J. 2012 1, 2","author":"Ranadive Adit","year":"2012","unstructured":"Adit Ranadive and Bhavesh Davda . 2012. Toward a paravirtual vRDMA device for VMware ESXi guests. VMware Techn. J. 2012 1, 2 ( 2012 ). Adit Ranadive and Bhavesh Davda. 2012. Toward a paravirtual vRDMA device for VMware ESXi guests. VMware Techn. J. 2012 1, 2 (2012)."},{"key":"e_1_2_1_120_1","doi-asserted-by":"publisher","DOI":"10.1145\/1996130.1996160"},{"key":"e_1_2_1_121_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER.2013.6702662"},{"key":"e_1_2_1_122_1","doi-asserted-by":"publisher","DOI":"10.1109\/HiPC.2012.6507485"},{"key":"e_1_2_1_123_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLUSTER.2015.76"},{"key":"e_1_2_1_124_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.3409"},{"key":"e_1_2_1_125_1","doi-asserted-by":"publisher","DOI":"10.1145\/2830013.2830015"},{"key":"e_1_2_1_126_1","doi-asserted-by":"publisher","DOI":"10.1145\/2043556.2043579"},{"key":"e_1_2_1_127_1","doi-asserted-by":"publisher","DOI":"10.1038\/nrg2857-c2"},{"key":"e_1_2_1_128_1","doi-asserted-by":"publisher","DOI":"10.1145\/2465829.2465830"},{"key":"e_1_2_1_129_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2014.47"},{"key":"e_1_2_1_130_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00450-011-0157-1"},{"key":"e_1_2_1_131_1","volume-title":"Proceedings of the Xen Project Developer Summit.","author":"Shan Haitao","year":"2013","unstructured":"Haitao Shan , Kevin Tian , Eddie Dong , and David Cowperthwaite . 2013 . XenGT: A software based intel graphics virtualization solution . Proceedings of the Xen Project Developer Summit. Haitao Shan, Kevin Tian, Eddie Dong, and David Cowperthwaite. 2013. XenGT: A software based intel graphics virtualization solution. Proceedings of the Xen Project Developer Summit."},{"key":"e_1_2_1_132_1","doi-asserted-by":"publisher","DOI":"10.5555\/2664633.2664640"},{"key":"e_1_2_1_133_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2009.5161020"},{"key":"e_1_2_1_134_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2011.112"},{"key":"e_1_2_1_135_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jnca.2010.06.005"},{"key":"e_1_2_1_136_1","doi-asserted-by":"publisher","DOI":"10.1109\/90.502236"},{"key":"e_1_2_1_137_1","unstructured":"Abraham Silberschatz Peter B. Galvin Greg Gagne and A. Silberschatz. 1998. Operating System Concepts. Vol. 4. Addison-Wesley Reading MA.  Abraham Silberschatz Peter B. Galvin Greg Gagne and A. Silberschatz. 1998. Operating System Concepts. Vol. 4. Addison-Wesley Reading MA."},{"key":"e_1_2_1_138_1","first-page":"2014","volume-title":"KVMGT: A full GPU virtualization solution. In KVM Forum","author":"Song Jike","year":"2014","unstructured":"Jike Song , Zhiyuan Lv , and Kevin Tian . 2014 . KVMGT: A full GPU virtualization solution. In KVM Forum 2014. http:\/\/www.linux-kvm.org\/page\/KVM_Forum_ 2014 . Jike Song, Zhiyuan Lv, and Kevin Tian. 2014. KVMGT: A full GPU virtualization solution. In KVM Forum 2014. http:\/\/www.linux-kvm.org\/page\/KVM_Forum_2014."},{"key":"e_1_2_1_139_1","volume-title":"Geng Daniel Liu, and Wen-mei W. Hwu","author":"Stratton John A.","year":"2012","unstructured":"John A. Stratton , Christopher Rodrigues , I- Jui Sung , Nady Obeid , Li-Wen Chang , Nasser Anssari , Geng Daniel Liu, and Wen-mei W. Hwu . 2012 . Parboil : A revised benchmark suite for scientific and commercial throughput computing. Center for Reliable and High-Performance Computing 127 (2012). John A. Stratton, Christopher Rodrigues, I-Jui Sung, Nady Obeid, Li-Wen Chang, Nasser Anssari, Geng Daniel Liu, and Wen-mei W. Hwu. 2012. Parboil: A revised benchmark suite for scientific and commercial throughput computing. Center for Reliable and High-Performance Computing 127 (2012)."},{"key":"e_1_2_1_140_1","volume-title":"Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914)","author":"Suzuki Yusuke","year":"2014","unstructured":"Yusuke Suzuki , Shinpei Kato , Hiroshi Yamada , and Kenji Kono . 2014 . GPUvm: Why not virtualizing GPUs at the hypervisor? . In Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914) . 109--120. Yusuke Suzuki, Shinpei Kato, Hiroshi Yamada, and Kenji Kono. 2014. GPUvm: Why not virtualizing GPUs at the hypervisor?. In Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914). 109--120."},{"key":"e_1_2_1_141_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2015.2506582"},{"key":"e_1_2_1_142_1","doi-asserted-by":"publisher","DOI":"10.1145\/2678373.2665702"},{"key":"e_1_2_1_143_1","volume-title":"Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914)","author":"Tian Kun","year":"2014","unstructured":"Kun Tian , Yaozu Dong , and David Cowperthwaite . 2014 . A full GPU virtualization solution with mediated pass-through . In Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914) . Kun Tian, Yaozu Dong, and David Cowperthwaite. 2014. A full GPU virtualization solution with mediated pass-through. In Proceedings of the 2014 USENIX Annual Technical Conference (USENIX ATC\u201914)."},{"key":"e_1_2_1_144_1","doi-asserted-by":"publisher","DOI":"10.1002\/spe.2166"},{"key":"e_1_2_1_145_1","unstructured":"Top500. 2016. TOP500 Supercomputer Sites. Retrieved from https:\/\/www.top500.org\/list\/2016\/06\/.  Top500. 2016. TOP500 Supercomputer Sites. Retrieved from https:\/\/www.top500.org\/list\/2016\/06\/."},{"key":"e_1_2_1_146_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2005.163"},{"key":"e_1_2_1_147_1","doi-asserted-by":"publisher","DOI":"10.1145\/1134760.1134762"},{"key":"e_1_2_1_148_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2006.393"},{"key":"e_1_2_1_149_1","volume-title":"Microsoft Virtualization with Hyper-V","author":"Velte Anthony","unstructured":"Anthony Velte and Toby Velte . 2009. Microsoft Virtualization with Hyper-V . McGraw-Hill, Inc. Anthony Velte and Toby Velte. 2009. Microsoft Virtualization with Hyper-V. McGraw-Hill, Inc."},{"key":"e_1_2_1_150_1","doi-asserted-by":"publisher","DOI":"10.1109\/PDGC.2012.6449892"},{"key":"e_1_2_1_151_1","volume-title":"Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 2.","author":"Vu Lan","year":"2014","unstructured":"Lan Vu , Hari Sivaraman , and Rishi Bidarkar . 2014 . GPU virtualization for high performance general purpose computing on the ESX hypervisor . In Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 2. Lan Vu, Hari Sivaraman, and Rishi Bidarkar. 2014. GPU virtualization for high performance general purpose computing on the ESX hypervisor. In Proceedings of the High Performance Computing Symposium. Society for Computer Simulation International, 2."},{"key":"e_1_2_1_152_1","doi-asserted-by":"publisher","DOI":"10.1109\/CLOUD.2014.90"},{"key":"e_1_2_1_153_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.future.2016.03.011"},{"key":"e_1_2_1_154_1","doi-asserted-by":"publisher","DOI":"10.1145\/1383422.1383437"},{"key":"e_1_2_1_155_1","doi-asserted-by":"publisher","DOI":"10.1145\/1456455.1456460"},{"key":"e_1_2_1_156_1","doi-asserted-by":"publisher","DOI":"10.1109\/MM.2011.24"},{"key":"e_1_2_1_157_1","volume-title":"Version 1.2","author":"Woo Mason","unstructured":"Mason Woo , Jackie Neider , Tom Davis , and Dave Shreiner . 1999. OpenGL Programming Guide: The Official Guide to Learning OpenGL , Version 1.2 . Addison-Wesley Longman Publishing Co., Inc. Mason Woo, Jackie Neider, Tom Davis, and Dave Shreiner. 1999. OpenGL Programming Guide: The Official Guide to Learning OpenGL, Version 1.2. Addison-Wesley Longman Publishing Co., Inc."},{"key":"e_1_2_1_158_1","doi-asserted-by":"publisher","DOI":"10.1109\/CCGrid.2011.51"},{"key":"e_1_2_1_159_1","unstructured":"Xenproject. 2016. Xen Project Release Features. Retrieved from https:\/\/wiki.xenproject.org\/wiki\/Xen_Project_Release_Features.  Xenproject. 2016. Xen Project Release Features. Retrieved from https:\/\/wiki.xenproject.org\/wiki\/Xen_Project_Release_Features."},{"key":"e_1_2_1_160_1","doi-asserted-by":"publisher","DOI":"10.1109\/InPar.2012.6339609"},{"key":"e_1_2_1_161_1","volume-title":"Nouveau: Accelerated Open Source driver for nVidia cards.","author":"OrgFoundation X.","year":"2011","unstructured":"X. OrgFoundation . 2011 . Nouveau: Accelerated Open Source driver for nVidia cards. Retrieved from https:\/\/nouveau.freedesktop.org\/wiki\/. X.OrgFoundation. 2011. Nouveau: Accelerated Open Source driver for nVidia cards. Retrieved from https:\/\/nouveau.freedesktop.org\/wiki\/."},{"key":"e_1_2_1_162_1","volume-title":"Proceedings of the 2016 USENIX Annual Technical Conference (USENIX ATC\u201916)","author":"Xue Mochi","year":"2016","unstructured":"Mochi Xue , Kun Tian , Yaozu Dong , Jiajun Wang , Zhengwei Qi , Bingsheng He , and Haibing Guan . 2016 . gScale: Scaling up GPU virtualization with dynamic sharing of graphics memory space . In Proceedings of the 2016 USENIX Annual Technical Conference (USENIX ATC\u201916) . Mochi Xue, Kun Tian, Yaozu Dong, Jiajun Wang, Zhengwei Qi, Bingsheng He, and Haibing Guan. 2016. gScale: Scaling up GPU virtualization with dynamic sharing of graphics memory space. In Proceedings of the 2016 USENIX Annual Technical Conference (USENIX ATC\u201916)."},{"key":"e_1_2_1_163_1","doi-asserted-by":"publisher","DOI":"10.1007\/s11227-013-1034-4"},{"key":"e_1_2_1_164_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-35606-3_53"},{"key":"e_1_2_1_165_1","doi-asserted-by":"publisher","DOI":"10.1109\/CloudCom.2012.6427531"},{"key":"e_1_2_1_166_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-38027-3_45"},{"key":"e_1_2_1_167_1","doi-asserted-by":"publisher","DOI":"10.1145\/2688500.2688505"},{"key":"e_1_2_1_168_1","doi-asserted-by":"publisher","DOI":"10.1109\/CCGrid.2014.93"},{"key":"e_1_2_1_169_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPSW.2014.97"},{"key":"e_1_2_1_170_1","doi-asserted-by":"publisher","DOI":"10.1145\/2731186.2731194"},{"key":"e_1_2_1_171_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2013.288"},{"key":"e_1_2_1_172_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2015.2433916"},{"key":"e_1_2_1_173_1","doi-asserted-by":"publisher","DOI":"10.1109\/RTAS.2015.7108420"}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3068281","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3068281","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T03:03:44Z","timestamp":1750215824000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3068281"}},"subtitle":["A Comprehensive Survey"],"short-title":[],"issued":{"date-parts":[[2017,6,29]]},"references-count":171,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2018,5,31]]}},"alternative-id":["10.1145\/3068281"],"URL":"https:\/\/doi.org\/10.1145\/3068281","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2017,6,29]]},"assertion":[{"value":"2016-10-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-03-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2017-06-29","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}