{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:28:01Z","timestamp":1750307281482,"version":"3.41.0"},"reference-count":18,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2011,2,18]],"date-time":"2011-02-18T00:00:00Z","timestamp":1297987200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGOPS Oper. Syst. Rev."],"published-print":{"date-parts":[[2011,2,18]]},"abstract":"<jats:p>The client computing platform is moving towards a heterogeneous architecture that combines scalar-oriented CPU cores and throughput-oriented accelerator cores. Recognizing that existing programming models for such heterogeneous platforms are still difficult for most programmers, we advocate a shared virtual memory programming model to improve programmability. In this paper, we focus on performance, and demonstrate that users need not sacrifice performance for programmability. We describe our approaches, experiences, and results in optimizing MYO on a heterogeneous platform consisting of a CPU and an Aubrey Isle accelerator. Our efforts involve the whole system software stack including the OS, runtime, and application.<\/jats:p>","DOI":"10.1145\/1945023.1945035","type":"journal-article","created":{"date-parts":[[2011,3,1]],"date-time":"2011-03-01T20:14:26Z","timestamp":1299010466000},"page":"92-100","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Optimizing a shared virtual memory system for a heterogeneous CPU-accelerator platform"],"prefix":"10.1145","volume":"45","author":[{"given":"Shoumeng","family":"Yan","sequence":"first","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Xiaocheng","family":"Zhou","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ying","family":"Gao","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Hu","family":"Chen","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Gansha","family":"Wu","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Sai","family":"Luo","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Bratin","family":"Saha","sequence":"additional","affiliation":[{"name":"Intel Corporation"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2011,2,18]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1360612.1360617"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/1542476.1542525"},{"key":"e_1_2_1_3_1","unstructured":"Nvidia Corp CUDA Programming Environment www.nvidia.com\/object\/cuda_what_is.html.  Nvidia Corp CUDA Programming Environment www.nvidia.com\/object\/cuda_what_is.html."},{"key":"e_1_2_1_4_1","unstructured":"AMD CTM http:\/\/ati.amd.com\/companyinfo\/researcher\/documents\/ATI_CTM_Guide.pdf  AMD CTM http:\/\/ati.amd.com\/companyinfo\/researcher\/documents\/ATI_CTM_Guide.pdf"},{"key":"e_1_2_1_5_1","unstructured":"AMD Stream SDK ati.amd.com\/technology\/streamcomputing.  AMD Stream SDK ati.amd.com\/technology\/streamcomputing."},{"key":"e_1_2_1_6_1","unstructured":"Dubey P. Recognition Mining and Synthesis moves computers to the era of tera. Technology@Intel Feb 2005.  Dubey P. Recognition Mining and Synthesis moves computers to the era of tera. Technology@Intel Feb 2005."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/1103900.1103933"},{"key":"e_1_2_1_8_1","first-page":"150","volume-title":"High Performance Computing (HiPC), 2009 International Conference on, vol., no.","author":"Shoumeng Yan","year":"2009"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1145\/844128.844138"},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1229428.1229483"},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1109\/71.159038"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/1736020.1736059"},{"key":"e_1_2_1_13_1","unstructured":"Zhigang Wang Gansha Wu Zhaohui Du Zhanglin Liu Yongjian Chen Peng Guo Dan Zhang Anwar Ghuloum and Chris Newburn. Compiler-guided smart data transfer for accelerators Submitted to CGO 2011 April 2-6 2011.  Zhigang Wang Gansha Wu Zhaohui Du Zhanglin Liu Yongjian Chen Peng Guo Dan Zhang Anwar Ghuloum and Chris Newburn. Compiler-guided smart data transfer for accelerators Submitted to CGO 2011 April 2-6 2011."},{"key":"e_1_2_1_14_1","doi-asserted-by":"publisher","DOI":"10.1145\/237090.237181"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1023\/B:IJPP.0000023480.82632.87"},{"key":"e_1_2_1_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/248208.237185"},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/277830.277925"},{"key":"e_1_2_1_18_1","unstructured":"Intel \"Intel news release: Intel unveils new product plans for high-performance computing \" http:\/\/www.intel.com\/pressroom\/archive\/releases\/20100531comp.htm May 2010.  Intel \"Intel news release: Intel unveils new product plans for high-performance computing \" http:\/\/www.intel.com\/pressroom\/archive\/releases\/20100531comp.htm May 2010."}],"container-title":["ACM SIGOPS Operating Systems Review"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1945023.1945035","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/1945023.1945035","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T10:59:31Z","timestamp":1750244371000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1945023.1945035"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2011,2,18]]},"references-count":18,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2011,2,18]]}},"alternative-id":["10.1145\/1945023.1945035"],"URL":"https:\/\/doi.org\/10.1145\/1945023.1945035","relation":{},"ISSN":["0163-5980"],"issn-type":[{"type":"print","value":"0163-5980"}],"subject":[],"published":{"date-parts":[[2011,2,18]]},"assertion":[{"value":"2011-02-18","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}