{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,24]],"date-time":"2026-04-24T01:20:38Z","timestamp":1776993638898,"version":"3.51.4"},"publisher-location":"New York, NY, USA","reference-count":53,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,4,7]],"date-time":"2021-04-07T00:00:00Z","timestamp":1617753600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,4,16]]},"DOI":"10.1145\/3453933.3454019","type":"proceedings-article","created":{"date-parts":[[2021,4,8]],"date-time":"2021-04-08T05:33:57Z","timestamp":1617860037000},"page":"125-138","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":7,"title":["Multiple-tasks on multiple-devices (MTMD): exploiting concurrency in heterogeneous managed runtimes"],"prefix":"10.1145","author":[{"given":"Michail","family":"Papadimitriou","sequence":"first","affiliation":[{"name":"University of Manchester, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Eleni","family":"Markou","sequence":"additional","affiliation":[{"name":"BEAT, Greece"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Juan","family":"Fumero","sequence":"additional","affiliation":[{"name":"University of Manchester, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Athanasios","family":"Stratikopoulos","sequence":"additional","affiliation":[{"name":"University of Manchester, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Florin","family":"Blanaru","sequence":"additional","affiliation":[{"name":"University of Manchester, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Christos","family":"Kotselidis","sequence":"additional","affiliation":[{"name":"University of Manchester, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2021,4,7]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/3306346.3322967"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.parco.2016.05.006"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-24322-6_6"},{"key":"e_1_3_2_1_4_1","unstructured":"AMD. Accessed in 2020. Aparapi project. https:\/\/aparapi.github.io\/  AMD. Accessed in 2020. Aparapi project. https:\/\/aparapi.github.io\/"},{"key":"e_1_3_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/SBAC-PAD.2014.30"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-40994-3_29"},{"key":"e_1_3_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1016\/S0031-3203(96)00142-2"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/3431731"},{"key":"e_1_3_2_1_9_1","article-title":"SMOTE: Synthetic Minority over-Sampling Technique","volume":"16","author":"Chawla Nitesh V.","year":"2002","unstructured":"Nitesh V. Chawla , Kevin W. Bowyer , Lawrence O. Hall , and W. Philip Kegelmeyer . 2002 . SMOTE: Synthetic Minority over-Sampling Technique . J. Artif. Int. Res. 16 , 1 (2002). issn:1076-9757 Nitesh V. Chawla, Kevin W. Bowyer, Lawrence O. Hall, and W. Philip Kegelmeyer. 2002. SMOTE: Synthetic Minority over-Sampling Technique. J. Artif. Int. Res. 16, 1 (2002). issn:1076-9757","journal-title":"J. Artif. Int. Res."},{"key":"e_1_3_2_1_10_1","volume-title":"TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI).","author":"Chen Tianqi","year":"2018","unstructured":"Tianqi Chen , Thierry Moreau , Ziheng Jiang , Lianmin Zheng , Eddie Yan , Haichen Shen , Meghan Cowan , Leyuan Wang , Yuwei Hu , Luis Ceze , Carlos Guestrin , and Arvind Krishnamurthy . 2018 . TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI). Tianqi Chen, Thierry Moreau, Ziheng Jiang, Lianmin Zheng, Eddie Yan, Haichen Shen, Meghan Cowan, Leyuan Wang, Yuwei Hu, Luis Ceze, Carlos Guestrin, and Arvind Krishnamurthy. 2018. TVM: An Automated End-to-End Optimizing Compiler for Deep Learning. In 13th USENIX Symposium on Operating Systems Design and Implementation (OSDI)."},{"key":"e_1_3_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/3191697.3191730"},{"key":"e_1_3_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3237009.3237016"},{"key":"e_1_3_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/202529.202534"},{"key":"e_1_3_2_1_14_1","volume-title":"CUDA Programming: A Developer?s Guide to Parallel Computing with GPUs","author":"Cook Shane","unstructured":"Shane Cook . 2012. CUDA Programming: A Developer?s Guide to Parallel Computing with GPUs ( 1 st ed.). Morgan Kaufmann Publishers Inc . Shane Cook. 2012. CUDA Programming: A Developer?s Guide to Parallel Computing with GPUs (1st ed.). Morgan Kaufmann Publishers Inc.","edition":"1"},{"key":"e_1_3_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2542142.2542143"},{"key":"e_1_3_2_1_16_1","volume-title":"The Gini Index and Measures of Inequality. The American Mathematical Monthly 117, 10","author":"Farris Frank A.","year":"2010","unstructured":"Frank A. Farris . 2010. The Gini Index and Measures of Inequality. The American Mathematical Monthly 117, 10 ( 2010 ), 851?864. Frank A. Farris. 2010. The Gini Index and Measures of Inequality. The American Mathematical Monthly 117, 10 (2010), 851?864."},{"key":"e_1_3_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313808.3313819"},{"key":"e_1_3_2_1_18_1","unstructured":"V. M. Garcia A. Liberos A. M. Climent A. Vidal J. Millet and A. Gonz\\'alez. 2011. An adaptive step size GPU ODE solver for simulating the electric cardiac activity. In Computing in Cardiology. 233?236.  V. M. Garcia A. Liberos A. M. Climent A. Vidal J. Millet and A. Gonz\\'alez. 2011. An adaptive step size GPU ODE solver for simulating the electric cardiac activity. In Computing in Cardiology. 233?236."},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/1297105.1297033"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10994-006-6226-1"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"crossref","unstructured":"Anirban Ghose Siddharth Singh Vivek Kulaharia Lokesh Dokara Srijeeta Maity and Soumyajit Dey. 2020. PySchedCL: Leveraging Concurrency in Heterogeneous Data-Parallel Systems. arxiv:2009.07482 [cs.DC]  Anirban Ghose Siddharth Singh Vivek Kulaharia Lokesh Dokara Srijeeta Maity and Soumyajit Dey. 2020. PySchedCL: Leveraging Concurrency in Heterogeneous Data-Parallel Systems. arxiv:2009.07482 [cs.DC]","DOI":"10.1109\/TC.2021.3125792"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2008.5213922"},{"key":"e_1_3_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/2458523.2458536"},{"key":"e_1_3_2_1_24_1","volume-title":"O'Boyle","author":"Grewe Dominik","year":"2011","unstructured":"Dominik Grewe and Michael F. P . O'Boyle . 2011 . A Static Task Partitioning Approach for Heterogeneous Systems Using OpenCL. In Compiler Construction, Jens Knoop (Ed.). Springer Berlin Heidelberg . Dominik Grewe and Michael F. P. O'Boyle. 2011. A Static Task Partitioning Approach for Heterogeneous Systems Using OpenCL. In Compiler Construction, Jens Knoop (Ed.). Springer Berlin Heidelberg."},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1109\/CGO.2013.6494993"},{"key":"e_1_3_2_1_26_1","volume-title":"SOCL: An OpenCL Implementation with Automatic Multi-Device Adaptation Support. Research Report RR-8346. INRIA. 18 pages. https:\/\/hal.inria.fr\/hal-00853423","author":"Henry Sylvain","year":"2013","unstructured":"Sylvain Henry , Denis Barthou , Alexandre Denis , Raymond Namyst , and Marie-Christine Counilh . 2013 . SOCL: An OpenCL Implementation with Automatic Multi-Device Adaptation Support. Research Report RR-8346. INRIA. 18 pages. https:\/\/hal.inria.fr\/hal-00853423 Sylvain Henry, Denis Barthou, Alexandre Denis, Raymond Namyst, and Marie-Christine Counilh. 2013. SOCL: An OpenCL Implementation with Automatic Multi-Device Adaptation Support. Research Report RR-8346. INRIA. 18 pages. https:\/\/hal.inria.fr\/hal-00853423"},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1109\/TELFOR.2015.7377632"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/2145816.2145818"},{"key":"e_1_3_2_1_29_1","unstructured":"IBM. [n.d.]. https:\/\/www.ibm.com\/support\/knowledgecenter\/en\/SSYKE2_7.0.0\/com.ibm.java.win.70.doc\/user\/java_jvm.html  IBM. [n.d.]. https:\/\/www.ibm.com\/support\/knowledgecenter\/en\/SSYKE2_7.0.0\/com.ibm.java.win.70.doc\/user\/java_jvm.html"},{"key":"e_1_3_2_1_30_1","unstructured":"Intel. [n.d.]. oneAPI Specification. https:\/\/spec.oneapi.com\/versions\/latest\/index.html  Intel. [n.d.]. oneAPI Specification. https:\/\/spec.oneapi.com\/versions\/latest\/index.html"},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1109\/ECAI.2017.8166501"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1109\/PACT.2013.6618799"},{"key":"e_1_3_2_1_33_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.jpdc.2019.05.015"},{"key":"e_1_3_2_1_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2304576.2304623"},{"key":"e_1_3_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/2464996.2465007"},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1145\/3050748.3050764"},{"key":"e_1_3_2_1_37_1","volume-title":"42nd Annual IEEE\/ACM International Symposium on Microarchitecture (MICRO).","author":"Luk C.","unstructured":"C. Luk , S. Hong , and H. Kim . 2009. Qilin: Exploiting parallelism on heterogeneous multiprocessors with adaptive mapping . In 42nd Annual IEEE\/ACM International Symposium on Microarchitecture (MICRO). C. Luk, S. Hong, and H. Kim. 2009. Qilin: Exploiting parallelism on heterogeneous multiprocessors with adaptive mapping. In 42nd Annual IEEE\/ACM International Symposium on Microarchitecture (MICRO)."},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICRA.2015.7140009"},{"key":"e_1_3_2_1_39_1","doi-asserted-by":"publisher","DOI":"10.1109\/hpcs48598.2019.9188188"},{"key":"e_1_3_2_1_40_1","doi-asserted-by":"crossref","unstructured":"William F. Ogilvie Pavlos Petoumenos Zheng Wang and Hugh Leather. 2015. Fast Automatic Heuristic Construction Using Active Learning. In Languages and Compilers for Parallel Computing James Brodman and Peng Tu (Eds.).  William F. Ogilvie Pavlos Petoumenos Zheng Wang and Hugh Leather. 2015. Fast Automatic Heuristic Construction Using Active Learning. In Languages and Compilers for Parallel Computing James Brodman and Peng Tu (Eds.).","DOI":"10.1007\/978-3-319-17473-0_10"},{"key":"e_1_3_2_1_41_1","volume-title":"Supercomputing Frontiers","author":"Ohshima Satoshi","unstructured":"Satoshi Ohshima , Ichitaro Yamazaki , Akihiro Ida , and Rio Yokota . 2018. Optimization of Hierarchical Matrix Computation on GPU . In Supercomputing Frontiers . Springer International Publishing . Satoshi Ohshima, Ichitaro Yamazaki, Akihiro Ida, and Rio Yokota. 2018. Optimization of Hierarchical Matrix Computation on GPU. In Supercomputing Frontiers. Springer International Publishing."},{"key":"e_1_3_2_1_42_1","doi-asserted-by":"publisher","DOI":"10.1145\/2581122.2544163"},{"key":"e_1_3_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1109\/FCCM.2019.00051"},{"key":"e_1_3_2_1_44_1","doi-asserted-by":"publisher","DOI":"10.22152\/programming-journal.org\/2021\/5\/8"},{"key":"e_1_3_2_1_45_1","volume-title":"Santambrogio","author":"Parravicini Alberto","year":"2020","unstructured":"Alberto Parravicini , Arnaud Delamare , Marco Arnaboldi , and Marco D . Santambrogio . 2020 . DAG-based Scheduling with Resource Sharing for Multi-task Applications in a Polyglot GPU Runtime . arxiv:2012.09646 [cs.DC] Alberto Parravicini, Arnaud Delamare, Marco Arnaboldi, and Marco D. Santambrogio. 2020. DAG-based Scheduling with Resource Sharing for Multi-task Applications in a Polyglot GPU Runtime. arxiv:2012.09646 [cs.DC]"},{"key":"e_1_3_2_1_46_1","volume-title":"Proceedings of the International Conference on Computer Design (CDES).","author":"Playne DP","year":"2009","unstructured":"DP Playne , MGB Johnson , and KA Hawick . 2009 . Benchmarking GPU Devices with N-Body Simulations . In Proceedings of the International Conference on Computer Design (CDES). DP Playne, MGB Johnson, and KA Hawick. 2009. Benchmarking GPU Devices with N-Body Simulations. In Proceedings of the International Conference on Computer Design (CDES)."},{"key":"e_1_3_2_1_47_1","volume-title":"Kroese","author":"Rubinstein Reuven Y.","year":"2016","unstructured":"Reuven Y. Rubinstein and Dirk P . Kroese . 2016 . Simulation and the Monte Carlo Method (3rd ed.). Wiley Publishing . isbn:1118632168 Reuven Y. Rubinstein and Dirk P. Kroese. 2016. Simulation and the Monte Carlo Method (3rd ed.). Wiley Publishing. isbn:1118632168"},{"key":"e_1_3_2_1_48_1","doi-asserted-by":"publisher","DOI":"10.1145\/3126548"},{"key":"e_1_3_2_1_49_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCSE.2010.69"},{"key":"e_1_3_2_1_50_1","doi-asserted-by":"publisher","DOI":"10.1109\/CGO.2009.20"},{"key":"e_1_3_2_1_51_1","doi-asserted-by":"publisher","DOI":"10.1109\/SC.2008.5214359"},{"key":"e_1_3_2_1_52_1","doi-asserted-by":"publisher","DOI":"10.1109\/HiPC.2014.7116910"},{"key":"e_1_3_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1145\/2688500.2688505"}],"event":{"name":"VEE '21: 17th ACM SIGPLAN\/SIGOPS International Conference on Virtual Execution Environments","location":"Virtual USA","acronym":"VEE '21","sponsor":["SIGPLAN ACM Special Interest Group on Programming Languages","SIGOPS ACM Special Interest Group on Operating Systems"]},"container-title":["Proceedings of the 17th ACM SIGPLAN\/SIGOPS International Conference on Virtual Execution Environments"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3453933.3454019","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3453933.3454019","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:17:42Z","timestamp":1750191462000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3453933.3454019"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,4,7]]},"references-count":53,"alternative-id":["10.1145\/3453933.3454019","10.1145\/3453933"],"URL":"https:\/\/doi.org\/10.1145\/3453933.3454019","relation":{},"subject":[],"published":{"date-parts":[[2021,4,7]]},"assertion":[{"value":"2021-04-07","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}