{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,2]],"date-time":"2026-06-02T22:21:41Z","timestamp":1780438901861,"version":"3.54.1"},"reference-count":71,"publisher":"Association for Computing Machinery (ACM)","issue":"13","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. VLDB Endow."],"published-print":{"date-parts":[[2022,9]]},"abstract":"<jats:p>The ever-increasing demand for high performance Big Data analytics and data processing, has paved the way for heterogeneous hardware accelerators, such as Graphics Processing Units (GPUs) and Field Programmable Gate Arrays (FPGAs), to be integrated into modern Big Data platforms. Currently, this integration comes at the cost of programmability since the end-user Application Programming Interface (APIs) must be altered to access the underlying heterogeneous hardware. For example, current Big Data frameworks, such as Apache Spark, provide a new API that combines the existing Spark programming model with GPUs. For other Big Data frameworks, such as Flink, the integration of GPUs and FPGAs is achieved via external API calls that bypass their execution models completely.<\/jats:p>\n          <jats:p>In this paper, we rethink current Big Data frameworks from a systems and programming language perspective, and introduce a novel co-designed approach for integrating hardware acceleration into their execution models. The novelty of our approach is attributed to two key design decisions: a) support for arbitrary User Defined Functions (UDFs), and b) no modifications to the user level API. The proposed approach has been prototyped in the context of Apache Flink, and enables unmodified applications written in Java to run on heterogeneous hardware, such as GPU and FPGAs, transparently to the users. The performance evaluation of the proposed solution has shown performance speedups of up to 65x on GPUs and 184x on FPGAs for suitable workloads of standard benchmarks and industrial use cases against vanilla Flink running on traditional multi-core CPUs.<\/jats:p>","DOI":"10.14778\/3565838.3565842","type":"journal-article","created":{"date-parts":[[2023,1,20]],"date-time":"2023-01-20T23:09:56Z","timestamp":1674256196000},"page":"3869-3882","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":9,"title":["Enabling Transparent Acceleration of Big Data Frameworks Using Heterogeneous Hardware"],"prefix":"10.14778","volume":"15","author":[{"given":"Maria","family":"Xekalaki","sequence":"first","affiliation":[{"name":"The University of Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Juan","family":"Fumero","sequence":"additional","affiliation":[{"name":"The University of Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Athanasios","family":"Stratikopoulos","sequence":"additional","affiliation":[{"name":"The University of Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Katerina","family":"Doka","sequence":"additional","affiliation":[{"name":"National Technical University of Athens, Greece"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Christos","family":"Katsakioris","sequence":"additional","affiliation":[{"name":"National Technical University of Athens, Greece"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Constantinos","family":"Bitsakos","sequence":"additional","affiliation":[{"name":"National Technical University of Athens, Greece"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Nectarios","family":"Koziris","sequence":"additional","affiliation":[{"name":"National Technical University of Athens, Greece"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Christos","family":"Kotselidis","sequence":"additional","affiliation":[{"name":"The University of Manchester, UK"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2023,1,20]]},"reference":[{"key":"e_1_2_1_1_1","unstructured":"Accessed in October 2022. RAPIDS. https:\/\/rapids.ai\/  Accessed in October 2022. RAPIDS. https:\/\/rapids.ai\/"},{"key":"e_1_2_1_2_1","volume-title":"The 16th CSI International Symposium on Computer Architecture and Digital Systems (CADS)","author":"Abbasi Amin","year":"2012","unstructured":"Amin Abbasi , Farshad Khunjush , and Reza Azimi . 2012 . A preliminary study of incorporating GPUs in the Hadoop framework . The 16th CSI International Symposium on Computer Architecture and Digital Systems (CADS) (2012), 178--185. Amin Abbasi, Farshad Khunjush, and Reza Azimi. 2012. A preliminary study of incorporating GPUs in the Hadoop framework. The 16th CSI International Symposium on Computer Architecture and Digital Systems (CADS) (2012), 178--185."},{"key":"e_1_2_1_3_1","unstructured":"AMD. Accessed in October 2022. Aparapi project. https:\/\/aparapi.github.io\/  AMD. Accessed in October 2022. Aparapi project. https:\/\/aparapi.github.io\/"},{"key":"e_1_2_1_4_1","volume-title":"Accessed","year":"2022","unstructured":"Apache. Accessed in October 2022 . Apache Yarn GPU. https:\/\/hadoop.apache.org\/docs\/stable\/hadoop-yarn\/hadoop-yarn-site\/UsingGpus.html Apache. Accessed in October 2022. Apache Yarn GPU. https:\/\/hadoop.apache.org\/docs\/stable\/hadoop-yarn\/hadoop-yarn-site\/UsingGpus.html"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ISIE.2007.4374862"},{"key":"e_1_2_1_6_1","first-page":"28","article-title":"Apache Flink: Stream and Batch Processing in a Single Engine","volume":"38","author":"Carbone Paris","year":"2015","unstructured":"Paris Carbone , Asterios Katsifodimos , Stephan Ewen , Volker Markl , Seif Haridi , and Kostas Tzoumas . 2015 . Apache Flink: Stream and Batch Processing in a Single Engine . IEEE Data Eng. Bull. 38 (2015), 28 -- 38 . Paris Carbone, Asterios Katsifodimos, Stephan Ewen, Volker Markl, Seif Haridi, and Kostas Tzoumas. 2015. Apache Flink: Stream and Batch Processing in a Single Engine. IEEE Data Eng. Bull. 38 (2015), 28--38.","journal-title":"IEEE Data Eng. Bull."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1109\/TC.2018.2839719"},{"key":"e_1_2_1_8_1","doi-asserted-by":"crossref","first-page":"1275","DOI":"10.1109\/TPDS.2018.2794343","article-title":"GFlink: An In-Memory Computing Architecture on Heterogeneous CPU-GPU Clusters for Big Data","volume":"29","author":"Chen Cen","year":"2016","unstructured":"Cen Chen , Kenli Li , Aijia Ouyang , Zeng Zeng , and Keqin Li . 2016 . GFlink: An In-Memory Computing Architecture on Heterogeneous CPU-GPU Clusters for Big Data . IEEE Transactions on Parallel and Distributed Systems 29 (2016), 1275 -- 1288 . Cen Chen, Kenli Li, Aijia Ouyang, Zeng Zeng, and Keqin Li. 2016. GFlink: An In-Memory Computing Architecture on Heterogeneous CPU-GPU Clusters for Big Data. IEEE Transactions on Parallel and Distributed Systems 29 (2016), 1275--1288.","journal-title":"IEEE Transactions on Parallel and Distributed Systems"},{"key":"e_1_2_1_9_1","doi-asserted-by":"publisher","DOI":"10.1109\/TSMC.2017.2690673"},{"key":"e_1_2_1_10_1","volume-title":"2015 IEEE International Conference on Big Data (Big Data)","author":"Chen Zhenhua","year":"2015","unstructured":"Zhenhua Chen , Jielong Xu , Jian Tang , Kevin A. Kwiat , and Charles A. Kamhoua . 2015. G-Storm: GPU-enabled high-throughput online data processing in Storm . 2015 IEEE International Conference on Big Data (Big Data) ( 2015 ), 307--312. Zhenhua Chen, Jielong Xu, Jian Tang, Kevin A. Kwiat, and Charles A. Kamhoua. 2015. G-Storm: GPU-enabled high-throughput online data processing in Storm. 2015 IEEE International Conference on Big Data (Big Data) (2015), 307--312."},{"key":"e_1_2_1_11_1","volume-title":"Vispark: GPU-accelerated distributed visual computing using spark. In LDAV.","author":"Choi Woohyuk","year":"2015","unstructured":"Woohyuk Choi and Won-Ki Jeong . 2015 . Vispark: GPU-accelerated distributed visual computing using spark. In LDAV. Woohyuk Choi and Won-Ki Jeong. 2015. Vispark: GPU-accelerated distributed visual computing using spark. In LDAV."},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/3237009.3237016"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/508352.508353"},{"key":"e_1_2_1_14_1","volume-title":"Accessed","author":"NVIDIA Corporation","year":"2022","unstructured":"NVIDIA Corporation . Accessed in October 2022 . CUDA Toolkit Documentation . https:\/\/docs.nvidia.com\/cuda\/ NVIDIA Corporation. Accessed in October 2022. CUDA Toolkit Documentation. https:\/\/docs.nvidia.com\/cuda\/"},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1145\/2155620.2155676"},{"key":"e_1_2_1_16_1","volume-title":"IEEE International Conference on Cluster Computing and Workshops","author":"Farivar Reza","year":"2009","unstructured":"Reza Farivar , Abhishek Verma , Ellick Chan , and Roy H. Campbell . 2009. MITHRA: Multiple data independent tasks on a heterogeneous resource architecture . IEEE International Conference on Cluster Computing and Workshops ( 2009 ), 1--10. Reza Farivar, Abhishek Verma, Ellick Chan, and Roy H. Campbell. 2009. MITHRA: Multiple data independent tasks on a heterogeneous resource architecture. IEEE International Conference on Cluster Computing and Workshops (2009), 1--10."},{"key":"e_1_2_1_17_1","volume-title":"Accessed","author":"Flink Apache","year":"2022","unstructured":"Apache Flink . Accessed in October 2022 . Flink Checkpoints . https:\/\/nightlies.apache.org\/flink\/flink-docs-release-1.3\/setup\/checkpoints.html Apache Flink. Accessed in October 2022. Flink Checkpoints. https:\/\/nightlies.apache.org\/flink\/flink-docs-release-1.3\/setup\/checkpoints.html"},{"key":"e_1_2_1_18_1","volume-title":"Accessed","author":"Software Foundation The Apache","year":"2022","unstructured":"The Apache Software Foundation . Accessed in October 2022 . Accelerating your workload with GPU and other external resources. https:\/\/flink.apache.org\/news\/2020\/08\/06\/external-resource.html The Apache Software Foundation. Accessed in October 2022. Accelerating your workload with GPU and other external resources. https:\/\/flink.apache.org\/news\/2020\/08\/06\/external-resource.html"},{"key":"e_1_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/3313808.3313819"},{"key":"e_1_2_1_20_1","volume-title":"Accelerating Apache Spark Big Data Analysis with FPGAs","author":"Ghasemi Ehsan","year":"2016","unstructured":"Ehsan Ghasemi and Paul Chow . 2016. Accelerating Apache Spark Big Data Analysis with FPGAs . International IEEE Conferences on Ubiquitous Intelligence and Computing, Advanced and Trusted Computing, Scalable Computing and Communications, Cloud and Big Data Computing, Internet of People, and Smart World Congress (UIC\/ATC\/ScalCom\/CBDCom\/IoP\/SmartWorld) ( 2016 ), 737--744. Ehsan Ghasemi and Paul Chow. 2016. Accelerating Apache Spark Big Data Analysis with FPGAs. International IEEE Conferences on Ubiquitous Intelligence and Computing, Advanced and Trusted Computing, Scalable Computing and Communications, Cloud and Big Data Computing, Internet of People, and Smart World Congress (UIC\/ATC\/ScalCom\/CBDCom\/IoP\/SmartWorld) (2016), 737--744."},{"key":"e_1_2_1_21_1","volume-title":"Accelerating Apache Spark with FPGAs. Concurrency and Computation: Practice and Experience 31","author":"Ghasemi Ehsan","year":"2019","unstructured":"Ehsan Ghasemi and Paul Chow . 2019. Accelerating Apache Spark with FPGAs. Concurrency and Computation: Practice and Experience 31 ( 2019 ). Ehsan Ghasemi and Paul Chow. 2019. Accelerating Apache Spark with FPGAs. Concurrency and Computation: Practice and Experience 31 (2019)."},{"key":"e_1_2_1_22_1","volume-title":"IEEE International Symposium on Parallel and Distributed Processing, Workshops and Phd Forum (2013)","author":"Grossman Max","year":"2013","unstructured":"Max Grossman , Maur\u00edcio Breternitz , and Vivek Sarkar . 2013 . HadoopCL: MapReduce on Distributed Heterogeneous Platforms through Seamless Integration of Hadoop and OpenCL . IEEE International Symposium on Parallel and Distributed Processing, Workshops and Phd Forum (2013) , 1918--1927. Max Grossman, Maur\u00edcio Breternitz, and Vivek Sarkar. 2013. HadoopCL: MapReduce on Distributed Heterogeneous Platforms through Seamless Integration of Hadoop and OpenCL. IEEE International Symposium on Parallel and Distributed Processing, Workshops and Phd Forum (2013), 1918--1927."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2015.2414943"},{"key":"e_1_2_1_24_1","volume-title":"Proceedings of the Principles and Practices of Programming on The Java Platform","author":"Grossman Max","year":"2015","unstructured":"Max Grossman , Shams Mahmood Imam , and Vivek Sarkar . 2015 . HJ-OpenCL: Reducing the Gap Between the JVM and Accelerators . Proceedings of the Principles and Practices of Programming on The Java Platform (2015). Max Grossman, Shams Mahmood Imam, and Vivek Sarkar. 2015. HJ-OpenCL: Reducing the Gap Between the JVM and Accelerators. Proceedings of the Principles and Practices of Programming on The Java Platform (2015)."},{"key":"e_1_2_1_25_1","volume-title":"Proceedings of the 25th ACM International Symposium on High-Performance Parallel and Distributed Computing","author":"Grossman Max","year":"2016","unstructured":"Max Grossman and Vivek Sarkar . 2016 . SWAT: A Programmable, In-Memory, Distributed, High-Performance Computing Platform . Proceedings of the 25th ACM International Symposium on High-Performance Parallel and Distributed Computing (2016). Max Grossman and Vivek Sarkar. 2016. SWAT: A Programmable, In-Memory, Distributed, High-Performance Computing Platform. Proceedings of the 25th ACM International Symposium on High-Performance Parallel and Distributed Computing (2016)."},{"key":"e_1_2_1_26_1","volume-title":"Accessed","author":"Khronos OpneCL Working Group","year":"2022","unstructured":"Khronos OpneCL Working Group . Accessed in October 2022 . The OpenCL C Specification . https:\/\/www.khronos.org\/registry\/OpenCL\/specs\/3.0-unified\/html\/OpenCL_C.html Khronos OpneCL Working Group. Accessed in October 2022. The OpenCL C Specification. https:\/\/www.khronos.org\/registry\/OpenCL\/specs\/3.0-unified\/html\/OpenCL_C.html"},{"key":"e_1_2_1_27_1","volume-title":"Proceedings of the 29th ACM on International Conference on Supercomputing","author":"He Wenting","year":"2015","unstructured":"Wenting He , Huimin Cui , Binbin Lu , Jiacheng Zhao , Shengmei Li , G. Ruan , Jingling Xue , Xiaobing Feng , Wensen Yang , and Youliang Yan . 2015 . Hadoop+: Modeling and Evaluating the Heterogeneity for MapReduce Applications in Heterogeneous Clusters . Proceedings of the 29th ACM on International Conference on Supercomputing (2015). Wenting He, Huimin Cui, Binbin Lu, Jiacheng Zhao, Shengmei Li, G. Ruan, Jingling Xue, Xiaobing Feng, Wensen Yang, and Youliang Yan. 2015. Hadoop+: Modeling and Evaluating the Heterogeneity for MapReduce Applications in Heterogeneous Clusters. Proceedings of the 29th ACM on International Conference on Supercomputing (2015)."},{"key":"e_1_2_1_28_1","volume-title":"GPU in-Memory Processing Using Spark for Iterative Computation. 17th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (CCGRID)","author":"Hong Sumin","year":"2017","unstructured":"Sumin Hong , Woohyuk Choi , and Won-Ki Jeong . 2017 . GPU in-Memory Processing Using Spark for Iterative Computation. 17th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (CCGRID) (2017), 31--41. Sumin Hong, Woohyuk Choi, and Won-Ki Jeong. 2017. GPU in-Memory Processing Using Spark for Iterative Computation. 17th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (CCGRID) (2017), 31--41."},{"key":"e_1_2_1_29_1","volume-title":"2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications\/ 12th IEEE International Conference On Big Data Science And Engineering (TrustCom\/BigDataSE)","author":"Hou Junjie","year":"2018","unstructured":"Junjie Hou , Yongxin Zhu , L. Kong , Zhe Wang , Sen Du , Shijin Song , and Tian Huang . 2018 . A Case Study of Accelerating Apache Spark with FPGA . 2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications\/ 12th IEEE International Conference On Big Data Science And Engineering (TrustCom\/BigDataSE) (2018), 855--860. Junjie Hou, Yongxin Zhu, L. Kong, Zhe Wang, Sen Du, Shijin Song, and Tian Huang. 2018. A Case Study of Accelerating Apache Spark with FPGA. 2018 17th IEEE International Conference On Trust, Security And Privacy In Computing And Communications\/ 12th IEEE International Conference On Big Data Science And Engineering (TrustCom\/BigDataSE) (2018), 855--860."},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/2987550.2987569"},{"key":"e_1_2_1_31_1","volume-title":"Accessed","author":"Hutter Marco","year":"2022","unstructured":"Marco Hutter . Accessed in October 2022 . Java bindings for CUBLAS. http:\/\/javagl.de\/jcuda.org\/jcuda\/jcublas\/JCublas.html Marco Hutter. Accessed in October 2022. Java bindings for CUBLAS. http:\/\/javagl.de\/jcuda.org\/jcuda\/jcublas\/JCublas.html"},{"key":"e_1_2_1_32_1","volume-title":"Accessed","author":"Hutter Marco","year":"2022","unstructured":"Marco Hutter . Accessed in October 2022 . Java bindings for CUDA. http:\/\/www.jcuda.org\/ Marco Hutter. Accessed in October 2022. Java bindings for CUDA. http:\/\/www.jcuda.org\/"},{"key":"e_1_2_1_33_1","unstructured":"IBM. Accessed in October 2022. OpenJ9. https:\/\/github.com\/eclipse-openj9\/openj9  IBM. Accessed in October 2022. OpenJ9. https:\/\/github.com\/eclipse-openj9\/openj9"},{"key":"e_1_2_1_34_1","volume-title":"Accessed","author":"INTEL.","year":"2022","unstructured":"INTEL. Accessed in October 2022 . OneAPI. https:\/\/oneapi.io INTEL. Accessed in October 2022. OneAPI. https:\/\/oneapi.io"},{"key":"e_1_2_1_35_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICBNMT.2011.6156017"},{"key":"e_1_2_1_36_1","volume-title":"Accessed","author":"Kloeckner Andreas","year":"2022","unstructured":"Andreas Kloeckner . Accessed in October 2022 . PyCUDA. https:\/\/documen.tician.de\/pycuda Andreas Kloeckner. Accessed in October 2022. PyCUDA. https:\/\/documen.tician.de\/pycuda"},{"key":"e_1_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.5555\/3408352.3408392"},{"key":"e_1_2_1_38_1","volume-title":"IEEE International Conference on Networking, Architecture and Storage (NAS)","author":"Li Peilong","year":"2015","unstructured":"Peilong Li , Yan Luo , Ning Zhang , and Yu Cao . 2015 . HeteroSpark: A heterogeneous CPU\/GPU Spark platform for machine learning algorithms . IEEE International Conference on Networking, Architecture and Storage (NAS) (2015), 347--348. Peilong Li, Yan Luo, Ning Zhang, and Yu Cao. 2015. HeteroSpark: A heterogeneous CPU\/GPU Spark platform for machine learning algorithms. IEEE International Conference on Networking, Architecture and Storage (NAS) (2015), 347--348."},{"key":"e_1_2_1_39_1","volume-title":"Proceedings of the 26th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming","author":"Li Zhifang","year":"2021","unstructured":"Zhifang Li , Mingcong Han , Shangwei Wu , and Chuliang Weng . 2021 . ShadowVM: accelerating data plane for data analytics with bare metal CPUs and GPUs . Proceedings of the 26th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming (2021). Zhifang Li, Mingcong Han, Shangwei Wu, and Chuliang Weng. 2021. ShadowVM: accelerating data plane for data analytics with bare metal CPUs and GPUs. Proceedings of the 26th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming (2021)."},{"key":"e_1_2_1_40_1","unstructured":"Yu Lin Semih Okur and Cosmin Radoi. 2012. Hadoop+Aparapi: Making heterogenous MapReduce programming easier.  Yu Lin Semih Okur and Cosmin Radoi. 2012. Hadoop+Aparapi: Making heterogenous MapReduce programming easier."},{"key":"e_1_2_1_41_1","volume-title":"International Conference on Field-Programmable Technology (FPT)","author":"Lin Zhongduo","year":"2013","unstructured":"Zhongduo Lin and Paul Chow . 2013 . ZCluster: A Zynq-based Hadoop cluster . International Conference on Field-Programmable Technology (FPT) (2013), 450--453. Zhongduo Lin and Paul Chow. 2013. ZCluster: A Zynq-based Hadoop cluster. International Conference on Field-Programmable Technology (FPT) (2013), 450--453."},{"key":"e_1_2_1_42_1","volume-title":"Exploring GPU Acceleration of Apache Spark. 2016 IEEE International Conference on Cloud Engineering (IC2E)","author":"Manzi D.","year":"2016","unstructured":"D. Manzi and David Tompkins . 2016 . Exploring GPU Acceleration of Apache Spark. 2016 IEEE International Conference on Cloud Engineering (IC2E) (2016), 222--223. D. Manzi and David Tompkins. 2016. Exploring GPU Acceleration of Apache Spark. 2016 IEEE International Conference on Cloud Engineering (IC2E) (2016), 222--223."},{"key":"e_1_2_1_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/2854038.2854040"},{"key":"e_1_2_1_44_1","volume-title":"Accelerating Big Data Analytics Using FPGAs. 2015 IEEE 23rd Annual International Symposium on Field-Programmable Custom Computing Machines","author":"Neshatpour Katayoun","year":"2015","unstructured":"Katayoun Neshatpour , Maria Malik , Mohammad Ali Ghodrat , and Houman Homayoun . 2015 . Accelerating Big Data Analytics Using FPGAs. 2015 IEEE 23rd Annual International Symposium on Field-Programmable Custom Computing Machines (2015), 164--164. Katayoun Neshatpour, Maria Malik, Mohammad Ali Ghodrat, and Houman Homayoun. 2015. Accelerating Big Data Analytics Using FPGAs. 2015 IEEE 23rd Annual International Symposium on Field-Programmable Custom Computing Machines (2015), 164--164."},{"key":"e_1_2_1_45_1","doi-asserted-by":"publisher","DOI":"10.1109\/BigData.2015.7363748"},{"key":"e_1_2_1_46_1","volume-title":"Accelerating Machine Learning Kernel in Hadoop Using FPGAs. 2015 15th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing","author":"Neshatpour Katayoun","year":"2015","unstructured":"Katayoun Neshatpour , Maria Malik , and Houman Homayoun . 2015 . Accelerating Machine Learning Kernel in Hadoop Using FPGAs. 2015 15th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (2015), 1151--1154. Katayoun Neshatpour, Maria Malik, and Houman Homayoun. 2015. Accelerating Machine Learning Kernel in Hadoop Using FPGAs. 2015 15th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (2015), 1151--1154."},{"key":"e_1_2_1_47_1","volume-title":"2016 International Conference on Hardware\/Software Codesign and System Synthesis (CODES+ISSS)","author":"Neshatpour Katayoun","year":"2016","unstructured":"Katayoun Neshatpour , Avesta Sasan , and Houman Homayoun . 2016 . Big data analytics on heterogeneous accelerator architectures . 2016 International Conference on Hardware\/Software Codesign and System Synthesis (CODES+ISSS) (2016), 1--3. Katayoun Neshatpour, Avesta Sasan, and Houman Homayoun. 2016. Big data analytics on heterogeneous accelerator architectures. 2016 International Conference on Hardware\/Software Codesign and System Synthesis (CODES+ISSS) (2016), 1--3."},{"key":"e_1_2_1_48_1","volume-title":"Accessed","author":"NVIDIA.","year":"2022","unstructured":"NVIDIA. Accessed in October 2022 . CUDA. https:\/\/docs.nvidia.com\/cuda\/parallel-thread-execution\/index.html NVIDIA. Accessed in October 2022. CUDA. https:\/\/docs.nvidia.com\/cuda\/parallel-thread-execution\/index.html"},{"key":"e_1_2_1_49_1","volume-title":"Accessed","year":"2022","unstructured":"ObjectWeb. Accessed in October 2022 . ASM. https:\/\/asm.ow2.io ObjectWeb. Accessed in October 2022. ASM. https:\/\/asm.ow2.io"},{"key":"e_1_2_1_50_1","volume-title":"Accelerating Spark RDD Operations with Local and Remote GPU Devices. 2016 IEEE 22nd International Conference on Parallel and Distributed Systems (ICPADS)","author":"Ohno Yasuhiro","year":"2016","unstructured":"Yasuhiro Ohno , Shin Morishima , and Hiroki Matsutani . 2016 . Accelerating Spark RDD Operations with Local and Remote GPU Devices. 2016 IEEE 22nd International Conference on Parallel and Distributed Systems (ICPADS) (2016), 791--799. Yasuhiro Ohno, Shin Morishima, and Hiroki Matsutani. 2016. Accelerating Spark RDD Operations with Local and Remote GPU Devices. 2016 IEEE 22nd International Conference on Parallel and Distributed Systems (ICPADS) (2016), 791--799."},{"key":"e_1_2_1_51_1","volume-title":"Accessed","year":"2022","unstructured":"Oracle. Accessed in October 2022 . The Java Tutorials . https:\/\/docs.oracle.com\/javase\/tutorial\/jndi\/objects\/serial.html Oracle. Accessed in October 2022. The Java Tutorials. https:\/\/docs.oracle.com\/javase\/tutorial\/jndi\/objects\/serial.html"},{"key":"e_1_2_1_52_1","volume-title":"Accessed","year":"2022","unstructured":"Oracle. Accessed in October 2022 . The Java Virtual Machine Specification . https:\/\/docs.oracle.com\/javase\/specs\/jvms\/se7\/html\/ Oracle. Accessed in October 2022. The Java Virtual Machine Specification. https:\/\/docs.oracle.com\/javase\/specs\/jvms\/se7\/html\/"},{"key":"e_1_2_1_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/JPROC.2008.917757"},{"key":"e_1_2_1_54_1","doi-asserted-by":"publisher","DOI":"10.1145\/3453933.3454014"},{"key":"e_1_2_1_55_1","doi-asserted-by":"publisher","DOI":"10.22152\/programming-journal.org\/2021\/5\/8"},{"key":"e_1_2_1_56_1","doi-asserted-by":"crossref","first-page":"8","DOI":"10.22152\/programming-journal.org\/2021\/5\/8","article-title":"Transparent Compiler and Runtime Specializations for Accelerating Managed Languages on FPGAs. Art Sci","volume":"5","author":"Papadimitriou Michail","year":"2021","unstructured":"Michail Papadimitriou , Juan Jos\u00e9 Fumero , Athanasios Stratikopoulos , Foivos S. Zakkak , and Christos Kotselidis . 2021 . Transparent Compiler and Runtime Specializations for Accelerating Managed Languages on FPGAs. Art Sci . Eng. Program. 5 (2021), 8 . Michail Papadimitriou, Juan Jos\u00e9 Fumero, Athanasios Stratikopoulos, Foivos S. Zakkak, and Christos Kotselidis. 2021. Transparent Compiler and Runtime Specializations for Accelerating Managed Languages on FPGAs. Art Sci. Eng. Program. 5 (2021), 8.","journal-title":"Eng. Program."},{"key":"e_1_2_1_57_1","doi-asserted-by":"publisher","DOI":"10.1145\/3453933.3454019"},{"key":"e_1_2_1_58_1","doi-asserted-by":"crossref","first-page":"630","DOI":"10.1007\/s10766-017-0513-2","article-title":"Real-Time Big Data Stream Processing Using GPU with Spark Over Hadoop Ecosystem","volume":"46","author":"Rathore M. Mazhar","year":"2017","unstructured":"M. Mazhar Rathore , Hojae Son , Awais Ahmad , Anand Paul , and Gwanggil Jeon . 2017 . Real-Time Big Data Stream Processing Using GPU with Spark Over Hadoop Ecosystem . International Journal of Parallel Programming 46 (2017), 630 -- 646 . M. Mazhar Rathore, Hojae Son, Awais Ahmad, Anand Paul, and Gwanggil Jeon. 2017. Real-Time Big Data Stream Processing Using GPU with Spark Over Hadoop Ecosystem. International Journal of Parallel Programming 46 (2017), 630--646.","journal-title":"International Journal of Parallel Programming"},{"key":"e_1_2_1_60_1","volume-title":"Deep Dive into GPU Support in Apache Spark 3.x. https:\/\/databricks.com\/session_na20\/deep-dive-into-gpu-support-in-apache-spark-3-x SPARK+AI SUMMIT","author":"Robert Evans Jason Lowe","year":"2020","unstructured":"Jason Lowe Robert Evans . 2020. Deep Dive into GPU Support in Apache Spark 3.x. https:\/\/databricks.com\/session_na20\/deep-dive-into-gpu-support-in-apache-spark-3-x SPARK+AI SUMMIT 2020 . Jason Lowe Robert Evans. 2020. Deep Dive into GPU Support in Apache Spark 3.x. https:\/\/databricks.com\/session_na20\/deep-dive-into-gpu-support-in-apache-spark-3-x SPARK+AI SUMMIT 2020."},{"key":"e_1_2_1_61_1","doi-asserted-by":"publisher","DOI":"10.1145\/2749246.2749261"},{"key":"e_1_2_1_62_1","volume-title":"SparkCL: A Unified Programming Framework for Accelerators on Heterogeneous Clusters. ArXiv abs\/1505.01120","author":"Segal Oren","year":"2015","unstructured":"Oren Segal , Philip Colangelo , Nasibeh Nasiri , Zhuo Qian , and Martin Margala . 2015. SparkCL: A Unified Programming Framework for Accelerators on Heterogeneous Clusters. ArXiv abs\/1505.01120 ( 2015 ). Oren Segal, Philip Colangelo, Nasibeh Nasiri, Zhuo Qian, and Martin Margala. 2015. SparkCL: A Unified Programming Framework for Accelerators on Heterogeneous Clusters. ArXiv abs\/1505.01120 (2015)."},{"key":"e_1_2_1_63_1","doi-asserted-by":"crossref","first-page":"3165","DOI":"10.1007\/s11227-020-03390-z","article-title":"Ignite-GPU: A GPU-enabled in-memory computing architecture on clusters","volume":"77","author":"Sojoodi Amir Hossein","year":"2020","unstructured":"Amir Hossein Sojoodi , Majid Salimi Beni , and Farshad Khunjush . 2020 . Ignite-GPU: A GPU-enabled in-memory computing architecture on clusters . The Journal of Supercomputing 77 (2020), 3165 -- 3192 . Amir Hossein Sojoodi, Majid Salimi Beni, and Farshad Khunjush. 2020. Ignite-GPU: A GPU-enabled in-memory computing architecture on clusters. The Journal of Supercomputing 77 (2020), 3165--3192.","journal-title":"The Journal of Supercomputing"},{"key":"e_1_2_1_64_1","volume-title":"Accessed","author":"Spark Apache","year":"2022","unstructured":"Apache Spark . Accessed in October 2022 . Spoark Checkpoints . https:\/\/spark.apache.org\/docs\/latest\/streaming-programming-guide.html Apache Spark. Accessed in October 2022. Spoark Checkpoints. https:\/\/spark.apache.org\/docs\/latest\/streaming-programming-guide.html"},{"key":"e_1_2_1_65_1","volume-title":"2018 International Conference on High Performance Computing and Simulation (HPCS)","author":"Stamelos Ioannis","year":"2018","unstructured":"Ioannis Stamelos , Elias Koromilas , Christoforos Kachris , and Dimitrios Soudris . 2018 . A Novel Framework for the Seamless Integration of FPGA Accelerators with Big Data Analytics Frameworks in Heterogeneous Data Centers . 2018 International Conference on High Performance Computing and Simulation (HPCS) (2018), 539--545. Ioannis Stamelos, Elias Koromilas, Christoforos Kachris, and Dimitrios Soudris. 2018. A Novel Framework for the Seamless Integration of FPGA Accelerators with Big Data Analytics Frameworks in Heterogeneous Data Centers. 2018 International Conference on High Performance Computing and Simulation (HPCS) (2018), 539--545."},{"key":"e_1_2_1_66_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCSE.2010.69"},{"key":"e_1_2_1_67_1","doi-asserted-by":"publisher","DOI":"10.1145\/3426182.3426188"},{"key":"e_1_2_1_68_1","volume-title":"2012 12th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (ccgrid 2012)","author":"Tan Yu Shyang","year":"2012","unstructured":"Yu Shyang Tan , Bu-Sung Lee , Bingsheng He , and Roy H. Campbell . 2012. A Map-Reduce Based Framework for Heterogeneous Processing Element Cluster Environments . 2012 12th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (ccgrid 2012) ( 2012 ), 57--64. Yu Shyang Tan, Bu-Sung Lee, Bingsheng He, and Roy H. Campbell. 2012. A Map-Reduce Based Framework for Heterogeneous Processing Element Cluster Environments. 2012 12th IEEE\/ACM International Symposium on Cluster, Cloud and Grid Computing (ccgrid 2012) (2012), 57--64."},{"key":"e_1_2_1_69_1","volume-title":"SparkJNI: A Toolchain for Hardware Accelerated Big Data Apache Spark. 2019 IEEE 4th International Conference on Big Data Analytics (ICBDA)","author":"Voicu Tudor Alexandru","year":"2019","unstructured":"Tudor Alexandru Voicu and Zaid Al-Ars . 2019 . SparkJNI: A Toolchain for Hardware Accelerated Big Data Apache Spark. 2019 IEEE 4th International Conference on Big Data Analytics (ICBDA) (2019), 152--157. Tudor Alexandru Voicu and Zaid Al-Ars. 2019. SparkJNI: A Toolchain for Hardware Accelerated Big Data Apache Spark. 2019 IEEE 4th International Conference on Big Data Analytics (ICBDA) (2019), 152--157."},{"key":"e_1_2_1_70_1","volume-title":"When FPGA-Accelerator Meets Stream Data Processing in the Edge. IEEE 39th International Conference on Distributed Computing Systems (ICDCS) (2019)","author":"Wu Song","year":"2019","unstructured":"Song Wu , Die Hu , Shadi Ibrahim , Hai Jin , Jiang Xiao , Fei Chen , and Haikun Liu . 2019 . When FPGA-Accelerator Meets Stream Data Processing in the Edge. IEEE 39th International Conference on Distributed Computing Systems (ICDCS) (2019) , 1818--1829. Song Wu, Die Hu, Shadi Ibrahim, Hai Jin, Jiang Xiao, Fei Chen, and Haikun Liu. 2019. When FPGA-Accelerator Meets Stream Data Processing in the Edge. IEEE 39th International Conference on Distributed Computing Systems (ICDCS) (2019), 1818--1829."},{"key":"e_1_2_1_71_1","volume-title":"2016 IEEE International Conference on Big Data (Big Data)","author":"Yuan Yuan","year":"2016","unstructured":"Yuan Yuan , Meisam Fathi Salmi , Yin Huai , Kaibo Wang , Rubao Lee , and Xiaodong Zhang . 2016 . Spark-GPU: An accelerated in-memory data processing engine on clusters . 2016 IEEE International Conference on Big Data (Big Data) (2016), 273--283. Yuan Yuan, Meisam Fathi Salmi, Yin Huai, Kaibo Wang, Rubao Lee, and Xiaodong Zhang. 2016. Spark-GPU: An accelerated in-memory data processing engine on clusters. 2016 IEEE International Conference on Big Data (Big Data) (2016), 273--283."},{"key":"e_1_2_1_72_1","volume-title":"IEEE\/ACIS 13th International Conference on Computer and Information Science (ICIS)","author":"Zhu Jie","year":"2014","unstructured":"Jie Zhu , Juanjuan Li , Erikson Hardesty , Hai Jiang , and Kuan-Ching Li . 2014 . GPU-in-Hadoop: Enabling MapReduce across distributed heterogeneous platforms . IEEE\/ACIS 13th International Conference on Computer and Information Science (ICIS) (2014), 321--326. Jie Zhu, Juanjuan Li, Erikson Hardesty, Hai Jiang, and Kuan-Ching Li. 2014. GPU-in-Hadoop: Enabling MapReduce across distributed heterogeneous platforms. IEEE\/ACIS 13th International Conference on Computer and Information Science (ICIS) (2014), 321--326."}],"container-title":["Proceedings of the VLDB Endowment"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.14778\/3565838.3565842","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2023,1,20]],"date-time":"2023-01-20T23:18:04Z","timestamp":1674256684000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.14778\/3565838.3565842"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2022,9]]},"references-count":71,"journal-issue":{"issue":"13","published-print":{"date-parts":[[2022,9]]}},"alternative-id":["10.14778\/3565838.3565842"],"URL":"https:\/\/doi.org\/10.14778\/3565838.3565842","relation":{},"ISSN":["2150-8097"],"issn-type":[{"value":"2150-8097","type":"print"}],"subject":[],"published":{"date-parts":[[2022,9]]},"assertion":[{"value":"2023-01-20","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}