{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,14]],"date-time":"2026-07-14T10:13:35Z","timestamp":1784024015225,"version":"3.55.0"},"reference-count":134,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2013,7,11]],"date-time":"2013-07-11T00:00:00Z","timestamp":1373500800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"DOI":"10.13039\/100015539","name":"Australian Government","doi-asserted-by":"crossref","id":[{"id":"10.13039\/100015539","id-type":"DOI","asserted-by":"crossref"}]},{"name":"ICT Centre of Excellence program"},{"DOI":"10.13039\/501100000923","name":"Australian Research Council","doi-asserted-by":"publisher","id":[{"id":"10.13039\/501100000923","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Comput. Surv."],"published-print":{"date-parts":[[2013,10]]},"abstract":"<jats:p>\n            In the last two decades, the continuous increase of computational power has produced an overwhelming flow of data which has called for a paradigm shift in the computing architecture and large-scale data processing mechanisms. MapReduce is a simple and powerful programming model that enables easy development of scalable parallel applications to process vast amounts of data on large clusters of commodity machines. It isolates the application from the details of running a distributed program such as issues on data distribution, scheduling, and fault tolerance. However, the original implementation of the MapReduce framework had some limitations that have been tackled by many research efforts in several followup works after its introduction. This article provides a comprehensive survey for a\n            <jats:italic>family<\/jats:italic>\n            of approaches and mechanisms of large-scale data processing mechanisms that have been implemented based on the original idea of the MapReduce framework and are currently gaining a lot of momentum in both research and industrial communities. We also cover a set of introduced systems that have been implemented to provide declarative programming interfaces on top of the MapReduce framework. In addition, we review several large-scale data processing systems that resemble some of the ideas of the MapReduce framework for different purposes and application scenarios. Finally, we discuss some of the future research directions for implementing the next generation of MapReduce-like solutions.\n          <\/jats:p>","DOI":"10.1145\/2522968.2522979","type":"journal-article","created":{"date-parts":[[2013,11,6]],"date-time":"2013-11-06T14:09:19Z","timestamp":1383746959000},"page":"1-44","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":120,"title":["The family of mapreduce and large-scale data processing systems"],"prefix":"10.1145","volume":"46","author":[{"given":"Sherif","family":"Sakr","sequence":"first","affiliation":[{"name":"NICTA and University of New South Wales, Sydney, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Anna","family":"Liu","sequence":"additional","affiliation":[{"name":"NICTA and University of New South Wales, Sydney, Australia"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ayman G.","family":"Fayoumi","sequence":"additional","affiliation":[{"name":"King Abdulaziz University, Saudia Arabia"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2013,7,11]]},"reference":[{"key":"e_1_2_2_1_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00778-008-0125-y"},{"key":"e_1_2_2_2_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687627.1687731"},{"key":"e_1_2_2_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807294"},{"key":"e_1_2_2_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/1739041.1739056"},{"key":"e_1_2_2_5_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2012.66"},{"key":"e_1_2_2_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2011.47"},{"key":"e_1_2_2_7_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1921056"},{"key":"e_1_2_2_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/1755913.1755937"},{"key":"e_1_2_2_9_1","unstructured":"Armbrust M. Fox A. Rean G. Joseph A. Katz R. Konwinski A. Gunho L. David P. Rabkin A. Stoica I. and Zaharia M. 2009. Above the clouds: A berkeley view of cloud computing. http:\/\/www.cs.columbia.edu\/&sim;roxana\/teaching\/COMS-E6998-7-Fall-2011\/papers\/armbrust-tr09.pdf.  Armbrust M. Fox A. Rean G. Joseph A. Katz R. Konwinski A. Gunho L. David P. Rabkin A. Stoica I. and Zaharia M. 2009. Above the clouds: A berkeley view of cloud computing. http:\/\/www.cs.columbia.edu\/&sim;roxana\/teaching\/COMS-E6998-7-Fall-2011\/papers\/armbrust-tr09.pdf."},{"key":"e_1_2_2_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807128.1807150"},{"key":"e_1_2_2_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/2213836.2213938"},{"key":"e_1_2_2_12_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807128.1807148"},{"key":"e_1_2_2_13_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10619-011-7082-y"},{"key":"e_1_2_2_14_1","doi-asserted-by":"publisher","DOI":"10.1109\/MC.2006.29"},{"key":"e_1_2_2_15_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402755.3402761"},{"key":"e_1_2_2_16_1","doi-asserted-by":"publisher","DOI":"10.1145\/2038916.2038923"},{"key":"e_1_2_2_17_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807273"},{"key":"e_1_2_2_18_1","unstructured":"Boag S. Chamberlin D. Fernandez M. F. Florescu D. Robie J. and Simeon J. 2010. XQuery 1.0: An xml query language. http:\/\/www.w3.org\/TR\/xquery.  Boag S. Chamberlin D. Fernandez M. F. Florescu D. Robie J. and Simeon J. 2010. XQuery 1.0: An xml query language. http:\/\/www.w3.org\/TR\/xquery."},{"key":"e_1_2_2_19_1","first-page":"2","article-title":"ASTERIX: An open source system for big data management and analysis","volume":"5","author":"Borkar V.","year":"2012","unstructured":"Borkar , V. , Alsubaiee , S. , Altowim , Y. , Altwaijry , H. , Behm , A. , Bu , Y. , Carey , M. , Grover , R. , Heilbron , Z. , Kim , Y.-S. , Li , C. , Pirzadeh , P. , Onose , N. , Vernica , R. , and Wen , J. 2012 a. ASTERIX: An open source system for big data management and analysis . Proc. VLDB Endow. 5 , 2 . Borkar, V., Alsubaiee, S., Altowim, Y., Altwaijry, H., Behm, A., Bu, Y., Carey, M., Grover, R., Heilbron, Z., Kim, Y.-S., Li, C., Pirzadeh, P., Onose, N., Vernica, R., and Wen, J. 2012a. ASTERIX: An open source system for big data management and analysis. Proc. VLDB Endow. 5, 2.","journal-title":"Proc. VLDB Endow."},{"key":"e_1_2_2_20_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2011.5767921"},{"key":"e_1_2_2_21_1","doi-asserted-by":"publisher","DOI":"10.1145\/2247596.2247598"},{"key":"e_1_2_2_22_1","unstructured":"Bray T. Paoli J. Sperberg-Mcqueen C. M. Maler E. and Yergeau F. 2008. Extensible markup language (xml) 1.0 5th ed. http:\/\/www.w3.org\/TR\/REC-xml\/.  Bray T. Paoli J. Sperberg-Mcqueen C. M. Maler E. and Yergeau F. 2008. Extensible markup language (xml) 1.0 5 th ed. http:\/\/www.w3.org\/TR\/REC-xml\/."},{"key":"e_1_2_2_23_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920881"},{"key":"e_1_2_2_24_1","doi-asserted-by":"publisher","DOI":"10.1145\/1859127.1859141"},{"key":"e_1_2_2_25_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-02279-1_24"},{"key":"e_1_2_2_26_1","doi-asserted-by":"publisher","DOI":"10.14778\/1454159.1454166"},{"key":"e_1_2_2_27_1","doi-asserted-by":"publisher","DOI":"10.1145\/1806596.1806638"},{"key":"e_1_2_2_28_1","doi-asserted-by":"publisher","DOI":"10.1145\/1365815.1365816"},{"key":"e_1_2_2_29_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402755.3402765"},{"key":"e_1_2_2_30_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807297"},{"key":"e_1_2_2_31_1","doi-asserted-by":"publisher","DOI":"10.14778\/1453856.1453978"},{"key":"e_1_2_2_32_1","volume-title":"Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation (NSDI'10)","author":"Condie T.","unstructured":"Condie , T. , Conway , N. , Alvaro , P. , Hellerstein , J. M. , Elmeleegy , K. , and Sears , R . 2010a. MapReduce online . In Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation (NSDI'10) . 313--328. Condie, T., Conway, N., Alvaro, P., Hellerstein, J. M., Elmeleegy, K., and Sears, R. 2010a. MapReduce online. In Proceedings of the 7th USENIX Conference on Networked Systems Design and Implementation (NSDI'10). 313--328."},{"key":"e_1_2_2_33_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807295"},{"key":"e_1_2_2_34_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020516"},{"key":"e_1_2_2_35_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807275"},{"key":"e_1_2_2_36_1","volume-title":"Proceedings of the 6th Symposium on Operating System Design and Implementation (OSDI'04)","author":"Dean J.","unstructured":"Dean , J. and Ghemawat , S . 2004. MapReduce: Simplified data processing on large clusters . In Proceedings of the 6th Symposium on Operating System Design and Implementation (OSDI'04) . 137--150. Dean, J. and Ghemawat, S. 2004. MapReduce: Simplified data processing on large clusters. In Proceedings of the 6th Symposium on Operating System Design and Implementation (OSDI'04). 137--150."},{"key":"e_1_2_2_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/1327452.1327492"},{"key":"e_1_2_2_38_1","doi-asserted-by":"publisher","DOI":"10.1145\/1629175.1629198"},{"key":"e_1_2_2_39_1","doi-asserted-by":"publisher","DOI":"10.1145\/129888.129894"},{"key":"e_1_2_2_40_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920908"},{"key":"e_1_2_2_41_1","doi-asserted-by":"publisher","DOI":"10.1145\/1851476.1851593"},{"key":"e_1_2_2_42_1","doi-asserted-by":"publisher","DOI":"10.14778\/2168651.2168659"},{"key":"e_1_2_2_43_1","doi-asserted-by":"publisher","DOI":"10.1145\/2213836.2213937"},{"key":"e_1_2_2_44_1","doi-asserted-by":"publisher","DOI":"10.14778\/2002938.2002943"},{"key":"e_1_2_2_45_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020515"},{"key":"e_1_2_2_46_1","volume-title":"Proceedings of the International Workshop on the Web and Databases (WebDB).","author":"Fegaras L.","unstructured":"Fegaras , L. , Li , C. , Gupta , U. , and Philip , J . 2011. XML query optimization in map-reduce . In Proceedings of the International Workshop on the Web and Databases (WebDB). Fegaras, L., Li, C., Gupta, U., and Philip, J. 2011. XML query optimization in map-reduce. In Proceedings of the International Workshop on the Web and Databases (WebDB)."},{"key":"e_1_2_2_47_1","doi-asserted-by":"publisher","DOI":"10.14778\/1988776.1988778"},{"key":"e_1_2_2_48_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687553.1687567"},{"key":"e_1_2_2_49_1","volume-title":"Programming Pig","author":"Gates A.","unstructured":"Gates , A. 2011. Programming Pig . O'Reilly Media . Gates, A. 2011. Programming Pig. O'Reilly Media."},{"key":"e_1_2_2_50_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687553.1687568"},{"key":"e_1_2_2_51_1","doi-asserted-by":"publisher","DOI":"10.1145\/945445.945450"},{"key":"e_1_2_2_52_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020464"},{"key":"e_1_2_2_53_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2011.5767930"},{"key":"e_1_2_2_54_1","doi-asserted-by":"publisher","DOI":"10.1007\/s007780100054"},{"key":"e_1_2_2_55_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2011.5767933"},{"key":"e_1_2_2_56_1","unstructured":"Herodotou H. 2011. Hadoop performance models. CoRR abs\/1106.0940. http:\/\/arxiv.org\/abs\/1106.0940.  Herodotou H. 2011. Hadoop performance models. CoRR abs\/1106.0940. http:\/\/arxiv.org\/abs\/1106.0940."},{"key":"e_1_2_2_57_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402707.3402746"},{"key":"e_1_2_2_58_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402755.3402792"},{"key":"e_1_2_2_59_1","volume-title":"Proceedings of the 5th Conference on Innovative Data Systems Research (CIDR'11)","author":"Herodotou H.","unstructured":"Herodotou , H. , Lim , H. , Luo , G. , Borisov , N. , Dong , L. , Cetin , F. B. , and Babu , S . 2011b. Starfish: A self-tuning system for big data analytics . In Proceedings of the 5th Conference on Innovative Data Systems Research (CIDR'11) . 261--272. Herodotou, H., Lim, H., Luo, G., Borisov, N., Dong, L., Cetin, F. B., and Babu, S. 2011b. Starfish: A self-tuning system for big data analytics. In Proceedings of the 5th Conference on Innovative Data Systems Research (CIDR'11). 261--272."},{"key":"e_1_2_2_60_1","volume-title":"eds","author":"Hey T.","year":"2009","unstructured":"Hey , T. , Tansley , S. , and Tolle , K. , eds . 2009 . The fourth paradigm: Data-intensive scientific discovery. Microsoft Research . http:\/\/research.microsoft.com\/en-us\/collaboration\/fourthparadigm\/4th_paradigm_book_complete_lr.pdf. Hey, T., Tansley, S., and Tolle, K., eds. 2009. The fourth paradigm: Data-intensive scientific discovery. Microsoft Research. http:\/\/research.microsoft.com\/en-us\/collaboration\/fourthparadigm\/4th_paradigm_book_complete_lr.pdf."},{"key":"e_1_2_2_61_1","volume-title":"HotCloud Workshop held in conjunction with the USENIX Annual Technical Conference. https:\/\/www.usenix.org\/legacy\/event\/hotcloud09\/tech\/full_papers\/hindman.pdf.","author":"Hindman B.","unstructured":"Hindman , B. , Konwinski , A. , Zaharia , M. , and Stoica , I . 2009. A common substrate for cluster computing . In HotCloud Workshop held in conjunction with the USENIX Annual Technical Conference. https:\/\/www.usenix.org\/legacy\/event\/hotcloud09\/tech\/full_papers\/hindman.pdf. Hindman, B., Konwinski, A., Zaharia, M., and Stoica, I. 2009. A common substrate for cluster computing. In HotCloud Workshop held in conjunction with the USENIX Annual Technical Conference. https:\/\/www.usenix.org\/legacy\/event\/hotcloud09\/tech\/full_papers\/hindman.pdf."},{"key":"e_1_2_2_62_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402707.3402747"},{"key":"e_1_2_2_63_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2011.103"},{"key":"e_1_2_2_64_1","doi-asserted-by":"publisher","DOI":"10.1145\/1272996.1273005"},{"key":"e_1_2_2_65_1","doi-asserted-by":"publisher","DOI":"10.14778\/1978665.1978670"},{"key":"e_1_2_2_66_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920903"},{"key":"e_1_2_2_67_1","doi-asserted-by":"publisher","DOI":"10.1109\/TKDE.2010.248"},{"key":"e_1_2_2_68_1","doi-asserted-by":"publisher","DOI":"10.1145\/2038916.2038937"},{"key":"e_1_2_2_69_1","doi-asserted-by":"publisher","DOI":"10.1145\/2247596.2247600"},{"key":"e_1_2_2_70_1","volume-title":"Proceedings of the 15th Pacific-Asia conference on Advances in Knowledge Discovery and Data Mining (PAKDD'11)","author":"Kang U.","unstructured":"Kang , U. , Meeder , B. , and Faloutsos , C . 2011a. Spectral analysis for billion-scale graphs: Discoveries and implementation . In Proceedings of the 15th Pacific-Asia conference on Advances in Knowledge Discovery and Data Mining (PAKDD'11) . 13--25. Kang, U., Meeder, B., and Faloutsos, C. 2011a. Spectral analysis for billion-scale graphs: Discoveries and implementation. In Proceedings of the 15th Pacific-Asia conference on Advances in Knowledge Discovery and Data Mining (PAKDD'11). 13--25."},{"key":"e_1_2_2_71_1","doi-asserted-by":"publisher","DOI":"10.1145\/2020408.2020580"},{"key":"e_1_2_2_72_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDM.2009.14"},{"key":"e_1_2_2_73_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10115-010-0305-0"},{"key":"e_1_2_2_74_1","volume-title":"Proceedings of the 5th Alberto Mendelzon International Workshop on Foundations of Data Management (AMW'11)","author":"Khatchadourian S.","unstructured":"Khatchadourian , S. , Consens , M. P. , and Simeon , J . 2011. Having a chuql at xml on the cloud . In Proceedings of the 5th Alberto Mendelzon International Workshop on Foundations of Data Management (AMW'11) . Khatchadourian, S., Consens, M. P., and Simeon, J. 2011. Having a chuql at xml on the cloud. In Proceedings of the 5th Alberto Mendelzon International Workshop on Foundations of Data Management (AMW'11)."},{"key":"e_1_2_2_75_1","doi-asserted-by":"publisher","DOI":"10.14778\/3402755.3402787"},{"key":"e_1_2_2_76_1","doi-asserted-by":"publisher","DOI":"10.14778\/2367502.2367527"},{"key":"e_1_2_2_77_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2012.22"},{"key":"e_1_2_2_78_1","doi-asserted-by":"publisher","DOI":"10.1145\/1739041.1739120"},{"key":"e_1_2_2_79_1","doi-asserted-by":"publisher","DOI":"10.1145\/279227.279229"},{"key":"e_1_2_2_80_1","unstructured":"Large Synoptic Survey. 2013. http:\/\/www.lsst.org\/.  Large Synoptic Survey. 2013. http:\/\/www.lsst.org\/."},{"key":"e_1_2_2_81_1","doi-asserted-by":"publisher","DOI":"10.1145\/1989493.1989505"},{"key":"e_1_2_2_82_1","doi-asserted-by":"publisher","DOI":"10.14778\/2350229.2350239"},{"key":"e_1_2_2_83_1","doi-asserted-by":"publisher","DOI":"10.1145\/1571941.1571970"},{"key":"e_1_2_2_84_1","doi-asserted-by":"publisher","DOI":"10.1145\/1989323.1989424"},{"key":"e_1_2_2_85_1","doi-asserted-by":"publisher","DOI":"10.14778\/1454159.1454204"},{"key":"e_1_2_2_86_1","doi-asserted-by":"publisher","DOI":"10.1145\/1095810.1095818"},{"key":"e_1_2_2_87_1","volume-title":"Proceedings of the 26th Conference on Uncertainty in Artificial Intelligence (UAI'10)","author":"Low Y.","unstructured":"Low , Y. , Gonzalez , J. , Kyrola , A. , Bickson , D. , Guestrin , C. , and Hellerstein , J. M . 2010. GraphLab: A new framework for parallel machine learning . In Proceedings of the 26th Conference on Uncertainty in Artificial Intelligence (UAI'10) . 340--349. Low, Y., Gonzalez, J., Kyrola, A., Bickson, D., Guestrin, C., and Hellerstein, J. M. 2010. GraphLab: A new framework for parallel machine learning. In Proceedings of the 26th Conference on Uncertainty in Artificial Intelligence (UAI'10). 340--349."},{"key":"e_1_2_2_88_1","doi-asserted-by":"publisher","DOI":"10.14778\/2212351.2212354"},{"key":"e_1_2_2_89_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807184"},{"key":"e_1_2_2_90_1","unstructured":"Manola F. and Miller E. 2004. RDF Primer W3C Recommendation. http:\/\/www.w3.org\/TR\/REC-rdf-syntax\/.  Manola F. and Miller E. 2004. RDF Primer W3C Recommendation. http:\/\/www.w3.org\/TR\/REC-rdf-syntax\/."},{"key":"e_1_2_2_91_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920886"},{"key":"e_1_2_2_92_1","doi-asserted-by":"publisher","DOI":"10.14778\/2212351.2212353"},{"key":"e_1_2_2_93_1","doi-asserted-by":"publisher","DOI":"10.14778\/1988776.1988782"},{"key":"e_1_2_2_94_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807223"},{"key":"e_1_2_2_95_1","volume-title":"Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10)","author":"Morton K.","unstructured":"Morton , K. , Friesen , A. , Balazinska , M. , and Grossman , D . 2010b. Estimating the progress of mapreduce pipelines . In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10) . 681--684. Morton, K., Friesen, A., Balazinska, M., and Grossman, D. 2010b. Estimating the progress of mapreduce pipelines. In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10). 681--684."},{"key":"e_1_2_2_96_1","doi-asserted-by":"publisher","DOI":"10.1145\/1779599.1779605"},{"key":"e_1_2_2_97_1","doi-asserted-by":"publisher","DOI":"10.14778\/1453856.1453927"},{"key":"e_1_2_2_98_1","doi-asserted-by":"publisher","DOI":"10.14778\/1920841.1920906"},{"key":"e_1_2_2_99_1","volume-title":"Scala: A Comprehensive Step-by-Step Guide","author":"Odersky M.","year":"2011","unstructured":"Odersky , M. , Spoon , L. , and Venners , B . 2011 . Programming in Scala: A Comprehensive Step-by-Step Guide . Artima Inc . Odersky, M., Spoon, L., and Venners, B. 2011. Programming in Scala: A Comprehensive Step-by-Step Guide. Artima Inc."},{"key":"e_1_2_2_100_1","doi-asserted-by":"publisher","DOI":"10.1145\/1376616.1376726"},{"key":"e_1_2_2_101_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687553.1687569"},{"key":"e_1_2_2_102_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDM.2008.142"},{"key":"e_1_2_2_103_1","doi-asserted-by":"publisher","DOI":"10.1145\/1327452.1327491"},{"key":"e_1_2_2_104_1","doi-asserted-by":"publisher","DOI":"10.1145\/1559845.1559865"},{"key":"e_1_2_2_105_1","first-page":"277","article-title":"Interpreting the data: Parallel analysis with sawzall. Sci","volume":"13","author":"Pike R.","year":"2005","unstructured":"Pike , R. , Dorward , S. , Griesemer , R. , and Quinlan , S. 2005 . Interpreting the data: Parallel analysis with sawzall. Sci . Program. 13 , 4, 277 -- 298 . Pike, R., Dorward, S., Griesemer, R., and Quinlan, S. 2005. Interpreting the data: Parallel analysis with sawzall. Sci. Program. 13, 4, 277--298.","journal-title":"Program."},{"key":"e_1_2_2_106_1","unstructured":"Prudhommeaux E. and Seaborne A. 2008. SPARQL query language for rdf w3c recommendation. http:\/\/www.w3.org\/TR\/rdf-sparql-query\/.  Prudhommeaux E. and Seaborne A. 2008. SPARQL query language for rdf w3c recommendation. http:\/\/www.w3.org\/TR\/rdf-sparql-query\/."},{"key":"e_1_2_2_107_1","doi-asserted-by":"publisher","DOI":"10.1145\/1989323.1989460"},{"key":"e_1_2_2_108_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2011.5767877"},{"key":"e_1_2_2_109_1","volume-title":"Proceedings of the 8th Extended Semantic Web Conference on the Semanic Web: Research and Applications (ESWC'11)","author":"Ravindra P.","unstructured":"Ravindra , P. , Kim , H. , and Anyanwu , K . 2011. An intermediate algebra for optimizing rdf graph pattern matching on mapreduce . In Proceedings of the 8th Extended Semantic Web Conference on the Semanic Web: Research and Applications (ESWC'11) . 46--61. Ravindra, P., Kim, H., and Anyanwu, K. 2011. An intermediate algebra for optimizing rdf graph pattern matching on mapreduce. In Proceedings of the 8th Extended Semantic Web Conference on the Semanic Web: Research and Applications (ESWC'11). 46--61."},{"key":"e_1_2_2_110_1","doi-asserted-by":"publisher","DOI":"10.1145\/1999299.1999303"},{"key":"e_1_2_2_111_1","doi-asserted-by":"publisher","DOI":"10.1145\/582095.582099"},{"key":"e_1_2_2_112_1","first-page":"4","article-title":"The case for shared nothing","volume":"9","author":"Stonebraker M.","year":"1986","unstructured":"Stonebraker , M. 1986 . The case for shared nothing . IEEE Datab. Engin. Bull. 9 , 1, 4 -- 9 . Stonebraker, M. 1986. The case for shared nothing. IEEE Datab. Engin. Bull. 9, 1, 4--9.","journal-title":"IEEE Datab. Engin. Bull."},{"key":"e_1_2_2_113_1","doi-asserted-by":"publisher","DOI":"10.1145\/1629175.1629197"},{"key":"e_1_2_2_114_1","volume-title":"Proceedings of the International Semantic Web Conference. 764--780","author":"Stutz P.","unstructured":"Stutz , P. , Bernstein , A. , and Cohen , W. W . 2010. Signal\/collect: Graph algorithms for the (semantic) web . In Proceedings of the International Semantic Web Conference. 764--780 . Stutz, P., Bernstein, A., and Cohen, W. W. 2010. Signal\/collect: Graph algorithms for the (semantic) web. In Proceedings of the International Semantic Web Conference. 764--780."},{"key":"e_1_2_2_115_1","doi-asserted-by":"publisher","DOI":"10.14778\/1687553.1687609"},{"key":"e_1_2_2_116_1","volume-title":"Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10)","author":"Thusoo A.","unstructured":"Thusoo , A. , Sarma , J. , Jain , N. , Shao , Z. , Chakka , P. , Zhang , N. , Anthony , S. , Liu , H. , and Murthy , R . 2010a. Hive - A petabyte scale data warehouse using hadoop . In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10) . 996--1005. Thusoo, A., Sarma, J., Jain, N., Shao, Z., Chakka, P., Zhang, N., Anthony, S., Liu, H., and Murthy, R. 2010a. Hive - A petabyte scale data warehouse using hadoop. In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10). 996--1005."},{"key":"e_1_2_2_117_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807278"},{"key":"e_1_2_2_118_1","volume-title":"Principles of Database and Knowledge Base Systems: Volume II: The New Technologies","author":"Ullman J. D.","unstructured":"Ullman , J. D. 1990. Principles of Database and Knowledge Base Systems: Volume II: The New Technologies . W. H. Freeman and Co. , New York . Ullman, J. D. 1990. Principles of Database and Knowledge Base Systems: Volume II: The New Technologies. W. H. Freeman and Co., New York."},{"key":"e_1_2_2_119_1","doi-asserted-by":"publisher","DOI":"10.1145\/79173.79181"},{"key":"e_1_2_2_120_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807222"},{"key":"e_1_2_2_121_1","doi-asserted-by":"publisher","DOI":"10.1145\/1807167.1807296"},{"key":"e_1_2_2_122_1","volume-title":"Proceedings of the 7th Conference on Innovative Data Systems Research (CIDR'13)","author":"Wang G.","unstructured":"Wang , G. , Xie , W. , Demers , A. , and Gehrke , J . 2013. Asynchronous large-scale graph processing made easy . In Proceedings of the 7th Conference on Innovative Data Systems Research (CIDR'13) . Wang, G., Xie, W., Demers, A., and Gehrke, J. 2013. Asynchronous large-scale graph processing made easy. In Proceedings of the 7th Conference on Innovative Data Systems Research (CIDR'13)."},{"key":"e_1_2_2_123_1","volume-title":"Hadoop: The Definitive Guide","author":"White T.","year":"2012","unstructured":"White , T. 2012 . Hadoop: The Definitive Guide . O'Reilly Media . White, T. 2012. Hadoop: The Definitive Guide. O'Reilly Media."},{"key":"e_1_2_2_124_1","doi-asserted-by":"publisher","DOI":"10.1145\/2000824.2000825"},{"key":"e_1_2_2_125_1","doi-asserted-by":"publisher","DOI":"10.1145\/1247480.1247602"},{"key":"e_1_2_2_126_1","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-00887-0_27"},{"key":"e_1_2_2_127_1","volume-title":"Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08)","author":"Yu Y.","unstructured":"Yu , Y. , Isard , M. , Fetterly , D. , Budiu , M. , Erlingsson , U. , Gunda , P. , and Currey , J . 2008. DryadLINQ: A system for general-purpose distributed data-parallel computing using a high-level language . In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08) . 1--14. Yu, Y., Isard, M., Fetterly, D., Budiu, M., Erlingsson, U., Gunda, P., and Currey, J. 2008. DryadLINQ: A system for general-purpose distributed data-parallel computing using a high-level language. In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08). 1--14."},{"key":"e_1_2_2_128_1","doi-asserted-by":"publisher","DOI":"10.1145\/1755913.1755940"},{"key":"e_1_2_2_129_1","volume-title":"Proceedings of the 9th USENIX Conference on Networked Systems Design and Implementation (NSDI'12)","author":"Zaharia M.","unstructured":"Zaharia , M. , Chowdhury , M. , Das , T. , Dave , A. , Ma , J. , McCauley , M. , Franklin , M. J. , Shenker , S. , and Stoica , I . 2012. Resilient distributed datasets: A fault-tolerant abstraction for in-memory cluster computing . In Proceedings of the 9th USENIX Conference on Networked Systems Design and Implementation (NSDI'12) . Zaharia, M., Chowdhury, M., Das, T., Dave, A., Ma, J., McCauley, M., Franklin, M. J., Shenker, S., and Stoica, I. 2012. Resilient distributed datasets: A fault-tolerant abstraction for in-memory cluster computing. In Proceedings of the 9th USENIX Conference on Networked Systems Design and Implementation (NSDI'12)."},{"key":"e_1_2_2_130_1","volume-title":"Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing (HotCloud'10)","author":"Zaharia M.","unstructured":"Zaharia , M. , Chowdhury , M. , Franklin , M. J. , Shenker , S. , and Stoica , I . 2010b. Spark: Cluster computing with working sets . In Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing (HotCloud'10) . 10. Zaharia, M., Chowdhury, M., Franklin, M. J., Shenker, S., and Stoica, I. 2010b. Spark: Cluster computing with working sets. In Proceedings of the 2nd USENIX Conference on Hot Topics in Cloud Computing (HotCloud'10). 10."},{"key":"e_1_2_2_131_1","volume-title":"Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08)","author":"Zaharia M.","unstructured":"Zaharia , M. , Konwinski , A. , Joseph , A. , Katz , R. , and Stoica , I . 2008. Improving mapreduce performance in heterogeneous environments . In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08) . 29--42. Zaharia, M., Konwinski, A., Joseph, A., Katz, R., and Stoica, I. 2008. Improving mapreduce performance in heterogeneous environments. In Proceedings of the 8th USENIX Conference on Operating Systems Design and Implementation (OSDI'08). 29--42."},{"key":"e_1_2_2_132_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10723-012-9204-9"},{"key":"e_1_2_2_133_1","volume-title":"Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10)","author":"Zhou J.","unstructured":"Zhou , J. , Larson , P. , and Chaiken , R . 2010. Incorporating partitioning and parallel plans into the SCOPE optimizer . In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10) . 1060--1071. Zhou, J., Larson, P., and Chaiken, R. 2010. Incorporating partitioning and parallel plans into the SCOPE optimizer. In Proceedings of the 26th IEEE International Conference on Data Engineering (ICDE'10). 1060--1071."},{"key":"e_1_2_2_134_1","first-page":"17","article-title":"MonetDB\/X100 - A dbms in the cpu cache","volume":"28","author":"Zukowski M.","year":"2005","unstructured":"Zukowski , M. , Boncz , P. A. , Nes , N. , and Heman , S. 2005 . MonetDB\/X100 - A dbms in the cpu cache . IEEE Data Engin. Bull. 28 , 2, 17 -- 22 . Zukowski, M., Boncz, P. A., Nes, N., and Heman, S. 2005. MonetDB\/X100 - A dbms in the cpu cache. IEEE Data Engin. Bull. 28, 2, 17--22.","journal-title":"IEEE Data Engin. Bull."}],"container-title":["ACM Computing Surveys"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2522968.2522979","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2522968.2522979","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T07:34:53Z","timestamp":1750232093000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2522968.2522979"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,7,11]]},"references-count":134,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2013,10]]}},"alternative-id":["10.1145\/2522968.2522979"],"URL":"https:\/\/doi.org\/10.1145\/2522968.2522979","relation":{},"ISSN":["0360-0300","1557-7341"],"issn-type":[{"value":"0360-0300","type":"print"},{"value":"1557-7341","type":"electronic"}],"subject":[],"published":{"date-parts":[[2013,7,11]]},"assertion":[{"value":"2012-07-01","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2013-01-01","order":1,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2013-07-11","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}