{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,9,26]],"date-time":"2025-09-26T13:18:22Z","timestamp":1758892702843,"version":"3.41.0"},"reference-count":33,"publisher":"Association for Computing Machinery (ACM)","issue":"4","license":[{"start":{"date-parts":[[2014,2,28]],"date-time":"2014-02-28T00:00:00Z","timestamp":1393545600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGMOD Rec."],"published-print":{"date-parts":[[2014,2,28]]},"abstract":"<jats:p>\n            Many organizations today are faced with the challenge of processing and distilling information from huge and growing collections of data. Such organizations are increasingly deploying sophisticated mathematical algorithms to model the behavior of their business processes to discover correlations in the data, to predict trends and ultimately drive decisions to optimize their operations. These techniques, are known collectively as\n            <jats:italic>analytics<\/jats:italic>\n            , and draw upon multiple disciplines, including statistics, quantitative analysis, data mining, and machine learning. In this survey paper, we identify some of the key techniques employed in analytics both to serve as an introduction for the non-specialist and to explore the opportunity for greater optimizations for parallelization and acceleration using commodity and specialized multi-core processors. We are interested in isolating and documenting repeated patterns in analytical algorithms, data structures and data types, and in understanding howthese could be most effectively mapped onto parallel infrastructure. To this end, we focus on analytical models that can be executed using different algorithms. For most major model types, we study implementations of key algorithms to determine common computational and runtime patterns. We then use this information to characterize and recommend suitable parallelization strategies for these algorithms, specifically when used in data management workloads.\n          <\/jats:p>","DOI":"10.1145\/2590989.2590993","type":"journal-article","created":{"date-parts":[[2014,3,4]],"date-time":"2014-03-04T13:24:59Z","timestamp":1393939499000},"page":"17-28","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":6,"title":["Analyzing analytics"],"prefix":"10.1145","volume":"42","author":[{"given":"Rajesh","family":"Bordawekar","sequence":"first","affiliation":[{"name":"IBM Watson Research Center, Yorktown Heights, NY"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Bob","family":"Blainey","sequence":"additional","affiliation":[{"name":"IBM Toronto Software Lab, Markham, Ontario"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Chidanand","family":"Apte","sequence":"additional","affiliation":[{"name":"IBM Watson Research Center, Yorktown Heights, NY"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2014,2,28]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1287\/inte.1030.0060"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.5555\/2124408"},{"key":"e_1_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1080\/09332480.2010.10739787"},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/1557019.1557136"},{"key":"e_1_2_1_7_1","volume-title":"December","author":"Crosbie P.","year":"2003","unstructured":"P. Crosbie and J. Bohn . Modeling default risk , December 2003 . Moody's KMV. P. Crosbie and J. Bohn. Modeling default risk, December 2003. Moody's KMV."},{"key":"e_1_2_1_8_1","volume-title":"The New Science of Winning","author":"Davenport T.","year":"2007","unstructured":"T. Davenport and J. Harris . Competing on Analytics , The New Science of Winning . Harvard Business School Press , 2007 . T. Davenport and J. Harris. Competing on Analytics, The New Science of Winning. Harvard Business School Press, 2007."},{"key":"e_1_2_1_9_1","volume-title":"Smarter Decisions, Better Results","author":"Davenport T.","year":"2010","unstructured":"T. Davenport , J. Harris , and R. Morison . Analytics at Work , Smarter Decisions, Better Results . Harvard Business School Press , 2010 . T. Davenport, J. Harris, and R. Morison. Analytics at Work, Smarter Decisions, Better Results. Harvard Business School Press, 2010."},{"key":"e_1_2_1_10_1","doi-asserted-by":"publisher","DOI":"10.1145\/1629175.1629198"},{"key":"e_1_2_1_11_1","unstructured":"D. J. Feldman. Netezza Performance Architecture. Keynote Presentation at the DaMoN'13 Workshop.  D. J. Feldman. Netezza Performance Architecture. Keynote Presentation at the DaMoN'13 Workshop."},{"key":"e_1_2_1_12_1","volume-title":"Building Watson: An Overview of the DeepQA Project. AI Magazine, 59(Fall)","author":"Ferrucci D.","year":"2010","unstructured":"D. Ferrucci , E. Brown , J. Chu-Carroll , J. Fan , D. Gondek , A. A. Kalyanpur , A. Lally , J. W. Murdock , E. Nyberg , J. Prager , N. Schlaefer , and C. Welty . Building Watson: An Overview of the DeepQA Project. AI Magazine, 59(Fall) , 2010 . D. Ferrucci, E. Brown, J. Chu-Carroll, J. Fan, D. Gondek, A. A. Kalyanpur, A. Lally, J. W. Murdock, E. Nyberg, J. Prager, N. Schlaefer, and C. Welty. Building Watson: An Overview of the DeepQA Project. AI Magazine, 59(Fall), 2010."},{"key":"e_1_2_1_13_1","volume-title":"Computer Science Department","author":"Go A.","year":"2009","unstructured":"A. Go , R. Bhayani , and L. Huang . Twitter sentiment classification using distant supervision. Technical report , Computer Science Department , Stanford University , December 2009 . A. Go, R. Bhayani, and L. Huang. Twitter sentiment classification using distant supervision. Technical report, Computer Science Department, Stanford University, December 2009."},{"key":"e_1_2_1_14_1","volume-title":"The New York Times","author":"Goode E.","year":"2011","unstructured":"E. Goode . Sending the police before there's a crime . The New York Times , August 16, 2011 . E. Goode. Sending the police before there's a crime. The New York Times, August 16, 2011."},{"key":"e_1_2_1_15_1","volume-title":"Proc. of the VLDB Endowment","author":"Grosse P.","year":"2011","unstructured":"P. Grosse , W. Lehner , T. Weichert , F. Faerber , and W.-S. Li . Bridging Two Worlds with RICE: Integrating R into the SAP In-Memory Computing Engine . In Proc. of the VLDB Endowment , September 2011 . P. Grosse, W. Lehner, T. Weichert, F. Faerber, and W.-S. Li. Bridging Two Worlds with RICE: Integrating R into the SAP In-Memory Computing Engine. In Proc. of the VLDB Endowment, September 2011."},{"key":"e_1_2_1_16_1","volume-title":"Data Mining: Concepts and Techniques","author":"Han J.","year":"2006","unstructured":"J. Han and M. Kamber . Data Mining: Concepts and Techniques . Morgan Kaufmann Publishers , 2006 . J. Han and M. Kamber. Data Mining: Concepts and Techniques. Morgan Kaufmann Publishers, 2006."},{"key":"e_1_2_1_17_1","unstructured":"IBM Corp. Engineering and Scientific Subroutine Library (ESSL) and Parallel ESSL. www.ibm.com\/systems\/software\/essl.  IBM Corp. Engineering and Scientific Subroutine Library (ESSL) and Parallel ESSL. www.ibm.com\/systems\/software\/essl."},{"key":"e_1_2_1_18_1","unstructured":"Intel Inc. Intel Math Kernel Library. software.intel.com.  Intel Inc. Intel Math Kernel Library. software.intel.com."},{"issue":"10","key":"e_1_2_1_19_1","first-page":"40","article-title":"Pandora and the Music Genome Project","volume":"23","author":"Joyce J.","year":"2006","unstructured":"J. Joyce . Pandora and the Music Genome Project . Scientific Computing , 23 ( 10 ): 40 -- 41 , September 2006 . J. Joyce. Pandora and the Music Genome Project. Scientific Computing, 23(10):40--41, September 2006.","journal-title":"Scientific Computing"},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1287\/trsc.1030.0026"},{"key":"e_1_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1198\/jasa.2011.ap09546"},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/1183614.1183678"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2008.02.021"},{"key":"e_1_2_1_25_1","volume-title":"Oracle Advanced Analytics","author":"Oracle Inc.","year":"2013","unstructured":"Oracle Inc. Oracle Advanced Analytics , 2013 . Oracle Data Sheet . Oracle Inc. Oracle Advanced Analytics, 2013. Oracle Data Sheet."},{"key":"e_1_2_1_26_1","volume-title":"Mining Massive Datasets","author":"Rajaraman A.","year":"2010","unstructured":"A. Rajaraman and J. Ullman . Mining Massive Datasets . Cambridge University Press , 2010 . Free version available at infolab.stanford.edu\/~ullman\/mmds.html. A. Rajaraman and J. Ullman. Mining Massive Datasets. Cambridge University Press, 2010. Free version available at infolab.stanford.edu\/~ullman\/mmds.html."},{"key":"e_1_2_1_27_1","volume-title":"Procs. of the Predictive Analytics World","author":"Rexer K.","year":"2010","unstructured":"K. Rexer , H. Allen , and P. Gearan . 2010 data miner survey summary . In Procs. of the Predictive Analytics World , October 2010 . K. Rexer, H. Allen, and P. Gearan. 2010 data miner survey summary. In Procs. of the Predictive Analytics World, October 2010."},{"key":"e_1_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1137\/1.9781611972801.64"},{"key":"e_1_2_1_29_1","volume-title":"February","author":"Issue Science Special","year":"2011","unstructured":"Science Special Issue . Dealing with data. Science, 331(6018) , February 2011 . Science Special Issue. Dealing with data. Science, 331(6018), February 2011."},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.1147\/JRD.2011.2163281"},{"key":"e_1_2_1_31_1","unstructured":"Sony Pictures Inc. Jeopardy! The IBM Challenge. www.jeopardy.com\/minisites\/watson.  Sony Pictures Inc. Jeopardy! The IBM Challenge. www.jeopardy.com\/minisites\/watson."},{"key":"e_1_2_1_32_1","unstructured":"Splunk Inc. Splunk Tutorial 2011. www.splunk.com.  Splunk Inc. Splunk Tutorial 2011. www.splunk.com."},{"key":"e_1_2_1_33_1","volume-title":"October","author":"Suman A.","year":"2006","unstructured":"A. Suman . Automated face recognition , applications within law enforcement: Market and technology review , October 2006 . A. Suman. Automated face recognition, applications within law enforcement: Market and technology review, October 2006."},{"key":"e_1_2_1_34_1","volume-title":"13th","author":"Economist The","year":"2007","unstructured":"The Economist . Algorithms : Business by numbers. Print Edition , 13th September , 2007 . The Economist. Algorithms: Business by numbers. Print Edition, 13th September, 2007."},{"key":"e_1_2_1_35_1","volume-title":"Data, data everywhere. Print Edition, 25th","author":"Economist The","year":"2010","unstructured":"The Economist . Data, data everywhere. Print Edition, 25th February , 2010 . The Economist. Data, data everywhere. Print Edition, 25th February, 2010."},{"key":"e_1_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10115-007-0114-2"}],"container-title":["ACM SIGMOD Record"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2590989.2590993","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/2590989.2590993","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T20:13:52Z","timestamp":1750277632000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/2590989.2590993"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2014,2,28]]},"references-count":33,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2014,2,28]]}},"alternative-id":["10.1145\/2590989.2590993"],"URL":"https:\/\/doi.org\/10.1145\/2590989.2590993","relation":{},"ISSN":["0163-5808"],"issn-type":[{"type":"print","value":"0163-5808"}],"subject":[],"published":{"date-parts":[[2014,2,28]]},"assertion":[{"value":"2014-02-28","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}