{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,16]],"date-time":"2026-01-16T08:45:01Z","timestamp":1768553101375,"version":"3.49.0"},"reference-count":9,"publisher":"Association for Computing Machinery (ACM)","issue":"12","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["Proc. VLDB Endow."],"published-print":{"date-parts":[[2018,8]]},"abstract":"<jats:p>With Data Science continuing to emerge as a powerful differentiator across industries, organisations are now focused on transforming their data into actionable insights. This task is challenging as in today's knowledge-, service-, and cloud-based economy, businesses accumulate massive amounts of raw data from a variety of sources. Data Lakes introduced as a storage repository to organize this raw data in its native format (supporting from relational to NoSQL DBs) until it is needed. The rationale behind a Data Lake is to store raw data and let the data analyst decide how to cook\/curate them later. In this paper, we present the notion of Knowledge Lake, i.e. a contextualized Data Lake. The Knowledge Lake will provide the foundation for big data analytics by automatically curating the raw data in the Data Lake and to prepare them for deriving insights. We present CoreKG-an open source Data and Knowledge Lake service- which offers researchers and developers a single REST API to organize, curate, index and query their data and metadata in the Lake and over time. CoreKG manages multiple database technologies (from Relational to NoSQL) and offers a built-in design for data curation, security and provenance.<\/jats:p>","DOI":"10.14778\/3229863.3236230","type":"journal-article","created":{"date-parts":[[2018,9,10]],"date-time":"2018-09-10T12:12:28Z","timestamp":1536581548000},"page":"1942-1945","source":"Crossref","is-referenced-by-count":63,"title":["CoreKG"],"prefix":"10.14778","volume":"11","author":[{"given":"Amin","family":"Beheshti","sequence":"first","affiliation":[{"name":"Macquarie University, Sydney, Australia"}]},{"given":"Boualem","family":"Benatallah","sequence":"additional","affiliation":[{"name":"University of New South, Wales, Sydney, Australia"}]},{"given":"Reza","family":"Nouri","sequence":"additional","affiliation":[{"name":"University of New South, Wales, Sydney, Australia"}]},{"given":"Alireza","family":"Tabebordbar","sequence":"additional","affiliation":[{"name":"University of New South, Wales, Sydney, Australia"}]}],"member":"320","published-online":{"date-parts":[[2018,8]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.14778\/2732951.2732958"},{"key":"e_1_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/3132847.3133171"},{"key":"e_1_2_1_3_1","volume-title":"Proceedings of the 19th International Conference on Extending Database Technology, EDBT 2016","author":"Beheshti S.","year":"2016","unstructured":"S. Beheshti , B. Benatallah , and H. R. Motahari-Nezhad . Galaxy: A platform for explorative analysis of open data sources . In Proceedings of the 19th International Conference on Extending Database Technology, EDBT 2016 , Bordeaux, France, pages 640--643 , 2016 . S. Beheshti, B. Benatallah, and H. R. Motahari-Nezhad. Galaxy: A platform for explorative analysis of open data sources. In Proceedings of the 19th International Conference on Extending Database Technology, EDBT 2016, Bordeaux, France, pages 640--643, 2016."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1007\/s10619-014-7171-9"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1007\/s00607-016-0490-0"},{"key":"e_1_2_1_6_1","volume-title":"Temporal provenance model (TPM): model and query language. CoRR, abs\/1211.5009","author":"Beheshti S.","year":"2012","unstructured":"S. Beheshti , H. R. M. Nezhad , and B. Benatallah . Temporal provenance model (TPM): model and query language. CoRR, abs\/1211.5009 , 2012 . S. Beheshti, H. R. M. Nezhad, and B. Benatallah. Temporal provenance model (TPM): model and query language. CoRR, abs\/1211.5009, 2012."},{"key":"e_1_2_1_7_1","doi-asserted-by":"publisher","DOI":"10.1145\/3041021.3054726"},{"key":"e_1_2_1_8_1","volume-title":"Big data for all: Privacy and user control in the age of analytics. Nw. J. Tech. & Intell. Prop., 11:xxvii","author":"Tene O.","year":"2012","unstructured":"O. Tene and J. Polonetsky . Big data for all: Privacy and user control in the age of analytics. Nw. J. Tech. & Intell. Prop., 11:xxvii , 2012 . O. Tene and J. Polonetsky. Big data for all: Privacy and user control in the age of analytics. Nw. J. Tech. & Intell. Prop., 11:xxvii, 2012."},{"key":"e_1_2_1_9_1","volume-title":"CIDR 2015, Seventh Biennial Conference on Innovative Data Systems Research","author":"Terrizzano I. G.","year":"2015","unstructured":"I. G. Terrizzano , P. M. Schwarz , M. Roth , and J. E. Colino . Data wrangling: The challenging yourney from the wild to the lake . In CIDR 2015, Seventh Biennial Conference on Innovative Data Systems Research , Asilomar, CA, USA , 2015 . I. G. Terrizzano, P. M. Schwarz, M. Roth, and J. E. Colino. Data wrangling: The challenging yourney from the wild to the lake. In CIDR 2015, Seventh Biennial Conference on Innovative Data Systems Research, Asilomar, CA, USA, 2015."}],"container-title":["Proceedings of the VLDB Endowment"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.14778\/3229863.3236230","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2022,12,28]],"date-time":"2022-12-28T10:14:22Z","timestamp":1672222462000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.14778\/3229863.3236230"}},"subtitle":["a knowledge lake service"],"short-title":[],"issued":{"date-parts":[[2018,8]]},"references-count":9,"journal-issue":{"issue":"12","published-print":{"date-parts":[[2018,8]]}},"alternative-id":["10.14778\/3229863.3236230"],"URL":"https:\/\/doi.org\/10.14778\/3229863.3236230","relation":{},"ISSN":["2150-8097"],"issn-type":[{"value":"2150-8097","type":"print"}],"subject":[],"published":{"date-parts":[[2018,8]]}}}