{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T04:32:47Z","timestamp":1750307567011,"version":"3.41.0"},"reference-count":13,"publisher":"Association for Computing Machinery (ACM)","issue":"1","license":[{"start":{"date-parts":[[2010,3,12]],"date-time":"2010-03-12T00:00:00Z","timestamp":1268352000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGOPS Oper. Syst. Rev."],"published-print":{"date-parts":[[2010,3,12]]},"abstract":"<jats:p>Information management applications exhibit a wide range of query performance and result freshness goals. Some applications, such as web search, require interactive performance, but may safely operate on stale data. Others, such as policy violation detection, require up-to-date results, but can tolerate relaxed performance goals. Furthermore, information processing applications must be able to ingest updates at the scale of an entire organization. In this paper, we present LazyBase, a system that allows users to trade off query performance and result freshness in order to satisfy the full range of users' goals. LazyBase breaks up data ingestion into a pipeline of operations to minimize ingest time and uses models of processing and query performance to execute user queries. Initial results with LazyBase illustrate the feasibility of the pipelined model, highlight a rich space of trade-offs between result freshness and query performance, and often outperform existing solutions in the space.<\/jats:p>","DOI":"10.1145\/1740390.1740395","type":"journal-article","created":{"date-parts":[[2010,3,19]],"date-time":"2010-03-19T19:22:47Z","timestamp":1269026567000},"page":"15-19","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":13,"title":["LazyBase"],"prefix":"10.1145","volume":"44","author":[{"given":"Kimberly","family":"Keeton","sequence":"first","affiliation":[{"name":"Hewlett-Packard Laboratories"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"suffix":"III","given":"Charles B.","family":"Morrey","sequence":"additional","affiliation":[{"name":"Hewlett-Packard Laboratories"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Craig A.N.","family":"Soules","sequence":"additional","affiliation":[{"name":"Hewlett-Packard Laboratories"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Alistair","family":"Veitch","sequence":"additional","affiliation":[{"name":"Hewlett-Packard Laboratories"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2010,3,12]]},"reference":[{"key":"e_1_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.1145\/1496909.1496923"},{"key":"e_1_2_1_2_1","volume-title":"http:\/\/ lucene.apache.org\/","author":"Lucene Apache","year":"2009","unstructured":"Apache Lucene . http:\/\/ lucene.apache.org\/ . 2009 . Apache Lucene. http:\/\/ lucene.apache.org\/. 2009."},{"key":"e_1_2_1_3_1","volume-title":"Proc. CIDR","author":"Armbrust M.","year":"2009","unstructured":"M. Armbrust , A. Fox , D.A. Patterson , N. Lanham , B. Trushkowsky , J. Trutna , and H. Oh . SCADS: Scale-independent storage for social computing applications . In Proc. CIDR , January 2009 . M. Armbrust, A. Fox, D.A. Patterson, N. Lanham, B. Trushkowsky, J. Trutna, and H. Oh. SCADS: Scale-independent storage for social computing applications. In Proc. CIDR, January 2009."},{"key":"e_1_2_1_4_1","doi-asserted-by":"publisher","DOI":"10.1145\/301816.301823"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1145\/1099554.1099645"},{"key":"e_1_2_1_6_1","first-page":"205","volume-title":"Proc. OSDI","author":"Chang F.","year":"2006","unstructured":"F. Chang , J. Dean , S. Ghemawat , W.C. Hsieh , D.A. Wallach , M. Burrows , T. Chandra , A. Fikes , and R.E. Gruber . Bigtable: A distributed storage system for structured data . In Proc. OSDI , pages 205 -- 218 , November 2006 . F. Chang, J. Dean, S. Ghemawat, W.C. Hsieh, D.A. Wallach, M. Burrows, T. Chandra, A. Fikes, and R.E. Gruber. Bigtable: A distributed storage system for structured data. In Proc. OSDI, pages 205--218, November 2006."},{"key":"e_1_2_1_7_1","first-page":"137","volume-title":"Proc. OSDI","author":"Dean J.","year":"2004","unstructured":"J. Dean and S. Ghemawat . MapReduce: Simplified data processing on large clusters . In Proc. OSDI , pages 137 -- 150 , 2004 . J. Dean and S. Ghemawat. MapReduce: Simplified data processing on large clusters. In Proc. OSDI, pages 137--150, 2004."},{"volume-title":"http:\/\/ hadoop.apache.org\/","year":"2009","key":"e_1_2_1_8_1","unstructured":"Hadoop. http:\/\/ hadoop.apache.org\/ . 2009 . Hadoop. http:\/\/ hadoop.apache.org\/. 2009."},{"key":"e_1_2_1_9_1","volume-title":"Eidgenossiche Technische Hochschule Zurich (ETHZ)","author":"Hildenbrand S.","year":"2008","unstructured":"S. Hildenbrand . Performance tradeoffs in write-optimized databases. Technical report , Eidgenossiche Technische Hochschule Zurich (ETHZ) , 2008 . S. Hildenbrand. Performance tradeoffs in write-optimized databases. Technical report, Eidgenossiche Technische Hochschule Zurich (ETHZ), 2008."},{"key":"e_1_2_1_10_1","first-page":"15","volume-title":"Proc. 27th Australian Conf. on Computer Science (ACSC)","author":"Lester N.","year":"2004","unstructured":"N. Lester , J. Zobel , and H.E. Williams . In-place versus re-build versus re-merge: index maintenance strategies for text retrieval systems . Proc. 27th Australian Conf. on Computer Science (ACSC) , pages 15 -- 22 , 2004 . N. Lester, J. Zobel, and H.E. Williams. In-place versus re-build versus re-merge: index maintenance strategies for text retrieval systems. Proc. 27th Australian Conf. on Computer Science (ACSC), pages 15--22, 2004."},{"key":"e_1_2_1_11_1","doi-asserted-by":"publisher","DOI":"10.1145\/320473.320484"},{"key":"e_1_2_1_12_1","doi-asserted-by":"publisher","DOI":"10.1109\/ICDE.2003.1260779"},{"key":"e_1_2_1_13_1","doi-asserted-by":"publisher","DOI":"10.1145\/1519065.1519079"}],"container-title":["ACM SIGOPS Operating Systems Review"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1740390.1740395","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/1740390.1740395","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,18]],"date-time":"2025-06-18T12:40:55Z","timestamp":1750250455000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/1740390.1740395"}},"subtitle":["freshness vs. performance in information management"],"short-title":[],"issued":{"date-parts":[[2010,3,12]]},"references-count":13,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2010,3,12]]}},"alternative-id":["10.1145\/1740390.1740395"],"URL":"https:\/\/doi.org\/10.1145\/1740390.1740395","relation":{},"ISSN":["0163-5980"],"issn-type":[{"type":"print","value":"0163-5980"}],"subject":[],"published":{"date-parts":[[2010,3,12]]},"assertion":[{"value":"2010-03-12","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}