{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2023,1,10]],"date-time":"2023-01-10T11:08:33Z","timestamp":1673348913872},"reference-count":16,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2016,1,26]],"date-time":"2016-01-26T00:00:00Z","timestamp":1453766400000},"content-version":"unspecified","delay-in-days":0,"URL":"http:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Datenbank Spektrum"],"published-print":{"date-parts":[[2016,3]]},"abstract":"<jats:title>Abstract<\/jats:title>\n          <jats:p>Probabilistic Datalog (PDatalog, proposed in 1995) is a probabilistic variant of Datalog and a nice conceptual idea to model Information Retrieval in a logical, rule-based programming paradigm. Making PDatalog work in real-world applications requires more than probabilistic facts and rules, and the semantics associated with the evaluation of the programs. We report in this paper some of the key features of the HySpirit system required to scale the execution of PDatalog programs.<\/jats:p>\n          <jats:p>Firstly, there is the requirement to express <jats:italic>probability estimation<\/jats:italic> in PDatalog. Secondly, fuzzy-like predicates are required to model <jats:italic>vague predicates<\/jats:italic> (e.g. vague match of attributes such as age or price). Thirdly, to handle large data sets there are scalability issues to be addressed, and therefore, HySpirit provides <jats:italic>probabilistic relational indexes<\/jats:italic> and <jats:italic>parallel and distributed processing<\/jats:italic>. The main contribution of this paper is a consolidated view on the methods of the HySpirit system to make PDatalog applicable in real-scale applications that involve a wide range of requirements typical for data (information) management and analysis.<\/jats:p>","DOI":"10.1007\/s13222-015-0208-z","type":"journal-article","created":{"date-parts":[[2016,1,26]],"date-time":"2016-01-26T07:13:39Z","timestamp":1453792419000},"page":"39-48","update-policy":"http:\/\/dx.doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Scalable DB+IR Technology: Processing Probabilistic Datalog with HySpirit"],"prefix":"10.1007","volume":"16","author":[{"given":"Ingo","family":"Frommholz","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Thomas","family":"Roelleke","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2016,1,26]]},"reference":[{"key":"208_CR1","doi-asserted-by":"crossref","unstructured":"Azzam H, Yahyaei S, Bonzanini M, Roelleke T (2012) A schema-driven approach for knowledge-oriented retrieval and query formulation. In: Proceedings of the Third International Workshop on Keyword Search on Structured Data - KEYS '12. ACM, Scottsdale, AZ, USA. doi:10.1145\/2254736.2254746. URL http:\/\/dl.acm.org\/citation.cfm?doid=2254736.2254746","DOI":"10.1145\/2254736.2254746"},{"key":"208_CR2","unstructured":"Cornacchia R, Kamps J, Alink W, de Vries AP (2013) Searching political data by strategy. In: Lupu M, Salampasis M, Fuhr N, Hanbury A, Larsen B, Strindberg H (eds) Proceedings of the Integrating IR technologies for Professional Search Workshop. CEUR-WS.org, Moscow, pp\u00a088\u201391. http:\/\/ceur-ws.org\/Vol-968\/irps_15.pdf"},{"key":"208_CR3","doi-asserted-by":"crossref","unstructured":"Frommholz I, Fuhr N (2006) Probabilistic, object-oriented logics for annotation-based retrieval in digital libraries. In: Nelson M, Marshall C, Marchionini G (eds) Proc. of the 6th ACM\/IEEE Joint Conference on Digital Libraries (JCDL 2006). ACM, New York, pp\u00a055\u201364","DOI":"10.1145\/1141753.1141764"},{"key":"208_CR4","doi-asserted-by":"crossref","unstructured":"Fuhr N (2000) Probabilistic datalog: implementing logical information retrieval for advanced applications. J Am Soc Inf Sci 51:95\u2013110","DOI":"10.1002\/(SICI)1097-4571(2000)51:2<95::AID-ASI2>3.0.CO;2-H"},{"key":"208_CR5","doi-asserted-by":"crossref","unstructured":"Fuhr N (2014) Bridging information retrieval and databases. In: Ferro N (ed) Bridging between information retrieval and databases. Springer, Berlin, pp 97\u2013115. doi:10.1007\/978-3-642-54798-0fn{_}g5","DOI":"10.1007\/978-3-642-54798-0_5"},{"key":"208_CR6","doi-asserted-by":"crossref","unstructured":"Fuhr N, G\u00f6vert N, R\u00f6lleke T (1998) DOLORES: a system for logic-based retrieval of multimedia objects. In: Croft WB, Moffat A, van Rijsbergen C, Wilkinson R, Zobel J (eds) Proceedings of the 21st Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, pp\u00a0257\u2013265. ACM, New York (1998)","DOI":"10.1145\/290941.291005"},{"key":"208_CR7","doi-asserted-by":"crossref","unstructured":"Fuhr N, R\u00f6lleke T (1997) A probabilistic relational algebra for the Integration of information retrieval and database systems. ACM Transactions on Information Systems 14, 32\u201366","DOI":"10.1145\/239041.239045"},{"key":"208_CR8","doi-asserted-by":"crossref","unstructured":"Fuhr N, R\u00f6lleke T (1998) HySpirit \u2013 a probabilistic inference engine for hypermedia retrieval in large databases. In: Proceedings of the 6th International Conference on Extending Database Technology (EDBT), pp\u00a024\u201338. Springer, Heidelberg et al.","DOI":"10.1007\/BFb0100975"},{"key":"208_CR9","doi-asserted-by":"crossref","unstructured":"Klampanos I, Azzam H, Roelleke T (2009) A case for probabilistic logic for scalable patent retrieval. In: CIKM Workshop on Patent Retrieval","DOI":"10.1145\/1651343.1651345"},{"key":"208_CR10","doi-asserted-by":"crossref","unstructured":"Lalmas M, R\u00f6lleke T (2003) Four-valued knowledge augmentation for structured document retrieval. Int J Uncertain Fuzziness Knowledge- Based Syst 11:67\u201385","DOI":"10.1142\/S0218488503001953"},{"key":"208_CR11","doi-asserted-by":"crossref","unstructured":"Ounis I, Amati G, Plachouras V, He B, Macdonald C, Lioma C (2006) Terrier: A High Performance and Scalable Information Retrieval Platform. In: Proceedings of ACM SIGIR'06 Workshop on Open Source Information Retrieval (OSIR 2006)","DOI":"10.1007\/978-3-540-31865-1_37"},{"key":"208_CR12","doi-asserted-by":"crossref","unstructured":"Roelleke T (2003) A frequency-based and a Poisson-based probability of being informative. In: ACM SIGIR. Toronto, pp\u00a0227\u2013234","DOI":"10.1145\/860435.860478"},{"key":"208_CR13","unstructured":"Roelleke T (2003) The relational Bayes for frequency-based and information-theoretic probability estimation in a probabilistic relational algebra. Patent application 0322328.6"},{"key":"208_CR14","doi-asserted-by":"crossref","unstructured":"Roelleke T (2013) Information retrieval models: foundations and relationships. Morgan & Claypool. doi:10.2200\/S00494ED1V01Y201304ICR027","DOI":"10.1145\/2499178.2499203"},{"key":"208_CR15","doi-asserted-by":"crossref","unstructured":"Roelleke T, Bonzanini M, Martinez-Alvarez M (2013) On the modelling of ranking algorithms in probabilistic datalog categories and subject descriptors. In: Proceedings of the 7th International Workshop on Ranking in Databases, 1, pp\u00a04\u20139. Riva del Garda, Italy. doi:10.1145\/2524828.2524832","DOI":"10.1145\/2524828.2524832"},{"key":"208_CR16","doi-asserted-by":"crossref","unstructured":"Roelleke T, Wu H, Wang J, Azzam H (2008) Modelling retrieval models in a probabilistic relational algebra with a new operator: the relational Bayes. The VLDB Journal - The International Journal on Very Large Data Bases, Special Issue on DB & IR 17(1):5\u201337. http:\/\/portal.acm.org\/citation.cfm?id=1325167","DOI":"10.1007\/s00778-007-0073-y"}],"container-title":["Datenbank-Spektrum"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s13222-015-0208-z.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/article\/10.1007\/s13222-015-0208-z\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"http:\/\/link.springer.com\/content\/pdf\/10.1007\/s13222-015-0208-z","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s13222-015-0208-z.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2021,9,2]],"date-time":"2021-09-02T20:10:59Z","timestamp":1630613459000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s13222-015-0208-z"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2016,1,26]]},"references-count":16,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2016,3]]}},"alternative-id":["208"],"URL":"https:\/\/doi.org\/10.1007\/s13222-015-0208-z","relation":{},"ISSN":["1618-2162","1610-1995"],"issn-type":[{"value":"1618-2162","type":"print"},{"value":"1610-1995","type":"electronic"}],"subject":[],"published":{"date-parts":[[2016,1,26]]},"assertion":[{"value":"5 November 2015","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 December 2015","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"26 January 2016","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}}]}}