{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,2]],"date-time":"2025-08-02T17:37:56Z","timestamp":1754156276929,"version":"3.41.2"},"reference-count":41,"publisher":"Emerald","issue":"5","license":[{"start":{"date-parts":[[2018,5,14]],"date-time":"2018-05-14T00:00:00Z","timestamp":1526256000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.emerald.com\/insight\/site-policies"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":["JD"],"published-print":{"date-parts":[[2018,8,3]]},"abstract":"<jats:sec>\n<jats:title content-type=\"abstract-subheading\">Purpose<\/jats:title>\n<jats:p>Advanced usage of web analytics tools allows to capture the content of user queries. Despite their relevant nature, the manual analysis of large volumes of user queries is problematic. The purpose of this paper is to address the problem of named entity recognition in digital library user queries.<\/jats:p>\n<\/jats:sec>\n<jats:sec>\n<jats:title content-type=\"abstract-subheading\">Design\/methodology\/approach<\/jats:title>\n<jats:p>The paper presents a large-scale case study conducted at the Royal Library of Belgium in its online historical newspapers platform BelgicaPress. The object of the study is a data set of 83,854 queries resulting from 29,812 visits over a 12-month period. By making use of information extraction methods, knowledge bases (KBs) and various authority files, this paper presents the possibilities and limits to identify what percentage of end users are looking for person and place names.<\/jats:p>\n<\/jats:sec>\n<jats:sec>\n<jats:title content-type=\"abstract-subheading\">Findings<\/jats:title>\n<jats:p>Based on a quantitative assessment, the method can successfully identify the majority of person and place names from user queries. Due to the specific character of user queries and the nature of the KBs used, a limited amount of queries remained too ambiguous to be treated in an automated manner.<\/jats:p>\n<\/jats:sec>\n<jats:sec>\n<jats:title content-type=\"abstract-subheading\">Originality\/value<\/jats:title>\n<jats:p>This paper demonstrates in an empirical manner how user queries can be extracted from a web analytics tool and how named entities can then be mapped with KBs and authority files, in order to facilitate automated analysis of their content. Methods and tools used are generalisable and can be reused by other collection holders.<\/jats:p>\n<\/jats:sec>","DOI":"10.1108\/jd-09-2017-0133","type":"journal-article","created":{"date-parts":[[2018,5,14]],"date-time":"2018-05-14T09:17:27Z","timestamp":1526289447000},"page":"936-950","source":"Crossref","is-referenced-by-count":8,"title":["Mining user queries with information extraction methods and linked data"],"prefix":"10.1108","volume":"74","author":[{"given":"Anne","family":"Chardonnens","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Ettore","family":"Rizza","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mathias","family":"Coeckelbergs","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Seth","family":"van Hooland","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"140","published-online":{"date-parts":[[2018,5,14]]},"reference":[{"key":"key2021041507440929700_ref001","unstructured":"Alasiry, A.M. (2015), \u201cNamed entity recognition and classification in search queries\u201d, PhD thesis, Birkbeck, University of London, London."},{"first-page":"11","article-title":"L\u2019utilisation des entit\u00e9s nomm\u00e9es pour l\u2019expansion s\u00e9mantique des requ\u00eates web","year":"2014","key":"key2021041507440929700_ref002"},{"issue":"2","key":"key2021041507440929700_ref003","article-title":"Automatic classification of web queries using very large unlabeled query logs","volume":"25","year":"2007","journal-title":"ACM Transactions on Information Systems"},{"volume-title":"Informatique, normes et temps","year":"1999","key":"key2021041507440929700_ref004"},{"first-page":"3","article-title":"Context-aware query classification","year":"2009","key":"key2021041507440929700_ref005"},{"key":"key2021041507440929700_ref006","first-page":"384","article-title":"Improving Europeana search experience using query logs","year":"2011","journal-title":"Research and Advanced Technology for Digital Libraries"},{"first-page":"177","article-title":"Text mining for user query analysis. Everything changes, everything stays the same? Understanding information spaces","year":"2017","key":"key2021041507440929700_ref007"},{"first-page":"10","article-title":"Inside the user\u2019s mind: combiner analyses quantitatives et qualitatives afin de mieux comprendre les pratiques et besoins des utilisateurs des archives et biblioth\u00e8ques","year":"2017","key":"key2021041507440929700_ref008"},{"first-page":"1","article-title":"Impact of OCR errors on the use of digital libraries: towards a better access to information","year":"2017","key":"key2021041507440929700_ref009"},{"issue":"1","key":"key2021041507440929700_ref010","doi-asserted-by":"crossref","first-page":"37","DOI":"10.1177\/001316446002000104","article-title":"A coefficient of agreement for nominal scales","volume":"20","year":"1960","journal-title":"Educational and Psychological Measurement"},{"first-page":"567","article-title":"A piggyback system for joint entity mention detection and linking in web queries","year":"2016","key":"key2021041507440929700_ref011"},{"first-page":"3935","article-title":"Named entity recognition in travel-related search queries","year":"2015","key":"key2021041507440929700_ref012"},{"issue":"5","key":"key2021041507440929700_ref014","doi-asserted-by":"crossref","first-page":"534","DOI":"10.1007\/s10791-010-9135-7","article-title":"Why finding entities in Wikipedia is difficult, sometimes","volume":"13","year":"2010","journal-title":"Information Retrieval"},{"issue":"2","key":"key2021041507440929700_ref015","doi-asserted-by":"crossref","first-page":"32","DOI":"10.1016\/j.ipm.2014.10.006","article-title":"Analysis of named entity recognition and linking for tweets","volume":"51","year":"2015","journal-title":"Information Processing & Management"},{"issue":"4","key":"key2021041507440929700_ref013","article-title":"Semantic enrichment of a multilingual archive with linked open data","volume":"11","year":"2017","journal-title":"Digital Humanities Quarterly"},{"key":"key2021041507440929700_ref016","doi-asserted-by":"crossref","unstructured":"Dijkshoorn, C., Aroyo, L., Schreiber, G., Wielemaker, J. and Jongma, L. (2014), \u201cUsing linked data to diversify search results a case study in cultural heritage\u201d, EKAW, pp. 109-120.","DOI":"10.1007\/978-3-319-13704-9_9"},{"key":"key2021041507440929700_ref017","doi-asserted-by":"crossref","first-page":"178","DOI":"10.1016\/j.sbspro.2011.10.596","article-title":"Named entity recognition for short text messages","volume":"27","year":"2011","journal-title":"Procedia-Social and Behavioral Sciences"},{"key":"key2021041507440929700_ref018","unstructured":"Gei\u00df, J., Spitz, A. and Gertz, M. (2017), \u201cNECKAr: a named entity classifier for Wikidata\u201d, available at: http:\/\/event.ifi.uni-heidelberg.de\/?page_id=532 (accessed 20 September 2017)."},{"issue":"2","key":"key2021041507440929700_ref019","doi-asserted-by":"crossref","first-page":"232","DOI":"10.1108\/JD-10-2014-0149","article-title":"Exploring the information behaviour of users of Welsh newspapers online through web log analysis","volume":"72","year":"2016","journal-title":"Journal of Documentation"},{"first-page":"267","article-title":"Named entity recognition in query","year":"2009","key":"key2021041507440929700_ref020"},{"key":"key2021041507440929700_ref021","first-page":"13","article-title":"Text preparation through extended tokenization","volume":"37","year":"2006","journal-title":"WIT Transactions on Information and Communication Technologies"},{"key":"key2021041507440929700_ref022","unstructured":"Hickey, T. (2016), \u201cVIAF reflections\u201d, available at: www.oclc.org\/content\/dam\/oclc\/events\/2016\/IFLA2016\/presentations\/VIAF-Reflections.pdf (accessed 20 September 2017)."},{"key":"key2021041507440929700_ref023","first-page":"1","volume-title":"\u20189000: 2005\u2019, Quality Management Systems \u2013 Fundamentals and Vocabulary (ISO 9000: 2005)","author":"ISO","year":"2005"},{"issue":"4","key":"key2021041507440929700_ref024","doi-asserted-by":"crossref","first-page":"384","DOI":"10.1080\/19322909.2014.954740","article-title":"Assessment of digitized library and archives materials: a literature review","volume":"8","year":"2014","journal-title":"Journal of Web Librarianship"},{"issue":"1","key":"key2021041507440929700_ref025","first-page":"1","article-title":"Altmetrics and archives","volume":"4","year":"2017","journal-title":"Journal of Contemporary Archival Studies"},{"issue":"2","key":"key2021041507440929700_ref026","doi-asserted-by":"crossref","first-page":"143","DOI":"10.1504\/IJIIDS.2011.038969","article-title":"Query classification using Wikipedia","volume":"5","year":"2011","journal-title":"International Journal of Intelligent Information and Database Systems"},{"first-page":"879","article-title":"Joint entity recognition and disambiguation","year":"2015","key":"key2021041507440929700_ref027"},{"issue":"1","key":"key2021041507440929700_ref028","doi-asserted-by":"crossref","first-page":"3","DOI":"10.1075\/li.30.1.03nad","article-title":"A survey of named entity recognition and classification","volume":"30","year":"2007","journal-title":"Lingvisticae Investigationes"},{"first-page":"683","article-title":"Weakly-supervised discovery of named entities using web search queries","year":"2007","key":"key2021041507440929700_ref029"},{"key":"key2021041507440929700_ref030","doi-asserted-by":"crossref","unstructured":"Rao, D., McNamee, P. and Dredze, M. (2013), \u201cEntity linking: finding extracted entities in a knowledge base\u201d, in Poibeau, T., Saggion, H., Piskorski, J. and Yangarber, R. (Eds), Multi-source, Multilingual Information Extraction and Summarization, Springer, Berlin, Heidelberg, pp. 93-115.","DOI":"10.1007\/978-3-642-28569-1_5"},{"issue":"2","key":"key2021041507440929700_ref031","doi-asserted-by":"crossref","first-page":"443","DOI":"10.1109\/TKDE.2014.2327028","article-title":"Entity linking with a knowledge base: issues, techniques, and solutions","volume":"27","year":"2015","journal-title":"IEEE Transactions on Knowledge and Data Engineering"},{"first-page":"138","article-title":"Results of the WNUT16 named entity recognition shared task","year":"2016","key":"key2021041507440929700_ref032"},{"issue":"4","key":"key2021041507440929700_ref033","article-title":"The problem of \u2018userism\u2019, and how to overcome it in library theory","volume":"12","year":"2007","journal-title":"Information Research"},{"first-page":"640","article-title":"Trank: ranking entity types using the web of data","year":"2013","key":"key2021041507440929700_ref034"},{"key":"key2021041507440929700_ref035","first-page":"170","article-title":"Contextualized ranking of entity types based on knowledge graphs","volume":"37","year":"2016","journal-title":"Web Semantics: Science, Services and Agents on the World Wide Web"},{"volume-title":"Linked Data for Libraries, Archives and Museums: How to Clean, Link and Publish Your Metadata","year":"2014","key":"key2021041507440929700_ref036"},{"issue":"7\/8","key":"key2021041507440929700_ref037","article-title":"Semantic enrichment: a low-barrier infrastructure and proposal for alignment","volume":"21","year":"2015","journal-title":"D-Lib Magazine"},{"issue":"2","key":"key2021041507440929700_ref038","first-page":"232","article-title":"The fallacy of the multi-API culture: conceptual and practical benefits of representational state transfer (REST)","volume":"71","year":"2015","journal-title":"Journal of Documentation"},{"key":"key2021041507440929700_ref039","unstructured":"Wikidata (2017), \u201cStatistics \u2013 Wikidata\u201d, available at: www.wikidata.org\/wiki\/Special:Statistics (accessed 20 September 2017)."},{"issue":"1","key":"key2021041507440929700_ref040","first-page":"1","article-title":"Collection-level user searches in federated digital resource environment","volume":"44","year":"2007","journal-title":"Proceedings of the Association for Information Science and Technology"},{"issue":"2","key":"key2021041507440929700_ref041","doi-asserted-by":"crossref","first-page":"84","DOI":"10.5860\/lrts.58n2.84","article-title":"Understanding the information needs of large-scale digital library users","volume":"58","year":"2014","journal-title":"Library Resources & Technical Services"}],"container-title":["Journal of Documentation"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/JD-09-2017-0133\/full\/xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/JD-09-2017-0133\/full\/html","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T22:34:58Z","timestamp":1753396498000},"score":1,"resource":{"primary":{"URL":"http:\/\/www.emerald.com\/jd\/article\/74\/5\/936-950\/214110"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2018,5,14]]},"references-count":41,"journal-issue":{"issue":"5","published-online":{"date-parts":[[2018,5,14]]},"published-print":{"date-parts":[[2018,8,3]]}},"alternative-id":["10.1108\/JD-09-2017-0133"],"URL":"https:\/\/doi.org\/10.1108\/jd-09-2017-0133","relation":{},"ISSN":["0022-0418"],"issn-type":[{"type":"print","value":"0022-0418"}],"subject":[],"published":{"date-parts":[[2018,5,14]]}}}