{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,2]],"date-time":"2025-08-02T17:35:11Z","timestamp":1754156111165,"version":"3.41.2"},"reference-count":31,"publisher":"Emerald","issue":"3","license":[{"start":{"date-parts":[[2013,8,23]],"date-time":"2013-08-23T00:00:00Z","timestamp":1377216000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.emerald.com\/insight\/site-policies"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2013,8,23]]},"abstract":"<jats:sec><jats:title content-type=\"abstract-heading\">Purpose<\/jats:title><jats:p>The purpose of this paper is to focus on the problem of named entity disambiguation. The paper disambiguates named entities on a very detailed level. To each entity is assigned a concrete identifier of a corresponding Wikipedia article describing the entity.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Design\/methodology\/approach<\/jats:title><jats:p>For such a fine\u2010grained disambiguation a correct representation of the context is crucial. The authors compare various context representations: bag of words representation, linguistic representation and structured co\u2010occurrence representation. Models for each representation are described and evaluated. They also investigate the possibilities of multilingual named entity disambiguation.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Findings<\/jats:title><jats:p>Based on this evaluation, the structured co\u2010occurrence representation provides the best disambiguation results. It showed up that this method could be successfully applied also on other languages, not only on English.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Research limitations\/implications<\/jats:title><jats:p>Despite its good results the structured co\u2010occurrence context representation has several limitations. It trades precision for recall, which might not be desirable in some use cases. Also it is not able to disambiguate two different types of entities, which are mentioned under the same name in the same text. These limitations can be overcome by combination with other described methods.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Practical implications<\/jats:title><jats:p>The authors provide a ready\u2010made web service, which can be directly plugged in existing applications using a REST interface.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Originality\/value<\/jats:title><jats:p>The paper proposes a new approach to named entity disambiguation exploiting various context representation models (bag of words, linguistic and structural representation). The authors constructed a comprehensive dataset based on all English Wikipedia articles for named entity disambiguation. They evaluated and compared the individual context representation models on this dataset. They evaluate the support of multiple languages.<\/jats:p><\/jats:sec>","DOI":"10.1108\/ijwis-05-2013-0016","type":"journal-article","created":{"date-parts":[[2013,8,14]],"date-time":"2013-08-14T12:51:31Z","timestamp":1376484691000},"page":"242-259","source":"Crossref","is-referenced-by-count":2,"title":["Various approaches to text representation for named entity disambiguation"],"prefix":"10.1108","volume":"9","author":[{"given":"Ivo","family":"La\u0161ek","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Peter","family":"Vojt\u00e1\u0161","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"140","reference":[{"key":"key2022012419453998700_b1","doi-asserted-by":"crossref","unstructured":"Asahara, M. and Matsumoto, Y. (2003), \u201cJapanese named entity extraction with redundant morphological analysis\u201d, Proceedings of the 2003 Conference of the North American Chapter of the Association for Computational Linguistics on Human Language Technology, Vol. 1, Association for Computational Linguistics, Stroudsburg, PA, pp. 8\u201015.","DOI":"10.3115\/1073445.1073447"},{"key":"key2022012419453998700_b2","doi-asserted-by":"crossref","unstructured":"Auer, S., Bizer, C., Kobilarov, G., Lehmann, J., Cyganiak, R. and Ives, Z. (2007), \u201cDbpedia: a nucleus for a web of open data\u201d, The Semantic Web, Vol. 4825, pp. 722\u2010735.","DOI":"10.1007\/978-3-540-76298-0_52"},{"key":"key2022012419453998700_b3","doi-asserted-by":"crossref","unstructured":"Awang Iskandar, D., Pehcevski, J., Thom, J. and Tahaghoghi, S. (2007), \u201cSocial media retrieval using image features and structured text\u201d, Comparative Evaluation of XML Information Retrieval Systems, Lecture Notes in Computer Science, Vol. 4518, pp. 358\u2010372.","DOI":"10.1007\/978-3-540-73888-6_35"},{"key":"key2022012419453998700_b4","doi-asserted-by":"crossref","unstructured":"Berners\u2010Lee, T., Hendler, J. and Lassila, O. (2001), \u201cThe semantic web\u201d, Scientific American, Vol. 284 No. 5, pp. 28\u201037.","DOI":"10.1038\/scientificamerican0501-34"},{"key":"key2022012419453998700_b5","doi-asserted-by":"crossref","unstructured":"Bikel, D.M., Miller, S., Schwartz, R. and Weischedel, R. (1997), \u201cNymble: a high\u2010performance learning name\u2010finder\u201d, Proceedings of the Fifth Conference on Applied Natural Language Processing, Stroudsburg, PA, USA, pp. 194\u2010201.","DOI":"10.3115\/974557.974586"},{"key":"key2022012419453998700_b6","doi-asserted-by":"crossref","unstructured":"Bizer, C., Heath, T. and Berners\u2010Lee, T. (2009a), \u201cLinked data\u2010the story so far\u201d, International Journal on Semantic Web and Information Systems (IJSWIS), Vol. 5 No. 3, pp. 1\u201022.","DOI":"10.4018\/jswis.2009081901"},{"key":"key2022012419453998700_b7","doi-asserted-by":"crossref","unstructured":"Bizer, C., Lehmann, J., Kobilarov, G., Auer, S., Becker, C., Cyganiak, R. and Hellmann, S. (2009b), \u201cDBpedia \u2013 a crystallization point for the web of data\u201d, Web Semantics: Science, Services and Agents on the World Wide Web, Vol. 7 No. 3, pp. 154\u2010165.","DOI":"10.1016\/j.websem.2009.07.002"},{"key":"key2022012419453998700_b8","unstructured":"Borthwick, A., Sterling, J., Agichtein, E. and Grishman, R. (1998), \u201cDescription of the MENE named entity system as used in MUC\u20107\u201d, Proceedings of the Seventh Message Understanding Conference (MUC\u20107)."},{"key":"key2022012419453998700_b9","unstructured":"Carlson, A., Betteridge, J., Kisiel, B., Settles, B., Hruschk, E.R. and Mitchell, T.M. (2010), \u201cToward an architecture for never\u2010ending language learning\u201d, Proceedings of the Twenty\u2010fourth Conference on Artificial Intelligence (AAAI 2010), Vol. 2 No. 4, pp. 1306\u20101313."},{"key":"key2022012419453998700_b10","doi-asserted-by":"crossref","unstructured":"Finkel, J.R., Grenager, T. and Manning, C. (2005), \u201cIncorporating non\u2010local information into information extraction systems by gibbs sampling\u201d, Proceedings of the 43rd Annual Meeting on Association for Computational Linguistics, Association for Computational Linguistics, Stroudsburg, PA, pp. 363\u2010370.","DOI":"10.3115\/1219840.1219885"},{"key":"key2022012419453998700_b11","doi-asserted-by":"crossref","unstructured":"Gruhl, D., Nagarajan, M., Pieper, J., Robson, Ch. and Sheth, A. (2009), \u201cContext and domain knowledge enhanced entity spotting in informal text\u201d, The Semantic Web \u2013 ISWC 2009, Lecture Notes in Computer Science, Springer, Berlin, pp. 260\u2010276.","DOI":"10.1007\/978-3-642-04930-9_17"},{"key":"key2022012419453998700_b12","doi-asserted-by":"crossref","unstructured":"Hassell, J., Boanerges, A.M. and Arpinar, I. (2006), \u201cOntology\u2010driven automatic entity disambiguation in unstructured text\u201d, The Semantic Web \u2013 ISWC 2006, Lecture Notes in Computer Science, Springer, Berlin, pp. 44\u201057.","DOI":"10.1007\/11926078_4"},{"key":"key2022012419453998700_b13","unstructured":"Lafferty, J., McCallum, A. and Pereira, F.C. (2001), \u201cConditional random fields: probabilistic models for segmenting and labeling sequence data\u201d, Proceedings of the Eighteenth International Conference on Machine Learning, ICML '01, San Francisco, CA, USA, Morgan Kaufmann Publishers, San Francisco, CA, pp. 282\u2010289."},{"key":"key2022012419453998700_b14","doi-asserted-by":"crossref","unstructured":"Lesk, M. (1986), \u201cAutomatic sense disambiguation using machine readable dictionaries: how to tell a pine cone from an ice cream cone\u201d, Proceedings of the 5th Annual International Conference on Systems Documentation, ACM, pp. 24\u201026.","DOI":"10.1145\/318723.318728"},{"key":"key2022012419453998700_b15","doi-asserted-by":"crossref","unstructured":"McCallum, A. and Li, W. (2003), \u201cEarly results for named entity recognition with conditional random fields, feature induction and web\u2010enhanced lexicons\u201d, Proceedings of the Seventh Conference on Natural Language Learning at HLT\u2010NAACL 2003, Vol. 4, Association for Computational Linguistics, Stroudsburg, PA, pp. 188\u2010191.","DOI":"10.3115\/1119176.1119206"},{"key":"key2022012419453998700_b17","unstructured":"Medelyan, O., Witten, I.H. and Milne, D. (2008), \u201cTopic indexing with Wikipedia\u201d, Proceedings of the AAAI WikiAI Workshop."},{"key":"key2022012419453998700_b16","doi-asserted-by":"crossref","unstructured":"Mendes, P.N., Jakob, M., Garc\u00eda\u2010Silva, A. and Bizer, C. (2011), \u201cDbpedia spotlight: shedding light on the web of documents\u201d, Proceedings of the 7th International Conference on Semantic Systems, New York, NY, USA, ACM, pp. 1\u20108.","DOI":"10.1145\/2063518.2063519"},{"key":"key2022012419453998700_b18","doi-asserted-by":"crossref","unstructured":"Mihalcea, R. and Csomai, A. (2007), \u201cWikify!: linking documents to encyclopedic knowledge\u201d, Proceedings of the Sixteenth ACM Conference on Conference on Information and Knowledge Management, CIKM '07, New York, NY, USA, ACM, pp. 233\u2010242.","DOI":"10.1145\/1321440.1321475"},{"key":"key2022012419453998700_b19","doi-asserted-by":"crossref","unstructured":"Milne, D. and Witten, I.H. (2008), \u201cLearning to link with Wikipedia\u201d, Proceedings of the 17th ACM Conference on Information and Knowledge Management, CIKM'08, New York, NY, USA, ACM, pp. 509\u2010518.","DOI":"10.1145\/1458082.1458150"},{"key":"key2022012419453998700_b20","doi-asserted-by":"crossref","unstructured":"Ng, H.T. and Lee, H.B. (1996), \u201cIntegrating multiple knowledge sources to disambiguate word sense: an exemplar\u2010based approach\u201d, Proceedings of the 34th Annual Meeting on Association for Computational Linguistics, Association for Computational Linguistics, Stroudsburg, PA, pp. 40\u201047.","DOI":"10.3115\/981863.981869"},{"key":"key2022012419453998700_b21","unstructured":"Pasca, M., Lin, D., Bigham, J., Lifchits, A. and Jain, A. (2006), \u201cOrganizing and searching the world wide web of facts\u2010step one: the one\u2010million fact extraction challenge\u201d, Proceedings of the National Conference on Artificial Intelligence, Vol. 21 No. 2, p. 1400."},{"key":"key2022012419453998700_b22","unstructured":"Rizzo, G. and Troncy, R. (2011), \u201cNERD: a framework for evaluating named entity recognition tools in the web of data\u201d, 10th International Semantic Web Conference (ISWC'11), Demo Session, Bonn, Germany, pp. 1\u201016."},{"key":"key2022012419453998700_b23","unstructured":"Robertson, S.E. and Sparck, J.K. (1988), \u201cDocument retrieval systems\u201d, Relevance Weighting of Search Terms, Taylor Graham Publishing, London, pp. 143\u2010160."},{"key":"key2022012419453998700_b24","doi-asserted-by":"crossref","unstructured":"Rowe, M. (2009), \u201cApplying semantic social graphs to disambiguate identity references\u201d, The Semantic Web: Research and Applications, Lecture Notes in Computer Science, Springer, Berlin, pp. 461\u2010475.","DOI":"10.1007\/978-3-642-02121-3_35"},{"key":"key2022012419453998700_b25","unstructured":"Rusu, D., Dali, L., Fortuna, B., Grobelnik, M. and Mladenic, D. (2007), \u201cTriplet extraction from sentences\u201d, Proceedings of the 10th International Multiconference, Information Society\u2010IS, pp. 8\u201012."},{"key":"key2022012419453998700_b26","unstructured":"Sekine, S. (1998), \u201cDescription of the Japanese NE system used for MET\u20102\u201d, Proceedings of the Seventh Message Understanding Conference (MUC\u20107)."},{"key":"key2022012419453998700_b27","doi-asserted-by":"crossref","unstructured":"Small, H. (1973), \u201cCo\u2010citation in the scientific literature: a new measure of the relationship between two documents\u201d, Journal of the American Society for Information Science, Vol. 24 No. 4, pp. 265\u2010269.","DOI":"10.1002\/asi.4630240406"},{"key":"key2022012419453998700_b28","doi-asserted-by":"crossref","unstructured":"Sparck, J.K., Walker, S. and Robertson, S.E. (2000), \u201cA probabilistic model of information retrieval: development and comparative experiments: part 1\u201d, Information Processing & Management, Vol. 36 No. 6, pp. 779\u2010808.","DOI":"10.1016\/S0306-4573(00)00015-7"},{"key":"key2022012419453998700_b29","doi-asserted-by":"crossref","unstructured":"Tjong Kim Sang, E.F. and De Meulder, F. (2003), \u201cIntroduction to the CoNLL\u20102003 shared task: language\u2010independent named entity recognition\u201d, Proceedings of the Seventh Conference on Natural Language Learning at HLT\u2010NAACL 2003, Vol. 4, Association for Computational Linguistics, Stroudsburg, PA, pp. 142\u2010147.","DOI":"10.3115\/1119176.1119195"},{"key":"key2022012419453998700_b30","doi-asserted-by":"crossref","unstructured":"Velasquez, J.D., Rios, S.A., Bassi, A., Yasuda, H. and Aoki, T. (2005), \u201cTowards the identification of keywords in the web site text content: a methodological approach\u201d, International Journal of Web Information Systems, Vol. 1 No. 1, pp. 53\u201057.","DOI":"10.1108\/17440080580000083"},{"key":"key2022012419453998700_b31","unstructured":"Volz, R., Kleb, J. and Mueller, W. (2007), \u201cTowards ontology\u2010based disambiguation of geographical identifiers\u201d, I3: Identity, Identifiers, Identification, pp. 8\u201012."}],"container-title":["International Journal of Web Information Systems"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/www.emeraldinsight.com\/doi\/full-xml\/10.1108\/IJWIS-05-2013-0016","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/IJWIS-05-2013-0016\/full\/xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/IJWIS-05-2013-0016\/full\/html","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T22:24:03Z","timestamp":1753395843000},"score":1,"resource":{"primary":{"URL":"http:\/\/www.emerald.com\/ijwis\/article\/9\/3\/242-259\/166932"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2013,8,23]]},"references-count":31,"journal-issue":{"issue":"3","published-print":{"date-parts":[[2013,8,23]]}},"alternative-id":["10.1108\/IJWIS-05-2013-0016"],"URL":"https:\/\/doi.org\/10.1108\/ijwis-05-2013-0016","relation":{},"ISSN":["1744-0084"],"issn-type":[{"type":"print","value":"1744-0084"}],"subject":[],"published":{"date-parts":[[2013,8,23]]}}}