{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,8,2]],"date-time":"2025-08-02T17:08:40Z","timestamp":1754154520883,"version":"3.41.2"},"reference-count":32,"publisher":"Emerald","issue":"1","license":[{"start":{"date-parts":[[2012,1,13]],"date-time":"2012-01-13T00:00:00Z","timestamp":1326412800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/www.emerald.com\/insight\/site-policies"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2012,1,13]]},"abstract":"<jats:sec><jats:title content-type=\"abstract-heading\">Purpose<\/jats:title><jats:p>This paper seeks to focus on the problems of integrating information from open, distributed scholarly collections, and on the opportunities these collections represent for research communities in developing countries. The paper aims to introduce OntOAIr, a semi\u2010automatic method for constructing lightweight ontologies of documents in repositories such as those provided by the Open Archives Initiative (OAI).<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Design\/methodology\/approach<\/jats:title><jats:p>OntOAIr uses simplified document representations, a clustering algorithm, and ontological engineering techniques.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Findings<\/jats:title><jats:p>The paper presents experimental results of the potential positive impact of ontologies and specifically of OntOAIr on the use of collections provided by OAI.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Research limitations\/implications<\/jats:title><jats:p>By applying OntOAIr, scholars who frequently spend many hours organizing OAI information spaces will obtain support that will allow them to speed up the entire research cycle and, expectedly, participate more fully in global research communities.<\/jats:p><\/jats:sec><jats:sec><jats:title content-type=\"abstract-heading\">Originality\/value<\/jats:title><jats:p>The proposed method allows human and software agents to organize and retrieve groups of documents from multiple collections. Applications of OntOAIr include enhanced document retrieval. In this paper, the authors focus particularly on document retrieval applications.<\/jats:p><\/jats:sec>","DOI":"10.1108\/00012531211196701","type":"journal-article","created":{"date-parts":[[2012,1,14]],"date-time":"2012-01-14T07:38:29Z","timestamp":1326526709000},"page":"46-66","source":"Crossref","is-referenced-by-count":2,"title":["Organizing open archives via lightweight ontologies to facilitate the use of heterogeneous collections"],"prefix":"10.1108","volume":"64","author":[{"given":"J.","family":"Alfredo S\u00e1nchez","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Mar\u00eda","family":"Auxilio Medina","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Oleg","family":"Starostenko","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Antonio","family":"Benitez","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Eduardo","family":"L\u00f3pez Dom\u00ednguez","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"140","reference":[{"unstructured":"Aitken, S. and Reid, S. (2000), \u201cEvaluation of an ontology\u2010based information retrieval tool\u201d, in G\u00f3mez, A., Benjamins, V.R., Guarino, N. and Uschold, M. (Eds), Workshop on the Applications of Ontologies and Problem\u2010Solving Methods. European Conference on Artificial Intelligence (ECAI'00), Berlin, September, pp. 34\u201043.","key":"key2022020920381364000_b1"},{"unstructured":"Berry, W.M. and Castellanos, M. (2010), Survey of Text Mining II: Clustering, Classification, and Retrieval, Springer\u2010Verlag, London.","key":"key2022020920381364000_b2"},{"unstructured":"Borst, W. (1997), Construction of Engineering Ontologies for Knowledge Sharing and Reuse, Centre for Telematica and Information Technology, University of Twente, Enschede, technical report.","key":"key2022020920381364000_b4"},{"doi-asserted-by":"crossref","unstructured":"Brase, J. and Nejdl, W. (2004), \u201cOntologies and metadata for elearning\u201d, in Staab, S. and Studer, R. (Eds), Handbook on Ontologies, 2nd ed., International Handbooks on Information Systems, Springer, Dordrecht, Heidelberg, London and New York, NY, pp. 555\u201074.","key":"key2022020920381364000_b3","DOI":"10.1007\/978-3-540-24750-0_28"},{"unstructured":"Cui, Z. and O'Brien, P. (2000), \u201cDomain ontology management environment\u201d, Proceedings of the 33rd Hawaii International Conference on System Sciences '00 (HICSS'00), Hawaii, January, pp. 8\u201015.","key":"key2022020920381364000_b5"},{"unstructured":"Diederich, J. and Balke, W. (2007), \u201cThe semantic growbag algorithm: automatically deriving categorization systems\u201d, Proceedings of the 11th European Conference on Research and Advanced Technology for Digital Libraries '07 (ECDL'07, Budapest, Hungary, September). Lecture Notes in Computer Science, Vol. 4675, Springer, Berlin, pp. 33\u201040.","key":"key2022020920381364000_b6"},{"unstructured":"doccluster (2007), Clustering C++ Library Documentation, available at: http:\/\/wikipedia\u2010clustering.speedblue.org\/download\/ClusteringDoc\/main.html (accessed November 9, 2009).","key":"key2022020920381364000_b7"},{"doi-asserted-by":"crossref","unstructured":"Fung, B.C.M., Wang, K. and Ester, M. (2006), Hierarchical document clustering, The Encyclopedia of Data Warehousing and Mining, Volume I, Idea Group Reference, Hershey, PA, London, Melbourne and Singapore, pp. 555\u20109, available at: www.cs.sfu.ca\/\u223cwangk\/pub\/FWE05dwm.pdf.","key":"key2022020920381364000_b9","DOI":"10.4018\/978-1-59140-557-3.ch105"},{"doi-asserted-by":"crossref","unstructured":"Giunchiglia, F., Marchese, M. and Zaihrayeu, I. (2006), \u201cEncoding classifications as lightweight ontologies\u201d, Journal of Data Semantic, Vol. VIII, Winter.","key":"key2022020920381364000_b10","DOI":"10.1007\/11762256_9"},{"unstructured":"G\u00f3mez, A., Fern\u00e1ndez, M. and Corcho, O. (2004), Ontological Engineering, Springer\u2010Verlag, London.","key":"key2022020920381364000_b11"},{"doi-asserted-by":"crossref","unstructured":"Gruber, T.R. (1993), \u201cA translation approach to portable ontology specification\u201d, Knowledge Acquisition, Vol. 5 No. 2, pp. 199\u2010220.","key":"key2022020920381364000_b13","DOI":"10.1006\/knac.1993.1008"},{"unstructured":"Guojun, G., Chaoqun, M. and Jianhong, W. (2007), Data Clustering: Theory, Algorithms, and Applications, ASA\u2010ISAM, Philadelphia, PA, ASA\u2010SIAM Series on Statistics and Applied Probability.","key":"key2022020920381364000_b12"},{"doi-asserted-by":"crossref","unstructured":"Halgamuge, S.K. and Wang, L.P. (2005), Classification and Clustering for Knowledge Discovery, Springer\u2010Verlag, Berlin and Heidelberg.","key":"key2022020920381364000_b16","DOI":"10.1007\/b98152"},{"doi-asserted-by":"crossref","unstructured":"Hamel, L. (2009), Knowledge Discovery with Support Vector Machines, John Wiley & Sons, Chichester, Wiley Series on Methods and Applications in Data Mining.","key":"key2022020920381364000_b17","DOI":"10.1002\/9780470503065"},{"unstructured":"Harrison, T.L., Elango, A., Bollen, J. and Nelson, M. (2004), Initial Experiences Re\u2010exporting Duplicate and Similarity Computation with an OAI\u2010PMH Aggregator, available at: http:\/\/arxiv.org\/abs\/cs\/0401001.","key":"key2022020920381364000_b14"},{"unstructured":"Jain, A.K. and Dubes, R.C. (1988), Algorithms for Clustering Data, Prentice Hall, Upper Saddle River, NJ.","key":"key2022020920381364000_b15"},{"doi-asserted-by":"crossref","unstructured":"Karouri, L., Aufaure, M.A. and Bennacer, N. (2006), \u201cContext based hierarchial clustering for the ontology learning\u201d, Proceedings of the IEEE\/WIC\/ACM International Conference on Web Intelligence (WI'06), Hong Kong, December, pp. 420\u20107.","key":"key2022020920381364000_b18","DOI":"10.1109\/WI.2006.55"},{"doi-asserted-by":"crossref","unstructured":"Lagoze, C. and van de Sompel, H. (2001), \u201cThe open archives initiative: building a low\u2010barrier interoperability framework\u201d, Proceedings of the Joint Conference on Digital Libraries (JCDL'01), Roanoke, VA, pp. 54\u201062.","key":"key2022020920381364000_b22","DOI":"10.1145\/379437.379449"},{"unstructured":"Lambert, M.S., Tennoe, T.M. and Henssonow, F.S. (2010), UPGMA, VDM Verlag Dr Mueller Ag & Co. KG, Saarbr\u00fccken.","key":"key2022020920381364000_b19"},{"unstructured":"Lassila, O. and McGuinness, D.L. (2001), \u201cThe role of frame\u2010based representation on the semantic web\u201d, Electronic Transactions on Artificial Intelligence, Vol. 6 No. 5, available at: www.ep.liu.se\/ea\/cis\/2001\/005\/cis01005.pdf.","key":"key2022020920381364000_b23"},{"unstructured":"Lewis, D.D. (1997), \u201cReuters\u201021578 text categorization test collection\u201d, Distribution 1.0, September 26, available at: http:\/\/kdd.ics.uci.edu\/databases\/reuters21578\/reuters21578.html (accessed June 28, 2008).","key":"key2022020920381364000_b20"},{"unstructured":"Ljubi\u010d, P., Lavra\u010d, N., Plisson, J., Mladeniae, D., Bollhalter, S. and Jermol, M. (2005), \u201cAutomated structuring of company competencies in virtual organizations\u201d, Proceedings of the Conference on Data Mining and Data Warehouses 2005 (SiKDD 2005), Ljubljana, October, pp. 190\u20103.","key":"key2022020920381364000_b21"},{"unstructured":"McCandless, M., Hatcher, E. and Gospodneti\u0107, O. (2010), Lucene in Action, 2nd ed., Manning Publications, Stamford, CT.","key":"key2022020920381364000_b24"},{"doi-asserted-by":"crossref","unstructured":"Medina, M.A. and S\u00e1nchez, J.A. (2008), \u201cOntOAIr: a method to construct lightweight ontologies from document collections\u201d, Proceedings of the Ninth Mexican International Conference on ComputerScience 2008 (ENC 08), Baja California, M\u00e9xico, October.","key":"key2022020920381364000_b26","DOI":"10.1109\/ENC.2008.35"},{"doi-asserted-by":"crossref","unstructured":"Medina, M.A., S\u00e1nchez, J.A., Ch\u00e1vez, A. and Benitez, A. (2004), \u201cDesigning ontological agents: an alternative to improve information retrieval in federated digital libraries\u201d, Proceedings of the Atlantic Web Intelligence Conference 2004 (AWIC'04, Canc\u00fan, M\u00e9xico, May). Advances in Web Intelligence, Lecture Notes in Computer Science. Vol. 3034, Springer, Berlin, pp. 155\u201063.","key":"key2022020920381364000_b25","DOI":"10.1007\/978-3-540-24681-7_18"},{"doi-asserted-by":"crossref","unstructured":"Navigli, R. and Velardi, P. (2004), \u201cLearning domain ontologies from document warehouses and dedicated web sites\u201d, Computational Linguistics, Vol. 30 No. 2, pp. 151\u201079.","key":"key2022020920381364000_b27","DOI":"10.1162\/089120104323093276"},{"unstructured":"Plisson, J., Mladeniae, D., Ljubi\u010d, P., Lavrac, N. and Grobelnik, M. (2005), \u201cUsing machine learning to structure the expertise of companies: analysis of the Yahoo! business data\u201d, Conference on Data Mining and Data Warehouses (SiKDD 2005 Proceedings). 7th International Multi\u2010conference on Information Society IS'05, pp. 186\u20109.","key":"key2022020920381364000_b28"},{"unstructured":"Ryszard, K.S. and McDaniel, B. (2010), Semantic Digital Libraries, Springer\u2010Verlag, Berlin and Heidelberg.","key":"key2022020920381364000_b29"},{"doi-asserted-by":"crossref","unstructured":"Witten, I.H. and Bainbridge, D. (2007), \u201cA retrospective look at Greenstone: lessons from the first decade\u201d, Proceedings of the 7th ACM\/IEEE\u2010CS Joint Conference on Digital Libraries (JCDL '07), Vancouver, BC, June 18\u201023, ACM, New York, NY, pp. 147\u201056.","key":"key2022020920381364000_b30","DOI":"10.1145\/1255175.1255204"},{"doi-asserted-by":"crossref","unstructured":"Xu, R. and Wunsch, C.D. II (2008), Clustering, Wiley\/IEEE Press, New York, NY.","key":"key2022020920381364000_b32","DOI":"10.1002\/9780470382776"},{"unstructured":"Fung, B.C.M., Wang, K. and Ester, M. (2005), \u201cHierarchical document clustering using frequent itemsets\u201d, Proceedings of the Third SIAM International Conference on Data Mining (SDM'03), San Francisco, CA, May, pp. 59\u201070.","key":"key2022020920381364000_frg1"},{"unstructured":"Witten, I.H., Bainbridge, D. and Nichols, D. (2009), How to Build a Digital Library, 2nd ed., Morgan Kaufmann Publishers, Amsterdam, Boston, MA, Heidelberg, London, New York, NY, Oxford, Paris, San Diego, CA, San Francisco, CA, Singapore, Sydney and Tokyo.","key":"key2022020920381364000_frg2"}],"container-title":["Aslib Proceedings"],"original-title":[],"language":"en","link":[{"URL":"http:\/\/www.emeraldinsight.com\/doi\/full-xml\/10.1108\/00012531211196701","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/00012531211196701\/full\/xml","content-type":"application\/xml","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.emerald.com\/insight\/content\/doi\/10.1108\/00012531211196701\/full\/html","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T11:37:16Z","timestamp":1753357036000},"score":1,"resource":{"primary":{"URL":"http:\/\/www.emerald.com\/ajim\/article\/64\/1\/46-66\/38837"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2012,1,13]]},"references-count":32,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2012,1,13]]}},"alternative-id":["10.1108\/00012531211196701"],"URL":"https:\/\/doi.org\/10.1108\/00012531211196701","relation":{},"ISSN":["0001-253X"],"issn-type":[{"type":"print","value":"0001-253X"}],"subject":[],"published":{"date-parts":[[2012,1,13]]}}}