{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,10,24]],"date-time":"2025-10-24T07:51:06Z","timestamp":1761292266763,"version":"3.32.0"},"reference-count":22,"publisher":"Oxford University Press (OUP)","issue":"18","content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2005,9,15]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>Motivation: Biological literature contains many abbreviations with one particular sense in each document. However, most abbreviations do not have a unique sense across the literature. Furthermore, many documents do not contain the long forms of the abbreviations. Resolving an abbreviation in a document consists of retrieving its sense in use. Abbreviation resolution improves accuracy of document retrieval engines and of information extraction systems.<\/jats:p><jats:p>Results: We combine an automatic analysis of Medline abstracts and linguistic methods to build a dictionary of abbreviation\/sense pairs. The dictionary is used for the resolution of abbreviations occurring with their long forms. Ambiguous global abbreviations are resolved using support vector machines that have been trained on the context of each instance of the abbreviation\/sense pairs, previously extracted for the dictionary set-up. The system disambiguates abbreviations with a precision of 98.9% for a recall of 98.2% (98.5% accuracy). This performance is superior in comparison with previously reported research work.<\/jats:p><jats:p>Availability: The abbreviation resolution module is available at http:\/\/www.ebi.ac.uk\/Rebholz\/software.html<\/jats:p><jats:p>Contact: \u00a0gaudan@ebi.ac.uk<\/jats:p>","DOI":"10.1093\/bioinformatics\/bti586","type":"journal-article","created":{"date-parts":[[2005,7,22]],"date-time":"2005-07-22T00:24:38Z","timestamp":1121991878000},"page":"3658-3664","source":"Crossref","is-referenced-by-count":69,"title":["Resolving abbreviations to their senses in Medline"],"prefix":"10.1093","volume":"21","author":[{"given":"S.","family":"Gaudan","sequence":"first","affiliation":[{"name":"European Bioinformatics Institute Wellcome Trust Genome Campus, Hinxton, Cambridge CB10 1SD, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"H.","family":"Kirsch","sequence":"additional","affiliation":[{"name":"European Bioinformatics Institute Wellcome Trust Genome Campus, Hinxton, Cambridge CB10 1SD, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"D.","family":"Rebholz-Schuhmann","sequence":"additional","affiliation":[{"name":"European Bioinformatics Institute Wellcome Trust Genome Campus, Hinxton, Cambridge CB10 1SD, UK"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"286","published-online":{"date-parts":[[2005,7,21]]},"reference":[{"key":"2023060912061601400_B1","doi-asserted-by":"crossref","unstructured":"Aho, A.V. and Corasick, J.M. 1975Efficient string matching: an aid to bibliographic search. Commun. ACM.18333\u2013340","DOI":"10.1145\/360825.360855"},{"key":"2023060912061601400_B2","unstructured":"Aronson, A. 2001Effective mapping of biomedical text to the UMLS meta-thesaurus: the MetaMap Program. Proc. AMIA Symp.200117\u201321"},{"key":"2023060912061601400_B3","unstructured":"Adar, E. 2004SaRAD: a Simple and Robust Abbreviation Dictionary. Bioinformatics20527\u2013533"},{"key":"2023060912061601400_B4","unstructured":"Chen, L., et al. 2005Gene name ambiguity of eukaryotic nomenclature. Bioinformatics21248\u2013256"},{"key":"2023060912061601400_B5","doi-asserted-by":"crossref","unstructured":"Dingare, S., Finkel, J., Manning, C., Nissim, M., Alex, B. 2004Exploring the boundaries: gene and protein identification in biomedical text. Proceedings of the BioCreative Workshop, Granada","DOI":"10.1186\/1471-2105-6-S1-S5"},{"key":"2023060912061601400_B6","doi-asserted-by":"crossref","unstructured":"Frantzi, K. and Ananiadou, S. 1999The C value domain independent method for multiword term extraction. JNLP6, pp. 145\u2013179","DOI":"10.5715\/jnlp.6.3_145"},{"key":"2023060912061601400_B7","unstructured":"Fred, H.L. and Cheng, T.O. 2003Acronymesis: the exploding misuse of acronyms. Tex. Heart Inst. J.30255\u2013257"},{"key":"2023060912061601400_B8","unstructured":"Hoffmann, R. and Valencia, A. 2003Life cycles of successful genes. Trends Genet.1979\u201381"},{"key":"2023060912061601400_B9","doi-asserted-by":"crossref","unstructured":"Joachims, T. 1997Text categorization with support vector machines: learning with many relevant features. Machine Learning: ECML-98, Tenth European Conference on Machine Learning , pp. 137\u2013142","DOI":"10.1007\/BFb0026683"},{"key":"2023060912061601400_B10","unstructured":"Liu, H., Aronson, A.R., Friedman, C. 2002A study of abbreviations in MEDLINE abstracts. Proc. AMIA Symp.2002464\u2013468"},{"key":"2023060912061601400_B11","unstructured":"Liu, H., Johnson, S.B., Friedman, C. 2002Automatic resolution of ambiguous terms based on machine learning and conceptual relations in the UMLS. J. Am. Med. Inform. Assoc.9621\u2013636"},{"key":"2023060912061601400_B12","doi-asserted-by":"crossref","unstructured":"Pakhomov, S. 2002Semi-Supervised maximum entropy based approach to acronym and abbreviation normalization in medical texts. Proceedings of the 40th Annual Meeting of the ACL , Philadelphia University of Pennsylvania, pp. 160\u2013167","DOI":"10.3115\/1073083.1073111"},{"key":"2023060912061601400_B13","unstructured":"Sentinel Event Alert. 2001Medication errors related to potentially dangerous abbreviations. JCAHO. Issue 23, September 2001"},{"key":"2023060912061601400_B14","doi-asserted-by":"crossref","unstructured":"Schwartz, A. and Hearst, M. 2003A simple algorithm for identifying abbreviation definitions in biomedical text. Proceedings of PSB'03Kauai 8, pp. , pp. 451\u2013462","DOI":"10.1142\/9789812776303_0042"},{"key":"2023060912061601400_B15","unstructured":"Word Sense Disambiguation, The Case for Combinations of Knowledge Sources Stevenson, M. 2002 CLSI Studies in Computational Linguistics. CLSI publications, Centre for the study of language and information, California"},{"key":"2023060912061601400_B16","doi-asserted-by":"crossref","unstructured":"Taghva, K. and Gilbreth, J. 1999Recognizing acronyms and their definitions. Int. J. Document Anal. Recogn.191\u2013198","DOI":"10.1007\/s100320050018"},{"key":"2023060912061601400_B17","doi-asserted-by":"crossref","unstructured":"Tsuruoka, Y. and Tsujii, J. 2003Probabilistic term variant generator for biomedical terms. Proceedings of the 26th ACM SIGIRToronto, Canada , pp. 167\u2013173","DOI":"10.1145\/860435.860467"},{"key":"2023060912061601400_B18","doi-asserted-by":"crossref","unstructured":"Wren, J.D., et al. 2005Biomedical term mapping databases. Nucleic Acid Res.33D289\u2013293","DOI":"10.1093\/nar\/gki137"},{"key":"2023060912061601400_B19","doi-asserted-by":"crossref","unstructured":"Yarowsky, D. 1995Unsupervised word sense disambiguation rivaling supervised methods. Proceedings of the 33rd Annual Meeting of the ACL , Massachusetts, USA Cambridge, pp. 189\u2013196","DOI":"10.3115\/981658.981684"},{"key":"2023060912061601400_B20","unstructured":"Yu, H. and Friedman, C. 2002Mapping abbreviations to full forms in biomedical articles. J. Am. Med. Inform. Assoc.9262\u2013272"},{"key":"2023060912061601400_B21","unstructured":"Yu, Z., Tsuruoka, Y., Tsujii, J. 2003Automatic resolution of ambiguous abbreviations in biomedical texts using support vector machines and one sense per discourse hypothesis. Proceedings of the SIGIR'03 , pp. 57\u201362"},{"key":"2023060912061601400_B22","doi-asserted-by":"crossref","unstructured":"Yoshida, M., Fukuda, K., Takagi, T. 2000PNAD-CSS: a workbench for constructing a protein name abbreviation dictionary. Bioinformatics16169\u2013175","DOI":"10.1093\/bioinformatics\/16.2.169"}],"container-title":["Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/21\/18\/3658\/50554859\/bioinformatics_21_18_3658.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article-pdf\/21\/18\/3658\/50554859\/bioinformatics_21_18_3658.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,1,2]],"date-time":"2025-01-02T16:08:16Z","timestamp":1735834096000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/bioinformatics\/article\/21\/18\/3658\/202307"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2005,7,21]]},"references-count":22,"journal-issue":{"issue":"18","published-print":{"date-parts":[[2005,9,15]]}},"URL":"https:\/\/doi.org\/10.1093\/bioinformatics\/bti586","relation":{},"ISSN":["1367-4811","1367-4803"],"issn-type":[{"type":"electronic","value":"1367-4811"},{"type":"print","value":"1367-4803"}],"subject":[],"published-other":{"date-parts":[[2005,9]]},"published":{"date-parts":[[2005,7,21]]}}}