{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,18]],"date-time":"2026-08-18T02:26:02Z","timestamp":1787019962117,"version":"3.56.0"},"reference-count":100,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2025,9,1]],"date-time":"2025-09-01T00:00:00Z","timestamp":1756684800000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2025,9,1]],"date-time":"2025-09-01T00:00:00Z","timestamp":1756684800000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Research Foundation\u2013 Flanders","award":["200021E_189458"],"award-info":[{"award-number":["200021E_189458"]}]},{"name":"Research Foundation\u2013 Flanders","award":["200021E_189458"],"award-info":[{"award-number":["200021E_189458"]}]},{"DOI":"10.13039\/501100001711","name":"Swiss National Science Foundation","doi-asserted-by":"crossref","award":["G094020N"],"award-info":[{"award-number":["G094020N"]}],"id":[{"id":"10.13039\/501100001711","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100001711","name":"Swiss National Science Foundation","doi-asserted-by":"crossref","award":["G094020N"],"award-info":[{"award-number":["G094020N"]}],"id":[{"id":"10.13039\/501100001711","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["BMC Bioinformatics"],"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:sec>\n                    <jats:title>Background<\/jats:title>\n                    <jats:p>Knowledge discovery in scientific literature is hindered by the increasing volume of publications and the scarcity of extensive annotated data. To tackle the challenge of information overload, it is essential to employ automated methods for knowledge extraction and processing. Finding the right balance between the level of supervision and the effectiveness of models poses a significant challenge. While supervised techniques generally result in better performance, they have the major drawback of demanding labeled data. This requirement is labor-intensive, time-consuming, and hinders scalability when exploring new domains.<\/jats:p>\n                  <\/jats:sec>\n                  <jats:sec>\n                    <jats:title>Methods and Results<\/jats:title>\n                    <jats:p>In this context, our study addresses the challenge of identifying semantic relationships between biomedical entities (e.g., diseases, proteins, medications) in unstructured text while minimizing dependency on supervision. We introduce a suite of unsupervised algorithms based on dependency trees and attention mechanisms and employ a range of pointwise binary classification methods. Transitioning from weakly supervised to fully unsupervised settings, we assess the methods\u2019 ability to learn from data with noisy labels. The evaluation on four biomedical benchmark datasets explores the effectiveness of the methods, demonstrating their potential to enable scalable knowledge discovery systems less reliant on annotated datasets.<\/jats:p>\n                  <\/jats:sec>\n                  <jats:sec>\n                    <jats:title>Conclusion<\/jats:title>\n                    <jats:p>Our approach tackles a central issue in knowledge discovery: balancing performance with minimal supervision which is crucial to adapting models to varied and changing domains. This study also investigates the use of pointwise binary classification techniques within a weakly supervised framework for knowledge discovery. By gradually decreasing supervision, we assess the robustness of these techniques in handling noisy labels, revealing their capability to shift from weakly supervised to entirely unsupervised scenarios. Comprehensive benchmarking offers insights into the effectiveness of these techniques, examining how unsupervised methods can reliably capture complex relationships in biomedical texts. These results suggest an encouraging direction toward scalable, adaptable knowledge discovery systems, representing progress in creating data-efficient methodologies for extracting useful insights when annotated data is limited.<\/jats:p>\n                  <\/jats:sec>","DOI":"10.1186\/s12859-025-06187-0","type":"journal-article","created":{"date-parts":[[2025,9,1]],"date-time":"2025-09-01T12:24:39Z","timestamp":1756729479000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":1,"title":["Reduction of supervision for biomedical knowledge discovery"],"prefix":"10.1186","volume":"26","author":[{"given":"Christos","family":"Theodoropoulos","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Andrei Catalin","family":"Coman","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"James","family":"Henderson","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Marie-Francine","family":"Moens","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,9,1]]},"reference":[{"key":"6187_CR1","doi-asserted-by":"crossref","unstructured":"K\u00fcbler S, McDonald R, Nivre J. Dependency parsing, 2009;11\u201320.","DOI":"10.1007\/978-3-031-02131-2_2"},{"key":"6187_CR2","doi-asserted-by":"publisher","first-page":"30046","DOI":"10.1073\/pnas.1907367117","volume":"117","author":"CD Manning","year":"2020","unstructured":"Manning CD, Clark K, Hewitt J, Khandelwal U, Levy O. Emergent linguistic structure in artificial neural networks trained by self-supervision. Proc Natl Acad Sci. 2020;117:30046\u201354.","journal-title":"Proc Natl Acad Sci"},{"key":"6187_CR3","unstructured":"Vaswani A, et\u00a0al. Attention is all you need. Adv Neural Inf Process Syst. 2017;30. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/3f5ee243547dee91fbd053c1c4a845aa-Paper.pdf."},{"key":"6187_CR4","unstructured":"Feng L, et\u00a0al. Pointwise binary classification with pairwise confidence comparisons. In: Meila M, Zhang T editors. Proceedings of the 38th international conference on machine learning, Vol. 139, Proceedings of machine learning research. p. 3252\u201362. PMLR; 2021. https:\/\/proceedings.mlr.press\/v139\/feng21d.html."},{"key":"6187_CR5","unstructured":"Proux D, Rechenmann F, Julliard L. A pragmatic information extraction strategy for gathering data on genetic interactions. In: Bourne PE, et\u00a0al. editors. Proceedings of the 8th international conference on intelligent systems for molecular biology. p.279\u201385. AAAI Press, San Diego; 2000. http:\/\/www.aaai.org\/Library\/ISMB\/2000\/ismb00-029.php."},{"key":"6187_CR6","doi-asserted-by":"publisher","first-page":"155","DOI":"10.1093\/bioinformatics\/17.2.155","volume":"17","author":"T Ono","year":"2001","unstructured":"Ono T, Hishigaki H, Tanigami A, Takagi T. Automated extraction of information on protein-protein interactions from the biological literature. Bioinformatics. 2001;17:155\u201361.","journal-title":"Bioinformatics"},{"key":"6187_CR7","doi-asserted-by":"crossref","unstructured":"Phuong TM, Lee D, Lee KH. Learning rules to extract protein interactions from biomedical text. In: Whang KY, Jeon J, Shim K, Srivastava J editors. Proceedings of the 7th Pacific-Asia conference on knowledge discovery and data mining, vol. 2637. Heidelberg: Springer; 2003. p. 148\u201358.","DOI":"10.1007\/3-540-36175-8_15"},{"key":"6187_CR8","doi-asserted-by":"publisher","first-page":"3604","DOI":"10.1093\/bioinformatics\/bth451","volume":"20","author":"M Huang","year":"2004","unstructured":"Huang M, et al. Discovering patterns to extract protein-protein interactions from full texts. Bioinformatics. 2004;20:3604\u201312.","journal-title":"Bioinformatics"},{"key":"6187_CR9","doi-asserted-by":"publisher","first-page":"3294","DOI":"10.1093\/bioinformatics\/bti493","volume":"21","author":"Y Hao","year":"2005","unstructured":"Hao Y, Zhu X, Huang M, Li M. Discovering patterns to extract protein-protein interactions from the literature: Part ii. Bioinformatics. 2005;21:3294\u2013300.","journal-title":"Bioinformatics"},{"key":"6187_CR10","doi-asserted-by":"crossref","unstructured":"Chun HW, Hwang YS, Rim HC. Unsupervised event extraction from biomedical literature using co-occurrence information and basic patterns. In: Su KY, Tsujii J, Lee JH, Kwong OY editors. Proceedings of the 1st international conference on natural language processing. Berlin Heidelberg: Springer; 2004. p. 777\u201386.","DOI":"10.1007\/978-3-540-30211-7_83"},{"key":"6187_CR11","doi-asserted-by":"crossref","unstructured":"Liu H, Blouin C, Ke\u0161elj V. Identifying interaction sentences from biological literature using automatically extracted patterns. In: Cohen KB et\u00a0al. editors. Proceedings of the BioNLP 2009 workshop. p. 133\u201341. Association for Computational Linguistics, Boulder, Colorado; 2009. https:\/\/aclanthology.org\/W09-1317.","DOI":"10.3115\/1572364.1572383"},{"key":"6187_CR12","doi-asserted-by":"publisher","first-page":"481","DOI":"10.1109\/TCBB.2010.51","volume":"7","author":"J Hakenberg","year":"2010","unstructured":"Hakenberg J, et al. Efficient extraction of protein-protein interactions from full-text articles. IEEE\/ACM Trans Comput Biol Bioinf. 2010;7:481\u201394.","journal-title":"IEEE\/ACM Trans Comput Biol Bioinf"},{"key":"6187_CR13","unstructured":"Thomas P, Pietschmann S, Solt I, Tikk D, Leser U. Not all links are equal: Exploiting dependency types for the extraction of protein-protein interactions from text. In: Cohen KB, et\u00a0al. editors. Proceedings of BioNLP 2011 workshop. p. 1\u20139. Association for Computational Linguistics, Portland, Oregon; 2011. https:\/\/aclanthology.org\/W11-0201."},{"key":"6187_CR14","first-page":"14","volume":"17","author":"C Blaschke","year":"2002","unstructured":"Blaschke C, Valencia A. The frame-based module of the SUISEKI information extraction system. IEEE Intell Syst. 2002;17:14\u201320.","journal-title":"IEEE Intell Syst"},{"key":"6187_CR15","doi-asserted-by":"crossref","unstructured":"Raja K, Subramani S, Natarajan J. PPInterFinder-a mining tool for extracting causal relations on human proteins from literature. Database. 2013;2013.","DOI":"10.1093\/database\/bas052"},{"key":"6187_CR16","unstructured":"Craven M, Kumlien J. Constructing biological knowledge bases by extracting information from text sources. In: Lengauer T, et al. editors. Proceedings of the 7th international conference on intelligent systems for molecular biology. p. 77\u201386. AAAI Press; 1999."},{"key":"6187_CR17","unstructured":"Bunescu R, Mooney R. Learning to extract relations from the web using minimal supervision. In: Zaenen A, van\u00a0den Bosch A editors. Proceedings of the 45th annual meeting of the association of computational linguistics. p. 576\u201383. Association for Computational Linguistics, Prague; 2007. https:\/\/aclanthology.org\/P07-1073."},{"key":"6187_CR18","unstructured":"Riedel S, Yao L, McCallum A. Modeling relations and their mentions without labeled text. In: Balc\u00e1zar JL, Bonchi F, Gionis A, Sebag M editors. Machine learning and knowledge discovery in databases: European conference, ECML PKDD 2010, Proceedings, Part III 21. New York: Springer; 2010. p. 148\u2013163."},{"key":"6187_CR19","unstructured":"Hoffmann R, Zhang C, Ling X, Zettlemoyer L, Weld DS. Knowledge-based weak supervision for information extraction of overlapping relations. In: Lin D, Matsumoto Y, Mihalcea R editors. Proceedings of the 49th annual meeting of the association for computational linguistics: human language technologies. p. 541\u2013550. Association for Computational Linguistics, Portland; 2011. https:\/\/aclanthology.org\/P11-1055."},{"key":"6187_CR20","doi-asserted-by":"crossref","unstructured":"Zeng D, Liu K, Chen Y, Zhao J. Distant supervision for relation extraction via piecewise convolutional neural networks. In: M\u00e0rquez L, Callison-Burch C, Su J editors. Proceedings of the 2015 conference on empirical methods in natural language processing. p. 1753\u201362. Association for Computational Linguistics, Lisbon; 2015. https:\/\/aclanthology.org\/D15-1203.","DOI":"10.18653\/v1\/D15-1203"},{"key":"6187_CR21","doi-asserted-by":"crossref","unstructured":"Lin Y, Shen S, Liu Z, Luan H, Sun M. Neural relation extraction with selective attention over instances. In: Erk K, Smith NA editors. Proceedings of the 54th annual meeting of the association for computational linguistics (Vol. 1 Long Papers). p. 430\u20139. Association for Computational Linguistics, Vancouver; 2017.","DOI":"10.18653\/v1\/P16-1200"},{"key":"6187_CR22","doi-asserted-by":"crossref","unstructured":"Luo B, et\u00a0al. Learning with noise: enhance distantly supervised relation extraction with dynamic transition matrix. In: Barzilay R, Kan MY editors. Proceedings of the 55th annual meeting of the association for computational linguistics (Vol. 1: Long Papers), p. 430\u20139. Association for Computational Linguistics, Vancouver; 2017. https:\/\/aclanthology.org\/P17-1040.","DOI":"10.18653\/v1\/P17-1040"},{"key":"6187_CR23","doi-asserted-by":"crossref","unstructured":"Han X, Liu Z, Sun M. Neural knowledge acquisition via mutual attention between knowledge graph and text. In: McIlraith SA, Weinberger KQ editors. Proceedings of the 32nd AAAI conference on artificial intelligence. vol. 32. 2018.","DOI":"10.1609\/aaai.v32i1.11927"},{"key":"6187_CR24","doi-asserted-by":"crossref","unstructured":"Alt C, H\u00fcbner M, Hennig L. Fine-tuning pre-trained transformer language models to distantly supervised relation extraction. In: Korhonen A, Traum D, M\u00e0rquez L editors. Proceedings of the 57th annual meeting of the association for computational linguistics. p. 1388\u201398. Association for Computational Linguistics, Florence; 2019. https:\/\/aclanthology.org\/P19-1134.","DOI":"10.18653\/v1\/P19-1134"},{"key":"6187_CR25","doi-asserted-by":"crossref","unstructured":"Dai Q, Inoue N, Reisert P, Takahashi R, Inui K. Distantly supervised biomedical knowledge acquisition via knowledge graph based attention. In: Nastase V, Roth B, Dietz L, McCallum A editors. Proceedings of the workshop on extracting structured knowledge from scientific publications. p. 1\u201310. Association for Computational Linguistics, Minneapolis; 2019. https:\/\/aclanthology.org\/W19-2601.","DOI":"10.18653\/v1\/W19-2601"},{"key":"6187_CR26","doi-asserted-by":"crossref","unstructured":"Amin S, Dunfield KA, Vechkaeva A, Neumann G. A data-driven approach for noise reduction in distantly supervised biomedical relation extraction. In: Demner-Fushman D, Cohen KB, Ananiadou S, Tsujii J editors. Proceedings of the 19th SIGBioMed workshop on biomedical language processing. p. 187\u201394. Association for Computational Linguistics; 2020. https:\/\/aclanthology.org\/2020.bionlp-1.20.","DOI":"10.18653\/v1\/2020.bionlp-1.20"},{"key":"6187_CR27","doi-asserted-by":"publisher","first-page":"1234","DOI":"10.1093\/bioinformatics\/btz682","volume":"36","author":"J Lee","year":"2020","unstructured":"Lee J, et al. BioBERT: a pre-trained biomedical language representation model for biomedical text mining. Bioinformatics. 2020;36:1234\u201340.","journal-title":"Bioinformatics"},{"key":"6187_CR28","doi-asserted-by":"crossref","unstructured":"Wu S, He Y. Enriching pre-trained language model with entity information for relation classification. In: Zhu W et\u00a0al. editors. Proceedings of the 28th ACM international conference on information and knowledge management, CIKM \u201919. p. 2361\u201364. Association for Computing Machinery, New York; 2019. https:\/\/doi.org\/10.1145\/3357384.3358119.","DOI":"10.1145\/3357384.3358119"},{"key":"6187_CR29","doi-asserted-by":"publisher","first-page":"D267","DOI":"10.1093\/nar\/gkh061","volume":"32","author":"O Bodenreider","year":"2004","unstructured":"Bodenreider O. The unified medical language system (umls): integrating biomedical terminology. Nucleic Acids Res. 2004;32:D267\u201370.","journal-title":"Nucleic Acids Res"},{"key":"6187_CR30","unstructured":"Hogan WP et\u00a0al. Abstractified multi-instance learning (AMIL) for biomedical relation extraction. 2021."},{"key":"6187_CR31","doi-asserted-by":"crossref","unstructured":"Beltagy I, Lo K, Cohan A. SciBERT: a pretrained language model for scientific text. In: Inui K, Jiang J, Ng V, Wan X editors. Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP). p. 3615\u20133620. Association for Computational Linguistics, Hong Kong; 2019. https:\/\/aclanthology.org\/D19-1371.","DOI":"10.18653\/v1\/D19-1371"},{"key":"6187_CR32","doi-asserted-by":"publisher","first-page":"1739","DOI":"10.1093\/bioinformatics\/btaa907","volume":"37","author":"M Asada","year":"2021","unstructured":"Asada M, Miwa M, Sasaki Y. Using drug descriptions and molecular structures for drug-drug interaction extraction from literature. Bioinformatics. 2021;37:1739\u201346.","journal-title":"Bioinformatics"},{"key":"6187_CR33","doi-asserted-by":"crossref","unstructured":"Yasunaga M, Leskovec J, Liang P. LinkBERT: Pretraining language models with document links. In: Muresan S, Nakov P, Villavicencio A editors. Proceedings of the 60th annual meeting of the association for computational linguistics. Vol. 1: Long Papers. p. 8003\u201316. Association for Computational Linguistics, Dublin; 2022. https:\/\/aclanthology.org\/2022.acl-long.551.","DOI":"10.18653\/v1\/2022.acl-long.551"},{"key":"6187_CR34","doi-asserted-by":"crossref","unstructured":"Wadhwa S, Amir S, Wallace B. Revisiting relation extraction in the era of large language models. In: Rogers A, Boyd-Graber J, Okazaki N editors. Proceedings of the 61st annual meeting of the association for computational linguistics. Vol. 1: Long Papers. p. 15566\u201315589. Association for Computational Linguistics, Toronto; 2023. https:\/\/aclanthology.org\/2023.acl-long.868.","DOI":"10.18653\/v1\/2023.acl-long.868"},{"key":"6187_CR35","unstructured":"Brown TB. Language models are few-shot learners. 2020."},{"key":"6187_CR36","first-page":"1","volume":"25","author":"HW Chung","year":"2024","unstructured":"Chung HW, et al. Scaling instruction-finetuned language models. J Mach Learn Res. 2024;25:1\u201353.","journal-title":"J Mach Learn Res"},{"key":"6187_CR37","doi-asserted-by":"publisher","DOI":"10.1016\/j.patcog.2024.110779","volume":"156","author":"C Gao","year":"2024","unstructured":"Gao C, et al. Few-shot relational triple extraction with hierarchical prototype optimization. Pattern Recogn. 2024;156: 110779.","journal-title":"Pattern Recogn"},{"key":"6187_CR38","first-page":"18661","volume":"33","author":"P Khosla","year":"2020","unstructured":"Khosla P, et al. Supervised contrastive learning. Adv Neural Inf Process Syst. 2020;33:18661\u201373.","journal-title":"Adv Neural Inf Process Syst"},{"key":"6187_CR39","doi-asserted-by":"crossref","unstructured":"Theodoropoulos C, Henderson J, Coman AC, Moens MF. Imposing relation structure in language-model embeddings using contrastive learning. In: Bisazza A, Abend O editors. Proceedings of the 25th conference on computational natural language learning. p. 337\u2013348. Association for Computational Linguistics; 2021. https:\/\/aclanthology.org\/2021.conll-1.27.","DOI":"10.18653\/v1\/2021.conll-1.27"},{"key":"6187_CR40","first-page":"857","volume":"35","author":"X Liu","year":"2021","unstructured":"Liu X, et al. Self-supervised learning: generative or contrastive. IEEE Trans Knowl Data Eng. 2021;35:857\u201376.","journal-title":"IEEE Trans Knowl Data Eng"},{"key":"6187_CR41","doi-asserted-by":"publisher","first-page":"btad557","DOI":"10.1093\/bioinformatics\/btad557","volume":"39","author":"Q Chen","year":"2023","unstructured":"Chen Q, et al. An extensive benchmark study on biomedical text generation and mining with ChatGPT. Bioinformatics. 2023;39:btad557.","journal-title":"Bioinformatics"},{"key":"6187_CR42","doi-asserted-by":"crossref","unstructured":"Asada M, Fukuda K. Enhancing relation extraction from biomedical texts by large language models. In: Degen H, Ntoa S editors. Proceedings of the 5th international conference of artificial intelligence in human-computer interaction, held as part of the 26th human-computer interaction international conference. p. 3\u201314. New York: Springer; 2024.","DOI":"10.1007\/978-3-031-60615-1_1"},{"key":"6187_CR43","first-page":"391","volume":"2024","author":"J Zhang","year":"2024","unstructured":"Zhang J, et al. A study of biomedical relation extraction using GPT models. AMIA Summits Transl Sci Proc. 2024;2024:391.","journal-title":"AMIA Summits Transl Sci Proc"},{"key":"6187_CR44","unstructured":"Achiam J, et\u00a0al. Gpt-4 technical report 2023."},{"key":"6187_CR45","doi-asserted-by":"crossref","unstructured":"Zhang K, Jimenez\u00a0Gutierrez B, Su Y, Rogers A. Aligning instruction tasks unlocks large language models as zero-shot relation extractors. In: Rogers A, Boyd-Graber J, Okazaki N editors. Findings of the association for computational linguistics: ACL 2023. p. 794\u2013812. Association for Computational Linguistics, Toronto; 2023. https:\/\/aclanthology.org\/2023.findings-acl.50.","DOI":"10.18653\/v1\/2023.findings-acl.50"},{"key":"6187_CR46","doi-asserted-by":"crossref","unstructured":"Ji Z, et\u00a0al. Survey of hallucination in natural language generation. ACM Comput Surveys. 2023. https:\/\/doi.org\/10.1145\/3571730.","DOI":"10.1145\/3571730"},{"key":"6187_CR47","doi-asserted-by":"publisher","first-page":"1930","DOI":"10.1038\/s41591-023-02448-8","volume":"29","author":"AJ Thirunavukarasu","year":"2023","unstructured":"Thirunavukarasu AJ, et al. Large language models in medicine. Nat Med. 2023;29:1930\u201340.","journal-title":"Nat Med"},{"key":"6187_CR48","doi-asserted-by":"crossref","unstructured":"Ceballos-Arroyo AM, et\u00a0al. Open (clinical) LLMs are sensitive to instruction phrasings. In: Demner-Fushman D, Ananiadou S, Miwa M, Roberts K, Tsujii J editors. Proceedings of the 23rd workshop on biomedical natural language processing. p. 50\u201371. Association for Computational Linguistics, Bangkok; 2024. https:\/\/aclanthology.org\/2024.bionlp-1.5.","DOI":"10.18653\/v1\/2024.bionlp-1.5"},{"key":"6187_CR49","doi-asserted-by":"crossref","unstructured":"Petroni F, et\u00a0al. Language models as knowledge bases? In: Inui K, Jiang J, Ng V, Wan X editors. Proceedings of the 2019 conference on empirical methods in natural language processing and the 9th international joint conference on natural language processing (EMNLP-IJCNLP). p. 2463\u20132473. Association for Computational Linguistics, Hong Kong; 2019.","DOI":"10.18653\/v1\/D19-1250"},{"key":"6187_CR50","volume-title":"Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition","author":"D Jurafsky","year":"2000","unstructured":"Jurafsky D, Martin JH. Speech and Language Processing: An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition. 1st ed. USA: Prentice Hall PTR; 2000.","edition":"1"},{"key":"6187_CR51","doi-asserted-by":"crossref","unstructured":"Bunescu R, Mooney R. A shortest path dependency kernel for relation extraction. In: Mooney R, Brew C, Chien LF, Kirchhoff K editors. Proceedings of human language technology conference and conference on empirical methods in natural language processing. p. 724\u2013731. Association for Computational Linguistics, Vancouver, British Columbia; 2005. https:\/\/aclanthology.org\/H05-1091.","DOI":"10.3115\/1220575.1220666"},{"key":"6187_CR52","doi-asserted-by":"crossref","unstructured":"Purpura A, Bonin F, Bettencourt-silva J. Accelerating the discovery of semantic associations from medical literature: mining relations between diseases and symptoms. In: Li Y, Lazaridou A editors. Proceedings of the 2022 conference on empirical methods in natural language processing: industry track. p. 77\u201389. Association for computational linguistics, Abu Dhabi; 2022. https:\/\/aclanthology.org\/2022.emnlp-industry.6.","DOI":"10.18653\/v1\/2022.emnlp-industry.6"},{"key":"6187_CR53","doi-asserted-by":"crossref","unstructured":"Toutanvoa K, Manning CD. Enriching the knowledge sources used in a maximum entropy part-of-speech tagger. In: Toutanvoa K, Manning CD editors. 2000 Joint SIGDAT Conference on Empirical Methods in Natural Language Processing and Very Large Corpora. Association for Computational Linguistics, Hong Kong; 2000.","DOI":"10.3115\/1117794.1117802"},{"key":"6187_CR54","doi-asserted-by":"crossref","unstructured":"Toutanova K, Klein D, Manning CD, Singer Y. Feature-rich part-of-speech tagging with a cyclic dependency network. In: Toutanova K, Klein D, Manning CD, Singer Y editors. Proceedings of the 2003 Conference of the North American Chapter of the Association for Computational Linguistics on Human Language Technology - Volume 1, NAACL \u201903. p. 173\u201380. Association for Computational Linguistics; 2003. https:\/\/doi.org\/10.3115\/1073445.1073478.","DOI":"10.3115\/1073445.1073478"},{"key":"6187_CR55","unstructured":"Dozat T, Manning CD. Deep biaffine attention for neural dependency parsing. In: Dozat T, Manning CD editors. International conference on learning representations. 2017. https:\/\/openreview.net\/forum?id=Hk95PK9le."},{"key":"6187_CR56","doi-asserted-by":"crossref","unstructured":"Neumann M, King D, Beltagy I, Ammar W. ScispaCy: fast and robust models for biomedical natural language processing. In: Demner-Fushman D, Cohen KB, Ananiadou S, Tsujii J editors. Proceedings of the 18th BioNLP workshop and shared task. p. 319\u2013327. Association for Computational Linguistics, Florence; 2019. https:\/\/aclanthology.org\/W19-5034.","DOI":"10.18653\/v1\/W19-5034"},{"key":"6187_CR57","unstructured":"Akbik A, et\u00a0al. FLAIR: an easy-to-use framework for state-of-the-art NLP. In: Ammar W, Louis A, Mostafazadeh N editors. Proceedings of the 2019 conference of the North American chapter of the association for computational linguistics (Demonstrations). p. 54\u20139. Association for Computational Linguistics, Minneapolis; 2019. https:\/\/aclanthology.org\/N19-4010."},{"key":"6187_CR58","doi-asserted-by":"crossref","unstructured":"Amini A, Liu T, Cotterell R. Hexatagging: Projective dependency parsing as tagging. In: Rogers A, Boyd-Graber J, Okazaki N editors. Proceedings of the 61st annual meeting of the association for computational linguistics. Vol. 2: Short Papers. p. 1453\u201364. Association for Computational Linguistics, Toronto; 2023. https:\/\/aclanthology.org\/2023.acl-short.124.","DOI":"10.18653\/v1\/2023.acl-short.124"},{"key":"6187_CR59","doi-asserted-by":"crossref","unstructured":"Zhou W, Huang K, Ma T, Huang J. Document-level relation extraction with adaptive thresholding and localized context pooling. 2021.","DOI":"10.1609\/aaai.v35i16.17717"},{"key":"6187_CR60","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3674501","volume":"56","author":"X Zhao","year":"2024","unstructured":"Zhao X, et al. A comprehensive survey on relation extraction: recent advances and new frontiers. ACM Comput Surv. 2024;56:1\u201339.","journal-title":"ACM Comput Surv"},{"key":"6187_CR61","doi-asserted-by":"crossref","unstructured":"Coman A, Theodoropoulos C, Moens M-F, Henderson J. GADePo: graph-assisted declarative pooling transformers for document-level relation extraction. In: Yu W, et\u00a0al. editors. Proceedings of the 3rd workshop on knowledge augmented methods for NLP. p. 1\u201314. Association for Computational Linguistics, Bangkok; 2024. https:\/\/aclanthology.org\/2024.knowledgenlp-1.1.","DOI":"10.18653\/v1\/2024.knowledgenlp-1.1"},{"key":"6187_CR62","doi-asserted-by":"crossref","unstructured":"Horn RA, Johnson CR. Matrix analysis. Cambridge University Press; 2012.","DOI":"10.1017\/CBO9781139020411"},{"key":"6187_CR63","doi-asserted-by":"publisher","first-page":"79","DOI":"10.1214\/aoms\/1177729694","volume":"22","author":"S Kullback","year":"1951","unstructured":"Kullback S, Leibler RA. On information and sufficiency. Ann Math Stat. 1951;22:79\u201386.","journal-title":"Ann Math Stat"},{"key":"6187_CR64","unstructured":"Kullback S. Information theory and statistics. Courier Corporation; 1997."},{"key":"6187_CR65","unstructured":"Lu N, Niu G, Menon AK, Sugiyama M. On the minimal supervision for training any binary classifier from only unlabeled data. 2019."},{"key":"6187_CR66","unstructured":"Lu N, Zhang T, Niu G, Sugiyama M. Mitigating overfitting in supervised classification from two unlabeled datasets: a consistent risk correction approach. In: Chiappa S, Calandra R. Proceedings of the 23rd international conference on artificial intelligence and statistics. Vol. 108 of Proceedings of machine learning research. p. 1115\u201325. PMLR; 2020. https:\/\/proceedings.mlr.press\/v108\/lu20c.html."},{"key":"6187_CR67","unstructured":"Xu Y, Zhang H, Miller K, Singh A, Dubrawski A. Noise-tolerant interactive learning using pairwise comparisons. In: Guyon I, et\u00a0al. editors. Advances in neural information processing systems. 30 (Curran Associates, Inc.; 2017. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/e11943a6031a0e6114ae69c257617980-Paper.pdf."},{"key":"6187_CR68","unstructured":"Xu L, Honda J, Niu G, Sugiyama M. Uncoupled regression from pairwise comparison data. In: Wallach H, et\u00a0al. editors. Advances in neural information processing systems. vol. 32. Curran Associates, Inc.; 2019. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/6832a7b24bc06775d02b7406880b93fc-Paper.pdf."},{"key":"6187_CR69","doi-asserted-by":"publisher","first-page":"659","DOI":"10.1162\/neco_a_01262","volume":"32","author":"Z Cui","year":"2020","unstructured":"Cui Z, Charoenphakdee N, Sato I, Sugiyama M. Classification from triplet comparison data. Neural Comput. 2020;32:659\u201381.","journal-title":"Neural Comput"},{"key":"6187_CR70","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1021\/ci0342472","volume":"44","author":"DM Hawkins","year":"2004","unstructured":"Hawkins DM. The problem of overfitting. J Chem Inf Comput Sci. 2004;44:1\u201312.","journal-title":"J Chem Inf Comput Sci"},{"key":"6187_CR71","unstructured":"Nair V, Hinton GE. Rectified linear units improve restricted boltzmann machines. In: F\u00fcrnkranz, J, Joachims T editors. Proceedings of the 27th international conference on machine learning. p. 807\u2013814. 2010."},{"key":"6187_CR72","unstructured":"Natarajan N, Dhillon IS, Ravikumar PK, Tewari A. Learning with noisy labels. In: Burges C, Bottou L, Welling M, Ghahramani Z, Weinberger K editors. Advances in neural information processing systems. vol. 26. Curran Associates, Inc.; 2013. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2013\/file\/3871bd64012152bfb53fdf04b401193f-Paper.pdf."},{"key":"6187_CR73","unstructured":"Northcutt CG, Wu T, Chuang IL. Learning with confident examples: rank pruning for robust classification with noisy labels. In: Zhalama JZ, Eberhardt F, Mayer W editors. Proceedings of the 33rd conference on uncertainty in artificial intelligence (UAI). 2017. https:\/\/auai.org\/uai2017\/proceedings\/papers\/35.pdf."},{"key":"6187_CR74","unstructured":"Laine S, Aila T. Temporal ensembling for semi-supervised learning. 2017. https:\/\/openreview.net\/forum?id=BJ6oOfqge."},{"key":"6187_CR75","unstructured":"Tarvainen A, Valpola H. Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results. In: Guyon I, et\u00a0al. editors. Advances in neural information processing systems. vol. 30. Curran Associates, Inc.; 2017. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2017\/file\/68053af2923e00204c3ca7c6a3150cf7-Paper.pdf."},{"key":"6187_CR76","unstructured":"Menon A, Van\u00a0Rooyen B, Ong CS, Williamson B. Learning from corrupted binary labels via class-probability estimation. In: Bach F, Blei D editors. Proceedings of the 32nd international conference on machine learning. p. 125\u201334. PMLR; 2015."},{"key":"6187_CR77","doi-asserted-by":"publisher","first-page":"180652","DOI":"10.1109\/ACCESS.2024.3509714","volume":"12","author":"C Theodoropoulos","year":"2024","unstructured":"Theodoropoulos C, Coman AC, Henderson J, Moens M-F. Enhancing biomedical knowledge discovery for diseases: an open-source framework applied on RETT syndrome and alzheimer\u2019s disease. IEEE Access. 2024;12:180652\u201373.","journal-title":"IEEE Access"},{"key":"6187_CR78","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1186\/s12859-015-0472-9","volume":"16","author":"\u00c0 Bravo","year":"2015","unstructured":"Bravo \u00c0, Pi\u00f1ero J, Queralt-Rosinach N, Rautschka M, Furlong LI. Extraction of relations between genes and diseases from text and large-scale data analysis: implications for translational research. BMC Bioinf. 2015;16:1\u201317.","journal-title":"BMC Bioinf"},{"key":"6187_CR79","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1186\/1471-2105-8-50","volume":"8","author":"S Pyysalo","year":"2007","unstructured":"Pyysalo S, et al. BioInfer: a corpus for information extraction in the biomedical domain. BMC Bioinf. 2007;8:1\u201324.","journal-title":"BMC Bioinf"},{"key":"6187_CR80","doi-asserted-by":"publisher","first-page":"5","DOI":"10.1186\/s13643-023-02169-6","volume":"12","author":"U Petriti","year":"2023","unstructured":"Petriti U, Dudman DC, Scosyrev E, Lopez-Leon S. Global prevalence of RETT syndrome: systematic review and meta-analysis. Syst Rev. 2023;12:5.","journal-title":"Syst Rev"},{"key":"6187_CR81","doi-asserted-by":"publisher","first-page":"1577","DOI":"10.1016\/S0140-6736(20)32205-4","volume":"397","author":"P Scheltens","year":"2021","unstructured":"Scheltens P, et al. Alzheimer\u2019s disease. Lancet. 2021;397:1577\u201390.","journal-title":"Lancet"},{"key":"6187_CR82","doi-asserted-by":"publisher","first-page":"173","DOI":"10.1007\/s13311-021-01146-y","volume":"19","author":"JA Trejo-Lopez","year":"2023","unstructured":"Trejo-Lopez JA, Yachnis AT, Prokop S. Neuropathology of Alzheimer\u2019s disease. Neurotherapeutics. 2023;19:173\u201385.","journal-title":"Neurotherapeutics"},{"key":"6187_CR83","doi-asserted-by":"crossref","unstructured":"Lu Z. Pubmed and beyond: a survey of web tools for searching biomedical literature. Database. 2011;2011, baq036.","DOI":"10.1093\/database\/baq036"},{"key":"6187_CR84","first-page":"1","volume":"3","author":"Y Gu","year":"2021","unstructured":"Gu Y, et al. Domain-specific language model pretraining for biomedical natural language processing. ACM Trans Comput Healthcare (HEALTH). 2021;3:1\u201323.","journal-title":"ACM Trans Comput Healthcare (HEALTH)"},{"key":"6187_CR85","doi-asserted-by":"crossref","unstructured":"Hagberg A, Swart PJ, Schult DA. Exploring network structure, dynamics, and function using networkX. In: Varoquaux G, Vaught T, Millman J editors. Proceedings of the 7th python in science conference. p. 11\u20135. Pasadena; 2008.","DOI":"10.25080\/TCWV9851"},{"key":"6187_CR86","doi-asserted-by":"crossref","unstructured":"Tinn R, et\u00a0al. Fine-tuning large neural language models for biomedical natural language processing. Patterns. 2023;4.","DOI":"10.1016\/j.patter.2023.100729"},{"key":"6187_CR87","unstructured":"Wolf T, et\u00a0al. Transformers: state-of-the-art natural language processing. In: Liu Q, Schlangen D editors. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations. p. 38\u201345. Association for Computational Linguistics, Online; 2020. https:\/\/aclanthology.org\/2020.emnlp-demos.6."},{"key":"6187_CR88","first-page":"1929","volume":"15","author":"N Srivastava","year":"2014","unstructured":"Srivastava N, Hinton G, Krizhevsky A, Sutskever I, Salakhutdinov R. Dropout: a simple way to prevent neural networks from overfitting. J Mach Learn Res. 2014;15:1929\u201358.","journal-title":"J Mach Learn Res"},{"key":"6187_CR89","unstructured":"Ioffe S, Szegedy C. Batch normalization: accelerating deep network training by reducing internal covariate shift. In: Bach F, Blei D editors. Proceedings of the 32nd international conference on machine learning. Vol.\u00a037 of Proceedings of machine learning research. p. 448\u201356. PMLR, Lille; 2015. https:\/\/proceedings.mlr.press\/v37\/ioffe15.html."},{"key":"6187_CR90","unstructured":"Kingma DP, Ba J. Adam: A method for stochastic optimization; 2014."},{"key":"6187_CR91","unstructured":"Paszke A, et\u00a0al. Pytorch: an imperative style, high-performance deep learning library. In: Wallach H et\u00a0al. Advances in neural information processing systems. vol. 32. Curran Associates, Inc.; 2019. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2019\/file\/bdbca288fee7f92f2bfa9f7012727740-Paper.pdf."},{"key":"6187_CR92","doi-asserted-by":"crossref","unstructured":"Yuan Z, Liu Y, Tan C, Huang S, Huang F. Improving biomedical pretrained language models with knowledge. In: Demner-Fushman D, Cohen KB, Ananiadou S, Tsujii J editors. Proceedings of the 20th workshop on biomedical language processing. p. 180\u201390. Association for Computational Linguistics, Online; 2021. https:\/\/aclanthology.org\/2021.bionlp-1.20.","DOI":"10.18653\/v1\/2021.bionlp-1.20"},{"key":"6187_CR93","doi-asserted-by":"crossref","unstructured":"Park G, McCorkle S, Soto C, Blaby I, Yoo S. Extracting protein-protein interactions (PPIs) from biomedical literature using attention-based relational context information. In: Tsumoto S et\u00a0al. editors. Proceedings of the 10th IEEE international conference on big data (big data), p. 2052\u201361. IEEE; 2022.","DOI":"10.1109\/BigData55660.2022.10021099"},{"key":"6187_CR94","doi-asserted-by":"crossref","unstructured":"Labrak Y, et\u00a0al. BioMistral: a collection of open-source pretrained large language models for medical domains. In: Ku LW, Martins A, Srikumar V editors. Findings of the association for computational linguistics ACL 2024. p. 5848\u201364. Association for computational linguistics, Bangkok, Thailand and virtual meeting; 2024. https:\/\/aclanthology.org\/2024.findings-acl.348.","DOI":"10.18653\/v1\/2024.findings-acl.348"},{"key":"6187_CR95","unstructured":"Jiang AQ, et\u00a0al. Mistral 7b 2023."},{"key":"6187_CR96","doi-asserted-by":"publisher","first-page":"245","DOI":"10.1145\/325165.325242","volume":"19","author":"K Shoemake","year":"1985","unstructured":"Shoemake K. Animating rotation with quaternion curves. ACM SIGGRAPH Comput Graph. 1985;19:245\u201354. https:\/\/doi.org\/10.1145\/325165.325242.","journal-title":"ACM SIGGRAPH Comput Graph"},{"key":"6187_CR97","unstructured":"Yadav P, Tam D, Choshen L, Raffel CA, Bansal M. Ties-merging: resolving interference when merging models. In: Oh A, et\u00a0al. editors. Advances in neural information processing systems. vol. 36, p. 7093\u2013115. Curran Associates, Inc.; 2023. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2023\/file\/1644c9af28ab7916874f6fd6228a9bcf-Paper-Conference.pdf."},{"key":"6187_CR98","unstructured":"Yu L, Yu B, Yu H, Huang F, Li Y. Language models are super mario: Absorbing abilities from homologous models as a free lunch. In: Salakhutdinov R, et\u00a0al. editors. Proceedings of the 41st international conference on machine learning. 2024. https:\/\/openreview.net\/forum?id=fq0NaiU8Ex."},{"key":"6187_CR99","doi-asserted-by":"crossref","unstructured":"Pal A, Umapathi LK, Sankarasubbu M. Med-HALT: Medical domain hallucination test for large language models. In: Jiang J, Reitter D, Deng S editors. Proceedings of the 27th conference on computational natural language learning (CoNLL). p. 314\u201334. Association for Computational Linguistics, Singapore; 2023. https:\/\/aclanthology.org\/2023.conll-1.21.","DOI":"10.18653\/v1\/2023.conll-1.21"},{"key":"6187_CR100","doi-asserted-by":"publisher","unstructured":"Huang L, et\u00a0al. A survey on hallucination in large language models: principles, taxonomy, challenges, and open questions. ACM Trans Inf Syst. 2025;43. https:\/\/doi.org\/10.1145\/3703155.","DOI":"10.1145\/3703155"}],"container-title":["BMC Bioinformatics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s12859-025-06187-0.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1186\/s12859-025-06187-0\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1186\/s12859-025-06187-0.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,9,10]],"date-time":"2025-09-10T03:31:34Z","timestamp":1757475094000},"score":1,"resource":{"primary":{"URL":"https:\/\/bmcbioinformatics.biomedcentral.com\/articles\/10.1186\/s12859-025-06187-0"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,9,1]]},"references-count":100,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2025,12]]}},"alternative-id":["6187"],"URL":"https:\/\/doi.org\/10.1186\/s12859-025-06187-0","relation":{},"ISSN":["1471-2105"],"issn-type":[{"value":"1471-2105","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,9,1]]},"assertion":[{"value":"26 December 2024","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"11 June 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"1 September 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"Not applicable.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"Not applicable.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The authors declare that they have no competing interests.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"225"}}