{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,4,28]],"date-time":"2026-04-28T14:01:43Z","timestamp":1777384903779,"version":"3.51.4"},"reference-count":46,"publisher":"Association for Computing Machinery (ACM)","issue":"10","license":[{"start":{"date-parts":[[2024,10,23]],"date-time":"2024-10-23T00:00:00Z","timestamp":1729641600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"Key Program of the National Natural Science Foundation of China","award":["U23A20316"],"award-info":[{"award-number":["U23A20316"]}]},{"name":"Key R&D Project of Hubei Province","award":["2021BAA029"],"award-info":[{"award-number":["2021BAA029"]}]},{"name":"Major Science and Technology Project of Yunnan Province","award":["202102AA100021"],"award-info":[{"award-number":["202102AA100021"]}]},{"DOI":"10.13039\/501100001809","name":"National Natural Science Foundation of China","doi-asserted-by":"crossref","award":["62306284"],"award-info":[{"award-number":["62306284"]}],"id":[{"id":"10.13039\/501100001809","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100002858","name":"China Postdoctoral Science Foundation","doi-asserted-by":"crossref","award":["2023M743189"],"award-info":[{"award-number":["2023M743189"]}],"id":[{"id":"10.13039\/501100002858","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100006407","name":"Natural Science Foundation of Henan Province","doi-asserted-by":"crossref","award":["232300421386"],"award-info":[{"award-number":["232300421386"]}],"id":[{"id":"10.13039\/501100006407","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["ACM Trans. Asian Low-Resour. Lang. Inf. Process."],"published-print":{"date-parts":[[2024,10,31]]},"abstract":"<jats:p>The Biomedical Entity Normalization (BEN) task aims to align raw, unstructured medical entities to standard entities, thus promoting data coherence and facilitating better downstream medical applications. Recently, prompt learning methods have shown promising results in the natural language processing field. However, existing research falls short in tackling the more complex Chinese BEN task, especially in the few-shot scenario with limited medical data, and the vast potential of the external medical knowledge base has not yet been fully exploited. To address these challenges, this article proposes a novel Knowledge-injected Prompt Learning (PL-Knowledge) method. Specifically, the approach consists of five stages: candidate entity matching, knowledge extraction, knowledge encoding, knowledge injection, and prediction output. By effectively encoding the knowledge items contained in medical entities and incorporating them into tailor-made knowledge-injected templates, the additional knowledge enhances the model\u2019s ability to capture latent relationships between medical entities, thus achieving a better match with the standard entities. Comprehensive experiments are conducted on a benchmark dataset in both few-shot and full-scale settings. This method outperforms existing baselines, with an average accuracy improvement of 12.96 percentage points in few-shot and 0.94 percentage points in full-data cases, showcasing its excellence in the BEN task.<\/jats:p>","DOI":"10.1145\/3689629","type":"journal-article","created":{"date-parts":[[2024,8,23]],"date-time":"2024-08-23T12:24:42Z","timestamp":1724415882000},"page":"1-21","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":3,"title":["Knowledge-injected Prompt Learning for Chinese Biomedical Entity Normalization"],"prefix":"10.1145","volume":"23","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-7301-4425","authenticated-orcid":false,"given":"Songhua","family":"Yang","sequence":"first","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-3555-1553","authenticated-orcid":false,"given":"Chenghao","family":"Zhang","sequence":"additional","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-9169-9553","authenticated-orcid":false,"given":"Chenyuan","family":"He","sequence":"additional","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8397-1459","authenticated-orcid":false,"given":"Hongfei","family":"Xu","sequence":"additional","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8874-1262","authenticated-orcid":false,"given":"Hongying","family":"Zan","sequence":"additional","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-0481-0740","authenticated-orcid":false,"given":"Yuxiang","family":"Jia","sequence":"additional","affiliation":[{"name":"Zhengzhou University, Zhengzhou, China"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2024,10,23]]},"reference":[{"key":"e_1_3_3_2_2","first-page":"72","article-title":"Publicly available clinical BERT embeddings","author":"Alsentzer Emily","year":"2019","unstructured":"Emily Alsentzer, John R. Murphy, Willie Boag, Wei-Hung Weng, Di Jin, Tristan Naumann, and Matthew B. A. McDermott. 2019. Publicly available clinical BERT embeddings. In Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201919) (2019), 72.","journal-title":"Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201919)"},{"key":"e_1_3_3_3_2","first-page":"17","volume-title":"Proceedings of the AMIA Symposium","author":"Aronson Alan R.","year":"2001","unstructured":"Alan R. Aronson. 2001. Effective mapping of biomedical text to the UMLS Metathesaurus: The MetaMap program. In Proceedings of the AMIA Symposium. American Medical Informatics Association, 17."},{"issue":"1","key":"e_1_3_3_4_2","first-page":"D267\u2013D270","article-title":"The unified medical language system (UMLS): Integrating biomedical terminology","volume":"32","author":"Bodenreider Olivier","year":"2004","unstructured":"Olivier Bodenreider. 2004. The unified medical language system (UMLS): Integrating biomedical terminology. Nucleic Acids Res. 32, suppl_1 (2004), D267\u2013D270.","journal-title":"Nucleic Acids Res."},{"issue":"10","key":"e_1_3_3_5_2","first-page":"1","article-title":"Preliminary study on the construction of Chinese medical knowledge graph","volume":"33","author":"Byambasuren Odmaa","year":"2019","unstructured":"Odmaa Byambasuren, Yunfei Yang, Zhifang Sui, Damai Dai, Baobao Chang, Sujian Li, and Hongying Zan. 2019. Preliminary study on the construction of Chinese medical knowledge graph. J. Chin. Inf. Process. 33, 10 (2019), 1\u20139.","journal-title":"J. Chin. Inf. Process."},{"key":"e_1_3_3_6_2","first-page":"12657","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Chen Lihu","year":"2021","unstructured":"Lihu Chen, Ga\u00ebl Varoquaux, and Fabian M. Suchanek. 2021. A lightweight neural model for biomedical entity linking. In Proceedings of the AAAI Conference on Artificial Intelligence. 12657\u201312665."},{"key":"e_1_3_3_7_2","volume-title":"Convolutional Neural Network for Sentence Classification","author":"Chen Yahui","year":"2015","unstructured":"Yahui Chen. 2015. Convolutional Neural Network for Sentence Classification. Master\u2019s Thesis. University of Waterloo."},{"key":"e_1_3_3_8_2","doi-asserted-by":"publisher","DOI":"10.1007\/BF00994018"},{"key":"e_1_3_3_9_2","first-page":"297","volume-title":"Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing","author":"D\u2019Souza Jennifer","year":"2015","unstructured":"Jennifer D\u2019Souza and Vincent Ng. 2015. Sieve-based entity linking for the biomedical domain. In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing. 297\u2013302."},{"key":"e_1_3_3_10_2","first-page":"345","volume-title":"Proceedings of the Conference on Empirical Methods in Natural Language Processing","author":"ElSherief Mai","year":"2021","unstructured":"Mai ElSherief, Caleb Ziems, David Muchlinski, Vaishnavi Anupindi, Jordyn Seybolt, Munmun De Choudhury, and Diyi Yang. 2021. Latent hatred: A benchmark for understanding implicit hate speech. In Proceedings of the Conference on Empirical Methods in Natural Language Processing. 345\u2013363."},{"key":"e_1_3_3_11_2","first-page":"3816","volume-title":"Proceedings of the Joint Conference of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL-IJCNLP\u201921)","author":"Gao Tianyu","year":"2021","unstructured":"Tianyu Gao, Adam Fisch, and Danqi Chen. 2021. Making pre-trained language models better few-shot learners. In Proceedings of the Joint Conference of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (ACL-IJCNLP\u201921). Association for Computational Linguistics (ACL), 3816\u20133830."},{"key":"e_1_3_3_12_2","doi-asserted-by":"publisher","DOI":"10.1093\/bioinformatics\/bti586"},{"issue":"1","key":"e_1_3_3_13_2","first-page":"139159","article-title":"Frequent item-set mining and clustering based ranked biomedical text summarization","volume":"79","author":"Gupta Supriya","year":"2022","unstructured":"Supriya Gupta, Aakanksha Sharaff, and Naresh Kumar Nagwani. 2022. Frequent item-set mining and clustering based ranked biomedical text summarization. J. Supercomput. 79, 1 (2022), 139159.","journal-title":"J. Supercomput."},{"key":"e_1_3_3_14_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.aiopen.2022.11.003"},{"issue":"3","key":"e_1_3_3_15_2","first-page":"94","article-title":"Overview of the CHIP2019 shared task track1: Normalization of Chinese clinical terminology","volume":"35","author":"Huang Yuanhang","year":"2021","unstructured":"Yuanhang Huang, Xiaokang Jiao, Buzhou Tang, Qingcai Chen, and Jun Yan. 2021. Overview of the CHIP2019 shared task track1: Normalization of Chinese clinical terminology. J. Chin. Inf. Process. 35, 3 (2021), 94\u201399.","journal-title":"J. Chin. Inf. Process."},{"key":"e_1_3_3_16_2","first-page":"269","article-title":"BERT-based ranking for biomedical entity normalization","author":"Ji Zongcheng","year":"2020","unstructured":"Zongcheng Ji, Qiang Wei, and Hua Xu. 2020. BERT-based ranking for biomedical entity normalization. In Proceedings of the AMIA Summits on Translational Science Proceedings. 269.","journal-title":"Proceedings of the AMIA Summits on Translational Science Proceedings"},{"key":"e_1_3_3_17_2","doi-asserted-by":"publisher","DOI":"10.1093\/jamia\/ocv108"},{"key":"e_1_3_3_18_2","first-page":"4171","volume-title":"Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201919)","author":"Kenton Jacob Devlin Ming-Wei Chang","year":"2019","unstructured":"Jacob Devlin Ming-Wei Chang Kenton and Lee Kristina Toutanova. 2019. BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201919). 4171\u20134186."},{"key":"e_1_3_3_19_2","doi-asserted-by":"crossref","first-page":"277","DOI":"10.1201\/9781003102380-14","volume-title":"Data Science and Its Applications","author":"Kumar Ashutosh","year":"2021","unstructured":"Ashutosh Kumar and Aakanksha Sharaff. 2021. Deep parallel-embedded BioNER model for biomedical entity extraction. In Data Science and Its Applications. Chapman and Hall\/CRC, 277\u2013294."},{"key":"e_1_3_3_20_2","doi-asserted-by":"publisher","DOI":"10.1093\/comjnl\/bxad051"},{"key":"e_1_3_3_21_2","first-page":"61","volume-title":"Proceedings of the 1st CCF International Conference on Natural Language Processing and Chinese Computing (NLPCC\u201922)","author":"Lai Zhaohong","year":"2022","unstructured":"Zhaohong Lai, Biao Fu, Shangfei Wei, and Xiaodong Shi. 2022. Continuous prompt enhanced biomedical entity normalization. In Proceedings of the 1st CCF International Conference on Natural Language Processing and Chinese Computing (NLPCC\u201922). Springer, 61\u201372."},{"key":"e_1_3_3_22_2","doi-asserted-by":"publisher","DOI":"10.1093\/bioinformatics\/btt474"},{"key":"e_1_3_3_23_2","doi-asserted-by":"publisher","DOI":"10.1093\/bioinformatics\/btz682"},{"key":"e_1_3_3_24_2","first-page":"3045","volume-title":"Proceedings of the Conference on Empirical Methods in Natural Language Processing","author":"Lester Brian","year":"2021","unstructured":"Brian Lester, Rami Al-Rfou, and Noah Constant. 2021. The power of scale for parameter-efficient prompt tuning. In Proceedings of the Conference on Empirical Methods in Natural Language Processing. 3045\u20133059."},{"key":"e_1_3_3_25_2","first-page":"79","article-title":"CNN-based ranking for biomedical entity normalization","volume":"18","author":"Li Haodi","year":"2017","unstructured":"Haodi Li, Qingcai Chen, Buzhou Tang, Xiaolong Wang, Hua Xu, Baohua Wang, and Dong Huang. 2017. CNN-based ranking for biomedical entity normalization. BMC Bioinform. 18 (2017), 79\u201386.","journal-title":"BMC Bioinform."},{"key":"e_1_3_3_26_2","first-page":"1018","article-title":"Stacking-BERT model for Chinese medical procedure entity normalization.","volume":"20","author":"Li Luqi","year":"2023","unstructured":"Luqi Li, Yun kai Zhai, Jinghong Gao, Linlin Wang, Li Hou, and Jie Zhao. 2023. Stacking-BERT model for Chinese medical procedure entity normalization. Math. Biosci. Eng. 20 1 (2023), 1018\u20131036.","journal-title":"Math. Biosci. Eng."},{"key":"e_1_3_3_27_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2021.acl-long.353"},{"key":"e_1_3_3_28_2","first-page":"4228","volume-title":"Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201921)","author":"Liu Fangyu","year":"2021","unstructured":"Fangyu Liu, Ehsan Shareghi, Zaiqiao Meng, Marco Basaldella, and Nigel Collier. 2021. Self-alignment pretraining for biomedical entity representations. In Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT\u201921). 4228\u20134238."},{"key":"e_1_3_3_29_2","doi-asserted-by":"publisher","DOI":"10.1145\/3560815"},{"key":"e_1_3_3_30_2","first-page":"380","volume-title":"Proceedings of the International Multiconference of Engineers and Computer Scientists","author":"Niwattanakul Suphakit","year":"2013","unstructured":"Suphakit Niwattanakul, Jatsada Singthongchai, Ekkachai Naenudorn, and Supachanun Wanapu. 2013. Using of Jaccard coefficient for keywords similarity. In Proceedings of the International Multiconference of Engineers and Computer Scientists. 380\u2013384."},{"key":"e_1_3_3_31_2","doi-asserted-by":"publisher","DOI":"10.1007\/s11042-023-14980-3"},{"key":"e_1_3_3_32_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2019.06.002"},{"key":"e_1_3_3_33_2","volume-title":"Introduction to Modern Information Retrieval","author":"Salton Gerard","year":"1983","unstructured":"Gerard Salton and Michael J. McGill. 1983. Introduction to Modern Information Retrieval. McGraw-Hill."},{"key":"e_1_3_3_34_2","doi-asserted-by":"publisher","DOI":"10.1016\/j.eswa.2021.116475"},{"key":"e_1_3_3_35_2","first-page":"8337","volume-title":"Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201922)","author":"Sui Xuhui","year":"2022","unstructured":"Xuhui Sui, Kehui Song, Baohang Zhou, Ying Zhang, and Xiaojie Yuan. 2022. A multi-task learning framework for Chinese medical procedure entity normalization. In Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP\u201922). IEEE, 8337\u20138341."},{"key":"e_1_3_3_36_2","unstructured":"Yixuan Weng. 2021. ChineseWord2vecMedicine. Retrieved from https:\/\/github.com\/WENGSYX\/Chinese-Word2vec-Medicine"},{"key":"e_1_3_3_37_2","doi-asserted-by":"publisher","DOI":"10.1093\/bioinformatics\/btp071"},{"key":"e_1_3_3_38_2","volume-title":"Proceedings of the Conference on Automated Knowledge Base Construction","author":"Wright Dustin","year":"2019","unstructured":"Dustin Wright, Yannis Katsis, Raghav Mehta, and Chun-Nan Hsu. 2019. NormCo: Deep disease normalization for biomedical knowledge base construction. In Proceedings of the Conference on Automated Knowledge Base Construction."},{"key":"e_1_3_3_39_2","first-page":"8452","volume-title":"Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics","author":"Xu Dongfang","year":"2020","unstructured":"Dongfang Xu, Zeyu Zhang, and Steven Bethard. 2020. A generate-and-rank framework with semantic type regularization for biomedical concept normalization. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics. 8452\u20138464."},{"key":"e_1_3_3_40_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.emnlp-main.116"},{"key":"e_1_3_3_41_2","article-title":"GraphPrompt: Biomedical entity normalization using graph-based prompt templates","author":"Zhang Jiayou","year":"2021","unstructured":"Jiayou Zhang, Zhirui Wang, Shizhuo Zhang, Megh Manoj Bhalerao, Yucong Liu, Dawei Zhu, and Sheng Wang. 2021. GraphPrompt: Biomedical entity normalization using graph-based prompt templates. arXiv preprint arXiv:2112.03002 (2021).","journal-title":"arXiv preprint arXiv:2112.03002"},{"key":"e_1_3_3_42_2","article-title":"Conceptualized representation learning for Chinese biomedical text mining","author":"Zhang Ningyu","year":"2020","unstructured":"Ningyu Zhang, Qianghuai Jia, Kangping Yin, Liang Dong, Feng Gao, and Nengwei Hua. 2020. Conceptualized representation learning for Chinese biomedical text mining. arXiv preprint arXiv:2008.10813 (2020).","journal-title":"arXiv preprint arXiv:2008.10813"},{"key":"e_1_3_3_43_2","first-page":"1290","volume-title":"Proceedings of the 23rd International Conference on Computational Linguistics (COLING\u201910)","author":"Zhang Wei","year":"2010","unstructured":"Wei Zhang, Jian Su, Chew Lim Tan, and Wen Ting Wang. 2010. Entity linking leveraging automatically generated annotation. In Proceedings of the 23rd International Conference on Computational Linguistics (COLING\u201910). 1290\u20131298."},{"key":"e_1_3_3_44_2","doi-asserted-by":"crossref","first-page":"243","DOI":"10.18653\/v1\/D19-6127","volume-title":"Proceedings of the 2nd Workshop on Deep Learning Approaches for Low-Resource NLP (DeepLo\u201919)","author":"Zhou Shuyan","year":"2019","unstructured":"Shuyan Zhou, Shruti Rijhwani, and Graham Neubig. 2019. Towards zero-resource cross-lingual entity linking. In Proceedings of the 2nd Workshop on Deep Learning Approaches for Low-Resource NLP (DeepLo\u201919). 243\u2013252."},{"key":"e_1_3_3_45_2","doi-asserted-by":"publisher","DOI":"10.1162\/tacl_a_00303"},{"key":"e_1_3_3_46_2","first-page":"9757","volume-title":"Proceedings of the AAAI Conference on Artificial Intelligence","author":"Zhu Ming","year":"2020","unstructured":"Ming Zhu, Busra Celikkaya, Parminder Bhatia, and Chandan K. Reddy. 2020. LATTE: Latent type modeling for biomedical entity linking. In Proceedings of the AAAI Conference on Artificial Intelligence. 9757\u20139764."},{"key":"e_1_3_3_47_2","volume-title":"Proceedings of the International Joint Conference on Artificial Intelligence","author":"Zhu Tiantian","year":"2022","unstructured":"Tiantian Zhu, Yang Qin, Qingcai Chen, Baotian Hu, and Yang Xiang. 2022. Enhancing entity representations with prompt learning for biomedical entity linking. In Proceedings of the International Joint Conference on Artificial Intelligence."}],"container-title":["ACM Transactions on Asian and Low-Resource Language Information Processing"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3689629","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3689629","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,19]],"date-time":"2025-06-19T01:09:47Z","timestamp":1750295387000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3689629"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,10,23]]},"references-count":46,"journal-issue":{"issue":"10","published-print":{"date-parts":[[2024,10,31]]}},"alternative-id":["10.1145\/3689629"],"URL":"https:\/\/doi.org\/10.1145\/3689629","relation":{},"ISSN":["2375-4699","2375-4702"],"issn-type":[{"value":"2375-4699","type":"print"},{"value":"2375-4702","type":"electronic"}],"subject":[],"published":{"date-parts":[[2024,10,23]]},"assertion":[{"value":"2023-06-05","order":0,"name":"received","label":"Received","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-08-18","order":2,"name":"accepted","label":"Accepted","group":{"name":"publication_history","label":"Publication History"}},{"value":"2024-10-23","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}