{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,23]],"date-time":"2026-07-23T01:17:57Z","timestamp":1784769477927,"version":"3.55.0"},"reference-count":118,"publisher":"Oxford University Press (OUP)","issue":"10","license":[{"start":{"date-parts":[[2024,8,29]],"date-time":"2024-08-29T00:00:00Z","timestamp":1724889600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/academic.oup.com\/pages\/standard-publication-reuse-rights"}],"funder":[{"DOI":"10.13039\/100000002","name":"National Institutes of Health","doi-asserted-by":"publisher","id":[{"id":"10.13039\/100000002","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100008460","name":"National Center for Complementary and Integrative Health","doi-asserted-by":"publisher","award":["R01AT009457"],"award-info":[{"award-number":["R01AT009457"]}],"id":[{"id":"10.13039\/100008460","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000049","name":"National Institute on Aging","doi-asserted-by":"publisher","award":["R01AG078154"],"award-info":[{"award-number":["R01AG078154"]}],"id":[{"id":"10.13039\/100000049","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/100000054","name":"National Cancer Institute","doi-asserted-by":"publisher","award":["R01CA287413"],"award-info":[{"award-number":["R01CA287413"]}],"id":[{"id":"10.13039\/100000054","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2024,10,1]]},"abstract":"<jats:title>Abstract<\/jats:title>\n               <jats:sec>\n                  <jats:title>Importance<\/jats:title>\n                  <jats:p>Reinforcement learning (RL) represents a pivotal avenue within natural language processing (NLP), offering a potent mechanism for acquiring optimal strategies in task completion. This literature review studies various NLP applications where RL has demonstrated efficacy, with notable applications in healthcare settings.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Objectives<\/jats:title>\n                  <jats:p>To systematically explore the applications of RL in NLP, focusing on its effectiveness in acquiring optimal strategies, particularly in healthcare settings, and provide a comprehensive understanding of RL\u2019s potential in NLP tasks.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Materials and Methods<\/jats:title>\n                  <jats:p>Adhering to the PRISMA guidelines, an exhaustive literature review was conducted to identify instances where RL has exhibited success in NLP applications, encompassing dialogue systems, machine translation, question-answering, text summarization, and information extraction. Our methodological approach involves closely examining the technical aspects of RL methodologies employed in these applications, analyzing algorithms, states, rewards, actions, datasets, and encoder-decoder architectures.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Results<\/jats:title>\n                  <jats:p>The review of 93 papers yields insights into RL algorithms, prevalent techniques, emergent trends, and the fusion of RL methods in NLP healthcare applications. It clarifies the strategic approaches employed, datasets utilized, and the dynamic terrain of RL-NLP systems, thereby offering a roadmap for research and development in RL and machine learning techniques in healthcare. The review also addresses ethical concerns to ensure equity, transparency, and accountability in the evolution and application of RL-based NLP technologies, particularly within sensitive domains such as healthcare.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Discussion<\/jats:title>\n                  <jats:p>The findings underscore the promising role of RL in advancing NLP applications, particularly in healthcare, where its potential to optimize decision-making and enhance patient outcomes is significant. However, the ethical challenges and technical complexities associated with RL demand careful consideration and ongoing research to ensure responsible and effective implementation.<\/jats:p>\n               <\/jats:sec>\n               <jats:sec>\n                  <jats:title>Conclusions<\/jats:title>\n                  <jats:p>By systematically exploring RL\u2019s applications in NLP and providing insights into technical analysis, ethical implications, and potential advancements, this review contributes to a deeper understanding of RL\u2019s role for language processing.<\/jats:p>\n               <\/jats:sec>","DOI":"10.1093\/jamia\/ocae215","type":"journal-article","created":{"date-parts":[[2024,8,29]],"date-time":"2024-08-29T22:11:05Z","timestamp":1724969465000},"page":"2379-2393","source":"Crossref","is-referenced-by-count":19,"title":["A review of reinforcement learning for natural language processing and applications in healthcare"],"prefix":"10.1093","volume":"31","author":[{"given":"Ying","family":"Liu","sequence":"first","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Haozhu","family":"Wang","sequence":"additional","affiliation":[{"name":"Amazon Web Service , Seattle, WA 98109, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-6524-5506","authenticated-orcid":false,"given":"Huixue","family":"Zhou","sequence":"additional","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"},{"name":"Institute for Health Informatics, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Mingchen","family":"Li","sequence":"additional","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Yu","family":"Hou","sequence":"additional","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Sicheng","family":"Zhou","sequence":"additional","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"},{"name":"Institute for Health Informatics, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Fang","family":"Wang","sequence":"additional","affiliation":[{"name":"Amazon Web Service , Seattle, WA 98109, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Rama","family":"Hoetzlein","sequence":"additional","affiliation":[{"name":"R&D, Quanta Sciences , Ithaca, NY 14850, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-8258-3585","authenticated-orcid":false,"given":"Rui","family":"Zhang","sequence":"additional","affiliation":[{"name":"Department of Surgery, University of Minnesota , Minneapolis, MN 55455, United States"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"286","published-online":{"date-parts":[[2024,8,29]]},"reference":[{"key":"2024092007540951300_ocae215-B1","author":"OpenAI","year":"2022"},{"issue":"3-4","key":"2024092007540951300_ocae215-B2","doi-asserted-by":"crossref","first-page":"229","DOI":"10.1007\/BF00992696","article-title":"Simple statistical gradient-following algorithms for connectionist reinforcement learning","volume":"8","author":"Williams","year":"1992","journal-title":"Mach Learn"},{"issue":"7540","key":"2024092007540951300_ocae215-B3","doi-asserted-by":"crossref","first-page":"529","DOI":"10.1038\/nature14236","article-title":"Human-level control through deep reinforcement learning","volume":"518","author":"Mnih","year":"2015","journal-title":"Nature"},{"key":"2024092007540951300_ocae215-B4","author":"Schulman","year":"2017"},{"key":"2024092007540951300_ocae215-B5","author":"Radford"},{"issue":"2","key":"2024092007540951300_ocae215-B6","doi-asserted-by":"crossref","first-page":"1543","DOI":"10.1007\/s10462-022-10205-5","article-title":"Survey on reinforcement learning for language processing","volume":"56","author":"Uc-Cetina","year":"2023","journal-title":"Artif Intell Rev"},{"key":"2024092007540951300_ocae215-B7","first-page":"19","author":"Wang","year":"2018"},{"key":"2024092007540951300_ocae215-B8","author":"Lin","year":"2023"},{"issue":"12","key":"2024092007540951300_ocae215-B9","doi-asserted-by":"crossref","first-page":"2049","DOI":"10.1016\/j.infsof.2013.07.010","article-title":"A systematic review of systematic review process research in software engineering","volume":"155","author":"Kitchenham","year":"2013","journal-title":"Inf Softw Technol"},{"key":"2024092007540951300_ocae215-B10","author":"Zhou","year":"2021"},{"issue":"1","key":"2024092007540951300_ocae215-B11","doi-asserted-by":"crossref","first-page":"186","DOI":"10.1038\/s41746-022-00730-6","article-title":"A survey on clinical natural language processing in the United Kingdom from 2007 to 2022","volume":"215","author":"Wu","year":"2022","journal-title":"NPJ Digit Med"},{"key":"2024092007540951300_ocae215-B12","doi-asserted-by":"crossref","first-page":"101964","DOI":"10.1016\/j.artmed.2020.101964","article-title":"Reinforcement learning for intelligent healthcare applications: a survey","volume":"109","author":"Coronato","year":"2020","journal-title":"Artif Intell Med"},{"issue":"1","key":"2024092007540951300_ocae215-B13","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3477600","article-title":"Reinforcement learning in healthcare: a survey","volume":"55","author":"Yu","year":"2023","journal-title":"ACM Comput Surv"},{"key":"2024092007540951300_ocae215-B14","author":"Abdellatif","year":"2021"},{"key":"2024092007540951300_ocae215-B15","doi-asserted-by":"crossref","first-page":"102193","DOI":"10.1016\/j.media.2021.102193","article-title":"Deep reinforcement learning in medical imaging: a literature review","volume":"73","author":"Zhou","year":"2021","journal-title":"Med Image Anal"},{"issue":"7","key":"2024092007540951300_ocae215-B16","doi-asserted-by":"crossref","first-page":"e18477","DOI":"10.2196\/18477","article-title":"Reinforcement learning for clinical decision support in critical care: comprehensive review","volume":"22","author":"Liu","year":"2020","journal-title":"J Med Internet Res"},{"issue":"4","key":"2024092007540951300_ocae215-B17","doi-asserted-by":"crossref","first-page":"2733","DOI":"10.1007\/s10462-021-10061-9","article-title":"Deep reinforcement learning in computer vision: a comprehensive survey","volume":"55","author":"Le","year":"2022","journal-title":"Artif Intell Rev"},{"key":"2024092007540951300_ocae215-B18","author":"Liu","year":"2019"},{"key":"2024092007540951300_ocae215-B19","first-page":"549","volume-title":"Reinforcement Learning: An Introduction","author":"Sutton","year":"2018","edition":"2nd ed."},{"key":"2024092007540951300_ocae215-B20","first-page":"2338","author":"Gigioli","year":"2018"},{"issue":"13","key":"2024092007540951300_ocae215-B21","doi-asserted-by":"crossref","first-page":"1891","DOI":"10.1093\/bioinformatics\/btab024","article-title":"Efficient multiple biomedical events extraction via reinforcement learning","volume":"37","author":"Zhao","year":"2021","journal-title":"Bioinformatics"},{"key":"2024092007540951300_ocae215-B22","doi-asserted-by":"crossref","first-page":"5557184","DOI":"10.1155\/2021\/5557184","article-title":"A sentence-level joint relation classification model based on reinforcement learning","volume":"2021","author":"Liu","year":"2021","journal-title":"Comput Intell Neurosci"},{"key":"2024092007540951300_ocae215-B23","author":"Feng","year":"2018"},{"key":"2024092007540951300_ocae215-B24","first-page":"95","author":"Xu","year":"2020"},{"issue":"1","key":"2024092007540951300_ocae215-B25","first-page":"5658","article-title":"Large scaled relation extraction with reinforcement learning","volume":"32","author":"Zeng","year":"2018","journal-title":"Proc AAAI Conf Artif Intell"},{"key":"2024092007540951300_ocae215-B26","doi-asserted-by":"crossref","first-page":"597","DOI":"10.1007\/978-3-030-92310-5_69","volume-title":"Neural Information Processing","author":"Nguyen","year":"2021"},{"key":"2024092007540951300_ocae215-B27","doi-asserted-by":"crossref","first-page":"11279","DOI":"10.1109\/ACCESS.2020.2965575","article-title":"Text summarization method based on double attention pointer network","volume":"8","author":"Li","year":"2020","journal-title":"IEEE Access"},{"key":"2024092007540951300_ocae215-B28","author":"Sharma","year":"2019"},{"key":"2024092007540951300_ocae215-B29","first-page":"2061","author":"Tian","year":"2019"},{"issue":"11","key":"2024092007540951300_ocae215-B30","doi-asserted-by":"crossref","first-page":"e38095","DOI":"10.2196\/38095","article-title":"Medical Text Simplification Using Reinforcement Learning (TESLEA): deep learning-based text simplification approach","volume":"10","author":"Phatak","year":"2022","journal-title":"JMIR Med Inform"},{"key":"2024092007540951300_ocae215-B31","first-page":"5602","author":"Wu","year":"2018"},{"key":"2024092007540951300_ocae215-B32","first-page":"194","author":"Sharma","year":"2021"},{"key":"2024092007540951300_ocae215-B33","first-page":"3612","author":"Wu","year":"2018"},{"key":"2024092007540951300_ocae215-B34","first-page":"523","author":"Geng","year":"2018"},{"key":"2024092007540951300_ocae215-B35","first-page":"3022","author":"Alinejad","year":"2018"},{"key":"2024092007540951300_ocae215-B36","first-page":"1368","author":"Tebbifakhr","year":"2019"},{"key":"2024092007540951300_ocae215-B37","first-page":"235","author":"Tebbifakhr","year":"2020"},{"key":"2024092007540951300_ocae215-B38","first-page":"120","author":"Dong","year":"2021"},{"key":"2024092007540951300_ocae215-B39","author":"Huang","year":"2021"},{"key":"2024092007540951300_ocae215-B40","author":"Buck","year":"2018"},{"key":"2024092007540951300_ocae215-B41","first-page":"5981","author":"Wang","year":"2018"},{"key":"2024092007540951300_ocae215-B42","doi-asserted-by":"crossref","first-page":"22988","DOI":"10.1109\/ACCESS.2019.2894438","article-title":"A home service-oriented question answering system with high accuracy and stability","volume":"7","author":"Zhang","year":"2019","journal-title":"IEEE Access"},{"issue":"3","key":"2024092007540951300_ocae215-B43","doi-asserted-by":"crossref","first-page":"252","DOI":"10.1016\/j.ipm.2015.01.002","article-title":"A reinforcement learning formulation to the complex question answering problem","volume":"51","author":"Chali","year":"2015","journal-title":"Inf Process Manag"},{"key":"2024092007540951300_ocae215-B44","author":"Kandasamy","year":"2017"},{"key":"2024092007540951300_ocae215-B45","first-page":"87","author":"Chou","year":"2019"},{"key":"2024092007540951300_ocae215-B46","first-page":"895","author":"Ling","year":"2017"},{"key":"2024092007540951300_ocae215-B47","first-page":"2137","author":"Qin","year":"2018"},{"key":"2024092007540951300_ocae215-B48","first-page":"271","author":"Ling","year":"2017"},{"key":"2024092007540951300_ocae215-B49","first-page":"397","author":"Wan","year":"2018"},{"key":"2024092007540951300_ocae215-B50","first-page":"1192","author":"Li","year":"2016"},{"key":"2024092007540951300_ocae215-B51","doi-asserted-by":"crossref","first-page":"124","DOI":"10.1007\/978-981-16-8656-6_11","volume-title":"LISS 2021","author":"Wang","year":"2022"},{"issue":"1","key":"2024092007540951300_ocae215-B52","doi-asserted-by":"crossref","first-page":"14841","DOI":"10.1149\/10701.14841ecst","article-title":"An artificial intelligent-based Chatbot for dosage prediction of medicine using noval deep reinforcement learning with natural language processing","volume":"107","author":"B","year":"2022","journal-title":"ECS Trans"},{"key":"2024092007540951300_ocae215-B53","first-page":"1","author":"Liu","year":"2019"},{"key":"2024092007540951300_ocae215-B54","doi-asserted-by":"crossref","first-page":"81320","DOI":"10.1109\/ACCESS.2020.2990680","article-title":"Reducing wrong labels for distantly supervised relation extraction with reinforcement learning","volume":"8","author":"Chen","year":"2020","journal-title":"IEEE Access"},{"key":"2024092007540951300_ocae215-B55","first-page":"2311","author":"Xu","year":"2022"},{"key":"2024092007540951300_ocae215-B56","author":"Shaham","year":"2020"},{"key":"2024092007540951300_ocae215-B57","author":"Yuan","year":"2019"},{"key":"2024092007540951300_ocae215-B58","first-page":"201","author":"Wei","year":"2018"},{"key":"2024092007540951300_ocae215-B59","doi-asserted-by":"crossref","first-page":"282","DOI":"10.1016\/j.neunet.2022.06.018","article-title":"A universal adversarial policy for text classifiers","volume":"153","author":"Maimon","year":"2022","journal-title":"Neural Netw"},{"key":"2024092007540951300_ocae215-B60","first-page":"27730","article-title":"Training language models to follow instructions with human feedback","volume":"35","author":"Ouyang","year":"2022","journal-title":"Adv Neural Inf Process Syst"},{"key":"2024092007540951300_ocae215-B61","author":"Wang","year":"2019"},{"key":"2024092007540951300_ocae215-B62","author":"Lu","year":"2022"},{"key":"2024092007540951300_ocae215-B63","author":"Jie","year":"2023"},{"key":"2024092007540951300_ocae215-B64","first-page":"1464","author":"Nguyen","year":"2017"},{"issue":"4","key":"2024092007540951300_ocae215-B65","doi-asserted-by":"crossref","first-page":"1","DOI":"10.1145\/3450970","article-title":"Reinforced NMT for sentiment and content preservation in low-resource scenario","volume":"20","author":"Kumari","year":"2021","journal-title":"ACM Trans Asian Low-Resour Lang Inf Process"},{"issue":"16","key":"2024092007540951300_ocae215-B66","doi-asserted-by":"crossref","first-page":"3391","DOI":"10.3390\/electronics12163391","article-title":"Research on the application of prompt learning pretrained language model in machine translation task with reinforcement learning","volume":"12","author":"Wang","year":"2023","journal-title":"Electronics"},{"key":"2024092007540951300_ocae215-B67","author":"Zeng"},{"key":"2024092007540951300_ocae215-B68","doi-asserted-by":"crossref","first-page":"1335","DOI":"10.1016\/j.procs.2023.01.112","article-title":"Natural language processing for Covid-19 consulting system","volume":"218","author":"Tripathy","year":"2023","journal-title":"Procedia Comput Sci"},{"key":"2024092007540951300_ocae215-B69","first-page":"9241","author":"Zeng","year":"2020"},{"key":"2024092007540951300_ocae215-B70","first-page":"1342","author":"Grissom Ii","year":"2014"},{"key":"2024092007540951300_ocae215-B71","first-page":"4586","author":"Naseem","year":"2019"},{"issue":"11","key":"2024092007540951300_ocae215-B72","doi-asserted-by":"crossref","first-page":"2980","DOI":"10.14778\/3551793.3551846","article-title":"BABOONS: black-box optimization of data summaries in natural language","volume":"15","author":"Trummer","year":"2022","journal-title":"Proc VLDB Endow"},{"key":"2024092007540951300_ocae215-B73","author":"Ouyang","year":"2022"},{"key":"2024092007540951300_ocae215-B74","first-page":"13","author":"Chinaei","year":"2014"},{"issue":"1","key":"2024092007540951300_ocae215-B75","doi-asserted-by":"crossref","first-page":"7072","DOI":"10.1609\/aaai.v33i01.33017072","article-title":"A hierarchical framework for relation extraction with reinforcement learning","volume":"33","author":"Takanobu","year":"2019","journal-title":"AAAI"},{"key":"2024092007540951300_ocae215-B76","first-page":"223","author":"Zhu","year":"2021"},{"key":"2024092007540951300_ocae215-B77","first-page":"634","author":"Camara","year":"2015"},{"key":"2024092007540951300_ocae215-B78","first-page":"311","author":"Papineni","year":"2001"},{"key":"2024092007540951300_ocae215-B79","first-page":"74","author":"Lin","year":"2004"},{"issue":"19","key":"2024092007540951300_ocae215-B80","doi-asserted-by":"crossref","first-page":"1061","DOI":"10.21037\/atm-22-3991","article-title":"Entity relation extraction in the medical domain: based on data augmentation","volume":"10","author":"Wang","year":"2022","journal-title":"Ann Transl Med"},{"key":"2024092007540951300_ocae215-B81","first-page":"47","author":"Shim","year":"2021"},{"key":"2024092007540951300_ocae215-B82","author":"Kreyssig","year":"2018"},{"key":"2024092007540951300_ocae215-B83","author":"Shi","year":"2022"},{"key":"2024092007540951300_ocae215-B84","doi-asserted-by":"crossref","first-page":"121","DOI":"10.1016\/j.iotcps.2023.04.003","article-title":"ChatGPT: a comprehensive review on background, applications, key challenges, bias, ethics, limitations and future scope","volume":"3","author":"Ray","year":"2023","journal-title":"Internet Things Cyber-Phys Syst"},{"issue":"2","key":"2024092007540951300_ocae215-B85","first-page":"186","article-title":"Review of artificial intelligence-based question-answering systems in healthcare","volume":"13","author":"Budler","year":"2023","journal-title":"WIREs Data Min Knowl Discov"},{"key":"2024092007540951300_ocae215-B86","author":"Jin","year":"2019"},{"issue":"3","key":"2024092007540951300_ocae215-B87","doi-asserted-by":"crossref","first-page":"102933","DOI":"10.1016\/j.ipm.2022.102933","article-title":"ARL: an adaptive reinforcement learning framework for complex question answering over knowledge base","volume":"59","author":"Zhang","year":"2022","journal-title":"Inf Process Manag"},{"key":"2024092007540951300_ocae215-B88","first-page":"474","author":"Qiu","year":"2020"},{"key":"2024092007540951300_ocae215-B89","author":"Hua","year":"2020"},{"issue":"5","key":"2024092007540951300_ocae215-B90","doi-asserted-by":"crossref","first-page":"1275","DOI":"10.1007\/s11606-021-07164-y","article-title":"A research agenda for using machine translation in clinical medicine","volume":"37","author":"Khoong","year":"2022","journal-title":"J Gen Intern Med"},{"issue":"4","key":"2024092007540951300_ocae215-B91","doi-asserted-by":"crossref","first-page":"580","DOI":"10.1001\/jamainternmed.2018.7653","article-title":"Assessing the use of Google translate for Spanish and Chinese translations of emergency department discharge instructions","volume":"179","author":"Khoong","year":"2019","journal-title":"JAMA Intern Med"},{"key":"2024092007540951300_ocae215-B92","first-page":"2016","author":"Mehandru","year":"2022"},{"key":"2024092007540951300_ocae215-B93","first-page":"48","author":"Tang","year":"2023"},{"key":"2024092007540951300_ocae215-B94","first-page":"259","author":"Pineau","year":"2010"},{"key":"2024092007540951300_ocae215-B95","first-page":"1","author":"Mugoye","year":"2019"},{"key":"2024092007540951300_ocae215-B96","first-page":"137","author":"Morbini","year":"2012"},{"key":"2024092007540951300_ocae215-B97","author":"Yunxiang"},{"key":"2024092007540951300_ocae215-B98","doi-asserted-by":"crossref","first-page":"233","DOI":"10.1007\/978-981-19-2416-3_13","volume-title":"Next Generation Healthcare Informatics","author":"Kulkarni","year":"2022"},{"issue":"1","key":"2024092007540951300_ocae215-B99","doi-asserted-by":"crossref","first-page":"49","DOI":"10.1109\/TNNLS.2020.2975035","article-title":"Multitask learning and reinforcement learning for personalized dialog generation: an empirical study","volume":"32","author":"Yang","year":"2021","journal-title":"IEEE Trans Neural Netw Learn Syst"},{"issue":"7","key":"2024092007540951300_ocae215-B100","doi-asserted-by":"crossref","first-page":"e0234894","DOI":"10.1371\/journal.pone.0234894","article-title":"A reinforcement-learning approach to efficient communication","volume":"15","author":"K\u00e5geb\u00e4ck","year":"2020","journal-title":"PLoS One"},{"issue":"4","key":"2024092007540951300_ocae215-B101","doi-asserted-by":"crossref","first-page":"24","DOI":"10.1109\/MSPEC.2019.8678513","article-title":"IBM Watson, heal thyself: how IBM overpromised and underdelivered on AI health care","volume":"56","author":"Strickland","year":"2019","journal-title":"IEEE Spectr"},{"key":"2024092007540951300_ocae215-B102","author":"Roy","year":"2022"},{"key":"2024092007540951300_ocae215-B103","author":"Singhal","year":"2022"},{"key":"2024092007540951300_ocae215-B104","first-page":"1","author":"Erraki","year":"2020"},{"issue":"1","key":"2024092007540951300_ocae215-B105","doi-asserted-by":"crossref","first-page":"1488","DOI":"10.3934\/mbe.2023067","article-title":"Extractive text summarization model based on advantage actor-critic and graph matrix methodology","volume":"20","author":"Yang","year":"2023","journal-title":"Math Biosci Eng"},{"key":"2024092007540951300_ocae215-B106","first-page":"4120","author":"Gao","year":"2018"},{"key":"2024092007540951300_ocae215-B107","doi-asserted-by":"crossref","first-page":"888","DOI":"10.1162\/tacl_a_00496","article-title":"Dependency parsing with backtracking using deep reinforcement learning","volume":"10","author":"Dary","year":"2022","journal-title":"Trans Assoc Comput Linguist"},{"key":"2024092007540951300_ocae215-B108","first-page":"5419","author":"Lu","year":"2022"},{"key":"2024092007540951300_ocae215-B109","first-page":"677","author":"L\u00ea","year":"2017"},{"key":"2024092007540951300_ocae215-B110","author":"Guan","year":"2023"},{"key":"2024092007540951300_ocae215-B111","author":"Yuan","year":"2022"},{"key":"2024092007540951300_ocae215-B112","first-page":"2223","author":"Nishino","year":"2020"},{"issue":"8","key":"2024092007540951300_ocae215-B113","doi-asserted-by":"crossref","first-page":"e37818","DOI":"10.2196\/37818","article-title":"Emotion-based reinforcement attention network for depression detection on social media: algorithm development and validation","volume":"10","author":"Cui","year":"2022","journal-title":"JMIR Med Inform"},{"key":"2024092007540951300_ocae215-B114","first-page":"5580","author":"Wang","year":"2019"},{"issue":"1","key":"2024092007540951300_ocae215-B115","doi-asserted-by":"crossref","first-page":"173","DOI":"10.1109\/TCBB.2019.2948985","article-title":"A method for generating synthetic electronic medical record text","volume":"18","author":"Guan","journal-title":"IEEE\/ACM Trans Comput Biol Bioinform"},{"key":"2024092007540951300_ocae215-B116","first-page":"2","author":"Shaham","year":"126"},{"key":"2024092007540951300_ocae215-B117","author":"Sharma","year":"2019"},{"key":"2024092007540951300_ocae215-B118","first-page":"123","author":"Henderson","year":"2018"}],"container-title":["Journal of the American Medical Informatics Association"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/academic.oup.com\/jamia\/advance-article-pdf\/doi\/10.1093\/jamia\/ocae215\/59206439\/ocae215.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"syndication"},{"URL":"https:\/\/academic.oup.com\/jamia\/advance-article-pdf\/doi\/10.1093\/jamia\/ocae215\/59206439\/ocae215.pdf","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,9,20]],"date-time":"2024-09-20T07:54:39Z","timestamp":1726818879000},"score":1,"resource":{"primary":{"URL":"https:\/\/academic.oup.com\/jamia\/article\/31\/10\/2379\/7745290"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,29]]},"references-count":118,"journal-issue":{"issue":"10","published-online":{"date-parts":[[2024,8,29]]},"published-print":{"date-parts":[[2024,10,1]]}},"URL":"https:\/\/doi.org\/10.1093\/jamia\/ocae215","relation":{},"ISSN":["1067-5027","1527-974X"],"issn-type":[{"value":"1067-5027","type":"print"},{"value":"1527-974X","type":"electronic"}],"subject":[],"published-other":{"date-parts":[[2024,10]]},"published":{"date-parts":[[2024,8,29]]}}}