{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T22:59:48Z","timestamp":1781650788477,"version":"3.54.5"},"reference-count":29,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2026,4,26]],"date-time":"2026-04-26T00:00:00Z","timestamp":1777161600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,4,26]],"date-time":"2026-04-26T00:00:00Z","timestamp":1777161600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100012681","name":"Universit\u00e4tsklinikum K\u00f6ln","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100012681","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Discov Artif Intell"],"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:sec>\n                    <jats:title>Objective<\/jats:title>\n                    <jats:p>This pilot study evaluated ChatGPT\u2019s ability to answer standard emergency room (ER) case questions compared with that of junior residents at a level 1 trauma centre. Since its release in March 2023, ChatGPT\u20114 has seen widespread adoption in education, research, and clinical practice, prompting interest in its medical utility. Despite concerns regarding accountability and trust, large language models (LLMs) have shown potential by-passing medical exams and generating differential diagnoses.<\/jats:p>\n                  <\/jats:sec>\n                  <jats:sec>\n                    <jats:title>Materials and methods<\/jats:title>\n                    <jats:p>The study was conducted at a level 1 trauma centre with an interdisciplinary ER from 21st to 27th February 2025. 21 fictional medical cases commonly encountered in the surgical ER were used. Seven junior residents with varying levels of training were included for comparison. ChatGPT and the doctors were prompted to answer these cases under identical conditions. Two board\u2011certified trauma surgeons independently rated responses for correctness, completeness, and adaptability; results were statistically analyzed.<\/jats:p>\n                  <\/jats:sec>\n                  <jats:sec>\n                    <jats:title>Results<\/jats:title>\n                    <jats:p>ChatGPT achieved mean scores of 4.62\u2009\u00b1\u20090.79 (correctness), 4.67\u2009\u00b1\u20090.56 (completeness), and 4.52\u2009\u00b1\u20090.66 (adaptability). Residents in years 1\u20133 scored 4.19\u2009\u00b1\u20090.05, 3.95\u2009\u00b1\u20090.48, and 4.05\u2009\u00b1\u20090.1, respectively; years 4\u20136 scored 3.99\u2009\u00b1\u20090.19, 3.71\u2009\u00b1\u20090.3, and 3.96\u2009\u00b1\u20090.2. ChatGPT and residents both performed best in lower extremity cases (14.5 vs. 12.9 points).<\/jats:p>\n                  <\/jats:sec>\n                  <jats:sec>\n                    <jats:title>Conclusions<\/jats:title>\n                    <jats:p>This pilot study indicates that ChatGPT can generate ER case responses comparable in quality to those of junior residents, particularly in completeness and correctness. Given the small sample size and vignette-based design, these findings do not support replacing physicians but highlight the potential of large language models as supportive or educational tools requiring clinician oversight. Establishing safe and ethical frameworks for AI use in medicine remains essential.<\/jats:p>\n                  <\/jats:sec>","DOI":"10.1007\/s44163-026-01276-2","type":"journal-article","created":{"date-parts":[[2026,4,26]],"date-time":"2026-04-26T01:54:30Z","timestamp":1777168470000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["ChatGPT-4 in comparison with traumatological junior doctors in emergency room cases at a level 1 trauma centre \u2013 a pilot study"],"prefix":"10.1007","volume":"6","author":[{"given":"Sebastian","family":"Wegmann","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Jannick","family":"Leyendecker","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Maximilian","family":"Lenz","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Michael","family":"Sarter","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Lars-Peter","family":"M\u00fcller","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Maximilian","family":"Weber","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,4,26]]},"reference":[{"issue":"4","key":"1276_CR1","doi-asserted-by":"publisher","first-page":"Doc54","DOI":"10.3205\/zma001636","volume":"40","author":"S Moritz","year":"2023","unstructured":"Moritz S, Romeike B, Stosch C, Tolks D. Generative AI (gAI) in medical education: Chat-GPT and co. GMS J Med Educ. 2023;40(4):Doc54. https:\/\/doi.org\/10.3205\/zma001636.","journal-title":"GMS J Med Educ"},{"key":"1276_CR2","doi-asserted-by":"publisher","first-page":"1076740747","DOI":"10.1016\/j.chb.2023.107674","volume":"143","author":"Q Zhou","year":"2023","unstructured":"Zhou Q, Li B, Han L, Jou M. Talking to a bot or a wall? How chatbots vs. human agents affect anticipated communication quality. Comput Hum Behav. 2023;143:1076740747\u20135632.","journal-title":"Comput Hum Behav"},{"key":"1276_CR3","doi-asserted-by":"publisher","first-page":"1019250736","DOI":"10.1016\/j.tele.2022.101925","volume":"77","author":"S Kelly","year":"2023","unstructured":"Kelly S, Kaye S-A, Oviedo-Trespalacios O. What factors contribute to the acceptance of artificial intelligence? A systematic review. Telematics Inform. 2023;77:1019250736\u20135853.","journal-title":"Telematics Inform"},{"key":"1276_CR4","doi-asserted-by":"crossref","unstructured":"Dignum V. Responsibility and artificial intelligence. Oxf Handb ethics AI. 2020;4698(215). 0190067403.","DOI":"10.1093\/oxfordhb\/9780190067397.013.12"},{"key":"1276_CR5","unstructured":"Antonovsky A, Franke A, Schulte N. Salutogenese: zur entmystifizierung der gesundheit. dgvt; 1997."},{"issue":"2","key":"1276_CR6","doi-asserted-by":"publisher","first-page":"e0000198","DOI":"10.1371\/journal.pdig.0000198","volume":"2","author":"TH Kung","year":"2023","unstructured":"Kung TH, Cheatham M, Medenilla A, Sillos C, De Leon L, Elepano C, et al. Performance of ChatGPT on USMLE: potential for AI-assisted medical education using large language models. PLOS Digit Health. 2023;2(2):e0000198. https:\/\/doi.org\/10.1371\/journal.pdig.0000198.","journal-title":"PLOS Digit Health"},{"issue":"23","key":"1276_CR7","doi-asserted-by":"publisher","first-page":"1173","DOI":"10.5435\/JAAOS-D-23-00396","volume":"31","author":"PA Massey","year":"2023","unstructured":"Massey PA, Montgomery C, Zhang AS. Comparison of ChatGPT\u20133.5, ChatGPT-4, and orthopaedic resident performance on orthopaedic assessment examinations %U https\/\/. JAAOS - J Am Acad Orthop Surg. 2023;31(23):1173\u20139. https:\/\/doi.org\/10.5435\/jaaos-d-23-00396. http:\/\/journals.lww.com\/jaaos\/fulltext\/2023\/12010\/comparison_of_chatgpt_3_5,_chatgpt_4,_and.1.aspx.","journal-title":"JAAOS - J Am Acad Orthop Surg"},{"key":"1276_CR8","doi-asserted-by":"publisher","first-page":"e67696","DOI":"10.2196\/67696","volume":"4","author":"M Pastrak","year":"2025","unstructured":"Pastrak M, Kajitani S, Goodings AJ, Drewek A, LaFree A, Murphy A. Evaluation of ChatGPT performance on emergency medicine board examination questions: observational study. JMIR AI. 2025;4:e67696.","journal-title":"JMIR AI"},{"issue":"1","key":"1276_CR9","doi-asserted-by":"publisher","first-page":"83","DOI":"10.1016\/j.annemergmed.2023.08.003","volume":"83","author":"H Ten Berg","year":"2024","unstructured":"Ten Berg H, van Bakel B, van de Wouw L, Jie KE, Schipper A, Jansen H, et al. ChatGPT and generating a differential diagnosis early in an emergency department presentation. Ann Emerg Med. 2024;83(1):83\u201360196.","journal-title":"Ann Emerg Med"},{"issue":"11","key":"1276_CR10","doi-asserted-by":"publisher","first-page":"1645","DOI":"10.1016\/j.jsurg.2024.08.002","volume":"81","author":"DS Hayes","year":"2024","unstructured":"Hayes DS, Foster BK, Makar G, Manzar S, Ozdag Y, Shultz M, et al. Artificial Intelligence in orthopaedics: performance of ChatGPT on text and image questions on a complete AAOS orthopaedic in-training examination (OITE). J Surg Educ. 2024;81(11):1645\u20139. https:\/\/doi.org\/10.1016\/j.jsurg.2024.08.002.","journal-title":"J Surg Educ"},{"issue":"2","key":"1276_CR11","doi-asserted-by":"publisher","first-page":"241","DOI":"10.1016\/j.surg.2024.04.003","volume":"176","author":"DL Palenzuela","year":"2024","unstructured":"Palenzuela DL, Mullen JT, Phitayakorn RAI, Versus MD. Evaluating the surgical decision-making accuracy of ChatGPT-4. Surgery. 2024;176(2):241\u20135. https:\/\/doi.org\/10.1016\/j.surg.2024.04.003.","journal-title":"Surgery"},{"issue":"6","key":"1276_CR12","doi-asserted-by":"publisher","first-page":"349","DOI":"10.1093\/intqhc\/mzm042","volume":"19","author":"A Tong","year":"2007","unstructured":"Tong A, Sainsbury P, Craig J. Consolidated criteria for reporting qualitative research (COREQ): a 32-item checklist for interviews and focus groups. Int J Qual Health Care. 2007;19(6):349\u201357. https:\/\/doi.org\/10.1093\/intqhc\/mzm042.","journal-title":"Int J Qual Health Care"},{"issue":"10","key":"1276_CR13","doi-asserted-by":"publisher","first-page":"3179","DOI":"10.1007\/s11606-021-06737-1","volume":"36","author":"A Sharma","year":"2021","unstructured":"Sharma A, Minh Duc NT, Luu Lam Thang T, Nam NH, Ng SJ, Abbas KS, et al. A consensus-based checklist for reporting of survey studies (CROSS). J Gen Intern Med. 2021;36(10):3179\u201387. https:\/\/doi.org\/10.1007\/s11606-021-06737-1.","journal-title":"J Gen Intern Med"},{"key":"1276_CR14","doi-asserted-by":"publisher","first-page":"1199350","DOI":"10.3389\/frai.2023.1199350","volume":"6","author":"G Orru","year":"2023","unstructured":"Orru G, Piarulli A, Conversano C, Gemignani A. Human-like problem-solving abilities in large language models using ChatGPT. Front Artif Intell. 2023;6:1199350. https:\/\/doi.org\/10.3389\/frai.2023.1199350.","journal-title":"Front Artif Intell"},{"key":"1276_CR15","unstructured":"White J, Fu Q, Hays S, Sandborn M, Olea C, Gilbert H et al. A prompt pattern catalog to enhance prompt engineering with chatgpt. arXiv preprint arXiv:230211382. 2023."},{"issue":"2","key":"1276_CR16","doi-asserted-by":"publisher","first-page":"148","DOI":"10.3928\/01913913-20240124-02","volume":"61","author":"JS Chen","year":"2024","unstructured":"Chen JS, Granet DB. Prompt engineering: helping ChatGPT respond better to patients and parents. J Pediatr Ophthalmol Strabismus. 2024;61(2):148\u20139. https:\/\/doi.org\/10.3928\/01913913-20240124-02.","journal-title":"J Pediatr Ophthalmol Strabismus"},{"key":"1276_CR17","doi-asserted-by":"publisher","unstructured":"Indran IR, Paranthaman P, Gupta N, Mustafa N. Twelve tips to leverage AI for efficient and effective medical question generation: a guide for educators using Chat GPT. Med Teach. 2023;1\u20136. https:\/\/doi.org\/10.1080\/0142159X.2023.2294703.","DOI":"10.1080\/0142159X.2023.2294703"},{"issue":"11","key":"1276_CR18","doi-asserted-by":"publisher","first-page":"5190","DOI":"10.1007\/s00167-023-07529-2","volume":"31","author":"J Kaarre","year":"2023","unstructured":"Kaarre J, Feldt R, Keeling LE, Dadoo S, Zsidai B, Hughes JD, et al. Exploring the potential of ChatGPT as a supplementary tool for providing orthopaedic information. Knee Surg Sports Traumatol Arthrosc. 2023;31(11):5190\u20138. https:\/\/doi.org\/10.1007\/s00167-023-07529-2.","journal-title":"Knee Surg Sports Traumatol Arthrosc"},{"issue":"5","key":"1276_CR19","doi-asserted-by":"publisher","first-page":"e248895","DOI":"10.1001\/jamanetworkopen.2024.8895","volume":"7","author":"CYK Williams","year":"2024","unstructured":"Williams CYK, Zack T, Miao BY, Sushil M, Wang M, Kornblith AE, et al. Use of a large language model to assess clinical acuity of adults in the emergency department. JAMA Netw Open. 2024;7(5):e248895. https:\/\/doi.org\/10.1001\/jamanetworkopen.2024.8895.","journal-title":"JAMA Netw Open"},{"issue":"1","key":"1276_CR20","doi-asserted-by":"publisher","first-page":"8236","DOI":"10.1038\/s41467-024-52415-1","volume":"15","author":"CYK Williams","year":"2024","unstructured":"Williams CYK, Miao BY, Kornblith AE, Butte AJ. Evaluating the use of large language models to provide clinical recommendations in the emergency department. Nat Commun. 2024;15(1):8236. https:\/\/doi.org\/10.1038\/s41467-024-52415-1.","journal-title":"Nat Commun"},{"issue":"4","key":"1276_CR21","doi-asserted-by":"publisher","first-page":"745","DOI":"10.1007\/s10439-023-03318-7","volume":"52","author":"M Alessandri Bonetti","year":"2024","unstructured":"Alessandri Bonetti M, Giorgino R, Gallo Afflitto G, De Lorenzi F, Egro FM. How does ChatGPT perform on the Italian residency admission national exam compared to 15,869 medical graduates? Ann Biomed Eng. 2024;52(4):745\u20139. https:\/\/doi.org\/10.1007\/s10439-023-03318-7.","journal-title":"Ann Biomed Eng"},{"key":"1276_CR22","unstructured":"Jung Lb Fau -, Gudera JA. Gudera Ja Fau - Wiegand TLT, Wiegand Tlt Fau - Allmendinger S, Allmendinger S Fau - Dimitriadis K, Dimitriadis K Fau - Koerte IK, Koerte IK. ChatGPT Passes German State Examination in Medicine With Picture Questions Omitted. (1866\u2009\u2013\u20090452 (Electronic))."},{"issue":"2","key":"1276_CR23","doi-asserted-by":"publisher","first-page":"155","DOI":"10.1272\/jnms.JNMS.2024_91-205","volume":"91","author":"Y Igarashi","year":"2024","unstructured":"Igarashi Y, Nakahara K, Norii T, Miyake N, Tagami T, Yokobori S. Performance of a large language model on japanese emergency medicine board certification examinations. J Nippon Med Sch. 2024;91(2):155\u201361. https:\/\/doi.org\/10.1272\/jnms.JNMS.2024_91-205.","journal-title":"J Nippon Med Sch"},{"key":"1276_CR24","doi-asserted-by":"publisher","first-page":"238212052412386","DOI":"10.1177\/23821205241238641","volume":"11","author":"A Sumbal","year":"2024","unstructured":"Sumbal A, Sumbal R, Amir A. Can ChatGPT-3.5 Pass a medical exam? A systematic review of ChatGPT\u2019s performance in academic testing. J Med Educ Curric Dev. 2024;11:23821205241238641. https:\/\/journals.sagepub.com\/doi\/abs\/10.1177\/23821205241238641. doi: 10.1177\/23821205241238641%U.","journal-title":"J Med Educ Curric Dev"},{"key":"1276_CR25","doi-asserted-by":"publisher","first-page":"e49183","DOI":"10.2196\/49183","volume":"9","author":"CR Buhr","year":"2023","unstructured":"Buhr CR, Smith H, Huppertz T, Bahr-Hamm K, Matthias C, Blaikie A, et al. ChatGPT Versus consultants: blinded evaluation on answering otorhinolaryngology case-based questions. JMIR Med Educ. 2023;9:e49183. https:\/\/doi.org\/10.2196\/49183.","journal-title":"JMIR Med Educ"},{"issue":"3","key":"1276_CR26","doi-asserted-by":"publisher","first-page":"191","DOI":"10.4103\/tjem.tjem_277_24","volume":"25","author":"A Batur","year":"2025","unstructured":"Batur A. Analysis of factors affecting fatigue in emergency medicine residents: a nationwide, cross-sectional, descriptive study. Turk J Emerg Med. 2025;25(3):191\u20138. https:\/\/doi.org\/10.4103\/tjem.tjem_277_24.","journal-title":"Turk J Emerg Med"},{"issue":"3","key":"1276_CR27","doi-asserted-by":"publisher","first-page":"e10527","DOI":"10.1002\/aet2.10527","volume":"5","author":"LZ Vanyo","year":"2021","unstructured":"Vanyo LZ, Goyal DG, Dhaliwal RS, Sorge RM, Nelson LS, Beeson MS, et al. Emergency medicine resident burnout and examination performance. AEM Educ Train. 2021;5(3):e10527.","journal-title":"AEM Educ Train"},{"issue":"6","key":"1276_CR28","doi-asserted-by":"publisher","first-page":"589","DOI":"10.1001\/jamainternmed.2023.1838","volume":"183","author":"JW Ayers","year":"2023","unstructured":"Ayers JW, Poliak A, Dredze M, Leas EC, Zhu Z, Kelley JB, et al. Comparing physician and Artificial Intelligence Chatbot responses to patient questions posted to a public social media forum. JAMA Intern Med. 2023;183(6):589\u201396. https:\/\/doi.org\/10.1001\/jamainternmed.2023.1838.","journal-title":"JAMA Intern Med"},{"key":"1276_CR29","doi-asserted-by":"publisher","unstructured":"Buhr CR, Smith H, Huppertz T, Bahr-Hamm K, Matthias C, Cuny C et al. Assessing unknown potential\u2014quality and limitations of different large language models in the field of otorhinolaryngology. Acta Oto-Laryngologica.1\u20136. https:\/\/doi.org\/10.1080\/00016489.2024.2352843","DOI":"10.1080\/00016489.2024.2352843"}],"container-title":["Discover Artificial Intelligence"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s44163-026-01276-2","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44163-026-01276-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s44163-026-01276-2.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,16]],"date-time":"2026-06-16T22:45:17Z","timestamp":1781649917000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s44163-026-01276-2"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,26]]},"references-count":29,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2026,12]]}},"alternative-id":["1276"],"URL":"https:\/\/doi.org\/10.1007\/s44163-026-01276-2","relation":{},"ISSN":["2731-0809"],"issn-type":[{"value":"2731-0809","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,4,26]]},"assertion":[{"value":"21 July 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 April 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"26 April 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The local ethics committee of the University of Cologne waived ethical approval. This study was conducted in accordance with the principles of the Declaration of Helsinki. All participating physicians (junior residents and assessors) voluntarily agreed to take part in the study and provided written informed consent prior to participation.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval and consent to participate"}},{"value":"No identifiable patient\u2011level data are presented in the manuscript.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Consent for publication"}},{"value":"The results\/data\/figures of this study are not published or under consideration elsewhere.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Dual publication"}},{"value":"The authors declare no competing interests.","order":5,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"367"}}