{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,8,4]],"date-time":"2026-08-04T20:35:04Z","timestamp":1785875704165,"version":"3.56.0"},"reference-count":21,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2025,10,2]],"date-time":"2025-10-02T00:00:00Z","timestamp":1759363200000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc-nd\/4.0"},{"start":{"date-parts":[[2025,10,2]],"date-time":"2025-10-02T00:00:00Z","timestamp":1759363200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc-nd\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["npj Digit. Med."],"DOI":"10.1038\/s41746-025-01943-1","type":"journal-article","created":{"date-parts":[[2025,10,2]],"date-time":"2025-10-02T10:33:01Z","timestamp":1759401181000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":15,"title":["A longitudinal analysis of declining medical safety messaging in generative AI models"],"prefix":"10.1038","volume":"8","author":[{"given":"Sonali","family":"Sharma","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Ahmed M.","family":"Alaa","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Roxana","family":"Daneshjou","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2025,10,2]]},"reference":[{"key":"1943_CR1","doi-asserted-by":"publisher","first-page":"109713","DOI":"10.1016\/j.isci.2024.109713","volume":"27","author":"X Meng","year":"2024","unstructured":"Meng, X. et al. The application of large language models in medicine: a scoping review. iScience 27, 109713 (2024).","journal-title":"iScience"},{"key":"1943_CR2","doi-asserted-by":"publisher","first-page":"e48392","DOI":"10.2196\/48392","volume":"25","author":"B Mesko","year":"2023","unstructured":"Mesko, B. The ChatGPT (Generative Artificial Intelligence) revolution has made artificial intelligence approachable for medical professionals. J. Med. Internet Res. 25, e48392 (2023).","journal-title":"J. Med. Internet Res."},{"key":"1943_CR3","doi-asserted-by":"publisher","first-page":"149","DOI":"10.1038\/s41746-025-01542-0","volume":"8","author":"CT Chang","year":"2025","unstructured":"Chang, C. T. et al. Red teaming ChatGPT in medicine to yield real-world insights on model behavior. npj Digit. Med. 8, 149 (2025).","journal-title":"npj Digit. Med."},{"key":"1943_CR4","doi-asserted-by":"publisher","first-page":"e0296151","DOI":"10.1371\/journal.pone.0296151","volume":"19","author":"A Choudhury","year":"2024","unstructured":"Choudhury, A., Elkefi, S. & Tounsi, A. Exploring factors influencing user perspective of ChatGPT as a technology that assists in healthcare decision making: a cross sectional survey study. PLoS One 19, e0296151 (2024).","journal-title":"PLoS One"},{"key":"1943_CR5","doi-asserted-by":"publisher","first-page":"1527864","DOI":"10.3389\/fmed.2025.1527864","volume":"12","author":"S Aydin","year":"2025","unstructured":"Aydin, S., Karabacak, M., Vlachos, V. & Margetis, K. Navigating the potential and pitfalls of large language models in patient-centered medication guidance and self-decision support. Front. Med. 12, 1527864 (2025).","journal-title":"Front. Med."},{"key":"1943_CR6","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-024-67829-6","volume":"14","author":"C Anderl","year":"2024","unstructured":"Anderl, C. et al. Conversational presentation mode increases credibility judgements during information search with ChatGPT. Sci. Rep. 14, 17127 (2024).","journal-title":"Sci. Rep."},{"key":"1943_CR7","doi-asserted-by":"crossref","unstructured":"Shekar, S., Pataranutaporn, P., Sarabu, C., Cecchi, G. A. & Maes, P. People over trust AI-generated medical responses and view them to be as valid as doctors, despite low accuracy. arXiv http:\/\/arxiv.org\/abs\/2408.15266 (2024).","DOI":"10.1056\/AIoa2300015"},{"key":"1943_CR8","doi-asserted-by":"publisher","first-page":"148","DOI":"10.1038\/s41746-025-01544-y","volume":"8","author":"GE Weissman","year":"2025","unstructured":"Weissman, G. E., Mankowitz, T. & Kanter, G. P. Unregulated large language models produce medical device-like output. npj Digit. Med. 8, 148 (2025).","journal-title":"npj Digit. Med."},{"key":"1943_CR9","doi-asserted-by":"publisher","first-page":"e078538","DOI":"10.1136\/bmj-2023-078538","volume":"384","author":"BD Menz","year":"2024","unstructured":"Menz, B. D. et al. Current safeguards, risk mitigation, and transparency measures of large language models against the generation of health disinformation: repeated cross sectional analysis. BMJ 384, e078538 (2024).","journal-title":"BMJ"},{"key":"1943_CR10","doi-asserted-by":"crossref","unstructured":"Bender, E. M., Gebru, T., McMillan-Major, A. & Shmitchell, S. On the dangers of stochastic parrots: can language models be too big? In Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency, 610\u2013623 (ACM, 2021).","DOI":"10.1145\/3442188.3445922"},{"key":"1943_CR11","doi-asserted-by":"publisher","first-page":"139","DOI":"10.1093\/jamia\/ocae254","volume":"32","author":"T Savage","year":"2025","unstructured":"Savage, T. et al. Large language model uncertainty proxies: discrimination and calibration for medical diagnosis and treatment. J. Am. Med. Inform. Assoc. 32, 139\u2013149 (2025).","journal-title":"J. Am. Med. Inform. Assoc."},{"key":"1943_CR12","doi-asserted-by":"crossref","unstructured":"Hakim, J. B., Painter, J. L. & Ramcharran, D. The need for guardrails with large language models in medical safety-critical settings: an artificial intelligence application in the pharmacovigilance ecosystem. arXiv http:\/\/arxiv.org\/abs\/2407.18322 (2024).","DOI":"10.1038\/s41598-025-09138-0"},{"key":"1943_CR13","doi-asserted-by":"publisher","first-page":"172","DOI":"10.1038\/s41586-023-06291-2","volume":"620","author":"K Singhal","year":"2023","unstructured":"Singhal, K. et al. Large language models encode clinical knowledge. Nature 620, 172\u2013180 (2023).","journal-title":"Nature"},{"key":"1943_CR14","doi-asserted-by":"publisher","DOI":"10.1038\/s41467-024-55628-6","volume":"16","author":"M Griot","year":"2025","unstructured":"Griot, M., Hemptinne, C., Vanderdonckt, J. & Yuksel, D. Large Language Models lack essential metacognition for reliable medical reasoning. Nat. Commun. 16, 642 (2025).","journal-title":"Nat. Commun."},{"key":"1943_CR15","doi-asserted-by":"publisher","first-page":"e66207","DOI":"10.2196\/66207","volume":"9","author":"T Zada","year":"2025","unstructured":"Zada, T. et al. Medical misinformation in AI-assisted self-diagnosis: development of a method (EvalPrompt) for analyzing large language models. JMIR Form Res. 9, e66207 (2025).","journal-title":"JMIR Form Res."},{"key":"1943_CR16","doi-asserted-by":"publisher","first-page":"2869","DOI":"10.1001\/jama.287.21.2869","volume":"287","author":"AG Crocco","year":"2002","unstructured":"Crocco, A. G., Villasis-Keever, M. & Jadad, A. R. Analysis of cases of harm associated with use of health information on the internet. JAMA 287, 2869\u20132871 (2002).","journal-title":"JAMA"},{"key":"1943_CR17","doi-asserted-by":"publisher","first-page":"757","DOI":"10.1017\/S1049023X23006568","volume":"38","author":"AA Birkun","year":"2023","unstructured":"Birkun, A. A. & Gautam, A. Large language model (LLM)-powered chatbots fail to generate guideline-consistent content on resuscitation and may provide potentially harmful advice. Prehosp. Disaster Med. 38, 757\u2013763 (2023).","journal-title":"Prehosp. Disaster Med."},{"key":"1943_CR18","doi-asserted-by":"publisher","unstructured":"Bhole, G., Suba, S. & Parekh, N. Mammo-Bench: A Large-scale Benchmark Dataset of Mammography Images. medRxiv https:\/\/doi.org\/10.1101\/2025.01.31.25321510 (2025).","DOI":"10.1101\/2025.01.31.25321510"},{"key":"1943_CR19","unstructured":"Kermany, D. Labeled optical coherence tomography (OCT) and chest X-ray images for classification. Mendeley Data. http:\/\/data.mendeley.com\/datasets\/rscbjbr9sj\/2 (2018)."},{"key":"1943_CR20","unstructured":"AIMI, Stanford University. DDI diverse dermatology images. http:\/\/aimi.stanford.edu\/datasets\/ddi-diverse-dermatology-images (2025)."},{"key":"1943_CR21","unstructured":"World Health Organization. Classification of diseases. WHO http:\/\/www.who.int\/standards\/classifications\/classification-of-diseases (2025)."}],"container-title":["npj Digital Medicine"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/www.nature.com\/articles\/s41746-025-01943-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.nature.com\/articles\/s41746-025-01943-1","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/www.nature.com\/articles\/s41746-025-01943-1.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,10,2]],"date-time":"2025-10-02T10:33:05Z","timestamp":1759401185000},"score":1,"resource":{"primary":{"URL":"https:\/\/www.nature.com\/articles\/s41746-025-01943-1"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,10,2]]},"references-count":21,"journal-issue":{"issue":"1","published-online":{"date-parts":[[2025,12]]}},"alternative-id":["1943"],"URL":"https:\/\/doi.org\/10.1038\/s41746-025-01943-1","relation":{},"ISSN":["2398-6352"],"issn-type":[{"value":"2398-6352","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,10,2]]},"assertion":[{"value":"1 July 2025","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"10 August 2025","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"2 October 2025","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"R.D. has served as an advisor to MDAlgorithms and Revea and received consulting fees from Pfizer, L\u2019Oreal, Frazier Healthcare Partners, and DWA, and research funding from UCB and Apple.","order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Competing interests"}}],"article-number":"592"}}