{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,21]],"date-time":"2026-06-21T14:49:20Z","timestamp":1782053360711,"version":"3.54.5"},"reference-count":44,"publisher":"Springer Science and Business Media LLC","issue":"1","license":[{"start":{"date-parts":[[2026,6,21]],"date-time":"2026-06-21T00:00:00Z","timestamp":1782000000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,6,21]],"date-time":"2026-06-21T00:00:00Z","timestamp":1782000000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"DOI":"10.13039\/501100004837","name":"Ministerio de Ciencia e Innovaci\u00f3n","doi-asserted-by":"publisher","award":["DOTT-HEALTH\/PAT-MED PID2019-106942RB-C31"],"award-info":[{"award-number":["DOTT-HEALTH\/PAT-MED PID2019-106942RB-C31"]}],"id":[{"id":"10.13039\/501100004837","id-type":"DOI","asserted-by":"publisher"}]},{"DOI":"10.13039\/501100003086","name":"Eusko Jaurlaritza","doi-asserted-by":"publisher","award":["IXA IT-1570-22"],"award-info":[{"award-number":["IXA IT-1570-22"]}],"id":[{"id":"10.13039\/501100003086","id-type":"DOI","asserted-by":"publisher"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Int J Data Sci Anal"],"published-print":{"date-parts":[[2026,12]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Preventive medicine aims to detect potential health risks early, thereby reducing morbidity and healthcare expenses. Beyond the mere anticipation of possible diagnoses, there is a demand for systems that justify predictions with clinical evidence and allow a quantitative assessment of the quality of the explanation. In this work, we present a novel approach: first, to address the prognosis of subsequent diagnoses in Spanish clinical data and, second, to generate explanations for each estimated diagnosis using generative models. By structuring patient data with the ICD-10 coding standard, and with the aid of cross-lingual data augmentation, substantial prognosis estimation improvements were achieved. We also put special attention to the assessment of the quality of the explanations generated. In this line, we propose the FIR framework (i.e. faithfulness, interpretability, robustness), in an attempt to seize the ability of each explanation to address core clinical questions relevant to the estimated ICD-10 diagnoses. Combining ICD-10 timelines coming from Osa (Spanish EHRs dataset) with MIMIC-IV (English EHRs dataset) significantly improved our Disease Risk Identifier\u2019s accuracy (by 11% in Spanish and 6% in English). Two LLMs (Mixtral and BioMistral) were used to generate explanations. Mixtral achieved higher faithfulness and interpretability (FidIn score of 74.59) compared to BioMistral (45.25), yet both maintained stable performance under minor textual perturbations. These findings demonstrate that leveraging ICD-10 diagnoses timelines coming from different data sources notably enhances next diagnosis predictions for Spanish EHRs. The proposed Disease Risk Explainer successfully produces clinically relevant textual justifications, and the FIR framework provides an objective metric to measure explanation quality.<\/jats:p>","DOI":"10.1007\/s41060-026-01167-w","type":"journal-article","created":{"date-parts":[[2026,6,21]],"date-time":"2026-06-21T14:37:44Z","timestamp":1782052664000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["Generative explainers in Spanish healthcare prognosis: a novel assessment framework"],"prefix":"10.1007","volume":"22","author":[{"ORCID":"https:\/\/orcid.org\/0000-0001-9093-4680","authenticated-orcid":false,"given":"Nuria","family":"Lebe\u00f1a","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Arantza","family":"Casillas","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Alicia","family":"P\u00e9rez","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,6,21]]},"reference":[{"key":"1167_CR1","volume-title":"International Statistical Classification of Diseases and Related Health Problems","author":"World Health Organization","year":"2004","unstructured":"World Health Organization: International Statistical Classification of Diseases and Related Health Problems, vol. 1. World Health Organization, Geneva (2004)"},{"key":"1167_CR2","doi-asserted-by":"crossref","unstructured":"Faggioli, G., Guazzo, A., Marchesin, S., Menotti, L., Trescato, I., Helena, A., Bergamaschi, R., Birolo, G., Cavalla, P., Chi\u00f2, A.: Overview of idpp@ clef 2023: the intelligent disease progression prediction challenge. In: CEUR workshop proceedings, vol. 3497, pp. 1123\u20131164. CEUR-WS (2023)","DOI":"10.1007\/978-3-031-42448-9_24"},{"key":"1167_CR3","unstructured":"Chiruzzo, L., Jim\u00e9nez-Zafra, S.M., Rangel, F.: Overview of iberlef 2024: natural language processing challenges for spanish and other Iberian languages. In: Proceedings of the Iberian Languages Evaluation Forum (IberLEF 2024), Co-located with the 40th Conference of the Spanish Society for Natural Language Processing (SEPLN 2024), CEUR-WS. Org (2024)"},{"issue":"1","key":"1167_CR4","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1186\/s12911-022-01857-y","volume":"22","author":"M Pishgar","year":"2022","unstructured":"Pishgar, M., Theis, J., Del Rios, M., Ardati, A., Anahideh, H., Darabi, H.: Prediction of unplanned 30-day readmission for ICU patients with heart failure. BMC Med. Inform. Decis. Mak. 22(1), 1\u201312 (2022)","journal-title":"BMC Med. Inform. Decis. Mak."},{"key":"1167_CR5","doi-asserted-by":"crossref","unstructured":"Harerimana, G., Kim, G.I., Kim, J.W., Jang, B.: HSGA: A hybrid lstm-cnn self-guided attention to predict the future diagnosis from discharge narratives. IEEE Access (2023)","DOI":"10.2139\/ssrn.4499140"},{"key":"1167_CR6","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1186\/1472-6963-13-269","volume":"13","author":"JF Orueta","year":"2013","unstructured":"Orueta, J.F., Nu\u00f1o-Solinis, R., Mateos, M., Vergara, I., Grandes, G., Esnaola, S.: Predictive risk modelling in the Spanish population: a cross-sectional study. BMC Health Serv. Res. 13, 1\u20139 (2013)","journal-title":"BMC Health Serv. Res."},{"key":"1167_CR7","doi-asserted-by":"publisher","first-page":"40846","DOI":"10.2196\/40846","volume":"25","author":"R Gonz\u00e1lez-Colom","year":"2023","unstructured":"Gonz\u00e1lez-Colom, R., Herranz, C., Vela, E., Monterde, D., Contel, J.C., Sis\u00f3-Almirall, A., Piera-Jim\u00e9nez, J., Roca, J., Cano, I.: Prevention of unplanned hospital admissions in multimorbid patients using computational modeling: observational retrospective cohort study. J. Med. Internet Res. 25, 40846 (2023)","journal-title":"J. Med. Internet Res."},{"issue":"2","key":"1167_CR8","doi-asserted-by":"publisher","first-page":"13","DOI":"10.26417\/ejis.v3i2.p13-25","volume":"3","author":"A Do\u00f1ate-Mart\u00ednez","year":"2017","unstructured":"Do\u00f1ate-Mart\u00ednez, A., R\u00f3denas-Rigla, F., Garc\u00e9s-Ferrer, J.: Development of a predictive model of elderly patients at risk of future hospital admission at primary care centres in valencia (spain). Eur J Interdiscip Stud 3(2), 13\u201325 (2017)","journal-title":"Eur J Interdiscip Stud"},{"key":"1167_CR9","doi-asserted-by":"crossref","unstructured":"Rosenfeld, A.: Better metrics for evaluating explainable artificial intelligence. In: Proceedings of the 20th International Conference on Autonomous Agents and Multiagent Systems, pp. 45\u201350. (2021)","DOI":"10.65109\/GNWD5518"},{"key":"1167_CR10","unstructured":"Union., E.: European General Data Protection Regulation (2018). https:\/\/commission.europa.eu\/law\/law-topic\/data-protection\/data-protection-eu_en"},{"key":"1167_CR11","doi-asserted-by":"crossref","unstructured":"Danilevsky, M., Qian, K., Aharonov, R., Katsis, Y., Kawas, B., Sen, P.: A survey of the state of explainable ai for natural language processing, (2020). arXiv preprint arXiv:2010.00711","DOI":"10.18653\/v1\/2020.aacl-main.46"},{"key":"1167_CR12","doi-asserted-by":"publisher","first-page":"100700","DOI":"10.1109\/ACCESS.2022.3207765","volume":"10","author":"A Theissler","year":"2022","unstructured":"Theissler, A., Spinnato, F., Schlegel, U., Guidotti, R.: Explainable ai for time series classification: a review, taxonomy and research directions. IEEE Access 10, 100700\u2013100724 (2022)","journal-title":"IEEE Access"},{"key":"1167_CR13","doi-asserted-by":"crossref","unstructured":"Blinov, P., Kokh, V.: Medical profile model: Scientific and practical applications in healthcare. IEEE J. Biomed. Health Inform. , (2023)","DOI":"10.1109\/JBHI.2023.3321132"},{"key":"1167_CR14","doi-asserted-by":"crossref","unstructured":"Peng, X., Long, G., Shen, T., Wang, S., Jiang, J., Zhang, C.: Bitenet: bidirectional temporal encoder network to predict medical outcomes. In: 2020 IEEE International Conference on Data Mining (ICDM), pp. 412\u2013421. IEEE (2020)","DOI":"10.1109\/ICDM50108.2020.00050"},{"key":"1167_CR15","doi-asserted-by":"crossref","unstructured":"Peng, X., Long, G., Shen, T., Wang, S., Jiang, J.: Self-attention enhanced patient journey understanding in healthcare system. In: Machine Learning and Knowledge Discovery in Databases: European Conference, ECML PKDD 2020, Ghent, Belgium, 2020, Proceedings, Part III, pp. 719\u2013735 (2021). Springer","DOI":"10.1007\/978-3-030-67664-3_43"},{"key":"1167_CR16","doi-asserted-by":"publisher","DOI":"10.1016\/j.imu.2025.101637","volume":"55","author":"N Lebe\u00f1a","year":"2025","unstructured":"Lebe\u00f1a, N., Casillas, A., P\u00e9rez, A.: Large language models aided patient progression documentation according to the ICD standard. Inf. Med. Unlocked 55, 101637 (2025)","journal-title":"Inf. Med. Unlocked"},{"key":"1167_CR17","doi-asserted-by":"crossref","unstructured":"Lebe\u00f1a, N., Borg, C., Casillas, A., P\u00e9rez, A.: Medical prognosis from electronic health records in Spanish. In: International Conference on Artificial Intelligence in Medicine, pp. 202\u2013211. Springer (2025)","DOI":"10.1007\/978-3-031-95838-0_20"},{"key":"1167_CR18","doi-asserted-by":"publisher","DOI":"10.1016\/j.ijmedinf.2021.104615","volume":"157","author":"O Trigueros","year":"2022","unstructured":"Trigueros, O., Blanco, A., Lebena, N., Casillas, A., P\u00e9rez, A.: Explainable ICD multi-label classification of EHRS in Spanish with convolutional attention. Int. J. Med. Informatics 157, 104615 (2022)","journal-title":"Int. J. Med. Informatics"},{"key":"1167_CR19","doi-asserted-by":"publisher","unstructured":"L\u00f3pez-Garc\u00eda, G., Jerez, J.M., Ribelles, N., Alba, E., Veredas, F.J.: Explainable clinical coding with in-domain adapted transformers. J. Biomed. Inf. 139, 104323 (2023) https:\/\/doi.org\/10.1016\/j.jbi.2023.104323","DOI":"10.1016\/j.jbi.2023.104323"},{"key":"1167_CR20","doi-asserted-by":"publisher","DOI":"10.1016\/j.compbiomed.2024.109127","volume":"182","author":"N Lebe\u00f1a","year":"2024","unstructured":"Lebe\u00f1a, N., P\u00e9rez, A., Casillas, A.: Quantifying decision support level of explainable automatic classification of diagnoses in Spanish medical records. Comput. Biol. Med. 182, 109127 (2024)","journal-title":"Comput. Biol. Med."},{"key":"1167_CR21","doi-asserted-by":"publisher","DOI":"10.1016\/j.knosys.2023.110866","volume":"278","author":"F Sovrano","year":"2023","unstructured":"Sovrano, F., Vitali, F.: An objective metric for explainable ai: how and why to estimate the degree of explainability. Knowl.-Based Syst. 278, 110866 (2023)","journal-title":"Knowl.-Based Syst."},{"key":"1167_CR22","unstructured":"Zhu, Y., Wang, Z., Gao, J., Tong, Y., An, J., Liao, W., Harrison, E.M., Ma, L., Pan, C.: Prompting large language models for zero-shot clinical prediction with structured longitudinal electronic health record data, (2024). arXiv preprint arXiv:2402.01713"},{"key":"1167_CR23","doi-asserted-by":"crossref","unstructured":"Hager, P., Jungmann, F., Bhagat, K., Hubrecht, I., Knauer, M., Vielhauer, J., Holland, R., Braren, R., Makowski, M., Kaisis, G., et al.: Evaluating and mitigating limitations of large language models in clinical decision making. medRxiv, 2024\u201301 (2024)","DOI":"10.1101\/2024.01.26.24301810"},{"key":"1167_CR24","unstructured":"Touvron, H., Martin, L., Stone, K., Albert, P., Almahairi, A., Babaei, Y., Bashlykov, N., Batra, S., Bhargava, P., Bhosale, S.: Llama 2: Open foundation and fine-tuned chat models, (2023). arXiv preprint arXiv:2307.09288"},{"key":"1167_CR25","doi-asserted-by":"crossref","unstructured":"K\u00f6pf, A., Kilcher, Y., R\u00fctte, D., Anagnostidis, S., Tam, Z.R., Stevens, K., Barhoum, A., Nguyen, D., Stanley, O., Nagyfi, R.: Openassistant conversations-democratizing large language model alignment. Adv. Neural. Inf. Process. Syst. 36, (2024)","DOI":"10.52202\/075280-2064"},{"key":"1167_CR26","unstructured":"Xu, C., Sun, Q., Zheng, K., Geng, X., Zhao, P., Feng, J., Tao, C., Jiang, D.: Wizardlm: Empowering large language models to follow complex instructions, (2023). arXiv preprint arXiv:2304.12244"},{"key":"1167_CR27","unstructured":"Toma, A., Lawler, P.R., Ba, J., Krishnan, R.G., Rubin, B.B., Wang, B.: Clinical camel: an open-source expert-level medical language model with dialogue-based knowledge encoding, CoRR (2023)"},{"key":"1167_CR28","unstructured":"Chen, Z., Cano, A.H., Romanou, A., Bonnet, A., Matoba, K., Salvi, F., Pagliardini, M., Fan, S., K\u00f6pf, A., Mohtashami, A.: Meditron-70b: Scaling medical pretraining for large language models, (2023). arXiv preprint arXiv:2311.16079"},{"key":"1167_CR29","unstructured":"Jiang, A.Q., Sablayrolles, A., Mensch, A., Bamford, C., Chaplot, D.S., Casas, D.d.l., Bressand, F., Lengyel, G., Lample, G., Saulnier, L., et al.: Mistral 7b. arXiv preprint arXiv:2310.06825 (2023)"},{"key":"1167_CR30","doi-asserted-by":"crossref","unstructured":"Wu, C., Lin, W., Zhang, X., Zhang, Y., Xie, W., Wang, Y.: Pmc-llama: toward building open-source language models for medicine. J. Am. Med. Inf. Assoc. 045 (2024)","DOI":"10.1093\/jamia\/ocae045"},{"key":"1167_CR31","doi-asserted-by":"crossref","unstructured":"Labrak, Y., Bazoge, A., Morin, E., Gourraud, P.-A., Rouvier, M., Dufour, R.: Biomistral: A collection of open-source pretrained large language models for medical domains, (2024). arXiv preprint arXiv:2402.10373","DOI":"10.18653\/v1\/2024.findings-acl.348"},{"key":"1167_CR32","doi-asserted-by":"crossref","unstructured":"Alonso, I., Oronoz, M., Agerri, R.: Medexpqa: Multilingual benchmarking of large language models for medical question answering, (2024). arXiv preprint arXiv:2404.05590","DOI":"10.2139\/ssrn.4780937"},{"key":"1167_CR33","unstructured":"Devlin, J., Chang, M.-W., Lee, K., Toutanova, K.: BERT: Pre-training of deep bidirectional transformers for language understanding, (2018). arXiv preprint arXiv:1810.04805"},{"key":"1167_CR34","doi-asserted-by":"crossref","unstructured":"Yang, X., Chen, A., PourNejatian, N., Shin, H.C., Smith, K.E., Parisien, C., Compas, C., Martin, C., Costa, A.B., Flores, M.G., et al.: A large language model for electronic health records. NPJ Dig. med. 5(1), 194 (2022)","DOI":"10.1038\/s41746-022-00742-2"},{"key":"1167_CR35","unstructured":"Jiang, A.Q., Sablayrolles, A., Roux, A., Mensch, A., Savary, B., Bamford, C., Chaplot, D.S., Casas, D.d.l., Hanna, E.B., Bressand, F., et al.: Mixtral of experts. arXiv preprint arXiv:2401.04088 (2024)"},{"key":"1167_CR36","doi-asserted-by":"crossref","unstructured":"Garc\u00eda-Ferrero, I., Agerri, R., Atutxa\u00a0Salazar, A., Cabrio, E., Iglesia, I., Lavelli, A., Magnini, B., Molinet, B., Ramirez-Romero, J., Rigau, G., Villa-Gonzalez, J.M., Villata, S., Zaninello, A.: Medical mT5: an open-source multilingual text-to-text LLM for the medical domain. arXiv:2404.07613 (2024)","DOI":"10.63317\/2kopd98p2rcz"},{"key":"1167_CR37","unstructured":"Honnibal, M., Montani, I.: spaCy: industrial-strength natural language processing in Python (2023). https:\/\/spacy.io\/"},{"key":"1167_CR38","doi-asserted-by":"crossref","unstructured":"Reimers, N., Gurevych, I.: Sentence-bert: sentence embeddings using siamese bert-networks. In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing (2019). https:\/\/arxiv.org\/abs\/1908.10084","DOI":"10.18653\/v1\/D19-1410"},{"issue":"4","key":"1167_CR39","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1145\/3519265","volume":"12","author":"F Sovrano","year":"2022","unstructured":"Sovrano, F., Vitali, F.: Generating user-centred explanations via illocutionary question answering: From philosophy to interfaces. ACM Trans. Interact. Intell. Syst. 12(4), 1\u201332 (2022)","journal-title":"ACM Trans. Interact. Intell. Syst."},{"key":"1167_CR40","unstructured":"Honnibal, M., Montani, I.: Spacy: Natural language understanding with bloom embeddings convolutional neural networks and incremental parsing. (2017). https:\/\/spacy.io"},{"key":"1167_CR41","unstructured":"Bird, S., Klein, E., Loper, E.: Natural language processing with Python. O\u2019Reilly Media, Inc., (2009). https:\/\/www.nltk.org"},{"key":"1167_CR42","unstructured":"Bond, F., Foster, R.: Linking and extending an open multilingual wordnet. In: Proceedings of the 51st annual meeting of the association for computational linguistics (2013). http:\/\/compling.hss.ntu.edu.sg\/omw\/"},{"key":"1167_CR43","unstructured":"Johnson, A., Bulgarelli, L., Pollard, T., Horng, S., Celi, L.A., Mark, R.: Mimic-iv. PhysioNet. Available online at: https:\/\/physionet.org\/content\/mimiciv\/1.0\/ (accessed August 23, 2021) (2020)"},{"key":"1167_CR44","first-page":"191498","volume":"12","author":"N Scarpato","year":"2024","unstructured":"Scarpato, N., Ferroni, P., Guadagni, F.: Xai unveiled: revealing the potential of explainable AI in medicine: a systematic review. IEEE Access 12, 191498\u2013191516 (2024)","journal-title":"IEEE Access"}],"container-title":["International Journal of Data Science and Analytics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s41060-026-01167-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s41060-026-01167-w","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s41060-026-01167-w.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,6,21]],"date-time":"2026-06-21T14:37:56Z","timestamp":1782052676000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s41060-026-01167-w"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,6,21]]},"references-count":44,"journal-issue":{"issue":"1","published-print":{"date-parts":[[2026,12]]}},"alternative-id":["1167"],"URL":"https:\/\/doi.org\/10.1007\/s41060-026-01167-w","relation":{},"ISSN":["2364-415X","2364-4168"],"issn-type":[{"value":"2364-415X","type":"print"},{"value":"2364-4168","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,6,21]]},"assertion":[{"value":"23 March 2026","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"17 May 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"21 June 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"The authors declare no conflict of interest.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}},{"value":"This article does not contain any studies with human participants or animals performed by any of the authors.","order":3,"name":"Ethics","group":{"name":"EthicsHeading","label":"Ethics approval"}},{"value":"During the preparation of this work, the authors used ChatGPT in order to assist with the writing of the paper. After using this tool\/service, the authors reviewed and edited the content as needed and take full responsibility for the content of the published article.","order":4,"name":"Ethics","group":{"name":"EthicsHeading","label":"Generative AI declaration"}}],"article-number":"202"}}