{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2025,7,24]],"date-time":"2025-07-24T11:12:14Z","timestamp":1753355534998,"version":"3.40.5"},"reference-count":0,"publisher":"IOS Press","isbn-type":[{"type":"electronic","value":"9781643685335"}],"license":[{"start":{"date-parts":[[2024,8,22]],"date-time":"2024-08-22T00:00:00Z","timestamp":1724284800000},"content-version":"unspecified","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by-nc\/4.0\/"}],"content-domain":{"domain":[],"crossmark-restriction":false},"short-container-title":[],"published-print":{"date-parts":[[2024,8,22]]},"abstract":"<jats:p>Electronic Health Records (EHRs) contain a wealth of unstructured patient data, making it challenging for physicians to do informed decisions. In this paper, we introduce a Natural Language Processing (NLP) approach for the extraction of therapies, diagnosis, and symptoms from ambulatory EHRs of patients with chronic Lupus disease. We aim to demonstrate the effort of a comprehensive pipeline where a rule-based system is combined with text segmentation, transformer-based topic analysis and clinical ontology, in order to enhance text preprocessing and automate rules\u2019 identification. Our approach is applied on a sub-cohort of 56 patients, with a total of 750 EHRs written in Italian language, achieving an Accuracy and an F-score over 97% and 90% respectively, in the three extracted domains. This work has the potential to be integrated with EHR systems to automate information extraction, minimizing the human intervention, and providing personalized digital solutions in the chronic Lupus disease domain.<\/jats:p>","DOI":"10.3233\/shti240559","type":"book-chapter","created":{"date-parts":[[2024,8,23]],"date-time":"2024-08-23T09:49:31Z","timestamp":1724406571000},"source":"Crossref","is-referenced-by-count":3,"title":["A Comprehensive Natural Language Processing Pipeline for the Chronic Lupus Disease"],"prefix":"10.3233","author":[{"ORCID":"https:\/\/orcid.org\/0009-0005-3319-7211","authenticated-orcid":false,"given":"Livia","family":"Lilli","sequence":"first","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"},{"name":"Catholic University of the Sacred Heart, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-4837-447X","authenticated-orcid":false,"given":"Silvia Laura","family":"Bosello","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"given":"Laura","family":"Antenucci","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"},{"name":"Catholic University of the Sacred Heart, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-2765-5935","authenticated-orcid":false,"given":"Stefano","family":"Patarnello","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"given":"Augusta","family":"Ortolan","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-8366-1474","authenticated-orcid":false,"given":"Jacopo","family":"Lenkowicz","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0009-0008-1455-3884","authenticated-orcid":false,"given":"Marco","family":"Gorini","sequence":"additional","affiliation":[{"name":"AstraZeneca Italy, MIND, Milan, Italy"}]},{"given":"Gabriella","family":"Castellino","sequence":"additional","affiliation":[{"name":"AstraZeneca Italy, MIND, Milan, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-4687-0709","authenticated-orcid":false,"given":"Alfredo","family":"Cesario","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0000-0002-5347-0060","authenticated-orcid":false,"given":"Maria Antonietta","family":"D\u2019Agostino","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-6415-7267","authenticated-orcid":false,"given":"Carlotta","family":"Masciocchi","sequence":"additional","affiliation":[{"name":"Fondazione Policlinico Universitario Agostino Gemelli IRCCS, Rome, Italy"}]}],"member":"7437","container-title":["Studies in Health Technology and Informatics","Digital Health and Informatics Innovations for Sustainable Health Care Systems"],"original-title":[],"link":[{"URL":"https:\/\/ebooks.iospress.nl\/pdf\/doi\/10.3233\/SHTI240559","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,8,23]],"date-time":"2024-08-23T09:49:32Z","timestamp":1724406572000},"score":1,"resource":{"primary":{"URL":"https:\/\/ebooks.iospress.nl\/doi\/10.3233\/SHTI240559"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2024,8,22]]},"ISBN":["9781643685335"],"references-count":0,"URL":"https:\/\/doi.org\/10.3233\/shti240559","relation":{},"ISSN":["0926-9630","1879-8365"],"issn-type":[{"type":"print","value":"0926-9630"},{"type":"electronic","value":"1879-8365"}],"subject":[],"published":{"date-parts":[[2024,8,22]]}}}