{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,5,1]],"date-time":"2026-05-01T16:07:47Z","timestamp":1777651667118,"version":"3.51.4"},"reference-count":43,"publisher":"Springer Science and Business Media LLC","issue":"8","license":[{"start":{"date-parts":[[2023,12,8]],"date-time":"2023-12-08T00:00:00Z","timestamp":1701993600000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2023,12,8]],"date-time":"2023-12-08T00:00:00Z","timestamp":1701993600000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"funder":[{"name":"Austrian Research Promotion Agency","award":["885377"],"award-info":[{"award-number":["885377"]}]},{"DOI":"10.13039\/501100002341","name":"Academy of Finland","doi-asserted-by":"crossref","award":["331197"],"award-info":[{"award-number":["331197"]}],"id":[{"id":"10.13039\/501100002341","id-type":"DOI","asserted-by":"crossref"}]},{"DOI":"10.13039\/501100002666","name":"Aalto University","doi-asserted-by":"crossref","id":[{"id":"10.13039\/501100002666","id-type":"DOI","asserted-by":"crossref"}]}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["Neural Comput &amp; Applic"],"published-print":{"date-parts":[[2024,3]]},"abstract":"<jats:title>Abstract<\/jats:title><jats:p>The successful application of machine learning (ML) methods increasingly depends on their interpretability or explainability. Designing explainable ML (XML) systems is instrumental for ensuring transparency of automated decision-making that targets humans. The explainability of ML methods is also an essential ingredient for trustworthy artificial intelligence. A key challenge in ensuring explainability is its dependence on the specific human end user of an ML system. The users of ML methods might have vastly different background knowledge about ML principles, with some having formal training in the specific field and others having none. We use information-theoretic concepts to develop a novel measure for the subjective explainability of predictions delivered by a ML method. We construct this measure via the conditional entropy of predictions, given the user signal. Our approach allows for a wide range of user signals, ranging from responses to surveys to biophysical measurements. We use this measure of subjective explainability as a regularizer for model training. The resulting explainable empirical risk minimization (EERM) principle strives to balance subjective explainability and risk. The EERM principle is flexible and can be combined with arbitrary ML models. We present several practical implementations of EERM for linear models and decision trees. Numerical experiments demonstrate the application of EERM to weather prediction and detecting inappropriate language in social media.<\/jats:p>","DOI":"10.1007\/s00521-023-09269-3","type":"journal-article","created":{"date-parts":[[2023,12,8]],"date-time":"2023-12-08T10:01:58Z","timestamp":1702029718000},"page":"3983-3996","update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":5,"title":["Explainable empirical risk minimization"],"prefix":"10.1007","volume":"36","author":[{"given":"Linli","family":"Zhang","sequence":"first","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Georgios","family":"Karakasidis","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Arina","family":"Odnoblyudova","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Leyla","family":"Dogruel","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"ORCID":"https:\/\/orcid.org\/0000-0001-9467-9410","authenticated-orcid":false,"given":"Yu","family":"Tian","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Alex","family":"Jung","sequence":"additional","affiliation":[],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"297","published-online":{"date-parts":[[2023,12,8]]},"reference":[{"key":"9269_CR1","unstructured":"High-Level Expert Group on AI (2019) Ethics guidelines for trustworthy AI. Technical report, European Comission"},{"issue":"1","key":"9269_CR2","doi-asserted-by":"publisher","first-page":"18","DOI":"10.3390\/e23010018","volume":"23","author":"P Linardatos","year":"2021","unstructured":"Linardatos P, Papastefanopoulos V, Kotsiantis S (2021) Explainable AI: a review of machine learning interpretability methods. Entropy 23(1):18. https:\/\/doi.org\/10.3390\/e23010018","journal-title":"Entropy"},{"issue":"5","key":"9269_CR3","doi-asserted-by":"publisher","first-page":"593","DOI":"10.3390\/electronics10050593","volume":"10","author":"J Zhou","year":"2021","unstructured":"Zhou J, Gandomi AH, Chen F, Holzinger A (2021) Evaluating the quality of machine learning explanations: a survey on methods and metrics. Electronics 10(5):593. https:\/\/doi.org\/10.3390\/electronics10050593","journal-title":"Electronics"},{"key":"9269_CR4","unstructured":"ISO (2020) Information technology\u2014artificial intelligence\u2014overview of trustworthiness in artificial intelligence, vol. ISO\/IEC TR 24028:2020(E), 1st edn. ISO\/IEC"},{"key":"9269_CR5","doi-asserted-by":"crossref","unstructured":"Ribeiro MT, Singh S, Guestrin C (2016) \u201cWhy should i trust you?\u201d: explaining the predictions of any classifier. In: Proceedings of the 22nd ACM SIGKDD, pp 1135\u20131144","DOI":"10.1145\/2939672.2939778"},{"key":"9269_CR6","doi-asserted-by":"publisher","first-page":"825","DOI":"10.1109\/LSP.2020.2993176","volume":"27","author":"A Jung","year":"2020","unstructured":"Jung A, Nardelli PHJ (2020) An information-theoretic approach to personalized explainable machine learning. IEEE Signal Process Lett 27:825\u2013829","journal-title":"IEEE Signal Process Lett"},{"key":"9269_CR7","doi-asserted-by":"publisher","DOI":"10.3389\/fdata.2021.688969","author":"V Belle","year":"2021","unstructured":"Belle V, Papantonis I (2021) Principles and practice of explainable machine learning. Front Big Data. https:\/\/doi.org\/10.3389\/fdata.2021.688969","journal-title":"Front. Big Data"},{"issue":"7","key":"9269_CR8","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1371\/journal.pone.0130140","volume":"10","author":"S Bach","year":"2015","unstructured":"Bach S, Binder A, Montavon G, Klauschen F, M\u00fcller K-R, Samek W (2015) On pixel-wise explanations for non-linear classifier decisions by layer-wise relevance propagation. PLoS ONE 10(7):1\u201346","journal-title":"PLoS ONE"},{"key":"9269_CR9","doi-asserted-by":"publisher","DOI":"10.1016\/j.media.2022.102364","volume":"77","author":"MS Ayhan","year":"2022","unstructured":"Ayhan MS, K\u00fcmmerle LB, K\u00fchlewein L, Inhoffen W, Aliyeva G, Ziemssen F, Berens P (2022) Clinical validation of saliency maps for understanding deep neural networks in ophthalmology. Med Image Anal 77:102364. https:\/\/doi.org\/10.1016\/j.media.2022.102364","journal-title":"Med Image Anal"},{"key":"9269_CR10","volume-title":"Semi-supervised learning","year":"2006","unstructured":"Chapelle O, Sch\u00f6lkopf B, Zien A (eds) (2006) Semi-supervised learning. The MIT Press, Cambridge"},{"key":"9269_CR11","doi-asserted-by":"publisher","DOI":"10.1007\/978-981-16-8193-6","volume-title":"Machine learning: the basics","author":"A Jung","year":"2022","unstructured":"Jung A (2022) Machine learning: the basics. Springer, HHH, Cham"},{"key":"9269_CR12","doi-asserted-by":"publisher","first-page":"1","DOI":"10.1016\/j.dsp.2017.10.011","volume":"73","author":"G Montavon","year":"2018","unstructured":"Montavon G, Samek W, M\u00fcller K (2018) Methods for interpreting and understanding deep neural networks. Digit Signal Process 73:1\u201315","journal-title":"Digit Signal Process"},{"issue":"9","key":"9269_CR13","doi-asserted-by":"publisher","first-page":"28","DOI":"10.1109\/MC.2018.3620965","volume":"51","author":"H Hagras","year":"2018","unstructured":"Hagras H (2018) Toward human-understandable, explainable AI. Computer 51(9):28\u201336","journal-title":"Computer"},{"issue":"5","key":"9269_CR14","doi-asserted-by":"publisher","first-page":"206","DOI":"10.1038\/s42256-019-0048-x","volume":"1","author":"C Rudin","year":"2019","unstructured":"Rudin C (2019) Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead. Nat Mach Intell 1(5):206\u2013215. https:\/\/doi.org\/10.1038\/s42256-019-0048-x","journal-title":"Nat Mach Intell"},{"key":"9269_CR15","unstructured":"Molnar C (2019) Interpretable machine learning: a guide for making black box models explainable. [online] Available: https:\/\/christophm.github.io\/interpretable-ml-book\/"},{"key":"9269_CR16","volume-title":"The elements of statistical learning. Springer series in statistics","author":"T Hastie","year":"2001","unstructured":"Hastie T, Tibshirani R, Friedman J (2001) The elements of statistical learning. Springer series in statistics. Springer, New York"},{"key":"9269_CR17","volume-title":"Elements of information theory","author":"TM Cover","year":"2006","unstructured":"Cover TM, Thomas JA (2006) Elements of information theory, 2nd edn. Wiley, Hoboken","edition":"2"},{"key":"9269_CR18","unstructured":"Chen J, Song L, Wainwright MJ, Jordan MI (2018) Learning to explain: an information-theoretic perspective on model interpretation. In: Proceedings of the 35th International conference on machine learning, Stockholm, Sweden"},{"key":"9269_CR19","volume-title":"Pattern recognition and machine learning","author":"CM Bishop","year":"2006","unstructured":"Bishop CM (2006) Pattern recognition and machine learning. Springer, Cham"},{"issue":"5","key":"9269_CR20","doi-asserted-by":"publisher","first-page":"699","DOI":"10.1109\/TPAMI.2005.93","volume":"27","author":"Y Zhang","year":"2005","unstructured":"Zhang Y, Ji Q (2005) Active and dynamic information fusion for facial expression understanding from image sequences. IEEE Trans Pattern Anal Mach Intell 27(5):699\u2013714","journal-title":"IEEE Trans Pattern Anal Mach Intell"},{"key":"9269_CR21","volume-title":"Deep learning","author":"I Goodfellow","year":"2016","unstructured":"Goodfellow I, Bengio Y, Courville A (2016) Deep learning. MIT Press, Cambridge"},{"issue":"85","key":"9269_CR22","first-page":"2825","volume":"12","author":"F Pedregosa","year":"2011","unstructured":"Pedregosa F (2011) Scikit-learn: machine learning in python. J Mach Learn Res 12(85):2825\u20132830","journal-title":"J Mach Learn Res"},{"key":"9269_CR23","doi-asserted-by":"publisher","DOI":"10.1201\/b18401","volume-title":"Statistical learning with sparsity: the lasso and its generalizations","author":"T Hastie","year":"2015","unstructured":"Hastie T, Tibshirani R, Wainwright M (2015) Statistical learning with sparsity: the lasso and its generalizations. CRC Press, Boca Raton"},{"key":"9269_CR24","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-642-20192-9","volume-title":"Statistics for high-dimensional data","author":"P B\u00fchlmann","year":"2011","unstructured":"B\u00fchlmann P, van de Geer S (2011) Statistics for high-dimensional data. Springer, New York"},{"key":"9269_CR25","doi-asserted-by":"publisher","DOI":"10.1017\/9781108627771","volume-title":"High-dimensional statistics: a non-asymptotic viewpoint","author":"M Wainwright","year":"2019","unstructured":"Wainwright M (2019) High-dimensional statistics: a non-asymptotic viewpoint. Cambridge University Press, Cambridge"},{"key":"9269_CR26","volume-title":"Nonlinear programming","author":"DP Bertsekas","year":"1999","unstructured":"Bertsekas DP (1999) Nonlinear programming, 2nd edn. Athena Scientific, Belmont","edition":"2"},{"key":"9269_CR27","doi-asserted-by":"publisher","DOI":"10.1017\/CBO9780511804441","volume-title":"Convex optimization","author":"S Boyd","year":"2004","unstructured":"Boyd S, Vandenberghe L (2004) Convex optimization. Cambridge University Press, Cambridge"},{"key":"9269_CR28","doi-asserted-by":"crossref","unstructured":"Davidson T, Warmsley D, Macy M, Weber I (2017) Automated hate speech detection and the problem of offensive language. In: Proceedings of the 11th international AAAI conference on web and social media (ICWSM), vol 11, no 1, pp 512\u2013515","DOI":"10.1609\/icwsm.v11i1.14955"},{"key":"9269_CR29","volume-title":"Introduction to probability","author":"DP Bertsekas","year":"2008","unstructured":"Bertsekas DP, Tsitsiklis JN (2008) Introduction to probability, 2nd edn. Athena Scientific, Belmont","edition":"2"},{"key":"9269_CR30","volume-title":"Theory of point estimation","author":"EL Lehmann","year":"1998","unstructured":"Lehmann EL, Casella G (1998) Theory of point estimation, 2nd edn. Springer, New York","edition":"2"},{"key":"9269_CR31","volume-title":"Gaussian processes for machine learning","author":"CE Rasmussen","year":"2006","unstructured":"Rasmussen CE, Williams CKI (2006) Gaussian processes for machine learning. MIT Press, Cambridge"},{"key":"9269_CR32","doi-asserted-by":"publisher","unstructured":"Wang X, Wei F, Liu X, Zhou M, Zhang M (2011) Topic sentiment analysis in twitter: a graph-based hashtag sentiment classification approach. In: Proceedings of the 20th ACM international conference on information and knowledge management. CIKM \u201911. Association for Computing Machinery, New York, NY, USA, pp 1031\u20131040. https:\/\/doi.org\/10.1145\/2063576.2063726","DOI":"10.1145\/2063576.2063726"},{"key":"9269_CR33","doi-asserted-by":"publisher","first-page":"3","DOI":"10.3389\/fdata.2020.00003","volume":"3","author":"SM Laaksonen","year":"2020","unstructured":"Laaksonen SM, Haapoja J, Kinnunen T, Nelimarkka M, P\u00f6yht\u00e4ri R (2020) The datafication of hate: expectations and challenges in automated hate speech monitoring. Front Big Data 3:3","journal-title":"Front Big Data"},{"key":"9269_CR34","doi-asserted-by":"crossref","unstructured":"Hardage D, Peyman N (2020) Hate and toxic speech detection in the context of Covid-19 pandemic using XAI: ongoing applied research. In: Proceedings of the 1st workshop on NLP for COVID-19 (Part 2) at EMNLP 2020","DOI":"10.18653\/v1\/2020.nlpcovid19-2.36"},{"key":"9269_CR35","unstructured":"Gagliardone I, Gal D, Alves T, Mart\u00ednez G (2015) Countering online hate speech. UNESCO"},{"issue":"6","key":"9269_CR36","doi-asserted-by":"publisher","first-page":"899","DOI":"10.1080\/15205436.2011.619679","volume":"15","author":"K Erjavec","year":"2012","unstructured":"Erjavec K, Kova\u010di\u010d MP (2012) You don\u2018t understand, this is a new war! Mass Commun Soc 15(6):899\u2013920","journal-title":"Mass Commun Soc"},{"key":"9269_CR37","doi-asserted-by":"publisher","DOI":"10.1007\/s40747-021-00561-0","author":"J Papcunov\u00e1","year":"2021","unstructured":"Papcunov\u00e1 J, Marton\u010dik M, Fed\u00e1kov\u00e1 D, Kento\u0161 M, Bozog\u00e1\u0148ov\u00e1 M, Srba I, Moro R, Pikuliak M, \u0160imko M, Adamkovi\u010d M (2021) Hate speech operationalization: a preliminary examination of hate speech indicators and their structure. Complex Intell Syst. https:\/\/doi.org\/10.1007\/s40747-021-00561-0","journal-title":"Complex Intell Syst"},{"key":"9269_CR38","doi-asserted-by":"publisher","unstructured":"Liao QV, Gruen D, Miller S (2020) Questioning the AI: informing design practices for explainable AI user experiences. In: Proceedings of the 2020 CHI conference on human factors in computing systems. CHI \u201920. Association for Computing Machinery, New York, NY, USA, pp 1\u201315. https:\/\/doi.org\/10.1145\/3313831.3376590","DOI":"10.1145\/3313831.3376590"},{"key":"9269_CR39","doi-asserted-by":"crossref","unstructured":"Bunde E (2021) AI-assisted and explainable hate speech detection for social media moderators: a design science approach. In: Proceedings of the 54th Hawaii international conference on systems sciences 2021","DOI":"10.24251\/HICSS.2021.154"},{"key":"9269_CR40","volume-title":"Modern information retrieval","author":"R Baeza-Yates","year":"2011","unstructured":"Baeza-Yates R, Ribeiro-Neto B (2011) Modern information retrieval. Addison Wesley, Boston"},{"issue":"3","key":"9269_CR41","doi-asserted-by":"publisher","first-page":"717","DOI":"10.1109\/TCDS.2020.3044366","volume":"13","author":"KJ Rohlfing","year":"2021","unstructured":"Rohlfing KJ, Cimiano P, Scharlau I, Matzner T, Buhl HM, Buschmeier H, Esposito E, Grimminger A, Hammer B, H\u00e4b-Umbach R, Horwath I, H\u00fcllermeier E, Kern F, Kopp S, Thommes K, Ngonga Ngomo A-C, Schulte C, Wachsmuth H, Wagner P, Wrede B (2021) Explanation as a social practice: toward a conceptual framework for the social design of AI systems. IEEE Trans Cogn Dev Syst 13(3):717\u2013728. https:\/\/doi.org\/10.1109\/TCDS.2020.3044366","journal-title":"IEEE Trans Cogn Dev Syst"},{"key":"9269_CR42","doi-asserted-by":"publisher","DOI":"10.14763\/2020.2.1469","author":"S Larsson","year":"2020","unstructured":"Larsson S, Heintz F (2020) Transparency in artificial intelligence. Internet Policy Rev. https:\/\/doi.org\/10.14763\/2020.2.1469","journal-title":"Internet Policy Rev"},{"key":"9269_CR43","doi-asserted-by":"publisher","first-page":"235","DOI":"10.1007\/s13218-020-00637-y","volume":"34","author":"K Sokol","year":"2020","unstructured":"Sokol K, Flach P (2020) One explanation does not fit all. KI-K\u00fcnstliche Intell 34:235\u2013250","journal-title":"KI-K\u00fcnstliche Intell"}],"container-title":["Neural Computing and Applications"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-023-09269-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s00521-023-09269-3\/fulltext.html","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s00521-023-09269-3.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2024,2,12]],"date-time":"2024-02-12T10:07:12Z","timestamp":1707732432000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s00521-023-09269-3"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2023,12,8]]},"references-count":43,"journal-issue":{"issue":"8","published-print":{"date-parts":[[2024,3]]}},"alternative-id":["9269"],"URL":"https:\/\/doi.org\/10.1007\/s00521-023-09269-3","relation":{},"ISSN":["0941-0643","1433-3058"],"issn-type":[{"value":"0941-0643","type":"print"},{"value":"1433-3058","type":"electronic"}],"subject":[],"published":{"date-parts":[[2023,12,8]]},"assertion":[{"value":"5 March 2023","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"6 November 2023","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"8 December 2023","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"order":1,"name":"Ethics","group":{"name":"EthicsHeading","label":"Declarations"}},{"value":"There are no conflicts of interest or competing interests.","order":2,"name":"Ethics","group":{"name":"EthicsHeading","label":"Conflict of interest"}}]}}