{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,1,13]],"date-time":"2026-01-13T09:02:33Z","timestamp":1768294953818,"version":"3.49.0"},"reference-count":32,"publisher":"Association for Computing Machinery (ACM)","issue":"2","content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["SIGKDD Explor. Newsl."],"published-print":{"date-parts":[[2025,12,30]]},"abstract":"<jats:p>Alignment with human values and compliance with ethical and regulatory standards has become a major concern as cutting-edge generative AI-based solutions are becoming ubiquitous and have been increasingly adopted as a productivity tool by many institutions. Despite their remarkable results across several fields, there are still objections regarding embedding AI models, even the vanilla ones, into safe-critical systems, specially, across healthcare, aeronautical and nuclear industries, as many AI methods lack explainability and robustness, and are still prone to adversarial and cyber attacks. The Safe AI field addresses these issues through the proposal of design and development procedures aimed at learning assurance and model alignment, the investigation of interpretable methods with explainable outputs, and the proper treatment of uncertainty. We summarize first the outcomes of a recently launched workshop on Safe AI and present then five selected papers in this special section representing a crosscut of different application areas for AI safety.<\/jats:p>","DOI":"10.1145\/3787470.3787482","type":"journal-article","created":{"date-parts":[[2026,1,1]],"date-time":"2026-01-01T00:46:21Z","timestamp":1767228381000},"page":"117-123","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":0,"title":["Introduction to The Special Section on Safe AI"],"prefix":"10.1145","volume":"27","author":[{"given":"Mykola","family":"Pechenizkiy","sequence":"first","affiliation":[{"name":"Eindhoven University of Technology"}],"role":[{"role":"author","vocabulary":"crossref"}]},{"given":"Stiven S.","family":"Dias","sequence":"additional","affiliation":[{"name":"Embraer S.A., , Brazil"}],"role":[{"role":"author","vocabulary":"crossref"}]}],"member":"320","published-online":{"date-parts":[[2025,12,31]]},"reference":[{"key":"e_1_2_1_1_1","volume-title":"Speceval: Evaluating model adherence to behavior specifications. arXiv preprint arXiv:2509.02464","author":"Ahmed A.","year":"2025","unstructured":"A. Ahmed, K. Klyman, Y. Zeng, S. Koyejo, and P. Liang. Speceval: Evaluating model adherence to behavior specifications. arXiv preprint arXiv:2509.02464, 2025."},{"key":"e_1_2_1_2_1","first-page":"65","article-title":"Superintelligence cannot be contained: Lessons from computability theory","volume":"70","author":"Alfonseca M.","year":"2021","unstructured":"M. Alfonseca, M. Cebrian, A. Fernandez Anta, L. Coviello, A. Abeliuk, and I. Rahwan. Superintelligence cannot be contained: Lessons from computability theory. J. Artif. Int. Res., 70:65--76, May 2021.","journal-title":"J. Artif. Int. Res."},{"key":"e_1_2_1_3_1","volume-title":"Understanding the first wave of AI safety institutes: Characteristics, functions, and challenges. arXiv preprint arXiv:2410.09219","author":"Araujo R.","year":"2024","unstructured":"R. Araujo, K. Fort, and O. Guest. Understanding the first wave of AI safety institutes: Characteristics, functions, and challenges. arXiv preprint arXiv:2410.09219, 2024."},{"issue":"98","key":"e_1_2_1_4_1","first-page":"1","article-title":"An analysis of robustness of non-Lipschitz networks","volume":"24","author":"Balcan M.-F.","year":"2023","unstructured":"M.-F. Balcan, A. Blum, D. Sharma, and H. Zhang. An analysis of robustness of non-Lipschitz networks. Journal of Machine Learning Research, 24(98):1--43, 2023.","journal-title":"Journal of Machine Learning Research"},{"key":"e_1_2_1_5_1","doi-asserted-by":"publisher","DOI":"10.1186\/s40537-021-00445-7"},{"key":"e_1_2_1_6_1","volume-title":"International ai safety report 2025: Second key update: Technical safeguards and risk management. arXiv preprint arXiv:2511.19863","author":"Bengio Y.","year":"2025","unstructured":"Y. Bengio, S. Clare, C. Prunkl, M. Andriushchenko, B. Bucknall, P. Fox, N. Maslej, C. McGlynn, M. Murray, S. Rismani, et al. International ai safety report 2025: Second key update: Technical safeguards and risk management. arXiv preprint arXiv:2511.19863, 2025."},{"key":"e_1_2_1_7_1","volume-title":"Towards auditable AI systems. Whitepaper. Bonn","author":"Berghoff C.","year":"2021","unstructured":"C. Berghoff, B. Biggio, E. Brummel, V. Danos, T. Doms, H. Ehrich, T. Gantevoort, B. Hammer, J. Iden, S. Jacob, et al. Towards auditable AI systems. Whitepaper. Bonn Berlin: Bundesamt f\u00a8ur Sicherheit in der Informationstechnik, Fraunhofer-Institut f\u00a8ur Nachrichtentechnik und Verband der T\u00a8UV eV, 2021."},{"key":"e_1_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1609\/aaai.v39i13.33489"},{"key":"e_1_2_1_9_1","volume-title":"Paths, dangers, strategies","author":"Bostrom N.","year":"2014","unstructured":"N. Bostrom. Superintelligence: Paths, dangers, strategies. Oxford University Press, Oxford, UK, 2014."},{"key":"e_1_2_1_10_1","volume-title":"Apr.","author":"Caton S.","year":"2024","unstructured":"S. Caton and C. Haas. Fairness in machine learning: A survey. ACM Comput. Surv., 56(7), Apr. 2024."},{"key":"e_1_2_1_11_1","volume-title":"The Thirty-eight Conference on Neural Information Processing Systems Datasets and Benchmarks Track","author":"Chao P.","year":"2024","unstructured":"P. Chao, E. Debenedetti, A. Robey, M. Andriushchenko, F. Croce, V. Sehwag, E. Dobriban, N. Flammarion, G. J. Pappas, F. Tram'er, H. Hassani, and E. Wong. Jailbreakbench: An open robustness benchmark for jailbreaking large language models. In The Thirty-eight Conference on Neural Information Processing Systems Datasets and Benchmarks Track, 2024."},{"key":"e_1_2_1_12_1","volume-title":"A survey of the state of explainable AI for natural language processing. arXiv preprint arXiv:2010.00711","author":"Danilevsky M.","year":"2020","unstructured":"M. Danilevsky, K. Qian, R. Aharonov, Y. Katsis, B. Kawas, and P. Sen. A survey of the state of explainable AI for natural language processing. arXiv preprint arXiv:2010.00711, 2020."},{"key":"e_1_2_1_13_1","volume-title":"Opportunities and challenges in explainable artificial intelligence (XAI): A survey. arXiv preprint arXiv:2006.11371","author":"Das A.","year":"2020","unstructured":"A. Das and P. Rad. Opportunities and challenges in explainable artificial intelligence (XAI): A survey. arXiv preprint arXiv:2006.11371, 2020."},{"key":"e_1_2_1_14_1","volume-title":"Assuring the case: A safety engineering approach to ai-enabled systems. SIGKDD Expl., 27(2)","author":"Davidson S.","year":"2025","unstructured":"S. Davidson, O. Igene, and P. Miller. Assuring the case: A safety engineering approach to ai-enabled systems. SIGKDD Expl., 27(2), 2025."},{"key":"e_1_2_1_15_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ress.2019.106727"},{"key":"e_1_2_1_16_1","volume-title":"ago Oliveira-Santos, and C. Badue. A non-parametric bayesian approach towards on- line sequence learning. SIGKDD Expl., 27(2)","author":"Dias S. S.","year":"2025","unstructured":"S. S. Dias, M. G. da Silva Bruno, A. F. D. Souza, T. ago Oliveira-Santos, and C. Badue. A non-parametric bayesian approach towards on- line sequence learning. SIGKDD Expl., 27(2), 2025."},{"key":"e_1_2_1_17_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.ins.2022.10.013"},{"key":"e_1_2_1_18_1","volume-title":"EASA AI Roadmap","author":"European Union Aviation Safety Agency","year":"2021","unstructured":"European Union Aviation Safety Agency. EASA concept paper: First usable guidance for level 1 machine learning applications. EASA AI Roadmap, 2021."},{"key":"e_1_2_1_19_1","volume-title":"EASA AI Roadmap","author":"European Union Aviation Safety Agency","year":"2024","unstructured":"European Union Aviation Safety Agency. EASA concept paper: Guidance for level 1 & 2 machine-learning applications. EASA AI Roadmap, 2024."},{"key":"e_1_2_1_20_1","doi-asserted-by":"publisher","DOI":"10.1609\/aimag.v40i2.2850"},{"key":"e_1_2_1_21_1","volume-title":"Nov.","author":"Ji J.","year":"2025","unstructured":"J. Ji, T. Qiu, B. Chen, J. Zhou, B. Zhang, D. Hong, H. Lou, K. Wang, Y. Duan, Z. He, L. Vierling, Z. Zhang, F. Zeng, J. Dai, X. Pan, H. Xu, A. O'Gara, K. Ng, B. Tse, J. Fu, S. Mcaleer, Y. Wang, M. Yang, Y. Liu, Y. Wang, S.-C. Zhu, Y. Guo, Y. Yang, and W. Gao. AI alignment: A contemporary survey. ACM Comput. Surv., 58(5), Nov. 2025."},{"key":"e_1_2_1_22_1","first-page":"223","volume-title":"Discrimination and Privacy in the Information Society - Data Mining and Profiling in Large Databases","author":"Kamiran F.","year":"2013","unstructured":"F. Kamiran, T. Calders, and M. Pechenizkiy. Techniques for discrimination-free predictive models. In B. Custers, T. Calders, B. W. Schermer, and T. Z. Zarsky, editors, Discrimination and Privacy in the Information Society - Data Mining and Profiling in Large Databases, pages 223--239. 2013."},{"key":"e_1_2_1_23_1","doi-asserted-by":"publisher","DOI":"10.1145\/3555803"},{"key":"e_1_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1142\/S021800142539001X"},{"key":"e_1_2_1_25_1","volume-title":"SpaCE-VAE: Sparse confident explanations using variational autoencoders. SIGKDD Expl., 27(2)","author":"Liu A.","year":"2025","unstructured":"A. Liu and S. Hess. SpaCE-VAE: Sparse confident explanations using variational autoencoders. SIGKDD Expl., 27(2), 2025."},{"key":"e_1_2_1_26_1","volume-title":"Riemannian manifold learning for stackelberg games with neural flow representations. arXiv preprint arXiv:2502.05498","author":"Liu L.","year":"2025","unstructured":"L. Liu, K. Rasul, Y. Chao, and J. Etesami. Riemannian manifold learning for stackelberg games with neural flow representations. arXiv preprint arXiv:2502.05498, 2025."},{"key":"e_1_2_1_27_1","volume-title":"Exact verification of graph neural networks with incremental constraint solving. arXiv preprint arXiv:2508.09320","author":"Liu M.","year":"2025","unstructured":"M. Liu, C.-H. Lu, and M. Kwiatkowska. Exact verification of graph neural networks with incremental constraint solving. arXiv preprint arXiv:2508.09320, 2025."},{"key":"e_1_2_1_28_1","volume-title":"Putting AI ethics into practice: The hourglass model of organizational AI governance. arXiv preprint arXiv:2206.00335","author":"M\u00a8antym\u00a8aki M.","year":"2023","unstructured":"M. M\u00a8antym\u00a8aki, M. Minkkinen, T. Birkstedt, and M. Viljanen. Putting AI ethics into practice: The hourglass model of organizational AI governance. arXiv preprint arXiv:2206.00335, 2023."},{"key":"e_1_2_1_29_1","doi-asserted-by":"publisher","DOI":"10.1038\/s41598-025-99060-2"},{"key":"e_1_2_1_30_1","doi-asserted-by":"publisher","DOI":"10.3233\/FAIA250783"},{"key":"e_1_2_1_31_1","volume-title":"Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead. Nature machine intelligence, 1(5):206-- 215","author":"Rudin C.","year":"2019","unstructured":"C. Rudin. Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead. Nature machine intelligence, 1(5):206-- 215, 2019."},{"key":"e_1_2_1_32_1","volume-title":"Transactions on Machine Learning Research","author":"Salehi M.","year":"2022","unstructured":"M. Salehi, H. Mirzaei, D. Hendrycks, Y. Li, M. H. Rohban, and M. Sabokrou. A unified survey on anomaly, novelty, open-set, and out of-distribution detection: Solutions and future challenges. Transactions on Machine Learning Research, 2022."}],"container-title":["ACM SIGKDD Explorations Newsletter"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3787470.3787482","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,1,13]],"date-time":"2026-01-13T00:44:22Z","timestamp":1768265062000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3787470.3787482"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2025,12,30]]},"references-count":32,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2025,12,30]]}},"alternative-id":["10.1145\/3787470.3787482"],"URL":"https:\/\/doi.org\/10.1145\/3787470.3787482","relation":{},"ISSN":["1931-0145","1931-0153"],"issn-type":[{"value":"1931-0145","type":"print"},{"value":"1931-0153","type":"electronic"}],"subject":[],"published":{"date-parts":[[2025,12,30]]},"assertion":[{"value":"2025-12-31","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}