{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T01:03:51Z","timestamp":1784163831856,"version":"3.55.0"},"reference-count":89,"publisher":"Springer Science and Business Media LLC","issue":"4","license":[{"start":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:00:00Z","timestamp":1784160000000},"content-version":"tdm","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"},{"start":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:00:00Z","timestamp":1784160000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0"}],"content-domain":{"domain":["link.springer.com"],"crossmark-restriction":false},"short-container-title":["AI Ethics"],"published-print":{"date-parts":[[2026,8]]},"abstract":"<jats:title>Abstract<\/jats:title>\n                  <jats:p>Large language models (LLMs) have achieved remarkable gains in cognitive performance through attention mechanisms functionally inspired by human attention. This paper asks philosophically whether a comparable architectural insight could technically advance moral processing. We argue that current alignment techniques primarily shape outputs after representations have been formed. They therefore cannot realise, using Iris Murdoch\u2019s loving attention approach, a just, reality-sensitive orientation toward others that operates at the level of representation. Drawing on Murdoch\u2019s moral philosophy, we identify three substrate-neutral features of loving attention, namely locational, representational, and dispositional, that survive translation from human moral phenomenology to computational systems. On this basis, we propose that moral processing in LLMs should be studied not only as a problem of output control but also as a problem of representational architecture. We further formulate the loving attention geometry hypothesis, suggesting that if morally improved perception involves a systematic shift from ego distorted to more just representations, this transformation may leave detectable structure in LLM embedding and activation spaces. With this focus on Murdoch, the paper contributes a novel philosophical technical approach that links moral attention, representation learning, and AI alignment. It clarifies why architectural considerations matter for moral AI, distinguishes between the tractability and validation of moral geometry, and outlines design constraints for responsible research. We argue that exploring representational forms of moral attention is a promising and necessary direction for AI ethics research. If transformer attention has been able to scale intelligence, it is worth inquiring if representations and architectures can scale morality.<\/jats:p>","DOI":"10.1007\/s43681-026-01239-4","type":"journal-article","created":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:45:34Z","timestamp":1784162734000},"update-policy":"https:\/\/doi.org\/10.1007\/springer_crossmark_policy","source":"Crossref","is-referenced-by-count":0,"title":["If attention scaled intelligence, can it scale morality?"],"prefix":"10.1007","volume":"6","author":[{"ORCID":"https:\/\/orcid.org\/0000-0002-8006-1617","authenticated-orcid":false,"given":"Gunter","family":"Bombaerts","sequence":"first","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Bram","family":"Delisse","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Uzay","family":"Kaymak","sequence":"additional","affiliation":[],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"297","published-online":{"date-parts":[[2026,7,16]]},"reference":[{"key":"1239_CR1","doi-asserted-by":"publisher","unstructured":"Abdulhai, M., Serapio-Garcia, G., Crepy, C., Valter, D., Canny, J., Jaques, N.: Moral Foundations of Large Language Models (arXiv:2310.15337). arXiv. (2023). https:\/\/doi.org\/10.48550\/arXiv.2310.15337","DOI":"10.48550\/arXiv.2310.15337"},{"issue":"1","key":"1239_CR2","doi-asserted-by":"publisher","first-page":"13","DOI":"10.1609\/aies.v7i1.31613","volume":"7","author":"C Akbulut","year":"2024","unstructured":"Akbulut, C., Weidinger, L., Manzini, A., Gabriel, I., Rieser, V.: All too human? Mapping and mitigating the risk from anthropomorphic AI. Proc. AAAI\/ACM Conf. AI Ethics Soc. 7(1), 13\u201326 (2024). https:\/\/doi.org\/10.1609\/aies.v7i1.31613","journal-title":"Proc. AAAI\/ACM Conf. AI Ethics Soc."},{"issue":"5","key":"1239_CR3","doi-asserted-by":"publisher","first-page":"1131","DOI":"10.1007\/s12671-019-01286-5","volume":"11","author":"B Analayo","year":"2020","unstructured":"Analayo, B.: Attention and mindfulness. Mindfulness. 11(5), 1131\u20131138 (2020). https:\/\/doi.org\/10.1007\/s12671-019-01286-5","journal-title":"Mindfulness"},{"key":"1239_CR4","doi-asserted-by":"crossref","unstructured":"Antonaccio, M.: Picturing the human: the moral thought of Iris Murdoch. Oxford University Press (2000)","DOI":"10.1093\/oso\/9780195131710.001.0001"},{"key":"1239_CR5","doi-asserted-by":"publisher","first-page":"105184","DOI":"10.1016\/j.knosys.2019.105184","volume":"191","author":"O Araque","year":"2020","unstructured":"Araque, O., Gatti, L., Kalimeri, K.: MoralStrength: exploiting a moral lexicon and embedding similarity for moral foundations prediction. Knowl. Based Syst. 191, 105184 (2020)","journal-title":"Knowl. Based Syst."},{"issue":"7729","key":"1239_CR6","doi-asserted-by":"publisher","first-page":"59","DOI":"10.1038\/s41586-018-0637-6","volume":"563","author":"E Awad","year":"2018","unstructured":"Awad, E., Dsouza, S., Kim, R., Schulz, J., Henrich, J., Shariff, A., Bonnefon, J.-F., Rahwan, I.: The moral machine experiment. Nature. 563(7729), 59\u201364 (2018)","journal-title":"Nature"},{"key":"1239_CR7","unstructured":"Bai, Y., Kadavath, S., Kundu, S., Askell, A., Kernion, J., Jones, A., Chen, A., Goldie, A., Mirhoseini, A., McKinnon, C.: Constitutional ai: harmlessness from ai feedback. arXiv Preprint arXiv:2212.08073. (2022). https:\/\/ai-plans.com\/file_storage\/4f32fa39-3a01-46c7-878e-c92b7aa7165f_2212.08073v1.pdf"},{"key":"1239_CR8","doi-asserted-by":"publisher","first-page":"101316","DOI":"10.1016\/j.cogsys.2024.101316","volume":"89","author":"P Bello","year":"2025","unstructured":"Bello, P., Bridewell, W.: Self-control on the path toward artificial moral agency. Cogn. Syst. Res. 89, 101316 (2025)","journal-title":"Cogn. Syst. Res."},{"key":"1239_CR9","doi-asserted-by":"publisher","unstructured":"Bender, E.M., Gebru, T., McMillan-Major, A., Shmitchell, S.: On the dangers of stochastic parrots: can language models be too big? \u5217. Proc. 2021 ACM Conf. Fairness Account. Transpar. 610\u2013623 (2021). https:\/\/doi.org\/10.1145\/3442188.3445922","DOI":"10.1145\/3442188.3445922"},{"key":"1239_CR10","doi-asserted-by":"crossref","unstructured":"Blodgett, S.L., Barocas, S., Daum\u00e9 Iii, H., Wallach, H.: Language (technology) is power: a critical survey of bias in NLP. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, 5454\u20135476. (2020). https:\/\/aclanthology.org\/2020.acl-main.485\/","DOI":"10.18653\/v1\/2020.acl-main.485"},{"issue":"4","key":"1239_CR11","doi-asserted-by":"publisher","first-page":"701","DOI":"10.1086\/293340","volume":"101","author":"L Blum","year":"1991","unstructured":"Blum, L.: Moral perception and particularity. Ethics. 101(4), 701\u2013725 (1991). https:\/\/doi.org\/10.1086\/293340","journal-title":"Ethics"},{"issue":"3","key":"1239_CR12","doi-asserted-by":"publisher","first-page":"343","DOI":"10.1007\/BF00353837","volume":"50","author":"LA Blum","year":"1986","unstructured":"Blum, L.A.: Iris murdoch and the domain of the moral. Philos. Stud. 50(3), 343\u2013367 (1986)","journal-title":"Philos. Stud."},{"issue":"4","key":"1239_CR13","doi-asserted-by":"publisher","first-page":"632","DOI":"10.1007\/s11747-020-00762-y","volume":"49","author":"M Blut","year":"2021","unstructured":"Blut, M., Wang, C., W\u00fcnderlich, N.V., Brock, C.: Understanding anthropomorphism in service provision: a meta-analysis of physical robots, chatbots, and other AI. J. Acad. Mark. Sci. 49(4), 632\u2013658 (2021). https:\/\/doi.org\/10.1007\/s11747-020-00762-y","journal-title":"J. Acad. Mark. Sci."},{"key":"1239_CR14","unstructured":"Bolukbasi, T., Chang, K.-W., Zou, J.Y., Saligrama, V., Kalai, A.T.: Man is to computer programmer as woman is to homemaker? Debiasing word embeddings. Advances in Neural Information Processing Systems, 29. (2016). https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2016\/hash\/a486cd07e4ac3d270571622f4f316ec5-Abstract.html"},{"key":"1239_CR15","doi-asserted-by":"publisher","unstructured":"Bombaerts, G.: Harnessing Ambiguity in Challenge-Based Learning as a Practice of Deep Engagement with Large Language Models. Digital Society, 4(3), 80. (2025). https:\/\/doi.org\/10.1007\/s44206-025-00233-3","DOI":"10.1007\/s44206-025-00233-3"},{"key":"1239_CR16","doi-asserted-by":"publisher","unstructured":"Bombaerts, G., & Botin, L.: From Individual Intentionality to Sympoiesis in System Phenomenology. Philosophy & Technology, 38(1), 33. (2025). https:\/\/doi.org\/10.1007\/s13347-025-00859-8","DOI":"10.1007\/s13347-025-00859-8"},{"key":"1239_CR17","doi-asserted-by":"crossref","unstructured":"Bombaerts, G.: Relational positionism: A constructive interpretation of morality in Luhmann\u2019s social systems theory. Kybernetes, 52(13), 29\u201344. (2023).","DOI":"10.1108\/K-03-2023-0429"},{"key":"1239_CR18","doi-asserted-by":"publisher","unstructured":"Bombaerts, G., Anderson, J., Dennis, M., Gerola, A., Frank, L., Hannes, T., Hopster, J., Marin, L., & Spahn, A.: Attention as Practice: Buddhist Ethics Responses to Persuasive Technologies. Global Philosophy, 33(2), 25. (2023). https:\/\/doi.org\/10.1007\/s10516-023-09680-4","DOI":"10.1007\/s10516-023-09680-4"},{"key":"1239_CR19","doi-asserted-by":"publisher","unstructured":"Bombaerts, G.: From Morality Installation in LLMs to LLMs in Morality-as-a-System (arXiv:2603.22944). arXiv. (2026). https:\/\/doi.org\/10.48550\/arXiv.2603.22944","DOI":"10.48550\/arXiv.2603.22944"},{"key":"1239_CR20","doi-asserted-by":"publisher","unstructured":"Bombaerts, G., Delisse, B., & Kaymak, U.: Morality in AI. A plea to embed morality in LLM architectures and frameworks (arXiv:2511.20689). arXiv. (2025). https:\/\/doi.org\/10.48550\/arXiv.2511.20689","DOI":"10.48550\/arXiv.2511.20689"},{"key":"1239_CR21","doi-asserted-by":"crossref","unstructured":"Bombaerts, G., Jenkins, K., Sanusi, Y. A., & Guoyu, W.: Expanding ethics justice across borders: The role of global philosophy. Energy Justice across Borders, 3\u201321. (2020).","DOI":"10.1007\/978-3-030-24021-9_1"},{"key":"1239_CR22","doi-asserted-by":"publisher","unstructured":"Bombaerts, G., Hannes, T., Adam, M., Aloisi, A., Anderson, J., Berger, L., Bettera, S. D., Campo, E., Candiotto, L., Panizza, S. C., Citton, Y., D\u00e2\u20ac\u2122Angelo, D., Dennis, M., Depraz, N., Doran, P., Drechsler, W., Duane, B., Edelglass, W., Eisenberger, I., \u2026 Zheng, Y.: From an attention economy to an ecology of attending. A manifesto (arXiv:2410.17421). arXiv. (2024). https:\/\/doi.org\/10.48550\/arXiv.2410.17421","DOI":"10.48550\/arXiv.2410.17421"},{"key":"1239_CR23","doi-asserted-by":"publisher","unstructured":"Bommasani, R., Hudson, D.A., Adeli, E., Altman, R., Arora, S., von Arx, S., Bernstein, M.S., Bohg, J., Bosselut, A., Brunskill, E., Brynjolfsson, E., Buch, S., Card, D., Castellon, R., Chatterji, N., Chen, A., Creel, K., Davis, J.Q., Demszky, D., Liang, P.: On the opportunities and risks of foundation models (arXiv:2108.07258). arXiv. (2022). https:\/\/doi.org\/10.48550\/arXiv.2108.07258","DOI":"10.48550\/arXiv.2108.07258"},{"key":"1239_CR24","doi-asserted-by":"publisher","unstructured":"Brunner, G., Liu, Y., Pascual, D., Richter, O., Ciaramita, M., Wattenhofer, R.: On Identifiability in transformers (arXiv:1908.04211). arXiv. (2020). https:\/\/doi.org\/10.48550\/arXiv.1908.04211","DOI":"10.48550\/arXiv.1908.04211"},{"key":"1239_CR25","doi-asserted-by":"publisher","unstructured":"Chefer, H., Gur, S., Wolf, L.: Transformer interpretability beyond attention visualization (arXiv:2012.09838). arXiv. (2021). https:\/\/doi.org\/10.48550\/arXiv.2012.09838","DOI":"10.48550\/arXiv.2012.09838"},{"issue":"3","key":"1239_CR26","doi-asserted-by":"publisher","first-page":"181","DOI":"10.1017\/S0140525X12000477","volume":"36","author":"A Clark","year":"2013","unstructured":"Clark, A.: Whatever next? Predictive brains, situated agents, and the future of cognitive science. Behav. Brain Sci. 36(3), 181\u2013204 (2013). https:\/\/doi.org\/10.1017\/S0140525X12000477","journal-title":"Behav. Brain Sci."},{"issue":"1","key":"1239_CR27","doi-asserted-by":"publisher","first-page":"59","DOI":"10.1080\/02691728.2025.2466164","volume":"40","author":"M Coeckelbergh","year":"2026","unstructured":"Coeckelbergh, M.: AI and epistemic agency: how ai influences belief revision and its normative implications. Soci Epistemol. 40(1), 59\u201371 (2026). https:\/\/doi.org\/10.1080\/02691728.2025.2466164","journal-title":"Soci Epistemol"},{"key":"1239_CR28","unstructured":"Crawford, M.B.: The world beyond your head: on becoming an individual in an age of distraction. Farrar, Straus and Giroux. (2015). https:\/\/books.google.com\/books?hl=en&lr=&id=uyilBAAAQBAJ&oi=fnd&pg=PP2&dq=the+world+beyond+your+head+crawford&ots=jnRd63NTxT&sig=vlthLLnUudDRcF2T6ZQPRJRfAvU"},{"key":"1239_CR29","unstructured":"De Beauvoir, S.: The ethics of ambiguity. Open Road Media. (2018). https:\/\/books.google.com\/books?hl=en&lr=&id=WhRZDwAAQBAJ&oi=fnd&pg=PT17&dq=ambiguity+ethics+beauvoir&ots=9_io5aYFnL&sig=13oWcmwUo_VbojXf1pof6piyk1Q"},{"key":"1239_CR30","first-page":"101434","volume":"11","author":"S Di Plinio","year":"2025","unstructured":"Di Plinio, S.: Panta Rh-AI: assessing multifaceted AI threats on human agency and identity. Social Sci. Humanit. Open. 11, 101434 (2025)","journal-title":"Social Sci. Humanit. Open."},{"issue":"6","key":"1239_CR31","doi-asserted-by":"publisher","first-page":"e70088","DOI":"10.1111\/psyp.70088","volume":"62","author":"S Di Plinio","year":"2025","unstructured":"Di Plinio, S., Aquino, A., Haddock, G., Alparone, F.R., Ebisch, S.J.H.: Brain and behavior in persuasion: the role of affective-cognitive matching. Psychophysiology. 62(6), e70088 (2025). https:\/\/doi.org\/10.1111\/psyp.70088","journal-title":"Psychophysiology"},{"key":"1239_CR32","doi-asserted-by":"publisher","unstructured":"Doris, J.M.: Lack of character: personality and moral behavior. Cambridge University Press. https:\/\/books.google.com\/books?hl=en&lr=&id=hnrQcJLaHVEC&oi=fnd&pg=PA1&dq=Doris,+J.+M.+(2002).+Lack+of+character:+Personality+and+moral+behavior.+Cambridge+University+Press. (2002). https:\/\/doi.org\/10.1017\/CBO9781139878364","DOI":"10.1017\/CBO9781139878364"},{"key":"1239_CR33","doi-asserted-by":"publisher","unstructured":"Fitz, S.: Do large GPT models discover moral dimensions in language representations? A topological study of sentence embeddings (arXiv:2309.09397). arXiv. (2023). https:\/\/doi.org\/10.48550\/arXiv.2309.09397","DOI":"10.48550\/arXiv.2309.09397"},{"issue":"4","key":"1239_CR34","doi-asserted-by":"publisher","first-page":"681","DOI":"10.1007\/s11023-020-09548-1","volume":"30","author":"L Floridi","year":"2020","unstructured":"Floridi, L., Chiriatti, M.: GPT-3: Its nature, scope, limits, and consequences. Mind. Mach. 30(4), 681\u2013694 (2020). https:\/\/doi.org\/10.1007\/s11023-020-09548-1","journal-title":"Mind. Mach."},{"key":"1239_CR35","doi-asserted-by":"crossref","unstructured":"Frankfurt, H.G.: The importance of what we care about: philosophical essays. Cambridge University Press. (1988). https:\/\/books.google.com\/books?hl=en&lr=&id=sSuqlGRrPRkC&oi=fnd&pg=PR7&dq=The+Importance+of+What+We+Care+About+frankfurt+harry&ots=Dj7CTOnurg&sig=pDnZElFJbUQ7ALnD5f4deJH-4vo","DOI":"10.1017\/CBO9780511818172"},{"issue":"1","key":"1239_CR36","doi-asserted-by":"publisher","first-page":"24","DOI":"10.1080\/00071773.2020.1836978","volume":"53","author":"A Fredriksson","year":"2022","unstructured":"Fredriksson, A., Panizza, S.: Ethical attention and the self in iris murdoch and maurice merleau-ponty. J. Br. Soc. Phenomenol. 53(1), 24\u201339 (2022). https:\/\/doi.org\/10.1080\/00071773.2020.1836978","journal-title":"J. Br. Soc. Phenomenol."},{"issue":"2","key":"1239_CR37","doi-asserted-by":"publisher","first-page":"127","DOI":"10.1038\/nrn2787","volume":"11","author":"K Friston","year":"2010","unstructured":"Friston, K.: The free-energy principle: a unified brain theory? Nat. Rev. Neurosci. 11(2), 127\u2013138 (2010)","journal-title":"Nat. Rev. Neurosci."},{"issue":"3","key":"1239_CR38","doi-asserted-by":"publisher","first-page":"411","DOI":"10.1007\/s11023-020-09539-2","volume":"30","author":"I Gabriel","year":"2020","unstructured":"Gabriel, I.: Artificial intelligence, values, and alignment. Mind. Mach. 30(3), 411\u2013437 (2020). https:\/\/doi.org\/10.1007\/s11023-020-09539-2","journal-title":"Mind. Mach."},{"key":"1239_CR39","unstructured":"Graham, J., Haidt, J., Koleva, S., Motyl, M., Iyer, R., Wojcik, S.P., Ditto, P.H.: Moral foundations theory: the pragmatic validity of moral pluralism. In: Advances in Experimental Social Psychology (Vol. 47, pp. 55\u2013130). Elsevier. (2013). https:\/\/www.sciencedirect.com\/science\/article\/pii\/B9780124072367000024?casa_token=FgYa7LrWB1QAAAAA:k2rAHxTnM2y_Q0B6bGue3aGtXbCry6W7GkwEQQvSn4TPVkMU9NLIHXXO-0bqsZPbezQIps1ekA"},{"issue":"2","key":"1239_CR40","doi-asserted-by":"publisher","first-page":"241","DOI":"10.1080\/14746700.2025.2472118","volume":"23","author":"M Graves","year":"2025","unstructured":"Graves, M.: Moral attention is all you need. Theol. Sci. 23(2), 241\u2013248 (2025). https:\/\/doi.org\/10.1080\/14746700.2025.2472118","journal-title":"Theol. Sci."},{"key":"1239_CR41","doi-asserted-by":"publisher","unstructured":"Hannes, T., & Bombaerts, G.: What does it mean that all is aflame? Non-axial Buddhist inspiration for an Anthropocene ontology. The Anthropocene Review, 10(3), 771\u2013786. (2023). https:\/\/doi.org\/10.1177\/20530196231153929","DOI":"10.1177\/20530196231153929"},{"issue":"4","key":"1239_CR42","doi-asserted-by":"publisher","first-page":"1051","DOI":"10.1007\/s43681-023-00305-5","volume":"4","author":"JKG Hopster","year":"2024","unstructured":"Hopster, J.K.G., Maas, M.M.: The technology triad: disruptive AI, regulatory gaps and value change. AI Ethics. 4(4), 1051\u20131069 (2024). https:\/\/doi.org\/10.1007\/s43681-023-00305-5","journal-title":"AI Ethics"},{"issue":"1","key":"1239_CR43","doi-asserted-by":"publisher","first-page":"299","DOI":"10.1007\/s00146-021-01167-3","volume":"37","author":"A Izzidien","year":"2022","unstructured":"Izzidien, A.: Word vector embeddings hold social ontological relations capable of reflecting meaningful fairness assessments. AI Soc. 37(1), 299\u2013318 (2022). https:\/\/doi.org\/10.1007\/s00146-021-01167-3","journal-title":"AI Soc."},{"key":"1239_CR44","first-page":"3543","volume":"1 Long and Shor","author":"S Jain","year":"2019","unstructured":"Jain, S., Wallace, B.C.: Attention is not explanation. Proc. 2019 Conf. North. Am. Chapter Association Comput. Linguistics: Hum. Lang. Technol. 1 Long and Short Papers, 3543\u20133556 (2019). https:\/\/aclanthology.org\/N19-1357\/","journal-title":"Proc. 2019 Conf. North. Am. Chapter Association Comput. Linguistics: Hum. Lang. Technol."},{"key":"1239_CR45","doi-asserted-by":"publisher","unstructured":"Jiang, L., Hwang, J.D., Bhagavatula, C., Bras, R.L., Liang, J., Dodge, J., Sakaguchi, K., Forbes, M., Borchardt, J., Gabriel, S., Tsvetkov, Y., Etzioni, O., Sap, M., Rini, R., Choi, Y.: Can machines learn morality? The delphi experiment (arXiv:2110.07574). arXiv. (2022). https:\/\/doi.org\/10.48550\/arXiv.2110.07574","DOI":"10.48550\/arXiv.2110.07574"},{"issue":"9","key":"1239_CR46","doi-asserted-by":"publisher","first-page":"389","DOI":"10.1038\/s42256-019-0088-2","volume":"1","author":"A Jobin","year":"2019","unstructured":"Jobin, A., Ienca, M., Vayena, E.: The global landscape of AI ethics guidelines. Nat. Mach. Intell. 1(9), 389\u2013399 (2019)","journal-title":"Nat. Mach. Intell."},{"key":"1239_CR47","unstructured":"Kahneman, D.: Thinking, fast and slow. macmillan (2011)"},{"key":"1239_CR48","doi-asserted-by":"publisher","first-page":"104696","DOI":"10.1016\/j.cognition.2021.104696","volume":"212","author":"B Kennedy","year":"2021","unstructured":"Kennedy, B., Atari, M., Davani, A.M., Hoover, J., Omrani, A., Graham, J., Dehghani, M.: Moral concerns are differentially observable in language. Cognition. 212, 104696 (2021)","journal-title":"Cognition"},{"issue":"2","key":"1239_CR49","doi-asserted-by":"publisher","first-page":"11","DOI":"10.14394\/filnau.2021.0007","volume":"29","author":"J Knobe","year":"2021","unstructured":"Knobe, J.: Philosophical intuitions are surprisingly stable across both demographic groups and situations. Filozofia Nauki. 29(2), 11\u201376 (2021)","journal-title":"Filozofia Nauki"},{"key":"1239_CR50","doi-asserted-by":"crossref","unstructured":"Lamberti, P. M., Bombaerts, G., & IJsselsteijn, W. (2025). Mind the gap: Bridging the divide between computer scientists and ethicists in shaping moral machines. Ethics and Information Technology, 27(1), 1\u201311.","DOI":"10.1007\/s10676-024-09806-1"},{"key":"1239_CR51","doi-asserted-by":"crossref","unstructured":"Latour, B.: Reassembling the social: an introduction to actor-network-theory. Oxford University Press (2005)","DOI":"10.1093\/oso\/9780199256044.001.0001"},{"key":"1239_CR52","unstructured":"Lave, J., Wenger, E.: Situated learning: legitimate peripheral participation. Cambridge university press. (1991). https:\/\/books.google.com\/books?hl=en&lr=&id=CAVIOrW3vYAC&oi=fnd&pg=PA11&dq=Lave,+J.,+%26+Wenger,+E.+(1991).+Situated+learning:+Legitimate+peripheral+participation.+Cambridge+University+Press.&ots=OEszqs1IDl&sig=OvaoB8th1JG81hYvCd9QYML0fXc"},{"key":"1239_CR53","unstructured":"Leshinskaya, A., Chakroff, A.: Value as semantics: representations of human moral and hedonic value in large language models. NeurIPS 2023 Workshop: AI Meets Moral Philosophy and Moral Psychology. (2023). https:\/\/ai.objectives.institute\/s\/Leshinskaya_Chakroff_2023_Value_as_Semantics.pdf"},{"key":"1239_CR54","doi-asserted-by":"publisher","first-page":"29","DOI":"10.3389\/fncom.2020.00029","volume":"14","author":"GW Lindsay","year":"2020","unstructured":"Lindsay, G.W.: Attention in psychology, neuroscience, and machine learning. Front. Comput. Neurosci. 14, 29 (2020)","journal-title":"Front. Comput. Neurosci."},{"key":"1239_CR55","doi-asserted-by":"crossref","unstructured":"Machery, E.: Philosophy within its proper bounds. Oxford University Press. (2017). https:\/\/books.google.com\/books?hl=en&lr=&id=DF4vDwAAQBAJ&oi=fnd&pg=PP1&dq=Machery,+E.+(2017).+Philosophy+within+its+proper+bounds.+Oxford+University+Press.&ots=QRKc5u6VvC&sig=8OK-_OfKuR6fkuwHvsY5xVKGL8Q","DOI":"10.1093\/oso\/9780198807520.001.0001"},{"key":"1239_CR56","doi-asserted-by":"crossref","unstructured":"MacIntyre, A.: Whose justice? Which rationality? In: The New Social Theory Reader (pp. 130\u2013137). Routledge. (2020). https:\/\/www.taylorfrancis.com\/chapters\/edit\/10.4324\/9781003060963-20\/whose-justice-rationality-alasdair-macintyre","DOI":"10.4324\/9781003060963-20"},{"key":"1239_CR57","unstructured":"McMahan, D.L.: The making of Buddhist modernism. Oxford University Press. (2008). https:\/\/books.google.com\/books?hl=en&lr=&id=GMNViIBK0wQC&oi=fnd&pg=PP8&dq=McMahan,+D.+L.+(2008).+The+making+of+Buddhist+modernism.+Oxford+University+Press.&ots=fJJ3CJcCFw&sig=5VEN3RZ1IE6zTPug9p-4byKlvPE"},{"issue":"4","key":"1239_CR58","doi-asserted-by":"publisher","first-page":"143","DOI":"10.1016\/j.tics.2006.12.007","volume":"11","author":"J Mikhail","year":"2007","unstructured":"Mikhail, J.: Universal moral grammar: theory, evidence and the future. Trends Cogn. Sci. 11(4), 143\u2013152 (2007)","journal-title":"Trends Cogn. Sci."},{"issue":"11","key":"1239_CR59","doi-asserted-by":"publisher","first-page":"501","DOI":"10.1038\/s42256-019-0114-4","volume":"1","author":"B Mittelstadt","year":"2019","unstructured":"Mittelstadt, B.: Principles alone cannot guarantee ethical AI. Nat. Mach. Intell. 1(11), 501\u2013507 (2019)","journal-title":"Nat. Mach. Intell."},{"issue":"2","key":"1239_CR60","doi-asserted-by":"publisher","first-page":"205395171667967","DOI":"10.1177\/2053951716679679","volume":"3","author":"BD Mittelstadt","year":"2016","unstructured":"Mittelstadt, B.D., Allo, P., Taddeo, M., Wachter, S., Floridi, L.: The ethics of algorithms: mapping the debate. Big Data Soc. 3(2), 2053951716679679 (2016). https:\/\/doi.org\/10.1177\/2053951716679679","journal-title":"Big Data Soc."},{"key":"1239_CR61","doi-asserted-by":"crossref","unstructured":"Murdoch, I.: On \u2018god\u2019 and \u2018good\u2019. In: The Sovereignty of Good, pp. 45\u201374. Routledge (2013a)","DOI":"10.4324\/9781003416586-2"},{"key":"1239_CR62","doi-asserted-by":"crossref","unstructured":"Murdoch, I.: The idea of perfection. In: The Sovereignty of Good, pp. 1\u201344. Routledge (2013b)","DOI":"10.4324\/9781003416586-1"},{"key":"1239_CR63","doi-asserted-by":"crossref","unstructured":"Murdoch, I.: The sovereignty of good over other concepts. In: The Sovereignty of Good, pp. 75\u2013102. Routledge (2013c)","DOI":"10.4324\/9781003416586-3"},{"key":"1239_CR64","unstructured":"Merleau-Ponty, M., Landes, D., Carman, T., Lefort, C.: Phenomenology of perception. Routledge. (2013). https:\/\/api.taylorfrancis.com\/content\/books\/mono\/download?identifierName=doi&identifierValue=10.4324\/9780203720714&type=googlepdf"},{"key":"1239_CR65","unstructured":"Mikolov, T., Yih, W., Zweig, G.: Linguistic regularities in continuous space word representations. Proceedings of the 2013 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, 746\u2013751. (2013). https:\/\/aclanthology.org\/N13-1090.pdf"},{"issue":"10","key":"1239_CR66","doi-asserted-by":"publisher","first-page":"516","DOI":"10.2307\/2026358","volume":"82","author":"M Nussbaum","year":"1985","unstructured":"Nussbaum, M.: Finely aware and richly responsible\u2019: moral attention and the moral task of literature. J. Philos. 82(10), 516\u2013529 (1985). https:\/\/doi.org\/10.2307\/2026358","journal-title":"J. Philos."},{"key":"1239_CR67","unstructured":"Nussbaum, M.C.: Love\u2019s knowledge: essays on philosophy and literature. OUP USA https:\/\/books.google.com\/books?hl=en (1990). &lr=&id=CXQ8DwAAQBAJ&oi=fnd&pg=PP1&dq=Nussbaum,+M.+C.+(1990).+Love%E2%80%99s+Knowledge:+Essays+on+Philosophy+and+Literature.+Oxford+University+Press.&ots=p7PuQlvpHv&sig=gkUqC2GrjfC2-8-Y9eNhWS4ruk4"},{"key":"1239_CR68","doi-asserted-by":"publisher","first-page":"27730","DOI":"10.52202\/068431-2011","volume":"35","author":"L Ouyang","year":"2022","unstructured":"Ouyang, L., Wu, J., Jiang, X., Almeida, D., Wainwright, C., Mishkin, P., Zhang, C., Agarwal, S., Slama, K., Ray, A.: Training language models to follow instructions with human feedback. Adv. Neural. Inf. Process. Syst. 35, 27730\u201327744 (2022)","journal-title":"Adv. Neural. Inf. Process. Syst."},{"key":"1239_CR69","unstructured":"Pan, A., Chan, J.S., Zou, A., Li, N., Basart, S., Woodside, T., Zhang, H., Emmons, S., Hendrycks, D.: Do the rewards justify the means? Measuring trade-offs between rewards and ethical behavior in the machiavelli benchmark. International Conference on Machine Learning, 26837\u201326867. (2023). https:\/\/proceedings.mlr.press\/v202\/pan23a.html"},{"issue":"2","key":"1239_CR70","doi-asserted-by":"publisher","first-page":"273","DOI":"10.1007\/s10790-019-09695-4","volume":"54","author":"S Panizza","year":"2020","unstructured":"Panizza, S.: Moral perception beyond supervenience: iris murdoch\u2019s radical perspective. J. Value Inq. 54(2), 273\u2013288 (2020). https:\/\/doi.org\/10.1007\/s10790-019-09695-4","journal-title":"J. Value Inq."},{"issue":"1","key":"1239_CR71","doi-asserted-by":"publisher","first-page":"73","DOI":"10.1146\/annurev-neuro-062111-150525","volume":"35","author":"SE Petersen","year":"2012","unstructured":"Petersen, S.E., Posner, M.I.: The attention system of the human brain: 20 years after. Annu. Rev. Neurosci. 35(1), 73\u201389 (2012). https:\/\/doi.org\/10.1146\/annurev-neuro-062111-150525","journal-title":"Annu. Rev. Neurosci."},{"issue":"4","key":"1239_CR72","doi-asserted-by":"publisher","first-page":"294","DOI":"10.1016\/j.tics.2018.01.009","volume":"22","author":"G Pezzulo","year":"2018","unstructured":"Pezzulo, G., Rigoli, F., Friston, K.J.: Hierarchical active inference: a theory of motivated control. Trends Cogn. Sci. 22(4), 294\u2013306 (2018)","journal-title":"Trends Cogn. Sci."},{"key":"1239_CR73","doi-asserted-by":"crossref","unstructured":"Qian, P., Qiu, X., Huang, X.-J.: Investigating language universal and specific properties in word embeddings. Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 1478\u20131488. (2016). https:\/\/aclanthology.org\/P16-1140.pdf","DOI":"10.18653\/v1\/P16-1140"},{"key":"1239_CR74","first-page":"15504","volume":"1: Long Papers","author":"N Rimsky","year":"2024","unstructured":"Rimsky, N., Gabrieli, N., Schulz, J., Tong, M., Hubinger, E., Turner, A.: Steering llama 2 via contrastive activation addition. Proc. 62nd Annual Meeting Association Comput. Linguist. 1: Long Papers, 15504\u201315522 (2024). https:\/\/aclanthology.org\/2024.acl-long.828.pdf","journal-title":"Proc. 62nd Annual Meeting Associ Comput. Linguist"},{"key":"1239_CR75","doi-asserted-by":"publisher","unstructured":"Selbst, A.D., Boyd, D., Friedler, S.A., Venkatasubramanian, S., Vertesi, J.: Fairness and abstraction in sociotechnical systems. Proc. Conf. Fairness Account. Transpar. 59\u201368 (2019). https:\/\/doi.org\/10.1145\/3287560.3287598","DOI":"10.1145\/3287560.3287598"},{"key":"1239_CR76","doi-asserted-by":"crossref","unstructured":"Serrano, S., Smith, N.A.: Is attention interpretable? Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, 2931\u20132951. (2019). https:\/\/aclanthology.org\/P19-1282\/","DOI":"10.18653\/v1\/P19-1282"},{"key":"1239_CR77","doi-asserted-by":"crossref","unstructured":"Susser, D., Roessler, B., Nissenbaum, H.: Online manipulation: Hidden influences in a digital world. (2019). https:\/\/philpapers.org\/rec\/SUSOMH","DOI":"10.2139\/ssrn.3306006"},{"key":"1239_CR78","unstructured":"Templeton, A., Conerly, T., Marcus, J., Lindsey, J., Bricken, T., Chen, B., Pearce, A., Citro, C., Ameisen, E., Jones, H.C.A.: Scaling monosemanticity: extracting interpretable features from claude 3 sonnet\u2014transformer-circuits. pub. (2024). Https:\/\/Transformer-Circuits. Pub\/2024\/Scaling-Monosemanticity\/Index. Html"},{"key":"1239_CR79","doi-asserted-by":"publisher","DOI":"10.5281\/zenodo.10908367","author":"L Teitelbaum","year":"2024","unstructured":"Teitelbaum, L., Simchon, A.: Data science for psychology: natural language (Version v0.1.0). Comput. Social Psychol. Lab. (2024). https:\/\/doi.org\/10.5281\/zenodo.10908367","journal-title":"Comput. Social Psychol. Lab."},{"key":"1239_CR80","unstructured":"Teitelbaum, L., Simchon, A.: Neural text embeddings in psychological research: a guide with examples in R. (2025). https:\/\/files.osf.io\/v1\/resources\/j9g4a_v1\/providers\/osfstorage\/682c5b6bcb8f1afaa8cacdc4?action=download&direct&version=1"},{"issue":"1","key":"1239_CR81","doi-asserted-by":"publisher","first-page":"107","DOI":"10.1007\/s13347-014-0156-9","volume":"28","author":"S Vallor","year":"2015","unstructured":"Vallor, S.: Moral deskilling and upskilling in a new machine age: reflections on the ambiguous future of character. Philos. Technol. 28(1), 107\u2013124 (2015)","journal-title":"Philos. Technol."},{"issue":"3","key":"1239_CR82","doi-asserted-by":"publisher","first-page":"385","DOI":"10.1007\/s11023-020-09537-4","volume":"30","author":"I van de Poel","year":"2020","unstructured":"van de Poel, I.: Embedding values in artificial intelligence (AI) systems. Mind. Mach. 30(3), 385\u2013409 (2020). https:\/\/doi.org\/10.1007\/s11023-020-09537-4","journal-title":"Mind. Mach."},{"issue":"3","key":"1239_CR83","doi-asserted-by":"publisher","first-page":"2071","DOI":"10.1007\/s43681-024-00536-0","volume":"5","author":"A Vijayaraghavan","year":"2025","unstructured":"Vijayaraghavan, A., Badea, C.: Minimum levels of interpretability for artificial moral agents. AI Ethics. 5(3), 2071\u20132087 (2025). https:\/\/doi.org\/10.1007\/s43681-024-00536-0","journal-title":"AI Ethics"},{"key":"1239_CR84","unstructured":"Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, \u0141., Polosukhin, I.: Attention is all you need. Advances in Neural Information Processing Systems, 30. (2017). https:\/\/proceedings.neurips.cc\/paper\/2017\/hash\/3f5ee243547dee91fbd053c1c4a845aa-Abstract.html"},{"key":"1239_CR85","unstructured":"Weil, S.: Attention and will. Gravity Grace, 116\u2013122. (1986)"},{"key":"1239_CR86","doi-asserted-by":"crossref","unstructured":"Widdows, H.: The moral vision of Iris Murdoch. Routledge (2017). https:\/\/www.taylorfrancis.com\/books\/mono\/10.4324\/9781315238265\/moral-vision-iris-murdoch-heather-widdows","DOI":"10.4324\/9781315238265"},{"key":"1239_CR87","unstructured":"Wong, D.B.: Natural moralities: a defense of pluralistic relativism. Oxford University Press (2009)"},{"key":"1239_CR88","doi-asserted-by":"publisher","unstructured":"Yaacov, D.-D.: Normative moral pluralism for AI: a framework for deliberation in complex moral contexts (arXiv:2508.08333). arXiv. (2025). https:\/\/doi.org\/10.48550\/arXiv.2508.08333","DOI":"10.48550\/arXiv.2508.08333"},{"key":"1239_CR89","doi-asserted-by":"publisher","unstructured":"Zou, A., Phan, L., Chen, S., Campbell, J., Guo, P., Ren, R., Pan, A., Yin, X., Mazeika, M., Dombrowski, A.-K., Goel, S., Li, N., Byun, M.J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Hendrycks, D.: Representation engineering: a top-down approach to ai transparency (arXiv:2310.01405). arXiv. (2025). https:\/\/doi.org\/10.48550\/arXiv.2310.01405","DOI":"10.48550\/arXiv.2310.01405"}],"container-title":["AI and Ethics"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s43681-026-01239-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/article\/10.1007\/s43681-026-01239-4","content-type":"text\/html","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/link.springer.com\/content\/pdf\/10.1007\/s43681-026-01239-4.pdf","content-type":"application\/pdf","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T00:45:41Z","timestamp":1784162741000},"score":1,"resource":{"primary":{"URL":"https:\/\/link.springer.com\/10.1007\/s43681-026-01239-4"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,7,16]]},"references-count":89,"journal-issue":{"issue":"4","published-print":{"date-parts":[[2026,8]]}},"alternative-id":["1239"],"URL":"https:\/\/doi.org\/10.1007\/s43681-026-01239-4","relation":{},"ISSN":["2730-5953","2730-5961"],"issn-type":[{"value":"2730-5953","type":"print"},{"value":"2730-5961","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,7,16]]},"assertion":[{"value":"23 January 2026","order":1,"name":"received","label":"Received","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"20 June 2026","order":2,"name":"accepted","label":"Accepted","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"16 July 2026","order":3,"name":"first_online","label":"First Online","group":{"name":"ArticleHistory","label":"Article History"}},{"value":"The authors declare no conflict of interest.","order":1,"name":"Ethics","label":"Conflict of interest","group":{"name":"EthicsHeading","label":"Declarations"}}],"article-number":"403"}}