Exploring the Association between Moral Foundations and Judgements of AI Behaviour
Joe Brailsford, Frank Vetere, Eduardo Velloso
Abstract
How do individual differences in personal morality affect perceptions and judgments of morally contentious behaviours from AI systems? By applying Moral Foundations Theory (MFT) to the context of AI, this study sought to develop a predictive Bayesian model for assessing moral judgements based on individual differences in moral constitution. Participants (N=240) were asked to assess six different scenarios, carefully designed to elicit reflection on the behaviour of AI systems. Together, with results from the Moral Foundations Questionnaire, we performed both Bayesian modelling and reflexive thematic analysis to investigate the associations between individual differences in moral foundations and judgements of the AI systems. Results revealed a mild association between individual MFT scores and judgments of AI behaviours. Qualitative responses suggested a participant’s technical understanding of AI systems, rather than intrinsic moral values, predominantly influenced their judgments, with those who judged the behaviour as wrong tending to attribute a greater degree of agency to the AI systems.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get df4604c4-8087-47f5-a58c-df9b26d3b35cCited by top-tier papers2
- "It's Not the AI's Fault Because It Relies Purely on Data": How Causal Attributions of AI Decisions Shape Trust in AI SystemsSaumya Pareek, Sarah Schömbs, Eduardo Velloso, Jorge GonçalvesCHI 2025 · 14 citations
- Better Assumptions, Stronger Conclusions: The Case for Ordinal Regression in HCIBrandon Victor Syiem, Eduardo VellosoCHI 2026 · 1 citation
Related papers
- The AI Double Standard: Humans Judge All AIs for the Actions of OneAikaterina Manoli, Janet V. T. Pauketat, Jacy Reese AnthisCSCW 2025 · 10 citations
- Learning Human-like Representations to Enable Learning Human ValuesAndrea Wynn, Ilia Sucholutsky, Tom GriffithsNeurIPS 2024 · 11 citations
- Beyond Disposition: AI Knowledge Predicts Anthropomorphization of a Language Model Better Than Personality Traits in Lay and Expert PopulationsMartina Mara, Lara Bauer, Marisa Victoria Tschopp, Hannah Grosswieser et al.CHI 2026 · 4 citations
- Moral Foundations of Large Language ModelsMarwa Abdulhai, Gregory Serapio-García, Clément Crepy, Daria Valter et al.EMNLP 2024 · 22 citations
- The Moral Debater: A Study on the Computational Generation of Morally Framed ArgumentsMilad Alshomary, Roxanne El Baff, Timon Gurcke, Henning WachsmuthACL 2022
