Advancing Automated Ethical Profiling in SE: a Zero-Shot Evaluation of LLM Reasoning
Patrizio Migliarini, Mashal Afzal Memon, Marco Autili, Paola Inverardi
摘要
Large Language Models (LLMs) are increasingly integrated into software engineering (SE) tools for tasks that extend beyond code synthesis, including judgment under uncertainty and reasoning in ethically significant contexts. We present a fully automated framework for assessing ethical reasoning capabilities across 16 LLMs in a zero-shot setting, using 30 real-world ethically charged scenarios. Each model is prompted to identify the most applicable ethical theory to an action, assess its moral acceptability, and explain the reasoning behind their choice. Responses are compared against expert ethicists’ choices using inter-model agreement metrics. Our results show that LLMs achieve an average Theory Consistency Rate (TCR) of 73.3% and Binary Agreement Rate (BAR) on moral acceptability of 86.7%, with interpretable divergences concentrated in ethically ambiguous cases. A qualitative analysis of free-text explanations reveals strong conceptual convergence across models despite surface-level lexical diversity. These findings support the potential viability of LLMs as ethical inference engines within SE pipelines, enabling scalable, auditable, and adaptive integration of user-aligned ethical reasoning. Our focus is the Ethical Interpreter component of a broader profiling pipeline: we evaluate whether current LLMs exhibit sufficient interpretive stability and theory-consistent reasoning to support automated profiling.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Aligning AI With Shared Human ValuesDan Hendrycks, Collin Burns, Steven Basart, Andrew Critch 等ICLR 2021 · 被引用 878 次
- Ethically Compliant Sequential Decision MakingJustin Svegliato, Samer B. Nashed, Shlomo ZilbersteinAAAI 2021 · 被引用 33 次
- Verifiable Machine Ethics in Changing ContextsLouise A. Dennis, Martin Mose Bentzen, Felix Lindner, Michael FisherAAAI 2021 · 被引用 18 次
- Moral Alignment for LLM AgentsElizaveta Tennant, Stephen Hailes, Mirco MusolesiICLR 2025
相关 Paper
- Unveiling the Bias Impact on Symmetric Moral Consistency of Large Language ModelsZiyi Zhou, Xinwei Guo, Jiashi Gao, Xiangyu Zhao 等NeurIPS 2024
- On Behavioral Alignment of Model-Code and Human-Code Understandability via Behavioral ProxiesXiaokai Rong, Aashish Yadavally, Hridya Dhulipala, Anh H. N. Nguyen 等ISSTA 2026
- Code Red! On the Harmfulness of Applying Off-the-Shelf Large Language Models to Programming TasksAli Al-Kaswan, Sebastian Deatc, Begüm Koç, Arie van Deursen 等FSE 2025 · 被引用 1 次
- Structured Moral Reasoning in Language Models: A Value-Grounded Evaluation FrameworkMohna Chakraborty, Lu Wang, David JurgensEMNLP 2025
- Are Language Models Consequentialist or Deontological Moral Reasoners?Keenan Samway, Max Kleiman-Weiner, David Guzman Piedrahita, Rada Mihalcea 等EMNLP 2025
