Insights into Moral Reasoning of AI: A Comparative Study Between Humans and Large Language Models

Journal of Media Ethics:1-15 (forthcoming)
  Copy   BIBTEX

Abstract

This study investigates the moral reasoning capabilities of large language models (LLMs), focusing on biases and the extent to which outputs reflect training data patterns rather than genuine reasoning. Using the Moral Competence Test (MCT) and the Moral Foundations Questionnaire (MFQ), we compared responses from human participants and LLM-based chatbots like ChatGPT. MCT results show that humans consistently outperform LLMs, indicating higher moral competence. MFQ responses from LLMs emphasize harm/care and fairness/reciprocity, but under-represent loyalty, authority, and purity. This pattern suggests a data-proportionality effect, where moral emphasis mirrors the prevalence of certain values in training data. Additionally, fine-tuning methods such as reinforcement learning with human feedback may amplify specific moral norms. These imbalances could unintentionally shape users’ moral intuitions and societal norms when LLMs are widely deployed. Our findings underscore the need for continuous auditing and alignment to ensure that LLMs provide ethically balanced and socially responsible guidance in morally sensitive applications.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

Beyond Ethical Alignment: Evaluating LLMs as Artificial Moral Assistants.Luca Alberto Rappuoli, Alessio Galatolo, Katie Winkle & Meriem Beloucif - 2025 - Proceedings of the 28Th European Conference on Artificial Intelligence (Ecai25) 413 (1):1213-1220.

Analytics

Added to PP
2025-09-15

Downloads
51 (#1,160,126)

6 months
41 (#207,492)

Historical graph of downloads
How can I increase my downloads?

References found in this work

No references found.

Add more references