Exploring large language models’ responses to moral reasoning dilemmas

Ethics and Behavior (forthcoming)
  Copy   BIBTEX

Abstract

This study investigates how various large language models (LLMs) generate responses to moral reasoning dilemmas. It specifically examines LLM-generated responses using the Defining Issues Test (DIT-2), which measures abstract moral reasoning schemas, and the Intermediate Concepts Measure (ICM Educational leaders’ version), which assesses domain-specific professional moral reasoning. For DIT-2, Claude prioritizes the highest post-conventional moral reasoning, followed by Gemini Advanced and Gemini. For the ICM Educational Leaders version, Gemini Advanced had the highest total ICM score, followed by Claude 3.5 Sonnet and Gemini. The findings indicate that some LLMs can generate responses consistent with sophisticated moral reasoning patterns, producing scores comparable to or exceeding graduate-level human participants; however, no direct comparisons with human participants were made in this study. This study provides a methodological framework for guiding larger-scale research into AI-generated and human moral reasoning patterns.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

Do negative mood states impact moral reasoning?Brian Barger & W. Pitt Derryberry - 2013 - Journal of Moral Education 42 (4):443-459.

Analytics

Added to PP
2025-12-16

Downloads
58 (#1,036,161)

6 months
46 (#184,227)

Historical graph of downloads
How can I increase my downloads?