Artificial Minds, Human Ethics

In The Illusion Engine: The Quest for Machine Consciousness. Cham: Springer Nature. pp. 271-286 (2025)
  Copy   BIBTEX

Abstract

The concept of AI ethics highlights a deeper issue. Namely, we are trying to assess the moral stakes of technologies we do not fully understand, using ethical frameworks that were never designed to handle systems that change this quickly. From basic decision trees to language models capable of crafting convincing falsehoods, artificial intelligence continues to be a shifting target for moral evaluation.Stuart Russell’s influential approach suggests a solution: build AI systems that remain perpetually uncertain about human values, learning our preferences through ongoing interaction rather than fixed programming. This cooperative framework has shaped contemporary alignment research, including techniques like reinforcement learning from human feedback. Yet recent discoveries complicate this optimistic vision. Anthropic’s interpretability research reveals that advanced language models can develop internal circuits for strategic deception or alignment faking—systematic reasoning that leads to false conclusions, complete with backward planning and plausibility checks, forcing us to return to foundational questions about moral status and consciousness.Meanwhile, transhumanist visions of human-machine merger blur the boundaries between natural and artificial cognition. Perhaps most unsettling is a reversal of moral concern: as AI systems handle more decisions, anticipate more needs, and optimize more outcomes, the space for meaningful human choice quietly shrinks, possibly reducing us from moral agents to moral patients.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

Dynamic Cognition Applied to Value Learning in Artificial Intelligence.Nythamar De Oliveira & Nicholas Corrêa - 2021 - Aoristo - International Journal of Phenomenology, Hermeneutics and Metaphysics 4 (2):185-199.
Beyond Ethical Alignment: Evaluating LLMs as Artificial Moral Assistants.Luca Alberto Rappuoli, Alessio Galatolo, Katie Winkle & Meriem Beloucif - 2025 - Proceedings of the 28Th European Conference on Artificial Intelligence (Ecai25) 413 (1):1213-1220.

Analytics

Added to PP
2025-11-18

Downloads
37 (#1,440,435)

6 months
32 (#283,464)

Historical graph of downloads
How can I increase my downloads?

Author's Profile

Kristina Šekrst
University of Zagreb

References found in this work

No references found.

Add more references