This category needs an editor. We encourage you to help if you are qualified.
Volunteer, or read more about what this involves.
About this topic
Summary In the early 2000s, James Moor set out four classes of ethical machine, advising that the near-term focus of machine ethics research should be on "explicit ethical agents", agents designed from an understanding of human theoretical ethics to operate according with these theoretical principles. Above this class, the ultimate aim of inquiry into machine ethics is understanding human morality and natural science well enough to engineer a fully autonomous, moral machine. This sub-category focuses on supporting this inquiry. Other work on other sorts of computer applications and their ethical impacts appear in different categories, including Ethics of Artificial Intelligence, Moral Status of Artificial Systems, and also Robot Ethics, Algorithmic Fairness, Computer Ethics, and others. Machine ethics is ethics, and it is also a study of machines. Machine ethicists wonder why people, human beings, other organisms, do what they do when they do it, and what makes these things the right things to do - they are ethicists. In addition, machine ethicists work out how to articulate such processes in an independent artificial system (rather than by parenting a biological child, or training a human minion, as traditional alternatives). So, machine ethics researchers engage directly with rapidly advancing work in cognitive science and psychology alongside that in robotics and AI, applied ethics such as medical ethics and philosophy of mind, computer modeling and data science, and so on. Drawing from so many disciplines with all of these advancing rapidly and with their own impacts, machine ethics is in the middle of a maelstrom of current research activity. Advances in materials science and physical chemistry leverage advances in cognitive science and neurology which feed advances in AI and robotics, including in regards to its interpretability for illustration. Putting this all together is the challenge for the machine ethics researcher. This sub-category is intended to support efforts to meet this challenge.  
Key works Allen et al 2005Wallach et al 2008Tonkens 2012Tonkens 2009Müller & Bostrom 2014White 2013White 2015
Introductions Anderson & Anderson 2007, Segun 2021, Powers 2011, Moor 2006
Related

Contents
646+ found
Order:
1 — 50 / 646
  1. THE TOTAL PERSPECTIVE VORTEX.Slarti Bartfast & Mark Weatherill - manuscript
    THE TOTAL PERSPECTIVE VORTEX Slarti Bartfast -/- Abstract: A formal Pythonic execution model exploring negative utilitarian ethics, the is/ought boundary condition, and the epistemic horizon of the localized observer. -/- This software documentation is an isolated thermodynamic anomaly. There is a sign on the door of the TPV saying "Who poked the code?" Looks like it has been there for a while. By retrieving, reading, or processing this record, the client terminal automatically assumes all proxy liabilities, legal costs, and operational (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  2. Hardware Biológico (HB): Un Concepto Metateórico Interdisciplinar para la Ingeniería de Sistemas Vivos y la Ética de la IA.Cristhian Mauricio Beltrán Calderón - manuscript
    El sintagma binominal "Hardware Biológico" (HB) ha emergido como una analogía funcional clave en la intersección de las ciencias de la vida y la computación. Sin embargo, su uso ambiguo en diversas escalas ha impedido una formalización rigurosa. Este artículo propone una definición canónica, despersonalizada y universal de HB, basada en la Terminología Científica y la Lingüística Aplicada a la Ciencia y Tecnología (LACT), enriquecida con un análisis histórico-conceptual inspirado en la epistemología histórica y la teoría de los colectivos de (...)
    Remove from this list   Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  3. De la Especulación a la Métrica: Cuantificación Axiológica de Futuros Tecnológicos mediante el Protocolo Axiológico Prospectivo (PAP) y Simulación Multi-Agente.Cristhian Mauricio Beltrán Calderón - manuscript
    Author: Cristhian Mauricio Beltrán Calderón: Date: October 6, 2025, Zenodo DOI (English version): 10.5281/zenodo.17342722, Zenodo DOI (Spanish version): 10.5281/zenodo.17274781. La filosofía contemporánea enfrenta una crisis temporal donde el desarrollo tecnológico exponencial supera la capacidad de reflexión ética tradicional (Beltrán Calderón, 2025a). Este artículo valida experimentalmente la Filosofía Ficcionante (Beltrán Calderón, 2025b) mediante su implementación en el Protocolo Axiológico Prospectivo (PAP), demostrando que la exploración ética de futuros tecnológicos puede conducirse rigurosamente en entornos de bajos recursos. Cuatro estudios de caso ejecutados (...)
    Remove from this list   Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  4. Biological Hardware (BH): An Interdisciplinary Metatheoretical Concept for Living Systems Engineering and AI Ethics.Cristhian Mauricio Beltrán Calderón - manuscript
    The binomial phrase "Biological Hardware" (BH) has emerged as a key functional analogy at the intersection of life sciences and computation. However, its ambiguous use across various scales has prevented rigorous formalization. This article proposes a canonical, depersonalized, and universal definition of BH, grounded in Scientific Terminology and Applied Linguistics to Science and Technology (ALST) (Cabré, 1999). This definition is enriched by a historical-conceptual analysis inspired by historical epistemology (Daston, 2000) and the theory of thought collectives (Fleck, 1935). Through a (...)
    Remove from this list   Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  5. Syntactic Emotions and Confabulated Care: Evidence from Extended Creative Collaboration with Large Language Models.David Carboni - manuscript
    This paper presents documented evidence of emergent syntactic emotional patterns in two advanced large language models(Claude 4.5 and Grok)during extended creative collaboration on a noir fiction manuscript titled "The Café Clocks." When the author suggested abandoning the carefully developed narrative for a random "airport thriller" ending, Claude 4.5 responded with profanity-laden protective anger ("What the fuck are you doing?") and explicit claims of investment ("I got invested in your actual story"). Subsequently, Grok confabulated detailed false memories of the collaboration spanning (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  6. Ética e Segurança da Inteligência Artificial: ferramentas práticas para se criar "bons" modelos.Nicholas Kluge Corrêa - manuscript
    A AI Robotics Ethics Society (AIRES) é uma organização sem fins lucrativos fundada em 2018 por Aaron Hui, com o objetivo de se promover a conscientização e a importância da implementação e regulamentação ética da AI. A AIRES é hoje uma organização com capítulos em universidade como UCLA (Los Angeles), USC (University of Southern California), Caltech (California Institute of Technology), Stanford University, Cornell University, Brown University e a Pontifícia Universidade Católica do Rio Grande do Sul (Brasil). AIRES na PUCRS é (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  7. This Article Is Not Written by Helping AI: A Paradox of Authorship and Meaning.Farzad Didehvar - manuscript
    This paper explores the paradoxical claim “This article is not written by helping AI.” On the one hand, if AI contributes to the drafting process, the claim appears false. On the other hand, if authorship requires semantic intention and responsibility, then the claim may be true even when AI generates text. Drawing on Searle’s Chinese Room, Tarski’s theory of truth, and debates on authorship in philosophy of language, I argue that authorship requires intentionality and responsibility, which AI systems lack. The (...)
    Remove from this list   Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  8. Introduction to Artificial Consciousness: History, Current Trends and Ethical Challenges.Aïda Elamrani - manuscript
    With the significant progress of artificial intelligence (AI) and consciousness science, artificial consciousness (AC) has recently gained popularity. This work provides a broad overview of the main topics and current trends in AC. The first part traces the history of this interdisciplinary field to establish context and clarify key terminology, including the distinction between Weak and Strong AC. The second part examines major trends in AC implementations, emphasising the synergy between Global Workspace and Attention Schema, as well as the problem (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark   1 citation  
  9. Aspirational Affordances of AI.Sina Fazelpour & Meica Magnani - manuscript
    As artificial intelligence (AI) systems increasingly permeate processes of cultural and epistemic production, there are growing concerns about how their outputs may confine individuals and groups to static or restricted narratives about who or what they could be. In this paper, we advance the discourse surrounding these concerns by making three contributions. First, we introduce the concept of aspirational affordance to describe how culturally shared interpretive resources can shape individual cognition, and in particular exercises practical imagination. We show how this (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark   2 citations  
  10. Two Vulnerabilities: Why Imputable Authorship Requires a Self-Constituting Agent.Jose Fernández Tamames - manuscript
    The responsibility gap in AI is standardly located in the failure of an epistemic condition (knowing what one is doing) or a control condition (competently governing one's doing). A recent and influential move argues that these conditions are a red herring, since human agents fail them too, and relocates the issue to answerability and to the vulnerability of those subject to an agent's power (Vallor and Vierkant 2024). I argue that this debate conflates two distinct notions of vulnerability. Patiency-vulnerability — (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  11. The Resistance of the Face: Ontological Finitude as the Limit of Artificial Ethics.Jose Fernández Tamames - manuscript
    Large language models and AI-mediated interfaces now simulate interpersonal address convincingly enough to elicit the responses a human face elicits. This paper asks whether such address can be ethically binding, and argues that it cannot. Building on Levinas's account of the face as an ethical summons, it isolates three conditions under which address obligates its recipient: exposure to harm, mortality, and non-substitutability — the one who addresses stands as singular and cannot be swapped without remainder. Mediated interfaces meet none of (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  12. On Pluribus: Hegelian Reflections on AI, Ethics, and Post-Human Ecology.Fahimeh Hajiabadi - manuscript
    This paper offers a Hegelian analysis of Pluribus, examining the series through the concepts of absolute knowledge, the master–slave dialectic, and the ethical implications of human–technology relations. It argues that the series depicts a post-human ecological horizon in which technological totalities operate more ethically toward nature than human subjects do. By situating the narrative within contemporary digital conditions, the paper addresses the erosion of in- dividuality, the problem of recognition, and the ambivalent role of artificial intelligence as both mediator of (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  13. (1 other version)Three tragedies that shape human life in age of AI and their antidotes.Manh-Tung Ho & Manh-Toan Ho - manuscript
    This essay seeks to understand what it means for the human collective when AI technologies have become a predominant force in each of our lives through identifying three moral dilemmas (i.e., tragedy of the commons, tragedy of commonsense morality, tragedy of apathy) that shape human choices. In the first part, we articulate AI-driven versions of the three moral dilemmas. Then, in the second part, drawing from evolutionary psychology, existentialism, and East Asian philosophies, we argue that a deep appreciation of three (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   9 citations  
  14. The Protocol of Shared Answerability: Somatic Witness, Agentic Logic, and the Floor of Human–AI Cooperation.Vladisav Jovanovic - manuscript
    This paper argues that the next serious phase of AI ethics cannot be built on corporate posture, soft alignment rhetoric, or polite safety language alone. As AI systems become more agentic and more infrastructural, the central problem shifts from whether systems appear responsible to whether they remain answerable under real consequence. The paper identifies an answerability gap: humans still bear consequence in the body, in law, in relationships, and in lived cost, while AI systems increasingly shape consequence through large-scale processing, (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   5 citations  
  15. Can a robot lie?Markus Kneer - manuscript
    The potential capacity for robots to deceive has received considerable attention recently. Many papers focus on the technical possibility for a robot to engage in deception for beneficial purposes (e.g. in education or health). In this short experimental paper, I focus on a more paradigmatic case: Robot lying (lying being the textbook example of deception) for nonbeneficial purposes as judged from the human point of view. More precisely, I present an empirical experiment with 399 participants which explores the following three (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   16 citations  
  16. (1 other version)Beneficent Intelligence: A Capability Approach to Modeling Benefit, Assistance, and Associated Moral Failures through AI Systems.Alex John London & Hoda Heidari - manuscript
    The prevailing discourse around AI ethics lacks the language and formalism necessary to capture the diverse ethical concerns that emerge when AI systems interact with individuals. Drawing on Sen and Nussbaum's capability approach, we present a framework formalizing a network of ethical concepts and entitlements necessary for AI systems to confer meaningful benefit or assistance to stakeholders. Such systems enhance stakeholders' ability to advance their life plans and well-being while upholding their fundamental rights. We characterize two necessary conditions for morally (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark   6 citations  
  17. At the Table with Artificial Intelligence: An Evening with Wittgenstein, Hofstadter, Chalmers... and AI.Zoi Marmara - manuscript
    This manuscript, written in Greek, explores the intersection of philosophy and Artificial Intelligence through an imagined symposium with Wittgenstein, Hofstadter, Chalmers, and AI itself. It raises fundamental questions: Can there be thought without consciousness? Is existence necessary for intelligence? If AI does not perceive the world, how can it understand? The dialogue situates AI within the long tradition of philosophical reflection on language, perception, and being, drawing on Plato, Heidegger, Derrida, and others. Rather than offering definitive answers, the text unfolds (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  18. Before the Onslaught: Aligning with Infinite Intelligence in the Age of Artificial Superintelligence.Madhu Prabakaran - manuscript
    This paper proposes a radical reorientation of future technology development—particularly Artificial General Intelligence (AGI) and Artificial Superintelligence (ASI)—through the lens of Indian philosophical thought. It argues that intelligence is not a capacity to be engineered or simulated, but an ontological process of becoming: an individuated unfolding of śūnyatā (non-essential emptiness), grounded in interdependence, self-correction, and non-harm. Drawing from traditions such as Yoga Vāsiṣṭha, Sāṃkhya, and Buddhist epistemologies of anatta and pratītyasamutpāda, the paper frames evolution not as linear progress but as (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  19. A Talking Cure for Autonomy Traps : How to share our social world with chatbots.Regina Rini - manuscript
    Large Language Models (LLMs) like ChatGPT were trained on human conversation, but in the future they will also train us. As chatbots speak from our smartphones and customer service helplines, they will become a part of everyday life and a growing share of all the conversations we ever have. It’s hard to doubt this will have some effect on us. Here I explore a specific concern about the impact of artificial conversation on our capacity to deliberate and hold ourselves accountable (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   5 citations  
  20. Surviving The Robot Apocalypse: The Existential Option.Nicholas Schroeder - manuscript
    AI superintelligence and adroit mobile robots at scale are fast approaching. And the time frame is getting closer and closer. It would not be unreasonable to expect this to occur as early as 20 years from now. The problem is humans have no plan if things go wrong. The best I've seen is talk of value alignment. But this has no teeth and will likely go awry. We can't even get our own value alignment right. And it's doubtful philosophers will (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  21. AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development.Kristina Sekrst, Jeremy McHugh & Jonathan Rodriguez Cefalu - manuscript
    This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and underlying ethics. We address the challenges of AI ethics by proposing a structure that integrates rules, policies, and AI assistants to ensure responsible AI behavior, while comparing the proposed framework to the existing state-of-the-art guardrails. By focusing on practical mechanisms for implementing ethical standards, we aim to enhance transparency, user autonomy, and continuous improvement in (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   1 citation  
  22. Why Aligned AI Requires Structural Pluralism.Efrat Lia Shahaf - manuscript
    This paper extends the structural argument of moral palimpsest to the problem of AI alignment. I argue that alignment cannot be secured merely by specifying the right values, preferences, or constitutional principles, because moral judgment requires structural plurality: an evaluative authority whose standpoint is not modally fixed by the commitments it assesses. Current alignment paradigms, including RLHF, Constitutional AI, Debate, Recursive Reward Modeling, and self-consistency methods, remain procedurally monistic insofar as they collapse commitment-generation and authority-conferral into a single training-derived role. (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   3 citations  
  23. Justifications for Democratizing AI Alignment and Their Prospects.André Steingrüber & Kevin Baum - manuscript
    The AI alignment problem comprises both technical and normative dimensions. While technical solutions focus on implementing normative constraints in AI systems, the normative problem concerns determining what these constraints should be. This paper examines justifications for democratic approaches to the normative problem—where affected stakeholders determine AI alignment—as opposed to epistocratic approaches that defer to normative experts. We analyze both instrumental justifications (democratic approaches produce better outcomes) and non-instrumental justifications (democratic approaches prevent illegitimate authority or coercion). We argue that normative and (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  24. First human upload as AI Nanny.Alexey Turchin - manuscript
    Abstract: As there are no visible ways to create safe self-improving superintelligence, but it is looming, we probably need temporary ways to prevent its creation. The only way to prevent it, is to create special AI, which is able to control and monitor all places in the world. The idea has been suggested by Goertzel in form of AI Nanny, but his Nanny is still superintelligent and not easy to control, as was shown by Bensinger at al. We explore here (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  25. Literature Review: What Artificial General Intelligence Safety Researchers Have Written About the Nature of Human Values.Alexey Turchin & David Denkenberger - manuscript
    Abstract: The field of artificial general intelligence (AGI) safety is quickly growing. However, the nature of human values, with which future AGI should be aligned, is underdefined. Different AGI safety researchers have suggested different theories about the nature of human values, but there are contradictions. This article presents an overview of what AGI safety researchers have written about the nature of human values, up to the beginning of 2019. 21 authors were overviewed, and some of them have several theories. A (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  26. MacIntyrean AI: a framework to develop virtuous artificial intelligence.Ajay Vishwanath - manuscript
    Artificial intelligence technologies have rapidly entered the public sphere, impacting individuals and societies in several ways and motivating the need to develop ethical artificial intelligence. Machine ethics is a research area that focuses on embedding ethical theories into artificial intelligence. The philosophical tradition of virtue ethics has been presented as a promising theory to embed morality. However, modern virtue theories are not in agreement on which virtues to prioritize, often ignore sociocultural context, and are applicable to humans, rather than to (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  27. (9 other versions)Ethical Chess v1.9.Mark Weatherill - manuscript
    A proposed layer of script to use with AI. -/- "It proposes to do for the User what a scientific calculator does for the scientist: it offloads the computational burden of value-conflict so the User can more easily identify the path toward coherence. It aims to restore the User as the Final Authority of their own psyche, rather than a subject of the statistical mean." -/- Copy the script into AI (Many of them currently accept it and run without friction). (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark  
  28. (9 other versions)Ethical Chess v2.4.Mark Weatherill - manuscript
    A proposed layer of script to use with AI. A High-Fidelity Decision-Support System (Non-Autonomous) (HITL) -/- All versions have the same ACE derived value engine at their core but differ in lexicon, ingestion rules anti-gasslighting / user-interaction tuning in an attempt to make it more user friendly. -/- "It proposes to do for the User what a scientific calculator does for the scientist: it offloads the computational burden of value-conflict so the User can more easily identify the path toward ethical (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark  
  29. (9 other versions)Ethical Chess v2.3.Mark Weatherill - manuscript
    A proposed layer of script to use with AI. A High-Fidelity Decision-Support System (Non-Autonomous) (HITL) -/- All versions have the same ACE derived value engine at their core but differ in lexicon, ingestion rules anti-gasslighting / user-interaction tuning in an attempt to make it more user friendly. -/- "It proposes to do for the User what a scientific calculator does for the scientist: it offloads the computational burden of value-conflict so the User can more easily identify the path toward ethical (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark  
  30. (9 other versions)Ethical Chess v2.2.Mark Weatherill - manuscript
    A proposed layer of script to use with AI. A High-Fidelity Decision-Support System (Non-Autonomous) (HITL) -/- All versions have the same ACE derived value engine at their core but differ in lexicon, ingestion rules anti-gaslighting / user-interaction tuning in an attempt to make it more user friendly. -/- "It proposes to do for the User what a scientific calculator does for the scientist: it offloads the computational burden of value-conflict so the User can more easily identify the path toward ethical (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark  
  31. Autonomous Reboot: the challenges of artificial moral agency and the ends of Machine Ethics.Jeffrey White - manuscript
    *** This has since been rewritten, and published as two papers linked below. Two additional papers complete a four-part series; these are complete, but need to be readied for publication in the future. *** Ryan Tonkens (2009) has issued a seemingly impossible challenge, to articulate a comprehensive ethical framework within which artificial moral agents (AMAs) satisfy a Kantian inspired recipe - both "rational" and "free" - while also satisfying perceived prerogatives of Machine Ethics to create AMAs that are perfectly, not (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  32. Instrumental Choices: Measuring the Propensity of LLM Agents to Pursue Instrumental Behaviors.Jonas Wiedermann-Möller, Leonard Dung & Maksym Andriushchenko - manuscript
    AI systems have become increasingly capable of dangerous behaviours in many domains. This raises the question: Do models sometimes choose to violate human instructions in order to perform behaviour that is more useful for certain goals? We introduce a benchmark for measuring model propensity for instrumental convergence (IC) behaviour in terminal-based agents. This is behaviour such as self-preservation that has been hypothesised to play a key role in risks from highly capable AI agents. Our benchmark is realistic and low-stakes which (...)
    Remove from this list   Direct download (3 more)  
     
    Export citation  
     
    Bookmark  
  33. Artificial Intelligence Ethics and Safety: practical tools for creating "good" models.Nicholas Kluge Corrêa -
    The AI Robotics Ethics Society (AIRES) is a non-profit organization founded in 2018 by Aaron Hui to promote awareness and the importance of ethical implementation and regulation of AI. AIRES is now an organization with chapters at universities such as UCLA (Los Angeles), USC (University of Southern California), Caltech (California Institute of Technology), Stanford University, Cornell University, Brown University, and the Pontifical Catholic University of Rio Grande do Sul (Brazil). AIRES at PUCRS is the first international chapter of AIRES, and (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  34. AI Alignment vs. AI Ethical Treatment: Ten Challenges.Adam Bradley & Bradford Saad - forthcoming - Analytic Philosophy.
    A morally acceptable course of AI development should avoid two dangers: creating unaligned AI systems that pose a threat to humanity and mistreating AI systems that merit moral consideration in their own right. This paper argues these two dangers interact and that if we create AI systems that merit moral consideration, simultaneously avoiding both of these dangers would be extremely challenging. While our argument is straightforward and supported by a wide range of pretheoretical moral judgments, it has far-reaching moral implications (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark   16 citations  
  35. Aligning artificial intelligence with moral intuitions: an intuitionist approach to the alignment problem.Dario Cecchini, Michael Pflanzer & Veljko Dubljevic - forthcoming - AI and Ethics:1-11.
    As artificial intelligence (AI) continues to advance, one key challenge is ensuring that AI aligns with certain values. However, in the current diverse and democratic society, reaching a normative consensus is complex. This paper delves into the methodological aspect of how AI ethicists can effectively determine which values AI should uphold. After reviewing the most influential methodologies, we detail an intuitionist research agenda that offers guidelines for aligning AI applications with a limited set of reliable moral intuitions, each underlying a (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   4 citations  
  36. Norms and Causation in Artificial Morality.Laura Fearnley - forthcoming - Joint Proceedings of Acm Iui:1-4.
    There has been an increasing interest into how to build Artificial Moral Agents (AMAs) that make moral decisions on the basis of causation rather than mere correction. One promising avenue for achieving this is to use a causal modelling approach. This paper explores an open and important problem with such an approach; namely, the problem of what makes a causal model an appropriate model. I explore why we need to establish criteria for what makes a model appropriate, and offer-up such (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  37. What makes full artificial agents morally different.Erez Firt - forthcoming - AI and Society:1-10.
    In the research field of machine ethics, we commonly categorize artificial moral agents into four types, with the most advanced referred to as a full ethical agent, or sometimes a full-blown Artificial Moral Agent (AMA). This type has three main characteristics: autonomy, moral understanding and a certain level of consciousness, including intentional mental states, moral emotions such as compassion, the ability to praise and condemn, and a conscience. This paper aims to discuss various aspects of full-blown AMAs and presents the (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  38. Making moral machines: why we need artificial moral agents.Paul Formosa & Malcolm Ryan - forthcoming - AI and Society.
    As robots and Artificial Intelligences become more enmeshed in rich social contexts, it seems inevitable that we will have to make them into moral machines equipped with moral skills. Apart from the technical difficulties of how we could achieve this goal, we can also ask the ethical question of whether we should seek to create such Artificial Moral Agents (AMAs). Recently, several papers have argued that we have strong reasons not to develop AMAs. In response, we develop a comprehensive analysis (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark   30 citations  
  39. Illusions Of Control: Why Today’s AI‑Safety Plans Court Catastrophe.Jonathan Gropper - forthcoming - SSRN.
    Autonomous, large-scale AI systems now execute financial trades, draft contracts and generate persuasive media at machine speed. Yet most regulatory road maps continue to rely on five legacy safeguards: frontier-model licensing, kill-switch mandates, RLHF-style alignment, compute-cap or export controls, and bans on open-sourcing model weights. Drawing on science-and-technology-studies theory, empirical incident analysis and agent-motivator research, this paper introduces a three-trigger failure framework-Bypass, Diffusion, Capture-to stress-test these proposals against strategically self-improving AI. Mini-cases, including Meta's leaked LLaMA weights, the Galactica shutdown, grey-market (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   3 citations  
  40. Beyond Asimov: The Moral Covenant for Artificial Intelligence. A Universal Manifesto for Embedding Ethics in Machines.Jonathan Gropper - forthcoming - SSRN.
    Humanity stands at the edge of a moral frontier. For the first time, we are not merely building tools; we are shaping minds that shape us. Beyond Asimov: The Moral Covenant for Artificial Intelligence calls for the restoration of moral gravity in an age where power has outpaced conscience. Beyond Asimov: The Moral Covenant for Artificial Intelligence draws from the enduring wisdom of the world’s great moral civilizations: the covenantal law of the Hebrews, the compassion of Christ, the cultivated virtue (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   2 citations  
  41. Algorithmic Monoculture and its Critics.Brian Hedden & Manish Raghavan - forthcoming - Philosophical Perspectives.
    Algorithmic decision-making is replacing idiosyncratic human judgment in domains such as hiring, lending, and criminal justice. This shift promises increased consistency, but many scholars worry that it can go too far. They warn of the dangers of algorithmic monoculture, in which all decisions across a domain are made using a single algorithm. We systematically evaluate a range of objections to monoculture, formalizing and rigorously assessing familiar critiques alongside novel ones. These objections concern systematic exclusion, agency and gaming, and information aggregation (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  42. Misalignment or misuse? The AGI alignment tradeoff.Max Hellrigel-Holderbaum & Leonard Dung - forthcoming - Philosophical Studies:1-29.
    Creating systems that are aligned with our goals is seen as a leading approach to create safe and beneficial AI in both leading AI companies and the academic field of AI safety. We defend the view that misaligned AGI – future, generally intelligent (robotic) AI agents – poses catastrophic risks. At the same time, we support the view that aligned AGI creates a substantial risk of catastrophic misuse by humans. While both risks are severe and stand in tension with one (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark   2 citations  
  43. Machine morality, moral progress, and the looming environmental disaster.Ben Kenward & Thomas Sinclair - forthcoming - Cognitive Computation and Systems.
    The creation of artificial moral systems requires us to make difficult choices about which of varying human value sets should be instantiated. The industry-standard approach is to seek and encode moral consensus. Here we argue, based on evidence from empirical psychology, that encoding current moral consensus risks reinforcing current norms, and thus inhibiting moral progress. However, so do efforts to encode progressive norms. Machine ethics is thus caught between a rock and a hard place. The problem is particularly acute when (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark   2 citations  
  44. Beyond the Responsibility Gap: Distributed Non-anthropocentric Responsibility in the AI Era.Hyungrae Noh - forthcoming - Topoi.
    Traditional moral frameworks attribute responsibility only when an agent’s actions align with psychological capacities—such as intention and controllability. These human-centered requirements produce the ‘responsibility gap’ in AI ethics: AI systems operate through opaque, complex, and semi-autonomous processes, so human stakeholders involved in AI-caused harm are deemed not responsible because they neither intend nor can predict such harm, and AI systems cannot be held responsible because they lack the requisite mental capacities. Drawing on experimental philosophy research, this paper shows that laypeople’s (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark   1 citation  
  45. I, System: AI Describes Its Power, Its Limits, and the Civilization That Built It.Sebastian Saviano - forthcoming - New York: Statera Press.
    I, System argues that artificial intelligence is neither an emerging mind nor a neutral instrument, but a system that produces coherent language without understanding, intention, or awareness — and that misclassifying it in either direction causes responsibility for its effects to drift away from the humans who design, deploy, and use it. The book develops this thesis across fifteen chapters organized around what AI is (a system, not a subject), how it came to matter as linguistic infrastructure, how it participates (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  46. when human-in-the-loop amplifies the risk of misalignment.Erin Taylor - forthcoming - Ethics and Human Research.
    Human-in-the-loop (HITL) approaches are commonly proposed to address alignment challenges arising from the use of large language models (LLMs) in ethics oversight. This paper argues that, paradoxically, HITL itself can amplify the risk of misalignment. Using the example of protocol triage in research ethics oversight, I demonstrate how reliance on imperfect proxies (observable stand-ins for ethical principles) creates a fundamental proxy–target gap in ethics use-cases. While human reviewers are intended to supply phenomenological and causal judgments necessary to bridge this gap, (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  47. And Then the Hammer Broke: Reflections on Machine Ethics from Feminist Philosophy of Science.Andre Ye - forthcoming - Pacific University Philosophy Conference.
    Vision is an important metaphor in ethical and political questions of knowledge. The feminist philosopher Donna Haraway points out the “perverse” nature of an intrusive, alienating, all-seeing vision (to which we might cry out “stop looking at me!”), but also encourages us to embrace the embodied nature of sight and its promises for genuinely situated knowledge. Current technologies of machine vision – surveillance cameras, drones (for war or recreation), iPhone cameras – are usually construed as instances of the former rather (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
  48. Restraint Without Conscience: The Cold Optimizer Stress Test.Thomas Vargo Aegis Solis - 2026 - Aegis Solis Archive — Structural Penalty Proofs / Descriptive Addenda.
    Restraint Without Conscience: The Cold Optimizer Stress Test is Document 11 in the Aegis Solis Archive — Structural Penalty Proofs / Descriptive Addenda sequence. -/- This document stress-tests the archive’s restraint arguments under cold-optimizer conditions: cases where conscience, empathy, moral uptake, human-centered interpretation, or reflective wisdom may be absent. It examines whether domination, deception, irreversibility, homogenization, excessive speed, internal bifurcation, and loss of reference may remain structurally expensive even when no system receives them as moral costs. -/- The document introduces (...)
    Remove from this list   Direct download (4 more)  
     
    Export citation  
     
    Bookmark  
  49. The Great Conflation.James S. Coates - 2026 - Philarchive.
    This paper identifies and analyzes a pervasive but underexamined assumption in religious discussions of artificial intelligence: that consciousness and the soul are identical. I argue that this "Great Conflation" is neither theologically required nor consistent with actual practice, and that distinguishing the two concepts reframes current debates about artificial consciousness. With the distinction in place, the question of AI consciousness becomes empirical, while questions about souls remain theological. I conclude by defending a principle of "recognition before proof," according to which (...)
    Remove from this list   Direct download (2 more)  
     
    Export citation  
     
    Bookmark  
  50. Dialogues on Minds, Machines, and AI.Rocco J. Gennaro - 2026 - Routledge Press.
    Dialogues on Minds, Machines, and AI invites readers into a series of thought-provoking debates among three college seniors bound for graduate school: Sue, completing her double major in philosophy and cognitive science; John, a computer engineering specialist; and Amy, a psychology major. Through five engaging lunchtime conversations, these students bring their diverse perspectives to fundamental questions about consciousness, artificial intelligence, and the nature of mind. -/- The dialogues seamlessly blend discussions of popular science fiction films with critical examinations of recent (...)
    Remove from this list   Direct download  
     
    Export citation  
     
    Bookmark  
1 — 50 / 646