Bias, machine learning, and conceptual engineering

Philosophical Studies 182 (7):1889-1917 (2025)
  Copy   BIBTEX

Abstract

Large language models (LLMs) such as OpenAI’s ChatGPT reflect, and can potentially perpetuate, social biases in language use. Conceptual engineering aims to revise our concepts to eliminate such bias. We show how machine learning and conceptual engineering can be fruitfully brought together to offer new insights to both conceptual engineers and LLM designers. Specifically, we suggest that LLMs can be used to detect and expose bias in the prototypes associated with concepts, and that LLM de-biasing can serve conceptual engineering projects that aim to revise such conceptual prototypes. At present, these de-biasing techniques primarily involve approaches requiring bespoke interventions based on choices of the algorithm’s designers. Thus, conceptual engineering through de-biasing will include making choices about what kind of normative training an LLM should receive, especially with respect to different notions of bias. This offers a new perspective on what conceptual engineering involves and how it can be implemented. And our conceptual engineering approach also offers insight, to those engaged in LLM de-biasing, into the normative distinctions that are needed for that work.

Other Versions

No versions found

Links

PhilArchive

External links

Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Similar books and articles

Analytics

Added to PP
2025-02-19

Downloads
130 (#345,782)

6 months
31 (#298,472)

Historical graph of downloads
How can I increase my downloads?

Author Profiles

Rachel Etta Rudolph
University of California, San Diego
Elay Shech
Auburn University

Citations of this work

Cognitive bias in large language models: A vindicatory approach.David Thorstad - forthcoming - British Journal for the Philosophy of Science.
Engineering Social Concepts: Labels and the Science of Categorization.Eleonore Neufeld - forthcoming - In Sally Haslanger, Karen Jones, Greg Restall, Francois Schroeter & Laura Schroeter, Mind, Language, and Social Hierarchy: Constructing a Shared Social World. Oxford University Press.
On travelling concepts.Martin Stokhof - 2025 - Proceedings of the Paris Institute for Advanced Study.

Add more citations