Formats of Representation in Large Language Models

Philosophy and the Mind Sciences (forthcoming)
  Copy   BIBTEX

Abstract

This paper argues for a pluralist approach to representation in large language models. There are two parts to this pluralism, the first is that we should recognise more than one vehicle of representation in transformer models. Call this vehicle pluralism. Rather than identifying the vehicles of representation with a single component of a system, e.g. individual neurons, patterns of activation, regions in the activation space, we should acknowledge multiple systems of representation within a network operating with different vehicles. The second claim is that we should recognise that there are different formats of representation in transformer models. Transformer models do not operate with a purely analogue, structural, or symbolic architecture but are a hybrid system of representation. Finally, I will discuss how this relates to several working hypotheses about representation that have become adopted in the field of mechanistic interpretability.

Other Versions

No versions found

Links

PhilArchive

External links

  • This entry has no external links. Add one.
Setup an account with your affiliations in order to access resources via your University's proxy server

Through your library

Analytics

Added to PP
2025-10-19

Downloads
716 (#81,246)

6 months
375 (#15,535)

Historical graph of downloads
How can I increase my downloads?