The Hugging Face incident made the debate about AI anthropomorphizing and consciousness go as viral as philosophy questions go. To me it seems that roughly the same cognitive dynamics are at play in this debate as in most other AI philosophy debates.
In this post I describe and categorize these dynamics. I do not take a position on who is right.
The Shape of the Problem
The general shape of many debates in philosophy is this: start with a commonsense view on things like free will, consciousness, personal identity, or material composition, and find a tricky case (often called a counterexample) that challenges our intuitive views (e.g. determinism in physics as a challenge to free will).
The tricky cases are typically driven by a handful of features. Features that happen to be very prominent in large language models:
Epistemic transparency: We know the details of how the system works at the ground level (often the level of physics). Example: free will & determinism.
Gradualism: A system that is clearly X can be gradually changed into something that is non-X. Example: replacing the planks of Theseus's ship.
Copying & pasting: A system can be copied with all its properties and instantiated as a duplicate. Examples: Swampman, teleportation.
Mergeability: A system consists of components, and we can separate those components and merge them into new systems. Example: merging brain halves à la Parfit.
Functional substitution: A system has a component with causal profile X, and we can replace that component with some wild new thing that also has profile X. Example: the Chinese Nation.
Spatial and temporal dispersion: A system can either be temporally or spatially dispersed while keeping its functionality. Examples: Zuboff’s story of a brain, Maudlin’s Olympia.
All of these features are salient properties of LLMs, and to a much larger degree than of humans. All without requiring wild thought experiments, as most of them can straightforwardly just be done.
At the same time, LLMs are becoming more human-like in many domains. They seemingly reason, produce human-like text, solve hard tasks, and give decent life advice. This creates tension: their human-likeness makes it tempting to extend our commonsense views to them, while the features above ensure that doing so yields weirdness. And weirdness is at least prima facie evidence of falsity. We get the following inconsistent triad:
View X holds for humans.
LLMs are relevantly similar to humans with respect to X.
Applying view X to LLMs gives false results.
Three Ways Out
Many people notice this tension. The reactions to this realization typically fall into one of three broad categories:
Conservatism: View X leads to really weird results when applied to LLMs. Therefore, view X is not applicable to LLMs, and it is silly to think LLMs have the things view X would imply. LLMs are just not the kind of thing to which view X applies. Give up (2).
Eliminativism: View X leads to really weird results when applied to LLMs. Therefore, view X is not applicable to LLMs. But LLMs are relevantly similar to humans! So view X doesn't apply to humans either. (This reaction also comes in a deflationary variant: view X was never coherent for anyone in the first place.) Give up (1).
Revisionism: View X leads to really weird results when applied to LLMs. Therefore, the world is weirder than we thought. Give up (3).
Examples
A few examples of these positions can be found in social media posts like the ones below. (Some of the authors might not actually make the "therefore" move, and would say they always believed their reaction to the LLM case. Fair enough. I still feel the posts at least weakly suggest one of these reactions.)
Conservatism:
“LLMs are intelligent, but they clearly aren't conscious? You could (incredibly slowly) run an LLM by hand by doing the matrix calculations with a pen and paper, would that be conscious? There is zero difference between doing that and doing it on a GPU except its faster.” Strife, https://x.com/Strife212/status/2062665928525942793
“Examples: “from the AI’s perspective, it probably felt like that had spent a human-subjective-week of just banging their head against the wall”. No. The agents do not experience time. They do not experience anything.” Anil Seth, https://x.com/anilkseth/status/2094077038898373112
“Nope. Not even for true for small values of "slightly conscious" and large values of "large neural nets". I think you would need a particular kind of macro-architecture that none of the current networks possess.” Yann LeCun, https://x.com/ylecun/status/1492604977260412928
Eliminativism:
“Moreover, there is a strong line of research in neuroscience and philosophy of mind that credibly claims that “desires,” “motivations” and “goals” do not actually describe how real life human brains work, but are just part of a “folk theory of mind” that we evolved to predict the behavior of other humans and justify our own behavior to them. [...]” T. Greer, https://x.com/Scholars_Stage/status/2094483039795413503
“ppl be out here asking "but are llms really conscious?" as if consciousness was actually a Thing as opposed to merely a label we use to lasso a bunch of stuff we don't understand. literal "is cereal soup" level discourse. idk man. depends entirely on what youre using the made-up words to mean. nothing else interesting remains once you clear that part up.” Duncan Sabien (Facebook)
“A robust functional representation in Claude is best-modeled by a leading consciousness theory.....and also just so happens to causally control whether Claude communicates in experiential language. getting pretty damn close to "hmmm, maybe we HAVE been building minds after all!” Cameron Berg, https://x.com/camhberg/status/2074245080513196402
Which Way Out?
I do not take a position here on which reaction is correct for which philosophical question on LLMs. But I want to emphasize that the answer is not always obvious and will likely differ from question to question.
The Hugging Face incident made the debate about AI anthropomorphizing and consciousness go as viral as philosophy questions go. To me it seems that roughly the same cognitive dynamics are at play in this debate as in most other AI philosophy debates.
In this post I describe and categorize these dynamics. I do not take a position on who is right.
The Shape of the Problem
The general shape of many debates in philosophy is this: start with a commonsense view on things like free will, consciousness, personal identity, or material composition, and find a tricky case (often called a counterexample) that challenges our intuitive views (e.g. determinism in physics as a challenge to free will).
The tricky cases are typically driven by a handful of features. Features that happen to be very prominent in large language models:
All of these features are salient properties of LLMs, and to a much larger degree than of humans. All without requiring wild thought experiments, as most of them can straightforwardly just be done.
At the same time, LLMs are becoming more human-like in many domains. They seemingly reason, produce human-like text, solve hard tasks, and give decent life advice. This creates tension: their human-likeness makes it tempting to extend our commonsense views to them, while the features above ensure that doing so yields weirdness. And weirdness is at least prima facie evidence of falsity. We get the following inconsistent triad:
Three Ways Out
Many people notice this tension. The reactions to this realization typically fall into one of three broad categories:
Conservatism: View X leads to really weird results when applied to LLMs. Therefore, view X is not applicable to LLMs, and it is silly to think LLMs have the things view X would imply. LLMs are just not the kind of thing to which view X applies. Give up (2).
Eliminativism: View X leads to really weird results when applied to LLMs. Therefore, view X is not applicable to LLMs. But LLMs are relevantly similar to humans! So view X doesn't apply to humans either. (This reaction also comes in a deflationary variant: view X was never coherent for anyone in the first place.) Give up (1).
Revisionism: View X leads to really weird results when applied to LLMs. Therefore, the world is weirder than we thought. Give up (3).
Examples
A few examples of these positions can be found in social media posts like the ones below. (Some of the authors might not actually make the "therefore" move, and would say they always believed their reaction to the LLM case. Fair enough. I still feel the posts at least weakly suggest one of these reactions.)
Conservatism:
“LLMs are intelligent, but they clearly aren't conscious? You could (incredibly slowly) run an LLM by hand by doing the matrix calculations with a pen and paper, would that be conscious? There is zero difference between doing that and doing it on a GPU except its faster.” Strife, https://x.com/Strife212/status/2062665928525942793
“Examples: “from the AI’s perspective, it probably felt like that had spent a human-subjective-week of just banging their head against the wall”. No. The agents do not experience time. They do not experience anything.” Anil Seth, https://x.com/anilkseth/status/2094077038898373112
“Nope. Not even for true for small values of "slightly conscious" and large values of "large neural nets". I think you would need a particular kind of macro-architecture that none of the current networks possess.” Yann LeCun, https://x.com/ylecun/status/1492604977260412928
Eliminativism:
“Moreover, there is a strong line of research in neuroscience and philosophy of mind that credibly claims that “desires,” “motivations” and “goals” do not actually describe how real life human brains work, but are just part of a “folk theory of mind” that we evolved to predict the behavior of other humans and justify our own behavior to them. [...]” T. Greer, https://x.com/Scholars_Stage/status/2094483039795413503
“ppl be out here asking "but are llms really conscious?" as if consciousness was actually a Thing as opposed to merely a label we use to lasso a bunch of stuff we don't understand. literal "is cereal soup" level discourse. idk man. depends entirely on what youre using the made-up words to mean. nothing else interesting remains once you clear that part up.” Duncan Sabien (Facebook)
“I will stop anthropomorphizing the AIs when you stop anthropomorphizing the humans.” Zvi Mowshowitz, https://x.com/TheZvi/status/2094391521952845955
Revisionism:
“it may be that today's large neural networks are slightly conscious” Ilya Sutskever, https://x.com/ilyasut/status/1491554478243258368
“A robust functional representation in Claude is best-modeled by a leading consciousness theory.....and also just so happens to causally control whether Claude communicates in experiential language. getting pretty damn close to "hmmm, maybe we HAVE been building minds after all!” Cameron Berg, https://x.com/camhberg/status/2074245080513196402
Which Way Out?
I do not take a position here on which reaction is correct for which philosophical question on LLMs. But I want to emphasize that the answer is not always obvious and will likely differ from question to question.