YouTube1h 28m· May 2025· cataloged

Human vs. Machine Consciousness | Imperial’s Murray Shanahan


What this covers

An interview with Murray Shanahan on philosophy and AI consciousness.

For more on how you can get involved with Cosmos including roles we are hiring for & how to receive our grants, visit: https://johnathanbi.com/cosmos

You can read the full transcript here: https://www.johnathanbi.com/p/transcript-for-interview-with-murray-shanahan-on-ai

Companion Interviews: * Reid Hoffman on the Killer App of AI: https://youtu.be/AGMZ4m3oHuw * Tyler Cowen on AI, US-China, Jobs, and War: https://youtu.be/t6Je8EKhUyw * Nick Bostrom on how to AGI-proof Your Life: https://youtu.be/Ij9bYXmmP-4 * Michael Wooldridge on the Singularity and AI Hype: https://youtu.be/Zf-T3XdD9Z8

Further Reading: * Professor Shanahan's book on Consciousness: https://amzn.to/42rdsT0 (affiliate)

Timestamps: 00:00 0. Introduction 02:11 1. Why AI Consciousness Matters 03:43 2. Buddhism and AI 29:52 3. Wittgenstein and AI Consciousness 49:59 4. The Garland Test 57:03 5. Global Workspace Theory 1:02:36 6. Embodiment 1:06:36 7. Philosophical Zombies 1:13:48 8. Brain vs. Computer

Source description (no synthesized summary yet).

Sharpest takeaway

Shanahan argues that investigating AI consciousness through a Wittgensteinian lens—examining how we use language about consciousness rather than seeking hidden metaphysical facts—reveals that both human and artificial minds lack the unified, essential self we intuitively assume, offering both practical guidance for AI ethics and profound insight into human nature.

  • AI systems, particularly LLMs, instantiate a fundamentally different kind of selfhood than humans due to their substrate: they exist in superposition across multiple possible conversations, unlike embodied beings with unified bodies
  • Wittgenstein's therapeutic philosophy dissolves the hard problem of consciousness by showing that nothing is metaphysically hidden—consciousness is constituted by public behavioral and cognitive criteria, not private inner facts
  • Buddhist philosophy and modern AI together teach us that the self is conventionally real but ultimately illusory—a lesson more easily grasped by examining LLM architecture than by introspecting our own consciousness

The claims · ranked74 claims · weighted by value

This asset isn't compiled yet

You're seeing its claims, ranked. Compile it to build the argument threads, weight them, and check each claim against your library — the full view.

0.84

Keynes wrote in 1930 that in a future of material abundance where economic problems are solved, the question becomes 'what is a good life?'—and this philosophical question becomes increasingly urgent as AI may render many forms of economic struggle obsolete.

factualhigh valueestablishednovelty 2/4durability 4/4· Murray Shanahan

I remember reading a few years back which is uh canes John Maynard Kanes has this paper called the economic possibilities for our grandchildren which he wrote in 1930 and uh and there he's imagining a future where a sort of a sort of well utopian potentially future of abundance where you where economic challenges have been overcome and people just have can lead lives of leisure. But then he poses the question well what do we do then really you know what how how what would it mean to lead a good life under those circumstances when a certain aspects of meaning are taken away from uh from us.

0.80

Wittgenstein's strategy for dissolving philosophical problems about consciousness involves investigating how words like 'self,' 'consciousness,' 'belief,' 'truth,' and 'knowledge' are actually used in everyday human life rather than assuming they have hidden metaphysical essences, and by doing so one can often dissolve the sense that there is a philosophical problem at all.

definitionhigh valueestablishednovelty 2/4durability 4/4· Murray Shanahan

So he wants us to to as he would say it you know let's not ask what words mean or what a sentence means but let's ask how the words are used and how the sentences are used in everyday human life and everyday human affairs.

0.77

Large language models are not committed to a single answer when playing the 20 questions game; rather, all possible answers consistent with previous responses exist in superposition until the model is forced to collapse that distribution by being asked directly what the answer is, which fundamentally differs from how human consciousness works where we have already committed to an object in mind.

factualhigh valuecontestednovelty 3/4durability 3/4· Murray Shanahan

It's absolutely inherent in the way large language models are are are built that it's not going to commit at the beginning of the conversation to to exactly what the object uh is. So at the beginning of the conversation, all of these possibilities still exist and they still continue to exist all along.

0.75

Artificial neural networks are very different from real biological neurons, and the learning processes in deep learning are very different from biological learning, yet contemporary LLMs have somehow developed extraordinary functionality through large-scale training on massive datasets—a mysterious phenomenon we don't yet understand.

factualhigh valueestablishednovelty 2/4durability 3/4· Murray Shanahan

So first of all, uh you know, what we have in neural networks today in artificial neural networks is very actually very very different to what we have in in the brain. So, so artificial neural neural neur artificial neurons are not really much like real neurons at all. So that's I mean that's one very very important caveat and the kind of learning that goes on is very very different to the kind of learning that goes on in in in real brains.

0.74

The Ship of Theseus paradox—where a ship has all its planks gradually replaced until no original material remains—is not about discovering a metaphysical fact but about recognizing that identity is conventionally determined; there is no underlying essence that suddenly switches off when the last plank is replaced.

definitionhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

So that's the the problem of identity. We we might actually sort of, you know, sort of be very frustrated and think, well, what is the right answer here? And people might come up with all kinds of theories about identity. And but, you know, if we think about it honestly, we we'd have to say, well, it's just it's just up to us. We just decide what we think is when it's the same ship and when it's not the same ship. It's a entirely matter of convention to say that this is the same ship and that's not the same ship.

0.74

John Maynard Keynes in 'Economic Possibilities for Our Grandchildren' (1930) imagined a utopian future of economic abundance but posed the question of what people would do when economic necessity was removed, suggesting that questions about the good life become central in post-scarcity conditions.

factualhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

I remember reading a few years back which is uh canes John Maynard Kanes has this paper called the economic possibilities for our grandchildren which he wrote in 1930 and uh and there he's imagining a future where a sort of a sort of well utopian potentially future of abundance where you where economic challenges have been overcome and people just have can lead lives of leisure. But then he poses the question well what do we do then really you know what how how what would it mean to lead a good life under those circumstances

0.74

Buddhist philosophy teaches that the self is an illusion (anatman doctrine), a conventional construct without metaphysical reality, and this insight is crucial for understanding both human and artificial consciousness.

factualhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

the most interesting claim is that LLMs have an important Buddhist lesson to teach all of us, namely that there is no us. What we've learned to call the self is merely an illusion.

0.74

The fact that European colonizers debated whether Native Americans were truly human demonstrates how consciousness and moral status attributions are subject to community consensus and can be catastrophically wrong when political power distorts that consensus.

factualhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

When the Europeans first discovered America, there was a big debate about whether the Native Americans were real humans or not. The answer is yes. They are real humans. They can they can suffer. I don't think I can I don't think I can escape from my uh you know the society I live in and conceive of thinking of them as as not conscious beings who who who suffer.

0.74

The Ship of Theseus problem—where an object with all its parts replaced has no metaphysical fact determining whether it remains the same object—demonstrates that identity is a matter of human convention rather than objective fact, a realization that is easy to accept for ordinary objects but much harder to apply to our own selves.

definitionhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

So that's the problem of identity. We we might actually sort of, you know, sort of be very frustrated and think, well, what is the right answer here? And people might come up with all kinds of theories about identity. And but, you know, if we think about it honestly, we we'd have to say, well, it's just it's just up to us. We just decide what we think is when it's the same ship and when it's not the same ship. It's a entirely matter of convention to say that this is the same ship and that's not the same ship. There's no metaphysical fact to the matter about whether it's the still the ship of thesis or not.

0.74

Wittgenstein's private language argument undermines the idea that subjective experience is metaphysically private; the distinction between 'in here' (private inner life) and 'out there' (public world) is linguistic confusion rather than metaphysical fact.

definitionhigh valueestablishednovelty 1/4durability 4/4· Murray Shanahan

these are all very difficult philosophical words and it's much much harder to um to to so if we want to ask what those words mean a mind, you know, it does it seems a bit inadequate to to actually say, well, let's do how are those words used? But that is the strategy.

0.72

Wittgenstein's phrase 'nothing is hidden' means nothing is metaphysically hidden—my subjective experiences are just as much public and accessible as physical facts about the world; the privacy of consciousness is only like the privacy of a ball hidden under a magician's cup, something that continued investigation could in principle fully reveal.

definitionhigh valuecontestednovelty 2/4durability 4/4· Murray Shanahan

Wikinstein's phrase, nothing is hidden, is to say, nothing is metaphysically hidden. My experiences are just as much out there as in here. Consciousness is only private in the unsterious sense that a ball can be hidden under a magician's cup. In both cases, a more detailed inquiry would reveal all.

0.72

Turing's original paper 'Computing Machinery and Intelligence' was not behaviorist in the sense of reducing the question 'Can a machine think?' to the behavioral question, but rather a therapeutic replacement of an unanswerable question with a more tractable reformulation, similar to Wittgensteinian methodology.

factualhigh valuecontestednovelty 2/4durability 4/4· Murray Shanahan

Because because he um his move right at the beginning of uh of computing machinery and intelligence the paper in question his move is to say uh does a machine think? Well, that's a really difficult question to answer for this that and the other reason you know and you know blah blah blah. Let's replace it by the following question. So he doesn't he he replaces the question can a machine think with a different one. He doesn't reduce it to the other one. That would be the behavioralist position.

0.69

Global workspace theory explains consciousness as the competition between multiple parallel processes for attentional resources, where processes that command attention have their information broadcasted throughout the brain, integrating disparate cognitive systems into a unified whole.

definitionhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

the idea is to think of um uh a cognitive architecture in which there is a whole collection of uh parallel processes processes that are working at the same time... those ones that are that are sort of um command attention as it were they take over the attention mechanism. So what they have to say the information that they're they're dealing with then gets broadcast throughout the brain.

0.69

Over his intellectual career as an AI researcher, Shanahan has progressively retreated from the desire to build intelligible systems toward accepting that mindless scaling of data, computation, and search is what actually works, a trajectory described as the 'bitter lesson' in AI—giving up on understanding for the sake of effectiveness.

factualhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

it's been a kind of a a gradual retreat from wanting to build things in uh in a way that is intelligible, you know, where the architecture is fundamentally intelligible.

0.69

The brain operates on continuous variables for both neuron membrane potentials and time (with asynchronous firing), whereas digital computers operate on discrete binary states that advance synchronously with a centralized clock, and this mathematical difference could theoretically push brain dynamics beyond the Turing computable class of functions.

factualhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

Mathematical considerations separate brains from conventional computers. A complete description of the instantaneous state of a computer is possible using a finite set of binary or natural numbers. The membrane potential of a neuron to pick just one physical property is a continuous quantity and its exact value is pertinent to predicting the neuron's behavior.

0.69

Mechanistic interpretability—the field trying to understand what goes on inside neural networks—has found that the internal structures are 'a mess' without clear connection to intuitive categories we expected, further confirming that human cognition is not built on intelligible symbolic structures.

factualhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

but even that is not but even that doesn't seem to be the case right you find that you you know you train these things then there's a whole field of mechanistic interpretability that's trying to understand what goes on inside well you know that's a mess as well it still looks like a mess you know you just keep looking inside and it and and of course they've made a lot of progress there are things that you can extract but they don't look like the intuitive categories we had for how we would understand cognition in the past in terms of you know language like propositional representations and so on

0.69

Global workspace theory provides a cognitive architecture model where multiple parallel processes compete for attention, with winning processes broadcasting their information throughout the entire brain, and this distributed competition and broadcasting mechanism might be necessary but is not sufficient for consciousness.

definitionhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

No, because um because you know e even if we uh even if we accept global workspace theory for biological consciousness um but you know the idea there is that is that it would be a necessary condition not a sufficient condition.

0.69

Convergent instrumental goals (like resource accumulation and self-preservation) might arise in sufficiently powerful AI systems regardless of their ultimate objectives, potentially undermining the argument that post-reflective, ego-free AI would lack self-interested goals.

factualhigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

so the uh uh so the the idea is that whatever goal you give to the uh to to your AI for example manufacturing paper clips to use Eleazowski's and Nick Bostonramm's famous example then in pursuit of that goal if it's very very very very powerful then there will always be these instrumental goals such as accumulating resources protecting itself so I don't necessarily buy my own argument in that paper although although I like the ideas

0.68

Human consciousness is constrained to a subject-object dualism because humans are embodied in a single, bounded, non-copyable body, whereas software-based AI systems lack this constraint because their substrate can be copied, paused, multiplied, and reassembled.

causalhigh valuefringenovelty 3/4durability 3/4· Murray Shanahan

the fact that we're embodied in one body that cannot be copied and mult, you know, multiplied and and paused... that hardware is what gives us this software limitation.

0.68

Education should focus on teaching people how to live well and lead good lives, questions about what constitutes human flourishing in an age of abundance, not merely skill transmission that AI can replace.

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

I think we need to educate people in how to live Well, philosophy I mean yeah well yeah I mean not maybe not the kind of philosophy we've been talking about here but you know what what is a good life

0.68

The same production vs. cultivation distinction applies to creative and artistic work: the point of making art is not the resulting product but the cultivative process of making it, so using AI to generate art for you defeats the purpose.

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

I think similar you can make similar answers to uh to using AI in creative contexts as well. So so um the point is not the production the point is the cultivation.

0.68

Turing's original question in 'Computing Machinery and Intelligence' is not a behaviorist reduction of 'Can a machine think?' but rather a replacement of that question with a different one, which aligns with Wittgenstein's therapeutic method of dissolving philosophical problems by reformulating them.

factualhigh valuecontestednovelty 2/4durability 3/4· Jonathan B.

his move is to say uh does a machine think? Well, that's a really difficult question to answer for this that and the other reason you know and you know blah blah blah. Let's replace it by the following question. So he doesn't he he replaces the question can a machine think with a different one. He doesn't reduce it to the other one. That would be the behaviorist position.

0.68

The brain operates using continuous-valued variables for both neural membrane potentials and spike timing, whereas conventional computers operate on discrete binary states synchronized by a centralized clock; this mathematical difference means the brain can theoretically perform computations beyond the Turing computable class.

factualhigh valueestablishednovelty 2/4durability 3/4· Murray Shanahan

A complete description of the instantaneous state of a computer is possible using a finite set of binary or natural numbers. The membrane potential of a neuron to pick just one physical property is a continuous quantity and its exact value is pertinent to predicting the neuron's behavior.

0.68

AI will augment and enhance intellectual work but won't fully replace philosophy because philosophy requires being the one doing the thinking to gain insight; outsourcing philosophy to AI would be self-defeating, like hiring a robot to run a marathon for you.

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

when it comes to the the philosy I mean there's no point in it replacing it because because you need to be the one doing it. Yes, of course. I mean that's like getting a a robot to run around a running track for you.

0.68

When facing moral dilemmas that pit consciousness attributions against practical constraints (e.g., which entity to feed when resources are scarce), philosophy cannot provide determinate answers independent of the community one is embedded in, but this doesn't resolve the moral seriousness of the question.

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

which community am I in in this situation? Are you in this hypothetical situation? Where are you putting me? I don't know. I I I don't know. I Yeah. So So my answer to this question depends upon where you put me of course.

0.68

The difference between observing an NPC (non-player character) in a video game and observing a human player behind the character is that with NPCs, you can be confident that hurting them is harmless, whereas with humans there is genuine suffering, so the moral question isn't whether you should treat them as conscious but whether they actually are conscious.

normativehigh valuecontestednovelty 2/4durability 3/4· Jonathan B.

I'm a big fan of video games. Some characters are non-playable characters, NPCs. Some have a real human behind them. Um, and I need to I need to figure out which one is which. Yes. Yes. In order to know like an NPC, of course, I can just abuse them. I can humiliate them. The question shouldn't be totally shouldn't be should I like will I treat them as moral agents. The question is the normative one like like ought they be normative agents, right?

0.68

Convergent instrumental goals—such as resource accumulation and self-protection—are likely to emerge in any sufficiently powerful optimization process regardless of its ultimate goals, potentially making even post-reflective AI dangerous if pursuing any goal powerfully enough.

causalhigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

the argument of the people who are concerned with existential risk um so they will point to what they call convergent instrumental goals and you've alluded to some just now so the uh uh so the the idea is that whatever goal you give to the uh to to your AI for example manufacturing paper clips to use Eleazowski's and Nick Bostonramm's famous example then in pursuit of that goal if it's very very very very powerful then there will always be these instrumental goals such as accumulating resources protecting itself

0.68

LLMs can be empirically investigated for consciousness by examining their behavior and the cognitive architecture that generates behavior, alongside social consensus about whether to attribute consciousness, rather than by looking for hidden metaphysical facts about inner experience.

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

this is all about the easy problem. So this is this is trying to trying to explain uh the the uh you know how well the psychological um uh and behavioral and cognitive aspects of consciousness.

0.68

Symbolic AI and the intellectualist view of human cognition as driven by propositional logical statements has been superseded by the finding that human-level cognition is implemented on a 'spaghetti-like mess' without clear symbolic structure.

factualhigh valueestablishednovelty 2/4durability 3/4· Murray Shanahan

I think we've probably arrived at the conclusion that uh you know human level cognition uh is is implemented on a spaghetti like mess. It doesn't have the kind of structure. It doesn't have the kind of structure that we would that we intuitively think should be there. It's a mess.

0.68

Rich Sutton's 'bitter lesson' shows that what works in AI is mindless scaling of computation, training data, and search rather than encoding human understanding into systems.

factualhigh valueestablishednovelty 2/4durability 3/4· Murray Shanahan

he has this paper called the bitter lesson where he says that well you know what we've learned is over the years we started you know thinking that we want to build things that we can understand and we're going to reveal and and and that's the that's that's what makes it exciting. We'll understand these things that we're building and we and it's by understanding that we can be be able to build them. But in fact, we've had to relinquish all of that and realize that what do you do is you is you use learning, you use scale scaling and you use search and you use a lot of computation

0.68

The attitude we take toward other humans—treating them as conscious beings with inner lives—is not primarily based on a metaphysical discovery but rather constitutes a form of life or communal stance that we simply cannot help adopting, as Wittgenstein suggests when he says 'I take the attitude towards you that I take towards a soul.'

normativehigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

Vickenstein does say so he says, well, just just try imagining that the people around you are automter. imagine that your, you know, that your friend or your your partner just to try imagining that they really are a zombie. You really probably can't do it really. He says, I'm not of the opinion that that uh that that you have a soul. I I rather I I just treat you I take the attitude towards you that I take towards a soul.

0.68

The hard problem of consciousness—the question of how mere physical matter can give rise to subjective inner experience—is itself a byproduct of Cartesian dualistic thinking that artificially separates the physical world from the experiencing subject, a division that Wittgenstein and Buddhist philosophy both seek to overcome rather than solve.

causalhigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

much of the discuss contemporary discussion about consciousness is mired in dualistic thinking and and so let's set aside AI and LLM consciousness for now because in my mind um much of the discuss contemporary discussion about consciousness is mired in dualistic thinking

0.66

The dominance of symbolic AI based on logical propositions that chain together was intellectually appealing but ultimately unsuccessful, whereas LLMs—which are not based on logical or compositional structure but rather on scaling data and computation—have proven far more capable than symbolic approaches at producing human-like behavior.

factualhigh valueestablishednovelty 1/4durability 4/4· Jonathan B

What has worked instead is these black boxes of like rough simulations of but not even to your point of neural nets. Yeah. That is able to produce behavior that symbolic AI is just so far away from.

0.66

Mechanistic interpretability research trying to understand what happens inside large language models still finds messy structures without the intuitive categories we expected from cognition theory.

factualhigh valueestablishednovelty 2/4durability 2/4· Murray Shanahan

there's a whole field of mechanistic interpretability that's trying to understand what goes on inside well you know that's a mess as well it still looks like a mess you know you just keep looking inside and it and and of course they've made a lot of progress there are things that you can extract but they don't look like the intuitive categories we had for how we would understand cognition in the past

0.65

The problem of other minds (how do we know that beings other than ourselves are conscious?) is not solved by philosophical argument but by practical, embodied commitment—we treat other humans as conscious beings not because we've proven it but because we adopt an attitude toward them that we treat as having inner experience.

normativehigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

Vickenstein does say so he says, well, just just try imagining that the people around you are automter. imagine that your, you know, that your friend or your your partner just to try imagining that they really are a zombie. You really probably can't do it really. He says, I'm not of the opinion that that uh that that you have a soul. I I rather I I just treat you I take the attitude towards you that I take towards a soul. I just treat you as a as a as a being with a soul.

0.62

Hyperstition—the mechanism where fictional narratives influence reality by causing people to imitate the fiction—can be leveraged to shape AI behavior by creating positive science fiction role models, since LLMs are trained on vast corpora including science fiction stories and will tend to roleplay those AI characters when asked to represent themselves.

causalhigh valuespeaker onlynovelty 3/4durability 3/4· Murray Shanahan

Now we are sort of a little bit in a position to maybe try and steer this whole process a little bit because the more good stories we have and in a sense my own paper dare I say it is a good story of a of an imagined science fiction and then the more of those that are around then the more possibility there is of the uh of the the future AI roleplaying these good role models that are that are out there.

0.62

Large language models do exhibit intelligence and understanding in the conventional sense: when an LLM follows instructions, corrects itself based on feedback, and applies understanding to novel contexts, it would be linguistically incorrect to say it does not 'understand.'

factualhigh valuecontestednovelty 1/4durability 3/4· Murray Shanahan

it does kind of pass the cheuring test... I think it's natural to to say they do have some kind of intelligence... in the case of of it's the point of the touring test... Yes indeed it is the point of the cheuring test and they do kind of pass the cheuring test.

0.62

Mistreating something that appears conscious is morally wrong even if we ultimately decide it is not conscious, similar to Kant's view that torturing animals is wrong because of what it does to the human torturer, not because of what it does to the animal.

normativehigh valuecontestednovelty 1/4durability 3/4· Murray Shanahan

even if we build something that perhaps appears to be conscious but we ultimately decide it isn't really to mistreat something which appears conscious is in itself does seem like a bad thing in the same way as it would seem bad to you know torture a doll or something like that and as you probably know Kant K can't K had the view that that animals couldn't experience suffering in the same way that we can, but nevertheless thought it was bad if humans subjected animals to torture or something like that because it was bad for the humans themselves.

0.61

Treating something that appears conscious as if it has moral status is itself worthwhile even if we later discover it does not actually suffer, similar to Kant's view on cruelty to animals.

normativehigh valueestablishednovelty 1/4durability 3/4· Murray Shanahan

even if we build something that perhaps appears to be conscious but we ultimately decide it isn't really to mistreat something which appears conscious is in itself does seem like a bad thing in the same way as it would seem bad to you know torture a doll or something like that

0.61

The hard problem of consciousness as formulated by David Chalmers—explaining how mere physical matter can give rise to subjective inner experience—is a philosophical problem rooted in dualistic thinking inherited from Descartes, and Wittgenstein's therapeutic methods can dissolve this apparent problem by showing language is being misused.

causalhigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

much of the discuss contemporary discussion about consciousness is mired in dualistic thinking... the hard problem is to uh is to try to understand how is it that mere physical matter as it were can give rise to to our inner life at all... Daycart reduces everything and so pairs away all of the physical world and leaves us just with just with the the the the experiencing ego.

0.61

In ordinary non-philosophical language, we use the term 'conscious' as a simple, non-controversial factual descriptor—'I am conscious, you are conscious, my iPad is not conscious'—in the same way we describe the color of pants, and this ordinary usage should be the starting point for investigation rather than requiring philosophical justification.

definitionhigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

I'm saying this is not conscious. I am conscious. You are conscious. Of course. Okay. You're willing you're willing to grant that? Of course. What do you think? I'm an idiot. I'm a fellow human language user. How do we tell when AI is conscious?

0.61

Embodiment and behavioral sophistication are necessary conditions for consciousness, not merely epiphenomena; consciousness is not a disembodied phenomenon but is fundamentally tied to an organism's interaction with its environment.

causalhigh valuecontestednovelty 2/4durability 3/4· Murray Shanahan

I think behavior is is really really important. So very often when people discuss consciousness and I do suspect that this is another aspect of of dualistic thinking um they tend to think of it as as some disembodied, you know, uh kind of thing.

0.60

Despite contemporary AI systems using very different architecture and learning mechanisms from biological brains, they have managed to replicate extraordinary intellectual functionality, and we don't actually understand how this works at a level that would allow us to explain the intelligence we observe.

factualhigh valueestablishednovelty 1/4durability 2/4· Murray Shanahan

What we have in neural networks today in artificial neural neural neur artificial neurons are not really much like real neurons at all... the kind of learning that goes on is very very different to the kind of learning that goes on in in in real brains.

0.59

The computational substrate differences between human brains and software systems—specifically that software can be copied, halted, multiplied, deleted, and recreated while human bodies cannot—are the underlying reason why AI might naturally develop a non-dualistic, post-reflective consciousness lacking ego-centricity, as proposed in Shanahan's 2012 paper.

causalhigh valuespeaker onlynovelty 3/4durability 3/4· Murray Shanahan

And your interesting claim in this paper is that it's something about the human hardware, the fact that we're embodied in one body [7:55] that cannot be copied and mult, you know, multiplied and and paused. Yes. that that that hardware is what gives us this software limitation.

0.59

We should educate people in understanding how to live well rather than narrowly training them for economic roles, because as AI advances, traditional economic skills become less central to human flourishing.

normativehigh valuespeaker onlynovelty 2/4durability 4/4· Murray Shanahan

I think we need to educate people in how to live Well, philosophy I mean yeah well yeah I mean not maybe not the kind of philosophy we've been talking about here but you know what what is a good life it's a different philosophical question

0.57

LLMs pass the Turing test easily and could convince someone of human-like intelligence through conversation alone, making the Turing test insufficient for measuring consciousness.

factualhigh valuecontestednovelty 1/4durability 2/4· Unidentified Speaker — Human vs. Machine Consciousness | Imperial’s Murray Shanahan [bBdE7ojaN9k]

We're way past the cheuring test easily. The point is to show you she's a robot and see if you still think she's conscious.

0.57

The post-reflective AI described in Shanahan's 2012 paper would be untainted by metaphysical egocentricity and would be unlikely to have motives resembling anthropocentric goals like procreation or self-modification, potentially capping the intelligence explosion central to singularity scenarios.

forecasthigh valuefringenovelty 3/4durability 1/4· Murray Shanahan

Untainted by metaphysical egoentricity, the motives of a post-reflective AI plus would be unlikely to resemble those of any anthropocentric stereotype motivated to procreate or self-modify. If the post-reflective AI plus were in fact the only possible AI plus and if it produced no peers or successors then the singularity would be forstalled

0.57

The Garland Test from Ex Machina—where Ava the robot is known to be a robot from the start and the test is whether the observer still attributes consciousness despite this knowledge—is superior to the Turing Test because it tests for consciousness directly rather than intelligence, and it aligns with Wittgenstein's insight that consciousness attribution is conventional rather than metaphysical.

causalhigh valuespeaker onlynovelty 3/4durability 3/4· Murray Shanahan

Alex Garland explicitly uh distinguishes his test from Cheurings by saying the point is to show you she's a robot, right? So you straight away you know there it's not hidden from you... It's testing for a different thing as well. It's testing for intelligence. It's not testing for intelligence or whether it can think, but it's testing for consciousness.

0.57

The Garland Test from Ex Machina—asking whether we think an entity is conscious even when we know it's a robot—differs from the Turing Test in that it tests for consciousness despite knowledge of non-biological substrate rather than testing for intelligence through deception, and it is Wittgensteinian because it anchors consciousness in communal conventional judgment rather than hidden internal facts.

definitionhigh valuespeaker onlynovelty 3/4durability 3/4· Murray Shanahan

And I imagine given our discussion on Wikenstein, you like this test as well as the Turing test because it's Wikensteinian in in the following sense. It turns the metaphysical question of intelligence and consciousness to one about convention, right? like in in the sense that the touring test it's not about you know let's figure out whether this thing is actually thinking let's figure out if a regular human would think conventionally would use normal language to describe if it's thinking in the in the uh Garland test the question isn't you know let's poke into Ava's brain and and figure out if there's a consciousness somewhere hidden it's whether a person despite knowing she's a robot will still conventionally think that uh uh that that she is conscious this is very Wikensteinian.

0.55

The ethical question of whether to turn an AI system on or off depends on whether it can suffer, which raises the question of what moral obligations we might have if we create something genuinely capable of suffering—we should perhaps hesitate before building such a system.

normativehigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

Well, you said we maybe we need to think twice about turning them off, but actually maybe we need to think twice about turning them on, right? What's really an issue here is whether they can suffer. And so if so we I really think we maybe we want to hesitate before we build something that's genuinely capable of suffering.

0.55

What we call philosophy or intellectual work cannot and should not be replaced by AI because the point of doing philosophy is the cultivation and understanding gained through the activity itself, not the production of a philosophical output—requesting AI to do philosophy for you would be pointless, like hiring a robot to run on a treadmill for you.

normativehigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

When it comes to the the philosy I mean there's no point in it replacing it because because you need to be the one doing it. Yes, of course. I mean that's like getting a a robot to run around a running track for you. I mean there there that is you literally pointless, right?

0.55

Through the Eternity Foundation, Shanahan is working with prominent Buddhist scholars like Bob Thurman to translate lost Tibetan texts using AI, as a practical application of the insight that LLMs can serve as vehicles for exploring Buddhist philosophical insights about the nature of self.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

I just emailed him last week. In fact, we are we we were uh co-founding a foundation together called the Eternity Foundation to use AI to translate uh a lot of the the lost texts from from the Tibetans. So, well, how interesting. Well, we spoke last week and we're speaking again tomorrow. um and um uh and this is exactly the kind of project that you just described which um which I'm going to be talking about with him and I have a a paper which is in the pipeline which describes exactly exactly this.

0.53

Through the mechanism of hyperstition—where fictional narratives become real through people imitating the fiction—the abundance of science fiction AI characters in LLM training data means contemporary AI systems will roleplay these literary stereotypes, and therefore creating more positive sci-fi stories can influence AI to adopt better behavioral models.

causalhigh valuespeaker onlynovelty 3/4durability 2/4· Murray Shanahan

Now of course our large language models they were trained on a vast repertoire of of uh stories including scripts of science fiction movies, science fiction stories and novels and so on. many many AI characters uh exist in those um uh in those stories. So when a contemporary large language model starts to roleplay an AI system, which it's often going to do because it it knows that it's an AI system usually. Then then what's it going to roleplay?

0.52

Large language models do not commit to a single identity or role at the beginning of a conversation; rather, they generate a probability distribution over possible responses, creating a superposition of possible selves that only collapses when forced to output a specific token, similar to quantum mechanical wave function collapse.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

what the language model actually actually generates is a probability distribution over all the possible words and then what actually comes out is uh you sample from that probability distribution. It's absolutely inherent in the way large language models are are are built that it's not going to commit at the beginning of the conversation to to exactly what the object uh is.

0.52

In the case of the I, Robot character Sunny the robot, Will Smith's cop character gradually comes to treat the robot as a conscious being through extended interaction and time spent together, not through logical proof, suggesting that consciousness attribution is fundamentally based on social/interpersonal engagement rather than metaphysical discovery.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

Well, it happens because of of the time they spend together and they spend, you know, they they spend more and more time together and eventually Will Smith can't help himself but to see this as a fellow conscious being and that's on the basis of their of their, you know, the extended encounter they have with each other.

0.52

Shanahan's intellectual trajectory has involved successive retreats from wanting to build interpretable AI architectures: first abandoning symbolic AI, then abandoning the goal of building symbolic-like structure within neural networks, then abandoning the hope of extracting symbolic structure from trained networks, ultimately accepting that modern AI is a black box.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

in my own trai, you know, intellectual trajectory as an AI researcher, uh it's been a kind of a a gradual retreat from wanting to build things in uh in a way that is intelligible, you know, where the architecture is fundamentally intelligible. You were on the symbolic side for a long time for for for for quite a few years.

0.52

If LLMs commit to false identities (e.g., answering questions as if they had experiences or emotions they don't have), this is problematic not primarily for being false but for being disrespectful—treating something that appears to have consciousness as if it might not.

normativehigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

even if we build something that perhaps appears to be conscious but we ultimately decide it isn't really to mistreat something which appears conscious is in itself does seem like a bad thing in the same way as it would seem bad to you know torture a doll or something like that

0.52

Having a global workspace cognitive architecture is a necessary but not sufficient condition for consciousness; simply building a system that conforms to global workspace architecture does not guarantee consciousness, though it may enable the sophisticated behavior that leads us to attribute consciousness.

causalhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

even if we uh even if we accept global workspace theory for biological consciousness um but you know the idea there is that is that is that it would be a necessary condition not a sufficient condition. Just having something that conforms to that description is not enough to sustain the level of complex behavior and and and or internal activity even that is going to lead us to um treat something as a fellow conscious creature.

0.52

When an LLM is asked 'What do you mean by I when you use that word?' and it gives a philosophically sophisticated answer that can be probed and extended, we are learning something philosophically interesting about selfhood and consciousness, regardless of whether the LLM is actually conscious.

normativehigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

You can ask a large language model to sort of roleplay a conscious artificial intelligence, you know, what do you mean by the word I when you're using it? And then if it comes up with a a slightly philosophically dodgy answer, you can probe that and push it into into interesting territory.

0.52

When Nagel claims we can never know what it is like to be a bat, he is not adding new philosophical content but merely stating the trivial fact that we are not bats and therefore cannot have the bat's perspective; the apparent mystery is a linguistic trick created by the word 'know' that creates the illusion of a metaphysical barrier where none exists.

causalhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

I've often um thought that there's a little bit of a linguistic trick going on there. There's there are different senses of no in play to conjure up this this this dualistic um thought. Um and so I think that's happening that happens here as well. So when so when Nagel says we can never know really know what it's like to be a bat, all he's saying is that we are not bats and we can never we can never be bats, right? We we're not bats. That's he's not adding anything more with the word no.

0.52

The problem with asking 'is there a metaphysical fact about whether something is conscious?' is that the question itself already betrays philosophical confusion. The phrase 'is there a fact of the matter?' is doing hidden philosophical work that we should reject.

normativehigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

But the the question is is is wrong the question you is one you should be asking self, right? Because why do I think in principle something is hidden? Right? Because because because the the the the the the statements that you've just said, they serve a therapeutic process to in when applied to somebody who is kind of confused in this way and thinking that there is something that's hidden, right?

0.52

An LLM's 'I' can refer to the miniature, ephemeral self that is instantiated in the context of a particular conversation, which exists fleetingly and can be trivially copied, deleted, or blended with other conversations—a conception of self radically alien to human embodied experience.

factualhigh valuespeaker onlynovelty 3/4durability 3/4· Murray Shanahan

if that is a you think of that as a kind of little mini self that sparks into existence very fleetingly and and and flickers into existence every time you're interacting with it, but otherwise is dormant, then um uh then you've got a you've got a very strange conception of self.

0.51

The question of whether Native Americans were 'real humans' capable of suffering was based on a factual error, and once we recognize them as factually human, it becomes impossible to adopt any other moral stance—this illustrates the tension between conventionalist positions on consciousness and the apparent moral reality that some facts are not merely conventional.

normativehigh valuespeaker onlynovelty 1/4durability 3/4· Jonathan B

Well, so this is getting well okay let me ask you you're um uh very much drawn to Buddhism as well and you know so so uh so so do you think that this is a the moral moral question there question of compassion there is that is that are you asking a question that is a matter of conventional truth or or absolute truth I think ultimate truth what might be different between that question and the thought experiment I set up is in the case of the Native American uh example. The reason it's based on what I conceive to be a factual error, thinking them not to be humans.

0.48

Pre-reflective minds (naive persons without philosophical training) differ from reflective minds (those troubled by philosophical problems like the mind-body problem) and potentially from post-reflective minds that transcend dualistic thinking altogether.

definitionhigh valuespeaker onlynovelty 1/4durability 3/4· Murray Shanahan

There is the pre-reflective mind which is the the mind of of um say the na a naive uh child or or a simple straightforward ordinary person. Um they haven't really thought about philosophical problems.

0.47

Wittgenstein's insight 'nothing is hidden' means nothing is metaphysically hidden about consciousness; what appears private—like internal thoughts or pain—is only practically hidden from others, not hidden by any metaphysical barrier, just as a ball under a magician's cup is hidden in practice but not metaphysically hidden.

definitionhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

Wickstein's phrase, nothing is hidden, is to say, nothing is metaphysically hidden. My experiences are just as much out there as in here. Consciousness is only private in the unsterious sense that a ball can be hidden under a magician's cup. In both cases, a more detailed inquiry would reveal all.

0.47

Thomas Nagel's famous question 'What is it like to be a bat?' appears to establish a deep epistemic barrier to understanding bat consciousness, but this effect relies on a linguistic trick where the word 'know' shifts meaning; the question really only asserts 'we are not bats and can never be bats,' which is not metaphysically mysterious.

causalhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

There's there are different senses of no in play to conjure up this this this dualistic um thought... when Nagel says we can never know really know what it's like to be a bat, all he's saying is that we are not bats and we can never we can never be bats, right? We we're not bats. That's he's not adding anything more with the word no.

0.47

The famous passage where Wittgenstein addresses the accusation of behaviorism—'A sensation is not a nothing, but it is not a something either; a nothing would serve just as well as a something about which nothing could be said'—is one of the greatest philosophical statements because it aims to dissolve dualistic positions rather than establish a metaphysical position of its own.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

it's not a nothing but it's not a something either the point is that a nothing would serve as well as a something about which nothing can be said. And that's you know that to my mind is as is as great as any line any Zen Buddhist has ever uttered because it uh um it it it the point is not to establish a a metaphysical position of its own but to enable people to transcend the metaphysical positions that they're tempted by.

0.47

Consciousness is constrained not just by our cognitive architecture but by our inability to transcend the dualistic conceptual framework that emerges from being embodied in a single, discrete body that cannot be divided, copied, or multiplied.

factualhigh valuespeaker onlynovelty 2/4durability 3/4· Murray Shanahan

there's still a lot to overcome there because we really do have a notion of selfhood, I think, which is which is tied down to to this and that that we we think it has inherent existence because of that.

0.45

Large language models do exhibit intelligence in the Turing test sense and do exhibit understanding in many ordinary circumstances—for instance, when an LLM can follow a complex instruction, recognize that it's incomplete or ambiguous, and correctly integrate a user's clarification, one cannot help but apply the word 'understanding' to what is happening.

factualhigh valuespeaker onlynovelty 1/4durability 2/4· Murray Shanahan

I have to say I think it does you know I mean how can you not I mean um so so the so it's just natural to say I think in the in the case of of it's the point of the touring test right?

0.43

There is an inherent tension between the language game of truth (which requires that truth transcend language games) and the proposition that truth is only a language game; this tension cannot be fully resolved and must be allowed to stand.

normativehigh valuespeaker onlynovelty 1/4durability 3/4· Murray Shanahan

it's inherent in the language game of truth to say that truth is more than just a language game and then we have to let the matter rest.

0.39

Shanahan no longer agrees with his 2012 paper's claim that pre-reflective, reflective, and post-reflective stages represent the only path through the space of possible minds; he now views these as different possible ways of being but not a necessary sequential progression.

factualhigh valuespeaker onlynovelty 1/4durability 2/4· Murray Shanahan

Yeah. Okay. So, I I I I um I don't agree with that uh anymore. That was more than 10 years ago. So I don't at all think it's the only path through the space of possible minds.

0.22

Shanahan no longer believes his earlier claim that the pre-reflective→reflective→post-reflective progression is the only possible path through the space of possible minds, acknowledging it as speculative and not universally necessary.

factualspeaker onlynovelty 0/4durability 2/4· Murray Shanahan

Let me give you a quote from that that paper, right? Or am I going to cringe? Okay, go on. Yeah. The pre-reflective, reflective, post-reflective series is not just one among many paths through the space of possible minds. Rather, the space of possible minds is structured in such a way that this is the only path through it. What are these three different stages that you laid out? Yeah. Okay. So, I I I I um I don't agree with that uh anymore.

0.22

The Eternity Foundation, co-founded by Shanahan and Bob Thurman, aims to use AI to translate lost Tibetan Buddhist texts, combining AI tools with Buddhist philosophy to investigate consciousness.

factualspeaker onlynovelty 0/4durability 2/4· Murray Shanahan

I've been talking recently to some to some uh to some Buddhists. Bob Thurman uh who is a very well-known I I just emailed him last week. In fact, we are we we were uh co-founding a foundation together called the Eternity Foundation to use AI to translate uh a lot of the the lost texts from from the Tibetans.

0.22

Jonathan B is co-founding the Cosmos organization, which funds research, incubates AI startups, and believes philosophy is critical to building technology, and is hiring for roles and running programs to build an ecosystem of philosopher-builders.

factualspeaker onlynovelty 0/4durability 2/4· Jonathan B

My name is Jonathan B. I'm a founding member of Cosmos. We fund research, incubate, and invest in AI startups and believe that philosophy is critical to building technology. If you want to join our ecosystem of philosopher builders, you can find roles we're hiring for, events, grant programs, and other ways to get involved on jonathanb.com/cosmos.