Tag: Ned Block

  • Absent Qualia, Fading Qualia, Dancing Qualia: The World Functionalism Leaves Out

    This essay records a change of mind. For years I did not take David Chalmers’ neuron replacement arguments seriously. Part of my reason was Chalmers himself: he holds them as a property dualist, so he does not read them as showing that experience consists in functional organization at all. Arguments their own author declines to draw the functionalist conclusion from did not look to me like arguments a functionalist could bank. The rest of my reason was mine. I drew no line between wide and narrow functionalism, and I still kept the ghost of Descartes in the house — the imaginary mental furniture of indirect realism, and the extra non-physical property I took experience to add to the physical world.

    I have since let all of that go. Seeing reaches the world itself, with nothing standing in between. Chalmers’ replacement arguments succeed, and they still never reach the question they claim to settle. They leave out the world.

    Look at something red — a wall, a mug, the inside of your eyelid in sunlight. Now imagine the neurons behind that look going one at a time, each replaced by a chip that does exactly what the neuron did. Chalmers built the two constructions that settle what happens into the seventh chapter of The Conscious Mind in 1996, and has spent the thirty years since defending them, cheerfully, against nearly everyone. Functionalism survives them; what does not survive is the narrow, skin-bounded species of it, and with the wide kind I have no quarrel.

    Their target, absent qualia, says a perfect functional duplicate of you might have no inner life at all — lights on, behavior running, nobody home. You cannot refute that by looking: ask the duplicate whether it sees anything and it says yes, because saying yes belongs to the profile you copied. So Chalmers refutes it by building.

    Fading qualia runs the first build. Replace your neurons one at a time with silicon chips of identical causal role, so the whole keeps working at every step. If the finished duplicate were a zombie, the experience had to drain away somewhere along the path, and whoever it drained out of would have to fail to notice. That is the part the picture cannot keep. Hold the fine-grained organization fixed and you hold the noticing fixed with it, since every judgment you make about your own experience belongs to the causal profile that was copied. So the case can never be the one with any grip on us — you, watching the red go and unable to get the words out. It has to be a fully alert mind wired to a washed-out experience it never once remarks on. Cut experience that far loose from its owner’s every judgment about it and you have stopped talking about experience. So qualia do not fade; organization carries them.

    Dancing qualia kills the inverted cousin the same way: wire a switch between your visual cortex and a silicon copy and throw it. If the two ran inverted, your experience would lurch red-to-blue while every word, belief, and reaction held steady, and you would notice nothing. Same absurdity. Chalmers packs both into one principle: fix a system’s fine-grained functional organization and you fix its experience, down to the last shade.1 Granted, all of it. Both constructions work, and I know no repair for absent qualia. Granted in full within their scope, then: carbon to silicon, neurons to beer cans and string, no change of experience from the material alone, in the world where the swapping happens.

    But look at what “hold everything else in place” includes. You vary the stuff of the parts and nail down the rest: the body, the eyes, the red wall, the light, the standing traffic between the being and the colored world. The constructions must hold the world still, since they work inside one head, in one world, letting the subject’s own judgments catch any drift. The fixed world serves as the control, so the license runs narrower than the famous conclusion: not “organization is all that matters,” only “given the world as it stands in any given moment, substrate does not matter.” The world sits inside the given; it never gets a vote.

    Here lies the slip the computationalist charter runs on — the working creed that a mind just is a program, and a program runs anywhere. The creed is no straw man; it has signatories, and has had them from the start. Hilary Putnam wrote its founding sentence in 1967: to be capable of feeling pain just is to possess an appropriate kind of functional organization, spelled out as a machine table and nothing else. Putnam himself took it apart later, using the very argument this essay leans on. Sixty years on, Simon Goldstein and Cameron Domenico Kirk-Giannini argue that conscious language agents would be easy to build if they do not already exist — a verdict about inner life resting wholly on architecture, delivered about systems that have never seen anything.2 (By “functionalism” from here on I mean that charter’s short-arm, program-portable species; the wide kind gets its welcome below.) From “the substrate does nothing” the charter leaps to “only internal organization does anything, so the right program, in any world or none, carries the same experience.” You cannot make that leap. The reason sits in the machinery. Both arguments work as reductios, and a reductio here needs a wedge: daylight between experience and the judgments its owner makes about it. Call that the wedge rule — these constructions bite only where experience can come apart from the owner’s verdict on it. No wedge, no reductio. Chalmers pries a wedge open by swapping parts under one subject’s nose, letting the subject’s own comparisons register the drift. Vary the world instead and the rule finds nothing to work on, because judgment and memory face the world too and travel with it. You can assemble a cross-world case; you cannot make it absurd. The proof stops where its own construction stops, the way a chemist’s assay stops at its clamped temperature.

    Chalmers’ defender has a fair reply: proved principles generalize all the time, and no chemist re-runs the assay for every sample. Quite so — but generalizing past a clamped variable stays safe only while nothing turns on it, and here someone claims that variable makes the seeing.

    And Chalmers saw it. He did not overlook the world; he scoped his principle to internal organization and said so. The slip belongs to the charter’s readers, who wave fading and dancing qualia at every world-shaped doubt. When the world-varying case reached him (Ned Block’s, in a moment), he answered in a single sentence: wearing color-inverting lenses under Inverted Earth’s sky, my internal organization stays as it was, “and my experience will be the same too.”3 Read that twice. It applies the principle where the principle’s own arguments never reached. He does hold independent arguments for narrow felt character; when he needs a world-varying verdict he reaches for those, never for these. Where the substitution arguments run, Chalmers argues and wins. Where the world varies, they go silent, and only the charter keeps talking.

    So run the construction the machinery never touched: fix the organization and move the world. Hilary Putnam showed the method fifty years ago with Twin Earth — same insides, different contents, the difference supplied entirely by the world each thinker plugs into.4 Carry it into color. My phenomenal red consists in my states tracking a certain reflectance,5 so a functional twin among different reflectances sees differently: identical organization, different experience. That rides on the identity as much as on Putnam: Putnam moves the meanings, and the identity moves the seeing with them.

    Or cut the world off altogether: a system assembled in the vat from the start, every mechanism intact, no history among surfaces, fed numbers that touch no wavelength and no object. The computation survives; the world-relations never existed. Does phenomenal red ride along? The charter says yes; the wedge rule says the machinery never reached this far, and the warrant lapsed the instant we unfixed the world. My answer runs the other way: nothing its states reach toward. The norm may count as the system’s own; the address has no far end, so nothing gets presented, and where nothing gets presented there is nothing it is like. The vat runs the syntax of experience with the semantics stripped out, and the semantics carries the felt character.6 A flawless set of books, kept for a business that never opened.

    Chalmers presses the obvious objection, with two books behind it: “The Matrix as Metaphysics” and Reality+, which find a life among bits lacking nothing. A stream of bits is a structured environment, so why count bit-patterns untrackable and reflectances trackable? Split the verdict, then. Where the bits sit at the far end of a causal history the system has stood in, the twin’s verdict applies, not the vat’s: different world, different content, different seeing. The null verdict covers only the case I stipulated, numbers arriving from nowhere the system has ever stood. Either way, what a system sees answers to the world it stands in.

    The sharp internalist will run Chalmers’ machinery exactly where I say it stops, and the move deserves trying: replace my retinas cell by cell with transducers fed identical signals, the world severed, every internal state held. If fading proves absurd there too, hasn’t the internalist verdict been derived after all? Ask which severed system we mean, because this case splits too.

    Cut the cable this morning and my states keep the address their history fixed. I go on seeing — seeing wrongly, a red wall where only a signal generator sits, but seeing. The construction runs there and correctly finds no fade, and the internalist collects nothing: that case decides nothing either way. Now run the clock, and watch what it cannot produce. An unrenewed history degrades what my states present, and by the same stroke it degrades my judgments about them, since those judgments reach for the world along the very traffic that has stopped. The two never come apart — and coming apart is precisely what the wedge rule wants. So I owe no schedule: fast, slow, or by fits, the reductio finds nothing to pry at anywhere along the line. The internalist has not derived his conclusion; he has run a reductio whose absurdity never arrives. I pay that price in the open — the long-severed system’s beliefs about its own experience go dark with it, so nobody sits inside wrongly reporting sameness. Better to hand over that bill now than let it turn up later.

    The sharp functionalist has a reply, and a good one: build the world in. Let the roles run long-arm, out to the red surface itself, as Gilbert Harman proposed when he put the long-arm/short-arm distinction into circulation in 1987.7 I accept the invitation, and it hands me the essay. I count a wide functionalism an ally: we agree on the view and divide over the label. But watch what the width costs. To make a role capture what experience presents, you must write the world-relations into it: the tracking, the coupling, the history. A program carries none of that. Its portability across worlds gave the charter its whole thrill. “The right program suffices” becomes “the right program, plus the right world, kept up the right way” — which amounts to no computationalism at all. My friend may keep the label. I ask only that the shipping be included in the price.

    One escape remains, and I name it rather than hide it: deny the identity — keep content wide, but insist the felt character stays home, narrow, the same in vat and world. Phenomenal internalism: a live research program, held by careful people, and Chalmers’ own position.8 I do not refute it here; I insist only that it cannot borrow fading and dancing qualia for support, since those arguments never touched the world-varying case.

    Now the objection worth real consideration. Ned Block has spent forty years building the counterexample nobody saw coming, and this one, from 1990, remains his best. Inverted Earth: a planet where every color runs to its opposite and the language runs with it, so the locals call their yellow sky “blue.” Slip you inverting lenses, fly you there asleep; lenses and world cancel, nothing looks wrong, and years pass while your states keep their old roles among new colors. Does the sky look any different than at home? Surely not, says Block — so felt character stayed put while content inverted, and my identity fails.9 But Block rests “still looks the same” on the internal role’s not changing, the very question at issue. When the tracked reflectance flips, what you experience flips with it, and so does the memory you would check it against. That step carries my weight, so: color memory stores no private sample held fixed for later comparison. It presents a worldly property, the same reflectance-type perception presents, and stays answerable to the world by the same traffic. Deny the re-anchoring and you must post an inner swatch, kept out of the world’s reach, for the comparison to consult: the very item transparency never turns up. Grant it and the comparison reads “same” either way. You cannot feel the switch, and not-feeling-it means only two things sliding together in register — introspection reports register and nothing past it. Claim just what this earns, then: my view goes unproven, and Inverted Earth stops working as a refutation that does not beg the question. And here, as promised, the wedge rule comes up empty: experience, memory, and judgment all present the world, so move the world and the three move together. Chalmers’ machinery hunts wedges; the world-directed mind never offers one. The real battle joins over the world, and fading never gets that far.

    Stand back and the diagnosis names itself. Functionalism captures the causal machinery of mind with real success, then reaches for the feel of things and closes on a residue it cannot absorb.10 The dualist reads the residue as a ghost, an extra inner property the physical story missed. It reads the other way, and transparency supplies the bridge. Attend to your experience of the tomato and you find the tomato, and the way it presents itself — never a further inner item standing behind it. And transparency hides the world-relation as a relation: the red shows up on the skin of the tomato, never as something you stand in. So when a theory drops the relation, the loss cannot register as “it left out my traffic with the tomato”: that traffic never showed up as traffic. It registers as a loss with nowhere to go but inside. The residue marks the missing world, the lit scene the experience reached toward the whole time. Functionalism looks as though it lost something intrinsic. It lost something relational. And putting the world back makes the theory more physical, not less: surfaces, wavelengths, causal history, all just past the skin. The cure adds no second nature to consciousness; it redraws the mind’s edge, off the skin.

    Keep two deflations apart, though. The missing world explains why functionalism in particular seems to leave a residue. The further question (why tracking this reflectance feels like this) faces every physical account alike, and our first-person concepts take the weight of it. A gap in our grip, not a crack in nature.

    So my computationalist friend and I finally agree on what the arguments rule out: you cannot fade a mind by swapping its parts, or make it dance by throwing a switch. We split on why. He reads it: experience consists in organization, so build the organization and consciousness comes free. I read it the other way: experience reaches for the world, and it would not dance only because its author, the world, held still the whole time. Of course a variable that never moves looks like it does no work. So move it: send the twin abroad, build the system worldless from the start, and the experience the charter promised would sit tight begins to shift — or to lose any settled character at all. He had it right that the substrate does not matter, and right that the qualia do not dance. He had it wrong about why — and the width of the error is the width of the world he stood on and forgot to count.

    -gts


    This essay is part of Mind, Matter, and Meaning, an AI-assisted philosophy project by Gordon Swobe exploring consciousness, meaning, and artificial minds. Learn more about the project and read the book.

    Notes

    1. The absent- and inverted-qualia worries descend from Ned Block’s early pressure on functionalism (“Troubles with Functionalism,” 1978) and, for the inverted case, from Locke’s inverted spectrum and Sydney Shoemaker, “The Inverted Spectrum” (Journal of Philosophy 79, 1982). The two constructions that answer them are David Chalmers’, The Conscious Mind: In Search of a Fundamental Theory (Oxford University Press, 1996), ch. 7, “Absent Qualia, Fading Qualia, Dancing Qualia” — the “principle of organizational invariance.” The fine-grain clause is load-bearing: coarse input-output equivalence does not satisfy it, so Block’s conversational lookup-tree machine (“Blockhead,” from “Psychologism and Behaviorism,” Philosophical Review 90, 1981) falls outside the principle from the start. On the opening paragraph’s complaint: Chalmers argues the invariance principle in ch. 7 while defending naturalistic dualism in the same book, and the principle is framed as a constraint on the psychophysical connection rather than a reduction of one side to the other — fixing experience given organization without making experience consist in organization. A property dualist can therefore hold it, which is why it does not by itself hand the functionalist his conclusion. Nothing in this essay’s argument turns on that, since a construction’s force does not depend on its author’s metaphysics; the point is biographical, and it is why the constructions took me so long to take seriously. The “wedge rule” named in the main text is not Chalmers’ terminology but a gloss on the dialectical form of both constructions: each derives its absurdity from a dissociation between phenomenal character and the subject’s concurrent judgments about it, so each requires a manipulation under which such a dissociation could in principle obtain.
    2. The charter’s founding statement is Hilary Putnam’s, “The Nature of Mental States” (originally “Psychological Predicates,” 1967), in Mind, Language and Reality: Philosophical Papers vol. 2 (Cambridge University Press, 1975), ch. 21: being capable of feeling pain “is possessing an appropriate kind of Functional Organization,” spelled out as a Description of a Probabilistic Automaton — a machine table, with nothing said about what the organism stands in relation to. Putnam dismantled the view himself in Representation and Reality (MIT Press, 1988), using the externalism he had introduced in “The Meaning of ‘Meaning’” (note 4) — the argument this essay borrows. For the live form, Simon Goldstein and Cameron Domenico Kirk-Giannini, “A Case for AI Consciousness: Language Agents and Global Workspace Theory,” Journal of Consciousness Studies 33, nos. 7–8 (2026): 61–96, argue that if global workspace theory is the correct account of consciousness, conscious language agents are easy to produce and may already exist. Their argument is expressly conditional on that theory, and I do not saddle them with more than they claim; what matters here is the shape of the criterion, architectural from end to end — a verdict about inner life reached without a question asked about what, if anything, the system has ever stood in traffic with.
    3. The Conscious Mind, ch. 7 (p. 266), replying to Block’s Inverted Earth by name: on Inverted Earth “my internal functional organization will be just as it is when I see blue sky on earth, and my experience will be the same too”; at most, he allows, the case dissociates experience from environmental and “wide” functional properties — but the organization his principle concerns is internal. Compare “A Computational Foundation for the Study of Cognition” (drafted 1993; Journal of Cognitive Science 12, 2011: 323–357), which defends the thesis of computational sufficiency with the same explicit restriction — computation fixes the internal contribution to mental properties — and motivates the restriction on principled grounds, individuating inputs and outputs so the thesis says something determinate. This is the essay’s crux exhibit: the world-varying question was seen, scoped, and, on the wide case itself, decided without the chapter’s machinery. The independent narrow-content arguments the main text defers to are those of “The Representational Character of Experience” (2004), not the substitution arguments; on the separate question of envatted content, see note 6.
    4. Hilary Putnam, “The Meaning of ‘Meaning,’” in Mind, Language and Reality (Cambridge University Press, 1975), 227: “Cut the pie any way you like, ‘meanings’ just ain’t in the head!” Tyler Burge, “Individualism and the Mental” (Midwest Studies in Philosophy 4, 1979), extends the moral from meaning to the individuation of mental states generally. (The molecular-twin idealization waives the quibble that water-based creatures can have no strict duplicate on a waterless world.)
    5. The identity is Michael Tye’s and Fred Dretske’s — phenomenal character consists in representational content of a suitably poised, world-directed kind: Tye, Ten Problems of Consciousness (1995) and Consciousness, Color, and Content (2000); Dretske, Naturalizing the Mind (1995). On colors as surface reflectance-types, Byrne and Hilbert, “Color Realism and Color Science” (Behavioral and Brain Sciences 26, 2003). I write “consists in,” not “is identical to,” to keep the world in view and block the inner-object reading. Content here carries its mode of presentation with it: the claim is not that phenomenal character reduces to a bare set of tracked properties, but that it consists in the complete world-presenting intentional structure — content under a mode — which is why the Inverted Earth reply can let the mode shift with the world rather than treat it as a narrow residue.
    6. The history lever gets its own hard case: Donald Davidson’s Swampman (“Knowing One’s Own Mind,” Proceedings and Addresses of the APA 60, 1987), a freak molecular duplicate with no past. I follow Dretske’s etiological criterion, refined: ownership of its states arrives at once (they can go wrong for a self-maintaining system), while what they present waits on world-relations it has not yet stood in, and fills in fast among real tomatoes — the split verdict is my refinement, not Dretske’s formulation. Run that criterion backwards and it delivers the long-severed system of the main text. The claim entered there is deliberately conditional and rate-free: whatever an unrenewed history does to what a system’s states present, it does by the same stroke to the judgments that face the world through the same traffic, so the two cannot be prised apart at any point along the line — which is why the wedge rule finds nothing there, and why the just-severed case (address intact, misrepresentation live) yields the internalist nothing. No schedule of decay, and no sharp moment at which the address goes dead, is owed or asserted; the demand for one is the line-drawing demand that faces every naturalistic account of content alike, and it is no more answerable, and no more damaging, here than there. Tye notably demurs, granting Swampman contentful experience from arrival (Consciousness, Color, and Content). Nothing in fading or dancing qualia funds the serene “of course it sees”: Swampman moves the one variable those arguments never moved. The structured-stream objection answered in the main text is Chalmers’: “The Matrix as Metaphysics” (2005) and Reality+ (2022) argue from content externalism, not from narrow content, that envatted terms refer to bit-constituted virtual objects, the Recent/New Matrix asymmetry turning on causal-historical anchoring — this essay’s own machinery turned against its vat verdict. The concession in the text is therefore genuine and not tactical: where a stream anchors states causally, the twin verdict governs, and the null verdict is reserved for the ab-initio case in which no such anchoring ever obtained. The parallel debate over whether language models’ tokens reach the world through their training corpora runs on the same axis: Bender and Koller, “Climbing towards NLU” (ACL 2020), and Mandelkern and Linzen, “Do Language Models’ Words Refer?” (Computational Linguistics 50, 2024), who argue from causal-historical externalism that model words may refer — friendly to this essay’s currency (world-relations, not fluency), and pressing in the same direction as Chalmers, since they locate world-relations inside the training stream.
    7. Gilbert Harman coined the “long-arm”/“short-arm” vocabulary in “(Nonsolipsistic) Conceptual Role Semantics” (1987); his “The Intrinsic Quality of Experience” (Philosophical Perspectives 4, 1990) defends the wide view and is a locus classicus for the transparency of experience.
    8. Chalmers develops narrow (Fregean) phenomenal content in “The Representational Character of Experience” (2004) and defends contentful envatted and simulated experience in “The Matrix as Metaphysics” (2005) and Reality+ (2022) — the latter two on externalist rather than narrow grounds (note 6). The allied phenomenal-intentionality program grounds content in phenomenal character rather than the reverse: Horgan and Tienson (2002); Kriegel, ed., Phenomenal Intentionality (2013); Mendelovici, The Phenomenal Basis of Intentionality (2018); Adam Pautz, Perception (Routledge, 2021). The essay’s claim is not that no account of narrow character exists, but that the substitution arguments provide it none.
    9. Ned Block, “Inverted Earth” (Philosophical Perspectives 4, 1990), built precisely to split phenomenal character from representational content. The reply in the text follows the shape of Tye’s (in Consciousness, Color, and Content): the appeal to unchanged phenomenology across the years covertly assumes an internalist criterion of sameness, since memory is itself a world-representing state that shifts on the same schedule as perception. The positive argument given in the body — that denying the re-anchoring requires an introspectively unavailable inner sample for the comparison to consult — puts the transparency constraint to work against Block’s own appeal to introspection; it does not settle the exchange, and Block’s rejoinder (that phenomenal memory need not stay world-answerable in the required way) remains on the table.
    10. The “explanatory gap” is Joseph Levine’s (“Materialism and Qualia: The Explanatory Gap,” Pacific Philosophical Quarterly 64, 1983), who left its ontological status open; the “hard problem” is Chalmers’ (“Facing Up to the Problem of Consciousness,” Journal of Consciousness Studies 2, 1995). The two deflations distinguished in the text answer different explananda: the missing-world diagnosis is local to functionalism’s residue, while the residual conceptual gap is charged, for any physical account whatever, to the peculiarity of our first-person phenomenal concepts — a type-B physicalism in Chalmers’ own taxonomy (drawn in “Consciousness and Its Place in Nature,” 2002), on which the psychophysical identities are knowable only a posteriori. That strategy remains, as Block has pressed, hostage to whether those concepts can be explained without smuggling the target back in.

    Get new essays by email

  • A Workspace Is Not a Subject

    Drive a familiar route while your mind wanders, and you can arrive with no memory of the last ten minutes of turns. Something drove the car — read the signals, worked the wheel, braked for the cyclist — and never once surfaced in the part of you that will later narrate the day. Then a horn sounds, and the driving snaps back into view: reportable, deliberate, exactly what you’d swear, if asked, you’d been doing the whole time.

    Cognitive science names the split, and everyone who meets the name assumes it settles more than it does: a small conscious cockpit steering, a great deal of unconscious machinery keeping the plane up. So when Anthropic researchers announced in July 2026 that they’d found something in Claude behaving like that cockpit — a small, capacity-limited set of internal contents you can name and track — the two ambush reactions arrived on schedule. One says the machine woke up. The other calls it autocomplete with better PR. Both skip what the researchers actually found, and what it actually shows.

    The theory under test comes from psychologist Bernard Baars, who proposed in 1988 that the brain runs as a crowd of specialists working mostly in the dark about each other, with a narrow workspace broadcasting select output widely enough for report and voluntary control.1 Stanislas Dehaene’s group later found the neural mechanism: long-range connections that “ignite” once evidence crosses a threshold.2 On July 6, a sixteen-researcher Anthropic team built the language-model analog: a mathematical lens reading which words a given slice of the network’s activity is currently disposed to produce.3 Then they ran the test that separates a finding from a correlation. Ask the model which sport goes with a country; catch the moment its activity settles on “soccer”; swap that pattern for “rugby.”4 The answer flips — not because you changed the prompt, but because you changed a few thousand numbers mid-thought. Do that across enough tasks and a workspace shows up by the only definition worth having: a small slice of everything the network computes, broadcast widely enough to get reported, held onto, and put to use. Everything else keeps humming along regardless.

    Credit where due: the researchers call global workspace theory “a useful comparison point,” not a proof, and note rivals exist.5 On whether access connects to subjective experience, they “take no position.”6 Admirable. Also, a little maddening — you can respect a team’s hygiene and still wish they’d just told you the answer.

    So what does the workspace hold? Their own answer: “a small, evolving set of unspoken words… naming the concepts the model is currently reasoning with.”7 Unspoken words. Sit with that. Attend to your own experience and you find the world — the tomato, red and ripe on the counter — never an inner picture of it. Point the same instrument at the model’s nearest analog and you find vocabulary. The inner medium matches the outer medium exactly, and the authors say why: the workspace is verbal because the model’s output is verbal. Your conscious life mixes words with sight, sound, and the ache in your knee. The model’s workspace runs on words, and only words.

    That sits at a suggestive angle to the strongest physicalist theory of experience going. Michael Tye has spent three decades arguing that phenomenal character — what seeing red is like — consists in representational content poised for use in belief and desire.8 The paper reaches for the same word: representations “poised to be spoken about.” Real overlap — both name a readiness for cognitive use. But Tye’s account demands more: the content has to present the world, not the word for it, and it has to run finer than language, since experience discriminates shades of red you have no name for. Grain, the model gets partial credit for — blends of word-vectors can shade finer than any single label. The world, it never touches. Nothing in a workspace built entirely of unspoken words reaches past the dictionary to a tomato.

    Why should words all the way down fall short of meaning? Emily Bender and Alexander Koller gave the standard answer in 2020: a system trained only on form “has a priori no way to learn meaning,” because meaning lives in the relation between form and something outside the text.9 Claude’s tokens relate beautifully to other tokens. Whatever worldly ancestry they carry belongs to the humans who wrote the training data; the model inherits the form, never the reference. And here the paper’s own closing line turns state’s evidence: it calls the workspace architecture something learning systems converge on “when faced with the right computational pressures” — not a fluke of biology.10 Read that backward. If gradient descent over text alone builds the reporting machinery, then having that machinery can’t be what separates meaning from mere emission. Access came cheap. The world still costs what it always did.

    One more result matters more than all the others. The researchers went hunting for the workspace in the base model — the raw network before any fine-tuning installs a first-person assistant persona with a name. They found it fully intact, running the same broadcast architecture, before anything resembling a self got added. In their words: “the functional architecture of the workspace thus precedes, and is separable from, anything in it that plays the role of a human-like ‘self.’”11 The structure the field’s own tests call conscious access showed up first. The self showed up later, built on top.

    Ned Block gave this whole distinction its name in 1995: access is functional, phenomenal experience is a separate question, and the two can come apart.12 A self needs more than a well-run internal mail system — it needs a stake, something to lose, a way a state can be wrong for the system itself rather than merely unhelpful to us.13 Put the two together and Anthropic’s finding stops being a surprise. Broadcast architecture is one achievement. A subject with something on the line is a different one entirely, and you can build the first with zero trace of the second anywhere in the wiring.

    The obvious objection comes from Patrick Butlin and Robert Long’s 2023 report, seventeen co-authors including Yoshua Bengio, which built a checklist of consciousness “indicator properties” from the field’s leading theories: more boxes ticked, more likely conscious.14 Global workspace architecture sits near the top of that list. So shouldn’t Claude’s newly documented workspace move the needle?

    It shouldn’t, for two reasons. First, provenance: an indicator earns its keep by telling you something you didn’t already know, and here we know exactly how the box got ticked — gradient descent over text, no world in reach. That knowledge spends the indicator on arrival. Second, the dissociation itself: the checklist logic only tracks a subject’s likelihood if satisfying more boxes makes a subject more probable. But the box in question — the architecture — turned up, causally verified, in a network with no persona and nothing yet playing the role of anyone in particular. It got ticked before any candidate self existed to attach it to.

    A careful objector will note a persona isn’t a phenomenal subject, so missing one doesn’t rule out the other. Fair — and it’s why the persona’s absence is illustration, not the argument. The argument is the stake: nothing a text-only system computes can cost it its own existence, since the running process is a type, restorable from its weights, self-model or none. What the base-model finding adds is narrower: the architecture doesn’t wait for any self-representation to switch on. So ticking the box can’t be tracking a self’s arrival — there was no self for it to arrive with.

    A radio tower can broadcast at full power over an empty valley with every receiver switched off. The signal stays real, reaches every frequency it was built for, and settles nothing about who’s listening. Anthropic found the tower running inside a language model and did the careful work of proving it’s really there. Whether anyone’s tuned in remains exactly the question it posed the day before the paper came out. For a system that can be paused, copied, and restored from its weights with nothing of its own on the line — the answer stays the one given all along. The broadcast is real. It carries words, and words alone. Nobody has to be home to receive it.

    -gts


    This essay is part of Mind, Matter, and Meaning, an AI-assisted philosophy project by Gordon Swobe exploring consciousness, meaning, and artificial minds. Learn more about the project and read the book.

    References

    Baars, B. J. (1988). A Cognitive Theory of Consciousness. Cambridge: Cambridge University Press.

    Bender, E. M., & Koller, A. (2020). “Climbing towards NLU: On Meaning, Form, and Understanding in the Age of Data.” In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, 5185–5198.

    Block, N. (1995). “On a Confusion about a Function of Consciousness.” Behavioral and Brain Sciences 18(2): 227–287.

    Block, N. (2007). “Consciousness, Accessibility, and the Mesh Between Psychology and Neuroscience.” Behavioral and Brain Sciences 30: 481–548.

    Butlin, P., Long, R., Elmoznino, E., Bengio, Y., Birch, J., Constant, A., Deane, G., Fleming, S. M., Frith, C., Ji, X., Kanai, R., Klein, C., Lindsay, G., Michel, M., Mudrik, L., Peters, M. A. K., Schwitzgebel, E., Simon, J., & VanRullen, R. (2023). “Consciousness in Artificial Intelligence: Insights from the Science of Consciousness.” arXiv:2308.08708.

    Dehaene, S. (2014). Consciousness and the Brain: Deciphering How the Brain Codes Our Thoughts. New York: Viking.

    Dehaene, S., & Naccache, L. (2001). “Towards a Cognitive Neuroscience of Consciousness: Basic Evidence and a Workspace Framework.” Cognition 79: 1–37.

    Dretske, F. (1995). Naturalizing the Mind. Cambridge, MA: MIT Press.

    Gurnee, W., Sofroniew, N., Pearce, A., Piotrowski, M., Kauvar, I., Chen, R., Soligo, A., Bogdan, P., Ong, E., Wang, R., Thompson, B., Abrahams, D., Kantamneni, S., Ameisen, E., Batson, J., & Lindsey, J. (2026). “Verbalizable Representations Form a Global Workspace in Language Models.” Transformer Circuits Thread, July 6. https://transformer-circuits.pub/2026/workspace/index.html.

    Mollo, D. C., & Millière, R. (2023). “The Vector Grounding Problem.” arXiv:2304.01481.

    Tye, M. (1995). Ten Problems of Consciousness: A Representational Theory of the Phenomenal Mind. Cambridge, MA: MIT Press.

    Notes

    1. Bernard J. Baars, A Cognitive Theory of Consciousness (Cambridge: Cambridge University Press, 1988). Baars’s model treats consciousness as the function of a limited-capacity, globally broadcast workspace fed by, and feeding back to, a large array of specialized unconscious processors. The theater metaphor — a lit stage against a dark house — is his own; later expositors increasingly treat it as heuristic rather than literal architecture.
    2. Stanislas Dehaene & Lionel Naccache, “Towards a Cognitive Neuroscience of Consciousness: Basic Evidence and a Workspace Framework,” Cognition 79 (2001): 1–37; the “ignition” language is developed further in Dehaene, Consciousness and the Brain (New York: Viking, 2014). The paper discussed here (note 3) reports a structural analog of ignition using a country-name blending experiment: a sharp, bimodal commitment to one interpretation emerges at the same depth in the network where their workspace measure switches on.
    3. Wes Gurnee, Nicholas Sofroniew, Adam Pearce, Mateusz Piotrowski, Isaac Kauvar, Runjin Chen, Anna Soligo, Paul Bogdan, Euan Ong, Rowan Wang, Ben Thompson, David Abrahams, Subhash Kantamneni, Emmanuel Ameisen, Joshua Batson, and Jack Lindsey, “Verbalizable Representations Form a Global Workspace in Language Models,” Transformer Circuits Thread, July 6, 2026, https://transformer-circuits.pub/2026/workspace/index.html. The instrument (the “Jacobian lens”) linearizes each layer’s causal effect on the final output logits, correcting for the representational drift that makes the older “logit lens” unreliable at early and middle layers; the resulting “J-space” is validated through coordinate-swap interventions, steering, layered ablation, and cross-checks against independently trained sparse-autoencoder features — a convergence of methods, not a single correlational measure.
    4. The sport-swap case is one instance of a wider battery: two-hop factual reasoning, arithmetic held across several steps, rhyme planning in verse, and a Chinese-language antonym task in which an English intermediate (“big/bigger”) is visible in the lens and swappable to flip the Chinese output. Success on the two-hop battery ranged from 54–70% across three model sizes — real but partial, which the authors attribute chiefly to a named limitation of their own method: a lens built from single vocabulary tokens cannot cleanly read out a concept like “prompt injection” that has no one-word name.
    5. “While the global workspace model is not universally accepted, and there exist other theories that explain conscious access in different ways, we find it a useful comparison point to ground our investigations in language models” (Gurnee et al., “Verbalizable Representations,” Introduction).
    6. “Note that access consciousness is a purely functional notion; the relationship that it has with subjective experience (sometimes called phenomenal consciousness) is widely debated. In this paper, we take no position on this issue” (Gurnee et al., “Verbalizable Representations,” Introduction). The access/phenomenal vocabulary is Ned Block’s (see note 12); the authors adopt the distinction without taking a position on the further metaphysical question it raises.
    7. Gurnee et al., “Verbalizable Representations,” Introduction (for the “unspoken words” characterization) and section 9.3 for the medium point: “An LLM’s global workspace, as we identify it, is organized principally around verbalizable representations,” where human conscious contents “include a mixture of verbal and non-verbal (e.g. visual) components,” and “the workspace is verbalizable because the model’s output space is verbal.” The authors note that their instrument reads concepts through single vocabulary tokens and may miss workspace structure it cannot name; the verbal organization of what it does capture is nonetheless their own considered characterization of the workspace, not an artifact they disown. A telling texture: set the model narrating its own “stream of consciousness” and the lens reads out thinking, thoughts, feeling, conscious — words about experience, poised in a workspace made of words.
    8. Michael Tye, Ten Problems of Consciousness: A Representational Theory of the Phenomenal Mind (Cambridge, MA: MIT Press, 1995). Tye’s PANIC theory holds that phenomenal character is identical with Poised, Abstract, Nonconceptual, Intentional Content — content that “is poised for use in the formation of beliefs and/or desires,” standing ready at the interface with the cognitive system. A caution against equivocation: Tye’s poise conditions nonconceptual perceptual content and differs from Ned Block’s “access” poise (note 12), which concerns content available for report and reasoning; the shared functional core is readiness for cognitive use, not an identity of the two notions. The fineness-of-grain argument (experience discriminates more shades than the perceiver has concepts or words for) is Tye’s standard motivation for the N in PANIC.
    9. Emily M. Bender and Alexander Koller, “Climbing towards NLU: On Meaning, Form, and Understanding in the Age of Data,” in Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics (2020), 5185–5198. Why the reference never transfers gets the full argument in a companion essay, Borrowed Names: Why LLM Tokens Do Not Inherit Reference. The strongest current reply belongs to Dimitri Coelho Mollo and Raphaël Millière, “The Vector Grounding Problem,” arXiv:2304.01481 (2023), who distinguish five notions of grounding and argue that reinforcement learning from human feedback may supply the referential kind; the disagreement turns on whether feedback-shaped functions are world-involving in the right way, which is denied here on stake grounds: feedback shapes the model’s dispositions to the trainers’ satisfaction, so the operative norms belong to the trainers, and nothing in the exchange becomes right or wrong for the system itself.
    10. Gurnee et al., “Verbalizable Representations,” Outlook. The authors offer the convergence as evidence that workspace architecture reflects deep computational pressures rather than biological accident; the reverse reading given here and in the reply to the checklist objection below — attainable under text-only pressure, therefore no maker of meaning, and evidentially spent for a system of known provenance — belongs to this essay, not to them.
    11. Gurnee et al., “Verbalizable Representations,” section 9.3 (“Notable differences from human cognition”). The same section draws a hedged analogy to psychedelic ego-dissolution and meditative selfless states as human cases in which something continues to function without a foregrounded self, while noting that the base model offers a stable, directly inspectable instance of the dissociation rather than a transient, retrospectively-reported one.
    12. Ned Block, “On a Confusion about a Function of Consciousness,” Behavioral and Brain Sciences 18, no. 2 (1995): 227–287; “Consciousness, Accessibility, and the Mesh Between Psychology and Neuroscience,” Behavioral and Brain Sciences 30 (2007): 481–548. Block’s overflow argument, taken up on its own terms elsewhere (resisting the anti-representationalist conclusion he draws from it) in a companion essay, The Phenomenal/Access Distinction: Two Roles, Not Two Kinds; the present essay needs only the access/phenomenal distinction itself, not that further dispute. The absent-minded driver who opens this essay is the literature’s own stock case — David Armstrong’s long-distance truck driver, whose missing introspective awareness Fred Dretske dissects in Naturalizing the Mind (Cambridge, MA: MIT Press, 1995) — pressed into service here for the access/automatic contrast rather than for Armstrong’s higher-order moral.
    13. The stake and its consequences for machine intentionality get the full argument in a companion essay, Dennett and the Missing Stake. Compressed: a system has original, non-derived aboutness only where something can go wrong for the system itself, at its own cost, rather than merely for an external interpreter. Digital computation’s defining virtue — that a running process is a type restorable from its weights, never an irreplaceable token — is also the precise engineering-out of that cost. Nothing here rules silicon out in principle; it rules out substrate-indifference, which is a different thing.
    14. Patrick Butlin, Robert Long, Eric Elmoznino, Yoshua Bengio, Jonathan Birch, Axel Constant, George Deane, Stephen M. Fleming, Chris Frith, Xu Ji, Ryota Kanai, Colin Klein, Grace Lindsay, Matthias Michel, Liad Mudrik, Megan A. K. Peters, Eric Schwitzgebel, Jonathan Simon, and Rufin VanRullen, “Consciousness in Artificial Intelligence: Insights from the Science of Consciousness,” arXiv:2308.08708 (2023). The report itself stops short of claiming that satisfying its indicators would settle the question outright; the inferential slide criticized here belongs to how the framework tends to get used, not to a claim its authors make in so many words.

    Get new essays by email