TOPIC 14 · PHASE B CONSCIOUSNESS

Reasoning Backwards from "What It's Like"

The last few issues asked "which brain region has to light up for consciousness." This school does the opposite: first nail down the ironclad laws every experience must obey, then deduce backwards — what kind of physical thing could ever deserve them.

2026-07-20 · BigCat

A million-pixel camera "sees" far more sharply than you do, yet feels nothing. And the reason you have experience at all may be just this: the stuff in your head has fused into one whole that can no longer be pulled apart.

Last issue's global workspace started from function: broadcast the information, make it usable, and that counts as conscious. This issue's Integrated Information Theory (IIT — proposed by Giulio Tononi, championed by Christof Koch) takes an almost opposite, and bolder, road. It doesn't ask "what does the brain do to produce consciousness." It turns around and asks: does "experience" itself, wherever it occurs, have a few properties it can never escape? For instance — your experience right now is one whole. You can't feel the left half of your visual scene and its right half as two separate, independent experiences; it comes fused, un-splittable. IIT lifts out these "any experience must be like this" properties and treats them as axioms, then reasons backwards: what structure must a physical system have to actually possess them? The counterintuitive landing point is that consciousness turns out to have nothing to do with being smart or being able to talk, and everything to do with one thing: whether this pile of stuff has fused into a single whole that can't be taken apart.

// 01

Flip the arrow: don't start from the brain, start from experience

Almost every other theory works outside-in: scan the brain, find which activity shows up together with "you reported seeing it," and call that the marker of consciousness. IIT thinks that road is wrong at the root — it says the one thing you're 100% certain of is not the brain, but that you are having experience. If experience is the only bedrock, you should start from it, not from a lump of brain cells you can only observe by way of experience in the first place.

So IIT sits down and asks: does any experience, whatever its content, share a few always-true properties? It lists several (it calls them axioms); take the three easiest ones. One, your experience really exists, and exists for itself — it doesn't need anyone watching from outside to count. Two, it's one whole, indivisible — what you experience is one complete scene of "blue, round, on the left," not a jigsaw of "blue" + "round" + "left"; you simply can't experience those separately. Three, it's extraordinarily specific — this exact scene amounts to simultaneously ruling out nearly infinitely many other possible scenes, and it's that "this one and not the others" that gives it content. Having written these down, IIT makes the key move: reason backwards — a physical system that truly deserves "exists, unsplittable, ultra-specific" must satisfy matching hard conditions in its causal makeup. For the first time, consciousness research runs from the conclusion back to the premises.

Usual scan brain find correlate = marker IIT · backwards axioms of exp. what substrate must be exist·unsplit two roads, opposite directions — brain→experience vs experience→substrate
Standard research scans from brain toward experience; IIT flips the arrow — first fix the axioms of "experience must be like this," then deduce what the physical substrate must satisfy
// 02

Φ: the part the whole has beyond its parts

So how do you measure "un-splittable"? IIT gives it a number, Φ (the Greek letter phi). The intuition is simple: look at the causal power a system has as a whole, versus the causal power you get by cutting it into pieces that each act alone — and ask how much is left over. That leftover — causal structure the whole has that no part does — is Φ.

Put a camera next to a brain and it's obvious at a glance. A camera has millions of pixels, each measuring its own light, none caring about the others. Snip that sensor in half and the left half still images, the right half still images — nothing is lost, because there never was a "whole," just a million mutually indifferent little islands stacked together. Such a system has Φ≈0: it records, yet comes apart cleanly, with no "only-exists-together" thing — so (per IIT) it experiences nothing. Your brain is the reverse: in vision, "red," "shape," and "location" are woven together over and over by the circuitry, so tightly that you can't even feel them apart. Actually cut that web at its weakest seam and the causal structure lost is enormous — this "cut it and a big chunk collapses" indivisibility is high Φ. The amount of consciousness = Φ; and its specific content (why red feels like red) = the shape that cause-effect structure takes.

Camera · islands Φ ≈ 0 cut anywhere · lose nothing no whole · no experience Brain · woven web Φ > 0 cut the weakest seam · a chunk collapses an unsplittable whole · has experience
Φ measures "the causal power the whole has beyond its parts": a camera's millions of pixels are independent, so cutting loses nothing (Φ≈0); the brain's circuits are woven, so even the cheapest cut collapses a big chunk (Φ>0)

AI cross-read

Here's IIT's sharpest cut at AI: it measures how the physical substrate itself is wired, not what it's computing. Two machines can compute the same function and both get it right, but one is highly interwoven and the other is an assembly line — and their Φ can differ wildly. The chips running today's large models are essentially a fetch-an-instruction, push-through-the-gates one-at-a-time pipeline: at each step causation flows almost purely forward, rarely looping back into an indivisible whole. So by IIT, the Φ of the hardware is negligibly low. That collides head-on with last issue's functionalism: computing correctly ≠ having experience. The question was never "what does it compute," but "has this physical stuff fused into one piece."

// 03

Follow Φ and you run into some startling conclusions

Take this recipe seriously and a string of unsettling conclusions follows — and some of them actually match clinical fact.

First: more neurons doesn't mean more consciousness. Your cerebellum holds roughly 80% of the brain's neurons, yet it's wired like rows of separate conveyor belts with little cross-looping — low Φ. And the clinic agrees: destroy the entire cerebellum and a person stays awake and conscious, just uncoordinated. What actually carries consciousness is the highly looped, interwoven "hot zone" of posterior cortex — high Φ there. Counting parts is useless; count how much they're fused into one.

Second, and harshest: a purely feedforward system must have Φ=0. As long as information runs strictly forward and never loops back (each layer feeds the next, never returning), then however beautifully it does the task, IIT says it has zero shred of experience — a flat-out "philosophical zombie." Third, even wilder: Φ is continuous, so any tiny bit of irreducible causal wholeness carries a tiny flicker of experience. Consciousness is then no longer a human monopoly but a gradient smeared across all of nature — a conclusion that pushes IIT all the way into "panpsychism" territory (more on that below).

Φ (integration) → cameramegapixels·Φ≈0 feedforwardsmart·Φ=0 cerebellummost neurons·low Φ post. hot zonewoven·high Φ counterintuitive: not count, not cleverness — only how fused it is
Along the Φ spectrum the placements are all counterintuitive: a megapixel camera and any clever feedforward net sit at zero, the neuron-richest cerebellum is low, and only the highly looped posterior hot zone stands high

AI cross-read

Aim this at today's AI and the conclusion is uncomfortably extreme: a language model can write poems, reason, and pass any behavioral test — yet as long as its physical execution is a feedforward pipeline, IIT flatly declares it Φ=0, not a shred of inner experience — however human-like, an empty shell. This is the exact opposite of last issue's functionalism: that camp says "can broadcast, can recruit" earns consciousness; this camp says "behavior can fool everyone, but it can't fake a single unit of Φ." The clash forces a real question: should consciousness be scored by "what it manages to do," or by "how its physical substrate is built"? Which side you pick decides whether you believe AI could ever truly "wake up."

// 04

Uncomputable, and branded pseudoscience by 124 signatories — so why is it still standing?

IIT's troubles are plain. First, Φ is essentially uncomputable for a real brain: to find that "cheapest cut" you'd have to try every possible way of partitioning the system, and the cost explodes with neuron count — a few hundred units is already astronomical, and 86 billion neurons is out of the question. So the full theory can't be tested directly on a human brain.

Second, the picture it implies is too counterintuitive: since only causal structure matters, a big grid of barely-active logic gates, wired loopily enough, could in principle have higher Φ than you — which computer scientist Scott Aaronson wielded as a reductio, "that's just absurd." In September 2023, 124 scholars co-signed a letter branding IIT outright "pseudoscience," on the grounds that it's unfalsifiable and its panpsychist implications untestable; another camp fired back at once: precisely because it dares to make predictions that specific and that counterintuitive, it's the least pseudoscientific of the bunch. That fight isn't over.

And yet it isn't just talk. The core intuition IIT forces out — "consciousness = integrated and differentiated" — has landed as a working clinical tool: the Perturbational Complexity Index (PCI). Give the brain a sharp magnetic "zap," watch whether the electrical activity it stirs up spreads into a complex echo or just dully passes through and dies; then compress that echo and measure how complex it is. This "zap it, zip it" method can measure whether consciousness is still present in patients who can't speak or move — it tells a vegetative state from a minimally conscious one, and general anesthesia from wakefulness. So don't bury IIT yet: it's the theory currently roasted hardest over the fire, yet one that has genuinely produced clinical use. Set back at Topic 11's "front vs back" fork, it stands on the posterior side, betting head-to-head against the workspace theory's wager on prefrontal cortex — and so far neither has fully won.

🌀 Crossing over · interdisciplinary echoes

"The whole can't be split; each existence is its own point of view" — several old traditions had already felt their way to this door:

// Going deeper

If Φ can't be computed for a real brain, does IIT even count as "falsifiable science"?
This is the heart of the whole dispute. Critics say: how do you get an experiment's hands around the throat of a theory that can't even measure its own core quantity? But defenders hit back hard: uncomputable is not unfalsifiable — IIT has made a string of predictions specific enough to be slapped down (low-consciousness cerebellum, posterior hot zone, feedforward-must-be-zero, the front-vs-back bet), and those are being tested one by one. The truly thorny part isn't "can't compute Φ," but that its most startling implications (a heap of logic gates being conscious) may never touch an experiment — the line where science ends and metaphysics begins runs right through IIT.
A big grid of nearly-motionless logic gates could have higher Φ than a human, hence be "more conscious" — believable?
This is Aaronson's famous counter. Your gut almost certainly shouts "absurd." But saying exactly where the absurdity lies is hard: if you say "it does nothing, behaviorally it's like a rock," IIT replies — consciousness has nothing to do with behavior, only with causal structure; that's its thesis, not a bug. So you're cornered into a choice: either accept that "a dumb-looking thing can be highly conscious" (swallow the counterintuition), or accept that "experience must be tied to behavior/function" (in which case you've quietly walked back to last issue's functionalism). No free lunch — it forces you to show where you've really staked consciousness.
IIT says a feedforward net is necessarily Φ=0, a zombie however smart; do you believe that, or last issue's "broadcast = conscious" functionalism?
This is Phase B's sharpest fork, and the two camps hand AI opposite verdicts. Functionalism: consciousness is about what it manages to do — an agent that can broadcast and recruit across modules qualifies. IIT: consciousness is about how the physical substrate is built — the same behavior, run on a feedforward pipeline, is a Φ=0 shell. The trouble is that these two verdicts can never be told apart from the outside: a "genuinely experiencing system" and a "behaviorally identical Φ=0 zombie" would say and do exactly the same things. Which you believe rests in the end not on data but on whether you think consciousness is essentially a function, or a physical way of existing.
Reasoning from "experience must be like this" axioms back to a physical substrate — is that direction even legitimate, or does it smuggle the conclusion into the premise?
This is the deepest challenge to IIT's method. Critics ask: what guarantees those few "axioms" hold for all experience rather than just being a bias of human experience? Pick the axioms wrong and everything deduced afterward, however rigorous, is wrong. Supporters counter that starting from "I'm certain I'm experiencing" — the one indubitable point — is firmer than starting from "I'll just trust that an external brain really exists." The disagreement is ancient: it's a modern replay of Descartes' "I think therefore I am" versus natural science's "first trust the world" — only this time the stakes are a formula that claims to compute consciousness.

// Further reading