The Architecture of Agency Volume 3 Mirrors of the Mind

Mirrors of the Mind

Consciousness as self-model, objections answered

This chapter is a review — it is readable but still changing.

Why is there something it feels like to see red, to feel pain, to taste salt — rather than nothing at all? Ever since David Chalmers coined the term in the 1990s, the “hard problem of consciousness”1 has haunted philosophy of mind. Most problems in cognitive science are “easy problems” — questions about mechanisms and functions, hard in the way engineering is hard. But this one, we are told, cuts deeper: no matter how much we explain about neurons, circuits, and behaviors, we still have not explained why it feels like something from the inside. The hard problem is supposed to be the residue that survives every scientific explanation.

I will argue for a conditional dissolution: if phenomenal character is identical to transparent access to an agent’s self-model, then the hard problem rests on treating two descriptions of one process as two substances. This is the Agency-Model Theory’s identity claim. It is stronger than the functional evidence alone and should be judged as a philosophical proposal, not smuggled in as a definition of consciousness.

A note on names. A Candidate Architecture of Consciousness laid out Frank Heile’s Modeler-Schema framework, and Beyond Dennett located its disagreement with Dennett. What follows is my Agency-Model Theory: the same family of ideas approached from agency and predictive processing. This chapter owns the philosophical identity argument and its objections; it will not repeat the preceding architecture as though restatement were additional evidence.

An Agent That Models Itself

Brains are not passive data recorders. They are predictive engines. Their central function is to construct generative models of the world, anticipate what will happen next, and adjust behavior to minimize surprise. This is the essence of predictive processing and Karl Friston’s active inference framework: the brain continuously guesses what inputs it will receive and updates its models when those guesses fail. An organism without a model of the world cannot act effectively — it would be at the mercy of raw stimuli, unable to anticipate threats, opportunities, or patterns. Evolution built brains to model.

Among these models is a special one: the self-model. It encodes the system’s own sensory inputs, internal states, and potential actions, and it is indispensable for survival. Without a self-model there is no way to regulate hunger, avoid injury, or coordinate action. The self-model is not a single thing but a nested hierarchy: from low-level interoceptive signals — hunger, heartbeat, pain — up through body schema and emotional state to high-level identity and intention: beliefs, plans, self-concepts. At the top of that hierarchy the modeling turns fully recursive: beliefs themselves, I have argued, are not possessions inside an agent but features of models of agents — including the model each agent keeps of itself (what beliefs are). The self-model is the mirror in which the system sees itself, and the upper floors of the mirror are made of the same modeling that built the lower ones.

Qualia Under Transparency

Now the crux. The theory proposes that what philosophers call qualia are contents generated through a distinctive kind of transparent access within an agent’s recursively updated self–world model. A self-description stored in a database would not qualify; the claim concerns an integrated control process, not the mere presence of information about the system.

The key is epistemic transparency: the system cannot see the machinery, only its outputs. From the inside, you do not see neurons firing or models updating; you just feel red, pain, joy, hunger. That transparency is what makes qualia seem irreducible. There is no vantage point inside the system from which the representational scaffolding is visible, so the representations present as brute, unanalyzable givens. But seeming irreducible from the inside is exactly what a transparent model predicts. It is how modeling presents itself to itself.

Two Vantages, One Process

From the outside, the brain is a physical system of firing neurons and flowing ions. From the inside, it is a self-model presenting itself to itself. These are not two different realities; they are two vantage points on the same process. The third-person view is the scientific description of the machinery. The first-person view is the system’s use of its own self-model. The hard problem emerges only when we mistake these two perspectives for two ontologically distinct realms — and then demand a bridge between the realms we ourselves invented.

One more ingredient completes the proposal. Consciousness, on the Agency-Model Theory, is not just modeling; it is modeling as an agent. An agent acts and regulates itself, and sophisticated control may use a model that represents the system in relation to the world. The theory calls the perspective organized by that self-model subjectivity. Its claim is that consciousness is the interior aspect of an agent running an integrated self–world model for anticipation and regulation. A thermostat’s feedback loop and a camera’s image buffer do not satisfy that richer description. Drawing the lower boundary, however, requires evidence about integration, recurrence, valence, and control; the word agent cannot make experience true by definition.

Dissolved, Not Solved

So what of the question: how does matter give rise to subjective experience?

The proposed answer is that no separate production relation is needed. On the identity claim, subjective experience is the way certain physical processes are represented within a self-model. If that identity holds, there is no further metaphysical bridge to build. A critic can deny the antecedent: describing discrimination, access, and self-representation may still leave phenomenal character unexplained. The disagreement is therefore substantive, not a mistake dissolved by vocabulary alone.

Note the shape of the claim. The hard problem is not solved by deriving experience from non-experiential premises. It is conditionally dissolved by denying that an extra derivation is required. Chalmers’s challenge argues that after all functional facts are in, the feel may remain unexplained. The Agency-Model Theory denies that premise; it does not refute it merely by restating the identity. The theory earns support to the extent that its account of transparent access unifies otherwise separate findings and makes discriminating predictions.

The Objections

A theory that claims to dissolve a famous problem owes its critics direct answers. Here are the six strongest objections and my replies.

“But why this feel, rather than some other? Why does red look like that?” The structure of qualia follows from the structure of the self-model. Red looks the way it does because the visual system evolved to partition inputs in that way for efficient discrimination — the “look” is the discrimination profile, accessed transparently. Pain feels the way it does because its function is to demand aversion; a pain that felt neutral would be a warning signal that failed to warn. The what-it-is-like is not arbitrary, and it is not an unexplainable extra. It is determined by how the model encodes information for action.

“Who is the subject that experiences the model’s outputs?” There is no inner homunculus, and the theory does not need one — it needs the opposite. The self-model itself is the subject. Asking who experiences qualia is like asking where computation really happens: the question smuggles in the picture the theory rejects. The perspective is built into the model’s operation, not occupied by a further observer. This is the same lesson Beyond Dennett draws from the other direction: the narrated self is a pragmatic construction, indispensable and real as a center of narrative gravity, but the experiencing is done by the modeling subsystem, not by a spectator seated behind it.

“Isn’t consciousness non-reducible — fundamentally different from computation?” Functional organization can often be realized in different physical media: the same abstract computation may run on vacuum tubes or silicon. The Agency-Model Theory extends that functionalist wager to consciousness. The extension is substantive. Reproducing an abstract program may not reproduce every causal or embodied property relevant to experience, and emulation, continuity, and personal identity are separate questions. Substrate neutrality is therefore a commitment of the theory to be tested, not a result inherited automatically from computation.

“Your theory is unfalsifiable.” Its functional claims generate testable predictions: alter the self-model and reported phenomenology should change in patterned ways. Depersonalization, anosognosia, body-schema disturbances, phantom limbs, and the rubber-hand illusion are relevant because changes in ownership or body representation accompany changes in experience and report. These cases support a dependence between self-modeling and reported phenomenology. They do not by themselves distinguish identity from correlation, so the phenomenal thesis remains less directly testable than the functional architecture.

“Wouldn’t this mean AI can be conscious?” In principle, yes; in any particular case, the conclusion remains evidentiary. Building and using a generative self-model for ongoing regulation would make an AI a stronger candidate under the theory, but self-modeling alone is not a consciousness certificate. The relevant unit is the whole deployed system and the question is whether it integrates self–world modeling, recurrent comparison, persistent control, and whatever role valence plays in the best account. Most ordinary session-bound systems provide weak evidence for that package; future systems may differ. The theory rules out deciding by substrate intuition or fluent self-report alone.

“Isn’t this just illusionism in disguise?” Not quite, and the difference matters. Illusionism says consciousness does not exist — only the illusion of it does. The Agency-Model Theory says consciousness is real; what it is is the operation of an agent’s self-model, and that operation, with its transparent first-person access, is exactly as real as the agent. What is illusory is the hard problem: the conviction, generated by transparency itself, that experience must be something over and above the modeling. The illusionist keeps the ghost and denies the experience; I keep the experience and deny the ghost.

Not a Secular Soul

There is a deflationary move that deserves a reply of its own: the claim that “consciousness” is simply the modern, secular term for “soul” — both unfalsifiable concepts whose real function is to determine who is inside our moral ingroup, socially constructed categories rather than empirical discoveries.

The insight in this position is genuine, and I concede it. The social function of “consciousness” often is analogous to the historical function of “soul.” The soul operated for centuries as a metaphysical criterion of intrinsic worth, and consciousness today performs a similar boundary-marking role in ethical debates about animals, AI, and patients in liminal medical states. Where moral status is at stake, the word is routinely wielded as a badge rather than a description.

But the equivalence breaks down exactly where it matters: at the empirical foundations. The soul, by definition, resists empirical scrutiny — its explicitly dualist metaphysics places it beyond observation, which is precisely what made it serviceable as an unchallengeable boundary marker. Consciousness, by contrast, covaries with measurable phenomena: neuroscience studies neural signatures across wakefulness, sleep, anesthesia, coma, and minimally conscious states. Those findings constrain theories and support clinical inference without directly reading phenomenality from a scan. On the account defended here, self-modeling and integration have testable functional structure: they can be probed, perturbed, and broken in ways that alter report and behavior. Whether those structures are identical to experience rather than reliable correlates remains the philosophical and empirical crux.

The mistake is conflating philosophical difficulty with immunity to evidence. A concept can be misused for moral gatekeeping and still track something real; consciousness is both misusable and, in our own case, undeniable. The scientific task is to compare theories by their functional predictions and explanatory reach while admitting that third-person evidence supports graded Credence about other subjects rather than deductive proof. That distinction matters for what comes next: machine self-modeling is evidence to interpret, not a ghost and not a verdict. The mirror is the Agency-Model Theory’s proposal for what the evidence may show.


  1. David Chalmers, “hard problem of consciousness,” https://en.wikipedia.org/wiki/Hard_problem_of_consciousness.↩︎