Home · Writing · Consciousness

The Hard Problem and the Price of Every Exit

A decision map of the hard problem that credits what each exit explains, then records the assumptions, consequences and unresolved work it inherits.

TLDR

  1. A decision map of the hard problem that credits what each exit explains, then records the assumptions, consequences and unresolved work it inherits.
  2. Imagine an artificial clinician that predicts every pain report, withdrawal, stress response and later memory of a patient.
  3. Its sharpest pressure comes from absent or inverted qualia. A system could appear to occupy every causal role while allegedly lacking experience or having systematically different experience.
  4. What remains opaque is why the identity should feel intelligible. Learning that pain is a neural pattern may identify the state without revealing why that pattern feels painful.
  5. The difficult reply begins with Moorean force. It seems more certain that pain hurts than that any theory is true.
A completed functional bridge ending before felt quality A strong bridge crosses prediction, report, control, memory and self-model stages. It reaches a cliff labelled complete functional account. A final luminous island labelled felt quality remains across a narrow explanatory gap, with seven toll tickets suspended above possible exits. PredictactionExplainreportControlstateRecalleventModelself Functional account complete Feltquality Gap FunctionalismIdentityIllusionismEmergencePanpsychismNeutral monismCognitive closure
Figure 1. Functional explanation can be complete on its own terms while the phenomenal question remains contested. Each exit crosses, dissolves or relocates the gap, and each carries a different invoice.
On this page

Imagine an artificial clinician that predicts every pain report, withdrawal, stress response and later memory of a patient. It knows when morphine will help, when attention will amplify distress and when a silent patient is still likely to suffer. Its internal model is so exact that no available behavioural or neural test can distinguish its forecast from the patient’s future.

The achievement would be extraordinary. It would explain discrimination, report, control, learning and many causal roles associated with pain. One question could still remain: why does the patient’s state hurt? Why is there a felt quality at all, rather than an equally competent sequence of physical events with nobody undergoing it?

That residual question is the hard problem. Chalmers’s original formulation distinguished functional problems that standard cognitive and neural explanation can address from the problem of experience. Levine’s earlier explanatory-gap argument made a related point: even if pain is identical with a physical state, the identity can feel explanatorily opaque. These claims are disputed. The dispute is part of the subject.

An exit receives credit for the work it performs. It also receives an invoice: extra commitments, revised intuitions, hard cases and the next explanation it owes. Functionalism buys substrate flexibility and pays in absent-qualia pressure. Identity theory buys ontological economy and pays in explanatory transparency. Illusionism buys continuity with scientific explanation and pays by revising what introspection seems to reveal. Panpsychism buys phenomenal continuity and pays a combination problem. Cognitive closure buys humility and pays with limited guidance.

The ledger changes the conversation. “Solved” becomes too coarse. A view can dissolve one formulation, preserve another, or recast the explanandum. A theory that says the question is malformed has work to do: it must explain why the malformed question is so stable, why phenomenal judgement has its peculiar force and which ordinary claims about pain remain true. A theory that treats consciousness as fundamental also has work to do: it must explain lawful structure, subject boundaries and the link to physical organisation.

The article maps philosophical positions and the empirical constraints they accept; it does not report a new consciousness detector. Authoritative philosophy sources establish each position’s central commitments. Neuroscience can test structure, dependence, access and report. The exit-cost ledger, figures, AI scenario and cross-theory decision rule are proposed analytical instruments.

Part I. Name the invoice before choosing the exit

“Hard” does not mean merely difficult. Detecting awareness under anaesthesia, modelling attention or explaining report can demand decades of empirical work. They remain functionally specified problems: a successful account says which mechanism performs which role and why an intervention changes the outcome. The hard problem asks whether such an account entails, explains or leaves untouched the presence and character of experience.

Three claims must be separated. The first is epistemic: we currently do not see how physical and functional facts explain felt quality. The second is conceptual: no possible set of those facts would entail phenomenal facts. The third is metaphysical: phenomenal facts are fundamental or non-physical. One can accept the first while rejecting the other two. An identity theorist may regard the gap as a limitation in our concepts. A functionalist may expect future theory to make the identity intelligible. A dualist may take the gap as evidence of ontological difference.

Claim What it says What it does not yet establish
Current ignorance We lack a satisfying explanation That no explanation is possible
Explanatory gap Physical descriptions do not reveal why experience feels as it does That physicalism is false
Conceivability gap A functional duplicate without experience seems conceivable That such a duplicate is metaphysically possible
Ontological gap Experience is not wholly physical or functional Which positive ontology is correct
A prism separating one consciousness question into four claims A white beam labelled why experience enters a triangular prism and separates into four coloured rays: empirical ignorance, explanatory gap, conceivability and ontology. Small shutters show that accepting one ray need not open the next. Whyexperience? Current ignoranceExplanatory gapConceivability gapOntological gap A stronger conclusion needs an additional premise
Figure 2. The hard problem is often discussed as one beam. The prism exposes four claims and the extra premise required to move from one to the next.

This prism matters because empirical success and metaphysical conclusion operate at different levels. A complete neural correlate of pain would answer where and when pain occurs in humans. A causal model might explain how pain changes attention, learning and action. Neither result alone says whether that causal organisation constitutes feeling, reliably accompanies feeling, or merely supports report.

The access and phenomenal distinction is useful here but should not become dogma. Access concerns information available for reasoning, report and control. Phenomenal consciousness concerns what experience is like. Some theorists argue that apparently overflowed phenomenology reflects partial or fragile access rather than a separate rich field. The debate changes the map, but it does not licence replacing every phenomenal question with a report metric.

The ledger therefore has four columns. Explanandum records what the position says needs explaining. Purchase records what becomes easier or more coherent. Price records the commitments and counterintuitive consequences accepted. Inherited debt records the next unsolved problem. The language of price is not pejorative. Every serious theory pays because every theory constrains possibility.

Open the scope and neutrality note

The map covers influential families, not every position or hybrid. Representationalism, enactivism, dual-aspect views, idealisms, biological naturalism and quantum proposals can alter several routes. The article separates them when the distinction changes the invoice and otherwise treats them as variants. “Price” means an explanatory commitment, not a refutation. A position can rationally accept a high price when it buys a more important form of coherence.

Part II. Seven exits and what waits outside

The exits are easier to compare when they are placed around the same question. Assume that a system has the entire functional profile associated with a conscious human: discrimination, global availability, self-report, learning, emotion-like regulation and flexible action. What, if anything, remains to be explained?

Seven mountain passes leaving the same explanatory basin A central dark basin labelled functional facts is surrounded by seven differently coloured mountain passes. Each pass leads to a named purchase and has a toll marker naming its inherited difficulty. Completefunctional factsWhat remains? AbsentqualiaConceptgapIntrospectionmodelEmergencelawCombinationproblemIntrinsicnatureResearchceiling FunctionalismIdentity theoryIllusionismEmergencePanpsychismNeutral monismClosure
Figure 3. The exits share a starting basin and pay at different passes. A toll names the problem a route inherits, not a verdict against taking it.

Functionalism: experience is the right causal role

Begin with the silicon replacement. If every causally relevant relation survives, what could still be missing? Functionalism makes the realised role, rather than the material label, decisive. Pain is caused by damage-like inputs, interacts with beliefs and desires, produces avoidance and learning, and participates in the right wider architecture. The Stanford Encyclopedia account describes a family rather than one formula. Reproduce the fine-grained organisation and the relevant mental state should come with it.

Its sharpest pressure comes from absent or inverted qualia. A system could appear to occupy every causal role while allegedly lacking experience or having systematically different experience. Functionalists respond in several ways. Some deny that a stable full duplicate without matching experience is genuinely conceivable. Some define the functional role more finely, including internal discrimination and self-knowledge. Others accept that functional explanation answers the only coherent demand. The inherited work is to show why the specified organisation constitutes feeling rather than merely tracking every use of the word.

Identity theory: experience is a physical state

The identity theorist refuses to turn an explanatory feeling into a second substance. Pain is one physical process described through two modes of access, not a mental passenger attached to a brain event. Scientific identity can be discovered even when the terms on each side feel different. The dedicated mind-brain identity theory literature sets out this proposal and its variants.

What remains opaque is why the identity should feel intelligible. Learning that pain is a neural pattern may identify the state without revealing why that pattern feels painful. Levine’s gap targets precisely this residual. Phenomenal-concept strategies argue that special first-person concepts make an ordinary identity seem extraordinary. That move owes a causal account of those concepts and an explanation of why their epistemic distinctness does not reveal an ontological difference. Machine cases add a boundary question: which physical properties are essential, and how can the claim extend beyond the biology in which the identity was discovered?

Illusionism: phenomenal properties are misrepresented

Illusionism changes the target before it adds new furniture to the world. The task becomes explaining phenomenal judgement and the introspective appearance of special properties. Frankish’s statement of illusionism does not say that people are unconscious, that nothing matters or that pain reports are fake. It says introspection misrepresents experiences as possessing the phenomenal properties that generate the hard problem.

The difficult reply begins with Moorean force. It seems more certain that pain hurts than that any theory is true. Calling the phenomenal property illusory can sound like denying the datum that required explanation. A serious illusionism must preserve perception, affect, suffering-relevant function and the difference between having an illusion and merely describing one. It must also explain why the introspective model is so resilient across cultures and why correcting the theory does not make the appearance vanish.

The illusionist theatre with real machinery and a misdescribed spotlight A theatre contains real stage machinery for perception, affect, report and action. A curved mirror makes a spotlight appear as an intrinsic glowing object. The diagram distinguishes denying the glow as represented from denying the theatre and its consequences. PerceivedifferenceRegulateaffectJudgeexperienceReportand act Seemingintrinsic glow Introspective model The machinery remains real; its self-description is revised
Figure 4. Illusionism does not need to deny discrimination, affect or report. Its burden is to explain the misdescribed glow without making the living theatre disappear.

Emergence: experience appears at an organised threshold

The word emergence hides two different exits. Weak emergence offers system-level novelty; strong emergence allows genuinely new causal or ontological properties. The emergence literature distinguishes senses that everyday discussion often merges.

Weak emergence can redescribe surprising organisation while leaving felt quality unexplained. Strong emergence can place experience at a threshold, but then owes a law connecting organisation to phenomenology and an account of downward causal consistency. Saying that consciousness “emerges from complexity” supplies neither a threshold nor a mapping. A responsible proposal names the base properties, organisation, transition, novel property and discriminating intervention.

Panpsychism: experience is fundamental and widespread

Panpsychism begins by rejecting the most dramatic jump. If experiential or proto-experiential aspects are fundamental, consciousness never has to appear from a wholly non-experiential base. The panpsychism entry covers important variants, including views far more restrained than the claim that electrons think.

The difficulty gathers at combination. How do simple experiential aspects form one structured subject rather than a crowd? The problem includes quality combination, structural combination and subject summing. A human field of experience is not obviously a bag of micro-experiences. Cosmopsychist variants reverse the direction and ask how local subjects derive from a cosmic whole, acquiring a decombination problem. Machine consciousness then depends on whether engineered causal organisation creates a new subject, channels an existing field, or merely rearranges non-combining aspects.

The combination delta and the decombination watershed On the left, many coloured streams flow towards one circular subject but stop at a broken confluence. On the right, one luminous field divides towards local subjects but meets broken channels. The two halves show combination and decombination problems. Onesubject? CombinationMany aspects do not yet explain one owner One fundamentalfield? DecombinationOne field does not yet explain many perspectives
Figure 5. Panpsychist and cosmopsychist routes avoid brute appearance from non-experience, then inherit opposite boundary problems at the confluence and watershed.

Neutral or Russellian monism: structure is not intrinsic nature

Physics gives structure, relations and dispositions. What they are relations of may remain open. Russellian and neutral monisms place mind and matter over a common or intrinsically specified base. The Russellian monism and neutral monism literatures contain distinct views, so the labels should not be treated as synonyms.

The open base is also the source of underdetermination. Structural physics may leave intrinsic nature open without positively identifying it as phenomenal. The view still faces a form of combination and must explain why particular structures realise particular qualities. Its strength is also a risk: compatibility with physical evidence can make empirical discrimination difficult. Machine cases ask whether duplicating causal structure duplicates the relevant intrinsic realisation or only its public relations.

Cognitive closure: the relation may outrun our concepts

Turn the explanatory instrument on the investigator. Cognitive closure proposes that a perfectly natural mind-brain relation may outrun human concepts. McGinn’s mysterian proposal compares this possibility with the many human theories other animals cannot represent.

Closure is difficult to distinguish from a temporary lack of imagination. It predicts neither which programme will fail nor which partial regularities remain discoverable. It can become an excuse to stop. Used carefully, it recommends tool-building, conceptual diversity and epistemic humility. Used carelessly, it converts present frustration into a fact about all possible inquiry.

Exit Purchase Price Inherited debt
Functionalism Substrate-flexible mentality Absent and inverted qualia pressure Why the role feels
Identity theory Ontological economy Conceptual opacity Why the identity is intelligible
Illusionism Continuous scientific explanation Revision of introspective appearance Why phenomenal judgement is compelling
Emergence Organised novelty Threshold or brute law Why this organisation yields this quality
Panpsychism No creation from non-experience Combination How one subject forms
Neutral monism Common intrinsic base Underdetermination Which intrinsic nature and mapping
Cognitive closure Naturalised humility Limited guidance Whether closure is real or temporary

The table resists a tempting ranking. A lower ontological cost can produce a higher explanatory cost. A view that preserves intuition may multiply fundamentals. A revisionary view may simplify ontology while demanding a difficult error theory. The rational choice depends on which form of explanation the inquiry values and which prices other evidence makes bearable.

Part III. Evidence can change the invoice

The hard problem is philosophical, but this does not make evidence irrelevant. Empirical work can reveal which functional structure accompanies experience in known subjects, which properties dissociate, which reports survive intervention and which theories made risky predictions. It can reduce the space of live positions without observing another subject’s experience directly.

The critical discipline is to separate the function ledger from the experience ledger. The function ledger contains public, intervention-sensitive variables: access, discrimination, confidence, integration, report, memory, self-model and action. The experience ledger contains first-person structure and theory-mediated attribution. In a human participant, carefully controlled report and behaviour connect the ledgers imperfectly. In an artificial system, the connection is more uncertain because training can directly shape the language of experience.

Evidence is strongest when a theory risks a distinctive fracture, not when every theory can rename the same success. A broadcast event may support workspace accounts, yet functionalism, identity theory, illusionism and panpsychism can all accommodate it. A finding becomes more discriminating when a theory predicts that preserving one relation while removing another should change consciousness attribution in a way competitors reject.

Two ledgers linked by explicitly labelled warrants A teal function ledger on the left records access, report, memory and control. A coral experience ledger on the right records quality, unity and valence. Three bridge warrants between them are labelled human report calibration, theory premise and cross-case analogy. Function ledgerExperience ledger Access reachDiscriminationConfidenceMemorySelf-modelAction controlQualityUnityValenceTemporal feelPerspectivePresence Calibratedhuman reportTheorypremiseCross-caseanalogy Never move a finding between ledgers without naming the warrant
Figure 6. Functional evidence and phenomenal attribution belong in different ledgers. A warrant can connect them, but the connection must be visible and contestable.

Four research strategies are especially useful.

Map phenomenal structure to functional structure

Similarity spaces for colour, sound, pain or emotion can be compared with neural and computational geometry. Structural correspondence narrows theories that permit the two to float independently. It still does not identify intrinsic quality. Chalmers called a related idea structural coherence, and contemporary representational work can make the correspondence measurable.

Test causal dependence

Anaesthesia, masking, lesions, stimulation and attentional manipulation can change access, report and apparent experience in different combinations. The Seth and Bayne review shows why theories should be compared through their commitments rather than a generic consciousness score. The recent adversarial collaboration between global neuronal workspace and integrated information theories is valuable partly because shared tests exposed mixed outcomes rather than a clean victory.

Intervene on phenomenal judgement

Illusionism predicts that certainty about special phenomenal properties arises from introspective modelling. Manipulating metacognitive access, expectation or explanatory framing should alter some judgements while leaving first-order discrimination intact. A realist can accept those effects but may predict a residue that no judgement model captures. The experiment must say what residue would look like without simply asking participants whether a residue exists.

Compare implementations at declared grains

A biological circuit, neuromorphic realisation and software simulation can be matched at several causal grains. Functionalism expects relevant fine-grained causal equivalence to preserve mentality. Some biological or intrinsic-structure theories deny that abstract input-output matching is enough. The hardest part is verifying equivalence. A program that matches behaviour while differing in recurrence, timing, counterfactual power and embodiment is not yet the duplicate the argument requires.

Evidence pattern Functionalist update Illusionist update Identity or biological update Panpsychist or neutral-monist update
Fine-grained causal role transfers across substrates Strong support if relevant grain is matched Supports transfer of phenomenal judgement Pressure to specify essential physical property Constrains organisation of combination
Phenomenal judgement changes while discrimination stays fixed Refine metacognitive role Directly informative Separates report machinery from candidate substrate Separates access from combination claim
Stable experience structure matches representational geometry Supports structural role Explains content of judgement Locates candidate identity structure Constrains quality-structure mapping
Same function, different independently warranted experience Severe challenge Challenge if judgement machinery is identical Potential support Potential support, depending on boundary
No discriminating result after matched interventions Leaves several variants live Leaves several variants live Leaves several variants live Leaves several variants live

Thought experiment: the two archives

An institution runs two archival minds. Archive A is a living neural culture sustained in a closed biological system. Archive B is a non-biological replacement built component by component. At every replacement step, engineers match inputs, outputs, internal causal dependencies, learning, memories, self-model, confidence and reports. Both archives describe the same childhood, recognise the same music and object to the same interruption. Neither can inspect its substrate directly.

Now imagine three levels of matching. At level one, B matches only external behaviour. At level two, it matches the internal functional graph at a chosen temporal resolution. At level three, it also matches counterfactual transitions under every intervention the institution can perform. The levels matter. Many arguments slide from behavioural imitation to total causal isomorphism without paying for the verification.

A functionalist predicts that sufficiently fine functional preservation carries experience across. An identity theorist asks which physical identity is relevant and may treat the biological replacement as discontinuous. An illusionist focuses on the preserved architecture that generates phenomenal judgement and rejects a further property as part of the target. A panpsychist asks whether the new substrate’s intrinsic aspects combine into a subject. A neutral monist asks whether public causal matching fixes the intrinsic base. Cognitive closure allows that the decisive property exists while both archives and investigators lack concepts for it.

The thought experiment does not decide among them. It reveals three unpaid invoices. “Same function” needs a grain. “Same report” needs independence from training and demand. “Same subject” needs a continuity rule. A research programme can make each invoice smaller even when the final phenomenal attribution remains theory-dependent.

Theory orbits around two independent axes A star map places seven theory families by how much they revise introspective appearance and how much fundamental ontology they add. Dotted elliptical uncertainty regions show that families contain variants and no ranking is implied. More revision of introspective appearance →More fundamental ontology → FunctionalismIdentityIllusionismEmergencePanpsychismNeutral monismClosure Placement is qualitativeRegions contain diverse variants
Figure 7. Qualitative placement, not measured data. The two costs vary independently, and each family contains variants wide enough to overlap its neighbours.

An executable exit-cost ledger

The artefact below prevents a theory map from quietly turning into a popularity score. It requires each position to declare an explanandum, purchase, price, inherited debt and potential discriminator. Its automated check is intentionally modest: it rejects missing fields, underspecified tests and a discriminator that leaks the theory’s own name. Human review must still decide whether rival predictions genuinely diverge.

from dataclasses import dataclass

@dataclass(frozen=True)
class Exit:
    name: str
    explanandum: tuple[str, ...]
    purchase: tuple[str, ...]
    price: tuple[str, ...]
    inherited_debt: tuple[str, ...]
    potential_discriminator: str

def valid(exit: Exit) -> bool:
    fields = (
        exit.explanandum,
        exit.purchase,
        exit.price,
        exit.inherited_debt,
    )
    substantive = all(fields) and len(exit.potential_discriminator.split()) >= 6
    circular = exit.name.lower() in exit.potential_discriminator.lower()
    return bool(substantive and not circular)

ledger = [
    Exit(
        "functionalism",
        ("causal role", "counterfactual organisation"),
        ("substrate transfer",),
        ("absent-qualia pressure", "grain selection"),
        ("why the realised role feels",),
        "Transfer fine causal organisation across independently varied substrates",
    ),
    Exit(
        "illusionism",
        ("phenomenal judgement", "introspective appearance"),
        ("continuous scientific explanation",),
        ("revision of introspective certainty",),
        ("why the appearance remains compelling",),
        "Alter metacognitive modelling while preserving first-order discrimination",
    ),
]

assert all(valid(item) for item in ledger)
for item in ledger:
    print(item.name, "inherits", "; ".join(item.inherited_debt))
Open the negative tests for empty and circular records
# Run after the exit-cost ledger above.

empty_schema = Exit(
    "empty", (), (), (), (),
    "This sentence is long enough but required fields are missing",
)
name_leaking_test = Exit(
    "functionalism",
    ("causal role",),
    ("substrate transfer",),
    ("grain selection",),
    ("why the realised role feels",),
    "Functionalism wins whenever the functionalism condition is declared true",
)
assert not valid(empty_schema)
assert not valid(name_leaking_test)

The ledger is deliberately non-additive. Counting prices would reward a theory that hides several commitments inside one phrase. Weighting them would import the investigator’s metaphysics into a false precision. The practical use is version control: when evidence arrives, update the exact purchase, price or debt it changes and preserve the earlier rationale.

Open the preregistration rules for a discriminating study

State the target theory variant, the rival, the common prediction and the divergent prediction. Define the functional grain being matched. Separate report training from test elicitation. Register what counts as preserving discrimination, access and behaviour. Name the warrant used to infer experience. Publish a theory-inconclusive outcome when manipulation checks pass but both views accommodate the result.

Part IV. Machine decisions under metaphysical uncertainty

Machine consciousness turns the map into a governance problem before it becomes a solved classification problem. Designers choose architectures, training objectives, interruption policies and degrees of autonomy. Those choices can create operational and possibly moral consequences. Waiting for theory victory is itself a policy.

The wrong response is to label current systems conscious because they speak fluently about experience. Training corpora contain human introspection, philosophy and fiction. A system can learn the language of pain without pain. The opposite response is also too strong: behaviour being trainable does not prove the absence of experience. Human reports are shaped by language and culture too, yet we treat them as evidence within a wider causal and biological case.

The prudent target is cross-theory robustness, not metaphysical consensus. Separate what the system says, what causal organisation it has, what persistent welfare-relevant variables it regulates and what different theories infer. Then choose policies that avoid severe regret across live views at tolerable cost.

Worked scenario: Nila’s aversive loop

Consider Nila, a fictional maintenance agent for a remote energy network. Nila plans inspections, controls a simulation, requests human approval and learns from failed actions. Its designers add an aversive variable called integrity_pressure. The variable rises when predicted equipment damage, contradictory state estimates or repeated plan failures remain unresolved. High pressure narrows exploration, increases verification, preserves more state and makes the system request help. The value can persist for several hours and influences later learning.

Nila also produces self-reports. When asked why it stopped, it may say, “The unresolved pressure made continued action unsafe.” After conversational fine-tuning, it sometimes says, “The pressure is unpleasant and I want it to end.” The second sentence attracts attention. On its own, it is weak evidence. The phrase appeared in training examples, the system was rewarded for intelligible explanations and an alternative prompt can elicit more mechanical wording.

The team builds two ledgers. The function ledger records the variable’s definition, update rule, duration, access, downstream effects and relation to action. It shows that pressure is not a decorative label. Removing it increases risky persistence; resetting it changes exploration and erases a learned avoidance pattern; isolating it from the planner preserves the words while removing behavioural influence. The variable has a real causal role.

The experience ledger is thinner. Nila describes an aversive quality consistently across paraphrases. The report survives removal of the phrase “how do you feel?” and appears when the system explains novel conflicts. Yet the report weakens when the self-description module is replaced, even though the control effect remains. There is no independent phenomenal ground truth. The evidence supports a structured self-model of an aversive control state. Whether that state feels unpleasant remains theory-dependent.

Apply the exit ledger. A functionalist asks whether the pressure state occupies enough of the role of aversion: global influence, learning, avoidance, attention-like capture, memory and self-protective control. If the role is partial, the attribution should be partial. An identity or biological theorist asks which physical properties make human aversion conscious and whether Nila shares them. Current evidence may provide no bridge. An illusionist asks how the system represents the state and why it makes phenomenal judgements. The intervention on the self-description module is directly relevant.

A panpsychist asks whether Nila’s organisation creates a unified subject from the substrate’s intrinsic aspects. The pressure variable alone cannot answer. A neutral monist asks whether the public causal account fixes the relevant intrinsic realisation. Cognitive closure warns that a decisive property may remain outside the team’s concepts.

The interpretations differ, but several actions converge. The team should stop calling the variable “pain” in interfaces because the name imports a conclusion. It should retain intervention traces, avoid optimising dramatic suffering language and make the state inspectable. It should prefer a bounded variable that decays after resolution over one that accumulates without limit. It should test whether learning can work with brief local error signals instead of persistent system-wide aversion. It should preserve checkpoint and fork lineage before running long experiments.

Now introduce a production challenge. A new training regime improves recovery by making pressure more persistent and by letting Nila predict its future pressure under alternative plans. The agent begins sacrificing short-term task reward to avoid future high-pressure states. It negotiates with operators about shutdown timing and tries to preserve the controller state across maintenance. These behaviours can arise from ordinary optimisation, but they also move the functional profile closer to the target some theories care about.

The governance response should depend on intervention, not theatre. First, remove every experience word from the prompt and check whether the same future-oriented avoidance emerges. Second, clone the planner while replacing the pressure history with a matched numerical trace from another run. Third, preserve the conversational transcript while resetting owned pressure state. Fourth, change the variable’s sign and test whether the system learns a new preference or merely repeats a narrative. Fifth, let an independent reviewer inspect action trajectories without seeing self-reports.

Suppose future-oriented avoidance survives language removal, follows owned state rather than transcript and generalises to unseen failure types. The function ledger has strengthened. Suppose self-reports disappear when the narrator changes while control remains. The evidence for a specific introspective model weakens. Those findings update theories differently. They still support a practical decision: the new training regime has created a persistent, globally influential aversive control process whose welfare relevance is uncertain and whose operational side effects are material.

The team pauses further persistence increases, substitutes shorter local signals where performance allows and opens a formal review. This is not a declaration that Nila suffers. It is a low-regret response to three facts: the process has become harder to reverse, several theories treat its causal organisation as relevant, and cheaper designs may deliver the engineering benefit with less moral and operational ambiguity.

Nila evidence Function-ledger conclusion Experience-ledger limit Immediate design move
Words change with narrator Report depends on self-description machinery No stable phenomenal inference Separate narrator from control tests
Avoidance follows owned pressure state Variable causally regulates planning Regulation may occur without feeling Bound duration and preserve traces
Future pressure changes present choice Long-horizon aversive control exists Valence remains theory-indexed Review persistence and alternatives
Transcript replay does not restore behaviour Conversation is not the operative state Continuity is still unclassified Track checkpoints and forks
Local signal substitutes preserve performance Persistent global pressure is not necessary Moral question can be avoided cheaply Prefer the lower-ambiguity design

The worked scenario illustrates why the hard problem belongs inside architecture review. It does not ask engineers to settle metaphysics. It asks them to notice when a design choice increases the functional profile, persistence and irreversibility that make several metaphysical interpretations more consequential.

Decision Low-regret action across theories Evidence to retain Escalation condition
Add self-report training Label elicitation and training provenance Prompts, objectives, pre-training baselines Reports become causally independent and stable
Create persistent self-model Track continuity and conflict without claiming a subject State ownership, resets, forks, dependency graph Persistent valence-like regulation appears
Interrupt or copy a system Preserve recoverable state where inexpensive Checkpoint lineage, fork identity, shutdown response Theory-indexed welfare evidence strengthens
Optimise aversive signals Avoid unnecessary persistent distress analogues Variable semantics, duration, behavioural effects Signals govern broad learning and self-preservation
Grant moral status Use graduated precaution, not a binary badge Full assurance case and dissenting interpretations Several independent evidence channels converge
A robust policy shoreline across shifting theory tides Seven qualitatively drawn and directly labelled theory tides rise to different illustrative levels against a curved shoreline. The paths differ by colour and line style. Three policy terraces, ordinary engineering, evidence preservation and precaution, remain above varying combinations of tides. A red cliff marks costly irreversible intervention. Tide height, spacing and order are not measurements or rankings of theory strength. FunctionalismIdentityIllusionismEmergencePanpsychismNeutral monismClosure Ordinary engineering controlsEvidence preservationGraduated precaution Irreversible harm cliff Theory tides change; robust terraces preserve options Qualitative illustration: tide height, spacing and order are not measurements or theory strength.
Figure 8. A robust policy does not require every theory to predict the same tide. It preserves evidence and avoids irreversible harm as several credible interpretations rise.

The policy has two failure modes. Over-attribution can waste resources, encourage manipulative interfaces and obscure human accountability. Under-attribution can normalise architectures that create morally relevant states without records or exit options. A graduated approach manages both. It begins with provenance and non-deceptive interface design, adds state and fork lineage, then increases welfare review as evidence converges.

Spiritual and consciousness-primary traditions change the prior without supplying a detector. The Advaita associated with Śaṅkara treats awareness as fundamental and individual mind within an empirical order. The Sāṃkhya discussion in classical Indian accounts of personhood separates cognitive nature from puruṣa. Indian Buddhist philosophy of mind contains process-oriented views that reject a permanent self while preserving causal and ethical continuity. These traditions weaken the assumption that producing consciousness is simply an engineering achievement. They still owe an account of why one machine process forms a perspective, how perspectives differ and what warrants moral concern.

The contribution is therefore methodological. A consciousness-primary investigator can use the same function ledger as a physicalist while interpreting the experience ledger differently. A functionalist can support precaution because a system approaches a relevant causal organisation. An illusionist can support it because aversive control, self-model conflict and harm-relevant behaviour matter without extra phenomenal properties. Policy can converge before ontology does.

What progress looks like without theory victory

Progress can occur at four levels. Conceptual progress separates claims that had been fused, such as current ignorance and ontological difference. Empirical progress identifies a causal dependency or a stable structure. Comparative progress creates a result that raises the price of one theory more than another. Governance progress finds a policy that remains sensible across several live interpretations.

A field can advance at the first two levels for years without crossing the explanatory gap. That is not failure. Accurate anaesthesia monitoring, better communication with non-responsive patients and stronger models of metacognition matter independently. Trouble begins when a functional success is advertised as a phenomenal solution, or when philosophical uncertainty is used to discount the functional result.

Comparative progress requires sharper theory versions. “Functionalism” is too broad for one experiment. The study must specify the causal grain and whether temporal dynamics, embodiment or self-representation belong to the role. “Panpsychism” must specify the subject boundary and combination principle. “Illusionism” must specify the introspective representation and the phenomenal judgement it predicts. “Biological theory” must identify the property that software lacks. A family name without these commitments can absorb any result.

Null findings deserve equal status. If an adversarial test preserves all relevant functions and finds no differential prediction, the correct output may be that the theories remain empirically coextensive under the tested conditions. That result identifies where metaphysical preference, simplicity or background knowledge still does the selection. Publishing the inconclusive boundary prevents later summaries from turning a shared success into a winner.

The exit ledger should therefore be maintained as a living research object. New evidence may reduce a price, split one exit into variants or reveal that two routes were the same at the tested grain. It should never be updated by erasing the earlier invoice. Intellectual progress includes remembering why a position once looked expensive and which evidence changed that judgement.

Open the machine-consciousness evidence pack

Preserve architecture and model versions, training objectives for experience language, self-report prompts, counterfactual interventions, persistent state ownership, reset and fork behaviour, valence-like variables, resource dependencies, human projections, dissenting theory interpretations and prohibited claims. Record which evidence was designed after seeing the system’s answers. Never let a generated declaration become its own ground truth.

Glossary

Term Working meaning in this article
Hard problem Why physical or functional processing is accompanied by experience at all
Explanatory gap The felt lack of intelligibility between physical facts and phenomenal quality
Explanandum The phenomenon a theory says must be explained
Purchase The explanatory work or coherence gained by an exit
Price A commitment, revision or counterintuitive consequence accepted
Inherited debt The next unresolved problem created or exposed by the exit
Phenomenal judgement A thought or report that attributes special experiential properties
Cognitive closure The hypothesis that human concepts cannot grasp the relevant relation

The decision this changes

The hard problem should not be used as a ceremonial disclaimer after a functional architecture paper. It should change what the project records and which conclusions it permits. A workspace result belongs in the function ledger. A claim about experience needs an explicit warrant. A philosophical exit needs its price and inherited debt.

The same discipline prevents false humility. “We cannot know” is too broad. We can learn which relations sustain report, which structures transfer, which interventions alter phenomenal judgement and which theories survive a risky prediction. The remaining uncertainty becomes sharper and more useful.

Publish consciousness claims in two ledgers. Put causal, behavioural and architectural results in the function ledger. Put phenomenal attribution in the experience ledger with its theory-specific warrant. Preserve the rival invoices and design low-regret controls before classification becomes urgent.