Overview

This note states, as honestly as possible, what kind of thing Thresholds Theory currently is and what would have to happen for it to become something more.

Current standing: Thresholds is an interpretive instrument with aspirations toward being a predictive map. It generates readings of people, characters, and cultural artifacts that feel illuminating and that its authors find internally consistent. It has not been empirically tested. No prediction derived from it has been registered in advance and checked. There is no inter-rater reliability data. Every applied diagnostic in this vault was produced by people who already believed the framework.

This is not a disqualification. Most of the parent theories were in exactly this position for decades, and some still are. But the distinction between “coheres and illuminates” and “is true” is the one this note exists to hold open.


Inherited Standing

Thresholds does not inherit a single evidential status, because its parents differ sharply in theirs.

  • Lacanian psychoanalysis — the primary parent and the most radically revised. Essentially no quantitative empirical support; defended on hermeneutic and clinical grounds rather than experimental ones. Whatever Thresholds inherits here, it inherits the epistemology too.
  • Attachment theory (Bowlby, Ainsworth) — the strongest parent. Large empirical literature, replicated instruments, cross-cultural work. Thresholds borrows less from it than it could.
  • Piaget — the stage sequence survived; the ages did not. Later research consistently found competencies earlier than Piaget claimed. This is a live warning for any age-anchored stage model, including this one.
  • Erikson — widely taught, thinly tested. Closer to Lacan than to Ainsworth in standing.
  • Kegan — has a scoring instrument (the Subject-Object Interview) and some validation work, which is more than Thresholds has. Worth studying as a model for how a stage theory becomes measurable.
  • Mahler — see below. The weakest link, and load-bearing.

Load-Bearing Weaknesses

The Mahler problem claude-synthesis

The Detached structure is anchored at 0–3 months and its preliminal phase is named Autistic. Both derive, directly or indirectly, from Mahler’s proposed “normal autistic phase” — an early period of objectlessness preceding symbiosis.

Infant research from the 1980s onward largely abandoned this. Neonates show preferential face-tracking, cross-modal matching, and contingency detection from the first days of life. Stern’s argument in The Interpersonal World of the Infant is that there is no objectless phase to emerge from; some sense of an emergent self and of others is present from the start. Mahler’s own later collaborators softened or dropped the autistic phase.

What is at risk, precisely: not the existence of a Detached structure — the adult phenomenology described in the Detached note stands or falls on its own evidence. What is at risk is (a) the 0–3 month anchoring, (b) the claim that Splitting is developmentally first, and (c) the recapitulation logic that lets an adult structure be explained by reference to an infant stage. If there is no objectless infancy, the framework’s origin story for its first structure needs rebuilding, and the age table needs to be treated as metaphor rather than chronology.

Possible repairs, in ascending order of cost: rename the phase to drop the Mahler inheritance; decouple the structures from literal ages entirely and treat the sequence as ordinal only; or accept the ages as approximate and defend them independently.

Applied diagnostics are not evidence claude-synthesis

Character and public-figure analysis is the framework’s main activity and its main epistemic hazard. The procedure — observe behaviour, assign structure, note fit — has no failure condition. Every structure has a plausible account of almost any behaviour, because each structure includes phases spanning wildly different presentations, plus the rule that later structures retain earlier mechanisms.

Two concrete symptoms of this problem already visible in the vault: several figures are listed with question marks or in more than one place, and the same figure can be read differently depending on which evidence is foregrounded.

Minimum fix: before analysing a new figure, write down what evidence would rule out your leading hypothesis. If nothing would, the analysis is not doing epistemic work — it’s doing illustrative work, which is fine but should be labelled as such.

Better fix: blind inter-rater testing. Two analysts, same dossier, independent assignment, compare. Even a small sample would tell you whether the framework is a shared instrument or a private idiom. This is the single highest-value empirical step available and it costs almost nothing.

Forward-only progression risks unfalsifiability

Progression states that regression is not possible absent brain damage. This is one of the framework’s most distinctive claims and, as currently stated, one of its least testable — because any apparent regression can be absorbed by reinterpretation (“they were always that structure,” “that’s a later structure using an earlier mechanism”).

To keep it falsifiable, state in advance what would count. A candidate: a person independently rated at structure N by two blind raters at time 1, and at structure N−1 by two blind raters at time 2, with no neurological event. Without a criterion like this, forward-only is a rule of interpretation rather than a claim about the world.

The normative question, and where the burden actually falls

The framework insists that structures are developmental achievements rather than pathologies, and that later is not “healthier” but is more. It is worth being clear about what this does and does not commit the framework to.

The claim is hypothetical, not categorical. Thresholds does not assert that later structures are better simpliciter. It asserts that if what you want is increased capability, structural progression is the terrain on which that happens. See Who This Is For. This is the same form of claim a training discipline makes: the physiology is true whether or not you want to be strong, and the discipline addresses those who do.

This resolves the apparent conflict with contemplative traditions rather than papering over it. Thresholds is not a rival account of the good life; it is a capability discipline that can be placed in service of one — including, explicitly, the reduction of suffering. Note that Mahāyāna Buddhism’s own critique of the arahant ideal is a capability critique, which puts a large part of that tradition on the same side of this question rather than the opposite one.

Where the burden relocates. The normative question does not disappear; it becomes an empirical one. Everything now rests on whether progression is a strict dominance relation or a set of trade-offs.

The vault currently asserts dominance: mechanism mastery frees background processing bandwidth, information processing speed increases with structure, and later-structure skills acquired by earlier structures are “of a different quality” (Progression, Information Processing Mechanisms). If that holds, progression is convergently instrumental — good relative to almost any goal — and the framework needs no value premise at all.

If it does not hold, “more capable” is false and the honest formulation is “differently capable,” which would require re-specifying the discipline as capability at what. Candidate trade-offs worth taking seriously: Detached sustained solitary attention and hyperfocus; Psychotic intersubjective merger as a substrate for exceptional empathy and care work; Neurotic moral seriousness as an engine of sustained ethical effort.

Testable form of the dominance claim claude-synthesis: individuals rated at later structures should outperform those rated at earlier structures on capability tasks unrelated to the framework’s own concerns, with no domain in which the ordering reverses. A single robust reversal would falsify dominance and force the trade-off framing. This is the load-bearing prediction behind the framework’s entire normative posture, and it has not been tested. open-problem

Practical consequence. Because the claim is hypothetical, the framework’s judgements about practices, relationships, and interventions are indexed to the aim of progression and carry no authority outside it. A practice that reduces suffering while preventing progression is counterproductive relative to that aim and may be entirely correct relative to another. See Contemplative Attainment.

Unresolved internal tensions open-problem

These are not objections from outside; they are places where the vault is currently inconsistent with itself.

  • The two-axis ambiguity: the overview table lists adult lifespan rows (Instrumental 20–40, Middle Age 40–60, Old Age 60–80) alongside structure rows. Are these structures, universal stages everyone passes through regardless of committed structure, or something else? The framework reads differently depending on the answer — one version is a stage theory, the other is a typology with a developmental origin story.
  • The Klein gap: Splitting as a foundational organising mechanism is Klein’s contribution, and the vault does not acknowledge the convergence. Either the framework is closer to object relations than it presents itself as being, or it needs to state the difference.
  • Ego-ideal terminology: used variously as a function, a compromise formation, and an identification. These are not the same thing.
  • Structure 7 and 8 are placeholders. A cycle model that requires eight positions and can describe six is making a structural promise it has not kept.

Terminological hazard

“Psychotic,” “Perverse,” “Schizoid,” “Borderline,” “Bipolar,” “Narcissistic” are borrowed clinical terms carrying non-clinical meanings here. Every note that uses them has to spend paragraphs saying what they don’t mean. This is a real cost in adoption and a real risk of harm if the framework circulates. The in-progress renaming is not cosmetic — it is partly an epistemic correction, since the borrowed terms import diagnostic authority the framework has not earned.


Falsifiable Predictions

These are the claims that distinguish Thresholds from its parents and could, in principle, be checked. Listed roughly in order of tractability.

  1. Committal window — structure commitment occurs when support fails during the corresponding age window. Prediction: a documented support collapse at 18mo–3yr (parental death, institutionalisation, war displacement) should produce a disproportionate rate of Neurotic commitment relative to collapses at other ages. Retrospective cohort work on natural experiments could test this without new data collection.
  2. Mechanism availability — earlier structures cannot use later mechanisms. Prediction: individuals rated Detached should fail context-dependent attribution tasks (2D, Repression-adjacent) that individuals rated Neurotic pass, independent of IQ. This is the most conventionally experimental prediction the framework makes.
  3. Symptom vocabulary mismatch — earlier structures use later-structure words for different experiences (Detached people say “anxious” and mean something closer to Derealization). Testable by structured interview with phenomenological probes.
  4. Forward-only — see criterion above.
  5. Practice-container divergence claude-synthesis — among long-term contemplative practitioners with matched technique and duration, those embedded in a lineage/sangha/teacher relationship should be rated at later structures than solitary practitioners. Predicts that the container rather than the technique does the developmental work.
  6. SensitivitySpencer’s hypothesis (sensitivity correlates with earlier structures) and Tiffani’s hypothesis (sensitivity is evenly distributed and expressed structurally) make opposite predictions about the joint distribution of HSP scores and structure ratings. This is the cleanest internal fork in the vault and could be settled with a survey.

What Would Falsify the Framework

Stated plainly, so it can be held to:

  • Reliable documented regression without neurological cause.
  • Failure of inter-rater reliability — if trained raters cannot agree on structure assignment above chance, the framework describes the rater rather than the rated.
  • No association between support-failure timing and adult structure — this would remove the framework’s causal story while possibly leaving the typology intact.
  • Demonstration that mechanisms are independently available — if people rated at early structures can be trained to use later mechanisms without structural change, the whole sequencing claim collapses.

Note the asymmetry: the fourth would be fatal, the second would be nearly fatal, and the first and third would each force major revision without necessarily killing the descriptive content.


Standard of Evidence Within the Vault

Proposed convention, extending the existing tags:

  • Untagged prose in a Theory or Structure note = vault canon: claims the authors are prepared to defend as part of the framework.
  • #spencer-hypothesis / #tiffani-hypothesis = authored speculation, not yet canon, possibly contested between the authors.
  • #spencer-question = open question, no position taken.
  • #claude-synthesis = derived in conversation with an LLM, not independently verified, should be treated as a suggestion rather than a source.
  • #open-problem = known internal inconsistency, flagged for resolution.

Nothing in this vault is #empirical. That tag should not be used until something is.


See Also