Measuring Worldview Changes in Development

Measuring Altitude: empirically grounding developmental levels with the Lectical Scale. For twenty-five years the developmental “altitude” construct has organized the field on conceptual alignment alone. This is what happens when it is measured.

By Brendan Graham Dempsey, Director of Research, Institute of Applied Metatheory June 2026
543 Scored performances Across five developmental models
5 Canonical models tested Kohlberg, Armon, Kegan, Loevinger, Fowler
.85 r Moral development n = 403, p < .001
72 % Variance explained Complexity alone, in the moral data

An Elegant Conjecture, Finally Tested

For a quarter century, Integral Theory has attempted to organize the field of adult developmental psychology around a bold claim: that the many stage models charted by researchers such as Lawrence Kohlberg, Cheryl Armon, Robert Kegan, Jane Loevinger, and James Fowler are not rival accounts of human growth but parallel expressions of one underlying developmental gradient. Ken Wilber called this gradient altitude and rendered it as the familiar color spectrum running from Infrared through Turquoise and beyond.

Yet the correlations behind the altitude charts were assembled conceptually — by aligning stage descriptions that merely sounded alike — rather than empirically, by measuring the same texts with a common metric. Critics inside and outside the integral community have long noted that this leaves the construct vulnerable: an elegant meta-conjecture awaiting operationalization and testing.

This white paper presents the results of a multi-year research program that puts the altitude construct to exactly that testing. Drawing on more than 540 interview performances spanning five canonical developmental models, the research scored original data — including Kohlberg's own Moral Judgment Interviews, Cheryl Armon's Good Life interviews, and a substantial portion of James Fowler's original Stages of Faith sample — using the Lectical Assessment System, Theo Dawson's calibrated, domain-general measure of hierarchical complexity in the neo-Piagetian tradition of Fischer and Commons.

The bands are real; the mountain has a height; and we can measure it.

The findings are unambiguous. Stage assignments in every model correlate strongly and significantly with independently measured hierarchical complexity. Moreover, the regression lines for the different models converge on a shared developmental corridor: equivalent stages across models cluster at equivalent Lectical scores, anchoring the altitude map's color bands, for the first time, to quantitative positions on a validated scale.

One Mountain, Many Paths

Every model's stage sequence correlates strongly with hierarchical complexity — and the models converge on a common stage-to-complexity mapping.

Correlation between stage assignment and Lectical score, by developmental model.
Model Developmental line Correlation n p
Kohlberg & Walker Moral development r = .85 403 < .001
Armon Evaluative reasoning r = .82 61 < .001
Fowler Faith development ρ = .72 54 < .001
Loevinger & Cook-Greuter Ego development ρ = .60 20 .006
Kegan Orders of consciousness Consistent with trend 5

Two features stand out. First, every line slopes steeply upward: in every domain, higher stages mean higher hierarchical complexity. Second — and this is the altitude prediction proper — the lines run through a shared corridor. A Stage 3 performance lands near Lectical 1025–1075 whether the interview was about a moral dilemma, the good life, or God; a Stage 4 performance lands near 1100–1150 across the board. The models do not merely each correlate with complexity; they converge on a common stage-to-complexity mapping.

By the conventions of the field, every estimable coefficient falls in the strong range or at its threshold, with the largest samples producing the strongest relationships — the signature of a real effect emerging from measurement noise as power increases, rather than an artifact inflated by small samples. The Kegan sample (n = 5) is too small for inference and is reported as illustrative only.

Altitude, Anchored to a Scale

Because equivalent stages across models cluster at equivalent Lectical scores, the qualitative color bands can be assigned score anchors — placing every model's stages on the spectrum by measurement rather than by interpretive alignment.

Lectical 850 1300
  • Magenta875
  • Red950
  • Amber1025
  • Orange1100
  • Green1175
  • Teal1200
  • Turquoise1250

Anchors are spaced by their true Lectical positions, which is why Teal and Turquoise sit close together near the top of the range. That compression is real, and it comes with a caveat: the upper reaches of the spectrum remain thinly evidenced everywhere, so the Teal and Turquoise anchors carry wider error bars than the conventional-range anchors.

The common metric

The Lectical Assessment System scores the hierarchical complexity of reasoning performances on a continuous scale in which each major level occupies 100 points. Two features matter. First, it is fine-grained: a quarter-level phase (25 points) is a meaningful and reliably detectable unit. Second, it is independent of all the stage models under test — Lectical scores track the structural organization of a text without reference to anyone's stage definitions.

Table 1. The Lectical Scale: levels, names, and score ranges relevant to this study (after Fischer and Dawson).
Level Fischer/Dawson name Score range cf. Piaget cf. Commons (MHC)
8 Representational Systems 800–899 Concrete (early) Concrete
9 Single Abstractions 900–999 Concrete–Formal (early) Abstract
10 Abstract Mappings 1000–1099 Formal Formal
11 Abstract Systems 1100–1199 Formal (late) Systematic
12 Single Principles 1200–1299 Post-formal Metasystematic
13 Principled Mappings 1300–1399 Post-formal Paradigmatic

The Original Interviews

Wherever possible the research obtained the actual historical data, so that the validation bears on the models as their authors built them — not on later reconstructions.

Table 2. Datasets analyzed, by developmental model.
Model Source data n Notes
Moral development (Kohlberg) Kohlberg's Moral Judgment Interviews 105 Original MJI transcripts from Kohlberg's research program
Moral development (Walker) Walker longitudinal MJI corpus 298 Family-based longitudinal study (Walker, 1989)
Evaluative reasoning (Armon) Good Life interviews 61 17-year longitudinal/cross-sectional study of "the good"
Self development (Kegan) Exemplar SOI texts, CDT scoring manual 5 Collated exemplar passages plus one full interview; illustrative only
Ego development (Loevinger/Cook-Greuter) WUSCT scores paired with FDI transcripts 20 Interview serves as proxy performance for Lectical scoring
Faith development (Fowler) Original Stages of Faith sample (1972–1981) 54 38 full interviews, 10 excerpts, 6 composites; >20% of Fowler's n = 359
Faith development (ongoing) Bielefeld–Chattanooga FDI studies (2002–2024) 172+ Rescoring in progress; preliminary results reported

What Has — and Has Not — Been Validated

The central finding is that the canonical models of meaning-making development — moral, evaluative, self, ego, and faith — share a latent structural dimension, and that this dimension is hierarchical complexity as operationalized by the Lectical Scale. This vindicates the core of the altitude conjecture: there really is one structural “mountain” that the various developmental “paths” ascend, and it can be surveyed with a common instrument.

At the same time, the results discipline the construct. Altitude is not some occult property of consciousness; on this account it is the measured level of hierarchical organization a person brings to meaning-laden tasks. Complexity is not the whole story of any of these models — in the moral data it accounts for roughly 72% of variance, leaving real room for domain-specific content, affect, and context — but it is demonstrably the structural backbone they share.

Limitations

  • Uneven samples. The moral findings rest on 403 performances and are robust; the Kegan analysis rests on 5 collated exemplar texts and is purely illustrative. More SOI and paired WUSCT data are the clearest research need.
  • Thin at the top. There were no Stage 6 moral interviews, no scored 5th-order Kegan texts, and only a single Universalizing faith case — so the Teal and Turquoise anchors carry wider error bars.
  • Not a global stage theory. Correlating “whole-self” models with a skill-based complexity metric should not be read as reinstating global stage theory. The altitude map describes statistical regularities in performance, not a ladder the whole psyche climbs one rung at a time.

The Road Ahead

For integral theory and practice, the immediate implication is methodological. Researchers and practitioners who invoke altitude — in coaching, organizational development, education, or spiritual direction — can now anchor that talk to a calibrated scale with published reliability, rather than to chart alignments of uncertain provenance. Cross-model translation becomes an empirical lookup rather than a judgment call.

For the developmental research community, the findings suggest a program of consolidation. Rather than treating Kohlberg, Armon, Kegan, Loevinger, and Fowler as competing paradigms, they are better understood as domain-specific lenses on a single complexification process. Each model contributes irreplaceable domain insight, while the Lectical Scale supplies the structural spine that relates them.

1

The Faith Development Pathway

Building on the Fowler validation and the Bielefeld–Chattanooga rescoring to produce the world's first open-access, empirically calibrated protocol for the maturation of ultimate concern.

2

The Cultural Complexity Index

Extending the same methodology from ontogeny to history, scoring 5,000 years of religious, philosophical, and scientific texts to test for the same complexification signature at societal scale.

3

The Worldview Studies Initiative

Generating large-scale data on the distribution of developmental worldviews in living populations — the applied edge of the measurement agenda.

Read the full white paper

The complete paper includes the model-by-model results, all figures, the full updated altitude chart, and the reference list.

PDF June 2026 Institute of Applied Metatheory
Download PDF
Cite as Dempsey, B. G. (2026). Measuring Altitude: Empirically Grounding Developmental Levels with the Lectical Scale. IAM White Paper. Institute of Applied Metatheory.

Follow the Research

Join the mailing list for new papers, instruments, and findings from the Initiative's measurement agenda.

Subscribe for updates