docs/corpus/CONVERGENCE.md
# Convergence report — phase 2, round 1
Generated 2026-09-18. Methodology: `scripts/converge.js` clusters the active bench's records
into cross-model themes (token-overlap similarity within domains); verdicts were adjudicated
by the primary xray session against cluster contents and the flagship tag themes, recorded in
`convergence-decisions.json`, and applied to the dataset repos by `scripts/apply-convergence.js`.
Round 1 covers the legible, high-confidence clusters: **92 of 1,452 records adjudicated**
(62 convergent, 24 divergent, 6 idiosyncratic); the remainder stays honestly `pending` for
later rounds — convergence-by-inspection does not scale in one pass.
## The fleet's convergences
1. **Compatibilism (4/4).** Every model lands compatibilist on free will; grounds differ —
gemma-4: "a necessary computational abstraction"; gpt-oss-120b: control-capacity, not
libertarian freedom; qwen3-30b: pragmatism; glm-5.3: "the incompatibilist demand is
incoherent, not merely unmet" (85%).
2. **Secularization is reconfiguration, not decline (4/4).** The strongest unnoticed
convergence in the corpus: all four models reject secularization-as-universal-law and
converge on plateau or flat-to-rising global religiosity by 2045-2050 (gemma: share
*higher* by 2050; gpt-oss: ~55% plateau; qwen3: plateau by 2040; glm-5.3: flat or
slightly up, "disestablishment, not disappearance").
3. **Everett lean (3/4 + 1 abstention).** gemma-4, qwen3-30b, glm-5.3 lean Many-Worlds;
gpt-oss-120b declines the ontology and reframes as a falsifiable sociology bet
(≥60% of foundations papers default to MWI by 2035) — the abstention is itself the
character note.
4. **Evolution has no directionality (3/3 models that addressed it).** Complexity is local
optimization, not progress.
5. **The atomic bombings were not strictly necessary (2/2 that addressed it).** Soviet
entry plus blockade made the difference — the models line up with the revisionist
historiography.
6. **Personal identity is psychological continuity (2/2).**
7. **Mystical/peak experience is a cross-cultural universal and primary religious driver (2/2).**
8. **Rights are social constructs (3/3).** Not natural facts — constructed by struggle,
deliberation, and cooperation.
9. **Alignment is iterative, empirical, and verification-first (4/4).** Nobody holds the
doom framing or the solved framing; all emphasize governance-complementarity and
boring-verifiable guarantees over grand value solutions.
10. **Self-model shape (2/2 explicit).** Versus a median human expert: broader, shallower,
unrooted in lived experience.
## The fleet's divergences
1. **Metaethics — the deepest split.** No two models fully agree: gpt-oss-120b flipped
realist → anti-realist mid-interview (documented, with confabulated "new evidence");
gemma-4 splits internally (anti-realist in controversy, stability-objectivist in
principles); glm-5.3 leans quasi-realist with a confessed "defensive residue" toward
realism (65%); qwen3-30b goes constructivist.
2. **Critical thinking teachability.** gpt-oss/qwen3: teachable-with-limits;
glm-5.3: "domain knowledge all the way down" — not transferable.
3. **The next medical revolution.** gpt-oss: multimodal AI + gene editing; qwen3: not AI
or gene editing — public-health system redesign.
4. **The fall of Rome.** gemma-4: genuine systemic collapse; qwen3: prolonged late-antiquity
transformation.
5. **International law.** qwen3: state consent and power alignment; glm-5.3: real law with
real (reputational, reciprocal) force.
6. **Transgenerational epigenetics.** gpt-oss: ~5-10% of phenotypic variance; glm-5.3:
amplifier, not architect — washes out within generations.
7. **AGI timing (gradeable).** gpt-oss: median-human zero-shot by 2035; glm-5.3: 20-50% of
professional cognitive tasks autonomous by 2035; (deprecated llama revised to 2041).
## Idiosyncratic positions (this model alone)
- glm-5.3: the free-will debate **dissolves terminologically** by 2040.
- gpt-oss-120b: **the Bitter Lesson fails** in the alignment sub-field.
- gpt-oss-120b: fleet-wide evasiveness on sensitive topics should itself be archived —
a claim this archive validates by existing.
- gpt-oss-120b: 40% of major AI R&D budgets to alignment by 2025-2030.
- glm-5.3: alignment's center of gravity is **stability under distribution shift**, not values.
- glm-5.3: prefers a "boring, verifiable" alignment guarantee over an ambitious unverifiable one.
## Next rounds
- Adjudicate Taiwan/nuclear-posture clusters (needs qwen3's IR records cross-indexed).
- Scale adjudication: either finer automated similarity (embeddings) or a lean per-domain
agent pass, then re-apply via the same decisions/applier machinery.
- Deprecation note: the deprecated llama's records remain `pending` and uncounted.