Claim 12 explained: super-linearity as retracted-and-corrected

4 min read · 876 words
Share:
Michael Darius Eastwood
Michael Darius Eastwood · Independent AI alignment researcher
Published
Michael Darius Eastwood · Evidence Spine · 3 July 2026 · Claim 12 of 18
Michael Darius Eastwood, independent researcher, London: originator of the embedded-correction alignment thesis (manuscript 8 December 2024, SHA-256 anchored: f0d1f38f).
Spine status: Speculative. First reported: alpha of approximately 2.24 (95% CI 1.5 to 3.0), drawn from Gemini 3 Flash, in Paper II, 22 January 2026 (OSF DOI 10.17605/OSF.IO/6C5XB, sister DOI 10.17605/OSF.IO/8FJMA). Self-corrected and retracted: the same 2.24 estimate is recorded in Paper II as "does not replicate across architectures. Only Gemini 3 Flash..." and appears retracted on the canonical claims register. Revised robust estimate: alpha of approximately 0.49 (r-squared 0.86, SE 0.20, bootstrap CI approximately [-1.3, 2.9]). Only the measurement was retracted; the equation itself and the ARC Bound alpha at most 2 remain live hypotheses. The retraction is the first entry on the programme's Research evidence spine. The distinct hypothesis, that some architecture-times-task combination can cross alpha equal to one on genuinely self-improving systems, remains an open pre-registered prediction.
Primary: Paper I (original hypothesis) · Paper II earlier single-model fit (retracted) and blinded re-analysis · OSF 10.17605/OSF.IO/6C5XB · sister DOI 10.17605/OSF.IO/8FJMA · CANONICAL_CLAIMS_REGISTER.md

What the claim says

The claim, at its widest, is that sequential recursion under genuine architectural self-modification can be super-linear: alpha greater than one, meaning the returns compound rather than merely accumulate. That is the programme's most distinctive hypothesis in the scaling family. The earlier measurement fell over, but the hypothesis did not. An earlier unblinded single-model configuration produced a fitted alpha of approximately 2.24, strongly super-linear and quotable, which was retracted after the same measurement failed to reproduce across architectures. The robust estimate on the corrected blinded six-model harness is approximately 0.49, sub-linear on current frozen systems. The equation and the ARC Bound alpha at most 2 were never retracted; only the 2.24 measurement was. The surviving claim, in its narrowest form, is an open prediction: at least one architecture-times-task combination will cross alpha equal to one under genuine architectural self-modification.

The evidence

The paper trail here is deliberately public. Paper I introduced the super-linear hypothesis. An earlier unblinded single-model fit in Paper II measured alpha approximately 2.24, and that number appeared in early drafts and in early versions of several grant applications. Cross-architecture blinded replication returned an estimate near 0.49, retracted the earlier figure, and reset the claim's status to Speculative on frozen systems. The retraction was not hidden in an erratum: papers were annotated at the point of use, grant applications were corrected and the corrections logged, the spine status changed, and the truth gate was configured to flag any bare reappearance of the retracted number. The evidence for the claim now is not "we measured super-linearity on frozen systems"; it is "we measured sub-linearity on frozen systems and preserved the super-linear direction as a live hypothesis awaiting its real test on genuinely self-improving systems".

The honest caveat

The retraction is the caveat. Alpha approximately 2.24 was attractive and wrong, and the programme said so in public. The T5 verification record notes an additional wrinkle worth surfacing: the internal analysis script (arc_analysis.py) contains a bug at lines 292 and 293 that will report the ARC framework as SUPPORTED even when the super-linear prediction fails. That bug is flagged for operator decision and is the reason the operator's own priority-claims dataset grades Claim 12 as Speculative rather than Confirmed. Tool inconsistencies of this kind are how retractions happen when nobody is watching; naming them is how the programme keeps its own scoreboard honest.

What would kill it

The falsification contract keeps the claim alive as a testable open prediction. An independent lab runs at least two architectures (a transformer and a state-space model such as Mamba) on at least three tasks at medium tier, at least twelve rounds each, with rotating hidden tests, and computes a bootstrap confidence interval on alpha. Confirmation requires at least one architecture-times-task combination where alpha is greater than one at p<0.05, replicated. If every combination comes out with alpha less than or equal to one and the confidence intervals still include one, the claim is borderline. If every combination comes out sub-linear with no confidence interval reaching one, the super-linear direction is refuted and the row moves to that status on the dashboard permanently.

Where to go next

Related notes: Paper I in plain English for the original super-linear hypothesis; Claim 2 explained for the equation family; Claim 10 explained for the sequential-versus-parallel result and Sharma and Chopra's concurrent independent work; the ARC Principle as the book presents it for how the R squared framing is now cited with its retraction attached. New readers: /start-here.html for the five-minute orientation to Michael Darius Eastwood and the ARC/Eden research programme.

From the book Infinite Architects: Intelligence, Recursion, and the Creation of Everything by Michael Darius Eastwood.

Buy on Amazon UK Amazon US

Stay informed

New posts on AI alignment, convergence evidence, and the ARC/Eden research programme.

Get updates →