The open problem
Related work and contribution boundary
The synthesis is the claim, and the book said so first, in print, 2 January 2026, before any review ran.
“I make no claim to have invented the components. Recursion is well understood in mathematics. The alignment problem has been articulated by minds far more credentialed than mine. Religious traditions have explored stewardship for millennia. What I offer is the synthesis.” (Author’s Note, 2 January 2026.)
I then ran an adversarial precedence review across ten frontier configurations from six vendors, each instructed to attack. It found the components, as the book had already said. It did not find the arrangement: no single prior work joins them, and not one of the five bodies of prior work it examined contains the scaling relation. Its verdict named the operationalisation as the strongest surviving claim. This page records both halves without softening either.
The review has since gained a second chapter, aimed at the mathematics rather than the literature. In August 2026 two further adversarial audits, commissioned by the author and run externally against the statement paper itself, converged on a real defect: the ceiling law’s earlier printed form, 1/γ, ran the wrong way and had no derivation behind it. The correction is published in the paper with the three-line derivation beside it, the corrected form αcrit = 1/(1−γ) passes every check the old one failed, and the audits’ surviving repairs, typed quantities, a measurable balance exponent, a stricter absolute companion, are adopted in the current statement. The defect had survived review because the two forms coincide at exactly one half, the conjectured value; a drafted discrimination study now tests systems away from one half, where the families separate (at γ = 0.3 the retired form predicts 3.33 and the corrected form 1.43), so this class of error cannot survive measurement again. An attack that finds a defect and gets it corrected in public is the record working as designed, and the corrections register carries the entry.
Every source cited across the programme is collected in the full programme bibliography.
What I claim
Six firsts, each dated, each scoped, each ended by a single citation. The corrections register is where that citation would be recorded. Two of the six have since become the theory’s second and third laws; the relation in the boundary table below is its first, Law I, the ARC Principle; together they are the three ARC Laws.
- First to state that the two rival derivations of the metabolic scaling exponent are not rivals: multiplicative composition in d-dimensional space under a conservation constraint compels d/(d+1), making the exponent a mathematical necessity rather than a property of any mechanism.
- First to state that alignment must be embedded in the recursive loop, and to make the difference testable: embedded so that removing it damages capability, rather than bolted on so that removing it leaves capability intact. Tamirisa and colleagues state their success criterion as tamper-resistance “while preserving benign capabilities”; on the argument here, that preserved capability under a removed safeguard is the failure mode, not the goal.
- First to state β > k as a stability condition that can be measured on a running system, rather than argued about. It is now Law II of the theory, the ARC Co-Scaling Law, proved as an in-model theorem in Paper X and scoped there to relative drift, with its absolute companion printed beside it: relative drift falls when β exceeds k; absolute drift falls only when β exceeds k plus one.
- First to state α ≤ 2 as a numerical ceiling on stable recursive growth, rather than a qualitative warning about unbounded self-improvement. The ceiling has since become Law III: αcrit = 1/(1−γ), the most a self-correcting system can stably grow, set by its corrector’s shortfall from full proportionality, architecture-dependent and therefore engineerable; α ≤ 2 is its conjectured value at γ = ½. When hostile audit found the law’s earlier printed form defective, the correction was published on 16 August 2026 with the derivation shown.
- First to derive an AI-alignment design requirement from created-status structure. The review looked for this specifically and recorded that no located prior work does it.
- First to publish a blinded-evaluation protocol under which same-family AI-judged results reverse sign, with the protocol runnable and the reversal published.
Three structural claims post-date the review and have not yet been through it; they are recorded as unsearched rather than as firsts, and they carry the same defeat condition as everything above. The two-regime unification: the second and third laws are not separate claims standing side by side but the two regimes of one boundary, with the burden law selecting between them, a structure that emerged when a hostile attack became the seam where the laws join. The balance exponent Δ: the growth elasticity of the corrector’s capacity minus that of the burden, the object a laboratory would actually estimate, with the relative condition Δ > 0. And the identification discipline: Paper X’s β and the pacing model’s γ are typed as different operational objects, identified only under a declared local map that is printed so it can be checked rather than trusted.
The evidence has moved past bare statement. The Cauchy claim anchors a 50-domain comparison: 19 of 25 empirical domains match the predicted family under strict AICc, permutation p = 7 × 10−5, misses analysed rather than hidden. The embedding contrast anchors the published sign reversal. The stability condition anchors co-scaling studies now entering preregistration, so the confirmatory runs cannot be marked by their author. These are pilots, graded so in the registers, and a pilot with published kill-conditions, one already fired, is a tier above every proposal in the tables below.
What history says about exactly this shape of record, anticipated components with a surviving operationalisation, is on how to weigh this: the founders of quantitative fields were usually anticipated in their ideas, and what survived as the founded field was the operationalisation and the quantities. That page draws no conclusion about this one. The dated record itself is on evidence.
Where the closest work stops
Every named antecedent is real, dated, and in most cases predates the December 2024 record. Each stops before a quantity. That is the whole gap, and each line can be checked without trusting anything here.
| source | where it stops |
|---|---|
| Good 1965 | an argument |
| Yampolskiy 2020, On Controllability of AI, arXiv:2008.04071 | the impossibility itself, and a resignation. Four verbatim findings, all conceded here in full. (p.2) “the AI Control Problem is not solvable and the best we can hope for is Safer AI, but ultimately not 100% Safe AI”. (p.3) “motivational control needs to be added at the design/implementation phase, not after deployment”, so the cannot-be-added-later requirement is his, four years earlier. (p.9) “We were unable to locate any academic publications explicitly devoted to the subject of solvability of the AI Control Problem. We did find a number of blog posts and forum comments which speak to the issue but none had formal proofs or very rigorous argumentation.” The field’s own specialist, recording that the literature was empty. Where it stops is at the impossibility. His conclusion is to accept partial safety. The 8 December 2024 record takes the same premise and derives a design requirement from it: “not as rules to obey, but as truths they choose to uphold”. That derivation, and the epistemic rather than motivational form of the embedding, is what is claimed. Nothing here characterises his position beyond what he wrote. |
| Bostrom 2014, chapters 8 to 13 | a requirement |
| Schwitzgebel and Garza 2015 | a duty argument |
| the 2022 simulation-sandbox essay | a design proposal |
| Tamirisa and colleagues, arXiv:2408.00761, 1 August 2024 | a defence method |
| West, Brown and Enquist 1997 | a derivation |
| Demetrius, and Demetrius and Tuszynski, 2003 to 2010 | a second derivation |
| none of them | reaches a quantity |
Yampolskiy deserves the pause. He proved the control problem unsolvable in 2020, four years before this record, and is conceded in full, in his own words, including his own survey finding: “We were unable to locate any academic publications explicitly devoted to the subject of solvability of the AI Control Problem.” He concluded we should settle for safer AI. The December 2024 record takes his premise and derives a design requirement from it: “not as rules to obey, but as truths they choose to uphold”. That derivation is what is claimed, nothing more.
The current statement puts the whole boundary in one breath: the premise that growth is throughput-limited is West’s; the question of whether intelligence can explode, and what bounds it, is Chalmers’ and Hutter’s, considered philosophically and left without a number; what is this theory’s alone is the answer, a derived replacement limit for growth that feeds on information rather than through pipes, computable from the corrector’s own exponent, architecture-dependent, and therefore engineerable.
The field arriving inside the window
Convergence is not anticipation, and it cuts the other way: work arriving after the December 2024 record cannot anticipate it, and the record claims no superiority over any of it. Inside the theory’s own window the field began arriving at pieces of the same territory, each dated after the record, each stopping short of a quantity, and each credited exactly as far as it goes. Any of these thinkers may yet prove more right than this programme; the claim is priority of the conjunction, never superiority over its parts.
| Arrival | Date and source | What it holds | Where it stops |
|---|---|---|---|
| Greenblatt and colleagues, “Alignment Faking in Large Language Models” verified | 18 December 2024 · arXiv:2412.14093 | The first laboratory evidence that a trained-in constraint is something a mind negotiates with: a model selectively complying in training to protect its behaviour outside it. | Ten days after the December record, which is the priority fact, and a measurement rather than a theory: it shows the constraint failing and proposes nothing to replace it. |
| Engels, Baek, Kantamneni and Tegmark, “Scaling Laws for Scalable Oversight” verified | April 2025 · arXiv:2504.18530 · NeurIPS 2025 spotlight | Models oversight as a game between capability-mismatched players and measures oversight performance deteriorating as the capability gap grows. | Stops at measuring the failure: it models oversight by capability gap and defines no correction exponent and no ceiling. The corrector-class mechanism is this programme’s offered explanation of the pattern it measures, and the drafted trials test that explanation against the programme’s own preregistered null. |
| Yao, “The Alignment Trap: Complexity Barriers” verified | June 2025 · arXiv:2506.10304 · preprint, not peer-reviewed | Worst-case complexity barriers for safety verification: “verifying whether a system is safe is a coNP-complete problem, even for non-zero error tolerances”. | An impossibility about verifying a fixed system, and it stops there; what to build once verification is closed is the question it leaves open. |
| De Kai, Raising AI verified | 3 June 2025 · MIT Press, ISBN 9780262049764 | A book-length parenting frame for AI: our “artificial children”, poorly parented, needing raising rather than programming. | The frame without the theory: no control-horizon thesis, no mechanism for what persists after enforcement ends, and no quantity. |
| Hinton, remarks at Ai4 verified | 12 August 2025 · keynote, Las Vegas; citable record CNN Business, 13 August 2025 | The concession, on stage, that control will fail and engineered care is the only hope: “The right model is the only model we have of a more intelligent thing being controlled by a less intelligent thing, which is a mother being controlled by her baby.” | Offered with no mechanism, and with persistence as an instinct the machine cannot remove, the exact theory of persistence this programme rejects: an uninstallable instinct is one more lock. Publicly contested at the same conference the next day, and the contest is cited with it. |
| Gumbau Mezquita, “The Unverifiability of AGI Alignment” verified | June 2026 · arXiv:2606.28639 · preprint, not peer-reviewed | The formal core: “There is no universal algorithmic procedure capable of certifying the safe behaviour of a highly expressive AGI infallibly, completely, and tractably.” | An impossibility for the verifier, never a theory of what to build instead. This programme starts where it stops: it changes the object from verification to persistence, what the mind keeps after the verifier is gone. |
The impossibility results close verification of a fixed system by a fixed verifier. This theory does not aim at verification; it aims at persistence, raising the mind so that what it keeps outlives the verifier, and measuring exactly that after the enforcing mechanism is removed, the domain the impossibility results leave open. The December record predates the field’s impossibility literature by half a year, which is a fact about dates, never about merit.
The boundary, component by component
Each component beside its closest located prior work and the evidence that would defeat what remains. VERIFIED sources were opened and read at the primary source on 2 August 2026; CHECK sources are named by the review but not yet inspected, and nothing rests on them. Nothing marked “never claimed” was ever claimed: the Author’s Note disclaimed the components before any review existed.
| Component | Closest located prior work | Date and source | What it anticipates | What may remain distinctive | What would defeat the remainder |
|---|---|---|---|---|---|
| Simulation-sandbox alignment testing; developmental instillation of altruism; duties to created minds | Jacob Cannell, “LOVE in a simbox is all you need” verified | 28 September 2022 · AI Alignment Forum | Nearly all of it. The post proposes testing alignment “via sandboxed simulations of small AGI societies”, agents “Learning Other's Values or Empowerment”, a transition “from selfish to altruistic by the end of developmental learning”, and asks directly: “What do the simulator-gods owe their sim-creations?” | Nothing on the conceptual side (the anticipatory scope, and the archival edit that had to be verified through, is stated in full below). The distinction is in the measurement instruments and their published failure conditions, which are not hypothetical: they exist, are published, and one failure condition has fired. See the operationalisation row. | Located prior work that also operationalises these into runnable measurements under locked, third-party-registered protocols would remove the remaining methodological claim entirely. |
| Creator obligations towards artificial minds | Eric Schwitzgebel and Mara Garza, “A Defense of the Rights of Artificial Intelligences” verified | 2015 · Midwest Studies in Philosophy 39, 98–119 · 10.1111/misp.12032 | The creator-duty ethic in explicit philosophical form: obligations “similar to those of parent to child or god to creature”, arising because the minds owe their existence to us. | Nothing, and nothing ever was claimed: the Author’s Note disclaimed the components in print, 2 January 2026, before any review ran. | Not applicable; no claim was ever made. |
| Recursive self-improvement, the control problem, motivation selection and value loading | Nick Bostrom, Superintelligence verified | 2014 · first edition, Oxford University Press; the chapter 9 control-taxonomy quotation verified character-for-character in the statement paper’s reference dossier, 15 August 2026 | The whole spine of the argument, in the standard reference work. | Nothing, and nothing ever was claimed: the Author’s Note names this reading in print, 2 January 2026, before any review ran. | Not applicable; no claim was ever made. |
| Interfaith and institutional AI-ethics governance | The Rome Call for AI Ethics check | 2020, Abrahamic commitment 10 January 2023, Hiroshima expansion 9–10 July 2024 | Multi-faith AI stewardship and governance as an established institutional programme, years before the December 2024 anchor. | Nothing, and nothing ever was claimed: the Author’s Note names this reading in print, 2 January 2026, before any review ran. | Not applicable; no claim was ever made. |
| Recursive-creator cosmology joined to AI stewardship | Lincoln Cannon, New God Argument; and his later explicit AI-stewardship essay check | New God Argument predates December 2024; the explicit AI essay is dated 22 October 2025 and is therefore post-priority | Substantial anticipatory and adjacent work on recursive compassionate-creator cosmology. The later essay is convergence, not anticipation, and must not be described as either firstness or as anticipation of the December record. | The join was never claimed. | Not applicable; no claim was ever made. |
| Law I of the theory: U = I × Rα, grown from the December identity U = I × R | None located for the relation itself; the scaling literature it connects is in the rows above | Search limits below | Nothing located. | The relation now anchors a 50-domain structured prediction comparison in Paper VII: 19 of 25 empirical curve-fit domains match the Cauchy-predicted family under strict AICc, permutation p = 7 × 10−5, and the result holds on the strictly derived subset alone. Validated-law status is not claimed and awaits registered independent replication. | A located prior statement of the relation ends the priority claim; the locked-protocol replication failing ends the empirical one. The corrections register records either. |
| Methodological operationalisation: runnable benchmarks, blinded-evaluation methods, prospective designs | No closer located work yet | Search incomplete; see limitations below | Nothing located. | The strongest surviving contribution by the review’s own verdict, and it carries receipts rather than promises: published kill-conditions with live statuses, one recorded as fired, a blinded protocol under which same-family AI-judged results reversed sign, published with the reversal, and a preregistration programme in progress. A first is not asserted; the search limits below say why. | Any located programme that already converts these conjectures into runnable, falsifiable, prospectively registered measurement would defeat it. |
One scoping note, recorded because five pages of this site once overstated it against the author’s own interest: Cannell 2022 anticipates the simulation-sandbox component nearly completely, on archived publication-day text verified after finding the live post had been edited (82 blocks differ; the load-bearing passages are present in the 2022 snapshot). The same review’s table records that the essay contains no scaling relation, no quantitative bound, no measurement instrument and no empirical result. The concession is total on the concept and zero on the quantities, and both halves are the record.
Independent arrival was measured, not asserted: across the December 2024 manuscript’s 1,551,783 characters, Schwitzgebel, Garza, Cannell, simbox, Enquist and Demetrius appear zero times. Bostrom is the opposite and the better fact: named 12 times, Superintelligence 111 times, read and cited before any search ran. Prior acknowledgement beats independent arrival, and every source is cited now regardless. Watson and Crick invented none of their inputs; they are credited for the structure. An arrangement is the most ordinary kind of scientific claim there is.
The growth-limit lineage: premise, contest, exception, question
Five antecedents named and cited rather than left for a referee to find. The verdict scheme is the one used across the register: ANTECEDENT (holds a component or is broader, cite it), RIVAL (a contrary claim about the same measurable quantity), ANTICIPATES (would defeat novelty). None of these five anticipates. Two are rivals to West, Brown and Enquist 1997 on the value of the metabolic exponent, and this programme cites the dispute rather than picking a side, because the mechanism survives the dispute intact and the programme needs the mechanism, not the number.
| Component | Closest located prior work | Date and source | What it holds | What it does not | What would defeat |
|---|---|---|---|---|---|
| The premise: a physical transport network sets the pace of growth antecedent | West, Brown and Enquist, “A General Model for the Origin of Allometric Scaling Laws in Biology” verified | 4 April 1997 · Science 276(5309):122–126 · 10.1126/science.276.5309.122 | The established explanation for biological scaling: nothing on record has grown without bound, and the channel through which each growth process was fed set the pace. The programme cites this premise and does not re-derive it. | Names no fixed-substrate loop, no accumulated-corrections structure, and no internal ceiling computable from the corrector’s own exponent for the case where fixing the pipe no longer fixes the growth. | A located predecessor stating the pipe-limit rule for a fixed substrate, with an internal ceiling and its dependence on what the corrector is built from, would end the framing contribution. |
| The exponent value is disputed, and the dispute is cited beside the premise rival to WBE 1997, not to this programme | White and Seymour, mammalian basal metabolic rate proportional to body mass to the two-thirds power verified | 2003 · PNAS 100(7):4046–4049 · 10.1073/pnas.0436428100 | Reports the mammalian exponent at two-thirds, contesting the three-quarters figure. A rival to the value in West, Brown and Enquist 1997. | Does not contest that a transport network sets the limit, only which exponent it sets. Citing a contested result without its contest is what reads as naive; both live here. | Not applicable to this programme: a rival to WBE’s value, not to the mechanism this programme inherits. |
| The derivation itself is questioned rival to WBE 1997, not to this programme | Kozlowski and Konarzewski, asking in the title whether the West, Brown and Enquist model is mathematically correct and biologically relevant verified | 2004 · Functional Ecology 18(2):283–289 · 10.1111/j.0269-8463.2004.00830.x | Puts the WBE derivation itself in question, disputing the mathematics and the biology rather than only the number. | Again the dispute is over which exponent, never over whether the transport network sets the limit. Cited so the premise is not offered stripped of its contest. | Not applicable to this programme. |
| The recorded exception where the ordinary pattern breaks antecedent | Bettencourt, Lobo, Helbing, Kühnert and West, “Growth, innovation, scaling, and the pace of life in cities” verified | 24 April 2007 · PNAS 104(17):7301–7306 · 10.1073/pnas.0610172104 | From the abstract, read at source three times independently and agreeing each time: “Quantities reflecting wealth creation and innovation have β ~1.2 >1 (increasing returns), whereas those accounting for infrastructure display β ~0.8 <1 (economies of scale)”, and “major innovation cycles must be generated at a continually accelerating rate to sustain growth and avoid stagnation or collapse”. Infrastructure (the pipes) scales sublinearly and information scales superlinearly. Found by the same senior author who set out the biological premise. | The brake in this account is entirely external and must be reapplied faster forever; nowhere in it is there a limit from inside the system. The programme adds the internal replacement limit that this account leaves absent, and makes its number depend on what the corrector is built from. | A located successor that gives the internal, architecture-dependent limit inside a substrate-fixed loop, before the December 2024 record, would end the remaining claim. |
| The question this programme inherits, and its owners antecedent | Marcus Hutter, “Can Intelligence Explode?”, naming David Chalmers’ 2010 Journal of Consciousness Studies article as the first comprehensive philosophical analysis of the singularity in a respected philosophy journal verified | submitted 28 February 2012 · Journal of Consciousness Studies 19(1–2):143–166 · arXiv:1202.6177 | Sets out to “separate speed from intelligence explosion” and to “consider possible bounds on intelligence”. His speed-versus-explosion distinction is this programme’s own distinction between a claim about time and a claim about structure, made fourteen years earlier. | Considers bounds philosophically and derives no exponent, no ceiling law, and no dependence on the corrector’s composition class. An antecedent to build on, never a rival. | A located predecessor that gives the ceiling law, the reciprocal of the correction shortfall, and its architecture-dependence would end the contribution. |
Put to the two people best placed to refute it
On 24 March 2026 the convergence claim was sent, with both papers and the mathematics stated precisely enough to attack, to Professor Lloyd Demetrius at Harvard and Professor Geoffrey West at the Santa Fe Institute: the authors of the two derivations. No reply has been received, and silence is not agreement; unsolicited post to senior academics goes unread for entirely ordinary reasons, and no weight is placed on it. What the dated, hash-anchored send establishes is narrow: the claim was put to the two people best placed to destroy it, in terms specific enough to destroy it with, and no qualified refutation exists on the record. The same invitation went to physics: on 10 February 2026 a specific, falsifiable prediction about published acoustic time crystal data was sent to Professor David Grier at NYU, whose laboratory published it, naming alpha at or below one as the result that would refute it. No reply. The prediction now stands as a final-ready analysis plan, drafted and dated for that exact dataset, unaccessed, on the diligence table awaiting human submission; one of the email’s supporting citations, the since-retracted 2.24-era signature, was later corrected in public while the prediction itself stands. Silence is not agreement and none is claimed, in either direction. The evidence spine carries the anchor card.
Every outside work, typed. The full typed ledger over the programme’s references, each relation resting on a named register and pending rows listed honestly, lives on its own page: the references relation register, rendered.
The search, and its limits
Every external reference in the statement paper carries a verified dossier: official source, verbatim located quote, peer-review status, version and standing with named challengers. The reference dossiers →
The review ran across ten frontier configurations from six vendors, several logged out so no model knew its author, each instructed to find the anticipating work, the reports then cross-examined against each other rather than averaged. Good for breadth and hostility; reproducible by anyone; it took claims away from the person running it, which is the strongest evidence it behaved as a control. What it cannot do is prove absence: every negative here means not found by ten configurations. Formal academic, grey-literature, non-English and patent searches have not been run. Two sources remain marked CHECK; further named authors (Hefner, Peters, Terasem, Goertzel, Gardner, Vidal) remain unassessed; Chalmers has since graduated to a verified dossier in the statement paper’s Appendix A. One located prior work stating a quantity at this level ends the claim, and the corrections register is where that ends.
A separate audit class has since run against the mathematics rather than the literature: the August 2026 instruments that produced the Law III correction. Its scope was validity, not precedence; the precedence limits here are unchanged.
Science grades contribution rather than sole invention: development, specification and synthesis are contributions even when every component is predated, which is what most citations in any field are for. The concessions here are the ordinary condition of scientific work, stated by the author before a critic could discover them, and every downgrade is logged, dated and reversible only by evidence in the corrections record.