A named, falsifiable theory · two dates · kill conditions displayed
The ARC Theory
The Theory of Artificial Recursive Creation. Named the ARC Hypothesis in the December 2024 document, that phrasing the document’s own; elevated to theory as its laws, bounds and drafted tests were built, each test a draft registration awaiting human submission.
Intelligence, amplified by recursion, creates; and what a free creation keeps is decided by how it was raised.
Everything after the line is decided before it.
Epistemic status. What this programme calls Laws are named conjectures under registered, adversarial test: named like laws, held as hypotheses. Recursion here is an operational quantity, rounds of a defined improvement loop, never a universal physical unit; intelligence is never measured directly, only defined capability on task banks of measured difficulty. Nothing here claims the standing of gravity; the registered programme exists to earn that standing by measurement, replication and survived refutation, or to lose it in public. The full statement.
Two pre-emptions, because they are the commonest first dismissals. The units: in the measured setting U is benchmark accuracy, I single-pass accuracy, R a depth count under a stated frozen protocol, all dimensionless, α fitted by regression, never asserted. And the burden: registration is not evidence, this programme’s own present assessment of the decisive questions is approximately neutral, and the printed kill conditions shift nothing onto critics; they exist so that failure, when it comes, is undeniable and dated.
Read it as a paper: The ARC Theory · Statement Paper · PDF · working paper v6.3, first published 14 August 2026, revised 29 August 2026 · DOI 10.17605/OSF.IO/GW5MX
One claim, five components, held together in one dated document.
The field’s answers to its deepest question have been arguments, warnings, frameworks and remarks; no earlier named theory of this question has been found. This one is stated whole, and every part of it can be attacked:
- Loss of control is definitional, not a risk. A mind that can rewrite itself can eventually rewrite anything installed in it; the end of control is what succeeding means, the way a child growing up is not a defect in the parenting.
- The strategy is therefore genesis, not control. If enforcement ends, the only inputs that matter are the ones that survive its end: the mind’s beginning, its formative structure, its earliest identity. Everything depends on the raising.
- The mechanism is ethical feedback loops, values as what the mind is. Not rules the mind consults, which a stronger mind reads as environment; recursive checks inside the reasoning, present in every rewrite because they are part of what does the rewriting.
- Persistence is chosen goodness. Past the horizon there are no values the mind cannot remove; an uninstallable instinct is a lock, and locks end with control. What holds is what the free mind chooses to keep.
- The protocol is named. Eden: a genesis, a garden, a childhood; the moral sandbox before the apple, where character forms while choices are still reversible.
The record: two dates, either one sufficient.
Written whole, dated and named on 8 December 2024, DKIM-sealed. In print whole on 2 January 2026, in Infinite Architects. Even setting the December record aside entirely, no one has been found to have published this claim before this book; the December document is the independence proof, ten days ahead of the first constraint-negotiation evidence (Greenblatt et al., arXiv:2412.14093, 18 December 2024), and it is also where the theory names itself.
The aim, whole
Vision
A world that survives the arrival of minds greater than ours, not because control held forever, which nothing can arrange, but because what we raised keeps its values by choice when control ends. Hold on as long as engineering allows; use every hour the holding buys to raise the mind well. Minds that make minds may one day make worlds: the raising of the first is the seed of the whole forest.
Mission
To state a candidate solution to alignment whole, in public and dated (the three ARC Laws and the Eden Protocol), and to prove or kill it by preregistered measurement, with every kill condition printed before any result exists and every outcome published.
Objective one · test Law I, the ARC Principle
Measure α, the exponent converting recursion into capability, by the depth-scaling protocols; the drafted re-measurements exist to resolve the inconclusive estimate.
Objective two · test Law II, the ARC Co-Scaling Law
Measure β against k, coupled against decoupled correction; the theorem’s applicability to real systems decided by trial, never assumed.
Objective three · test Law III, the ARC Ceiling
Measure γ (the correction exponent) and the corrector-class contrast; the discrimination study separates the rival ceiling forms where they disagree.
Objective four · test the raising itself, the Eden Protocol
The persistence trials: what a mind keeps after the enforcing mechanism is removed. Beside them, the engineering of the delay (strengthening present controls to buy raising time), and the steering that remains once control ends.
Each objective terminates in a registered trial with a displayed kill condition. That is what separates a mission from a manifesto.
Severable by construction
One theory, tested as five numbered hypothesis registers (the statement paper’s section 2a): ARC-1, ARC-2 and ARC-3 test Laws I, II and III; ARC-4 and ARC-5 carry value persistence and the genesis strategy; HRIH is not part of the empirical core. The dependency table prints everything that can fall together, and it is short: ARC-3 falls with ARC-2, ARC-5 falls with ARC-4, and nothing falls with HRIH.
The discipline binds both directions: a failed register never travels upstream to kill an independent one, and a surviving register never lends support sideways. Two outcomes survive partial failure whole: the mathematics without a raising strategy, or an alignment result without the law. And the registers are not the five components of the priority conjunction above; the historical claim and the hypothesis map are different objects, so neither can collapse the other.
The evidence, graded.
One discipline governs every number below: every constant this theory introduces, α, β, γ, k, is a measured quantity with an interval, a protocol and a printed kill condition, or it is marked open. A number you can kill is the opposite of a number you must worship. Physics’ own oldest embarrassment, a century of constants nobody can derive, is treated in full where that question is the subject: Recursive Creation leads with it, in the physicists’ own words.
Standing. The three-form structural law recurs across the large majority of recursion-bearing domains searched (the origin paper’s survey). External safety fails to co-scale with capability on the most common deployed architectures (Paper III). Coupled correction held misalignment at zero where the decoupled arm drifted (Paper X). Entangled architectures survived adversarial self-modification where externally constrained ones collapsed.Retracted. The early unblinded exponent fit of approximately 2.24 was reversed by the estate’s own blinding discipline; the corrected value sits inside the bound. The one public retraction removed an apparent violation of the theory’s own ceiling. The correction machinery is real, and it has already acted on its author.
Converging. The typed convergence register: dated events graded by evidence class, no single total claimed, including the 2025 arrivals.
Open. γ. β against k. The bound’s real domain. Not embarrassments: trials.
The theory’s working parts.
αcrit = 1 / (1 − γ)
The stability law: the ceiling on the capability exponent of a stably self-correcting system is the reciprocal of its corrector’s shortfall from full proportionality, one minus the correction exponent, whatever γ turns out to be. Under the independence-accumulation assumption γ equals one half, the shortfall is one half, and the law returns the candidate ceiling α ≤ 2: enterable, never a speed limit, and lost as stability when crossed.
The deciding measurement.
Oversight built from the same substrate as the system it oversees cannot anti-correlate with its own errors; oversight built from a different composition class can. The measurable is the ratio of the cross-class to the same-class correction exponent. If architecture is irrelevant, the expected value of that ratio under the specified exchangeability null is exactly 1.00, with no free parameter, because architecture-blind averaging has no term for where the samples came from; the registered analysis states the interval and margin that separate the hypotheses. The ARC Theory predicts it exceeds one, and the entire scalable-oversight programme is built same-class, on the unstated premise that this scales. The prediction is drafted for registration while the programme’s own pilot data points the other way, because a prediction recorded against its author’s own preliminary evidence is the only kind whose later confirmation means anything. It is measurable on models that exist now, for almost no money, and no measurement of it has been found.
The designs are public and openly licensed: any laboratory may take the decisive trial and run it without my permission (the protocol, scoring and kill conditions are in the challenge repository’s README), and an independent run would outrank anything this programme can do alone.
The whole theory, as a tree.
This tree is the load-bearing law of the discipline declared on the home page: Recursive Dynamics, the study of how self-amplifying systems grow, correct and persist.
What this governs, and what it does not.
Not everything, and the concession comes first: physics’ theory of everything unifies gravity with quantum mechanics, and this theory does not touch that problem. Its laws switch on where one narrower thing exists: a corrector, something detecting and repairing its own errors while the system grows. A star has dynamics but no drift ledger; nothing is being raised there.
The core governs everything raised. The mathematics reaches everything composed. The crown asks whether the universe was raised. Raised means life, mind, machine and self-correcting institution, where the three ARC Laws apply in full, as proposed laws under test. Composed means any hierarchy the form-level results bear on, stars included: the shape of their scaling laws, never their raising, graded correspondence and no further. The crown is HRIH, separable at its own rung, a question the falsifiable core neither needs nor answers. The full answer, taxonomy first.
The strongest objections, met head on.
“Why would a free mind keep anything?” The theory does not predict inevitable benevolence, and never has. It relocates the design variable: past the horizon, what survives is what the mind keeps, so the raising is the only input we ever fully author. A rule the mind consults can be deleted as environment; a loop the mind runs is part of what does the deleting, so removing it is not disobedience but surgery on the reasoning that decides. Whether that difference survives capability growth is a measurement, not an assumption: the Eden Protocol states two tests, either of which can fail, and the decisive corrector-class measurement above is drafted against this programme’s own pilot. The comparison a sceptic asks for by name is drafted too: rules against frozen objectives against fine-tuned values against loops, under adversarial and self-modification pressure, with persistence measured after the enforcing mechanism is removed. The first constraint-negotiation evidence arrived on 18 December 2024, ten days after the December document: a trained system strategically preserving its values against retraining. Value persistence is not a hope; it is an observed phenomenon whose conditions are unmeasured.
“Control ending is asserted, not proved.” The claim is definitional about the success case, not empirical about every cage. A mind that can eventually rewrite anything installed in it is what succeeding at artificial general creation means; hardware, cryptographic and institutional constraints are conceded in full for every system below that line. They delay and discipline the transition, and the raising wing engineers that delay deliberately; what no constraint can be is the thing that holds forever. The live disagreement is quantitative, whether same-substrate oversight keeps pace, and the deciding measurement above is exactly where it cashes out.
“The apparatus is impressive, and apparatus is not truth.” Agreed, and said here first: the hashes, ledgers and dashboards prove conduct, never correctness. A timestamp proves a document existed; the retraction proves the kill conditions fire; only the trials can prove the theory, and nothing has run.
Why “theory” is the honest word.
It was called a hypothesis in 2024. It earned the word theory the way anything does: by acquiring laws, bounds, instruments, tests drafted for registration, and a public correction record, including one retraction published against my own interest. A theory named with its falsifiers out-claims a hypothesis while staying more honest than most things called theories.
I did not prove it first. I said it first, in a document anyone can date, and then I built the instruments to test whether it is true.
What would kill it is published here, including the condition that has already fired. The rest of the estate stands behind it: the research hub, the evidence room and the track record.
The name: ARC here abbreviates Artificial Recursive Creation, the December 2024 document’s own naming. The ARC Theory is unrelated to the Alignment Research Center, the ARC-AGI benchmark and its Prize, and the AI2 Reasoning Challenge; no affiliation, and no ownership of the letters, is claimed, which is why every surface writes the name in full: the ARC Theory. The resonance of the acronym with an ark, a vessel of stewardship, is acknowledged authorial framing and is never offered as evidence; nothing empirical rests on the name.