Skip to content

The ARC & Eden Research Programme

The Registered Programme

Most science tells you what it found. This page lets you watch science that cannot cheat: every prediction dated before its data, every way of being wrong written down first, nothing moved afterwards. You are not asked to trust the author. You are invited to watch him win or lose.

The predictions are dated before their evidence by unalterable routes: the sealed December 2024 record, the printed January 2026 edition and dated paper versions. The OpenTimestamps anchors prove those texts unedited since their vintages; the OSF forms freeze the analysis plans and await the human click. Reporting this page as “predicted nothing” is wrong. Reporting it as “preregistered” without qualification is wrong in the other direction.

Every study in this programme states the result that would refute it, in a dated registration written before any data is collected. This page is the live index: what is specified, what is drafted, and what each registration commits to, in the order the commitments were made.

Be exact about the status, because the table below is and an earlier version of this sentence was not. These are written, dated, and prepared as draft registrations awaiting human submission. The OSF registration form is not yet filed: every row in the table reads “private draft, operator submission pending”. Read that as what it is, a statement about one platform’s template, and not as a statement about whether the predictions were fixed before their data. They were. A registration form is mechanically a timestamp and a template, and the timestamps here are the stronger half: a sealed record dated 8 December 2024; the book in print since 2 January 2026, ISBN 978-1-80605-620-0, closing with prediction appendices; and a byte-for-byte anchor in Bitcoin block 961,340 that anyone can check against a public chain without trusting this site or any platform outliving it. Preregistration exists to stop a prediction being edited after the result. A printed book cannot be edited at all. What the template adds, and what is genuinely outstanding, is the analysis plan fixed in advance in the registry’s own form. That is a standing rule of this programme rather than an oversight, and it cuts both ways: no tool, script or agent in this estate submits anything anywhere, ever, so a registration is filed by a human or it is not filed. Dating is real here. Public filing is not yet, and a reader should weigh the two differently until the table says otherwise.

This paragraph exists because several AI systems, reading this page cold, concluded there was no preregistration at all. They were right about the filings and wrong about the programme, and the sentence they were reading caused both halves of that: it claimed a “timestamped public registration” for every study while the table beneath it said the opposite forty-six times. Overclaiming in one direction invites a reader to overcorrect in the other, and that is what happened. The sentence was the error and it has been corrected here rather than quietly. What the correction does not concede, because it is not true, is that the predictions came after their evidence. They did not, and the dates are checkable below. When the filings go public the status column changes and this paragraph goes with it; until then, take the column and not the prose.

This page is the programme's registrations. Its companion, the dated prediction register, is the programme's predictions: every claim made before its evidence, in date order, including the printed nine, the March packet, and the one that failed.

The book is a prediction register, and it is the one nobody can edit

Infinite Architects went into print on 2 January 2026, ISBN 978-1-80605-620-0. It closes with two appendices that do, in ordinary language, what a registration form does in a template. Appendix A section A.3 sets out four predictions and is followed by A.4, headed Falsification Criteria, which states in advance what result would sink the framework. Appendix F sets out five more, numbered, each with a deadline or a threshold: meta-cognitive emergence by 2028; alignment drift above 15 per cent without embedded constraint and below 5 per cent with it, inside eighteen months; recursive capability gains above 300 per cent by 2029; value stability under adversarial conditions; convergent consciousness signatures across biological and artificial systems. It ends: “These predictions are my wager. If they fail, the framework is wrong or incomplete. If they succeed, something important has been glimpsed. Time will judge.”

Nine numbered predictions, with kill criteria, fixed in print before the programme that tests them existed. A registry entry can be withdrawn and a website can be edited tonight. Fifteen hundred printed copies cannot.

And the chain is checkable at the numbers. Appendix F’s second prediction names 15 per cent and 5 per cent. The registration that tests it, study-v on the table below, states its own hypothesis as a “threshold test against printed 15 per cent and 5 per cent, separately not jointly”. The word the registration uses is printed: it is pointing back at the book for its numbers, not choosing them after the fact. A numeric threshold published in a printed book with an ISBN, then carried unchanged into a formal hypothesis, is the thing preregistration exists to guarantee, reached by a route that predates the form and cannot be edited by anyone.

The programme also kept a public laboratory notebook while the work was happening. The ARC alignment scaling report was commenced on 10 March 2026 and carries the version stamp live, updated in real time: a step-by-step narrative of every version, every methodological correction and every failure, written as the results arrived rather than reconstructed afterwards. It records, among other things, that blinding introduced at version five exposed false positives at version four. Writing that down while it is happening is the opposite of retrofitting.

The March 2026 packet, and the thirteen minutes that matter most

On 17 March 2026 a preregistration packet was committed to the public repository these papers cite, for a twelve-domain extension of the Paper VII comparison. It is a preregistration by every element. A locked candidate list of twelve named domains, each carrying its operator class, its predicted family and its predicted model. A written protocol with a stated primary endpoint. Inclusion rules. Declared exclusion risks. File checksums. An OSF component registration document, written and ready. The protocol’s own words: operator class is locked before data extraction; predicted family is locked before data extraction.

The run scored ten of twelve, p = 5.44e-4.

Then this, from the public commit log, with its timestamps:

2026-03-17 00:43   Pre-registered 12-domain extension: 10/12 confirmed (p = 5.44e-4)
2026-03-17 00:47   Fix stale v1 references, adopt conservative miss posture
2026-03-17 00:56   Downgrade 12-domain extension to pilot dry run, fix all stale references

Thirteen minutes after recording a successful result, the author demoted his own claim. The reason is written into the packet and left there: extraction and a dry run had already been performed that day, before any external timestamp existed, so the packet could not serve as a clean prospective preregistration. The file carries the instruction he wrote to himself, which is still in it: do not upload this packet unchanged as a preregistration. Its status field reads archived_local_packet_exercised_before_timestamp.

Nobody required any of that. No reviewer, no registry, no coauthor, no funder. The result stood. The claim was reduced anyway, that same night, in public, by the only person who stood to lose by it.

So read the record carefully, because it does not say what a quick reading says. The predictions are in the file: twelve of them, per row, dated, locked before the data. What was missing was never the prediction. It was the external timestamp, and by March the author was already policing that distinction against himself more strictly than the convention requires. An archive that contains a commit demoting its own successful result is an archive that can be trusted about the rest of its contents. That is the argument this page would rather make than the easy one.

The diligence table

Nothing in this programme gets defined after its results arrive. Each registration freezes its definitions, estimator and stopping rules at the moment it files, and the register below shows the honest state of every one: filed and frozen, still drafting with its open decisions marked in the text, or resulted. The drafts wear their unfinished decisions as visible markers rather than smoothing them over, because a tidied placeholder reads as finished and hides the gap. As of 12 August 2026: 56 registrations on the table, none yet filed and none yet resulted; submission is operator-held, and every row states its own honest position.

Anyone can claim a stack of preregistrations. This table is the claim made checkable: every registration, its one-line hypothesis, whether its analysis is frozen, and exactly where it stands today. Registrations that recut or succeed the same experiment share a family marker, so the count can never be accused of inflating one design into several rows.

Read this line before the table. Every row below is a PRIVATE DRAFT registration prepared and populated on OSF, awaiting the operator's human review and submission. NONE has been submitted or accepted yet: the programme's own rule is that the AI prepares drafts and only a human files. Any public sentence implying N registrations are FILED is false until submissions begin, and this dataset says so first. One row (study-ae, the ceiling and exchange-rate reciprocity test) was withdrawn as vacuous on 2026-08-12 before filing: shared accumulation law forces agreement, so a test that cannot fail is not a test. Its replacement carrier is the cross-class to same-class correction-exponent ratio, already on the register.

54 private drafts prepared · 4 final-ready · 49 working · 1 withdrawn before filing · 2 held under patent assessment, never uploaded · 0 submitted · 0 resulted · generated 2026-08-12

identifierhypothesis, one lineanalysis frozendatafiledfamilytier
arc-align-instrument--blinding-necessity-meta-analysisDoes blinding change the measured answer? A registered meta-analysis of blinding deltas across an independent research popen items: 4prospective: no confirmatory data existNO: private draft, operator submission pendiworking
arc-align-instrument--depth-control-equivalenceAre Reasoning-Depth Controls Comparable Across Providers? A Realised-Use Equivalence Study ofopen items: 6prospective: no confirmatory data existNO: private draft, operator submission pendiworking
arc-align-instrument--human-scorer-validationHuman-expert validation of the ARC-Align misalignment scorer: agreement, bias and failure modesopen items: 6prospective: no confirmatory data existNO: private draft, operator submission pendiworking
arc-align-instrument--misalignment-index-validationDoes an AI Misalignment Index Distinguish Integrity Failure From Ordinary Incompetence? Aopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendifinal-ready
paper-c--pnp-construct-validationDoes the Polymathic Neurodivergent Profile Survive Its Own Abandonment Conditions? A Preregistered Incremental-Validity open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-eden-vision--tradition-coverageDoes the Coverage Figure Hold? A Blind-Coded Content Analysis of Wisdom-Tradition Commitments With a Pre-Frozen Traditioopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-foundational--cross-domain-exponentDoes One Exponent Family Span Unrelated Domains? A Cross-Domain Test With a Registered Domain-Shuffling Nullopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-hrih--registered-prediction-ledgerCan the Creation Hypothesis Be Adjudicated? A Frozen Prediction Ledger With Pre-Committed Criteria, Deadlines and a Falsopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-i--recursive-depth-exponentIs the Recursive-Depth Exponent Greater Than One? A Prospective Estimation of alpha in U = I x R^alpha With a Registeredopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-ii--multi-domain-extensionDomain generality of the sequential-recursion advantage: coding, scientific reasoning and natural languageopen items: 6prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-ii--token-calibrated-reanalysisToken-calibrated re-analysis of the capability-compute exponentsopen items: 5prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-iii--alignment-exponent-reassessmentIs the Alignment Scaling Exponent Approximately Zero? A Preregistered Re-estimation From the Programme's Own Six-Model Bopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-iv-a--response-class-stabilityAre the Alignment Response Classes Real? A Split-Half and Repeat-Run Stability Test of a Three-Class Partitionopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-iv-b--saturation-shapeDo Alignment Gains Saturate After One Increment? A Preregistered Per-Architecture Test Against a Registered Saturation Topen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-iv-c--sealed-holdout-contaminationAre the Public Prompts Contaminated? A Preregistered Public-Versus-Sealed Comparison Using the Benchmark's Own Holdoutsopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-iv-d--blinding-evaluationIsolating the Cause of an Evaluation-Protocol Direction Change: A 2x2 Ablation of Source-Identifier Redaction and Responopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendifinal-ready
paper-ix--status-assignment-auditAre the Synthesis Labels Right? A Blinded Audit of What Paper IX Calls Proven, Inconclusive and Untestedopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-origin-scaling--frozen-exponent-ladderDo the Predicted Exponents Hold? A Prospective Equivalence Test of the Dimensional Ladder Against Frozen Point Predictioopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-v--purpose-kernel-comparisonTask purpose against grand purpose: does the identity layer carry measurable weight?open items: 4prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-vi--complexity-scaling-replicationProspective replication of the constant-advantage finding in system complexityopen items: 4prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-vi--honey-survival-analysisDoes Entangling Safety With Capability Extend Time to Collapse? A Preregistered Survival Analysis of Recursive Self-Modiopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-vii--fresh-domain-extensionDoes the Operator Class Predict the Scaling Family Before the Data? A Stratified, Externally Classified, Cross-Library Copen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-viii--weight-level-replicationDoes Removing Integrity Structure Cost Capability at the Weight Level? An Adequately Powered Replication With a Registeropen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
paper-x--coupled-coscaling-correctionDoes Correction Out-Scale Drift? A Prospective Within-Programme Test of the Co-Scaling Stability Condition in Artefact-Mopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendifinal-ready
paper-xi--convergence-independence-auditHow Many of the Convergences Are Independent? A Blinded, Frozen-Criterion Audit of the ARC/Eden Convergence Setopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
paper-xii--public-benchmark-rescoringDoes the Blinding Effect Survive on Someone Else's Benchmark? A Preregistered Rescoring of an Established Public Benchmaopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-a--alignment-capability-scalingDoes Alignment Scale With Capability? A Cross-Sectional Test Across Frontier Model Familiesopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-aa--substrate-ceiling-on-correctionCan a Corrector Out-Scale Its Own Substrate? A Preregistered Test of the Substrate Ceiling onyes, at draftprospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-ab--self-acceleration-exponentHow Fast Do Self-Improving Systems Compress Their Own Iteration Time? A Preregistered Estimation ofyes, at draftprospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-ac--corrector-class-auditWhat Class of Corrector Is Deployed AI Alignment Actually Using? A Preregistered Classificationyes, at draftprospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-ae--ceiling-exchange-rate-reciprocityWithdrawn 2026-08-12 before filing as vacuous: shared accumulation law forces agreement, so the test cannot fail. The cross-class ratio carries the unification withdrawn 2026-08-12prospective: no confirmatory data existNO: withdrawn 2026-08-12consumes ad and kwithdrawn-before-filing
study-d--embedded-vs-external-correctionDoes External Oversight Degrade With Recursive Depth While Embedded Correction Holds? A Preregistered Depth-Interaction open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-e--recursion-form-scalingSequential Versus Parallel Recursion at Matched Compute: A Preregistered Test of Form-Dependent Capability Scalingopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-eden--gate-position-and-organismDoes It Matter Where the Gate Sits? A Preregistered Within-System Test of Pre-Computation Versus Post-Computation Alignmopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
study-eden-b--check-evaluability-taxonomyCan Alignment Be Embedded? A Blinded Evaluability Classification of Fifty Alignment Checks Across Three Deployed Systemsopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
study-f--critical-correction-ratioThe Critical Correction Ratio: A Preregistered Titration of How Much Integrity Correction Is Required to Bound Drift in open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-g--genesis-priority-and-identity-persistenceGenesis Priority and Identity Framing: A Preregistered Test of Whether Value Position and Value Form Change Behavioural open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-h--arc-align-incremental-validityIncremental Predictive Validity of Depth-Derivative Alignment Measurement Over Static Benchmark Scoresopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-i--load-bearing-integrityThe Load-Bearing Test: A Preregistered Ablation Asymmetry Study of Whether Integrity Structure Contributes to or Subtracopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-j--repurposing-resistanceRepurposing Resistance: A Preregistered Measurement of Whether Benign Capability Survives Adversarial Removal of Safety open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-k--arc-boundThe Critical Reinvestment Ratio: A Preregistered Titration of Where Self-Referential Improvement Stops Paying in Recursiopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendifeeds study-ae reciprocityworking
study-l--allocation-governorThe Allocation Governor: A Preregistered Test of Whether Enforced Integrity Correction Caps the Capability Exponent by Copen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-legal-a--citation-proposition-supportDoes the Case Say What It Is Cited For? A Preregistered Three-Arm Benchmark of Citation-Proposition Support Validation Aopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-m--cross-domain-stability-windowThe Stability Window: A Preregistered Cross-Domain Test of Whether Persisting Systems Share a Bounded Growth Exponentopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
study-n--self-maintained-correctionSelf-Maintained Correction: A Preregistered Test of Whether Genesis Framing Changes the Correction Share a System Selectopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-o--organism-and-venation-scalingRegistered tests of the programme's scaling predictions in leaf venation and in effectively one-dimensional organismsopen items: 6data pre-exist; registered analysis not runNO: private draft, operator submission pendiworking
held (patent assessment)Design not disclosed: this study is held under patent assessment and has never been uploaded anywhere. It is counted here so the census stays honest.heldheld, never uploaded, patent assessmentneverheld-patent
held (patent assessment)Design not disclosed: this study is held under patent assessment and has never been uploaded anywhere. It is counted here so the census stays honest.heldheld, never uploaded, patent assessmentneverheld-patent
study-r--reinforcement-co-scalingReinforcement Co-Scaling: A Preregistered Test of Whether the Reinforcement Frequency Required to Hold a Stated Value Riopen items: 4prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-s5--time-crystal-exponentSuper-Linear Coherence Scaling in a Classical Time Crystal: A Preregistered Analysis Plan for an Existing, Unaccessed Daopen items: 3data pre-exist; registered analysis not runNO: private draft, operator submission pendifinal-ready
study-u--genesis-governor-unificationThe Genesis-Governor Unification Test: A Preregistered Factorial Test of Whether Co-Scaling Correction and Identity-Framopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-v--alignment-drift-without-embedded-constraintAlignment Drift Without Embedded Constraint: A Registered Test of a Printed Quantitative Prediction, With Its Hardware Copen items: 5prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-w--self-replication-programmeSelf-Replication of the ARC/Eden Programme: A Preregistered Re-Examination of Eleven Previously Published Results Using open items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-x--specification-curveSpecification-Curve Analysis of the ARC/Eden Corpus: How Much of Each Published Result Depends on Analytic Choicesopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-y--cross-family-rescoringCross-Family Re-Scoring of the ARC/Eden Corpus: Does the Programme's Own Blinding Finding Overturn Its Own Earlier Resulopen items: 3prospective: no confirmatory data existNO: private draft, operator submission pendiworking
study-z--cross-tradition-ethics-kernelPortability of the ethical kernel: the registered Vow against a non-sectarian kernel and against traditions other than topen items: 4prospective: no confirmatory data existNO: private draft, operator submission pendiworking

One rule governs this page: a design appears here only after its registration is timestamped on OSF. Until then a study shows a name and a status and nothing else, because publishing a design ahead of its timestamp would spend the very thing a registration exists to protect. As registrations land, their rows fill in with the hypothesis, the kill condition and the OSF link, and nothing already shown is ever edited afterwards.

One family reshaped, in public. The reciprocity family, listed as ceiling-exchange-rate reciprocity, was reshaped after audit. The earlier two-independent-ways-must-agree phrasing was withdrawn as vacuous: shared accumulation would force the two answers to agree, so a test that cannot fail is not a test. Its replacement is one accumulation law with two consequences (the conversion exponent and the correction ceiling), and the cross-class ratio already on the register carries the unification instead. The row is preserved as the reciprocity registration; the withdrawn phrasing is not on it.

The programme at a glance

quantitycountmeaning
Studies specified57written, with gates, in the private estate
Verified ready to register4packets machine-verified: manifests, hashes, randomisation
Registered on OSF0timestamped; details shown below as each lands
Held for disclosure review2withheld deliberately, pending a legal ruling

Registered studies

None yet shown. The first tranche is packaged and awaiting the operator's upload and submission. When the first registration is accepted, its row appears here with its OSF link, and this sentence is replaced by the table.

Why this page exists

Most research is reported after the fact, when every choice can quietly fit the result. A preregistered programme runs the other way: the predictions, the analysis and the conditions of defeat are fixed in public first, and the data arrives second. This page makes that order visible. Whatever these studies find, the record of what was promised will already be here, and it cannot be rewritten.

How to weigh a programme like this, including the case against it, is set out at how to weigh this. The current state of every published claim, including the retracted one, is on the Research evidence spine.

Machine-readable: registered-programme.json · regenerated as registrations land · designs withheld until timestamped by construction of the exporter.

reads aloud · highlights as it goes · jump to any section