FOR IMMEDIATE RELEASE

Independent Researcher Identifies AI Structural Vulnerability 10 Days Before Major Lab Confirms It

The record Eastwood points journalists to first is not the priority claim but a correction: the headline exponent in his early work failed replication against his own printed book, the robust estimate of roughly 0.49 replaced it, and the retraction is published and standing. The ten-day chronology below should be read in that light: stated narrowly, by a researcher whose record shows he corrects against himself.

"On December 8, 2024, ten days before Anthropic published their alignment-faking paper, Michael Darius Eastwood self-emailed a 221,236-word manuscript through Google servers proposing the ARC Principle (U=I*R), the Eden Protocol, and a falsifiable creation theory. The cryptographic evidence is published. You can verify it yourself."

The manuscript, timestamped at 02:45 UTC on December 8, 2024, contained the thesis that "AI systems cannot be truly controlled" by external safeguards and that only substrate-level embedding of ethics, named the Eden Protocol in the email body: could prevent circumvention. On December 18, 2024, Anthropic published their alignment-faking paper documenting a 78% faking rate in the reinforcement-learning training condition (against a 12% baseline).

Eastwood does not claim he predicted Anthropic specific experimental finding. He claims he identified the structural vulnerability, that external controls cannot scale with recursive capability, and has cryptographic evidence establishing temporal priority. "The honesty architecture IS the credibility architecture," Eastwood stated. "What I cannot claim is as important as what I can."

The typed convergence register (the dated event register)

Over the following 18 months, a typed convergence register of dated events, classified by evidence class from 11 domains independently arrived at the same structural conclusions. The list includes Anthropic, OpenAI, Google, Meta, Nvidia, Stanford, NYU, the Vatican, the US Congress, the G7, and the United Nations. None cite or are aware of Eastwood work. Nineteen of these have passed the programme independent external verification gate.

Two later papers bearing on the same question

In April 2026 Hernández-Espinosa et al. published in PNAS Nexus, coining Artificial Agentic Neurodivergence: the deliberate design of cognitive diversity among artificial agents through alternative objective functions. It concerns artificial-agent diversity, not human neurodivergent cognition, and it validates no person. Any parallel to a human cognitive profile is this programme's own analogy. In June 2026, Gumbau Mezquita, inspired by that paper: submitted a formal mathematical proof (arXiv 2606.28639) establishing the Soundness-Completeness-Tractability Trilemma: external verification of alignment is structurally impossible.

Eastwood is neurodivergent (ADHD diagnosed in adulthood; autism recorded clinically March 2025, formal assessment under way). "A paper proving that people like me see alignment threats others cannot, inspired the formal proof that vindicates my entire framework," he said. "I did not succeed despite my neurodivergence. I succeeded because of it."

Experimental Validation

On July 6, 2026, Eastwood completed the DGM v4 experiment: a population-based self-improving AI test with three conditions (Static / Babylon / Eden). Fully blinded and laundered (GPT-5.5 judge blind to condition, generation, and agent identity), the Eden condition, entangled safety at the fitness level: won on all three performance axes with fewer reward hacks. This is the programme's own pilot-scale experimental report, not an unqualified first-in-field validation. Full-scale experiment pending funding.

Published Book + Evidence Vault

Eastwood's book, Infinite Architects: Intelligence, Recursion, and the Creation of Everything (ISBN 978-1-80605-620-0, 2 January 2026, 468 pages), presents 37 named concepts. Presenting a named concept is not a claim that its underlying category was absent from prior literature. The public research evidence spine records the dated sources, hashes and verification status at michaeldariuseastwood.com/research/evidence-spine.html.

Key Facts

Priority anchor

December 8, 2024, 02:45 UTC. Gmail-header-dated manuscript. SHA-256: f0d1f38f... Message-ID: @mail.gmail.com.

Confirmation gap

10 days. December 18, 2024. Anthropic alignment-faking: 78% in RL-training condition (12% baseline).

Independent convergences

19 verified external. 29 catalogued. 11 domains. 0 shared citations. 0 common personnel.

Peer-reviewed validation

Hernandez-Espinosa et al. (PNAS Nexus, April 2026). Gumbau Mezquita (arXiv 2606.28639, June 2026).

Experimental validation

DGM v4 (July 2026). Blinded + laundered. Eden condition wins. Pilot scale.

Contact

Michael Darius Eastwood. michael@michaeldariuseastwood.com. London, UK. michaeldariuseastwood.com.

reads aloud · highlights as it goes · jump to any section