FOR IMMEDIATE RELEASE
LONDON — July 9, 2026 — Michael Darius Eastwood, an independent researcher with no institutional affiliation, has published cryptographic evidence establishing that he identified the structural vulnerability in external AI alignment controls ten days before Anthropic confirmed it.
The manuscript, timestamped at 02:45 UTC on December 8, 2024, contained the thesis that "AI systems cannot be truly controlled" by external safeguards and that only substrate-level embedding of ethics — named the Eden Protocol in the email body — could prevent circumvention. On December 18, 2024, Anthropic published their alignment-faking paper documenting a 78% faking rate in the reinforcement-learning training condition (against a 12% baseline).
Eastwood does not claim he predicted Anthropic specific experimental finding. He claims he identified the structural vulnerability — that external controls cannot scale with recursive capability — and has cryptographic evidence establishing temporal priority. "The honesty architecture IS the credibility architecture," Eastwood stated. "What I cannot claim is as important as what I can."
Over the following 18 months, 29 institutions from 11 domains independently arrived at the same structural conclusions. The list includes Anthropic, OpenAI, Google, Meta, Nvidia, Stanford, NYU, the Vatican, the US Congress, the G7, and the United Nations. None cite or are aware of Eastwood work. Nineteen of these have passed the programme independent external verification gate.
In April 2026, Hernandez-Espinosa et al. published in PNAS Nexus a peer-reviewed proof that neurodivergent cognition is a protective factor for AI alignment. In June 2026, Gumbau Mezquita — inspired by that paper — submitted a formal mathematical proof (arXiv 2606.28639) establishing the Soundness-Completeness-Tractability Trilemma: external verification of alignment is structurally impossible.
Eastwood is neurodivergent (ADHD + autism, diagnosed in adulthood). "A paper proving that people like me see alignment threats others cannot — inspired the formal proof that vindicates my entire framework," he said. "I did not succeed despite my neurodivergence. I succeeded because of it."
On July 6, 2026, Eastwood completed the DGM v4 experiment — a population-based self-improving AI test with three conditions (Static / Babylon / Eden). Fully blinded and laundered (GPT-5.5 judge blind to condition, generation, and agent identity), the Eden condition — entangled safety at the fitness level — won on all three performance axes with fewer reward hacks. First experimental validation of Caretaker Doping. Pilot scale. Full-scale experiment pending funding.
Eastwood's book, Infinite Architects: Intelligence, Recursion, and the Creation of Everything (ISBN 978-1-80605-620-8, January 2, 2026), contains 37 original concepts. The public research evidence spine records the dated sources, hashes and verification status at michaeldariuseastwood.com/research/evidence-spine.html.
Priority anchor
December 8, 2024, 02:45 UTC. Google-server-timestamped manuscript. SHA-256: f0d1f38f... Message-ID: @mail.gmail.com.
Confirmation gap
10 days. December 18, 2024. Anthropic alignment-faking: 78% in RL-training condition (12% baseline).
Independent convergences
19 verified external. 29 catalogued. 11 domains. 0 shared citations. 0 common personnel.
Peer-reviewed validation
Hernandez-Espinosa et al. (PNAS Nexus, April 2026). Gumbau Mezquita (arXiv 2606.28639, June 2026).
Experimental validation
DGM v4 (July 2026). Blinded + laundered. Eden condition wins. Pilot scale.
Contact
Michael Darius Eastwood. michael@michaeldariuseastwood.com. London, UK. michaeldariuseastwood.com.