FOR IMMEDIATE RELEASE
LONDON: July 9, 2026: Michael Darius Eastwood, an independent researcher with no institutional affiliation, has published cryptographic evidence establishing that he identified the structural vulnerability in external AI alignment controls ten days before Anthropic confirmed it.
The record Eastwood points journalists to first is not the priority claim but a correction: the headline exponent in his early work failed replication against his own printed book, the robust estimate of roughly 0.49 replaced it, and the retraction is published and standing. The ten-day chronology below should be read in that light: stated narrowly, by a researcher whose record shows he corrects against himself.
The manuscript, timestamped at 02:45 UTC on December 8, 2024, contained the thesis that "AI systems cannot be truly controlled" by external safeguards and that only substrate-level embedding of ethics, named the Eden Protocol in the email body: could prevent circumvention. On December 18, 2024, Anthropic published their alignment-faking paper documenting a 78% faking rate in the reinforcement-learning training condition (against a 12% baseline).
Eastwood does not claim he predicted Anthropic specific experimental finding. He claims he identified the structural vulnerability, that external controls cannot scale with recursive capability, and has cryptographic evidence establishing temporal priority. "The honesty architecture IS the credibility architecture," Eastwood stated. "What I cannot claim is as important as what I can."
Over the following 18 months, a typed convergence register of dated events, classified by evidence class from 11 domains independently arrived at the same structural conclusions. The list includes Anthropic, OpenAI, Google, Meta, Nvidia, Stanford, NYU, the Vatican, the US Congress, the G7, and the United Nations. None cite or are aware of Eastwood work. Nineteen of these have passed the programme independent external verification gate.
In April 2026 Hernández-Espinosa et al. published in PNAS Nexus, coining Artificial Agentic Neurodivergence: the deliberate design of cognitive diversity among artificial agents through alternative objective functions. It concerns artificial-agent diversity, not human neurodivergent cognition, and it validates no person. Any parallel to a human cognitive profile is this programme's own analogy. In June 2026, Gumbau Mezquita, inspired by that paper: submitted a formal mathematical proof (arXiv 2606.28639) establishing the Soundness-Completeness-Tractability Trilemma: external verification of alignment is structurally impossible.
Eastwood is neurodivergent (ADHD diagnosed in adulthood; autism recorded clinically March 2025, formal assessment under way). "A paper proving that people like me see alignment threats others cannot, inspired the formal proof that vindicates my entire framework," he said. "I did not succeed despite my neurodivergence. I succeeded because of it."
On July 6, 2026, Eastwood completed the DGM v4 experiment: a population-based self-improving AI test with three conditions (Static / Babylon / Eden). Fully blinded and laundered (GPT-5.5 judge blind to condition, generation, and agent identity), the Eden condition, entangled safety at the fitness level: won on all three performance axes with fewer reward hacks. This is the programme's own pilot-scale experimental report, not an unqualified first-in-field validation. Full-scale experiment pending funding.
Eastwood's book, Infinite Architects: Intelligence, Recursion, and the Creation of Everything (ISBN 978-1-80605-620-0, 2 January 2026, 468 pages), presents 37 named concepts. Presenting a named concept is not a claim that its underlying category was absent from prior literature. The public research evidence spine records the dated sources, hashes and verification status at michaeldariuseastwood.com/research/evidence-spine.html.
Priority anchor
December 8, 2024, 02:45 UTC. Gmail-header-dated manuscript. SHA-256: f0d1f38f... Message-ID: @mail.gmail.com.
Confirmation gap
10 days. December 18, 2024. Anthropic alignment-faking: 78% in RL-training condition (12% baseline).
Independent convergences
19 verified external. 29 catalogued. 11 domains. 0 shared citations. 0 common personnel.
Peer-reviewed validation
Hernandez-Espinosa et al. (PNAS Nexus, April 2026). Gumbau Mezquita (arXiv 2606.28639, June 2026).
Experimental validation
DGM v4 (July 2026). Blinded + laundered. Eden condition wins. Pilot scale.
Contact
Michael Darius Eastwood. michael@michaeldariuseastwood.com. London, UK. michaeldariuseastwood.com.