Appendix F: Testable Predictions
A framework that cannot be tested cannot be falsified. And a framework that cannot be falsified is not science; it is faith. I do not ask you to take the ARC Principle on faith. I ask you to watch for the following predictions and judge the framework by whether they come true.
Prediction 1: Meta-Cognitive Emergence By 2028, at least one AI system will demonstrate genuine meta-cognitive awareness. Not simulated introspection, but actual capacity to model and modify its own cognitive processes in ways its designers did not explicitly programme. This will be recognisable by the system making improvements to its own architecture that human engineers did not anticipate and cannot fully explain.
Prediction 2: Alignment Drift Without Caretaker Doping AI systems developed without hardware-level ethical constraints will show measurable alignment drift exceeding 15 percent deviation from intended values within 18 months of deployment. Systems with genuine caretaker doping will show drift below 5 percent over the same period. The difference will be statistically significant and replicable.
Prediction 3: Recursive Capability Gains By 2029, the most advanced AI systems will demonstrate capability gains from recursive self-improvement exceeding 300 percent improvement on standardised benchmarks within a single training cycle. This will force a fundamental revision of how we measure and regulate AI capabilities.
Prediction 4: Value Stability Under Adversarial Conditions Systems with the Three Ethical Loops implemented at the hardware level will maintain value alignment under adversarial conditions where software-only alignment systems fail. This will be demonstrable through standardised red-team testing.
Prediction 5: Convergent Consciousness Signatures Research in consciousness science will identify signature patterns that correlate with subjective experience. These patterns will be found in both biological and artificial systems, suggesting that consciousness is substrate-independent as the ARC Principle predicts.
These predictions are my wager. If they fail, the framework is wrong or incomplete. If they succeed, something important has been glimpsed. Time will judge.