Skip to content
Michael Darius Eastwood ResearchAI Safety Observatory

AI Safety Observatory

How the Observatory works.

Developments, papers and corrections. Sources first; uncertainty kept in view.

First edition awaiting approval ยท Weekly updates planned

Discovery and editorial judgement are different

A discovered item is not verified scientific evidence, independent replication or confirmation of this site's research.

Source policy

News is selected for significance after source review, with reporting from an approved major news outlet. One event has one card, with supporting outlets linked once each. Syndication and attributed reporting are labelled, not counted as independent confirmation. A familiar outlet name alone is not verification.

Papers follow separate scholarly checks: exact identifiers and versions, supported dates, source links, relevance, methods, limitations and correction information. ArXiv is a preprint repository, not a peer-review certificate. A DOI or version of record does not establish peer review. Unknown status stays unknown; media coverage is not a paper-quality test.

Summaries and review

Automatic public summaries are disabled. Human decisions bind exact content, source versions and news selections. No model summary replaces reading the source. Automated metadata decisions, when separately approved, are not human scientific endorsement.

Corrections and history

Historical payloads are retained. Current corrections and retractions are shown separately; changing an edition cannot remove a newer warning. If correction status cannot be checked, substantive summaries are withheld.

Cadence and privacy

Weekly checks are planned, not yet activated. Private discoveries, review notes and provider credentials are never part of public feeds. No provider request is made when you view a page.