Anthropic's embedded evaluators publish no incident report in year one
Why I called it
Amodei's essay names a concrete commitment: evaluators with ongoing access comparable to employees, checking safety practices, reporting incidents, and examining training before a model ships. That is a stronger claim than the enrollment-number gap Prediction 2026-09-06-T1 tracks, because it promises an outside voice with something to publish. Anthropic's own week supplies the stakes: a fourth cybersecurity incident, a declined UK safety review, and a resigned researcher all landed before the essay did, so whether an embedded evaluator ever surfaces something publicly is the difference between a governance structure and a talking point.
The call, in full. Within twelve months of Anthropic's September 12th 2026 commitment to give outside evaluators ongoing, employee-comparable access, no evaluator operating under that program has publicly published an incident report, safety-practice finding, or training-process review naming Anthropic.
Scoring criterion. RESOLVES WRONG if, by September 12th 2027, an evaluator operating under Anthropic's embedded-access program publishes, or a credible outlet reports the evaluator authored, a specific incident report, safety-practice finding, or training-process review concerning Anthropic. RESOLVES CORRECT otherwise.
The criterion is the machine-checkable version: a prediction that cannot be settled by a third party against a public source fails the build before it reaches this page.