AI safety · 19 Aug 2026
Anthropic raises its own misalignment risk rating and keeps a stronger model unreleased
The company says the change reflects rising uncertainty about its evaluations, not new evidence that either its released or unreleased systems are more dangerous.