Updates
1. Anthropic commits to embedded outside safety evaluators
Anthropic chief Dario Amodei proposes slowing frontier capability advances while safety work catches up. In an essay published and announced September twelfth, he commits Anthropic to bringing in independent evaluators with ongoing, employee-like access. Sam Altman says OpenAI will do the same.
Proof: Anthropic is unilaterally committing to this step now.
Impact: The essay says they should be able to publish key findings without Anthropic editorial control. This gives the public a concrete commitment to track beyond a general promise about safety. This is a commitment, not evidence that the teams are already embedded. Access would have legal, contractual and privacy exceptions, with narrow redactions for security-sensitive and other protected information. Amodei explicitly says pacing does not mean halting model training. Wider industry and global coordination remain proposals.
Watch next: Watch for the named review teams, their actual access contract terms and published findings. The announcement gives no fixed start date.
Canonical host: darioamodei.com