Anthropic CEO Dario Amodei published a proposal for embedded safety evaluators.
He wants independent groups to report incidents and assess model alignment.
OpenAI CEO Sam Altman agreed to commit to this practice.
Evaluators like METR and Redwood Research would get unprecedented system access.
They could check intermediate training versions called checkpoints instead of just finished models.
This helps find how models behave during training, not just when released.



