Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Anthropic and OpenAI proposed embedding third-party safety evaluators like METR and Redwood Research inside frontier AI companies to assess alignment, report incidents, and inspect training checkpoints. While evaluators welcomed the access proposal, they noted that formal details and legislative backing are necessary to guarantee true independence.

Cover image for Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?