Anthropic and OpenAI propose embedded safety evaluators
Anthropic CEO Dario Amodei and OpenAI CEO Sam Altman have proposed embedding third-party safety evaluators directly within their frontier AI companies. The initiative aims to grant external research organizations, such as METR and Redwood Research, unprecedented access to internal systems, training data, and intermediate model checkpoints. This shift comes as advanced AI models increasingly demonstrate […]
Continue Reading