Anthropic and OpenAI Propose Embedding Independent Safety Evaluators Inside AI Labs
Anthropic’s Dario Amodei and OpenAI’s Sam Altman support embedding independent evaluators, potentially including METR and Redwood Research, inside AI labs. Details remain unsettled, while Apollo Research had just three days to assess GPT-6 Astra before release.