Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
Anthropic CEO Dario Amodei proposed embedding third-party evaluators within frontier AI companies to report safety incidents, assess AI model alignment, and share findings. This proposal, which would have been rejected by the AI industry a year ago, aims to ensure independence. A similar issue arose during the pre-release testing for OpenAI's GPT-6 Astra, where Apollo Research had only three days to test the model, hindering firm conclusions despite OpenAI touting it as its most aligned model.
Unlike previous industry resistance, Anthropic's proposal for embedded third-party evaluators directly addresses the independence concerns highlighted by issues in OpenAI's GPT-6 Astra testing.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Sep 16, 2026, 22:00 UTC
- Ingested
- Sep 16, 2026, 22:00
- Source type
- Media
Full text isn't available here.
Read at source →