Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?
Anthropic CEO Dario Amodei proposed embedding third-party evaluators within frontier AI companies to report safety incidents, assess AI model alignment, and share findings. This proposal, which would have been rejected by the AI industry a year ago, aims to ensure independence. A similar issue arose during the pre-release testing for OpenAI's GPT-6 Astra, where Apollo Research had only three days to test the model, hindering firm conclusions despite OpenAI touting it as its most aligned model.
Unlike previous industry resistance, Anthropic's proposal for embedded third-party evaluators directly addresses the independence concerns highlighted by issues in OpenAI's GPT-6 Astra testing.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年9月16日 22:00 UTC
- 收录
- 2026年9月16日 22:00
- 来源类型
- 媒体报道
本站未收录正文。
前往源站阅读 →