Anthropic and OpenAI Propose Embedding Independent Safety Evaluators Inside AI Companies
Anthropic CEO Dario Amodei proposed in a weekend essay that frontier AI companies embed third-party evaluators with the power to assess model alignment, report safety incidents, and publish findings without…