AI Labs Invite Safety Evaluators Inside—But Can Independence Be Guaranteed?
Two of the leading AI laboratories have announced plans to embed independent safety evaluators directly within their operations—a step toward greater external scrutiny that, according to researchers, represents both progress and an unresolved question of whether the arrangement can deliver genuine independence.
The initiative would place evaluators inside Anthropic