Agents & InferenceTechCrunch

Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

OpenAI and Anthropic are committing to embedded third-party safety evaluators, potentially giving groups like METR and Redwood access beyond final model testing into checkpoints, training logs, evaluation transcripts, and post-training environments. For teams shipping frontier or agentic systems, this shifts safety work toward auditable training-time evidence, not just pre-release eval scores; expect pressure to preserve logs, expose intermediate model behavior, and prove models didn’t learn to game evaluations.