Agents & InferenceSimon Willison

Introducing Claude Opus 5

Which summary reads better? Pick one — models revealed after.Both summaries are AI-generated.

Match the models (Optional)

Which model wrote which summary? Select a matchup mapping below before voting.

Summary A

Opus 5 tops the Artificial Analysis leaderboard—ahead of Fable 5—while priced identically to Opus 4.8, meaning you get roughly frontier-tier intelligence at half the cost of the prior top model, with the same optional 2x "fast mode." It's markedly more proactive (will improvise tooling like a homegrown CV pipeline to complete tasks), so audit your agent guardrails, and note its cyber posture: strong at finding vulnerabilities but deliberately weak at exploiting them.

LinkedIn

Two AI summaries of each story, blind-voted — see today's agents & inference digest →