Anthropic Opus 5 Official Announcement!
Opus 5 = Approaching Fable 5 frontier intelligence while the price is half!
Aiming for the model you use every day, the new default model for Claude Max and the strongest model available in Claude Pro.
SOTA in coding and knowledge work evaluations like Frontier-Bench and GDPval-AA 👍🏻
More than 2x performance compared to Opus 4.8 at lower cost per task!
In max effort, it's within 0.5% difference of Fable 5's top score while the cost is half.
Surpassing Fable 5's top performance at just over 1/3 the cost. Also outperforms Opus 4.8 across all items in life sciences.
A structure that adjusts intelligence priority/token savings via effort settings. And they talk a ton about verification and judgment cases.
↓
Boris comment..💬
"Opus 5 is an excellent model across coding, data analysis, design, biology, and knowledge work in general.
But what's most exciting to me isn't these eval scores—it's something else.
Opus 5 is the strongest model we've ever made against prompt injection.
It's a bit buried in the system card, but across prompt injection evals and red team testing overall, Opus 5 makes it extremely difficult to succeed with injections.
When you layer defenses on top of this—strong model alignment, prompt injection detection probes, and applying Claude Code's Auto Mode together—the success rate of prompt injection attacks converges to nearly 0.
This is new, and a really exciting result!"