
It comes in at half the cost of Fable 5, and the same price as Opus 4.8 at $5/million inputs and $25/million outputs. Anthropic also says it is more cost efficient, using fewer tokens per task and achieving more with less effort.
Especially on ARC-AGI-3, which tests models with never-before seen puzzles, Opus 5 scores 30.2%, well above the previous record by GPT-5.6 Sol at 7.8%. Similarly, it gets 43.3% on Frontier-Bench to Opus 4.8’s 21.1% and the list goes on.
On cybersecurity, it falls far behind Mythos 5 on offensive ability, being much less able to exploit security vulnerabilities, but scores about equally on finding these bugs. That could make it useful for cyber research, and less so for writing exploits.
On general security, it is less likely to get tricked into offering dual-use responses, and it is the most aligned model from Anthropic ever, they say.
As of today, the model becomes the default model on Anthropic’s Max plan, and is the strongest model available on Claude Pro.
Read more: Anthropic’s announcement, launch post on X. More on TechCrunch, Axios, and CNBC. Discussion on r/Singularity and Hacker News.













