新闻
Introducing Claude Opus 5
Anthropic · 发布于 · 阅读约3分钟
30秒读懂
- 发生了什么
- Anthropic released Claude Opus 5, a new mid-tier model matching Fable 5 performance at half the cost for coding and knowledge work.
- 为何重要
- Engineers building software, analyzing data, or automating workflows should evaluate it as a potential default model for daily development tasks.
- 注意
- Opus 5 remains behind Mythos 5 on cybersecurity tasks, and performance varies by effort setting; benchmark results may not reflect your specific use case.
收听本摘要
文章节选
Product Announcements
Introducing Claude Opus 5
Jul 24, 2026
Claude Opus 5 is available today. It’s a thoughtful and proactive model that comes close to the frontier intelligence of Claude Fable 5 at half the price.
On coding and knowledge work evaluations like Frontier-Bench and GDPval-AA , Opus 5 is the new state-of-the-art, though it remains behind Mythos 5 on cybersecurity tasks.
Opus 5 is designed to be used every day: it works more efficiently than other models. It’s the new default model on Claude Max, and the strongest model on Claude Pro.
Performance and cost-effectiveness
Claude Opus 5 provides greatly improved performance for the same cost as its predecessor, Opus 4.8. The charts in this section show how performance changes according to the model’s effort setting, which customers can use to optimize for intelligence or conserve tokens for faster and cheaper results.
Opus 5 excels on valuable software engineering tasks. For example, on Frontier-Bench v0.1, Opus 5 surpasses all other models, and more than doubles Opus 4.8’s performance at a lower cost per task. On CursorBench 3.2 , at max effort, the model performs within 0.5% of Fable 5’s peak score, but at half the cost per task; it also achieves greater performance at a given cost than all other models on high, xhigh, and max effort.
Frontier-Bench v0.1 CursorBench AA Coding Agent Index
We see similar resu
节选自原文。请前往来源阅读全文。
- claude
The Agent Architect
每周一个模式、一个权衡、一个生产事故案例。为构建智能体系统的人准备的每周简报。
每周一封邮件,一键退订。您的地址仅用于发送简报。