Signal Brief

Opus 5 发布:价格减半性能领先

Anthropic 发布 Claude Opus 5,定价 $5/$25 每百万 token(与 Opus 4.8 相同),仅为 Fable 5 的一半,但在绝大多数基准测试中击败了 Fable 5。它支持五种努力设置,默认启用推理,并在编码、操作计算机等经济价值高的任务上表现更强,具有自主验证和恢...

twitter关注列表 Rohan Paul (@rohanpaul_ai) 发布 2026-07-24 收录 2026-07-24 值得跟进

一句话判断

与 Fable 5 相比,Opus 5 以半价提供接近的智能,并引入五级努力设置和默认推理,显著提升自主任务完成能力,值得关注其安全分类器放宽和迁移陷阱。

核心信息

Anthropic 发布 Claude Opus 5,定价 $5/$25 每百万 token(与 Opus 4.8 相同),仅为 Fable 5 的一半,但在绝大多数基准测试中击败了 Fable 5。它支持五种努力设置,默认启用推理,并在编码、操作计算机等经济价值高的任务上表现更强,具有自主验证和恢复能力。

原始内容

Anthropic launched Claude Opus 5 at half the price of its most powerful model Fable 5, while at the same time beating Feble 5 in almost all benchmarks. Costs the same as Opus 4.8, $5/$25 per mn input/ output tokens, but exactly half as much as Fable 5. Frankly I am trying to figure out at this price of Opus 5 and with this great benchmark result why would I ever need to use Fable 5. Opus 5 also has less restrictive cyber classifiers than Fable, automatic fallback options, mid-conversation tool changes and no special data-retention requirement for general access. The cheaper model is winning specifically on economically valuable activities—coding, operating computers, research and completing business processes—not merely trivia or academic question answering. Anthropic repeatedly describes Opus 5 as more willing to verify, recover and continue until a task is actually complete. Examples include: - Creating its own computer-vision pipeline to reconstruct a 3D machine component from raw pixels - Finding the root cause of an open-source bug and fixing an edge case missed by the existing community patch - Building a market-data integration and then creating its own test harness when no live validation feed was available - Completing Zapier’s full churn-prevention workflow when previous models failed These are signs of adaptive problem-solving: the model notices that a required capability or validation mechanism is missing and constructs one instead of stopping. Anthropic also describes Opus 5 as more thorough about validating its work and recovering when necessary. The launch gives examples in which it constructed its own vision pipeline, identified the underlying cause of a software bug and built a test harness when no live data feed existed. That is potentially more commercially meaningful than a modest benchmark gain. An agent that finishes 70% of jobs without intervention can be dramatically more valuable than one that produces slightly better individual responses but frequently stalls. GPT-5.6 Sol leads DeepSWE: 72.7% versus Opus 5’s 68.8% ![photo](https://pbs.twimg.com/media/HOA03Etb0AAThHz.jpg) > **引用原帖 Claude (@claudeai):** > Introducing Claude Opus 5. > It's a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price. https://t.co/GQWhcq2CQL > https://x.com/claudeai/status/2080699495453528290 Rohan Paul (@rohanpaul_ai): Opus 5 now is effectively a family of 5 models behind one endpoint: supports five effort settings: low, medium, high, xhigh and max. Reasoning is enabled by default, with the effort setting controlling how deeply it thinks. Anthropic says the lower levels retain strong quality while reducing tokens and latency, whereas max is intended for the deepest reasoning and extended agent work. That can simplify routing systems. Developers may no longer need to constantly switch between a cheap model and a premium reasoning model. Rohan Paul (@rohanpaul_ai): There is one migration trap: thinking is now on by default, and disabling it while selecting xhigh or max returns an API error. Opus 5 also tends to produce longer deliverables, narrate agent progress more frequently and delegate to subagents more readily than Opus 4.8.

相关动态

01

Cursor宣布Claude Opus 5上线

Cursor平台上线Claude Opus 5模型,在CursorBench基准测试中得分66.7,与Fable 5的66.5分持平,但价格仅为后者一半(输入5美元、输出25美元每百万token),且支持零数据保留,而Fable 5仍保留数据30天。

twitter关注列表2026-07-24#模型发布#产品发布#AI
观察
05

Opus 5 提示注入防御取得突破

Claude 团队称 Opus 5 在编码、数据分析、设计、生物学、知识工作等评测中表现优异,同时强调该模型是迄今为止最不易被提示注入的模型,在结合模型对齐、提示注入探针和 Claude Code 中的 Auto Mode 后,提示注入攻击成功率降至约 0%。

twitter关注列表2026-07-24#AI安全#模型发布#技术更新
观察