Anthropic releases Opus 5.5 with lower prices and Fable-level performance
Anthropic releases Opus 5.5 with lower prices and Fable-level performance
Anthropic 发布 Opus 5.5:价格更低,性能媲美 Fable
Anthropic’s newest model, Opus 5.5, was released on Tuesday, setting a new state-of-the-art in coding and knowledge work performance, according to the company. Opus is the most capable and expensive tier in Anthropic’s three-tier Claude lineup; Sonnet sits in the middle, while Haiku is the fastest and cheapest.
Anthropic 于周二发布了其最新模型 Opus 5.5。据该公司称,该模型在编程和知识工作表现方面树立了新的行业标杆。Opus 是 Anthropic 三级 Claude 产品线中能力最强、价格最高的一档;Sonnet 居于中间,而 Haiku 则是速度最快、价格最低的一档。
Notably, Anthropic says, the release outpaces the larger Fable model in many benchmarks and succeeded in a number of informal tasks that Fable failed to complete. The new model is also significantly cheaper than its predecessor. Output tokens will be charged at $20 per million tokens for Opus 5.5, compared to $25 for the previous model. Other metrics have similar price drops. The model is also faster to run, reflecting an overall drop in the compute required to serve it.
值得注意的是,Anthropic 表示,该版本在多项基准测试中超越了规模更大的 Fable 模型,并成功完成了 Fable 未能完成的一些非正式任务。新模型的价格也比前代产品大幅降低。Opus 5.5 的输出 Token 价格为每百万 Token 20 美元,而上一代模型为 25 美元。其他指标的价格也有类似的下调。此外,该模型的运行速度更快,反映出其所需的计算资源整体有所减少。
The new version also makes significant changes to how Opus communicates, with the Opus 5.5 less likely to use jargon and more likely to put important information at the start of its messages. The launch comes just two months after the release of Opus 5 on July 24. According to the announcement, Sonnet 5.5 and Haiku 5.5, which are the next tiers in the lineup, will be released “in the coming weeks,” with similar performance improvements.
新版本还对 Opus 的沟通方式进行了重大改进:Opus 5.5 更少使用专业术语,并倾向于将重要信息放在回复的开头。此次发布距离 7 月 24 日 Opus 5 的推出仅过去两个月。根据公告,产品线中的后续梯队 Sonnet 5.5 和 Haiku 5.5 将在“未来几周内”发布,并带来类似的性能提升。
Anthropic says that Opus 5.5 is comparable to Mythos in its biology and cybersecurity capabilities, so its release is subject to the same safeguards as the company’s Fable model. Those safeguards limit how much the models can be used to discover exploits in compiled programs or developing recognizable biological weapons, among other tasks.
Anthropic 表示,Opus 5.5 在生物学和网络安全能力方面与 Mythos 相当,因此其发布受到与该公司 Fable 模型相同的安全防护措施限制。这些防护措施限制了模型在发现已编译程序漏洞或开发可识别生物武器等任务中的使用程度。
Opus 5.5 is Anthropic’s first model release since CEO Dario Amodei embraced calls to pace the frontier, deliberately slowing down progress on AI capabilities to match the rate of progress on alignment. “I have become convinced that fully addressing the risks requires even more prudence,” Amodei wrote in a post earlier this month, “not just investing in risk prevention, but pacing the rate of capabilities advancement so that risk prevention has time to keep up.”
Opus 5.5 是 Anthropic 自首席执行官 Dario Amodei 响应“放缓前沿研究”呼吁以来的首个模型发布,旨在刻意减缓 AI 能力的进步速度,以匹配对齐技术(alignment)的进展。Amodei 在本月初的一篇文章中写道:“我深信,要全面应对风险,需要更加审慎。不仅要投资于风险预防,还要控制能力提升的速度,以便让风险预防有时间跟上。”
Opus 5.5’s safety training was broadly similar to its predecessors, with alignment testing and pre-release evaluation by outside organizations like METR and Frontier Design. But Anthropic emphasized that more advanced training and evaluation systems were already being prepared for future models, including improved security and monitoring systems.
Opus 5.5 的安全训练与前代产品大致相似,包括对齐测试以及由 METR 和 Frontier Design 等外部机构进行的发布前评估。但 Anthropic 强调,针对未来模型,更先进的训练和评估系统已经在筹备中,包括改进的安全和监控系统。
“As AI becomes more capable, public policy should play a larger role in making sure the systems people rely on are safe. That capacity takes time to build, and we’ve started to put the infrastructure in place to support it,” the blog post reads. “We expect to share more details on these efforts soon.”
博客文章写道:“随着 AI 能力的增强,公共政策应在确保人们所依赖的系统安全方面发挥更大作用。这种能力建设需要时间,我们已经开始部署相关基础设施来提供支持。我们期待很快分享关于这些工作的更多细节。”