- 开始
- 2026年6月30日可信度 90%来自来源
Claude Sonnet 5 发布
由原文自动翻译
原标题: Claude Sonnet 5 Released
- “Claude Sonnet 5 的打造目标是成为迄今最具智能体能力的 Sonnet 模型”:它会做计划、使用浏览器和终端等工具并自主运行,表现接近 Opus 4.8,价格却更低[1][3][4]
- 从发布起,它就是免费版和 Pro 套餐的默认模型,并向 Max、Team 和 Enterprise 套餐开放,在 Claude API 上以 claude-sonnet-5 的名称提供,在 Chat、Cowork、Claude Code 和 Claude 平台上均享有更高的速率限制[1]
- 发布时定价为每百万输入 token 2 美元、每百万输出 token 10 美元,“截至 8 月 31 日”有效,此后将调整为每百万 3 美元和 15 美元;2026 年 8 月 10 日,Anthropic 将这一入门价定为永久价格[1][3][4]
- 其失调行为比 Sonnet 4.6 少,但比 Opus 4.8 和 Claude Mythos 预览版多;由于从未接受过网络安全方面的训练,它无法构建可用的 Firefox 漏洞利用程序,但仍配备了 Opus 4.7 和 4.8 所使用的网络安全防护措施[1][5]
- 作为 Sonnet 4.6 的直接升级版:自适应思考默认开启,手动开启扩展思考或使用非默认的 temperature、top_p 和 top_k 参数现在会返回错误[2]
值得关注的功能
- 规格:100 万 token 的上下文窗口,最多 128,000 个输出 token(Batch API 测试版为 300,000 个),自适应思考默认为高强度,支持文本和图像输入、文本输出,知识截止日期为 2026 年 1 月[2]
- Anthropic 的表格显示,它在 SWE-bench Pro 上得分 63.2%,在 Terminal-Bench 2.1 上得分 80.4%,在 OSWorld-Verified 上得分 81.2%,而 Sonnet 4.6 分别为 58.1%、67.0% 和 78.5%,Opus 4.8 分别为 69.2%、82.7% 和 83.4%[1][3]
- 在知识型工作(GDPval-AA v2)上得分 1618,略高于 Opus 4.8 的 1615,远高于 Sonnet 4.6 的 1395[1][3]
- 与 Opus 4.7 一样采用了新的分词器,同样的输入映射出的 token 数量大约是原来的 1.0 到 1.35 倍[1]
- 强度等级从低到最高不等,让用户可以用成本换取准确度;在最高强度下,它在某些任务上能与 Opus 4.8 匹敌[1]
基准测试[1]
| 基准测试 | Sonnet 5 | Sonnet 4.6 | Opus 4.8(供参考) |
|---|---|---|---|
| 智能体编程(SWE-bench Pro) | 63.2% | 58.1% | 69.2% |
| 智能体编程(Terminal-Bench 2.1) | 80.4% | 67.0% | 82.7% |
| 跨学科推理(Humanity's Last Exam),不使用工具 | 43.2% | 34.6% | 49.8% |
| 跨学科推理(Humanity's Last Exam),使用工具 | 57.4% | 46.8% | 57.9% |
| 计算机操作(OSWorld-Verified) | 81.2% | 78.5% | 83.4% |
| 知识型工作(GDPval-AA v2) | 1618 | 1395 | 1615 |
参考资料 5可信度 88%总体可信度: 88%该图钉的来源和参考资料对其日期的支持程度。包括来源在内的 5 份资料对图钉开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半显示所有可信度不低于 75% 的图钉
第一项始终是图钉的来源。总体可信度是各份资料对上方所用开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半。
- [1]90%anthropic.com/news/claude-sonnet-5anthropic.com· 发布于 2026年9月28日· 开始 2026年6月30日 ✓· 占评分 24%
该文章标注日期为“2026 年 6 月 30 日”,并写道:“从今天起,Claude[2] Sonnet 5 已在所有套餐中可用”;文档中的模型页面标注“发布日期:2026 年 6 月 30 日”,TechCrunch[3] 2026-06-30 的报道称它“从周二起”成为默认模型。
- [2]90%Claude Sonnet 5 - Claude Platform Docsplatform.claude.com· 添加于 2026年9月28日· 占评分 24%
Anthropic's[1][5] model page: claude-sonnet-5, a drop-in upgrade for Sonnet 4.6 with adaptive thinking on by default, 1M-token context, 128K max output (300K Batch beta), $2/$10 per million tokens, text and images in, a January 2026 knowledge cutoff, "Released June 30, 2026".
- [3]88%Anthropic launches Claude Sonnet 5 as a cheaper way to run agentstechcrunch.com· 发表于 2026年6月30日· 开始 2026年6月30日· 占评分 17%
TechCrunch, 2026-06-30: default for Free and Pro "Starting Tuesday"; $2/$10 per million tokens through August 31, then $3/$15, cheaper than Opus 4.8, GPT-5.5 and Gemini 3.1 Pro but dearer than Gemini 3.5 Flash; 63.2% on agentic coding against Opus 4.8's 69.2% and Sonnet 4.6's 58.1%.
- [4]80%Anthropic Launches Claude Sonnet 5 With Near-Opus Performance at a Lower Pricemacrumors.com· 发表于 2026年6月30日· 开始 2026年6月30日· 占评分 17%
MacRumors, 2026-06-30: "Anthropic[1][5] today introduced Claude[2] Sonnet 5", its most agentic Sonnet, with performance similar to Opus 4.8, lower hallucination and sycophancy, and the launch pricing of $2/$10 before a planned rise to $3/$15.
- [5]90%System Card: Claude Sonnet 5anthropic.com· 发表于 2026年6月30日· 占评分 17%
Anthropic's[1] system card dated "June 30, 2026", describing "the latest model in Anthropic's Sonnet family" as an upgrade to Claude[2] Sonnet 4.6, with the behavioural audit and Firefox exploit results the launch post cites.
建议更正
有遗漏或错误吗?用你自己的话说明:能佐证此图钉的链接、不同的开始或结束日期及理由,或缺失、有误的信息。AI 会对照此图钉的来源进行核实,搜索更好的来源,并添加任何支持你说法的页面。图钉自身的来源仍然最重要。AI 也会查看图片:显示的是别的东西或显示效果差的图片会被移到后面或替换。