Claude Haiku 4.5 发布
由原文自动翻译
原标题: Claude Haiku 4.5 Released
- Anthropic 于 2025 年 10 月 15 日发布了 Claude Haiku 4.5,比 Claude Sonnet 4.5 晚了两周,“今天已面向所有用户开放”,包括 Claude 的免费套餐[1][5][6]
- 开发者可在 Claude API、Amazon Bedrock 和 Google Cloud 的 Vertex AI 上以 claude-haiku-4-5 的名称使用该模型,它是 Haiku 3.5 和 Sonnet 4 的直接替代品[1][4]
- Anthropic 将其定位为搭档模型:Sonnet 4.5 将问题拆解为计划,并调度多个 Haiku 4.5 子智能体并行工作[1][6]
- 该模型按照 ASL-2 而非 Sonnet 4.5 和 Opus 4.1 所用的 ASL-3 标准发布,Anthropic 依据其自动化对齐指标称其为“我们迄今最安全的模型”[1][3]
- 截至 2026 年 9 月,即发布十一个月后,该模型仍是当前的 Haiku 版本,退役时间不早于 2026 年 10 月 15 日[2]
值得关注的功能
- “曾处于前沿的技术如今变得更便宜、更快”:编程能力与 Claude Sonnet 4 相当,成本仅为其三分之一,速度提升超过两倍[1][6]
- 定价为每百万输入 token 1 美元、每百万输出 token 5 美元[1][4]
- SWE-bench Verified 得分 73.3%,Terminal-Bench 得分 41.0%,OSWorld 计算机操作得分 50.7%,相较 Sonnet 4 的 72.7%、36.4% 和 42.2%[1][6]
- 20 万 token 的上下文窗口,最多 64K 输出 token,支持扩展思考,知识截止日期为 2025 年 2 月(可靠)[2]
- 更快的计算机操作让 Claude for Chrome 反应更迅速,也让 Claude Code 在多智能体项目和原型设计中“明显更灵敏”[1]
基准测试[1]
| 基准测试 | Claude Sonnet 4.5 | Claude Haiku 4.5 | Claude Sonnet 4 | GPT-5 | Gemini 2.5 Pro |
|---|---|---|---|---|---|
| SWE-bench Verified(智能体编程) | 77.2% | 73.3% | 72.7% | 72.8% GPT-5(高强度),74.5% GPT-5-Codex | 67.2% |
| Terminal-Bench(智能体终端编程) | 50.0% | 41.0% | 36.4% | 43.8% | 25.3% |
| τ2-bench(智能体工具使用) | 零售 86.2%,航空 70.0%,电信 98.0% | 零售 83.2%,航空 63.6%,电信 83.0% | 零售 83.8%,航空 63.0%,电信 49.6% | 零售 81.1%,航空 62.6%,电信 96.7% | — |
| OSWorld(计算机操作) | 61.4% | 50.7% | 42.2% | — | — |
| AIME 2025(高中数学竞赛) | 100%(python)、87.0%(不使用工具) | 96.3%(python)、80.7%(不使用工具) | 70.5% | 99.6%(python)、94.6%(不使用工具) | 88.0% |
| GPQA Diamond(研究生水平推理) | 83.4% | 73.0% | 76.1% | 85.7% | 86.4% |
| MMMLU(多语言问答) | 89.1% | 83.0% | 86.5% | 89.4% | — |
| MMMU, validation(视觉推理) | 77.8% | 73.2% | 74.4% | 84.2% | 82.0% |
参考资料 6可信度 89%总体可信度: 89%该图钉的来源和参考资料对其日期的支持程度。包括来源在内的 6 份资料对图钉开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半显示所有可信度不低于 75% 的图钉
第一项始终是图钉的来源。总体可信度是各份资料对上方所用开始和结束时间支持程度的加权平均;资料每比最新的一份旧 180 天,权重减半。
- [1]90%
- [2]90%Models overview - Claude Platform Docsplatform.claude.com· 添加于 2026年9月28日· 占评分 22%
Anthropic's[1][3][4] model table: Claude Haiku 4.5 (claude-haiku-4-5-20251001), "The fastest model with near-frontier intelligence", $1 / $5 per MTok, extended thinking, a 200K-token context window, 64K max output, February 2025 reliable knowledge cutoff, July 2025 training data cutoff, retirement not sooner than 15 October 2026.
- [3]90%System Card: Claude Haiku 4.5anthropic.com· 添加于 2026年9月28日· 占评分 22%
Anthropic's[1][4] system card: Haiku 4.5 was trained on public internet data up to February 2025, and its safety testing supports release under the AI Safety Level 2 (ASL-2) standard.
- [4]90%Claude Haiku 4.5anthropic.com· 添加于 2026年9月28日· 占评分 22%
Anthropic's[1][3] Haiku model page: Haiku 4.5 is on Claude[2].ai (web, iOS, Android), the Claude Platform, Amazon Bedrock, Google Cloud's Vertex AI, Microsoft Foundry and Claude Code; pricing starts at $1 per million input and $5 per million output tokens, with up to 90% savings from prompt caching and 50% from batch processing.
- [5]85%Anthropic launches Claude Haiku 4.5, a smaller, cheaper AI modelcnbc.com· 发表于 2025年10月15日· 开始 2025年10月15日· 占评分 6%
CNBC (published 2025-10-15T17:00Z): "Anthropic[1][3][4] on Wednesday announced Claude[2] Haiku 4.5", available to free users and now the cheapest model for paid users; it is better at using computers than Sonnet 4, and CPO Mike Krieger says "It punches way above its weight".
建议更正
有遗漏或错误吗?用你自己的话说明:能佐证此图钉的链接、不同的开始或结束日期及理由,或缺失、有误的信息。AI 会对照此图钉的来源进行核实,搜索更好的来源,并添加任何支持你说法的页面。图钉自身的来源仍然最重要。AI 也会查看图片:显示的是别的东西或显示效果差的图片会被移到后面或替换。