- Start
- Aug 5, 202590% CONFIDENCEfrom the source
Claude Opus 4.1 Released
- Anthropic released Claude Opus 4.1 on 5 August 2025 to paid Claude users, in Claude Code, and on its API, Amazon Bedrock and Google Cloud's Vertex AI, at the same price as Opus 4[1][2]
- It came out the same day OpenAI released its first open-weight reasoning models since 2019; Anthropic said it planned "substantially larger improvements to our models in the coming weeks"[1][6]
- GitHub put it into Copilot the same day as a public preview for Enterprise and Pro+ plans[5]
- Like Opus 4 it is deployed under the ASL-3 safety standard[3]
- Retired from the Claude API on 5 August 2026, exactly a year after launch, replaced by Claude Opus 4.8[4]
Notable features
- An upgrade to Claude Opus 4 "on agentic tasks, real-world coding, and reasoning", with better in-depth research and data analysis, "especially around detail tracking and agentic search"[1]
- 74.5% on SWE-bench Verified against Opus 4's 72.5%, and 43.3% on Terminal-Bench against 39.2%[1][6]
- Other scores in Anthropic's table: 80.9% on GPQA Diamond, 82.4% on TAU-bench retail, 89.5% on MMMLU and 78.0% on AIME 2025[1]
- Customer reports: GitHub sees notable gains in multi-file refactoring, Rakuten precise fixes in large codebases "without making unnecessary adjustments", and Windsurf a one-standard-deviation jump over Opus 4[1]
- Priced like Opus 4 at $15 per million input and $75 per million output tokens; API id claude-opus-4-1-20250805[1][6]
Benchmarks[1]
| Benchmark | Claude Opus 4.1 | Claude Opus 4 | Claude Sonnet 4 | OpenAI o3 | Gemini 2.5 Pro |
|---|---|---|---|---|---|
| SWE-bench Verified¹ (agentic coding) | 74.5% | 72.5% | 72.7% | 69.1% | 67.2% |
| Terminal-Bench² (agentic terminal coding) | 43.3% | 39.2% | 35.5% | 30.2% | 25.3% |
| GPQA Diamond (graduate-level reasoning) | 80.9% | 79.6% | 75.4% | 83.3% | 86.4% |
| TAU-bench (agentic tool use) | Retail 82.4%, Airline 56.0% | Retail 81.4%, Airline 59.6% | Retail 80.5%, Airline 60.0% | Retail 70.4%, Airline 52.0% | — |
| MMMLU³ (multilingual Q&A) | 89.5% | 88.8% | 86.5% | 88.8% | — |
| MMMU, validation (visual reasoning) | 77.1% | 76.5% | 74.4% | 82.9% | 82% |
| AIME 2025⁴ (high school math competition) | 78.0% | 75.5% | 70.5% | 88.9% | 88% |
¹ Opus 4.1, Opus 4 and Sonnet 4 run pass@1 with bash/editor tools, averaged over 10 trials, single-attempt patches, no test-time compute. ² Default agent framework (Terminus 1), averaged over 5 trials. ³ Claude scores are the average over 14 non-English languages. ⁴ Run with nucleus sampling, top_p 0.95.
References 689% CONFIDENCEOverall confidence: 89%How well the pin's source and references back up its dates.Weighted average of how firmly 6 references, the source included, support the pin's start and end times; a reference counts half as much for every 180 days older than the newestShow all pins at 75% confidence or better
The first entry is always the pin's source. Overall confidence is a weighted average of how firmly each reference supports the start and end times used above; a reference counts half as much for every 180 days older than the newest.
- [1]90%anthropic.com/news/claude-opus-4-1anthropic.com· Posted Sep 28, 2026· Starts Aug 5, 2025 ✓· 23% of score
Anthropic's[3] post is dated "Aug 5, 2025": "Today we're releasing Claude[2][4] Opus 4.1 ... now available to paid Claude users and in Claude Code. It's also on our API, Amazon Bedrock, and Google Cloud's Vertex AI"; the Claude release notes and GitHub's[5] changelog carry the same date.
- [2]90%Claude Platform release notes - Claude Platform Docsplatform.claude.com· Added Sep 28, 2026· 23% of score
"August 5, 2025 We've launched Claude[4] Opus 4.1, an incremental update to Claude Opus 4", which does not allow both temperature and top_p to be set; later entries record structured outputs launching for Opus 4.1 (14 November 2025) and its retirement.
- [3]90%System Card Addendum: Claude Opus 4.1anthropic.com· Added Sep 28, 2026· 23% of score
Anthropic's[1] system card addendum: "Like Claude[2][4] Opus 4, Claude Opus 4.1 is deployed under the AI Safety Level 3 (ASL-3) Standard"; it is not "notably more capable" than Opus 4 under the RSP, so new evaluations were voluntary.
- [4]90%Model deprecations - Claude Platform Docsplatform.claude.com· Added Sep 28, 2026· 23% of score
Anthropic[1][3] notified developers on 5 June 2026 that Claude[2] Opus 4.1 (claude-opus-4-1-20250805) would be retired; it was retired on 5 August 2026, one year after release, with claude-opus-4-8 as the replacement.
- [5]85%Anthropic Claude Opus 4.1 is now in public preview in GitHub Copilotgithub.blog· Published Aug 5, 2025· Starts Aug 5, 2025· 5% of score
GitHub's changelog, "Release August 5, 2025": Claude[2][4] Opus 4.1, "the successor to Claude Opus 4", is available in GitHub Copilot Chat for Copilot Enterprise and Pro+ plans on github.com, Visual Studio Code and GitHub Mobile.
Suggest a correction
Something missing or wrong? Say it in your own words: a link that backs this pin up, a different start or end date and why, or a fact it lacks or gets wrong. The AI checks it against this pin's sources, searches for better ones, and adds any page that backs you up. The pin's own sources still count most. A picture that shows something else, or shows it badly, is looked at too, and moved down or replaced.