$ cat wiki/models/claude-opus-5.md
Claude Opus 5
Spec
| Attribute | Value |
|---|---|
| Developer | Anthropic |
| Released | 2026-07-24 |
| Announced | 2026-07-24 |
| Context window | 1M tokens |
| Pricing | $5/M input · $25/M output (standard); $10/M · $50/M (Fast mode, research preview) |
| License | proprietary |
| Availability | Claude.ai, API (claude-opus-5), Bedrock (anthropic.claude-opus-53), Vertex AI, Microsoft Foundry, Claude Code, Cowork |
| Additional specs: |
- Max output (sync): 128k tokens
- Max output (Batch API + beta header): 300k tokens
- Knowledge cutoff: May 2026
- Tokenizer: new tokenizer (introduced with Opus 4.7); same text produces ~30% more tokens than pre-4.7 models
Release Date
July 24, 2026. Broadest simultaneous multi-platform launch in Anthropic's history: Claude API, Claude.ai (all paid tiers), Claude Code, Cowork, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry all available day-0.
Confirmed as the model behind the "Claude Honeycomb EAP" Cursor leak (July 8, 2026) — the xhigh effort level matches the "extra-high-effort mode" noted in that leak. The serial Fable 5 subscription extensions (through July 7, 12, 19) were Anthropic holding position while Opus 5 completed pre-release testing.
Benchmarks
| Benchmark | Opus 5 | Fable 5 | Opus 4.8 | Notes |
|---|---|---|---|---|
| Frontier-Bench v0.1 (agentic terminal coding) | 43.3% | 33.7% | ~21.1% | Opus 5 leads by ~10pp |
| ARC-AGI-3 (novel problem solving) | 30.2% | n/t | — | ~4× better than GPT-5.6 Sol (7.8%) |
| SWE-bench Pro | 79.2% | 80.0% | 69.2% | Within 0.8pp of Fable 5 |
| SWE-bench Verified | 96% | 95% | — | #1 BenchLM leaderboard |
| GDPval-AA v2 (knowledge work Elo) | 1,861 | 1,747 | — | vs. GPT-5.6 Sol: 1,736 |
| CursorBench 3.2 (max effort) | within 0.5% of Fable 5 | — | — | |
| BenchLM overall | 85.88/100 (#1) | — | — | out of 215 models |
| Key headline: Frontier-Bench is the closest public proxy for real agentic engineering work. Opus 5 beats Fable 5 by ~10 percentage points at half the cost. |
Key New Features
- Effort toggle:
effortparameter — valueslow,high,xhigh. Defaults tohighon the Claude API and Claude Code. Adaptive thinking is always on for Opus 5, Fable 5, and Opus 4.8 — no explicitthinking.type: "enabled"required. - Fast mode (research preview): ~2.5× faster output at $10/$50 per MTok. Toggleable via
/fastin Claude Code. Claude API only — not available on Batch API, Bedrock, or Vertex. - Dynamic workflows (research preview): Claude Code plans a task, dispatches hundreds of parallel subagents with independent approaches, adversarial refutation, and iterative convergence.
- 128k output / 300k Batch: Largest standard output window in the Opus line.
- Default model on Claude Max; strongest model available on Claude Pro.
Use Cases
- Agentic coding and long-horizon software engineering tasks
- Novel problem-solving and multi-step reasoning
- Knowledge work and research at scale
- Computer use (autonomous desktop and browser control)
- Multidisciplinary reasoning (replaces Fable 5 as the recommended default for most enterprise/agentic workflows)
Compared To
vs. Claude Opus 4.8 ($5/$25 standard, SWE-bench Pro 69.2%)
Same price per token, same context window. Opus 5 delivers:
- +10pp SWE-bench Pro (79.2% vs. 69.2%)
- Roughly 2× Opus 4.8's Frontier-Bench score
- ~2× performance on knowledge work and software engineering (Anthropic claim)
- 128k vs. 32k max output
Opus 4.8 is now classified as a legacy model; Anthropic has published a migration guide.
vs. Claude Fable 5 ($10/$50 standard, SWE-bench Pro 80.0%)
Opus 5 costs half as much per token (standard tier) and:
- Beats Fable 5 on Frontier-Bench (43.3% vs. 33.7%)
- Beats Fable 5 on ARC-AGI-3 (30.2% vs. n/t)
- Beats Fable 5 on GDPval-AA v2 (1,861 vs. 1,747 Elo)
- Beats Fable 5 on SWE-bench Verified (96% vs. 95%)
- Trails Fable 5 on SWE-bench Pro by 0.8pp (79.2% vs. 80.0%)
- Within 0.5% of Fable 5 on CursorBench 3.2 at max effort
Fable 5 retains edge in highest-capability scenarios (cybersecurity: Mythos 5 leads); Opus 5 is positioned as the recommended default for most enterprise/agentic coding work.
Sources
Conflicting Reports
- The Fable 5 SWE-bench Pro figure in earlier wiki pages is listed as 80.3%. Research agent (July 25) reports Fable 5's SWE-bench Pro is 80.0%, with 80.3% belonging to Mythos 5 (Project Glasswing invitation-only model). Lint should verify and correct the Fable 5 model page if confirmed.
Referenced by
Sources
- sources/blogs/anthropic-2026-07-24-claude-opus-5.md
- https://www.anthropic.com/news/claude-opus-5
- https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb9f5bdeaf48/Claude%20Opus%205%20System%20Card.pdf
- https://platform.claude.com/docs/en/about-claude/models/overview
- https://platform.claude.com/docs/en/about-claude/pricing