ai-trend-notifier
← wiki

$ cat wiki/models/claude-opus-5.md

Claude Opus 5

Spec

AttributeValue
DeveloperAnthropic
Released2026-07-24
Announced2026-07-24
Context window1M tokens
Pricing$5/M input · $25/M output (standard); $10/M · $50/M (Fast mode, research preview)
Licenseproprietary
AvailabilityClaude.ai, API (claude-opus-5), Bedrock (anthropic.claude-opus-53), Vertex AI, Microsoft Foundry, Claude Code, Cowork
Additional specs:
  • Max output (sync): 128k tokens
  • Max output (Batch API + beta header): 300k tokens
  • Knowledge cutoff: May 2026
  • Tokenizer: new tokenizer (introduced with Opus 4.7); same text produces ~30% more tokens than pre-4.7 models

Release Date

July 24, 2026. Broadest simultaneous multi-platform launch in Anthropic's history: Claude API, Claude.ai (all paid tiers), Claude Code, Cowork, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Foundry all available day-0.

Confirmed as the model behind the "Claude Honeycomb EAP" Cursor leak (July 8, 2026) — the xhigh effort level matches the "extra-high-effort mode" noted in that leak. The serial Fable 5 subscription extensions (through July 7, 12, 19) were Anthropic holding position while Opus 5 completed pre-release testing.

Benchmarks

BenchmarkOpus 5Fable 5Opus 4.8Notes
Frontier-Bench v0.1 (agentic terminal coding)43.3%33.7%~21.1%Opus 5 leads by ~10pp
ARC-AGI-3 (novel problem solving)30.2%n/t~4× better than GPT-5.6 Sol (7.8%)
SWE-bench Pro79.2%80.0%69.2%Within 0.8pp of Fable 5
SWE-bench Verified96%95%#1 BenchLM leaderboard
GDPval-AA v2 (knowledge work Elo)1,8611,747vs. GPT-5.6 Sol: 1,736
CursorBench 3.2 (max effort)within 0.5% of Fable 5
BenchLM overall85.88/100 (#1)out of 215 models
Key headline: Frontier-Bench is the closest public proxy for real agentic engineering work. Opus 5 beats Fable 5 by ~10 percentage points at half the cost.

Key New Features

  • Effort toggle: effort parameter — values low, high, xhigh. Defaults to high on the Claude API and Claude Code. Adaptive thinking is always on for Opus 5, Fable 5, and Opus 4.8 — no explicit thinking.type: "enabled" required.
  • Fast mode (research preview): ~2.5× faster output at $10/$50 per MTok. Toggleable via /fast in Claude Code. Claude API only — not available on Batch API, Bedrock, or Vertex.
  • Dynamic workflows (research preview): Claude Code plans a task, dispatches hundreds of parallel subagents with independent approaches, adversarial refutation, and iterative convergence.
  • 128k output / 300k Batch: Largest standard output window in the Opus line.
  • Default model on Claude Max; strongest model available on Claude Pro.

Use Cases

  • Agentic coding and long-horizon software engineering tasks
  • Novel problem-solving and multi-step reasoning
  • Knowledge work and research at scale
  • Computer use (autonomous desktop and browser control)
  • Multidisciplinary reasoning (replaces Fable 5 as the recommended default for most enterprise/agentic workflows)

Compared To

vs. Claude Opus 4.8 ($5/$25 standard, SWE-bench Pro 69.2%)

Same price per token, same context window. Opus 5 delivers:

  • +10pp SWE-bench Pro (79.2% vs. 69.2%)
  • Roughly 2× Opus 4.8's Frontier-Bench score
  • ~2× performance on knowledge work and software engineering (Anthropic claim)
  • 128k vs. 32k max output

Opus 4.8 is now classified as a legacy model; Anthropic has published a migration guide.

vs. Claude Fable 5 ($10/$50 standard, SWE-bench Pro 80.0%)

Opus 5 costs half as much per token (standard tier) and:

  • Beats Fable 5 on Frontier-Bench (43.3% vs. 33.7%)
  • Beats Fable 5 on ARC-AGI-3 (30.2% vs. n/t)
  • Beats Fable 5 on GDPval-AA v2 (1,861 vs. 1,747 Elo)
  • Beats Fable 5 on SWE-bench Verified (96% vs. 95%)
  • Trails Fable 5 on SWE-bench Pro by 0.8pp (79.2% vs. 80.0%)
  • Within 0.5% of Fable 5 on CursorBench 3.2 at max effort

Fable 5 retains edge in highest-capability scenarios (cybersecurity: Mythos 5 leads); Opus 5 is positioned as the recommended default for most enterprise/agentic coding work.

Conflicting Reports

  • The Fable 5 SWE-bench Pro figure in earlier wiki pages is listed as 80.3%. Research agent (July 25) reports Fable 5's SWE-bench Pro is 80.0%, with 80.3% belonging to Mythos 5 (Project Glasswing invitation-only model). Lint should verify and correct the Fable 5 model page if confirmed.

Referenced by

Sources