ModelsCompareBest forBenchmarksStatusPricingAPI
Compare / head-to-head

Claude Opus 4.7vsClaude Sonnet 4.5

Claude Opus 4.7 leads 10 of 10 shared benchmarks. Claude Sonnet 4.5 is 1.7x cheaper per token. Claude Opus 4.7 has the larger context window (1M tokens).

Benchmarks from cited public sources; pricing from official pages; status from official provider feeds.

Shared benchmarks
10 – 0
Claude Opus 4.7 leads
Cheaper per token
Claude Sonnet 4.5
1.7x cheaper, input + output
Larger context
Claude Opus 4.7
1M tokens
Provider uptime (30d)
100%
Anthropic
Claude Opus 4.7
Anthropic · released 2026-04-16
textvision
Context
1M
Max out
128K
Input /1M
$5
Output /1M
$25
Cached /1M
$0.5
Scores
18 · 14 core
Claude Sonnet 4.5
Anthropic · released 2025-09-29
textvision
Context
200K
Max out
64K
Input /1M
$3
Output /1M
$15
Cached /1M
$0.3
Scores
21 · 10 core
Quality

Benchmark matrix

BenchmarkClaude Opus 4.7Claude Sonnet 4.5ΔEdge
Reported by both · 10
SWE-bench Verified87.6% 77.2% +10.4 ptClaude Opus 4.7
GPQA Diamond94.2% 83.4% +10.8 ptClaude Opus 4.7
Humanity's Last Exam (no tools)46.9% 17.7% +29.2 ptClaude Opus 4.7
FrontierMath Tier 4 v2 (Epoch AI run)31.7% epoch2.4% epoch+29.3 ptClaude Opus 4.7
FrontierMath Tiers 1-3 v2 (Epoch AI run)70.2% epoch23.9% epoch+46.3 ptClaude Opus 4.7
Humanity's Last Exam (with tools)54.7% 33.6% +21.1 ptClaude Opus 4.7
OSWorld-Verified82.8% 61.4% +21.4 ptClaude Opus 4.7
OTIS Mock AIME 2024-2025 (Epoch AI run)97.8% epoch77.8% epoch+20 ptClaude Opus 4.7
SimpleQA Verified51.7% epoch30.7% epoch+21 ptClaude Opus 4.7
SWE-bench Verified (Epoch AI run)83.5% epoch71.3% epoch+12.2 ptClaude Opus 4.7
Only Claude Opus 4.7 reports · 7
AIME 202695.8% matharenanot reported
LMArena Elo1483.4 not reported
BrowseComp79.8% not reported
SWE-bench Multilingual80.5% not reported
SWE-bench Multimodal34.5% not reported
SWE-Bench Pro64.3% not reported
Terminal-Bench 2.166.1% not reported
Only Claude Sonnet 4.5 reports · 10
AIME 2025not reported84.2% matharena
ARC-AGI-2 (Verified)not reported13.6%
GDPval-AAnot reported1276
MCP Atlasnot reported43.8%
MMMLUnot reported89.5%
MMMU-Pro (no tools)not reported63.4%
MMMU-Pro (with tools)not reported68.9%
Tau2-bench Retailnot reported86.2%
Tau2-bench Telecomnot reported98%
Terminal-Bench 2.0not reported51%
Scores tagged "epoch" or "matharena" are independent runs, used only where the lab has not published its own; ⚠ marks rows MathArena flags as released after the competition. Higher is better on every row. Δ is Claude Opus 4.7 minus Claude Sonnet 4.5 in the benchmark's own unit. "Not reported" means the lab has not published that figure; it is not a zero. ↗ opens the source.
Specs & pricing

Side by side

SpecClaude Opus 4.7Claude Sonnet 4.5Edge
Context window1M tokens200K tokensClaude Opus 4.7
Max output128K tokens64K tokensClaude Opus 4.7
Input price / 1M$5 $3 Claude Sonnet 4.5
Output price / 1M$25 $15 Claude Sonnet 4.5
Cached input / 1M$0.5 $0.3 Claude Sonnet 4.5
Input + output / 1M
Lower is cheaper. List prices; batch, tool and regional fees excluded.
$30$18Claude Sonnet 4.5
Modalitiestext · visiontext · visionTie
Released2026-04-162025-09-29
Cited benchmark scores1821
Reliability

Anthropic status

All providers
More matchups

Claude Opus 4.7 vs …

More matchups

Claude Sonnet 4.5 vs …

Built by Respan
Which one wins on your data?

Public benchmarks are a starting point. Run Claude Opus 4.7 and Claude Sonnet 4.5 on your own prompts with Respan evals, or route to either through one gateway key with automatic failover.