An

claude-opus-4-7 Benchmark & Insights

Anthropic Claude API
Updated Jul 18, 2026 All models
Sample size
156 runs
in window
Accuracy
88.8%
consensus match · 160d
Confidence
86%
over 161 runs
Window end
Jul 18, 2026
most recent run
Input price
$5.00/MTok
prompt tokens
Output price
$25.00/MTok
completion tokens
Model insights
  • 01 The weakest Opus: same $5/$25 price as opus-4-8 but 4.5 points less accurate, with a heavy "optimistic" tilt (17 under vs 1 over) and false "safe" calls on 2026-02-18, 02-25 and 07-05.
  • 02 Its misses cluster in the Feb-Mar stretch where it kept answering "low" against a "medium" consensus.
Recent forecasts