An
claude-opus-4-7 Benchmark & Insights
Anthropic Claude API
Updated Jul 31, 2026 All models
Sample size
169 runs
in window
Accuracy
92.1%
consensus match · 202d
Confidence
86%
over 203 runs
Window end
Jul 31, 2026
most recent run
Input price
$5.00/MTok
prompt tokens
Output price
$25.00/MTok
completion tokens
Model insights
- 01 The weakest of the three Opus generations at an identical price, and the most one-sided — fourteen downward misses against one upward.
- 02 It repeats the family failure of reading a "high" day as "medium" and "safe", twice in July.
Recent forecasts