Leaderboard / anthropic/claude-opus-4-7
claude-opus-4-7
anthropic · Tasks passed 93.2% · Finance Index 92.8
Tasks passed
93.2%
Overall accuracy
91.8%
Finance Index
92.8
Consistency
2.8/3 avg
Est. eval cost
$1.14
Median latency
—
Median output speed
—
- Harness version
- 0.1.0
- Task set
- v2
- Runs per task
- 3
- Evaluated
- 6/15/2026, 7:44:50 AM
- Run ID
- 1c90c402-c758-49c0-a969-2898f4e3c1b9
- Results file
- Download JSON
Pass@1 by Difficulty
Easy
—
Medium
—
Hard
—
Scores by domain
Tasks passed (%) · top 1 models
Quant
Per-attempt drill-down requires a local result JSON in `results/`. Aggregate scores are shown from the published leaderboard.