Model Release Radar
Price/capability frontier position, benchmark scores, and community signal
Model Release Radar
Claude Sonnet 5
anthropic·closed weights·Proprietary·released Tuesday, Jun 30, 2026
Frontier position
- Behind frontier DeepSWE pass@1 (agentic coding) - Gemini 3.7 Flash, Claude Opus 5, GLM-5.3, GPT-5.6 Terra, GLM-5.2 are priced the same or lower and at least as capable.
Cost vs. capability
Loading chart…
Benchmark scores
Measured cost per task (DeepSWE)
| Benchmark | Pass@1 | Measured cost / task | Frontier |
|---|---|---|---|
| DeepSWE (agentic coding) (high effort, 95% CI 43.7%-52.7%, 4 runs) | 48.2% | $7.43 (median $5.83) | behind |
Artificial Analysis scores (no matched cost)
No Artificial Analysis benchmark scores are tracked for this model yet.
Variants
This page covers 6 reasoning-effort variants of the same model.
| Variant | AA coding index | AA intelligence index | Price /1M |
|---|---|---|---|
| High effort (shown above) | - | - | $4.00/1M |
| Extra-high effort | - | - | $4.00/1M |
| Low effort | - | - | $4.00/1M |
| Max effort | 71.5 | 55.3 | $4.00/1M |
| Medium effort | - | - | $4.00/1M |
| Non-reasoning | 66.4 | 42.6 | $4.00/1M |
Community signal
| Variant | Coding Elo | Overall Elo | Votes |
|---|---|---|---|
| High effort | 1,486 | 1,442 | 29,477 |