Model Release Radar

Price/capability frontier position, benchmark scores, and community signal

View as JSON

Model Release Radar

Claude Opus 4.6 (Non reasoning,

anthropicยทclosed weightsยทProprietaryยทreleased Thursday, Feb 5, 2026

Token prices

No matched provider pricing is available yet.

Compare cost and intelligence โ†’

Frontier position

Not enough paired price and capability data exists yet to place this model on the price/capability frontier.

Cost vs. capability

This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet, so no cost/capability chart is drawn. Its Artificial Analysis benchmark scores are listed below.

Benchmark scores

Measured cost per task (DeepSWE)

This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet.

Artificial Analysis scores

Artificial Analysis scores describe capability. Compare them against estimated token spend with your cache mix; these estimates are not measured task costs.

BenchmarkScore
AA intelligence index26.4
GPQA Diamond (science)84.0%
Humanity's Last Exam19.1%
IFBench (instruction following)44.6%
LCR67.0%
Tau2 (tool use)84.8%
Terminal-Bench (hard)48.5%
Variants

This page covers 2 reasoning-effort variants of the same model.

VariantAA coding indexAA intelligence indexBlended /1M (3:1 input/output)
Standard (shown above)-26.4undisclosed
High effort--$10/1M
Community signal
VariantCoding EloOverall EloVotes
High effort1,5351,50477,193
Standard1,5351,49881,769