Model Release Radar

Price/capability frontier position, benchmark scores, and community signal

View as JSON

Model Release Radar

Qwen3 Next 80B A3B Instruct

alibabaยทopen weightsยทApache 2.0ยทreleased Thursday, Sep 11, 2025

Frontier position

Not enough paired price and capability data exists yet to place this model on the price/capability frontier.

Cost vs. capability

This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet, so no cost/capability chart is drawn. Its Artificial Analysis benchmark scores are listed below.

Benchmark scores

Measured cost per task (DeepSWE)

This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet.

Artificial Analysis scores (no matched cost)

Artificial Analysis scores below have no measured per-task cost, so no frontier claim is made for them here.

BenchmarkScore
AA intelligence index13.8
AIME 2025 (math)66.3%
GPQA Diamond (science)73.8%
Humanity's Last Exam7.6%
IFBench (instruction following)39.7%
LCR52.0%
LiveCodeBench (coding)68.4%
MMLU-Pro (knowledge)81.9%
SciCode (coding)30.7%
Tau2 (tool use)21.6%
Terminal-Bench (hard)7.6%
Community signal
VariantCoding EloOverall EloVotes
Rating1,4411,41722,523