LLM Digest
Subscribe

Model Release Radar

Price/capability frontier position, benchmark scores, and community signal

View as JSON

Model Release Radar

GLM-5.3

zai·open weights·MIT·released Tuesday, Aug 18, 2026

Frontier position
  • Behind frontier DeepSWE pass@1 (agentic coding) - GPT-5.6 Terra, GPT-5.6 Sol are priced the same or lower and at least as capable.
Cost vs. capability

Loading chart…

Benchmark scores

Measured cost per task (DeepSWE)

BenchmarkPass@1Measured cost / taskFrontier
DeepSWE (agentic coding) (max effort, 95% CI 65.9%-72.0%, 4 runs)69.0%$3.99 (median $3.13)behind

Artificial Analysis scores (no matched cost)

Artificial Analysis scores below have no measured per-task cost, so no frontier claim is made for them here.

BenchmarkScore
AA intelligence index59.5
AA coding index74.8
GPQA Diamond (science)91.7%
Humanity's Last Exam42.3%
LCR76.3%
SciCode (coding)56.5%
Tau2 (banking)50.3%
Terminal-Bench 2.1 (agentic)83.9%
Community signal
VariantCoding EloOverall EloVotes
Rating1,5051,4755,820