Model Release Radar
Price/capability frontier position, benchmark scores, and community signal
Model Release Radar
Grok 4.1 Fast
xaiยทclosed weightsยทProprietaryยทreleased Wednesday, Nov 19, 2025
Frontier position
Not enough paired price and capability data exists yet to place this model on the price/capability frontier.
Cost vs. capability
This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet, so no cost/capability chart is drawn. Its Artificial Analysis benchmark scores are listed below.
Benchmark scores
Measured cost per task (DeepSWE)
This model has not been measured on DeepSWE (or any other benchmark with a real measured per-task cost) yet.
Artificial Analysis scores (no matched cost)
Artificial Analysis scores below have no measured per-task cost, so no frontier claim is made for them here.
| Benchmark | Score |
|---|---|
| AA intelligence index | 31.3 |
| AIME 2025 (math) | 89.3% |
| GPQA Diamond (science) | 85.3% |
| Humanity's Last Exam | 19.3% |
| IFBench (instruction following) | 52.7% |
| LCR | 70.3% |
| LiveCodeBench (coding) | 82.2% |
| MMLU-Pro (knowledge) | 85.4% |
| SciCode (coding) | 44.2% |
| Tau2 (tool use) | 93.3% |
| Terminal-Bench (hard) | 24.2% |
Community signal
| Variant | Coding Elo | Overall Elo | Votes |
|---|---|---|---|
| Rating | 1,411 | 1,408 | 55,434 |