Compare · ModelsLive · 2 picked · head to head
Grok-3 mini vs Qwen3 32B
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Grok-3 mini wins on 3/4 benchmarks
Grok-3 mini wins 3 of 4 shared benchmarks. Leads in coding · math.
Category leads
coding·Grok-3 miniknowledge·Qwen3 32Bmath·Grok-3 mini
Hype vs Reality
Attention vs performance
Grok-3 mini
#93 by perf·no signal
Qwen3 32B
#109 by perf·#2 by attention
Vendor risk
Who is behind the model
xAI
$250.0B·Tier 1
Alibaba (Qwen)
$293.0B·Tier 1
Head to head
4 benchmarks · 2 models
Grok-3 miniQwen3 32B
Aider polyglot
Grok-3 mini leads by +9.3
Aider Polyglot · measures how well AI models can edit code across multiple programming languages using the Aider coding assistant framework.
Grok-3 mini
49.3
Qwen3 32B
40.0
Fiction.LiveBench
Qwen3 32B leads by +7.5
Fiction.LiveBench · a continuously updated benchmark using recently published fiction to test reading comprehension and reasoning, preventing data contamination.
Grok-3 mini
66.7
Qwen3 32B
74.2
GPQA diamond
Grok-3 mini leads by +14.1
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Grok-3 mini
68.3
Qwen3 32B
54.3
OTIS Mock AIME 2024-2025
Grok-3 mini leads by +10.9
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Grok-3 mini
77.8
Qwen3 32B
66.9
Full benchmark table
| Benchmark | Grok-3 mini | Qwen3 32B |
|---|---|---|
Aider polyglot Aider Polyglot · measures how well AI models can edit code across multiple programming languages using the Aider coding assistant framework. | 49.3 | 40.0 |
Fiction.LiveBench Fiction.LiveBench · a continuously updated benchmark using recently published fiction to test reading comprehension and reasoning, preventing data contamination. | 66.7 | 74.2 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 68.3 | 54.3 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 77.8 | 66.9 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| — | — | — | — | |
| $0.08 | $0.28 | 131K tokens (~66 books) | $1.30 |