Compare · ModelsLive · 2 picked · head to head

Grok 4 vs Qwen2-72B

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

Grok 4 wins 2 of 2 shared benchmarks. Leads in knowledge · coding.

Category leads
knowledge·Grok 4coding·Grok 4
Hype vs Reality
Grok 4
#88 by perf·no signal
QUIET
Qwen2-72B
#177 by perf·no signal
QUIET
Best value
Grok 4
no price
Qwen2-72B
no price
Vendor risk
xAI logo
xAI
$250.0B·Tier 1
Medium risk
Alibaba Qwen logo
Alibaba (Qwen)
$293.0B·Tier 1
Low risk
Head to head
Grok 4Qwen2-72B
GPQA diamond
Grok 4 leads by +61.6
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Grok 4
82.7
Qwen2-72B
21.0
WeirdML
Grok 4 leads by +34.4
WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns.
Grok 4
45.7
Qwen2-72B
11.3
Full benchmark table
BenchmarkGrok 4Qwen2-72B
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
82.721.0
WeirdML
WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns.
45.711.3
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
xAI logoGrok 4
Alibaba Qwen logoQwen2-72B