Compare · ModelsLive · 2 picked · head to head
Gpt2 vs DeepSeek R1 Distill Qwen 1.5B
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
DeepSeek R1 Distill Qwen 1.5B wins on 4/6 benchmarks
DeepSeek R1 Distill Qwen 1.5B wins 4 of 6 shared benchmarks. Leads in general · language · math.
Category leads
general·DeepSeek R1 Distill Qwen 1.5Bknowledge·Gpt2language·DeepSeek R1 Distill Qwen 1.5Bmath·DeepSeek R1 Distill Qwen 1.5Breasoning·Gpt2
Hype vs Reality
Attention vs performance
Gpt2
#274 by perf·no signal
DeepSeek R1 Distill Qwen 1.5B
#268 by perf·no signal
Vendor risk
Mixed exposure
One or more vendors flagged
OpenAI
$840.0B·Tier 1
DeepSeek
$3.4B·Tier 1
Head to head
6 benchmarks · 2 models
Gpt2DeepSeek R1 Distill Qwen 1.5B
BBH (HuggingFace)
DeepSeek R1 Distill Qwen 1.5B leads by +2.1
Gpt2
2.7
DeepSeek R1 Distill Qwen 1.5B
4.7
GPQA
Gpt2 leads by +0.3
Gpt2
1.1
DeepSeek R1 Distill Qwen 1.5B
0.8
IFEval
DeepSeek R1 Distill Qwen 1.5B leads by +16.7
Gpt2
17.9
DeepSeek R1 Distill Qwen 1.5B
34.6
MATH Level 5
DeepSeek R1 Distill Qwen 1.5B leads by +16.7
Gpt2
0.2
DeepSeek R1 Distill Qwen 1.5B
16.9
MMLU-PRO
DeepSeek R1 Distill Qwen 1.5B leads by +0.3
Gpt2
1.8
DeepSeek R1 Distill Qwen 1.5B
2.1
MUSR
Gpt2 leads by +12.4
Gpt2
15.3
DeepSeek R1 Distill Qwen 1.5B
3.0
Full benchmark table
| Benchmark | Gpt2 | DeepSeek R1 Distill Qwen 1.5B |
|---|---|---|
BBH (HuggingFace) | 2.7 | 4.7 |
GPQA | 1.1 | 0.8 |
IFEval | 17.9 | 34.6 |
MATH Level 5 | 0.2 | 16.9 |
MMLU-PRO | 1.8 | 2.1 |
MUSR | 15.3 | 3.0 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| — | — | — | — | |
| — | — | — | — |