Compare · ModelsLive · 2 picked · head to head
Claude 3 Haiku vs Dolly 2.0-12b
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Claude 3 Haiku wins on 2/2 benchmarks
Claude 3 Haiku wins 2 of 2 shared benchmarks. Leads in knowledge.
Category leads
knowledge·Claude 3 Haiku
Hype vs Reality
Attention vs performance
Claude 3 Haiku
#271 by perf·no signal
Dolly 2.0-12b
#253 by perf·no signal
Vendor risk
Who is behind the model
Anthropic
$965.0B·Tier 1
U
Unknown
private · undisclosed
Head to head
2 benchmarks · 2 models
Claude 3 HaikuDolly 2.0-12b
MMLU
Claude 3 Haiku leads by +63.5
Massive Multitask Language Understanding · 57 subjects spanning STEM, humanities, social sciences, and more. The standard benchmark for broad knowledge.
Claude 3 Haiku
65.1
Dolly 2.0-12b
1.6
Winogrande
Claude 3 Haiku leads by +24.8
WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs.
Claude 3 Haiku
48.4
Dolly 2.0-12b
23.6
Full benchmark table
| Benchmark | Claude 3 Haiku | Dolly 2.0-12b |
|---|---|---|
MMLU Massive Multitask Language Understanding · 57 subjects spanning STEM, humanities, social sciences, and more. The standard benchmark for broad knowledge. | 65.1 | 1.6 |
Winogrande WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs. | 48.4 | 23.6 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| $0.25 | $1.25 | 200K tokens (~100 books) | $5.00 | |
U Dolly 2.0-12b | — | — | — | — |