Compare · ModelsLive · 2 picked · head to head
Llama 2-13B vs Qwen3.5-Flash
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Qwen3.5-Flash wins on 4/4 benchmarks
Qwen3.5-Flash wins 4 of 4 shared benchmarks. Leads in knowledge · general · math.
Category leads
knowledge·Qwen3.5-Flashgeneral·Qwen3.5-Flashmath·Qwen3.5-Flash
Hype vs Reality
Attention vs performance
Llama 2-13B
#224 by perf·#18 by attention
Qwen3.5-Flash
#232 by perf·#2 by attention
Vendor risk
Who is behind the model
Meta AI
$1.87T·Tier 1
Alibaba (Qwen)
$293.0B·Tier 1
Head to head
4 benchmarks · 2 models
Llama 2-13BQwen3.5-Flash
Chess Puzzles
Qwen3.5-Flash leads by +16.9
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
Llama 2-13B
0.0
Qwen3.5-Flash
16.9
Dtbench
Qwen3.5-Flash leads by +67.8
Llama 2-13B
3.7
Qwen3.5-Flash
71.5
GPQA diamond
Qwen3.5-Flash leads by +74.7
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Llama 2-13B
1.8
Qwen3.5-Flash
76.4
OTIS Mock AIME 2024-2025
Qwen3.5-Flash leads by +84.4
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Llama 2-13B
0.0
Qwen3.5-Flash
84.4
Full benchmark table
| Benchmark | Llama 2-13B | Qwen3.5-Flash |
|---|---|---|
Chess Puzzles Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities. | 0.0 | 16.9 |
Dtbench | 3.7 | 71.5 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 1.8 | 76.4 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 0.0 | 84.4 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| — | — | — | — | |
| $0.07 | $0.26 | 1.0M tokens (~500 books) | $1.14 |