Compare · ModelsLive · 2 picked · head to head

Llama 3 8B Instruct vs Qwen 3.5 Plus (hosted 397B-A17B)

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

Qwen 3.5 Plus (hosted 397B-A17B) wins 2 of 2 shared benchmarks. Leads in knowledge · math.

Category leads
knowledge·Qwen 3.5 Plus (hosted 397B-A17B)math·Qwen 3.5 Plus (hosted 397B-A17B)
Hype vs Reality
Llama 3 8B Instruct
#217 by perf·no signal
QUIET
Qwen 3.5 Plus (hosted 397B-A17B)
#202 by perf·no signal
QUIET
Best value
Llama 3 8B Instruct
220.0 pts/$
$0.14/M
Qwen 3.5 Plus (hosted 397B-A17B)
no price
Vendor risk
Meta logo
Meta AI
$1.50T·Tier 1
Low risk
Alibaba Qwen logo
Alibaba (Qwen)
$293.0B·Tier 1
Low risk
Head to head
Llama 3 8B InstructQwen 3.5 Plus (hosted 397B-A17B)
GPQA diamond
Qwen 3.5 Plus (hosted 397B-A17B) leads by +77.4
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Llama 3 8B Instruct
1.4
Qwen 3.5 Plus (hosted 397B-A17B)
78.9
OTIS Mock AIME 2024-2025
Qwen 3.5 Plus (hosted 397B-A17B) leads by +84.3
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Llama 3 8B Instruct
0.7
Qwen 3.5 Plus (hosted 397B-A17B)
85.0
Full benchmark table
BenchmarkLlama 3 8B InstructQwen 3.5 Plus (hosted 397B-A17B)
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
1.478.9
OTIS Mock AIME 2024-2025
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
0.785.0
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
Meta logoLlama 3 8B Instruct$0.14$0.148K tokens (~4 books)$1.40
Alibaba Qwen logoQwen 3.5 Plus (hosted 397B-A17B)