Compare · ModelsLive · 2 picked · head to head

Claude 2 vs Qwen 3.5 Plus (hosted 397B-A17B)

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

Qwen 3.5 Plus (hosted 397B-A17B) wins 2 of 2 shared benchmarks. Leads in knowledge · math.

Category leads
knowledge·Qwen 3.5 Plus (hosted 397B-A17B)math·Qwen 3.5 Plus (hosted 397B-A17B)
Hype vs Reality
Claude 2
#186 by perf·no signal
QUIET
Qwen 3.5 Plus (hosted 397B-A17B)
#202 by perf·no signal
QUIET
Best value
Claude 2
no price
Qwen 3.5 Plus (hosted 397B-A17B)
no price
Vendor risk
Anthropic logo
Anthropic
$380.0B·Tier 1
Medium risk
Alibaba Qwen logo
Alibaba (Qwen)
$293.0B·Tier 1
Low risk
Head to head
Claude 2Qwen 3.5 Plus (hosted 397B-A17B)
GPQA diamond
Qwen 3.5 Plus (hosted 397B-A17B) leads by +66.0
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Claude 2
12.9
Qwen 3.5 Plus (hosted 397B-A17B)
78.9
OTIS Mock AIME 2024-2025
Qwen 3.5 Plus (hosted 397B-A17B) leads by +82.6
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Claude 2
2.4
Qwen 3.5 Plus (hosted 397B-A17B)
85.0
Full benchmark table
BenchmarkClaude 2Qwen 3.5 Plus (hosted 397B-A17B)
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
12.978.9
OTIS Mock AIME 2024-2025
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
2.485.0
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
Anthropic logoClaude 2
Alibaba Qwen logoQwen 3.5 Plus (hosted 397B-A17B)