Compare · ModelsLive · 2 picked · head to head

PaLM 2-L vs GPT-3.5 Turbo (older v0613)

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

PaLM 2-L wins 2 of 3 shared benchmarks. Leads in knowledge.

Category leads
knowledge·PaLM 2-L
Hype vs Reality
PaLM 2-L
#14 by perf·no signal
QUIET
GPT-3.5 Turbo (older v0613)
#139 by perf·no signal
QUIET
Best value
PaLM 2-L
no price
GPT-3.5 Turbo (older v0613)
30.5 pts/$
$1.50/M
Vendor risk
U
Unknown
private · undisclosed
Unknown
OpenAI logo
OpenAI
$840.0B·Tier 1
Medium risk
Head to head
PaLM 2-LGPT-3.5 Turbo (older v0613)
ARC AI2
GPT-3.5 Turbo (older v0613) leads by +24.3
AI2 Reasoning Challenge · tests grade-school level science knowledge with multiple-choice questions requiring reasoning beyond simple retrieval.
PaLM 2-L
58.9
GPT-3.5 Turbo (older v0613)
83.2
TriviaQA
PaLM 2-L leads by +0.3
TriviaQA · reading comprehension benchmark with trivia questions, requiring models to find and reason over evidence from provided documents.
PaLM 2-L
86.1
GPT-3.5 Turbo (older v0613)
85.8
Winogrande
PaLM 2-L leads by +2.8
WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs.
PaLM 2-L
66.0
GPT-3.5 Turbo (older v0613)
63.2
Full benchmark table
BenchmarkPaLM 2-LGPT-3.5 Turbo (older v0613)
ARC AI2
AI2 Reasoning Challenge · tests grade-school level science knowledge with multiple-choice questions requiring reasoning beyond simple retrieval.
58.983.2
TriviaQA
TriviaQA · reading comprehension benchmark with trivia questions, requiring models to find and reason over evidence from provided documents.
86.185.8
Winogrande
WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs.
66.063.2
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
U
PaLM 2-L
OpenAI logoGPT-3.5 Turbo (older v0613)$1.00$2.004K tokens (~2 books)$12.50