Compare · ModelsLive · 2 picked · head to head

GPT-3.5 Turbo (older v0613) vs PaLM 2-L

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

PaLM 2-L wins 2 of 3 shared benchmarks. Leads in knowledge.

Category leads
knowledge·PaLM 2-L
Hype vs Reality
GPT-3.5 Turbo (older v0613)
#235 by perf·no signal
QUIET
PaLM 2-L
#12 by perf·no signal
QUIET
Best value
GPT-3.5 Turbo (older v0613)
21.9 pts/$
$1.50/M
PaLM 2-L
n/a
no price
Vendor risk
OpenAI logo
OpenAI
$840.0B·Tier 1
Medium risk
Unknown
private · undisclosed
Unknown
Head to head
GPT-3.5 Turbo (older v0613)PaLM 2-L
ARC AI2
GPT-3.5 Turbo (older v0613) leads by +24.3
AI2 Reasoning Challenge · tests grade-school level science knowledge with multiple-choice questions requiring reasoning beyond simple retrieval.
GPT-3.5 Turbo (older v0613)
83.2
PaLM 2-L
58.9
TriviaQA
PaLM 2-L leads by +0.3
TriviaQA · reading comprehension benchmark with trivia questions, requiring models to find and reason over evidence from provided documents.
GPT-3.5 Turbo (older v0613)
85.8
PaLM 2-L
86.1
Winogrande
PaLM 2-L leads by +2.8
WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs.
GPT-3.5 Turbo (older v0613)
63.2
PaLM 2-L
66.0
Full benchmark table
BenchmarkGPT-3.5 Turbo (older v0613)PaLM 2-L
ARC AI2
AI2 Reasoning Challenge · tests grade-school level science knowledge with multiple-choice questions requiring reasoning beyond simple retrieval.
83.258.9
TriviaQA
TriviaQA · reading comprehension benchmark with trivia questions, requiring models to find and reason over evidence from provided documents.
85.886.1
Winogrande
WinoGrande · large-scale commonsense reasoning benchmark where models must resolve ambiguous pronouns in carefully constructed sentence pairs.
63.266.0
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
OpenAI logoGPT-3.5 Turbo (older v0613)$1.00$2.004K tokens (~2 books)$12.50
PaLM 2-L————