Compare · ModelsLive · 2 picked · head to head

Gemini 3.1 Pro Preview vs Gemma 4 31B

Side by side · benchmarks, pricing, and signals you can act on.

Winner summary

Gemini 3.1 Pro Preview wins 20 of 20 shared benchmarks. Leads in speed · arena · knowledge.

Category leads
speed·Gemini 3.1 Pro Previewarena·Gemini 3.1 Pro Previewknowledge·Gemini 3.1 Pro Previewgeneral·Gemini 3.1 Pro Previewmath·Gemini 3.1 Pro Previewcoding·Gemini 3.1 Pro Preview
Hype vs Reality
Gemini 3.1 Pro Preview
#137 by perf·#8 by attention
DESERVED
Gemma 4 31B
#101 by perf·#7 by attention
DESERVED
Best value
35.7x better value than Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview
6.9 pts/$
$7.00/M
Gemma 4 31B
245.6 pts/$
$0.22/M
Vendor risk
Google DeepMind logo
Google DeepMind
$4.20T·Tier 1
Low risk
Google DeepMind logo
Google DeepMind
$4.20T·Tier 1
Low risk
Head to head
Gemini 3.1 Pro PreviewGemma 4 31B
Artificial Analysis · CritPt
Gemini 3.1 Pro Preview leads by +16.3
Gemini 3.1 Pro Preview
17.7
Gemma 4 31B
1.4
Artificial Analysis · GDPval
Gemini 3.1 Pro Preview leads by +8.1
Gemini 3.1 Pro Preview
13.8
Gemma 4 31B
5.7
Artificial Analysis · GPQA Diamond
Gemini 3.1 Pro Preview leads by +8.4
Gemini 3.1 Pro Preview
94.1
Gemma 4 31B
85.7
Artificial Analysis · Humanity's Last Exam
Gemini 3.1 Pro Preview leads by +23.4
Gemini 3.1 Pro Preview
47.0
Gemma 4 31B
23.6
Artificial Analysis · IFBench
Gemini 3.1 Pro Preview leads by +1.5
Gemini 3.1 Pro Preview
77.1
Gemma 4 31B
75.6
Artificial Analysis · Long Context Reasoning
Gemini 3.1 Pro Preview leads by +12.3
Gemini 3.1 Pro Preview
82.0
Gemma 4 31B
69.7
Artificial Analysis · MMMU Pro
Gemini 3.1 Pro Preview leads by +9.0
Gemini 3.1 Pro Preview
82.4
Gemma 4 31B
73.4
Artificial Analysis · Quality Index
Gemini 3.1 Pro Preview leads by +15.0
Gemini 3.1 Pro Preview
29.7
Gemma 4 31B
14.7
Artificial Analysis · SciCode
Gemini 3.1 Pro Preview leads by +13.2
Gemini 3.1 Pro Preview
58.7
Gemma 4 31B
45.5
Artificial Analysis · tau2-Bench Telecom
Gemini 3.1 Pro Preview leads by +35.7
Gemini 3.1 Pro Preview
95.6
Gemma 4 31B
59.9
Artificial Analysis · Terminal-Bench Hard
Gemini 3.1 Pro Preview leads by +17.4
Gemini 3.1 Pro Preview
53.8
Gemma 4 31B
36.4
Chatbot Arena Elo · Coding
Gemini 3.1 Pro Preview leads by +81.5
Gemini 3.1 Pro Preview
1446.2
Gemma 4 31B
1364.7
Chatbot Arena Elo · Overall
Gemini 3.1 Pro Preview leads by +34.1
Gemini 3.1 Pro Preview
1487.0
Gemma 4 31B
1452.8
Chess Puzzles
Gemini 3.1 Pro Preview leads by +52.6
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
Gemini 3.1 Pro Preview
52.6
Gemma 4 31B
0.0
Dtbench
Gemini 3.1 Pro Preview leads by +24.0
Gemini 3.1 Pro Preview
95.1
Gemma 4 31B
71.1
GPQA diamond
Gemini 3.1 Pro Preview leads by +24.9
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
Gemini 3.1 Pro Preview
92.6
Gemma 4 31B
67.7
Lmca
Gemini 3.1 Pro Preview leads by +17.1
Gemini 3.1 Pro Preview
63.3
Gemma 4 31B
46.2
OTIS Mock AIME 2024-2025
Gemini 3.1 Pro Preview leads by +22.3
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
Gemini 3.1 Pro Preview
95.6
Gemma 4 31B
73.3
SimpleQA Verified
Gemini 3.1 Pro Preview leads by +63.1
SimpleQA Verified · short factual questions with verified answers, measuring factual accuracy and the tendency to hallucinate or provide incorrect information.
Gemini 3.1 Pro Preview
73.5
Gemma 4 31B
10.4
WeirdML
Gemini 3.1 Pro Preview leads by +19.8
WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns.
Gemini 3.1 Pro Preview
72.1
Gemma 4 31B
52.3
Full benchmark table
BenchmarkGemini 3.1 Pro PreviewGemma 4 31B
Artificial Analysis · CritPt
17.71.4
Artificial Analysis · GDPval
13.85.7
Artificial Analysis · GPQA Diamond
94.185.7
Artificial Analysis · Humanity's Last Exam
47.023.6
Artificial Analysis · IFBench
77.175.6
Artificial Analysis · Long Context Reasoning
82.069.7
Artificial Analysis · MMMU Pro
82.473.4
Artificial Analysis · Quality Index
29.714.7
Artificial Analysis · SciCode
58.745.5
Artificial Analysis · tau2-Bench Telecom
95.659.9
Artificial Analysis · Terminal-Bench Hard
53.836.4
Chatbot Arena Elo · Coding
1446.21364.7
Chatbot Arena Elo · Overall
1487.01452.8
Chess Puzzles
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
52.60.0
Dtbench
95.171.1
GPQA diamond
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
92.667.7
Lmca
63.346.2
OTIS Mock AIME 2024-2025
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
95.673.3
SimpleQA Verified
SimpleQA Verified · short factual questions with verified answers, measuring factual accuracy and the tendency to hallucinate or provide incorrect information.
73.510.4
WeirdML
WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns.
72.152.3
Pricing · per 1M tokens · projected $/mo at 10M tokens
ModelInputOutputContextProjected $/mo
Google DeepMind logoGemini 3.1 Pro Preview$2.00$12.001.0M tokens (~524 books)$45.00
Google DeepMind logoGemma 4 31B$0.09$0.34262K tokens (~131 books)$1.53