Compare · ModelsLive · 2 picked · head to head
GLM 5 vs Kimi K2.7 Code
Side by side · benchmarks, pricing, and signals you can act on.
Winner summary
Kimi K2.7 Code wins on 12/14 benchmarks
Kimi K2.7 Code wins 12 of 14 shared benchmarks. Leads in agentic · knowledge · coding.
Category leads
agentic·Kimi K2.7 Codeknowledge·Kimi K2.7 Codecoding·Kimi K2.7 Codereasoning·Kimi K2.7 Codelanguage·Kimi K2.7 Codemath·GLM 5
Hype vs Reality
Attention vs performance
GLM 5
#79 by perf·#3 by attention
Kimi K2.7 Code
#69 by perf·#17 by attention
Best value
GLM 5
1.4x better value than Kimi K2.7 Code
GLM 5
43.7 pts/$
$1.26/M
Kimi K2.7 Code
30.5 pts/$
$1.84/M
Vendor risk
Who is behind the model
z-ai
private · undisclosed
Moonshot AI
$18.0B·Tier 1
Head to head
14 benchmarks · 2 models
GLM 5Kimi K2.7 Code
APEX-Agents
Kimi K2.7 Code leads by +20.4
APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments.
GLM 5
17.2
Kimi K2.7 Code
37.6
Chess Puzzles
Kimi K2.7 Code leads by +11.6
Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities.
GLM 5
5.3
Kimi K2.7 Code
16.9
GPQA diamond
Kimi K2.7 Code leads by +0.1
Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs.
GLM 5
83.8
Kimi K2.7 Code
83.8
LiveBench · Agentic Coding
Kimi K2.7 Code leads by +15.0
GLM 5
55.0
Kimi K2.7 Code
70.0
LiveBench · Coding
Kimi K2.7 Code leads by +0.3
GLM 5
73.6
Kimi K2.7 Code
74.0
LiveBench · Data Analysis
GLM 5 leads by +5.2
GLM 5
67.9
Kimi K2.7 Code
62.7
LiveBench · If
Kimi K2.7 Code leads by +1.0
GLM 5
55.3
Kimi K2.7 Code
56.3
LiveBench · Language
Kimi K2.7 Code leads by +0.4
GLM 5
77.5
Kimi K2.7 Code
77.9
LiveBench · Mathematics
GLM 5 leads by +3.9
GLM 5
83.5
Kimi K2.7 Code
79.6
LiveBench · Overall
Kimi K2.7 Code leads by +3.0
GLM 5
68.8
Kimi K2.7 Code
71.9
LiveBench · Reasoning
Kimi K2.7 Code leads by +13.7
GLM 5
69.1
Kimi K2.7 Code
82.8
OTIS Mock AIME 2024-2025
Kimi K2.7 Code leads by +15.6
OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills.
GLM 5
80.0
Kimi K2.7 Code
95.5
SimpleBench
Kimi K2.7 Code leads by +5.6
SimpleBench · tests fundamental reasoning capabilities with straightforward problems designed to expose gaps in basic logical and spatial thinking.
GLM 5
43.8
Kimi K2.7 Code
49.5
WeirdML
Kimi K2.7 Code leads by +5.9
WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns.
GLM 5
48.2
Kimi K2.7 Code
54.1
Full benchmark table
| Benchmark | GLM 5 | Kimi K2.7 Code |
|---|---|---|
APEX-Agents APEX-Agents · evaluates AI agents on complex, multi-step tasks requiring planning, tool use, and autonomous decision-making in realistic environments. | 17.2 | 37.6 |
Chess Puzzles Chess Puzzles · tests strategic and tactical reasoning by having models solve chess puzzle positions, evaluating lookahead and pattern recognition abilities. | 5.3 | 16.9 |
GPQA diamond Graduate-Level Google-Proof QA (Diamond set) · expert-crafted questions in physics, biology, and chemistry that are difficult even for domain PhDs. | 83.8 | 83.8 |
LiveBench · Agentic Coding | 55.0 | 70.0 |
LiveBench · Coding | 73.6 | 74.0 |
LiveBench · Data Analysis | 67.9 | 62.7 |
LiveBench · If | 55.3 | 56.3 |
LiveBench · Language | 77.5 | 77.9 |
LiveBench · Mathematics | 83.5 | 79.6 |
LiveBench · Overall | 68.8 | 71.9 |
LiveBench · Reasoning | 69.1 | 82.8 |
OTIS Mock AIME 2024-2025 OTIS Mock AIME 2024-2025 · simulated American Invitational Mathematics Examination problems testing advanced problem-solving skills. | 80.0 | 95.5 |
SimpleBench SimpleBench · tests fundamental reasoning capabilities with straightforward problems designed to expose gaps in basic logical and spatial thinking. | 43.8 | 49.5 |
WeirdML WeirdML · tests models on unusual and adversarial machine learning tasks that require creative problem-solving beyond standard patterns. | 48.2 | 54.1 |
Pricing · per 1M tokens · projected $/mo at 10M tokens
| Model | Input | Output | Context | Projected $/mo |
|---|---|---|---|---|
| $0.60 | $1.92 | 205K tokens (~102 books) | $9.30 | |
| $0.61 | $3.07 | 262K tokens (~131 books) | $12.26 |