实时
Apr 7Claude Mythos Preview · Anthropic's most capable model arrives·Mar 31GPT-5.4 Nano launched on OpenAI·Mar 31GPT-5.4 Mini joins the OpenAI lineup·Mar 30Claude Opus 4.5 input price dropped · $5.00 per 1M tokens·Mar 30Mistral Small 4 available via Mistral AI·Mar 29Gemini 2.5 Pro scores 94.1% on MMLU·Mar 29Grok 4.20 Multi-Agent Beta enters agent rankings·Mar 28DeepSeek V3.2 output price dropped · $0.38 per 1M tokens·Mar 287 new MCP servers added in dev-tools category·Mar 27Claude Sonnet 4.6 released by Anthropic·Mar 27Claude Opus 4.6 released by Anthropic·Mar 26OTIS Mock AIME 2024-2025 benchmark added·Mar 26Claude Opus 4.1 pricing increased · $15/$75 per 1M tokens·Mar 25Grok 4.20 Beta launched by xAI·Mar 25Inception added as a tracked provider·Mar 24DeepSeek R1 0528 posted 87.2% on GPQA Diamond·Mar 243 new MCP servers in AI/ML category·Mar 23GPT-4o Audio Preview marked as deprecated·Mar 23Mistral Medium 3.1 input price cut to $0.40 per 1M tokens·Mar 22DeepSeek V3.2 Speciale released·Mar 22WeirdML benchmark now tracked on BenchGecko·Mar 20Nemotron 3 Super (120B) launched by NVIDIA·Mar 20Gemini 2.5 Flash Lite priced at $0.10/$0.40 per 1M tokens·Mar 18Mistral Large 3 2512 released by Mistral AI·Mar 18Grok Code Fast 1 added to agent rankings·Mar 16Claude Sonnet 4.5 scores 91.7% on MMLU·Mar 1612 new MCP servers added across 5 categories·Mar 14GPT-5.4 Pro launched · OpenAI's new flagship·Mar 14GPT-5.4 standard tier released by OpenAI·Mar 12Grok 3 Mini marked as deprecated by xAI·Mar 12Llama 3.3 Nemotron Super 49B pricing dropped·Mar 10Liquid added as a tracked provider·Mar 10MiniMax M2.7 released by MiniMax·Mar 8Grok 4 posted 89.4% on GPQA Diamond·Mar 8LAMBADA benchmark scores now tracked·Mar 5Gemini 2.5 Flash output price reduced to $2.50 per 1M tokens·Mar 5Mercury 2 launched by Inception·Mar 3Qwen3.5-Flash released by Alibaba Qwen·Mar 35 new MCP servers added · finance and auth categories·Apr 7Claude Mythos Preview · Anthropic's most capable model arrives·Mar 31GPT-5.4 Nano launched on OpenAI·Mar 31GPT-5.4 Mini joins the OpenAI lineup·Mar 30Claude Opus 4.5 input price dropped · $5.00 per 1M tokens·Mar 30Mistral Small 4 available via Mistral AI·Mar 29Gemini 2.5 Pro scores 94.1% on MMLU·Mar 29Grok 4.20 Multi-Agent Beta enters agent rankings·Mar 28DeepSeek V3.2 output price dropped · $0.38 per 1M tokens·Mar 287 new MCP servers added in dev-tools category·Mar 27Claude Sonnet 4.6 released by Anthropic·Mar 27Claude Opus 4.6 released by Anthropic·Mar 26OTIS Mock AIME 2024-2025 benchmark added·Mar 26Claude Opus 4.1 pricing increased · $15/$75 per 1M tokens·Mar 25Grok 4.20 Beta launched by xAI·Mar 25Inception added as a tracked provider·Mar 24DeepSeek R1 0528 posted 87.2% on GPQA Diamond·Mar 243 new MCP servers in AI/ML category·Mar 23GPT-4o Audio Preview marked as deprecated·Mar 23Mistral Medium 3.1 input price cut to $0.40 per 1M tokens·Mar 22DeepSeek V3.2 Speciale released·Mar 22WeirdML benchmark now tracked on BenchGecko·Mar 20Nemotron 3 Super (120B) launched by NVIDIA·Mar 20Gemini 2.5 Flash Lite priced at $0.10/$0.40 per 1M tokens·Mar 18Mistral Large 3 2512 released by Mistral AI·Mar 18Grok Code Fast 1 added to agent rankings·Mar 16Claude Sonnet 4.5 scores 91.7% on MMLU·Mar 1612 new MCP servers added across 5 categories·Mar 14GPT-5.4 Pro launched · OpenAI's new flagship·Mar 14GPT-5.4 standard tier released by OpenAI·Mar 12Grok 3 Mini marked as deprecated by xAI·Mar 12Llama 3.3 Nemotron Super 49B pricing dropped·Mar 10Liquid added as a tracked provider·Mar 10MiniMax M2.7 released by MiniMax·Mar 8Grok 4 posted 89.4% on GPQA Diamond·Mar 8LAMBADA benchmark scores now tracked·Mar 5Gemini 2.5 Flash output price reduced to $2.50 per 1M tokens·Mar 5Mercury 2 launched by Inception·Mar 3Qwen3.5-Flash released by Alibaba Qwen·Mar 35 new MCP servers added · finance and auth categories·
脉搏19·健康
泡沫278%·波动
GPT-5.5 Pro+4.0
Open Source16.2%

The Pulse

脉搏
healthy
7d · +3 分
泡沫指数 · 分项
Valuation Premiumhealthy+2.1
Funding Accelerationhealthy+1.5
Concentration Riskhealthy0
Revenue Qualityhealthy+1.4
Capex Gaphealthy+0.3
最大变动 · Valuation Premium 上升 2.1
AI 泡沫指数
健康泡沫初现过热泡沫
已更新 May 21·方法论·研究·免费 API·开发者

跨层洞察

完整矩阵
#基准测试
1OpenAI logoGPT-5.5 Pro99.9$30.00400K3
2Anthropic logoClaude Mythos Preview99.81000K14
3Alibaba Qwen logoQwen3.5 397B A17B96.3$0.39262K11
4DeepSeek logoDeepSeek V3.2 Speciale95.2$0.40164K9
5OpenAI logoGPT-5.4 Pro93.0$30.001050K8
6OpenAI logoGPT-5.1-Codex-Max91.2$1.25400K8
7Google DeepMind logoGemini 3.1 Pro Preview90.0$2.001049K23
8stepfun logoStep 3.5 Flash89.5$0.10262K10
9OpenAI logoGPT-5 Chat89.0$1.25128K7
10Alibaba Qwen logoQwen3.6 Plus88.7$0.331000K11
11DeepSeek logoDeepSeek R1 Distill Qwen 14B88.311
12
HA
Qwen2.5 72B Instruct Abliterated
87.56
13z-ai logoGLM 5.187.0$1.05203K12
14OpenAI logoGPT-5.2-Codex85.4$1.75400K9
15Anthropic logoClaude Instant84.64
16DeepSeek logoDeepSeek-V2 (MoE-236B, May 2024)84.47
17OpenAI logoGPT-5.483.4$2.501050K16
18Anthropic logoClaude Opus 4.6 (Fast)83.3$30.001000K12
19OpenAI logoGPT-5.1-Codex82.8$1.25400K8
20xiaomi logoMiMo-V2-Flash81.7$0.09262K11
搜索 297 AI 术语 · 从 transformer 到注意力溢价打开
完整方法论
BenchGecko 数据多久更新一次?

模型和基准测试数据每日从主要来源刷新。定价从每个供应商 API 滚动拉取。热度 信号每周汇总。脉搏 在 UTC 00:00 重新计算。

什么是 脉搏?

一个 0-100 的 AI 经济健康综合分数。融合了反向 泡沫指数、基准测试速度、价格压缩、热度 多样性和供应链紧张度。数字越低越健康。

基准测试分数如何标准化?

每项基准测试在所有已评分模型中进行最小最大归一化。排名取每个模型在 3 项以上基准测试中的归一化分数平均值,以避免过度加权单项测试。

定价数据来自哪里?

直接来自供应商 API 响应 · OpenRouter、OpenAI、Anthropic、Google、xAI、DeepSeek、Mistral 等。每份快照都在模型详情页附有来源归属缓存。

可以引用 BenchGecko 的数据吗?

可以。每个页面都提供分享与引用栏,含 APA、MLA、BibTeX、Chicago 和纯文本格式。免费 API 层要求归属,所有场景均建议归属。

数据来源 ·OpenRouterEpoch AISWE-benchMCP RegistryChatbot ArenaHuggingFaceLiveBenchArtificial AnalysisSEALAider
2 小时前更新 · 10+ 权威来源 · 零编辑内容·Learn · Glossary·研究·Developers