Qwen3 4B Instruct 2507
开源来自 Alibaba · 发布于 2025-08-05
47.2
平均分
N/A
输入价格
N/A
输出价格
N/A
上下文窗口
text-generation
类型
Tested on 6 benchmarks with 47.2% average. Top scores: OpenCompass — IFEval (82.4%), OpenCompass — MMLU-Pro (63.0%), OpenCompass — GPQA-Diamond (52.3%).
基准测试分数
| 基准测试 | 类别 | 分数 | Bar |
|---|---|---|---|
| OpenCompass — IFEval | language | 82.4 | |
| OpenCompass — MMLU-Pro | knowledge | 63.0 | |
| OpenCompass — GPQA-Diamond | knowledge | 52.3 | |
| OpenCompass — AIME2025 | math | 46.9 | |
| OpenCompass — LiveCodeBenchV6 | coding | 33.5 | |
| OpenCompass — HLE | knowledge | 5.1 |
相似模型
Alibaba
47.3
Google DeepMind
47.4
Meta
46.9
OpenAI
46.9