Better than 77% of all models
Context
N/A
Input $/1M
TBD
Output $/1M
TBD
Type
text-generation
License
Open Source
Benchmarks
12 tested
Data updated today
About
Qwen text generation model. 758K downloads on HuggingFace.
Tested on 12 benchmarks with 50.0% average. Top scores: GSM8K (88.7%), HellaSwag (73.6%), Aider — Code Editing (69.2%).
Capabilities
coding
69.2
#23 globally
reasoning
7.0
#178 globally
math
60.6
#60 globally
knowledge
48.1
#131 globally
language
69.1
#80 globally
general
44.2
#14 globally
Benchmark Scores
Compare AllTested on 12 benchmarks · Ranked across 6 categories
Score Distribution (all 274 models)
0255075100
▲ You are here
codingCompare coding →
Aider — Code Editing
69.2—Code editing benchmark from the Aider project. Measures ability to apply targeted code changes while maintaining correctness and style.
reasoningCompare reasoning →
MUSR
7.0—HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.
mathCompare math →
GSM8K
88.7—Grade school math word problems. 8,500 problems testing multi-step arithmetic reasoning. A foundational math benchmark.
MATH Level 5
32.5—HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Research
Documentation
Community
Source Code
BenchGecko API
qwen-qwen25-coder-14b-instruct
Specifications
- Typetext-generation
- ContextN/A
- ReleasedNov 2024
- LicenseOpen Source
- StatusActive
Available On
Learn More
Share & Export
Frequently Asked Questions
Qwen2.5 Coder 14B Instruct is an open-source text-generation AI model by Alibaba, released in November 2024. It has an average benchmark score of 64.4.