Better than 13% of all models
Context
N/A
Input $/1M
TBD
Output $/1M
TBD
Type
text-generation
License
Open Source
Benchmarks
11 tested
Data updated today
About
Meta-llama text generation model. 865K downloads on HuggingFace.
Tested on 11 benchmarks with 23.6% average. Top scores: JSQuAD (79.9%), LLM-JP — Overall (37.2%), JNLI (36.1%).
Capabilities
reasoning
3.8
#166 globally
math
1.7
#208 globally
knowledge
5.9
#216 globally
language
38.7
#114 globally
general
10.3
#50 globally
Benchmark Scores
Compare AllTested on 11 benchmarks · Ranked across 5 categories
Score Distribution (all 231 models)
0255075100
▲ You are here
reasoningCompare reasoning →
MUSR
3.8—HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.
mathCompare math →
MATH Level 5
1.7—HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.
knowledgeCompare knowledge →
MMLU-PRO
9.6—HuggingFace MMLU-Pro. Harder version of MMLU with 10 answer choices instead of 4 and more challenging questions.
GPQA
2.2—HuggingFace evaluation of GPQA (Graduate-Level Google-Proof Q&A). PhD-level science questions that cannot be easily searched.
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Info
Research
Documentation
Community
Source Code
BenchGecko API
meta-llama-llama-2-7b-hf
Specifications
- Typetext-generation
- ContextN/A
- ReleasedJul 2023
- LicenseOpen Source
- StatusActive
Available On
Learn More
Share & Export
Frequently Asked Questions
Llama 2 7b Hf is an open-source text-generation AI model by Meta, released in July 2023. It has an average benchmark score of 22.2.