Home/Models/Llama 3-70B
Meta logo

Llama 3-70B

by Meta · Released Jan 2024

Open Source
34.2
avg score
Rank #201
Compare
Better than 27% of all models
Context
N/A
Input $/1M
TBD
Output $/1M
TBD
Type
text
License
Open Source
Benchmarks
12 tested
Data updated today
About

Tested on 12 benchmarks with 24.7% average. Top scores: MMLU (72.4%), Winogrande (67.0%), IFEval (33.2%).

Capabilities
coding
5.0
#164 globally
reasoning
7.1
#176 globally
math
10.8
#217 globally
knowledge
38.4
#186 globally
language
33.2
#139 globally
general
26.7
#36 globally
Benchmark Scores
Compare All
Tested on 12 benchmarks · Ranked across 6 categories
Score Distribution (all 274 models)
0255075100
▲ You are here
Cybench

Capture-the-flag cybersecurity challenges. Tests vulnerability analysis, reverse engineering, cryptography, and exploitation skills.

5.0
MUSR

HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.

7.1
MATH level 5

Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.

22.6
MATH Level 5

HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.

5.7
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

4.2
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
llama-3-70b
Specifications
  • Typetext
  • ContextN/A
  • ReleasedJan 2024
  • LicenseOpen Source
  • Statusbenchmark-only
Available On
Meta logoMetaTBD
Share & Export
Tweet
Llama 3-70B is an open-source text AI model by Meta, released in January 2024. It has an average benchmark score of 34.2.