Home/Models/Llama 3 8B Instruct
Meta logo

Llama 3 8B Instruct

by Meta · Released Apr 2024

Open Source
41.8
avg score
Rank #189
Compare
Better than 39% of all models
Context
8K tokens (~4 books)
Input $/1M
$0.14
Output $/1M
$0.14
Type
text
License
Open Source
Benchmarks
16 tested
Data as of
About

Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 8B instruct-tuned version was optimized for high quality dialogue usecases. It has demonstrated strong...

Tested on 16 benchmarks · BenchGecko score 41.8. Top scores: Chatbot Arena Elo — Overall (1222.8%), ARC AI2 (77.1%), OpenBookQA (76.8%).

Capabilities
reasoning
19.9
#141 globally
math
3.6
#265 globally
knowledge
43.2
#178 globally
language
24.0
#153 globally
general
18.4
#168 globally
Benchmark Scores
Compare All
Tested on 16 benchmarks · Ranked across 6 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
MUSR

HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.

19.9·
MATH level 5

Competition-level math from AMC, AIME, and olympiad problems. Level 5 is the hardest tier, requiring creative problem-solving.

6.1·
MATH Level 5

HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.

3.9·
OTIS Mock AIME 2024-2025

Mock AIME (American Invitational Mathematics Exam) problems from OTIS. Tests mathematical competition performance.

0.7·
ARC AI2

AI2 Reasoning Challenge. Grade-school science questions requiring multi-step reasoning. Easy and Challenge sets test different difficulty levels.

77.1·
OpenBookQA

Elementary science questions with access to a small book of core science facts. Tests reasoning beyond memorization.

76.8·
TriviaQA

Trivia questions sourced from trivia enthusiasts and quiz websites. Tests breadth of general knowledge.

67.7·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Links
Documentation
Community
BenchGecko API
llama-3-8b-instruct
Specifications
  • Typetext
  • Context8K tokens (~4 books)
  • ReleasedApr 2024
  • LicenseOpen Source
  • StatusActive
  • Cost / Message~$0.000
Available On
Meta logoMeta$0.14
Share & Export
Tweet
Llama 3 8B Instruct is an open-source text AI model by Meta, released in April 2024. It has an average benchmark score of 41.8. Context window: 8K tokens.

Key facts · as of 2026-07-10

  • Llama 3 8B Instruct by Meta. BenchGecko score 41.8, rank 189 of 312 scored models (normalized average of public benchmark scores).
  • List price $0.14 input · $0.14 output per 1M tokens (as of 2026-07-10).

How to cite · data as of 2026-07-10

Llama 3 8B Instruct · benchmarks, pricing and providers. BenchGecko, data as of 2026-07-10. https://benchgecko.ai/model/llama-3-8b-instruct

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP