Home/Models/Hermes 3 70B Instruct
nousresearch logo

Hermes 3 70B Instruct

by nousresearch · Released Aug 2024

Open Source
73.2
avg score
Rank #30
Compare
Better than 90% of all models
Context
131K tokens (~66 books)
Input $/1M
$0.70
Output $/1M
$0.70
Type
text
License
Open Source
Benchmarks
6 tested
Data as of
About

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

Tested on 6 benchmarks · BenchGecko score 73.2. Top scores: IFEval (76.6%), BBH (HuggingFace) (53.8%), MMLU-PRO (41.4%).

Looking for similar performance at lower cost?
gpt-oss-120b (free) scores 74.2 (101% as good) at $0.00/1M input · 100% cheaper
Capabilities
reasoning
23.4
#138 globally
math
21.0
#221 globally
knowledge
28.1
#245 globally
general
53.8
#25 globally
language
76.6
#63 globally
Benchmark Scores
Compare All
Tested on 6 benchmarks · Ranked across 5 categories
Score Distribution (all 312 models)
0255075100
▲ You are here
MUSR

HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.

23.4·
MATH Level 5

HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.

21.0·
MMLU-PRO

HuggingFace MMLU-Pro. Harder version of MMLU with 10 answer choices instead of 4 and more challenging questions.

41.4·
GPQA

HuggingFace evaluation of GPQA (Graduate-Level Google-Proof Q&A). PhD-level science questions that cannot be easily searched.

14.9·
Excellent (85+) Good (70-85) Average (50-70) Below (<50)
Recently Happened
Hermes 3 70B Instruct pricing increased 133%
Jun 26, 2026
Links
Documentation
Community
BenchGecko API
hermes-3-llama-3-1-70b
Specifications
  • Typetext
  • Context131K tokens (~66 books)
  • ReleasedAug 2024
  • LicenseOpen Source
  • StatusActive
  • Cost / Message~$0.002
Available On
nousresearch logonousresearch$0.70
Share & Export
Tweet
Hermes 3 70B Instruct is an open-source text AI model by nousresearch, released in August 2024. It has an average benchmark score of 73.2. Context window: 131K tokens.

Key facts · as of 2026-10-05

  • Hermes 3 70B Instruct by nousresearch. BenchGecko score 73.2, rank 30 of 312 scored models (normalized average of public benchmark scores).
  • List price $0.70 input · $0.70 output per 1M tokens (as of 2026-10-05).
  • Sold by 1 provider (as of 2026-10-05): DeepInfra (fp8) $0.70 in / $0.70 out.

How to cite · data as of 2026-10-05

Hermes 3 70B Instruct · benchmarks, pricing and providers. BenchGecko, data as of 2026-10-05. https://benchgecko.ai/model/hermes-3-llama-3-1-70b

Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP