DeepSeek R1 Distill Llama 70B is a distilled large language model based on Llama-3.3-70B-Instruct, using outputs from DeepSeek R1. The model combines advanced distillation techniques to achieve high performance across...
Tested on 8 benchmarks · BenchGecko score 47.3. Top scores: IFEval (43.4%), MMLU-PRO (41.6%), BBH (HuggingFace) (35.8%).
gpt-oss-120b scores 46.7 (99% as good) at $0.04/1M input · 95% cheaper
HuggingFace MuSR (Multi-Step Reasoning). Tests multi-hop reasoning requiring chaining multiple facts together.
HuggingFace evaluation of MATH Level 5 problems. Competition math requiring advanced reasoning and proof construction.
HuggingFace MMLU-Pro. Harder version of MMLU with 10 answer choices instead of 4 and more challenging questions.
HuggingFace evaluation of GPQA (Graduate-Level Google-Proof Q&A). PhD-level science questions that cannot be easily searched.
- Typetext
- Context8K tokens (~4 books)
- ReleasedJan 2025
- LicenseOpen Source
- Statusexpired
- Cost / Message~$0.002
Frequently Asked Questions
Key facts · as of 2026-09-28
- R1 Distill Llama 70B by DeepSeek. BenchGecko score 47.3, rank 161 of 312 scored models (normalized average of public benchmark scores).
- List price $0.80 input · $0.80 output per 1M tokens (as of 2026-09-28).
How to cite · data as of 2026-09-28
R1 Distill Llama 70B · benchmarks, pricing and providers. BenchGecko, data as of 2026-09-28. https://benchgecko.ai/model/deepseek-r1-distill-llama-70b
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP