NVIDIA's Llama 3.1 Nemotron 70B is a language model designed for generating precise and useful responses. Leveraging Llama 3.1 70B architecture and Reinforcement Learning from Human Feedback (RLHF), it excels...
Tested on 1 benchmarks. Top scores: Chatbot Arena Elo — Overall (1298.5%).
Chatbot Arena overall Elo rating. Crowdsourced human preference ranking from blind head-to-head comparisons across all topics.
- Typetext
- Context131K tokens (~66 books)
- ReleasedOct 2024
- LicenseOpen Source
- StatusActive
- Cost / Message~$0.004
Frequently Asked Questions
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenBenchmarks
Chatbot Arena Elo — OverallKey facts · as of 2026-05-03
- Llama 3.1 Nemotron 70B Instruct by NVIDIA. Not enough public benchmark scores to rank yet.
- List price $1.20 input · $1.20 output per 1M tokens (as of 2026-05-03).
How to cite · data as of 2026-05-03
Llama 3.1 Nemotron 70B Instruct · benchmarks, pricing and providers. BenchGecko, data as of 2026-05-03. https://benchgecko.ai/model/llama-3-1-nemotron-70b-instruct
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP