Context
512K tokens (~256 books)
Input $/1M
$0.60
Output $/1M
$3.60
Type
text
License
Proprietary
Benchmarks
0 tested
Data as of
About
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
No benchmark data available yet.
Recently Happened
Nemotron 3 Ultra (batch) pricing increased 100%
Aug 19, 2026
Links
Research
Documentation
Community
BenchGecko API
nemotron-3-ultra-550b-a55b-batch
Specifications
- Typetext
- Context512K tokens (~256 books)
- ReleasedJun 2026
- LicenseProprietary
- StatusActive
- Cost / Message~$0.005
Available On
Learn More
Share & Export
Frequently Asked Questions
Nemotron 3 Ultra (batch) is a proprietary text AI model by NVIDIA, released in June 2026. Context window: 512K tokens.
Related Models
Qwen3 Coder Next FP8 · AlibabaQwen2.5 Coder 32B Instruct AWQ · AlibabaSpeaker Diarization Community 1 · PyannoteGemma 3 12B (free) · Google DeepMindQwen3 Next 80B A3B Instruct (free) · Alibaba QwenBenchmarks
Key facts · as of 2026-09-02
- Nemotron 3 Ultra (batch) by NVIDIA. Not enough public benchmark scores to rank yet.
- List price $0.60 input · $3.60 output per 1M tokens (as of 2026-09-02).
How to cite · data as of 2026-09-02
Nemotron 3 Ultra (batch) · benchmarks, pricing and providers. BenchGecko, data as of 2026-09-02. https://benchgecko.ai/model/nemotron-3-ultra-550b-a55b-batch
Credit "Source: BenchGecko" with a link. Prices per provider and Gecko Tests are BenchGecko data (CC BY 4.0); benchmark scores keep their original source, listed in the JSON. JSON · llms.txt · MCP