Nemotron 3 Ultra
nvidia/nemotron-3-ultra-550b-a55b-20260604
- tool use
- structured outputs
- reasoning
- open-weight
At a glance
providernvidia
releasedJun 4, 2026
context512K
input / M$0.60
Specifications
What the API gives you
- Context window
- 512K tokens
- Max output
- —
- Input price
- $0.60 / 1M tokens
- Output price
- $3.60 / 1M tokens
- Cache read
- $0.20 / 1M tokens
- Modalities
- text → text
- Knowledge cutoff
- —
- Released
- Jun 4, 2026
Recommendation coverage
Where this model appears
Not currently included in a published recommendation. This means the available task-specific evidence did not support a ranking yet—not that the model was reviewed and rejected.
Benchmarks
No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.
Open-weight model on Hugging Face.
Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.