Nemotron 3 Ultra

nvidia/nemotron-3-ultra-550b-a55b-20260604

  • tool use
  • structured outputs
  • reasoning
  • open-weight

At a glance

providernvidia
releasedJun 4, 2026
context512K
input / M$0.60

Specifications

What the API gives you

Context window
512K tokens
Max output
Input price
$0.60 / 1M tokens
Output price
$3.60 / 1M tokens
Cache read
$0.20 / 1M tokens
Modalities
text → text
Knowledge cutoff
Released
Jun 4, 2026

Recommendation coverage

Where this model appears

Not currently included in a published recommendation. This means the available task-specific evidence did not support a ranking yet—not that the model was reviewed and rejected.

Benchmarks

No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.

Open-weight model on Hugging Face.

Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.