Nemotron 3 Ultra 550B A55B
by OpenRouter · family nemotron · listed Jun 2026RTW
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Current rates USD per 1M tokens
Input$0.60
Output$3.60
Cache read$0.20
Cache write—
Against the market
On the record
- Model id
nvidia/nemotron-3-ultra-550b-a55b- Context window
- 512K tokens
- Max output
- 16K tokens
- Input modalities
- text
- Output modalities
- text
- Reasoning
- yes
- Tool calls
- yes
- Open weights
- yes
- Knowledge cutoff
- —
- Released
- 2026-06-04
- Last updated
- 2026-06-04
- Provider docs
- openrouter.ai/models
Back of the envelope
≈
Also listed at
- Kenarifree in · free out
- Requestyfree in · free out
- Nvidia$0.50 in · $2.50 out
- NanoGPT$0.50 in · $2.50 out
- Eden AI$0.50 in · $2.20 out
- Kilo Gateway$0.50 in · $2.20 out
- Together AI$0.60 in · $3.60 out
- Vercel AI Gateway$0.60 in · $2.40 out