Nemotron 3 Ultra
Updated Sep 11, 2026
nemotron1.0M context262k max outputreleased Jun 4, 2026Open weightsReasoningTool calling
Official
—
output per 1M tokens
Providers
10
offering this model
Spread
1.8×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| Requesty | Cheapest$0.5 | $2.5 | — | — | 262k |
| Vercel AI Gatewayopen | $0.6 | $2.4 | $0.12 | — | 1.0M |
| DigitalOcean | $0.9 | Cheapest$1.7 | — | — | 131k |
| Venice AIopen | $0.625 | $3.125 | $0.1875 | — | 256k |
| Weights & Biasesopen | $0.75 | $2.75 | $0.15 | — | 262k |
| Bothubfree | free | free | — | — | 1.0M |
| Kilo Gatewayfree | free | free | — | — | 1.0M |
| OpenCode Zenfree | free | free | free | — | 1.0M |
| OpenRouterfree | free | free | — | — | 1.0M |
Other nemotron models
- Llama 3.1 Nemotron 70B Instruct
- Nemotron 3.5 Content Safety
- Nemotron 3.5 Lightning
- Nemotron 3.5 Lightning 30B A3B
- nemotron-3-nano:30b
- Nemotron 3 Nano 30B A3B
- Nemotron 3 Nano Omni
- Nemotron 3 Nano Omni 30B A3B Reasoning
- Nemotron 3 Super
- Nemotron 3 Super 120B
- Nemotron 3 Super 120B A12B
- Nemotron 3 Ultra 550B
- Nemotron 3 Ultra 550B A55B
- Nvidia Nemotron Nano 12B V2 VL
- Nvidia Nemotron Nano 9B V2