Llama 3.1 Nemotron 70B Instruct
Updated Sep 11, 2026
nemotron131k context8k max outputreleased Apr 15, 2025Open weightsTool calling
Official
—
output per 1M tokens
Providers
2
offering this model
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|
| Eden AIopen | Cheapest$0.6 | Cheapest$0.6 | — | — | 131k |
| Nvidiafree | free | free | — | — | 128k |