Llama 3.3 70B
Updated Sep 11, 2026
llama131k context131k max outputreleased Dec 6, 2024Open weightsTool calling
Official
—
output per 1M tokens
Providers
6
offering this model
Spread
3.9×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| STACKITopen | Cheapest$0.53 | $0.76 | — | — | 128k |
| Groqopen | $0.59 | $0.79 | — | — | 131k |
| Weights & Biasesopen | $0.71 | Cheapest$0.71 | $0.71 | — | 128k |
| Together AIopen | $1.04 | $1.04 | — | — | 131k |
| Venice AIopen | $0.7 | $2.8 | — | — | 128k |
| NanoGPTopen | $1.75 | $2.75 | $1.75 | — | 128k |
Other llama models
- Hermes 2 Pro Llama 3 8B
- Llama 3.1 70B
- Llama 3.1 70B Instruct
- Llama 3.1 8B
- Llama 3.1 8B Instruct
- Llama 3.2 11B Instruct
- Llama 3.2 11B Vision Instruct
- Llama 3.2 1B Instruct
- Llama 3.2 3B
- Llama 3.2 3B Instruct
- Llama 3.2 90B Vision Instruct
- Llama-3.3-70B-Instruct
- Llama 3.3 70B Versatile
- Llama 3.3 Euryale 70B
- Llama 3 70B Instruct
- Llama 4 Maverick
- Llama 4 Maverick 17B 128E Instruct
- Llama 4 Maverick 17B 128E Instruct FP8
- Llama 4 Maverick 17B Instruct
- Llama 4 Scout
- Llama 4 Scout 17B 16E Instruct
- Llama 4 Scout 17B Instruct
- Llama-Guard-3-8B
- Llama Guard 4 12B
- Llama Prompt Guard 2 22M
- Magnum v4 72B
- MythoMax 13B
- ReMM SLERP 13B