Llama 3.1 70B Instruct
Updated Sep 11, 2026
llama131k context16k max outputreleased Jul 23, 2024Open weightsTool calling
Official
—
output per 1M tokens
Providers
7
offering this model
Spread
1.8×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context | Changed |
|---|---|---|---|---|---|---|
| Kilo Gatewayopen | Cheapest$0.4 | Cheapest$0.4 | — | — | 131k | — |
| OpenRouteropen | $0.4 | $0.4 | — | — | 131k | todayout −44.4% |
| Amazon Bedrockopen | $0.72 | $0.72 | — | — | 128k | — |
| Amazon Bedrockusopen | $0.72 | $0.72 | — | — | 128k | — |
| DevPass (LLM Gateway)open | $0.72 | $0.72 | — | — | 128k | — |
| LLM Gateway | $0.72 | $0.72 | — | — | 128k | — |
| Vercel AI Gateway | $0.72 | $0.72 | — | — | 128k | — |
| Nvidiafree | free | free | — | — | 128k | — |
Other llama models
- Hermes 2 Pro Llama 3 8B
- Llama 3.1 70B
- Llama 3.1 8B
- Llama 3.1 8B Instruct
- Llama 3.2 11B Instruct
- Llama 3.2 11B Vision Instruct
- Llama 3.2 1B Instruct
- Llama 3.2 3B
- Llama 3.2 3B Instruct
- Llama 3.2 90B Vision Instruct
- Llama 3.3 70B
- Llama-3.3-70B-Instruct
- Llama 3.3 70B Versatile
- Llama 3.3 Euryale 70B
- Llama 3 70B Instruct
- Llama 4 Maverick
- Llama 4 Maverick 17B 128E Instruct
- Llama 4 Maverick 17B 128E Instruct FP8
- Llama 4 Maverick 17B Instruct
- Llama 4 Scout
- Llama 4 Scout 17B 16E Instruct
- Llama 4 Scout 17B Instruct
- Llama-Guard-3-8B
- Llama Guard 4 12B
- Llama Prompt Guard 2 22M
- Magnum v4 72B
- MythoMax 13B
- ReMM SLERP 13B