Llama 4 Maverick 17B 128E Instruct FP8
Updated Sep 11, 2026
llama1.0M context16k max outputreleased Apr 5, 2025Open weightsTool calling
Official
—
output per 1M tokens
Providers
5
offering this model
Spread
1.5×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| Azureopen | Cheapest$0.25 | Cheapest$1 | — | — | 1.0M |
| Azure Cognitive Servicesopen | $0.25 | $1 | — | — | 1.0M |
| watsonx.aiopen | $0.371 | $1.484 | — | — | 131k |
| Llamafree | free | free | — | — | 128k |
| Vercel AI Gatewayfree | free | free | — | — | 128k |
Other llama models
- Hermes 2 Pro Llama 3 8B
- Llama 3.1 70B
- Llama 3.1 70B Instruct
- Llama 3.1 8B
- Llama 3.1 8B Instruct
- Llama 3.2 11B Instruct
- Llama 3.2 11B Vision Instruct
- Llama 3.2 1B Instruct
- Llama 3.2 3B
- Llama 3.2 3B Instruct
- Llama 3.2 90B Vision Instruct
- Llama 3.3 70B
- Llama-3.3-70B-Instruct
- Llama 3.3 70B Versatile
- Llama 3.3 Euryale 70B
- Llama 3 70B Instruct
- Llama 4 Maverick
- Llama 4 Maverick 17B 128E Instruct
- Llama 4 Maverick 17B Instruct
- Llama 4 Scout
- Llama 4 Scout 17B 16E Instruct
- Llama 4 Scout 17B Instruct
- Llama-Guard-3-8B
- Llama Guard 4 12B
- Llama Prompt Guard 2 22M
- Magnum v4 72B
- MythoMax 13B
- ReMM SLERP 13B