Magnum v4 72B
Updated Sep 11, 2026
llama33k context8k max outputreleased Oct 22, 2024Open weights
Official
—
output per 1M tokens
Providers
3
offering this model
Spread
1.7×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| NanoGPTopen | Cheapest$2.006 | Cheapest$2.992 | $1.003 | — | 16k |
| Kilo Gateway | $2.5 | $5 | — | — | 33k |
| OpenRouteropen | $2.5 | $5 | — | — | 33k |
Other llama models
- Hermes 2 Pro Llama 3 8B
- Llama 3.1 70B
- Llama 3.1 70B Instruct
- Llama 3.1 8B
- Llama 3.1 8B Instruct
- Llama 3.2 11B Instruct
- Llama 3.2 11B Vision Instruct
- Llama 3.2 1B Instruct
- Llama 3.2 3B
- Llama 3.2 3B Instruct
- Llama 3.2 90B Vision Instruct
- Llama 3.3 70B
- Llama-3.3-70B-Instruct
- Llama 3.3 70B Versatile
- Llama 3.3 Euryale 70B
- Llama 3 70B Instruct
- Llama 4 Maverick
- Llama 4 Maverick 17B 128E Instruct
- Llama 4 Maverick 17B 128E Instruct FP8
- Llama 4 Maverick 17B Instruct
- Llama 4 Scout
- Llama 4 Scout 17B 16E Instruct
- Llama 4 Scout 17B Instruct
- Llama-Guard-3-8B
- Llama Guard 4 12B
- Llama Prompt Guard 2 22M
- MythoMax 13B
- ReMM SLERP 13B