Skip to content
llmprice

Magnum v4 72B

Updated Sep 11, 2026

llama33k context8k max outputreleased Oct 22, 2024Open weights
Cheapest input
$2.006
Cheapest output
$2.992
Official
output per 1M tokens
Providers
3
offering this model
Spread
1.7×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
NanoGPTopenCheapest$2.006Cheapest$2.992$1.00316k
Kilo Gateway$2.5$533k
OpenRouteropen$2.5$533k

Other llama models