Skip to content
llmprice

Llama 4 Maverick 17B 128E Instruct

Updated Sep 11, 2026

llama524k context8k max outputreleased Jan 15, 2025Open weightsTool calling
Cheapest input
$0.15
Cheapest output
$0.6
Official
output per 1M tokens
Providers
3
offering this model
Spread
1.9×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
IO.NETopenCheapest$0.15Cheapest$0.6$0.075$0.3430k
Vertexopen$0.35$1.15524k
Nvidiafreefreefree128k

Other llama models