Skip to content
llmprice

Llama 4 Scout 17B 16E Instruct

Updated Sep 11, 2026

llama131k context16k max outputreleased Apr 5, 2025Open weightsTool calling
Cheapest input
$0.2
Cheapest output
$0.78
Official
output per 1M tokens
Providers
3
offering this model
Spread
1.1×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
AzureopenCheapest$0.2Cheapest$0.78128k
Azure Cognitive Servicesopen$0.2$0.78128k
Cloudflare Workers AIopen$0.27$0.85131k

Other llama models