Skip to content
llmprice

Llama 3.1 8B

Updated Sep 11, 2026

llama131k context131k max outputreleased Jul 23, 2024Open weightsTool calling
Cheapest input
$0.05
Cheapest output
$0.08
Official
output per 1M tokens
Providers
3
offering this model
Spread
2.8×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
GroqopenCheapest$0.05Cheapest$0.08131k
Merge Gatewayopen$0.22$0.22128k
Weights & Biasesopen$0.22$0.22$0.22131k

Other llama models