Skip to content
llmprice

Llama 3.1 70B

Updated Sep 11, 2026

llama131k context131k max outputreleased Jul 23, 2024Open weightsTool calling
Cheapest input
$0.8
Cheapest output
$0.8
Official
output per 1M tokens
Providers
2
offering this model
Spread
1.2×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
Weights & BiasesopenCheapest$0.8Cheapest$0.8$0.8131k
Merge Gatewayopen$0.99$0.99128k

Other llama models