Skip to content
llmprice

Llama 3.1 70B Instruct

Updated Sep 11, 2026

llama131k context16k max outputreleased Jul 23, 2024Open weightsTool calling
Cheapest input
$0.4
Cheapest output
$0.4
Official
output per 1M tokens
Providers
7
offering this model
Spread
1.8×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContextChanged
Kilo GatewayopenCheapest$0.4Cheapest$0.4131k
OpenRouteropen$0.4$0.4131ktodayout −44.4%
Amazon Bedrockopen$0.72$0.72128k
Amazon Bedrockusopen$0.72$0.72128k
DevPass (LLM Gateway)open$0.72$0.72128k
LLM Gateway$0.72$0.72128k
Vercel AI Gateway$0.72$0.72128k
Nvidiafreefreefree128k

Other llama models