Skip to content
llmprice

Llama 3.3 70B

Updated Sep 11, 2026

llama131k context131k max outputreleased Dec 6, 2024Open weightsTool calling
Cheapest input
$0.53
Cheapest output
$0.71
Official
output per 1M tokens
Providers
6
offering this model
Spread
3.9×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
STACKITopenCheapest$0.53$0.76128k
Groqopen$0.59$0.79131k
Weights & Biasesopen$0.71Cheapest$0.71$0.71128k
Together AIopen$1.04$1.04131k
Venice AIopen$0.7$2.8128k
NanoGPTopen$1.75$2.75$1.75128k

Other llama models