Skip to content
llmprice

Llama 3.2 3B

Updated Sep 11, 2026

llama131k context131k max outputreleased Sep 25, 2024Open weightsTool calling
Cheapest input
$0.1
Cheapest output
$0.1
Official
output per 1M tokens
Providers
2
offering this model
Spread
6.0×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
PioneeropenCheapest$0.1Cheapest$0.1$0.1$0.1131k
Venice AIopen$0.15$0.6128k

Other llama models