Skip to content
llmprice

Llama 3.2 3B Instruct

Updated Sep 11, 2026

llama131k context118k max outputreleased Sep 18, 2024Open weightsReasoningTool calling
Cheapest input
$0.02
Cheapest output
$0.02
Official
output per 1M tokens
Providers
10
offering this model
Spread
16.8×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
InferenceopenCheapest$0.02Cheapest$0.0216k
DevPass (LLM Gateway)open$0.03$0.0533k
LLM Gateway$0.03$0.0533k
NovitaAIopen$0.03$0.0533k
NanoGPTopen$0.0306$0.0493$0.0153131k
Kilo Gateway$0.05$0.33131k
OpenRouteropen$0.05$0.33131k
Cloudflare Workers AIopen$0.0509$0.33580k
Pioneeropen$0.1$0.335$0.1$0.1131k
Nvidiafreefreefree33k

Other llama models