Skip to content
llmprice

Llama 3.2 1B Instruct

Updated Sep 11, 2026

llama131k context60k max outputreleased Sep 25, 2024Open weightsReasoningTool calling
Cheapest input
$0.01
Cheapest output
$0.01
Official
output per 1M tokens
Providers
6
offering this model
Spread
20.1×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
InferenceopenCheapest$0.01Cheapest$0.0116k
Cloudflare Workers AIopen$0.027$0.20160k
Kilo Gateway$0.027$0.20160k
OpenRouteropen$0.027$0.20160k
Pioneeropen$0.1$0.201$0.1$0.1131k
Nvidiafreefreefree128k

Other llama models