Skip to content
llmprice

Llama Guard 4 12B

Updated Sep 11, 2026

llama164k context16k max outputreleased Apr 30, 2025Open weights
Cheapest input
$0.18
Cheapest output
$0.18
Official
output per 1M tokens
Providers
4
offering this model
Spread
1.2×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
Kilo GatewayCheapest$0.18Cheapest$0.18164k
OpenRouteropen$0.18$0.18164k
Helicone$0.21$0.21131k
Nvidiafreefreefree128k

Other llama models