Skip to content
llmprice

DeepSeek R1 Distill Llama 70B

Updated Sep 11, 2026

deepseek-thinking131k context131k max outputreleased Jan 23, 2025Open weightsReasoningTool calling
Cheapest input
$0.03
Cheapest output
$0.13
Official
output per 1M tokens
Providers
7
offering this model
Spread
7.6×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
HeliconeCheapest$0.03Cheapest$0.13128k
FastRouteropen$0.03$0.14131k
Alibaba (China)$0.287$0.86133k
Kilo Gateway$0.8$0.88k
NovitaAIopen$0.8$0.88k
OpenRouteropen$0.8$0.88k
DigitalOceanopen$0.99$0.9933k

Other deepseek-thinking models