Skip to content
llmprice

Qwen3.5 4B

Updated Sep 11, 2026

qwen262k context33k max outputreleased Nov 1, 2025Open weightsReasoningTool calling
Cheapest input
$0.04
Cheapest output
$0.07
Official
output per 1M tokens
Providers
3
offering this model
Spread
2.9×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
EmpirioLabs AIopenCheapest$0.04Cheapest$0.07$0.02262k
NanoGPTopen$0.1$0.2$0.05262k
QVACfreefreefree33k

Other qwen models