Skip to content
llmprice

Gemini 3.8 Flash

Updated Sep 11, 2026

gemini-flash1.0M context1.0M max outputreleased Sep 2, 2026ReasoningTool calling
Cheapest input
$0.75
Cheapest output
$3.75
Official (Google)
$3.75
output per 1M tokens
Providers
19
offering this model
Spread
2.0×
max / min output

Offers

Metered per-token prices. Cheapest input and output are highlighted; marks a context-tier surcharge.

Provider1M input1M output1M cache read1M cache writeContext
302.AICheapest$0.75Cheapest$3.751.0M
CrossModel$0.75$3.75$0.075$0.751.0M
Eden AI$0.75$3.75$0.075$0.04171.0M
Eden AI$0.75$3.75$0.075$0.04171.0M
Google$0.75$3.75$0.0751.0M
Vertex$0.75$3.75$0.0751.0M
Kilo Gateway$0.75$3.75$0.075$0.04171.0M
DevPass (LLM Gateway)$0.75$3.75$0.075$0.08331.0M
LLM Gateway$0.75$3.75$0.075$0.08331.0M
LLM Gateway$0.75$3.75$0.075$0.08331.0M
Merge Gateway$0.75$3.75$0.0751.0M
NanoGPT$0.75$3.75$0.075$0.04171.0M
Ofox$0.75$3.75$0.075$0.04151.0M
OpenRouter$0.75$3.75$0.075$0.04171.0M
Requesty$0.75$3.75$0.0751.0M
Vercel AI Gateway$0.75$3.75$0.0751.0M
Vivgrid$0.75$3.75$0.151.0M
Cortecs$0.825$4.125$0.082$0.0841.0M
Requestyeu$0.825$4.125$0.08251.0M
Venice AI$0.9375$4.6875$0.09381.0M
OpenCode Zen$1.5$7.5$0.151.0M

Subscription plans (not per-token)

Billed per month, not per token — never counted as the cheapest offer.

Provider1M input1M outputContext
GitHub Copilotplan$0.75$3.751.0M

Other gemini-flash models