Grok 4.1 Fast (Non-Reasoning)
Updated Sep 11, 2026
grok2.0M context2.0M max outputreleased Nov 19, 2025Tool calling
Official
—
output per 1M tokens
Providers
11
offering this model
Spread
1.1×
max / min output
Offers
Metered per-token prices. Cheapest input and output are highlighted; † marks a context-tier surcharge.
| Provider | 1M input | 1M output | 1M cache read | 1M cache write | Context |
|---|---|---|---|---|---|
| Jiekou.AI | Cheapest$0.18 | Cheapest$0.45 | — | — | 2.0M |
| Abacus | $0.2 | $0.5 | — | — | 2.0M |
| Azure | $0.2 | $0.5 | $0.05 | — | 128k |
| FrogBot | $0.2 | $0.5 | $0.05 | — | 2.0M |
| Helicone | $0.2 | $0.5 | $0.05 | — | 2.0M |
| DevPass (LLM Gateway) | $0.2 | $0.5 | $0.05 | — | 2.0M |
| LLM Gateway | $0.2 | $0.5 | — | — | 2.0M |
| Perplexity Agent | $0.2 | $0.5 | $0.05 | — | 2.0M |
| Vercel AI Gateway | $0.2 | $0.5 | $0.05 | — | 1.0M |