grok-4-1-fast
- Input / 1M
- $0.20
- Output / 1M
- $0.50
- Cached input
- $0.05
- Context
- 2M
A 1,000-token prompt costs $0.00020 to send. Token counts for xAI models are estimated — approximated, because this family has no public tokenizer.
- Vision
- Tool calling
- Reasoning
- Max output 2M
Long-context rate
Past 128K tokens in a single request, grok-4-1-fast switches to $0.40 input and $1.00 output per million — 2.0× the base input rate. Most calculators quote only the base figure.
Priced closest to this
| Model | Provider | Input | Output | |
|---|---|---|---|---|
| gpt-5.4-nano | OpenAI | $0.20 | $1.25 | |
| gpt-5.6-luna | OpenAI | $0.20 | $1.20 | |
| jamba-1.5 | AI21 | $0.20 | $0.40 | |
| jamba-1.5-mini | AI21 | $0.20 | $0.40 | |
| jamba-mini-1.6 | AI21 | $0.20 | $0.40 | |
| jamba-mini-1.7 | AI21 | $0.20 | $0.40 | compare |
Compare grok-4-1-fast
- vs claude-fable-5
- vs claude-opus-5
- vs claude-sonnet-5
- vs command-r
- vs command-r-plus
- vs deepseek-v3.2
- vs deepseek-v4-flash
- vs deepseek-v4-pro
- vs gemini-3.5-flash
- vs gemini-3.5-flash-lite
- vs gemini-3.6-flash
- vs gpt-5.6
- vs gpt-5.6-sol
- vs gpt-5.6-terra
- vs gpt-oss-120b
- vs grok-4-1-fast-reasoning
- vs grok-4.20-0309-reasoning
- vs jamba-large-1.6
- vs jamba-large-1.7
- vs jamba-mini-1.7
- vs kimi-k2-0905-preview
- vs kimi-k2.5
- vs kimi-k2.6
- vs llama-3.1-70b-instruct
- vs llama-3.1-8b-instant
- vs llama-3.1-8b-instruct
- vs llama-3.3-70b-versatile
- vs ministral-8b-latest
- vs mistral-large-3
- vs mistral-medium-3-5
- vs sonar-pro
- vs zai-glm-4.6
- vs zai-glm-4.7