gpt-5.4-mini
- Input / 1M
- $0.75
- Output / 1M
- $4.50
- Cached input
- $0.075
- Context
- 272K
A 1,000-token prompt costs $0.00075 to send. Token counts for OpenAI models are exact — run through the published tokenizer.
- Vision
- Tool calling
- Reasoning
- Max output 128K
Pinned versions at the same rate: gpt-5.4-mini-2026-03-17
Priced closest to this
| Model | Provider | Input | Output | |
|---|---|---|---|---|
| gemini-3.1-flash-live-preview | $0.75 | $4.50 | ||
| open-mixtral-8x7b | Mistral | $0.70 | $0.70 | |
| codellama-70b-instruct | Perplexity | $0.70 | $2.80 | |
| llama-2-70b-chat | Perplexity | $0.70 | $2.80 | |
| pplx-70b-chat | Perplexity | $0.70 | $2.80 | |
| llama-3.3-70b | Cerebras | $0.85 | $1.20 |