gemini-2.0-flash
- Input / 1M
- $0.10
- Output / 1M
- $0.40
- Cached input
- $0.025
- Context
- 1M
A 1,000-token prompt costs $0.00010 to send. Token counts for Google models are estimated — approximated, because this family has no public tokenizer.
- Vision
- Tool calling
- Max output 8K
Scheduled for deprecation on 2026-06-01. Plan a migration before then.
Priced closest to this
| Model | Provider | Input | Output | |
|---|---|---|---|---|
| llama3.1-8b | Cerebras | $0.10 | $0.10 | |
| gemini-2.0-flash-001 | $0.10 | $0.40 | ||
| gemini-2.5-flash-lite | $0.10 | $0.40 | ||
| gemini-flash-lite-latest | $0.10 | $0.40 | ||
| gpt-4.1-nano | OpenAI | $0.10 | $0.40 | |
| devstral-small-2505 | Mistral | $0.10 | $0.30 |