codellama-70b-instruct
- Input / 1M
- $0.70
- Output / 1M
- $2.80
- Cached input
- —
- Context
- 16K
A 1,000-token prompt costs $0.00070 to send. Token counts for Perplexity models are estimated — approximated, because this family has no public tokenizer.
- Max output 16K
Priced closest to this
| Model | Provider | Input | Output | |
|---|---|---|---|---|
| open-mixtral-8x7b | Mistral | $0.70 | $0.70 | |
| llama-2-70b-chat | Perplexity | $0.70 | $2.80 | |
| pplx-70b-chat | Perplexity | $0.70 | $2.80 | |
| gemini-3.1-flash-live-preview | $0.75 | $4.50 | ||
| gpt-5.4-mini | OpenAI | $0.75 | $4.50 | |
| llama3.1-70b | Cerebras | $0.60 | $0.60 |