gpt-oss-120b
- Input / 1M
- $0.35
- Output / 1M
- $0.75
- Cached input
- —
- Context
- 131K
A 1,000-token prompt costs $0.00035 to send. Token counts for Cerebras models are estimated — approximated, because this family has no public tokenizer.
- Tool calling
- Reasoning
- Max output 33K
Priced closest to this
| Model | Provider | Input | Output | |
|---|---|---|---|---|
| gemini-gemma-2-27b-it | $0.35 | $1.05 | ||
| gemini-gemma-2-9b-it | $0.35 | $1.05 | ||
| codellama-34b-instruct | Perplexity | $0.35 | $1.40 | |
| command-light | Cohere | $0.30 | $0.60 | |
| gemini-2.5-flash-native-audio-latest | $0.30 | $2.50 | ||
| gemini-2.5-flash-native-audio-preview-09-2025 | $0.30 | $2.50 |
Compare gpt-oss-120b
- vs claude-fable-5
- vs claude-opus-5
- vs claude-sonnet-5
- vs command-r
- vs command-r-plus
- vs deepseek-v3.2
- vs deepseek-v4-flash
- vs deepseek-v4-pro
- vs gemini-3.5-flash
- vs gemini-3.5-flash-lite
- vs gemini-3.6-flash
- vs gpt-5.6
- vs gpt-5.6-sol
- vs gpt-5.6-terra
- vs grok-4-1-fast
- vs grok-4-1-fast-reasoning
- vs grok-4.20-0309-reasoning
- vs jamba-large-1.6
- vs jamba-large-1.7
- vs jamba-mini-1.7
- vs kimi-k2-0905-preview
- vs kimi-k2.5
- vs kimi-k2.6
- vs llama-3.1-70b-instruct
- vs llama-3.1-8b-instant
- vs llama-3.1-8b-instruct
- vs llama-3.3-70b-versatile
- vs ministral-8b-latest
- vs mistral-large-3
- vs mistral-medium-3-5
- vs sonar-pro
- vs zai-glm-4.6
- vs zai-glm-4.7