gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.
function_callinghttps://api.deepinfra.com/v1/openai/v1/chat/completions| Date | Input | Output | Change |
|---|---|---|---|
| Apr 26, 2026 | $0.039 | $0.19 | — |
| Jul 9, 202653ef452 | $0.037 | $0.17 | ↓5% |
Input is ↓5% since Apr 26, 2026, from $0.039 to $0.037 per 1M tokens.
| Provider | Input | Output | Since launch |
|---|---|---|---|
| DeepInfracheapest | $0.04 | $0.17 | ↓5% |
| OpenRoutercheapest | $0.04 | $0.17 | ↓5% |
| Cloudflare AI Gateway | $0.04 | $0.19 | — |
| Novita AI | $0.05 | $0.25 | — |
| Groq | $0.15 | $0.6 | — |
| SambaNova Cloud | $0.22 | $0.59 | — |
| Cerebras | $0.35 | $0.75 | — |
| Vercel AI Gateway | $0.35 | $0.75 | — |
| Model | Context | Max out | Pricing |
|---|---|---|---|
| gpt-oss-20b | 131K | 131K | $0.03$0.14 |
| gpt-oss-120b | 131K | 131K | $0.04$0.17 |
| gpt-oss-120b-Turbo | 131K | — | $0.15$0.6 |
| gpt-oss-120b-Ultra | 131K | — | $0.2$0.95 |
/v1/models/deepinfra/openai/gpt-oss-120bgpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. The model supports configurable reasoning depth, full chain-of-thought access, and native tool use, including function calling, browsing, and structured output generation.