GPT OSS 120B

apache-2.0120B

`gpt-oss-120b`is our most powerful open-weight model, which fits into a single H100 GPU (117B parameters with 5.1B active parameters).

Intelligence
Reasoning
Speed
Context
131K
Max output
66K
Input price
$0.15/1M tokens
Output price
$0.6/1M tokens

Capabilities

vision
tool call
structured output
reasoning
json mode
streaming
fine tuning
batch

Details

Model IDopenai/gpt-oss-120b
Provider Groq
Creator OpenAI
Familygpt-oss
Licenseapache-2.0
Parameters120B
Statusactive
Input modalitiestext
Output modalitiestext
Architecture—
Knowledge cutoff
Training data cutoff—
Release date—
Deprecation date—
Typechat
Reasoning tokensYes
Max input—
Open weightYes
Sourceofficial
Last updated

Tools

Function Callingfunction_calling
Call external functions and APIs

Pricing

Input
$0.15
Output
$0.6
Cache write
—
Cache read
—
Batch in
—
Batch out
—

Price across providers

ProviderInputOutputSince launch
DeepInfracheapest$0.04$0.17↓5%
OpenRoutercheapest$0.04$0.17↓5%
Cloudflare AI Gateway$0.04$0.19—
Novita AI$0.05$0.25—
Groq$0.15$0.6—
SambaNova Cloud$0.22$0.59—
Cerebras$0.35$0.75—
Vercel AI Gateway$0.35$0.75—

Family Comparison: gpt-oss

ModelContextMax outPricing
GPT OSS 20B131K66K$0.08$0.3
Safety GPT OSS 20B131K66K$0.08$0.3
GPT OSS 120B131K66K$0.15$0.6

API

GET/v1/models/groq/openai/gpt-oss-120b