Skip to content

gpt-oss-120b

OpenAI · Open weights · 131K context · released Aug 5, 2025 · checked 10h ago

Ways to buy 26

Way to buyIn / out per 1M
AkashML
bf16 · cached $0.03
$0.03 / $0.17Cheapest
CoreWeave
fp4 · cached $0.03
$0.03 / $0.17Cheapest
DekaLLM
bf16
$0.03 / $0.18+4% vs cheapest
DeepInfra
bf16
$0.037 / $0.17+8% vs cheapest
Crusoe
bf16 · cached $0.05
$0.05 / $0.25+54% vs cheapest
Novita AI
fp4
$0.05 / $0.25+54% vs cheapest
Mancer 2
fp8
$0.05 / $0.3+73% vs cheapest
DigitalOcean
Default region · cached $0.012
$0.06 / $0.42+131% vs cheapest
Google Vertex AI
global
$0.09 / $0.36+142% vs cheapest
Baseten
fp4 · cached $0.1
$0.1 / $0.5+208% vs cheapest
AWS Bedrock
Default region
$0.15 / $0.6+304% vs cheapest
AWS Bedrock
eu-west-1
$0.15 / $0.6+304% vs cheapest
Databricks Mosaic AI Gateway + Foundation Model APIs
From the platform pricing page
$0.15 / $0.6+304% vs cheapest
DeepInfra
turbo, bf16
$0.15 / $0.6+304% vs cheapest
Groq
Default region · cached $0.075
$0.15 / $0.6+304% vs cheapest
Nebius
fp4
$0.15 / $0.6+304% vs cheapest
OCI Generative AI
From the platform pricing page
$0.15 / $0.6+304% vs cheapest
Phala
Default region
$0.15 / $0.6+304% vs cheapest
SiliconFlow
fp8 · cached $0.075
$0.15 / $0.6+304% vs cheapest
Together AI
Default region
$0.15 / $0.6+304% vs cheapest
Parasail
fp4 · cached $0.055
$0.1 / $0.75+304% vs cheapest
IBM watsonx.ai
From the platform pricing page
$0.159 / $0.636+328% vs cheapest
Mara
Default region
$0.15 / $0.75+362% vs cheapest
SambaNova
Default region
$0.14 / $0.95+427% vs cheapest
DeepInfra
fp8
$0.2 / $0.95+496% vs cheapest
Cerebras
fp16 · cached $0.35
$0.35 / $0.75+592% vs cheapest

USD per 1M tokens. Regions are separate routes.

Through a gateway

Your own numbers

100M input and 20M output tokens a month: $6.40 direct.

No published token fee: Martian Gateway, Portkey, TrueFoundry AI Gateway.

Related models

In / out per 1M
GPT-6 Luna
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex)
$0.1 / $0.5
GPT-6 Luna Pro
OpenAI · 1.05MCheapest: $0.1 on OpenAI (flex)
$0.1 / $0.5
GPT-6 Sol
OpenAI · 1.05MCheapest: $2 on OpenAI (flex)
$2 / $10
GPT-6 Sol Pro
OpenAI · 1.05MCheapest: $2 on OpenAI (flex)
$2 / $10
Muse Spark 1.3 Contributor
Meta · 1M
$0.1 / $0.2
GLM 5.3 Flash
Z.ai · 1.31M
$0.045 / $0.14
Muse Spark 1.2 Contributor
Meta · 1M
$0.1 / $0.2
DeepSeek V4 Flash 0731
DeepSeek · 1.31M
$0.053 / $0.158

About gpt-oss-120b

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases.

Takes text in and returns text. First listed Sep 23, 2026. Up to 66K output tokens. Blended price $0.065 per 1M tokens at 3 input to 1 output. Model id gpt-oss-120b.

Price history

No list-price history: this model is sold by hosts rather than its lab.

Questions

How much does gpt-oss-120b cost?

$0.03 per million input tokens and $0.17 per million output tokens, with cached input at $0.03.

What is the context window of gpt-oss-120b?

131K tokens, with up to 66K tokens of output.

Where is gpt-oss-120b cheapest?

Of 26 routes tracked, AkashML is cheapest at $0.03 input and $0.17 output per million tokens.

List prices from providers' pages on the date shown; your contract may differ. Logos via logo.dev; trademarks belong to their owners.

Weekly: LLM API list prices that changed, Thursday mornings.