Exact model · openai/gpt-oss-120b
GPT-OSS 120B API pricing and monthly cost by provider
Enter your monthly token workload and compare the exact openai/gpt-oss-120b route across every provider with complete matching evidence.
Exact model · openai/gpt-oss-120b
GPT-OSS 120B monthly cost comparison
Enter your monthly token workload and compare the exact openai/gpt-oss-120b route across every provider with complete matching evidence.
| Rank status | Provider | Input price | Cache-read price | Output price | Monthly cost | Exact configuration | Sources and freshness |
|---|---|---|---|---|---|---|---|
| —Evidence is aging or stale | Fireworks AIaggregator | $0.15 / 1M input tokens | $0.015 / 1M cached input tokens | $0.6 / 1M output tokens | UnavailableEvidence is aging or stale | accounts/fireworks/models/gpt-oss-120b · serverless standard · chat completions |
|
| —Evidence is aging or stale | OpenRouterrouter | $0.037 / 1M input tokens | $0.03 / 1M cached input tokens | $0.17 / 1M output tokens | UnavailableEvidence is aging or stale | openai/gpt-oss-120b · standard · chat completions |
|
| —Evidence is aging or stale | Together AIaggregator | $0.15 / 1M input tokens | Not published | $0.6 / 1M output tokens | UnavailableEvidence is aging or stale | openai/gpt-oss-120b · serverless · chat completions |
|
Missing and excluded providers (6)
- Fireworks AI: Evidence is aging or stale.
- OpenRouter: Evidence is aging or stale.
- Together AI: Evidence is aging or stale.
- OpenAI direct: No accepted owner-hosted GPT-OSS 120B API price is on file; the open model release is not treated as a hosted API offer.
- DeepInfra: No accepted exact-route price observation was found in this review.
- Novita AI: No accepted exact-route price observation was found in this review.
Strict boundary: Shared hosted routes only. Dedicated capacity, throughput, latency, rate limits, retries, taxes, support, and service quality are outside this price order. Together has no accepted cached-input meter, so it is excluded when cached usage is greater than zero.
Embeddable result
Put this exact-model comparison on your site
The SVG card is cookie-free and links back to the same source-auditable workload. It shows a rank only while at least three routes qualify.
<script async referrerpolicy="no-referrer" src="https://apidir.dev/embed/model-cost.js?model=gpt-oss-120b&input=100000000&cached=0&output=20000000"></script>
Exact-model alert · controlled rollout
Join the GPT-OSS 120B price alert waitlist.
Confirm your address to join the waitlist. Delivery remains paused until sender, suppression, and evidence checks are ready; ambiguous or stale changes stay quarantined.
Evidence boundary
What the monthly comparison proves
Shared hosted routes only. Dedicated capacity, throughput, latency, rate limits, retries, taxes, support, and service quality are outside this price order. Together has no accepted cached-input meter, so it is excluded when cached usage is greater than zero.
A row qualifies only when the exact model route, channel, region, mode, tier, and input/output meters match the reviewed definition. Cached input becomes mandatory when your cached-input volume is above zero.
Questions developers ask
Which exact GPT-OSS 120B route is compared?
This page is locked to openai/gpt-oss-120b. Family-only names, near matches, different model versions, and unreviewed aliases are excluded.
When is a lowest-cost label shown?
Only when at least three unique providers have complete fresh evidence for every token meter used by the submitted workload. Equal monthly costs share a rank.
What is excluded from monthly cost?
Retries, failed requests, networking, storage, taxes, support, dedicated capacity, latency, rate limits, and provider-specific service quality are outside this token-meter estimate.