Exact model · openai/gpt-oss-120b

GPT-OSS 120B API pricing and monthly cost by provider

Enter your monthly token workload and compare the exact openai/gpt-oss-120b route across every provider with complete matching evidence.

Exact model · openai/gpt-oss-120b

GPT-OSS 120B monthly cost comparison

Enter your monthly token workload and compare the exact openai/gpt-oss-120b route across every provider with complete matching evidence.

No complete comparison
Versioned service or route ID; family-only names are excluded.
A route without a reviewed cache meter is excluded only when this value is above zero.
No complete comparison. No provider has complete, fresh evidence for this exact workload. APIDir does not estimate missing meters.

USD-normalized monthly cost for the entered workload; every meter is per 1 million tokens.
Rank statusProviderInput priceCache-read priceOutput priceMonthly costExact configurationSources and freshness
—Evidence is aging or staleFireworks AIaggregator$0.15 / 1M input tokens$0.015 / 1M cached input tokens$0.6 / 1M output tokensUnavailableEvidence is aging or staleaccounts/fireworks/models/gpt-oss-120b · serverless standard · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleOpenRouterrouter$0.037 / 1M input tokens$0.03 / 1M cached input tokens$0.17 / 1M output tokensUnavailableEvidence is aging or staleopenai/gpt-oss-120b · standard · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Cached input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
—Evidence is aging or staleTogether AIaggregator$0.15 / 1M input tokensNot published$0.6 / 1M output tokensUnavailableEvidence is aging or staleopenai/gpt-oss-120b · serverless · chat completions
  • Input · checked · fresh through staleFirst-party source
  • Output · checked · fresh through staleFirst-party source
Row eligibility uses only meters with non-zero workload.
Missing and excluded providers (6)
  • Fireworks AI: Evidence is aging or stale.
  • OpenRouter: Evidence is aging or stale.
  • Together AI: Evidence is aging or stale.
  • OpenAI direct: No accepted owner-hosted GPT-OSS 120B API price is on file; the open model release is not treated as a hosted API offer.
  • DeepInfra: No accepted exact-route price observation was found in this review.
  • Novita AI: No accepted exact-route price observation was found in this review.

Strict boundary: Shared hosted routes only. Dedicated capacity, throughput, latency, rate limits, retries, taxes, support, and service quality are outside this price order. Together has no accepted cached-input meter, so it is excluded when cached usage is greater than zero.

Embeddable result

Put this exact-model comparison on your site

The SVG card is cookie-free and links back to the same source-auditable workload. It shows a rank only while at least three routes qualify.

<script async referrerpolicy="no-referrer" src="https://apidir.dev/embed/model-cost.js?model=gpt-oss-120b&input=100000000&cached=0&output=20000000"></script>

One exact AI model compared across three provider routes with separate token meters and an incomplete evidence route
AI concept diagramConcept diagram: exact identity and complete meters decide eligibility. An incomplete route stays visible as a gap and cannot become a winner.

Exact-model alert · controlled rollout

Join the GPT-OSS 120B price alert waitlist.

Confirm your address to join the waitlist. Delivery remains paused until sender, suppression, and evidence checks are ready; ambiguous or stale changes stay quarantined.

Evidence boundary

What the monthly comparison proves

Shared hosted routes only. Dedicated capacity, throughput, latency, rate limits, retries, taxes, support, and service quality are outside this price order. Together has no accepted cached-input meter, so it is excluded when cached usage is greater than zero.

A row qualifies only when the exact model route, channel, region, mode, tier, and input/output meters match the reviewed definition. Cached input becomes mandatory when your cached-input volume is above zero.

Questions developers ask

Which exact GPT-OSS 120B route is compared?

This page is locked to openai/gpt-oss-120b. Family-only names, near matches, different model versions, and unreviewed aliases are excluded.

When is a lowest-cost label shown?

Only when at least three unique providers have complete fresh evidence for every token meter used by the submitted workload. Equal monthly costs share a rank.

What is excluded from monthly cost?

Retries, failed requests, networking, storage, taxes, support, dedicated capacity, latency, rate limits, and provider-specific service quality are outside this token-meter estimate.