OpenAI · GPT OSS

gpt-oss-120b

An open-weight mixture-of-experts reasoning model designed for agentic and general-purpose workloads.

Context window131.1K
Providers5
Input from$0.030
Output from$0.17

Model Markets treats this immutable version as the base entity. Provider endpoints, pricing, quantization, regions, service tiers, and observations remain separate so the record can change without rewriting the model itself.

Provider comparison

Available endpoints

Comparison view →
ProviderEndpointQuant.Input / 1MOutput / 1MRegionEvidence
CoreWeaveAvailablecoreweave/fp4fp4$0.030$0.17Provider global routingProvider-reported
DeepInfraAvailabledeepinfra/bf16bf16$0.037$0.17Provider global routingProvider-reported
NovitaAvailablenovita/fp4fp4$0.050$0.25Provider global routingProvider-reported
DigitalOceanAvailabledigitaloceanunknown$0.055$0.385Provider global routingProvider-reported
Google Vertex AIAvailablegoogle-vertex/globalunknown$0.090$0.36Provider global routingProvider-reported

Model specification

Canonical IDopenai/gpt-oss-120b
ReleasedAug 5, 2025
Maximum output131.1K tokens
Modalitiestext
LicenseApache 2.0

Capability record

Capabilities below come from the cited provider catalog. Independent correctness probes are intentionally separate.

reasoningtoolsstructured outputs

Questions

gpt-oss-120b questions

How much does gpt-oss-120b cost?+

The lowest provider-reported input price in this snapshot is $0.030 per million tokens. Output prices begin at $0.17 per million tokens. Compare provider pricing

Which providers offer gpt-oss-120b?+

5 provider endpoints appear in this source snapshot: CoreWeave, DeepInfra, Novita, DigitalOcean, Google Vertex AI. Browse providers

Which provider is cheapest for gpt-oss-120b?+

CoreWeave has the lowest listed input price in this snapshot. This is a provider-reported price, not an execution quote.

Which provider is fastest for gpt-oss-120b?+

Model Markets does not yet have approved independent probe results for this model, so it does not name a fastest provider.

What is the gpt-oss-120b context window?+

The cited model record reports a 131.1K-token context window.

Does gpt-oss-120b support tools and structured output?+

The cited provider catalog reports reasoning, tools, structured outputs. These are provider-reported capabilities until independently tested.

Is gpt-oss-120b open-weight?+

gpt-oss-120b is recorded as open-weight under the Apache 2.0 license.

How current is this data?+

The source snapshot was observed Aug 18, 2026. The page is explicitly marked non-live until production collectors are activated.

How does Model Markets measure latency?+

Latency uses disclosed-region active probes, preserving model version, provider, region, workload, service tier, request parameters, and timestamp. No paid probes were run for this delivery. Read the methodology