xpu liveBETA
Models›gpt-oss-120b

gpt-oss-120b

Model reference ↗

Provider prices 4

Lowest input price per billing unit. Conditions in Details.

ProviderPriceProvider model IDStatusContext / max outputSource
OpenRouter
Router catalog
Input $0.037USD / 1M tokensOutput $0.17USD / 1M tokens
openai/gpt-oss-120bCatalog-listed
Conditions

Catalog-listed; account access and regional availability may apply.

Context tokens 131072Output tokens 65536Official source ↗2026-09-28
Groq
Standard
Input $0.15USD / 1M tokensOutput $0.6USD / 1M tokens
openai/gpt-oss-120bCatalog-listed
Conditions

Catalog-listed; account access and regional availability may apply.

Context tokens 131072Output tokens 65536Official source ↗2026-09-28
Together AI
Standard
Input $0.15USD / 1M tokensOutput $0.6USD / 1M tokens
openai/gpt-oss-120bCatalog-listed
Conditions

Catalog-listed; account access and regional availability may apply.

Context tokens 131072Official source ↗2026-09-28
Fireworks AINot collectedAwaiting verified pricingNot yet capturedCatalog-listed
Conditions

Serverless catalog listing verified; provider-specific API ID not yet captured.

Not suppliedOfficial source ↗2026-09-28

gpt-oss-120b accepts text and produces text. Tool calling and structured outputs are available on supported routes.

CONTEXT TOKENS131,072Developer documentation
MAX OUTPUT TOKENS131,072Developer documentation
PROVIDERS4Linked in this catalog
SOURCE CHECKED2026-09-29Sources reviewed; unknown fields labeled

Specifications

Catalog model ID
openai/gpt-oss-120b
Developer model ID / revision
gpt-oss-120bDeveloper documentation ↗
Input modalities
textDeveloper documentation ↗
Output modalities
textDeveloper documentation ↗
Context / input token limit
131,072Developer documentation ↗
Maximum output tokens
131,072Developer documentation ↗
Tool calling
SupportedDeveloper documentation ↗
Structured output
SupportedDeveloper documentation ↗
API parameters (source-specific)
frequency_penalty, include_reasoning, logit_bias, logprobs, max_tokens, min_p, presence_penalty, reasoning, reasoning_effort, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, tool_choice, tools, top_a, top_k, top_logprobs, top_pOpenRouter catalog ↗
API parameter scope
OpenRouter API; support varies by route
Weight access
Open weightsDeveloper documentation ↗
Weight license / access terms
Apache 2.0Developer documentation ↗
Total parameters (billions)
117Developer documentation ↗
Active parameters (billions)
5.1Developer documentation ↗
Download size (GB)
Depends on checkpoint and precision; no verified file size
Release date
Not verified in cited sources
Added to OpenRouter
2025-08-05OpenRouter catalog ↗
Notes
API parameter names are reported by OpenRouter and do not describe the developer’s native API. Token limits labeled OpenRouter are route catalog limits, not universal model limits. Media limits on this page describe the linked provider endpoint; other deployments may differ. Output tokens share the model context budget with input tokens. Hosted providers may impose lower output caps.

Catalog entries link to their specification and availability sources. This is a curated selection, not every model each provider offers. Best prices are the lowest collected USD tariffs, not a total-cost estimate. Token offers are selected by input price, with output from the same tariff; workload mix may change the cheapest provider. Listed variants may require batch processing, off-peak hours, or specific media settings. Price coverage varies by provider; collected tariffs include their source and observation time. Benchmark scores are a dated snapshot, not a live feed or a universal measure of quality. They apply to the shown evaluation configuration; listed prices may use different settings. Models without verified scores are not assumed to be weaker.