OpenAI logo

OpenAI: GPT OSS 120B

Open GPT reasoning model for self-hosted agents and controllable deployments

Source-linked gpt-oss Open weights Released 2025-08-05
API record Report
InputT
OutputT
Input price$0.03/M
Output price$0.17/M
Context131.072K
Max output32.768K
Providers67
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering GPT OSS 120B
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Nvidia openai/gpt-oss-120b 128K 8.192K ReasoningToolsJSON Docs ↗
Pendra gpt-oss:120b 131.072K 32.768K ReasoningToolsJSON Docs ↗
Kenari gpt-oss-120b 131.072K 32.768K ReasoningToolsJSON Docs ↗
QVAC gpt-oss-120b 131.072K 32.768K ReasoningToolsJSON Docs ↗
Weights & Biases openai/gpt-oss-120b 131.072K 131.072K $0.03 $0.17 $0.03 ReasoningToolsJSON Docs ↗
Kilo Gateway openai/gpt-oss-120b 131.072K 117.964K $0.03 $0.17 $0.03 ReasoningToolsJSON Docs ↗
OrcaRouter openai/gpt-oss-120b 131.072K 32.768K $0.03 $0.17 ReasoningToolsJSON Docs ↗
DevPass (LLM Gateway) gpt-oss-120b 131.072K 32.766K $0.032 $0.14 $0.032 ReasoningToolsJSON Docs ↗
OpenRouter openai/gpt-oss-120b 131.072K 117.964K $0.037 $0.17 ReasoningToolsJSON Docs ↗
Eden AI deepinfra/openai/gpt-oss-120b 131.072K 32.768K $0.037 $0.17 ReasoningToolsJSON Docs ↗
Deep Infra openai/gpt-oss-120b 131.072K 16.384K $0.037 $0.17 ReasoningToolsJSON Docs ↗
Eden AI flexai/gpt-oss-120b 131.072K 32.768K $0.039 $0.1 Reasoning Docs ↗
TensorX openai/gpt-oss-120b 131.072K 32.768K $0.04 $0.2 $0.01 ReasoningToolsJSON Docs ↗
IO.NET openai/gpt-oss-120b 131.072K 4.096K $0.04 $0.4 $0.02 Tools Docs ↗
Crusoe openai/gpt-oss-120b 131.072K 32.768K $0.05 $0.2 $0.05 ReasoningToolsJSON Docs ↗
NovitaAI openai/gpt-oss-120b 131.072K 32.768K $0.05 $0.25 ReasoningToolsJSON Docs ↗
SiliconFlow openai/gpt-oss-120b 131K 8K $0.05 $0.45 ReasoningToolsJSON Docs ↗
Databricks databricks-gpt-oss-120b 131.072K 32.768K $0.072 $0.28 ReasoningToolsJSON Docs ↗
Abacus openai/gpt-oss-120b 128K 16.384K $0.08 $0.44 ReasoningToolsJSON Docs ↗
Cortecs gpt-oss-120b 131K 131K $0.089 $0.446 $0.01 ReasoningToolsJSON Docs ↗
Eden AI ovhcloud/gpt-oss-120b 131.072K 32.768K $0.09 $0.47 ReasoningJSON Docs ↗
Merge Gateway openai/gpt-oss-120b 131.072K 32.768K $0.09 $0.36 Reasoning Docs ↗
Vertex openai/gpt-oss-120b-maas 131.072K 32.768K $0.09 $0.36 ReasoningTools Docs ↗
Vercel AI Gateway openai/gpt-oss-120b 131.072K 131.072K $0.1 $0.5 ReasoningToolsJSON Docs ↗
Synthetic hf:openai/gpt-oss-120b 131.072K 32.768K $0.1 $0.1 $0.1 ReasoningTools Docs ↗
submodel openai/gpt-oss-120b 131.072K 32.768K $0.1 $0.5 ReasoningTools Docs ↗
Baseten openai/gpt-oss-120b 128.072K 128.072K $0.1 $0.5 ReasoningToolsJSON Docs ↗
OpenReason openai/gpt-oss-120b 131.072K 32.768K $0.105 $0.422 ReasoningToolsJSON Docs ↗
CoralBricks gpt-oss-120b 131.072K 32.768K $0.12 $0.6 ReasoningToolsJSON Docs ↗
Eden AI fireworks_ai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.014 ReasoningTools Docs ↗
Impossibl fireworks/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.015 ReasoningToolsJSON Docs ↗
Impossibl groq/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.075 ReasoningToolsJSON Docs ↗
FastRouter openai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 ReasoningTools Docs ↗
NEAR AI Cloud openai/gpt-oss-120b 131K 32.768K $0.15 $0.55 ReasoningToolsJSON Docs ↗
ai& openai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 ReasoningToolsJSON Docs ↗
Pioneer openai/gpt-oss-120b 131.072K 131.072K $0.15 $0.6 $0.015 ReasoningToolsJSON Docs ↗
Eden AI fireworks_ai/accounts/fireworks/models/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.015 ReasoningToolsJSON Docs ↗
Nebius Token Factory openai/gpt-oss-120b 131.072K 8.192K $0.15 $0.6 $0.015 ReasoningToolsJSON Docs ↗
Fireworks AI accounts/fireworks/models/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.015 ReasoningTools Docs ↗
Groq openai/gpt-oss-120b 131.072K 65.536K $0.15 $0.6 $0.075 ReasoningToolsJSON Docs ↗
FrogBot gpt-oss-120b 131.072K 32.768K $0.15 $0.6 ReasoningToolsJSON Docs ↗
AKI.IO gpt-oss-120b 128K 32.768K $0.15 $0.55 ReasoningToolsJSON Docs ↗
Eden AI nebius/openai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.15 ReasoningToolsJSON Docs ↗
Neon gpt-oss-120b 131.072K 25K $0.15 $0.6 ReasoningToolsJSON Docs ↗
Eden AI together_ai/openai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 ReasoningToolsJSON Docs ↗
Eden AI groq/openai/gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.075 ReasoningToolsJSON Docs ↗
Together AI openai/gpt-oss-120b 131.072K 131.072K $0.15 $0.6 ReasoningTools Docs ↗
Eden AI databricks/databricks-gpt-oss-120b 131.072K 32.768K $0.15 $0.6 $0.015 ReasoningTools Docs ↗
watsonx.ai openai/gpt-oss-120b 131.072K 32.768K $0.159 $0.636 ReasoningToolsJSON Docs ↗
SCX.ai gpt-oss-120b 131.072K 131.072K $0.17 $0.55 ReasoningToolsJSON Docs ↗
SCX.ai gpt-oss-120b 131.072K 131.072K $0.17 $0.55 ReasoningToolsJSON Docs ↗
Eden AI ionos/openai/gpt-oss-120b 131.072K 32.768K $0.174 $0.755 Reasoning Docs ↗
Eden AI scaleway/gpt-oss-120b 128K 32.768K $0.174 $0.697 ReasoningTools Docs ↗
Charm Hyper gpt-oss-120b 128.072K 13.107K $0.178 $0.68 $0.089 ReasoningToolsJSON Docs ↗
Berget.AI openai/gpt-oss-120b 128K 8.192K $0.22 $0.83 ReasoningToolsJSON Docs ↗
GreenPT gpt-oss-120b 131.072K 32.768K $0.228 $0.798 ReasoningToolsJSON Docs ↗
evroc openai/gpt-oss-120b 65.536K 65.536K $0.23 $0.92 ReasoningTools Docs ↗
Hugging Face openai/gpt-oss-120b 131.072K 32.768K $0.25 $0.69 ReasoningToolsJSON Docs ↗
Cerebras gpt-oss-120b 131.072K 40.96K $0.35 $0.75 ReasoningToolsJSON Docs ↗
Eden AI cerebras/gpt-oss-120b 131.072K 32.768K $0.35 $0.75 $0.35 ReasoningToolsJSON Docs ↗
Impossibl cerebras/gpt-oss-120b 131.072K 32.768K $0.35 $0.75 ReasoningToolsJSON Docs ↗
NanoGPT openai/gpt-oss-120b 128K 16.384K $0.35 $0.75 ReasoningToolsJSON Docs ↗
Cloudflare Workers AI @cf/openai/gpt-oss-120b 128K 16.384K $0.35 $0.75 ReasoningToolsJSON Docs ↗
Cloudflare AI Gateway workers-ai/@cf/openai/gpt-oss-120b 128K 16.384K $0.35 $0.75 ReasoningToolsJSON Docs ↗
Eden AI cloudflare/@cf/openai/gpt-oss-120b 128K 32.768K $0.35 $0.75 ReasoningTools Docs ↗
STACKIT openai/gpt-oss-120b 131K 8.192K $0.53 $0.76 ReasoningTools Docs ↗
CloudFerro Sherlock openai/gpt-oss-120b 131K 131K $2.92 $2.92 ReasoningToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.03/M

Across 63 priced providers

Median input $0.15/M

Midpoint of listed rates

Highest input $2.92/M

97.3× the lowest listed rate

Output range $0.1 – $2.92

Per million output tokens

Weights & Biases $0.03/M
Kilo Gateway $0.03/M
OrcaRouter $0.03/M
DevPass (LLM Gateway) $0.032/M
OpenRouter $0.037/M
Eden AI $0.037/M
Deep Infra $0.037/M
Eden AI $0.039/M
TensorX $0.04/M
IO.NET $0.04/M
Crusoe $0.05/M
NovitaAI $0.05/M
SiliconFlow $0.05/M
Databricks $0.072/M
Abacus $0.08/M
Cortecs $0.089/M
Eden AI $0.09/M
Merge Gateway $0.09/M
Vertex $0.09/M
Vercel AI Gateway $0.1/M
Synthetic $0.1/M
submodel $0.1/M
Baseten $0.1/M
OpenReason $0.105/M
CoralBricks $0.12/M
Eden AI $0.15/M
Impossibl $0.15/M
Impossibl $0.15/M
FastRouter $0.15/M
NEAR AI Cloud $0.15/M
ai& $0.15/M
Pioneer $0.15/M
Eden AI $0.15/M
Nebius Token Factory $0.15/M
Fireworks AI $0.15/M
Groq $0.15/M
FrogBot $0.15/M
AKI.IO $0.15/M
Eden AI $0.15/M
Neon $0.15/M
Eden AI $0.15/M
Eden AI $0.15/M
Together AI $0.15/M
Eden AI $0.15/M
watsonx.ai $0.159/M
SCX.ai $0.17/M
SCX.ai $0.17/M
Eden AI $0.174/M
Eden AI $0.174/M
Charm Hyper $0.178/M
Berget.AI $0.22/M
GreenPT $0.228/M
evroc $0.23/M
Hugging Face $0.25/M
Cerebras $0.35/M
Eden AI $0.35/M
Impossibl $0.35/M
NanoGPT $0.35/M
Cloudflare Workers AI $0.35/M
Cloudflare AI Gateway $0.35/M
Eden AI $0.35/M
STACKIT $0.53/M
CloudFerro Sherlock $2.92/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

3 providers list the identical $0.03 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Context limits also differ by provider, from 65.536K to 131.072K tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

Reasoning Yes
Tool calling Yes
Structured output Yes
× Attachments No
× Vision input No
Open weights Yes
Creator
OpenAI
Model family
gpt-oss
Knowledge cutoff
Not documented
License
Not documented
Release date
2025-08-05
Model ID
openai/gpt-oss-120b

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
Coding IndexCoding Index Agentic IndexAgentic Index IntelligenceIntelligence Index Design Arena: w…Design Arena: website Design Arena: c…Design Arena: codecategories Design Arena: g…Design Arena: gamedev Design Arena: d…Design Arena: dataviz Design Arena: u…Design Arena: uicomponent This model — Coding Index: 30.4 index (31.0th percentile) This model — Agentic Index: 6.2 index (29.5th percentile) This model — Intelligence Index: 12.3 index (17.7th percentile) This model — Design Arena: website: 980 elo (7.4th percentile) This model — Design Arena: codecategories: 978 elo (4.8th percentile) This model — Design Arena: gamedev: 1015 elo (11.4th percentile) This model — Design Arena: dataviz: 1002 elo (8.5th percentile) This model — Design Arena: uicomponent: 941 elo (5.0th percentile)

Each axis is this model's percentile among the 8 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

Coding Index index Source ↗
Coding Index #141 of 205
Agentic Index index Source ↗
Agentic Index #136 of 193
Intelligence Index index Source ↗
Intelligence Index #158 of 192
Design Arena: website elo Source ↗
Design Arena: website #162 of 175
Design Arena: codecategories elo Source ↗
Design Arena: codecategories #158 of 167
Design Arena: gamedev elo Source ↗
Design Arena: gamedev #147 of 166
Design Arena: dataviz elo Source ↗
Design Arena: dataviz #151 of 165
Design Arena: uicomponent elo Source ↗
Design Arena: uicomponent #153 of 161
Design Arena: 3d elo Source ↗
Design Arena: 3d #146 of 157

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Coding Index, the benchmark with the widest published coverage that this model appears in.

0 43 86 $0.01 $0.1 $1 $10 Qwen: Qwen3.7 Plus — $0.32/M, score 55.9 OpenAI: GPT-5.4 (batch) — $1.25/M, score 71.1 Qwen: Qwen3.6 Plus — $0.325/M, score 54.5 Qwen: Qwen3 Next 80B A3B Thinking — $0.15/M, score 17.4 OpenAI: GPT-5.1 (batch) — $0.625/M, score 49.4 Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 43.9 Z.ai: GLM 4.6 — $0.43/M, score 45.8 OpenAI: GPT-4 Turbo (batch) — $5/M, score 21.5 Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 52.1 GPT-5 Mini — $0.25/M, score 15.6 Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 9.7 Anthropic: Claude Sonnet 4 — $3/M, score 37.6 MiMo-V2.5-Pro — $0.435/M, score 60.2 Gemini 3.1 Pro Preview — $2/M, score 68.8 OpenAI: o3 Mini High (batch) — $0.55/M, score 16.3 DeepSeek V4 Flash 0731 — $0.05/M, score 69.1 GPT-4o (2024-05-13) — $5/M, score 24.2 GPT OSS 20B — $0.02/M, score 20.7 GPT-4 Turbo — $10/M, score 21.5 o1 — $15/M, score 39.7 Qwen: Qwen3.8 Max (0902) — $2/M, score 71.8 Gemini 3.5 Flash — $1.5/M, score 70.1 Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 71.5 Gemma 3 12B IT — $0.05/M, score 5.8 GPT-4 — $30/M, score 13.1 GPT-4.1 mini — $0.4/M, score 20.2 DeepSeek V4 Pro — $0.435/M, score 59.4 Claude Sonnet 5 — $2/M, score 71.5 Mistral: Ministral 3 8B 2512 — $0.15/M, score 9.7 Gemini 3.1 Flash Lite Preview — $0.25/M, score 34.7 OpenAI: o1 (batch) — $7.5/M, score 39.7 Anthropic: Claude Fable 5.1 (batch) — $5/M, score 81.6 Gemini 3.7 Flash — $0.75/M, score 76.1 LongCat-2.0 — $0.3/M, score 45.3 Z.ai: GLM 5.1 — $0.966/M, score 55.8 Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 74.3 Mistral: Mistral Large 3 2512 — $0.5/M, score 20.1 DeepSeek V3.2 — $0.18/M, score 44.2 Anthropic: Claude Opus 4.7 — $5/M, score 73.6 Google: Gemini 3.5 Flash (batch) — $0.75/M, score 70.1 OpenAI: GPT-5.4 Nano (batch) — $0.1/M, score 56.1 Inception: Mercury 2 — $0.25/M, score 31.1 OpenAI: GPT-4o-mini (batch) — $0.075/M, score 11.4 Qwen: Qwen3.5-9B (batch) — $0.17/M, score 28.7 Gemini 3.6 Flash — $0.75/M, score 69.2 Gemma 4 26B A4B IT — $0.042/M, score 39.3 Claude Fable 5 — $10/M, score 76.5 Gemma 4 31B IT — $0.09/M, score 43.4 Z.ai: GLM 5.3 Flash — $0.15/M, score 71.5 OpenAI: GPT-3.5 Turbo (batch) — $0.25/M, score 10.7 Gemma 3 27B IT — $0.08/M, score 10.1 MiniMax: MiniMax M3 — $0.3/M, score 58.6 DeepSeek: DeepSeek V4 Pro 0813 (batch) — $0.66/M, score 68.8 Thinking Machines: Inkling Small (batch) — $0.5/M, score 52.9 Anthropic: Claude Fable 5.1 — $10/M, score 81.6 Google: Gemma 4 31B (batch) — $0.39/M, score 43.4 Qwen: Qwen3.8 2.4T A95B (batch) — $2/M, score 71.9 Mistral: Devstral 2 2512 — $0.4/M, score 31.3 MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 60.8 Inkling — $1.87/M, score 52.1 OpenAI: gpt-oss-20b (batch) — $0.05/M, score 20.7 Muse Spark 1.1 — $1.25/M, score 71.3 MiMo-V2.5 — $0.14/M, score 56.8 Google: Gemini 3.5 Flash Lite (batch) — $0.15/M, score 49.3 Kimi K2 Thinking — $0.4/M, score 21.0 Mistral: Mistral Medium 3.5 (batch) — $0.75/M, score 46.9 GPT-5.6 Luna — $0.2/M, score 71.4 Qwen: Qwen3.8 Max (0803) — $2/M, score 68.9 Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 20.5 GPT-5.1 — $1.25/M, score 49.4 Qwen: Qwen3.8 27B — $0.42/M, score 68.1 Google: Gemini 3.8 Flash (batch) — $0.375/M, score 76.3 Google: Gemini 3.7 Flash (batch) — $0.375/M, score 76.1 Mistral: Mistral Small 4 (batch) — $0.075/M, score 26.6 inclusionAI: Ling 3.0 Flash VL — $0.06/M, score 57.0 Z.ai: GLM 5.3 (batch) — $0.7/M, score 74.8 DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 69.1 OpenAI: GPT-5.5 (batch) — $2.5/M, score 74.9 MoonshotAI: Kimi K3 (batch) — $3/M, score 76.2 Inkling Small — $0.45/M, score 52.9 Anthropic: Claude Sonnet 5 (batch) — $1/M, score 71.5 Anthropic: Claude Fable 5 (batch) — $5/M, score 76.5 Nemotron 3 Super 120B A12B — $0.2/M, score 37.7 IBM: Granite 4.2 8B — $0.06/M, score 22.4 OpenAI: GPT-5 (batch) — $0.625/M, score 37.8 Google: Gemini 2.5 Pro (batch) — $0.625/M, score 33.3 GPT OSS 120B — $0.03/M, score 30.4 OpenAI: gpt-oss-120b (batch) — $0.15/M, score 30.4 OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 20.2 OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 11.1 SpaceXAI: Grok 4.6 — $2/M, score 76.8 DeepSeek-R1 — $0.7/M, score 24.6 DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 43.5 OpenAI: GPT-6 Astra (batch) — $5/M, score 76.9 OpenAI: GPT-5.6 Terra (batch) — $1/M, score 76.7 Gemma 3 4B IT — $0.04/M, score 2.7 DeepSeek V4 Pro 0813 — $0.442/M, score 68.8 Z.ai: GLM 5.2 (batch) — $0.7/M, score 68.8 Google: Gemini 3.6 Flash (batch) — $0.375/M, score 69.2 GPT-6 Astra — $10/M, score 76.9 Nemotron 3 Nano 30B A3B — $0.05/M, score 14.4 Z.ai: GLM 5.3 — $1.4/M, score 74.8 Anthropic: Claude Opus 4.8 — $5/M, score 74.3 SpaceXAI: Grok 4.3 — $1.25/M, score 42.2 GPT-5.6 Terra — $2/M, score 76.7 MiniMax: MiniMax M3 (batch) — $0.3/M, score 58.6 Gemini 3.8 Flash — $0.75/M, score 76.3 NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 49.3 MiniMax: MiniMax M2.7 — $0.3/M, score 52.6 Kimi K2.6 — $0.95/M, score 61.8 Kimi K3 — $3/M, score 76.2 Claude Opus 5 — $5/M, score 78.0 Claude Opus 5 (batch) — $2.5/M, score 78.0 SpaceXAI: Grok 4.5 — $2/M, score 72.4 Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 20.1 Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 73.6 Anthropic: Claude Sonnet 4.5 — $3/M, score 52.1 Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 68.8 Mistral: Mistral Medium 3.1 — $0.4/M, score 20.5 Thinking Machines: Inkling (batch) — $1/M, score 52.1 GPT-5.5 — $5/M, score 74.9 Anthropic: Claude Sonnet 4.6 — $3/M, score 63.0 Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 63.0 GPT-5 — $1.25/M, score 37.8 OpenAI: GPT-5 Mini (batch) — $0.125/M, score 15.6 Gemini 2.5 Pro — $1.25/M, score 33.3 inclusionAI: Ling 3.0 Flash — $0.021/M, score 50.6 SpaceXAI: Grok 4.3 (batch) — $1/M, score 42.2 Nemotron 3 Ultra 550B A55B — $0.5/M, score 49.3 OpenAI: GPT-5.6 Luna (batch) — $0.1/M, score 71.4 OpenAI: GPT-5.4 Mini (batch) — $0.375/M, score 56.1 Qwen: Qwen3.8 2.4T A95B — $2/M, score 71.9 OpenAI: GPT-5.6 Sol (batch) — $1/M, score 77.4 GPT-5.4 mini — $0.75/M, score 56.1 Kimi K2.7 Code — $0.95/M, score 60.8 Gemini 3.5 Flash Lite — $0.3/M, score 49.3 Hy3 preview — $0.066/M, score 58.8 Muse Spark 1.2 — $1.25/M, score 72.2 DeepSeek V4 Flash — $0.15/M, score 52.0 Nemotron 3.5 Lightning 30B A3B — $0.05/M, score 26.8 GPT-5.6 Sol — $4/M, score 77.4 GPT-4.1 nano — $0.1/M, score 11.1 GPT-5.4 — $2.5/M, score 71.1 GPT-4o mini — $0.15/M, score 11.4 GPT-5.4 nano — $0.2/M, score 56.1 GPT-3.5-turbo — $0.5/M, score 10.7 Solar Pro 4 — $0.3/M, score 52.7 Kimi K2.5 — $0.3/M, score 46.8 Trinity Large Thinking — $0.25/M, score 25.8 Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 12.1 Z.ai: GLM 5.2 — $0.966/M, score 68.8 Qwen: Qwen3.7 Max — $1.475/M, score 66.0 SpaceXAI: Grok Build 0.1 — $1/M, score 51.5 Qwen: Qwen3.6 35B A3B — $0.1/M, score 41.9 Qwen: Qwen3.6 27B — $0.3/M, score 53.7 inclusionAI: Ling-2.6-flash — $0.01/M, score 25.3 Mistral: Mistral Small 4 — $0.15/M, score 26.6 Kwaipilot: KAT-Coder-Pro V2 — $0.3/M, score 59.5 Qwen: Qwen3.5-9B — $0.1/M, score 28.7 Qwen: Qwen3.5-35B-A3B — $0.312/M, score 37.0 Qwen: Qwen3.5-122B-A10B — $0.26/M, score 45.7 Upstage: Solar Pro 3 — $0.15/M, score 16.2 Z.ai: GLM 4.7 — $0.4/M, score 45.3 Amazon: Nova 2 Lite — $0.3/M, score 23.0 Mistral: Ministral 3 3B 2512 — $0.1/M, score 4.8 Google: Gemma 3n 4B — $0.06/M, score 3.2 Meta: Llama 4 Maverick — $0.2/M, score 16.3 OpenAI: o3 Mini High — $1.1/M, score 16.3 Meta: Llama 3.3 70B Instruct — $0.1/M, score 11.9 Meta: Llama 3.1 8B Instruct — $0.05/M, score 5.4 Nex AGI: Nex-N2-Pro — $0.25/M, score 59.1 Mistral: Mistral Medium 3.5 — $1.5/M, score 46.9 inclusionAI: Ring-2.6-1T — $0.075/M, score 42.8 IBM: Granite 4.1 8B — $0.05/M, score 9.5 Qwen: Qwen3.5 397B A17B — $0.55/M, score 48.2 Qwen: Qwen3 Coder Next — $0.12/M, score 36.2 Mistral: Ministral 3 14B 2512 — $0.2/M, score 14.4 Anthropic: Claude Haiku 4.5 — $1/M, score 43.9 Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 22.1 Qwen: Qwen3 8B — $0.117/M, score 9.0 Qwen: Qwen3 14B — $0.227/M, score 13.8 Qwen: Qwen3 32B — $0.08/M, score 15.3 Meta: Llama 4 Scout — $0.1/M, score 8.2 DeepSeek: DeepSeek V3 0324 — $0.25/M, score 21.2 Cohere: Command A — $2.5/M, score 27.8 Step 3.7 Flash — $0.185/M, score 39.6 GPT OSS 120B Input price per million tokens (log scale) Index

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Max Output Tokens117964 → 32768
Price Completion0.16999999999999998 → 0.17
Price Prompt0.037 → 0.03
Max Output Tokens32768 → 117964
Price Completion0.17 → 0.16999999999999998
Price Prompt0.03 → 0.037
Max Output Tokens117964 → 32768
Price Completion0.16999999999999998 → 0.17
Price Prompt0.037 → 0.03
Max Output Tokens32768 → 117964
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/openai/gpt-oss-120b
curl
curl "https://model.kyssta.lol/api/v1/models/openai/gpt-oss-120b"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is GPT OSS 120B?

Open GPT reasoning model for self-hosted agents and controllable deployments. It is published by OpenAI and catalogued here from Models.dev.

How much does GPT OSS 120B cost?

Listed input pricing starts at $0.03 per million tokens from Weights & Biases, rising to $2.92 across 63 listed providers.

What is the context length of GPT OSS 120B?

GPT OSS 120B accepts up to 131.072K tokens of context and returns up to 32.768K output tokens.

Does GPT OSS 120B support tool calling and structured output?

Provider catalogs list support for tool calling, structured output, and reasoning.

Which providers serve GPT OSS 120B?

52 providers list this model: Nvidia, Pendra, Kenari, QVAC, Weights & Biases, Kilo Gateway and 46 more.

Are the weights for GPT OSS 120B open?

Yes. The weights are published and downloadable from Hugging Face.

When was GPT OSS 120B released?

The catalog records a release date of 2025-08-05, last verified Sep 11, 2026.

More models from OpenAI

View all →