Models

gpt-oss-120b

runware · 131K ctx · 33 providers · updated 2026-09-19 02:15 UTC

Cheapest
$0.059 via Requesty
Avg price
$0.237 /1M
providers
33
Tokens (24h)
67B
Volume (24h)
$17.7K est.
(3 × $0.032 + $0.140) ÷ 4 = $0.059/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.032, 3:1 $0.059, 1:1 $0.086, output only $0.140 per 1M.

The board shows this model's cheapest measured offer ($0.059/1M); the smaller figure under it is the mean of the 55 offers listed below ($0.237/1M). Add those up and divide — it matches. 24h token volume (67B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
requesty Requesty $0.032$0.140 API-measured
wandb Pricebook (LiteLLM) $0.030$0.170 Official list
AkashML (bf16) OpenRouter $0.030$0.170 703ms 36 99.9% API-measured
CoreWeave (fp4) OpenRouter $0.030$0.170 274ms 81 100.0% API-measured
DekaLLM (bf16) OpenRouter $0.030$0.180 593ms 36 99.9% API-measured
deepinfra Pricebook (LiteLLM) $0.037$0.170 Official list
deepinfra HuggingFace Router $0.037$0.170 370ms 52 API-measured
DeepInfra (bf16) OpenRouter $0.037$0.170 420ms 47 100.0% API-measured
novita Pricebook (LiteLLM) $0.050$0.250 Official list
novita HuggingFace Router $0.050$0.250 516ms 162 API-measured
novita-ai Novita AI $0.050$0.250 API-measured
Crusoe (bf16) OpenRouter $0.050$0.250 325ms 220 99.7% API-measured
Novita (fp4) OpenRouter $0.050$0.250 690ms 109 100.0% API-measured
Mancer 2 (fp8) OpenRouter $0.050$0.300 448ms 61 100% API-measured
DigitalOcean OpenRouter $0.060$0.420 326ms 41 100.0% API-measured
Google (global) OpenRouter $0.090$0.360 3129ms 79 85.3% API-measured
ovhcloud Pricebook (LiteLLM) $0.080$0.400 Official list
nscale HuggingFace Router $0.100$0.400 661ms 69 API-measured
ovhcloud HuggingFace Router $0.090$0.470 346ms 149 API-measured
baseten Pricebook (LiteLLM) $0.100$0.500 Official list
baseten HuggingFace Router $0.100$0.500 415ms 206 API-measured
BaseTen (fp4) OpenRouter $0.100$0.500 203ms 208 99.9% API-measured
azure_ai Pricebook (LiteLLM) $0.150$0.600 Official list
fireworks_ai Pricebook (LiteLLM) $0.150$0.600 Official list
groq Pricebook (LiteLLM) $0.150$0.600 Official list
nebius Pricebook (LiteLLM) $0.150$0.600 Official list
openrouter Pricebook (LiteLLM) $0.150$0.600 Official list
together_ai Pricebook (LiteLLM) $0.150$0.600 Official list
scaleway Pricebook (LiteLLM) $0.150$0.600 Official list
tensormesh Pricebook (LiteLLM) $0.150$0.600 Official list
together HuggingFace Router $0.150$0.600 278ms 222 API-measured
fireworks-ai HuggingFace Router $0.150$0.600 206ms 128 API-measured
Amazon Bedrock (eu-west-1) OpenRouter $0.150$0.600 100% API-measured
Nebius (fp4) OpenRouter $0.150$0.600 197ms 267 100% API-measured
Amazon Bedrock OpenRouter $0.150$0.600 594ms 134 100% API-measured
DeepInfra (turbo) OpenRouter $0.150$0.600 419ms 136 99.8% API-measured
SiliconFlow (fp8) OpenRouter $0.150$0.600 1092ms 38 100% API-measured
Phala OpenRouter $0.150$0.600 876ms 114 100% API-measured
Together OpenRouter $0.150$0.600 245ms 159 92.2% API-measured
Groq OpenRouter $0.150$0.600 226ms 334 100.0% API-measured

Who burns this model

Top apps routing traffic to this model over the last 30 days, via OpenRouter's public stats. Only apps that opted into tracking appear, and only the top few are published — this is not the full demand picture. These figures are a 30-day window and are NOT comparable with the cumulative totals on the Apps board.

AppTokens (30d)Requests
Mira is the leading AI agent inside Telegram 113B 15M
cvi-magic-tasks-ai-tagging 37B 12M
IntelliProcure-NLP 17B 1M
Craft 13B 3M
MakeMyCV 10B 6M

Volume trend

151B 30d d-29 · 85B tokensd-28 · 70B tokensd-27 · 72B tokensd-26 · 54B tokensd-25 · 58B tokensd-24 · 65B tokensd-23 · 116B tokensd-22 · 83B tokensd-21 · 76B tokensd-20 · 36B tokensd-19 · 38B tokensd-18 · 65B tokensd-17 · 90B tokensd-16 · 82B tokensd-15 · 63B tokensd-14 · 64B tokensd-13 · 47B tokensd-12 · 51B tokensd-11 · 58B tokensd-10 · 74B tokensd-9 · 83B tokensd-8 · 92B tokensd-7 · 112B tokensd-6 · 72B tokensd-5 · 56B tokensd-4 · 84B tokensd-3 · 128B tokensd-2 · 151B tokensd-1 · 82B tokensd-0 · 67B tokens

Last 30 days: 85B → 67B (-21%); range 36B–151B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/gpt-oss-120b.json. Free reuse requires attribution to tkx.org.

About

GPT-OSS-120b is a large language model served by multiple inference providers on AI model marketplaces, known for its capabilities in natural language processing tasks.

AI-generated summary from public information — is this yours?

FAQ

How much does gpt-oss-120b cost?

Cheapest measured offer right now: $0.032/1M input, $0.14/1M output via requesty — across 33 tracked providers, refreshed hourly.

Who serves gpt-oss-120b?

33 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is gpt-oss-120b used?

gpt-oss-120b routed 67B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings