Models

gpt-5-mini:flex

openai ↗ · 400K ctx · 6 providers · updated 2026-08-05 00:15 UTC

Cheapest
$0.344 via Requesty
Avg price
$0.626 /1M
providers
6
Tokens (24h)
29B
Volume (24h)
$20.2K est.
(3 × $0.125 + $1.00) ÷ 4 = $0.344/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.125, 3:1 $0.344, 1:1 $0.563, output only $1.00 per 1M.

The board shows this model's cheapest measured offer ($0.344/1M); the smaller figure under it is the mean of the 10 offers listed below ($0.626/1M). Add those up and divide — it matches. 24h token volume (29B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
requesty Requesty $0.125$1.00 API-measured
OpenAI (flex) OpenRouter $0.125$1.00 4718ms 114 98.9% API-measured
azure Pricebook (LiteLLM) $0.250$2.00 Official list
openai Pricebook (LiteLLM) $0.250$2.00 Official list
openrouter Pricebook (LiteLLM) $0.250$2.00 Official list
replicate Pricebook (LiteLLM) $0.250$2.00 Official list
burncloud burncloud $0.250$2.00 API-measured
OpenAI OpenRouter $0.250$2.00 5242ms 74 98.9% API-measured
Azure OpenRouter $0.250$2.00 4471ms 80 100% API-measured
Azure (swedencentral) OpenRouter $0.275$2.20 100% API-measured

Volume trend

29B 30d d-29 · 22B tokensd-28 · 23B tokensd-27 · 21B tokensd-26 · 20B tokensd-25 · 21B tokensd-24 · 15B tokensd-23 · 13B tokensd-22 · 21B tokensd-21 · 28B tokensd-20 · 25B tokensd-19 · 24B tokensd-18 · 20B tokensd-17 · 14B tokensd-16 · 13B tokensd-15 · 22B tokensd-14 · 22B tokensd-13 · 0 tokensd-12 · 22B tokensd-11 · 27B tokensd-10 · 0 tokensd-9 · 21B tokensd-8 · 0 tokensd-7 · 24B tokensd-6 · 23B tokensd-5 · 0 tokensd-4 · 0 tokensd-3 · 18B tokensd-2 · 0 tokensd-1 · 29B tokensd-0 · 29B tokens

Last 30 days: 22B → 29B (+36%); range 0–29B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/gpt-5-mini.json. Free reuse requires attribution to tkx.org.

About

GPT-5-Mini is a language model designed for natural language processing tasks. It is served by multiple inference providers on AI model marketplaces and is listed under the organization tag "OpenAI".

AI-generated summary from public information — is this yours?

FAQ

How much does gpt-5-mini:flex cost?

Cheapest measured offer right now: $0.125/1M input, $1/1M output via requesty — across 6 tracked providers, refreshed hourly.

Who serves gpt-5-mini:flex?

6 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is gpt-5-mini:flex used?

gpt-5-mini:flex routed 29B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings