Models

gpt-4o-mini

azure ↗ · 131K ctx · 7 providers · updated 2026-08-05 00:15 UTC

Cheapest
$0.263 Official list · Azure
Avg price
$0.265 /1M
providers
7
Tokens (24h)
32B
Volume (24h)
$8.4K est.
(3 × $0.150 + $0.600) ÷ 4 = $0.263/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.150, 3:1 $0.262, 1:1 $0.375, output only $0.600 per 1M.

The board shows this model's cheapest measured offer ($0.263/1M); the smaller figure under it is the mean of the 10 offers listed below ($0.265/1M). Add those up and divide — it matches. 24h token volume (32B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
azure Pricebook (LiteLLM) $0.150$0.600 Official list
gmi Pricebook (LiteLLM) $0.150$0.600 Official list
openai Pricebook (LiteLLM) $0.150$0.600 Official list
replicate Pricebook (LiteLLM) $0.150$0.600 Official list
vercel_ai_gateway Pricebook (LiteLLM) $0.150$0.600 Official list
requesty Requesty $0.150$0.600 API-measured
burncloud burncloud $0.150$0.600 API-measured
Azure OpenRouter $0.150$0.600 1166ms 23 100.0% API-measured
OpenAI OpenRouter $0.150$0.600 507ms 32 99.9% API-measured
Azure (swedencentral) OpenRouter $0.165$0.660 100% API-measured

Who burns this model

Top apps routing traffic to this model over the last 30 days, via OpenRouter's public stats. Only apps that opted into tracking appear, and only the top few are published — this is not the full demand picture. These figures are a 30-day window and are NOT comparable with the cumulative totals on the Apps board.

AppTokens (30d)Requests
Hermes Agent 12B 536K
pi 10B 152K
OpenClaw 9B 326K
unit-reason-graph 9B 1M
Portkey AI 9B 477K

Volume trend

88B 30d d-29 · 88B tokensd-28 · 48B tokensd-27 · 48B tokensd-26 · 35B tokensd-25 · 35B tokensd-24 · 25B tokensd-23 · 24B tokensd-22 · 30B tokensd-21 · 43B tokensd-20 · 41B tokensd-19 · 35B tokensd-18 · 35B tokensd-17 · 36B tokensd-16 · 62B tokensd-15 · 48B tokensd-14 · 47B tokensd-13 · 35B tokensd-12 · 36B tokensd-11 · 37B tokensd-10 · 23B tokensd-9 · 25B tokensd-8 · 35B tokensd-7 · 35B tokensd-6 · 35B tokensd-5 · 33B tokensd-4 · 30B tokensd-3 · 23B tokensd-2 · 25B tokensd-1 · 33B tokensd-0 · 32B tokens

Last 30 days: 88B → 32B (-64%); range 23B–88B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/gpt-4o-mini.json. Free reuse requires attribution to tkx.org.

About

GPT-4O-Mini is a language model designed for natural language processing tasks, served by multiple inference providers on AI model marketplaces and listed under the Azure serving-platform tag.

AI-generated summary from public information — is this yours?

FAQ

How much does gpt-4o-mini cost?

Cheapest measured offer right now: $0.15/1M input, $0.6/1M output via azure — across 7 tracked providers, refreshed hourly.

Who serves gpt-4o-mini?

7 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is gpt-4o-mini used?

gpt-4o-mini routed 32B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings