Models

Kimi K3 (batch)

moonshotai ↗ · 1049K ctx · 23 providers · updated 2026-09-19 00:16 UTC

Cheapest
$4.19 via OpenRouter
Avg price
$5.85 /1M
providers
23
Tokens (24h)
230B
Volume (24h)
$965.1K est.
(3 × $1.95 + $10.92) ÷ 4 = $4.19/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $1.95, 3:1 $4.19, 1:1 $6.43, output only $10.92 per 1M.

The board shows this model's cheapest measured offer ($4.19/1M); the smaller figure under it is the mean of the 35 offers listed below ($5.85/1M). Add those up and divide — it matches. 24h token volume (230B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here. Cross-check — Vercel AI Gateway: 2.49% of gateway token volume, 9.13% of spend (2026-09-18; share-of-traffic, daily, data CC BY 4.0).

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
Morph (fp8) OpenRouter $1.95$10.92 5934ms 7 98.6% API-measured
Sail Research (fp4) OpenRouter $2.14$10.76 1744ms 64 99.7% API-measured
openrouter Pricebook (LiteLLM) $2.10$10.95 Official list
InferenceNet (fp4) OpenRouter $2.10$10.95 1870ms 31 99.7% API-measured
Relace (fp4) OpenRouter $2.10$10.95 417ms 101 99.1% API-measured
Wafer OpenRouter $2.50$10.95 345ms 22 100% API-measured
Phala OpenRouter $2.40$12.00 4232ms 21 100.0% API-measured
Makora OpenRouter $2.55$12.75 824ms 32 97.6% API-measured
DigitalOcean OpenRouter $2.55$12.95 2364ms 29 99.6% API-measured
deepinfra Pricebook (LiteLLM) $2.85$14.25 Official list
deepinfra HuggingFace Router $2.85$14.25 517ms 14 API-measured
DeepInfra (bf16) OpenRouter $2.85$14.25 2924ms 21 88.1% API-measured
moonshot Pricebook (LiteLLM) $3.00$15.00 Official list
nebius Pricebook (LiteLLM) $3.00$15.00 Official list
together_ai Pricebook (LiteLLM) $3.00$15.00 Official list
fireworks_ai Pricebook (LiteLLM) $3.00$15.00 Official list
novita Pricebook (LiteLLM) $3.00$15.00 Official list
aihubmix Pricebook (LiteLLM) $3.00$15.00 Official list
together HuggingFace Router $3.00$15.00 706ms 91 API-measured
fireworks-ai HuggingFace Router $3.00$15.00 890ms 40 API-measured
baseten HuggingFace Router $3.00$15.00 537ms 73 API-measured
0g-teetls 0G Router $3.00$15.00 API-measured
novita-ai Novita AI $3.00$15.00 API-measured
requesty Requesty $3.00$15.00 API-measured
Chutes (mxfp4) OpenRouter $3.00$15.00 3015ms 26 99.2% API-measured
Parasail (fp4) OpenRouter $3.00$15.00 1070ms 68 97.9% API-measured
Modal (mxfp4) OpenRouter $3.00$15.00 1465ms 52 99.9% API-measured
Together OpenRouter $3.00$15.00 1093ms 67 99.9% API-measured
Fireworks OpenRouter $3.00$15.00 1563ms 35 100% API-measured
BaseTen (fp8) OpenRouter $3.00$15.00 3154ms 34 100% API-measured
Moonshot AI (mxfp4) OpenRouter $3.00$15.00 4055ms 31 100.0% API-measured
Fireworks (us) OpenRouter $3.30$16.50 2108ms 49 98.6% API-measured
Alibaba OpenRouter $3.45$17.25 3187ms 32 98.6% API-measured
Fireworks (fast) OpenRouter $4.50$22.50 864ms 33 99.9% API-measured
Morph (fast) OpenRouter $6.00$22.50 7492ms 6 99.5% API-measured

Volume trend

341B 30d d-29 · 210B tokensd-28 · 214B tokensd-27 · 172B tokensd-26 · 160B tokensd-25 · 225B tokensd-24 · 237B tokensd-23 · 238B tokensd-22 · 248B tokensd-21 · 226B tokensd-20 · 170B tokensd-19 · 231B tokensd-18 · 241B tokensd-17 · 305B tokensd-16 · 296B tokensd-15 · 341B tokensd-14 · 321B tokensd-13 · 296B tokensd-12 · 180B tokensd-11 · 224B tokensd-10 · 259B tokensd-9 · 254B tokensd-8 · 197B tokensd-7 · 203B tokensd-6 · 167B tokensd-5 · 167B tokensd-4 · 204B tokensd-3 · 204B tokensd-2 · 226B tokensd-1 · 222B tokensd-0 · 230B tokens

Last 30 days: 210B → 230B (+10%); range 160B–341B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/kimi-k3.json. Free reuse requires attribution to tkx.org.

About

Kimi K3 is a machine-learning system that serves as a large language model, designed to process and generate human-like text based on input data. It is listed under the organization tag "moonshotai" and is available on AI model marketplaces for various inference providers to utilize.

AI-generated summary from public information — is this yours?

FAQ

How much does Kimi K3 (batch) cost?

Cheapest measured offer right now: $1.95/1M input, $10.92/1M output via Morph (fp8) — across 23 tracked providers, refreshed hourly.

Who serves Kimi K3 (batch)?

23 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is Kimi K3 (batch) used?

Kimi K3 (batch) routed 230B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings