Models

DeepSeek V4.1 Flash

deepseek ↗ · 1049K ctx · 4 providers · updated 2026-09-12 02:15 UTC

Cheapest
$0.263 via OpenRouter
Avg price
$0.433 /1M
providers
4
Tokens (24h)
1.5T
Volume (24h)
$403.3K est.
(3 × $0.150 + $0.600) ÷ 4 = $0.263/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.150, 3:1 $0.262, 1:1 $0.375, output only $0.600 per 1M.

The board shows this model's cheapest measured offer ($0.263/1M); the smaller figure under it is the mean of the 5 offers listed below ($0.433/1M). Add those up and divide — it matches. 24h token volume (1.5T) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here. Cross-check — Vercel AI Gateway: 43.54% of gateway token volume, 1.94% of spend (2026-09-12; share-of-traffic, daily, data CC BY 4.0).

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
openrouter-best OpenRouter $0.150$0.600 API-measured
requesty Requesty $0.220$0.660 API-measured
novita HuggingFace Router $0.300$1.20 543ms 131 API-measured
baseten HuggingFace Router $0.300$1.20 375ms 191 API-measured
novita-ai Novita AI $0.300$1.20 API-measured

Volume trend

1.5T 2d d-1 · 679B tokensd-0 · 1.5T tokens

Last 2 days: 679B → 1.5T (+126%); range 679B–1.5T tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/deepseek-v4.1-flash.json. Free reuse requires attribution to tkx.org.

FAQ

How much does DeepSeek V4.1 Flash cost?

Cheapest measured offer right now: $0.15/1M input, $0.6/1M output via openrouter-best — across 4 tracked providers, refreshed hourly.

Who serves DeepSeek V4.1 Flash?

4 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is DeepSeek V4.1 Flash used?

DeepSeek V4.1 Flash routed 1.5T tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings