Models

DeepSeek V4 Flash 0423

deepseek ↗ · 1442K ctx · 30 providers · updated 2026-09-19 00:16 UTC

Cheapest
$0.061 via OpenRouter
Avg price
$0.196 /1M
providers
30
Tokens (24h)
2.4T
Volume (24h)
$101.8K est.
(3 × $0.048 + $0.097) ÷ 4 = $0.061/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.048, 3:1 $0.061, 1:1 $0.073, output only $0.097 per 1M.

The board shows this model's cheapest measured offer ($0.061/1M); the smaller figure under it is the mean of the 37 offers listed below ($0.196/1M). Add those up and divide — it matches. 24h token volume (2.4T) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
StreamLake (fp8) OpenRouter $0.048$0.097 1443ms 42 99.5% API-measured
Baidu (fp8) OpenRouter $0.049$0.097 698ms 82 77.8% API-measured
openrouter Pricebook (LiteLLM) $0.050$0.100 Official list
OpenInference (fp8) OpenRouter $0.050$0.140 1095ms 25 100.0% API-measured
deepinfra Pricebook (LiteLLM) $0.090$0.180 Official list
deepinfra HuggingFace Router $0.090$0.180 496ms 16 API-measured
DeepInfra (fp8) OpenRouter $0.090$0.180 7466ms 14 99.6% API-measured
GMICloud (fp8) OpenRouter $0.091$0.182 3169ms 30 100.0% API-measured
Venice OpenRouter $0.097$0.193 5514ms 17 96.6% API-measured
DigitalOcean OpenRouter $0.098$0.196 1045ms 14 100.0% API-measured
pinstripes Pricebook (LiteLLM) $0.100$0.200 Official list
SiliconFlow (fp8) OpenRouter $0.130$0.280 1733ms 39 98.5% API-measured
Alibaba (fp8) OpenRouter $0.134$0.268 856ms 88 99.3% API-measured
burncloud burncloud $0.137$0.274 API-measured
fireworks_ai Pricebook (LiteLLM) $0.140$0.280 Official list
nebius Pricebook (LiteLLM) $0.140$0.280 Official list
tensormesh Pricebook (LiteLLM) $0.140$0.280 Official list
tencent Pricebook (LiteLLM) $0.140$0.280 Official list
novita Pricebook (LiteLLM) $0.140$0.280 Official list
wandb Pricebook (LiteLLM) $0.140$0.280 Official list
novita HuggingFace Router $0.140$0.280 1347ms 90 API-measured
novita-ai Novita AI $0.140$0.280 API-measured
Novita (fp8) OpenRouter $0.140$0.280 1274ms 48 100.0% API-measured
AtlasCloud (fp4) OpenRouter $0.140$0.280 1012ms 36 100% API-measured
Parasail (fp8) OpenRouter $0.140$0.280 895ms 34 99.4% API-measured
aihubmix Pricebook (LiteLLM) $0.142$0.284 Official list
NextBit (fp8) OpenRouter $0.150$0.350 2413ms 63 98.9% API-measured
dashscope Pricebook (LiteLLM) $0.200$0.400 Official list
qwencloud Pricebook (LiteLLM) $0.200$0.400 Official list
qwen_ai_platform Pricebook (LiteLLM) $0.200$0.400 Official list
Phala OpenRouter $0.200$0.400 1403ms 59 100% API-measured
Mancer 2 (fp8) OpenRouter $0.190$0.500 1302ms 22 98.0% API-measured
azure_ai Pricebook (LiteLLM) $0.190$0.510 Official list
Azure (us) OpenRouter $0.210$0.560 1024ms 25 96.1% API-measured
0g-teetls 0G Router $0.264$0.792 API-measured
deepseek Pricebook (LiteLLM) $0.300$1.20 Official list
libertai Pricebook (LiteLLM) $0.250$1.75 Official list

Volume trend

3.2T 30d d-29 · 2.7T tokensd-28 · 2.6T tokensd-27 · 2.1T tokensd-26 · 2.3T tokensd-25 · 2.5T tokensd-24 · 2.6T tokensd-23 · 3.2T tokensd-22 · 2.8T tokensd-21 · 2.3T tokensd-20 · 2.0T tokensd-19 · 2.1T tokensd-18 · 2.3T tokensd-17 · 2.4T tokensd-16 · 2.5T tokensd-15 · 2.7T tokensd-14 · 3.0T tokensd-13 · 2.4T tokensd-12 · 2.1T tokensd-11 · 2.3T tokensd-10 · 2.4T tokensd-9 · 2.3T tokensd-8 · 2.7T tokensd-7 · 2.5T tokensd-6 · 1.9T tokensd-5 · 2.0T tokensd-4 · 2.4T tokensd-3 · 2.2T tokensd-2 · 1.9T tokensd-1 · 1.9T tokensd-0 · 2.4T tokens

Last 30 days: 2.7T → 2.4T (-10%); range 1.9T–3.2T tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/deepseek-v4-flash.json. Free reuse requires attribution to tkx.org.

About

DeepSeek-V4-Flash is an AI model, a machine-learning system, served by multiple inference providers on AI model marketplaces. It is known for its role in natural language processing tasks.

AI-generated summary from public information — is this yours?

FAQ

How much does DeepSeek V4 Flash 0423 cost?

Cheapest measured offer right now: $0.0484/1M input, $0.0969/1M output via StreamLake (fp8) — across 30 tracked providers, refreshed hourly.

Who serves DeepSeek V4 Flash 0423?

30 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is DeepSeek V4 Flash 0423 used?

DeepSeek V4 Flash 0423 routed 2.4T tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings