Models

Qwen3.8-Flash

together_ai · 1049K ctx · 6 providers · updated 2026-09-12 06:15 UTC

Cheapest
$0.230 Official list · Together AI
Avg price
$0.233 /1M
providers
6
Tokens (24h)
51B
Volume (24h)
$11.7K est.
(3 × $0.150 + $0.470) ÷ 4 = $0.230/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.150, 3:1 $0.230, 1:1 $0.310, output only $0.470 per 1M.

The board shows this model's cheapest measured offer ($0.230/1M); the smaller figure under it is the mean of the 6 offers listed below ($0.233/1M). Add those up and divide — it matches. 24h token volume (51B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
together_ai Pricebook (LiteLLM) $0.150$0.470 Official list
openrouter Pricebook (LiteLLM) $0.150$0.470 Official list
openrouter-best OpenRouter $0.150$0.470 API-measured
novita-ai Novita AI $0.150$0.470 API-measured
requesty Requesty $0.160$0.470 API-measured
0g-teetls 0G Router $0.150$0.506 API-measured

Volume trend

51B 14d d-13 · 25B tokensd-12 · 30B tokensd-11 · 0 tokensd-10 · 0 tokensd-9 · 0 tokensd-8 · 0 tokensd-7 · 0 tokensd-6 · 37B tokensd-5 · 33B tokensd-4 · 0 tokensd-3 · 0 tokensd-2 · 0 tokensd-1 · 0 tokensd-0 · 51B tokens

Last 14 days: 25B → 51B (+101%); range 0–51B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen3.8-flash.json. Free reuse requires attribution to tkx.org.

FAQ

How much does Qwen3.8-Flash cost?

Cheapest measured offer right now: $0.15/1M input, $0.47/1M output via together_ai — across 6 tracked providers, refreshed hourly.

Who serves Qwen3.8-Flash?

6 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is Qwen3.8-Flash used?

Qwen3.8-Flash routed 51B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings