Models

Qwen3-32B

ovhcloud · 131K ctx · 11 providers · updated 2026-08-05 00:15 UTC

Cheapest
$0.117 Official list · Ovhcloud
Avg price
$0.261 /1M · 1 outlier(s) excluded
providers
11
(3 × $0.080 + $0.230) ÷ 4 = $0.117/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.080, 3:1 $0.117, 1:1 $0.155, output only $0.230 per 1M.

The board shows this model's cheapest measured offer ($0.117/1M); the smaller figure under it is the mean of the 14 offers listed below ($0.261/1M). Add those up and divide — it matches. 1 offer(s) above 10x the median were left out of the average and flagged ⚠ in the table below — they are still listed, and still compete for the cheapest price.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
ovhcloud Pricebook (LiteLLM) $0.080$0.230 Official list
nscale HuggingFace Router $0.080$0.250 628ms 48 API-measured
deepinfra HuggingFace Router $0.080$0.280 510ms 47 API-measured
DeepInfra (fp8) OpenRouter $0.080$0.280 272ms 31 99.5% API-measured
deepinfra Pricebook (LiteLLM) $0.100$0.280 Official list
nebius Pricebook (LiteLLM) $0.100$0.300 Official list
Nebius (base) OpenRouter $0.100$0.300 477ms 29 99.6% API-measured
requesty Requesty $0.100$0.300 API-measured
Alibaba (fp8) OpenRouter $0.104$0.416 API-measured
SiliconFlow (fp8) OpenRouter $0.140$0.570 1543ms 17 97.3% API-measured
groq Pricebook (LiteLLM) $0.290$0.590 Official list
Groq OpenRouter $0.290$0.590 255ms 297 100% API-measured
sambanova Pricebook (LiteLLM) $0.400$0.800 Official list
fireworks_ai Pricebook (LiteLLM) $0.900$0.900 Official list
burncloud ⚠ outlier burncloud $2.00$20.00 API-measured

Price trend

$0.1175 $0.1175 2026-07-06 → 2026-08-05 · 31d 2026-07-06 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-07 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-08 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-09 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-10 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-11 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-12 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-13 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-14 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-15 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-16 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-17 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-18 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-19 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-20 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-21 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-22 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-23 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-24 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-25 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-26 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-27 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-28 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-29 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-30 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-07-31 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-08-01 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-08-02 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-08-03 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-08-04 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M2026-08-05 · O 0.1175 H 0.1175 L 0.1175 C 0.1175 $/1M

Cheapest blended price, last 31 days (USD/1M): $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen3-32b.json. Free reuse requires attribution to tkx.org.

About

Qwen3-32B is a language model AI, designed for natural language processing tasks. It is served by multiple inference providers on AI model marketplaces and is listed under the "ovhcloud" serving-platform tag.

AI-generated summary from public information — is this yours?

FAQ

How much does Qwen3-32B cost?

Cheapest measured offer right now: $0.08/1M input, $0.23/1M output via ovhcloud — across 11 tracked providers, refreshed hourly.

Who serves Qwen3-32B?

11 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

Is the price of Qwen3-32B going up or down?

The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.

Related analysis

Back to rankings