Qwen3-32B
ovhcloud · 131K ctx · 11 providers · updated 2026-08-05 00:15 UTC
The board shows this model's cheapest measured offer ($0.117/1M); the smaller figure under it is the mean of the 14 offers listed below ($0.261/1M). Add those up and divide — it matches. 1 offer(s) above 10x the median were left out of the average and flagged ⚠ in the table below — they are still listed, and still compete for the cheapest price.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| ovhcloud | Pricebook (LiteLLM) | $0.080 | $0.230 | — | — | — | Official list |
| nscale | HuggingFace Router | $0.080 | $0.250 | 628ms | 48 | — | API-measured |
| deepinfra | HuggingFace Router | $0.080 | $0.280 | 510ms | 47 | — | API-measured |
| DeepInfra (fp8) | OpenRouter | $0.080 | $0.280 | 272ms | 31 | 99.5% | API-measured |
| deepinfra | Pricebook (LiteLLM) | $0.100 | $0.280 | — | — | — | Official list |
| nebius | Pricebook (LiteLLM) | $0.100 | $0.300 | — | — | — | Official list |
| Nebius (base) | OpenRouter | $0.100 | $0.300 | 477ms | 29 | 99.6% | API-measured |
| requesty | Requesty | $0.100 | $0.300 | — | — | — | API-measured |
| Alibaba (fp8) | OpenRouter | $0.104 | $0.416 | — | — | — | API-measured |
| SiliconFlow (fp8) | OpenRouter | $0.140 | $0.570 | 1543ms | 17 | 97.3% | API-measured |
| groq | Pricebook (LiteLLM) | $0.290 | $0.590 | — | — | — | Official list |
| Groq | OpenRouter | $0.290 | $0.590 | 255ms | 297 | 100% | API-measured |
| sambanova | Pricebook (LiteLLM) | $0.400 | $0.800 | — | — | — | Official list |
| fireworks_ai | Pricebook (LiteLLM) | $0.900 | $0.900 | — | — | — | Official list |
| burncloud ⚠ outlier | burncloud | $2.00 | $20.00 | — | — | — | API-measured |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175 → $0.1175
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen3-32b.json. Free reuse requires attribution to tkx.org.
About
Qwen3-32B is a language model AI, designed for natural language processing tasks. It is served by multiple inference providers on AI model marketplaces and is listed under the "ovhcloud" serving-platform tag.
AI-generated summary from public information — is this yours?
FAQ
How much does Qwen3-32B cost?
Cheapest measured offer right now: $0.08/1M input, $0.23/1M output via ovhcloud — across 11 tracked providers, refreshed hourly.
Who serves Qwen3-32B?
11 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of Qwen3-32B going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.