Qwen3-32B:flex
deepinfra · 131K ctx · 11 providers · updated 2026-09-19 02:15 UTC
The board shows this model's cheapest measured offer ($0.104/1M); the smaller figure under it is the mean of the 12 offers listed below ($0.252/1M). Add those up and divide — it matches. 1 offer(s) above 10x the median were left out of the average and flagged ⚠ in the table below — they are still listed, and still compete for the cheapest price.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| requesty | Requesty ↗ | $0.064 | $0.224 | — | — | — | API-measured |
| ovhcloud | Pricebook (LiteLLM) ↗ | $0.080 | $0.230 | — | — | — | Official list |
| nscale | HuggingFace Router ↗ | $0.080 | $0.250 | 763ms | 38 | — | API-measured |
| deepinfra | Pricebook (LiteLLM) ↗ | $0.080 | $0.280 | — | — | — | Official list |
| openrouter | Pricebook (LiteLLM) ↗ | $0.080 | $0.280 | — | — | — | Official list |
| deepinfra | HuggingFace Router ↗ | $0.080 | $0.280 | 565ms | 25 | — | API-measured |
| DeepInfra (fp8) | OpenRouter ↗ | $0.080 | $0.280 | 328ms | 30 | 100.0% | API-measured |
| nebius | Pricebook (LiteLLM) ↗ | $0.100 | $0.300 | — | — | — | Official list |
| SiliconFlow (fp8) | OpenRouter ↗ | $0.140 | $0.570 | 846ms | 25 | 99.5% | API-measured |
| groq | Pricebook (LiteLLM) ↗ | $0.290 | $0.590 | — | — | — | Official list |
| sambanova | Pricebook (LiteLLM) ↗ | $0.400 | $0.800 | — | — | — | Official list |
| fireworks_ai | Pricebook (LiteLLM) ↗ | $0.900 | $0.900 | — | — | — | Official list |
| burncloud ⚠ outlier | burncloud ↗ | $2.00 | $20.00 | — | — | — | API-measured |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.0936 → $0.0936 → $0.0936 → $0.0936 → $0.0936 → $0.0936 → $0.0936 → $0.0936 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104 → $0.104
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen3-32b.json. Free reuse requires attribution to tkx.org.
About
Qwen3-32B is a language model AI, designed for natural language processing tasks. It is served by multiple inference providers on AI model marketplaces and is listed under the "ovhcloud" serving-platform tag.
AI-generated summary from public information — is this yours?
FAQ
How much does Qwen3-32B:flex cost?
Cheapest measured offer right now: $0.064/1M input, $0.224/1M output via requesty — across 11 tracked providers, refreshed hourly.
Who serves Qwen3-32B:flex?
11 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of Qwen3-32B:flex going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.