Qwen2.5-Coder-32B-Instruct
deepinfra · 131K ctx · 6 providers · updated 2026-08-05 00:15 UTC
The board shows this model's cheapest measured offer ($0.092/1M); the smaller figure under it is the mean of the 6 offers listed below ($0.344/1M). Add those up and divide — it matches. 1 offer(s) above 10x the median were left out of the average and flagged ⚠ in the table below — they are still listed, and still compete for the cheapest price.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| requesty | Requesty | $0.070 | $0.160 | — | — | — | API-measured |
| nscale | Pricebook (LiteLLM) | $0.060 | $0.200 | — | — | — | Official list |
| nscale | HuggingFace Router | $0.060 | $0.200 | 575ms | 48 | — | API-measured |
| hyperbolic | Pricebook (LiteLLM) | $0.120 | $0.300 | — | — | — | Official list |
| cloudflare | Pricebook (LiteLLM) | $0.660 | $1.00 | — | — | — | Official list |
| ovhcloud | Pricebook (LiteLLM) | $0.870 | $0.870 | — | — | — | Official list |
| burncloud ⚠ outlier | burncloud | $2.00 | $6.00 | — | — | — | API-measured |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.095 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen2.5-coder-32b-instruct.json. Free reuse requires attribution to tkx.org.
About
Qwen2.5-Coder-32B-Instruct is a large language model served by multiple inference providers on AI model marketplaces, designed for coding and development tasks. It is listed under the org tag "deepinfra" on serving-platform tags.
AI-generated summary from public information — is this yours?
FAQ
How much does Qwen2.5-Coder-32B-Instruct cost?
Cheapest measured offer right now: $0.07/1M input, $0.16/1M output via requesty — across 6 tracked providers, refreshed hourly.
Who serves Qwen2.5-Coder-32B-Instruct?
6 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of Qwen2.5-Coder-32B-Instruct going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.