Models

Qwen2.5-Coder-32B-Instruct

deepinfra · 131K ctx · 7 providers · updated 2026-09-19 02:15 UTC

Cheapest
$0.092 via Requesty
Avg price
$0.733 /1M
providers
7
(3 × $0.070 + $0.160) ÷ 4 = $0.092/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.070, 3:1 $0.092, 1:1 $0.115, output only $0.160 per 1M.

The board shows this model's cheapest measured offer ($0.092/1M); the smaller figure under it is the mean of the 8 offers listed below ($0.733/1M). Add those up and divide — it matches.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
requesty Requesty $0.070$0.160 API-measured
nscale Pricebook (LiteLLM) $0.060$0.200 Official list
nscale HuggingFace Router $0.060$0.200 732ms 33 API-measured
hyperbolic Pricebook (LiteLLM) $0.120$0.300 Official list
cloudflare Pricebook (LiteLLM) $0.660$1.00 Official list
together_ai Pricebook (LiteLLM) $0.800$0.800 Official list
ovhcloud Pricebook (LiteLLM) $0.870$0.870 Official list
burncloud burncloud $2.00$6.00 API-measured

Price trend

Cheapest blended price, last 31 days (USD/1M): $0.0833 → $0.0833 → $0.0833 → $0.0833 → $0.0833 → $0.0833 → $0.0833 → $0.0833 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925 → $0.0925

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen2.5-coder-32b-instruct.json. Free reuse requires attribution to tkx.org.

About

Qwen2.5-Coder-32B-Instruct is a large language model served by multiple inference providers on AI model marketplaces, designed for coding and development tasks. It is listed under the org tag "deepinfra" on serving-platform tags.

AI-generated summary from public information — is this yours?

FAQ

How much does Qwen2.5-Coder-32B-Instruct cost?

Cheapest measured offer right now: $0.07/1M input, $0.16/1M output via requesty — across 7 tracked providers, refreshed hourly.

Who serves Qwen2.5-Coder-32B-Instruct?

7 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

Is the price of Qwen2.5-Coder-32B-Instruct going up or down?

The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.

Related analysis

Back to rankings