Qwen2.5-7B-Instruct
deepinfra · 33K ctx · 4 providers · updated 2026-08-05 00:15 UTC
The board shows this model's cheapest measured offer ($0.055/1M); the smaller figure under it is the mean of the 5 offers listed below ($0.224/1M). Add those up and divide — it matches.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| deepinfra | Pricebook (LiteLLM) | $0.040 | $0.100 | — | — | — | Official list |
| novita | Pricebook (LiteLLM) | $0.070 | $0.070 | — | — | — | Official list |
| novita-ai | Novita AI | $0.070 | $0.070 | — | — | — | API-measured |
| together | HuggingFace Router | $0.300 | $0.300 | 201ms | 130 | — | API-measured |
| burncloud | burncloud | $0.500 | $1.00 | — | — | — | API-measured |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055 → $0.055
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/qwen2.5-7b-instruct.json. Free reuse requires attribution to tkx.org.
FAQ
How much does Qwen2.5-7B-Instruct cost?
Cheapest measured offer right now: $0.04/1M input, $0.1/1M output via deepinfra — across 4 tracked providers, refreshed hourly.
Who serves Qwen2.5-7B-Instruct?
4 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of Qwen2.5-7B-Instruct going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.