The board shows this model's cheapest measured offer ($0.155/1M); the smaller figure under it is the mean of the 28 offers listed below ($0.482/1M). Add those up and divide — it matches.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| DeepInfra (turbo) | OpenRouter | $0.100 | $0.320 | 482ms | 14 | 95.6% | API-measured |
| hyperbolic | Pricebook (LiteLLM) | $0.120 | $0.300 | — | — | — | Official list |
| nebius | Pricebook (LiteLLM) | $0.130 | $0.400 | — | — | — | Official list |
| requesty | Requesty | $0.130 | $0.400 | — | — | — | API-measured |
| Nebius (fp8) | OpenRouter | $0.130 | $0.400 | 486ms | 23 | 59.3% | API-measured |
| AkashML (fp8) | OpenRouter | $0.130 | $0.400 | 434ms | 29 | 99.6% | API-measured |
| crusoe | Pricebook (LiteLLM) | $0.200 | $0.200 | — | — | — | Official list |
| nscale | Pricebook (LiteLLM) | $0.200 | $0.200 | — | — | — | Official list |
| novita | Pricebook (LiteLLM) | $0.135 | $0.400 | — | — | — | Official list |
| Novita (bf16) | OpenRouter | $0.135 | $0.400 | 554ms | 37 | 98.4% | API-measured |
| novita | HuggingFace Router | $0.135 | $0.400 | 905ms | 41 | — | API-measured |
| novita-ai | Novita AI | $0.135 | $0.400 | — | — | — | API-measured |
| deepinfra | Pricebook (LiteLLM) | $0.230 | $0.400 | — | — | — | Official list |
| Parasail (fp8) | OpenRouter | $0.220 | $0.500 | 519ms | 39 | 97.9% | API-measured |
| Crusoe (bf16) | OpenRouter | $0.250 | $0.750 | 327ms | 60 | 99.6% | API-measured |
| SambaNova | OpenRouter | $0.450 | $0.900 | 697ms | 60 | 97.9% | API-measured |
| groq | HuggingFace Router | $0.590 | $0.790 | 129ms | 253 | — | API-measured |
| Groq | OpenRouter | $0.590 | $0.790 | 256ms | 35 | 100% | API-measured |
| azure_ai | Pricebook (LiteLLM) | $0.710 | $0.710 | — | — | — | Official list |
| CoreWeave (fp16) | OpenRouter | $0.710 | $0.710 | 210ms | 67 | 99.9% | API-measured |
| OpenRouter | $0.720 | $0.720 | 279ms | 53 | 100% | API-measured | |
| Google (us-central1) | OpenRouter | $0.720 | $0.720 | 303ms | 63 | 100% | API-measured |
| ovhcloud | HuggingFace Router | $0.740 | $0.740 | 622ms | 24 | — | API-measured |
| Cloudflare (fp8) | OpenRouter | $0.293 | $2.25 | 289ms | 41 | 98.9% | API-measured |
| scaleway | Pricebook (LiteLLM) | $0.900 | $0.900 | — | — | — | Official list |
| scaleway | HuggingFace Router | $1.03 | $1.03 | 314ms | 87 | — | API-measured |
| together | HuggingFace Router | $1.04 | $1.04 | 2011ms | 89 | — | API-measured |
| Together (fp8) | OpenRouter | $1.04 | $1.04 | 618ms | 26 | 99.5% | API-measured |
Who burns this model
Top apps routing traffic to this model over the last 30 days, via OpenRouter's public stats. Only apps that opted into tracking appear, and only the top few are published — this is not the full demand picture. These figures are a 30-day window and are NOT comparable with the cumulative totals on the Apps board.
| App | Tokens (30d) | Requests |
|---|---|---|
| DialOS Ask Merlin | 11B | 12M |
| shapes inc | 5B | 336K |
| Peezy Gateway | 5B | 397K |
| Secret Loyalties - Evaluation | 4B | 337K |
| FVChat | 3B | 133K |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155 → $0.155
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/llama-3.3-70b-instruct.json. Free reuse requires attribution to tkx.org.
About
Llama 3.3 70B Instruct is a large language model served by multiple inference providers on AI model marketplaces, known for its capabilities in natural language processing and understanding.
AI-generated summary from public information — is this yours?
FAQ
How much does Llama 3.3 70B Instruct cost?
Cheapest measured offer right now: $0.1/1M input, $0.32/1M output via DeepInfra (turbo) — across 18 tracked providers, refreshed hourly.
Who serves Llama 3.3 70B Instruct?
18 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of Llama 3.3 70B Instruct going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.