gemma-4-26B-A4B-it:flex
deepinfra · 262K ctx · 9 providers · updated 2026-09-19 02:15 UTC
The board shows this model's cheapest measured offer ($0.110/1M); the smaller figure under it is the mean of the 13 offers listed below ($0.196/1M). Add those up and divide — it matches.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| requesty | Requesty ↗ | $0.056 | $0.272 | — | — | — | API-measured |
| deepinfra | Pricebook (LiteLLM) ↗ | $0.070 | $0.340 | — | — | — | Official list |
| deepinfra | HuggingFace Router ↗ | $0.070 | $0.340 | 272ms | 48 | — | API-measured |
| openrouter | Pricebook (LiteLLM) ↗ | $0.090 | $0.300 | — | — | — | Official list |
| openrouter-best | OpenRouter ↗ | $0.090 | $0.300 | — | — | — | API-measured |
| cloudflare | Pricebook (LiteLLM) ↗ | $0.100 | $0.300 | — | — | — | Official list |
| novita | Pricebook (LiteLLM) ↗ | $0.130 | $0.400 | — | — | — | Official list |
| novita | HuggingFace Router ↗ | $0.130 | $0.400 | 906ms | 65 | — | API-measured |
| novita-ai | Novita AI ↗ | $0.130 | $0.400 | — | — | — | API-measured |
| aihubmix | Pricebook (LiteLLM) ↗ | $0.140 | $0.400 | — | — | — | Official list |
| vertex_ai | Pricebook (LiteLLM) ↗ | $0.150 | $0.600 | — | — | — | Official list |
| scaleway | Pricebook (LiteLLM) ↗ | $0.250 | $0.500 | — | — | — | Official list |
| scaleway | HuggingFace Router ↗ | $0.285 | $0.570 | 322ms | 175 | — | API-measured |
Price trend
Cheapest blended price, last 31 days (USD/1M): $0.099 → $0.099 → $0.099 → $0.099 → $0.099 → $0.099 → $0.099 → $0.099 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.11 → $0.0865 → $0.0865 → $0.0865 → $0.0865 → $0.0865 → $0.0865 → $0.0865 → $0.0865 → $0.11 → $0.11
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/gemma-4-26b-a4b-it.json. Free reuse requires attribution to tkx.org.
About
Gemma 4 26B A4B is a large language model served by multiple inference providers on AI model marketplaces. It is used for natural language processing tasks, leveraging its capabilities in understanding and generating human-like text.
AI-generated summary from public information — is this yours?
FAQ
How much does gemma-4-26B-A4B-it:flex cost?
Cheapest measured offer right now: $0.056/1M input, $0.272/1M output via requesty — across 9 tracked providers, refreshed hourly.
Who serves gemma-4-26B-A4B-it:flex?
9 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
Is the price of gemma-4-26B-A4B-it:flex going up or down?
The daily candles above track the cheapest blended price ($/1M, 3:1 in:out). TKX snapshots every hour, so drops and hikes show up the same day they happen.