gemini-3.1-flash-lite
vertex_ai-language-models ↗ · 1049K ctx · 5 providers · updated 2026-08-05 00:15 UTC
The board shows this model's cheapest measured offer ($0.563/1M); the smaller figure under it is the mean of the 5 offers listed below ($0.563/1M). Add those up and divide — it matches. 1 offer(s) above 10x the median were left out of the average and flagged ⚠ in the table below — they are still listed, and still compete for the cheapest price. 24h token volume (80B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.
| Provider | Router / Source | $ in /1M | $ out /1M | TTFT | TPS | Uptime | Evidence |
|---|---|---|---|---|---|---|---|
| vertex_ai-language-models | Pricebook (LiteLLM) | $0.250 | $1.50 | — | — | — | Official list |
| gemini | Pricebook (LiteLLM) | $0.250 | $1.50 | — | — | — | Official list |
| openrouter | Pricebook (LiteLLM) | $0.250 | $1.50 | — | — | — | Official list |
| openrouter-best | OpenRouter | $0.250 | $1.50 | — | — | — | API-measured |
| requesty | Requesty | $0.250 | $1.50 | — | — | — | API-measured |
| burncloud ⚠ outlier | burncloud | $75.00 | $300 | — | — | — | API-measured |
Volume trend
Last 30 days: 106B → 80B (-24%); range 50B–110B tokens/day. Source: OpenRouter public rankings; refreshed hourly.
Data & method
Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/gemini-3.1-flash-lite.json. Free reuse requires attribution to tkx.org.
About
Gemini-3.1-flash-lite is an AI model, specifically a machine-learning system designed for natural language processing tasks. It is served by multiple inference providers on AI model marketplaces and is listed under the org tag "vertex_ai-language-models".
AI-generated summary from public information — is this yours?
FAQ
How much does gemini-3.1-flash-lite cost?
Cheapest measured offer right now: $0.25/1M input, $1.5/1M output via vertex_ai-language-models — across 5 tracked providers, refreshed hourly.
Who serves gemini-3.1-flash-lite?
5 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.
How heavily is gemini-3.1-flash-lite used?
gemini-3.1-flash-lite routed 80B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.