Models

Nemotron 3 Ultra

nvidia ↗ · 512K ctx · 2 providers · updated 2026-08-05 00:15 UTC

Cheapest
$1.35 via OpenRouter
Avg price
$1.43 /1M
providers
2
Tokens (24h)
326B
Volume (24h)
$0 est.
(3 × $0.600 + $3.60) ÷ 4 = $1.35/1M — the $/1M figure blends input and output at 3:1 — that is the only way models with two different rates can be ranked on one axis. Your actual rate depends on your own input:output mix — input only $0.600, 3:1 $1.35, 1:1 $2.10, output only $3.60 per 1M.

The board shows this model's cheapest measured offer ($1.35/1M); the smaller figure under it is the mean of the 2 offers listed below ($1.43/1M). Add those up and divide — it matches. 24h token volume (326B) is sourced from OpenRouter public rankings — today's only public source of absolute per-model token counts (Vercel AI Gateway publishes share-of-traffic percentages); as more sources come online, the per-router breakdown will sum to this total here.

ProviderRouter / Source$ in /1M$ out /1MTTFTTPSUptimeEvidence
openrouter-best OpenRouter $0.600$3.60 API-measured
requesty Requesty $1.00$3.00 API-measured

Volume trend

499B 30d d-29 · 136B tokensd-28 · 126B tokensd-27 · 125B tokensd-26 · 244B tokensd-25 · 441B tokensd-24 · 492B tokensd-23 · 499B tokensd-22 · 480B tokensd-21 · 469B tokensd-20 · 453B tokensd-19 · 441B tokensd-18 · 426B tokensd-17 · 434B tokensd-16 · 442B tokensd-15 · 253B tokensd-14 · 220B tokensd-13 · 257B tokensd-12 · 366B tokensd-11 · 416B tokensd-10 · 424B tokensd-9 · 405B tokensd-8 · 394B tokensd-7 · 336B tokensd-6 · 358B tokensd-5 · 360B tokensd-4 · 306B tokensd-3 · 340B tokensd-2 · 333B tokensd-1 · 321B tokensd-0 · 326B tokens

Last 30 days: 136B → 326B (+140%); range 125B–499B tokens/day. Source: OpenRouter public rankings; refreshed hourly.

Data & method

Cheapest ranked by blended = (3×in + out) ÷ 4; performance columns show only when a source provides them, never estimated. “Official list” = LiteLLM open pricebook (MIT); “API-measured” = provider/aggregator public APIs. Machine-readable data: /api/v1/models/nemotron-3-ultra-550b-a55b.json. Free reuse requires attribution to tkx.org.

About

Nemotron 3 Ultra is an AI model designed for natural language processing tasks, served by multiple inference providers on AI model marketplaces and listed under the organization tag "NVIDIA".

AI-generated summary from public information — is this yours?

FAQ

How much does Nemotron 3 Ultra cost?

Cheapest measured offer right now: $0.6/1M input, $3.6/1M output via openrouter-best — across 2 tracked providers, refreshed hourly.

Who serves Nemotron 3 Ultra?

2 providers currently serve it. The table above lists each with input/output price, TTFT, TPS, uptime and an evidence tier.

How heavily is Nemotron 3 Ultra used?

Nemotron 3 Ultra routed 326B tokens in the last 24h. The daily bars above show routed volume over recent days. Source: OpenRouter public rankings — one of the few routers publishing per-model usage; TKX refreshes hourly.

Related analysis

Back to rankings