Insights Rankings ↗
🌐 EN
Insights

9.6 trillion tokens a day: ranking models by what actually runs

July 17, 2026 · TKX

Ask ten people which AI models "everyone uses" and you get ten guesses. TKX's heat board no longer guesses: it ranks models by measured token volume — real traffic, read daily from the largest public routing marketplace. On July 16, 2026, that meant 9.62 trillion tokens in a single day.

The top of the board, measured

The ten busiest models by routed token volume (July 16, 2026):

#ModelTokens / dayShare
1hy31,745 B18.1%
2mimo-v2.51,387 B14.4%
3deepseek-v4-flash777 B8.1%
4glm-5.2624 B6.5%
5minimax-m3564 B5.9%
6nemotron-3-ultra-550b441 B4.6%
7deepseek-v4-pro392 B4.1%
8claude-4.7-opus390 B4.1%
9claude-4.8-opus332 B3.5%
10claude-4.6-sonnet171 B1.8%

Live data, refreshed daily: /api/v1/usage.json.

What the data shows

The honest scope note

This is real usage on one marketplace, not global market share. OpenRouter's traffic skews price-sensitive and multi-model; most first-party API traffic (e.g. direct enterprise use of a vendor's own API) never touches a router and is invisible here. So closed flagship models are under-represented in these numbers. TKX labels the metric accordingly — a measured slice, honestly scoped, beats a guessed total.

Why rank by usage at all?

Because the alternative is worse. Popularity proxies — search volume, social buzz, benchmark scores — are all gameable and none of them clear at a price. A routed token is a paid, delivered unit of work. It is the same discipline that runs every TKX board: measure, tag the evidence, and never accept self-reported numbers.

← All insights
The data authority for AI tokens & compute · tkx.org