Skip to content
AI Tier ListField guide
01 · WEEKLY RANKING

LLM Rankings

Week of

50 models · 67.07T tokens this week · ranked by real usage

Key facts · week of 2026-08-10

  1. F1

    The most-used LLM for the week of is DeepSeek: DeepSeek V4 Flash 0731, with 15.8% of OpenRouter token traffic (10.69T tokens).

    §
  2. F2

    The top 3 models (DeepSeek: DeepSeek V4 Flash 0731, Tencent: Hy3, OpenAI: GPT-5.6 Luna (batch)) account for 39.2% of all usage.

    §
  3. F3

    This week's biggest riser is OpenAI: GPT-5.5 (batch), up from #50 to #39.

    §
  4. F4

    The cheapest model in the usage top 10 is OpenAI: GPT-5.6 Luna (batch) at $0.09999999999999999 per 1M input tokens (usage rank #3).

    §
  5. F5

    The largest context window in the top 10 is OpenAI: GPT-5.6 Luna (batch) at 1.1M tokens.

    §

Cite as: AI Tier List, LLM Usage Rankings (2026-08-10), www.aitierlist.xyz/en/models — data: OpenRouter

Which model should you pick?

When quality comes first
Anthropic: Claude Fable 5 (batch)anthropic
Artificial Analysis intelligence index score 62.1 — highest in this ranking
The proven default
DeepSeek: DeepSeek V4 Flash 0731deepseek
15.8% usage share — #1 in real usage
When cost matters
NVIDIA: Nemotron 3 Ultra (free)nvidia
Cheapest in the top 15 — Free/1M input
For long docs & codebases
OpenAI: GPT-5.6 Luna (batch)openai
1.1M token context window
To bet on the trending choice
OpenAI: GPT-5.5 (batch)openai
Up 11 spots (#50 → #39)

Full ranking · 50 models

RankModelShare
1
DeepSeek: DeepSeek V4 Flash 0731
deepseek·vs Tencent: Hy3
Intelligence 51.8 · in $0.14 · out $0.28 · 1M
15.8%
10.69T
2
Tencent: Hy3
in $0.13 · out $0.53 · 262K
15.7%
10.65T
32
OpenAI: GPT-5.6 Luna (batch)
in $0.10 · out $0.60 · 1.1M
7.7%
5.20T
41
DeepSeek: DeepSeek V4 Flash 0423
in $0.14 · out $0.28 · 1M
7.6%
5.11T
51
Xiaomi: MiMo-V2.5
Intelligence 38 · in $0.14 · out $0.28 · 1.1M
6.4%
4.30T
6
Z.ai: GLM 5.2 (batch)
in $0.70 · out $2.20 · 512K
5.5%
3.68T
7
DeepSeek: DeepSeek V4 Pro
Intelligence 45.3 · in $1.17 · out $2.34 · 1M
4.1%
2.76T
8
NVIDIA: Nemotron 3 Ultra (free)
in Free · out Free · 1M
2.9%
1.94T
9
Google: Gemini 3.6 Flash (batch)
in $0.38 · out $1.88 · 1M
2.5%
1.70T
10
Poolside: Laguna S 2.1 (free)
in Free · out Free · 262K
2.5%
1.68T
113
Claude Opus 5 (batch)
in $2.50 · out $12.50 · 1M
2.4%
1.60T
121
MiniMax: MiniMax M3 (batch)
in $0.15 · out $0.60 · 524K
2.3%
1.56T
131
MoonshotAI: Kimi K3
Intelligence 59.7 · in $3 · out $15 · 1M
2.1%
1.45T
141
StepFun: Step 3.7 Flash
Intelligence 30.9 · in $0.20 · out $1.15 · 262K
1.7%
1.12T
15
Anthropic: Claude Sonnet 5 (batch)
in $1 · out $5 · 1M
1.6%
1.07T
16
Google: Gemini 3 Flash Preview (batch)
in $0.25 · out $1.50 · 1M
1.3%
877.3B
171
OpenAI: GPT-5.6 Terra (batch)
in $1 · out $6 · 1.1M
1.3%
875.5B
181
Anthropic: Claude Sonnet 4.6 (batch)
Intelligence 48.4 · in $1.50 · out $7.50 · 1M
1.1%
776.3B
192
OpenAI: GPT-5.6 Sol (batch)
in $2.50 · out $15 · 1.1M
1%
683.3B
20
Google: Gemini 2.5 Flash Lite (batch)
in $0.05 · out $0.20 · 1M
1%
677.7B
212
Anthropic: Claude Opus 4.8 (batch)
in $2.50 · out $12.50 · 1M
0.9%
602.6B
22
OpenAI: GPT-5.6 Luna Pro (batch)
in $0.10 · out $0.60 · 1.1M
0.9%
600.7B
235
Google: Gemma 4 31B (free)
Intelligence 29.7 · in Free · out Free · 262K
0.8%
507.5B
24
Google: Gemini 2.5 Flash (batch)
in $0.15 · out $1.25 · 1M
0.7%
499.6B
25
Xiaomi: MiMo-V2.5-Pro
Intelligence 42.9 · in $0.43 · out $0.87 · 1.1M
0.7%
467.2B
261
OpenAI: gpt-oss-120b
Intelligence 24.1 · in $0.03 · out $0.17 · 131K
0.6%
436.7B
271
Google: Gemini 3.1 Flash Lite (batch)
in $0.13 · out $0.75 · 1M
0.6%
425.1B
281
DeepSeek: DeepSeek V3.2
Intelligence 32.6 · in $0.27 · out $0.40 · 164K
0.6%
403.3B
291
Google: Gemma 4 26B A4B (free)
Intelligence 26.1 · in Free · out Free · 262K
0.6%
391.2B
301
NVIDIA: Nemotron 3 Super (free)
in Free · out Free · 262K
0.5%
336.2B
311
SpaceXAI: Grok 4.5
Intelligence 55.8 · in $2 · out $6 · 500K
0.5%
327.1B
321
Anthropic: Claude Opus 4.7 (batch)
in $2.50 · out $12.50 · 1M
0.5%
310.7B
336
Qwen: Qwen3.8 Max
Intelligence 58.1 · in $2 · out $6 · 1M
0.4%
300.2B
34NEW
NVIDIA: Nemotron 3.5 Lightning (free)
in Free · out Free · 1M
0.4%
255.2B
351
Anthropic: Claude Fable 5 (batch)
Intelligence 62.1 · in $5 · out $25 · 1M
0.4%
254.7B
361
Anthropic: Claude Haiku 4.5 (batch)
Intelligence 29.9 · in $0.50 · out $2.50 · 200K
0.3%
234.4B
37
inclusionAI: Ling-2.6-flash
Intelligence 14.2 · in $0.01 · out $0.03 · 262K
0.3%
229.2B
384
Cohere: North Mini Code (free)
Intelligence 20.2 · in Free · out Free · 256K
0.3%
217.6B
3911
OpenAI: GPT-5.5 (batch)
in $2.50 · out $15 · 1.1M
0.3%
211.5B
402
OpenAI: GPT-4o-mini (batch)
in $0.07 · out $0.30 · 128K
0.3%
208.9B
411
Google: Gemini 3.5 Flash Lite (batch)
in $0.15 · out $1.25 · 1M
0.3%
187.2B
422
MiniMax: MiniMax M2.7
Intelligence 38.9 · in $0.30 · out $1.20 · 205K
0.3%
184.7B
435
Anthropic: Claude Opus 4.6 (batch)
in $2.50 · out $12.50 · 1M
0.2%
155.4B
443
OpenAI: GPT-5 Mini (batch)
in $0.13 · out $1 · 400K
0.2%
153.6B
45NEW
DeepSeek: DeepSeek V4 Pro 0813
Intelligence 53.2 · in $0.43 · out $0.87 · 1M
0.2%
142.9B
461
OpenAI: GPT-5.4 (batch)
in $1.25 · out $7.50 · 1.1M
0.2%
142.1B
472
Poolside: Laguna XS 2.1 (free)
in Free · out Free · 262K
0.2%
140.3B
484
Google: Gemini 3.1 Pro Preview (batch)
in $1 · out $6 · 1M
0.2%
120.8B
493
OpenAI: gpt-oss-20b (free)
Intelligence 15.2 · in Free · out Free · 131K
0.2%
120.2B
507
Google: Gemini 3.5 Flash (batch)
Intelligence 52 · in $0.75 · out $4.50 · 1M
0.2%
107.0B
Share these rankingsXReddit

Get next week’s LLM rankings

These rankings refresh every week. We email you when they move — once a week, no spam.

FAQ

What is the most-used LLM?

DeepSeek: DeepSeek V4 Flash 0731 by deepseek ranks first by real token usage. It accounts for roughly 15.8% of tokens in this segment. Tencent: Hy3, OpenAI: GPT-5.6 Luna (batch), DeepSeek: DeepSeek V4 Flash 0423 follow.

How are these rankings calculated?

Models are ranked by tokens actually consumed through OpenRouter, not by benchmark scores. It refreshes weekly and currently tracks 50 models.

Which is the cheapest LLM?

Among the paid models in this ranking, inclusionAI: Ling-2.6-flash is cheapest at $0.01 per 1M input tokens. 9 models in the list are free to use.

Which model has the longest context window?

OpenAI: GPT-5.6 Luna (batch) has the longest context window in this ranking at 1.1M tokens.

Token usage and share data aggregated by OpenRouter, refreshed weekly. Prices are USD per 1M tokens. Benchmark scores come from Artificial Analysis; models it hasn't measured show —.