Skip to content
AI Tier ListField guide
01 · WEEKLY RANKING

Best LLMs for Math

Week of

10 models · ranked by real Math usage

Key facts · week of 2026-08-10

  1. F1

    The top 3 models (DeepSeek: DeepSeek V4 Flash 0423, DeepSeek: DeepSeek V4 Flash 0731, DeepSeek: DeepSeek V4 Pro) account for 32.2% of all usage.

    §
  2. F2

    This week's biggest riser is DeepSeek: DeepSeek V4 Flash 0731, up from #8 to #2.

    §
  3. F3

    The cheapest model in the usage top 10 is OpenAI: GPT-5.6 Luna (batch) at $0.09999999999999999 per 1M input tokens (usage rank #5).

    §
  4. F4

    The largest context window in the top 10 is OpenAI: GPT-5.6 Luna (batch) at 1.1M tokens.

    §

Cite as: AI Tier List, LLM Usage Rankings (2026-08-10), www.aitierlist.xyz/en/models/math — data: OpenRouter

Which model should you pick?

When quality comes first
MoonshotAI: Kimi K3moonshotai
Artificial Analysis intelligence index score 59.7 — highest in this ranking
The proven default
DeepSeek: DeepSeek V4 Flash 0423deepseek
12.7% usage share — #1 in real usage
When cost matters
DeepSeek: DeepSeek V4 Flash 0731deepseek
Cheapest in the top 15 — $0.14/1M input
For long docs & codebases
OpenAI: GPT-5.6 Luna (batch)openai
1.1M token context window

Full ranking · 10 models

RankModelShare
1
DeepSeek: DeepSeek V4 Flash 0423
in $0.14 · out $0.28 · 1M
12.7%
26
DeepSeek: DeepSeek V4 Flash 0731
Intelligence 51.8 · in $0.14 · out $0.28 · 1M
12.6%
31
DeepSeek: DeepSeek V4 Pro
Intelligence 45.3 · in $1.17 · out $2.34 · 1M
6.9%
41
Z.ai: GLM 5.2 (batch)
in $0.70 · out $2.20 · 512K
6.2%
52
OpenAI: GPT-5.6 Luna (batch)
in $0.10 · out $0.60 · 1.1M
4.4%
62
Anthropic: Claude Opus 4.8 (batch)
in $2.50 · out $12.50 · 1M
3.2%
72
Xiaomi: MiMo-V2.5
Intelligence 38 · in $0.14 · out $0.28 · 1.1M
3%
82
MoonshotAI: Kimi K3
Intelligence 59.7 · in $3 · out $15 · 1M
2.5%
9
MoonshotAI: Kimi K2.6
Intelligence 45.1 · in $0.95 · out $4 · 262K
2%
104
Tencent: Hy3
in $0.13 · out $0.53 · 262K
2%
Share these rankingsXReddit

Get notified when Math rankings move

This ranking refreshes every week. We email you when it moves — once a week, no spam.

FAQ

What is the most-used LLM for Math?

DeepSeek: DeepSeek V4 Flash 0423 by deepseek ranks first for Math by real token usage. It accounts for roughly 12.7% of tokens in this segment. DeepSeek: DeepSeek V4 Flash 0731, DeepSeek: DeepSeek V4 Pro, Z.ai: GLM 5.2 (batch) follow.

How are these rankings calculated?

Models are ranked by tokens actually consumed through OpenRouter, not by benchmark scores. This page counts only usage tagged as Math tasks. It refreshes weekly and currently tracks 10 models.

Which is the cheapest LLM for Math?

Among the paid models in this ranking, OpenAI: GPT-5.6 Luna (batch) is cheapest at $0.10 per 1M input tokens.

Is DeepSeek: DeepSeek V4 Flash 0423 still the most-used LLM for Math right now?

Yes. As of the week of 2026-08-10, DeepSeek: DeepSeek V4 Flash 0423 is #1 by real usage. It was #1 last week too. The ranking refreshes weekly, so this page changes as soon as it does.

Which model has the longest context window?

OpenAI: GPT-5.6 Luna (batch) has the longest context window in this ranking at 1.1M tokens.

Token usage and share data aggregated by OpenRouter, refreshed weekly. Prices are USD per 1M tokens. Benchmark scores come from Artificial Analysis; models it hasn't measured show —.