Best LLMs for Math
10 models · ranked by real Math usage
Key facts · week of 2026-08-10
The top 3 models (DeepSeek: DeepSeek V4 Flash 0423, DeepSeek: DeepSeek V4 Flash 0731, DeepSeek: DeepSeek V4 Pro) account for 32.2% of all usage.
§This week's biggest riser is DeepSeek: DeepSeek V4 Flash 0731, up from #8 to #2.
§The cheapest model in the usage top 10 is OpenAI: GPT-5.6 Luna (batch) at $0.09999999999999999 per 1M input tokens (usage rank #5).
§The largest context window in the top 10 is OpenAI: GPT-5.6 Luna (batch) at 1.1M tokens.
§
Cite as: AI Tier List, LLM Usage Rankings (2026-08-10), www.aitierlist.xyz/en/models/math — data: OpenRouter
Which model should you pick?
- When quality comes first
- MoonshotAI: Kimi K3
- Artificial Analysis intelligence index score 59.7 — highest in this ranking
- The proven default
- DeepSeek: DeepSeek V4 Flash 0423
- 12.7% usage share — #1 in real usage
- When cost matters
- DeepSeek: DeepSeek V4 Flash 0731
- Cheapest in the top 15 — $0.14/1M input
- For long docs & codebases
- OpenAI: GPT-5.6 Luna (batch)
- 1.1M token context window
Full ranking · 10 models
| Rank | Model | Share |
|---|---|---|
1 | DeepSeek: DeepSeek V4 Flash 0423 | |
2 | DeepSeek: DeepSeek V4 Flash 0731 | |
3 | DeepSeek: DeepSeek V4 Pro | |
4 | Z.ai: GLM 5.2 (batch) | |
5 | OpenAI: GPT-5.6 Luna (batch) | |
6 | Anthropic: Claude Opus 4.8 (batch) | |
7 | Xiaomi: MiMo-V2.5 | |
8 | MoonshotAI: Kimi K3 | |
9 | MoonshotAI: Kimi K2.6 | |
10 | Tencent: Hy3 |
Model Comparisons
- DeepSeek: DeepSeek V4 Flash 0423 vs DeepSeek: DeepSeek V4 Flash 0731
- DeepSeek: DeepSeek V4 Flash 0423 vs DeepSeek: DeepSeek V4 Pro
- DeepSeek: DeepSeek V4 Flash 0423 vs Z.ai: GLM 5.2 (batch)
- DeepSeek: DeepSeek V4 Flash 0731 vs DeepSeek: DeepSeek V4 Pro
- DeepSeek: DeepSeek V4 Flash 0731 vs Z.ai: GLM 5.2 (batch)
- DeepSeek: DeepSeek V4 Pro vs Z.ai: GLM 5.2 (batch)
- Full LLM API pricing →
Related AI Tool Tier Lists
Get notified when Math rankings move
This ranking refreshes every week. We email you when it moves — once a week, no spam.
FAQ
What is the most-used LLM for Math?
DeepSeek: DeepSeek V4 Flash 0423 by deepseek ranks first for Math by real token usage. It accounts for roughly 12.7% of tokens in this segment. DeepSeek: DeepSeek V4 Flash 0731, DeepSeek: DeepSeek V4 Pro, Z.ai: GLM 5.2 (batch) follow.
How are these rankings calculated?
Models are ranked by tokens actually consumed through OpenRouter, not by benchmark scores. This page counts only usage tagged as Math tasks. It refreshes weekly and currently tracks 10 models.
Which is the cheapest LLM for Math?
Among the paid models in this ranking, OpenAI: GPT-5.6 Luna (batch) is cheapest at $0.10 per 1M input tokens.
Is DeepSeek: DeepSeek V4 Flash 0423 still the most-used LLM for Math right now?
Yes. As of the week of 2026-08-10, DeepSeek: DeepSeek V4 Flash 0423 is #1 by real usage. It was #1 last week too. The ranking refreshes weekly, so this page changes as soon as it does.
Which model has the longest context window?
OpenAI: GPT-5.6 Luna (batch) has the longest context window in this ranking at 1.1M tokens.
Token usage and share data aggregated by OpenRouter, refreshed weekly. Prices are USD per 1M tokens. Benchmark scores come from Artificial Analysis; models it hasn't measured show —.