Skip to content
AI Tier ListField guide
01 · WEEKLY RANKING

Best LLMs for Roleplay & Fiction

Week of

10 models · ranked by real Roleplay & Fiction usage

Key facts · week of 2026-08-10

  1. F1

    The top 3 models (DeepSeek: DeepSeek V4 Flash 0423, DeepSeek: DeepSeek V4 Pro, Google: Gemma 4 31B (free)) account for 41.5% of all usage.

    §
  2. F2

    The cheapest model in the usage top 10 is Google: Gemini 2.5 Flash Lite (batch) at $0.05 per 1M input tokens (usage rank #7).

    §
  3. F3

    The largest context window in the top 10 is Xiaomi: MiMo-V2.5 at 1.1M tokens.

    §

Cite as: AI Tier List, LLM Usage Rankings (2026-08-10), www.aitierlist.xyz/en/models/roleplay-fiction — data: OpenRouter

Which model should you pick?

When quality comes first
DeepSeek: DeepSeek V4 Flash 0731deepseek
Artificial Analysis intelligence index score 51.8 — highest in this ranking
The proven default
DeepSeek: DeepSeek V4 Flash 0423deepseek
25.7% usage share — #1 in real usage
When cost matters
Google: Gemma 4 31B (free)google
Cheapest in the top 15 — Free/1M input
For long docs & codebases
Xiaomi: MiMo-V2.5xiaomi
1.1M token context window

Full ranking · 10 models

RankModelShare
1
DeepSeek: DeepSeek V4 Flash 0423
in $0.14 · out $0.28 · 1M
25.7%
2
DeepSeek: DeepSeek V4 Pro
Intelligence 45.3 · in $1.17 · out $2.34 · 1M
8.2%
3
Google: Gemma 4 31B (free)
Intelligence 29.7 · in Free · out Free · 262K
7.6%
4
DeepSeek: DeepSeek V3.2
Intelligence 32.6 · in $0.27 · out $0.40 · 164K
6.1%
5
Google: Gemini 3 Flash Preview (batch)
in $0.25 · out $1.50 · 1M
4.7%
6
Google: Gemma 4 26B A4B (free)
Intelligence 26.1 · in Free · out Free · 262K
4.2%
71
Google: Gemini 2.5 Flash Lite (batch)
in $0.05 · out $0.20 · 1M
3.7%
81
Xiaomi: MiMo-V2.5
Intelligence 38 · in $0.14 · out $0.28 · 1.1M
3.6%
9NEW
DeepSeek: DeepSeek V4 Flash 0731
Intelligence 51.8 · in $0.14 · out $0.28 · 1M
3.1%
10
Tencent: Hy3
in $0.13 · out $0.53 · 262K
2.9%
Share these rankingsXReddit

Get notified when Roleplay & Fiction rankings move

This ranking refreshes every week. We email you when it moves — once a week, no spam.

FAQ

What is the most-used LLM for Roleplay & Fiction?

DeepSeek: DeepSeek V4 Flash 0423 by deepseek ranks first for Roleplay & Fiction by real token usage. It accounts for roughly 25.7% of tokens in this segment. DeepSeek: DeepSeek V4 Pro, Google: Gemma 4 31B (free), DeepSeek: DeepSeek V3.2 follow.

How are these rankings calculated?

Models are ranked by tokens actually consumed through OpenRouter, not by benchmark scores. This page counts only usage tagged as Roleplay & Fiction tasks. It refreshes weekly and currently tracks 10 models.

Which is the cheapest LLM for Roleplay & Fiction?

Among the paid models in this ranking, Google: Gemini 2.5 Flash Lite (batch) is cheapest at $0.05 per 1M input tokens. 2 models in the list are free to use.

Is DeepSeek: DeepSeek V4 Flash 0423 still the most-used LLM for Roleplay & Fiction right now?

Yes. As of the week of 2026-08-10, DeepSeek: DeepSeek V4 Flash 0423 is #1 by real usage. It was #1 last week too. The ranking refreshes weekly, so this page changes as soon as it does.

Which model has the longest context window?

Xiaomi: MiMo-V2.5 has the longest context window in this ranking at 1.1M tokens.

Token usage and share data aggregated by OpenRouter, refreshed weekly. Prices are USD per 1M tokens. Benchmark scores come from Artificial Analysis; models it hasn't measured show —.