Google: Gemini 2.5 Flash Lite (batch) vs Google: Gemma 4 31B (free)
Based on weekly OpenRouter usage tracked by AI Tier List, Google: Gemini 2.5 Flash Lite (batch) ranks #20 overall — ahead of Google: Gemma 4 31B (free) at #23. Google: Gemma 4 31B (free) is cheaper on input tokens at Free/1M vs $0.05/1M for Google: Gemini 2.5 Flash Lite (batch). Google: Gemini 2.5 Flash Lite (batch) offers the larger context window at 1M tokens.
Google: Gemini 2.5 Flash Lite (batch) vs Google: Gemma 4 31B (free): which should you use?
Choose Google: Gemini 2.5 Flash Lite (batch) if…
- It sees more real traffic — #20 overall on OpenRouter's weekly ranking, against #23 for Google: Gemma 4 31B (free).
- Bigger context window — 1M vs 262K tokens.
- It accepts file, audio input, which Google: Gemma 4 31B (free) does not.
Choose Google: Gemma 4 31B (free) if…
- Cheaper to read — Free vs $0.05 per 1M input tokens, which is what dominates long-prompt and RAG workloads.
- Cheaper to write — Free vs $0.20 per 1M output tokens, the line item that decides generation-heavy costs.
- It is the newer model — released Apr 2026 against Jul 2025.
Spec Comparison
| Spec | Google: Gemini 2.5 Flash Lite (batch) | Google: Gemma 4 31B (free) |
|---|---|---|
| Author | ||
| Input price (per 1M tokens) | $0.05 | Free |
| Output price (per 1M tokens) | $0.20 | Free |
| Context window | 1M | 262K |
| Input modalities | text, image, file, audio, video | image, text, video |
| Output modalities | text | text |
| AA Intelligence | — | 29.7 |
| AA Coding | — | 43.4 |
| AA Agentic | — | 14.4 |
| Released | Jul 2025 | Apr 2026 |
What Google: Gemini 2.5 Flash Lite (batch) and Google: Gemma 4 31B (free) actually cost to run
On a read-heavy job — 1M tokens in, 100K out — Google: Gemini 2.5 Flash Lite (batch) costs $0.07 and Google: Gemma 4 31B (free) costs $0.00.
Flip the shape — 100K in, 1M out — and it becomes $0.21 for Google: Gemini 2.5 Flash Lite (batch) versus $0.00 for Google: Gemma 4 31B (free).
Both figures are the published per-token rates multiplied out; caching and batch discounts are not included.
Usage Rank by Task
| Task | Google: Gemini 2.5 Flash Lite (batch) | Google: Gemma 4 31B (free) |
|---|---|---|
| Overall | #20 | #23 |
| Classification & Tagging | #4 | — |
| Data Extraction | #3 | — |
| Data Transformation | #4 | — |
| Image Processing | #1 | #8 |
| Roleplay & Fiction | #7 | #3 |
| Summarization | #8 | — |
| Translation | #4 | #3 |
About Google: Gemini 2.5 Flash Lite (batch) and Google: Gemma 4 31B (free)
Google: Gemini 2.5 Flash Lite (batch)
Built by google · released Jul 2025 · 1M-token context
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Google: Gemma 4 31B (free)
Built by google · released Apr 2026 · 262K-token context
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Related comparisons
- Google: Gemini 2.5 Flash Lite (batch) vs DeepSeek: DeepSeek V4 Flash 0731
- Google: Gemini 2.5 Flash Lite (batch) vs Tencent: Hy3
- Google: Gemini 2.5 Flash Lite (batch) vs OpenAI: GPT-5.6 Luna (batch)
- Google: Gemma 4 31B (free) vs DeepSeek: DeepSeek V4 Flash 0731
- Google: Gemma 4 31B (free) vs Tencent: Hy3
- Google: Gemma 4 31B (free) vs OpenAI: GPT-5.6 Luna (batch)
Get notified when Google: Gemini 2.5 Flash Lite (batch) or Google: Gemma 4 31B (free) moves
Usage ranks and prices refresh weekly. We email you when they move — once a week, no spam.
FAQ
Which is better, Google: Gemini 2.5 Flash Lite (batch) or Google: Gemma 4 31B (free)?
Based on weekly OpenRouter usage tracked by AI Tier List, Google: Gemini 2.5 Flash Lite (batch) ranks #20 overall — ahead of Google: Gemma 4 31B (free) at #23. Google: Gemma 4 31B (free) is cheaper on input tokens at Free/1M vs $0.05/1M for Google: Gemini 2.5 Flash Lite (batch). Google: Gemini 2.5 Flash Lite (batch) offers the larger context window at 1M tokens.
Which is cheaper, Google: Gemini 2.5 Flash Lite (batch) or Google: Gemma 4 31B (free)?
Google: Gemma 4 31B (free) is cheaper at Free per 1M input tokens. Output pricing: Google: Gemini 2.5 Flash Lite (batch) $0.20, Google: Gemma 4 31B (free) Free.
Is Google: Gemini 2.5 Flash Lite (batch) still ahead of Google: Gemma 4 31B (free) right now?
Yes. For the week of 2026-08-10, Google: Gemini 2.5 Flash Lite (batch) ranks #20 overall on OpenRouter usage against #23 for Google: Gemma 4 31B (free). Ranks refresh weekly, and the verdict on this page flips if they cross.
Usage ranks are weekly data aggregated by OpenRouter. Prices are USD per 1M tokens.