NVIDIA: Nemotron 3 Ultra (free) vs Z.ai: GLM 5.2 (batch)
Based on weekly OpenRouter usage tracked by AI Tier List, Z.ai: GLM 5.2 (batch) ranks #6 overall — ahead of NVIDIA: Nemotron 3 Ultra (free) at #8. NVIDIA: Nemotron 3 Ultra (free) is cheaper on input tokens at Free/1M vs $0.70/1M for Z.ai: GLM 5.2 (batch). NVIDIA: Nemotron 3 Ultra (free) offers the larger context window at 1M tokens.
NVIDIA: Nemotron 3 Ultra (free) vs Z.ai: GLM 5.2 (batch): which should you use?
Choose NVIDIA: Nemotron 3 Ultra (free) if…
- Cheaper to read — Free vs $0.70 per 1M input tokens, which is what dominates long-prompt and RAG workloads.
- Cheaper to write — Free vs $2.20 per 1M output tokens, the line item that decides generation-heavy costs.
- Bigger context window — 1M vs 512K tokens.
Choose Z.ai: GLM 5.2 (batch) if…
- It sees more real traffic — #6 overall on OpenRouter's weekly ranking, against #8 for NVIDIA: Nemotron 3 Ultra (free).
- It is the newer model — released Jun 2026 against Jun 2026.
Spec Comparison
| Spec | NVIDIA: Nemotron 3 Ultra (free) | Z.ai: GLM 5.2 (batch) |
|---|---|---|
| Author | nvidia | z-ai |
| Input price (per 1M tokens) | Free | $0.70 |
| Output price (per 1M tokens) | Free | $2.20 |
| Context window | 1M | 512K |
| Input modalities | text | text |
| Output modalities | text | text |
| AA Intelligence | — | — |
| AA Coding | — | — |
| AA Agentic | — | — |
| Released | Jun 2026 | Jun 2026 |
What NVIDIA: Nemotron 3 Ultra (free) and Z.ai: GLM 5.2 (batch) actually cost to run
On a read-heavy job — 1M tokens in, 100K out — NVIDIA: Nemotron 3 Ultra (free) costs $0.00 and Z.ai: GLM 5.2 (batch) costs $0.92.
Flip the shape — 100K in, 1M out — and it becomes $0.00 for NVIDIA: Nemotron 3 Ultra (free) versus $2.27 for Z.ai: GLM 5.2 (batch).
Both figures are the published per-token rates multiplied out; caching and batch discounts are not included.
Usage Rank by Task
| Task | NVIDIA: Nemotron 3 Ultra (free) | Z.ai: GLM 5.2 (batch) |
|---|---|---|
| Overall | #8 | #6 |
| Memory Extraction | #4 | #7 |
| Multi-step Planning | #9 | #5 |
| Tool Dispatch | #9 | #5 |
| Web Search | #7 | #6 |
| Agent Workflows | #8 | #4 |
| Debugging | #6 | #2 |
| DevOps Config | #4 | #3 |
| File Read/Write | #7 | #5 |
| Frontend UI | #7 | #3 |
| Code Implementation | #7 | #2 |
| Repo Analysis | #6 | #7 |
| Code Review & Security | #6 | #1 |
| Shell Execution | #9 | #5 |
| SQL & Database | #7 | #3 |
About NVIDIA: Nemotron 3 Ultra (free) and Z.ai: GLM 5.2 (batch)
NVIDIA: Nemotron 3 Ultra (free)
Built by nvidia · released Jun 2026 · 1M-token context
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Z.ai: GLM 5.2 (batch)
Built by z-ai · released Jun 2026 · 512K-token context
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
Related comparisons
- NVIDIA: Nemotron 3 Ultra (free) vs DeepSeek: DeepSeek V4 Flash 0731
- NVIDIA: Nemotron 3 Ultra (free) vs Tencent: Hy3
- NVIDIA: Nemotron 3 Ultra (free) vs OpenAI: GPT-5.6 Luna (batch)
- Z.ai: GLM 5.2 (batch) vs DeepSeek: DeepSeek V4 Flash 0731
- Z.ai: GLM 5.2 (batch) vs Tencent: Hy3
- Z.ai: GLM 5.2 (batch) vs OpenAI: GPT-5.6 Luna (batch)
Get notified when NVIDIA: Nemotron 3 Ultra (free) or Z.ai: GLM 5.2 (batch) moves
Usage ranks and prices refresh weekly. We email you when they move — once a week, no spam.
FAQ
Which is better, NVIDIA: Nemotron 3 Ultra (free) or Z.ai: GLM 5.2 (batch)?
Based on weekly OpenRouter usage tracked by AI Tier List, Z.ai: GLM 5.2 (batch) ranks #6 overall — ahead of NVIDIA: Nemotron 3 Ultra (free) at #8. NVIDIA: Nemotron 3 Ultra (free) is cheaper on input tokens at Free/1M vs $0.70/1M for Z.ai: GLM 5.2 (batch). NVIDIA: Nemotron 3 Ultra (free) offers the larger context window at 1M tokens.
Which is cheaper, NVIDIA: Nemotron 3 Ultra (free) or Z.ai: GLM 5.2 (batch)?
NVIDIA: Nemotron 3 Ultra (free) is cheaper at Free per 1M input tokens. Output pricing: NVIDIA: Nemotron 3 Ultra (free) Free, Z.ai: GLM 5.2 (batch) $2.20.
Which is better for coding, NVIDIA: Nemotron 3 Ultra (free) or Z.ai: GLM 5.2 (batch)?
Z.ai: GLM 5.2 (batch). Best coding-task usage rank this week: NVIDIA: Nemotron 3 Ultra (free) #4, Z.ai: GLM 5.2 (batch) #1.
Is Z.ai: GLM 5.2 (batch) still ahead of NVIDIA: Nemotron 3 Ultra (free) right now?
Yes. For the week of 2026-08-10, Z.ai: GLM 5.2 (batch) ranks #6 overall on OpenRouter usage against #8 for NVIDIA: Nemotron 3 Ultra (free). Ranks refresh weekly, and the verdict on this page flips if they cross.
Usage ranks are weekly data aggregated by OpenRouter. Prices are USD per 1M tokens.