OpenAI: gpt-oss-120b vs StepFun: Step 3.7 Flash
Based on weekly OpenRouter usage tracked by AI Tier List, StepFun: Step 3.7 Flash ranks #14 overall — ahead of OpenAI: gpt-oss-120b at #26. OpenAI: gpt-oss-120b is cheaper on input tokens at $0.03/1M vs $0.20/1M for StepFun: Step 3.7 Flash. StepFun: Step 3.7 Flash offers the larger context window at 262K tokens.
OpenAI: gpt-oss-120b vs StepFun: Step 3.7 Flash: which should you use?
Choose OpenAI: gpt-oss-120b if…
- Cheaper to read — $0.03 vs $0.20 per 1M input tokens, which is what dominates long-prompt and RAG workloads.
- Cheaper to write — $0.17 vs $1.15 per 1M output tokens, the line item that decides generation-heavy costs.
Choose StepFun: Step 3.7 Flash if…
- It sees more real traffic — #14 overall on OpenRouter's weekly ranking, against #26 for OpenAI: gpt-oss-120b.
- Bigger context window — 262K vs 131K tokens.
- Higher Artificial Analysis intelligence index — 30.9 vs 24.1.
- Higher Artificial Analysis coding score — 39.6 vs 30.4.
- Higher Artificial Analysis agentic score — 21.7 vs 13.4.
- It accepts image, video input, which OpenAI: gpt-oss-120b does not.
- It is the newer model — released May 2026 against Aug 2025.
Spec Comparison
| Spec | OpenAI: gpt-oss-120b | StepFun: Step 3.7 Flash |
|---|---|---|
| Author | openai | stepfun |
| Input price (per 1M tokens) | $0.03 | $0.20 |
| Output price (per 1M tokens) | $0.17 | $1.15 |
| Context window | 131K | 262K |
| Input modalities | text | text, image, video |
| Output modalities | text | text |
| AA Intelligence | 24.1 | 30.9 |
| AA Coding | 30.4 | 39.6 |
| AA Agentic | 13.4 | 21.7 |
| Released | Aug 2025 | May 2026 |
What OpenAI: gpt-oss-120b and StepFun: Step 3.7 Flash actually cost to run
On a read-heavy job — 1M tokens in, 100K out — OpenAI: gpt-oss-120b costs $0.05 and StepFun: Step 3.7 Flash costs $0.32.
Flip the shape — 100K in, 1M out — and it becomes $0.17 for OpenAI: gpt-oss-120b versus $1.17 for StepFun: Step 3.7 Flash.
Both figures are the published per-token rates multiplied out; caching and batch discounts are not included.
Usage Rank by Task
| Task | OpenAI: gpt-oss-120b | StepFun: Step 3.7 Flash |
|---|---|---|
| Overall | #26 | #14 |
| Memory Extraction | — | #5 |
| Tool Dispatch | — | #8 |
| Agent Workflows | — | #8 |
| Classification & Tagging | #6 | — |
| Debugging | — | #9 |
| DevOps Config | — | #9 |
| File Read/Write | — | #8 |
| Frontend UI | — | #8 |
| Code Implementation | — | #10 |
| Repo Analysis | — | #9 |
| Code Review & Security | — | #10 |
| Shell Execution | — | #8 |
| SQL & Database | — | #8 |
| Conversation | — | #9 |
About OpenAI: gpt-oss-120b and StepFun: Step 3.7 Flash
OpenAI: gpt-oss-120b
Built by openai · released Aug 2025 · 131K-token context
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
StepFun: Step 3.7 Flash
Built by stepfun · released May 2026 · 262K-token context
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Related comparisons
- OpenAI: gpt-oss-120b vs DeepSeek: DeepSeek V4 Flash 0731
- OpenAI: gpt-oss-120b vs Tencent: Hy3
- OpenAI: gpt-oss-120b vs OpenAI: GPT-5.6 Luna (batch)
- StepFun: Step 3.7 Flash vs DeepSeek: DeepSeek V4 Flash 0731
- StepFun: Step 3.7 Flash vs Tencent: Hy3
- StepFun: Step 3.7 Flash vs OpenAI: GPT-5.6 Luna (batch)
Get notified when OpenAI: gpt-oss-120b or StepFun: Step 3.7 Flash moves
Usage ranks and prices refresh weekly. We email you when they move — once a week, no spam.
FAQ
Which is better, OpenAI: gpt-oss-120b or StepFun: Step 3.7 Flash?
Based on weekly OpenRouter usage tracked by AI Tier List, StepFun: Step 3.7 Flash ranks #14 overall — ahead of OpenAI: gpt-oss-120b at #26. OpenAI: gpt-oss-120b is cheaper on input tokens at $0.03/1M vs $0.20/1M for StepFun: Step 3.7 Flash. StepFun: Step 3.7 Flash offers the larger context window at 262K tokens.
Which is cheaper, OpenAI: gpt-oss-120b or StepFun: Step 3.7 Flash?
OpenAI: gpt-oss-120b is cheaper at $0.03 per 1M input tokens. Output pricing: OpenAI: gpt-oss-120b $0.17, StepFun: Step 3.7 Flash $1.15.
Which is better for coding, OpenAI: gpt-oss-120b or StepFun: Step 3.7 Flash?
StepFun: Step 3.7 Flash. Artificial Analysis coding scores: OpenAI: gpt-oss-120b 30.4, StepFun: Step 3.7 Flash 39.6. Best coding-task usage rank this week: OpenAI: gpt-oss-120b —, StepFun: Step 3.7 Flash #8.
Is StepFun: Step 3.7 Flash still ahead of OpenAI: gpt-oss-120b right now?
Yes. For the week of 2026-08-10, StepFun: Step 3.7 Flash ranks #14 overall on OpenRouter usage against #26 for OpenAI: gpt-oss-120b. Ranks refresh weekly, and the verdict on this page flips if they cross.
Usage ranks are weekly data aggregated by OpenRouter. Prices are USD per 1M tokens.