Nemotron 3 Super 120B A12B vs Qwen3 Max
Cost and capability comparison. For a sample workload of 1,000 input and 500 output tokens across 100,000 requests/month, Nemotron 3 Super 120B A12B is cheaper by about $376.25/month.
| Nemotron 3 Super 120B A12B | Qwen3 Max | |
|---|---|---|
| Provider | nvidia | alibaba |
| Input / 1M | $0.2100 | $1.20 |
| Output / 1M | $0.4550 | $6.00 |
| Per request (sample) | $0.000438 | $0.00420 |
| Monthly (sample) | $43.75 | $420.00 |
| Context window | 262,144 | 262,144 |
| Vision | No | No |
| Reasoning | Yes | No |
| Arena · Overall | — | 1404 |
Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 28 July 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.
Compare on your own workload
Nemotron 3 Super 120B A12B and Qwen3 Max are pre-loaded. Adjust the tokens and monthly volume to see which is cheaper for your usage.
Used unless you set active users below.
Results update automatically as you type.
| Nemotron 3 Super 120B A12B | Nvidia | $0.000438 | $43.75 | $525.00 | 262,144 |
| Qwen3 Max | Alibaba | $0.00420 | $420.00 | $5,040.00 | 262,144 |
FAQ
Which is cheaper, Nemotron 3 Super 120B A12B or Qwen3 Max?
For a workload of 1,000 input and 500 output tokens over 100,000 requests/month, Nemotron 3 Super 120B A12B is cheaper, costing about $43.75/month versus $420.00/month.
What is the context window of Nemotron 3 Super 120B A12B and Qwen3 Max?
Nemotron 3 Super 120B A12B supports 262,144 tokens, while Qwen3 Max supports 262,144 tokens.
Estimates for planning only. Verify current pricing with each provider.