AI Calculator Pro

Gemini 2.5 Flash vs Nemotron 3 Super 120B A12B

Cost and capability comparison. For a sample workload of 1,000 input and 500 output tokens across 100,000 requests/month, Nemotron 3 Super 120B A12B is cheaper by about $111.25/month.

Gemini 2.5 FlashNemotron 3 Super 120B A12B
Providergooglenvidia
Input / 1M$0.3000$0.2100
Output / 1M$2.50$0.4550
Per request (sample)$0.00155$0.000438
Monthly (sample)$155.00$43.75
Context window1,000,000262,144
VisionYesNo
ReasoningYesYes
Arena · Overall1417
Arena · Coding1370
Arena · Math & reasoning1385
Arena · Agentic / tool use1368

Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 28 July 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.

Compare on your own workload

Gemini 2.5 Flash and Nemotron 3 Super 120B A12B are pre-loaded. Adjust the tokens and monthly volume to see which is cheaper for your usage.

Used unless you set active users below.

Results update automatically as you type.

Result
Cheapest: Nemotron 3 Super 120B A12B at $43.75/month
Workload: 1,000 in / 500 out × 100,000 requests/month
Nemotron 3 Super 120B A12BNvidia$0.000438$43.75$525.00262,144
Gemini 2.5 FlashGoogle$0.00155$155.00$1,860.001,000,000

FAQ

Which is cheaper, Gemini 2.5 Flash or Nemotron 3 Super 120B A12B?

For a workload of 1,000 input and 500 output tokens over 100,000 requests/month, Nemotron 3 Super 120B A12B is cheaper, costing about $43.75/month versus $155.00/month.

What is the context window of Gemini 2.5 Flash and Nemotron 3 Super 120B A12B?

Gemini 2.5 Flash supports 1,000,000 tokens, while Nemotron 3 Super 120B A12B supports 262,144 tokens.

Estimates for planning only. Verify current pricing with each provider.