AI Calculator Pro

Gemini 2.5 Flash-Lite vs Llama 4 Maverick

Cost and capability comparison. For a sample workload of 1,000 input and 500 output tokens across 100,000 requests/month, Gemini 2.5 Flash-Lite is cheaper by about $34.50/month.

Gemini 2.5 Flash-LiteLlama 4 Maverick
Providergooglemeta
Input / 1M$0.1000$0.2200
Output / 1M$0.4000$0.8500
Per request (sample)$0.000300$0.000645
Monthly (sample)$30.00$64.50
Context window1,000,0001,000,000
VisionYesYes
ReasoningNoNo
Arena · Overall13301360
Arena · Coding13251358
Arena · Math & reasoning13321352
Arena · Agentic / tool use1350

Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 28 July 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.

Compare on your own workload

Gemini 2.5 Flash-Lite and Llama 4 Maverick are pre-loaded. Adjust the tokens and monthly volume to see which is cheaper for your usage.

Used unless you set active users below.

Results update automatically as you type.

Result
Cheapest: Gemini 2.5 Flash-Lite at $30.00/month
Workload: 1,000 in / 500 out × 100,000 requests/month
Gemini 2.5 Flash-LiteGoogle$0.000300$30.00$360.001,000,000
Llama 4 MaverickMeta$0.000645$64.50$774.001,000,000

FAQ

Which is cheaper, Gemini 2.5 Flash-Lite or Llama 4 Maverick?

For a workload of 1,000 input and 500 output tokens over 100,000 requests/month, Gemini 2.5 Flash-Lite is cheaper, costing about $30.00/month versus $64.50/month.

What is the context window of Gemini 2.5 Flash-Lite and Llama 4 Maverick?

Gemini 2.5 Flash-Lite supports 1,000,000 tokens, while Llama 4 Maverick supports 1,000,000 tokens.

Estimates for planning only. Verify current pricing with each provider.