Llama-4-Maverick-17B-128E-Instruct-FP8
Meta · model · context window 128,000 tokens.
Pricing (per 1M tokens)
Input$0.2000
Output$0.8000
Specifications
Context window128,000 tokens
Max output4,096 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo
Estimate cost with Llama-4-Maverick-17B-128E-Instruct-FP8
Llama-4-Maverick-17B-128E-Instruct-FP8 is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.
Results update automatically as you type.
Result
$0.000600 per request
~$18.00/month at 1,000 requests/day
- Input cost
- $0.000200
- Output cost
- $0.000400
- Per request
- $0.000600
- Per day
- $0.6000
- Per month
- $18.00
Source: representative host (deepinfra via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 19 July 2026. A recent date here just means the price is still current, not that it moved.