AI Calculator Pro

Llama-4-Maverick-17B-128E-Instruct-FP8

Meta · model · context window 128,000 tokens.

Pricing (per 1M tokens)

Input$0.2000
Output$0.8000

Specifications

Context window128,000 tokens
Max output4,096 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Estimate cost with Llama-4-Maverick-17B-128E-Instruct-FP8

Llama-4-Maverick-17B-128E-Instruct-FP8 is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000600 per request
~$18.00/month at 1,000 requests/day
Input cost
$0.000200
Output cost
$0.000400
Per request
$0.000600
Per day
$0.6000
Per month
$18.00

Source: representative host (deepinfra via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 19 July 2026. A recent date here just means the price is still current, not that it moved.