AI Calculator Pro

Llama 3.2 11B Vision Instruct

Meta · model · context window 16,000 tokens.

Pricing (per 1M tokens)

Input$0.0550
Output$0.0550

Specifications

Context window16,000 tokens
Max output4,096 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Llama 3.2 11B Vision Instruct

The same model is resold by 3 providers we track. The typical one charges 3.7× what the cheapest does — $0.2060 against $0.0550 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Inference$0.0550$0.0550$0.0550
Cloudflare Workers AI$0.0490$0.6760$0.2060
Edenai$0.3450$0.3450$0.3450

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Llama 3.2 11B Vision Instruct

Llama 3.2 11B Vision Instruct is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.0000825 per request
~$2.48/month at 1,000 requests/day
Input cost
$0.0000550
Output cost
$0.0000275
Per request
$0.0000825
Per day
$0.0825
Per month
$2.48

Source: community-hosted (cheapest of 3 providers, inference via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.