AI Calculator Pro

Qwen3 VL 30B A3B Instruct

Alibaba · model · context window 262,144 tokens.

Pricing (per 1M tokens)

Input$0.1300
Output$0.5200

Specifications

Context window262,144 tokens
Max output32,768 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3 VL 30B A3B Instruct

The same model is resold by 8 providers we track. The typical one charges 1.1× what the cheapest does — $0.2620 against $0.2280 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Kilo$0.1300$0.5200$0.2280
OpenRouter$0.1300$0.5200$0.2280
Nearai$0.1500$0.5500$0.2500
LLM Gateway$0.1500$0.6000$0.2620
Llmgateway Providers$0.1500$0.6000$0.2620
Novita AI$0.2000$0.7000$0.3250
Siliconflow$0.2900$1.00$0.4670
Siliconflow CN$0.2900$1.00$0.4670

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3 VL 30B A3B Instruct

Qwen3 VL 30B A3B Instruct is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000390 per request
~$11.70/month at 1,000 requests/day
Input cost
$0.000130
Output cost
$0.000260
Per request
$0.000390
Per day
$0.3900
Per month
$11.70

Source: community-hosted (cheapest of 8 providers, openrouter via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.