AI Calculator Pro

Qwen3 VL 235B A22B Thinking

Alibaba · model · context window 131,072 tokens.

Pricing (per 1M tokens)

Input$0.2870
Output$2.87
Cache read$0.0574

Specifications

Context window131,072 tokens
Max output32,768 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Qwen3 VL 235B A22B Thinking has had 2 tracked price changes since we began monitoring. Most recent: input $0.4500$0.2870 on 18 August 2026. All-time low input rate: $0.2870/1M.

Input price over time ($/1M tokens)
Low $0.287High $0.450

Recorded from our daily pricing checks. See the full LLM price tracker.

Intelligence & value

Measured
Overall1401 · #74

Worth it? Step 3.5 Flash matches or beats Qwen3 VL 235B A22B Thinking on the overall Arena and costs about $30.00/month versus $186.40/month at 100,000 calls — roughly $156.40/month cheaper.

Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 27 August 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.

Where else you can run Qwen3 VL 235B A22B Thinking

The same model is resold by 12 providers we track. The typical one charges 1.6× what the cheapest does — $1.51 against $0.9320 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Merge Gateway$0.2870$2.87$0.9320
Siliconflow$0.4500$3.50$1.21
Siliconflow CN$0.4500$3.50$1.21
Edenai$0.4000$4.00$1.30
Kilo$0.4000$4.00$1.30
OpenRouter$0.4000$4.00$1.30
Huggingface$0.9800$3.95$1.72
Jalapeno$0.9800$3.95$1.72
LLM Gateway$0.9800$3.95$1.72
Llmgateway Providers$0.9800$3.95$1.72

Showing the 10 cheapest of 12. The full list for every model is in the open endpoint dataset.

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3 VL 235B A22B Thinking

Qwen3 VL 235B A22B Thinking is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.00172 per request
~$51.61/month at 1,000 requests/day
Input cost
$0.000287
Output cost
$0.00143
Per request
$0.00172
Per day
$1.72
Per month
$51.61

Source: community-hosted (cheapest of 12 providers, merge-gateway via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.