AI Calculator Pro

Qwen3.5 Flash

Alibaba · model · context window 1,000,000 tokens.

Pricing (per 1M tokens)

Input$0.0290
Output$0.2870
Cache read$0.00580

Specifications

Context window1,000,000 tokens
Max output250,000 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Qwen3.5 Flash has had 3 tracked price changes since we began monitoring. Most recent: input $0.0900$0.0290 on 10 August 2026. All-time low input rate: $0.0290/1M.

Input price over time ($/1M tokens)
Low $0.029High $0.090

Recorded from our daily pricing checks. See the full LLM price tracker.

Intelligence & value

Measured
Overall1398 · #76

Worth it? Qwen3.5 Flash is on the price/quality frontier — nothing we track matches its overall Arena score for less. If you can accept a small quality drop, the downgrade calculator finds cheaper options.

Intelligence scores are community Arena ratings from LMArena / arena.ai, used under CC BY 4.0. Snapshot last refreshed 27 August 2026. Scores are a relative signal, not an absolute measure of capability. A Measured score was voted on directly; an Estimated score is inherited from an identical base model (a regional/creator-prefixed hosting duplicate). See our methodology.

Where else you can run Qwen3.5 Flash

The same model is resold by 7 providers we track. The typical one charges 1.9× what the cheapest does — $0.1750 against $0.0940 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Merge Gateway$0.0290$0.2870$0.0940
Empiriolabs$0.0900$0.3680$0.1600
NanoGPT$0.1000$0.4000$0.1750
Ofox$0.1000$0.4000$0.1750
Vercel AI Gateway$0.1000$0.4000$0.1750
Zenmux$0.1000$0.4000$0.1750
Alibaba CN$0.1720$1.72$0.5590

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3.5 Flash

Qwen3.5 Flash is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000172 per request
~$5.18/month at 1,000 requests/day
Input cost
$0.0000290
Output cost
$0.000143
Per request
$0.000172
Per day
$0.1725
Per month
$5.18

Source: community-hosted (cheapest of 7 providers, merge-gateway via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.