AI Calculator Pro

Qwen3 Coder Flash

Alibaba · model · context window 1,000,000 tokens.

Pricing (per 1M tokens)

Input$0.3000
Output$1.50

Specifications

Context window1,000,000 tokens
Max output65,536 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3 Coder Flash

The same model is resold by 12 providers we track. The typical one charges 2.4× what the cheapest does — $0.6000 against $0.2510 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Alibaba CN$0.1440$0.5740$0.2510
Merge Gateway$0.1440$0.5740$0.2510
Kilo$0.1950$0.9750$0.3900
OpenRouter$0.1950$0.9750$0.3900
Alibaba$0.3000$1.50$0.6000
Edenai$0.3000$1.50$0.6000
LLM Gateway$0.3000$1.50$0.6000
Llmgateway Providers$0.3000$1.50$0.6000
Llmtr$0.3000$1.50$0.6000
NanoGPT$0.3000$1.50$0.6000

Showing the 10 cheapest of 12. The full list for every model is in the open endpoint dataset.

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3 Coder Flash

Qwen3 Coder Flash is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.00105 per request
~$31.50/month at 1,000 requests/day
Input cost
$0.000300
Output cost
$0.000750
Per request
$0.00105
Per day
$1.05
Per month
$31.50

Source: auto-discovery (models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 19 July 2026. A recent date here just means the price is still current, not that it moved.