AI Calculator Pro

Qwen3.8 Omni Flash

Alibaba · model · context window 1,000,000 tokens.

Pricing (per 1M tokens)

Input$0.1500
Output$0.4700
Cache read$0.0160

Specifications

Context window1,000,000 tokens
Max output131,072 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3.8 Omni Flash

The same model is resold by 10 providers we track. The typical one charges 1.3× what the cheapest does — $0.2300 against $0.1800 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Aihubmix$0.1130$0.3800$0.1800
Alibaba CN$0.1130$0.3820$0.1800
Crossmodel$0.1300$0.4300$0.2050
Alibaba$0.1500$0.4700$0.2300
Edenai$0.1500$0.4700$0.2300
Kilo$0.1500$0.4700$0.2300
NanoGPT$0.1500$0.4700$0.2300
OpenRouter ↗$0.1500$0.4700$0.2300
Vercel AI Gateway$0.1500$0.4700$0.2300
Empiriolabs$0.3000$0.9400$0.4600

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3.8 Omni Flash

Qwen3.8 Omni Flash is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000385 per request
~$11.55/month at 1,000 requests/day
Input cost
$0.000150
Output cost
$0.000235
Per request
$0.000385
Per day
$0.3850
Per month
$11.55

Source: auto-discovery (models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 28 September 2026. A recent date here just means the price is still current, not that it moved.