AI Calculator Pro

Qwen3.5 9B

Alibaba · model · context window 262,144 tokens.

Pricing (per 1M tokens)

Input$0.0400
Output$0.1500
Cache read$0.00800

Specifications

Context window262,144 tokens
Max output262,144 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3.5 9B

The same model is resold by 21 providers we track. The typical one charges 1.7× what the cheapest does — $0.1130 against $0.0680 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Crof$0.0400$0.1500$0.0680
NanoGPT$0.0500$0.1500$0.0750
Empiriolabs$0.0900$0.1300$0.1000
Merge Gateway$0.0900$0.1300$0.1000
DeepInfra$0.1000$0.1500$0.1130
Kilo$0.1000$0.1500$0.1130
LLM Gateway$0.1000$0.1500$0.1130
Llmgateway Providers$0.1000$0.1500$0.1130
OpenRouter$0.1000$0.1500$0.1130
Siliconflow$0.1000$0.1500$0.1130

Showing the 10 cheapest of 21. The full list for every model is in the open endpoint dataset.

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3.5 9B

Qwen3.5 9B is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000115 per request
~$3.45/month at 1,000 requests/day
Input cost
$0.0000400
Output cost
$0.0000750
Per request
$0.000115
Per day
$0.1150
Per month
$3.45

Source: community-hosted (cheapest of 21 providers, crof via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.