AI Calculator Pro

Qwen3.7 Flash

Alibaba · model · context window 1,000,000 tokens.

Pricing (per 1M tokens)

Input$0.0296
Output$0.1185
Cache read$0.00296
Cache write$0.0370

Specifications

Context window1,000,000 tokens
Max output65,536 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3.7 Flash

The same model is resold by 10 providers we track. The typical one charges 1.1× what the cheapest does — $0.0550 against $0.0520 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Alibaba CN$0.0300$0.1190$0.0520
Empiriolabs$0.0300$0.1300$0.0550
Kilo$0.0300$0.1300$0.0550
LLM Gateway$0.0300$0.1300$0.0550
Llmgateway Providers$0.0300$0.1300$0.0550
NanoGPT$0.0300$0.1300$0.0550
OpenRouter$0.0300$0.1300$0.0550
Vercel AI Gateway$0.0300$0.1300$0.0550
Crossmodel$0.0400$0.1300$0.0630
Hyper$0.2000$0.8000$0.3500

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3.7 Flash

Qwen3.7 Flash is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.0000889 per request
~$2.67/month at 1,000 requests/day
Input cost
$0.0000296
Output cost
$0.0000592
Per request
$0.0000889
Per day
$0.0889
Per month
$2.67

Source: community-hosted (cheapest of 10 providers, alibaba-cn via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.