AI Calculator Pro

Qwen3 30B A3b fp8

Alibaba · model · context window 32,768 tokens.

Pricing (per 1M tokens)

Input$0.0509
Output$0.3350

Specifications

Context window32,768 tokens
Max output32,768 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run Qwen3 30B A3b fp8

The same model is resold by 3 providers we track. The typical one charges 1.5× what the cheapest does — $0.1800 against $0.1220 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Cloudflare Workers AI$0.0510$0.3350$0.1220
Jiekou$0.0900$0.4500$0.1800
Novita AI$0.0900$0.4500$0.1800

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with Qwen3 30B A3b fp8

Qwen3 30B A3b fp8 is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.000218 per request
~$6.55/month at 1,000 requests/day
Input cost
$0.0000509
Output cost
$0.000168
Per request
$0.000218
Per day
$0.2184
Per month
$6.55

Source: community-hosted (cheapest of 3 providers, cloudflare-workers-ai via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.