AI Calculator Pro

DeepSeek R1 Distill Llama 70B

DeepSeek · model · context window 128,000 tokens.

Pricing (per 1M tokens)

Input$0.0300
Output$0.1300

Specifications

Context window128,000 tokens
Max output4,096 tokens
Modalitiestext
VisionNo
Tool useYes
ReasoningNo
Prompt cachingNo

Where else you can run DeepSeek R1 Distill Llama 70B

The same model is resold by 7 providers we track. The typical one charges 14.5× what the cheapest does — $0.8000 against $0.0550 per 1M tokens, blended 3:1. Routing the same workload to a different host is often a bigger saving than switching model.

ProviderInput / 1MOutput / 1MBlended
Helicone$0.0300$0.1300$0.0550
Fastrouter$0.0300$0.1400$0.0580
Alibaba CN$0.2870$0.8610$0.4310
Kilo$0.8000$0.8000$0.8000
Novita AI$0.8000$0.8000$0.8000
OpenRouter$0.8000$0.8000$0.8000
DigitalOcean$0.9900$0.9900$0.9900

Third-party listings via models.dev, not rates we verify with each host the way we do first-party pricing. Treat them as a shortlist to check, not a quote — and confirm quantisation, context limits and rate limits before switching, since a cheaper host is not always serving the same thing.

Estimate cost with DeepSeek R1 Distill Llama 70B

DeepSeek R1 Distill Llama 70B is pre-selected below. Enter your token counts and volume to see the cost per request, per day and per month — then switch models to compare.

Results update automatically as you type.

Result
$0.0000950 per request
~$2.85/month at 1,000 requests/day
Input cost
$0.0000300
Output cost
$0.0000650
Per request
$0.0000950
Per day
$0.0950
Per month
$2.85

Source: community-hosted (cheapest of 7 providers, helicone via models.dev). Prices are re-checked daily against the provider’s official pricing; this rate was last changed on 27 August 2026. A recent date here just means the price is still current, not that it moved.