Output Token Cost Calculator
Estimate the cost of model output length.
With the default inputs, $0.00800 per response — ~$240.00/month at 1,000 calls/day. Enter your own numbers below to recompute instantly; the full step-by-step math is shown under the worked example.
Results update automatically as you type.
- Output tokens
- 800 tokens
- Output rate
- $10/M
- Per response
- $0.00800
- Per month
- $240.00
- Model max output
- 128,000 tokens
Output tokens usually cost several times more than input. Set the response length and volume to see what generation alone costs per response and per month.
How this is calculated
Cost = output tokens ÷ 1,000,000 × the model's output price, scaled by calls per day × 30 for the monthly figure. We isolate generation because output tokens are usually billed 3–5× higher than input, so response length — not prompt size — often drives the bill.
Is this a good result? What to do next
If output cost dominates your per-call total, response length is your biggest lever. A verbose model set to 'explain everything' can cost several times a concise one for the same answer. Judge the monthly figure against the value of the content generated.
Typical planning ranges
- Output vs input price ratio
- ~3–5×
- Mid-tier output
- ~$5–$15 per 1M tokens
- Frontier/reasoning output
- ~$15–$75 per 1M tokens
Ranges are typical planning figures to sanity-check your result, not authoritative benchmarks. Your numbers will vary with use case, volume, and vendor.
How to improve this number
- Cap max output tokens and prompt for concise answers.
- Route long generations to a cheaper model where quality allows.
- On reasoning models, account for hidden thinking tokens billed as output.
Common mistakes
- Budgeting only input tokens and ignoring the pricier output side.
- Forgetting reasoning models bill invisible thinking tokens as output.
When to use a different approach
To include input and volume, use the LLM API cost calculator. For hidden reasoning-token overhead, use the reasoning token cost calculator.
Worked example (defaults)
With the default inputs above, here is the result:
- Output tokens
- 800 tokens
- Output rate
- $10/M
- Per response
- $0.00800
- Per month
- $240.00
- Model max output
- 128,000 tokens
Sources & references
Frequently asked questions
Why is output more expensive than input?+
Generating tokens is more compute-intensive than reading them, so providers price output 3-5x higher. Reasoning models also bill hidden reasoning tokens as output.
Related calculators
Answering a real question?
This calculator powers these problem-solving guides: