DeepSeek
Lowest55 in · 11 out
Planning estimate
- Input
- $0.000008
- Cached input
- $0.00
- Output
- $0.000003
- Daily
- $0.0108
- Monthly
- $0.3234
- $ / 1k tok
- $0.000163
- Budget fit
- 46,382,189
- Per request
- $0.000011
Calculator
Paste text to count tokens and compare providers side by side—or use the advanced volume planner for request-scale budgets.
OpenAI / Grok / Perplexity cards use exact BPE (Planning estimate). Anthropic, Google, and DeepSeek use planning estimates — each card shows its own count.
20% of input when using presets
Project spend and see how many requests fit your monthly budget.
Actionable reads from this prompt + volume—not financial advice.
Special
Split traffic between your cheapest and most expensive selected models, then see monthly savings and how many days your budget lasts—share the scenario without pasting your prompt.
DeepSeek · $0.000011/req
Anthropic · $0.000333/req
At 1,000 req/day on this mix
55 in · 11 out
Planning estimate
53 in · 11 out
Planning estimate
$0.000033 more per request than the lowest option ($0.9786/mo).
54 in · 11 out
Planning estimate
$0.000054 more per request than the lowest option ($1.63/mo).
54 in · 11 out
Planning estimate
$0.000167 more per request than the lowest option ($5.00/mo).
54 in · 11 out
Planning estimate
$0.000316 more per request than the lowest option ($9.49/mo).
56 in · 11 out
Planning estimate
$0.000322 more per request than the lowest option ($9.67/mo).
Uncached input shown as 54 tokens at the OpenAI-style baseline; each card re-estimates with its family. Volume figures multiply the per-request total. Rates updated 2026-07-31. Approximate static list prices for planning only. Always verify against each provider’s official pricing page before production budgeting.
Plan multi-request workloads with cached input, batch pricing, and model multi-select—same engine as before.
Select up to 4
1,000 requests · 2,000 in / 500 out
OpenAI
$10.00
Effective rates: in $2.50/1M · out $10.00/1M
Rates last updated 2026-07-31. Approximate static list prices for planning only. Always verify against each provider’s official pricing page before production budgeting.
Guide
Estimate monthly API spend across major LLM providers before you ship. Enter input and output token volumes separately, because output tokens often cost 3–5x more than input on many models. Use this as a planning baseline and verify against each provider's official pricing page.
Total cost = (input tokens / 1,000,000 × input price per 1M) + (output tokens / 1,000,000 × output price per 1M)
Most teams underestimate LLM bills because they track requests, not tokens, and ignore the output-side price premium. A realistic input/output split prevents surprise overruns when you scale from prototype traffic to production monthly volume.
No. CentsPerToken uses approximate static list prices for planning only. Always confirm current rates on each provider's pricing page before committing to a budget.
Providers charge separately because generating output tokens consumes more compute than processing input. On many frontier models, output per-million pricing is several times higher than input.
Cached prompt tokens are billed at a reduced rate on supported providers. Use the cache savings calculator to model hit rates instead of counting cached tokens at full input price.
Multiply daily requests by average input and output tokens per request, then multiply by 30 for a rough monthly figure. Add peak-day buffer if traffic is spiky.