Cost formula
Cost for 1M input tokens = input price per 1M; Cost for 1M output tokens = output price per 1M; Tokens affordable = budget ÷ price per 1M × 1,000,000
Calculator
Pick an input/output mix (and optional cache or batch), then rank models by the cost of exactly one million tokens—and see how many tokens a fixed budget buys.
1M token reference
$0.102for 1M tokens at this mix
$0.000102 / 1K · Google
Reference
Pick an input/output/cache mix, then compare list prices across models.
Remaining 20% billed as output
Compare
Cheapest at this mix: Gemini 2.0 Flash-Lite
Showing top 8 of 71 ranked models. Spread from cheapest to priciest: $239.90.
“1M tokens” here means one million tokens under your chosen input/output/cache mix—not 1M input plus 1M output unless the mix is set that way. Updated 2026-07-31. Approximate static list prices for planning only. Always verify against each provider’s official pricing page before production budgeting.
Guide
See exactly how much one million input or output tokens costs by model, and invert the math to find how many tokens your budget buys. Useful for quick mental math when reading provider pricing pages or comparing per-million list rates across vendors.
Cost for 1M input tokens = input price per 1M; Cost for 1M output tokens = output price per 1M; Tokens affordable = budget ÷ price per 1M × 1,000,000
Provider pricing is published per million tokens, but product teams think in requests and budgets. This calculator bridges list prices and actionable volume planning without manual arithmetic.
Input and output are priced separately. One million input tokens and one million output tokens have different costs on most models.
CentsPerToken uses approximate list prices for planning. Providers update rates—confirm officially before budgeting.
Roughly 750,000 words in English for many models, but varies by language and content. Use the token counter for precise counts.
Embedding models have their own per-1M rates. Select an embedding model if available or use the embedding RAG calculator for pipeline context.