LLM API Cost Calculator

Price your actual workload, not a marketing table. Tokens per request, daily volume, cache hit rate, and batch discounts across Claude, GPT, and Gemini. Every price is editable. Runs entirely in your browser.

0% of input cached

Cached input is billed at roughly 10% of the input rate across providers; the slider applies that to the cached share. Long-context surcharges apply automatically when input tokens per request cross a provider's threshold. Prices are standard tier per 1M tokens, verified , and editable below if they have drifted.

Model$/1M in · outPer requestPer dayPer month

How to read this honestly

For the sizing side of the same conversation, pair this with the Context Window Budget Calculator. For why most AI project budgets die anyway: The $400B Lie.