AI API Cost Calculator
Estimate your monthly LLM spend in real time. Compare GPT-5, Claude, DeepSeek, Gemini and Llama pricing, then apply cached-input pricing and cost-based routing to see what you can save.
⚙️ Usage & Model
💰 Cost Optimization
📊 Your Estimated Cost
Prices are list prices (USD per 1M tokens) from official provider pricing. Your actual bill depends on provider discounts and usage patterns.
Model Comparison
Same workload priced across models — your selection, GPT-5 (premium) and DeepSeek Chat (budget), with your cache & routing options applied.
| Model | Input $/1M | Output $/1M | Cost / Request | Monthly Cost | Relative to GPT-5 |
|---|
How the calculation works
LLM APIs bill per token — roughly 4 characters or 0.75 English words. Input tokens (your prompt) and output tokens (the model's reply) are priced separately, and costs scale linearly with usage:
- Cached-input pricing: providers heavily discount input tokens already seen before (cache hits). This calculator applies the discount to the whole request cost at your chosen hit rate — e.g. 23% of requests get the cheaper path.
- Cost-based routing: a router like DrAI's sends each request to the cheapest model that can handle it (e.g. GPT-5-mini or DeepSeek Chat for simple tasks), cutting blended spend by up to 62%.
- The comparison baseline is GPT-5 at full list price with no discounts — the premium default most teams start with.
Frequently asked questions
Which models are included in the calculator?
How accurate is the estimate?
What is cached-input pricing?
How does cost-based routing save 62%?
Stop guessing your AI bill
Run the same workloads on DrAI with transparent per-token pricing, cached-input discounts and smart routing — and pay only for what you use.