LLM API cost calculator
One workload, every provider. Pick a model to see its line items, then read the table below to see what the same tokens cost elsewhere.
Gemini 2.5 Pro at 1,000 requests per day costs $0.009750 per request, $9.75 per day, $292.50 per month and $3,558.75 per year.
Prices effective 2026-09-06Verified 30 day month, 365 day year
| Line | Quantity per request | Unit price | Per request | Per month |
|---|---|---|---|---|
| input | 3,000 | $1.25 per 1m tokens | $0.003750 | $112.50 |
| output | 600 | $10.00 per 1m tokens | $0.006000 | $180.00 |
- Long-context summarisation (Gemini 2.5 Pro, 250k in)
- Support bot (Claude Haiku 4.5)
- Enterprise GPT-5 on Azure
Same workload on other models
| Model | Provider | Per request | Per month | Change from selected |
|---|---|---|---|---|
| GPT-5 nano | Azure OpenAI (Foundry Models) | $0.000390 | $11.70 | -$280.80 |
| GPT-5 nano | OpenAI | $0.000390 | $11.70 | -$280.80 |
| GPT-4.1 nano | Azure OpenAI (Foundry Models) | $0.000540 | $16.20 | -$276.30 |
| Gemini 2.5 Flash-Lite | Google (Gemini API) | $0.000540 | $16.20 | -$276.30 |
| GPT-4.1 nano | OpenAI | $0.000540 | $16.20 | -$276.30 |
| GPT-4o mini | Azure OpenAI (Foundry Models) | $0.000810 | $24.30 | -$268.20 |
Full ranking on the comparison page. Want CI to fail when a change pushes the monthly figure past a limit? See budget contracts.
How the number is computed
- Each token dimension is multiplied by its list price per 1M tokens in integer micro-dollars, then rounded half up once per line. No floating point.
- Cache write tokens use the 5 minute TTL price. Cache read tokens use the cache read price. Both are in addition to uncached input tokens.
- Batch API applies the model's published multiplier (50% on every current model) to every token line.
- Month = 30 days, year = 365 days. Taxes, volume discounts and provider commitments are excluded.
- Long-context tiers (Gemini above 200k input tokens) switch automatically when the input token count crosses the threshold.
Sources and verification
- Anthropic: https://platform.claude.com/docs/en/about-claude/pricing
- OpenAI: https://developers.openai.com/api/docs/pricing
- Google (Gemini API): https://ai.google.dev/gemini-api/docs/pricing
- Azure OpenAI (Foundry Models): https://azure.microsoft.com/en-us/pricing/details/cognitive-services/openai-service/