Quick Answer — How Much Will My LLM API Usage Cost?
The core formula: Cost per request = (Input tokens ÷ 1,000,000 × Input price) + (Output tokens ÷ 1,000,000 × Output price). Monthly cost = Cost per request × Requests per month.
Quick reference (no caching):
- 50,000 requests/month, 800 input + 300 output tokens, $0.15/$0.60 per 1M → $15.00/month ($180.00/year)
- 2,000,000 requests/month, 200 input + 20 output tokens, $0.15/$0.60 per 1M → $84.00/month ($1,008.00/year)
- 5,000 requests/month, 4,000 input + 1,000 output tokens, $3.00/$15.00 per 1M, 40% cache hit rate at 50% discount → $123.00/month ($1,476.00/year)
Worked example: a customer-support bot handling 50,000 requests/month, averaging 800 input tokens and 300 output tokens per request, on a model priced at $0.15 per 1M input tokens and $0.60 per 1M output tokens. Cost per request = (800/1,000,000 × $0.15) + (300/1,000,000 × $0.60) = $0.00012 + $0.00018 = $0.0003. At 50,000 requests: $15.00/month, or $180.00/year.