LLM API Pricing Comparison
LLM API Pricing Comparison
Pricing comparison for major LLM APIs. Prices are per 1M tokens (unless noted).
Frontier Models
| Model | Provider | Input (per 1M) | Output (per 1M) | Context |
|---|---|---|---|---|
| GPT-4o | OpenAI | $2.50 | $10.00 | 128K |
| GPT-4o-mini | OpenAI | $0.15 | $0.60 | 128K |
| o1 | OpenAI | $15.00 | $60.00 | 128K |
| Claude 4 Opus | Anthropic | $15.00 | $75.00 | 200K |
| Claude 4 Sonnet | Anthropic | $3.00 | $15.00 | 200K |
| Claude 4 Haiku | Anthropic | $0.25 | $1.25 | 200K |
| Gemini 2 Pro | $2.50 | $10.00 | 1M | |
| gemini-2-flash | $0.10 | $0.40 | 1M |
Open Models via API Providers
| Model | Provider | Input (per 1M) | Output (per 1M) |
|---|---|---|---|
| DeepSeek V3 | DeepSeek | $0.27 | $1.10 |
| DeepSeek R1 | DeepSeek | $0.55 | $2.19 |
| Llama 4 Scout | Various | ~$0.10 | ~$0.40 |
| Mistral Large | Mistral | $2.00 | $6.00 |
| qwen-2.5-72B | Alibaba | $0.90 | $0.90 |
Cost Efficiency Leaders
- Cheapest frontier: Gemini 2 Flash ($0.10/$0.40 per 1M)
- Cheapest open: DeepSeek V3 ($0.27/$1.10 per 1M)
- Best value: GPT-4o-mini ($0.15/$0.60 per 1M)
- Most expensive: Claude 4 Opus ($15.00/$75.00 per 1M)
Batch Pricing
Most providers offer ~50% discount for batch/async processing:
- OpenAI: 50% off batch API
- Anthropic: 50% off message batches
- Google: Pay-as-you-go (no batch discount currently)
Related
- Open-Source vs Closed-Source LLMs — Open vs closed comparison
- context-window-comparison — Context window sizes
- deployment-strategies — Self-hosting cost analysis