LLM API Pricing Comparison
This table lists 44 models from three providers. Prices are USD per million tokens, taken from each provider's pricing page. Use the cost calculator to turn them into a monthly bill.
| Model | Provider | Input | Cached input | Output |
|---|---|---|---|---|
| Claude Fable 5.1 | Anthropic (Claude) | $10.00 | $0.250 | $50.00 |
| Claude Fable 5 | Anthropic (Claude) | $10.00 | $1.00 | $50.00 |
| Claude Opus 5.5 | Anthropic (Claude) | $4.00 | $0.200 | $20.00 |
| Claude Opus 5 | Anthropic (Claude) | $5.00 | $0.500 | $25.00 |
| Claude Opus 4.8 | Anthropic (Claude) | $5.00 | $0.500 | $25.00 |
| Claude Opus 4.7 | Anthropic (Claude) | $5.00 | $0.500 | $25.00 |
| Claude Opus 4.6 | Anthropic (Claude) | $5.00 | $0.500 | $25.00 |
| Claude Opus 4.5 | Anthropic (Claude) | $5.00 | $0.500 | $25.00 |
| Claude Sonnet 5.5 | Anthropic (Claude) | $2.00 | $0.200 | $10.00 |
| Claude Sonnet 5 | Anthropic (Claude) | $2.00 | $0.200 | $10.00 |
| Claude Sonnet 4.6 | Anthropic (Claude) | $3.00 | $0.300 | $15.00 |
| Claude Sonnet 4.5 | Anthropic (Claude) | $3.00 | $0.300 | $15.00 |
| Claude Haiku 4.5 | Anthropic (Claude) | $1.00 | $0.100 | $5.00 |
| GPT-6 Astra | OpenAI | $10.00 | $1.00 | $50.00 |
| GPT-6.1 Sol | OpenAI | $2.00 | $0.100 | $10.00 |
| GPT-6 Sol | OpenAI | $2.00 | $0.200 | $10.00 |
| GPT-6 Luna | OpenAI | $0.100 | $0.010 | $0.500 |
| GPT-5.6 Sol | OpenAI | $4.00 | $0.400 | $20.00 |
| GPT-5.6 Terra | OpenAI | $2.00 | $0.200 | $12.00 |
| GPT-5.6 Luna | OpenAI | $0.200 | $0.020 | $1.20 |
| GPT-5.5 | OpenAI | $5.00 | $0.500 | $30.00 |
| GPT-5.5 Pro | OpenAI | $30.00 | — | $180 |
| GPT-5.4 | OpenAI | $2.50 | $0.250 | $15.00 |
| GPT-5.4 mini | OpenAI | $0.750 | $0.075 | $4.50 |
| GPT-5.4 nano | OpenAI | $0.200 | $0.020 | $1.25 |
| GPT-5.2 | OpenAI | $1.75 | $0.175 | $14.00 |
| GPT-5.1 | OpenAI | $1.25 | $0.125 | $10.00 |
| GPT-5 mini | OpenAI | $0.250 | $0.025 | $2.00 |
| GPT-5 nano | OpenAI | $0.050 | $0.0050 | $0.400 |
| GPT-4.1 | OpenAI | $2.00 | $0.500 | $8.00 |
| GPT-4.1 mini | OpenAI | $0.400 | $0.100 | $1.60 |
| GPT-4.1 nano | OpenAI | $0.100 | $0.025 | $0.400 |
| GPT-4o | OpenAI | $2.50 | $1.25 | $10.00 |
| GPT-4o mini | OpenAI | $0.150 | $0.075 | $0.600 |
| o3 | OpenAI | $2.00 | $0.500 | $8.00 |
| o4-mini | OpenAI | $1.10 | $0.275 | $4.40 |
| Gemini 3.8 Flash | Google (Gemini) | $0.750 → $1.50 from 2027-01-01 | $0.075 → $0.150 from 2027-01-01 | $3.75 → $7.50 from 2027-01-01 |
| Gemini 3.7 Flash | Google (Gemini) | $0.750 → $1.50 from 2027-01-01 | $0.075 → $0.150 from 2027-01-01 | $3.75 → $7.50 from 2027-01-01 |
| Gemini 3.6 Flash | Google (Gemini) | $0.750 → $1.50 from 2027-01-01 | $0.075 → $0.150 from 2027-01-01 | $3.75 → $7.50 from 2027-01-01 |
| Gemini 3.5 Flash | Google (Gemini) | $1.50 | $0.150 | $9.00 |
| Gemini 3.1 Pro (Preview) | Google (Gemini) | $2.00 ($4.00 above 200k) | $0.200 ($0.400 above 200k) | $12.00 ($18.00 above 200k) |
| Gemini 2.5 Pro | Google (Gemini) | $1.25 ($2.50 above 200k) | $0.125 ($0.250 above 200k) | $10.00 ($15.00 above 200k) |
| Gemini 2.5 Flash Input price shown is for text, image and video; audio input costs more. | Google (Gemini) | $0.300 | $0.030 | $2.50 |
| Gemini 2.5 Flash-Lite Input price shown is for text, image and video; audio input costs more. | Google (Gemini) | $0.100 | $0.010 | $0.400 |
Official sources
- Anthropic (Claude) pricing page. Batch API: 50% discount on input and output. Prompt caching: reads cost 0.1x the input price (0.025x on Fable 5.1, 0.05x on Opus 5.5).
- OpenAI pricing page. Cached input price is listed per model.
- Google (Gemini) pricing page. Some models change price on 2027-01-01. Pro models cost more above 200k prompt tokens.
How to read the table
- Input vs output: output tokens cost several times more than input tokens, so long answers cost more than long prompts.
- Cached input: repeated prompt prefixes can be read from cache at a fraction of the input price when the provider supports it.
- Tokens differ between providers: each provider counts tokens with its own tokenizer, so the same text is not exactly the same number of tokens everywhere. Anthropic says its newer tokenizer produces about 30% more tokens for the same text.
See the ranking in cheapest LLM APIs, or the provider pages for Claude, OpenAI and Gemini.