AI API prices per million tokens
Input and output token prices of OpenAI, Anthropic, Google, xAI, Mistral and DeepSeek models, converted at the ČNB rate, with a calculator: how much X messages a month cost, and when a subscription is cheaper.
Prices from official price lists, converted at the Czech National Bank rate of 02/10/2026. We check the price lists every week; a person confirms every change.
How much do I pay for the API?
Enter how many messages you send a month. A typical message is 1 500 tokens in (your question with context) and 600 tokens out (about 1 100 and 450 words). The table converts at the ČNB rate and compares with each provider's cheapest paid subscription.
| Model | A month | A message | Subscription or API? |
|---|---|---|---|
| GPT-6 Luna APIOpenAI | € 0,40 | € 0,0004 | API is cheaperChatGPT Go at € 7,13 pays off from 17.785 messages a month |
| Mistral Small 4 APIMistral | € 0,52 | € 0,00052 | No chat subscription from this provider |
| DeepSeek V4.1 Flash APIDeepSeek | € 1,04 | € 0,001 | No chat subscription from this provider |
| Mistral Large 3 APIMistral | € 1,47 | € 0,0015 | No chat subscription from this provider |
| Gemini 3.5 Flash-Lite APIGoogle | € 1,74 | € 0,0017 | No chat subscription from this provider |
| Grok Build 0.1 APIxAI | € 2,41 | € 0,0024 | API is cheaperGrok SuperGrok at € 26,73 pays off from 11.112 messages a month |
| Gemini 3.8 Flash APIGoogle | € 3,01 | € 0,003 | No chat subscription from this provider |
| Grok 4.3 APIxAI | € 3,01 | € 0,003 | API is cheaperGrok SuperGrok at € 26,73 pays off from 8.890 messages a month |
| DeepSeek V4 Pro APIDeepSeek | € 3,88 | € 0,0039 | No chat subscription from this provider |
| Claude Haiku 4.5 APIAnthropic | € 4,01 | € 0,004 | API is cheaperClaude Pro at € 18,00 pays off from 4.490 messages a month |
| Grok 4.7 APIxAI | € 5,88 | € 0,0059 | API is cheaperGrok SuperGrok at € 26,73 pays off from 4.546 messages a month |
| Mistral Medium 3.5 APIMistral | € 6,01 | € 0,006 | No chat subscription from this provider |
| Claude Sonnet 5.5 APIAnthropic | € 8,02 | € 0,008 | API is cheaperClaude Pro at € 18,00 pays off from 2.245 messages a month |
| GPT-6.1 Sol APIOpenAI | € 8,02 | € 0,008 | Subscription is cheaperChatGPT Go at € 7,13 pays off from 889 messages a month |
| Gemini 3.1 Pro (Preview) APIGoogle | € 9,09 | € 0,0091 | No chat subscription from this provider |
| Claude Opus 5.5 APIAnthropic | € 16,04 | € 0,016 | API is cheaperClaude Pro at € 18,00 pays off from 1.122 messages a month |
| Claude Fable 5.1 APIAnthropic | € 40,09 | € 0,04 | Subscription is cheaperClaude Pro at € 18,00 pays off from 449 messages a month |
| GPT-6 Astra APIOpenAI | € 40,09 | € 0,04 | Subscription is cheaperChatGPT Go at € 7,13 pays off from 177 messages a month |
1. OpenAI
GPT-6 Luna API
- Input, 1M tokens
- € 0,09
- US$ 0,10
- Output, 1M tokens
- € 0,45
- US$ 0,50
Context: 1.050.000- Cached input $0.01 / 1M tokens
- Cheapest, most efficient tier
Same >272K long-context caveat. Previous gpt-5.4-mini ($0.75/$4.50) and gpt-5.4-nano ($0.20/$1.25) are still listed.
2. Mistral
Mistral Small 4 API
- Input, 1M tokens
- € 0,13
- US$ 0,15
- Output, 1M tokens
- € 0,53
- US$ 0,60
- Cached input −90% (generic discount)
Context window not shown on the pricing page.
3. DeepSeek
DeepSeek V4.1 Flash API
- Input, 1M tokens
- € 0,27
- US$ 0,30
- Output, 1M tokens
- € 1,07
- US$ 1,20
Context: 1.000.000- Cached input $0.006 / 1M tokens (peak)
- Off-peak half price: input $0.15, output $0.60
- Max output 384K tokens
Peak price. Off-peak: input $0.15, cached $0.003, output $0.60.
4. Mistral
Mistral Large 3 API
- Input, 1M tokens
- € 0,45
- US$ 0,50
- Output, 1M tokens
- € 1,34
- US$ 1,50
- Open-weight model
- Cached input −90% (no per-model price shown)
- Batch at half price
Context window not shown on the pricing page.
5. Google
Gemini 3.5 Flash-Lite API
- Input, 1M tokens
- € 0,27
- US$ 0,30
- Output, 1M tokens
- € 2,23
- US$ 2,50
- Cached input $0.03 / 1M tokens
- Stable release
Context window not read on an official page.
6. xAI
Grok Build 0.1 API
- Input, 1M tokens
- € 0,89
- US$ 1
- Output, 1M tokens
- € 1,78
- US$ 2
Context: 256.000- Cached input $0.20 / 1M tokens
- Prompts from 200k: input $2 / output $4
Cheapest xAI model listed; positioning (coding?) not confirmed. From 200k tokens cached input is $0.40.
7. Google
Gemini 3.8 Flash API
- Input, 1M tokens
- € 0,67
- US$ 0,75
- Output, 1M tokens
- € 3,34
- US$ 3,75
Context: 1.048.576- Cached input $0.075 / 1M tokens
- Max output 65,536 tokens
Promotional price until 31 Dec 2026; from 1 Jan 2027 input $1.50, output $7.50, caching $0.15.
8. xAI
Grok 4.3 API
- Input, 1M tokens
- € 1,11
- US$ 1,25
- Output, 1M tokens
- € 2,23
- US$ 2,50
Context: 1.000.000- Cached input $0.20 / 1M tokens
- Prompts from 200k: input $2.50 / output $5
From 200k tokens cached input is $0.40.
9. DeepSeek
DeepSeek V4 Pro API
- Input, 1M tokens
- € 1,18
- US$ 1,32
- Output, 1M tokens
- € 3,53
- US$ 3,96
Context: 1.000.000- Cached input $0.044 / 1M tokens (peak)
- Off-peak half price: input $0.66, output $1.98
- Max output 384K tokens
Peak price (01–04 and 06–10 UTC Mon–Fri). Off-peak (other hours, weekends, Chinese holidays) is half: input $0.66, cached $0.022, output $1.98.
10. Anthropic
Claude Haiku 4.5 API
- Input, 1M tokens
- € 0,89
- US$ 1
- Output, 1M tokens
- € 4,45
- US$ 5
Context: 200.000- Cached input $0.10 / 1M tokens
- 200k token context
API id claude-haiku-4-5-20251001. Retirement not sooner than 15 Oct 2026 (watch for a successor).
11. xAI
Grok 4.7 API
- Input, 1M tokens
- € 1,78
- US$ 2
- Output, 1M tokens
- € 5,35
- US$ 6
Context: 500.000- Cached input $0.50 / 1M tokens
- Prompts from 200k: input $4 / output $12
From 200k tokens the higher rate (cached $1.00) applies to all tokens of the request.
12. Mistral
Mistral Medium 3.5 API
- Input, 1M tokens
- € 1,34
- US$ 1,50
- Output, 1M tokens
- € 6,68
- US$ 7,50
- Frontier-class, agentic and coding
- Cached input −90% (generic discount)
More expensive than Mistral Large 3. Context window not shown on the pricing page.
13. Anthropic
Claude Sonnet 5.5 API
- Input, 1M tokens
- € 1,78
- US$ 2
- Output, 1M tokens
- € 8,91
- US$ 10
Context: 1.000.000- Cached input $0.20 / 1M tokens
- No surcharge for long context
- Batch −50%
Cache write 5 min $2.50, 1 h $4.
14. OpenAI
GPT-6.1 Sol API
- Input, 1M tokens
- € 1,78
- US$ 2
- Output, 1M tokens
- € 8,91
- US$ 10
Context: 1.050.000- Cached input $0.10 / 1M tokens
- Near-Astra performance at lower cost
Same long-context caveat: input over 272K may be billed at double rates; not confirmed for this model.
15. Google
Gemini 3.1 Pro (Preview) API
- Input, 1M tokens
- € 1,78
- US$ 2
- Output, 1M tokens
- € 10,69
- US$ 12
- Cached input $0.20 / 1M tokens
- Prompts over 200k: input $4 / output $18
Preview status. Caching over 200k is $0.40. Context window not read on an official page.
16. Anthropic
Claude Opus 5.5 API
- Input, 1M tokens
- € 3,56
- US$ 4
- Output, 1M tokens
- € 17,82
- US$ 20
Context: 1.000.000- Cached input $0.20 / 1M tokens
- No surcharge for long context
- Fast mode $8 / $40; Batch −50%
Recommended default. Cache write 5 min $5, 1 h $8.
17. Anthropic
Claude Fable 5.1 API
- Input, 1M tokens
- € 8,91
- US$ 10
- Output, 1M tokens
- € 44,54
- US$ 50
Context: 1.000.000- Cached input $0.25 / 1M tokens
- Full 1M context at standard price
- Batch −50%; cache write 5 min $12.50, 1 h $20
Top tier above Opus. US-only inference (inference_geo us) costs 1.1x.
18. OpenAI
GPT-6 Astra API
- Input, 1M tokens
- € 8,91
- US$ 10
- Output, 1M tokens
- € 44,54
- US$ 50
Context: 1.050.000- Cached input $1.00 / 1M tokens
- Flagship, most capable model
Pricing page: input over 272K tokens is billed at double input/cache rates for models with long-context pricing; not confirmed for GPT-6 models.
Some of these services may pay us a commission when you subscribe through our link. It never changes the order: cheapest first, nothing else.