Skip to content

AI API prices per million tokens

Input and output token prices of OpenAI, Anthropic, Google, xAI, Mistral and DeepSeek models, converted at the ČNB rate, with a calculator: how much X messages a month cost, and when a subscription is cheaper.

Prices from official price lists, converted at the Czech National Bank rate of 02/10/2026. We check the price lists every week; a person confirms every change.

How much do I pay for the API?

Enter how many messages you send a month. A typical message is 1 500 tokens in (your question with context) and 600 tokens out (about 1 100 and 450 words). The table converts at the ČNB rate and compares with each provider's cheapest paid subscription.

ModelA monthA messageSubscription or API?
GPT-6 Luna APIOpenAI€ 0,40€ 0,0004API is cheaperChatGPT Go at € 7,13 pays off from 17.785 messages a month
Mistral Small 4 APIMistral€ 0,52€ 0,00052No chat subscription from this provider
DeepSeek V4.1 Flash APIDeepSeek€ 1,04€ 0,001No chat subscription from this provider
Mistral Large 3 APIMistral€ 1,47€ 0,0015No chat subscription from this provider
Gemini 3.5 Flash-Lite APIGoogle€ 1,74€ 0,0017No chat subscription from this provider
Grok Build 0.1 APIxAI€ 2,41€ 0,0024API is cheaperGrok SuperGrok at € 26,73 pays off from 11.112 messages a month
Gemini 3.8 Flash APIGoogle€ 3,01€ 0,003No chat subscription from this provider
Grok 4.3 APIxAI€ 3,01€ 0,003API is cheaperGrok SuperGrok at € 26,73 pays off from 8.890 messages a month
DeepSeek V4 Pro APIDeepSeek€ 3,88€ 0,0039No chat subscription from this provider
Claude Haiku 4.5 APIAnthropic€ 4,01€ 0,004API is cheaperClaude Pro at € 18,00 pays off from 4.490 messages a month
Grok 4.7 APIxAI€ 5,88€ 0,0059API is cheaperGrok SuperGrok at € 26,73 pays off from 4.546 messages a month
Mistral Medium 3.5 APIMistral€ 6,01€ 0,006No chat subscription from this provider
Claude Sonnet 5.5 APIAnthropic€ 8,02€ 0,008API is cheaperClaude Pro at € 18,00 pays off from 2.245 messages a month
GPT-6.1 Sol APIOpenAI€ 8,02€ 0,008Subscription is cheaperChatGPT Go at € 7,13 pays off from 889 messages a month
Gemini 3.1 Pro (Preview) APIGoogle€ 9,09€ 0,0091No chat subscription from this provider
Claude Opus 5.5 APIAnthropic€ 16,04€ 0,016API is cheaperClaude Pro at € 18,00 pays off from 1.122 messages a month
Claude Fable 5.1 APIAnthropic€ 40,09€ 0,04Subscription is cheaperClaude Pro at € 18,00 pays off from 449 messages a month
GPT-6 Astra APIOpenAI€ 40,09€ 0,04Subscription is cheaperChatGPT Go at € 7,13 pays off from 177 messages a month
  • 1. OpenAI

    GPT-6 Luna API

    Input, 1M tokens
    € 0,09
    US$ 0,10
    Output, 1M tokens
    € 0,45
    US$ 0,50
    Context: 1.050.000
    • Cached input $0.01 / 1M tokens
    • Cheapest, most efficient tier

    Same >272K long-context caveat. Previous gpt-5.4-mini ($0.75/$4.50) and gpt-5.4-nano ($0.20/$1.25) are still listed.

    Source: developers.openai.com, checked 05/10/2026

  • 2. Mistral

    Mistral Small 4 API

    Input, 1M tokens
    € 0,13
    US$ 0,15
    Output, 1M tokens
    € 0,53
    US$ 0,60
    • Cached input −90% (generic discount)

    Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 3. DeepSeek

    DeepSeek V4.1 Flash API

    Input, 1M tokens
    € 0,27
    US$ 0,30
    Output, 1M tokens
    € 1,07
    US$ 1,20
    Context: 1.000.000
    • Cached input $0.006 / 1M tokens (peak)
    • Off-peak half price: input $0.15, output $0.60
    • Max output 384K tokens

    Peak price. Off-peak: input $0.15, cached $0.003, output $0.60.

    Source: api-docs.deepseek.com, checked 05/10/2026

  • 4. Mistral

    Mistral Large 3 API

    Input, 1M tokens
    € 0,45
    US$ 0,50
    Output, 1M tokens
    € 1,34
    US$ 1,50
    • Open-weight model
    • Cached input −90% (no per-model price shown)
    • Batch at half price

    Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 5. Google

    Gemini 3.5 Flash-Lite API

    Input, 1M tokens
    € 0,27
    US$ 0,30
    Output, 1M tokens
    € 2,23
    US$ 2,50
    • Cached input $0.03 / 1M tokens
    • Stable release

    Context window not read on an official page.

    Source: ai.google.dev, checked 05/10/2026

  • 6. xAI

    Grok Build 0.1 API

    Input, 1M tokens
    € 0,89
    US$ 1
    Output, 1M tokens
    € 1,78
    US$ 2
    Context: 256.000
    • Cached input $0.20 / 1M tokens
    • Prompts from 200k: input $2 / output $4

    Cheapest xAI model listed; positioning (coding?) not confirmed. From 200k tokens cached input is $0.40.

    Source: docs.x.ai, checked 05/10/2026

  • 7. Google

    Gemini 3.8 Flash API

    Input, 1M tokens
    € 0,67
    US$ 0,75
    Output, 1M tokens
    € 3,34
    US$ 3,75
    Context: 1.048.576
    • Cached input $0.075 / 1M tokens
    • Max output 65,536 tokens

    Promotional price until 31 Dec 2026; from 1 Jan 2027 input $1.50, output $7.50, caching $0.15.

    Source: ai.google.dev, checked 05/10/2026

  • 8. xAI

    Grok 4.3 API

    Input, 1M tokens
    € 1,11
    US$ 1,25
    Output, 1M tokens
    € 2,23
    US$ 2,50
    Context: 1.000.000
    • Cached input $0.20 / 1M tokens
    • Prompts from 200k: input $2.50 / output $5

    From 200k tokens cached input is $0.40.

    Source: docs.x.ai, checked 05/10/2026

  • 9. DeepSeek

    DeepSeek V4 Pro API

    Input, 1M tokens
    € 1,18
    US$ 1,32
    Output, 1M tokens
    € 3,53
    US$ 3,96
    Context: 1.000.000
    • Cached input $0.044 / 1M tokens (peak)
    • Off-peak half price: input $0.66, output $1.98
    • Max output 384K tokens

    Peak price (01–04 and 06–10 UTC Mon–Fri). Off-peak (other hours, weekends, Chinese holidays) is half: input $0.66, cached $0.022, output $1.98.

    Source: api-docs.deepseek.com, checked 05/10/2026

  • 10. Anthropic

    Claude Haiku 4.5 API

    Input, 1M tokens
    € 0,89
    US$ 1
    Output, 1M tokens
    € 4,45
    US$ 5
    Context: 200.000
    • Cached input $0.10 / 1M tokens
    • 200k token context

    API id claude-haiku-4-5-20251001. Retirement not sooner than 15 Oct 2026 (watch for a successor).

    Source: platform.claude.com, checked 05/10/2026

  • 11. xAI

    Grok 4.7 API

    Input, 1M tokens
    € 1,78
    US$ 2
    Output, 1M tokens
    € 5,35
    US$ 6
    Context: 500.000
    • Cached input $0.50 / 1M tokens
    • Prompts from 200k: input $4 / output $12

    From 200k tokens the higher rate (cached $1.00) applies to all tokens of the request.

    Source: docs.x.ai, checked 05/10/2026

  • 12. Mistral

    Mistral Medium 3.5 API

    Input, 1M tokens
    € 1,34
    US$ 1,50
    Output, 1M tokens
    € 6,68
    US$ 7,50
    • Frontier-class, agentic and coding
    • Cached input −90% (generic discount)

    More expensive than Mistral Large 3. Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 13. Anthropic

    Claude Sonnet 5.5 API

    Input, 1M tokens
    € 1,78
    US$ 2
    Output, 1M tokens
    € 8,91
    US$ 10
    Context: 1.000.000
    • Cached input $0.20 / 1M tokens
    • No surcharge for long context
    • Batch −50%

    Cache write 5 min $2.50, 1 h $4.

    Source: platform.claude.com, checked 05/10/2026

  • 14. OpenAI

    GPT-6.1 Sol API

    Input, 1M tokens
    € 1,78
    US$ 2
    Output, 1M tokens
    € 8,91
    US$ 10
    Context: 1.050.000
    • Cached input $0.10 / 1M tokens
    • Near-Astra performance at lower cost

    Same long-context caveat: input over 272K may be billed at double rates; not confirmed for this model.

    Source: developers.openai.com, checked 05/10/2026

  • 15. Google

    Gemini 3.1 Pro (Preview) API

    Input, 1M tokens
    € 1,78
    US$ 2
    Output, 1M tokens
    € 10,69
    US$ 12
    • Cached input $0.20 / 1M tokens
    • Prompts over 200k: input $4 / output $18

    Preview status. Caching over 200k is $0.40. Context window not read on an official page.

    Source: ai.google.dev, checked 05/10/2026

  • 16. Anthropic

    Claude Opus 5.5 API

    Input, 1M tokens
    € 3,56
    US$ 4
    Output, 1M tokens
    € 17,82
    US$ 20
    Context: 1.000.000
    • Cached input $0.20 / 1M tokens
    • No surcharge for long context
    • Fast mode $8 / $40; Batch −50%

    Recommended default. Cache write 5 min $5, 1 h $8.

    Source: platform.claude.com, checked 05/10/2026

  • 17. Anthropic

    Claude Fable 5.1 API

    Input, 1M tokens
    € 8,91
    US$ 10
    Output, 1M tokens
    € 44,54
    US$ 50
    Context: 1.000.000
    • Cached input $0.25 / 1M tokens
    • Full 1M context at standard price
    • Batch −50%; cache write 5 min $12.50, 1 h $20

    Top tier above Opus. US-only inference (inference_geo us) costs 1.1x.

    Source: platform.claude.com, checked 05/10/2026

  • 18. OpenAI

    GPT-6 Astra API

    Input, 1M tokens
    € 8,91
    US$ 10
    Output, 1M tokens
    € 44,54
    US$ 50
    Context: 1.050.000
    • Cached input $1.00 / 1M tokens
    • Flagship, most capable model

    Pricing page: input over 272K tokens is billed at double input/cache rates for models with long-context pricing; not confirmed for GPT-6 models.

    Source: developers.openai.com, checked 05/10/2026

Some of these services may pay us a commission when you subscribe through our link. It never changes the order: cheapest first, nothing else.