Skip to content

AI API prices per million tokens

Input and output token prices of OpenAI, Anthropic, Google, xAI, Mistral and DeepSeek models, converted at the ČNB rate, with a calculator: how much X messages a month cost, and when a subscription is cheaper.

Prices from official price lists, converted at the Czech National Bank rate of 02/10/2026. We check the price lists every week; a person confirms every change.

How much do I pay for the API?

Enter how many messages you send a month. A typical message is 1 500 tokens in (your question with context) and 600 tokens out (about 1 100 and 450 words). The table converts at the ČNB rate and compares with each provider's cheapest paid subscription.

ModelA monthA messageSubscription or API?
GPT-6 Luna APIOpenAI€0.40€0.0004API is cheaperChatGPT Go at €7.13 pays off from 17 785 messages a month
Mistral Small 4 APIMistral€0.52€0.00052No chat subscription from this provider
DeepSeek V4.1 Flash APIDeepSeek€1.04€0.001No chat subscription from this provider
Mistral Large 3 APIMistral€1.47€0.0015No chat subscription from this provider
Gemini 3.5 Flash-Lite APIGoogle€1.74€0.0017No chat subscription from this provider
Grok Build 0.1 APIxAI€2.41€0.0024API is cheaperGrok SuperGrok at €26.73 pays off from 11 112 messages a month
Gemini 3.8 Flash APIGoogle€3.01€0.003No chat subscription from this provider
Grok 4.3 APIxAI€3.01€0.003API is cheaperGrok SuperGrok at €26.73 pays off from 8 890 messages a month
DeepSeek V4 Pro APIDeepSeek€3.88€0.0039No chat subscription from this provider
Claude Haiku 4.5 APIAnthropic€4.01€0.004API is cheaperClaude Pro at €18.00 pays off from 4 490 messages a month
Grok 4.7 APIxAI€5.88€0.0059API is cheaperGrok SuperGrok at €26.73 pays off from 4 546 messages a month
Mistral Medium 3.5 APIMistral€6.01€0.006No chat subscription from this provider
Claude Sonnet 5.5 APIAnthropic€8.02€0.008API is cheaperClaude Pro at €18.00 pays off from 2 245 messages a month
GPT-6.1 Sol APIOpenAI€8.02€0.008Subscription is cheaperChatGPT Go at €7.13 pays off from 889 messages a month
Gemini 3.1 Pro (Preview) APIGoogle€9.09€0.0091No chat subscription from this provider
Claude Opus 5.5 APIAnthropic€16.04€0.016API is cheaperClaude Pro at €18.00 pays off from 1 122 messages a month
Claude Fable 5.1 APIAnthropic€40.09€0.04Subscription is cheaperClaude Pro at €18.00 pays off from 449 messages a month
GPT-6 Astra APIOpenAI€40.09€0.04Subscription is cheaperChatGPT Go at €7.13 pays off from 177 messages a month
  • 1. OpenAI

    GPT-6 Luna API

    Input, 1M tokens
    €0.09
    US$0.10
    Output, 1M tokens
    €0.45
    US$0.50
    Context: 1 050 000
    • Cached input $0.01 / 1M tokens
    • Cheapest, most efficient tier

    Same >272K long-context caveat. Previous gpt-5.4-mini ($0.75/$4.50) and gpt-5.4-nano ($0.20/$1.25) are still listed.

    Source: developers.openai.com, checked 05/10/2026

  • 2. Mistral

    Mistral Small 4 API

    Input, 1M tokens
    €0.13
    US$0.15
    Output, 1M tokens
    €0.53
    US$0.60
    • Cached input −90% (generic discount)

    Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 3. DeepSeek

    DeepSeek V4.1 Flash API

    Input, 1M tokens
    €0.27
    US$0.30
    Output, 1M tokens
    €1.07
    US$1.20
    Context: 1 000 000
    • Cached input $0.006 / 1M tokens (peak)
    • Off-peak half price: input $0.15, output $0.60
    • Max output 384K tokens

    Peak price. Off-peak: input $0.15, cached $0.003, output $0.60.

    Source: api-docs.deepseek.com, checked 05/10/2026

  • 4. Mistral

    Mistral Large 3 API

    Input, 1M tokens
    €0.45
    US$0.50
    Output, 1M tokens
    €1.34
    US$1.50
    • Open-weight model
    • Cached input −90% (no per-model price shown)
    • Batch at half price

    Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 5. Google

    Gemini 3.5 Flash-Lite API

    Input, 1M tokens
    €0.27
    US$0.30
    Output, 1M tokens
    €2.23
    US$2.50
    • Cached input $0.03 / 1M tokens
    • Stable release

    Context window not read on an official page.

    Source: ai.google.dev, checked 05/10/2026

  • 6. xAI

    Grok Build 0.1 API

    Input, 1M tokens
    €0.89
    US$1
    Output, 1M tokens
    €1.78
    US$2
    Context: 256 000
    • Cached input $0.20 / 1M tokens
    • Prompts from 200k: input $2 / output $4

    Cheapest xAI model listed; positioning (coding?) not confirmed. From 200k tokens cached input is $0.40.

    Source: docs.x.ai, checked 05/10/2026

  • 7. Google

    Gemini 3.8 Flash API

    Input, 1M tokens
    €0.67
    US$0.75
    Output, 1M tokens
    €3.34
    US$3.75
    Context: 1 048 576
    • Cached input $0.075 / 1M tokens
    • Max output 65,536 tokens

    Promotional price until 31 Dec 2026; from 1 Jan 2027 input $1.50, output $7.50, caching $0.15.

    Source: ai.google.dev, checked 05/10/2026

  • 8. xAI

    Grok 4.3 API

    Input, 1M tokens
    €1.11
    US$1.25
    Output, 1M tokens
    €2.23
    US$2.50
    Context: 1 000 000
    • Cached input $0.20 / 1M tokens
    • Prompts from 200k: input $2.50 / output $5

    From 200k tokens cached input is $0.40.

    Source: docs.x.ai, checked 05/10/2026

  • 9. DeepSeek

    DeepSeek V4 Pro API

    Input, 1M tokens
    €1.18
    US$1.32
    Output, 1M tokens
    €3.53
    US$3.96
    Context: 1 000 000
    • Cached input $0.044 / 1M tokens (peak)
    • Off-peak half price: input $0.66, output $1.98
    • Max output 384K tokens

    Peak price (01–04 and 06–10 UTC Mon–Fri). Off-peak (other hours, weekends, Chinese holidays) is half: input $0.66, cached $0.022, output $1.98.

    Source: api-docs.deepseek.com, checked 05/10/2026

  • 10. Anthropic

    Claude Haiku 4.5 API

    Input, 1M tokens
    €0.89
    US$1
    Output, 1M tokens
    €4.45
    US$5
    Context: 200 000
    • Cached input $0.10 / 1M tokens
    • 200k token context

    API id claude-haiku-4-5-20251001. Retirement not sooner than 15 Oct 2026 (watch for a successor).

    Source: platform.claude.com, checked 05/10/2026

  • 11. xAI

    Grok 4.7 API

    Input, 1M tokens
    €1.78
    US$2
    Output, 1M tokens
    €5.35
    US$6
    Context: 500 000
    • Cached input $0.50 / 1M tokens
    • Prompts from 200k: input $4 / output $12

    From 200k tokens the higher rate (cached $1.00) applies to all tokens of the request.

    Source: docs.x.ai, checked 05/10/2026

  • 12. Mistral

    Mistral Medium 3.5 API

    Input, 1M tokens
    €1.34
    US$1.50
    Output, 1M tokens
    €6.68
    US$7.50
    • Frontier-class, agentic and coding
    • Cached input −90% (generic discount)

    More expensive than Mistral Large 3. Context window not shown on the pricing page.

    Source: mistral.ai, checked 05/10/2026

  • 13. Anthropic

    Claude Sonnet 5.5 API

    Input, 1M tokens
    €1.78
    US$2
    Output, 1M tokens
    €8.91
    US$10
    Context: 1 000 000
    • Cached input $0.20 / 1M tokens
    • No surcharge for long context
    • Batch −50%

    Cache write 5 min $2.50, 1 h $4.

    Source: platform.claude.com, checked 05/10/2026

  • 14. OpenAI

    GPT-6.1 Sol API

    Input, 1M tokens
    €1.78
    US$2
    Output, 1M tokens
    €8.91
    US$10
    Context: 1 050 000
    • Cached input $0.10 / 1M tokens
    • Near-Astra performance at lower cost

    Same long-context caveat: input over 272K may be billed at double rates; not confirmed for this model.

    Source: developers.openai.com, checked 05/10/2026

  • 15. Google

    Gemini 3.1 Pro (Preview) API

    Input, 1M tokens
    €1.78
    US$2
    Output, 1M tokens
    €10.69
    US$12
    • Cached input $0.20 / 1M tokens
    • Prompts over 200k: input $4 / output $18

    Preview status. Caching over 200k is $0.40. Context window not read on an official page.

    Source: ai.google.dev, checked 05/10/2026

  • 16. Anthropic

    Claude Opus 5.5 API

    Input, 1M tokens
    €3.56
    US$4
    Output, 1M tokens
    €17.82
    US$20
    Context: 1 000 000
    • Cached input $0.20 / 1M tokens
    • No surcharge for long context
    • Fast mode $8 / $40; Batch −50%

    Recommended default. Cache write 5 min $5, 1 h $8.

    Source: platform.claude.com, checked 05/10/2026

  • 17. Anthropic

    Claude Fable 5.1 API

    Input, 1M tokens
    €8.91
    US$10
    Output, 1M tokens
    €44.54
    US$50
    Context: 1 000 000
    • Cached input $0.25 / 1M tokens
    • Full 1M context at standard price
    • Batch −50%; cache write 5 min $12.50, 1 h $20

    Top tier above Opus. US-only inference (inference_geo us) costs 1.1x.

    Source: platform.claude.com, checked 05/10/2026

  • 18. OpenAI

    GPT-6 Astra API

    Input, 1M tokens
    €8.91
    US$10
    Output, 1M tokens
    €44.54
    US$50
    Context: 1 050 000
    • Cached input $1.00 / 1M tokens
    • Flagship, most capable model

    Pricing page: input over 272K tokens is billed at double input/cache rates for models with long-context pricing; not confirmed for GPT-6 models.

    Source: developers.openai.com, checked 05/10/2026

Some of these services may pay us a commission when you subscribe through our link. It never changes the order: cheapest first, nothing else.