Skip to content

AI capacity and pricing

When a model runs, it consumes tokens. How many depends on how much material went into the request and how long the response is — this can’t be predicted exactly in advance. Roughly 95% of total token consumption is input tokens, with the remaining 5% output.

Different models price input and output tokens differently, so the same task can cost different amounts depending on which model handles it.

OptionDescriptionMonthly cost
Base levelBasic level, corresponding to roughly 8,000 queries€92
ExtendedAdditional capacity in increments of €92€184

Prices are exclusive of VAT and stated in Euro.

Model typeCreatorModelEUR / input MEUR / output M
Closed sourceOpenAIGPT-5.21.612.9
Closed sourceOpenAIGPT-51.29.2
Closed sourceOpenAIGPT-5 mini0.21.9
Closed sourceOpenAIGPT-427.755.4
Closed sourceOpenAIGPT-4o2.39.2
Closed sourceOpenAIGPT-4o mini0.10.6
Closed sourceOpenAIo3-mini1.04.1
Closed sourceOpenAIGPT-4.11.97.4
Closed sourceOpenAIGPT-4.1 mini0.41.5
Closed sourceOpenAIGPT-4.1 nano0.10.4
Open sourceOpenAIGPT-OSS-120B0.30.9
Closed sourceMistralMistral Large1.95.5
Closed sourceAnthropicClaude 4.5 Haiku0.94.6
Closed sourceAnthropicClaude Opus 413.869.2
Closed sourceAnthropicClaude 4 Sonnet2.813.8
Closed sourceAnthropicClaude 4.5 Sonnet2.813.8
Closed sourceAnthropicClaude Opus 4.113.869.2
Open sourceMetaLlama 3.3 70B0.80.8
Closed sourceOpenAIGPT-5.3 Instant1.612.9
Closed sourceOpenAIGPT-5.42.313.8
Closed sourceOpenAIGPT-5.54.627.7
Closed sourceOpenAIGPT-5.1 (Azure)1.29.2
Closed sourceOpenAIGPT-5 mini (Azure)0.21.9
Closed sourceOpenAIGPT-5 nano (Azure)0.10.4
Closed sourceOpenAIGPT-4.1 mini (Azure)0.41.5
Closed sourceOpenAIGPT-4.1 nano (Azure)0.10.4
Closed sourceAnthropicClaude 4.6 Sonnet2.813.8
Closed sourceAnthropicClaude 4.6 Opus4.623.1
Closed sourceAnthropicClaude 4.7 Opus4.623.1
Closed sourceMistralMistral 4 Small0.10.6
Closed sourceMistralMistral 3.5 Medium1.46.9
Closed sourceGoogleGemini 2.5 Flash0.32.3
Closed sourceGoogleGemini 2.5 Pro1.29.2
Closed sourceGoogleGemini 3.1 Pro Preview1.811.1
Closed sourceGoogleGemini 3.1 Flash-Lite Preview0.21.4
Open sourceGoogleSwedish Hosted - Managed by Intric0.00.0
Closed sourceOpenAIWhisper0.3
Open sourceNational Library of SwedenKB - Whisper *0.1
Closed sourceOpenAIGPT - Image-1Approx. €0.3 per image

For transcription, the price per query corresponds to 60 minutes of recording.

Intric doesn’t automatically know the price of a custom or self-hosted model, so an administrator sets its token cost directly on the model’s configuration. See Models for where that’s configured, and keep the value current — an outdated or missing cost leads to under- or over-reported figures for that model.

Choosing a more capable, more expensive model for a simple task costs more without necessarily improving the result. Reserve the highest-cost models for tasks that genuinely need their reasoning ability, and use an economical or fast option for everyday work — see Completion model for how that choice is made per assistant.