StackProofStackProof

Model pricing · Mistral · value tier · August 2026

Mistral Small API pricing, in numbers that matter

List price, real cost per task, rank against the field, and the self-hosting break-even, all computed from our open dataset.

Input, $/1M tokens

$0.15

Output, $/1M tokens

$0.6

Cost per task

$0.00158

Cost rank

#1 of 10

The task profile behind every figure: 1,500 input and 500 output tokens per call, 3 calls per task. On that profile Mistral Small costs $0.00158 per task and $1,575 per million tasks. Output is billed at 4x the input rate, so despite the profile sending three times more input than output, output makes up 57% of the bill. Capping response length is the highest-return prompt-level optimisation on this model.

Mistral Small against the field

#Model$/taskvs Mistral SmallHead-to-head
1Mistral Small$0.00158--
2DeepSeek-V4-Flash$0.003962.51xcompare
3Gemini 2.5 Flash$0.005103.24xcompare
4Grok 4.3$0.009385.95xcompare
5GPT-5.4 mini$0.010136.43xcompare
6Claude Haiku 4.5$0.012007.62xcompare
7Claude Sonnet 5$0.0240015.2xcompare
8Gemini 3.1 Pro$0.0270017.1xcompare
9Claude Opus 4.8$0.0600038.1xcompare
10GPT-5.5$0.0675042.9xcompare

Self-hosting break-even

Against a reference $2/hour dedicated instance ($1,460/month fixed, 788,400 tasks/month capacity at full utilisation), the break-even against Mistral Small is 926,984 tasks per month. That exceeds what the box can physically serve, so against this model that GPU never pays off at any volume. Work through your own numbers in the $/task calculator or the break-even guide.

Frequently asked

How much does the Mistral Small API cost?
Mistral Small lists at $0.15 per million input tokens and $0.6 per million output tokens as of August 2026. On a realistic task of 1,500 input and 500 output tokens across 3 calls, that is $0.00158 per task, or $1,575 per million tasks.
Is Mistral Small expensive compared to other models?
Mistral Small is the cheapest of the 10 models we track on cost per task. The most expensive tracked model, GPT-5.5, costs 42.9x more for the same work.
When does self-hosting beat the Mistral Small API?
Effectively never on commodity hardware: against a $2/hour GPU the break-even is 926,984 tasks per month, above the 788,400 tasks that box can physically serve, so the crossover is unreachable.

See also: all tracked models · every head-to-head comparison · download the dataset