StackProofStackProof

Model pricing · Google · frontier tier · August 2026

Gemini 3.1 Pro API pricing, in numbers that matter

List price, real cost per task, rank against the field, and the self-hosting break-even, all computed from our open dataset.

Input, $/1M tokens

$2

Output, $/1M tokens

$12

Cost per task

$0.02700

Cost rank

#8 of 10

The task profile behind every figure: 1,500 input and 500 output tokens per call, 3 calls per task. On that profile Gemini 3.1 Pro costs $0.02700 per task and $27,000 per million tasks. Output is billed at 6x the input rate, so despite the profile sending three times more input than output, output makes up 67% of the bill. Capping response length is the highest-return prompt-level optimisation on this model.

Gemini 3.1 Pro against the field

#Model$/taskvs Gemini 3.1 ProHead-to-head
1Mistral Small$0.001580.06xcompare
2DeepSeek-V4-Flash$0.003960.15xcompare
3Gemini 2.5 Flash$0.005100.19xcompare
4Grok 4.3$0.009380.35xcompare
5GPT-5.4 mini$0.010130.38xcompare
6Claude Haiku 4.5$0.012000.44xcompare
7Claude Sonnet 5$0.024000.89xcompare
8Gemini 3.1 Pro$0.02700--
9Claude Opus 4.8$0.060002.22xcompare
10GPT-5.5$0.067502.50xcompare

Self-hosting break-even

Against a reference $2/hour dedicated instance ($1,460/month fixed, 788,400 tasks/month capacity at full utilisation), the break-even against Gemini 3.1 Pro is 54,074 tasks per month. That sits inside the box's capacity, so past that volume self-hosting genuinely wins on cost. Work through your own numbers in the $/task calculator or the break-even guide.

Frequently asked

How much does the Gemini 3.1 Pro API cost?
Gemini 3.1 Pro lists at $2 per million input tokens and $12 per million output tokens as of August 2026. On a realistic task of 1,500 input and 500 output tokens across 3 calls, that is $0.02700 per task, or $27,000 per million tasks.
Is Gemini 3.1 Pro expensive compared to other models?
Gemini 3.1 Pro ranks #8 of 10 tracked models on cost per task. It costs 17.1x more than the cheapest tracked model (Mistral Small), while the most expensive (GPT-5.5) costs 2.50x more than it.
When does self-hosting beat the Gemini 3.1 Pro API?
Against a $2/hour dedicated GPU ($1,460/month), self-hosting becomes cheaper than Gemini 3.1 Pro above 54,074 tasks per month, within that box's physical capacity of 788,400 tasks.

See also: all tracked models · every head-to-head comparison · download the dataset