Gemini 3.1 Pro
The Vertex AI enterprise tier offers a fuller feature set.
01
Snapshot
Overall
83
rank #3 of 22
Coding
74
rank #9 of 22
Multimodal
92
rank #2 of 22
Input / Output
$2 / $12
USD per 1M tokens
Context
2M
usable recall 92
Speed
~60 tok/s
TTFT 0.7s
Aggregated from public sources and independently weighted; methodology on the Terms page. Scores within 3 points are treated as statistically tied.
02
Specs & pricing
Every tracked field. List API prices in USD per 1M tokens; verify current vendor pricing before purchase.
| Vendor | Google (US) |
| Released / current | 2026.05 |
| Licensing | Closed / managed API |
| Context window | 2M |
| Max output | 128K |
| Effective-context score | 92 |
| Input / output price | $2 / $12 per 1M tokens |
| Cache discount | No published cache discount |
| Free tier | Gemini App free with limited quota; API free tier available |
| Speed / TTFT | ~60 tok/s / 0.7s |
| Function calling | 78 |
| Refusal rate | ~12% |
| English / Chinese | 90 / 75 |
| Modalities | text, image, audio, video |
| Fine-tuning | Yes |
| Private deployment | No |
| SOC2 / no-train | yes / yes |
03
Strengths & watch-outs
Strengths
- Native audio and video
- Strong scientific computing
- Largest 2M context
- Search integration
Watch-outs
- Weaker coding
- Average Chinese
- Unstable function calling
04
Where it ranks among 22 models
Overall 83/100 (#3), coding 74 (#9), multimodal 92 (#2). See the full boards and side-by-side compare on the leaderboard.
05
Pricing in practice
Illustrative monthly bill for 100M input + 30M output tokens: $560 at list. Your mix and cache-hit ratio change this.
Illustrative model using public list rates before any enterprise agreement; cache discount: No published cache discount.
06
Frequently asked questions
How much does Gemini 3.1 Pro cost?
List API pricing is $2 input and $12 output per 1M tokens. No published cache discount. Gemini App free with limited quota; API free tier available
What is Gemini 3.1 Pro's context window, and is it usable end to end?
It advertises a 2M window with 128K max output; its effective-context score is 92/100, which is the better predictor of whether long-document details are actually retained.
Is Gemini 3.1 Pro good for coding?
Its coding aggregate is 74/100, rank #9 of 22. It is a strong choice for most product engineering.
Can I self-host Gemini 3.1 Pro?
No. It is a closed managed API; there is no weights download.
How strong is it in Chinese and on multimodal inputs?
Chinese score is 75 versus English 90; multimodal aggregate is 92/100 (rank #2), accepting text, image, audio, video.