MModelspectraIndependent AI Model Intelligence
Model profile · Updated 2026-09-08
Model Profile · Google · US

Gemini 3.5 Flash

Best forFast tasks, Multimodal, Low cost
LicensingClosed API
Updated2026-09-08
01

Snapshot

Overall
69
rank #14 of 22
Coding
64
rank #17 of 22
Multimodal
86
rank #7 of 22
Input / Output
$1.5 / $9
USD per 1M tokens
Context
1M
usable recall 88
Speed
~80 tok/s
TTFT 0.3s

Aggregated from public sources and independently weighted; methodology on the Terms page. Scores within 3 points are treated as statistically tied.

02

Specs & pricing

Every tracked field. List API prices in USD per 1M tokens; verify current vendor pricing before purchase.

VendorGoogle (US)
Released / current2026.07
LicensingClosed / managed API
Context window1M
Max output128K
Effective-context score88
Input / output price$1.5 / $9 per 1M tokens
Cache discountNo published cache discount
Free tierGemini App free; API free tier
Speed / TTFT~80 tok/s / 0.3s
Function calling75
Refusal rate~10%
English / Chinese85 / 72
Modalitiestext, image, audio, video
Fine-tuningYes
Private deploymentNo
SOC2 / no-trainyes / yes
03

Strengths & watch-outs

Strengths

  • Extremely fast
  • Low price
  • Native multimodal
  • Built for scale

Watch-outs

  • Average reasoning depth
  • Limited on hard tasks
04

Where it ranks among 22 models

Overall 69/100 (#14), coding 64 (#17), multimodal 86 (#7). See the full boards and side-by-side compare on the leaderboard.
05

Pricing in practice

Illustrative monthly bill for 100M input + 30M output tokens: $420 at list. Your mix and cache-hit ratio change this.

Illustrative model using public list rates before any enterprise agreement; cache discount: No published cache discount.

06

Frequently asked questions

How much does Gemini 3.5 Flash cost?
List API pricing is $1.5 input and $9 output per 1M tokens. No published cache discount. Gemini App free; API free tier
What is Gemini 3.5 Flash's context window, and is it usable end to end?
It advertises a 1M window with 128K max output; its effective-context score is 88/100, which is the better predictor of whether long-document details are actually retained.
Is Gemini 3.5 Flash good for coding?
Its coding aggregate is 64/100, rank #17 of 22. It is adequate for scripts and assisted completion rather than the most demanding SWE-style work.
Can I self-host Gemini 3.5 Flash?
No. It is a closed managed API; there is no weights download.
How strong is it in Chinese and on multimodal inputs?
Chinese score is 72 versus English 85; multimodal aggregate is 86/100 (rank #7), accepting text, image, audio, video.