Gemini 3.7 Flash pricing
Google ·
gemini-gemini-3.7-flash ·
1.05M context
· 38th cheapest of 66 current models
These are promotional rates, in force until 2026-12-31. The standard rates are $1.5 in / $7.5 out per 1M tokens.
This row is flagged for review: vendor override differs from the automated feed (50%): $0.75/$3.75 vs $1.5/$7.5. Check the vendor’s own page before relying on it.
What Gemini 3.7 Flash costs
| Rate | Per 1M tokens | Per 1K tokens | What it covers |
|---|---|---|---|
| Input | $0.75 | $0.00075 | Every token you send: prompt, history, documents. |
| Output | $3.75 | $0.00375 | Every token the model generates, including hidden reasoning tokens. |
| Cached input | $0.075 | $0.000075 | Input served from the provider’s prompt cache. |
| Cache storage / hour | $1 | $0.001 | Separate residency charge per 1M cached tokens; not included without a retention duration. |
What that means in practice
The same three workloads are costed on every model page, so these numbers are directly comparable across the catalog. No prompt caching and no batch discount is assumed — the figures are what you pay before you optimise anything.
Support chatbot
2,500 six-turn conversations a day
$3,021 / month
$0.040 per conversation · $36,248 a year
An 800-token system prompt, 400-token questions and 900-token answers, over six turns — each turn re-sending everything before it.
Document summariser
1,000 single-shot summaries a day
$364 / month
$0.012 per conversation · $4,374 a year
One 12,000-token document in, an 800-token summary out, no conversation history. The case where input dominates the bill.
Coding agent
200 twelve-turn sessions a day
$2,106 / month
$0.351 per conversation · $25,272 a year
A 3,000-token system prompt, 1,500-token instructions and 2,500-token responses over twelve turns. Long sessions are where the quadratic cost of re-sent history stops being theoretical.
Change the assumptions in the calculator
Cheaper models in the catalog
Nothing cheaper is scored as highly as Gemini 3.7 Flash, so these are simply the cheaper options, best-scored first. Trading down here is a real trade.
| Model | Input | Output | Blended saving | |
|---|---|---|---|---|
| Kimi K2.6 | $0.95 | $4 | 43% | |
| Grok 4.3 | $1.25 | $2.5 | 48% | |
| GLM-5.1 | $1.4 | $4.4 | 28% | |
| DeepSeek V4 Pro | $1.32 | $3.96 | 34% | |
| Kimi K2.5 | $0.6 | $3 | 60% |
Specification
| Context window | 1.05M tokens |
|---|---|
| Maximum output | 66K tokens |
| Provider | |
| Status | Current |
| Reasoning model | Yes |
| Vision | Yes |
| Token counting | Estimated at ~4 characters per token (Estimate — Gemini's tokenizer is server-side) |
Where these numbers come from
Last checked against its source on 2026-08-16. Source of record: the vendor’s own published pricing.
The vendor page a human checked
Regional and data-residency premiums, priority tiers, server-side tool fees and negotiated discounts are not included. Prices change without notice; this page is rebuilt every time the catalog does.