Gemini 3.6 Flash pricing

Google · gemini-gemini-3.6-flash · 1.05M context · 39th cheapest of 67 current models

These are promotional rates, in force until 2026-12-31. The standard rates are $1.5 in / $7.5 out per 1M tokens.

What Gemini 3.6 Flash costs

USD, standard tier, global endpoint
RatePer 1M tokensPer 1K tokensWhat it covers
Input $0.75 (Promotional rate until 2026-12-31 — standard rate $1.5/M) $0.00075 Every token you send: prompt, history, documents.
Output $3.75 (Promotional rate until 2026-12-31 — standard rate $7.5/M) $0.00375 Every token the model generates, including hidden reasoning tokens.
Cached input $0.075 (Promotional rate until 2026-12-31 — standard rate $0.15/M) $0.000075 Input served from the provider’s prompt cache.
Cache storage / hour $1 $0.001 Separate residency charge per 1M cached tokens; not included without a retention duration.

What that means in practice

The same three workloads are costed on every model page, so these numbers are directly comparable across the catalog. No prompt caching and no batch discount is assumed — the figures are what you pay before you optimise anything.

Support chatbot

2,500 six-turn conversations a day

$3,021 / month

$0.040 per conversation · $36,248 a year

An 800-token system prompt, 400-token questions and 900-token answers, over six turns — each turn re-sending everything before it.

Document summariser

1,000 single-shot summaries a day

$364 / month

$0.012 per conversation · $4,374 a year

One 12,000-token document in, an 800-token summary out, no conversation history. The case where input dominates the bill.

Coding agent

200 twelve-turn sessions a day

$2,106 / month

$0.351 per conversation · $25,272 a year

A 3,000-token system prompt, 1,500-token instructions and 2,500-token responses over twelve turns. Long sessions are where the quadratic cost of re-sent history stops being theoretical.

Change the assumptions in the calculator

Cheaper, and scored at least as capable

Every model below costs less on a blended rate and carries a capability score at least as high as Gemini 3.6 Flash’s. The score is a rough estimate, not a benchmark — treat it as a shortlist, not a verdict.

ModelInputOutputBlended saving
Kimi K2.6 $0.95 $4 43% side by side

Head to head

Specification

Context window1.05M tokens
Maximum output66K tokens
ProviderGoogle
StatusCurrent
Reasoning modelYes
VisionYes
Token countingEstimated at ~4 characters per token (Estimate — Gemini's tokenizer is server-side)

Where these numbers come from

Last checked against its source on 2026-09-16. Source of record: the vendor’s own published pricing.

The vendor page a human checked

Regional and data-residency premiums, priority tiers, server-side tool fees and negotiated discounts are not included. Prices change without notice; this page is rebuilt every time the catalog does.