Gemini 3.5 Flash-Lite vs o4-mini
Google against OpenAI, costed on the same three workloads. Gemini 3.5 Flash-Lite is cheaper on all three workloads.
Monthly bill, side by side
| Workload | Gemini 3.5 Flash-Lite | o4-mini | Difference |
|---|---|---|---|
| Support chatbot 2,500 six-turn conversations a day |
$1,613 | $3,985 | $2,371 (2.5×) |
| Document summariser 1,000 single-shot summaries a day |
$170 | $508 | $338 (3×) |
| Coding agent 200 twelve-turn sessions a day |
$1,022 | $2,891 | $1,868 (2.8×) |
No prompt caching and no batch discount on either side, so the comparison is like for like. Both figures assume 30 days a month.
The rate cards
| Gemini 3.5 Flash-Lite | o4-mini | |
|---|---|---|
| Input, per 1M tokens | $0.3 | $1.1 |
| Output, per 1M tokens | $2.5 | $4.4 |
| Cached input | $0.03 | $0.275 |
| Context window | 1.05M | 200K |
| Maximum output | 66K | 100K |
| Reasoning model | yes | yes |
| Vision | yes | yes |
Run this comparison on your own numbers
Full detail: Gemini 3.5 Flash-Lite pricing · o4-mini pricing