Head to head
gemini-3.1-pro-preview vs gemini-3.7-flash: API Pricing Compared
On a typical chat turn of 10K input and 1K output tokens, gemini-3.7-flash costs $0.011 against $0.032 for gemini-3.1-pro-preview, making it 65% cheaper. Both rates are read daily from the providers' own pricing pages.
Side by side
| gemini-3.1-pro-preview | gemini-3.7-flash | |
|---|---|---|
| Provider | ||
| Input / 1M tokens | $2.00 | $0.75 |
| Output / 1M tokens | $12.00 | $3.75 |
| Context window | Not published | Not published |
| Blended cost / chat turn | $0.032 | $0.011 |
| Price verified | 2026-06-11 | 2026-08-14 |
Cost by workload
The headline rate rarely decides the bill. These are the same four workloads priced against both models.
| Workload | Tokens | gemini-3.1-pro-preview | gemini-3.7-flash |
|---|---|---|---|
| Single short call | 2K in, 500 out | $0.010 | $0.00337 |
| Typical chat turn | 10K in, 1K out | $0.032 | $0.011 |
| Long document pass | 100K in, 2K out | $0.224 | $0.083 |
| 1,000 chat turns | 10M in, 1M out | $32.00 | $11.25 |
Which one actually wins
gemini-3.7-flash is cheaper on both the input and the output rate, so it never costs more no matter how input-heavy or output-heavy the workload is. There is no crossover point to reason about. Price is only one axis, so weigh it against context window, latency, and the quality difference on your own evaluation set before switching.
Common questions
Which is cheaper, gemini-3.1-pro-preview or gemini-3.7-flash?
gemini-3.7-flash is cheaper for a typical chat turn of 10K input and 1K output tokens, at $0.011 against $0.032 for gemini-3.1-pro-preview, a difference of 65%.
How much does gemini-3.1-pro-preview cost per 1M tokens?
Google charges $2.00 per 1M input tokens and $12.00 per 1M output tokens for gemini-3.1-pro-preview, verified 2026-06-11.
How much does gemini-3.7-flash cost per 1M tokens?
Google charges $0.75 per 1M input tokens and $3.75 per 1M output tokens for gemini-3.7-flash, verified 2026-08-14.
Does gemini-3.7-flash stay cheaper for every workload?
Yes. gemini-3.7-flash is cheaper on both the input and the output rate, so it never costs more regardless of how input-heavy or output-heavy the workload is.
Full pricing pages
Maintained by Jordan Hale, TechFastForward. Both prices are read daily from the official provider pages and recorded only when they move. Compare any other pair on the LLM API Pricing Tracker.