Gemini 3.6 Flash API pricing
Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens at the standard rate. Here are its cached and batch prices, and what it costs for your usage.
Gemini 3.6 Flash has exactly the same prices as Gemini 3.8 Flash.
| Token type | Standard | Batch (50% off) |
|---|---|---|
| Input | $0.75 | $0.375 |
| Cached input | $0.075 | $0.0375 |
| Output | $3.75 | $1.875 |
Gemini 3.6 Flash: promotional price until 31 Dec 2026.
- Input $22.50
- Cached input $2.25
- Output $33.75
| Workload | Standard | Batch | |
|---|---|---|---|
Support chatbot 2,000 in / 300 out tokens,
1,000 requests a day, 50% cached | $58.50 | $29.25 | 7th cheapest of 18 |
Document summaries 10,000 in / 800 out tokens,
200 requests a day | $63.00 | $31.50 | 7th cheapest of 18 |
Tagging / classifying 400 in / 20 out tokens,
20,000 requests a day | $225.00 | $112.50 | 7th cheapest of 18 |
Other Google models
FAQ
Frequently asked questions
How much does Gemini 3.6 Flash cost per million tokens?
Gemini 3.6 Flash costs $0.75 per million input tokens and $3.75 per million output tokens at the standard rate. Prices are in USD, before tax, from Gemini API pricing, last checked on 30 Sept 2026.
Does Gemini 3.6 Flash have a Batch API discount?
Yes. Through the Batch API, Gemini 3.6 Flash costs $0.375 per million input tokens and $1.875 per million output tokens, 50% less than the standard rate. Batch requests are processed asynchronously rather than straight away.
How much does prompt caching save on Gemini 3.6 Flash?
Input read from the prompt cache costs $0.075 per million tokens instead of $0.75, 90% less. Writing to or storing the cache can cost extra; those charges aren’t included here.
Is Gemini 3.6 Flash one of the cheaper API models?
For a support chatbot sending 1,000 requests a day (2,000 input and 300 output tokens each, 50% of the input cached), Gemini 3.6 Flash is the 7th cheapest of the 18 OpenAI, Anthropic and Google models we track, at $58.50 a month. Price alone doesn’t tell you how well a model will handle your task, so test the candidates on your own prompts.