gemini-3.6-flash
Google Vertex
gemini-3.6-flash
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and less hedging, while reducing token use and the number of model calls needed to complete a task.
Input / output per 1M
$0.75 / $3.75
Context Window
1M
Max output
65,536
Released
Jul 21, 2026
Provider pricing
The same model is offered by several providers. Prices below are per provider; select a row to switch provider.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output |
|---|---|---|---|---|---|
Current | $0.75 | $3.75 | $0.075 | 1M | 65,536 |
Pricing
Prices from the current provider, per 1M tokens.
| Input / 1M | Output / 1M | Cache read / 1M |
|---|---|---|
| $0.75 | $3.75 | $0.075 |
Performance
Data for this section is coming soon.
Uptime
Data for this section is coming soon.