gemini-3.7-flash
Google Vertex
gemini-3.7-flash
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step problem solving.
Input / output per 1M
$0.75 / $3.75
Context Window
1M
Max output
65,536
Released
Aug 13, 2026
Provider pricing
The same model is offered by several providers. Prices below are per provider; select a row to switch provider.
| Provider | Input / 1M | Output / 1M | Cache read / 1M | Context | Max output |
|---|---|---|---|---|---|
Current | $0.75 | $3.75 | $0.075 | 1M | 65,536 |
Pricing
Prices from the current provider, per 1M tokens.
| Input / 1M | Output / 1M | Cache read / 1M |
|---|---|---|
| $0.75 | $3.75 | $0.075 |
Performance
Data for this section is coming soon.
Uptime
Data for this section is coming soon.