All models
Gemini 3.7 Flash
google/gemini-3.7-flash
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Modalities
Text
Image
Video
File
Audio
In / Out price
$0.375 / $1.875 per 1M
Context
1.05M · 66K out
Released
Aug 13, 2026
Providers
Live list prices and uptime across 6 providers serving this model.
| Provider | Input /M | Output /M | Cache read /M | Context | Max output | Uptime 24h |
|---|---|---|---|---|---|---|
| $0.375 | $1.875 | $0.037 | 1.05M | 66K | 99.44% | |
| $0.188 | $0.938 | $0.019 | 1.05M | 66K | 99.44% | |
| $0.675 | $3.375 | $0.068 | 1.05M | 66K | 99.44% | |
| Google AI Studio | $0.75 | $3.75 | $0.075 | 1.05M | 66K | 99.70% |
| Google AI Studio | $0.375 | $1.875 | $0.037 | 1.05M | 66K | 99.70% |
| Google AI Studio | $1.35 | $6.75 | $0.135 | 1.05M | 66K | 99.70% |
Implicit prompt caching available — repeated context bills at the cache read rate.
Capabilities
Tool calling
Reasoning
include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
stop
structured_outputs
temperature
tool_choice
top_p
Pricing detail
- Input
- $0.375
- Output
- $1.875
- Cache read
- $0.037
- Cache write
- $0.021
USD per 1M tokens.