All models
Gemini 3.6 Flash
google/gemini-3.6-flash
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Modalities
Text
Image
Video
File
Audio
In / Out price
$1.50 / $7.50 per 1M
Context
1.05M · 66K out
Released
Jul 21, 2026
Providers
Live list prices and uptime across 6 providers serving this model.
| Provider | Input /M | Output /M | Cache read /M | Context | Max output | Uptime 24h |
|---|---|---|---|---|---|---|
| $1.50 | $7.50 | $0.15 | 1.05M | 66K | 99.38% | |
| $0.75 | $3.75 | $0.075 | 1.05M | 66K | 99.38% | |
| $2.70 | $13.50 | $0.27 | 1.05M | 66K | 99.38% | |
| Google AI Studio | $1.50 | $7.50 | $0.15 | 1.05M | 66K | 98.89% |
| Google AI Studio | $0.75 | $3.75 | $0.075 | 1.05M | 66K | 98.89% |
| Google AI Studio | $2.70 | $13.50 | $0.27 | 1.05M | 66K | 98.89% |
Implicit prompt caching available — repeated context bills at the cache read rate.
Capabilities
Tool calling
Reasoning
include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
stop
structured_outputs
temperature
tool_choice
top_p
Pricing detail
- Input
- $1.50
- Output
- $7.50
- Cache read
- $0.15
- Cache write
- $0.083
USD per 1M tokens.