All models

Gemini 3.6 Flash

google/gemini-3.6-flash

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

Modalities
Text
Image
Video
File
Audio
In / Out price
$1.50 / $7.50 per 1M
Context
1.05M · 66K out
Released
Jul 21, 2026

Providers

Live list prices and uptime across 6 providers serving this model.

ProviderInput /MOutput /MCache read /MContextMax outputUptime 24h
Google$1.50$7.50$0.151.05M66K99.38%
Google$0.75$3.75$0.0751.05M66K99.38%
Google$2.70$13.50$0.271.05M66K99.38%
Google AI Studio$1.50$7.50$0.151.05M66K98.89%
Google AI Studio$0.75$3.75$0.0751.05M66K98.89%
Google AI Studio$2.70$13.50$0.271.05M66K98.89%

Implicit prompt caching available — repeated context bills at the cache read rate.

Capabilities

Tool calling
Reasoning
include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
stop
structured_outputs
temperature
tool_choice
top_p

Pricing detail

Input
$1.50
Output
$7.50
Cache read
$0.15
Cache write
$0.083

USD per 1M tokens.