All models

Gemini 3.7 Flash

google/gemini-3.7-flash

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

Modalities
Text
Image
Video
File
Audio
In / Out price
$0.375 / $1.875 per 1M
Context
1.05M · 66K out
Released
Aug 13, 2026

Providers

Live list prices and uptime across 6 providers serving this model.

ProviderInput /MOutput /MCache read /MContextMax outputUptime 24h
Google$0.375$1.875$0.0371.05M66K99.44%
Google$0.188$0.938$0.0191.05M66K99.44%
Google$0.675$3.375$0.0681.05M66K99.44%
Google AI Studio$0.75$3.75$0.0751.05M66K99.70%
Google AI Studio$0.375$1.875$0.0371.05M66K99.70%
Google AI Studio$1.35$6.75$0.1351.05M66K99.70%

Implicit prompt caching available — repeated context bills at the cache read rate.

Capabilities

Tool calling
Reasoning
include_reasoning
max_tokens
reasoning
reasoning_effort
response_format
seed
stop
structured_outputs
temperature
tool_choice
top_p

Pricing detail

Input
$0.375
Output
$1.875
Cache read
$0.037
Cache write
$0.021

USD per 1M tokens.