🇺🇸 Global model
Gemini 3.7 Flash
Text Image Video Audio audio Function calling Structured output Reasoning Web search File input
Pricing & specs
Input / 1M tokens $0.75
Output / 1M tokens $3.75
Cache / 1M tokens $0.07
Context window 1.0M tokens
Max output 66K tokens
Updated 2026-08-13
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Google Official $0.75 / $3.75 NanoGPT lowest $0.38 / $1.88 OpenRouter $0.38 / $1.88 Requesty $0.60 / $3.00 Cortecs $0.75 / $3.75 CrossModel $0.75 / $3.75 Opper $0.75 / $3.75 6 more at higher prices
Similar models · Top 5
GPT-5.4 mini 🇺🇸 400K In $0.75 · Out $4.50 · Cache $0.07 Qwen3 Coder Plus 🇨🇳 1.0M In $1.00 · Out $5.00 · Cache — GLM-5.2 🇨🇳 1M In $1.19 · Out $4.17 · Cache $0.04 Qwen3.7 Plus 🇨🇳 1M In $0.50 · Out $3.00 · Cache $0.05 Qwen3.6 Plus 🇨🇳 1M In $0.50 · Out $3.00 · Cache $0.05
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.