🇺🇸 Global model
Gemini 3.1 Flash Lite
Text Image Video Audio audio Function calling Structured output Reasoning Web search File input
Pricing & specs
Input / 1M tokens $0.25
Output / 1M tokens $1.50
Cache / 1M tokens $0.03
Context window 1.0M tokens
Max output 66K tokens
Updated 2026-05-07
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Google Official $0.25 / $1.50 Requesty lowest $0.23 / $1.35 NanoGPT $0.25 / $1.50 Impossibl $0.25 / $1.50 OpenRouter $0.25 / $1.50 LLM Gateway $0.25 / $1.50 NEAR AI Cloud $0.25 / $1.50 6 more at higher prices
Similar models · Top 5
MiniMax-M3 🇨🇳 1.0M In $0.31 · Out $1.25 · Cache $0.06 GPT-4.1 mini 🇺🇸 1.0M In $0.40 · Out $1.60 · Cache $0.10 Qwen3.6 Flash 🇨🇳 1M In $0.19 · Out $1.13 · Cache — GPT-5 Mini 🇺🇸 400K In $0.25 · Out $2.00 · Cache $0.03 DeepSeek V4 Flash 🇨🇳 1M In $0.45 · Out $1.34 · Cache $0.01
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.