🇨🇳 Chinese model
Qwen3.8 Flash
Text Image Video Function calling Structured output Reasoning File input
Pricing & specs
Input / 1M tokens $0.12
Output / 1M tokens $0.40
Cache / 1M tokens $0.01
Context window 1M tokens
Max output 131K tokens
Updated 2026-08-26
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Alibaba Cloud Official $0.12 / $0.40 Vancine lowest $0.12 / $0.38 CrossModel $0.13 / $0.43 OpenRouter $0.15 / $0.47 OpenCode Go $0.15 / $0.47 LLM Gateway $0.15 / $0.47 Charm Hyper $0.15 / $0.47 5 more at higher prices
Similar models · Top 5
Gemini 2.5 Flash-Lite 🇺🇸 1.0M In $0.10 · Out $0.40 · Cache $0.01 DeepSeek V4 Flash Vision Exp 🇨🇳 1M In $0.14 · Out $0.28 · Cache $0.00 Doubao 1.5 Pro 🇨🇳 256K In $0.12 · Out $0.30 · Cache $0.02 Hunyuan TurboS 🇨🇳 256K In $0.12 · Out $0.30 · Cache — GLM-5.3-Flash 🇨🇳 1M In $0.07 · Out $0.25 · Cache $0.01
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.