π¨π³ Chinese model
GLM-5.3-Flash
Text Image Video Function calling Structured output Reasoning File input
Pricing & specs
Input / 1M tokens $0.07
Output / 1M tokens $0.25
Cache / 1M tokens $0.01
Context window 1M tokens
Max output 131K tokens
Updated 2026-08-26
Docs Official docs β
Price source View source β
Providers
Similar models Β· Top 5
Qwen Plus π¨π³ 1M In $0.12 Β· Out $0.29 Β· Cache $0.01 DeepSeek V4 Flash Vision Exp π¨π³ 1M In $0.14 Β· Out $0.28 Β· Cache $0.00 Gemini 2.5 Flash-Lite πΊπΈ 1.0M In $0.10 Β· Out $0.40 Β· Cache $0.01 Qwen3.5 Flash π¨π³ 1M In $0.03 Β· Out $0.30 Β· Cache β Qwen Long π¨π³ 10M In $0.07 Β· Out $0.29 Β· Cache β
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.