🇨🇳 Chinese model
GLM-4.5-Air
Text Function calling Reasoning
Pricing & specs
Input / 1M tokens $0.20
Output / 1M tokens $1.10
Cache / 1M tokens $0.03
Context window 131K tokens
Max output 98K tokens
Updated 2025-07-28
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Zhipu GLM Official $0.20 / $1.10 ZenMux lowest $0.11 / $0.56 302.AI $0.11 / $0.29 OpenRouter $0.13 / $0.85 LLM Gateway $0.13 / $0.85 NovitaAI $0.13 / $0.85 DevPass (LLM Gateway) $0.13 / $0.85 5 more at higher prices
Similar models · Top 5
Step 3.7 Flash 🇨🇳 131K In $0.20 · Out $1.21 · Cache — Qwen3 32B 🇨🇳 131K In $0.29 · Out $1.15 · Cache — Qwen3 235B-A22B 🇨🇳 131K In $0.29 · Out $1.15 · Cache — siliconflow/deepseek-v3-0324 🇨🇳 164K In $0.25 · Out $1.00 · Cache — siliconflow/deepseek-v3.1-terminus 🇨🇳 164K In $0.27 · Out $1.00 · Cache —
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.