🇨🇳 Chinese model
Qwen3 32B
Text Function calling Reasoning
Pricing & specs
Input / 1M tokens $0.29
Output / 1M tokens $1.15
Cache / 1M tokens —
Context window 131K tokens
Max output 16K tokens
Updated 2025-04
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Alibaba Cloud Official $0.29 / $1.15 Deep Infra lowest $0.08 / $0.28 Abacus $0.09 / $0.29 Nebius Token Factory $0.10 / $0.30 SiliconFlow $0.14 / $0.57 SiliconFlow (China) $0.14 / $0.57 Hugging Face $0.29 / $0.59 1 more at higher prices
Similar models · Top 5
GLM-4.7 🇨🇳 205K In $0.30 · Out $1.19 · Cache $0.06 MiniMax-M2.5 🇨🇳 205K In $0.30 · Out $1.20 · Cache — MiniMax-M2.1 🇨🇳 205K In $0.30 · Out $1.20 · Cache $0.03 MiniMax-M2 🇨🇳 205K In $0.30 · Out $1.20 · Cache — DeepSeek V3 🇨🇳 66K In $0.29 · Out $1.15 · Cache —
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.