🇨🇳 Chinese model
QwQ 32B
Text Function calling Reasoning
Pricing & specs
Input / 1M tokens $0.29
Output / 1M tokens $0.86
Cache / 1M tokens —
Context window 131K tokens
Max output 8K tokens
Updated 2024-12
Docs Official docs ↗
Price source View source ↗
Providers
Similar models · Top 5
GLM-4.6V 🇨🇳 128K In $0.30 · Out $0.90 · Cache — siliconflow/deepseek-v3.1-terminus 🇨🇳 164K In $0.27 · Out $1.00 · Cache — siliconflow/deepseek-v3-0324 🇨🇳 164K In $0.25 · Out $1.00 · Cache — GLM-4.7 🇨🇳 205K In $0.30 · Out $1.19 · Cache $0.06 DeepSeek R1 Distill Qwen 32B 🇨🇳 33K In $0.29 · Out $0.86 · Cache —
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.