🇨🇳 Chinese model
Qwen3-Next 80B-A3B (Thinking)
Text Function calling Reasoning
Pricing & specs
Input / 1M tokens $0.14
Output / 1M tokens $1.43
Cache / 1M tokens —
Context window 131K tokens
Max output 33K tokens
Updated 2025-09
Docs Official docs ↗
Price source View source ↗
Providers
USD / 1M tokens · input / output
Alibaba Cloud Official $0.14 / $1.43 Cortecs lowest $0.15 / $1.20 NanoGPT $0.15 / $0.65 Jiekou.AI $0.15 / $1.50 OpenRouter $0.15 / $1.20 LLM Gateway $0.15 / $1.20 Merge Gateway $0.15 / $1.20 5 more at higher prices
Similar models · Top 5
Step 3.7 Flash 🇨🇳 131K In $0.20 · Out $1.21 · Cache — GLM-4.5-Air 🇨🇳 131K In $0.20 · Out $1.10 · Cache $0.03 Doubao 1.6 🇨🇳 256K In $0.12 · Out $1.19 · Cache $0.02 GPT-5.4 nano 🇺🇸 400K In $0.20 · Out $1.25 · Cache $0.02 GPT-4o mini 🇺🇸 128K In $0.15 · Out $0.60 · Cache $0.07
Want to go deeper on AI and cloud cost management?
Explore more on AI infrastructure, model API cost and resource optimization at mofcloud.