GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.
GLM 5.2 vs Qwen Turbo
Output tokens cost $3.94 per million on GLM 5.2 and $0.09 per million on Qwen Turbo. Input tokens cost $1.13 per million on GLM 5.2 and $0.05 per million on Qwen Turbo. Cached input tokens are billed at $0.28 per million on GLM 5.2 and $0.01 per million on Qwen Turbo.
GLM 5.2The Qwen series model with the fastest speed and lowest cost, suitable for simple tasks. This model is a dynamically updated version, and updates will not be announced in advance. The model's overall Chinese and English abilities have been significantly improved, human preference alignment has been greatly enhanced, inference capability and complex instruction understanding have been substantially strengthened, performance on difficult tasks is better, and mathematics and coding skills have been significantly improved. The current version is qwen-turbo-2025-04-28.
Pricing & Specifications
Prices are per million tokens. Time to First Token and throughput are rolling averages measured on AIHubMix.
Promotional prices show the discounted rate; see each model page for promotion windows.
Activity Past 30 Days
Daily traffic served through AIHubMix — how demand for each model is trending.
Tokens / day
Requests / day
Performance Past 3 Days
Measured on real AIHubMix traffic, hourly buckets. Gaps mean no traffic in that hour.
Throughput (tok/s)
TTFT (s)
Uptime (%)
Cost calculator
Estimate your monthly bill for the same workload on each model.
Monthly = daily × 30. Discounted rates applied where a promotion is active.
FAQ
Which is cheaper: GLM 5.2, Qwen Turbo?
Qwen Turbo: $0.09/M output tokens; GLM 5.2: $3.94/M. Use the cost calculator above to estimate your own workload.
Can I call GLM 5.2 and Qwen Turbo with the same API key?
Yes. AIHubMix serves every model on this page behind one OpenAI-compatible endpoint, so switching between them is a one-line change to the model field — no second account, key or SDK.
Popular comparisons
Related model match-ups readers also look at.
