The Qwen3 series Turbo model effectively integrates thinking and non-thinking modes, allowing seamless switching between modes during conversations. With a smaller parameter size, its reasoning ability rivals that of QwQ-32B, and its general capabilities significantly surpass those of Qwen2.5-Turbo, reaching state-of-the-art (SOTA) levels among models of the same scale. This version is a snapshot model as of April 28, 2025.
Pricing
- Input Tokens: $0.046 /M tokens
- Output Tokens: $0.092 /M tokens
- Cache Read: $0.000 /M tokens
Input Modalities
Output Modalities
- Text
Try this model
Python
