Pricing
- Input Tokens: $0.000 /M tokens
- Output Tokens: $0.000 /M tokens
- Cache Read: $0.000 /M tokens
Input Modalities
- Text
Output Modalities
- Text
Capabilities
- Thinking
- Long context
Try this model
import os
from openai import OpenAI
client = OpenAI(
api_key=os.environ["AIHUBMIX_API_KEY"],
base_url="https://aihubmix.com/v1",
)
response = client.chat.completions.create(
model="qwen3.6-plus-preview-free",
messages=[
{
"role": "user",
"content": "Hello, how are you?"
}
],
max_tokens=1024,
stream=False,
)
print(response.choices[0].message.content)Frequently asked questions
What is Qwen3.6 Plus Preview (free)?
What is the context length of Qwen3.6 Plus Preview (free)?
What modalities does Qwen3.6 Plus Preview (free) support?
What capabilities does Qwen3.6 Plus Preview (free) support?
How do I call Qwen3.6 Plus Preview (free) via API?
Who created Qwen3.6 Plus Preview (free)?
Compare Qwen3.6 Plus Preview (free)
More models from Qwen
See all Qwen models →Qwen3.8-2.4T-A95B is Alibaba’s most powerful Qwen model to date. It is a 2.4‑trillion‑parameter sparse Mixture-of-Experts (MoE) model with approximately 95 billion active parameters. It is built for autonomous, long‑duration tasks: multi‑day code runs, reproducing research papers, and self‑improvement.
Image input: $0.00286 per image;Image generation: $0.02535 per image.
Qwen Image 3.0(qwen-image-3.0) is an image generation and editing model developed by Alibaba Cloud’s Qwen team. It supports text-to-image generation, reference-based creation, and image editing. It is well suited for social media, e-commerce, creative design, and everyday content production, offering a strong balance of image quality and speed. Compared with the Pro version, it is better suited for frequent and large-scale daily creation.
Image input: $0.00286 per image; 1K image generation: $0.03572 per image; 2K image generation: $0.07143 per image.
Qwen Image 3.0 Pro (qwen-image-3.0-pro) is Alibaba Cloud Qwen’s flagship image generation and editing model. It is designed for advertising, brand visuals, UI, presentations, product imagery, and professional design. Its strengths include complex layouts, accurate Chinese and English text rendering, realistic materials, and reference-based editing. Compared with the standard version, it delivers stronger detail, composition, and commercial-grade visual quality.
Qwen 3.8 Max(qwen3.8-max) is Alibaba Cloud’s flagship native vision-language model, built on a 2.4-trillion-parameter Mixture-of-Experts (MoE) architecture and supporting context windows of up to 1 million tokens. It is well suited for complex multimodal understanding, advanced reasoning, software development, agentic workflows, and long-context processing. At a similar price to Qwen3.7-Max, Qwen3.8-Max delivers significant improvements in reasoning, coding, and agent capabilities, with overall performance comparable to today’s leading models.
- Input: $ 0.338 /M
- Output: $ 1.014 /M
- Web Search: $0.000548/request
Qwen 3.8 Max Preview(Qwen3.8-Max-Preview) is the latest-generation foundation model in the Qwen family, packing 2.4T parameters and still evolving. Compared with the previous flagship Qwen 3.7 Max, it delivers major gains in core capabilities like Coding and Cowork (professional productivity), with world-leading performance on complex, long-horizon tasks such as full-stack development, data analysis, and Office workflows. Launch offer: Credits are consumed at just 20% of the standard rate, effectively 5× your usage. Limited time only.
$0.141 / 10K characters
qwen-audio-3.0-tts-flash is a high-performance speech synthesis large model optimized for real-time interactive scenarios. Compared with the previous version, the model supports more low-resource languages and Chinese dialects, improves the authenticity of dialect pronunciation, and enhances free-style instruction following and fine-grained label control, enabling more flexible control of expression such as emotion, tone, character, speaking rate, and volume. At the same time, the model exhibits stronger robustness under complex acoustic conditions like noise and reverberation, improving sound quality, clarity, and overall expressiveness. The Flash version focuses on optimizing the real-time synthesis experience, keeping first-packet latency under 200 ms, making it suitable for low-latency interactive scenarios such as voice assistants, real-time dialogue, and intelligent customer service.
AIHubMix© 2023 - 2026 AIHubMix, LLC