Qwen-Maxalibaba/qwen-max
Qwen-Max, based on Qwen2.5, provides the best inference performance among [Qwen models](/qwen), especially for complex multi-step tasks. It's a large-scale MoE model that has been pretrained on over 20 trillion...
Pricing across providers
| Provider | Input /1M | Output /1M | Blended /1M | Latency p50 | Format | Freshness | Action |
|---|
1 hosting hidden — prices older than 60 days are pending re-verification.
Affiliate disclosure: We may earn a commission from qualified signups. Pricing independence is enforced at the data layer — see our Editorial Independence Policy.
Works with
Point any of these clients at a hosting's base URL — they all speak at least one of this model's endpoint protocols (OPENAI_COMPATIBLE).
Capabilities
Code samples
Example using Alibaba Cloud DashScope — the cheapest hosting for this model as of last verification. Swap base_url and model to use a different provider from the matrix above.
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://dashscope.aliyun.com/compatible-mode/v1",
)
response = client.chat.completions.create(
model="qwen-max",
messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)
Technical specs
- Context
- 33K
- Max output
- 8K
- Parameters
- —
- Release
- —
- Training cutoff
- —
- License
- —
Similar models
Compare with
- Qwen-Max vs DeepSeek V3Comparison planned — not yet published
- Qwen-Max vs DeepSeek V3 0324Comparison planned — not yet published
- Qwen-Max vs DeepSeek V3.1Comparison planned — not yet published
Frequently asked
How much does Qwen-Max cost?+−
How do I access Qwen-Max from outside China?+−
Is Qwen-Max open-source? Can I fine-tune it?+−
Is Qwen-Max OpenAI-compatible?+−
openai SDK client at the Provider's base_url and use the Provider's model name. See the Code Samples above for a copy-pasteable example.