Qwen
Qwen: Qwen-Max
Qwen-Max, based on Qwen2.5, provides the best inference performance among [Qwen models](/qwen), especially for complex multi-step tasks. It's a large-scale MoE model that has been pretrained on over 20 trillion tokens and further post-trained with curated Supervised Fine-Tuning (SFT) and Reinforcement Learning from Human Feedback (RLHF) methodologies. The parameter count is unknown.
| Specifications | |
|---|---|
| Provider | Qwen |
| Input pricing | $1.04/M tokens |
| Output pricing | $4.16/M tokens |
| Context length | 33K tokens |
| Max output | 8K tokens |
| Input modalities | text |
| Output modalities | text |
| Tokenizer | Qwen |
| Knowledge cutoff | 2025-03-31 |
| Content moderated | No |
Supported Parameters
max_tokenspresence_penaltyresponse_formatseedtemperaturetool_choicetoolstop_p