Qwen
Qwen: Qwen3 235B A22B Thinking 2507
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
| Specifications | |
|---|---|
| Provider | Qwen |
| Input pricing | $0.30/M tokens |
| Output pricing | $3.00/M tokens |
| Context length | 262K tokens |
| Max output | 33K tokens |
| Input modalities | text |
| Output modalities | text |
| Tokenizer | Qwen3 |
| Knowledge cutoff | 2025-06-30 |
| Content moderated | No |
| HuggingFace | Qwen/Qwen3-235B-A22B-Thinking-2507 |
Supported Parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningrepetition_penaltyresponse_formatseedstoptemperaturetool_choicetoolstop_ktop_logprobstop_p