Alibaba's Qwen team has made Qwen3.8-Max broadly available via API, with open weights for both Qwen3.8-Max and the smaller Qwen3.8-27B checkpoint scheduled to ship next week. The flagship model is a 2.4-trillion-parameter mixture-of-experts architecture that accepts text, image, and video inputs.
- Qwen3.8-Max features a 1M-token context window, with maximum input of 991K tokens and output of 131K tokens.
- Pricing is set at $2.00 per 1M input tokens, $6.00 per 1M output tokens, and $0.25 for cached input reads.
- The model scores 86.6 on Terminal-Bench 2.1 and leads PaperBench at 93.0, showing significant gains in multimodal and agentic tasks over its predecessor.
- Supported capabilities include function calling, structured outputs, and five built-in tools such as code_interpreter and web_search.
The release provides a highly capable multimodal option for industries like software engineering and legal review, though the full 2.4T parameter count limits on-premise deployment to the smaller Qwen3.8-27B variant.