Alibaba has released the model weights for Qwen3.8-Flash-Next, serving as a preview of the upcoming Qwen4 architecture for developers to experiment with and evaluate.

  • The model is a multimodal mixture-of-experts (MoE) with 176B total parameters, including 51B N-gram embedding parameters.
  • It activates 6B parameters per token while maintaining a native 262,144-token context window.
  • The context window is extensible to 1M tokens using YaRN.

This release allows developers to evaluate the new architecture's capabilities for agentic coding tasks before the official Qwen4 launch.