The Zai organization has made the weights for the GLM-5.3 model available on Hugging Face, fulfilling a previous announcement.
- The GLM-5.3 checkpoint is now hosted.
- This release follows a prior promise to make the model publicly accessible.
The Zai organization has made the weights for the GLM-5.3 model available on Hugging Face, fulfilling a previous announcement.
Nvidia is acquiring HuggingFace for $13 billion, nearly double its initial January 2026 offer, as the platform doubles its customer base in 2026. Simultaneously, Z.ai has formally launched GLM-5.3-Flash, a natively multimodal open-weight model previously known as Ox Alpha.
Z.ai has formally launched GLM-5.3-Flash, revealing that the previously previewed "Ox Alpha" model is its public identity. The model features 320 billion total parameters with 18 billion active, a 1 million-token context window, and native multimodal capabilities.
Z.ai has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series, featuring a mixture-of-experts architecture with 320 billion total parameters and 18 billion active per token. The model supports image and video inputs within a 1,048,576-token context window and is available under an MIT license on Hugging Face.
Zhipu AI has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series and the first open-weight release of the glm5_next architecture. The 320B-parameter model is trained on a 30T-token multimodal corpus and features a hybrid sparse and linear attention mechanism to reduce long-context serving costs.
ZhiPu introduces GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series, featuring a hybrid architecture that combines sparse and linear attention to reduce long-context serving costs while preserving precision.
We use cookies to measure traffic and improve the site. You can accept or decline analytics cookies. Privacy policy