The article confirms that the model known as Ox Alpha is GLM-5.3-Flash.
- Multimodal (Vision)
- 1M Tokens Context Window
- DeepSWE ~63%
The article confirms that the model known as Ox Alpha is GLM-5.3-Flash.
| Benchmark | Model | Score |
|---|---|---|
| DeepSWE | GLM-5.3-Flash | 63% |
Zai has introduced GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series. It features a hybrid architecture combining sparse and linear attention to reduce long-context serving costs while maintaining precise capabilities.
Zhipu AI has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series and the first open-weight release of the glm5_next architecture. The 320B-parameter model is trained on a 30T-token multimodal corpus and features a hybrid sparse and linear attention mechanism to reduce long-context serving costs.
ZhiPu introduces GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series, featuring a hybrid architecture that combines sparse and linear attention to reduce long-context serving costs while preserving precision.
Z.ai has made the GLM-5.3 API available with unchanged pricing, offering stronger coding and long-horizon agent performance compared to its predecessor. Concurrently, Cerebras introduced its new CS-4 computer, which it claims is multiple times faster than its previous generation and more responsive than Nvidia-based systems.
Mistral is now hosting the GLM-5.2 model, which is priced even cheaper than its current flagship model, Mistral Medium 3.5.
We use cookies to measure traffic and improve the site. You can accept or decline analytics cookies. Privacy policy