Zhipu AI has released the GLM-4.5 family of open-weights large language models, designed to excel at tool use and coding while offering switchable reasoning capabilities.
- The GLM-4.5 model contains 355 billion total parameters with 32 billion active, while the smaller GLM-4.5-Air has 106 billion total and 12 billion active parameters.
- Both models outperform Anthropic Claude 4 Opus, DeepSeek-R1-0528, Google Gemini 2.5 Pro, Grok 4, Kimi K2, and OpenAI o3 on at least one reasoning, coding, or agentic benchmark.
- GLM-4.5 achieved 90.6 percent accuracy in tool-use benchmarks, surpassing Claude Sonnet 4 and Kimi K2, and matched Claude 4 Opus on MATH 500 with 98.2 percent accuracy.
- Weights are available via HuggingFace and ModelScope under an MIT license for commercial and noncommercial use.
Zhipu AI's approach distills three specialized variations of the base model into a single architecture, establishing momentum in open-weights models tuned for agentic behavior.