GLM-5.3 Flash distillation trades consistency for 17x cost reduction
A comparison of GLM-5.3 and its distilled variant, GLM-5.3 Flash, on the DeepSWE benchmark reveals that distillation preserves core coding capability while significantly reducing rollout costs. The full model achieves a 69.0% pass@1 rate at $3.99 per task, whereas the Flash variant scores 63.4% at just $0.24, representing a 17x price reduction.