The new DeepSeek V4-Flash model has achieved a score of 50 on the ArtificialAnalysis Index. This result places it just one point below GLM-5.2 and GPT-5.6 Luna.
DeepSeek V4-Flash achieves 50 on ArtificialAnalysis Index
Benchmarks
| Benchmark | Model | Score |
|---|---|---|
| Artificial Analysis Intelligence Index | DeepSeek V4-Flash | 50% |
Kimi K3, DeepSeek V4 Pro, and GLM-5.2 compared on benchmarks, license, and serving cost
A comparison of three open-weight sparse Mixture-of-Experts models—Moonshot AI's Kimi K3, DeepSeek V4 Pro, and Zhipu AI's GLM-5.2—evaluates their capabilities, licensing terms, and serving costs for long-horizon coding and agent workloads.
Base-10's Charlie O'Neill argues Kimi and GLM are better than Opus 5
Charlie O'Neill of Base-10 discusses why the Kimi and GLM models are "almost objectively" superior to Opus 5 in a recent episode featuring Dwarkesh Patel.
DeepSeek V4.1 Flash tops Artificial Analysis Intelligence Index v4.3
DeepSeek V4.1 Flash has taken first place on the new Astra benchmark within the Artificial Analysis Intelligence Index v4.3 update.
LiveBench adds DeepSeek v4.1 Flash
The LiveBench benchmarking platform has added the DeepSeek v4.1 Flash model to its evaluation suite.
GLM-5.3 Flash distillation trades consistency for 17x cost reduction
A comparison of GLM-5.3 and its distilled variant, GLM-5.3 Flash, on the DeepSWE benchmark reveals that distillation preserves core coding capability while significantly reducing rollout costs. The full model achieves a 69.0% pass@1 rate at $3.99 per task, whereas the Flash variant scores 63.4% at just $0.24, representing a 17x price reduction.