Kimi K3 has achieved the top position on AfterQuery's SpreadsheetBench 2 evaluation. This result places it ahead of Claude Fable 5 in this specific benchmark.
Kimi K3 ranks #1 on AfterQuery's SpreadsheetBench 2, surpassing Claude Fable 5
Kimi K3 ranks at same level as Opus thinking on Agent Arena
According to the Agent Arena leaderboard, Kimi K3 performs at a level comparable to Opus in non-vision tasks. While some users note that Opus may have an advantage in vision capabilities, the ranking indicates parity for other use cases.
Kimi K3 (max) beats Sonnet 5 on Simple Bench
The article reports that Kimi K3 (max) outperforms Sonnet 5 on the Simple Bench benchmark.
Moonshot AI plans Hong Kong IPO following Kimi K3 release
Following the release of the Kimi K3 model, Moonshot AI has informed investors that it plans an initial public offering in Hong Kong within the next six months.
Moonshot AI releases Kimi K3, a 2.8T parameter MoE model
On July 16th, Moonshot AI released its latest flagship model, Kimi K3, a 2.8 trillion parameter Mixture of Experts (MoE) architecture. The company has promised to release the model's weights on July 27th.
Kimi K3, DeepSeek V4 Pro, and GLM-5.2 compared on benchmarks, license, and serving cost
A comparison of three open-weight sparse Mixture-of-Experts models—Moonshot AI's Kimi K3, DeepSeek V4 Pro, and Zhipu AI's GLM-5.2—evaluates their capabilities, licensing terms, and serving costs for long-horizon coding and agent workloads.