DeepSeek has released DeepSeek-V3-0324, a model update that demonstrates notable improvements over its predecessor, DeepSeek-V3, particularly in reasoning capabilities, code executability, and Chinese writing proficiency.

  • MMLU-Pro score increased from 75.9 to 81.2, GPQA from 59.1 to 68.4, AIME from 39.6 to 59.4, and LiveCodeBench from 39.2 to 49.2.
  • Enhanced front-end web development output with improved code executability and aesthetics.
  • Optimized Chinese writing style to align with R1 and improved medium-to-long-form content quality.
  • Fixed function calling accuracy issues present in previous V3 versions and enhanced multi-turn interactive rewriting.

The update provides better benchmark performance and more reliable function calling for developers using libraries like Transformers, vLLM, or SGLang.