Upstage has released the Solar Open 2 family of open-weight language models, including a 250B-A15B variant and a 100B-A12B variant. The release highlights competitive benchmark scores against proprietary models like DeepSeek-V4-Flash.

  • Solar Open 2 (250B-A15B) achieves 86.2 on MMLU-Pro, 86.3 on GPQA-Diamond, and 92.4 on LiveCodeBench v6.
  • The smaller Solar Open 100B (102B-A12B) model scores 80.4 on MMLU-Pro and 56.5 on LiveCodeBench v6.
  • On agent benchmarks, the 250B model reaches 70.4 on SWE-Bench Verified and 16.6 on APEX-Agents.
  • Performance is compared directly against DeepSeek-V4-Flash, Command A+, Mistral Medium 3.5, and MiMo-V2.5 across knowledge, reasoning, and agent tasks.

The models are available for download via Hugging Face.