xAI released Grok 4.6, a 1.5T model focused on long-running agents and interactive work, while Alibaba released open-weight Qwen3.8-Max and DeepSeek made V4 Pro generally available with aggressive pricing.
- Grok 4.6 scored 61 on the Intelligence Index and 88.4% on Terminal-Bench v2.1, priced at $2/$6 per 1M tokens.
- Qwen3.8-Max is a 2.4T total / 95B active MoE model with day-0 support for NVIDIA B300 and AMD MI355X.
- DeepSeek V4 Pro GA offers pricing around $0.435/M input and $0.87/M output, roughly 57x cheaper than Fable 5.
- Microsoft announced MAI-Thinking-1, its first reasoning model built from scratch, now available in Foundry.
- Upstage's Solar Pro 4 jumped to rank 42 on the Intelligence Index with gains in agentic tasks.
These releases highlight a competitive frontier landscape where efficiency and cost are becoming as critical as raw benchmark scores.