vLLM has introduced a new streaming parser for Qwen3+ available in its nightly build, addressing issues like mid-turn stopping and failed streaming tool calls due to chunk boundaries. The update reportedly resolves these problems in limited testing, improving reliability for agentic workflows.
vLLM releases new streaming parser for Qwen3+ in nightly
CORTIS: Text-Only Adaptation of Spoken Language Models
CORTIS enables task-oriented voice agents to generate structured speech outputs by fine-tuning spoken language models using only text-form task supervision. It outperforms ASR-LLM cascades under acoustic degradation, especially in preserving high-level task semantics, without requiring paired speech-target annotations during training.
ExecCritic uses role-specific RL to improve coding agents via test-guided repair
ExecCritic introduces a framework that separates test construction from source-code repair to prevent false confidence in coding agents. The system employs a Test agent and a Repair agent, both backed by Qwen-3.5-35B-A3B, trained separately using reinforcement learning.
Alibaba releases Qwen3.8-Flash-Next 176B preview weights for agentic coding
Alibaba has released the model weights for Qwen3.8-Flash-Next, serving as a preview of the upcoming Qwen4 architecture for developers to experiment with and evaluate.
Qwen 3.8 Max release highlights planning, simplification, and experimental design
The full release of Qwen 3.8 Max confirms early impressions that the model is exceptionally fast, highly capable at planning, and skilled at identifying unnecessary complexity in problems.
Qwen3.6-27b-mtp-q8 creates A* pathfinding implementation via autonomous testing
The Qwen3.6-27b-mtp-q8 model successfully generated an A* pathfinding implementation for a Java-based test game using Claude Code locally. The process involved nearly 12 hours of iterative development where the model autonomously created and ran a testing suite.