Lab · Zhipu AI
arxiv arXiv cs.LG · 1d ago HealthBench · 50.1% · 9 views

Fathom-Vaidya improves medical reasoning with rubric-based rewards

The authors introduce Fathom-Vaidya, a 30B parameter model that uses synthetic data and rubric-based reinforcement learning to enhance diagnostic and clinical healthcare reasoning. The training framework first applies rule-guided RL to MedBullets-derived questions for diagnosis, then utilizes 5.3k synthetic multi-turn scenarios with multi-dimensional rubrics for interactive clinical tasks.

arxiv arXiv cs.CL · 13d ago · 28 views

Single-direction attack on GLM-5.3-Flash reveals safety alignment fragility in MoE models

Researchers demonstrate that directional ablation, a white-box attack removing refusal by projecting out a single direction from weights, remains effective on frontier mixture-of-experts (MoE) models like GLM-5.3-Flash. The study shows that while the attack survives the architecture, its effects are distributed across attention, dense, and routed-expert writers rather than concentrated in one location.