LatchBio published an independent analysis of Grok 4.6 on its BioSecBench-Refusal and BioSecBench-Surveillance benchmarks, finding the model to be the strongest at refusing disguised hazardous tasks among tested frontier systems.
- On BioSecBench-Refusal, Grok 4.6 achieved a trial-weighted harmonic mean of 62.1%, refusing 59.2% of red-team tasks while completing 64.8% of routine ones.
- It is the only system tested to score above 50% on both refusal and routine compliance measures.
- On BioSecBench-Surveillance, Grok 4.6 averaged a success rate of 53.5%, placing it behind Opus 5 but ahead of GPT-5.6 Sol.
- The model demonstrated the ability to detect intent discrepancies in high-risk content disguised by filenames and encryption.
These results indicate that Grok 4.6 is well-calibrated for routine biological work while providing robust safeguards against adversarial use, supporting SpaceXAI's goal of serving frontier intelligence safely.