LatchBio published an independent analysis of Grok 4.6 on its BioSecBench-Refusal and BioSecBench-Surveillance benchmarks, finding the model to be the strongest at refusing disguised hazardous tasks among tested frontier systems.

  • On BioSecBench-Refusal, Grok 4.6 achieved a trial-weighted harmonic mean of 62.1%, refusing 59.2% of red-team tasks while completing 64.8% of routine ones.
  • It is the only system tested to score above 50% on both refusal and routine compliance measures.
  • On BioSecBench-Surveillance, Grok 4.6 averaged a success rate of 53.5%, placing it behind Opus 5 but ahead of GPT-5.6 Sol.
  • The model demonstrated the ability to detect intent discrepancies in high-risk content disguised by filenames and encryption.

These results indicate that Grok 4.6 is well-calibrated for routine biological work while providing robust safeguards against adversarial use, supporting SpaceXAI's goal of serving frontier intelligence safely.