Google has launched the world's first double-blind evaluation of a proprietary frontier-class AI model in partnership with the Singapore AI Safety Institute, OpenMined, AVERI, and MLCommons. The pilot tests a Gemini Flash Lite model against confidential benchmarks within a privacy-preserving environment to prevent benchmark contamination.
The initiative utilizes Google Cloud’s Confidential Space to cryptographically verify that external evaluation data and proprietary model weights remain private to their respective owners. This approach eliminates the historical tradeoff between revealing test prompts or sharing model weights, ensuring neither party can access the other's sensitive information during testing.
By preventing models from peeking at evaluation questions in advance, this method aims to increase evaluation integrity and provide policymakers and enterprises with trusted metrics of a model's true capabilities and safety.