kapynResearch

AI benchmarks have a trust problem and Google wants to fix it

Google DeepMind is testing double-blind, cryptographically secured AI benchmark evaluations. The pilot with Singapore's AI Safety Institute uses Confidential Space to hide test questions from Google and model weights from evaluators. It could set a new standard for tamper-proof model evaluation.

The Decoder·Aug 28, 2026

Opening Kapyn…