13:15
2026-08-28
the-decoder.com
ai-safety
AI benchmarks have a trust problem and Google wants to fix it
Google DeepMind is piloting a double-blind evaluation of a frontier AI model for the first time, using cryptographic protection through Confidential Space to prevent Google from seeing test questions …