Loading…

Google Deepmind is testing a double-blind evaluation of a frontier AI model for the first time. Cryptographic protection through Confidential Space is meant to keep Google from seeing the test questions and keep evaluators from seeing the model weights. The pilot project with…
To respect copyright, we link to the source rather than republishing the full text. Read the complete article on The Decoder.