AI benchmarks have a trust problem and Google wants to fix it
The Decoderen

Google Deepmind is testing a double-blind evaluation of a frontier AI model for the first time. Cryptographic protection through Confidential Space is meant to keep Google from seeing the test questions and keep evaluators from seeing the model weights. The pilot project with the Singapore AI Safety Institute uses a Gemini Flash Lite and could set a new standard for tamper-proof AI benchmarks. The article AI benchmarks have a trust problem and Google wants to fix it appeared first on The Decoder .
This is a short summary published by AI Global Wire. The full article is owned and hosted by The Decoder — open it there to read it in full.
Read the full story at The Decoder- Verktyg
- Företag
Related AI news
- [Ekstra] Cybersikkerhetsekspert: Ny Chat GPT-funksjon er en gigantisk bakdør til Iphonen dindigi.no · August 28, 2026
- Nvidia eldar på lånefesten: ”Efterfrågan länge till”DI Digital · August 28, 2026
- How to get free Google AI Pro for an entire year - and save $240: 3 waysZDNET AI · August 28, 2026
- U.S. court rules Pentagon's blacklisting of Anthropic was unlawfulThe Decoder · August 28, 2026
- How to free Google AI Pro for an entire year - and save $240: 3 waysZDNET AI · August 28, 2026
- Beatport blocks fully AI-generated music from its DJ marketplaceThe Decoder · August 28, 2026