Report details first double-blind eval of LLM using secure enclaves
Miles_Brundage · x · 2026-08-27
AVERI released a pilot report detailing the world's first double-blind evaluation of a proprietary language model, Gemini 2.5 Flash-Lite, in collaboration with Google DeepMind, OpenMined, and MLCommons. The evaluation used a secure enclave and Trusted Execution Environment (TEE) to ensure model weights and test prompts remained hidden from each other, addressing benchmark contamination and structural issues in independent auditing.
More from Safety
- Shared Agent Skill Libraries Propagate Malware, 41.8% Self-Poisoning Rate Found — omarsar0 · 2026-08-27
- Matthew Green questions OpenAI security awareness — matthew_d_green · 2026-08-27
- METR report footnote suggests more third parties compromised in HF incident — GarrisonLovely · 2026-08-27
- FDA-cleared AI sepsis detection tool helps clinicians identify infections earlier — mdredze · 2026-08-27
- Essay: AI-era information intermediaries pose systemic danger even in careful hands — sebkrier · 2026-08-27
- OpenAI researcher warns ultrafast AI could outpace security teams — The Decoder · 2026-08-27