跳到正文
The Decoder· Maximilian Schreiner·· 2026-08-28AI 评分48

Google DeepMind 首次测试前沿 AI 模型双盲评测,用 Confidential Space 加密防篡改

AI benchmarks have a trust problem and Google wants to fix it

阅读原文

本站未展示全文,请前往来源网站阅读。

AI 导读

Google DeepMind 正首次测试前沿 AI 模型的双盲评测,通过 Confidential Space 加密保护,让 Google 看不到测试题目、评测方看不到模型权重。该试点项目与新加坡 AI Safety Institute 合作,使用 Gemini Flash Lite,有望为防篡改 AI 基准测试树立新标准。

来源:The Decoder · the-decoder.com