论文研究普通
AI 基准测试存在信任问题,Google DeepMind 用密码学方法修复它
原始标题:AI benchmarks have a trust problem and Google wants to fix it
内容摘要
Google DeepMind 启动首个专有前沿 AI 模型的双盲评估试点,与新加坡 AI 安全研究所等合作,在 Gemini Flash Lite 系列模型上运行。
内容分类AI 论文与研究
内容层级普通情报
发布时间(北京时间)
本站收录时间(北京时间)
信息来源The Decoder:AI News(RSS)
站内情报编号intel-664ba97ed9eed1faf6b72edd