跳到正文
@GoogleDeepMind· @GoogleDeepMind · X·· 2026-09-03AI 评分33
AI 导读

在 CyberGym 等基准测试中,该模型在自主发现弱点方面领先,同时保持快速高效。 在 @GoogleChrome 代码库的真实测试中,它产出了 2.6 倍的有效修复,帮助更快地保护软件。https://t.co/VqewNeIXQ8

正文

On benchmarks like CyberGym, the model leads in finding weaknesses autonomously while staying fast and efficient.

In real-world testing across @GoogleChrome codebases, it produced 2.6 times more valid fixes to help protect software faster. https://t.co/VqewNeIXQ8

来源:@GoogleDeepMind · x.com