跳到正文
@testingcatalog· @testingcatalog · X·· 24 天前AI 评分38
AI 导读

OpenAI 近期发布的"Defender's gap"图表显示防御方与攻击方的能力差距,这一差距也被类比为闭源与开源 AI 模型之间的差距。Anthropic 与 OpenAI 希望扩大自身 AI 与全球其他模型的能力鸿沟,通过蒸馏限制等手段在 3-5 年内显著拉大美国对中国的领先优势。前沿实验室将继续用 RSI 推进内部模型,但在差距足够大且对齐标准达标前不会公开发布。

正文

This is a "Defender's gap" chart that OpenAI published recently. It shows the gap between defenders' capabilities and attackers' capabilities from a cybersecurity POV. This also translates to a gap between proprietary and open AI models.

What Dario is proposing is closely related: "Thus, a key part of pacing within democracies is to keep democracies’ AI lead over autocracies as large as possible, to give us the breathing room we need in order to pace effectively."

In other words, Anthropic and OpenAI want to widen the "gap" between what their AI can do and what the rest of the world can do.

This doesn't necessarily mean that they will stop AI development and training.

What this leads to is:
- Anthropic, OpenAI, and other frontier labs will need to work together to make sure that every lab maintains alignment standards.
- These labs will continue using RSI to advance their internal models with "employee-like access".
- These models WON'T be released to the public until the "gap" is sufficient and until alignment standards are met.

What about China?

> "Distillation of frontier models allows lagging companies to narrow the gap using a fraction of the cost it would take to develop their own AI independently."

> "If we execute these measures well, I believe they would slow China’s progress enough to widen America’s lead significantly over the next 3–5 years — the window when AI becomes geopolitically most important."

The assumption behind these measures is simple: without being able to distill frontier models, it will take China significantly longer to close the gap with top-tier models.

All the above may fay fail. China may or may not take the lead in AI progress.

Yet, it's not a surprise that AI can already be used as a cybersecurity weapon. Note that the top point on the "frontier" line describes defenders' capabilities available to companies with Daybreak access and similar. However, the top point on the "open-weight" line is accessible to everyone.

Even with current levels of intelligence, we will start seeing more and more major security incidents around the globe.

> “Everything that makes it successful is exactly what makes it dangerous.”

Monitoring 👀👀👀

来源:@testingcatalog · x.com