跳到正文
@kimmonismus· @kimmonismus · X·· 19 天前精选AI 评分72
AI 导读

Anthropic 公布内部 AI 研发自动化指标,Claude 主导的模型研发任务占比在约六个月内从 1% 升至 26%。Anthropic 称 Claude 还参与或主导其超过 90% 的模型研发工作,公司主要内部智能体平台上任意时刻约有 3 万个 AI 智能体在从事研究与工程任务。Anthropic 表示任何前沿实验室都可发布同样的度量,第三方也可以验证,完整方法与文章已在官网发布。

推荐理由

Anthropic 公开内部研发自动化指标与方法,读者可据此观察 Claude 参与模型研发的比例变化。

正文

In around six months, the share of AI model R&D tasks at Anthropic led by Claude has risen from 1% to 26%.

Speaking about RSI, Claude also collaborates on or leads more than 90% of Anthropic’s model R&D work. Around 30,000 AI agents are working on research and engineering at any given time on its main internal agent platform.

引用@AnthropicAI@AnthropicAI
AI systems are getting more powerful, and they're increasingly being used to build the next version of themselves. We want to illuminate that progress for the public. Today, we're sharing three measurements that help track AI development: 1. How much AI R&D is done by AI. 2. How well AI agents are overseen. 3. How compute is allocated. We provide a snapshot of these metrics from inside Anthropic. Any frontier developer could publish the same measures, and third parties could verify them. As the world considers pacing the frontier, we should do everything possible to minimize the gap between what frontier labs know and what the public knows. This means better measuring the development of AI, publishing our findings, and giving society an opportunity to decide how to use this information. Read the full post and methodology: https://t.co/iPFz8Z4ugE
在 X 查看被引用的帖子

来源:@kimmonismus · x.com