跳到正文
@rohanpaul_ai· @rohanpaul_ai · X·· 19 天前AI 评分62
AI 导读

Anthropic 与 Accenture 合作开展前沿 AI 的独立评估,由 Accenture 旗下 Faculty 团队负责模型评估、红队测试、对齐评估与安全保障测试。双方预计未来五年各投入至少 10 亿美元建设相关能力,Anthropic 将直接向 Accenture 支付评估费用。作者指出,评估方由被评估公司付费,其独立性如何保持是尚未解决的问题。

正文

Anthropic is hiring Accenture to act like an outside safety inspector, but with unusually deep access inside Anthropic while Claude models are being built and tested.

The work, led by Accenture’s Faculty unit, will cover model evaluations, red-teaming, alignment assessments and safeguard testing.

Both companies expect to invest at least $1B each over 5 years in building AI-safety capacity, while Anthropic will directly pay Accenture for its evaluation work.

Accenture will try to break the models, check whether safety rules actually work, and examine whether Anthropic is following its own safety commitments.

The unusual part is that Anthropic is paying the evaluator itself, so the big unresolved question is how independent Accenture can remain while working so closely with the company it is judging.

引用@AnthropicAI@AnthropicAI
We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build capacity in this area over the next five years. https://t.co/SHVzjpgnfx
在 X 查看被引用的帖子

来源:@rohanpaul_ai · x.com