跳到正文
@testingcatalog· @testingcatalog · X·· 13 天前AI 评分64
AI 导读

Anthropic 宣布将恢复对安全机制在 Claude 回复前拦截的请求计费,目前仅适用于误报率较低的三类场景:生物学、蒸馏攻击和前沿 LLM 开发。官方称近期测试中 99.7% 使用 Claude Code、Claude.ai 或 Cowork 的账户未触发这些新增的可计费拦截,相关分类器误报率低于 0.1%。若认为请求被错误拦截,可通过 Claude Code 的 /feedback 反馈。

正文

Anthropic will resume charging for rejected requests in order to protect from distillation attacks.

This applies to requests related to biology, distillation attacks and LLM development. https://t.co/pTaOW503KW https://t.co/F9y7F0rJNM

引用@ClaudeDevs@ClaudeDevs
Today, we'll resume charging for requests our safeguards block before Claude responds. This only applies in categories with low false positive rates: biology, distillation attacks, and frontier LLM development. We've seen some coordinated attacks on our systems in recent weeks, and this is one layer of defense. In recent testing, 99.7% of accounts using Claude Code, Claude​.ai, or Cowork did not hit any of these newly "billable blocks." The classifiers behind the blocks we’re resuming charging for today are tuned to have a <0.1% false positive rate. We know that's not 0%, and we're going to keep improving them so they interrupt your work less often. If you think a request has been blocked incorrectly, please report it with /feedback in Claude Code. https://t.co/uX6JhvN6He
在 X 查看被引用的帖子

来源:@testingcatalog · x.com