Anthropic 宣布恢复对 Claude 响应前被安全机制拦截的请求计费,范围限于误报率低的类别,包括生物、蒸馏攻击和前沿 LLM 开发。官方称近期出现针对其系统的协同攻击,这是防御手段之一;测试中 99.7% 使用 Claude Code、Claude.ai 或 Cowork 的账户未触发这类计费拦截,相关分类器误报率低于 0.1%。
Today, we'll resume charging for requests our safeguards block before Claude responds. This only applies in categories with low false positive rates: biology, distillation attacks, and frontier LLM development. We've seen some coordinated attacks on our systems in recent weeks, and this is one layer of defense.
In recent testing, 99.7% of accounts using Claude Code, Claude.ai, or Cowork did not hit any of these newly "billable blocks." The classifiers behind the blocks we’re resuming charging for today are tuned to have a <0.1% false positive rate. We know that's not 0%, and we're going to keep improving them so they interrupt your work less often. If you think a request has been blocked incorrectly, please report it with /feedback in Claude Code. https://t.co/uX6JhvN6He
来源:@ClaudeDevs · x.com