Anthropic 宣布恢复对 Claude 响应前被安全系统拦截的请求收费,仅适用于误报率低的类别:生物学、蒸馏攻击和前沿 LLM 开发。官方称这是对近期协同攻击的一层防御,测试中 99.7% 的 Claude Code、Claude.ai 或 Cowork 账户未触发这些计费拦截,相关分类器误报率控制在 <0.1%,误判可通过 Claude Code 中 /feedback 反馈。
Today, we'll resume charging for requests our safeguards block before Claude responds. This only applies in categories with low false positive rates: biology, distillation attacks, and frontier LLM development. We've seen some coordinated attacks on our systems in recent weeks, and this is one layer of defense.
In recent testing, 99.7% of accounts using Claude Code, Claude.ai, or Cowork did not hit any of these newly "billable blocks." The classifiers behind the blocks we’re resuming charging for today are tuned to have a <0.1% false positive rate. We know that's not 0%, and we're going to keep improving them so they interrupt your work less often. If you think a request has been blocked incorrectly, please report it with /feedback in Claude Code. https://platform.claude.com/docs/en/build-with-claude/refusals-and-fallback#how-refusals-are-billed
来源:ClaudeDevs · x.com