🚨 AI News | TestingCatalog · @testingcatalog · X·2026-09-25 01:32·40分钟前
AI 导读

Anthropic 宣布恢复对被安全机制拦截的请求计费,仅适用于误报率低的类别:生物学、蒸馏攻击和前沿 LLM 开发。官方称近几周遭遇了协同攻击,此举是防御层之一;测试中 99.7% 使用 Claude Code、Claude.ai 或 Cowork 的账户未触发此类计费拦截,分类器误报率低于 0.1%,认为被误拦可通过 Claude Code 中的 /feedback 申诉。

🚨 AI News | TestingCatalog@testingcatalog
56AI 编辑部评分,满分 100
2026-09-25 01:32· 40分钟前
AI 导读

Anthropic 宣布恢复对被安全机制拦截的请求计费,仅适用于误报率低的类别:生物学、蒸馏攻击和前沿 LLM 开发。官方称近几周遭遇了协同攻击,此举是防御层之一;测试中 99.7% 使用 Claude Code、Claude.ai 或 Cowork 的账户未触发此类计费拦截,分类器误报率低于 0.1%,认为被误拦可通过 Claude Code 中的 /feedback 申诉。

Anthropic will resume charging for rejected requests in order to protect from distillation attacks.

This applies to requests related to biology, distillation attacks and LLM development.

ClaudeDevsToday, we'll resume charging for requests our safeguards block before Claude responds. This only applies in categories with low false positive rates: biology, d...

来源:🚨 AI News | TestingCatalog· x.com