注册并分享邀请链接,可获得视频播放与邀请奖励。

🚨 AI News | TestingCatalog
@testingcatalog
Reporting AI nonsense. Latest AI News on AI Agents, Model Releases, Tools, Leaks, and Rumors 🗞️
加入 October 2015
997 正在关注    77.7K 粉丝
Anthropic will resume charging for rejected requests in order to protect from distillation attacks. This applies to requests related to biology, distillation attacks and LLM development.
显示更多
Today, we'll resume charging for requests our safeguards block before Claude responds. This only applies in categories with low false positive rates: biology, distillation attacks, and frontier LLM development. We've seen some coordinated attacks on our systems in recent weeks, and this is one layer of defense. In recent testing, 99.7% of accounts using Claude Code, Claude​.ai, or Cowork did not hit any of these newly "billable blocks." The classifiers behind the blocks we’re resuming charging for today are tuned to have a <0.1% false positive rate. We know that's not 0%, and we're going to keep improving them so they interrupt your work less often. If you think a request has been blocked incorrectly, please report it with /feedback in Claude Code.
显示更多