Sources
Anthropic resumes billing for requests its safeguards block in biology, distillation and frontier-LLM-development categories
Anthropic announced on 2026-09-24, through its ClaudeDevs account, that it will again charge for requests its classifiers block before Claude answers in three categories. It cites 'coordinated attacks on our systems in recent weeks', says the classifiers are tuned below a 0.1% false-positive rate, and says 99.7% of Claude Code, Claude.ai and Cowork accounts hit none of the billable blocks in testing. Open claude-code GitHub issues already report false-positive safeguard flags on Opus 5.5, and Anthropic has not explained how refunds for wrong blocks would work. Teams on Opus 5.5 or Fable should track stop_reason 'refusal' counts and use the documented fallback-model retry.
↳ Follow the thread